# Agents

## List Agents

`client.agents.list(AgentListParamsquery?, RequestOptionsoptions?): ArrayPage<AgentState>`

**get** `/v1/agents/`

Get a list of all agents.

### Parameters

- `query: AgentListParams`

  - `after?: string | null`

    Cursor for pagination

  - `ascending?: boolean`

    Whether to sort agents oldest to newest (True) or newest to oldest (False, default)

  - `base_template_id?: string | null`

    Search agents by base template ID

  - `before?: string | null`

    Cursor for pagination

  - `created_by_id?: string | null`

    Filter agents by the user who created them.

  - `identifier_keys?: Array<string> | null`

    Search agents by identifier keys

  - `identity_id?: string | null`

    Search agents by identity ID

  - `include?: Array<"agent.blocks" | "agent.identities" | "agent.managed_group" | 5 more> | null`

    Specify which relational fields to include in the response. No relationships are included by default.

    - `"agent.blocks"`

    - `"agent.identities"`

    - `"agent.managed_group"`

    - `"agent.pending_approval"`

    - `"agent.secrets"`

    - `"agent.sources"`

    - `"agent.tags"`

    - `"agent.tools"`

  - `include_relationships?: Array<string> | null`

    Specify which relational fields (e.g., 'tools', 'sources', 'memory') to include in the response. If not provided, all relationships are loaded by default. Using this can optimize performance by reducing unnecessary joins.This is a legacy parameter, and no longer supported after 1.0.0 SDK versions.

  - `last_stop_reason?: StopReasonType | null`

    Filter agents by their last stop reason.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `limit?: number | null`

    Limit for pagination

  - `match_all_tags?: boolean`

    If True, only returns agents that match ALL given tags. Otherwise, return agents that have ANY of the passed-in tags.

  - `name?: string | null`

    Name of the agent

  - `order?: "asc" | "desc"`

    Sort order for agents by creation time. 'asc' for oldest first, 'desc' for newest first

    - `"asc"`

    - `"desc"`

  - `order_by?: "created_at" | "updated_at" | "last_run_completion"`

    Field to sort by

    - `"created_at"`

    - `"updated_at"`

    - `"last_run_completion"`

  - `project_id?: string | null`

    Search agents by project ID - this will default to your default project on cloud

  - `query_text?: string | null`

    Search agents by name

  - `sort_by?: string | null`

    Field to sort by. Options: 'created_at' (default), 'last_run_completion'

  - `tags?: Array<string> | null`

    List of tags to filter agents by

  - `template_id?: string | null`

    Search agents by template ID

### Returns

- `AgentState`

  Representation of an agent's state. This is the state of the agent at a given time, and is persisted in the DB backend. The state has all the information needed to recreate a persisted agent.

  - `id: string`

    The id of the agent. Assigned by the database.

  - `agent_type: AgentType`

    The type of agent.

    - `"memgpt_agent"`

    - `"memgpt_v2_agent"`

    - `"letta_v1_agent"`

    - `"react_agent"`

    - `"workflow_agent"`

    - `"split_thread_agent"`

    - `"sleeptime_agent"`

    - `"voice_convo_agent"`

    - `"voice_sleeptime_agent"`

  - `blocks: Array<Block>`

    The memory blocks used by the agent.

    - `value: string`

      Value of the block.

    - `id?: string`

      The human-friendly ID of the Block

    - `base_template_id?: string | null`

      The base template id of the block.

    - `created_by_id?: string | null`

      The id of the user that made this Block.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      Description of the block.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the block will be hidden.

    - `is_template?: boolean`

      Whether the block is a template (e.g. saved human/persona options).

    - `label?: string | null`

      Label of the block (e.g. 'human', 'persona') in the context window.

    - `last_updated_by_id?: string | null`

      The id of the user that last updated this Block.

    - `limit?: number`

      Character limit of the block.

    - `metadata?: Record<string, unknown> | null`

      Metadata of the block.

    - `preserve_on_migration?: boolean | null`

      Preserve the block on template migration.

    - `project_id?: string | null`

      The associated project id.

    - `read_only?: boolean`

      Whether the agent has read-only access to the block.

    - `tags?: Array<string> | null`

      The tags associated with the block.

    - `template_id?: string | null`

      The id of the template.

    - `template_name?: string | null`

      Name of the block if it is a template.

  - `llm_config: LlmConfig`

    Deprecated: Use `model` field instead. The LLM configuration used by the agent.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

        - `type?: "text"`

          The type of the response format.

          - `"text"`

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

        - `json_schema: Record<string, unknown>`

          The JSON schema of the response.

        - `type?: "json_schema"`

          The type of the response format.

          - `"json_schema"`

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

        - `type?: "json_object"`

          The type of the response format.

          - `"json_object"`

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `memory: Memory`

    Deprecated: Use `blocks` field instead. The in-context memory of the agent.

    - `blocks: Array<Block>`

      Memory blocks contained in the agent's in-context memory

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `agent_type?: AgentType | (string & {}) | null`

      Agent type controlling prompt rendering.

      - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

        Enum to represent the type of agent.

      - `(string & {})`

    - `file_blocks?: Array<FileBlock>`

      Special blocks representing the agent's in-context memory of an attached file

      - `file_id: string`

        Unique identifier of the file.

      - `is_open: boolean`

        True if the agent currently has the file open.

      - `source_id: string`

        Deprecated: Use `folder_id` field instead. Unique identifier of the source.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_accessed_at?: string | null`

        UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `git_enabled?: boolean`

      Whether this agent uses git-backed memory with structured labels.

    - `prompt_template?: string`

      Deprecated. Ignored for performance.

  - `name: string`

    The name of the agent.

  - `sources: Array<Source>`

    Deprecated: Use `folders` field instead. The sources used by the agent.

    - `id: string`

      The human-friendly ID of the Source

    - `embedding_config: EmbeddingConfig`

      The embedding configuration used by the source.

      - `embedding_dim: number`

        The dimension of the embedding.

      - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"bedrock"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"mistral"`

        - `"together"`

        - `"pinecone"`

      - `embedding_model: string`

        The model for the embedding.

      - `azure_deployment?: string | null`

        The Azure deployment for the model.

      - `azure_endpoint?: string | null`

        The Azure endpoint for the model.

      - `azure_version?: string | null`

        The Azure version for the model.

      - `batch_size?: number`

        The maximum batch size for processing embeddings.

      - `embedding_chunk_size?: number | null`

        The chunk size of the embedding.

      - `embedding_endpoint?: string | null`

        The endpoint for the model (`None` if local).

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

    - `name: string`

      The name of the source.

    - `created_at?: string | null`

      The timestamp when the source was created.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `description?: string | null`

      The description of the source.

    - `instructions?: string | null`

      Instructions for how to use the source.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata?: Record<string, unknown> | null`

      Metadata associated with the source.

    - `updated_at?: string | null`

      The timestamp when the source was last updated.

    - `vector_db_provider?: VectorDBProvider`

      The vector database provider used for this source's passages

      - `"native"`

      - `"tpuf"`

      - `"pinecone"`

  - `system: string`

    The system prompt used by the agent.

  - `tags: Array<string>`

    The tags associated with the agent.

  - `tools: Array<Tool>`

    The tools used by the agent.

    - `id: string`

      The human-friendly ID of the Tool

    - `args_json_schema?: Record<string, unknown> | null`

      The args JSON schema of the function.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `default_requires_approval?: boolean | null`

      Default value for whether or not executing this tool requires approval.

    - `description?: string | null`

      The description of the tool.

    - `enable_parallel_execution?: boolean | null`

      If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

    - `json_schema?: Record<string, unknown> | null`

      The JSON schema of the function.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata_?: Record<string, unknown> | null`

      A dictionary of additional metadata for the tool.

    - `name?: string | null`

      The name of the function.

    - `npm_requirements?: Array<NpmRequirement> | null`

      Optional list of npm packages required by this tool.

      - `name: string`

        Name of the npm package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `pip_requirements?: Array<PipRequirement> | null`

      Optional list of pip packages required by this tool.

      - `name: string`

        Name of the pip package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `project_id?: string | null`

      The project id of the tool.

    - `return_char_limit?: number`

      The maximum number of characters in the response.

    - `source_code?: string | null`

      The source code of the function.

    - `source_type?: string | null`

      The type of the source code.

    - `tags?: Array<string>`

      Metadata tags.

    - `tool_type?: ToolType`

      The type of the tool.

      - `"custom"`

      - `"letta_core"`

      - `"letta_memory_core"`

      - `"letta_multi_agent_core"`

      - `"letta_sleeptime_core"`

      - `"letta_voice_sleeptime_core"`

      - `"letta_builtin"`

      - `"letta_files_core"`

      - `"external_langchain"`

      - `"external_composio"`

      - `"external_mcp"`

  - `base_template_id?: string | null`

    The base template id of the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `created_at?: string | null`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the agent will be hidden.

  - `identities?: Array<Identity>`

    The identities associated with this agent.

    - `id: string`

      The human-friendly ID of the Identity

    - `agent_ids: Array<string>`

      The IDs of the agents associated with the identity.

    - `block_ids: Array<string>`

      The IDs of the blocks associated with the identity.

    - `identifier_key: string`

      External, user-generated identifier key of the identity.

    - `identity_type: "org" | "user" | "other"`

      The type of the identity.

      - `"org"`

      - `"user"`

      - `"other"`

    - `name: string`

      The name of the identity.

    - `project_id?: string | null`

      The project id of the identity, if applicable.

    - `properties?: Array<Property>`

      List of properties associated with the identity

      - `key: string`

        The key of the property

      - `type: "string" | "number" | "boolean" | "json"`

        The type of the property

        - `"string"`

        - `"number"`

        - `"boolean"`

        - `"json"`

      - `value: string | number | boolean | Record<string, unknown>`

        The value of the property

        - `string`

        - `number`

        - `boolean`

        - `Record<string, unknown>`

  - `identity_ids?: Array<string>`

    Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

  - `last_run_completion?: string | null`

    The timestamp when the agent last completed a run.

  - `last_run_duration_ms?: number | null`

    The duration in milliseconds of the agent's last run.

  - `last_stop_reason?: StopReasonType | null`

    The stop reason from the agent's last run.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `managed_group?: ManagedGroup | null`

    The multi-agent group that this agent manages

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `message_buffer_autoclear?: boolean`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `message_ids?: Array<string> | null`

    The ids of the messages in the agent's in-context memory.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle used by the agent (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings used by the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `multi_agent_group?: MultiAgentGroup | null`

    Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `pending_approval?: ApprovalRequestMessage | null`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project_id?: string | null`

    The id of the project the agent belongs to.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format used by the agent

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Array<AgentEnvironmentVariable>`

    The environment variables for tool execution specific to this agent.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `template_id?: string | null`

    The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

    Deprecated: use `secrets` field instead.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The list of tool rules.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

// Automatically fetches more pages as needed.
for await (const agentState of client.agents.list()) {
  console.log(agentState.id);
}
```

#### Response

```json
[
  {
    "id": "id",
    "agent_type": "memgpt_agent",
    "blocks": [
      {
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "llm_config": {
      "context_window": 0,
      "model": "model",
      "model_endpoint_type": "openai",
      "compatibility_type": "gguf",
      "display_name": "display_name",
      "effort": "low",
      "enable_reasoner": true,
      "frequency_penalty": 0,
      "handle": "handle",
      "max_reasoning_tokens": 0,
      "max_tokens": 0,
      "model_endpoint": "model_endpoint",
      "model_wrapper": "model_wrapper",
      "parallel_tool_calls": true,
      "provider_category": "base",
      "provider_name": "provider_name",
      "put_inner_thoughts_in_kwargs": true,
      "reasoning_effort": "none",
      "response_format": {
        "type": "text"
      },
      "return_logprobs": true,
      "return_token_ids": true,
      "strict": true,
      "temperature": 0,
      "tier": "tier",
      "tool_call_parser": "tool_call_parser",
      "top_logprobs": 0,
      "verbosity": "low"
    },
    "memory": {
      "blocks": [
        {
          "value": "value",
          "id": "block-123e4567-e89b-12d3-a456-426614174000",
          "base_template_id": "base_template_id",
          "created_by_id": "created_by_id",
          "deployment_id": "deployment_id",
          "description": "description",
          "entity_id": "entity_id",
          "hidden": true,
          "is_template": true,
          "label": "label",
          "last_updated_by_id": "last_updated_by_id",
          "limit": 0,
          "metadata": {
            "foo": "bar"
          },
          "preserve_on_migration": true,
          "project_id": "project_id",
          "read_only": true,
          "tags": [
            "string"
          ],
          "template_id": "template_id",
          "template_name": "template_name"
        }
      ],
      "agent_type": "memgpt_agent",
      "file_blocks": [
        {
          "file_id": "file_id",
          "is_open": true,
          "source_id": "source_id",
          "value": "value",
          "id": "block-123e4567-e89b-12d3-a456-426614174000",
          "base_template_id": "base_template_id",
          "created_by_id": "created_by_id",
          "deployment_id": "deployment_id",
          "description": "description",
          "entity_id": "entity_id",
          "hidden": true,
          "is_template": true,
          "label": "label",
          "last_accessed_at": "2019-12-27T18:11:19.117Z",
          "last_updated_by_id": "last_updated_by_id",
          "limit": 0,
          "metadata": {
            "foo": "bar"
          },
          "preserve_on_migration": true,
          "project_id": "project_id",
          "read_only": true,
          "tags": [
            "string"
          ],
          "template_id": "template_id",
          "template_name": "template_name"
        }
      ],
      "git_enabled": true,
      "prompt_template": "prompt_template"
    },
    "name": "name",
    "sources": [
      {
        "id": "source-123e4567-e89b-12d3-a456-426614174000",
        "embedding_config": {
          "embedding_dim": 0,
          "embedding_endpoint_type": "openai",
          "embedding_model": "embedding_model",
          "azure_deployment": "azure_deployment",
          "azure_endpoint": "azure_endpoint",
          "azure_version": "azure_version",
          "batch_size": 0,
          "embedding_chunk_size": 0,
          "embedding_endpoint": "embedding_endpoint",
          "handle": "handle"
        },
        "name": "name",
        "created_at": "2019-12-27T18:11:19.117Z",
        "created_by_id": "created_by_id",
        "description": "description",
        "instructions": "instructions",
        "last_updated_by_id": "last_updated_by_id",
        "metadata": {
          "foo": "bar"
        },
        "updated_at": "2019-12-27T18:11:19.117Z",
        "vector_db_provider": "native"
      }
    ],
    "system": "system",
    "tags": [
      "string"
    ],
    "tools": [
      {
        "id": "tool-123e4567-e89b-12d3-a456-426614174000",
        "args_json_schema": {
          "foo": "bar"
        },
        "created_by_id": "created_by_id",
        "default_requires_approval": true,
        "description": "description",
        "enable_parallel_execution": true,
        "json_schema": {
          "foo": "bar"
        },
        "last_updated_by_id": "last_updated_by_id",
        "metadata_": {
          "foo": "bar"
        },
        "name": "name",
        "npm_requirements": [
          {
            "name": "x",
            "version": "version"
          }
        ],
        "pip_requirements": [
          {
            "name": "x",
            "version": "version"
          }
        ],
        "project_id": "project_id",
        "return_char_limit": 1,
        "source_code": "source_code",
        "source_type": "source_type",
        "tags": [
          "string"
        ],
        "tool_type": "custom"
      }
    ],
    "base_template_id": "base_template_id",
    "compaction_settings": {
      "clip_chars": 0,
      "mode": "all",
      "model": "model",
      "model_settings": {
        "max_output_tokens": 0,
        "parallel_tool_calls": true,
        "provider_type": "openai",
        "reasoning": {
          "reasoning_effort": "none"
        },
        "response_format": {
          "type": "text"
        },
        "strict": true,
        "temperature": 0
      },
      "prompt": "prompt",
      "prompt_acknowledgement": true,
      "sliding_window_percentage": 0
    },
    "created_at": "2019-12-27T18:11:19.117Z",
    "created_by_id": "created_by_id",
    "deployment_id": "deployment_id",
    "description": "description",
    "embedding": "embedding",
    "embedding_config": {
      "embedding_dim": 0,
      "embedding_endpoint_type": "openai",
      "embedding_model": "embedding_model",
      "azure_deployment": "azure_deployment",
      "azure_endpoint": "azure_endpoint",
      "azure_version": "azure_version",
      "batch_size": 0,
      "embedding_chunk_size": 0,
      "embedding_endpoint": "embedding_endpoint",
      "handle": "handle"
    },
    "enable_sleeptime": true,
    "entity_id": "entity_id",
    "hidden": true,
    "identities": [
      {
        "id": "identity-123e4567-e89b-12d3-a456-426614174000",
        "agent_ids": [
          "string"
        ],
        "block_ids": [
          "string"
        ],
        "identifier_key": "identifier_key",
        "identity_type": "org",
        "name": "name",
        "project_id": "project_id",
        "properties": [
          {
            "key": "key",
            "type": "string",
            "value": "string"
          }
        ]
      }
    ],
    "identity_ids": [
      "string"
    ],
    "last_run_completion": "2019-12-27T18:11:19.117Z",
    "last_run_duration_ms": 0,
    "last_stop_reason": "end_turn",
    "last_updated_by_id": "last_updated_by_id",
    "managed_group": {
      "id": "id",
      "agent_ids": [
        "string"
      ],
      "description": "description",
      "manager_type": "round_robin",
      "base_template_id": "base_template_id",
      "deployment_id": "deployment_id",
      "hidden": true,
      "last_processed_message_id": "last_processed_message_id",
      "manager_agent_id": "manager_agent_id",
      "max_message_buffer_length": 0,
      "max_turns": 0,
      "min_message_buffer_length": 0,
      "project_id": "project_id",
      "shared_block_ids": [
        "string"
      ],
      "sleeptime_agent_frequency": 0,
      "template_id": "template_id",
      "termination_token": "termination_token",
      "turns_counter": 0
    },
    "max_files_open": 0,
    "message_buffer_autoclear": true,
    "message_ids": [
      "string"
    ],
    "metadata": {
      "foo": "bar"
    },
    "model": "model",
    "model_settings": {
      "max_output_tokens": 0,
      "parallel_tool_calls": true,
      "provider_type": "openai",
      "reasoning": {
        "reasoning_effort": "none"
      },
      "response_format": {
        "type": "text"
      },
      "strict": true,
      "temperature": 0
    },
    "multi_agent_group": {
      "id": "id",
      "agent_ids": [
        "string"
      ],
      "description": "description",
      "manager_type": "round_robin",
      "base_template_id": "base_template_id",
      "deployment_id": "deployment_id",
      "hidden": true,
      "last_processed_message_id": "last_processed_message_id",
      "manager_agent_id": "manager_agent_id",
      "max_message_buffer_length": 0,
      "max_turns": 0,
      "min_message_buffer_length": 0,
      "project_id": "project_id",
      "shared_block_ids": [
        "string"
      ],
      "sleeptime_agent_frequency": 0,
      "template_id": "template_id",
      "termination_token": "termination_token",
      "turns_counter": 0
    },
    "pending_approval": {
      "id": "id",
      "date": "2019-12-27T18:11:19.117Z",
      "tool_call": {
        "arguments": "arguments",
        "name": "name",
        "tool_call_id": "tool_call_id"
      },
      "is_err": true,
      "message_type": "approval_request_message",
      "name": "name",
      "otid": "otid",
      "run_id": "run_id",
      "sender_id": "sender_id",
      "seq_id": 0,
      "step_id": "step_id",
      "tool_calls": [
        {
          "arguments": "arguments",
          "name": "name",
          "tool_call_id": "tool_call_id"
        }
      ]
    },
    "per_file_view_window_char_limit": 0,
    "project_id": "project_id",
    "response_format": {
      "type": "text"
    },
    "secrets": [
      {
        "agent_id": "agent_id",
        "key": "key",
        "value": "value",
        "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
        "created_at": "2019-12-27T18:11:19.117Z",
        "created_by_id": "created_by_id",
        "description": "description",
        "last_updated_by_id": "last_updated_by_id",
        "updated_at": "2019-12-27T18:11:19.117Z",
        "value_enc": "value_enc"
      }
    ],
    "template_id": "template_id",
    "timezone": "timezone",
    "tool_exec_environment_variables": [
      {
        "agent_id": "agent_id",
        "key": "key",
        "value": "value",
        "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
        "created_at": "2019-12-27T18:11:19.117Z",
        "created_by_id": "created_by_id",
        "description": "description",
        "last_updated_by_id": "last_updated_by_id",
        "updated_at": "2019-12-27T18:11:19.117Z",
        "value_enc": "value_enc"
      }
    ],
    "tool_rules": [
      {
        "children": [
          "string"
        ],
        "tool_name": "tool_name",
        "child_arg_nodes": [
          {
            "name": "name",
            "args": {
              "foo": "bar"
            }
          }
        ],
        "prompt_template": "prompt_template",
        "type": "constrain_child_tools"
      }
    ],
    "updated_at": "2019-12-27T18:11:19.117Z"
  }
]
```

## Create Agent

`client.agents.create(AgentCreateParamsbody, RequestOptionsoptions?): AgentState`

**post** `/v1/agents/`

Create an agent.

### Parameters

- `body: AgentCreateParams`

  - `agent_type?: AgentType`

    The type of agent.

    - `"memgpt_agent"`

    - `"memgpt_v2_agent"`

    - `"letta_v1_agent"`

    - `"react_agent"`

    - `"workflow_agent"`

    - `"split_thread_agent"`

    - `"sleeptime_agent"`

    - `"voice_convo_agent"`

    - `"voice_sleeptime_agent"`

  - `base_template_id?: string | null`

    Deprecated: No longer used. The base template id of the agent.

  - `block_ids?: Array<string> | null`

    The ids of the blocks used by the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

            - `type?: "text"`

              The type of the response format.

              - `"text"`

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

            - `json_schema: Record<string, unknown>`

              The JSON schema of the response.

            - `type?: "json_schema"`

              The type of the response format.

              - `"json_schema"`

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

            - `type?: "json_object"`

              The type of the response format.

              - `"json_object"`

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `context_window_limit?: number | null`

    The context window limit used by the agent.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_chunk_size?: number | null`

    Deprecated: No longer used. The embedding chunk size used by the agent.

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

    - `embedding_dim: number`

      The dimension of the embedding.

    - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"bedrock"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"mistral"`

      - `"together"`

      - `"pinecone"`

    - `embedding_model: string`

      The model for the embedding.

    - `azure_deployment?: string | null`

      The Azure deployment for the model.

    - `azure_endpoint?: string | null`

      The Azure endpoint for the model.

    - `azure_version?: string | null`

      The Azure version for the model.

    - `batch_size?: number`

      The maximum batch size for processing embeddings.

    - `embedding_chunk_size?: number | null`

      The chunk size of the embedding.

    - `embedding_endpoint?: string | null`

      The endpoint for the model (`None` if local).

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

  - `enable_reasoner?: boolean | null`

    Deprecated: Use `model` field to configure reasoning instead. Whether to enable internal extended thinking step for a reasoner model.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `folder_ids?: Array<string> | null`

    The ids of the folders used by the agent.

  - `from_template?: string | null`

    Deprecated: please use the 'create agents from a template' endpoint instead.

  - `hidden?: boolean | null`

    Deprecated: No longer used. If set to True, the agent will be hidden.

  - `identity_ids?: Array<string> | null`

    The ids of the identities associated with this agent.

  - `include_base_tool_rules?: boolean | null`

    If true, attaches the Letta base tool rules (e.g. deny all tools not explicitly allowed).

  - `include_base_tools?: boolean`

    If true, attaches the Letta core tools (e.g. core_memory related functions).

  - `include_default_source?: boolean`

    If true, automatically creates and attaches a default data source for this agent.

  - `initial_message_sequence?: Array<MessageCreate> | null`

    The initial set of messages to put in the agent's in-context memory.

    - `content: Array<LettaMessageContentUnion> | string`

      The content of the message.

      - `Array<LettaMessageContentUnion>`

        - `TextContent`

          - `text: string`

            The text content of the message.

          - `signature?: string | null`

            Stores a unique identifier for any reasoning associated with this text content.

          - `type?: "text"`

            The type of the message.

            - `"text"`

        - `ImageContent`

          - `source: URLImage | Base64Image | LettaImage`

            The source of the image.

            - `URLImage`

              - `url: string`

                The URL of the image.

              - `type?: "url"`

                The source type for the image.

                - `"url"`

            - `Base64Image`

              - `data: string`

                The base64 encoded image data.

              - `media_type: string`

                The media type for the image.

              - `detail?: string | null`

                What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

              - `type?: "base64"`

                The source type for the image.

                - `"base64"`

            - `LettaImage`

              - `file_id: string`

                The unique identifier of the image file persisted in storage.

              - `data?: string | null`

                The base64 encoded image data.

              - `detail?: string | null`

                What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

              - `media_type?: string | null`

                The media type for the image.

              - `type?: "letta"`

                The source type for the image.

                - `"letta"`

          - `type?: "image"`

            The type of the message.

            - `"image"`

        - `ToolCallContent`

          - `id: string`

            A unique identifier for this specific tool call instance.

          - `input: Record<string, unknown>`

            The parameters being passed to the tool, structured as a dictionary of parameter names to values.

          - `name: string`

            The name of the tool being called.

          - `signature?: string | null`

            Stores a unique identifier for any reasoning associated with this tool call.

          - `type?: "tool_call"`

            Indicates this content represents a tool call event.

            - `"tool_call"`

        - `ToolReturnContent`

          - `content: string`

            The content returned by the tool execution.

          - `is_error: boolean`

            Indicates whether the tool execution resulted in an error.

          - `tool_call_id: string`

            References the ID of the ToolCallContent that initiated this tool call.

          - `type?: "tool_return"`

            Indicates this content represents a tool return event.

            - `"tool_return"`

        - `ReasoningContent`

          Sent via the Anthropic Messages API

          - `is_native: boolean`

            Whether the reasoning content was generated by a reasoner model that processed this step.

          - `reasoning: string`

            The intermediate reasoning or thought process content.

          - `signature?: string | null`

            A unique identifier for this reasoning step.

          - `type?: "reasoning"`

            Indicates this is a reasoning/intermediate step.

            - `"reasoning"`

        - `RedactedReasoningContent`

          Sent via the Anthropic Messages API

          - `data: string`

            The redacted or filtered intermediate reasoning content.

          - `type?: "redacted_reasoning"`

            Indicates this is a redacted thinking step.

            - `"redacted_reasoning"`

        - `OmittedReasoningContent`

          A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

          - `signature?: string | null`

            A unique identifier for this reasoning step.

          - `type?: "omitted_reasoning"`

            Indicates this is an omitted reasoning step.

            - `"omitted_reasoning"`

      - `string`

    - `role: "user" | "system" | "assistant"`

      The role of the participant.

      - `"user"`

      - `"system"`

      - `"assistant"`

    - `batch_item_id?: string | null`

      The id of the LLMBatchItem that this message is associated with

    - `group_id?: string | null`

      The multi-agent group that the message was sent in

    - `name?: string | null`

      The name of the participant.

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `sender_id?: string | null`

      The id of the sender of the message, can be an identity id or agent id

    - `type?: "message" | null`

      The message type to be created.

      - `"message"`

  - `llm_config?: LlmConfig | null`

    Configuration for Language Model (LLM) connection and generation parameters.

    .. deprecated::
    LLMConfig is deprecated and should not be used as an input or return type in API calls.
    Use the schemas in letta.schemas.model (ModelSettings, OpenAIModelSettings, etc.) instead.
    For conversion, use the _to_model() method or Model._from_llm_config() method.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `max_reasoning_tokens?: number | null`

    Deprecated: Use `model` field to configure reasoning tokens instead. The maximum number of tokens to generate for reasoning step.

  - `max_tokens?: number | null`

    Deprecated: Use `model` field to configure max output tokens instead. The maximum number of tokens to generate, including reasoning step.

  - `memory_blocks?: Array<CreateBlock> | null`

    The blocks to create in the agent's in-context memory.

    - `label: string`

      Label of the block.

    - `value: string`

      Value of the block.

    - `base_template_id?: string | null`

      The base template id of the block.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      Description of the block.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the block will be hidden.

    - `is_template?: boolean`

    - `limit?: number`

      Character limit of the block.

    - `metadata?: Record<string, unknown> | null`

      Metadata of the block.

    - `preserve_on_migration?: boolean | null`

      Preserve the block on template migration.

    - `project_id?: string | null`

      The associated project id.

    - `read_only?: boolean`

      Whether the agent has read-only access to the block.

    - `tags?: Array<string> | null`

      The tags to associate with the block.

    - `template_id?: string | null`

      The id of the template.

    - `template_name?: string | null`

      Name of the block if it is a template.

  - `memory_variables?: Record<string, string> | null`

    Deprecated: Only relevant for creating agents from a template. Use the 'create agents from a template' endpoint instead.

  - `message_buffer_autoclear?: boolean`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle for the agent to use (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings for the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `name?: string`

    The name of the agent.

  - `parallel_tool_calls?: boolean | null`

    Deprecated: Use `model_settings` to configure parallel tool calls instead. If set to True, enables parallel tool calling.

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project?: string | null`

    Deprecated: Project should now be passed via the X-Project header instead of in the request body. If using the SDK, this can be done via the x_project parameter.

  - `project_id?: string | null`

    Deprecated: No longer used. The id of the project the agent belongs to.

  - `reasoning?: boolean | null`

    Deprecated: Use `model` field to configure reasoning instead. Whether to enable reasoning for this agent.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    Deprecated: Use `model_settings` field to configure response format instead. The response format for the agent.

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Record<string, string> | null`

    The environment variables for tool execution specific to this agent.

  - `source_ids?: Array<string> | null`

    Deprecated: Use `folder_ids` field instead. The ids of the sources used by the agent.

  - `system?: string | null`

    The system prompt used by the agent.

  - `tags?: Array<string> | null`

    The tags associated with the agent.

  - `template?: boolean`

    Deprecated: No longer used.

  - `template_id?: string | null`

    Deprecated: No longer used. The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Record<string, string> | null`

    Deprecated: Use `secrets` field instead. Environment variables for tool execution.

  - `tool_ids?: Array<string> | null`

    The ids of the tools used by the agent.

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The tool rules governing the agent.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

  - `tools?: Array<string> | null`

    The tools used by the agent.

### Returns

- `AgentState`

  Representation of an agent's state. This is the state of the agent at a given time, and is persisted in the DB backend. The state has all the information needed to recreate a persisted agent.

  - `id: string`

    The id of the agent. Assigned by the database.

  - `agent_type: AgentType`

    The type of agent.

    - `"memgpt_agent"`

    - `"memgpt_v2_agent"`

    - `"letta_v1_agent"`

    - `"react_agent"`

    - `"workflow_agent"`

    - `"split_thread_agent"`

    - `"sleeptime_agent"`

    - `"voice_convo_agent"`

    - `"voice_sleeptime_agent"`

  - `blocks: Array<Block>`

    The memory blocks used by the agent.

    - `value: string`

      Value of the block.

    - `id?: string`

      The human-friendly ID of the Block

    - `base_template_id?: string | null`

      The base template id of the block.

    - `created_by_id?: string | null`

      The id of the user that made this Block.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      Description of the block.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the block will be hidden.

    - `is_template?: boolean`

      Whether the block is a template (e.g. saved human/persona options).

    - `label?: string | null`

      Label of the block (e.g. 'human', 'persona') in the context window.

    - `last_updated_by_id?: string | null`

      The id of the user that last updated this Block.

    - `limit?: number`

      Character limit of the block.

    - `metadata?: Record<string, unknown> | null`

      Metadata of the block.

    - `preserve_on_migration?: boolean | null`

      Preserve the block on template migration.

    - `project_id?: string | null`

      The associated project id.

    - `read_only?: boolean`

      Whether the agent has read-only access to the block.

    - `tags?: Array<string> | null`

      The tags associated with the block.

    - `template_id?: string | null`

      The id of the template.

    - `template_name?: string | null`

      Name of the block if it is a template.

  - `llm_config: LlmConfig`

    Deprecated: Use `model` field instead. The LLM configuration used by the agent.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

        - `type?: "text"`

          The type of the response format.

          - `"text"`

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

        - `json_schema: Record<string, unknown>`

          The JSON schema of the response.

        - `type?: "json_schema"`

          The type of the response format.

          - `"json_schema"`

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

        - `type?: "json_object"`

          The type of the response format.

          - `"json_object"`

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `memory: Memory`

    Deprecated: Use `blocks` field instead. The in-context memory of the agent.

    - `blocks: Array<Block>`

      Memory blocks contained in the agent's in-context memory

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `agent_type?: AgentType | (string & {}) | null`

      Agent type controlling prompt rendering.

      - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

        Enum to represent the type of agent.

      - `(string & {})`

    - `file_blocks?: Array<FileBlock>`

      Special blocks representing the agent's in-context memory of an attached file

      - `file_id: string`

        Unique identifier of the file.

      - `is_open: boolean`

        True if the agent currently has the file open.

      - `source_id: string`

        Deprecated: Use `folder_id` field instead. Unique identifier of the source.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_accessed_at?: string | null`

        UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `git_enabled?: boolean`

      Whether this agent uses git-backed memory with structured labels.

    - `prompt_template?: string`

      Deprecated. Ignored for performance.

  - `name: string`

    The name of the agent.

  - `sources: Array<Source>`

    Deprecated: Use `folders` field instead. The sources used by the agent.

    - `id: string`

      The human-friendly ID of the Source

    - `embedding_config: EmbeddingConfig`

      The embedding configuration used by the source.

      - `embedding_dim: number`

        The dimension of the embedding.

      - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"bedrock"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"mistral"`

        - `"together"`

        - `"pinecone"`

      - `embedding_model: string`

        The model for the embedding.

      - `azure_deployment?: string | null`

        The Azure deployment for the model.

      - `azure_endpoint?: string | null`

        The Azure endpoint for the model.

      - `azure_version?: string | null`

        The Azure version for the model.

      - `batch_size?: number`

        The maximum batch size for processing embeddings.

      - `embedding_chunk_size?: number | null`

        The chunk size of the embedding.

      - `embedding_endpoint?: string | null`

        The endpoint for the model (`None` if local).

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

    - `name: string`

      The name of the source.

    - `created_at?: string | null`

      The timestamp when the source was created.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `description?: string | null`

      The description of the source.

    - `instructions?: string | null`

      Instructions for how to use the source.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata?: Record<string, unknown> | null`

      Metadata associated with the source.

    - `updated_at?: string | null`

      The timestamp when the source was last updated.

    - `vector_db_provider?: VectorDBProvider`

      The vector database provider used for this source's passages

      - `"native"`

      - `"tpuf"`

      - `"pinecone"`

  - `system: string`

    The system prompt used by the agent.

  - `tags: Array<string>`

    The tags associated with the agent.

  - `tools: Array<Tool>`

    The tools used by the agent.

    - `id: string`

      The human-friendly ID of the Tool

    - `args_json_schema?: Record<string, unknown> | null`

      The args JSON schema of the function.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `default_requires_approval?: boolean | null`

      Default value for whether or not executing this tool requires approval.

    - `description?: string | null`

      The description of the tool.

    - `enable_parallel_execution?: boolean | null`

      If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

    - `json_schema?: Record<string, unknown> | null`

      The JSON schema of the function.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata_?: Record<string, unknown> | null`

      A dictionary of additional metadata for the tool.

    - `name?: string | null`

      The name of the function.

    - `npm_requirements?: Array<NpmRequirement> | null`

      Optional list of npm packages required by this tool.

      - `name: string`

        Name of the npm package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `pip_requirements?: Array<PipRequirement> | null`

      Optional list of pip packages required by this tool.

      - `name: string`

        Name of the pip package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `project_id?: string | null`

      The project id of the tool.

    - `return_char_limit?: number`

      The maximum number of characters in the response.

    - `source_code?: string | null`

      The source code of the function.

    - `source_type?: string | null`

      The type of the source code.

    - `tags?: Array<string>`

      Metadata tags.

    - `tool_type?: ToolType`

      The type of the tool.

      - `"custom"`

      - `"letta_core"`

      - `"letta_memory_core"`

      - `"letta_multi_agent_core"`

      - `"letta_sleeptime_core"`

      - `"letta_voice_sleeptime_core"`

      - `"letta_builtin"`

      - `"letta_files_core"`

      - `"external_langchain"`

      - `"external_composio"`

      - `"external_mcp"`

  - `base_template_id?: string | null`

    The base template id of the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `created_at?: string | null`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the agent will be hidden.

  - `identities?: Array<Identity>`

    The identities associated with this agent.

    - `id: string`

      The human-friendly ID of the Identity

    - `agent_ids: Array<string>`

      The IDs of the agents associated with the identity.

    - `block_ids: Array<string>`

      The IDs of the blocks associated with the identity.

    - `identifier_key: string`

      External, user-generated identifier key of the identity.

    - `identity_type: "org" | "user" | "other"`

      The type of the identity.

      - `"org"`

      - `"user"`

      - `"other"`

    - `name: string`

      The name of the identity.

    - `project_id?: string | null`

      The project id of the identity, if applicable.

    - `properties?: Array<Property>`

      List of properties associated with the identity

      - `key: string`

        The key of the property

      - `type: "string" | "number" | "boolean" | "json"`

        The type of the property

        - `"string"`

        - `"number"`

        - `"boolean"`

        - `"json"`

      - `value: string | number | boolean | Record<string, unknown>`

        The value of the property

        - `string`

        - `number`

        - `boolean`

        - `Record<string, unknown>`

  - `identity_ids?: Array<string>`

    Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

  - `last_run_completion?: string | null`

    The timestamp when the agent last completed a run.

  - `last_run_duration_ms?: number | null`

    The duration in milliseconds of the agent's last run.

  - `last_stop_reason?: StopReasonType | null`

    The stop reason from the agent's last run.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `managed_group?: ManagedGroup | null`

    The multi-agent group that this agent manages

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `message_buffer_autoclear?: boolean`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `message_ids?: Array<string> | null`

    The ids of the messages in the agent's in-context memory.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle used by the agent (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings used by the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `multi_agent_group?: MultiAgentGroup | null`

    Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `pending_approval?: ApprovalRequestMessage | null`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project_id?: string | null`

    The id of the project the agent belongs to.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format used by the agent

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Array<AgentEnvironmentVariable>`

    The environment variables for tool execution specific to this agent.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `template_id?: string | null`

    The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

    Deprecated: use `secrets` field instead.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The list of tool rules.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const agentState = await client.agents.create();

console.log(agentState.id);
```

#### Response

```json
{
  "id": "id",
  "agent_type": "memgpt_agent",
  "blocks": [
    {
      "value": "value",
      "id": "block-123e4567-e89b-12d3-a456-426614174000",
      "base_template_id": "base_template_id",
      "created_by_id": "created_by_id",
      "deployment_id": "deployment_id",
      "description": "description",
      "entity_id": "entity_id",
      "hidden": true,
      "is_template": true,
      "label": "label",
      "last_updated_by_id": "last_updated_by_id",
      "limit": 0,
      "metadata": {
        "foo": "bar"
      },
      "preserve_on_migration": true,
      "project_id": "project_id",
      "read_only": true,
      "tags": [
        "string"
      ],
      "template_id": "template_id",
      "template_name": "template_name"
    }
  ],
  "llm_config": {
    "context_window": 0,
    "model": "model",
    "model_endpoint_type": "openai",
    "compatibility_type": "gguf",
    "display_name": "display_name",
    "effort": "low",
    "enable_reasoner": true,
    "frequency_penalty": 0,
    "handle": "handle",
    "max_reasoning_tokens": 0,
    "max_tokens": 0,
    "model_endpoint": "model_endpoint",
    "model_wrapper": "model_wrapper",
    "parallel_tool_calls": true,
    "provider_category": "base",
    "provider_name": "provider_name",
    "put_inner_thoughts_in_kwargs": true,
    "reasoning_effort": "none",
    "response_format": {
      "type": "text"
    },
    "return_logprobs": true,
    "return_token_ids": true,
    "strict": true,
    "temperature": 0,
    "tier": "tier",
    "tool_call_parser": "tool_call_parser",
    "top_logprobs": 0,
    "verbosity": "low"
  },
  "memory": {
    "blocks": [
      {
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "agent_type": "memgpt_agent",
    "file_blocks": [
      {
        "file_id": "file_id",
        "is_open": true,
        "source_id": "source_id",
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_accessed_at": "2019-12-27T18:11:19.117Z",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "git_enabled": true,
    "prompt_template": "prompt_template"
  },
  "name": "name",
  "sources": [
    {
      "id": "source-123e4567-e89b-12d3-a456-426614174000",
      "embedding_config": {
        "embedding_dim": 0,
        "embedding_endpoint_type": "openai",
        "embedding_model": "embedding_model",
        "azure_deployment": "azure_deployment",
        "azure_endpoint": "azure_endpoint",
        "azure_version": "azure_version",
        "batch_size": 0,
        "embedding_chunk_size": 0,
        "embedding_endpoint": "embedding_endpoint",
        "handle": "handle"
      },
      "name": "name",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "instructions": "instructions",
      "last_updated_by_id": "last_updated_by_id",
      "metadata": {
        "foo": "bar"
      },
      "updated_at": "2019-12-27T18:11:19.117Z",
      "vector_db_provider": "native"
    }
  ],
  "system": "system",
  "tags": [
    "string"
  ],
  "tools": [
    {
      "id": "tool-123e4567-e89b-12d3-a456-426614174000",
      "args_json_schema": {
        "foo": "bar"
      },
      "created_by_id": "created_by_id",
      "default_requires_approval": true,
      "description": "description",
      "enable_parallel_execution": true,
      "json_schema": {
        "foo": "bar"
      },
      "last_updated_by_id": "last_updated_by_id",
      "metadata_": {
        "foo": "bar"
      },
      "name": "name",
      "npm_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "pip_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "project_id": "project_id",
      "return_char_limit": 1,
      "source_code": "source_code",
      "source_type": "source_type",
      "tags": [
        "string"
      ],
      "tool_type": "custom"
    }
  ],
  "base_template_id": "base_template_id",
  "compaction_settings": {
    "clip_chars": 0,
    "mode": "all",
    "model": "model",
    "model_settings": {
      "max_output_tokens": 0,
      "parallel_tool_calls": true,
      "provider_type": "openai",
      "reasoning": {
        "reasoning_effort": "none"
      },
      "response_format": {
        "type": "text"
      },
      "strict": true,
      "temperature": 0
    },
    "prompt": "prompt",
    "prompt_acknowledgement": true,
    "sliding_window_percentage": 0
  },
  "created_at": "2019-12-27T18:11:19.117Z",
  "created_by_id": "created_by_id",
  "deployment_id": "deployment_id",
  "description": "description",
  "embedding": "embedding",
  "embedding_config": {
    "embedding_dim": 0,
    "embedding_endpoint_type": "openai",
    "embedding_model": "embedding_model",
    "azure_deployment": "azure_deployment",
    "azure_endpoint": "azure_endpoint",
    "azure_version": "azure_version",
    "batch_size": 0,
    "embedding_chunk_size": 0,
    "embedding_endpoint": "embedding_endpoint",
    "handle": "handle"
  },
  "enable_sleeptime": true,
  "entity_id": "entity_id",
  "hidden": true,
  "identities": [
    {
      "id": "identity-123e4567-e89b-12d3-a456-426614174000",
      "agent_ids": [
        "string"
      ],
      "block_ids": [
        "string"
      ],
      "identifier_key": "identifier_key",
      "identity_type": "org",
      "name": "name",
      "project_id": "project_id",
      "properties": [
        {
          "key": "key",
          "type": "string",
          "value": "string"
        }
      ]
    }
  ],
  "identity_ids": [
    "string"
  ],
  "last_run_completion": "2019-12-27T18:11:19.117Z",
  "last_run_duration_ms": 0,
  "last_stop_reason": "end_turn",
  "last_updated_by_id": "last_updated_by_id",
  "managed_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "max_files_open": 0,
  "message_buffer_autoclear": true,
  "message_ids": [
    "string"
  ],
  "metadata": {
    "foo": "bar"
  },
  "model": "model",
  "model_settings": {
    "max_output_tokens": 0,
    "parallel_tool_calls": true,
    "provider_type": "openai",
    "reasoning": {
      "reasoning_effort": "none"
    },
    "response_format": {
      "type": "text"
    },
    "strict": true,
    "temperature": 0
  },
  "multi_agent_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "pending_approval": {
    "id": "id",
    "date": "2019-12-27T18:11:19.117Z",
    "tool_call": {
      "arguments": "arguments",
      "name": "name",
      "tool_call_id": "tool_call_id"
    },
    "is_err": true,
    "message_type": "approval_request_message",
    "name": "name",
    "otid": "otid",
    "run_id": "run_id",
    "sender_id": "sender_id",
    "seq_id": 0,
    "step_id": "step_id",
    "tool_calls": [
      {
        "arguments": "arguments",
        "name": "name",
        "tool_call_id": "tool_call_id"
      }
    ]
  },
  "per_file_view_window_char_limit": 0,
  "project_id": "project_id",
  "response_format": {
    "type": "text"
  },
  "secrets": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "template_id": "template_id",
  "timezone": "timezone",
  "tool_exec_environment_variables": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "tool_rules": [
    {
      "children": [
        "string"
      ],
      "tool_name": "tool_name",
      "child_arg_nodes": [
        {
          "name": "name",
          "args": {
            "foo": "bar"
          }
        }
      ],
      "prompt_template": "prompt_template",
      "type": "constrain_child_tools"
    }
  ],
  "updated_at": "2019-12-27T18:11:19.117Z"
}
```

## Update Agent

`client.agents.update(stringagentID, AgentUpdateParamsbody, RequestOptionsoptions?): AgentState`

**patch** `/v1/agents/{agent_id}`

Update an existing agent.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `body: AgentUpdateParams`

  - `base_template_id?: string | null`

    The base template id of the agent.

  - `block_ids?: Array<string> | null`

    The ids of the blocks used by the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

            - `type?: "text"`

              The type of the response format.

              - `"text"`

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

            - `json_schema: Record<string, unknown>`

              The JSON schema of the response.

            - `type?: "json_schema"`

              The type of the response format.

              - `"json_schema"`

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

            - `type?: "json_object"`

              The type of the response format.

              - `"json_object"`

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `context_window_limit?: number | null`

    The context window limit used by the agent.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

    - `embedding_dim: number`

      The dimension of the embedding.

    - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"bedrock"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"mistral"`

      - `"together"`

      - `"pinecone"`

    - `embedding_model: string`

      The model for the embedding.

    - `azure_deployment?: string | null`

      The Azure deployment for the model.

    - `azure_endpoint?: string | null`

      The Azure endpoint for the model.

    - `azure_version?: string | null`

      The Azure version for the model.

    - `batch_size?: number`

      The maximum batch size for processing embeddings.

    - `embedding_chunk_size?: number | null`

      The chunk size of the embedding.

    - `embedding_endpoint?: string | null`

      The endpoint for the model (`None` if local).

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `folder_ids?: Array<string> | null`

    The ids of the folders used by the agent.

  - `hidden?: boolean | null`

    If set to True, the agent will be hidden.

  - `identity_ids?: Array<string> | null`

    The ids of the identities associated with this agent.

  - `last_run_completion?: string | null`

    The timestamp when the agent last completed a run.

  - `last_run_duration_ms?: number | null`

    The duration in milliseconds of the agent's last run.

  - `last_stop_reason?: StopReasonType | null`

    The stop reason from the agent's last run.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `llm_config?: LlmConfig | null`

    Configuration for Language Model (LLM) connection and generation parameters.

    .. deprecated::
    LLMConfig is deprecated and should not be used as an input or return type in API calls.
    Use the schemas in letta.schemas.model (ModelSettings, OpenAIModelSettings, etc.) instead.
    For conversion, use the _to_model() method or Model._from_llm_config() method.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `max_tokens?: number | null`

    Deprecated: Use `model` field to configure max output tokens instead. The maximum number of tokens to generate, including reasoning step.

  - `message_buffer_autoclear?: boolean | null`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `message_ids?: Array<string> | null`

    The ids of the messages in the agent's in-context memory.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle used by the agent (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings for the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `name?: string | null`

    The name of the agent.

  - `parallel_tool_calls?: boolean | null`

    Deprecated: Use `model_settings` to configure parallel tool calls instead. If set to True, enables parallel tool calling.

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project_id?: string | null`

    The id of the project the agent belongs to.

  - `reasoning?: boolean | null`

    Deprecated: Use `model` field to configure reasoning instead. Whether to enable reasoning for this agent.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    Deprecated: Use `model_settings` field to configure response format instead. The response format for the agent.

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Record<string, string> | null`

    The environment variables for tool execution specific to this agent.

  - `source_ids?: Array<string> | null`

    Deprecated: Use `folder_ids` field instead. The ids of the sources used by the agent.

  - `system?: string | null`

    The system prompt used by the agent.

  - `tags?: Array<string> | null`

    The tags associated with the agent.

  - `template_id?: string | null`

    The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Record<string, string> | null`

    Deprecated: use `secrets` field instead

  - `tool_ids?: Array<string> | null`

    The ids of the tools used by the agent.

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The tool rules governing the agent.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

### Returns

- `AgentState`

  Representation of an agent's state. This is the state of the agent at a given time, and is persisted in the DB backend. The state has all the information needed to recreate a persisted agent.

  - `id: string`

    The id of the agent. Assigned by the database.

  - `agent_type: AgentType`

    The type of agent.

    - `"memgpt_agent"`

    - `"memgpt_v2_agent"`

    - `"letta_v1_agent"`

    - `"react_agent"`

    - `"workflow_agent"`

    - `"split_thread_agent"`

    - `"sleeptime_agent"`

    - `"voice_convo_agent"`

    - `"voice_sleeptime_agent"`

  - `blocks: Array<Block>`

    The memory blocks used by the agent.

    - `value: string`

      Value of the block.

    - `id?: string`

      The human-friendly ID of the Block

    - `base_template_id?: string | null`

      The base template id of the block.

    - `created_by_id?: string | null`

      The id of the user that made this Block.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      Description of the block.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the block will be hidden.

    - `is_template?: boolean`

      Whether the block is a template (e.g. saved human/persona options).

    - `label?: string | null`

      Label of the block (e.g. 'human', 'persona') in the context window.

    - `last_updated_by_id?: string | null`

      The id of the user that last updated this Block.

    - `limit?: number`

      Character limit of the block.

    - `metadata?: Record<string, unknown> | null`

      Metadata of the block.

    - `preserve_on_migration?: boolean | null`

      Preserve the block on template migration.

    - `project_id?: string | null`

      The associated project id.

    - `read_only?: boolean`

      Whether the agent has read-only access to the block.

    - `tags?: Array<string> | null`

      The tags associated with the block.

    - `template_id?: string | null`

      The id of the template.

    - `template_name?: string | null`

      Name of the block if it is a template.

  - `llm_config: LlmConfig`

    Deprecated: Use `model` field instead. The LLM configuration used by the agent.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

        - `type?: "text"`

          The type of the response format.

          - `"text"`

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

        - `json_schema: Record<string, unknown>`

          The JSON schema of the response.

        - `type?: "json_schema"`

          The type of the response format.

          - `"json_schema"`

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

        - `type?: "json_object"`

          The type of the response format.

          - `"json_object"`

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `memory: Memory`

    Deprecated: Use `blocks` field instead. The in-context memory of the agent.

    - `blocks: Array<Block>`

      Memory blocks contained in the agent's in-context memory

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `agent_type?: AgentType | (string & {}) | null`

      Agent type controlling prompt rendering.

      - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

        Enum to represent the type of agent.

      - `(string & {})`

    - `file_blocks?: Array<FileBlock>`

      Special blocks representing the agent's in-context memory of an attached file

      - `file_id: string`

        Unique identifier of the file.

      - `is_open: boolean`

        True if the agent currently has the file open.

      - `source_id: string`

        Deprecated: Use `folder_id` field instead. Unique identifier of the source.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_accessed_at?: string | null`

        UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `git_enabled?: boolean`

      Whether this agent uses git-backed memory with structured labels.

    - `prompt_template?: string`

      Deprecated. Ignored for performance.

  - `name: string`

    The name of the agent.

  - `sources: Array<Source>`

    Deprecated: Use `folders` field instead. The sources used by the agent.

    - `id: string`

      The human-friendly ID of the Source

    - `embedding_config: EmbeddingConfig`

      The embedding configuration used by the source.

      - `embedding_dim: number`

        The dimension of the embedding.

      - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"bedrock"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"mistral"`

        - `"together"`

        - `"pinecone"`

      - `embedding_model: string`

        The model for the embedding.

      - `azure_deployment?: string | null`

        The Azure deployment for the model.

      - `azure_endpoint?: string | null`

        The Azure endpoint for the model.

      - `azure_version?: string | null`

        The Azure version for the model.

      - `batch_size?: number`

        The maximum batch size for processing embeddings.

      - `embedding_chunk_size?: number | null`

        The chunk size of the embedding.

      - `embedding_endpoint?: string | null`

        The endpoint for the model (`None` if local).

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

    - `name: string`

      The name of the source.

    - `created_at?: string | null`

      The timestamp when the source was created.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `description?: string | null`

      The description of the source.

    - `instructions?: string | null`

      Instructions for how to use the source.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata?: Record<string, unknown> | null`

      Metadata associated with the source.

    - `updated_at?: string | null`

      The timestamp when the source was last updated.

    - `vector_db_provider?: VectorDBProvider`

      The vector database provider used for this source's passages

      - `"native"`

      - `"tpuf"`

      - `"pinecone"`

  - `system: string`

    The system prompt used by the agent.

  - `tags: Array<string>`

    The tags associated with the agent.

  - `tools: Array<Tool>`

    The tools used by the agent.

    - `id: string`

      The human-friendly ID of the Tool

    - `args_json_schema?: Record<string, unknown> | null`

      The args JSON schema of the function.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `default_requires_approval?: boolean | null`

      Default value for whether or not executing this tool requires approval.

    - `description?: string | null`

      The description of the tool.

    - `enable_parallel_execution?: boolean | null`

      If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

    - `json_schema?: Record<string, unknown> | null`

      The JSON schema of the function.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata_?: Record<string, unknown> | null`

      A dictionary of additional metadata for the tool.

    - `name?: string | null`

      The name of the function.

    - `npm_requirements?: Array<NpmRequirement> | null`

      Optional list of npm packages required by this tool.

      - `name: string`

        Name of the npm package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `pip_requirements?: Array<PipRequirement> | null`

      Optional list of pip packages required by this tool.

      - `name: string`

        Name of the pip package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `project_id?: string | null`

      The project id of the tool.

    - `return_char_limit?: number`

      The maximum number of characters in the response.

    - `source_code?: string | null`

      The source code of the function.

    - `source_type?: string | null`

      The type of the source code.

    - `tags?: Array<string>`

      Metadata tags.

    - `tool_type?: ToolType`

      The type of the tool.

      - `"custom"`

      - `"letta_core"`

      - `"letta_memory_core"`

      - `"letta_multi_agent_core"`

      - `"letta_sleeptime_core"`

      - `"letta_voice_sleeptime_core"`

      - `"letta_builtin"`

      - `"letta_files_core"`

      - `"external_langchain"`

      - `"external_composio"`

      - `"external_mcp"`

  - `base_template_id?: string | null`

    The base template id of the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `created_at?: string | null`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the agent will be hidden.

  - `identities?: Array<Identity>`

    The identities associated with this agent.

    - `id: string`

      The human-friendly ID of the Identity

    - `agent_ids: Array<string>`

      The IDs of the agents associated with the identity.

    - `block_ids: Array<string>`

      The IDs of the blocks associated with the identity.

    - `identifier_key: string`

      External, user-generated identifier key of the identity.

    - `identity_type: "org" | "user" | "other"`

      The type of the identity.

      - `"org"`

      - `"user"`

      - `"other"`

    - `name: string`

      The name of the identity.

    - `project_id?: string | null`

      The project id of the identity, if applicable.

    - `properties?: Array<Property>`

      List of properties associated with the identity

      - `key: string`

        The key of the property

      - `type: "string" | "number" | "boolean" | "json"`

        The type of the property

        - `"string"`

        - `"number"`

        - `"boolean"`

        - `"json"`

      - `value: string | number | boolean | Record<string, unknown>`

        The value of the property

        - `string`

        - `number`

        - `boolean`

        - `Record<string, unknown>`

  - `identity_ids?: Array<string>`

    Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

  - `last_run_completion?: string | null`

    The timestamp when the agent last completed a run.

  - `last_run_duration_ms?: number | null`

    The duration in milliseconds of the agent's last run.

  - `last_stop_reason?: StopReasonType | null`

    The stop reason from the agent's last run.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `managed_group?: ManagedGroup | null`

    The multi-agent group that this agent manages

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `message_buffer_autoclear?: boolean`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `message_ids?: Array<string> | null`

    The ids of the messages in the agent's in-context memory.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle used by the agent (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings used by the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `multi_agent_group?: MultiAgentGroup | null`

    Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `pending_approval?: ApprovalRequestMessage | null`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project_id?: string | null`

    The id of the project the agent belongs to.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format used by the agent

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Array<AgentEnvironmentVariable>`

    The environment variables for tool execution specific to this agent.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `template_id?: string | null`

    The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

    Deprecated: use `secrets` field instead.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The list of tool rules.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const agentState = await client.agents.update('agent-123e4567-e89b-42d3-8456-426614174000');

console.log(agentState.id);
```

#### Response

```json
{
  "id": "id",
  "agent_type": "memgpt_agent",
  "blocks": [
    {
      "value": "value",
      "id": "block-123e4567-e89b-12d3-a456-426614174000",
      "base_template_id": "base_template_id",
      "created_by_id": "created_by_id",
      "deployment_id": "deployment_id",
      "description": "description",
      "entity_id": "entity_id",
      "hidden": true,
      "is_template": true,
      "label": "label",
      "last_updated_by_id": "last_updated_by_id",
      "limit": 0,
      "metadata": {
        "foo": "bar"
      },
      "preserve_on_migration": true,
      "project_id": "project_id",
      "read_only": true,
      "tags": [
        "string"
      ],
      "template_id": "template_id",
      "template_name": "template_name"
    }
  ],
  "llm_config": {
    "context_window": 0,
    "model": "model",
    "model_endpoint_type": "openai",
    "compatibility_type": "gguf",
    "display_name": "display_name",
    "effort": "low",
    "enable_reasoner": true,
    "frequency_penalty": 0,
    "handle": "handle",
    "max_reasoning_tokens": 0,
    "max_tokens": 0,
    "model_endpoint": "model_endpoint",
    "model_wrapper": "model_wrapper",
    "parallel_tool_calls": true,
    "provider_category": "base",
    "provider_name": "provider_name",
    "put_inner_thoughts_in_kwargs": true,
    "reasoning_effort": "none",
    "response_format": {
      "type": "text"
    },
    "return_logprobs": true,
    "return_token_ids": true,
    "strict": true,
    "temperature": 0,
    "tier": "tier",
    "tool_call_parser": "tool_call_parser",
    "top_logprobs": 0,
    "verbosity": "low"
  },
  "memory": {
    "blocks": [
      {
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "agent_type": "memgpt_agent",
    "file_blocks": [
      {
        "file_id": "file_id",
        "is_open": true,
        "source_id": "source_id",
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_accessed_at": "2019-12-27T18:11:19.117Z",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "git_enabled": true,
    "prompt_template": "prompt_template"
  },
  "name": "name",
  "sources": [
    {
      "id": "source-123e4567-e89b-12d3-a456-426614174000",
      "embedding_config": {
        "embedding_dim": 0,
        "embedding_endpoint_type": "openai",
        "embedding_model": "embedding_model",
        "azure_deployment": "azure_deployment",
        "azure_endpoint": "azure_endpoint",
        "azure_version": "azure_version",
        "batch_size": 0,
        "embedding_chunk_size": 0,
        "embedding_endpoint": "embedding_endpoint",
        "handle": "handle"
      },
      "name": "name",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "instructions": "instructions",
      "last_updated_by_id": "last_updated_by_id",
      "metadata": {
        "foo": "bar"
      },
      "updated_at": "2019-12-27T18:11:19.117Z",
      "vector_db_provider": "native"
    }
  ],
  "system": "system",
  "tags": [
    "string"
  ],
  "tools": [
    {
      "id": "tool-123e4567-e89b-12d3-a456-426614174000",
      "args_json_schema": {
        "foo": "bar"
      },
      "created_by_id": "created_by_id",
      "default_requires_approval": true,
      "description": "description",
      "enable_parallel_execution": true,
      "json_schema": {
        "foo": "bar"
      },
      "last_updated_by_id": "last_updated_by_id",
      "metadata_": {
        "foo": "bar"
      },
      "name": "name",
      "npm_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "pip_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "project_id": "project_id",
      "return_char_limit": 1,
      "source_code": "source_code",
      "source_type": "source_type",
      "tags": [
        "string"
      ],
      "tool_type": "custom"
    }
  ],
  "base_template_id": "base_template_id",
  "compaction_settings": {
    "clip_chars": 0,
    "mode": "all",
    "model": "model",
    "model_settings": {
      "max_output_tokens": 0,
      "parallel_tool_calls": true,
      "provider_type": "openai",
      "reasoning": {
        "reasoning_effort": "none"
      },
      "response_format": {
        "type": "text"
      },
      "strict": true,
      "temperature": 0
    },
    "prompt": "prompt",
    "prompt_acknowledgement": true,
    "sliding_window_percentage": 0
  },
  "created_at": "2019-12-27T18:11:19.117Z",
  "created_by_id": "created_by_id",
  "deployment_id": "deployment_id",
  "description": "description",
  "embedding": "embedding",
  "embedding_config": {
    "embedding_dim": 0,
    "embedding_endpoint_type": "openai",
    "embedding_model": "embedding_model",
    "azure_deployment": "azure_deployment",
    "azure_endpoint": "azure_endpoint",
    "azure_version": "azure_version",
    "batch_size": 0,
    "embedding_chunk_size": 0,
    "embedding_endpoint": "embedding_endpoint",
    "handle": "handle"
  },
  "enable_sleeptime": true,
  "entity_id": "entity_id",
  "hidden": true,
  "identities": [
    {
      "id": "identity-123e4567-e89b-12d3-a456-426614174000",
      "agent_ids": [
        "string"
      ],
      "block_ids": [
        "string"
      ],
      "identifier_key": "identifier_key",
      "identity_type": "org",
      "name": "name",
      "project_id": "project_id",
      "properties": [
        {
          "key": "key",
          "type": "string",
          "value": "string"
        }
      ]
    }
  ],
  "identity_ids": [
    "string"
  ],
  "last_run_completion": "2019-12-27T18:11:19.117Z",
  "last_run_duration_ms": 0,
  "last_stop_reason": "end_turn",
  "last_updated_by_id": "last_updated_by_id",
  "managed_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "max_files_open": 0,
  "message_buffer_autoclear": true,
  "message_ids": [
    "string"
  ],
  "metadata": {
    "foo": "bar"
  },
  "model": "model",
  "model_settings": {
    "max_output_tokens": 0,
    "parallel_tool_calls": true,
    "provider_type": "openai",
    "reasoning": {
      "reasoning_effort": "none"
    },
    "response_format": {
      "type": "text"
    },
    "strict": true,
    "temperature": 0
  },
  "multi_agent_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "pending_approval": {
    "id": "id",
    "date": "2019-12-27T18:11:19.117Z",
    "tool_call": {
      "arguments": "arguments",
      "name": "name",
      "tool_call_id": "tool_call_id"
    },
    "is_err": true,
    "message_type": "approval_request_message",
    "name": "name",
    "otid": "otid",
    "run_id": "run_id",
    "sender_id": "sender_id",
    "seq_id": 0,
    "step_id": "step_id",
    "tool_calls": [
      {
        "arguments": "arguments",
        "name": "name",
        "tool_call_id": "tool_call_id"
      }
    ]
  },
  "per_file_view_window_char_limit": 0,
  "project_id": "project_id",
  "response_format": {
    "type": "text"
  },
  "secrets": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "template_id": "template_id",
  "timezone": "timezone",
  "tool_exec_environment_variables": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "tool_rules": [
    {
      "children": [
        "string"
      ],
      "tool_name": "tool_name",
      "child_arg_nodes": [
        {
          "name": "name",
          "args": {
            "foo": "bar"
          }
        }
      ],
      "prompt_template": "prompt_template",
      "type": "constrain_child_tools"
    }
  ],
  "updated_at": "2019-12-27T18:11:19.117Z"
}
```

## Retrieve Agent

`client.agents.retrieve(stringagentID, AgentRetrieveParamsquery?, RequestOptionsoptions?): AgentState`

**get** `/v1/agents/{agent_id}`

Get the state of the agent.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `query: AgentRetrieveParams`

  - `include?: Array<"agent.blocks" | "agent.identities" | "agent.managed_group" | 5 more> | null`

    Specify which relational fields to include in the response. No relationships are included by default.

    - `"agent.blocks"`

    - `"agent.identities"`

    - `"agent.managed_group"`

    - `"agent.pending_approval"`

    - `"agent.secrets"`

    - `"agent.sources"`

    - `"agent.tags"`

    - `"agent.tools"`

  - `include_relationships?: Array<string> | null`

    Specify which relational fields (e.g., 'tools', 'sources', 'memory') to include in the response. If not provided, all relationships are loaded by default. Using this can optimize performance by reducing unnecessary joins.This is a legacy parameter, and no longer supported after 1.0.0 SDK versions.

### Returns

- `AgentState`

  Representation of an agent's state. This is the state of the agent at a given time, and is persisted in the DB backend. The state has all the information needed to recreate a persisted agent.

  - `id: string`

    The id of the agent. Assigned by the database.

  - `agent_type: AgentType`

    The type of agent.

    - `"memgpt_agent"`

    - `"memgpt_v2_agent"`

    - `"letta_v1_agent"`

    - `"react_agent"`

    - `"workflow_agent"`

    - `"split_thread_agent"`

    - `"sleeptime_agent"`

    - `"voice_convo_agent"`

    - `"voice_sleeptime_agent"`

  - `blocks: Array<Block>`

    The memory blocks used by the agent.

    - `value: string`

      Value of the block.

    - `id?: string`

      The human-friendly ID of the Block

    - `base_template_id?: string | null`

      The base template id of the block.

    - `created_by_id?: string | null`

      The id of the user that made this Block.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      Description of the block.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the block will be hidden.

    - `is_template?: boolean`

      Whether the block is a template (e.g. saved human/persona options).

    - `label?: string | null`

      Label of the block (e.g. 'human', 'persona') in the context window.

    - `last_updated_by_id?: string | null`

      The id of the user that last updated this Block.

    - `limit?: number`

      Character limit of the block.

    - `metadata?: Record<string, unknown> | null`

      Metadata of the block.

    - `preserve_on_migration?: boolean | null`

      Preserve the block on template migration.

    - `project_id?: string | null`

      The associated project id.

    - `read_only?: boolean`

      Whether the agent has read-only access to the block.

    - `tags?: Array<string> | null`

      The tags associated with the block.

    - `template_id?: string | null`

      The id of the template.

    - `template_name?: string | null`

      Name of the block if it is a template.

  - `llm_config: LlmConfig`

    Deprecated: Use `model` field instead. The LLM configuration used by the agent.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

        - `type?: "text"`

          The type of the response format.

          - `"text"`

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

        - `json_schema: Record<string, unknown>`

          The JSON schema of the response.

        - `type?: "json_schema"`

          The type of the response format.

          - `"json_schema"`

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

        - `type?: "json_object"`

          The type of the response format.

          - `"json_object"`

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `memory: Memory`

    Deprecated: Use `blocks` field instead. The in-context memory of the agent.

    - `blocks: Array<Block>`

      Memory blocks contained in the agent's in-context memory

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `agent_type?: AgentType | (string & {}) | null`

      Agent type controlling prompt rendering.

      - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

        Enum to represent the type of agent.

      - `(string & {})`

    - `file_blocks?: Array<FileBlock>`

      Special blocks representing the agent's in-context memory of an attached file

      - `file_id: string`

        Unique identifier of the file.

      - `is_open: boolean`

        True if the agent currently has the file open.

      - `source_id: string`

        Deprecated: Use `folder_id` field instead. Unique identifier of the source.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_accessed_at?: string | null`

        UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `git_enabled?: boolean`

      Whether this agent uses git-backed memory with structured labels.

    - `prompt_template?: string`

      Deprecated. Ignored for performance.

  - `name: string`

    The name of the agent.

  - `sources: Array<Source>`

    Deprecated: Use `folders` field instead. The sources used by the agent.

    - `id: string`

      The human-friendly ID of the Source

    - `embedding_config: EmbeddingConfig`

      The embedding configuration used by the source.

      - `embedding_dim: number`

        The dimension of the embedding.

      - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"bedrock"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"mistral"`

        - `"together"`

        - `"pinecone"`

      - `embedding_model: string`

        The model for the embedding.

      - `azure_deployment?: string | null`

        The Azure deployment for the model.

      - `azure_endpoint?: string | null`

        The Azure endpoint for the model.

      - `azure_version?: string | null`

        The Azure version for the model.

      - `batch_size?: number`

        The maximum batch size for processing embeddings.

      - `embedding_chunk_size?: number | null`

        The chunk size of the embedding.

      - `embedding_endpoint?: string | null`

        The endpoint for the model (`None` if local).

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

    - `name: string`

      The name of the source.

    - `created_at?: string | null`

      The timestamp when the source was created.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `description?: string | null`

      The description of the source.

    - `instructions?: string | null`

      Instructions for how to use the source.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata?: Record<string, unknown> | null`

      Metadata associated with the source.

    - `updated_at?: string | null`

      The timestamp when the source was last updated.

    - `vector_db_provider?: VectorDBProvider`

      The vector database provider used for this source's passages

      - `"native"`

      - `"tpuf"`

      - `"pinecone"`

  - `system: string`

    The system prompt used by the agent.

  - `tags: Array<string>`

    The tags associated with the agent.

  - `tools: Array<Tool>`

    The tools used by the agent.

    - `id: string`

      The human-friendly ID of the Tool

    - `args_json_schema?: Record<string, unknown> | null`

      The args JSON schema of the function.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `default_requires_approval?: boolean | null`

      Default value for whether or not executing this tool requires approval.

    - `description?: string | null`

      The description of the tool.

    - `enable_parallel_execution?: boolean | null`

      If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

    - `json_schema?: Record<string, unknown> | null`

      The JSON schema of the function.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata_?: Record<string, unknown> | null`

      A dictionary of additional metadata for the tool.

    - `name?: string | null`

      The name of the function.

    - `npm_requirements?: Array<NpmRequirement> | null`

      Optional list of npm packages required by this tool.

      - `name: string`

        Name of the npm package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `pip_requirements?: Array<PipRequirement> | null`

      Optional list of pip packages required by this tool.

      - `name: string`

        Name of the pip package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `project_id?: string | null`

      The project id of the tool.

    - `return_char_limit?: number`

      The maximum number of characters in the response.

    - `source_code?: string | null`

      The source code of the function.

    - `source_type?: string | null`

      The type of the source code.

    - `tags?: Array<string>`

      Metadata tags.

    - `tool_type?: ToolType`

      The type of the tool.

      - `"custom"`

      - `"letta_core"`

      - `"letta_memory_core"`

      - `"letta_multi_agent_core"`

      - `"letta_sleeptime_core"`

      - `"letta_voice_sleeptime_core"`

      - `"letta_builtin"`

      - `"letta_files_core"`

      - `"external_langchain"`

      - `"external_composio"`

      - `"external_mcp"`

  - `base_template_id?: string | null`

    The base template id of the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `created_at?: string | null`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the agent will be hidden.

  - `identities?: Array<Identity>`

    The identities associated with this agent.

    - `id: string`

      The human-friendly ID of the Identity

    - `agent_ids: Array<string>`

      The IDs of the agents associated with the identity.

    - `block_ids: Array<string>`

      The IDs of the blocks associated with the identity.

    - `identifier_key: string`

      External, user-generated identifier key of the identity.

    - `identity_type: "org" | "user" | "other"`

      The type of the identity.

      - `"org"`

      - `"user"`

      - `"other"`

    - `name: string`

      The name of the identity.

    - `project_id?: string | null`

      The project id of the identity, if applicable.

    - `properties?: Array<Property>`

      List of properties associated with the identity

      - `key: string`

        The key of the property

      - `type: "string" | "number" | "boolean" | "json"`

        The type of the property

        - `"string"`

        - `"number"`

        - `"boolean"`

        - `"json"`

      - `value: string | number | boolean | Record<string, unknown>`

        The value of the property

        - `string`

        - `number`

        - `boolean`

        - `Record<string, unknown>`

  - `identity_ids?: Array<string>`

    Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

  - `last_run_completion?: string | null`

    The timestamp when the agent last completed a run.

  - `last_run_duration_ms?: number | null`

    The duration in milliseconds of the agent's last run.

  - `last_stop_reason?: StopReasonType | null`

    The stop reason from the agent's last run.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `managed_group?: ManagedGroup | null`

    The multi-agent group that this agent manages

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `message_buffer_autoclear?: boolean`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `message_ids?: Array<string> | null`

    The ids of the messages in the agent's in-context memory.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle used by the agent (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings used by the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `multi_agent_group?: MultiAgentGroup | null`

    Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `pending_approval?: ApprovalRequestMessage | null`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project_id?: string | null`

    The id of the project the agent belongs to.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format used by the agent

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Array<AgentEnvironmentVariable>`

    The environment variables for tool execution specific to this agent.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `template_id?: string | null`

    The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

    Deprecated: use `secrets` field instead.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The list of tool rules.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const agentState = await client.agents.retrieve('agent-123e4567-e89b-42d3-8456-426614174000');

console.log(agentState.id);
```

#### Response

```json
{
  "id": "id",
  "agent_type": "memgpt_agent",
  "blocks": [
    {
      "value": "value",
      "id": "block-123e4567-e89b-12d3-a456-426614174000",
      "base_template_id": "base_template_id",
      "created_by_id": "created_by_id",
      "deployment_id": "deployment_id",
      "description": "description",
      "entity_id": "entity_id",
      "hidden": true,
      "is_template": true,
      "label": "label",
      "last_updated_by_id": "last_updated_by_id",
      "limit": 0,
      "metadata": {
        "foo": "bar"
      },
      "preserve_on_migration": true,
      "project_id": "project_id",
      "read_only": true,
      "tags": [
        "string"
      ],
      "template_id": "template_id",
      "template_name": "template_name"
    }
  ],
  "llm_config": {
    "context_window": 0,
    "model": "model",
    "model_endpoint_type": "openai",
    "compatibility_type": "gguf",
    "display_name": "display_name",
    "effort": "low",
    "enable_reasoner": true,
    "frequency_penalty": 0,
    "handle": "handle",
    "max_reasoning_tokens": 0,
    "max_tokens": 0,
    "model_endpoint": "model_endpoint",
    "model_wrapper": "model_wrapper",
    "parallel_tool_calls": true,
    "provider_category": "base",
    "provider_name": "provider_name",
    "put_inner_thoughts_in_kwargs": true,
    "reasoning_effort": "none",
    "response_format": {
      "type": "text"
    },
    "return_logprobs": true,
    "return_token_ids": true,
    "strict": true,
    "temperature": 0,
    "tier": "tier",
    "tool_call_parser": "tool_call_parser",
    "top_logprobs": 0,
    "verbosity": "low"
  },
  "memory": {
    "blocks": [
      {
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "agent_type": "memgpt_agent",
    "file_blocks": [
      {
        "file_id": "file_id",
        "is_open": true,
        "source_id": "source_id",
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_accessed_at": "2019-12-27T18:11:19.117Z",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "git_enabled": true,
    "prompt_template": "prompt_template"
  },
  "name": "name",
  "sources": [
    {
      "id": "source-123e4567-e89b-12d3-a456-426614174000",
      "embedding_config": {
        "embedding_dim": 0,
        "embedding_endpoint_type": "openai",
        "embedding_model": "embedding_model",
        "azure_deployment": "azure_deployment",
        "azure_endpoint": "azure_endpoint",
        "azure_version": "azure_version",
        "batch_size": 0,
        "embedding_chunk_size": 0,
        "embedding_endpoint": "embedding_endpoint",
        "handle": "handle"
      },
      "name": "name",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "instructions": "instructions",
      "last_updated_by_id": "last_updated_by_id",
      "metadata": {
        "foo": "bar"
      },
      "updated_at": "2019-12-27T18:11:19.117Z",
      "vector_db_provider": "native"
    }
  ],
  "system": "system",
  "tags": [
    "string"
  ],
  "tools": [
    {
      "id": "tool-123e4567-e89b-12d3-a456-426614174000",
      "args_json_schema": {
        "foo": "bar"
      },
      "created_by_id": "created_by_id",
      "default_requires_approval": true,
      "description": "description",
      "enable_parallel_execution": true,
      "json_schema": {
        "foo": "bar"
      },
      "last_updated_by_id": "last_updated_by_id",
      "metadata_": {
        "foo": "bar"
      },
      "name": "name",
      "npm_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "pip_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "project_id": "project_id",
      "return_char_limit": 1,
      "source_code": "source_code",
      "source_type": "source_type",
      "tags": [
        "string"
      ],
      "tool_type": "custom"
    }
  ],
  "base_template_id": "base_template_id",
  "compaction_settings": {
    "clip_chars": 0,
    "mode": "all",
    "model": "model",
    "model_settings": {
      "max_output_tokens": 0,
      "parallel_tool_calls": true,
      "provider_type": "openai",
      "reasoning": {
        "reasoning_effort": "none"
      },
      "response_format": {
        "type": "text"
      },
      "strict": true,
      "temperature": 0
    },
    "prompt": "prompt",
    "prompt_acknowledgement": true,
    "sliding_window_percentage": 0
  },
  "created_at": "2019-12-27T18:11:19.117Z",
  "created_by_id": "created_by_id",
  "deployment_id": "deployment_id",
  "description": "description",
  "embedding": "embedding",
  "embedding_config": {
    "embedding_dim": 0,
    "embedding_endpoint_type": "openai",
    "embedding_model": "embedding_model",
    "azure_deployment": "azure_deployment",
    "azure_endpoint": "azure_endpoint",
    "azure_version": "azure_version",
    "batch_size": 0,
    "embedding_chunk_size": 0,
    "embedding_endpoint": "embedding_endpoint",
    "handle": "handle"
  },
  "enable_sleeptime": true,
  "entity_id": "entity_id",
  "hidden": true,
  "identities": [
    {
      "id": "identity-123e4567-e89b-12d3-a456-426614174000",
      "agent_ids": [
        "string"
      ],
      "block_ids": [
        "string"
      ],
      "identifier_key": "identifier_key",
      "identity_type": "org",
      "name": "name",
      "project_id": "project_id",
      "properties": [
        {
          "key": "key",
          "type": "string",
          "value": "string"
        }
      ]
    }
  ],
  "identity_ids": [
    "string"
  ],
  "last_run_completion": "2019-12-27T18:11:19.117Z",
  "last_run_duration_ms": 0,
  "last_stop_reason": "end_turn",
  "last_updated_by_id": "last_updated_by_id",
  "managed_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "max_files_open": 0,
  "message_buffer_autoclear": true,
  "message_ids": [
    "string"
  ],
  "metadata": {
    "foo": "bar"
  },
  "model": "model",
  "model_settings": {
    "max_output_tokens": 0,
    "parallel_tool_calls": true,
    "provider_type": "openai",
    "reasoning": {
      "reasoning_effort": "none"
    },
    "response_format": {
      "type": "text"
    },
    "strict": true,
    "temperature": 0
  },
  "multi_agent_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "pending_approval": {
    "id": "id",
    "date": "2019-12-27T18:11:19.117Z",
    "tool_call": {
      "arguments": "arguments",
      "name": "name",
      "tool_call_id": "tool_call_id"
    },
    "is_err": true,
    "message_type": "approval_request_message",
    "name": "name",
    "otid": "otid",
    "run_id": "run_id",
    "sender_id": "sender_id",
    "seq_id": 0,
    "step_id": "step_id",
    "tool_calls": [
      {
        "arguments": "arguments",
        "name": "name",
        "tool_call_id": "tool_call_id"
      }
    ]
  },
  "per_file_view_window_char_limit": 0,
  "project_id": "project_id",
  "response_format": {
    "type": "text"
  },
  "secrets": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "template_id": "template_id",
  "timezone": "timezone",
  "tool_exec_environment_variables": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "tool_rules": [
    {
      "children": [
        "string"
      ],
      "tool_name": "tool_name",
      "child_arg_nodes": [
        {
          "name": "name",
          "args": {
            "foo": "bar"
          }
        }
      ],
      "prompt_template": "prompt_template",
      "type": "constrain_child_tools"
    }
  ],
  "updated_at": "2019-12-27T18:11:19.117Z"
}
```

## Delete Agent

`client.agents.delete(stringagentID, RequestOptionsoptions?): AgentDeleteResponse`

**delete** `/v1/agents/{agent_id}`

Delete an agent.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `AgentDeleteResponse = unknown`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const agent = await client.agents.delete('agent-123e4567-e89b-42d3-8456-426614174000');

console.log(agent);
```

#### Response

```json
{}
```

## Export Agent

`client.agents.exportFile(stringagentID, AgentExportFileParamsquery?, RequestOptionsoptions?): AgentExportFileResponse`

**get** `/v1/agents/{agent_id}/export`

Export the serialized JSON representation of an agent, formatted with indentation.

### Parameters

- `agentID: string`

- `query: AgentExportFileParams`

  - `conversation_id?: string | null`

    Conversation ID to export. If provided, uses messages from this conversation instead of the agent's global message history.

  - `max_steps?: number`

  - `scrub_messages?: boolean`

    If True, excludes all messages from the export. Useful for sharing agent configs without conversation history.

  - `use_legacy_format?: boolean`

    If True, exports using the legacy single-agent 'v1' format with inline tools/blocks. If False, exports using the new multi-entity 'v2' format, with separate agents, tools, blocks, files, etc.

### Returns

- `AgentExportFileResponse = string`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const response = await client.agents.exportFile('agent_id');

console.log(response);
```

#### Response

```json
"string"
```

## Import Agent

`client.agents.importFile(AgentImportFileParamsparams, RequestOptionsoptions?): AgentImportFileResponse`

**post** `/v1/agents/import`

Import a serialized agent file and recreate the agent(s) in the system.
Returns the IDs of all imported agents.

### Parameters

- `params: AgentImportFileParams`

  - `file: Uploadable`

    Body param

  - `append_copy_suffix?: boolean`

    Body param: If set to True, appends "_copy" to the end of the agent name.

  - `embedding?: string | null`

    Body param: Embedding handle to override with.

  - `env_vars_json?: string | null`

    Body param: Environment variables as a JSON string to pass to the agent for tool execution. Use 'secrets' instead.

  - `model?: string | null`

    Body param: Model handle to override the agent's default model. This allows the imported agent to use a different model while keeping other defaults (e.g., context size) from the original configuration.

  - `name?: string | null`

    Body param: If provided, overrides the agent name with this value.

  - `override_embedding_handle?: string | null`

    Body param: Override import with specific embedding handle. Use 'embedding' instead.

  - `override_existing_tools?: boolean`

    Body param: If set to True, existing tools can get their source code overwritten by the uploaded tool definitions. Note that Letta core tools can never be updated externally.

  - `override_model_handle?: string | null`

    Body param: Model handle to override the agent's default model. Use 'model' instead.

  - `override_name?: string | null`

    Body param: If provided, overrides the agent name with this value. Use 'name' instead.

  - `project_id?: string | null`

    Body param: The project ID to associate the uploaded agent with. This is now passed via headers.

  - `secrets?: string | null`

    Body param: Secrets as a JSON string to pass to the agent for tool execution.

  - `strip_messages?: boolean`

    Body param: If set to True, strips all messages from the agent before importing.

  - `xOverrideEmbeddingModel?: string`

    Header param

### Returns

- `AgentImportFileResponse`

  Response model for imported agents

  - `agent_ids: Array<string>`

    List of IDs of the imported agents

### Example

```typescript
import fs from 'fs';
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const response = await client.agents.importFile({ file: fs.createReadStream('path/to/file') });

console.log(response.agent_ids);
```

#### Response

```json
{
  "agent_ids": [
    "string"
  ]
}
```

## Recompile Agent

`client.agents.recompile(stringagentID, AgentRecompileParamsparams?, RequestOptionsoptions?): AgentRecompileResponse`

**post** `/v1/agents/{agent_id}/recompile`

Manually trigger system prompt recompilation for an agent.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `params: AgentRecompileParams`

  - `dry_run?: boolean`

    If True, do not persist changes; still returns the compiled system prompt.

  - `update_timestamp?: boolean`

    If True, update the in-context memory last edit timestamp embedded in the system prompt.

### Returns

- `AgentRecompileResponse = string`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const response = await client.agents.recompile('agent-123e4567-e89b-42d3-8456-426614174000');

console.log(response);
```

#### Response

```json
"string"
```

## Domain Types

### Agent Environment Variable

- `AgentEnvironmentVariable`

  - `agent_id: string`

    The ID of the agent this environment variable belongs to.

  - `key: string`

    The name of the environment variable.

  - `value: string`

    The value of the environment variable.

  - `id?: string`

    The human-friendly ID of the Agent-env

  - `created_at?: string | null`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `description?: string | null`

    An optional description of the environment variable.

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

  - `value_enc?: string | null`

    Encrypted secret value (stored as encrypted string)

### Agent State

- `AgentState`

  Representation of an agent's state. This is the state of the agent at a given time, and is persisted in the DB backend. The state has all the information needed to recreate a persisted agent.

  - `id: string`

    The id of the agent. Assigned by the database.

  - `agent_type: AgentType`

    The type of agent.

    - `"memgpt_agent"`

    - `"memgpt_v2_agent"`

    - `"letta_v1_agent"`

    - `"react_agent"`

    - `"workflow_agent"`

    - `"split_thread_agent"`

    - `"sleeptime_agent"`

    - `"voice_convo_agent"`

    - `"voice_sleeptime_agent"`

  - `blocks: Array<Block>`

    The memory blocks used by the agent.

    - `value: string`

      Value of the block.

    - `id?: string`

      The human-friendly ID of the Block

    - `base_template_id?: string | null`

      The base template id of the block.

    - `created_by_id?: string | null`

      The id of the user that made this Block.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      Description of the block.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the block will be hidden.

    - `is_template?: boolean`

      Whether the block is a template (e.g. saved human/persona options).

    - `label?: string | null`

      Label of the block (e.g. 'human', 'persona') in the context window.

    - `last_updated_by_id?: string | null`

      The id of the user that last updated this Block.

    - `limit?: number`

      Character limit of the block.

    - `metadata?: Record<string, unknown> | null`

      Metadata of the block.

    - `preserve_on_migration?: boolean | null`

      Preserve the block on template migration.

    - `project_id?: string | null`

      The associated project id.

    - `read_only?: boolean`

      Whether the agent has read-only access to the block.

    - `tags?: Array<string> | null`

      The tags associated with the block.

    - `template_id?: string | null`

      The id of the template.

    - `template_name?: string | null`

      Name of the block if it is a template.

  - `llm_config: LlmConfig`

    Deprecated: Use `model` field instead. The LLM configuration used by the agent.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

        - `type?: "text"`

          The type of the response format.

          - `"text"`

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

        - `json_schema: Record<string, unknown>`

          The JSON schema of the response.

        - `type?: "json_schema"`

          The type of the response format.

          - `"json_schema"`

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

        - `type?: "json_object"`

          The type of the response format.

          - `"json_object"`

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `memory: Memory`

    Deprecated: Use `blocks` field instead. The in-context memory of the agent.

    - `blocks: Array<Block>`

      Memory blocks contained in the agent's in-context memory

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `agent_type?: AgentType | (string & {}) | null`

      Agent type controlling prompt rendering.

      - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

        Enum to represent the type of agent.

      - `(string & {})`

    - `file_blocks?: Array<FileBlock>`

      Special blocks representing the agent's in-context memory of an attached file

      - `file_id: string`

        Unique identifier of the file.

      - `is_open: boolean`

        True if the agent currently has the file open.

      - `source_id: string`

        Deprecated: Use `folder_id` field instead. Unique identifier of the source.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_accessed_at?: string | null`

        UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `git_enabled?: boolean`

      Whether this agent uses git-backed memory with structured labels.

    - `prompt_template?: string`

      Deprecated. Ignored for performance.

  - `name: string`

    The name of the agent.

  - `sources: Array<Source>`

    Deprecated: Use `folders` field instead. The sources used by the agent.

    - `id: string`

      The human-friendly ID of the Source

    - `embedding_config: EmbeddingConfig`

      The embedding configuration used by the source.

      - `embedding_dim: number`

        The dimension of the embedding.

      - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"bedrock"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"mistral"`

        - `"together"`

        - `"pinecone"`

      - `embedding_model: string`

        The model for the embedding.

      - `azure_deployment?: string | null`

        The Azure deployment for the model.

      - `azure_endpoint?: string | null`

        The Azure endpoint for the model.

      - `azure_version?: string | null`

        The Azure version for the model.

      - `batch_size?: number`

        The maximum batch size for processing embeddings.

      - `embedding_chunk_size?: number | null`

        The chunk size of the embedding.

      - `embedding_endpoint?: string | null`

        The endpoint for the model (`None` if local).

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

    - `name: string`

      The name of the source.

    - `created_at?: string | null`

      The timestamp when the source was created.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `description?: string | null`

      The description of the source.

    - `instructions?: string | null`

      Instructions for how to use the source.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata?: Record<string, unknown> | null`

      Metadata associated with the source.

    - `updated_at?: string | null`

      The timestamp when the source was last updated.

    - `vector_db_provider?: VectorDBProvider`

      The vector database provider used for this source's passages

      - `"native"`

      - `"tpuf"`

      - `"pinecone"`

  - `system: string`

    The system prompt used by the agent.

  - `tags: Array<string>`

    The tags associated with the agent.

  - `tools: Array<Tool>`

    The tools used by the agent.

    - `id: string`

      The human-friendly ID of the Tool

    - `args_json_schema?: Record<string, unknown> | null`

      The args JSON schema of the function.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `default_requires_approval?: boolean | null`

      Default value for whether or not executing this tool requires approval.

    - `description?: string | null`

      The description of the tool.

    - `enable_parallel_execution?: boolean | null`

      If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

    - `json_schema?: Record<string, unknown> | null`

      The JSON schema of the function.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata_?: Record<string, unknown> | null`

      A dictionary of additional metadata for the tool.

    - `name?: string | null`

      The name of the function.

    - `npm_requirements?: Array<NpmRequirement> | null`

      Optional list of npm packages required by this tool.

      - `name: string`

        Name of the npm package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `pip_requirements?: Array<PipRequirement> | null`

      Optional list of pip packages required by this tool.

      - `name: string`

        Name of the pip package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `project_id?: string | null`

      The project id of the tool.

    - `return_char_limit?: number`

      The maximum number of characters in the response.

    - `source_code?: string | null`

      The source code of the function.

    - `source_type?: string | null`

      The type of the source code.

    - `tags?: Array<string>`

      Metadata tags.

    - `tool_type?: ToolType`

      The type of the tool.

      - `"custom"`

      - `"letta_core"`

      - `"letta_memory_core"`

      - `"letta_multi_agent_core"`

      - `"letta_sleeptime_core"`

      - `"letta_voice_sleeptime_core"`

      - `"letta_builtin"`

      - `"letta_files_core"`

      - `"external_langchain"`

      - `"external_composio"`

      - `"external_mcp"`

  - `base_template_id?: string | null`

    The base template id of the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `created_at?: string | null`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the agent will be hidden.

  - `identities?: Array<Identity>`

    The identities associated with this agent.

    - `id: string`

      The human-friendly ID of the Identity

    - `agent_ids: Array<string>`

      The IDs of the agents associated with the identity.

    - `block_ids: Array<string>`

      The IDs of the blocks associated with the identity.

    - `identifier_key: string`

      External, user-generated identifier key of the identity.

    - `identity_type: "org" | "user" | "other"`

      The type of the identity.

      - `"org"`

      - `"user"`

      - `"other"`

    - `name: string`

      The name of the identity.

    - `project_id?: string | null`

      The project id of the identity, if applicable.

    - `properties?: Array<Property>`

      List of properties associated with the identity

      - `key: string`

        The key of the property

      - `type: "string" | "number" | "boolean" | "json"`

        The type of the property

        - `"string"`

        - `"number"`

        - `"boolean"`

        - `"json"`

      - `value: string | number | boolean | Record<string, unknown>`

        The value of the property

        - `string`

        - `number`

        - `boolean`

        - `Record<string, unknown>`

  - `identity_ids?: Array<string>`

    Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

  - `last_run_completion?: string | null`

    The timestamp when the agent last completed a run.

  - `last_run_duration_ms?: number | null`

    The duration in milliseconds of the agent's last run.

  - `last_stop_reason?: StopReasonType | null`

    The stop reason from the agent's last run.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `managed_group?: ManagedGroup | null`

    The multi-agent group that this agent manages

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `message_buffer_autoclear?: boolean`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `message_ids?: Array<string> | null`

    The ids of the messages in the agent's in-context memory.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle used by the agent (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings used by the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `multi_agent_group?: MultiAgentGroup | null`

    Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `pending_approval?: ApprovalRequestMessage | null`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project_id?: string | null`

    The id of the project the agent belongs to.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format used by the agent

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Array<AgentEnvironmentVariable>`

    The environment variables for tool execution specific to this agent.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `template_id?: string | null`

    The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

    Deprecated: use `secrets` field instead.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The list of tool rules.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Agent Type

- `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

  Enum to represent the type of agent.

  - `"memgpt_agent"`

  - `"memgpt_v2_agent"`

  - `"letta_v1_agent"`

  - `"react_agent"`

  - `"workflow_agent"`

  - `"split_thread_agent"`

  - `"sleeptime_agent"`

  - `"voice_convo_agent"`

  - `"voice_sleeptime_agent"`

### Anthropic Model Settings

- `AnthropicModelSettings`

  - `effort?: "low" | "medium" | "high" | 2 more | null`

    Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

    - `"low"`

    - `"medium"`

    - `"high"`

    - `"xhigh"`

    - `"max"`

  - `max_output_tokens?: number`

    The maximum number of tokens the model can generate.

  - `parallel_tool_calls?: boolean`

    Whether to enable parallel tool calling.

  - `provider_type?: "anthropic"`

    The type of the provider.

    - `"anthropic"`

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format for the model.

    - `TextResponseFormat`

      Response format for plain text responses.

      - `type?: "text"`

        The type of the response format.

        - `"text"`

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

      - `json_schema: Record<string, unknown>`

        The JSON schema of the response.

      - `type?: "json_schema"`

        The type of the response format.

        - `"json_schema"`

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

      - `type?: "json_object"`

        The type of the response format.

        - `"json_object"`

  - `strict?: boolean`

    Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

  - `temperature?: number`

    The temperature of the model.

  - `thinking?: Thinking`

    The thinking configuration for the model.

    - `budget_tokens?: number`

      The maximum number of tokens the model can use for extended thinking.

    - `type?: "enabled" | "disabled"`

      The type of thinking to use.

      - `"enabled"`

      - `"disabled"`

  - `verbosity?: "low" | "medium" | "high" | null`

    Soft control for how verbose model output should be, used for GPT-5 models.

    - `"low"`

    - `"medium"`

    - `"high"`

### Azure Model Settings

- `AzureModelSettings`

  Azure OpenAI model configuration (OpenAI-compatible).

  - `max_output_tokens?: number`

    The maximum number of tokens the model can generate.

  - `parallel_tool_calls?: boolean`

    Whether to enable parallel tool calling.

  - `provider_type?: "azure"`

    The type of the provider.

    - `"azure"`

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format for the model.

    - `TextResponseFormat`

      Response format for plain text responses.

      - `type?: "text"`

        The type of the response format.

        - `"text"`

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

      - `json_schema: Record<string, unknown>`

        The JSON schema of the response.

      - `type?: "json_schema"`

        The type of the response format.

        - `"json_schema"`

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

      - `type?: "json_object"`

        The type of the response format.

        - `"json_object"`

  - `temperature?: number`

    The temperature of the model.

### Bedrock Model Settings

- `BedrockModelSettings`

  AWS Bedrock model configuration.

  - `max_output_tokens?: number`

    The maximum number of tokens the model can generate.

  - `parallel_tool_calls?: boolean`

    Whether to enable parallel tool calling.

  - `provider_type?: "bedrock"`

    The type of the provider.

    - `"bedrock"`

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format for the model.

    - `TextResponseFormat`

      Response format for plain text responses.

      - `type?: "text"`

        The type of the response format.

        - `"text"`

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

      - `json_schema: Record<string, unknown>`

        The JSON schema of the response.

      - `type?: "json_schema"`

        The type of the response format.

        - `"json_schema"`

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

      - `type?: "json_object"`

        The type of the response format.

        - `"json_object"`

  - `temperature?: number`

    The temperature of the model.

### Child Tool Rule

- `ChildToolRule`

  A ToolRule represents a tool that can be invoked by the agent.

  - `children: Array<string>`

    The children tools that can be invoked.

  - `tool_name: string`

    The name of the tool. Must exist in the database for the user's organization.

  - `child_arg_nodes?: Array<ChildArgNode> | null`

    Optional list of typed child argument overrides. Each node must reference a child in 'children'.

    - `name: string`

      The name of the child tool to invoke next.

    - `args?: Record<string, unknown> | null`

      Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

  - `prompt_template?: string | null`

    Optional template string (ignored).

  - `type?: "constrain_child_tools"`

    - `"constrain_child_tools"`

### Conditional Tool Rule

- `ConditionalToolRule`

  A ToolRule that conditionally maps to different child tools based on the output.

  - `child_output_mapping: Record<string, string>`

    The output case to check for mapping

  - `tool_name: string`

    The name of the tool. Must exist in the database for the user's organization.

  - `default_child?: string | null`

    The default child tool to be called. If None, any tool can be called.

  - `prompt_template?: string | null`

    Optional template string (ignored).

  - `require_output_mapping?: boolean`

    Whether to throw an error when output doesn't match any case

  - `type?: "conditional"`

    - `"conditional"`

### Continue Tool Rule

- `ContinueToolRule`

  Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

  - `tool_name: string`

    The name of the tool. Must exist in the database for the user's organization.

  - `prompt_template?: string | null`

    Optional template string (ignored).

  - `type?: "continue_loop"`

    - `"continue_loop"`

### Deepseek Model Settings

- `DeepseekModelSettings`

  Deepseek model configuration (OpenAI-compatible).

  - `max_output_tokens?: number`

    The maximum number of tokens the model can generate.

  - `parallel_tool_calls?: boolean`

    Whether to enable parallel tool calling.

  - `provider_type?: "deepseek"`

    The type of the provider.

    - `"deepseek"`

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format for the model.

    - `TextResponseFormat`

      Response format for plain text responses.

      - `type?: "text"`

        The type of the response format.

        - `"text"`

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

      - `json_schema: Record<string, unknown>`

        The JSON schema of the response.

      - `type?: "json_schema"`

        The type of the response format.

        - `"json_schema"`

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

      - `type?: "json_object"`

        The type of the response format.

        - `"json_object"`

  - `temperature?: number`

    The temperature of the model.

### Google AI Model Settings

- `GoogleAIModelSettings`

  - `max_output_tokens?: number`

    The maximum number of tokens the model can generate.

  - `parallel_tool_calls?: boolean`

    Whether to enable parallel tool calling.

  - `provider_type?: "google_ai"`

    The type of the provider.

    - `"google_ai"`

  - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response schema for the model.

    - `TextResponseFormat`

      Response format for plain text responses.

      - `type?: "text"`

        The type of the response format.

        - `"text"`

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

      - `json_schema: Record<string, unknown>`

        The JSON schema of the response.

      - `type?: "json_schema"`

        The type of the response format.

        - `"json_schema"`

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

      - `type?: "json_object"`

        The type of the response format.

        - `"json_object"`

  - `temperature?: number`

    The temperature of the model.

  - `thinking_config?: ThinkingConfig`

    The thinking configuration for the model.

    - `include_thoughts?: boolean`

      Whether to include thoughts in the model's response.

    - `thinking_budget?: number`

      The thinking budget for the model.

### Google Vertex Model Settings

- `GoogleVertexModelSettings`

  - `max_output_tokens?: number`

    The maximum number of tokens the model can generate.

  - `parallel_tool_calls?: boolean`

    Whether to enable parallel tool calling.

  - `provider_type?: "google_vertex"`

    The type of the provider.

    - `"google_vertex"`

  - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response schema for the model.

    - `TextResponseFormat`

      Response format for plain text responses.

      - `type?: "text"`

        The type of the response format.

        - `"text"`

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

      - `json_schema: Record<string, unknown>`

        The JSON schema of the response.

      - `type?: "json_schema"`

        The type of the response format.

        - `"json_schema"`

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

      - `type?: "json_object"`

        The type of the response format.

        - `"json_object"`

  - `temperature?: number`

    The temperature of the model.

  - `thinking_config?: ThinkingConfig`

    The thinking configuration for the model.

    - `include_thoughts?: boolean`

      Whether to include thoughts in the model's response.

    - `thinking_budget?: number`

      The thinking budget for the model.

### Groq Model Settings

- `GroqModelSettings`

  Groq model configuration (OpenAI-compatible).

  - `max_output_tokens?: number`

    The maximum number of tokens the model can generate.

  - `parallel_tool_calls?: boolean`

    Whether to enable parallel tool calling.

  - `provider_type?: "groq"`

    The type of the provider.

    - `"groq"`

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format for the model.

    - `TextResponseFormat`

      Response format for plain text responses.

      - `type?: "text"`

        The type of the response format.

        - `"text"`

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

      - `json_schema: Record<string, unknown>`

        The JSON schema of the response.

      - `type?: "json_schema"`

        The type of the response format.

        - `"json_schema"`

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

      - `type?: "json_object"`

        The type of the response format.

        - `"json_object"`

  - `temperature?: number`

    The temperature of the model.

### Init Tool Rule

- `InitToolRule`

  Represents the initial tool rule configuration.

  - `tool_name: string`

    The name of the tool. Must exist in the database for the user's organization.

  - `args?: Record<string, unknown> | null`

    Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

  - `prompt_template?: string | null`

    Optional template string (ignored). Rendering uses fast built-in formatting for performance.

  - `type?: "run_first"`

    - `"run_first"`

### Json Object Response Format

- `JsonObjectResponseFormat`

  Response format for JSON object responses.

  - `type?: "json_object"`

    The type of the response format.

    - `"json_object"`

### Json Schema Response Format

- `JsonSchemaResponseFormat`

  Response format for JSON schema-based responses.

  - `json_schema: Record<string, unknown>`

    The JSON schema of the response.

  - `type?: "json_schema"`

    The type of the response format.

    - `"json_schema"`

### Letta Message Content Union

- `LettaMessageContentUnion = TextContent | ImageContent | ToolCallContent | 4 more`

  Sent via the Anthropic Messages API

  - `TextContent`

    - `text: string`

      The text content of the message.

    - `signature?: string | null`

      Stores a unique identifier for any reasoning associated with this text content.

    - `type?: "text"`

      The type of the message.

      - `"text"`

  - `ImageContent`

    - `source: URLImage | Base64Image | LettaImage`

      The source of the image.

      - `URLImage`

        - `url: string`

          The URL of the image.

        - `type?: "url"`

          The source type for the image.

          - `"url"`

      - `Base64Image`

        - `data: string`

          The base64 encoded image data.

        - `media_type: string`

          The media type for the image.

        - `detail?: string | null`

          What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

        - `type?: "base64"`

          The source type for the image.

          - `"base64"`

      - `LettaImage`

        - `file_id: string`

          The unique identifier of the image file persisted in storage.

        - `data?: string | null`

          The base64 encoded image data.

        - `detail?: string | null`

          What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

        - `media_type?: string | null`

          The media type for the image.

        - `type?: "letta"`

          The source type for the image.

          - `"letta"`

    - `type?: "image"`

      The type of the message.

      - `"image"`

  - `ToolCallContent`

    - `id: string`

      A unique identifier for this specific tool call instance.

    - `input: Record<string, unknown>`

      The parameters being passed to the tool, structured as a dictionary of parameter names to values.

    - `name: string`

      The name of the tool being called.

    - `signature?: string | null`

      Stores a unique identifier for any reasoning associated with this tool call.

    - `type?: "tool_call"`

      Indicates this content represents a tool call event.

      - `"tool_call"`

  - `ToolReturnContent`

    - `content: string`

      The content returned by the tool execution.

    - `is_error: boolean`

      Indicates whether the tool execution resulted in an error.

    - `tool_call_id: string`

      References the ID of the ToolCallContent that initiated this tool call.

    - `type?: "tool_return"`

      Indicates this content represents a tool return event.

      - `"tool_return"`

  - `ReasoningContent`

    Sent via the Anthropic Messages API

    - `is_native: boolean`

      Whether the reasoning content was generated by a reasoner model that processed this step.

    - `reasoning: string`

      The intermediate reasoning or thought process content.

    - `signature?: string | null`

      A unique identifier for this reasoning step.

    - `type?: "reasoning"`

      Indicates this is a reasoning/intermediate step.

      - `"reasoning"`

  - `RedactedReasoningContent`

    Sent via the Anthropic Messages API

    - `data: string`

      The redacted or filtered intermediate reasoning content.

    - `type?: "redacted_reasoning"`

      Indicates this is a redacted thinking step.

      - `"redacted_reasoning"`

  - `OmittedReasoningContent`

    A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

    - `signature?: string | null`

      A unique identifier for this reasoning step.

    - `type?: "omitted_reasoning"`

      Indicates this is an omitted reasoning step.

      - `"omitted_reasoning"`

### Max Count Per Step Tool Rule

- `MaxCountPerStepToolRule`

  Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

  - `max_count_limit: number`

    The max limit for the total number of times this tool can be invoked in a single step.

  - `tool_name: string`

    The name of the tool. Must exist in the database for the user's organization.

  - `prompt_template?: string | null`

    Optional template string (ignored).

  - `type?: "max_count_per_step"`

    - `"max_count_per_step"`

### Message Create

- `MessageCreate`

  Request to create a message

  - `content: Array<LettaMessageContentUnion> | string`

    The content of the message.

    - `Array<LettaMessageContentUnion>`

      - `TextContent`

        - `text: string`

          The text content of the message.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this text content.

        - `type?: "text"`

          The type of the message.

          - `"text"`

      - `ImageContent`

        - `source: URLImage | Base64Image | LettaImage`

          The source of the image.

          - `URLImage`

            - `url: string`

              The URL of the image.

            - `type?: "url"`

              The source type for the image.

              - `"url"`

          - `Base64Image`

            - `data: string`

              The base64 encoded image data.

            - `media_type: string`

              The media type for the image.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `type?: "base64"`

              The source type for the image.

              - `"base64"`

          - `LettaImage`

            - `file_id: string`

              The unique identifier of the image file persisted in storage.

            - `data?: string | null`

              The base64 encoded image data.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `media_type?: string | null`

              The media type for the image.

            - `type?: "letta"`

              The source type for the image.

              - `"letta"`

        - `type?: "image"`

          The type of the message.

          - `"image"`

      - `ToolCallContent`

        - `id: string`

          A unique identifier for this specific tool call instance.

        - `input: Record<string, unknown>`

          The parameters being passed to the tool, structured as a dictionary of parameter names to values.

        - `name: string`

          The name of the tool being called.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this tool call.

        - `type?: "tool_call"`

          Indicates this content represents a tool call event.

          - `"tool_call"`

      - `ToolReturnContent`

        - `content: string`

          The content returned by the tool execution.

        - `is_error: boolean`

          Indicates whether the tool execution resulted in an error.

        - `tool_call_id: string`

          References the ID of the ToolCallContent that initiated this tool call.

        - `type?: "tool_return"`

          Indicates this content represents a tool return event.

          - `"tool_return"`

      - `ReasoningContent`

        Sent via the Anthropic Messages API

        - `is_native: boolean`

          Whether the reasoning content was generated by a reasoner model that processed this step.

        - `reasoning: string`

          The intermediate reasoning or thought process content.

        - `signature?: string | null`

          A unique identifier for this reasoning step.

        - `type?: "reasoning"`

          Indicates this is a reasoning/intermediate step.

          - `"reasoning"`

      - `RedactedReasoningContent`

        Sent via the Anthropic Messages API

        - `data: string`

          The redacted or filtered intermediate reasoning content.

        - `type?: "redacted_reasoning"`

          Indicates this is a redacted thinking step.

          - `"redacted_reasoning"`

      - `OmittedReasoningContent`

        A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

        - `signature?: string | null`

          A unique identifier for this reasoning step.

        - `type?: "omitted_reasoning"`

          Indicates this is an omitted reasoning step.

          - `"omitted_reasoning"`

    - `string`

  - `role: "user" | "system" | "assistant"`

    The role of the participant.

    - `"user"`

    - `"system"`

    - `"assistant"`

  - `batch_item_id?: string | null`

    The id of the LLMBatchItem that this message is associated with

  - `group_id?: string | null`

    The multi-agent group that the message was sent in

  - `name?: string | null`

    The name of the participant.

  - `otid?: string | null`

    The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

  - `sender_id?: string | null`

    The id of the sender of the message, can be an identity id or agent id

  - `type?: "message" | null`

    The message type to be created.

    - `"message"`

### OpenAI Model Settings

- `OpenAIModelSettings`

  - `max_output_tokens?: number`

    The maximum number of tokens the model can generate.

  - `parallel_tool_calls?: boolean`

    Whether to enable parallel tool calling.

  - `provider_type?: "openai"`

    The type of the provider.

    - `"openai"`

  - `reasoning?: Reasoning`

    The reasoning configuration for the model.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format for the model.

    - `TextResponseFormat`

      Response format for plain text responses.

      - `type?: "text"`

        The type of the response format.

        - `"text"`

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

      - `json_schema: Record<string, unknown>`

        The JSON schema of the response.

      - `type?: "json_schema"`

        The type of the response format.

        - `"json_schema"`

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

      - `type?: "json_object"`

        The type of the response format.

        - `"json_object"`

  - `strict?: boolean`

    Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

  - `temperature?: number`

    The temperature of the model.

### Parent Tool Rule

- `ParentToolRule`

  A ToolRule that only allows a child tool to be called if the parent has been called.

  - `children: Array<string>`

    The children tools that can be invoked.

  - `tool_name: string`

    The name of the tool. Must exist in the database for the user's organization.

  - `prompt_template?: string | null`

    Optional template string (ignored).

  - `type?: "parent_last_tool"`

    - `"parent_last_tool"`

### Required Before Exit Tool Rule

- `RequiredBeforeExitToolRule`

  Represents a tool rule configuration where this tool must be called before the agent loop can exit.

  - `tool_name: string`

    The name of the tool. Must exist in the database for the user's organization.

  - `prompt_template?: string | null`

    Optional template string (ignored).

  - `type?: "required_before_exit"`

    - `"required_before_exit"`

### Requires Approval Tool Rule

- `RequiresApprovalToolRule`

  Represents a tool rule configuration which requires approval before the tool can be invoked.

  - `tool_name: string`

    The name of the tool. Must exist in the database for the user's organization.

  - `prompt_template?: string | null`

    Optional template string (ignored). Rendering uses fast built-in formatting for performance.

  - `type?: "requires_approval"`

    - `"requires_approval"`

### Terminal Tool Rule

- `TerminalToolRule`

  Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

  - `tool_name: string`

    The name of the tool. Must exist in the database for the user's organization.

  - `prompt_template?: string | null`

    Optional template string (ignored).

  - `type?: "exit_loop"`

    - `"exit_loop"`

### Text Response Format

- `TextResponseFormat`

  Response format for plain text responses.

  - `type?: "text"`

    The type of the response format.

    - `"text"`

### Together Model Settings

- `TogetherModelSettings`

  Together AI model configuration (OpenAI-compatible).

  - `max_output_tokens?: number`

    The maximum number of tokens the model can generate.

  - `parallel_tool_calls?: boolean`

    Whether to enable parallel tool calling.

  - `provider_type?: "together"`

    The type of the provider.

    - `"together"`

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format for the model.

    - `TextResponseFormat`

      Response format for plain text responses.

      - `type?: "text"`

        The type of the response format.

        - `"text"`

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

      - `json_schema: Record<string, unknown>`

        The JSON schema of the response.

      - `type?: "json_schema"`

        The type of the response format.

        - `"json_schema"`

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

      - `type?: "json_object"`

        The type of the response format.

        - `"json_object"`

  - `temperature?: number`

    The temperature of the model.

### Xai Model Settings

- `XaiModelSettings`

  xAI model configuration (OpenAI-compatible).

  - `max_output_tokens?: number`

    The maximum number of tokens the model can generate.

  - `parallel_tool_calls?: boolean`

    Whether to enable parallel tool calling.

  - `provider_type?: "xai"`

    The type of the provider.

    - `"xai"`

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format for the model.

    - `TextResponseFormat`

      Response format for plain text responses.

      - `type?: "text"`

        The type of the response format.

        - `"text"`

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

      - `json_schema: Record<string, unknown>`

        The JSON schema of the response.

      - `type?: "json_schema"`

        The type of the response format.

        - `"json_schema"`

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

      - `type?: "json_object"`

        The type of the response format.

        - `"json_object"`

  - `temperature?: number`

    The temperature of the model.

### Agent Delete Response

- `AgentDeleteResponse = unknown`

### Agent Export File Response

- `AgentExportFileResponse = string`

### Agent Import File Response

- `AgentImportFileResponse`

  Response model for imported agents

  - `agent_ids: Array<string>`

    List of IDs of the imported agents

### Agent Recompile Response

- `AgentRecompileResponse = string`

# Messages

## List Messages

`client.agents.messages.list(stringagentID, MessageListParamsquery?, RequestOptionsoptions?): ArrayPage<Message>`

**get** `/v1/agents/{agent_id}/messages`

Retrieve message history for an agent.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `query: MessageListParams`

  - `after?: string | null`

    Cursor for pagination (message ID). Returns results relative to this ID in the specified sort order. Expected format: 'message-<uuid4>'

  - `assistant_message_tool_kwarg?: string`

    The name of the message argument.

  - `assistant_message_tool_name?: string`

    The name of the designated message tool.

  - `before?: string | null`

    Cursor for pagination (message ID). Returns results relative to this ID in the specified sort order. Expected format: 'message-<uuid4>'

  - `conversation_id?: string | null`

    Conversation ID to filter messages by.

  - `group_id?: string | null`

    Group ID to filter messages by.

  - `include_err?: boolean | null`

    Whether to include error messages and error statuses. For debugging purposes only.

  - `include_return_message_types?: Array<MessageType> | null`

    Message types to include in response. When null, all message types are returned.

    - `"system_message"`

    - `"user_message"`

    - `"assistant_message"`

    - `"reasoning_message"`

    - `"hidden_reasoning_message"`

    - `"tool_call_message"`

    - `"tool_return_message"`

    - `"approval_request_message"`

    - `"approval_response_message"`

    - `"summary_message"`

    - `"event_message"`

  - `limit?: number | null`

    Maximum number of messages to return

  - `order?: "asc" | "desc"`

    Sort order for messages by creation time. 'asc' for oldest first, 'desc' for newest first

    - `"asc"`

    - `"desc"`

  - `order_by?: "created_at"`

    Field to sort by

    - `"created_at"`

  - `use_assistant_message?: boolean`

    Whether to use assistant messages

### Returns

- `Message = SystemMessage | UserMessage | ReasoningMessage | 8 more`

  A message generated by the system. Never streamed back on a response, only used for cursor pagination.

  Args:
  id (str): The ID of the message
  date (datetime): The date the message was created in ISO format
  name (Optional[str]): The name of the sender of the message
  content (str): The message content sent by the system

  - `SystemMessage`

    A message generated by the system. Never streamed back on a response, only used for cursor pagination.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    content (str): The message content sent by the system

    - `id: string`

    - `content: string`

      The message content sent by the system

    - `date: string`

    - `is_err?: boolean | null`

    - `message_type?: "system_message"`

      The type of the message.

      - `"system_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `UserMessage`

    A message sent by the user. Never streamed back on a response, only used for cursor pagination.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    content (Union[str, List[LettaUserMessageContentUnion]]): The message content sent by the user (can be a string or an array of multi-modal content parts)

    - `id: string`

    - `content: Array<LettaUserMessageContentUnion> | string`

      The message content sent by the user (can be a string or an array of multi-modal content parts)

      - `Array<LettaUserMessageContentUnion>`

        - `TextContent`

          - `text: string`

            The text content of the message.

          - `signature?: string | null`

            Stores a unique identifier for any reasoning associated with this text content.

          - `type?: "text"`

            The type of the message.

            - `"text"`

        - `ImageContent`

          - `source: URLImage | Base64Image | LettaImage`

            The source of the image.

            - `URLImage`

              - `url: string`

                The URL of the image.

              - `type?: "url"`

                The source type for the image.

                - `"url"`

            - `Base64Image`

              - `data: string`

                The base64 encoded image data.

              - `media_type: string`

                The media type for the image.

              - `detail?: string | null`

                What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

              - `type?: "base64"`

                The source type for the image.

                - `"base64"`

            - `LettaImage`

              - `file_id: string`

                The unique identifier of the image file persisted in storage.

              - `data?: string | null`

                The base64 encoded image data.

              - `detail?: string | null`

                What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

              - `media_type?: string | null`

                The media type for the image.

              - `type?: "letta"`

                The source type for the image.

                - `"letta"`

          - `type?: "image"`

            The type of the message.

            - `"image"`

      - `string`

    - `date: string`

    - `is_err?: boolean | null`

    - `message_type?: "user_message"`

      The type of the message.

      - `"user_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `ReasoningMessage`

    Representation of an agent's internal reasoning.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    source (Literal["reasoner_model", "non_reasoner_model"]): Whether the reasoning
    content was generated natively by a reasoner model or derived via prompting
    reasoning (str): The internal reasoning of the agent
    signature (Optional[str]): The model-generated signature of the reasoning step

    - `id: string`

    - `date: string`

    - `reasoning: string`

    - `is_err?: boolean | null`

    - `message_type?: "reasoning_message"`

      The type of the message.

      - `"reasoning_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `signature?: string | null`

    - `source?: "reasoner_model" | "non_reasoner_model"`

      - `"reasoner_model"`

      - `"non_reasoner_model"`

    - `step_id?: string | null`

  - `HiddenReasoningMessage`

    Representation of an agent's internal reasoning where reasoning content
    has been hidden from the response.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    state (Literal["redacted", "omitted"]): Whether the reasoning
    content was redacted by the provider or simply omitted by the API
    hidden_reasoning (Optional[str]): The internal reasoning of the agent

    - `id: string`

    - `date: string`

    - `state: "redacted" | "omitted"`

      - `"redacted"`

      - `"omitted"`

    - `hidden_reasoning?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "hidden_reasoning_message"`

      The type of the message.

      - `"hidden_reasoning_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `ToolCallMessage`

    A message representing a request to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (Union[ToolCall, ToolCallDelta]): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "tool_call_message"`

      The type of the message.

      - `"tool_call_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `ToolReturnMessage`

    A message representing the return value of a tool call (generated by Letta executing the requested tool).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_return (str): The return value of the tool (deprecated, use tool_returns)
    status (Literal["success", "error"]): The status of the tool call (deprecated, use tool_returns)
    tool_call_id (str): A unique identifier for the tool call that generated this message (deprecated, use tool_returns)
    stdout (Optional[List(str)]): Captured stdout (e.g. prints, logs) from the tool invocation (deprecated, use tool_returns)
    stderr (Optional[List(str)]): Captured stderr from the tool invocation (deprecated, use tool_returns)
    tool_returns (Optional[List[ToolReturn]]): List of tool returns for multi-tool support

    - `id: string`

    - `date: string`

    - `status: "success" | "error"`

      - `"success"`

      - `"error"`

    - `tool_call_id: string`

    - `tool_return: string`

    - `is_err?: boolean | null`

    - `message_type?: "tool_return_message"`

      The type of the message.

      - `"tool_return_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `stderr?: Array<string> | null`

    - `stdout?: Array<string> | null`

    - `step_id?: string | null`

    - `tool_returns?: Array<ToolReturn> | null`

      - `status: "success" | "error"`

        - `"success"`

        - `"error"`

      - `tool_call_id: string`

      - `tool_return: Array<TextContent | ImageContent> | string`

        The tool return value - either a string or list of content parts (text/image)

        - `Array<TextContent | ImageContent>`

          - `TextContent`

          - `ImageContent`

        - `string`

      - `stderr?: Array<string> | null`

      - `stdout?: Array<string> | null`

      - `type?: "tool"`

        The message type to be created.

        - `"tool"`

  - `AssistantMessage`

    A message sent by the LLM in response to user input. Used in the LLM context.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    content (Union[str, List[LettaAssistantMessageContentUnion]]): The message content sent by the agent (can be a string or an array of content parts)

    - `id: string`

    - `content: Array<LettaAssistantMessageContentUnion> | string`

      The message content sent by the agent (can be a string or an array of content parts)

      - `Array<LettaAssistantMessageContentUnion>`

        - `text: string`

          The text content of the message.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this text content.

        - `type?: "text"`

          The type of the message.

          - `"text"`

      - `string`

    - `date: string`

    - `is_err?: boolean | null`

    - `message_type?: "assistant_message"`

      The type of the message.

      - `"assistant_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `ApprovalRequestMessage`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

      - `ToolCallDelta`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `ApprovalResponseMessage`

    A message representing a response form the user indicating whether a tool has been approved to run.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    approve: (bool) Whether the tool has been approved
    approval_request_id: The ID of the approval request
    reason: (Optional[str]) An optional explanation for the provided approval status

    - `id: string`

    - `date: string`

    - `approval_request_id?: string | null`

      The message ID of the approval request

    - `approvals?: Array<ApprovalReturn | ToolReturn> | null`

      The list of approval responses

      - `ApprovalReturn`

        - `approve: boolean`

          Whether the tool has been approved

        - `tool_call_id: string`

          The ID of the tool call that corresponds to this approval

        - `reason?: string | null`

          An optional explanation for the provided approval status

        - `type?: "approval"`

          The message type to be created.

          - `"approval"`

      - `ToolReturn`

        - `status: "success" | "error"`

        - `tool_call_id: string`

        - `tool_return: Array<TextContent | ImageContent> | string`

          The tool return value - either a string or list of content parts (text/image)

        - `stderr?: Array<string> | null`

        - `stdout?: Array<string> | null`

        - `type?: "tool"`

          The message type to be created.

    - `approve?: boolean | null`

      Whether the tool has been approved

    - `is_err?: boolean | null`

    - `message_type?: "approval_response_message"`

      The type of the message.

      - `"approval_response_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `reason?: string | null`

      An optional explanation for the provided approval status

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `SummaryMessage`

    A message representing a summary of the conversation. Sent to the LLM as a user or system message depending on the provider.

    - `id: string`

    - `date: string`

    - `summary: string`

    - `compaction_stats?: CompactionStats | null`

      Statistics about a memory compaction operation.

      - `context_window: number`

        The model's context window size

      - `messages_count_after: number`

        Number of messages after compaction

      - `messages_count_before: number`

        Number of messages before compaction

      - `trigger: string`

        What triggered the compaction (e.g., 'context_window_exceeded', 'post_step_context_check')

      - `context_tokens_after?: number | null`

        Token count after compaction (message tokens only, does not include tool definitions)

      - `context_tokens_before?: number | null`

        Token count before compaction (from LLM usage stats, includes full context sent to LLM)

    - `is_err?: boolean | null`

    - `message_type?: "summary_message"`

      - `"summary_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `EventMessage`

    A message for notifying the developer that an event that has occured (e.g. a compaction). Events are NOT part of the context window.

    - `id: string`

    - `date: string`

    - `event_data: Record<string, unknown>`

    - `event_type: "compaction"`

      - `"compaction"`

    - `is_err?: boolean | null`

    - `message_type?: "event_message"`

      - `"event_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

// Automatically fetches more pages as needed.
for await (const message of client.agents.messages.list(
  'agent-123e4567-e89b-42d3-8456-426614174000',
)) {
  console.log(message);
}
```

#### Response

```json
[
  {
    "id": "id",
    "content": "content",
    "date": "2019-12-27T18:11:19.117Z",
    "is_err": true,
    "message_type": "system_message",
    "name": "name",
    "otid": "otid",
    "run_id": "run_id",
    "sender_id": "sender_id",
    "seq_id": 0,
    "step_id": "step_id"
  }
]
```

## Create Message

`client.agents.messages.create(stringagentID, MessageCreateParamsbody, RequestOptionsoptions?): LettaResponse | Stream<LettaStreamingResponse>`

**post** `/v1/agents/{agent_id}/messages`

Process a user message and return the agent's response.
This endpoint accepts a message from a user and processes it through the agent.

**Note:** Sending multiple concurrent requests to the same agent can lead to undefined behavior.
Each agent processes messages sequentially, and concurrent requests may interleave in unexpected ways.
Wait for each request to complete before sending the next one. Use separate agents or conversations for parallel processing.

The response format is controlled by the `streaming` field in the request body:

- If `streaming=false` (default): Returns a complete LettaResponse with all messages
- If `streaming=true`: Returns a Server-Sent Events (SSE) stream

Additional streaming options (only used when streaming=true):

- `stream_tokens`: Stream individual tokens instead of complete steps
- `include_pings`: Include keepalive pings to prevent connection timeouts
- `background`: Process the request in the background

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `MessageCreateParams = MessageCreateParamsNonStreaming | MessageCreateParamsStreaming`

  - `MessageCreateParamsBase`

    - `assistant_message_tool_kwarg?: string`

      The name of the message argument in the designated message tool. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

    - `assistant_message_tool_name?: string`

      The name of the designated message tool. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

    - `background?: boolean`

      Whether to process the request in the background (only used when streaming=true).

    - `client_skills?: Array<ClientSkill> | null`

      Client-side skills available in the environment. These are rendered in the system prompt's available skills section alongside agent-scoped skills from MemFS.

      - `description: string`

        Description of what the skill does

      - `location: string`

        Path or location hint for the skill (e.g. skills/my-skill/SKILL.md)

      - `name: string`

        The name of the skill

    - `client_tools?: Array<ClientTool> | null`

      Client-side tools that the agent can call. When the agent calls a client-side tool, execution pauses and returns control to the client to execute the tool and provide the result via a ToolReturn.

      - `name: string`

        The name of the tool function

      - `description?: string | null`

        Description of what the tool does

      - `parameters?: Record<string, unknown> | null`

        JSON Schema for the function parameters

    - `enable_thinking?: string`

      If set to True, enables reasoning before responses or tool calls from the agent.

    - `include_compaction_messages?: boolean`

      If True, compaction events emit structured `SummaryMessage` and `EventMessage` types. If False (default), compaction messages are not included in the response.

    - `include_pings?: boolean`

      Whether to include periodic keepalive ping messages in the stream to prevent connection timeouts (only used when streaming=true).

    - `include_return_message_types?: Array<MessageType> | null`

      Only return specified message types in the response. If `None` (default) returns all messages.

      - `"system_message"`

      - `"user_message"`

      - `"assistant_message"`

      - `"reasoning_message"`

      - `"hidden_reasoning_message"`

      - `"tool_call_message"`

      - `"tool_return_message"`

      - `"approval_request_message"`

      - `"approval_response_message"`

      - `"summary_message"`

      - `"event_message"`

    - `input?: string | Array<TextContent | ImageContent | ToolCallContent | 5 more> | null`

      Syntactic sugar for a single user message. Equivalent to messages=[{'role': 'user', 'content': input}].

      - `string`

      - `Array<TextContent | ImageContent | ToolCallContent | 5 more>`

        - `TextContent`

          - `text: string`

            The text content of the message.

          - `signature?: string | null`

            Stores a unique identifier for any reasoning associated with this text content.

          - `type?: "text"`

            The type of the message.

            - `"text"`

        - `ImageContent`

          - `source: URLImage | Base64Image | LettaImage`

            The source of the image.

            - `URLImage`

              - `url: string`

                The URL of the image.

              - `type?: "url"`

                The source type for the image.

                - `"url"`

            - `Base64Image`

              - `data: string`

                The base64 encoded image data.

              - `media_type: string`

                The media type for the image.

              - `detail?: string | null`

                What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

              - `type?: "base64"`

                The source type for the image.

                - `"base64"`

            - `LettaImage`

              - `file_id: string`

                The unique identifier of the image file persisted in storage.

              - `data?: string | null`

                The base64 encoded image data.

              - `detail?: string | null`

                What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

              - `media_type?: string | null`

                The media type for the image.

              - `type?: "letta"`

                The source type for the image.

                - `"letta"`

          - `type?: "image"`

            The type of the message.

            - `"image"`

        - `ToolCallContent`

          - `id: string`

            A unique identifier for this specific tool call instance.

          - `input: Record<string, unknown>`

            The parameters being passed to the tool, structured as a dictionary of parameter names to values.

          - `name: string`

            The name of the tool being called.

          - `signature?: string | null`

            Stores a unique identifier for any reasoning associated with this tool call.

          - `type?: "tool_call"`

            Indicates this content represents a tool call event.

            - `"tool_call"`

        - `ToolReturnContent`

          - `content: string`

            The content returned by the tool execution.

          - `is_error: boolean`

            Indicates whether the tool execution resulted in an error.

          - `tool_call_id: string`

            References the ID of the ToolCallContent that initiated this tool call.

          - `type?: "tool_return"`

            Indicates this content represents a tool return event.

            - `"tool_return"`

        - `ReasoningContent`

          Sent via the Anthropic Messages API

          - `is_native: boolean`

            Whether the reasoning content was generated by a reasoner model that processed this step.

          - `reasoning: string`

            The intermediate reasoning or thought process content.

          - `signature?: string | null`

            A unique identifier for this reasoning step.

          - `type?: "reasoning"`

            Indicates this is a reasoning/intermediate step.

            - `"reasoning"`

        - `RedactedReasoningContent`

          Sent via the Anthropic Messages API

          - `data: string`

            The redacted or filtered intermediate reasoning content.

          - `type?: "redacted_reasoning"`

            Indicates this is a redacted thinking step.

            - `"redacted_reasoning"`

        - `OmittedReasoningContent`

          A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

          - `signature?: string | null`

            A unique identifier for this reasoning step.

          - `type?: "omitted_reasoning"`

            Indicates this is an omitted reasoning step.

            - `"omitted_reasoning"`

        - `SummarizedReasoningContent`

          The style of reasoning content returned by the OpenAI Responses API

          - `id: string`

            The unique identifier for this reasoning step.

          - `summary: Array<Summary>`

            Summaries of the reasoning content.

            - `index: number`

              The index of the summary part.

            - `text: string`

              The text of the summary part.

          - `encrypted_content?: string`

            The encrypted reasoning content.

          - `type?: "summarized_reasoning"`

            Indicates this is a summarized reasoning step.

            - `"summarized_reasoning"`

    - `max_steps?: number`

      Maximum number of steps the agent should take to process the request.

    - `messages?: Array<MessageCreate | ApprovalCreate | ToolReturnCreate> | null`

      The messages to be sent to the agent.

      - `MessageCreate`

        Request to create a message

        - `content: Array<LettaMessageContentUnion> | string`

          The content of the message.

          - `Array<LettaMessageContentUnion>`

            - `TextContent`

            - `ImageContent`

            - `ToolCallContent`

            - `ToolReturnContent`

            - `ReasoningContent`

              Sent via the Anthropic Messages API

            - `RedactedReasoningContent`

              Sent via the Anthropic Messages API

            - `OmittedReasoningContent`

              A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

          - `string`

        - `role: "user" | "system" | "assistant"`

          The role of the participant.

          - `"user"`

          - `"system"`

          - `"assistant"`

        - `batch_item_id?: string | null`

          The id of the LLMBatchItem that this message is associated with

        - `group_id?: string | null`

          The multi-agent group that the message was sent in

        - `name?: string | null`

          The name of the participant.

        - `otid?: string | null`

          The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

        - `sender_id?: string | null`

          The id of the sender of the message, can be an identity id or agent id

        - `type?: "message" | null`

          The message type to be created.

          - `"message"`

      - `ApprovalCreate`

        Input to approve or deny a tool call request

        - `approval_request_id?: string | null`

          The message ID of the approval request

        - `approvals?: Array<ApprovalReturn | ToolReturn> | null`

          The list of approval responses

          - `ApprovalReturn`

            - `approve: boolean`

              Whether the tool has been approved

            - `tool_call_id: string`

              The ID of the tool call that corresponds to this approval

            - `reason?: string | null`

              An optional explanation for the provided approval status

            - `type?: "approval"`

              The message type to be created.

              - `"approval"`

          - `ToolReturn`

            - `status: "success" | "error"`

              - `"success"`

              - `"error"`

            - `tool_call_id: string`

            - `tool_return: Array<TextContent | ImageContent> | string`

              The tool return value - either a string or list of content parts (text/image)

              - `Array<TextContent | ImageContent>`

                - `TextContent`

                - `ImageContent`

              - `string`

            - `stderr?: Array<string> | null`

            - `stdout?: Array<string> | null`

            - `type?: "tool"`

              The message type to be created.

              - `"tool"`

        - `approve?: boolean | null`

          Whether the tool has been approved

        - `group_id?: string | null`

          The multi-agent group that the message was sent in

        - `otid?: string | null`

          The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

        - `reason?: string | null`

          An optional explanation for the provided approval status

        - `type?: "approval"`

          The message type to be created.

          - `"approval"`

      - `ToolReturnCreate`

        Submit tool return(s) from client-side tool execution.

        This is the preferred way to send tool results back to the agent after
        client-side tool execution. It is equivalent to sending an ApprovalCreate
        with tool return approvals, but provides a cleaner API for the common case.

        - `tool_returns: Array<ToolReturn>`

          List of tool returns from client-side execution

          - `status: "success" | "error"`

          - `tool_call_id: string`

          - `tool_return: Array<TextContent | ImageContent> | string`

            The tool return value - either a string or list of content parts (text/image)

          - `stderr?: Array<string> | null`

          - `stdout?: Array<string> | null`

          - `type?: "tool"`

            The message type to be created.

        - `group_id?: string | null`

          The multi-agent group that the message was sent in

        - `otid?: string | null`

          The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

        - `type?: "tool_return"`

          The message type to be created.

          - `"tool_return"`

    - `override_model?: string | null`

      Model handle to use for this request instead of the agent's default model. This allows sending a message to a different model without changing the agent's configuration.

    - `override_system?: string | null`

      Optional per-request system prompt override. When set, this is passed directly to the underlying LLM request and bypasses the persisted/compiled system message for that request.

    - `return_logprobs?: boolean`

      If True, returns log probabilities of the output tokens in the response. Useful for RL training. Only supported for OpenAI-compatible providers (including SGLang).

    - `return_token_ids?: boolean`

      If True, returns token IDs and logprobs for ALL LLM generations in the agent step, not just the last one. Uses SGLang native /generate endpoint. Returns 'turns' field with TurnTokenData for each assistant/tool turn. Required for proper multi-turn RL training with loss masking.

    - `stream_tokens?: boolean`

      Flag to determine if individual tokens should be streamed, rather than streaming per step (only used when streaming=true).

    - `streaming?: false`

      If True, returns a streaming response (Server-Sent Events). If False (default), returns a complete response.

      - `false`

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `use_assistant_message?: boolean`

      Whether the server should parse specific tool call arguments (default `send_message`) as `AssistantMessage` objects. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

  - `MessageCreateParamsNonStreaming extends MessageCreateParamsBase`

    - `streaming?: false`

      If True, returns a streaming response (Server-Sent Events). If False (default), returns a complete response.

  - `MessageCreateParamsStreaming extends MessageCreateParamsBase`

    - `streaming: true`

      If True, returns a streaming response (Server-Sent Events). If False (default), returns a complete response.

      - `true`

### Returns

- `LettaResponse`

  Response object from an agent interaction, consisting of the new messages generated by the agent and usage statistics.
  The type of the returned messages can be either `Message` or `LettaMessage`, depending on what was specified in the request.

  Attributes:
  messages (List[Union[Message, LettaMessage]]): The messages returned by the agent.
  usage (LettaUsageStatistics): The usage statistics

  - `messages: Array<Message>`

    The messages returned by the agent.

    - `SystemMessage`

      A message generated by the system. Never streamed back on a response, only used for cursor pagination.

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      content (str): The message content sent by the system

      - `id: string`

      - `content: string`

        The message content sent by the system

      - `date: string`

      - `is_err?: boolean | null`

      - `message_type?: "system_message"`

        The type of the message.

        - `"system_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

    - `UserMessage`

      A message sent by the user. Never streamed back on a response, only used for cursor pagination.

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      content (Union[str, List[LettaUserMessageContentUnion]]): The message content sent by the user (can be a string or an array of multi-modal content parts)

      - `id: string`

      - `content: Array<LettaUserMessageContentUnion> | string`

        The message content sent by the user (can be a string or an array of multi-modal content parts)

        - `Array<LettaUserMessageContentUnion>`

          - `TextContent`

            - `text: string`

              The text content of the message.

            - `signature?: string | null`

              Stores a unique identifier for any reasoning associated with this text content.

            - `type?: "text"`

              The type of the message.

              - `"text"`

          - `ImageContent`

            - `source: URLImage | Base64Image | LettaImage`

              The source of the image.

              - `URLImage`

                - `url: string`

                  The URL of the image.

                - `type?: "url"`

                  The source type for the image.

                  - `"url"`

              - `Base64Image`

                - `data: string`

                  The base64 encoded image data.

                - `media_type: string`

                  The media type for the image.

                - `detail?: string | null`

                  What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

                - `type?: "base64"`

                  The source type for the image.

                  - `"base64"`

              - `LettaImage`

                - `file_id: string`

                  The unique identifier of the image file persisted in storage.

                - `data?: string | null`

                  The base64 encoded image data.

                - `detail?: string | null`

                  What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

                - `media_type?: string | null`

                  The media type for the image.

                - `type?: "letta"`

                  The source type for the image.

                  - `"letta"`

            - `type?: "image"`

              The type of the message.

              - `"image"`

        - `string`

      - `date: string`

      - `is_err?: boolean | null`

      - `message_type?: "user_message"`

        The type of the message.

        - `"user_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

    - `ReasoningMessage`

      Representation of an agent's internal reasoning.

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      source (Literal["reasoner_model", "non_reasoner_model"]): Whether the reasoning
      content was generated natively by a reasoner model or derived via prompting
      reasoning (str): The internal reasoning of the agent
      signature (Optional[str]): The model-generated signature of the reasoning step

      - `id: string`

      - `date: string`

      - `reasoning: string`

      - `is_err?: boolean | null`

      - `message_type?: "reasoning_message"`

        The type of the message.

        - `"reasoning_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `signature?: string | null`

      - `source?: "reasoner_model" | "non_reasoner_model"`

        - `"reasoner_model"`

        - `"non_reasoner_model"`

      - `step_id?: string | null`

    - `HiddenReasoningMessage`

      Representation of an agent's internal reasoning where reasoning content
      has been hidden from the response.

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      state (Literal["redacted", "omitted"]): Whether the reasoning
      content was redacted by the provider or simply omitted by the API
      hidden_reasoning (Optional[str]): The internal reasoning of the agent

      - `id: string`

      - `date: string`

      - `state: "redacted" | "omitted"`

        - `"redacted"`

        - `"omitted"`

      - `hidden_reasoning?: string | null`

      - `is_err?: boolean | null`

      - `message_type?: "hidden_reasoning_message"`

        The type of the message.

        - `"hidden_reasoning_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

    - `ToolCallMessage`

      A message representing a request to call a tool (generated by the LLM to trigger tool execution).

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      tool_call (Union[ToolCall, ToolCallDelta]): The tool call

      - `id: string`

      - `date: string`

      - `tool_call: ToolCall | ToolCallDelta`

        - `ToolCall`

          - `arguments: string`

          - `name: string`

          - `tool_call_id: string`

        - `ToolCallDelta`

          - `arguments?: string | null`

          - `name?: string | null`

          - `tool_call_id?: string | null`

      - `is_err?: boolean | null`

      - `message_type?: "tool_call_message"`

        The type of the message.

        - `"tool_call_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

      - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

        - `Array<ToolCall>`

          - `arguments: string`

          - `name: string`

          - `tool_call_id: string`

        - `ToolCallDelta`

    - `ToolReturnMessage`

      A message representing the return value of a tool call (generated by Letta executing the requested tool).

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      tool_return (str): The return value of the tool (deprecated, use tool_returns)
      status (Literal["success", "error"]): The status of the tool call (deprecated, use tool_returns)
      tool_call_id (str): A unique identifier for the tool call that generated this message (deprecated, use tool_returns)
      stdout (Optional[List(str)]): Captured stdout (e.g. prints, logs) from the tool invocation (deprecated, use tool_returns)
      stderr (Optional[List(str)]): Captured stderr from the tool invocation (deprecated, use tool_returns)
      tool_returns (Optional[List[ToolReturn]]): List of tool returns for multi-tool support

      - `id: string`

      - `date: string`

      - `status: "success" | "error"`

        - `"success"`

        - `"error"`

      - `tool_call_id: string`

      - `tool_return: string`

      - `is_err?: boolean | null`

      - `message_type?: "tool_return_message"`

        The type of the message.

        - `"tool_return_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `stderr?: Array<string> | null`

      - `stdout?: Array<string> | null`

      - `step_id?: string | null`

      - `tool_returns?: Array<ToolReturn> | null`

        - `status: "success" | "error"`

          - `"success"`

          - `"error"`

        - `tool_call_id: string`

        - `tool_return: Array<TextContent | ImageContent> | string`

          The tool return value - either a string or list of content parts (text/image)

          - `Array<TextContent | ImageContent>`

            - `TextContent`

            - `ImageContent`

          - `string`

        - `stderr?: Array<string> | null`

        - `stdout?: Array<string> | null`

        - `type?: "tool"`

          The message type to be created.

          - `"tool"`

    - `AssistantMessage`

      A message sent by the LLM in response to user input. Used in the LLM context.

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      content (Union[str, List[LettaAssistantMessageContentUnion]]): The message content sent by the agent (can be a string or an array of content parts)

      - `id: string`

      - `content: Array<LettaAssistantMessageContentUnion> | string`

        The message content sent by the agent (can be a string or an array of content parts)

        - `Array<LettaAssistantMessageContentUnion>`

          - `text: string`

            The text content of the message.

          - `signature?: string | null`

            Stores a unique identifier for any reasoning associated with this text content.

          - `type?: "text"`

            The type of the message.

            - `"text"`

        - `string`

      - `date: string`

      - `is_err?: boolean | null`

      - `message_type?: "assistant_message"`

        The type of the message.

        - `"assistant_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

    - `ApprovalRequestMessage`

      A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      tool_call (ToolCall): The tool call

      - `id: string`

      - `date: string`

      - `tool_call: ToolCall | ToolCallDelta`

        The tool call that has been requested by the llm to run

        - `ToolCall`

        - `ToolCallDelta`

      - `is_err?: boolean | null`

      - `message_type?: "approval_request_message"`

        The type of the message.

        - `"approval_request_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

      - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

        The tool calls that have been requested by the llm to run, which are pending approval

        - `Array<ToolCall>`

          - `arguments: string`

          - `name: string`

          - `tool_call_id: string`

        - `ToolCallDelta`

    - `ApprovalResponseMessage`

      A message representing a response form the user indicating whether a tool has been approved to run.

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      approve: (bool) Whether the tool has been approved
      approval_request_id: The ID of the approval request
      reason: (Optional[str]) An optional explanation for the provided approval status

      - `id: string`

      - `date: string`

      - `approval_request_id?: string | null`

        The message ID of the approval request

      - `approvals?: Array<ApprovalReturn | ToolReturn> | null`

        The list of approval responses

        - `ApprovalReturn`

          - `approve: boolean`

            Whether the tool has been approved

          - `tool_call_id: string`

            The ID of the tool call that corresponds to this approval

          - `reason?: string | null`

            An optional explanation for the provided approval status

          - `type?: "approval"`

            The message type to be created.

            - `"approval"`

        - `ToolReturn`

          - `status: "success" | "error"`

          - `tool_call_id: string`

          - `tool_return: Array<TextContent | ImageContent> | string`

            The tool return value - either a string or list of content parts (text/image)

          - `stderr?: Array<string> | null`

          - `stdout?: Array<string> | null`

          - `type?: "tool"`

            The message type to be created.

      - `approve?: boolean | null`

        Whether the tool has been approved

      - `is_err?: boolean | null`

      - `message_type?: "approval_response_message"`

        The type of the message.

        - `"approval_response_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `reason?: string | null`

        An optional explanation for the provided approval status

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

    - `SummaryMessage`

      A message representing a summary of the conversation. Sent to the LLM as a user or system message depending on the provider.

      - `id: string`

      - `date: string`

      - `summary: string`

      - `compaction_stats?: CompactionStats | null`

        Statistics about a memory compaction operation.

        - `context_window: number`

          The model's context window size

        - `messages_count_after: number`

          Number of messages after compaction

        - `messages_count_before: number`

          Number of messages before compaction

        - `trigger: string`

          What triggered the compaction (e.g., 'context_window_exceeded', 'post_step_context_check')

        - `context_tokens_after?: number | null`

          Token count after compaction (message tokens only, does not include tool definitions)

        - `context_tokens_before?: number | null`

          Token count before compaction (from LLM usage stats, includes full context sent to LLM)

      - `is_err?: boolean | null`

      - `message_type?: "summary_message"`

        - `"summary_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

    - `EventMessage`

      A message for notifying the developer that an event that has occured (e.g. a compaction). Events are NOT part of the context window.

      - `id: string`

      - `date: string`

      - `event_data: Record<string, unknown>`

      - `event_type: "compaction"`

        - `"compaction"`

      - `is_err?: boolean | null`

      - `message_type?: "event_message"`

        - `"event_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

  - `stop_reason: StopReason`

    The stop reason from Letta indicating why agent loop stopped execution.

    - `stop_reason: StopReasonType`

      The reason why execution stopped.

      - `"end_turn"`

      - `"error"`

      - `"llm_api_error"`

      - `"invalid_llm_response"`

      - `"invalid_tool_call"`

      - `"max_steps"`

      - `"max_tokens_exceeded"`

      - `"no_tool_call"`

      - `"tool_rule"`

      - `"cancelled"`

      - `"insufficient_credits"`

      - `"requires_approval"`

      - `"context_window_overflow_in_system_prompt"`

    - `message_type?: "stop_reason"`

      The type of the message.

      - `"stop_reason"`

  - `usage: Usage`

    The usage statistics of the agent.

    - `cache_write_tokens?: number | null`

      The number of input tokens written to cache (Anthropic only). None if not reported by provider.

    - `cached_input_tokens?: number | null`

      The number of input tokens served from cache. None if not reported by provider.

    - `completion_tokens?: number`

      The number of tokens generated by the agent.

    - `context_tokens?: number | null`

      Estimate of tokens currently in the context window.

    - `message_type?: "usage_statistics"`

      - `"usage_statistics"`

    - `prompt_tokens?: number`

      The number of tokens in the prompt.

    - `reasoning_tokens?: number | null`

      The number of reasoning/thinking tokens generated. None if not reported by provider.

    - `run_ids?: Array<string> | null`

      The background task run IDs associated with the agent interaction

    - `step_count?: number`

      The number of steps taken by the agent.

    - `total_tokens?: number`

      The total number of tokens processed by the agent.

  - `logprobs?: Logprobs | null`

    Log probabilities of the output tokens from the last LLM call. Only present if return_logprobs was enabled.

    - `content?: Array<Content> | null`

      - `token: string`

      - `logprob: number`

      - `top_logprobs: Array<TopLogprob>`

        - `token: string`

        - `logprob: number`

        - `bytes?: Array<number> | null`

      - `bytes?: Array<number> | null`

    - `refusal?: Array<Refusal> | null`

      - `token: string`

      - `logprob: number`

      - `top_logprobs: Array<TopLogprob>`

        - `token: string`

        - `logprob: number`

        - `bytes?: Array<number> | null`

      - `bytes?: Array<number> | null`

  - `turns?: Array<Turn> | null`

    Token data for all LLM generations in multi-turn agent interaction. Includes token IDs and logprobs for each assistant turn, plus tool result content. Only present if return_token_ids was enabled. Used for RL training with loss masking.

    - `role: "assistant" | "tool"`

      Role of this turn: 'assistant' for LLM generations (trainable), 'tool' for tool results (non-trainable).

      - `"assistant"`

      - `"tool"`

    - `content?: string | null`

      Text content. For tool turns, client tokenizes this with loss_mask=0.

    - `output_ids?: Array<number> | null`

      Token IDs from SGLang native endpoint. Only present for assistant turns.

    - `output_token_logprobs?: Array<Array<unknown>> | null`

      Logprobs from SGLang: [[logprob, token_id, top_logprob_or_null], ...]. Only present for assistant turns.

    - `tool_name?: string | null`

      Name of the tool called. Only present for tool turns.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const lettaResponse = await client.agents.messages.create(
  'agent-123e4567-e89b-42d3-8456-426614174000',
);

console.log(lettaResponse.messages);
```

#### Response

```json
{
  "messages": [
    {
      "id": "id",
      "content": "content",
      "date": "2019-12-27T18:11:19.117Z",
      "is_err": true,
      "message_type": "system_message",
      "name": "name",
      "otid": "otid",
      "run_id": "run_id",
      "sender_id": "sender_id",
      "seq_id": 0,
      "step_id": "step_id"
    }
  ],
  "stop_reason": {
    "stop_reason": "end_turn",
    "message_type": "stop_reason"
  },
  "usage": {
    "cache_write_tokens": 0,
    "cached_input_tokens": 0,
    "completion_tokens": 0,
    "context_tokens": 0,
    "message_type": "usage_statistics",
    "prompt_tokens": 0,
    "reasoning_tokens": 0,
    "run_ids": [
      "string"
    ],
    "step_count": 0,
    "total_tokens": 0
  },
  "logprobs": {
    "content": [
      {
        "token": "token",
        "logprob": 0,
        "top_logprobs": [
          {
            "token": "token",
            "logprob": 0,
            "bytes": [
              0
            ]
          }
        ],
        "bytes": [
          0
        ]
      }
    ],
    "refusal": [
      {
        "token": "token",
        "logprob": 0,
        "top_logprobs": [
          {
            "token": "token",
            "logprob": 0,
            "bytes": [
              0
            ]
          }
        ],
        "bytes": [
          0
        ]
      }
    ]
  },
  "turns": [
    {
      "role": "assistant",
      "content": "content",
      "output_ids": [
        0
      ],
      "output_token_logprobs": [
        [
          {}
        ]
      ],
      "tool_name": "tool_name"
    }
  ]
}
```

## Create Message Streaming

`client.agents.messages.stream(stringagentID, MessageStreamParamsbody, RequestOptionsoptions?): LettaStreamingResponse | Stream<LettaStreamingResponse>`

**post** `/v1/agents/{agent_id}/messages/stream`

Process a user message and return the agent's response.

Deprecated: Use the `POST /{agent_id}/messages` endpoint with `streaming=true` in the request body instead.

**Note:** Sending multiple concurrent requests to the same agent can lead to undefined behavior.
Each agent processes messages sequentially, and concurrent requests may interleave in unexpected ways.
Wait for each request to complete before sending the next one. Use separate agents or conversations for parallel processing.

This endpoint accepts a message from a user and processes it through the agent.
It will stream the steps of the response always, and stream the tokens if 'stream_tokens' is set to True.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `body: MessageStreamParams`

  - `assistant_message_tool_kwarg?: string`

    The name of the message argument in the designated message tool. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

  - `assistant_message_tool_name?: string`

    The name of the designated message tool. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

  - `background?: boolean`

    Whether to process the request in the background (only used when streaming=true).

  - `client_skills?: Array<ClientSkill> | null`

    Client-side skills available in the environment. These are rendered in the system prompt's available skills section alongside agent-scoped skills from MemFS.

    - `description: string`

      Description of what the skill does

    - `location: string`

      Path or location hint for the skill (e.g. skills/my-skill/SKILL.md)

    - `name: string`

      The name of the skill

  - `client_tools?: Array<ClientTool> | null`

    Client-side tools that the agent can call. When the agent calls a client-side tool, execution pauses and returns control to the client to execute the tool and provide the result via a ToolReturn.

    - `name: string`

      The name of the tool function

    - `description?: string | null`

      Description of what the tool does

    - `parameters?: Record<string, unknown> | null`

      JSON Schema for the function parameters

  - `enable_thinking?: string`

    If set to True, enables reasoning before responses or tool calls from the agent.

  - `include_compaction_messages?: boolean`

    If True, compaction events emit structured `SummaryMessage` and `EventMessage` types. If False (default), compaction messages are not included in the response.

  - `include_pings?: boolean`

    Whether to include periodic keepalive ping messages in the stream to prevent connection timeouts (only used when streaming=true).

  - `include_return_message_types?: Array<MessageType> | null`

    Only return specified message types in the response. If `None` (default) returns all messages.

    - `"system_message"`

    - `"user_message"`

    - `"assistant_message"`

    - `"reasoning_message"`

    - `"hidden_reasoning_message"`

    - `"tool_call_message"`

    - `"tool_return_message"`

    - `"approval_request_message"`

    - `"approval_response_message"`

    - `"summary_message"`

    - `"event_message"`

  - `input?: string | Array<TextContent | ImageContent | ToolCallContent | 5 more> | null`

    Syntactic sugar for a single user message. Equivalent to messages=[{'role': 'user', 'content': input}].

    - `string`

    - `Array<TextContent | ImageContent | ToolCallContent | 5 more>`

      - `TextContent`

        - `text: string`

          The text content of the message.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this text content.

        - `type?: "text"`

          The type of the message.

          - `"text"`

      - `ImageContent`

        - `source: URLImage | Base64Image | LettaImage`

          The source of the image.

          - `URLImage`

            - `url: string`

              The URL of the image.

            - `type?: "url"`

              The source type for the image.

              - `"url"`

          - `Base64Image`

            - `data: string`

              The base64 encoded image data.

            - `media_type: string`

              The media type for the image.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `type?: "base64"`

              The source type for the image.

              - `"base64"`

          - `LettaImage`

            - `file_id: string`

              The unique identifier of the image file persisted in storage.

            - `data?: string | null`

              The base64 encoded image data.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `media_type?: string | null`

              The media type for the image.

            - `type?: "letta"`

              The source type for the image.

              - `"letta"`

        - `type?: "image"`

          The type of the message.

          - `"image"`

      - `ToolCallContent`

        - `id: string`

          A unique identifier for this specific tool call instance.

        - `input: Record<string, unknown>`

          The parameters being passed to the tool, structured as a dictionary of parameter names to values.

        - `name: string`

          The name of the tool being called.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this tool call.

        - `type?: "tool_call"`

          Indicates this content represents a tool call event.

          - `"tool_call"`

      - `ToolReturnContent`

        - `content: string`

          The content returned by the tool execution.

        - `is_error: boolean`

          Indicates whether the tool execution resulted in an error.

        - `tool_call_id: string`

          References the ID of the ToolCallContent that initiated this tool call.

        - `type?: "tool_return"`

          Indicates this content represents a tool return event.

          - `"tool_return"`

      - `ReasoningContent`

        Sent via the Anthropic Messages API

        - `is_native: boolean`

          Whether the reasoning content was generated by a reasoner model that processed this step.

        - `reasoning: string`

          The intermediate reasoning or thought process content.

        - `signature?: string | null`

          A unique identifier for this reasoning step.

        - `type?: "reasoning"`

          Indicates this is a reasoning/intermediate step.

          - `"reasoning"`

      - `RedactedReasoningContent`

        Sent via the Anthropic Messages API

        - `data: string`

          The redacted or filtered intermediate reasoning content.

        - `type?: "redacted_reasoning"`

          Indicates this is a redacted thinking step.

          - `"redacted_reasoning"`

      - `OmittedReasoningContent`

        A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

        - `signature?: string | null`

          A unique identifier for this reasoning step.

        - `type?: "omitted_reasoning"`

          Indicates this is an omitted reasoning step.

          - `"omitted_reasoning"`

      - `SummarizedReasoningContent`

        The style of reasoning content returned by the OpenAI Responses API

        - `id: string`

          The unique identifier for this reasoning step.

        - `summary: Array<Summary>`

          Summaries of the reasoning content.

          - `index: number`

            The index of the summary part.

          - `text: string`

            The text of the summary part.

        - `encrypted_content?: string`

          The encrypted reasoning content.

        - `type?: "summarized_reasoning"`

          Indicates this is a summarized reasoning step.

          - `"summarized_reasoning"`

  - `max_steps?: number`

    Maximum number of steps the agent should take to process the request.

  - `messages?: Array<MessageCreate | ApprovalCreate | ToolReturnCreate> | null`

    The messages to be sent to the agent.

    - `MessageCreate`

      Request to create a message

      - `content: Array<LettaMessageContentUnion> | string`

        The content of the message.

        - `Array<LettaMessageContentUnion>`

          - `TextContent`

          - `ImageContent`

          - `ToolCallContent`

          - `ToolReturnContent`

          - `ReasoningContent`

            Sent via the Anthropic Messages API

          - `RedactedReasoningContent`

            Sent via the Anthropic Messages API

          - `OmittedReasoningContent`

            A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

        - `string`

      - `role: "user" | "system" | "assistant"`

        The role of the participant.

        - `"user"`

        - `"system"`

        - `"assistant"`

      - `batch_item_id?: string | null`

        The id of the LLMBatchItem that this message is associated with

      - `group_id?: string | null`

        The multi-agent group that the message was sent in

      - `name?: string | null`

        The name of the participant.

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `sender_id?: string | null`

        The id of the sender of the message, can be an identity id or agent id

      - `type?: "message" | null`

        The message type to be created.

        - `"message"`

    - `ApprovalCreate`

      Input to approve or deny a tool call request

      - `approval_request_id?: string | null`

        The message ID of the approval request

      - `approvals?: Array<ApprovalReturn | ToolReturn> | null`

        The list of approval responses

        - `ApprovalReturn`

          - `approve: boolean`

            Whether the tool has been approved

          - `tool_call_id: string`

            The ID of the tool call that corresponds to this approval

          - `reason?: string | null`

            An optional explanation for the provided approval status

          - `type?: "approval"`

            The message type to be created.

            - `"approval"`

        - `ToolReturn`

          - `status: "success" | "error"`

            - `"success"`

            - `"error"`

          - `tool_call_id: string`

          - `tool_return: Array<TextContent | ImageContent> | string`

            The tool return value - either a string or list of content parts (text/image)

            - `Array<TextContent | ImageContent>`

              - `TextContent`

              - `ImageContent`

            - `string`

          - `stderr?: Array<string> | null`

          - `stdout?: Array<string> | null`

          - `type?: "tool"`

            The message type to be created.

            - `"tool"`

      - `approve?: boolean | null`

        Whether the tool has been approved

      - `group_id?: string | null`

        The multi-agent group that the message was sent in

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `reason?: string | null`

        An optional explanation for the provided approval status

      - `type?: "approval"`

        The message type to be created.

        - `"approval"`

    - `ToolReturnCreate`

      Submit tool return(s) from client-side tool execution.

      This is the preferred way to send tool results back to the agent after
      client-side tool execution. It is equivalent to sending an ApprovalCreate
      with tool return approvals, but provides a cleaner API for the common case.

      - `tool_returns: Array<ToolReturn>`

        List of tool returns from client-side execution

        - `status: "success" | "error"`

        - `tool_call_id: string`

        - `tool_return: Array<TextContent | ImageContent> | string`

          The tool return value - either a string or list of content parts (text/image)

        - `stderr?: Array<string> | null`

        - `stdout?: Array<string> | null`

        - `type?: "tool"`

          The message type to be created.

      - `group_id?: string | null`

        The multi-agent group that the message was sent in

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `type?: "tool_return"`

        The message type to be created.

        - `"tool_return"`

  - `override_model?: string | null`

    Model handle to use for this request instead of the agent's default model. This allows sending a message to a different model without changing the agent's configuration.

  - `override_system?: string | null`

    Optional per-request system prompt override. When set, this is passed directly to the underlying LLM request and bypasses the persisted/compiled system message for that request.

  - `return_logprobs?: boolean`

    If True, returns log probabilities of the output tokens in the response. Useful for RL training. Only supported for OpenAI-compatible providers (including SGLang).

  - `return_token_ids?: boolean`

    If True, returns token IDs and logprobs for ALL LLM generations in the agent step, not just the last one. Uses SGLang native /generate endpoint. Returns 'turns' field with TurnTokenData for each assistant/tool turn. Required for proper multi-turn RL training with loss masking.

  - `stream_tokens?: boolean`

    Flag to determine if individual tokens should be streamed, rather than streaming per step (only used when streaming=true).

  - `streaming?: boolean`

    If True, returns a streaming response (Server-Sent Events). If False (default), returns a complete response.

  - `top_logprobs?: number | null`

    Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

  - `use_assistant_message?: boolean`

    Whether the server should parse specific tool call arguments (default `send_message`) as `AssistantMessage` objects. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

### Returns

- `LettaStreamingResponse = SystemMessage | UserMessage | ReasoningMessage | 10 more`

  Streaming response type for Server-Sent Events (SSE) endpoints.
  Each event in the stream will be one of these types.

  - `SystemMessage`

    A message generated by the system. Never streamed back on a response, only used for cursor pagination.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    content (str): The message content sent by the system

    - `id: string`

    - `content: string`

      The message content sent by the system

    - `date: string`

    - `is_err?: boolean | null`

    - `message_type?: "system_message"`

      The type of the message.

      - `"system_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `UserMessage`

    A message sent by the user. Never streamed back on a response, only used for cursor pagination.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    content (Union[str, List[LettaUserMessageContentUnion]]): The message content sent by the user (can be a string or an array of multi-modal content parts)

    - `id: string`

    - `content: Array<LettaUserMessageContentUnion> | string`

      The message content sent by the user (can be a string or an array of multi-modal content parts)

      - `Array<LettaUserMessageContentUnion>`

        - `TextContent`

          - `text: string`

            The text content of the message.

          - `signature?: string | null`

            Stores a unique identifier for any reasoning associated with this text content.

          - `type?: "text"`

            The type of the message.

            - `"text"`

        - `ImageContent`

          - `source: URLImage | Base64Image | LettaImage`

            The source of the image.

            - `URLImage`

              - `url: string`

                The URL of the image.

              - `type?: "url"`

                The source type for the image.

                - `"url"`

            - `Base64Image`

              - `data: string`

                The base64 encoded image data.

              - `media_type: string`

                The media type for the image.

              - `detail?: string | null`

                What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

              - `type?: "base64"`

                The source type for the image.

                - `"base64"`

            - `LettaImage`

              - `file_id: string`

                The unique identifier of the image file persisted in storage.

              - `data?: string | null`

                The base64 encoded image data.

              - `detail?: string | null`

                What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

              - `media_type?: string | null`

                The media type for the image.

              - `type?: "letta"`

                The source type for the image.

                - `"letta"`

          - `type?: "image"`

            The type of the message.

            - `"image"`

      - `string`

    - `date: string`

    - `is_err?: boolean | null`

    - `message_type?: "user_message"`

      The type of the message.

      - `"user_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `ReasoningMessage`

    Representation of an agent's internal reasoning.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    source (Literal["reasoner_model", "non_reasoner_model"]): Whether the reasoning
    content was generated natively by a reasoner model or derived via prompting
    reasoning (str): The internal reasoning of the agent
    signature (Optional[str]): The model-generated signature of the reasoning step

    - `id: string`

    - `date: string`

    - `reasoning: string`

    - `is_err?: boolean | null`

    - `message_type?: "reasoning_message"`

      The type of the message.

      - `"reasoning_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `signature?: string | null`

    - `source?: "reasoner_model" | "non_reasoner_model"`

      - `"reasoner_model"`

      - `"non_reasoner_model"`

    - `step_id?: string | null`

  - `HiddenReasoningMessage`

    Representation of an agent's internal reasoning where reasoning content
    has been hidden from the response.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    state (Literal["redacted", "omitted"]): Whether the reasoning
    content was redacted by the provider or simply omitted by the API
    hidden_reasoning (Optional[str]): The internal reasoning of the agent

    - `id: string`

    - `date: string`

    - `state: "redacted" | "omitted"`

      - `"redacted"`

      - `"omitted"`

    - `hidden_reasoning?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "hidden_reasoning_message"`

      The type of the message.

      - `"hidden_reasoning_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `ToolCallMessage`

    A message representing a request to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (Union[ToolCall, ToolCallDelta]): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "tool_call_message"`

      The type of the message.

      - `"tool_call_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `ToolReturnMessage`

    A message representing the return value of a tool call (generated by Letta executing the requested tool).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_return (str): The return value of the tool (deprecated, use tool_returns)
    status (Literal["success", "error"]): The status of the tool call (deprecated, use tool_returns)
    tool_call_id (str): A unique identifier for the tool call that generated this message (deprecated, use tool_returns)
    stdout (Optional[List(str)]): Captured stdout (e.g. prints, logs) from the tool invocation (deprecated, use tool_returns)
    stderr (Optional[List(str)]): Captured stderr from the tool invocation (deprecated, use tool_returns)
    tool_returns (Optional[List[ToolReturn]]): List of tool returns for multi-tool support

    - `id: string`

    - `date: string`

    - `status: "success" | "error"`

      - `"success"`

      - `"error"`

    - `tool_call_id: string`

    - `tool_return: string`

    - `is_err?: boolean | null`

    - `message_type?: "tool_return_message"`

      The type of the message.

      - `"tool_return_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `stderr?: Array<string> | null`

    - `stdout?: Array<string> | null`

    - `step_id?: string | null`

    - `tool_returns?: Array<ToolReturn> | null`

      - `status: "success" | "error"`

        - `"success"`

        - `"error"`

      - `tool_call_id: string`

      - `tool_return: Array<TextContent | ImageContent> | string`

        The tool return value - either a string or list of content parts (text/image)

        - `Array<TextContent | ImageContent>`

          - `TextContent`

          - `ImageContent`

        - `string`

      - `stderr?: Array<string> | null`

      - `stdout?: Array<string> | null`

      - `type?: "tool"`

        The message type to be created.

        - `"tool"`

  - `AssistantMessage`

    A message sent by the LLM in response to user input. Used in the LLM context.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    content (Union[str, List[LettaAssistantMessageContentUnion]]): The message content sent by the agent (can be a string or an array of content parts)

    - `id: string`

    - `content: Array<LettaAssistantMessageContentUnion> | string`

      The message content sent by the agent (can be a string or an array of content parts)

      - `Array<LettaAssistantMessageContentUnion>`

        - `text: string`

          The text content of the message.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this text content.

        - `type?: "text"`

          The type of the message.

          - `"text"`

      - `string`

    - `date: string`

    - `is_err?: boolean | null`

    - `message_type?: "assistant_message"`

      The type of the message.

      - `"assistant_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `ApprovalRequestMessage`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

      - `ToolCallDelta`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `ApprovalResponseMessage`

    A message representing a response form the user indicating whether a tool has been approved to run.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    approve: (bool) Whether the tool has been approved
    approval_request_id: The ID of the approval request
    reason: (Optional[str]) An optional explanation for the provided approval status

    - `id: string`

    - `date: string`

    - `approval_request_id?: string | null`

      The message ID of the approval request

    - `approvals?: Array<ApprovalReturn | ToolReturn> | null`

      The list of approval responses

      - `ApprovalReturn`

        - `approve: boolean`

          Whether the tool has been approved

        - `tool_call_id: string`

          The ID of the tool call that corresponds to this approval

        - `reason?: string | null`

          An optional explanation for the provided approval status

        - `type?: "approval"`

          The message type to be created.

          - `"approval"`

      - `ToolReturn`

        - `status: "success" | "error"`

        - `tool_call_id: string`

        - `tool_return: Array<TextContent | ImageContent> | string`

          The tool return value - either a string or list of content parts (text/image)

        - `stderr?: Array<string> | null`

        - `stdout?: Array<string> | null`

        - `type?: "tool"`

          The message type to be created.

    - `approve?: boolean | null`

      Whether the tool has been approved

    - `is_err?: boolean | null`

    - `message_type?: "approval_response_message"`

      The type of the message.

      - `"approval_response_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `reason?: string | null`

      An optional explanation for the provided approval status

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `LettaPing`

    A ping message used as a keepalive to prevent SSE streams from timing out during long running requests.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format

    - `id: string`

    - `date: string`

    - `is_err?: boolean | null`

    - `message_type?: "ping"`

      The type of the message. Ping messages are a keep-alive to prevent SSE streams from timing out during long running requests.

      - `"ping"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `LettaErrorMessage`

    Error messages are used to notify the client of an error that occurred during the agent's execution.

    - `error_type: string`

      The type of error.

    - `message: string`

      The error message.

    - `message_type: "error_message"`

      The type of the message.

      - `"error_message"`

    - `run_id: string`

      The ID of the run.

    - `detail?: string`

      An optional error detail.

    - `seq_id?: number`

      The sequence ID for cursor-based pagination.

  - `LettaStopReason`

    The stop reason from Letta indicating why agent loop stopped execution.

    - `stop_reason: StopReasonType`

      The reason why execution stopped.

      - `"end_turn"`

      - `"error"`

      - `"llm_api_error"`

      - `"invalid_llm_response"`

      - `"invalid_tool_call"`

      - `"max_steps"`

      - `"max_tokens_exceeded"`

      - `"no_tool_call"`

      - `"tool_rule"`

      - `"cancelled"`

      - `"insufficient_credits"`

      - `"requires_approval"`

      - `"context_window_overflow_in_system_prompt"`

    - `message_type?: "stop_reason"`

      The type of the message.

      - `"stop_reason"`

  - `LettaUsageStatistics`

    Usage statistics for the agent interaction.

    Attributes:
    completion_tokens (int): The number of tokens generated by the agent.
    prompt_tokens (int): The number of tokens in the prompt.
    total_tokens (int): The total number of tokens processed by the agent.
    step_count (int): The number of steps taken by the agent.
    cached_input_tokens (Optional[int]): The number of input tokens served from cache. None if not reported.
    cache_write_tokens (Optional[int]): The number of input tokens written to cache. None if not reported.
    reasoning_tokens (Optional[int]): The number of reasoning/thinking tokens generated. None if not reported.

    - `cache_write_tokens?: number | null`

      The number of input tokens written to cache (Anthropic only). None if not reported by provider.

    - `cached_input_tokens?: number | null`

      The number of input tokens served from cache. None if not reported by provider.

    - `completion_tokens?: number`

      The number of tokens generated by the agent.

    - `context_tokens?: number | null`

      Estimate of tokens currently in the context window.

    - `message_type?: "usage_statistics"`

      - `"usage_statistics"`

    - `prompt_tokens?: number`

      The number of tokens in the prompt.

    - `reasoning_tokens?: number | null`

      The number of reasoning/thinking tokens generated. None if not reported by provider.

    - `run_ids?: Array<string> | null`

      The background task run IDs associated with the agent interaction

    - `step_count?: number`

      The number of steps taken by the agent.

    - `total_tokens?: number`

      The total number of tokens processed by the agent.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const lettaStreamingResponse = await client.agents.messages.stream(
  'agent-123e4567-e89b-42d3-8456-426614174000',
);

console.log(lettaStreamingResponse);
```

#### Response

```json
{
  "id": "id",
  "content": "content",
  "date": "2019-12-27T18:11:19.117Z",
  "is_err": true,
  "message_type": "system_message",
  "name": "name",
  "otid": "otid",
  "run_id": "run_id",
  "sender_id": "sender_id",
  "seq_id": 0,
  "step_id": "step_id"
}
```

## Cancel Message

`client.agents.messages.cancel(stringagentID, MessageCancelParamsbody?, RequestOptionsoptions?): MessageCancelResponse`

**post** `/v1/agents/{agent_id}/messages/cancel`

Cancel runs associated with an agent. If run_ids are passed in, cancel those in particular.

Note to cancel active runs associated with an agent, redis is required.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `body: MessageCancelParams`

  - `run_ids?: Array<string> | null`

    Optional list of run IDs to cancel

### Returns

- `MessageCancelResponse = Record<string, unknown>`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const response = await client.agents.messages.cancel('agent-123e4567-e89b-42d3-8456-426614174000');

console.log(response);
```

#### Response

```json
{
  "foo": "bar"
}
```

## Create Message Async

`client.agents.messages.createAsync(stringagentID, MessageCreateAsyncParamsbody, RequestOptionsoptions?): Run`

**post** `/v1/agents/{agent_id}/messages/async`

Asynchronously process a user message and return a run object.
The actual processing happens in the background, and the status can be checked using the run ID.

This is "asynchronous" in the sense that it's a background run and explicitly must be fetched by the run ID.

**Note:** Sending multiple concurrent requests to the same agent can lead to undefined behavior.
Each agent processes messages sequentially, and concurrent requests may interleave in unexpected ways.
Wait for each request to complete before sending the next one. Use separate agents or conversations for parallel processing.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `body: MessageCreateAsyncParams`

  - `assistant_message_tool_kwarg?: string`

    The name of the message argument in the designated message tool. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

  - `assistant_message_tool_name?: string`

    The name of the designated message tool. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

  - `callback_url?: string | null`

    Optional callback URL to POST to when the job completes

  - `client_skills?: Array<ClientSkill> | null`

    Client-side skills available in the environment. These are rendered in the system prompt's available skills section alongside agent-scoped skills from MemFS.

    - `description: string`

      Description of what the skill does

    - `location: string`

      Path or location hint for the skill (e.g. skills/my-skill/SKILL.md)

    - `name: string`

      The name of the skill

  - `client_tools?: Array<ClientTool> | null`

    Client-side tools that the agent can call. When the agent calls a client-side tool, execution pauses and returns control to the client to execute the tool and provide the result via a ToolReturn.

    - `name: string`

      The name of the tool function

    - `description?: string | null`

      Description of what the tool does

    - `parameters?: Record<string, unknown> | null`

      JSON Schema for the function parameters

  - `enable_thinking?: string`

    If set to True, enables reasoning before responses or tool calls from the agent.

  - `include_compaction_messages?: boolean`

    If True, compaction events emit structured `SummaryMessage` and `EventMessage` types. If False (default), compaction messages are not included in the response.

  - `include_return_message_types?: Array<MessageType> | null`

    Only return specified message types in the response. If `None` (default) returns all messages.

    - `"system_message"`

    - `"user_message"`

    - `"assistant_message"`

    - `"reasoning_message"`

    - `"hidden_reasoning_message"`

    - `"tool_call_message"`

    - `"tool_return_message"`

    - `"approval_request_message"`

    - `"approval_response_message"`

    - `"summary_message"`

    - `"event_message"`

  - `input?: string | Array<TextContent | ImageContent | ToolCallContent | 5 more> | null`

    Syntactic sugar for a single user message. Equivalent to messages=[{'role': 'user', 'content': input}].

    - `string`

    - `Array<TextContent | ImageContent | ToolCallContent | 5 more>`

      - `TextContent`

        - `text: string`

          The text content of the message.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this text content.

        - `type?: "text"`

          The type of the message.

          - `"text"`

      - `ImageContent`

        - `source: URLImage | Base64Image | LettaImage`

          The source of the image.

          - `URLImage`

            - `url: string`

              The URL of the image.

            - `type?: "url"`

              The source type for the image.

              - `"url"`

          - `Base64Image`

            - `data: string`

              The base64 encoded image data.

            - `media_type: string`

              The media type for the image.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `type?: "base64"`

              The source type for the image.

              - `"base64"`

          - `LettaImage`

            - `file_id: string`

              The unique identifier of the image file persisted in storage.

            - `data?: string | null`

              The base64 encoded image data.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `media_type?: string | null`

              The media type for the image.

            - `type?: "letta"`

              The source type for the image.

              - `"letta"`

        - `type?: "image"`

          The type of the message.

          - `"image"`

      - `ToolCallContent`

        - `id: string`

          A unique identifier for this specific tool call instance.

        - `input: Record<string, unknown>`

          The parameters being passed to the tool, structured as a dictionary of parameter names to values.

        - `name: string`

          The name of the tool being called.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this tool call.

        - `type?: "tool_call"`

          Indicates this content represents a tool call event.

          - `"tool_call"`

      - `ToolReturnContent`

        - `content: string`

          The content returned by the tool execution.

        - `is_error: boolean`

          Indicates whether the tool execution resulted in an error.

        - `tool_call_id: string`

          References the ID of the ToolCallContent that initiated this tool call.

        - `type?: "tool_return"`

          Indicates this content represents a tool return event.

          - `"tool_return"`

      - `ReasoningContent`

        Sent via the Anthropic Messages API

        - `is_native: boolean`

          Whether the reasoning content was generated by a reasoner model that processed this step.

        - `reasoning: string`

          The intermediate reasoning or thought process content.

        - `signature?: string | null`

          A unique identifier for this reasoning step.

        - `type?: "reasoning"`

          Indicates this is a reasoning/intermediate step.

          - `"reasoning"`

      - `RedactedReasoningContent`

        Sent via the Anthropic Messages API

        - `data: string`

          The redacted or filtered intermediate reasoning content.

        - `type?: "redacted_reasoning"`

          Indicates this is a redacted thinking step.

          - `"redacted_reasoning"`

      - `OmittedReasoningContent`

        A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

        - `signature?: string | null`

          A unique identifier for this reasoning step.

        - `type?: "omitted_reasoning"`

          Indicates this is an omitted reasoning step.

          - `"omitted_reasoning"`

      - `SummarizedReasoningContent`

        The style of reasoning content returned by the OpenAI Responses API

        - `id: string`

          The unique identifier for this reasoning step.

        - `summary: Array<Summary>`

          Summaries of the reasoning content.

          - `index: number`

            The index of the summary part.

          - `text: string`

            The text of the summary part.

        - `encrypted_content?: string`

          The encrypted reasoning content.

        - `type?: "summarized_reasoning"`

          Indicates this is a summarized reasoning step.

          - `"summarized_reasoning"`

  - `max_steps?: number`

    Maximum number of steps the agent should take to process the request.

  - `messages?: Array<MessageCreate | ApprovalCreate | ToolReturnCreate> | null`

    The messages to be sent to the agent.

    - `MessageCreate`

      Request to create a message

      - `content: Array<LettaMessageContentUnion> | string`

        The content of the message.

        - `Array<LettaMessageContentUnion>`

          - `TextContent`

          - `ImageContent`

          - `ToolCallContent`

          - `ToolReturnContent`

          - `ReasoningContent`

            Sent via the Anthropic Messages API

          - `RedactedReasoningContent`

            Sent via the Anthropic Messages API

          - `OmittedReasoningContent`

            A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

        - `string`

      - `role: "user" | "system" | "assistant"`

        The role of the participant.

        - `"user"`

        - `"system"`

        - `"assistant"`

      - `batch_item_id?: string | null`

        The id of the LLMBatchItem that this message is associated with

      - `group_id?: string | null`

        The multi-agent group that the message was sent in

      - `name?: string | null`

        The name of the participant.

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `sender_id?: string | null`

        The id of the sender of the message, can be an identity id or agent id

      - `type?: "message" | null`

        The message type to be created.

        - `"message"`

    - `ApprovalCreate`

      Input to approve or deny a tool call request

      - `approval_request_id?: string | null`

        The message ID of the approval request

      - `approvals?: Array<ApprovalReturn | ToolReturn> | null`

        The list of approval responses

        - `ApprovalReturn`

          - `approve: boolean`

            Whether the tool has been approved

          - `tool_call_id: string`

            The ID of the tool call that corresponds to this approval

          - `reason?: string | null`

            An optional explanation for the provided approval status

          - `type?: "approval"`

            The message type to be created.

            - `"approval"`

        - `ToolReturn`

          - `status: "success" | "error"`

            - `"success"`

            - `"error"`

          - `tool_call_id: string`

          - `tool_return: Array<TextContent | ImageContent> | string`

            The tool return value - either a string or list of content parts (text/image)

            - `Array<TextContent | ImageContent>`

              - `TextContent`

              - `ImageContent`

            - `string`

          - `stderr?: Array<string> | null`

          - `stdout?: Array<string> | null`

          - `type?: "tool"`

            The message type to be created.

            - `"tool"`

      - `approve?: boolean | null`

        Whether the tool has been approved

      - `group_id?: string | null`

        The multi-agent group that the message was sent in

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `reason?: string | null`

        An optional explanation for the provided approval status

      - `type?: "approval"`

        The message type to be created.

        - `"approval"`

    - `ToolReturnCreate`

      Submit tool return(s) from client-side tool execution.

      This is the preferred way to send tool results back to the agent after
      client-side tool execution. It is equivalent to sending an ApprovalCreate
      with tool return approvals, but provides a cleaner API for the common case.

      - `tool_returns: Array<ToolReturn>`

        List of tool returns from client-side execution

        - `status: "success" | "error"`

        - `tool_call_id: string`

        - `tool_return: Array<TextContent | ImageContent> | string`

          The tool return value - either a string or list of content parts (text/image)

        - `stderr?: Array<string> | null`

        - `stdout?: Array<string> | null`

        - `type?: "tool"`

          The message type to be created.

      - `group_id?: string | null`

        The multi-agent group that the message was sent in

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `type?: "tool_return"`

        The message type to be created.

        - `"tool_return"`

  - `override_model?: string | null`

    Model handle to use for this request instead of the agent's default model. This allows sending a message to a different model without changing the agent's configuration.

  - `override_system?: string | null`

    Optional per-request system prompt override. When set, this is passed directly to the underlying LLM request and bypasses the persisted/compiled system message for that request.

  - `return_logprobs?: boolean`

    If True, returns log probabilities of the output tokens in the response. Useful for RL training. Only supported for OpenAI-compatible providers (including SGLang).

  - `return_token_ids?: boolean`

    If True, returns token IDs and logprobs for ALL LLM generations in the agent step, not just the last one. Uses SGLang native /generate endpoint. Returns 'turns' field with TurnTokenData for each assistant/tool turn. Required for proper multi-turn RL training with loss masking.

  - `top_logprobs?: number | null`

    Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

  - `use_assistant_message?: boolean`

    Whether the server should parse specific tool call arguments (default `send_message`) as `AssistantMessage` objects. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

### Returns

- `Run`

  Representation of a run - a conversation or processing session for an agent. Runs track when agents process messages and maintain the relationship between agents, steps, and messages.

  - `id: string`

    The human-friendly ID of the Run

  - `agent_id: string`

    The unique identifier of the agent associated with the run.

  - `background?: boolean | null`

    Whether the run was created in background mode.

  - `base_template_id?: string | null`

    The base template ID that the run belongs to.

  - `callback_error?: string | null`

    Optional error message from attempting to POST the callback endpoint.

  - `callback_sent_at?: string | null`

    Timestamp when the callback was last attempted.

  - `callback_status_code?: number | null`

    HTTP status code returned by the callback endpoint.

  - `callback_url?: string | null`

    If set, POST to this URL when the run completes.

  - `completed_at?: string | null`

    The timestamp when the run was completed.

  - `conversation_id?: string | null`

    The unique identifier of the conversation associated with the run.

  - `created_at?: string`

    The timestamp when the run was created.

  - `metadata?: Record<string, unknown> | null`

    Additional metadata for the run.

  - `request_config?: RequestConfig | null`

    The request configuration for the run.

    - `assistant_message_tool_kwarg?: string`

      The name of the message argument in the designated message tool.

    - `assistant_message_tool_name?: string`

      The name of the designated message tool.

    - `include_return_message_types?: Array<MessageType> | null`

      Only return specified message types in the response. If `None` (default) returns all messages.

      - `"system_message"`

      - `"user_message"`

      - `"assistant_message"`

      - `"reasoning_message"`

      - `"hidden_reasoning_message"`

      - `"tool_call_message"`

      - `"tool_return_message"`

      - `"approval_request_message"`

      - `"approval_response_message"`

      - `"summary_message"`

      - `"event_message"`

    - `use_assistant_message?: boolean`

      Whether the server should parse specific tool call arguments (default `send_message`) as `AssistantMessage` objects.

  - `status?: "created" | "running" | "completed" | 2 more`

    The current status of the run.

    - `"created"`

    - `"running"`

    - `"completed"`

    - `"failed"`

    - `"cancelled"`

  - `stop_reason?: StopReasonType | null`

    The reason why the run was stopped.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `total_duration_ns?: number | null`

    Total run duration in nanoseconds

  - `ttft_ns?: number | null`

    Time to first token for a run in nanoseconds

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const run = await client.agents.messages.createAsync('agent-123e4567-e89b-42d3-8456-426614174000');

console.log(run.id);
```

#### Response

```json
{
  "id": "run-123e4567-e89b-12d3-a456-426614174000",
  "agent_id": "agent_id",
  "background": true,
  "base_template_id": "base_template_id",
  "callback_error": "callback_error",
  "callback_sent_at": "2019-12-27T18:11:19.117Z",
  "callback_status_code": 0,
  "callback_url": "callback_url",
  "completed_at": "2019-12-27T18:11:19.117Z",
  "conversation_id": "conversation_id",
  "created_at": "2019-12-27T18:11:19.117Z",
  "metadata": {
    "foo": "bar"
  },
  "request_config": {
    "assistant_message_tool_kwarg": "assistant_message_tool_kwarg",
    "assistant_message_tool_name": "assistant_message_tool_name",
    "include_return_message_types": [
      "system_message"
    ],
    "use_assistant_message": true
  },
  "status": "created",
  "stop_reason": "end_turn",
  "total_duration_ns": 0,
  "ttft_ns": 0
}
```

## Reset Messages

`client.agents.messages.reset(stringagentID, MessageResetParamsbody, RequestOptionsoptions?): AgentState | null`

**patch** `/v1/agents/{agent_id}/reset-messages`

Resets the messages for an agent

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `body: MessageResetParams`

  - `add_default_initial_messages?: boolean`

    If true, adds the default initial messages after resetting.

### Returns

- `AgentState | null`

  - `id: string`

    The id of the agent. Assigned by the database.

  - `agent_type: AgentType`

    The type of agent.

    - `"memgpt_agent"`

    - `"memgpt_v2_agent"`

    - `"letta_v1_agent"`

    - `"react_agent"`

    - `"workflow_agent"`

    - `"split_thread_agent"`

    - `"sleeptime_agent"`

    - `"voice_convo_agent"`

    - `"voice_sleeptime_agent"`

  - `blocks: Array<Block>`

    The memory blocks used by the agent.

    - `value: string`

      Value of the block.

    - `id?: string`

      The human-friendly ID of the Block

    - `base_template_id?: string | null`

      The base template id of the block.

    - `created_by_id?: string | null`

      The id of the user that made this Block.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      Description of the block.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the block will be hidden.

    - `is_template?: boolean`

      Whether the block is a template (e.g. saved human/persona options).

    - `label?: string | null`

      Label of the block (e.g. 'human', 'persona') in the context window.

    - `last_updated_by_id?: string | null`

      The id of the user that last updated this Block.

    - `limit?: number`

      Character limit of the block.

    - `metadata?: Record<string, unknown> | null`

      Metadata of the block.

    - `preserve_on_migration?: boolean | null`

      Preserve the block on template migration.

    - `project_id?: string | null`

      The associated project id.

    - `read_only?: boolean`

      Whether the agent has read-only access to the block.

    - `tags?: Array<string> | null`

      The tags associated with the block.

    - `template_id?: string | null`

      The id of the template.

    - `template_name?: string | null`

      Name of the block if it is a template.

  - `llm_config: LlmConfig`

    Deprecated: Use `model` field instead. The LLM configuration used by the agent.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

        - `type?: "text"`

          The type of the response format.

          - `"text"`

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

        - `json_schema: Record<string, unknown>`

          The JSON schema of the response.

        - `type?: "json_schema"`

          The type of the response format.

          - `"json_schema"`

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

        - `type?: "json_object"`

          The type of the response format.

          - `"json_object"`

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `memory: Memory`

    Deprecated: Use `blocks` field instead. The in-context memory of the agent.

    - `blocks: Array<Block>`

      Memory blocks contained in the agent's in-context memory

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `agent_type?: AgentType | (string & {}) | null`

      Agent type controlling prompt rendering.

      - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

        Enum to represent the type of agent.

      - `(string & {})`

    - `file_blocks?: Array<FileBlock>`

      Special blocks representing the agent's in-context memory of an attached file

      - `file_id: string`

        Unique identifier of the file.

      - `is_open: boolean`

        True if the agent currently has the file open.

      - `source_id: string`

        Deprecated: Use `folder_id` field instead. Unique identifier of the source.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_accessed_at?: string | null`

        UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `git_enabled?: boolean`

      Whether this agent uses git-backed memory with structured labels.

    - `prompt_template?: string`

      Deprecated. Ignored for performance.

  - `name: string`

    The name of the agent.

  - `sources: Array<Source>`

    Deprecated: Use `folders` field instead. The sources used by the agent.

    - `id: string`

      The human-friendly ID of the Source

    - `embedding_config: EmbeddingConfig`

      The embedding configuration used by the source.

      - `embedding_dim: number`

        The dimension of the embedding.

      - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"bedrock"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"mistral"`

        - `"together"`

        - `"pinecone"`

      - `embedding_model: string`

        The model for the embedding.

      - `azure_deployment?: string | null`

        The Azure deployment for the model.

      - `azure_endpoint?: string | null`

        The Azure endpoint for the model.

      - `azure_version?: string | null`

        The Azure version for the model.

      - `batch_size?: number`

        The maximum batch size for processing embeddings.

      - `embedding_chunk_size?: number | null`

        The chunk size of the embedding.

      - `embedding_endpoint?: string | null`

        The endpoint for the model (`None` if local).

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

    - `name: string`

      The name of the source.

    - `created_at?: string | null`

      The timestamp when the source was created.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `description?: string | null`

      The description of the source.

    - `instructions?: string | null`

      Instructions for how to use the source.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata?: Record<string, unknown> | null`

      Metadata associated with the source.

    - `updated_at?: string | null`

      The timestamp when the source was last updated.

    - `vector_db_provider?: VectorDBProvider`

      The vector database provider used for this source's passages

      - `"native"`

      - `"tpuf"`

      - `"pinecone"`

  - `system: string`

    The system prompt used by the agent.

  - `tags: Array<string>`

    The tags associated with the agent.

  - `tools: Array<Tool>`

    The tools used by the agent.

    - `id: string`

      The human-friendly ID of the Tool

    - `args_json_schema?: Record<string, unknown> | null`

      The args JSON schema of the function.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `default_requires_approval?: boolean | null`

      Default value for whether or not executing this tool requires approval.

    - `description?: string | null`

      The description of the tool.

    - `enable_parallel_execution?: boolean | null`

      If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

    - `json_schema?: Record<string, unknown> | null`

      The JSON schema of the function.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata_?: Record<string, unknown> | null`

      A dictionary of additional metadata for the tool.

    - `name?: string | null`

      The name of the function.

    - `npm_requirements?: Array<NpmRequirement> | null`

      Optional list of npm packages required by this tool.

      - `name: string`

        Name of the npm package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `pip_requirements?: Array<PipRequirement> | null`

      Optional list of pip packages required by this tool.

      - `name: string`

        Name of the pip package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `project_id?: string | null`

      The project id of the tool.

    - `return_char_limit?: number`

      The maximum number of characters in the response.

    - `source_code?: string | null`

      The source code of the function.

    - `source_type?: string | null`

      The type of the source code.

    - `tags?: Array<string>`

      Metadata tags.

    - `tool_type?: ToolType`

      The type of the tool.

      - `"custom"`

      - `"letta_core"`

      - `"letta_memory_core"`

      - `"letta_multi_agent_core"`

      - `"letta_sleeptime_core"`

      - `"letta_voice_sleeptime_core"`

      - `"letta_builtin"`

      - `"letta_files_core"`

      - `"external_langchain"`

      - `"external_composio"`

      - `"external_mcp"`

  - `base_template_id?: string | null`

    The base template id of the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `created_at?: string | null`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the agent will be hidden.

  - `identities?: Array<Identity>`

    The identities associated with this agent.

    - `id: string`

      The human-friendly ID of the Identity

    - `agent_ids: Array<string>`

      The IDs of the agents associated with the identity.

    - `block_ids: Array<string>`

      The IDs of the blocks associated with the identity.

    - `identifier_key: string`

      External, user-generated identifier key of the identity.

    - `identity_type: "org" | "user" | "other"`

      The type of the identity.

      - `"org"`

      - `"user"`

      - `"other"`

    - `name: string`

      The name of the identity.

    - `project_id?: string | null`

      The project id of the identity, if applicable.

    - `properties?: Array<Property>`

      List of properties associated with the identity

      - `key: string`

        The key of the property

      - `type: "string" | "number" | "boolean" | "json"`

        The type of the property

        - `"string"`

        - `"number"`

        - `"boolean"`

        - `"json"`

      - `value: string | number | boolean | Record<string, unknown>`

        The value of the property

        - `string`

        - `number`

        - `boolean`

        - `Record<string, unknown>`

  - `identity_ids?: Array<string>`

    Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

  - `last_run_completion?: string | null`

    The timestamp when the agent last completed a run.

  - `last_run_duration_ms?: number | null`

    The duration in milliseconds of the agent's last run.

  - `last_stop_reason?: StopReasonType | null`

    The stop reason from the agent's last run.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `managed_group?: ManagedGroup | null`

    The multi-agent group that this agent manages

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `message_buffer_autoclear?: boolean`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `message_ids?: Array<string> | null`

    The ids of the messages in the agent's in-context memory.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle used by the agent (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings used by the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `multi_agent_group?: MultiAgentGroup | null`

    Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `pending_approval?: ApprovalRequestMessage | null`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project_id?: string | null`

    The id of the project the agent belongs to.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format used by the agent

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Array<AgentEnvironmentVariable>`

    The environment variables for tool execution specific to this agent.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `template_id?: string | null`

    The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

    Deprecated: use `secrets` field instead.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The list of tool rules.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const agentState = await client.agents.messages.reset('agent-123e4567-e89b-42d3-8456-426614174000');

console.log(agentState.id);
```

#### Response

```json
{
  "id": "id",
  "agent_type": "memgpt_agent",
  "blocks": [
    {
      "value": "value",
      "id": "block-123e4567-e89b-12d3-a456-426614174000",
      "base_template_id": "base_template_id",
      "created_by_id": "created_by_id",
      "deployment_id": "deployment_id",
      "description": "description",
      "entity_id": "entity_id",
      "hidden": true,
      "is_template": true,
      "label": "label",
      "last_updated_by_id": "last_updated_by_id",
      "limit": 0,
      "metadata": {
        "foo": "bar"
      },
      "preserve_on_migration": true,
      "project_id": "project_id",
      "read_only": true,
      "tags": [
        "string"
      ],
      "template_id": "template_id",
      "template_name": "template_name"
    }
  ],
  "llm_config": {
    "context_window": 0,
    "model": "model",
    "model_endpoint_type": "openai",
    "compatibility_type": "gguf",
    "display_name": "display_name",
    "effort": "low",
    "enable_reasoner": true,
    "frequency_penalty": 0,
    "handle": "handle",
    "max_reasoning_tokens": 0,
    "max_tokens": 0,
    "model_endpoint": "model_endpoint",
    "model_wrapper": "model_wrapper",
    "parallel_tool_calls": true,
    "provider_category": "base",
    "provider_name": "provider_name",
    "put_inner_thoughts_in_kwargs": true,
    "reasoning_effort": "none",
    "response_format": {
      "type": "text"
    },
    "return_logprobs": true,
    "return_token_ids": true,
    "strict": true,
    "temperature": 0,
    "tier": "tier",
    "tool_call_parser": "tool_call_parser",
    "top_logprobs": 0,
    "verbosity": "low"
  },
  "memory": {
    "blocks": [
      {
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "agent_type": "memgpt_agent",
    "file_blocks": [
      {
        "file_id": "file_id",
        "is_open": true,
        "source_id": "source_id",
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_accessed_at": "2019-12-27T18:11:19.117Z",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "git_enabled": true,
    "prompt_template": "prompt_template"
  },
  "name": "name",
  "sources": [
    {
      "id": "source-123e4567-e89b-12d3-a456-426614174000",
      "embedding_config": {
        "embedding_dim": 0,
        "embedding_endpoint_type": "openai",
        "embedding_model": "embedding_model",
        "azure_deployment": "azure_deployment",
        "azure_endpoint": "azure_endpoint",
        "azure_version": "azure_version",
        "batch_size": 0,
        "embedding_chunk_size": 0,
        "embedding_endpoint": "embedding_endpoint",
        "handle": "handle"
      },
      "name": "name",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "instructions": "instructions",
      "last_updated_by_id": "last_updated_by_id",
      "metadata": {
        "foo": "bar"
      },
      "updated_at": "2019-12-27T18:11:19.117Z",
      "vector_db_provider": "native"
    }
  ],
  "system": "system",
  "tags": [
    "string"
  ],
  "tools": [
    {
      "id": "tool-123e4567-e89b-12d3-a456-426614174000",
      "args_json_schema": {
        "foo": "bar"
      },
      "created_by_id": "created_by_id",
      "default_requires_approval": true,
      "description": "description",
      "enable_parallel_execution": true,
      "json_schema": {
        "foo": "bar"
      },
      "last_updated_by_id": "last_updated_by_id",
      "metadata_": {
        "foo": "bar"
      },
      "name": "name",
      "npm_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "pip_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "project_id": "project_id",
      "return_char_limit": 1,
      "source_code": "source_code",
      "source_type": "source_type",
      "tags": [
        "string"
      ],
      "tool_type": "custom"
    }
  ],
  "base_template_id": "base_template_id",
  "compaction_settings": {
    "clip_chars": 0,
    "mode": "all",
    "model": "model",
    "model_settings": {
      "max_output_tokens": 0,
      "parallel_tool_calls": true,
      "provider_type": "openai",
      "reasoning": {
        "reasoning_effort": "none"
      },
      "response_format": {
        "type": "text"
      },
      "strict": true,
      "temperature": 0
    },
    "prompt": "prompt",
    "prompt_acknowledgement": true,
    "sliding_window_percentage": 0
  },
  "created_at": "2019-12-27T18:11:19.117Z",
  "created_by_id": "created_by_id",
  "deployment_id": "deployment_id",
  "description": "description",
  "embedding": "embedding",
  "embedding_config": {
    "embedding_dim": 0,
    "embedding_endpoint_type": "openai",
    "embedding_model": "embedding_model",
    "azure_deployment": "azure_deployment",
    "azure_endpoint": "azure_endpoint",
    "azure_version": "azure_version",
    "batch_size": 0,
    "embedding_chunk_size": 0,
    "embedding_endpoint": "embedding_endpoint",
    "handle": "handle"
  },
  "enable_sleeptime": true,
  "entity_id": "entity_id",
  "hidden": true,
  "identities": [
    {
      "id": "identity-123e4567-e89b-12d3-a456-426614174000",
      "agent_ids": [
        "string"
      ],
      "block_ids": [
        "string"
      ],
      "identifier_key": "identifier_key",
      "identity_type": "org",
      "name": "name",
      "project_id": "project_id",
      "properties": [
        {
          "key": "key",
          "type": "string",
          "value": "string"
        }
      ]
    }
  ],
  "identity_ids": [
    "string"
  ],
  "last_run_completion": "2019-12-27T18:11:19.117Z",
  "last_run_duration_ms": 0,
  "last_stop_reason": "end_turn",
  "last_updated_by_id": "last_updated_by_id",
  "managed_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "max_files_open": 0,
  "message_buffer_autoclear": true,
  "message_ids": [
    "string"
  ],
  "metadata": {
    "foo": "bar"
  },
  "model": "model",
  "model_settings": {
    "max_output_tokens": 0,
    "parallel_tool_calls": true,
    "provider_type": "openai",
    "reasoning": {
      "reasoning_effort": "none"
    },
    "response_format": {
      "type": "text"
    },
    "strict": true,
    "temperature": 0
  },
  "multi_agent_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "pending_approval": {
    "id": "id",
    "date": "2019-12-27T18:11:19.117Z",
    "tool_call": {
      "arguments": "arguments",
      "name": "name",
      "tool_call_id": "tool_call_id"
    },
    "is_err": true,
    "message_type": "approval_request_message",
    "name": "name",
    "otid": "otid",
    "run_id": "run_id",
    "sender_id": "sender_id",
    "seq_id": 0,
    "step_id": "step_id",
    "tool_calls": [
      {
        "arguments": "arguments",
        "name": "name",
        "tool_call_id": "tool_call_id"
      }
    ]
  },
  "per_file_view_window_char_limit": 0,
  "project_id": "project_id",
  "response_format": {
    "type": "text"
  },
  "secrets": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "template_id": "template_id",
  "timezone": "timezone",
  "tool_exec_environment_variables": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "tool_rules": [
    {
      "children": [
        "string"
      ],
      "tool_name": "tool_name",
      "child_arg_nodes": [
        {
          "name": "name",
          "args": {
            "foo": "bar"
          }
        }
      ],
      "prompt_template": "prompt_template",
      "type": "constrain_child_tools"
    }
  ],
  "updated_at": "2019-12-27T18:11:19.117Z"
}
```

## Summarize Messages

`client.agents.messages.compact(stringagentID, MessageCompactParamsbody?, RequestOptionsoptions?): CompactionResponse`

**post** `/v1/agents/{agent_id}/summarize`

Summarize an agent's conversation history.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `body: MessageCompactParams`

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

            - `type?: "text"`

              The type of the response format.

              - `"text"`

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

            - `json_schema: Record<string, unknown>`

              The JSON schema of the response.

            - `type?: "json_schema"`

              The type of the response format.

              - `"json_schema"`

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

            - `type?: "json_object"`

              The type of the response format.

              - `"json_object"`

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

### Returns

- `CompactionResponse`

  - `num_messages_after: number`

  - `num_messages_before: number`

  - `summary: string`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const compactionResponse = await client.agents.messages.compact(
  'agent-123e4567-e89b-42d3-8456-426614174000',
);

console.log(compactionResponse.num_messages_after);
```

#### Response

```json
{
  "num_messages_after": 0,
  "num_messages_before": 0,
  "summary": "summary"
}
```

## Domain Types

### Approval Create

- `ApprovalCreate`

  Input to approve or deny a tool call request

  - `approval_request_id?: string | null`

    The message ID of the approval request

  - `approvals?: Array<ApprovalReturn | ToolReturn> | null`

    The list of approval responses

    - `ApprovalReturn`

      - `approve: boolean`

        Whether the tool has been approved

      - `tool_call_id: string`

        The ID of the tool call that corresponds to this approval

      - `reason?: string | null`

        An optional explanation for the provided approval status

      - `type?: "approval"`

        The message type to be created.

        - `"approval"`

    - `ToolReturn`

      - `status: "success" | "error"`

        - `"success"`

        - `"error"`

      - `tool_call_id: string`

      - `tool_return: Array<TextContent | ImageContent> | string`

        The tool return value - either a string or list of content parts (text/image)

        - `Array<TextContent | ImageContent>`

          - `TextContent`

            - `text: string`

              The text content of the message.

            - `signature?: string | null`

              Stores a unique identifier for any reasoning associated with this text content.

            - `type?: "text"`

              The type of the message.

              - `"text"`

          - `ImageContent`

            - `source: URLImage | Base64Image | LettaImage`

              The source of the image.

              - `URLImage`

                - `url: string`

                  The URL of the image.

                - `type?: "url"`

                  The source type for the image.

                  - `"url"`

              - `Base64Image`

                - `data: string`

                  The base64 encoded image data.

                - `media_type: string`

                  The media type for the image.

                - `detail?: string | null`

                  What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

                - `type?: "base64"`

                  The source type for the image.

                  - `"base64"`

              - `LettaImage`

                - `file_id: string`

                  The unique identifier of the image file persisted in storage.

                - `data?: string | null`

                  The base64 encoded image data.

                - `detail?: string | null`

                  What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

                - `media_type?: string | null`

                  The media type for the image.

                - `type?: "letta"`

                  The source type for the image.

                  - `"letta"`

            - `type?: "image"`

              The type of the message.

              - `"image"`

        - `string`

      - `stderr?: Array<string> | null`

      - `stdout?: Array<string> | null`

      - `type?: "tool"`

        The message type to be created.

        - `"tool"`

  - `approve?: boolean | null`

    Whether the tool has been approved

  - `group_id?: string | null`

    The multi-agent group that the message was sent in

  - `otid?: string | null`

    The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

  - `reason?: string | null`

    An optional explanation for the provided approval status

  - `type?: "approval"`

    The message type to be created.

    - `"approval"`

### Approval Request Message

- `ApprovalRequestMessage`

  A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

  Args:
  id (str): The ID of the message
  date (datetime): The date the message was created in ISO format
  name (Optional[str]): The name of the sender of the message
  tool_call (ToolCall): The tool call

  - `id: string`

  - `date: string`

  - `tool_call: ToolCall | ToolCallDelta`

    The tool call that has been requested by the llm to run

    - `ToolCall`

      - `arguments: string`

      - `name: string`

      - `tool_call_id: string`

    - `ToolCallDelta`

      - `arguments?: string | null`

      - `name?: string | null`

      - `tool_call_id?: string | null`

  - `is_err?: boolean | null`

  - `message_type?: "approval_request_message"`

    The type of the message.

    - `"approval_request_message"`

  - `name?: string | null`

  - `otid?: string | null`

    The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

  - `run_id?: string | null`

  - `sender_id?: string | null`

  - `seq_id?: number | null`

  - `step_id?: string | null`

  - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

    The tool calls that have been requested by the llm to run, which are pending approval

    - `Array<ToolCall>`

      - `arguments: string`

      - `name: string`

      - `tool_call_id: string`

    - `ToolCallDelta`

### Approval Response Message

- `ApprovalResponseMessage`

  A message representing a response form the user indicating whether a tool has been approved to run.

  Args:
  id (str): The ID of the message
  date (datetime): The date the message was created in ISO format
  name (Optional[str]): The name of the sender of the message
  approve: (bool) Whether the tool has been approved
  approval_request_id: The ID of the approval request
  reason: (Optional[str]) An optional explanation for the provided approval status

  - `id: string`

  - `date: string`

  - `approval_request_id?: string | null`

    The message ID of the approval request

  - `approvals?: Array<ApprovalReturn | ToolReturn> | null`

    The list of approval responses

    - `ApprovalReturn`

      - `approve: boolean`

        Whether the tool has been approved

      - `tool_call_id: string`

        The ID of the tool call that corresponds to this approval

      - `reason?: string | null`

        An optional explanation for the provided approval status

      - `type?: "approval"`

        The message type to be created.

        - `"approval"`

    - `ToolReturn`

      - `status: "success" | "error"`

        - `"success"`

        - `"error"`

      - `tool_call_id: string`

      - `tool_return: Array<TextContent | ImageContent> | string`

        The tool return value - either a string or list of content parts (text/image)

        - `Array<TextContent | ImageContent>`

          - `TextContent`

            - `text: string`

              The text content of the message.

            - `signature?: string | null`

              Stores a unique identifier for any reasoning associated with this text content.

            - `type?: "text"`

              The type of the message.

              - `"text"`

          - `ImageContent`

            - `source: URLImage | Base64Image | LettaImage`

              The source of the image.

              - `URLImage`

                - `url: string`

                  The URL of the image.

                - `type?: "url"`

                  The source type for the image.

                  - `"url"`

              - `Base64Image`

                - `data: string`

                  The base64 encoded image data.

                - `media_type: string`

                  The media type for the image.

                - `detail?: string | null`

                  What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

                - `type?: "base64"`

                  The source type for the image.

                  - `"base64"`

              - `LettaImage`

                - `file_id: string`

                  The unique identifier of the image file persisted in storage.

                - `data?: string | null`

                  The base64 encoded image data.

                - `detail?: string | null`

                  What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

                - `media_type?: string | null`

                  The media type for the image.

                - `type?: "letta"`

                  The source type for the image.

                  - `"letta"`

            - `type?: "image"`

              The type of the message.

              - `"image"`

        - `string`

      - `stderr?: Array<string> | null`

      - `stdout?: Array<string> | null`

      - `type?: "tool"`

        The message type to be created.

        - `"tool"`

  - `approve?: boolean | null`

    Whether the tool has been approved

  - `is_err?: boolean | null`

  - `message_type?: "approval_response_message"`

    The type of the message.

    - `"approval_response_message"`

  - `name?: string | null`

  - `otid?: string | null`

    The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

  - `reason?: string | null`

    An optional explanation for the provided approval status

  - `run_id?: string | null`

  - `sender_id?: string | null`

  - `seq_id?: number | null`

  - `step_id?: string | null`

### Approval Return

- `ApprovalReturn`

  - `approve: boolean`

    Whether the tool has been approved

  - `tool_call_id: string`

    The ID of the tool call that corresponds to this approval

  - `reason?: string | null`

    An optional explanation for the provided approval status

  - `type?: "approval"`

    The message type to be created.

    - `"approval"`

### Assistant Message

- `AssistantMessage`

  A message sent by the LLM in response to user input. Used in the LLM context.

  Args:
  id (str): The ID of the message
  date (datetime): The date the message was created in ISO format
  name (Optional[str]): The name of the sender of the message
  content (Union[str, List[LettaAssistantMessageContentUnion]]): The message content sent by the agent (can be a string or an array of content parts)

  - `id: string`

  - `content: Array<LettaAssistantMessageContentUnion> | string`

    The message content sent by the agent (can be a string or an array of content parts)

    - `Array<LettaAssistantMessageContentUnion>`

      - `text: string`

        The text content of the message.

      - `signature?: string | null`

        Stores a unique identifier for any reasoning associated with this text content.

      - `type?: "text"`

        The type of the message.

        - `"text"`

    - `string`

  - `date: string`

  - `is_err?: boolean | null`

  - `message_type?: "assistant_message"`

    The type of the message.

    - `"assistant_message"`

  - `name?: string | null`

  - `otid?: string | null`

    The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

  - `run_id?: string | null`

  - `sender_id?: string | null`

  - `seq_id?: number | null`

  - `step_id?: string | null`

### Event Message

- `EventMessage`

  A message for notifying the developer that an event that has occured (e.g. a compaction). Events are NOT part of the context window.

  - `id: string`

  - `date: string`

  - `event_data: Record<string, unknown>`

  - `event_type: "compaction"`

    - `"compaction"`

  - `is_err?: boolean | null`

  - `message_type?: "event_message"`

    - `"event_message"`

  - `name?: string | null`

  - `otid?: string | null`

    The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

  - `run_id?: string | null`

  - `sender_id?: string | null`

  - `seq_id?: number | null`

  - `step_id?: string | null`

### Hidden Reasoning Message

- `HiddenReasoningMessage`

  Representation of an agent's internal reasoning where reasoning content
  has been hidden from the response.

  Args:
  id (str): The ID of the message
  date (datetime): The date the message was created in ISO format
  name (Optional[str]): The name of the sender of the message
  state (Literal["redacted", "omitted"]): Whether the reasoning
  content was redacted by the provider or simply omitted by the API
  hidden_reasoning (Optional[str]): The internal reasoning of the agent

  - `id: string`

  - `date: string`

  - `state: "redacted" | "omitted"`

    - `"redacted"`

    - `"omitted"`

  - `hidden_reasoning?: string | null`

  - `is_err?: boolean | null`

  - `message_type?: "hidden_reasoning_message"`

    The type of the message.

    - `"hidden_reasoning_message"`

  - `name?: string | null`

  - `otid?: string | null`

    The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

  - `run_id?: string | null`

  - `sender_id?: string | null`

  - `seq_id?: number | null`

  - `step_id?: string | null`

### Image Content

- `ImageContent`

  - `source: URLImage | Base64Image | LettaImage`

    The source of the image.

    - `URLImage`

      - `url: string`

        The URL of the image.

      - `type?: "url"`

        The source type for the image.

        - `"url"`

    - `Base64Image`

      - `data: string`

        The base64 encoded image data.

      - `media_type: string`

        The media type for the image.

      - `detail?: string | null`

        What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

      - `type?: "base64"`

        The source type for the image.

        - `"base64"`

    - `LettaImage`

      - `file_id: string`

        The unique identifier of the image file persisted in storage.

      - `data?: string | null`

        The base64 encoded image data.

      - `detail?: string | null`

        What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

      - `media_type?: string | null`

        The media type for the image.

      - `type?: "letta"`

        The source type for the image.

        - `"letta"`

  - `type?: "image"`

    The type of the message.

    - `"image"`

### Internal Message

- `InternalMessage`

  Letta's internal representation of a message. Includes methods to convert to/from LLM provider formats.

  Attributes:
  id (str): The unique identifier of the message.
  role (MessageRole): The role of the participant.
  text (str): The text of the message.
  user_id (str): The unique identifier of the user.
  agent_id (str): The unique identifier of the agent.
  model (str): The model used to make the function call.
  name (str): The name of the participant.
  created_at (datetime): The time the message was created.
  tool_calls (List[OpenAIToolCall,]): The list of tool calls requested.
  tool_call_id (str): The id of the tool call.
  step_id (str): The id of the step that this message was created in.
  otid (str): The offline threading id associated with this message.
  tool_returns (List[ToolReturn]): The list of tool returns requested.
  group_id (str): The multi-agent group that the message was sent in.
  sender_id (str): The id of the sender of the message, can be an identity id or agent id.
  conversation_id (str): The conversation this message belongs to.
  t

  - `id: string`

    The human-friendly ID of the Message

  - `role: MessageRole`

    The role of the participant.

    - `"assistant"`

    - `"user"`

    - `"tool"`

    - `"function"`

    - `"system"`

    - `"approval"`

    - `"summary"`

  - `agent_id?: string | null`

    The unique identifier of the agent.

  - `approval_request_id?: string | null`

    The id of the approval request if this message is associated with a tool call request.

  - `approvals?: Array<ApprovalReturn | LettaSchemasMessageToolReturnOutput> | null`

    The list of approvals for this message.

    - `ApprovalReturn`

      - `approve: boolean`

        Whether the tool has been approved

      - `tool_call_id: string`

        The ID of the tool call that corresponds to this approval

      - `reason?: string | null`

        An optional explanation for the provided approval status

      - `type?: "approval"`

        The message type to be created.

        - `"approval"`

    - `LettaSchemasMessageToolReturnOutput`

      - `status: "success" | "error"`

        The status of the tool call

        - `"success"`

        - `"error"`

      - `func_response?: string | Array<TextContent | ImageContent> | null`

        The function response - either a string or list of content parts (text/image)

        - `string`

        - `Array<TextContent | ImageContent>`

          - `TextContent`

            - `text: string`

              The text content of the message.

            - `signature?: string | null`

              Stores a unique identifier for any reasoning associated with this text content.

            - `type?: "text"`

              The type of the message.

              - `"text"`

          - `ImageContent`

            - `source: URLImage | Base64Image | LettaImage`

              The source of the image.

              - `URLImage`

                - `url: string`

                  The URL of the image.

                - `type?: "url"`

                  The source type for the image.

                  - `"url"`

              - `Base64Image`

                - `data: string`

                  The base64 encoded image data.

                - `media_type: string`

                  The media type for the image.

                - `detail?: string | null`

                  What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

                - `type?: "base64"`

                  The source type for the image.

                  - `"base64"`

              - `LettaImage`

                - `file_id: string`

                  The unique identifier of the image file persisted in storage.

                - `data?: string | null`

                  The base64 encoded image data.

                - `detail?: string | null`

                  What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

                - `media_type?: string | null`

                  The media type for the image.

                - `type?: "letta"`

                  The source type for the image.

                  - `"letta"`

            - `type?: "image"`

              The type of the message.

              - `"image"`

      - `stderr?: Array<string> | null`

        Captured stderr from the tool invocation

      - `stdout?: Array<string> | null`

        Captured stdout (e.g. prints, logs) from the tool invocation

      - `tool_call_id?: unknown`

        The ID for the tool call

  - `approve?: boolean | null`

    Whether tool call is approved.

  - `batch_item_id?: string | null`

    The id of the LLMBatchItem that this message is associated with

  - `content?: Array<TextContent | ImageContent | ToolCallContent | 5 more> | null`

    The content of the message.

    - `TextContent`

    - `ImageContent`

    - `ToolCallContent`

      - `id: string`

        A unique identifier for this specific tool call instance.

      - `input: Record<string, unknown>`

        The parameters being passed to the tool, structured as a dictionary of parameter names to values.

      - `name: string`

        The name of the tool being called.

      - `signature?: string | null`

        Stores a unique identifier for any reasoning associated with this tool call.

      - `type?: "tool_call"`

        Indicates this content represents a tool call event.

        - `"tool_call"`

    - `ToolReturnContent`

      - `content: string`

        The content returned by the tool execution.

      - `is_error: boolean`

        Indicates whether the tool execution resulted in an error.

      - `tool_call_id: string`

        References the ID of the ToolCallContent that initiated this tool call.

      - `type?: "tool_return"`

        Indicates this content represents a tool return event.

        - `"tool_return"`

    - `ReasoningContent`

      Sent via the Anthropic Messages API

      - `is_native: boolean`

        Whether the reasoning content was generated by a reasoner model that processed this step.

      - `reasoning: string`

        The intermediate reasoning or thought process content.

      - `signature?: string | null`

        A unique identifier for this reasoning step.

      - `type?: "reasoning"`

        Indicates this is a reasoning/intermediate step.

        - `"reasoning"`

    - `RedactedReasoningContent`

      Sent via the Anthropic Messages API

      - `data: string`

        The redacted or filtered intermediate reasoning content.

      - `type?: "redacted_reasoning"`

        Indicates this is a redacted thinking step.

        - `"redacted_reasoning"`

    - `OmittedReasoningContent`

      A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

      - `signature?: string | null`

        A unique identifier for this reasoning step.

      - `type?: "omitted_reasoning"`

        Indicates this is an omitted reasoning step.

        - `"omitted_reasoning"`

    - `SummarizedReasoningContent`

      The style of reasoning content returned by the OpenAI Responses API

      - `id: string`

        The unique identifier for this reasoning step.

      - `summary: Array<Summary>`

        Summaries of the reasoning content.

        - `index: number`

          The index of the summary part.

        - `text: string`

          The text of the summary part.

      - `encrypted_content?: string`

        The encrypted reasoning content.

      - `type?: "summarized_reasoning"`

        Indicates this is a summarized reasoning step.

        - `"summarized_reasoning"`

  - `conversation_id?: string | null`

    The conversation this message belongs to

  - `created_at?: string`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `denial_reason?: string | null`

    The reason the tool call request was denied.

  - `group_id?: string | null`

    The multi-agent group that the message was sent in

  - `is_err?: boolean | null`

    Whether this message is part of an error step. Used only for debugging purposes.

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `model?: string | null`

    The model used to make the function call.

  - `name?: string | null`

    For role user/assistant: the (optional) name of the participant. For role tool/function: the name of the function called.

  - `otid?: string | null`

    The offline threading id associated with this message

  - `run_id?: string | null`

    The id of the run that this message was created in.

  - `sender_id?: string | null`

    The id of the sender of the message, can be an identity id or agent id

  - `step_id?: string | null`

    The id of the step that this message was created in.

  - `tool_call_id?: string | null`

    The ID of the tool call. Only applicable for role tool.

  - `tool_calls?: Array<ToolCall> | null`

    The list of tool calls requested. Only applicable for role assistant.

    - `id: string`

    - `function: Function`

      The function that the model called.

      - `arguments: string`

      - `name: string`

    - `type: "function"`

      - `"function"`

  - `tool_returns?: Array<ToolReturn> | null`

    Tool execution return information for prior tool calls

    - `status: "success" | "error"`

      The status of the tool call

      - `"success"`

      - `"error"`

    - `func_response?: string | Array<TextContent | ImageContent> | null`

      The function response - either a string or list of content parts (text/image)

      - `string`

      - `Array<TextContent | ImageContent>`

        - `TextContent`

        - `ImageContent`

    - `stderr?: Array<string> | null`

      Captured stderr from the tool invocation

    - `stdout?: Array<string> | null`

      Captured stdout (e.g. prints, logs) from the tool invocation

    - `tool_call_id?: unknown`

      The ID for the tool call

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Job Status

- `JobStatus = "created" | "running" | "completed" | 4 more`

  Status of the job.

  - `"created"`

  - `"running"`

  - `"completed"`

  - `"failed"`

  - `"pending"`

  - `"cancelled"`

  - `"expired"`

### Job Type

- `JobType = "job" | "run" | "batch"`

  - `"job"`

  - `"run"`

  - `"batch"`

### Letta Assistant Message Content Union

- `LettaAssistantMessageContentUnion`

  - `text: string`

    The text content of the message.

  - `signature?: string | null`

    Stores a unique identifier for any reasoning associated with this text content.

  - `type?: "text"`

    The type of the message.

    - `"text"`

### Letta Request

- `LettaRequest`

  - `assistant_message_tool_kwarg?: string`

    The name of the message argument in the designated message tool. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

  - `assistant_message_tool_name?: string`

    The name of the designated message tool. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

  - `client_skills?: Array<ClientSkill> | null`

    Client-side skills available in the environment. These are rendered in the system prompt's available skills section alongside agent-scoped skills from MemFS.

    - `description: string`

      Description of what the skill does

    - `location: string`

      Path or location hint for the skill (e.g. skills/my-skill/SKILL.md)

    - `name: string`

      The name of the skill

  - `client_tools?: Array<ClientTool> | null`

    Client-side tools that the agent can call. When the agent calls a client-side tool, execution pauses and returns control to the client to execute the tool and provide the result via a ToolReturn.

    - `name: string`

      The name of the tool function

    - `description?: string | null`

      Description of what the tool does

    - `parameters?: Record<string, unknown> | null`

      JSON Schema for the function parameters

  - `enable_thinking?: string`

    If set to True, enables reasoning before responses or tool calls from the agent.

  - `include_compaction_messages?: boolean`

    If True, compaction events emit structured `SummaryMessage` and `EventMessage` types. If False (default), compaction messages are not included in the response.

  - `include_return_message_types?: Array<MessageType> | null`

    Only return specified message types in the response. If `None` (default) returns all messages.

    - `"system_message"`

    - `"user_message"`

    - `"assistant_message"`

    - `"reasoning_message"`

    - `"hidden_reasoning_message"`

    - `"tool_call_message"`

    - `"tool_return_message"`

    - `"approval_request_message"`

    - `"approval_response_message"`

    - `"summary_message"`

    - `"event_message"`

  - `input?: string | Array<TextContent | ImageContent | ToolCallContent | 5 more> | null`

    Syntactic sugar for a single user message. Equivalent to messages=[{'role': 'user', 'content': input}].

    - `string`

    - `Array<TextContent | ImageContent | ToolCallContent | 5 more>`

      - `TextContent`

        - `text: string`

          The text content of the message.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this text content.

        - `type?: "text"`

          The type of the message.

          - `"text"`

      - `ImageContent`

        - `source: URLImage | Base64Image | LettaImage`

          The source of the image.

          - `URLImage`

            - `url: string`

              The URL of the image.

            - `type?: "url"`

              The source type for the image.

              - `"url"`

          - `Base64Image`

            - `data: string`

              The base64 encoded image data.

            - `media_type: string`

              The media type for the image.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `type?: "base64"`

              The source type for the image.

              - `"base64"`

          - `LettaImage`

            - `file_id: string`

              The unique identifier of the image file persisted in storage.

            - `data?: string | null`

              The base64 encoded image data.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `media_type?: string | null`

              The media type for the image.

            - `type?: "letta"`

              The source type for the image.

              - `"letta"`

        - `type?: "image"`

          The type of the message.

          - `"image"`

      - `ToolCallContent`

        - `id: string`

          A unique identifier for this specific tool call instance.

        - `input: Record<string, unknown>`

          The parameters being passed to the tool, structured as a dictionary of parameter names to values.

        - `name: string`

          The name of the tool being called.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this tool call.

        - `type?: "tool_call"`

          Indicates this content represents a tool call event.

          - `"tool_call"`

      - `ToolReturnContent`

        - `content: string`

          The content returned by the tool execution.

        - `is_error: boolean`

          Indicates whether the tool execution resulted in an error.

        - `tool_call_id: string`

          References the ID of the ToolCallContent that initiated this tool call.

        - `type?: "tool_return"`

          Indicates this content represents a tool return event.

          - `"tool_return"`

      - `ReasoningContent`

        Sent via the Anthropic Messages API

        - `is_native: boolean`

          Whether the reasoning content was generated by a reasoner model that processed this step.

        - `reasoning: string`

          The intermediate reasoning or thought process content.

        - `signature?: string | null`

          A unique identifier for this reasoning step.

        - `type?: "reasoning"`

          Indicates this is a reasoning/intermediate step.

          - `"reasoning"`

      - `RedactedReasoningContent`

        Sent via the Anthropic Messages API

        - `data: string`

          The redacted or filtered intermediate reasoning content.

        - `type?: "redacted_reasoning"`

          Indicates this is a redacted thinking step.

          - `"redacted_reasoning"`

      - `OmittedReasoningContent`

        A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

        - `signature?: string | null`

          A unique identifier for this reasoning step.

        - `type?: "omitted_reasoning"`

          Indicates this is an omitted reasoning step.

          - `"omitted_reasoning"`

      - `SummarizedReasoningContent`

        The style of reasoning content returned by the OpenAI Responses API

        - `id: string`

          The unique identifier for this reasoning step.

        - `summary: Array<Summary>`

          Summaries of the reasoning content.

          - `index: number`

            The index of the summary part.

          - `text: string`

            The text of the summary part.

        - `encrypted_content?: string`

          The encrypted reasoning content.

        - `type?: "summarized_reasoning"`

          Indicates this is a summarized reasoning step.

          - `"summarized_reasoning"`

  - `max_steps?: number`

    Maximum number of steps the agent should take to process the request.

  - `messages?: Array<MessageCreate | ApprovalCreate | ToolReturnCreate> | null`

    The messages to be sent to the agent.

    - `MessageCreate`

      Request to create a message

      - `content: Array<LettaMessageContentUnion> | string`

        The content of the message.

        - `Array<LettaMessageContentUnion>`

          - `TextContent`

          - `ImageContent`

          - `ToolCallContent`

          - `ToolReturnContent`

          - `ReasoningContent`

            Sent via the Anthropic Messages API

          - `RedactedReasoningContent`

            Sent via the Anthropic Messages API

          - `OmittedReasoningContent`

            A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

        - `string`

      - `role: "user" | "system" | "assistant"`

        The role of the participant.

        - `"user"`

        - `"system"`

        - `"assistant"`

      - `batch_item_id?: string | null`

        The id of the LLMBatchItem that this message is associated with

      - `group_id?: string | null`

        The multi-agent group that the message was sent in

      - `name?: string | null`

        The name of the participant.

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `sender_id?: string | null`

        The id of the sender of the message, can be an identity id or agent id

      - `type?: "message" | null`

        The message type to be created.

        - `"message"`

    - `ApprovalCreate`

      Input to approve or deny a tool call request

      - `approval_request_id?: string | null`

        The message ID of the approval request

      - `approvals?: Array<ApprovalReturn | ToolReturn> | null`

        The list of approval responses

        - `ApprovalReturn`

          - `approve: boolean`

            Whether the tool has been approved

          - `tool_call_id: string`

            The ID of the tool call that corresponds to this approval

          - `reason?: string | null`

            An optional explanation for the provided approval status

          - `type?: "approval"`

            The message type to be created.

            - `"approval"`

        - `ToolReturn`

          - `status: "success" | "error"`

            - `"success"`

            - `"error"`

          - `tool_call_id: string`

          - `tool_return: Array<TextContent | ImageContent> | string`

            The tool return value - either a string or list of content parts (text/image)

            - `Array<TextContent | ImageContent>`

              - `TextContent`

              - `ImageContent`

            - `string`

          - `stderr?: Array<string> | null`

          - `stdout?: Array<string> | null`

          - `type?: "tool"`

            The message type to be created.

            - `"tool"`

      - `approve?: boolean | null`

        Whether the tool has been approved

      - `group_id?: string | null`

        The multi-agent group that the message was sent in

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `reason?: string | null`

        An optional explanation for the provided approval status

      - `type?: "approval"`

        The message type to be created.

        - `"approval"`

    - `ToolReturnCreate`

      Submit tool return(s) from client-side tool execution.

      This is the preferred way to send tool results back to the agent after
      client-side tool execution. It is equivalent to sending an ApprovalCreate
      with tool return approvals, but provides a cleaner API for the common case.

      - `tool_returns: Array<ToolReturn>`

        List of tool returns from client-side execution

        - `status: "success" | "error"`

        - `tool_call_id: string`

        - `tool_return: Array<TextContent | ImageContent> | string`

          The tool return value - either a string or list of content parts (text/image)

        - `stderr?: Array<string> | null`

        - `stdout?: Array<string> | null`

        - `type?: "tool"`

          The message type to be created.

      - `group_id?: string | null`

        The multi-agent group that the message was sent in

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `type?: "tool_return"`

        The message type to be created.

        - `"tool_return"`

  - `override_model?: string | null`

    Model handle to use for this request instead of the agent's default model. This allows sending a message to a different model without changing the agent's configuration.

  - `override_system?: string | null`

    Optional per-request system prompt override. When set, this is passed directly to the underlying LLM request and bypasses the persisted/compiled system message for that request.

  - `return_logprobs?: boolean`

    If True, returns log probabilities of the output tokens in the response. Useful for RL training. Only supported for OpenAI-compatible providers (including SGLang).

  - `return_token_ids?: boolean`

    If True, returns token IDs and logprobs for ALL LLM generations in the agent step, not just the last one. Uses SGLang native /generate endpoint. Returns 'turns' field with TurnTokenData for each assistant/tool turn. Required for proper multi-turn RL training with loss masking.

  - `top_logprobs?: number | null`

    Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

  - `use_assistant_message?: boolean`

    Whether the server should parse specific tool call arguments (default `send_message`) as `AssistantMessage` objects. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

### Letta Response

- `LettaResponse`

  Response object from an agent interaction, consisting of the new messages generated by the agent and usage statistics.
  The type of the returned messages can be either `Message` or `LettaMessage`, depending on what was specified in the request.

  Attributes:
  messages (List[Union[Message, LettaMessage]]): The messages returned by the agent.
  usage (LettaUsageStatistics): The usage statistics

  - `messages: Array<Message>`

    The messages returned by the agent.

    - `SystemMessage`

      A message generated by the system. Never streamed back on a response, only used for cursor pagination.

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      content (str): The message content sent by the system

      - `id: string`

      - `content: string`

        The message content sent by the system

      - `date: string`

      - `is_err?: boolean | null`

      - `message_type?: "system_message"`

        The type of the message.

        - `"system_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

    - `UserMessage`

      A message sent by the user. Never streamed back on a response, only used for cursor pagination.

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      content (Union[str, List[LettaUserMessageContentUnion]]): The message content sent by the user (can be a string or an array of multi-modal content parts)

      - `id: string`

      - `content: Array<LettaUserMessageContentUnion> | string`

        The message content sent by the user (can be a string or an array of multi-modal content parts)

        - `Array<LettaUserMessageContentUnion>`

          - `TextContent`

            - `text: string`

              The text content of the message.

            - `signature?: string | null`

              Stores a unique identifier for any reasoning associated with this text content.

            - `type?: "text"`

              The type of the message.

              - `"text"`

          - `ImageContent`

            - `source: URLImage | Base64Image | LettaImage`

              The source of the image.

              - `URLImage`

                - `url: string`

                  The URL of the image.

                - `type?: "url"`

                  The source type for the image.

                  - `"url"`

              - `Base64Image`

                - `data: string`

                  The base64 encoded image data.

                - `media_type: string`

                  The media type for the image.

                - `detail?: string | null`

                  What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

                - `type?: "base64"`

                  The source type for the image.

                  - `"base64"`

              - `LettaImage`

                - `file_id: string`

                  The unique identifier of the image file persisted in storage.

                - `data?: string | null`

                  The base64 encoded image data.

                - `detail?: string | null`

                  What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

                - `media_type?: string | null`

                  The media type for the image.

                - `type?: "letta"`

                  The source type for the image.

                  - `"letta"`

            - `type?: "image"`

              The type of the message.

              - `"image"`

        - `string`

      - `date: string`

      - `is_err?: boolean | null`

      - `message_type?: "user_message"`

        The type of the message.

        - `"user_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

    - `ReasoningMessage`

      Representation of an agent's internal reasoning.

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      source (Literal["reasoner_model", "non_reasoner_model"]): Whether the reasoning
      content was generated natively by a reasoner model or derived via prompting
      reasoning (str): The internal reasoning of the agent
      signature (Optional[str]): The model-generated signature of the reasoning step

      - `id: string`

      - `date: string`

      - `reasoning: string`

      - `is_err?: boolean | null`

      - `message_type?: "reasoning_message"`

        The type of the message.

        - `"reasoning_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `signature?: string | null`

      - `source?: "reasoner_model" | "non_reasoner_model"`

        - `"reasoner_model"`

        - `"non_reasoner_model"`

      - `step_id?: string | null`

    - `HiddenReasoningMessage`

      Representation of an agent's internal reasoning where reasoning content
      has been hidden from the response.

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      state (Literal["redacted", "omitted"]): Whether the reasoning
      content was redacted by the provider or simply omitted by the API
      hidden_reasoning (Optional[str]): The internal reasoning of the agent

      - `id: string`

      - `date: string`

      - `state: "redacted" | "omitted"`

        - `"redacted"`

        - `"omitted"`

      - `hidden_reasoning?: string | null`

      - `is_err?: boolean | null`

      - `message_type?: "hidden_reasoning_message"`

        The type of the message.

        - `"hidden_reasoning_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

    - `ToolCallMessage`

      A message representing a request to call a tool (generated by the LLM to trigger tool execution).

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      tool_call (Union[ToolCall, ToolCallDelta]): The tool call

      - `id: string`

      - `date: string`

      - `tool_call: ToolCall | ToolCallDelta`

        - `ToolCall`

          - `arguments: string`

          - `name: string`

          - `tool_call_id: string`

        - `ToolCallDelta`

          - `arguments?: string | null`

          - `name?: string | null`

          - `tool_call_id?: string | null`

      - `is_err?: boolean | null`

      - `message_type?: "tool_call_message"`

        The type of the message.

        - `"tool_call_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

      - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

        - `Array<ToolCall>`

          - `arguments: string`

          - `name: string`

          - `tool_call_id: string`

        - `ToolCallDelta`

    - `ToolReturnMessage`

      A message representing the return value of a tool call (generated by Letta executing the requested tool).

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      tool_return (str): The return value of the tool (deprecated, use tool_returns)
      status (Literal["success", "error"]): The status of the tool call (deprecated, use tool_returns)
      tool_call_id (str): A unique identifier for the tool call that generated this message (deprecated, use tool_returns)
      stdout (Optional[List(str)]): Captured stdout (e.g. prints, logs) from the tool invocation (deprecated, use tool_returns)
      stderr (Optional[List(str)]): Captured stderr from the tool invocation (deprecated, use tool_returns)
      tool_returns (Optional[List[ToolReturn]]): List of tool returns for multi-tool support

      - `id: string`

      - `date: string`

      - `status: "success" | "error"`

        - `"success"`

        - `"error"`

      - `tool_call_id: string`

      - `tool_return: string`

      - `is_err?: boolean | null`

      - `message_type?: "tool_return_message"`

        The type of the message.

        - `"tool_return_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `stderr?: Array<string> | null`

      - `stdout?: Array<string> | null`

      - `step_id?: string | null`

      - `tool_returns?: Array<ToolReturn> | null`

        - `status: "success" | "error"`

          - `"success"`

          - `"error"`

        - `tool_call_id: string`

        - `tool_return: Array<TextContent | ImageContent> | string`

          The tool return value - either a string or list of content parts (text/image)

          - `Array<TextContent | ImageContent>`

            - `TextContent`

            - `ImageContent`

          - `string`

        - `stderr?: Array<string> | null`

        - `stdout?: Array<string> | null`

        - `type?: "tool"`

          The message type to be created.

          - `"tool"`

    - `AssistantMessage`

      A message sent by the LLM in response to user input. Used in the LLM context.

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      content (Union[str, List[LettaAssistantMessageContentUnion]]): The message content sent by the agent (can be a string or an array of content parts)

      - `id: string`

      - `content: Array<LettaAssistantMessageContentUnion> | string`

        The message content sent by the agent (can be a string or an array of content parts)

        - `Array<LettaAssistantMessageContentUnion>`

          - `text: string`

            The text content of the message.

          - `signature?: string | null`

            Stores a unique identifier for any reasoning associated with this text content.

          - `type?: "text"`

            The type of the message.

            - `"text"`

        - `string`

      - `date: string`

      - `is_err?: boolean | null`

      - `message_type?: "assistant_message"`

        The type of the message.

        - `"assistant_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

    - `ApprovalRequestMessage`

      A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      tool_call (ToolCall): The tool call

      - `id: string`

      - `date: string`

      - `tool_call: ToolCall | ToolCallDelta`

        The tool call that has been requested by the llm to run

        - `ToolCall`

        - `ToolCallDelta`

      - `is_err?: boolean | null`

      - `message_type?: "approval_request_message"`

        The type of the message.

        - `"approval_request_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

      - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

        The tool calls that have been requested by the llm to run, which are pending approval

        - `Array<ToolCall>`

          - `arguments: string`

          - `name: string`

          - `tool_call_id: string`

        - `ToolCallDelta`

    - `ApprovalResponseMessage`

      A message representing a response form the user indicating whether a tool has been approved to run.

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      approve: (bool) Whether the tool has been approved
      approval_request_id: The ID of the approval request
      reason: (Optional[str]) An optional explanation for the provided approval status

      - `id: string`

      - `date: string`

      - `approval_request_id?: string | null`

        The message ID of the approval request

      - `approvals?: Array<ApprovalReturn | ToolReturn> | null`

        The list of approval responses

        - `ApprovalReturn`

          - `approve: boolean`

            Whether the tool has been approved

          - `tool_call_id: string`

            The ID of the tool call that corresponds to this approval

          - `reason?: string | null`

            An optional explanation for the provided approval status

          - `type?: "approval"`

            The message type to be created.

            - `"approval"`

        - `ToolReturn`

          - `status: "success" | "error"`

          - `tool_call_id: string`

          - `tool_return: Array<TextContent | ImageContent> | string`

            The tool return value - either a string or list of content parts (text/image)

          - `stderr?: Array<string> | null`

          - `stdout?: Array<string> | null`

          - `type?: "tool"`

            The message type to be created.

      - `approve?: boolean | null`

        Whether the tool has been approved

      - `is_err?: boolean | null`

      - `message_type?: "approval_response_message"`

        The type of the message.

        - `"approval_response_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `reason?: string | null`

        An optional explanation for the provided approval status

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

    - `SummaryMessage`

      A message representing a summary of the conversation. Sent to the LLM as a user or system message depending on the provider.

      - `id: string`

      - `date: string`

      - `summary: string`

      - `compaction_stats?: CompactionStats | null`

        Statistics about a memory compaction operation.

        - `context_window: number`

          The model's context window size

        - `messages_count_after: number`

          Number of messages after compaction

        - `messages_count_before: number`

          Number of messages before compaction

        - `trigger: string`

          What triggered the compaction (e.g., 'context_window_exceeded', 'post_step_context_check')

        - `context_tokens_after?: number | null`

          Token count after compaction (message tokens only, does not include tool definitions)

        - `context_tokens_before?: number | null`

          Token count before compaction (from LLM usage stats, includes full context sent to LLM)

      - `is_err?: boolean | null`

      - `message_type?: "summary_message"`

        - `"summary_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

    - `EventMessage`

      A message for notifying the developer that an event that has occured (e.g. a compaction). Events are NOT part of the context window.

      - `id: string`

      - `date: string`

      - `event_data: Record<string, unknown>`

      - `event_type: "compaction"`

        - `"compaction"`

      - `is_err?: boolean | null`

      - `message_type?: "event_message"`

        - `"event_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

  - `stop_reason: StopReason`

    The stop reason from Letta indicating why agent loop stopped execution.

    - `stop_reason: StopReasonType`

      The reason why execution stopped.

      - `"end_turn"`

      - `"error"`

      - `"llm_api_error"`

      - `"invalid_llm_response"`

      - `"invalid_tool_call"`

      - `"max_steps"`

      - `"max_tokens_exceeded"`

      - `"no_tool_call"`

      - `"tool_rule"`

      - `"cancelled"`

      - `"insufficient_credits"`

      - `"requires_approval"`

      - `"context_window_overflow_in_system_prompt"`

    - `message_type?: "stop_reason"`

      The type of the message.

      - `"stop_reason"`

  - `usage: Usage`

    The usage statistics of the agent.

    - `cache_write_tokens?: number | null`

      The number of input tokens written to cache (Anthropic only). None if not reported by provider.

    - `cached_input_tokens?: number | null`

      The number of input tokens served from cache. None if not reported by provider.

    - `completion_tokens?: number`

      The number of tokens generated by the agent.

    - `context_tokens?: number | null`

      Estimate of tokens currently in the context window.

    - `message_type?: "usage_statistics"`

      - `"usage_statistics"`

    - `prompt_tokens?: number`

      The number of tokens in the prompt.

    - `reasoning_tokens?: number | null`

      The number of reasoning/thinking tokens generated. None if not reported by provider.

    - `run_ids?: Array<string> | null`

      The background task run IDs associated with the agent interaction

    - `step_count?: number`

      The number of steps taken by the agent.

    - `total_tokens?: number`

      The total number of tokens processed by the agent.

  - `logprobs?: Logprobs | null`

    Log probabilities of the output tokens from the last LLM call. Only present if return_logprobs was enabled.

    - `content?: Array<Content> | null`

      - `token: string`

      - `logprob: number`

      - `top_logprobs: Array<TopLogprob>`

        - `token: string`

        - `logprob: number`

        - `bytes?: Array<number> | null`

      - `bytes?: Array<number> | null`

    - `refusal?: Array<Refusal> | null`

      - `token: string`

      - `logprob: number`

      - `top_logprobs: Array<TopLogprob>`

        - `token: string`

        - `logprob: number`

        - `bytes?: Array<number> | null`

      - `bytes?: Array<number> | null`

  - `turns?: Array<Turn> | null`

    Token data for all LLM generations in multi-turn agent interaction. Includes token IDs and logprobs for each assistant turn, plus tool result content. Only present if return_token_ids was enabled. Used for RL training with loss masking.

    - `role: "assistant" | "tool"`

      Role of this turn: 'assistant' for LLM generations (trainable), 'tool' for tool results (non-trainable).

      - `"assistant"`

      - `"tool"`

    - `content?: string | null`

      Text content. For tool turns, client tokenizes this with loss_mask=0.

    - `output_ids?: Array<number> | null`

      Token IDs from SGLang native endpoint. Only present for assistant turns.

    - `output_token_logprobs?: Array<Array<unknown>> | null`

      Logprobs from SGLang: [[logprob, token_id, top_logprob_or_null], ...]. Only present for assistant turns.

    - `tool_name?: string | null`

      Name of the tool called. Only present for tool turns.

### Letta Streaming Request

- `LettaStreamingRequest`

  - `assistant_message_tool_kwarg?: string`

    The name of the message argument in the designated message tool. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

  - `assistant_message_tool_name?: string`

    The name of the designated message tool. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

  - `background?: boolean`

    Whether to process the request in the background (only used when streaming=true).

  - `client_skills?: Array<ClientSkill> | null`

    Client-side skills available in the environment. These are rendered in the system prompt's available skills section alongside agent-scoped skills from MemFS.

    - `description: string`

      Description of what the skill does

    - `location: string`

      Path or location hint for the skill (e.g. skills/my-skill/SKILL.md)

    - `name: string`

      The name of the skill

  - `client_tools?: Array<ClientTool> | null`

    Client-side tools that the agent can call. When the agent calls a client-side tool, execution pauses and returns control to the client to execute the tool and provide the result via a ToolReturn.

    - `name: string`

      The name of the tool function

    - `description?: string | null`

      Description of what the tool does

    - `parameters?: Record<string, unknown> | null`

      JSON Schema for the function parameters

  - `enable_thinking?: string`

    If set to True, enables reasoning before responses or tool calls from the agent.

  - `include_compaction_messages?: boolean`

    If True, compaction events emit structured `SummaryMessage` and `EventMessage` types. If False (default), compaction messages are not included in the response.

  - `include_pings?: boolean`

    Whether to include periodic keepalive ping messages in the stream to prevent connection timeouts (only used when streaming=true).

  - `include_return_message_types?: Array<MessageType> | null`

    Only return specified message types in the response. If `None` (default) returns all messages.

    - `"system_message"`

    - `"user_message"`

    - `"assistant_message"`

    - `"reasoning_message"`

    - `"hidden_reasoning_message"`

    - `"tool_call_message"`

    - `"tool_return_message"`

    - `"approval_request_message"`

    - `"approval_response_message"`

    - `"summary_message"`

    - `"event_message"`

  - `input?: string | Array<TextContent | ImageContent | ToolCallContent | 5 more> | null`

    Syntactic sugar for a single user message. Equivalent to messages=[{'role': 'user', 'content': input}].

    - `string`

    - `Array<TextContent | ImageContent | ToolCallContent | 5 more>`

      - `TextContent`

        - `text: string`

          The text content of the message.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this text content.

        - `type?: "text"`

          The type of the message.

          - `"text"`

      - `ImageContent`

        - `source: URLImage | Base64Image | LettaImage`

          The source of the image.

          - `URLImage`

            - `url: string`

              The URL of the image.

            - `type?: "url"`

              The source type for the image.

              - `"url"`

          - `Base64Image`

            - `data: string`

              The base64 encoded image data.

            - `media_type: string`

              The media type for the image.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `type?: "base64"`

              The source type for the image.

              - `"base64"`

          - `LettaImage`

            - `file_id: string`

              The unique identifier of the image file persisted in storage.

            - `data?: string | null`

              The base64 encoded image data.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `media_type?: string | null`

              The media type for the image.

            - `type?: "letta"`

              The source type for the image.

              - `"letta"`

        - `type?: "image"`

          The type of the message.

          - `"image"`

      - `ToolCallContent`

        - `id: string`

          A unique identifier for this specific tool call instance.

        - `input: Record<string, unknown>`

          The parameters being passed to the tool, structured as a dictionary of parameter names to values.

        - `name: string`

          The name of the tool being called.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this tool call.

        - `type?: "tool_call"`

          Indicates this content represents a tool call event.

          - `"tool_call"`

      - `ToolReturnContent`

        - `content: string`

          The content returned by the tool execution.

        - `is_error: boolean`

          Indicates whether the tool execution resulted in an error.

        - `tool_call_id: string`

          References the ID of the ToolCallContent that initiated this tool call.

        - `type?: "tool_return"`

          Indicates this content represents a tool return event.

          - `"tool_return"`

      - `ReasoningContent`

        Sent via the Anthropic Messages API

        - `is_native: boolean`

          Whether the reasoning content was generated by a reasoner model that processed this step.

        - `reasoning: string`

          The intermediate reasoning or thought process content.

        - `signature?: string | null`

          A unique identifier for this reasoning step.

        - `type?: "reasoning"`

          Indicates this is a reasoning/intermediate step.

          - `"reasoning"`

      - `RedactedReasoningContent`

        Sent via the Anthropic Messages API

        - `data: string`

          The redacted or filtered intermediate reasoning content.

        - `type?: "redacted_reasoning"`

          Indicates this is a redacted thinking step.

          - `"redacted_reasoning"`

      - `OmittedReasoningContent`

        A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

        - `signature?: string | null`

          A unique identifier for this reasoning step.

        - `type?: "omitted_reasoning"`

          Indicates this is an omitted reasoning step.

          - `"omitted_reasoning"`

      - `SummarizedReasoningContent`

        The style of reasoning content returned by the OpenAI Responses API

        - `id: string`

          The unique identifier for this reasoning step.

        - `summary: Array<Summary>`

          Summaries of the reasoning content.

          - `index: number`

            The index of the summary part.

          - `text: string`

            The text of the summary part.

        - `encrypted_content?: string`

          The encrypted reasoning content.

        - `type?: "summarized_reasoning"`

          Indicates this is a summarized reasoning step.

          - `"summarized_reasoning"`

  - `max_steps?: number`

    Maximum number of steps the agent should take to process the request.

  - `messages?: Array<MessageCreate | ApprovalCreate | ToolReturnCreate> | null`

    The messages to be sent to the agent.

    - `MessageCreate`

      Request to create a message

      - `content: Array<LettaMessageContentUnion> | string`

        The content of the message.

        - `Array<LettaMessageContentUnion>`

          - `TextContent`

          - `ImageContent`

          - `ToolCallContent`

          - `ToolReturnContent`

          - `ReasoningContent`

            Sent via the Anthropic Messages API

          - `RedactedReasoningContent`

            Sent via the Anthropic Messages API

          - `OmittedReasoningContent`

            A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

        - `string`

      - `role: "user" | "system" | "assistant"`

        The role of the participant.

        - `"user"`

        - `"system"`

        - `"assistant"`

      - `batch_item_id?: string | null`

        The id of the LLMBatchItem that this message is associated with

      - `group_id?: string | null`

        The multi-agent group that the message was sent in

      - `name?: string | null`

        The name of the participant.

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `sender_id?: string | null`

        The id of the sender of the message, can be an identity id or agent id

      - `type?: "message" | null`

        The message type to be created.

        - `"message"`

    - `ApprovalCreate`

      Input to approve or deny a tool call request

      - `approval_request_id?: string | null`

        The message ID of the approval request

      - `approvals?: Array<ApprovalReturn | ToolReturn> | null`

        The list of approval responses

        - `ApprovalReturn`

          - `approve: boolean`

            Whether the tool has been approved

          - `tool_call_id: string`

            The ID of the tool call that corresponds to this approval

          - `reason?: string | null`

            An optional explanation for the provided approval status

          - `type?: "approval"`

            The message type to be created.

            - `"approval"`

        - `ToolReturn`

          - `status: "success" | "error"`

            - `"success"`

            - `"error"`

          - `tool_call_id: string`

          - `tool_return: Array<TextContent | ImageContent> | string`

            The tool return value - either a string or list of content parts (text/image)

            - `Array<TextContent | ImageContent>`

              - `TextContent`

              - `ImageContent`

            - `string`

          - `stderr?: Array<string> | null`

          - `stdout?: Array<string> | null`

          - `type?: "tool"`

            The message type to be created.

            - `"tool"`

      - `approve?: boolean | null`

        Whether the tool has been approved

      - `group_id?: string | null`

        The multi-agent group that the message was sent in

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `reason?: string | null`

        An optional explanation for the provided approval status

      - `type?: "approval"`

        The message type to be created.

        - `"approval"`

    - `ToolReturnCreate`

      Submit tool return(s) from client-side tool execution.

      This is the preferred way to send tool results back to the agent after
      client-side tool execution. It is equivalent to sending an ApprovalCreate
      with tool return approvals, but provides a cleaner API for the common case.

      - `tool_returns: Array<ToolReturn>`

        List of tool returns from client-side execution

        - `status: "success" | "error"`

        - `tool_call_id: string`

        - `tool_return: Array<TextContent | ImageContent> | string`

          The tool return value - either a string or list of content parts (text/image)

        - `stderr?: Array<string> | null`

        - `stdout?: Array<string> | null`

        - `type?: "tool"`

          The message type to be created.

      - `group_id?: string | null`

        The multi-agent group that the message was sent in

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `type?: "tool_return"`

        The message type to be created.

        - `"tool_return"`

  - `override_model?: string | null`

    Model handle to use for this request instead of the agent's default model. This allows sending a message to a different model without changing the agent's configuration.

  - `override_system?: string | null`

    Optional per-request system prompt override. When set, this is passed directly to the underlying LLM request and bypasses the persisted/compiled system message for that request.

  - `return_logprobs?: boolean`

    If True, returns log probabilities of the output tokens in the response. Useful for RL training. Only supported for OpenAI-compatible providers (including SGLang).

  - `return_token_ids?: boolean`

    If True, returns token IDs and logprobs for ALL LLM generations in the agent step, not just the last one. Uses SGLang native /generate endpoint. Returns 'turns' field with TurnTokenData for each assistant/tool turn. Required for proper multi-turn RL training with loss masking.

  - `stream_tokens?: boolean`

    Flag to determine if individual tokens should be streamed, rather than streaming per step (only used when streaming=true).

  - `streaming?: boolean`

    If True, returns a streaming response (Server-Sent Events). If False (default), returns a complete response.

  - `top_logprobs?: number | null`

    Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

  - `use_assistant_message?: boolean`

    Whether the server should parse specific tool call arguments (default `send_message`) as `AssistantMessage` objects. Still supported for legacy agent types, but deprecated for letta_v1_agent onward.

### Letta Streaming Response

- `LettaStreamingResponse = SystemMessage | UserMessage | ReasoningMessage | 10 more`

  Streaming response type for Server-Sent Events (SSE) endpoints.
  Each event in the stream will be one of these types.

  - `SystemMessage`

    A message generated by the system. Never streamed back on a response, only used for cursor pagination.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    content (str): The message content sent by the system

    - `id: string`

    - `content: string`

      The message content sent by the system

    - `date: string`

    - `is_err?: boolean | null`

    - `message_type?: "system_message"`

      The type of the message.

      - `"system_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `UserMessage`

    A message sent by the user. Never streamed back on a response, only used for cursor pagination.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    content (Union[str, List[LettaUserMessageContentUnion]]): The message content sent by the user (can be a string or an array of multi-modal content parts)

    - `id: string`

    - `content: Array<LettaUserMessageContentUnion> | string`

      The message content sent by the user (can be a string or an array of multi-modal content parts)

      - `Array<LettaUserMessageContentUnion>`

        - `TextContent`

          - `text: string`

            The text content of the message.

          - `signature?: string | null`

            Stores a unique identifier for any reasoning associated with this text content.

          - `type?: "text"`

            The type of the message.

            - `"text"`

        - `ImageContent`

          - `source: URLImage | Base64Image | LettaImage`

            The source of the image.

            - `URLImage`

              - `url: string`

                The URL of the image.

              - `type?: "url"`

                The source type for the image.

                - `"url"`

            - `Base64Image`

              - `data: string`

                The base64 encoded image data.

              - `media_type: string`

                The media type for the image.

              - `detail?: string | null`

                What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

              - `type?: "base64"`

                The source type for the image.

                - `"base64"`

            - `LettaImage`

              - `file_id: string`

                The unique identifier of the image file persisted in storage.

              - `data?: string | null`

                The base64 encoded image data.

              - `detail?: string | null`

                What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

              - `media_type?: string | null`

                The media type for the image.

              - `type?: "letta"`

                The source type for the image.

                - `"letta"`

          - `type?: "image"`

            The type of the message.

            - `"image"`

      - `string`

    - `date: string`

    - `is_err?: boolean | null`

    - `message_type?: "user_message"`

      The type of the message.

      - `"user_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `ReasoningMessage`

    Representation of an agent's internal reasoning.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    source (Literal["reasoner_model", "non_reasoner_model"]): Whether the reasoning
    content was generated natively by a reasoner model or derived via prompting
    reasoning (str): The internal reasoning of the agent
    signature (Optional[str]): The model-generated signature of the reasoning step

    - `id: string`

    - `date: string`

    - `reasoning: string`

    - `is_err?: boolean | null`

    - `message_type?: "reasoning_message"`

      The type of the message.

      - `"reasoning_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `signature?: string | null`

    - `source?: "reasoner_model" | "non_reasoner_model"`

      - `"reasoner_model"`

      - `"non_reasoner_model"`

    - `step_id?: string | null`

  - `HiddenReasoningMessage`

    Representation of an agent's internal reasoning where reasoning content
    has been hidden from the response.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    state (Literal["redacted", "omitted"]): Whether the reasoning
    content was redacted by the provider or simply omitted by the API
    hidden_reasoning (Optional[str]): The internal reasoning of the agent

    - `id: string`

    - `date: string`

    - `state: "redacted" | "omitted"`

      - `"redacted"`

      - `"omitted"`

    - `hidden_reasoning?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "hidden_reasoning_message"`

      The type of the message.

      - `"hidden_reasoning_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `ToolCallMessage`

    A message representing a request to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (Union[ToolCall, ToolCallDelta]): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "tool_call_message"`

      The type of the message.

      - `"tool_call_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `ToolReturnMessage`

    A message representing the return value of a tool call (generated by Letta executing the requested tool).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_return (str): The return value of the tool (deprecated, use tool_returns)
    status (Literal["success", "error"]): The status of the tool call (deprecated, use tool_returns)
    tool_call_id (str): A unique identifier for the tool call that generated this message (deprecated, use tool_returns)
    stdout (Optional[List(str)]): Captured stdout (e.g. prints, logs) from the tool invocation (deprecated, use tool_returns)
    stderr (Optional[List(str)]): Captured stderr from the tool invocation (deprecated, use tool_returns)
    tool_returns (Optional[List[ToolReturn]]): List of tool returns for multi-tool support

    - `id: string`

    - `date: string`

    - `status: "success" | "error"`

      - `"success"`

      - `"error"`

    - `tool_call_id: string`

    - `tool_return: string`

    - `is_err?: boolean | null`

    - `message_type?: "tool_return_message"`

      The type of the message.

      - `"tool_return_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `stderr?: Array<string> | null`

    - `stdout?: Array<string> | null`

    - `step_id?: string | null`

    - `tool_returns?: Array<ToolReturn> | null`

      - `status: "success" | "error"`

        - `"success"`

        - `"error"`

      - `tool_call_id: string`

      - `tool_return: Array<TextContent | ImageContent> | string`

        The tool return value - either a string or list of content parts (text/image)

        - `Array<TextContent | ImageContent>`

          - `TextContent`

          - `ImageContent`

        - `string`

      - `stderr?: Array<string> | null`

      - `stdout?: Array<string> | null`

      - `type?: "tool"`

        The message type to be created.

        - `"tool"`

  - `AssistantMessage`

    A message sent by the LLM in response to user input. Used in the LLM context.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    content (Union[str, List[LettaAssistantMessageContentUnion]]): The message content sent by the agent (can be a string or an array of content parts)

    - `id: string`

    - `content: Array<LettaAssistantMessageContentUnion> | string`

      The message content sent by the agent (can be a string or an array of content parts)

      - `Array<LettaAssistantMessageContentUnion>`

        - `text: string`

          The text content of the message.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this text content.

        - `type?: "text"`

          The type of the message.

          - `"text"`

      - `string`

    - `date: string`

    - `is_err?: boolean | null`

    - `message_type?: "assistant_message"`

      The type of the message.

      - `"assistant_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `ApprovalRequestMessage`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

      - `ToolCallDelta`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `ApprovalResponseMessage`

    A message representing a response form the user indicating whether a tool has been approved to run.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    approve: (bool) Whether the tool has been approved
    approval_request_id: The ID of the approval request
    reason: (Optional[str]) An optional explanation for the provided approval status

    - `id: string`

    - `date: string`

    - `approval_request_id?: string | null`

      The message ID of the approval request

    - `approvals?: Array<ApprovalReturn | ToolReturn> | null`

      The list of approval responses

      - `ApprovalReturn`

        - `approve: boolean`

          Whether the tool has been approved

        - `tool_call_id: string`

          The ID of the tool call that corresponds to this approval

        - `reason?: string | null`

          An optional explanation for the provided approval status

        - `type?: "approval"`

          The message type to be created.

          - `"approval"`

      - `ToolReturn`

        - `status: "success" | "error"`

        - `tool_call_id: string`

        - `tool_return: Array<TextContent | ImageContent> | string`

          The tool return value - either a string or list of content parts (text/image)

        - `stderr?: Array<string> | null`

        - `stdout?: Array<string> | null`

        - `type?: "tool"`

          The message type to be created.

    - `approve?: boolean | null`

      Whether the tool has been approved

    - `is_err?: boolean | null`

    - `message_type?: "approval_response_message"`

      The type of the message.

      - `"approval_response_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `reason?: string | null`

      An optional explanation for the provided approval status

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `LettaPing`

    A ping message used as a keepalive to prevent SSE streams from timing out during long running requests.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format

    - `id: string`

    - `date: string`

    - `is_err?: boolean | null`

    - `message_type?: "ping"`

      The type of the message. Ping messages are a keep-alive to prevent SSE streams from timing out during long running requests.

      - `"ping"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `LettaErrorMessage`

    Error messages are used to notify the client of an error that occurred during the agent's execution.

    - `error_type: string`

      The type of error.

    - `message: string`

      The error message.

    - `message_type: "error_message"`

      The type of the message.

      - `"error_message"`

    - `run_id: string`

      The ID of the run.

    - `detail?: string`

      An optional error detail.

    - `seq_id?: number`

      The sequence ID for cursor-based pagination.

  - `LettaStopReason`

    The stop reason from Letta indicating why agent loop stopped execution.

    - `stop_reason: StopReasonType`

      The reason why execution stopped.

      - `"end_turn"`

      - `"error"`

      - `"llm_api_error"`

      - `"invalid_llm_response"`

      - `"invalid_tool_call"`

      - `"max_steps"`

      - `"max_tokens_exceeded"`

      - `"no_tool_call"`

      - `"tool_rule"`

      - `"cancelled"`

      - `"insufficient_credits"`

      - `"requires_approval"`

      - `"context_window_overflow_in_system_prompt"`

    - `message_type?: "stop_reason"`

      The type of the message.

      - `"stop_reason"`

  - `LettaUsageStatistics`

    Usage statistics for the agent interaction.

    Attributes:
    completion_tokens (int): The number of tokens generated by the agent.
    prompt_tokens (int): The number of tokens in the prompt.
    total_tokens (int): The total number of tokens processed by the agent.
    step_count (int): The number of steps taken by the agent.
    cached_input_tokens (Optional[int]): The number of input tokens served from cache. None if not reported.
    cache_write_tokens (Optional[int]): The number of input tokens written to cache. None if not reported.
    reasoning_tokens (Optional[int]): The number of reasoning/thinking tokens generated. None if not reported.

    - `cache_write_tokens?: number | null`

      The number of input tokens written to cache (Anthropic only). None if not reported by provider.

    - `cached_input_tokens?: number | null`

      The number of input tokens served from cache. None if not reported by provider.

    - `completion_tokens?: number`

      The number of tokens generated by the agent.

    - `context_tokens?: number | null`

      Estimate of tokens currently in the context window.

    - `message_type?: "usage_statistics"`

      - `"usage_statistics"`

    - `prompt_tokens?: number`

      The number of tokens in the prompt.

    - `reasoning_tokens?: number | null`

      The number of reasoning/thinking tokens generated. None if not reported by provider.

    - `run_ids?: Array<string> | null`

      The background task run IDs associated with the agent interaction

    - `step_count?: number`

      The number of steps taken by the agent.

    - `total_tokens?: number`

      The total number of tokens processed by the agent.

### Letta User Message Content Union

- `LettaUserMessageContentUnion = TextContent | ImageContent`

  - `TextContent`

    - `text: string`

      The text content of the message.

    - `signature?: string | null`

      Stores a unique identifier for any reasoning associated with this text content.

    - `type?: "text"`

      The type of the message.

      - `"text"`

  - `ImageContent`

    - `source: URLImage | Base64Image | LettaImage`

      The source of the image.

      - `URLImage`

        - `url: string`

          The URL of the image.

        - `type?: "url"`

          The source type for the image.

          - `"url"`

      - `Base64Image`

        - `data: string`

          The base64 encoded image data.

        - `media_type: string`

          The media type for the image.

        - `detail?: string | null`

          What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

        - `type?: "base64"`

          The source type for the image.

          - `"base64"`

      - `LettaImage`

        - `file_id: string`

          The unique identifier of the image file persisted in storage.

        - `data?: string | null`

          The base64 encoded image data.

        - `detail?: string | null`

          What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

        - `media_type?: string | null`

          The media type for the image.

        - `type?: "letta"`

          The source type for the image.

          - `"letta"`

    - `type?: "image"`

      The type of the message.

      - `"image"`

### Message

- `Message = SystemMessage | UserMessage | ReasoningMessage | 8 more`

  A message generated by the system. Never streamed back on a response, only used for cursor pagination.

  Args:
  id (str): The ID of the message
  date (datetime): The date the message was created in ISO format
  name (Optional[str]): The name of the sender of the message
  content (str): The message content sent by the system

  - `SystemMessage`

    A message generated by the system. Never streamed back on a response, only used for cursor pagination.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    content (str): The message content sent by the system

    - `id: string`

    - `content: string`

      The message content sent by the system

    - `date: string`

    - `is_err?: boolean | null`

    - `message_type?: "system_message"`

      The type of the message.

      - `"system_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `UserMessage`

    A message sent by the user. Never streamed back on a response, only used for cursor pagination.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    content (Union[str, List[LettaUserMessageContentUnion]]): The message content sent by the user (can be a string or an array of multi-modal content parts)

    - `id: string`

    - `content: Array<LettaUserMessageContentUnion> | string`

      The message content sent by the user (can be a string or an array of multi-modal content parts)

      - `Array<LettaUserMessageContentUnion>`

        - `TextContent`

          - `text: string`

            The text content of the message.

          - `signature?: string | null`

            Stores a unique identifier for any reasoning associated with this text content.

          - `type?: "text"`

            The type of the message.

            - `"text"`

        - `ImageContent`

          - `source: URLImage | Base64Image | LettaImage`

            The source of the image.

            - `URLImage`

              - `url: string`

                The URL of the image.

              - `type?: "url"`

                The source type for the image.

                - `"url"`

            - `Base64Image`

              - `data: string`

                The base64 encoded image data.

              - `media_type: string`

                The media type for the image.

              - `detail?: string | null`

                What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

              - `type?: "base64"`

                The source type for the image.

                - `"base64"`

            - `LettaImage`

              - `file_id: string`

                The unique identifier of the image file persisted in storage.

              - `data?: string | null`

                The base64 encoded image data.

              - `detail?: string | null`

                What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

              - `media_type?: string | null`

                The media type for the image.

              - `type?: "letta"`

                The source type for the image.

                - `"letta"`

          - `type?: "image"`

            The type of the message.

            - `"image"`

      - `string`

    - `date: string`

    - `is_err?: boolean | null`

    - `message_type?: "user_message"`

      The type of the message.

      - `"user_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `ReasoningMessage`

    Representation of an agent's internal reasoning.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    source (Literal["reasoner_model", "non_reasoner_model"]): Whether the reasoning
    content was generated natively by a reasoner model or derived via prompting
    reasoning (str): The internal reasoning of the agent
    signature (Optional[str]): The model-generated signature of the reasoning step

    - `id: string`

    - `date: string`

    - `reasoning: string`

    - `is_err?: boolean | null`

    - `message_type?: "reasoning_message"`

      The type of the message.

      - `"reasoning_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `signature?: string | null`

    - `source?: "reasoner_model" | "non_reasoner_model"`

      - `"reasoner_model"`

      - `"non_reasoner_model"`

    - `step_id?: string | null`

  - `HiddenReasoningMessage`

    Representation of an agent's internal reasoning where reasoning content
    has been hidden from the response.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    state (Literal["redacted", "omitted"]): Whether the reasoning
    content was redacted by the provider or simply omitted by the API
    hidden_reasoning (Optional[str]): The internal reasoning of the agent

    - `id: string`

    - `date: string`

    - `state: "redacted" | "omitted"`

      - `"redacted"`

      - `"omitted"`

    - `hidden_reasoning?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "hidden_reasoning_message"`

      The type of the message.

      - `"hidden_reasoning_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `ToolCallMessage`

    A message representing a request to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (Union[ToolCall, ToolCallDelta]): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "tool_call_message"`

      The type of the message.

      - `"tool_call_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `ToolReturnMessage`

    A message representing the return value of a tool call (generated by Letta executing the requested tool).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_return (str): The return value of the tool (deprecated, use tool_returns)
    status (Literal["success", "error"]): The status of the tool call (deprecated, use tool_returns)
    tool_call_id (str): A unique identifier for the tool call that generated this message (deprecated, use tool_returns)
    stdout (Optional[List(str)]): Captured stdout (e.g. prints, logs) from the tool invocation (deprecated, use tool_returns)
    stderr (Optional[List(str)]): Captured stderr from the tool invocation (deprecated, use tool_returns)
    tool_returns (Optional[List[ToolReturn]]): List of tool returns for multi-tool support

    - `id: string`

    - `date: string`

    - `status: "success" | "error"`

      - `"success"`

      - `"error"`

    - `tool_call_id: string`

    - `tool_return: string`

    - `is_err?: boolean | null`

    - `message_type?: "tool_return_message"`

      The type of the message.

      - `"tool_return_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `stderr?: Array<string> | null`

    - `stdout?: Array<string> | null`

    - `step_id?: string | null`

    - `tool_returns?: Array<ToolReturn> | null`

      - `status: "success" | "error"`

        - `"success"`

        - `"error"`

      - `tool_call_id: string`

      - `tool_return: Array<TextContent | ImageContent> | string`

        The tool return value - either a string or list of content parts (text/image)

        - `Array<TextContent | ImageContent>`

          - `TextContent`

          - `ImageContent`

        - `string`

      - `stderr?: Array<string> | null`

      - `stdout?: Array<string> | null`

      - `type?: "tool"`

        The message type to be created.

        - `"tool"`

  - `AssistantMessage`

    A message sent by the LLM in response to user input. Used in the LLM context.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    content (Union[str, List[LettaAssistantMessageContentUnion]]): The message content sent by the agent (can be a string or an array of content parts)

    - `id: string`

    - `content: Array<LettaAssistantMessageContentUnion> | string`

      The message content sent by the agent (can be a string or an array of content parts)

      - `Array<LettaAssistantMessageContentUnion>`

        - `text: string`

          The text content of the message.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this text content.

        - `type?: "text"`

          The type of the message.

          - `"text"`

      - `string`

    - `date: string`

    - `is_err?: boolean | null`

    - `message_type?: "assistant_message"`

      The type of the message.

      - `"assistant_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `ApprovalRequestMessage`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

      - `ToolCallDelta`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `ApprovalResponseMessage`

    A message representing a response form the user indicating whether a tool has been approved to run.

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    approve: (bool) Whether the tool has been approved
    approval_request_id: The ID of the approval request
    reason: (Optional[str]) An optional explanation for the provided approval status

    - `id: string`

    - `date: string`

    - `approval_request_id?: string | null`

      The message ID of the approval request

    - `approvals?: Array<ApprovalReturn | ToolReturn> | null`

      The list of approval responses

      - `ApprovalReturn`

        - `approve: boolean`

          Whether the tool has been approved

        - `tool_call_id: string`

          The ID of the tool call that corresponds to this approval

        - `reason?: string | null`

          An optional explanation for the provided approval status

        - `type?: "approval"`

          The message type to be created.

          - `"approval"`

      - `ToolReturn`

        - `status: "success" | "error"`

        - `tool_call_id: string`

        - `tool_return: Array<TextContent | ImageContent> | string`

          The tool return value - either a string or list of content parts (text/image)

        - `stderr?: Array<string> | null`

        - `stdout?: Array<string> | null`

        - `type?: "tool"`

          The message type to be created.

    - `approve?: boolean | null`

      Whether the tool has been approved

    - `is_err?: boolean | null`

    - `message_type?: "approval_response_message"`

      The type of the message.

      - `"approval_response_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `reason?: string | null`

      An optional explanation for the provided approval status

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `SummaryMessage`

    A message representing a summary of the conversation. Sent to the LLM as a user or system message depending on the provider.

    - `id: string`

    - `date: string`

    - `summary: string`

    - `compaction_stats?: CompactionStats | null`

      Statistics about a memory compaction operation.

      - `context_window: number`

        The model's context window size

      - `messages_count_after: number`

        Number of messages after compaction

      - `messages_count_before: number`

        Number of messages before compaction

      - `trigger: string`

        What triggered the compaction (e.g., 'context_window_exceeded', 'post_step_context_check')

      - `context_tokens_after?: number | null`

        Token count after compaction (message tokens only, does not include tool definitions)

      - `context_tokens_before?: number | null`

        Token count before compaction (from LLM usage stats, includes full context sent to LLM)

    - `is_err?: boolean | null`

    - `message_type?: "summary_message"`

      - `"summary_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

  - `EventMessage`

    A message for notifying the developer that an event that has occured (e.g. a compaction). Events are NOT part of the context window.

    - `id: string`

    - `date: string`

    - `event_data: Record<string, unknown>`

    - `event_type: "compaction"`

      - `"compaction"`

    - `is_err?: boolean | null`

    - `message_type?: "event_message"`

      - `"event_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

### Message Role

- `MessageRole = "assistant" | "user" | "tool" | 4 more`

  - `"assistant"`

  - `"user"`

  - `"tool"`

  - `"function"`

  - `"system"`

  - `"approval"`

  - `"summary"`

### Message Type

- `MessageType = "system_message" | "user_message" | "assistant_message" | 8 more`

  - `"system_message"`

  - `"user_message"`

  - `"assistant_message"`

  - `"reasoning_message"`

  - `"hidden_reasoning_message"`

  - `"tool_call_message"`

  - `"tool_return_message"`

  - `"approval_request_message"`

  - `"approval_response_message"`

  - `"summary_message"`

  - `"event_message"`

### Omitted Reasoning Content

- `OmittedReasoningContent`

  A placeholder for reasoning content we know is present, but isn't returned by the provider (e.g. OpenAI GPT-5 on ChatCompletions)

  - `signature?: string | null`

    A unique identifier for this reasoning step.

  - `type?: "omitted_reasoning"`

    Indicates this is an omitted reasoning step.

    - `"omitted_reasoning"`

### Reasoning Content

- `ReasoningContent`

  Sent via the Anthropic Messages API

  - `is_native: boolean`

    Whether the reasoning content was generated by a reasoner model that processed this step.

  - `reasoning: string`

    The intermediate reasoning or thought process content.

  - `signature?: string | null`

    A unique identifier for this reasoning step.

  - `type?: "reasoning"`

    Indicates this is a reasoning/intermediate step.

    - `"reasoning"`

### Reasoning Message

- `ReasoningMessage`

  Representation of an agent's internal reasoning.

  Args:
  id (str): The ID of the message
  date (datetime): The date the message was created in ISO format
  name (Optional[str]): The name of the sender of the message
  source (Literal["reasoner_model", "non_reasoner_model"]): Whether the reasoning
  content was generated natively by a reasoner model or derived via prompting
  reasoning (str): The internal reasoning of the agent
  signature (Optional[str]): The model-generated signature of the reasoning step

  - `id: string`

  - `date: string`

  - `reasoning: string`

  - `is_err?: boolean | null`

  - `message_type?: "reasoning_message"`

    The type of the message.

    - `"reasoning_message"`

  - `name?: string | null`

  - `otid?: string | null`

    The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

  - `run_id?: string | null`

  - `sender_id?: string | null`

  - `seq_id?: number | null`

  - `signature?: string | null`

  - `source?: "reasoner_model" | "non_reasoner_model"`

    - `"reasoner_model"`

    - `"non_reasoner_model"`

  - `step_id?: string | null`

### Redacted Reasoning Content

- `RedactedReasoningContent`

  Sent via the Anthropic Messages API

  - `data: string`

    The redacted or filtered intermediate reasoning content.

  - `type?: "redacted_reasoning"`

    Indicates this is a redacted thinking step.

    - `"redacted_reasoning"`

### Run

- `Run`

  Representation of a run - a conversation or processing session for an agent. Runs track when agents process messages and maintain the relationship between agents, steps, and messages.

  - `id: string`

    The human-friendly ID of the Run

  - `agent_id: string`

    The unique identifier of the agent associated with the run.

  - `background?: boolean | null`

    Whether the run was created in background mode.

  - `base_template_id?: string | null`

    The base template ID that the run belongs to.

  - `callback_error?: string | null`

    Optional error message from attempting to POST the callback endpoint.

  - `callback_sent_at?: string | null`

    Timestamp when the callback was last attempted.

  - `callback_status_code?: number | null`

    HTTP status code returned by the callback endpoint.

  - `callback_url?: string | null`

    If set, POST to this URL when the run completes.

  - `completed_at?: string | null`

    The timestamp when the run was completed.

  - `conversation_id?: string | null`

    The unique identifier of the conversation associated with the run.

  - `created_at?: string`

    The timestamp when the run was created.

  - `metadata?: Record<string, unknown> | null`

    Additional metadata for the run.

  - `request_config?: RequestConfig | null`

    The request configuration for the run.

    - `assistant_message_tool_kwarg?: string`

      The name of the message argument in the designated message tool.

    - `assistant_message_tool_name?: string`

      The name of the designated message tool.

    - `include_return_message_types?: Array<MessageType> | null`

      Only return specified message types in the response. If `None` (default) returns all messages.

      - `"system_message"`

      - `"user_message"`

      - `"assistant_message"`

      - `"reasoning_message"`

      - `"hidden_reasoning_message"`

      - `"tool_call_message"`

      - `"tool_return_message"`

      - `"approval_request_message"`

      - `"approval_response_message"`

      - `"summary_message"`

      - `"event_message"`

    - `use_assistant_message?: boolean`

      Whether the server should parse specific tool call arguments (default `send_message`) as `AssistantMessage` objects.

  - `status?: "created" | "running" | "completed" | 2 more`

    The current status of the run.

    - `"created"`

    - `"running"`

    - `"completed"`

    - `"failed"`

    - `"cancelled"`

  - `stop_reason?: StopReasonType | null`

    The reason why the run was stopped.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `total_duration_ns?: number | null`

    Total run duration in nanoseconds

  - `ttft_ns?: number | null`

    Time to first token for a run in nanoseconds

### Summary Message

- `SummaryMessage`

  A message representing a summary of the conversation. Sent to the LLM as a user or system message depending on the provider.

  - `id: string`

  - `date: string`

  - `summary: string`

  - `compaction_stats?: CompactionStats | null`

    Statistics about a memory compaction operation.

    - `context_window: number`

      The model's context window size

    - `messages_count_after: number`

      Number of messages after compaction

    - `messages_count_before: number`

      Number of messages before compaction

    - `trigger: string`

      What triggered the compaction (e.g., 'context_window_exceeded', 'post_step_context_check')

    - `context_tokens_after?: number | null`

      Token count after compaction (message tokens only, does not include tool definitions)

    - `context_tokens_before?: number | null`

      Token count before compaction (from LLM usage stats, includes full context sent to LLM)

  - `is_err?: boolean | null`

  - `message_type?: "summary_message"`

    - `"summary_message"`

  - `name?: string | null`

  - `otid?: string | null`

    The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

  - `run_id?: string | null`

  - `sender_id?: string | null`

  - `seq_id?: number | null`

  - `step_id?: string | null`

### System Message

- `SystemMessage`

  A message generated by the system. Never streamed back on a response, only used for cursor pagination.

  Args:
  id (str): The ID of the message
  date (datetime): The date the message was created in ISO format
  name (Optional[str]): The name of the sender of the message
  content (str): The message content sent by the system

  - `id: string`

  - `content: string`

    The message content sent by the system

  - `date: string`

  - `is_err?: boolean | null`

  - `message_type?: "system_message"`

    The type of the message.

    - `"system_message"`

  - `name?: string | null`

  - `otid?: string | null`

    The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

  - `run_id?: string | null`

  - `sender_id?: string | null`

  - `seq_id?: number | null`

  - `step_id?: string | null`

### Text Content

- `TextContent`

  - `text: string`

    The text content of the message.

  - `signature?: string | null`

    Stores a unique identifier for any reasoning associated with this text content.

  - `type?: "text"`

    The type of the message.

    - `"text"`

### Tool Call

- `ToolCall`

  - `arguments: string`

  - `name: string`

  - `tool_call_id: string`

### Tool Call Content

- `ToolCallContent`

  - `id: string`

    A unique identifier for this specific tool call instance.

  - `input: Record<string, unknown>`

    The parameters being passed to the tool, structured as a dictionary of parameter names to values.

  - `name: string`

    The name of the tool being called.

  - `signature?: string | null`

    Stores a unique identifier for any reasoning associated with this tool call.

  - `type?: "tool_call"`

    Indicates this content represents a tool call event.

    - `"tool_call"`

### Tool Call Delta

- `ToolCallDelta`

  - `arguments?: string | null`

  - `name?: string | null`

  - `tool_call_id?: string | null`

### Tool Call Message

- `ToolCallMessage`

  A message representing a request to call a tool (generated by the LLM to trigger tool execution).

  Args:
  id (str): The ID of the message
  date (datetime): The date the message was created in ISO format
  name (Optional[str]): The name of the sender of the message
  tool_call (Union[ToolCall, ToolCallDelta]): The tool call

  - `id: string`

  - `date: string`

  - `tool_call: ToolCall | ToolCallDelta`

    - `ToolCall`

      - `arguments: string`

      - `name: string`

      - `tool_call_id: string`

    - `ToolCallDelta`

      - `arguments?: string | null`

      - `name?: string | null`

      - `tool_call_id?: string | null`

  - `is_err?: boolean | null`

  - `message_type?: "tool_call_message"`

    The type of the message.

    - `"tool_call_message"`

  - `name?: string | null`

  - `otid?: string | null`

    The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

  - `run_id?: string | null`

  - `sender_id?: string | null`

  - `seq_id?: number | null`

  - `step_id?: string | null`

  - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

    - `Array<ToolCall>`

      - `arguments: string`

      - `name: string`

      - `tool_call_id: string`

    - `ToolCallDelta`

### Tool Return

- `ToolReturn`

  - `status: "success" | "error"`

    - `"success"`

    - `"error"`

  - `tool_call_id: string`

  - `tool_return: Array<TextContent | ImageContent> | string`

    The tool return value - either a string or list of content parts (text/image)

    - `Array<TextContent | ImageContent>`

      - `TextContent`

        - `text: string`

          The text content of the message.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this text content.

        - `type?: "text"`

          The type of the message.

          - `"text"`

      - `ImageContent`

        - `source: URLImage | Base64Image | LettaImage`

          The source of the image.

          - `URLImage`

            - `url: string`

              The URL of the image.

            - `type?: "url"`

              The source type for the image.

              - `"url"`

          - `Base64Image`

            - `data: string`

              The base64 encoded image data.

            - `media_type: string`

              The media type for the image.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `type?: "base64"`

              The source type for the image.

              - `"base64"`

          - `LettaImage`

            - `file_id: string`

              The unique identifier of the image file persisted in storage.

            - `data?: string | null`

              The base64 encoded image data.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `media_type?: string | null`

              The media type for the image.

            - `type?: "letta"`

              The source type for the image.

              - `"letta"`

        - `type?: "image"`

          The type of the message.

          - `"image"`

    - `string`

  - `stderr?: Array<string> | null`

  - `stdout?: Array<string> | null`

  - `type?: "tool"`

    The message type to be created.

    - `"tool"`

### Tool Return Content

- `ToolReturnContent`

  - `content: string`

    The content returned by the tool execution.

  - `is_error: boolean`

    Indicates whether the tool execution resulted in an error.

  - `tool_call_id: string`

    References the ID of the ToolCallContent that initiated this tool call.

  - `type?: "tool_return"`

    Indicates this content represents a tool return event.

    - `"tool_return"`

### Update Assistant Message

- `UpdateAssistantMessage`

  - `content: Array<LettaAssistantMessageContentUnion> | string`

    The message content sent by the assistant (can be a string or an array of content parts)

    - `Array<LettaAssistantMessageContentUnion>`

      - `text: string`

        The text content of the message.

      - `signature?: string | null`

        Stores a unique identifier for any reasoning associated with this text content.

      - `type?: "text"`

        The type of the message.

        - `"text"`

    - `string`

  - `message_type?: "assistant_message"`

    - `"assistant_message"`

### Update Reasoning Message

- `UpdateReasoningMessage`

  - `reasoning: string`

  - `message_type?: "reasoning_message"`

    - `"reasoning_message"`

### Update System Message

- `UpdateSystemMessage`

  - `content: string`

    The message content sent by the system (can be a string or an array of multi-modal content parts)

  - `message_type?: "system_message"`

    - `"system_message"`

### Update User Message

- `UpdateUserMessage`

  - `content: Array<LettaUserMessageContentUnion> | string`

    The message content sent by the user (can be a string or an array of multi-modal content parts)

    - `Array<LettaUserMessageContentUnion>`

      - `TextContent`

        - `text: string`

          The text content of the message.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this text content.

        - `type?: "text"`

          The type of the message.

          - `"text"`

      - `ImageContent`

        - `source: URLImage | Base64Image | LettaImage`

          The source of the image.

          - `URLImage`

            - `url: string`

              The URL of the image.

            - `type?: "url"`

              The source type for the image.

              - `"url"`

          - `Base64Image`

            - `data: string`

              The base64 encoded image data.

            - `media_type: string`

              The media type for the image.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `type?: "base64"`

              The source type for the image.

              - `"base64"`

          - `LettaImage`

            - `file_id: string`

              The unique identifier of the image file persisted in storage.

            - `data?: string | null`

              The base64 encoded image data.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `media_type?: string | null`

              The media type for the image.

            - `type?: "letta"`

              The source type for the image.

              - `"letta"`

        - `type?: "image"`

          The type of the message.

          - `"image"`

    - `string`

  - `message_type?: "user_message"`

    - `"user_message"`

### User Message

- `UserMessage`

  A message sent by the user. Never streamed back on a response, only used for cursor pagination.

  Args:
  id (str): The ID of the message
  date (datetime): The date the message was created in ISO format
  name (Optional[str]): The name of the sender of the message
  content (Union[str, List[LettaUserMessageContentUnion]]): The message content sent by the user (can be a string or an array of multi-modal content parts)

  - `id: string`

  - `content: Array<LettaUserMessageContentUnion> | string`

    The message content sent by the user (can be a string or an array of multi-modal content parts)

    - `Array<LettaUserMessageContentUnion>`

      - `TextContent`

        - `text: string`

          The text content of the message.

        - `signature?: string | null`

          Stores a unique identifier for any reasoning associated with this text content.

        - `type?: "text"`

          The type of the message.

          - `"text"`

      - `ImageContent`

        - `source: URLImage | Base64Image | LettaImage`

          The source of the image.

          - `URLImage`

            - `url: string`

              The URL of the image.

            - `type?: "url"`

              The source type for the image.

              - `"url"`

          - `Base64Image`

            - `data: string`

              The base64 encoded image data.

            - `media_type: string`

              The media type for the image.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `type?: "base64"`

              The source type for the image.

              - `"base64"`

          - `LettaImage`

            - `file_id: string`

              The unique identifier of the image file persisted in storage.

            - `data?: string | null`

              The base64 encoded image data.

            - `detail?: string | null`

              What level of detail to use when processing and understanding the image (low, high, or auto to let the model decide)

            - `media_type?: string | null`

              The media type for the image.

            - `type?: "letta"`

              The source type for the image.

              - `"letta"`

        - `type?: "image"`

          The type of the message.

          - `"image"`

    - `string`

  - `date: string`

  - `is_err?: boolean | null`

  - `message_type?: "user_message"`

    The type of the message.

    - `"user_message"`

  - `name?: string | null`

  - `otid?: string | null`

    The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

  - `run_id?: string | null`

  - `sender_id?: string | null`

  - `seq_id?: number | null`

  - `step_id?: string | null`

### Message Cancel Response

- `MessageCancelResponse = Record<string, unknown>`

# Schedule

## Schedule Agent Message

`client.agents.schedule.create(stringagentID, ScheduleCreateParamsbody, RequestOptionsoptions?): ScheduleCreateResponse`

**post** `/v1/agents/{agent_id}/schedule`

Schedule a message to be sent by the agent at a specified time or on a recurring basis.

### Parameters

- `agentID: string`

- `body: ScheduleCreateParams`

  - `messages: Array<Message>`

    - `content: Array<UnionMember0 | UnionMember1> | string`

      - `Array<UnionMember0 | UnionMember1>`

        - `UnionMember0`

          - `text: string`

          - `signature?: string | null`

          - `type?: "text"`

            - `"text"`

        - `UnionMember1`

          - `source: Source`

            - `data: string`

            - `media_type: string`

            - `detail?: string`

            - `type?: "base64"`

              - `"base64"`

          - `type: "image"`

            - `"image"`

      - `string`

    - `role: "user" | "assistant" | "system"`

      - `"user"`

      - `"assistant"`

      - `"system"`

    - `name?: string`

    - `otid?: string`

    - `sender_id?: string`

    - `type?: "message"`

      - `"message"`

  - `schedule: UnionMember0 | UnionMember1`

    - `UnionMember0`

      - `scheduled_at: number`

      - `type?: "one-time"`

        - `"one-time"`

    - `UnionMember1`

      - `cron_expression: string`

      - `type: "recurring"`

        - `"recurring"`

  - `callback_url?: string`

  - `include_return_message_types?: Array<"system_message" | "user_message" | "assistant_message" | 6 more>`

    - `"system_message"`

    - `"user_message"`

    - `"assistant_message"`

    - `"reasoning_message"`

    - `"hidden_reasoning_message"`

    - `"tool_call_message"`

    - `"tool_return_message"`

    - `"approval_request_message"`

    - `"approval_response_message"`

  - `max_steps?: number`

### Returns

- `ScheduleCreateResponse`

  - `id: string`

  - `next_scheduled_at?: string`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const schedule = await client.agents.schedule.create('agent_id', {
  messages: [{ content: [{ text: 'text' }], role: 'user' }],
  schedule: { scheduled_at: 0 },
});

console.log(schedule.id);
```

#### Response

```json
{
  "id": "id",
  "next_scheduled_at": "next_scheduled_at"
}
```

## List Scheduled Agent Messages

`client.agents.schedule.list(stringagentID, ScheduleListParamsquery?, RequestOptionsoptions?): ScheduleListResponse`

**get** `/v1/agents/{agent_id}/schedule`

List all scheduled messages for a specific agent.

### Parameters

- `agentID: string`

- `query: ScheduleListParams`

  - `after?: string`

  - `limit?: string`

### Returns

- `ScheduleListResponse`

  - `has_next_page: boolean`

  - `scheduled_messages: Array<ScheduledMessage>`

    - `id: string`

    - `agent_id: string`

    - `message: Message`

      - `messages: Array<Message>`

        - `content: Array<UnionMember0 | UnionMember1> | string`

          - `Array<UnionMember0 | UnionMember1>`

            - `UnionMember0`

              - `text: string`

              - `signature?: string | null`

              - `type?: "text"`

                - `"text"`

            - `UnionMember1`

              - `source: Source`

                - `data: string`

                - `media_type: string`

                - `detail?: string`

                - `type?: "base64"`

                  - `"base64"`

              - `type: "image"`

                - `"image"`

          - `string`

        - `role: "user" | "assistant" | "system"`

          - `"user"`

          - `"assistant"`

          - `"system"`

        - `name?: string`

        - `otid?: string`

        - `sender_id?: string`

        - `type?: "message"`

          - `"message"`

      - `callback_url?: string`

      - `include_return_message_types?: Array<"system_message" | "user_message" | "assistant_message" | 6 more>`

        - `"system_message"`

        - `"user_message"`

        - `"assistant_message"`

        - `"reasoning_message"`

        - `"hidden_reasoning_message"`

        - `"tool_call_message"`

        - `"tool_return_message"`

        - `"approval_request_message"`

        - `"approval_response_message"`

      - `max_steps?: number`

    - `next_scheduled_time: string | null`

    - `schedule: UnionMember0 | UnionMember1`

      - `UnionMember0`

        - `scheduled_at: number`

        - `type?: "one-time"`

          - `"one-time"`

      - `UnionMember1`

        - `cron_expression: string`

        - `type: "recurring"`

          - `"recurring"`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const schedules = await client.agents.schedule.list('agent_id');

console.log(schedules.has_next_page);
```

#### Response

```json
{
  "has_next_page": true,
  "scheduled_messages": [
    {
      "id": "id",
      "agent_id": "agent_id",
      "message": {
        "messages": [
          {
            "content": [
              {
                "text": "text",
                "signature": "signature",
                "type": "text"
              }
            ],
            "role": "user",
            "name": "name",
            "otid": "otid",
            "sender_id": "sender_id",
            "type": "message"
          }
        ],
        "callback_url": "https://example.com",
        "include_return_message_types": [
          "system_message"
        ],
        "max_steps": 0
      },
      "next_scheduled_time": "next_scheduled_time",
      "schedule": {
        "scheduled_at": 0,
        "type": "one-time"
      }
    }
  ]
}
```

## Retrieve Scheduled Agent Message

`client.agents.schedule.retrieve(stringscheduledMessageID, ScheduleRetrieveParamsparams, RequestOptionsoptions?): ScheduleRetrieveResponse`

**get** `/v1/agents/{agent_id}/schedule/{scheduled_message_id}`

Retrieve a scheduled message by its ID for a specific agent.

### Parameters

- `scheduledMessageID: string`

- `params: ScheduleRetrieveParams`

  - `agent_id: string`

### Returns

- `ScheduleRetrieveResponse`

  - `id: string`

  - `agent_id: string`

  - `message: Message`

    - `messages: Array<Message>`

      - `content: Array<UnionMember0 | UnionMember1> | string`

        - `Array<UnionMember0 | UnionMember1>`

          - `UnionMember0`

            - `text: string`

            - `signature?: string | null`

            - `type?: "text"`

              - `"text"`

          - `UnionMember1`

            - `source: Source`

              - `data: string`

              - `media_type: string`

              - `detail?: string`

              - `type?: "base64"`

                - `"base64"`

            - `type: "image"`

              - `"image"`

        - `string`

      - `role: "user" | "assistant" | "system"`

        - `"user"`

        - `"assistant"`

        - `"system"`

      - `name?: string`

      - `otid?: string`

      - `sender_id?: string`

      - `type?: "message"`

        - `"message"`

    - `callback_url?: string`

    - `include_return_message_types?: Array<"system_message" | "user_message" | "assistant_message" | 6 more>`

      - `"system_message"`

      - `"user_message"`

      - `"assistant_message"`

      - `"reasoning_message"`

      - `"hidden_reasoning_message"`

      - `"tool_call_message"`

      - `"tool_return_message"`

      - `"approval_request_message"`

      - `"approval_response_message"`

    - `max_steps?: number`

  - `next_scheduled_time: string | null`

  - `schedule: UnionMember0 | UnionMember1`

    - `UnionMember0`

      - `scheduled_at: number`

      - `type?: "one-time"`

        - `"one-time"`

    - `UnionMember1`

      - `cron_expression: string`

      - `type: "recurring"`

        - `"recurring"`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const schedule = await client.agents.schedule.retrieve('scheduled_message_id', {
  agent_id: 'agent_id',
});

console.log(schedule.id);
```

#### Response

```json
{
  "id": "id",
  "agent_id": "agent_id",
  "message": {
    "messages": [
      {
        "content": [
          {
            "text": "text",
            "signature": "signature",
            "type": "text"
          }
        ],
        "role": "user",
        "name": "name",
        "otid": "otid",
        "sender_id": "sender_id",
        "type": "message"
      }
    ],
    "callback_url": "https://example.com",
    "include_return_message_types": [
      "system_message"
    ],
    "max_steps": 0
  },
  "next_scheduled_time": "next_scheduled_time",
  "schedule": {
    "scheduled_at": 0,
    "type": "one-time"
  }
}
```

## Delete Scheduled Agent Message

`client.agents.schedule.delete(stringscheduledMessageID, ScheduleDeleteParamsparams, RequestOptionsoptions?): ScheduleDeleteResponse`

**delete** `/v1/agents/{agent_id}/schedule/{scheduled_message_id}`

Delete a scheduled message by its ID for a specific agent.

### Parameters

- `scheduledMessageID: string`

- `params: ScheduleDeleteParams`

  - `agent_id: string`

### Returns

- `ScheduleDeleteResponse`

  - `success: true`

    - `true`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const schedule = await client.agents.schedule.delete('scheduled_message_id', {
  agent_id: 'agent_id',
});

console.log(schedule.success);
```

#### Response

```json
{
  "success": true
}
```

## Domain Types

### Schedule Create Response

- `ScheduleCreateResponse`

  - `id: string`

  - `next_scheduled_at?: string`

### Schedule List Response

- `ScheduleListResponse`

  - `has_next_page: boolean`

  - `scheduled_messages: Array<ScheduledMessage>`

    - `id: string`

    - `agent_id: string`

    - `message: Message`

      - `messages: Array<Message>`

        - `content: Array<UnionMember0 | UnionMember1> | string`

          - `Array<UnionMember0 | UnionMember1>`

            - `UnionMember0`

              - `text: string`

              - `signature?: string | null`

              - `type?: "text"`

                - `"text"`

            - `UnionMember1`

              - `source: Source`

                - `data: string`

                - `media_type: string`

                - `detail?: string`

                - `type?: "base64"`

                  - `"base64"`

              - `type: "image"`

                - `"image"`

          - `string`

        - `role: "user" | "assistant" | "system"`

          - `"user"`

          - `"assistant"`

          - `"system"`

        - `name?: string`

        - `otid?: string`

        - `sender_id?: string`

        - `type?: "message"`

          - `"message"`

      - `callback_url?: string`

      - `include_return_message_types?: Array<"system_message" | "user_message" | "assistant_message" | 6 more>`

        - `"system_message"`

        - `"user_message"`

        - `"assistant_message"`

        - `"reasoning_message"`

        - `"hidden_reasoning_message"`

        - `"tool_call_message"`

        - `"tool_return_message"`

        - `"approval_request_message"`

        - `"approval_response_message"`

      - `max_steps?: number`

    - `next_scheduled_time: string | null`

    - `schedule: UnionMember0 | UnionMember1`

      - `UnionMember0`

        - `scheduled_at: number`

        - `type?: "one-time"`

          - `"one-time"`

      - `UnionMember1`

        - `cron_expression: string`

        - `type: "recurring"`

          - `"recurring"`

### Schedule Retrieve Response

- `ScheduleRetrieveResponse`

  - `id: string`

  - `agent_id: string`

  - `message: Message`

    - `messages: Array<Message>`

      - `content: Array<UnionMember0 | UnionMember1> | string`

        - `Array<UnionMember0 | UnionMember1>`

          - `UnionMember0`

            - `text: string`

            - `signature?: string | null`

            - `type?: "text"`

              - `"text"`

          - `UnionMember1`

            - `source: Source`

              - `data: string`

              - `media_type: string`

              - `detail?: string`

              - `type?: "base64"`

                - `"base64"`

            - `type: "image"`

              - `"image"`

        - `string`

      - `role: "user" | "assistant" | "system"`

        - `"user"`

        - `"assistant"`

        - `"system"`

      - `name?: string`

      - `otid?: string`

      - `sender_id?: string`

      - `type?: "message"`

        - `"message"`

    - `callback_url?: string`

    - `include_return_message_types?: Array<"system_message" | "user_message" | "assistant_message" | 6 more>`

      - `"system_message"`

      - `"user_message"`

      - `"assistant_message"`

      - `"reasoning_message"`

      - `"hidden_reasoning_message"`

      - `"tool_call_message"`

      - `"tool_return_message"`

      - `"approval_request_message"`

      - `"approval_response_message"`

    - `max_steps?: number`

  - `next_scheduled_time: string | null`

  - `schedule: UnionMember0 | UnionMember1`

    - `UnionMember0`

      - `scheduled_at: number`

      - `type?: "one-time"`

        - `"one-time"`

    - `UnionMember1`

      - `cron_expression: string`

      - `type: "recurring"`

        - `"recurring"`

### Schedule Delete Response

- `ScheduleDeleteResponse`

  - `success: true`

    - `true`

# Blocks

## Retrieve Block For Agent

`client.agents.blocks.retrieve(stringblockLabel, BlockRetrieveParamsparams, RequestOptionsoptions?): BlockResponse`

**get** `/v1/agents/{agent_id}/core-memory/blocks/{block_label}`

Retrieve a core memory block from an agent.

### Parameters

- `blockLabel: string`

- `params: BlockRetrieveParams`

  - `agent_id: string`

    The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `BlockResponse`

  - `id: string`

    The id of the block.

  - `value: string`

    Value of the block.

  - `base_template_id?: string | null`

    (Deprecated) The base template id of the block.

  - `created_by_id?: string | null`

    The id of the user that made this Block.

  - `deployment_id?: string | null`

    (Deprecated) The id of the deployment.

  - `description?: string | null`

    Description of the block.

  - `entity_id?: string | null`

    (Deprecated) The id of the entity within the template.

  - `hidden?: boolean | null`

    (Deprecated) If set to True, the block will be hidden.

  - `is_template?: boolean`

    Whether the block is a template (e.g. saved human/persona options).

  - `label?: string | null`

    Label of the block (e.g. 'human', 'persona') in the context window.

  - `last_updated_by_id?: string | null`

    The id of the user that last updated this Block.

  - `limit?: number`

    Character limit of the block.

  - `metadata?: Record<string, unknown> | null`

    Metadata of the block.

  - `preserve_on_migration?: boolean | null`

    (Deprecated) Preserve the block on template migration.

  - `project_id?: string | null`

    The associated project id.

  - `read_only?: boolean`

    (Deprecated) Whether the agent has read-only access to the block.

  - `tags?: Array<string> | null`

    The tags associated with the block.

  - `template_id?: string | null`

    (Deprecated) The id of the template.

  - `template_name?: string | null`

    (Deprecated) The name of the block template (if it is a template).

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const blockResponse = await client.agents.blocks.retrieve('block_label', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
});

console.log(blockResponse.id);
```

#### Response

```json
{
  "id": "id",
  "value": "value",
  "base_template_id": "base_template_id",
  "created_by_id": "created_by_id",
  "deployment_id": "deployment_id",
  "description": "description",
  "entity_id": "entity_id",
  "hidden": true,
  "is_template": true,
  "label": "label",
  "last_updated_by_id": "last_updated_by_id",
  "limit": 0,
  "metadata": {
    "foo": "bar"
  },
  "preserve_on_migration": true,
  "project_id": "project_id",
  "read_only": true,
  "tags": [
    "string"
  ],
  "template_id": "template_id",
  "template_name": "template_name"
}
```

## Update Block For Agent

`client.agents.blocks.update(stringblockLabel, BlockUpdateParamsparams, RequestOptionsoptions?): BlockResponse`

**patch** `/v1/agents/{agent_id}/core-memory/blocks/{block_label}`

Updates a core memory block of an agent.

### Parameters

- `blockLabel: string`

- `params: BlockUpdateParams`

  - `agent_id: string`

    Path param: The ID of the agent in the format 'agent-<uuid4>'

  - `base_template_id?: string | null`

    Body param: The base template id of the block.

  - `deployment_id?: string | null`

    Body param: The id of the deployment.

  - `description?: string | null`

    Body param: Description of the block.

  - `entity_id?: string | null`

    Body param: The id of the entity within the template.

  - `hidden?: boolean | null`

    Body param: If set to True, the block will be hidden.

  - `is_template?: boolean`

    Body param: Whether the block is a template (e.g. saved human/persona options).

  - `label?: string | null`

    Body param: Label of the block (e.g. 'human', 'persona') in the context window.

  - `limit?: number | null`

    Body param: Character limit of the block.

  - `metadata?: Record<string, unknown> | null`

    Body param: Metadata of the block.

  - `preserve_on_migration?: boolean | null`

    Body param: Preserve the block on template migration.

  - `project_id?: string | null`

    Body param: The associated project id.

  - `read_only?: boolean`

    Body param: Whether the agent has read-only access to the block.

  - `tags?: Array<string> | null`

    Body param: The tags to associate with the block.

  - `template_id?: string | null`

    Body param: The id of the template.

  - `template_name?: string | null`

    Body param: Name of the block if it is a template.

  - `value?: string | null`

    Body param: Value of the block.

### Returns

- `BlockResponse`

  - `id: string`

    The id of the block.

  - `value: string`

    Value of the block.

  - `base_template_id?: string | null`

    (Deprecated) The base template id of the block.

  - `created_by_id?: string | null`

    The id of the user that made this Block.

  - `deployment_id?: string | null`

    (Deprecated) The id of the deployment.

  - `description?: string | null`

    Description of the block.

  - `entity_id?: string | null`

    (Deprecated) The id of the entity within the template.

  - `hidden?: boolean | null`

    (Deprecated) If set to True, the block will be hidden.

  - `is_template?: boolean`

    Whether the block is a template (e.g. saved human/persona options).

  - `label?: string | null`

    Label of the block (e.g. 'human', 'persona') in the context window.

  - `last_updated_by_id?: string | null`

    The id of the user that last updated this Block.

  - `limit?: number`

    Character limit of the block.

  - `metadata?: Record<string, unknown> | null`

    Metadata of the block.

  - `preserve_on_migration?: boolean | null`

    (Deprecated) Preserve the block on template migration.

  - `project_id?: string | null`

    The associated project id.

  - `read_only?: boolean`

    (Deprecated) Whether the agent has read-only access to the block.

  - `tags?: Array<string> | null`

    The tags associated with the block.

  - `template_id?: string | null`

    (Deprecated) The id of the template.

  - `template_name?: string | null`

    (Deprecated) The name of the block template (if it is a template).

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const blockResponse = await client.agents.blocks.update('block_label', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
});

console.log(blockResponse.id);
```

#### Response

```json
{
  "id": "id",
  "value": "value",
  "base_template_id": "base_template_id",
  "created_by_id": "created_by_id",
  "deployment_id": "deployment_id",
  "description": "description",
  "entity_id": "entity_id",
  "hidden": true,
  "is_template": true,
  "label": "label",
  "last_updated_by_id": "last_updated_by_id",
  "limit": 0,
  "metadata": {
    "foo": "bar"
  },
  "preserve_on_migration": true,
  "project_id": "project_id",
  "read_only": true,
  "tags": [
    "string"
  ],
  "template_id": "template_id",
  "template_name": "template_name"
}
```

## List Blocks For Agent

`client.agents.blocks.list(stringagentID, BlockListParamsquery?, RequestOptionsoptions?): ArrayPage<BlockResponse>`

**get** `/v1/agents/{agent_id}/core-memory/blocks`

Retrieve the core memory blocks of a specific agent.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `query: BlockListParams`

  - `after?: string | null`

    Cursor for pagination (block ID). Returns results relative to this ID in the specified sort order. Expected format: 'block-<uuid4>'

  - `before?: string | null`

    Cursor for pagination (block ID). Returns results relative to this ID in the specified sort order. Expected format: 'block-<uuid4>'

  - `limit?: number | null`

    Maximum number of blocks to return

  - `order?: "asc" | "desc"`

    Sort order for blocks by creation time. 'asc' for oldest first, 'desc' for newest first

    - `"asc"`

    - `"desc"`

  - `order_by?: "created_at"`

    Field to sort by

    - `"created_at"`

### Returns

- `BlockResponse`

  - `id: string`

    The id of the block.

  - `value: string`

    Value of the block.

  - `base_template_id?: string | null`

    (Deprecated) The base template id of the block.

  - `created_by_id?: string | null`

    The id of the user that made this Block.

  - `deployment_id?: string | null`

    (Deprecated) The id of the deployment.

  - `description?: string | null`

    Description of the block.

  - `entity_id?: string | null`

    (Deprecated) The id of the entity within the template.

  - `hidden?: boolean | null`

    (Deprecated) If set to True, the block will be hidden.

  - `is_template?: boolean`

    Whether the block is a template (e.g. saved human/persona options).

  - `label?: string | null`

    Label of the block (e.g. 'human', 'persona') in the context window.

  - `last_updated_by_id?: string | null`

    The id of the user that last updated this Block.

  - `limit?: number`

    Character limit of the block.

  - `metadata?: Record<string, unknown> | null`

    Metadata of the block.

  - `preserve_on_migration?: boolean | null`

    (Deprecated) Preserve the block on template migration.

  - `project_id?: string | null`

    The associated project id.

  - `read_only?: boolean`

    (Deprecated) Whether the agent has read-only access to the block.

  - `tags?: Array<string> | null`

    The tags associated with the block.

  - `template_id?: string | null`

    (Deprecated) The id of the template.

  - `template_name?: string | null`

    (Deprecated) The name of the block template (if it is a template).

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

// Automatically fetches more pages as needed.
for await (const blockResponse of client.agents.blocks.list(
  'agent-123e4567-e89b-42d3-8456-426614174000',
)) {
  console.log(blockResponse.id);
}
```

#### Response

```json
[
  {
    "id": "id",
    "value": "value",
    "base_template_id": "base_template_id",
    "created_by_id": "created_by_id",
    "deployment_id": "deployment_id",
    "description": "description",
    "entity_id": "entity_id",
    "hidden": true,
    "is_template": true,
    "label": "label",
    "last_updated_by_id": "last_updated_by_id",
    "limit": 0,
    "metadata": {
      "foo": "bar"
    },
    "preserve_on_migration": true,
    "project_id": "project_id",
    "read_only": true,
    "tags": [
      "string"
    ],
    "template_id": "template_id",
    "template_name": "template_name"
  }
]
```

## Attach Block To Agent

`client.agents.blocks.attach(stringblockID, BlockAttachParamsparams, RequestOptionsoptions?): AgentState`

**patch** `/v1/agents/{agent_id}/core-memory/blocks/attach/{block_id}`

Attach a core memory block to an agent.

### Parameters

- `blockID: string`

  The ID of the block in the format 'block-<uuid4>'

- `params: BlockAttachParams`

  - `agent_id: string`

    The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `AgentState`

  Representation of an agent's state. This is the state of the agent at a given time, and is persisted in the DB backend. The state has all the information needed to recreate a persisted agent.

  - `id: string`

    The id of the agent. Assigned by the database.

  - `agent_type: AgentType`

    The type of agent.

    - `"memgpt_agent"`

    - `"memgpt_v2_agent"`

    - `"letta_v1_agent"`

    - `"react_agent"`

    - `"workflow_agent"`

    - `"split_thread_agent"`

    - `"sleeptime_agent"`

    - `"voice_convo_agent"`

    - `"voice_sleeptime_agent"`

  - `blocks: Array<Block>`

    The memory blocks used by the agent.

    - `value: string`

      Value of the block.

    - `id?: string`

      The human-friendly ID of the Block

    - `base_template_id?: string | null`

      The base template id of the block.

    - `created_by_id?: string | null`

      The id of the user that made this Block.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      Description of the block.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the block will be hidden.

    - `is_template?: boolean`

      Whether the block is a template (e.g. saved human/persona options).

    - `label?: string | null`

      Label of the block (e.g. 'human', 'persona') in the context window.

    - `last_updated_by_id?: string | null`

      The id of the user that last updated this Block.

    - `limit?: number`

      Character limit of the block.

    - `metadata?: Record<string, unknown> | null`

      Metadata of the block.

    - `preserve_on_migration?: boolean | null`

      Preserve the block on template migration.

    - `project_id?: string | null`

      The associated project id.

    - `read_only?: boolean`

      Whether the agent has read-only access to the block.

    - `tags?: Array<string> | null`

      The tags associated with the block.

    - `template_id?: string | null`

      The id of the template.

    - `template_name?: string | null`

      Name of the block if it is a template.

  - `llm_config: LlmConfig`

    Deprecated: Use `model` field instead. The LLM configuration used by the agent.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

        - `type?: "text"`

          The type of the response format.

          - `"text"`

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

        - `json_schema: Record<string, unknown>`

          The JSON schema of the response.

        - `type?: "json_schema"`

          The type of the response format.

          - `"json_schema"`

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

        - `type?: "json_object"`

          The type of the response format.

          - `"json_object"`

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `memory: Memory`

    Deprecated: Use `blocks` field instead. The in-context memory of the agent.

    - `blocks: Array<Block>`

      Memory blocks contained in the agent's in-context memory

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `agent_type?: AgentType | (string & {}) | null`

      Agent type controlling prompt rendering.

      - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

        Enum to represent the type of agent.

      - `(string & {})`

    - `file_blocks?: Array<FileBlock>`

      Special blocks representing the agent's in-context memory of an attached file

      - `file_id: string`

        Unique identifier of the file.

      - `is_open: boolean`

        True if the agent currently has the file open.

      - `source_id: string`

        Deprecated: Use `folder_id` field instead. Unique identifier of the source.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_accessed_at?: string | null`

        UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `git_enabled?: boolean`

      Whether this agent uses git-backed memory with structured labels.

    - `prompt_template?: string`

      Deprecated. Ignored for performance.

  - `name: string`

    The name of the agent.

  - `sources: Array<Source>`

    Deprecated: Use `folders` field instead. The sources used by the agent.

    - `id: string`

      The human-friendly ID of the Source

    - `embedding_config: EmbeddingConfig`

      The embedding configuration used by the source.

      - `embedding_dim: number`

        The dimension of the embedding.

      - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"bedrock"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"mistral"`

        - `"together"`

        - `"pinecone"`

      - `embedding_model: string`

        The model for the embedding.

      - `azure_deployment?: string | null`

        The Azure deployment for the model.

      - `azure_endpoint?: string | null`

        The Azure endpoint for the model.

      - `azure_version?: string | null`

        The Azure version for the model.

      - `batch_size?: number`

        The maximum batch size for processing embeddings.

      - `embedding_chunk_size?: number | null`

        The chunk size of the embedding.

      - `embedding_endpoint?: string | null`

        The endpoint for the model (`None` if local).

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

    - `name: string`

      The name of the source.

    - `created_at?: string | null`

      The timestamp when the source was created.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `description?: string | null`

      The description of the source.

    - `instructions?: string | null`

      Instructions for how to use the source.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata?: Record<string, unknown> | null`

      Metadata associated with the source.

    - `updated_at?: string | null`

      The timestamp when the source was last updated.

    - `vector_db_provider?: VectorDBProvider`

      The vector database provider used for this source's passages

      - `"native"`

      - `"tpuf"`

      - `"pinecone"`

  - `system: string`

    The system prompt used by the agent.

  - `tags: Array<string>`

    The tags associated with the agent.

  - `tools: Array<Tool>`

    The tools used by the agent.

    - `id: string`

      The human-friendly ID of the Tool

    - `args_json_schema?: Record<string, unknown> | null`

      The args JSON schema of the function.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `default_requires_approval?: boolean | null`

      Default value for whether or not executing this tool requires approval.

    - `description?: string | null`

      The description of the tool.

    - `enable_parallel_execution?: boolean | null`

      If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

    - `json_schema?: Record<string, unknown> | null`

      The JSON schema of the function.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata_?: Record<string, unknown> | null`

      A dictionary of additional metadata for the tool.

    - `name?: string | null`

      The name of the function.

    - `npm_requirements?: Array<NpmRequirement> | null`

      Optional list of npm packages required by this tool.

      - `name: string`

        Name of the npm package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `pip_requirements?: Array<PipRequirement> | null`

      Optional list of pip packages required by this tool.

      - `name: string`

        Name of the pip package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `project_id?: string | null`

      The project id of the tool.

    - `return_char_limit?: number`

      The maximum number of characters in the response.

    - `source_code?: string | null`

      The source code of the function.

    - `source_type?: string | null`

      The type of the source code.

    - `tags?: Array<string>`

      Metadata tags.

    - `tool_type?: ToolType`

      The type of the tool.

      - `"custom"`

      - `"letta_core"`

      - `"letta_memory_core"`

      - `"letta_multi_agent_core"`

      - `"letta_sleeptime_core"`

      - `"letta_voice_sleeptime_core"`

      - `"letta_builtin"`

      - `"letta_files_core"`

      - `"external_langchain"`

      - `"external_composio"`

      - `"external_mcp"`

  - `base_template_id?: string | null`

    The base template id of the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `created_at?: string | null`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the agent will be hidden.

  - `identities?: Array<Identity>`

    The identities associated with this agent.

    - `id: string`

      The human-friendly ID of the Identity

    - `agent_ids: Array<string>`

      The IDs of the agents associated with the identity.

    - `block_ids: Array<string>`

      The IDs of the blocks associated with the identity.

    - `identifier_key: string`

      External, user-generated identifier key of the identity.

    - `identity_type: "org" | "user" | "other"`

      The type of the identity.

      - `"org"`

      - `"user"`

      - `"other"`

    - `name: string`

      The name of the identity.

    - `project_id?: string | null`

      The project id of the identity, if applicable.

    - `properties?: Array<Property>`

      List of properties associated with the identity

      - `key: string`

        The key of the property

      - `type: "string" | "number" | "boolean" | "json"`

        The type of the property

        - `"string"`

        - `"number"`

        - `"boolean"`

        - `"json"`

      - `value: string | number | boolean | Record<string, unknown>`

        The value of the property

        - `string`

        - `number`

        - `boolean`

        - `Record<string, unknown>`

  - `identity_ids?: Array<string>`

    Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

  - `last_run_completion?: string | null`

    The timestamp when the agent last completed a run.

  - `last_run_duration_ms?: number | null`

    The duration in milliseconds of the agent's last run.

  - `last_stop_reason?: StopReasonType | null`

    The stop reason from the agent's last run.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `managed_group?: ManagedGroup | null`

    The multi-agent group that this agent manages

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `message_buffer_autoclear?: boolean`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `message_ids?: Array<string> | null`

    The ids of the messages in the agent's in-context memory.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle used by the agent (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings used by the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `multi_agent_group?: MultiAgentGroup | null`

    Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `pending_approval?: ApprovalRequestMessage | null`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project_id?: string | null`

    The id of the project the agent belongs to.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format used by the agent

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Array<AgentEnvironmentVariable>`

    The environment variables for tool execution specific to this agent.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `template_id?: string | null`

    The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

    Deprecated: use `secrets` field instead.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The list of tool rules.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const agentState = await client.agents.blocks.attach('block-123e4567-e89b-42d3-8456-426614174000', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
});

console.log(agentState.id);
```

#### Response

```json
{
  "id": "id",
  "agent_type": "memgpt_agent",
  "blocks": [
    {
      "value": "value",
      "id": "block-123e4567-e89b-12d3-a456-426614174000",
      "base_template_id": "base_template_id",
      "created_by_id": "created_by_id",
      "deployment_id": "deployment_id",
      "description": "description",
      "entity_id": "entity_id",
      "hidden": true,
      "is_template": true,
      "label": "label",
      "last_updated_by_id": "last_updated_by_id",
      "limit": 0,
      "metadata": {
        "foo": "bar"
      },
      "preserve_on_migration": true,
      "project_id": "project_id",
      "read_only": true,
      "tags": [
        "string"
      ],
      "template_id": "template_id",
      "template_name": "template_name"
    }
  ],
  "llm_config": {
    "context_window": 0,
    "model": "model",
    "model_endpoint_type": "openai",
    "compatibility_type": "gguf",
    "display_name": "display_name",
    "effort": "low",
    "enable_reasoner": true,
    "frequency_penalty": 0,
    "handle": "handle",
    "max_reasoning_tokens": 0,
    "max_tokens": 0,
    "model_endpoint": "model_endpoint",
    "model_wrapper": "model_wrapper",
    "parallel_tool_calls": true,
    "provider_category": "base",
    "provider_name": "provider_name",
    "put_inner_thoughts_in_kwargs": true,
    "reasoning_effort": "none",
    "response_format": {
      "type": "text"
    },
    "return_logprobs": true,
    "return_token_ids": true,
    "strict": true,
    "temperature": 0,
    "tier": "tier",
    "tool_call_parser": "tool_call_parser",
    "top_logprobs": 0,
    "verbosity": "low"
  },
  "memory": {
    "blocks": [
      {
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "agent_type": "memgpt_agent",
    "file_blocks": [
      {
        "file_id": "file_id",
        "is_open": true,
        "source_id": "source_id",
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_accessed_at": "2019-12-27T18:11:19.117Z",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "git_enabled": true,
    "prompt_template": "prompt_template"
  },
  "name": "name",
  "sources": [
    {
      "id": "source-123e4567-e89b-12d3-a456-426614174000",
      "embedding_config": {
        "embedding_dim": 0,
        "embedding_endpoint_type": "openai",
        "embedding_model": "embedding_model",
        "azure_deployment": "azure_deployment",
        "azure_endpoint": "azure_endpoint",
        "azure_version": "azure_version",
        "batch_size": 0,
        "embedding_chunk_size": 0,
        "embedding_endpoint": "embedding_endpoint",
        "handle": "handle"
      },
      "name": "name",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "instructions": "instructions",
      "last_updated_by_id": "last_updated_by_id",
      "metadata": {
        "foo": "bar"
      },
      "updated_at": "2019-12-27T18:11:19.117Z",
      "vector_db_provider": "native"
    }
  ],
  "system": "system",
  "tags": [
    "string"
  ],
  "tools": [
    {
      "id": "tool-123e4567-e89b-12d3-a456-426614174000",
      "args_json_schema": {
        "foo": "bar"
      },
      "created_by_id": "created_by_id",
      "default_requires_approval": true,
      "description": "description",
      "enable_parallel_execution": true,
      "json_schema": {
        "foo": "bar"
      },
      "last_updated_by_id": "last_updated_by_id",
      "metadata_": {
        "foo": "bar"
      },
      "name": "name",
      "npm_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "pip_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "project_id": "project_id",
      "return_char_limit": 1,
      "source_code": "source_code",
      "source_type": "source_type",
      "tags": [
        "string"
      ],
      "tool_type": "custom"
    }
  ],
  "base_template_id": "base_template_id",
  "compaction_settings": {
    "clip_chars": 0,
    "mode": "all",
    "model": "model",
    "model_settings": {
      "max_output_tokens": 0,
      "parallel_tool_calls": true,
      "provider_type": "openai",
      "reasoning": {
        "reasoning_effort": "none"
      },
      "response_format": {
        "type": "text"
      },
      "strict": true,
      "temperature": 0
    },
    "prompt": "prompt",
    "prompt_acknowledgement": true,
    "sliding_window_percentage": 0
  },
  "created_at": "2019-12-27T18:11:19.117Z",
  "created_by_id": "created_by_id",
  "deployment_id": "deployment_id",
  "description": "description",
  "embedding": "embedding",
  "embedding_config": {
    "embedding_dim": 0,
    "embedding_endpoint_type": "openai",
    "embedding_model": "embedding_model",
    "azure_deployment": "azure_deployment",
    "azure_endpoint": "azure_endpoint",
    "azure_version": "azure_version",
    "batch_size": 0,
    "embedding_chunk_size": 0,
    "embedding_endpoint": "embedding_endpoint",
    "handle": "handle"
  },
  "enable_sleeptime": true,
  "entity_id": "entity_id",
  "hidden": true,
  "identities": [
    {
      "id": "identity-123e4567-e89b-12d3-a456-426614174000",
      "agent_ids": [
        "string"
      ],
      "block_ids": [
        "string"
      ],
      "identifier_key": "identifier_key",
      "identity_type": "org",
      "name": "name",
      "project_id": "project_id",
      "properties": [
        {
          "key": "key",
          "type": "string",
          "value": "string"
        }
      ]
    }
  ],
  "identity_ids": [
    "string"
  ],
  "last_run_completion": "2019-12-27T18:11:19.117Z",
  "last_run_duration_ms": 0,
  "last_stop_reason": "end_turn",
  "last_updated_by_id": "last_updated_by_id",
  "managed_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "max_files_open": 0,
  "message_buffer_autoclear": true,
  "message_ids": [
    "string"
  ],
  "metadata": {
    "foo": "bar"
  },
  "model": "model",
  "model_settings": {
    "max_output_tokens": 0,
    "parallel_tool_calls": true,
    "provider_type": "openai",
    "reasoning": {
      "reasoning_effort": "none"
    },
    "response_format": {
      "type": "text"
    },
    "strict": true,
    "temperature": 0
  },
  "multi_agent_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "pending_approval": {
    "id": "id",
    "date": "2019-12-27T18:11:19.117Z",
    "tool_call": {
      "arguments": "arguments",
      "name": "name",
      "tool_call_id": "tool_call_id"
    },
    "is_err": true,
    "message_type": "approval_request_message",
    "name": "name",
    "otid": "otid",
    "run_id": "run_id",
    "sender_id": "sender_id",
    "seq_id": 0,
    "step_id": "step_id",
    "tool_calls": [
      {
        "arguments": "arguments",
        "name": "name",
        "tool_call_id": "tool_call_id"
      }
    ]
  },
  "per_file_view_window_char_limit": 0,
  "project_id": "project_id",
  "response_format": {
    "type": "text"
  },
  "secrets": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "template_id": "template_id",
  "timezone": "timezone",
  "tool_exec_environment_variables": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "tool_rules": [
    {
      "children": [
        "string"
      ],
      "tool_name": "tool_name",
      "child_arg_nodes": [
        {
          "name": "name",
          "args": {
            "foo": "bar"
          }
        }
      ],
      "prompt_template": "prompt_template",
      "type": "constrain_child_tools"
    }
  ],
  "updated_at": "2019-12-27T18:11:19.117Z"
}
```

## Detach Block From Agent

`client.agents.blocks.detach(stringblockID, BlockDetachParamsparams, RequestOptionsoptions?): AgentState`

**patch** `/v1/agents/{agent_id}/core-memory/blocks/detach/{block_id}`

Detach a core memory block from an agent.

### Parameters

- `blockID: string`

  The ID of the block in the format 'block-<uuid4>'

- `params: BlockDetachParams`

  - `agent_id: string`

    The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `AgentState`

  Representation of an agent's state. This is the state of the agent at a given time, and is persisted in the DB backend. The state has all the information needed to recreate a persisted agent.

  - `id: string`

    The id of the agent. Assigned by the database.

  - `agent_type: AgentType`

    The type of agent.

    - `"memgpt_agent"`

    - `"memgpt_v2_agent"`

    - `"letta_v1_agent"`

    - `"react_agent"`

    - `"workflow_agent"`

    - `"split_thread_agent"`

    - `"sleeptime_agent"`

    - `"voice_convo_agent"`

    - `"voice_sleeptime_agent"`

  - `blocks: Array<Block>`

    The memory blocks used by the agent.

    - `value: string`

      Value of the block.

    - `id?: string`

      The human-friendly ID of the Block

    - `base_template_id?: string | null`

      The base template id of the block.

    - `created_by_id?: string | null`

      The id of the user that made this Block.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      Description of the block.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the block will be hidden.

    - `is_template?: boolean`

      Whether the block is a template (e.g. saved human/persona options).

    - `label?: string | null`

      Label of the block (e.g. 'human', 'persona') in the context window.

    - `last_updated_by_id?: string | null`

      The id of the user that last updated this Block.

    - `limit?: number`

      Character limit of the block.

    - `metadata?: Record<string, unknown> | null`

      Metadata of the block.

    - `preserve_on_migration?: boolean | null`

      Preserve the block on template migration.

    - `project_id?: string | null`

      The associated project id.

    - `read_only?: boolean`

      Whether the agent has read-only access to the block.

    - `tags?: Array<string> | null`

      The tags associated with the block.

    - `template_id?: string | null`

      The id of the template.

    - `template_name?: string | null`

      Name of the block if it is a template.

  - `llm_config: LlmConfig`

    Deprecated: Use `model` field instead. The LLM configuration used by the agent.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

        - `type?: "text"`

          The type of the response format.

          - `"text"`

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

        - `json_schema: Record<string, unknown>`

          The JSON schema of the response.

        - `type?: "json_schema"`

          The type of the response format.

          - `"json_schema"`

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

        - `type?: "json_object"`

          The type of the response format.

          - `"json_object"`

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `memory: Memory`

    Deprecated: Use `blocks` field instead. The in-context memory of the agent.

    - `blocks: Array<Block>`

      Memory blocks contained in the agent's in-context memory

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `agent_type?: AgentType | (string & {}) | null`

      Agent type controlling prompt rendering.

      - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

        Enum to represent the type of agent.

      - `(string & {})`

    - `file_blocks?: Array<FileBlock>`

      Special blocks representing the agent's in-context memory of an attached file

      - `file_id: string`

        Unique identifier of the file.

      - `is_open: boolean`

        True if the agent currently has the file open.

      - `source_id: string`

        Deprecated: Use `folder_id` field instead. Unique identifier of the source.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_accessed_at?: string | null`

        UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `git_enabled?: boolean`

      Whether this agent uses git-backed memory with structured labels.

    - `prompt_template?: string`

      Deprecated. Ignored for performance.

  - `name: string`

    The name of the agent.

  - `sources: Array<Source>`

    Deprecated: Use `folders` field instead. The sources used by the agent.

    - `id: string`

      The human-friendly ID of the Source

    - `embedding_config: EmbeddingConfig`

      The embedding configuration used by the source.

      - `embedding_dim: number`

        The dimension of the embedding.

      - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"bedrock"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"mistral"`

        - `"together"`

        - `"pinecone"`

      - `embedding_model: string`

        The model for the embedding.

      - `azure_deployment?: string | null`

        The Azure deployment for the model.

      - `azure_endpoint?: string | null`

        The Azure endpoint for the model.

      - `azure_version?: string | null`

        The Azure version for the model.

      - `batch_size?: number`

        The maximum batch size for processing embeddings.

      - `embedding_chunk_size?: number | null`

        The chunk size of the embedding.

      - `embedding_endpoint?: string | null`

        The endpoint for the model (`None` if local).

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

    - `name: string`

      The name of the source.

    - `created_at?: string | null`

      The timestamp when the source was created.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `description?: string | null`

      The description of the source.

    - `instructions?: string | null`

      Instructions for how to use the source.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata?: Record<string, unknown> | null`

      Metadata associated with the source.

    - `updated_at?: string | null`

      The timestamp when the source was last updated.

    - `vector_db_provider?: VectorDBProvider`

      The vector database provider used for this source's passages

      - `"native"`

      - `"tpuf"`

      - `"pinecone"`

  - `system: string`

    The system prompt used by the agent.

  - `tags: Array<string>`

    The tags associated with the agent.

  - `tools: Array<Tool>`

    The tools used by the agent.

    - `id: string`

      The human-friendly ID of the Tool

    - `args_json_schema?: Record<string, unknown> | null`

      The args JSON schema of the function.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `default_requires_approval?: boolean | null`

      Default value for whether or not executing this tool requires approval.

    - `description?: string | null`

      The description of the tool.

    - `enable_parallel_execution?: boolean | null`

      If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

    - `json_schema?: Record<string, unknown> | null`

      The JSON schema of the function.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata_?: Record<string, unknown> | null`

      A dictionary of additional metadata for the tool.

    - `name?: string | null`

      The name of the function.

    - `npm_requirements?: Array<NpmRequirement> | null`

      Optional list of npm packages required by this tool.

      - `name: string`

        Name of the npm package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `pip_requirements?: Array<PipRequirement> | null`

      Optional list of pip packages required by this tool.

      - `name: string`

        Name of the pip package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `project_id?: string | null`

      The project id of the tool.

    - `return_char_limit?: number`

      The maximum number of characters in the response.

    - `source_code?: string | null`

      The source code of the function.

    - `source_type?: string | null`

      The type of the source code.

    - `tags?: Array<string>`

      Metadata tags.

    - `tool_type?: ToolType`

      The type of the tool.

      - `"custom"`

      - `"letta_core"`

      - `"letta_memory_core"`

      - `"letta_multi_agent_core"`

      - `"letta_sleeptime_core"`

      - `"letta_voice_sleeptime_core"`

      - `"letta_builtin"`

      - `"letta_files_core"`

      - `"external_langchain"`

      - `"external_composio"`

      - `"external_mcp"`

  - `base_template_id?: string | null`

    The base template id of the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `created_at?: string | null`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the agent will be hidden.

  - `identities?: Array<Identity>`

    The identities associated with this agent.

    - `id: string`

      The human-friendly ID of the Identity

    - `agent_ids: Array<string>`

      The IDs of the agents associated with the identity.

    - `block_ids: Array<string>`

      The IDs of the blocks associated with the identity.

    - `identifier_key: string`

      External, user-generated identifier key of the identity.

    - `identity_type: "org" | "user" | "other"`

      The type of the identity.

      - `"org"`

      - `"user"`

      - `"other"`

    - `name: string`

      The name of the identity.

    - `project_id?: string | null`

      The project id of the identity, if applicable.

    - `properties?: Array<Property>`

      List of properties associated with the identity

      - `key: string`

        The key of the property

      - `type: "string" | "number" | "boolean" | "json"`

        The type of the property

        - `"string"`

        - `"number"`

        - `"boolean"`

        - `"json"`

      - `value: string | number | boolean | Record<string, unknown>`

        The value of the property

        - `string`

        - `number`

        - `boolean`

        - `Record<string, unknown>`

  - `identity_ids?: Array<string>`

    Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

  - `last_run_completion?: string | null`

    The timestamp when the agent last completed a run.

  - `last_run_duration_ms?: number | null`

    The duration in milliseconds of the agent's last run.

  - `last_stop_reason?: StopReasonType | null`

    The stop reason from the agent's last run.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `managed_group?: ManagedGroup | null`

    The multi-agent group that this agent manages

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `message_buffer_autoclear?: boolean`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `message_ids?: Array<string> | null`

    The ids of the messages in the agent's in-context memory.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle used by the agent (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings used by the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `multi_agent_group?: MultiAgentGroup | null`

    Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `pending_approval?: ApprovalRequestMessage | null`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project_id?: string | null`

    The id of the project the agent belongs to.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format used by the agent

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Array<AgentEnvironmentVariable>`

    The environment variables for tool execution specific to this agent.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `template_id?: string | null`

    The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

    Deprecated: use `secrets` field instead.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The list of tool rules.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const agentState = await client.agents.blocks.detach('block-123e4567-e89b-42d3-8456-426614174000', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
});

console.log(agentState.id);
```

#### Response

```json
{
  "id": "id",
  "agent_type": "memgpt_agent",
  "blocks": [
    {
      "value": "value",
      "id": "block-123e4567-e89b-12d3-a456-426614174000",
      "base_template_id": "base_template_id",
      "created_by_id": "created_by_id",
      "deployment_id": "deployment_id",
      "description": "description",
      "entity_id": "entity_id",
      "hidden": true,
      "is_template": true,
      "label": "label",
      "last_updated_by_id": "last_updated_by_id",
      "limit": 0,
      "metadata": {
        "foo": "bar"
      },
      "preserve_on_migration": true,
      "project_id": "project_id",
      "read_only": true,
      "tags": [
        "string"
      ],
      "template_id": "template_id",
      "template_name": "template_name"
    }
  ],
  "llm_config": {
    "context_window": 0,
    "model": "model",
    "model_endpoint_type": "openai",
    "compatibility_type": "gguf",
    "display_name": "display_name",
    "effort": "low",
    "enable_reasoner": true,
    "frequency_penalty": 0,
    "handle": "handle",
    "max_reasoning_tokens": 0,
    "max_tokens": 0,
    "model_endpoint": "model_endpoint",
    "model_wrapper": "model_wrapper",
    "parallel_tool_calls": true,
    "provider_category": "base",
    "provider_name": "provider_name",
    "put_inner_thoughts_in_kwargs": true,
    "reasoning_effort": "none",
    "response_format": {
      "type": "text"
    },
    "return_logprobs": true,
    "return_token_ids": true,
    "strict": true,
    "temperature": 0,
    "tier": "tier",
    "tool_call_parser": "tool_call_parser",
    "top_logprobs": 0,
    "verbosity": "low"
  },
  "memory": {
    "blocks": [
      {
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "agent_type": "memgpt_agent",
    "file_blocks": [
      {
        "file_id": "file_id",
        "is_open": true,
        "source_id": "source_id",
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_accessed_at": "2019-12-27T18:11:19.117Z",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "git_enabled": true,
    "prompt_template": "prompt_template"
  },
  "name": "name",
  "sources": [
    {
      "id": "source-123e4567-e89b-12d3-a456-426614174000",
      "embedding_config": {
        "embedding_dim": 0,
        "embedding_endpoint_type": "openai",
        "embedding_model": "embedding_model",
        "azure_deployment": "azure_deployment",
        "azure_endpoint": "azure_endpoint",
        "azure_version": "azure_version",
        "batch_size": 0,
        "embedding_chunk_size": 0,
        "embedding_endpoint": "embedding_endpoint",
        "handle": "handle"
      },
      "name": "name",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "instructions": "instructions",
      "last_updated_by_id": "last_updated_by_id",
      "metadata": {
        "foo": "bar"
      },
      "updated_at": "2019-12-27T18:11:19.117Z",
      "vector_db_provider": "native"
    }
  ],
  "system": "system",
  "tags": [
    "string"
  ],
  "tools": [
    {
      "id": "tool-123e4567-e89b-12d3-a456-426614174000",
      "args_json_schema": {
        "foo": "bar"
      },
      "created_by_id": "created_by_id",
      "default_requires_approval": true,
      "description": "description",
      "enable_parallel_execution": true,
      "json_schema": {
        "foo": "bar"
      },
      "last_updated_by_id": "last_updated_by_id",
      "metadata_": {
        "foo": "bar"
      },
      "name": "name",
      "npm_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "pip_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "project_id": "project_id",
      "return_char_limit": 1,
      "source_code": "source_code",
      "source_type": "source_type",
      "tags": [
        "string"
      ],
      "tool_type": "custom"
    }
  ],
  "base_template_id": "base_template_id",
  "compaction_settings": {
    "clip_chars": 0,
    "mode": "all",
    "model": "model",
    "model_settings": {
      "max_output_tokens": 0,
      "parallel_tool_calls": true,
      "provider_type": "openai",
      "reasoning": {
        "reasoning_effort": "none"
      },
      "response_format": {
        "type": "text"
      },
      "strict": true,
      "temperature": 0
    },
    "prompt": "prompt",
    "prompt_acknowledgement": true,
    "sliding_window_percentage": 0
  },
  "created_at": "2019-12-27T18:11:19.117Z",
  "created_by_id": "created_by_id",
  "deployment_id": "deployment_id",
  "description": "description",
  "embedding": "embedding",
  "embedding_config": {
    "embedding_dim": 0,
    "embedding_endpoint_type": "openai",
    "embedding_model": "embedding_model",
    "azure_deployment": "azure_deployment",
    "azure_endpoint": "azure_endpoint",
    "azure_version": "azure_version",
    "batch_size": 0,
    "embedding_chunk_size": 0,
    "embedding_endpoint": "embedding_endpoint",
    "handle": "handle"
  },
  "enable_sleeptime": true,
  "entity_id": "entity_id",
  "hidden": true,
  "identities": [
    {
      "id": "identity-123e4567-e89b-12d3-a456-426614174000",
      "agent_ids": [
        "string"
      ],
      "block_ids": [
        "string"
      ],
      "identifier_key": "identifier_key",
      "identity_type": "org",
      "name": "name",
      "project_id": "project_id",
      "properties": [
        {
          "key": "key",
          "type": "string",
          "value": "string"
        }
      ]
    }
  ],
  "identity_ids": [
    "string"
  ],
  "last_run_completion": "2019-12-27T18:11:19.117Z",
  "last_run_duration_ms": 0,
  "last_stop_reason": "end_turn",
  "last_updated_by_id": "last_updated_by_id",
  "managed_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "max_files_open": 0,
  "message_buffer_autoclear": true,
  "message_ids": [
    "string"
  ],
  "metadata": {
    "foo": "bar"
  },
  "model": "model",
  "model_settings": {
    "max_output_tokens": 0,
    "parallel_tool_calls": true,
    "provider_type": "openai",
    "reasoning": {
      "reasoning_effort": "none"
    },
    "response_format": {
      "type": "text"
    },
    "strict": true,
    "temperature": 0
  },
  "multi_agent_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "pending_approval": {
    "id": "id",
    "date": "2019-12-27T18:11:19.117Z",
    "tool_call": {
      "arguments": "arguments",
      "name": "name",
      "tool_call_id": "tool_call_id"
    },
    "is_err": true,
    "message_type": "approval_request_message",
    "name": "name",
    "otid": "otid",
    "run_id": "run_id",
    "sender_id": "sender_id",
    "seq_id": 0,
    "step_id": "step_id",
    "tool_calls": [
      {
        "arguments": "arguments",
        "name": "name",
        "tool_call_id": "tool_call_id"
      }
    ]
  },
  "per_file_view_window_char_limit": 0,
  "project_id": "project_id",
  "response_format": {
    "type": "text"
  },
  "secrets": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "template_id": "template_id",
  "timezone": "timezone",
  "tool_exec_environment_variables": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "tool_rules": [
    {
      "children": [
        "string"
      ],
      "tool_name": "tool_name",
      "child_arg_nodes": [
        {
          "name": "name",
          "args": {
            "foo": "bar"
          }
        }
      ],
      "prompt_template": "prompt_template",
      "type": "constrain_child_tools"
    }
  ],
  "updated_at": "2019-12-27T18:11:19.117Z"
}
```

## Domain Types

### Block

- `Block`

  A Block represents a reserved section of the LLM's context window.

  - `value: string`

    Value of the block.

  - `id?: string`

    The human-friendly ID of the Block

  - `base_template_id?: string | null`

    The base template id of the block.

  - `created_by_id?: string | null`

    The id of the user that made this Block.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    Description of the block.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the block will be hidden.

  - `is_template?: boolean`

    Whether the block is a template (e.g. saved human/persona options).

  - `label?: string | null`

    Label of the block (e.g. 'human', 'persona') in the context window.

  - `last_updated_by_id?: string | null`

    The id of the user that last updated this Block.

  - `limit?: number`

    Character limit of the block.

  - `metadata?: Record<string, unknown> | null`

    Metadata of the block.

  - `preserve_on_migration?: boolean | null`

    Preserve the block on template migration.

  - `project_id?: string | null`

    The associated project id.

  - `read_only?: boolean`

    Whether the agent has read-only access to the block.

  - `tags?: Array<string> | null`

    The tags associated with the block.

  - `template_id?: string | null`

    The id of the template.

  - `template_name?: string | null`

    Name of the block if it is a template.

### Block Update

- `BlockUpdate`

  Update a block

  - `base_template_id?: string | null`

    The base template id of the block.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    Description of the block.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the block will be hidden.

  - `is_template?: boolean`

    Whether the block is a template (e.g. saved human/persona options).

  - `label?: string | null`

    Label of the block (e.g. 'human', 'persona') in the context window.

  - `limit?: number | null`

    Character limit of the block.

  - `metadata?: Record<string, unknown> | null`

    Metadata of the block.

  - `preserve_on_migration?: boolean | null`

    Preserve the block on template migration.

  - `project_id?: string | null`

    The associated project id.

  - `read_only?: boolean`

    Whether the agent has read-only access to the block.

  - `tags?: Array<string> | null`

    The tags to associate with the block.

  - `template_id?: string | null`

    The id of the template.

  - `template_name?: string | null`

    Name of the block if it is a template.

  - `value?: string | null`

    Value of the block.

# Tools

## List Tools For Agent

`client.agents.tools.list(stringagentID, ToolListParamsquery?, RequestOptionsoptions?): ArrayPage<Tool>`

**get** `/v1/agents/{agent_id}/tools`

Get tools from an existing agent.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `query: ToolListParams`

  - `after?: string | null`

    Cursor for pagination (tool ID). Returns results relative to this ID in the specified sort order. Expected format: 'tool-<uuid4>'

  - `before?: string | null`

    Cursor for pagination (tool ID). Returns results relative to this ID in the specified sort order. Expected format: 'tool-<uuid4>'

  - `limit?: number | null`

    Maximum number of tools to return

  - `order?: "asc" | "desc"`

    Sort order for tools by creation time. 'asc' for oldest first, 'desc' for newest first

    - `"asc"`

    - `"desc"`

  - `order_by?: "created_at"`

    Field to sort by

    - `"created_at"`

### Returns

- `Tool`

  Representation of a tool, which is a function that can be called by the agent.

  - `id: string`

    The human-friendly ID of the Tool

  - `args_json_schema?: Record<string, unknown> | null`

    The args JSON schema of the function.

  - `created_by_id?: string | null`

    The id of the user that made this Tool.

  - `default_requires_approval?: boolean | null`

    Default value for whether or not executing this tool requires approval.

  - `description?: string | null`

    The description of the tool.

  - `enable_parallel_execution?: boolean | null`

    If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

  - `json_schema?: Record<string, unknown> | null`

    The JSON schema of the function.

  - `last_updated_by_id?: string | null`

    The id of the user that made this Tool.

  - `metadata_?: Record<string, unknown> | null`

    A dictionary of additional metadata for the tool.

  - `name?: string | null`

    The name of the function.

  - `npm_requirements?: Array<NpmRequirement> | null`

    Optional list of npm packages required by this tool.

    - `name: string`

      Name of the npm package.

    - `version?: string | null`

      Optional version of the package, following semantic versioning.

  - `pip_requirements?: Array<PipRequirement> | null`

    Optional list of pip packages required by this tool.

    - `name: string`

      Name of the pip package.

    - `version?: string | null`

      Optional version of the package, following semantic versioning.

  - `project_id?: string | null`

    The project id of the tool.

  - `return_char_limit?: number`

    The maximum number of characters in the response.

  - `source_code?: string | null`

    The source code of the function.

  - `source_type?: string | null`

    The type of the source code.

  - `tags?: Array<string>`

    Metadata tags.

  - `tool_type?: ToolType`

    The type of the tool.

    - `"custom"`

    - `"letta_core"`

    - `"letta_memory_core"`

    - `"letta_multi_agent_core"`

    - `"letta_sleeptime_core"`

    - `"letta_voice_sleeptime_core"`

    - `"letta_builtin"`

    - `"letta_files_core"`

    - `"external_langchain"`

    - `"external_composio"`

    - `"external_mcp"`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

// Automatically fetches more pages as needed.
for await (const tool of client.agents.tools.list('agent-123e4567-e89b-42d3-8456-426614174000')) {
  console.log(tool.id);
}
```

#### Response

```json
[
  {
    "id": "tool-123e4567-e89b-12d3-a456-426614174000",
    "args_json_schema": {
      "foo": "bar"
    },
    "created_by_id": "created_by_id",
    "default_requires_approval": true,
    "description": "description",
    "enable_parallel_execution": true,
    "json_schema": {
      "foo": "bar"
    },
    "last_updated_by_id": "last_updated_by_id",
    "metadata_": {
      "foo": "bar"
    },
    "name": "name",
    "npm_requirements": [
      {
        "name": "x",
        "version": "version"
      }
    ],
    "pip_requirements": [
      {
        "name": "x",
        "version": "version"
      }
    ],
    "project_id": "project_id",
    "return_char_limit": 1,
    "source_code": "source_code",
    "source_type": "source_type",
    "tags": [
      "string"
    ],
    "tool_type": "custom"
  }
]
```

## Attach Tool To Agent

`client.agents.tools.attach(stringtoolID, ToolAttachParamsparams, RequestOptionsoptions?): AgentState | null`

**patch** `/v1/agents/{agent_id}/tools/attach/{tool_id}`

Attach a tool to an agent.

### Parameters

- `toolID: string`

  The ID of the tool in the format 'tool-<uuid4>'

- `params: ToolAttachParams`

  - `agent_id: string`

    The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `AgentState | null`

  - `id: string`

    The id of the agent. Assigned by the database.

  - `agent_type: AgentType`

    The type of agent.

    - `"memgpt_agent"`

    - `"memgpt_v2_agent"`

    - `"letta_v1_agent"`

    - `"react_agent"`

    - `"workflow_agent"`

    - `"split_thread_agent"`

    - `"sleeptime_agent"`

    - `"voice_convo_agent"`

    - `"voice_sleeptime_agent"`

  - `blocks: Array<Block>`

    The memory blocks used by the agent.

    - `value: string`

      Value of the block.

    - `id?: string`

      The human-friendly ID of the Block

    - `base_template_id?: string | null`

      The base template id of the block.

    - `created_by_id?: string | null`

      The id of the user that made this Block.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      Description of the block.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the block will be hidden.

    - `is_template?: boolean`

      Whether the block is a template (e.g. saved human/persona options).

    - `label?: string | null`

      Label of the block (e.g. 'human', 'persona') in the context window.

    - `last_updated_by_id?: string | null`

      The id of the user that last updated this Block.

    - `limit?: number`

      Character limit of the block.

    - `metadata?: Record<string, unknown> | null`

      Metadata of the block.

    - `preserve_on_migration?: boolean | null`

      Preserve the block on template migration.

    - `project_id?: string | null`

      The associated project id.

    - `read_only?: boolean`

      Whether the agent has read-only access to the block.

    - `tags?: Array<string> | null`

      The tags associated with the block.

    - `template_id?: string | null`

      The id of the template.

    - `template_name?: string | null`

      Name of the block if it is a template.

  - `llm_config: LlmConfig`

    Deprecated: Use `model` field instead. The LLM configuration used by the agent.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

        - `type?: "text"`

          The type of the response format.

          - `"text"`

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

        - `json_schema: Record<string, unknown>`

          The JSON schema of the response.

        - `type?: "json_schema"`

          The type of the response format.

          - `"json_schema"`

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

        - `type?: "json_object"`

          The type of the response format.

          - `"json_object"`

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `memory: Memory`

    Deprecated: Use `blocks` field instead. The in-context memory of the agent.

    - `blocks: Array<Block>`

      Memory blocks contained in the agent's in-context memory

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `agent_type?: AgentType | (string & {}) | null`

      Agent type controlling prompt rendering.

      - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

        Enum to represent the type of agent.

      - `(string & {})`

    - `file_blocks?: Array<FileBlock>`

      Special blocks representing the agent's in-context memory of an attached file

      - `file_id: string`

        Unique identifier of the file.

      - `is_open: boolean`

        True if the agent currently has the file open.

      - `source_id: string`

        Deprecated: Use `folder_id` field instead. Unique identifier of the source.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_accessed_at?: string | null`

        UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `git_enabled?: boolean`

      Whether this agent uses git-backed memory with structured labels.

    - `prompt_template?: string`

      Deprecated. Ignored for performance.

  - `name: string`

    The name of the agent.

  - `sources: Array<Source>`

    Deprecated: Use `folders` field instead. The sources used by the agent.

    - `id: string`

      The human-friendly ID of the Source

    - `embedding_config: EmbeddingConfig`

      The embedding configuration used by the source.

      - `embedding_dim: number`

        The dimension of the embedding.

      - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"bedrock"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"mistral"`

        - `"together"`

        - `"pinecone"`

      - `embedding_model: string`

        The model for the embedding.

      - `azure_deployment?: string | null`

        The Azure deployment for the model.

      - `azure_endpoint?: string | null`

        The Azure endpoint for the model.

      - `azure_version?: string | null`

        The Azure version for the model.

      - `batch_size?: number`

        The maximum batch size for processing embeddings.

      - `embedding_chunk_size?: number | null`

        The chunk size of the embedding.

      - `embedding_endpoint?: string | null`

        The endpoint for the model (`None` if local).

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

    - `name: string`

      The name of the source.

    - `created_at?: string | null`

      The timestamp when the source was created.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `description?: string | null`

      The description of the source.

    - `instructions?: string | null`

      Instructions for how to use the source.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata?: Record<string, unknown> | null`

      Metadata associated with the source.

    - `updated_at?: string | null`

      The timestamp when the source was last updated.

    - `vector_db_provider?: VectorDBProvider`

      The vector database provider used for this source's passages

      - `"native"`

      - `"tpuf"`

      - `"pinecone"`

  - `system: string`

    The system prompt used by the agent.

  - `tags: Array<string>`

    The tags associated with the agent.

  - `tools: Array<Tool>`

    The tools used by the agent.

    - `id: string`

      The human-friendly ID of the Tool

    - `args_json_schema?: Record<string, unknown> | null`

      The args JSON schema of the function.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `default_requires_approval?: boolean | null`

      Default value for whether or not executing this tool requires approval.

    - `description?: string | null`

      The description of the tool.

    - `enable_parallel_execution?: boolean | null`

      If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

    - `json_schema?: Record<string, unknown> | null`

      The JSON schema of the function.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata_?: Record<string, unknown> | null`

      A dictionary of additional metadata for the tool.

    - `name?: string | null`

      The name of the function.

    - `npm_requirements?: Array<NpmRequirement> | null`

      Optional list of npm packages required by this tool.

      - `name: string`

        Name of the npm package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `pip_requirements?: Array<PipRequirement> | null`

      Optional list of pip packages required by this tool.

      - `name: string`

        Name of the pip package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `project_id?: string | null`

      The project id of the tool.

    - `return_char_limit?: number`

      The maximum number of characters in the response.

    - `source_code?: string | null`

      The source code of the function.

    - `source_type?: string | null`

      The type of the source code.

    - `tags?: Array<string>`

      Metadata tags.

    - `tool_type?: ToolType`

      The type of the tool.

      - `"custom"`

      - `"letta_core"`

      - `"letta_memory_core"`

      - `"letta_multi_agent_core"`

      - `"letta_sleeptime_core"`

      - `"letta_voice_sleeptime_core"`

      - `"letta_builtin"`

      - `"letta_files_core"`

      - `"external_langchain"`

      - `"external_composio"`

      - `"external_mcp"`

  - `base_template_id?: string | null`

    The base template id of the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `created_at?: string | null`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the agent will be hidden.

  - `identities?: Array<Identity>`

    The identities associated with this agent.

    - `id: string`

      The human-friendly ID of the Identity

    - `agent_ids: Array<string>`

      The IDs of the agents associated with the identity.

    - `block_ids: Array<string>`

      The IDs of the blocks associated with the identity.

    - `identifier_key: string`

      External, user-generated identifier key of the identity.

    - `identity_type: "org" | "user" | "other"`

      The type of the identity.

      - `"org"`

      - `"user"`

      - `"other"`

    - `name: string`

      The name of the identity.

    - `project_id?: string | null`

      The project id of the identity, if applicable.

    - `properties?: Array<Property>`

      List of properties associated with the identity

      - `key: string`

        The key of the property

      - `type: "string" | "number" | "boolean" | "json"`

        The type of the property

        - `"string"`

        - `"number"`

        - `"boolean"`

        - `"json"`

      - `value: string | number | boolean | Record<string, unknown>`

        The value of the property

        - `string`

        - `number`

        - `boolean`

        - `Record<string, unknown>`

  - `identity_ids?: Array<string>`

    Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

  - `last_run_completion?: string | null`

    The timestamp when the agent last completed a run.

  - `last_run_duration_ms?: number | null`

    The duration in milliseconds of the agent's last run.

  - `last_stop_reason?: StopReasonType | null`

    The stop reason from the agent's last run.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `managed_group?: ManagedGroup | null`

    The multi-agent group that this agent manages

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `message_buffer_autoclear?: boolean`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `message_ids?: Array<string> | null`

    The ids of the messages in the agent's in-context memory.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle used by the agent (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings used by the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `multi_agent_group?: MultiAgentGroup | null`

    Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `pending_approval?: ApprovalRequestMessage | null`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project_id?: string | null`

    The id of the project the agent belongs to.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format used by the agent

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Array<AgentEnvironmentVariable>`

    The environment variables for tool execution specific to this agent.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `template_id?: string | null`

    The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

    Deprecated: use `secrets` field instead.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The list of tool rules.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const agentState = await client.agents.tools.attach('tool-123e4567-e89b-42d3-8456-426614174000', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
});

console.log(agentState.id);
```

#### Response

```json
{
  "id": "id",
  "agent_type": "memgpt_agent",
  "blocks": [
    {
      "value": "value",
      "id": "block-123e4567-e89b-12d3-a456-426614174000",
      "base_template_id": "base_template_id",
      "created_by_id": "created_by_id",
      "deployment_id": "deployment_id",
      "description": "description",
      "entity_id": "entity_id",
      "hidden": true,
      "is_template": true,
      "label": "label",
      "last_updated_by_id": "last_updated_by_id",
      "limit": 0,
      "metadata": {
        "foo": "bar"
      },
      "preserve_on_migration": true,
      "project_id": "project_id",
      "read_only": true,
      "tags": [
        "string"
      ],
      "template_id": "template_id",
      "template_name": "template_name"
    }
  ],
  "llm_config": {
    "context_window": 0,
    "model": "model",
    "model_endpoint_type": "openai",
    "compatibility_type": "gguf",
    "display_name": "display_name",
    "effort": "low",
    "enable_reasoner": true,
    "frequency_penalty": 0,
    "handle": "handle",
    "max_reasoning_tokens": 0,
    "max_tokens": 0,
    "model_endpoint": "model_endpoint",
    "model_wrapper": "model_wrapper",
    "parallel_tool_calls": true,
    "provider_category": "base",
    "provider_name": "provider_name",
    "put_inner_thoughts_in_kwargs": true,
    "reasoning_effort": "none",
    "response_format": {
      "type": "text"
    },
    "return_logprobs": true,
    "return_token_ids": true,
    "strict": true,
    "temperature": 0,
    "tier": "tier",
    "tool_call_parser": "tool_call_parser",
    "top_logprobs": 0,
    "verbosity": "low"
  },
  "memory": {
    "blocks": [
      {
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "agent_type": "memgpt_agent",
    "file_blocks": [
      {
        "file_id": "file_id",
        "is_open": true,
        "source_id": "source_id",
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_accessed_at": "2019-12-27T18:11:19.117Z",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "git_enabled": true,
    "prompt_template": "prompt_template"
  },
  "name": "name",
  "sources": [
    {
      "id": "source-123e4567-e89b-12d3-a456-426614174000",
      "embedding_config": {
        "embedding_dim": 0,
        "embedding_endpoint_type": "openai",
        "embedding_model": "embedding_model",
        "azure_deployment": "azure_deployment",
        "azure_endpoint": "azure_endpoint",
        "azure_version": "azure_version",
        "batch_size": 0,
        "embedding_chunk_size": 0,
        "embedding_endpoint": "embedding_endpoint",
        "handle": "handle"
      },
      "name": "name",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "instructions": "instructions",
      "last_updated_by_id": "last_updated_by_id",
      "metadata": {
        "foo": "bar"
      },
      "updated_at": "2019-12-27T18:11:19.117Z",
      "vector_db_provider": "native"
    }
  ],
  "system": "system",
  "tags": [
    "string"
  ],
  "tools": [
    {
      "id": "tool-123e4567-e89b-12d3-a456-426614174000",
      "args_json_schema": {
        "foo": "bar"
      },
      "created_by_id": "created_by_id",
      "default_requires_approval": true,
      "description": "description",
      "enable_parallel_execution": true,
      "json_schema": {
        "foo": "bar"
      },
      "last_updated_by_id": "last_updated_by_id",
      "metadata_": {
        "foo": "bar"
      },
      "name": "name",
      "npm_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "pip_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "project_id": "project_id",
      "return_char_limit": 1,
      "source_code": "source_code",
      "source_type": "source_type",
      "tags": [
        "string"
      ],
      "tool_type": "custom"
    }
  ],
  "base_template_id": "base_template_id",
  "compaction_settings": {
    "clip_chars": 0,
    "mode": "all",
    "model": "model",
    "model_settings": {
      "max_output_tokens": 0,
      "parallel_tool_calls": true,
      "provider_type": "openai",
      "reasoning": {
        "reasoning_effort": "none"
      },
      "response_format": {
        "type": "text"
      },
      "strict": true,
      "temperature": 0
    },
    "prompt": "prompt",
    "prompt_acknowledgement": true,
    "sliding_window_percentage": 0
  },
  "created_at": "2019-12-27T18:11:19.117Z",
  "created_by_id": "created_by_id",
  "deployment_id": "deployment_id",
  "description": "description",
  "embedding": "embedding",
  "embedding_config": {
    "embedding_dim": 0,
    "embedding_endpoint_type": "openai",
    "embedding_model": "embedding_model",
    "azure_deployment": "azure_deployment",
    "azure_endpoint": "azure_endpoint",
    "azure_version": "azure_version",
    "batch_size": 0,
    "embedding_chunk_size": 0,
    "embedding_endpoint": "embedding_endpoint",
    "handle": "handle"
  },
  "enable_sleeptime": true,
  "entity_id": "entity_id",
  "hidden": true,
  "identities": [
    {
      "id": "identity-123e4567-e89b-12d3-a456-426614174000",
      "agent_ids": [
        "string"
      ],
      "block_ids": [
        "string"
      ],
      "identifier_key": "identifier_key",
      "identity_type": "org",
      "name": "name",
      "project_id": "project_id",
      "properties": [
        {
          "key": "key",
          "type": "string",
          "value": "string"
        }
      ]
    }
  ],
  "identity_ids": [
    "string"
  ],
  "last_run_completion": "2019-12-27T18:11:19.117Z",
  "last_run_duration_ms": 0,
  "last_stop_reason": "end_turn",
  "last_updated_by_id": "last_updated_by_id",
  "managed_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "max_files_open": 0,
  "message_buffer_autoclear": true,
  "message_ids": [
    "string"
  ],
  "metadata": {
    "foo": "bar"
  },
  "model": "model",
  "model_settings": {
    "max_output_tokens": 0,
    "parallel_tool_calls": true,
    "provider_type": "openai",
    "reasoning": {
      "reasoning_effort": "none"
    },
    "response_format": {
      "type": "text"
    },
    "strict": true,
    "temperature": 0
  },
  "multi_agent_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "pending_approval": {
    "id": "id",
    "date": "2019-12-27T18:11:19.117Z",
    "tool_call": {
      "arguments": "arguments",
      "name": "name",
      "tool_call_id": "tool_call_id"
    },
    "is_err": true,
    "message_type": "approval_request_message",
    "name": "name",
    "otid": "otid",
    "run_id": "run_id",
    "sender_id": "sender_id",
    "seq_id": 0,
    "step_id": "step_id",
    "tool_calls": [
      {
        "arguments": "arguments",
        "name": "name",
        "tool_call_id": "tool_call_id"
      }
    ]
  },
  "per_file_view_window_char_limit": 0,
  "project_id": "project_id",
  "response_format": {
    "type": "text"
  },
  "secrets": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "template_id": "template_id",
  "timezone": "timezone",
  "tool_exec_environment_variables": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "tool_rules": [
    {
      "children": [
        "string"
      ],
      "tool_name": "tool_name",
      "child_arg_nodes": [
        {
          "name": "name",
          "args": {
            "foo": "bar"
          }
        }
      ],
      "prompt_template": "prompt_template",
      "type": "constrain_child_tools"
    }
  ],
  "updated_at": "2019-12-27T18:11:19.117Z"
}
```

## Detach Tool From Agent

`client.agents.tools.detach(stringtoolID, ToolDetachParamsparams, RequestOptionsoptions?): AgentState | null`

**patch** `/v1/agents/{agent_id}/tools/detach/{tool_id}`

Detach a tool from an agent.

### Parameters

- `toolID: string`

  The ID of the tool in the format 'tool-<uuid4>'

- `params: ToolDetachParams`

  - `agent_id: string`

    The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `AgentState | null`

  - `id: string`

    The id of the agent. Assigned by the database.

  - `agent_type: AgentType`

    The type of agent.

    - `"memgpt_agent"`

    - `"memgpt_v2_agent"`

    - `"letta_v1_agent"`

    - `"react_agent"`

    - `"workflow_agent"`

    - `"split_thread_agent"`

    - `"sleeptime_agent"`

    - `"voice_convo_agent"`

    - `"voice_sleeptime_agent"`

  - `blocks: Array<Block>`

    The memory blocks used by the agent.

    - `value: string`

      Value of the block.

    - `id?: string`

      The human-friendly ID of the Block

    - `base_template_id?: string | null`

      The base template id of the block.

    - `created_by_id?: string | null`

      The id of the user that made this Block.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      Description of the block.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the block will be hidden.

    - `is_template?: boolean`

      Whether the block is a template (e.g. saved human/persona options).

    - `label?: string | null`

      Label of the block (e.g. 'human', 'persona') in the context window.

    - `last_updated_by_id?: string | null`

      The id of the user that last updated this Block.

    - `limit?: number`

      Character limit of the block.

    - `metadata?: Record<string, unknown> | null`

      Metadata of the block.

    - `preserve_on_migration?: boolean | null`

      Preserve the block on template migration.

    - `project_id?: string | null`

      The associated project id.

    - `read_only?: boolean`

      Whether the agent has read-only access to the block.

    - `tags?: Array<string> | null`

      The tags associated with the block.

    - `template_id?: string | null`

      The id of the template.

    - `template_name?: string | null`

      Name of the block if it is a template.

  - `llm_config: LlmConfig`

    Deprecated: Use `model` field instead. The LLM configuration used by the agent.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

        - `type?: "text"`

          The type of the response format.

          - `"text"`

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

        - `json_schema: Record<string, unknown>`

          The JSON schema of the response.

        - `type?: "json_schema"`

          The type of the response format.

          - `"json_schema"`

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

        - `type?: "json_object"`

          The type of the response format.

          - `"json_object"`

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `memory: Memory`

    Deprecated: Use `blocks` field instead. The in-context memory of the agent.

    - `blocks: Array<Block>`

      Memory blocks contained in the agent's in-context memory

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `agent_type?: AgentType | (string & {}) | null`

      Agent type controlling prompt rendering.

      - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

        Enum to represent the type of agent.

      - `(string & {})`

    - `file_blocks?: Array<FileBlock>`

      Special blocks representing the agent's in-context memory of an attached file

      - `file_id: string`

        Unique identifier of the file.

      - `is_open: boolean`

        True if the agent currently has the file open.

      - `source_id: string`

        Deprecated: Use `folder_id` field instead. Unique identifier of the source.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_accessed_at?: string | null`

        UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `git_enabled?: boolean`

      Whether this agent uses git-backed memory with structured labels.

    - `prompt_template?: string`

      Deprecated. Ignored for performance.

  - `name: string`

    The name of the agent.

  - `sources: Array<Source>`

    Deprecated: Use `folders` field instead. The sources used by the agent.

    - `id: string`

      The human-friendly ID of the Source

    - `embedding_config: EmbeddingConfig`

      The embedding configuration used by the source.

      - `embedding_dim: number`

        The dimension of the embedding.

      - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"bedrock"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"mistral"`

        - `"together"`

        - `"pinecone"`

      - `embedding_model: string`

        The model for the embedding.

      - `azure_deployment?: string | null`

        The Azure deployment for the model.

      - `azure_endpoint?: string | null`

        The Azure endpoint for the model.

      - `azure_version?: string | null`

        The Azure version for the model.

      - `batch_size?: number`

        The maximum batch size for processing embeddings.

      - `embedding_chunk_size?: number | null`

        The chunk size of the embedding.

      - `embedding_endpoint?: string | null`

        The endpoint for the model (`None` if local).

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

    - `name: string`

      The name of the source.

    - `created_at?: string | null`

      The timestamp when the source was created.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `description?: string | null`

      The description of the source.

    - `instructions?: string | null`

      Instructions for how to use the source.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata?: Record<string, unknown> | null`

      Metadata associated with the source.

    - `updated_at?: string | null`

      The timestamp when the source was last updated.

    - `vector_db_provider?: VectorDBProvider`

      The vector database provider used for this source's passages

      - `"native"`

      - `"tpuf"`

      - `"pinecone"`

  - `system: string`

    The system prompt used by the agent.

  - `tags: Array<string>`

    The tags associated with the agent.

  - `tools: Array<Tool>`

    The tools used by the agent.

    - `id: string`

      The human-friendly ID of the Tool

    - `args_json_schema?: Record<string, unknown> | null`

      The args JSON schema of the function.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `default_requires_approval?: boolean | null`

      Default value for whether or not executing this tool requires approval.

    - `description?: string | null`

      The description of the tool.

    - `enable_parallel_execution?: boolean | null`

      If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

    - `json_schema?: Record<string, unknown> | null`

      The JSON schema of the function.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata_?: Record<string, unknown> | null`

      A dictionary of additional metadata for the tool.

    - `name?: string | null`

      The name of the function.

    - `npm_requirements?: Array<NpmRequirement> | null`

      Optional list of npm packages required by this tool.

      - `name: string`

        Name of the npm package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `pip_requirements?: Array<PipRequirement> | null`

      Optional list of pip packages required by this tool.

      - `name: string`

        Name of the pip package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `project_id?: string | null`

      The project id of the tool.

    - `return_char_limit?: number`

      The maximum number of characters in the response.

    - `source_code?: string | null`

      The source code of the function.

    - `source_type?: string | null`

      The type of the source code.

    - `tags?: Array<string>`

      Metadata tags.

    - `tool_type?: ToolType`

      The type of the tool.

      - `"custom"`

      - `"letta_core"`

      - `"letta_memory_core"`

      - `"letta_multi_agent_core"`

      - `"letta_sleeptime_core"`

      - `"letta_voice_sleeptime_core"`

      - `"letta_builtin"`

      - `"letta_files_core"`

      - `"external_langchain"`

      - `"external_composio"`

      - `"external_mcp"`

  - `base_template_id?: string | null`

    The base template id of the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `created_at?: string | null`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the agent will be hidden.

  - `identities?: Array<Identity>`

    The identities associated with this agent.

    - `id: string`

      The human-friendly ID of the Identity

    - `agent_ids: Array<string>`

      The IDs of the agents associated with the identity.

    - `block_ids: Array<string>`

      The IDs of the blocks associated with the identity.

    - `identifier_key: string`

      External, user-generated identifier key of the identity.

    - `identity_type: "org" | "user" | "other"`

      The type of the identity.

      - `"org"`

      - `"user"`

      - `"other"`

    - `name: string`

      The name of the identity.

    - `project_id?: string | null`

      The project id of the identity, if applicable.

    - `properties?: Array<Property>`

      List of properties associated with the identity

      - `key: string`

        The key of the property

      - `type: "string" | "number" | "boolean" | "json"`

        The type of the property

        - `"string"`

        - `"number"`

        - `"boolean"`

        - `"json"`

      - `value: string | number | boolean | Record<string, unknown>`

        The value of the property

        - `string`

        - `number`

        - `boolean`

        - `Record<string, unknown>`

  - `identity_ids?: Array<string>`

    Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

  - `last_run_completion?: string | null`

    The timestamp when the agent last completed a run.

  - `last_run_duration_ms?: number | null`

    The duration in milliseconds of the agent's last run.

  - `last_stop_reason?: StopReasonType | null`

    The stop reason from the agent's last run.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `managed_group?: ManagedGroup | null`

    The multi-agent group that this agent manages

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `message_buffer_autoclear?: boolean`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `message_ids?: Array<string> | null`

    The ids of the messages in the agent's in-context memory.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle used by the agent (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings used by the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `multi_agent_group?: MultiAgentGroup | null`

    Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `pending_approval?: ApprovalRequestMessage | null`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project_id?: string | null`

    The id of the project the agent belongs to.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format used by the agent

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Array<AgentEnvironmentVariable>`

    The environment variables for tool execution specific to this agent.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `template_id?: string | null`

    The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

    Deprecated: use `secrets` field instead.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The list of tool rules.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const agentState = await client.agents.tools.detach('tool-123e4567-e89b-42d3-8456-426614174000', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
});

console.log(agentState.id);
```

#### Response

```json
{
  "id": "id",
  "agent_type": "memgpt_agent",
  "blocks": [
    {
      "value": "value",
      "id": "block-123e4567-e89b-12d3-a456-426614174000",
      "base_template_id": "base_template_id",
      "created_by_id": "created_by_id",
      "deployment_id": "deployment_id",
      "description": "description",
      "entity_id": "entity_id",
      "hidden": true,
      "is_template": true,
      "label": "label",
      "last_updated_by_id": "last_updated_by_id",
      "limit": 0,
      "metadata": {
        "foo": "bar"
      },
      "preserve_on_migration": true,
      "project_id": "project_id",
      "read_only": true,
      "tags": [
        "string"
      ],
      "template_id": "template_id",
      "template_name": "template_name"
    }
  ],
  "llm_config": {
    "context_window": 0,
    "model": "model",
    "model_endpoint_type": "openai",
    "compatibility_type": "gguf",
    "display_name": "display_name",
    "effort": "low",
    "enable_reasoner": true,
    "frequency_penalty": 0,
    "handle": "handle",
    "max_reasoning_tokens": 0,
    "max_tokens": 0,
    "model_endpoint": "model_endpoint",
    "model_wrapper": "model_wrapper",
    "parallel_tool_calls": true,
    "provider_category": "base",
    "provider_name": "provider_name",
    "put_inner_thoughts_in_kwargs": true,
    "reasoning_effort": "none",
    "response_format": {
      "type": "text"
    },
    "return_logprobs": true,
    "return_token_ids": true,
    "strict": true,
    "temperature": 0,
    "tier": "tier",
    "tool_call_parser": "tool_call_parser",
    "top_logprobs": 0,
    "verbosity": "low"
  },
  "memory": {
    "blocks": [
      {
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "agent_type": "memgpt_agent",
    "file_blocks": [
      {
        "file_id": "file_id",
        "is_open": true,
        "source_id": "source_id",
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_accessed_at": "2019-12-27T18:11:19.117Z",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "git_enabled": true,
    "prompt_template": "prompt_template"
  },
  "name": "name",
  "sources": [
    {
      "id": "source-123e4567-e89b-12d3-a456-426614174000",
      "embedding_config": {
        "embedding_dim": 0,
        "embedding_endpoint_type": "openai",
        "embedding_model": "embedding_model",
        "azure_deployment": "azure_deployment",
        "azure_endpoint": "azure_endpoint",
        "azure_version": "azure_version",
        "batch_size": 0,
        "embedding_chunk_size": 0,
        "embedding_endpoint": "embedding_endpoint",
        "handle": "handle"
      },
      "name": "name",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "instructions": "instructions",
      "last_updated_by_id": "last_updated_by_id",
      "metadata": {
        "foo": "bar"
      },
      "updated_at": "2019-12-27T18:11:19.117Z",
      "vector_db_provider": "native"
    }
  ],
  "system": "system",
  "tags": [
    "string"
  ],
  "tools": [
    {
      "id": "tool-123e4567-e89b-12d3-a456-426614174000",
      "args_json_schema": {
        "foo": "bar"
      },
      "created_by_id": "created_by_id",
      "default_requires_approval": true,
      "description": "description",
      "enable_parallel_execution": true,
      "json_schema": {
        "foo": "bar"
      },
      "last_updated_by_id": "last_updated_by_id",
      "metadata_": {
        "foo": "bar"
      },
      "name": "name",
      "npm_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "pip_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "project_id": "project_id",
      "return_char_limit": 1,
      "source_code": "source_code",
      "source_type": "source_type",
      "tags": [
        "string"
      ],
      "tool_type": "custom"
    }
  ],
  "base_template_id": "base_template_id",
  "compaction_settings": {
    "clip_chars": 0,
    "mode": "all",
    "model": "model",
    "model_settings": {
      "max_output_tokens": 0,
      "parallel_tool_calls": true,
      "provider_type": "openai",
      "reasoning": {
        "reasoning_effort": "none"
      },
      "response_format": {
        "type": "text"
      },
      "strict": true,
      "temperature": 0
    },
    "prompt": "prompt",
    "prompt_acknowledgement": true,
    "sliding_window_percentage": 0
  },
  "created_at": "2019-12-27T18:11:19.117Z",
  "created_by_id": "created_by_id",
  "deployment_id": "deployment_id",
  "description": "description",
  "embedding": "embedding",
  "embedding_config": {
    "embedding_dim": 0,
    "embedding_endpoint_type": "openai",
    "embedding_model": "embedding_model",
    "azure_deployment": "azure_deployment",
    "azure_endpoint": "azure_endpoint",
    "azure_version": "azure_version",
    "batch_size": 0,
    "embedding_chunk_size": 0,
    "embedding_endpoint": "embedding_endpoint",
    "handle": "handle"
  },
  "enable_sleeptime": true,
  "entity_id": "entity_id",
  "hidden": true,
  "identities": [
    {
      "id": "identity-123e4567-e89b-12d3-a456-426614174000",
      "agent_ids": [
        "string"
      ],
      "block_ids": [
        "string"
      ],
      "identifier_key": "identifier_key",
      "identity_type": "org",
      "name": "name",
      "project_id": "project_id",
      "properties": [
        {
          "key": "key",
          "type": "string",
          "value": "string"
        }
      ]
    }
  ],
  "identity_ids": [
    "string"
  ],
  "last_run_completion": "2019-12-27T18:11:19.117Z",
  "last_run_duration_ms": 0,
  "last_stop_reason": "end_turn",
  "last_updated_by_id": "last_updated_by_id",
  "managed_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "max_files_open": 0,
  "message_buffer_autoclear": true,
  "message_ids": [
    "string"
  ],
  "metadata": {
    "foo": "bar"
  },
  "model": "model",
  "model_settings": {
    "max_output_tokens": 0,
    "parallel_tool_calls": true,
    "provider_type": "openai",
    "reasoning": {
      "reasoning_effort": "none"
    },
    "response_format": {
      "type": "text"
    },
    "strict": true,
    "temperature": 0
  },
  "multi_agent_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "pending_approval": {
    "id": "id",
    "date": "2019-12-27T18:11:19.117Z",
    "tool_call": {
      "arguments": "arguments",
      "name": "name",
      "tool_call_id": "tool_call_id"
    },
    "is_err": true,
    "message_type": "approval_request_message",
    "name": "name",
    "otid": "otid",
    "run_id": "run_id",
    "sender_id": "sender_id",
    "seq_id": 0,
    "step_id": "step_id",
    "tool_calls": [
      {
        "arguments": "arguments",
        "name": "name",
        "tool_call_id": "tool_call_id"
      }
    ]
  },
  "per_file_view_window_char_limit": 0,
  "project_id": "project_id",
  "response_format": {
    "type": "text"
  },
  "secrets": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "template_id": "template_id",
  "timezone": "timezone",
  "tool_exec_environment_variables": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "tool_rules": [
    {
      "children": [
        "string"
      ],
      "tool_name": "tool_name",
      "child_arg_nodes": [
        {
          "name": "name",
          "args": {
            "foo": "bar"
          }
        }
      ],
      "prompt_template": "prompt_template",
      "type": "constrain_child_tools"
    }
  ],
  "updated_at": "2019-12-27T18:11:19.117Z"
}
```

## Update Approval For Tool

`client.agents.tools.updateApproval(stringtoolName, ToolUpdateApprovalParamsparams, RequestOptionsoptions?): AgentState | null`

**patch** `/v1/agents/{agent_id}/tools/approval/{tool_name}`

Modify the approval requirement for a tool attached to an agent.

Accepts requires_approval via request body (preferred) or query parameter (deprecated).

### Parameters

- `toolName: string`

- `params: ToolUpdateApprovalParams`

  - `agent_id: string`

    Path param: The ID of the agent in the format 'agent-<uuid4>'

  - `query_requires_approval?: boolean | null`

    Query param: Whether the tool requires approval before execution

### Returns

- `AgentState | null`

  - `id: string`

    The id of the agent. Assigned by the database.

  - `agent_type: AgentType`

    The type of agent.

    - `"memgpt_agent"`

    - `"memgpt_v2_agent"`

    - `"letta_v1_agent"`

    - `"react_agent"`

    - `"workflow_agent"`

    - `"split_thread_agent"`

    - `"sleeptime_agent"`

    - `"voice_convo_agent"`

    - `"voice_sleeptime_agent"`

  - `blocks: Array<Block>`

    The memory blocks used by the agent.

    - `value: string`

      Value of the block.

    - `id?: string`

      The human-friendly ID of the Block

    - `base_template_id?: string | null`

      The base template id of the block.

    - `created_by_id?: string | null`

      The id of the user that made this Block.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      Description of the block.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the block will be hidden.

    - `is_template?: boolean`

      Whether the block is a template (e.g. saved human/persona options).

    - `label?: string | null`

      Label of the block (e.g. 'human', 'persona') in the context window.

    - `last_updated_by_id?: string | null`

      The id of the user that last updated this Block.

    - `limit?: number`

      Character limit of the block.

    - `metadata?: Record<string, unknown> | null`

      Metadata of the block.

    - `preserve_on_migration?: boolean | null`

      Preserve the block on template migration.

    - `project_id?: string | null`

      The associated project id.

    - `read_only?: boolean`

      Whether the agent has read-only access to the block.

    - `tags?: Array<string> | null`

      The tags associated with the block.

    - `template_id?: string | null`

      The id of the template.

    - `template_name?: string | null`

      Name of the block if it is a template.

  - `llm_config: LlmConfig`

    Deprecated: Use `model` field instead. The LLM configuration used by the agent.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

        - `type?: "text"`

          The type of the response format.

          - `"text"`

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

        - `json_schema: Record<string, unknown>`

          The JSON schema of the response.

        - `type?: "json_schema"`

          The type of the response format.

          - `"json_schema"`

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

        - `type?: "json_object"`

          The type of the response format.

          - `"json_object"`

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `memory: Memory`

    Deprecated: Use `blocks` field instead. The in-context memory of the agent.

    - `blocks: Array<Block>`

      Memory blocks contained in the agent's in-context memory

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `agent_type?: AgentType | (string & {}) | null`

      Agent type controlling prompt rendering.

      - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

        Enum to represent the type of agent.

      - `(string & {})`

    - `file_blocks?: Array<FileBlock>`

      Special blocks representing the agent's in-context memory of an attached file

      - `file_id: string`

        Unique identifier of the file.

      - `is_open: boolean`

        True if the agent currently has the file open.

      - `source_id: string`

        Deprecated: Use `folder_id` field instead. Unique identifier of the source.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_accessed_at?: string | null`

        UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `git_enabled?: boolean`

      Whether this agent uses git-backed memory with structured labels.

    - `prompt_template?: string`

      Deprecated. Ignored for performance.

  - `name: string`

    The name of the agent.

  - `sources: Array<Source>`

    Deprecated: Use `folders` field instead. The sources used by the agent.

    - `id: string`

      The human-friendly ID of the Source

    - `embedding_config: EmbeddingConfig`

      The embedding configuration used by the source.

      - `embedding_dim: number`

        The dimension of the embedding.

      - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"bedrock"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"mistral"`

        - `"together"`

        - `"pinecone"`

      - `embedding_model: string`

        The model for the embedding.

      - `azure_deployment?: string | null`

        The Azure deployment for the model.

      - `azure_endpoint?: string | null`

        The Azure endpoint for the model.

      - `azure_version?: string | null`

        The Azure version for the model.

      - `batch_size?: number`

        The maximum batch size for processing embeddings.

      - `embedding_chunk_size?: number | null`

        The chunk size of the embedding.

      - `embedding_endpoint?: string | null`

        The endpoint for the model (`None` if local).

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

    - `name: string`

      The name of the source.

    - `created_at?: string | null`

      The timestamp when the source was created.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `description?: string | null`

      The description of the source.

    - `instructions?: string | null`

      Instructions for how to use the source.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata?: Record<string, unknown> | null`

      Metadata associated with the source.

    - `updated_at?: string | null`

      The timestamp when the source was last updated.

    - `vector_db_provider?: VectorDBProvider`

      The vector database provider used for this source's passages

      - `"native"`

      - `"tpuf"`

      - `"pinecone"`

  - `system: string`

    The system prompt used by the agent.

  - `tags: Array<string>`

    The tags associated with the agent.

  - `tools: Array<Tool>`

    The tools used by the agent.

    - `id: string`

      The human-friendly ID of the Tool

    - `args_json_schema?: Record<string, unknown> | null`

      The args JSON schema of the function.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `default_requires_approval?: boolean | null`

      Default value for whether or not executing this tool requires approval.

    - `description?: string | null`

      The description of the tool.

    - `enable_parallel_execution?: boolean | null`

      If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

    - `json_schema?: Record<string, unknown> | null`

      The JSON schema of the function.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata_?: Record<string, unknown> | null`

      A dictionary of additional metadata for the tool.

    - `name?: string | null`

      The name of the function.

    - `npm_requirements?: Array<NpmRequirement> | null`

      Optional list of npm packages required by this tool.

      - `name: string`

        Name of the npm package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `pip_requirements?: Array<PipRequirement> | null`

      Optional list of pip packages required by this tool.

      - `name: string`

        Name of the pip package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `project_id?: string | null`

      The project id of the tool.

    - `return_char_limit?: number`

      The maximum number of characters in the response.

    - `source_code?: string | null`

      The source code of the function.

    - `source_type?: string | null`

      The type of the source code.

    - `tags?: Array<string>`

      Metadata tags.

    - `tool_type?: ToolType`

      The type of the tool.

      - `"custom"`

      - `"letta_core"`

      - `"letta_memory_core"`

      - `"letta_multi_agent_core"`

      - `"letta_sleeptime_core"`

      - `"letta_voice_sleeptime_core"`

      - `"letta_builtin"`

      - `"letta_files_core"`

      - `"external_langchain"`

      - `"external_composio"`

      - `"external_mcp"`

  - `base_template_id?: string | null`

    The base template id of the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `created_at?: string | null`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the agent will be hidden.

  - `identities?: Array<Identity>`

    The identities associated with this agent.

    - `id: string`

      The human-friendly ID of the Identity

    - `agent_ids: Array<string>`

      The IDs of the agents associated with the identity.

    - `block_ids: Array<string>`

      The IDs of the blocks associated with the identity.

    - `identifier_key: string`

      External, user-generated identifier key of the identity.

    - `identity_type: "org" | "user" | "other"`

      The type of the identity.

      - `"org"`

      - `"user"`

      - `"other"`

    - `name: string`

      The name of the identity.

    - `project_id?: string | null`

      The project id of the identity, if applicable.

    - `properties?: Array<Property>`

      List of properties associated with the identity

      - `key: string`

        The key of the property

      - `type: "string" | "number" | "boolean" | "json"`

        The type of the property

        - `"string"`

        - `"number"`

        - `"boolean"`

        - `"json"`

      - `value: string | number | boolean | Record<string, unknown>`

        The value of the property

        - `string`

        - `number`

        - `boolean`

        - `Record<string, unknown>`

  - `identity_ids?: Array<string>`

    Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

  - `last_run_completion?: string | null`

    The timestamp when the agent last completed a run.

  - `last_run_duration_ms?: number | null`

    The duration in milliseconds of the agent's last run.

  - `last_stop_reason?: StopReasonType | null`

    The stop reason from the agent's last run.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `managed_group?: ManagedGroup | null`

    The multi-agent group that this agent manages

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `message_buffer_autoclear?: boolean`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `message_ids?: Array<string> | null`

    The ids of the messages in the agent's in-context memory.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle used by the agent (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings used by the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `multi_agent_group?: MultiAgentGroup | null`

    Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `pending_approval?: ApprovalRequestMessage | null`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project_id?: string | null`

    The id of the project the agent belongs to.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format used by the agent

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Array<AgentEnvironmentVariable>`

    The environment variables for tool execution specific to this agent.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `template_id?: string | null`

    The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

    Deprecated: use `secrets` field instead.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The list of tool rules.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const agentState = await client.agents.tools.updateApproval('tool_name', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
  body_requires_approval: true,
});

console.log(agentState.id);
```

#### Response

```json
{
  "id": "id",
  "agent_type": "memgpt_agent",
  "blocks": [
    {
      "value": "value",
      "id": "block-123e4567-e89b-12d3-a456-426614174000",
      "base_template_id": "base_template_id",
      "created_by_id": "created_by_id",
      "deployment_id": "deployment_id",
      "description": "description",
      "entity_id": "entity_id",
      "hidden": true,
      "is_template": true,
      "label": "label",
      "last_updated_by_id": "last_updated_by_id",
      "limit": 0,
      "metadata": {
        "foo": "bar"
      },
      "preserve_on_migration": true,
      "project_id": "project_id",
      "read_only": true,
      "tags": [
        "string"
      ],
      "template_id": "template_id",
      "template_name": "template_name"
    }
  ],
  "llm_config": {
    "context_window": 0,
    "model": "model",
    "model_endpoint_type": "openai",
    "compatibility_type": "gguf",
    "display_name": "display_name",
    "effort": "low",
    "enable_reasoner": true,
    "frequency_penalty": 0,
    "handle": "handle",
    "max_reasoning_tokens": 0,
    "max_tokens": 0,
    "model_endpoint": "model_endpoint",
    "model_wrapper": "model_wrapper",
    "parallel_tool_calls": true,
    "provider_category": "base",
    "provider_name": "provider_name",
    "put_inner_thoughts_in_kwargs": true,
    "reasoning_effort": "none",
    "response_format": {
      "type": "text"
    },
    "return_logprobs": true,
    "return_token_ids": true,
    "strict": true,
    "temperature": 0,
    "tier": "tier",
    "tool_call_parser": "tool_call_parser",
    "top_logprobs": 0,
    "verbosity": "low"
  },
  "memory": {
    "blocks": [
      {
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "agent_type": "memgpt_agent",
    "file_blocks": [
      {
        "file_id": "file_id",
        "is_open": true,
        "source_id": "source_id",
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_accessed_at": "2019-12-27T18:11:19.117Z",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "git_enabled": true,
    "prompt_template": "prompt_template"
  },
  "name": "name",
  "sources": [
    {
      "id": "source-123e4567-e89b-12d3-a456-426614174000",
      "embedding_config": {
        "embedding_dim": 0,
        "embedding_endpoint_type": "openai",
        "embedding_model": "embedding_model",
        "azure_deployment": "azure_deployment",
        "azure_endpoint": "azure_endpoint",
        "azure_version": "azure_version",
        "batch_size": 0,
        "embedding_chunk_size": 0,
        "embedding_endpoint": "embedding_endpoint",
        "handle": "handle"
      },
      "name": "name",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "instructions": "instructions",
      "last_updated_by_id": "last_updated_by_id",
      "metadata": {
        "foo": "bar"
      },
      "updated_at": "2019-12-27T18:11:19.117Z",
      "vector_db_provider": "native"
    }
  ],
  "system": "system",
  "tags": [
    "string"
  ],
  "tools": [
    {
      "id": "tool-123e4567-e89b-12d3-a456-426614174000",
      "args_json_schema": {
        "foo": "bar"
      },
      "created_by_id": "created_by_id",
      "default_requires_approval": true,
      "description": "description",
      "enable_parallel_execution": true,
      "json_schema": {
        "foo": "bar"
      },
      "last_updated_by_id": "last_updated_by_id",
      "metadata_": {
        "foo": "bar"
      },
      "name": "name",
      "npm_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "pip_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "project_id": "project_id",
      "return_char_limit": 1,
      "source_code": "source_code",
      "source_type": "source_type",
      "tags": [
        "string"
      ],
      "tool_type": "custom"
    }
  ],
  "base_template_id": "base_template_id",
  "compaction_settings": {
    "clip_chars": 0,
    "mode": "all",
    "model": "model",
    "model_settings": {
      "max_output_tokens": 0,
      "parallel_tool_calls": true,
      "provider_type": "openai",
      "reasoning": {
        "reasoning_effort": "none"
      },
      "response_format": {
        "type": "text"
      },
      "strict": true,
      "temperature": 0
    },
    "prompt": "prompt",
    "prompt_acknowledgement": true,
    "sliding_window_percentage": 0
  },
  "created_at": "2019-12-27T18:11:19.117Z",
  "created_by_id": "created_by_id",
  "deployment_id": "deployment_id",
  "description": "description",
  "embedding": "embedding",
  "embedding_config": {
    "embedding_dim": 0,
    "embedding_endpoint_type": "openai",
    "embedding_model": "embedding_model",
    "azure_deployment": "azure_deployment",
    "azure_endpoint": "azure_endpoint",
    "azure_version": "azure_version",
    "batch_size": 0,
    "embedding_chunk_size": 0,
    "embedding_endpoint": "embedding_endpoint",
    "handle": "handle"
  },
  "enable_sleeptime": true,
  "entity_id": "entity_id",
  "hidden": true,
  "identities": [
    {
      "id": "identity-123e4567-e89b-12d3-a456-426614174000",
      "agent_ids": [
        "string"
      ],
      "block_ids": [
        "string"
      ],
      "identifier_key": "identifier_key",
      "identity_type": "org",
      "name": "name",
      "project_id": "project_id",
      "properties": [
        {
          "key": "key",
          "type": "string",
          "value": "string"
        }
      ]
    }
  ],
  "identity_ids": [
    "string"
  ],
  "last_run_completion": "2019-12-27T18:11:19.117Z",
  "last_run_duration_ms": 0,
  "last_stop_reason": "end_turn",
  "last_updated_by_id": "last_updated_by_id",
  "managed_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "max_files_open": 0,
  "message_buffer_autoclear": true,
  "message_ids": [
    "string"
  ],
  "metadata": {
    "foo": "bar"
  },
  "model": "model",
  "model_settings": {
    "max_output_tokens": 0,
    "parallel_tool_calls": true,
    "provider_type": "openai",
    "reasoning": {
      "reasoning_effort": "none"
    },
    "response_format": {
      "type": "text"
    },
    "strict": true,
    "temperature": 0
  },
  "multi_agent_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "pending_approval": {
    "id": "id",
    "date": "2019-12-27T18:11:19.117Z",
    "tool_call": {
      "arguments": "arguments",
      "name": "name",
      "tool_call_id": "tool_call_id"
    },
    "is_err": true,
    "message_type": "approval_request_message",
    "name": "name",
    "otid": "otid",
    "run_id": "run_id",
    "sender_id": "sender_id",
    "seq_id": 0,
    "step_id": "step_id",
    "tool_calls": [
      {
        "arguments": "arguments",
        "name": "name",
        "tool_call_id": "tool_call_id"
      }
    ]
  },
  "per_file_view_window_char_limit": 0,
  "project_id": "project_id",
  "response_format": {
    "type": "text"
  },
  "secrets": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "template_id": "template_id",
  "timezone": "timezone",
  "tool_exec_environment_variables": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "tool_rules": [
    {
      "children": [
        "string"
      ],
      "tool_name": "tool_name",
      "child_arg_nodes": [
        {
          "name": "name",
          "args": {
            "foo": "bar"
          }
        }
      ],
      "prompt_template": "prompt_template",
      "type": "constrain_child_tools"
    }
  ],
  "updated_at": "2019-12-27T18:11:19.117Z"
}
```

## Run Tool For Agent

`client.agents.tools.run(stringtoolName, ToolRunParamsparams, RequestOptionsoptions?): ToolExecutionResult`

**post** `/v1/agents/{agent_id}/tools/{tool_name}/run`

Trigger a tool by name on a specific agent, providing the necessary arguments.

This endpoint executes a tool that is attached to the agent, using the agent's
state and environment variables for execution context.

### Parameters

- `toolName: string`

- `params: ToolRunParams`

  - `agent_id: string`

    Path param: The ID of the agent in the format 'agent-<uuid4>'

  - `args?: Record<string, unknown>`

    Body param: Arguments to pass to the tool

### Returns

- `ToolExecutionResult`

  - `status: "success" | "error"`

    The status of the tool execution and return object

    - `"success"`

    - `"error"`

  - `agent_state?: AgentState | null`

    Representation of an agent's state. This is the state of the agent at a given time, and is persisted in the DB backend. The state has all the information needed to recreate a persisted agent.

    - `id: string`

      The id of the agent. Assigned by the database.

    - `agent_type: AgentType`

      The type of agent.

      - `"memgpt_agent"`

      - `"memgpt_v2_agent"`

      - `"letta_v1_agent"`

      - `"react_agent"`

      - `"workflow_agent"`

      - `"split_thread_agent"`

      - `"sleeptime_agent"`

      - `"voice_convo_agent"`

      - `"voice_sleeptime_agent"`

    - `blocks: Array<Block>`

      The memory blocks used by the agent.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `llm_config: LlmConfig`

      Deprecated: Use `model` field instead. The LLM configuration used by the agent.

      - `context_window: number`

        The context window size for the model.

      - `model: string`

        LLM model name.

      - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"lmstudio-chatcompletions"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"minimax"`

        - `"moonshot"`

        - `"moonshot_coding"`

        - `"mistral"`

        - `"together"`

        - `"bedrock"`

        - `"deepseek"`

        - `"xai"`

        - `"zai"`

        - `"zai_coding"`

        - `"baseten"`

        - `"fireworks"`

        - `"openrouter"`

        - `"chatgpt_oauth"`

      - `compatibility_type?: "gguf" | "mlx" | null`

        The framework compatibility type for the model.

        - `"gguf"`

        - `"mlx"`

      - `display_name?: string | null`

        A human-friendly display name for the model.

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `enable_reasoner?: boolean`

        Whether or not the model should use extended thinking if it is a 'reasoning' style model

      - `frequency_penalty?: number | null`

        Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

      - `max_reasoning_tokens?: number`

        Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

      - `max_tokens?: number | null`

        The maximum number of tokens to generate. If not set, the model will use its default value.

      - `model_endpoint?: string | null`

        The endpoint for the model.

      - `model_wrapper?: string | null`

        The wrapper for the model.

      - `parallel_tool_calls?: boolean | null`

        Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

      - `provider_category?: ProviderCategory | null`

        The provider category for the model.

        - `"base"`

        - `"byok"`

      - `provider_name?: string | null`

        The provider name for the model.

      - `put_inner_thoughts_in_kwargs?: boolean | null`

        Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

      - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

        The reasoning effort to use when generating text reasoning models

        - `"none"`

        - `"minimal"`

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

        - `TextResponseFormat`

          Response format for plain text responses.

          - `type?: "text"`

            The type of the response format.

            - `"text"`

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

          - `json_schema: Record<string, unknown>`

            The JSON schema of the response.

          - `type?: "json_schema"`

            The type of the response format.

            - `"json_schema"`

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

          - `type?: "json_object"`

            The type of the response format.

            - `"json_object"`

      - `return_logprobs?: boolean`

        Whether to return log probabilities of the output tokens. Useful for RL training.

      - `return_token_ids?: boolean`

        Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

      - `temperature?: number`

        The temperature to use when generating text with the model. A higher temperature will result in more random text.

      - `tier?: string | null`

        The cost tier for the model (cloud only).

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

      - `top_logprobs?: number | null`

        Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `memory: Memory`

      Deprecated: Use `blocks` field instead. The in-context memory of the agent.

      - `blocks: Array<Block>`

        Memory blocks contained in the agent's in-context memory

        - `value: string`

          Value of the block.

        - `id?: string`

          The human-friendly ID of the Block

        - `base_template_id?: string | null`

          The base template id of the block.

        - `created_by_id?: string | null`

          The id of the user that made this Block.

        - `deployment_id?: string | null`

          The id of the deployment.

        - `description?: string | null`

          Description of the block.

        - `entity_id?: string | null`

          The id of the entity within the template.

        - `hidden?: boolean | null`

          If set to True, the block will be hidden.

        - `is_template?: boolean`

          Whether the block is a template (e.g. saved human/persona options).

        - `label?: string | null`

          Label of the block (e.g. 'human', 'persona') in the context window.

        - `last_updated_by_id?: string | null`

          The id of the user that last updated this Block.

        - `limit?: number`

          Character limit of the block.

        - `metadata?: Record<string, unknown> | null`

          Metadata of the block.

        - `preserve_on_migration?: boolean | null`

          Preserve the block on template migration.

        - `project_id?: string | null`

          The associated project id.

        - `read_only?: boolean`

          Whether the agent has read-only access to the block.

        - `tags?: Array<string> | null`

          The tags associated with the block.

        - `template_id?: string | null`

          The id of the template.

        - `template_name?: string | null`

          Name of the block if it is a template.

      - `agent_type?: AgentType | (string & {}) | null`

        Agent type controlling prompt rendering.

        - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

          Enum to represent the type of agent.

        - `(string & {})`

      - `file_blocks?: Array<FileBlock>`

        Special blocks representing the agent's in-context memory of an attached file

        - `file_id: string`

          Unique identifier of the file.

        - `is_open: boolean`

          True if the agent currently has the file open.

        - `source_id: string`

          Deprecated: Use `folder_id` field instead. Unique identifier of the source.

        - `value: string`

          Value of the block.

        - `id?: string`

          The human-friendly ID of the Block

        - `base_template_id?: string | null`

          The base template id of the block.

        - `created_by_id?: string | null`

          The id of the user that made this Block.

        - `deployment_id?: string | null`

          The id of the deployment.

        - `description?: string | null`

          Description of the block.

        - `entity_id?: string | null`

          The id of the entity within the template.

        - `hidden?: boolean | null`

          If set to True, the block will be hidden.

        - `is_template?: boolean`

          Whether the block is a template (e.g. saved human/persona options).

        - `label?: string | null`

          Label of the block (e.g. 'human', 'persona') in the context window.

        - `last_accessed_at?: string | null`

          UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

        - `last_updated_by_id?: string | null`

          The id of the user that last updated this Block.

        - `limit?: number`

          Character limit of the block.

        - `metadata?: Record<string, unknown> | null`

          Metadata of the block.

        - `preserve_on_migration?: boolean | null`

          Preserve the block on template migration.

        - `project_id?: string | null`

          The associated project id.

        - `read_only?: boolean`

          Whether the agent has read-only access to the block.

        - `tags?: Array<string> | null`

          The tags associated with the block.

        - `template_id?: string | null`

          The id of the template.

        - `template_name?: string | null`

          Name of the block if it is a template.

      - `git_enabled?: boolean`

        Whether this agent uses git-backed memory with structured labels.

      - `prompt_template?: string`

        Deprecated. Ignored for performance.

    - `name: string`

      The name of the agent.

    - `sources: Array<Source>`

      Deprecated: Use `folders` field instead. The sources used by the agent.

      - `id: string`

        The human-friendly ID of the Source

      - `embedding_config: EmbeddingConfig`

        The embedding configuration used by the source.

        - `embedding_dim: number`

          The dimension of the embedding.

        - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

          The endpoint type for the model.

          - `"openai"`

          - `"anthropic"`

          - `"bedrock"`

          - `"google_ai"`

          - `"google_vertex"`

          - `"azure"`

          - `"groq"`

          - `"ollama"`

          - `"webui"`

          - `"webui-legacy"`

          - `"lmstudio"`

          - `"lmstudio-legacy"`

          - `"llamacpp"`

          - `"koboldcpp"`

          - `"vllm"`

          - `"hugging-face"`

          - `"mistral"`

          - `"together"`

          - `"pinecone"`

        - `embedding_model: string`

          The model for the embedding.

        - `azure_deployment?: string | null`

          The Azure deployment for the model.

        - `azure_endpoint?: string | null`

          The Azure endpoint for the model.

        - `azure_version?: string | null`

          The Azure version for the model.

        - `batch_size?: number`

          The maximum batch size for processing embeddings.

        - `embedding_chunk_size?: number | null`

          The chunk size of the embedding.

        - `embedding_endpoint?: string | null`

          The endpoint for the model (`None` if local).

        - `handle?: string | null`

          The handle for this config, in the format provider/model-name.

      - `name: string`

        The name of the source.

      - `created_at?: string | null`

        The timestamp when the source was created.

      - `created_by_id?: string | null`

        The id of the user that made this Tool.

      - `description?: string | null`

        The description of the source.

      - `instructions?: string | null`

        Instructions for how to use the source.

      - `last_updated_by_id?: string | null`

        The id of the user that made this Tool.

      - `metadata?: Record<string, unknown> | null`

        Metadata associated with the source.

      - `updated_at?: string | null`

        The timestamp when the source was last updated.

      - `vector_db_provider?: VectorDBProvider`

        The vector database provider used for this source's passages

        - `"native"`

        - `"tpuf"`

        - `"pinecone"`

    - `system: string`

      The system prompt used by the agent.

    - `tags: Array<string>`

      The tags associated with the agent.

    - `tools: Array<Tool>`

      The tools used by the agent.

      - `id: string`

        The human-friendly ID of the Tool

      - `args_json_schema?: Record<string, unknown> | null`

        The args JSON schema of the function.

      - `created_by_id?: string | null`

        The id of the user that made this Tool.

      - `default_requires_approval?: boolean | null`

        Default value for whether or not executing this tool requires approval.

      - `description?: string | null`

        The description of the tool.

      - `enable_parallel_execution?: boolean | null`

        If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

      - `json_schema?: Record<string, unknown> | null`

        The JSON schema of the function.

      - `last_updated_by_id?: string | null`

        The id of the user that made this Tool.

      - `metadata_?: Record<string, unknown> | null`

        A dictionary of additional metadata for the tool.

      - `name?: string | null`

        The name of the function.

      - `npm_requirements?: Array<NpmRequirement> | null`

        Optional list of npm packages required by this tool.

        - `name: string`

          Name of the npm package.

        - `version?: string | null`

          Optional version of the package, following semantic versioning.

      - `pip_requirements?: Array<PipRequirement> | null`

        Optional list of pip packages required by this tool.

        - `name: string`

          Name of the pip package.

        - `version?: string | null`

          Optional version of the package, following semantic versioning.

      - `project_id?: string | null`

        The project id of the tool.

      - `return_char_limit?: number`

        The maximum number of characters in the response.

      - `source_code?: string | null`

        The source code of the function.

      - `source_type?: string | null`

        The type of the source code.

      - `tags?: Array<string>`

        Metadata tags.

      - `tool_type?: ToolType`

        The type of the tool.

        - `"custom"`

        - `"letta_core"`

        - `"letta_memory_core"`

        - `"letta_multi_agent_core"`

        - `"letta_sleeptime_core"`

        - `"letta_voice_sleeptime_core"`

        - `"letta_builtin"`

        - `"letta_files_core"`

        - `"external_langchain"`

        - `"external_composio"`

        - `"external_mcp"`

    - `base_template_id?: string | null`

      The base template id of the agent.

    - `compaction_settings?: CompactionSettings | null`

      Configuration for conversation compaction / summarization.

      Per-model settings (temperature,
      max tokens, etc.) are derived from the default configuration for that handle.

      - `clip_chars?: number | null`

        The maximum length of the summary in characters. If none, no clipping is performed.

      - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

        The type of summarization technique use.

        - `"all"`

        - `"sliding_window"`

        - `"self_compact_all"`

        - `"self_compact_sliding_window"`

      - `model?: string | null`

        Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

      - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

        Optional model settings used to override defaults for the summarizer model.

        - `OpenAIModelSettings`

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "openai"`

            The type of the provider.

            - `"openai"`

          - `reasoning?: Reasoning`

            The reasoning configuration for the model.

            - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

              The reasoning effort to use when generating text reasoning models

              - `"none"`

              - `"minimal"`

              - `"low"`

              - `"medium"`

              - `"high"`

              - `"xhigh"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `strict?: boolean`

            Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

          - `temperature?: number`

            The temperature of the model.

        - `SgLangModelSettings`

          SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "sglang"`

            The type of the provider.

            - `"sglang"`

          - `reasoning?: Reasoning`

            The reasoning configuration for the model.

            - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

              The reasoning effort to use when generating text reasoning models

              - `"none"`

              - `"minimal"`

              - `"low"`

              - `"medium"`

              - `"high"`

              - `"xhigh"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `strict?: boolean`

            Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

          - `temperature?: number`

            The temperature of the model.

          - `tool_call_parser?: string | null`

            SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

        - `AnthropicModelSettings`

          - `effort?: "low" | "medium" | "high" | 2 more | null`

            Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

            - `"max"`

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "anthropic"`

            The type of the provider.

            - `"anthropic"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `strict?: boolean`

            Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

          - `temperature?: number`

            The temperature of the model.

          - `thinking?: Thinking`

            The thinking configuration for the model.

            - `budget_tokens?: number`

              The maximum number of tokens the model can use for extended thinking.

            - `type?: "enabled" | "disabled"`

              The type of thinking to use.

              - `"enabled"`

              - `"disabled"`

          - `verbosity?: "low" | "medium" | "high" | null`

            Soft control for how verbose model output should be, used for GPT-5 models.

            - `"low"`

            - `"medium"`

            - `"high"`

        - `GoogleAIModelSettings`

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "google_ai"`

            The type of the provider.

            - `"google_ai"`

          - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response schema for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

          - `thinking_config?: ThinkingConfig`

            The thinking configuration for the model.

            - `include_thoughts?: boolean`

              Whether to include thoughts in the model's response.

            - `thinking_budget?: number`

              The thinking budget for the model.

        - `GoogleVertexModelSettings`

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "google_vertex"`

            The type of the provider.

            - `"google_vertex"`

          - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response schema for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

          - `thinking_config?: ThinkingConfig`

            The thinking configuration for the model.

            - `include_thoughts?: boolean`

              Whether to include thoughts in the model's response.

            - `thinking_budget?: number`

              The thinking budget for the model.

        - `AzureModelSettings`

          Azure OpenAI model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "azure"`

            The type of the provider.

            - `"azure"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

        - `XaiModelSettings`

          xAI model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "xai"`

            The type of the provider.

            - `"xai"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

        - `MoonshotModelSettings`

          Moonshot/Kimi model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "moonshot"`

            The type of the provider.

            - `"moonshot"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `strict?: boolean`

            Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

          - `temperature?: number`

            The temperature of the model.

        - `ZaiModelSettings`

          Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "zai"`

            The type of the provider.

            - `"zai"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

          - `thinking?: Thinking`

            The thinking configuration for GLM-4.5+ models.

            - `clear_thinking?: boolean`

              If False, preserved thinking is used (recommended for agents).

            - `type?: "enabled" | "disabled"`

              Whether thinking is enabled or disabled.

              - `"enabled"`

              - `"disabled"`

        - `MoonshotCodingModelSettings`

          Kimi Code model configuration (Anthropic-compatible).

          - `effort?: "low" | "medium" | "high" | 2 more | null`

            Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

            - `"max"`

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "moonshot_coding"`

            The type of the provider.

            - `"moonshot_coding"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `strict?: boolean`

            Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

          - `temperature?: number`

            The temperature of the model.

          - `thinking?: Thinking`

            The thinking configuration for the model.

            - `budget_tokens?: number`

              The maximum number of tokens the model can use for extended thinking.

            - `type?: "enabled" | "disabled"`

              The type of thinking to use.

              - `"enabled"`

              - `"disabled"`

          - `verbosity?: "low" | "medium" | "high" | null`

            Soft control for how verbose model output should be, used for GPT-5 models.

            - `"low"`

            - `"medium"`

            - `"high"`

        - `GroqModelSettings`

          Groq model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "groq"`

            The type of the provider.

            - `"groq"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

        - `DeepseekModelSettings`

          Deepseek model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "deepseek"`

            The type of the provider.

            - `"deepseek"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

        - `TogetherModelSettings`

          Together AI model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "together"`

            The type of the provider.

            - `"together"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

        - `BedrockModelSettings`

          AWS Bedrock model configuration.

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "bedrock"`

            The type of the provider.

            - `"bedrock"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

        - `BasetenModelSettings`

          Baseten model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "baseten"`

            The type of the provider.

            - `"baseten"`

          - `temperature?: number`

            The temperature of the model.

        - `OpenRouterModelSettings`

          OpenRouter model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "openrouter"`

            The type of the provider.

            - `"openrouter"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

        - `ChatGptoAuthModelSettings`

          ChatGPT OAuth model configuration (uses ChatGPT backend API).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "chatgpt_oauth"`

            The type of the provider.

            - `"chatgpt_oauth"`

          - `reasoning?: Reasoning`

            The reasoning configuration for the model.

            - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

              The reasoning effort level for GPT-5.x and o-series models.

              - `"none"`

              - `"low"`

              - `"medium"`

              - `"high"`

              - `"xhigh"`

          - `temperature?: number`

            The temperature of the model.

      - `prompt?: string | null`

        The prompt to use for summarization. If None, uses mode-specific default.

      - `prompt_acknowledgement?: boolean`

        Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

      - `sliding_window_percentage?: number`

        The percentage of the context window to keep post-summarization (only used in sliding window modes).

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      The description of the agent.

    - `embedding?: string | null`

      The embedding model handle used by the agent (format: provider/model-name).

    - `embedding_config?: EmbeddingConfig | null`

      Configuration for embedding model connection and processing parameters.

    - `enable_sleeptime?: boolean | null`

      If set to True, memory management will move to a background agent thread.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the agent will be hidden.

    - `identities?: Array<Identity>`

      The identities associated with this agent.

      - `id: string`

        The human-friendly ID of the Identity

      - `agent_ids: Array<string>`

        The IDs of the agents associated with the identity.

      - `block_ids: Array<string>`

        The IDs of the blocks associated with the identity.

      - `identifier_key: string`

        External, user-generated identifier key of the identity.

      - `identity_type: "org" | "user" | "other"`

        The type of the identity.

        - `"org"`

        - `"user"`

        - `"other"`

      - `name: string`

        The name of the identity.

      - `project_id?: string | null`

        The project id of the identity, if applicable.

      - `properties?: Array<Property>`

        List of properties associated with the identity

        - `key: string`

          The key of the property

        - `type: "string" | "number" | "boolean" | "json"`

          The type of the property

          - `"string"`

          - `"number"`

          - `"boolean"`

          - `"json"`

        - `value: string | number | boolean | Record<string, unknown>`

          The value of the property

          - `string`

          - `number`

          - `boolean`

          - `Record<string, unknown>`

    - `identity_ids?: Array<string>`

      Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

    - `last_run_completion?: string | null`

      The timestamp when the agent last completed a run.

    - `last_run_duration_ms?: number | null`

      The duration in milliseconds of the agent's last run.

    - `last_stop_reason?: StopReasonType | null`

      The stop reason from the agent's last run.

      - `"end_turn"`

      - `"error"`

      - `"llm_api_error"`

      - `"invalid_llm_response"`

      - `"invalid_tool_call"`

      - `"max_steps"`

      - `"max_tokens_exceeded"`

      - `"no_tool_call"`

      - `"tool_rule"`

      - `"cancelled"`

      - `"insufficient_credits"`

      - `"requires_approval"`

      - `"context_window_overflow_in_system_prompt"`

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `managed_group?: ManagedGroup | null`

      The multi-agent group that this agent manages

      - `id: string`

        The id of the group. Assigned by the database.

      - `agent_ids: Array<string>`

      - `description: string`

      - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

        - `"round_robin"`

        - `"supervisor"`

        - `"dynamic"`

        - `"sleeptime"`

        - `"voice_sleeptime"`

        - `"swarm"`

      - `base_template_id?: string | null`

        The base template id.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `hidden?: boolean | null`

        If set to True, the group will be hidden.

      - `last_processed_message_id?: string | null`

      - `manager_agent_id?: string | null`

      - `max_message_buffer_length?: number | null`

        The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

      - `max_turns?: number | null`

      - `min_message_buffer_length?: number | null`

        The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

      - `project_id?: string | null`

        The associated project id.

      - `shared_block_ids?: Array<string>`

      - `sleeptime_agent_frequency?: number | null`

      - `template_id?: string | null`

        The id of the template.

      - `termination_token?: string | null`

      - `turns_counter?: number | null`

    - `max_files_open?: number | null`

      Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

    - `message_buffer_autoclear?: boolean`

      If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

    - `message_ids?: Array<string> | null`

      The ids of the messages in the agent's in-context memory.

    - `metadata?: Record<string, unknown> | null`

      The metadata of the agent.

    - `model?: string | null`

      The model handle used by the agent (format: provider/model-name).

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      The model settings used by the agent.

      - `OpenAIModelSettings`

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

      - `GoogleAIModelSettings`

      - `GoogleVertexModelSettings`

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `multi_agent_group?: MultiAgentGroup | null`

      Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

      - `id: string`

        The id of the group. Assigned by the database.

      - `agent_ids: Array<string>`

      - `description: string`

      - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

        - `"round_robin"`

        - `"supervisor"`

        - `"dynamic"`

        - `"sleeptime"`

        - `"voice_sleeptime"`

        - `"swarm"`

      - `base_template_id?: string | null`

        The base template id.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `hidden?: boolean | null`

        If set to True, the group will be hidden.

      - `last_processed_message_id?: string | null`

      - `manager_agent_id?: string | null`

      - `max_message_buffer_length?: number | null`

        The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

      - `max_turns?: number | null`

      - `min_message_buffer_length?: number | null`

        The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

      - `project_id?: string | null`

        The associated project id.

      - `shared_block_ids?: Array<string>`

      - `sleeptime_agent_frequency?: number | null`

      - `template_id?: string | null`

        The id of the template.

      - `termination_token?: string | null`

      - `turns_counter?: number | null`

    - `pending_approval?: ApprovalRequestMessage | null`

      A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      tool_call (ToolCall): The tool call

      - `id: string`

      - `date: string`

      - `tool_call: ToolCall | ToolCallDelta`

        The tool call that has been requested by the llm to run

        - `ToolCall`

          - `arguments: string`

          - `name: string`

          - `tool_call_id: string`

        - `ToolCallDelta`

          - `arguments?: string | null`

          - `name?: string | null`

          - `tool_call_id?: string | null`

      - `is_err?: boolean | null`

      - `message_type?: "approval_request_message"`

        The type of the message.

        - `"approval_request_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

      - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

        The tool calls that have been requested by the llm to run, which are pending approval

        - `Array<ToolCall>`

          - `arguments: string`

          - `name: string`

          - `tool_call_id: string`

        - `ToolCallDelta`

    - `per_file_view_window_char_limit?: number | null`

      The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

    - `project_id?: string | null`

      The id of the project the agent belongs to.

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format used by the agent

      - `TextResponseFormat`

        Response format for plain text responses.

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

    - `secrets?: Array<AgentEnvironmentVariable>`

      The environment variables for tool execution specific to this agent.

      - `agent_id: string`

        The ID of the agent this environment variable belongs to.

      - `key: string`

        The name of the environment variable.

      - `value: string`

        The value of the environment variable.

      - `id?: string`

        The human-friendly ID of the Agent-env

      - `created_at?: string | null`

        The timestamp when the object was created.

      - `created_by_id?: string | null`

        The id of the user that made this object.

      - `description?: string | null`

        An optional description of the environment variable.

      - `last_updated_by_id?: string | null`

        The id of the user that made this object.

      - `updated_at?: string | null`

        The timestamp when the object was last updated.

      - `value_enc?: string | null`

        Encrypted secret value (stored as encrypted string)

    - `template_id?: string | null`

      The id of the template the agent belongs to.

    - `timezone?: string | null`

      The timezone of the agent (IANA format).

    - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

      Deprecated: use `secrets` field instead.

      - `agent_id: string`

        The ID of the agent this environment variable belongs to.

      - `key: string`

        The name of the environment variable.

      - `value: string`

        The value of the environment variable.

      - `id?: string`

        The human-friendly ID of the Agent-env

      - `created_at?: string | null`

        The timestamp when the object was created.

      - `created_by_id?: string | null`

        The id of the user that made this object.

      - `description?: string | null`

        An optional description of the environment variable.

      - `last_updated_by_id?: string | null`

        The id of the user that made this object.

      - `updated_at?: string | null`

        The timestamp when the object was last updated.

      - `value_enc?: string | null`

        Encrypted secret value (stored as encrypted string)

    - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

      The list of tool rules.

      - `ChildToolRule`

        A ToolRule represents a tool that can be invoked by the agent.

        - `children: Array<string>`

          The children tools that can be invoked.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `child_arg_nodes?: Array<ChildArgNode> | null`

          Optional list of typed child argument overrides. Each node must reference a child in 'children'.

          - `name: string`

            The name of the child tool to invoke next.

          - `args?: Record<string, unknown> | null`

            Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

        - `prompt_template?: string | null`

          Optional template string (ignored).

        - `type?: "constrain_child_tools"`

          - `"constrain_child_tools"`

      - `InitToolRule`

        Represents the initial tool rule configuration.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

        - `prompt_template?: string | null`

          Optional template string (ignored). Rendering uses fast built-in formatting for performance.

        - `type?: "run_first"`

          - `"run_first"`

      - `TerminalToolRule`

        Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `prompt_template?: string | null`

          Optional template string (ignored).

        - `type?: "exit_loop"`

          - `"exit_loop"`

      - `ConditionalToolRule`

        A ToolRule that conditionally maps to different child tools based on the output.

        - `child_output_mapping: Record<string, string>`

          The output case to check for mapping

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `default_child?: string | null`

          The default child tool to be called. If None, any tool can be called.

        - `prompt_template?: string | null`

          Optional template string (ignored).

        - `require_output_mapping?: boolean`

          Whether to throw an error when output doesn't match any case

        - `type?: "conditional"`

          - `"conditional"`

      - `ContinueToolRule`

        Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `prompt_template?: string | null`

          Optional template string (ignored).

        - `type?: "continue_loop"`

          - `"continue_loop"`

      - `RequiredBeforeExitToolRule`

        Represents a tool rule configuration where this tool must be called before the agent loop can exit.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `prompt_template?: string | null`

          Optional template string (ignored).

        - `type?: "required_before_exit"`

          - `"required_before_exit"`

      - `MaxCountPerStepToolRule`

        Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

        - `max_count_limit: number`

          The max limit for the total number of times this tool can be invoked in a single step.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `prompt_template?: string | null`

          Optional template string (ignored).

        - `type?: "max_count_per_step"`

          - `"max_count_per_step"`

      - `ParentToolRule`

        A ToolRule that only allows a child tool to be called if the parent has been called.

        - `children: Array<string>`

          The children tools that can be invoked.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `prompt_template?: string | null`

          Optional template string (ignored).

        - `type?: "parent_last_tool"`

          - `"parent_last_tool"`

      - `RequiresApprovalToolRule`

        Represents a tool rule configuration which requires approval before the tool can be invoked.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `prompt_template?: string | null`

          Optional template string (ignored). Rendering uses fast built-in formatting for performance.

        - `type?: "requires_approval"`

          - `"requires_approval"`

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

  - `func_return?: unknown`

    The function return object

  - `sandbox_config_fingerprint?: string | null`

    The fingerprint of the config for the sandbox

  - `stderr?: Array<string> | null`

    Captured stderr from the function invocation

  - `stdout?: Array<string> | null`

    Captured stdout (prints, logs) from function invocation

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const toolExecutionResult = await client.agents.tools.run('tool_name', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
});

console.log(toolExecutionResult.status);
```

#### Response

```json
{
  "status": "success",
  "agent_state": {
    "id": "id",
    "agent_type": "memgpt_agent",
    "blocks": [
      {
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "llm_config": {
      "context_window": 0,
      "model": "model",
      "model_endpoint_type": "openai",
      "compatibility_type": "gguf",
      "display_name": "display_name",
      "effort": "low",
      "enable_reasoner": true,
      "frequency_penalty": 0,
      "handle": "handle",
      "max_reasoning_tokens": 0,
      "max_tokens": 0,
      "model_endpoint": "model_endpoint",
      "model_wrapper": "model_wrapper",
      "parallel_tool_calls": true,
      "provider_category": "base",
      "provider_name": "provider_name",
      "put_inner_thoughts_in_kwargs": true,
      "reasoning_effort": "none",
      "response_format": {
        "type": "text"
      },
      "return_logprobs": true,
      "return_token_ids": true,
      "strict": true,
      "temperature": 0,
      "tier": "tier",
      "tool_call_parser": "tool_call_parser",
      "top_logprobs": 0,
      "verbosity": "low"
    },
    "memory": {
      "blocks": [
        {
          "value": "value",
          "id": "block-123e4567-e89b-12d3-a456-426614174000",
          "base_template_id": "base_template_id",
          "created_by_id": "created_by_id",
          "deployment_id": "deployment_id",
          "description": "description",
          "entity_id": "entity_id",
          "hidden": true,
          "is_template": true,
          "label": "label",
          "last_updated_by_id": "last_updated_by_id",
          "limit": 0,
          "metadata": {
            "foo": "bar"
          },
          "preserve_on_migration": true,
          "project_id": "project_id",
          "read_only": true,
          "tags": [
            "string"
          ],
          "template_id": "template_id",
          "template_name": "template_name"
        }
      ],
      "agent_type": "memgpt_agent",
      "file_blocks": [
        {
          "file_id": "file_id",
          "is_open": true,
          "source_id": "source_id",
          "value": "value",
          "id": "block-123e4567-e89b-12d3-a456-426614174000",
          "base_template_id": "base_template_id",
          "created_by_id": "created_by_id",
          "deployment_id": "deployment_id",
          "description": "description",
          "entity_id": "entity_id",
          "hidden": true,
          "is_template": true,
          "label": "label",
          "last_accessed_at": "2019-12-27T18:11:19.117Z",
          "last_updated_by_id": "last_updated_by_id",
          "limit": 0,
          "metadata": {
            "foo": "bar"
          },
          "preserve_on_migration": true,
          "project_id": "project_id",
          "read_only": true,
          "tags": [
            "string"
          ],
          "template_id": "template_id",
          "template_name": "template_name"
        }
      ],
      "git_enabled": true,
      "prompt_template": "prompt_template"
    },
    "name": "name",
    "sources": [
      {
        "id": "source-123e4567-e89b-12d3-a456-426614174000",
        "embedding_config": {
          "embedding_dim": 0,
          "embedding_endpoint_type": "openai",
          "embedding_model": "embedding_model",
          "azure_deployment": "azure_deployment",
          "azure_endpoint": "azure_endpoint",
          "azure_version": "azure_version",
          "batch_size": 0,
          "embedding_chunk_size": 0,
          "embedding_endpoint": "embedding_endpoint",
          "handle": "handle"
        },
        "name": "name",
        "created_at": "2019-12-27T18:11:19.117Z",
        "created_by_id": "created_by_id",
        "description": "description",
        "instructions": "instructions",
        "last_updated_by_id": "last_updated_by_id",
        "metadata": {
          "foo": "bar"
        },
        "updated_at": "2019-12-27T18:11:19.117Z",
        "vector_db_provider": "native"
      }
    ],
    "system": "system",
    "tags": [
      "string"
    ],
    "tools": [
      {
        "id": "tool-123e4567-e89b-12d3-a456-426614174000",
        "args_json_schema": {
          "foo": "bar"
        },
        "created_by_id": "created_by_id",
        "default_requires_approval": true,
        "description": "description",
        "enable_parallel_execution": true,
        "json_schema": {
          "foo": "bar"
        },
        "last_updated_by_id": "last_updated_by_id",
        "metadata_": {
          "foo": "bar"
        },
        "name": "name",
        "npm_requirements": [
          {
            "name": "x",
            "version": "version"
          }
        ],
        "pip_requirements": [
          {
            "name": "x",
            "version": "version"
          }
        ],
        "project_id": "project_id",
        "return_char_limit": 1,
        "source_code": "source_code",
        "source_type": "source_type",
        "tags": [
          "string"
        ],
        "tool_type": "custom"
      }
    ],
    "base_template_id": "base_template_id",
    "compaction_settings": {
      "clip_chars": 0,
      "mode": "all",
      "model": "model",
      "model_settings": {
        "max_output_tokens": 0,
        "parallel_tool_calls": true,
        "provider_type": "openai",
        "reasoning": {
          "reasoning_effort": "none"
        },
        "response_format": {
          "type": "text"
        },
        "strict": true,
        "temperature": 0
      },
      "prompt": "prompt",
      "prompt_acknowledgement": true,
      "sliding_window_percentage": 0
    },
    "created_at": "2019-12-27T18:11:19.117Z",
    "created_by_id": "created_by_id",
    "deployment_id": "deployment_id",
    "description": "description",
    "embedding": "embedding",
    "embedding_config": {
      "embedding_dim": 0,
      "embedding_endpoint_type": "openai",
      "embedding_model": "embedding_model",
      "azure_deployment": "azure_deployment",
      "azure_endpoint": "azure_endpoint",
      "azure_version": "azure_version",
      "batch_size": 0,
      "embedding_chunk_size": 0,
      "embedding_endpoint": "embedding_endpoint",
      "handle": "handle"
    },
    "enable_sleeptime": true,
    "entity_id": "entity_id",
    "hidden": true,
    "identities": [
      {
        "id": "identity-123e4567-e89b-12d3-a456-426614174000",
        "agent_ids": [
          "string"
        ],
        "block_ids": [
          "string"
        ],
        "identifier_key": "identifier_key",
        "identity_type": "org",
        "name": "name",
        "project_id": "project_id",
        "properties": [
          {
            "key": "key",
            "type": "string",
            "value": "string"
          }
        ]
      }
    ],
    "identity_ids": [
      "string"
    ],
    "last_run_completion": "2019-12-27T18:11:19.117Z",
    "last_run_duration_ms": 0,
    "last_stop_reason": "end_turn",
    "last_updated_by_id": "last_updated_by_id",
    "managed_group": {
      "id": "id",
      "agent_ids": [
        "string"
      ],
      "description": "description",
      "manager_type": "round_robin",
      "base_template_id": "base_template_id",
      "deployment_id": "deployment_id",
      "hidden": true,
      "last_processed_message_id": "last_processed_message_id",
      "manager_agent_id": "manager_agent_id",
      "max_message_buffer_length": 0,
      "max_turns": 0,
      "min_message_buffer_length": 0,
      "project_id": "project_id",
      "shared_block_ids": [
        "string"
      ],
      "sleeptime_agent_frequency": 0,
      "template_id": "template_id",
      "termination_token": "termination_token",
      "turns_counter": 0
    },
    "max_files_open": 0,
    "message_buffer_autoclear": true,
    "message_ids": [
      "string"
    ],
    "metadata": {
      "foo": "bar"
    },
    "model": "model",
    "model_settings": {
      "max_output_tokens": 0,
      "parallel_tool_calls": true,
      "provider_type": "openai",
      "reasoning": {
        "reasoning_effort": "none"
      },
      "response_format": {
        "type": "text"
      },
      "strict": true,
      "temperature": 0
    },
    "multi_agent_group": {
      "id": "id",
      "agent_ids": [
        "string"
      ],
      "description": "description",
      "manager_type": "round_robin",
      "base_template_id": "base_template_id",
      "deployment_id": "deployment_id",
      "hidden": true,
      "last_processed_message_id": "last_processed_message_id",
      "manager_agent_id": "manager_agent_id",
      "max_message_buffer_length": 0,
      "max_turns": 0,
      "min_message_buffer_length": 0,
      "project_id": "project_id",
      "shared_block_ids": [
        "string"
      ],
      "sleeptime_agent_frequency": 0,
      "template_id": "template_id",
      "termination_token": "termination_token",
      "turns_counter": 0
    },
    "pending_approval": {
      "id": "id",
      "date": "2019-12-27T18:11:19.117Z",
      "tool_call": {
        "arguments": "arguments",
        "name": "name",
        "tool_call_id": "tool_call_id"
      },
      "is_err": true,
      "message_type": "approval_request_message",
      "name": "name",
      "otid": "otid",
      "run_id": "run_id",
      "sender_id": "sender_id",
      "seq_id": 0,
      "step_id": "step_id",
      "tool_calls": [
        {
          "arguments": "arguments",
          "name": "name",
          "tool_call_id": "tool_call_id"
        }
      ]
    },
    "per_file_view_window_char_limit": 0,
    "project_id": "project_id",
    "response_format": {
      "type": "text"
    },
    "secrets": [
      {
        "agent_id": "agent_id",
        "key": "key",
        "value": "value",
        "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
        "created_at": "2019-12-27T18:11:19.117Z",
        "created_by_id": "created_by_id",
        "description": "description",
        "last_updated_by_id": "last_updated_by_id",
        "updated_at": "2019-12-27T18:11:19.117Z",
        "value_enc": "value_enc"
      }
    ],
    "template_id": "template_id",
    "timezone": "timezone",
    "tool_exec_environment_variables": [
      {
        "agent_id": "agent_id",
        "key": "key",
        "value": "value",
        "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
        "created_at": "2019-12-27T18:11:19.117Z",
        "created_by_id": "created_by_id",
        "description": "description",
        "last_updated_by_id": "last_updated_by_id",
        "updated_at": "2019-12-27T18:11:19.117Z",
        "value_enc": "value_enc"
      }
    ],
    "tool_rules": [
      {
        "children": [
          "string"
        ],
        "tool_name": "tool_name",
        "child_arg_nodes": [
          {
            "name": "name",
            "args": {
              "foo": "bar"
            }
          }
        ],
        "prompt_template": "prompt_template",
        "type": "constrain_child_tools"
      }
    ],
    "updated_at": "2019-12-27T18:11:19.117Z"
  },
  "func_return": {},
  "sandbox_config_fingerprint": "sandbox_config_fingerprint",
  "stderr": [
    "string"
  ],
  "stdout": [
    "string"
  ]
}
```

## Domain Types

### Tool Execute Request

- `ToolExecuteRequest`

  Request to execute a tool.

  - `args?: Record<string, unknown>`

    Arguments to pass to the tool

### Tool Execution Result

- `ToolExecutionResult`

  - `status: "success" | "error"`

    The status of the tool execution and return object

    - `"success"`

    - `"error"`

  - `agent_state?: AgentState | null`

    Representation of an agent's state. This is the state of the agent at a given time, and is persisted in the DB backend. The state has all the information needed to recreate a persisted agent.

    - `id: string`

      The id of the agent. Assigned by the database.

    - `agent_type: AgentType`

      The type of agent.

      - `"memgpt_agent"`

      - `"memgpt_v2_agent"`

      - `"letta_v1_agent"`

      - `"react_agent"`

      - `"workflow_agent"`

      - `"split_thread_agent"`

      - `"sleeptime_agent"`

      - `"voice_convo_agent"`

      - `"voice_sleeptime_agent"`

    - `blocks: Array<Block>`

      The memory blocks used by the agent.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `llm_config: LlmConfig`

      Deprecated: Use `model` field instead. The LLM configuration used by the agent.

      - `context_window: number`

        The context window size for the model.

      - `model: string`

        LLM model name.

      - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"lmstudio-chatcompletions"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"minimax"`

        - `"moonshot"`

        - `"moonshot_coding"`

        - `"mistral"`

        - `"together"`

        - `"bedrock"`

        - `"deepseek"`

        - `"xai"`

        - `"zai"`

        - `"zai_coding"`

        - `"baseten"`

        - `"fireworks"`

        - `"openrouter"`

        - `"chatgpt_oauth"`

      - `compatibility_type?: "gguf" | "mlx" | null`

        The framework compatibility type for the model.

        - `"gguf"`

        - `"mlx"`

      - `display_name?: string | null`

        A human-friendly display name for the model.

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `enable_reasoner?: boolean`

        Whether or not the model should use extended thinking if it is a 'reasoning' style model

      - `frequency_penalty?: number | null`

        Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

      - `max_reasoning_tokens?: number`

        Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

      - `max_tokens?: number | null`

        The maximum number of tokens to generate. If not set, the model will use its default value.

      - `model_endpoint?: string | null`

        The endpoint for the model.

      - `model_wrapper?: string | null`

        The wrapper for the model.

      - `parallel_tool_calls?: boolean | null`

        Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

      - `provider_category?: ProviderCategory | null`

        The provider category for the model.

        - `"base"`

        - `"byok"`

      - `provider_name?: string | null`

        The provider name for the model.

      - `put_inner_thoughts_in_kwargs?: boolean | null`

        Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

      - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

        The reasoning effort to use when generating text reasoning models

        - `"none"`

        - `"minimal"`

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

        - `TextResponseFormat`

          Response format for plain text responses.

          - `type?: "text"`

            The type of the response format.

            - `"text"`

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

          - `json_schema: Record<string, unknown>`

            The JSON schema of the response.

          - `type?: "json_schema"`

            The type of the response format.

            - `"json_schema"`

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

          - `type?: "json_object"`

            The type of the response format.

            - `"json_object"`

      - `return_logprobs?: boolean`

        Whether to return log probabilities of the output tokens. Useful for RL training.

      - `return_token_ids?: boolean`

        Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

      - `temperature?: number`

        The temperature to use when generating text with the model. A higher temperature will result in more random text.

      - `tier?: string | null`

        The cost tier for the model (cloud only).

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

      - `top_logprobs?: number | null`

        Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `memory: Memory`

      Deprecated: Use `blocks` field instead. The in-context memory of the agent.

      - `blocks: Array<Block>`

        Memory blocks contained in the agent's in-context memory

        - `value: string`

          Value of the block.

        - `id?: string`

          The human-friendly ID of the Block

        - `base_template_id?: string | null`

          The base template id of the block.

        - `created_by_id?: string | null`

          The id of the user that made this Block.

        - `deployment_id?: string | null`

          The id of the deployment.

        - `description?: string | null`

          Description of the block.

        - `entity_id?: string | null`

          The id of the entity within the template.

        - `hidden?: boolean | null`

          If set to True, the block will be hidden.

        - `is_template?: boolean`

          Whether the block is a template (e.g. saved human/persona options).

        - `label?: string | null`

          Label of the block (e.g. 'human', 'persona') in the context window.

        - `last_updated_by_id?: string | null`

          The id of the user that last updated this Block.

        - `limit?: number`

          Character limit of the block.

        - `metadata?: Record<string, unknown> | null`

          Metadata of the block.

        - `preserve_on_migration?: boolean | null`

          Preserve the block on template migration.

        - `project_id?: string | null`

          The associated project id.

        - `read_only?: boolean`

          Whether the agent has read-only access to the block.

        - `tags?: Array<string> | null`

          The tags associated with the block.

        - `template_id?: string | null`

          The id of the template.

        - `template_name?: string | null`

          Name of the block if it is a template.

      - `agent_type?: AgentType | (string & {}) | null`

        Agent type controlling prompt rendering.

        - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

          Enum to represent the type of agent.

        - `(string & {})`

      - `file_blocks?: Array<FileBlock>`

        Special blocks representing the agent's in-context memory of an attached file

        - `file_id: string`

          Unique identifier of the file.

        - `is_open: boolean`

          True if the agent currently has the file open.

        - `source_id: string`

          Deprecated: Use `folder_id` field instead. Unique identifier of the source.

        - `value: string`

          Value of the block.

        - `id?: string`

          The human-friendly ID of the Block

        - `base_template_id?: string | null`

          The base template id of the block.

        - `created_by_id?: string | null`

          The id of the user that made this Block.

        - `deployment_id?: string | null`

          The id of the deployment.

        - `description?: string | null`

          Description of the block.

        - `entity_id?: string | null`

          The id of the entity within the template.

        - `hidden?: boolean | null`

          If set to True, the block will be hidden.

        - `is_template?: boolean`

          Whether the block is a template (e.g. saved human/persona options).

        - `label?: string | null`

          Label of the block (e.g. 'human', 'persona') in the context window.

        - `last_accessed_at?: string | null`

          UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

        - `last_updated_by_id?: string | null`

          The id of the user that last updated this Block.

        - `limit?: number`

          Character limit of the block.

        - `metadata?: Record<string, unknown> | null`

          Metadata of the block.

        - `preserve_on_migration?: boolean | null`

          Preserve the block on template migration.

        - `project_id?: string | null`

          The associated project id.

        - `read_only?: boolean`

          Whether the agent has read-only access to the block.

        - `tags?: Array<string> | null`

          The tags associated with the block.

        - `template_id?: string | null`

          The id of the template.

        - `template_name?: string | null`

          Name of the block if it is a template.

      - `git_enabled?: boolean`

        Whether this agent uses git-backed memory with structured labels.

      - `prompt_template?: string`

        Deprecated. Ignored for performance.

    - `name: string`

      The name of the agent.

    - `sources: Array<Source>`

      Deprecated: Use `folders` field instead. The sources used by the agent.

      - `id: string`

        The human-friendly ID of the Source

      - `embedding_config: EmbeddingConfig`

        The embedding configuration used by the source.

        - `embedding_dim: number`

          The dimension of the embedding.

        - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

          The endpoint type for the model.

          - `"openai"`

          - `"anthropic"`

          - `"bedrock"`

          - `"google_ai"`

          - `"google_vertex"`

          - `"azure"`

          - `"groq"`

          - `"ollama"`

          - `"webui"`

          - `"webui-legacy"`

          - `"lmstudio"`

          - `"lmstudio-legacy"`

          - `"llamacpp"`

          - `"koboldcpp"`

          - `"vllm"`

          - `"hugging-face"`

          - `"mistral"`

          - `"together"`

          - `"pinecone"`

        - `embedding_model: string`

          The model for the embedding.

        - `azure_deployment?: string | null`

          The Azure deployment for the model.

        - `azure_endpoint?: string | null`

          The Azure endpoint for the model.

        - `azure_version?: string | null`

          The Azure version for the model.

        - `batch_size?: number`

          The maximum batch size for processing embeddings.

        - `embedding_chunk_size?: number | null`

          The chunk size of the embedding.

        - `embedding_endpoint?: string | null`

          The endpoint for the model (`None` if local).

        - `handle?: string | null`

          The handle for this config, in the format provider/model-name.

      - `name: string`

        The name of the source.

      - `created_at?: string | null`

        The timestamp when the source was created.

      - `created_by_id?: string | null`

        The id of the user that made this Tool.

      - `description?: string | null`

        The description of the source.

      - `instructions?: string | null`

        Instructions for how to use the source.

      - `last_updated_by_id?: string | null`

        The id of the user that made this Tool.

      - `metadata?: Record<string, unknown> | null`

        Metadata associated with the source.

      - `updated_at?: string | null`

        The timestamp when the source was last updated.

      - `vector_db_provider?: VectorDBProvider`

        The vector database provider used for this source's passages

        - `"native"`

        - `"tpuf"`

        - `"pinecone"`

    - `system: string`

      The system prompt used by the agent.

    - `tags: Array<string>`

      The tags associated with the agent.

    - `tools: Array<Tool>`

      The tools used by the agent.

      - `id: string`

        The human-friendly ID of the Tool

      - `args_json_schema?: Record<string, unknown> | null`

        The args JSON schema of the function.

      - `created_by_id?: string | null`

        The id of the user that made this Tool.

      - `default_requires_approval?: boolean | null`

        Default value for whether or not executing this tool requires approval.

      - `description?: string | null`

        The description of the tool.

      - `enable_parallel_execution?: boolean | null`

        If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

      - `json_schema?: Record<string, unknown> | null`

        The JSON schema of the function.

      - `last_updated_by_id?: string | null`

        The id of the user that made this Tool.

      - `metadata_?: Record<string, unknown> | null`

        A dictionary of additional metadata for the tool.

      - `name?: string | null`

        The name of the function.

      - `npm_requirements?: Array<NpmRequirement> | null`

        Optional list of npm packages required by this tool.

        - `name: string`

          Name of the npm package.

        - `version?: string | null`

          Optional version of the package, following semantic versioning.

      - `pip_requirements?: Array<PipRequirement> | null`

        Optional list of pip packages required by this tool.

        - `name: string`

          Name of the pip package.

        - `version?: string | null`

          Optional version of the package, following semantic versioning.

      - `project_id?: string | null`

        The project id of the tool.

      - `return_char_limit?: number`

        The maximum number of characters in the response.

      - `source_code?: string | null`

        The source code of the function.

      - `source_type?: string | null`

        The type of the source code.

      - `tags?: Array<string>`

        Metadata tags.

      - `tool_type?: ToolType`

        The type of the tool.

        - `"custom"`

        - `"letta_core"`

        - `"letta_memory_core"`

        - `"letta_multi_agent_core"`

        - `"letta_sleeptime_core"`

        - `"letta_voice_sleeptime_core"`

        - `"letta_builtin"`

        - `"letta_files_core"`

        - `"external_langchain"`

        - `"external_composio"`

        - `"external_mcp"`

    - `base_template_id?: string | null`

      The base template id of the agent.

    - `compaction_settings?: CompactionSettings | null`

      Configuration for conversation compaction / summarization.

      Per-model settings (temperature,
      max tokens, etc.) are derived from the default configuration for that handle.

      - `clip_chars?: number | null`

        The maximum length of the summary in characters. If none, no clipping is performed.

      - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

        The type of summarization technique use.

        - `"all"`

        - `"sliding_window"`

        - `"self_compact_all"`

        - `"self_compact_sliding_window"`

      - `model?: string | null`

        Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

      - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

        Optional model settings used to override defaults for the summarizer model.

        - `OpenAIModelSettings`

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "openai"`

            The type of the provider.

            - `"openai"`

          - `reasoning?: Reasoning`

            The reasoning configuration for the model.

            - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

              The reasoning effort to use when generating text reasoning models

              - `"none"`

              - `"minimal"`

              - `"low"`

              - `"medium"`

              - `"high"`

              - `"xhigh"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `strict?: boolean`

            Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

          - `temperature?: number`

            The temperature of the model.

        - `SgLangModelSettings`

          SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "sglang"`

            The type of the provider.

            - `"sglang"`

          - `reasoning?: Reasoning`

            The reasoning configuration for the model.

            - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

              The reasoning effort to use when generating text reasoning models

              - `"none"`

              - `"minimal"`

              - `"low"`

              - `"medium"`

              - `"high"`

              - `"xhigh"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `strict?: boolean`

            Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

          - `temperature?: number`

            The temperature of the model.

          - `tool_call_parser?: string | null`

            SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

        - `AnthropicModelSettings`

          - `effort?: "low" | "medium" | "high" | 2 more | null`

            Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

            - `"max"`

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "anthropic"`

            The type of the provider.

            - `"anthropic"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `strict?: boolean`

            Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

          - `temperature?: number`

            The temperature of the model.

          - `thinking?: Thinking`

            The thinking configuration for the model.

            - `budget_tokens?: number`

              The maximum number of tokens the model can use for extended thinking.

            - `type?: "enabled" | "disabled"`

              The type of thinking to use.

              - `"enabled"`

              - `"disabled"`

          - `verbosity?: "low" | "medium" | "high" | null`

            Soft control for how verbose model output should be, used for GPT-5 models.

            - `"low"`

            - `"medium"`

            - `"high"`

        - `GoogleAIModelSettings`

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "google_ai"`

            The type of the provider.

            - `"google_ai"`

          - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response schema for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

          - `thinking_config?: ThinkingConfig`

            The thinking configuration for the model.

            - `include_thoughts?: boolean`

              Whether to include thoughts in the model's response.

            - `thinking_budget?: number`

              The thinking budget for the model.

        - `GoogleVertexModelSettings`

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "google_vertex"`

            The type of the provider.

            - `"google_vertex"`

          - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response schema for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

          - `thinking_config?: ThinkingConfig`

            The thinking configuration for the model.

            - `include_thoughts?: boolean`

              Whether to include thoughts in the model's response.

            - `thinking_budget?: number`

              The thinking budget for the model.

        - `AzureModelSettings`

          Azure OpenAI model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "azure"`

            The type of the provider.

            - `"azure"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

        - `XaiModelSettings`

          xAI model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "xai"`

            The type of the provider.

            - `"xai"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

        - `MoonshotModelSettings`

          Moonshot/Kimi model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "moonshot"`

            The type of the provider.

            - `"moonshot"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `strict?: boolean`

            Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

          - `temperature?: number`

            The temperature of the model.

        - `ZaiModelSettings`

          Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "zai"`

            The type of the provider.

            - `"zai"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

          - `thinking?: Thinking`

            The thinking configuration for GLM-4.5+ models.

            - `clear_thinking?: boolean`

              If False, preserved thinking is used (recommended for agents).

            - `type?: "enabled" | "disabled"`

              Whether thinking is enabled or disabled.

              - `"enabled"`

              - `"disabled"`

        - `MoonshotCodingModelSettings`

          Kimi Code model configuration (Anthropic-compatible).

          - `effort?: "low" | "medium" | "high" | 2 more | null`

            Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

            - `"max"`

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "moonshot_coding"`

            The type of the provider.

            - `"moonshot_coding"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `strict?: boolean`

            Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

          - `temperature?: number`

            The temperature of the model.

          - `thinking?: Thinking`

            The thinking configuration for the model.

            - `budget_tokens?: number`

              The maximum number of tokens the model can use for extended thinking.

            - `type?: "enabled" | "disabled"`

              The type of thinking to use.

              - `"enabled"`

              - `"disabled"`

          - `verbosity?: "low" | "medium" | "high" | null`

            Soft control for how verbose model output should be, used for GPT-5 models.

            - `"low"`

            - `"medium"`

            - `"high"`

        - `GroqModelSettings`

          Groq model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "groq"`

            The type of the provider.

            - `"groq"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

        - `DeepseekModelSettings`

          Deepseek model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "deepseek"`

            The type of the provider.

            - `"deepseek"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

        - `TogetherModelSettings`

          Together AI model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "together"`

            The type of the provider.

            - `"together"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

        - `BedrockModelSettings`

          AWS Bedrock model configuration.

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "bedrock"`

            The type of the provider.

            - `"bedrock"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

        - `BasetenModelSettings`

          Baseten model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "baseten"`

            The type of the provider.

            - `"baseten"`

          - `temperature?: number`

            The temperature of the model.

        - `OpenRouterModelSettings`

          OpenRouter model configuration (OpenAI-compatible).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "openrouter"`

            The type of the provider.

            - `"openrouter"`

          - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

            The response format for the model.

            - `TextResponseFormat`

              Response format for plain text responses.

            - `JsonSchemaResponseFormat`

              Response format for JSON schema-based responses.

            - `JsonObjectResponseFormat`

              Response format for JSON object responses.

          - `temperature?: number`

            The temperature of the model.

        - `ChatGptoAuthModelSettings`

          ChatGPT OAuth model configuration (uses ChatGPT backend API).

          - `max_output_tokens?: number`

            The maximum number of tokens the model can generate.

          - `parallel_tool_calls?: boolean`

            Whether to enable parallel tool calling.

          - `provider_type?: "chatgpt_oauth"`

            The type of the provider.

            - `"chatgpt_oauth"`

          - `reasoning?: Reasoning`

            The reasoning configuration for the model.

            - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

              The reasoning effort level for GPT-5.x and o-series models.

              - `"none"`

              - `"low"`

              - `"medium"`

              - `"high"`

              - `"xhigh"`

          - `temperature?: number`

            The temperature of the model.

      - `prompt?: string | null`

        The prompt to use for summarization. If None, uses mode-specific default.

      - `prompt_acknowledgement?: boolean`

        Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

      - `sliding_window_percentage?: number`

        The percentage of the context window to keep post-summarization (only used in sliding window modes).

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      The description of the agent.

    - `embedding?: string | null`

      The embedding model handle used by the agent (format: provider/model-name).

    - `embedding_config?: EmbeddingConfig | null`

      Configuration for embedding model connection and processing parameters.

    - `enable_sleeptime?: boolean | null`

      If set to True, memory management will move to a background agent thread.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the agent will be hidden.

    - `identities?: Array<Identity>`

      The identities associated with this agent.

      - `id: string`

        The human-friendly ID of the Identity

      - `agent_ids: Array<string>`

        The IDs of the agents associated with the identity.

      - `block_ids: Array<string>`

        The IDs of the blocks associated with the identity.

      - `identifier_key: string`

        External, user-generated identifier key of the identity.

      - `identity_type: "org" | "user" | "other"`

        The type of the identity.

        - `"org"`

        - `"user"`

        - `"other"`

      - `name: string`

        The name of the identity.

      - `project_id?: string | null`

        The project id of the identity, if applicable.

      - `properties?: Array<Property>`

        List of properties associated with the identity

        - `key: string`

          The key of the property

        - `type: "string" | "number" | "boolean" | "json"`

          The type of the property

          - `"string"`

          - `"number"`

          - `"boolean"`

          - `"json"`

        - `value: string | number | boolean | Record<string, unknown>`

          The value of the property

          - `string`

          - `number`

          - `boolean`

          - `Record<string, unknown>`

    - `identity_ids?: Array<string>`

      Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

    - `last_run_completion?: string | null`

      The timestamp when the agent last completed a run.

    - `last_run_duration_ms?: number | null`

      The duration in milliseconds of the agent's last run.

    - `last_stop_reason?: StopReasonType | null`

      The stop reason from the agent's last run.

      - `"end_turn"`

      - `"error"`

      - `"llm_api_error"`

      - `"invalid_llm_response"`

      - `"invalid_tool_call"`

      - `"max_steps"`

      - `"max_tokens_exceeded"`

      - `"no_tool_call"`

      - `"tool_rule"`

      - `"cancelled"`

      - `"insufficient_credits"`

      - `"requires_approval"`

      - `"context_window_overflow_in_system_prompt"`

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `managed_group?: ManagedGroup | null`

      The multi-agent group that this agent manages

      - `id: string`

        The id of the group. Assigned by the database.

      - `agent_ids: Array<string>`

      - `description: string`

      - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

        - `"round_robin"`

        - `"supervisor"`

        - `"dynamic"`

        - `"sleeptime"`

        - `"voice_sleeptime"`

        - `"swarm"`

      - `base_template_id?: string | null`

        The base template id.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `hidden?: boolean | null`

        If set to True, the group will be hidden.

      - `last_processed_message_id?: string | null`

      - `manager_agent_id?: string | null`

      - `max_message_buffer_length?: number | null`

        The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

      - `max_turns?: number | null`

      - `min_message_buffer_length?: number | null`

        The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

      - `project_id?: string | null`

        The associated project id.

      - `shared_block_ids?: Array<string>`

      - `sleeptime_agent_frequency?: number | null`

      - `template_id?: string | null`

        The id of the template.

      - `termination_token?: string | null`

      - `turns_counter?: number | null`

    - `max_files_open?: number | null`

      Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

    - `message_buffer_autoclear?: boolean`

      If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

    - `message_ids?: Array<string> | null`

      The ids of the messages in the agent's in-context memory.

    - `metadata?: Record<string, unknown> | null`

      The metadata of the agent.

    - `model?: string | null`

      The model handle used by the agent (format: provider/model-name).

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      The model settings used by the agent.

      - `OpenAIModelSettings`

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

      - `GoogleAIModelSettings`

      - `GoogleVertexModelSettings`

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `multi_agent_group?: MultiAgentGroup | null`

      Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

      - `id: string`

        The id of the group. Assigned by the database.

      - `agent_ids: Array<string>`

      - `description: string`

      - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

        - `"round_robin"`

        - `"supervisor"`

        - `"dynamic"`

        - `"sleeptime"`

        - `"voice_sleeptime"`

        - `"swarm"`

      - `base_template_id?: string | null`

        The base template id.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `hidden?: boolean | null`

        If set to True, the group will be hidden.

      - `last_processed_message_id?: string | null`

      - `manager_agent_id?: string | null`

      - `max_message_buffer_length?: number | null`

        The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

      - `max_turns?: number | null`

      - `min_message_buffer_length?: number | null`

        The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

      - `project_id?: string | null`

        The associated project id.

      - `shared_block_ids?: Array<string>`

      - `sleeptime_agent_frequency?: number | null`

      - `template_id?: string | null`

        The id of the template.

      - `termination_token?: string | null`

      - `turns_counter?: number | null`

    - `pending_approval?: ApprovalRequestMessage | null`

      A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

      Args:
      id (str): The ID of the message
      date (datetime): The date the message was created in ISO format
      name (Optional[str]): The name of the sender of the message
      tool_call (ToolCall): The tool call

      - `id: string`

      - `date: string`

      - `tool_call: ToolCall | ToolCallDelta`

        The tool call that has been requested by the llm to run

        - `ToolCall`

          - `arguments: string`

          - `name: string`

          - `tool_call_id: string`

        - `ToolCallDelta`

          - `arguments?: string | null`

          - `name?: string | null`

          - `tool_call_id?: string | null`

      - `is_err?: boolean | null`

      - `message_type?: "approval_request_message"`

        The type of the message.

        - `"approval_request_message"`

      - `name?: string | null`

      - `otid?: string | null`

        The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

      - `run_id?: string | null`

      - `sender_id?: string | null`

      - `seq_id?: number | null`

      - `step_id?: string | null`

      - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

        The tool calls that have been requested by the llm to run, which are pending approval

        - `Array<ToolCall>`

          - `arguments: string`

          - `name: string`

          - `tool_call_id: string`

        - `ToolCallDelta`

    - `per_file_view_window_char_limit?: number | null`

      The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

    - `project_id?: string | null`

      The id of the project the agent belongs to.

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format used by the agent

      - `TextResponseFormat`

        Response format for plain text responses.

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

    - `secrets?: Array<AgentEnvironmentVariable>`

      The environment variables for tool execution specific to this agent.

      - `agent_id: string`

        The ID of the agent this environment variable belongs to.

      - `key: string`

        The name of the environment variable.

      - `value: string`

        The value of the environment variable.

      - `id?: string`

        The human-friendly ID of the Agent-env

      - `created_at?: string | null`

        The timestamp when the object was created.

      - `created_by_id?: string | null`

        The id of the user that made this object.

      - `description?: string | null`

        An optional description of the environment variable.

      - `last_updated_by_id?: string | null`

        The id of the user that made this object.

      - `updated_at?: string | null`

        The timestamp when the object was last updated.

      - `value_enc?: string | null`

        Encrypted secret value (stored as encrypted string)

    - `template_id?: string | null`

      The id of the template the agent belongs to.

    - `timezone?: string | null`

      The timezone of the agent (IANA format).

    - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

      Deprecated: use `secrets` field instead.

      - `agent_id: string`

        The ID of the agent this environment variable belongs to.

      - `key: string`

        The name of the environment variable.

      - `value: string`

        The value of the environment variable.

      - `id?: string`

        The human-friendly ID of the Agent-env

      - `created_at?: string | null`

        The timestamp when the object was created.

      - `created_by_id?: string | null`

        The id of the user that made this object.

      - `description?: string | null`

        An optional description of the environment variable.

      - `last_updated_by_id?: string | null`

        The id of the user that made this object.

      - `updated_at?: string | null`

        The timestamp when the object was last updated.

      - `value_enc?: string | null`

        Encrypted secret value (stored as encrypted string)

    - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

      The list of tool rules.

      - `ChildToolRule`

        A ToolRule represents a tool that can be invoked by the agent.

        - `children: Array<string>`

          The children tools that can be invoked.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `child_arg_nodes?: Array<ChildArgNode> | null`

          Optional list of typed child argument overrides. Each node must reference a child in 'children'.

          - `name: string`

            The name of the child tool to invoke next.

          - `args?: Record<string, unknown> | null`

            Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

        - `prompt_template?: string | null`

          Optional template string (ignored).

        - `type?: "constrain_child_tools"`

          - `"constrain_child_tools"`

      - `InitToolRule`

        Represents the initial tool rule configuration.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

        - `prompt_template?: string | null`

          Optional template string (ignored). Rendering uses fast built-in formatting for performance.

        - `type?: "run_first"`

          - `"run_first"`

      - `TerminalToolRule`

        Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `prompt_template?: string | null`

          Optional template string (ignored).

        - `type?: "exit_loop"`

          - `"exit_loop"`

      - `ConditionalToolRule`

        A ToolRule that conditionally maps to different child tools based on the output.

        - `child_output_mapping: Record<string, string>`

          The output case to check for mapping

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `default_child?: string | null`

          The default child tool to be called. If None, any tool can be called.

        - `prompt_template?: string | null`

          Optional template string (ignored).

        - `require_output_mapping?: boolean`

          Whether to throw an error when output doesn't match any case

        - `type?: "conditional"`

          - `"conditional"`

      - `ContinueToolRule`

        Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `prompt_template?: string | null`

          Optional template string (ignored).

        - `type?: "continue_loop"`

          - `"continue_loop"`

      - `RequiredBeforeExitToolRule`

        Represents a tool rule configuration where this tool must be called before the agent loop can exit.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `prompt_template?: string | null`

          Optional template string (ignored).

        - `type?: "required_before_exit"`

          - `"required_before_exit"`

      - `MaxCountPerStepToolRule`

        Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

        - `max_count_limit: number`

          The max limit for the total number of times this tool can be invoked in a single step.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `prompt_template?: string | null`

          Optional template string (ignored).

        - `type?: "max_count_per_step"`

          - `"max_count_per_step"`

      - `ParentToolRule`

        A ToolRule that only allows a child tool to be called if the parent has been called.

        - `children: Array<string>`

          The children tools that can be invoked.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `prompt_template?: string | null`

          Optional template string (ignored).

        - `type?: "parent_last_tool"`

          - `"parent_last_tool"`

      - `RequiresApprovalToolRule`

        Represents a tool rule configuration which requires approval before the tool can be invoked.

        - `tool_name: string`

          The name of the tool. Must exist in the database for the user's organization.

        - `prompt_template?: string | null`

          Optional template string (ignored). Rendering uses fast built-in formatting for performance.

        - `type?: "requires_approval"`

          - `"requires_approval"`

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

  - `func_return?: unknown`

    The function return object

  - `sandbox_config_fingerprint?: string | null`

    The fingerprint of the config for the sandbox

  - `stderr?: Array<string> | null`

    Captured stderr from the function invocation

  - `stdout?: Array<string> | null`

    Captured stdout (prints, logs) from function invocation

# Folders

## Attach Folder To Agent

`client.agents.folders.attach(stringfolderID, FolderAttachParamsparams, RequestOptionsoptions?): AgentState | null`

**patch** `/v1/agents/{agent_id}/folders/attach/{folder_id}`

Attach a folder to an agent.

### Parameters

- `folderID: string`

  The ID of the source in the format 'source-<uuid4>'

- `params: FolderAttachParams`

  - `agent_id: string`

    The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `AgentState | null`

  - `id: string`

    The id of the agent. Assigned by the database.

  - `agent_type: AgentType`

    The type of agent.

    - `"memgpt_agent"`

    - `"memgpt_v2_agent"`

    - `"letta_v1_agent"`

    - `"react_agent"`

    - `"workflow_agent"`

    - `"split_thread_agent"`

    - `"sleeptime_agent"`

    - `"voice_convo_agent"`

    - `"voice_sleeptime_agent"`

  - `blocks: Array<Block>`

    The memory blocks used by the agent.

    - `value: string`

      Value of the block.

    - `id?: string`

      The human-friendly ID of the Block

    - `base_template_id?: string | null`

      The base template id of the block.

    - `created_by_id?: string | null`

      The id of the user that made this Block.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      Description of the block.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the block will be hidden.

    - `is_template?: boolean`

      Whether the block is a template (e.g. saved human/persona options).

    - `label?: string | null`

      Label of the block (e.g. 'human', 'persona') in the context window.

    - `last_updated_by_id?: string | null`

      The id of the user that last updated this Block.

    - `limit?: number`

      Character limit of the block.

    - `metadata?: Record<string, unknown> | null`

      Metadata of the block.

    - `preserve_on_migration?: boolean | null`

      Preserve the block on template migration.

    - `project_id?: string | null`

      The associated project id.

    - `read_only?: boolean`

      Whether the agent has read-only access to the block.

    - `tags?: Array<string> | null`

      The tags associated with the block.

    - `template_id?: string | null`

      The id of the template.

    - `template_name?: string | null`

      Name of the block if it is a template.

  - `llm_config: LlmConfig`

    Deprecated: Use `model` field instead. The LLM configuration used by the agent.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

        - `type?: "text"`

          The type of the response format.

          - `"text"`

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

        - `json_schema: Record<string, unknown>`

          The JSON schema of the response.

        - `type?: "json_schema"`

          The type of the response format.

          - `"json_schema"`

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

        - `type?: "json_object"`

          The type of the response format.

          - `"json_object"`

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `memory: Memory`

    Deprecated: Use `blocks` field instead. The in-context memory of the agent.

    - `blocks: Array<Block>`

      Memory blocks contained in the agent's in-context memory

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `agent_type?: AgentType | (string & {}) | null`

      Agent type controlling prompt rendering.

      - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

        Enum to represent the type of agent.

      - `(string & {})`

    - `file_blocks?: Array<FileBlock>`

      Special blocks representing the agent's in-context memory of an attached file

      - `file_id: string`

        Unique identifier of the file.

      - `is_open: boolean`

        True if the agent currently has the file open.

      - `source_id: string`

        Deprecated: Use `folder_id` field instead. Unique identifier of the source.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_accessed_at?: string | null`

        UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `git_enabled?: boolean`

      Whether this agent uses git-backed memory with structured labels.

    - `prompt_template?: string`

      Deprecated. Ignored for performance.

  - `name: string`

    The name of the agent.

  - `sources: Array<Source>`

    Deprecated: Use `folders` field instead. The sources used by the agent.

    - `id: string`

      The human-friendly ID of the Source

    - `embedding_config: EmbeddingConfig`

      The embedding configuration used by the source.

      - `embedding_dim: number`

        The dimension of the embedding.

      - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"bedrock"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"mistral"`

        - `"together"`

        - `"pinecone"`

      - `embedding_model: string`

        The model for the embedding.

      - `azure_deployment?: string | null`

        The Azure deployment for the model.

      - `azure_endpoint?: string | null`

        The Azure endpoint for the model.

      - `azure_version?: string | null`

        The Azure version for the model.

      - `batch_size?: number`

        The maximum batch size for processing embeddings.

      - `embedding_chunk_size?: number | null`

        The chunk size of the embedding.

      - `embedding_endpoint?: string | null`

        The endpoint for the model (`None` if local).

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

    - `name: string`

      The name of the source.

    - `created_at?: string | null`

      The timestamp when the source was created.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `description?: string | null`

      The description of the source.

    - `instructions?: string | null`

      Instructions for how to use the source.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata?: Record<string, unknown> | null`

      Metadata associated with the source.

    - `updated_at?: string | null`

      The timestamp when the source was last updated.

    - `vector_db_provider?: VectorDBProvider`

      The vector database provider used for this source's passages

      - `"native"`

      - `"tpuf"`

      - `"pinecone"`

  - `system: string`

    The system prompt used by the agent.

  - `tags: Array<string>`

    The tags associated with the agent.

  - `tools: Array<Tool>`

    The tools used by the agent.

    - `id: string`

      The human-friendly ID of the Tool

    - `args_json_schema?: Record<string, unknown> | null`

      The args JSON schema of the function.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `default_requires_approval?: boolean | null`

      Default value for whether or not executing this tool requires approval.

    - `description?: string | null`

      The description of the tool.

    - `enable_parallel_execution?: boolean | null`

      If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

    - `json_schema?: Record<string, unknown> | null`

      The JSON schema of the function.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata_?: Record<string, unknown> | null`

      A dictionary of additional metadata for the tool.

    - `name?: string | null`

      The name of the function.

    - `npm_requirements?: Array<NpmRequirement> | null`

      Optional list of npm packages required by this tool.

      - `name: string`

        Name of the npm package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `pip_requirements?: Array<PipRequirement> | null`

      Optional list of pip packages required by this tool.

      - `name: string`

        Name of the pip package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `project_id?: string | null`

      The project id of the tool.

    - `return_char_limit?: number`

      The maximum number of characters in the response.

    - `source_code?: string | null`

      The source code of the function.

    - `source_type?: string | null`

      The type of the source code.

    - `tags?: Array<string>`

      Metadata tags.

    - `tool_type?: ToolType`

      The type of the tool.

      - `"custom"`

      - `"letta_core"`

      - `"letta_memory_core"`

      - `"letta_multi_agent_core"`

      - `"letta_sleeptime_core"`

      - `"letta_voice_sleeptime_core"`

      - `"letta_builtin"`

      - `"letta_files_core"`

      - `"external_langchain"`

      - `"external_composio"`

      - `"external_mcp"`

  - `base_template_id?: string | null`

    The base template id of the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `created_at?: string | null`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the agent will be hidden.

  - `identities?: Array<Identity>`

    The identities associated with this agent.

    - `id: string`

      The human-friendly ID of the Identity

    - `agent_ids: Array<string>`

      The IDs of the agents associated with the identity.

    - `block_ids: Array<string>`

      The IDs of the blocks associated with the identity.

    - `identifier_key: string`

      External, user-generated identifier key of the identity.

    - `identity_type: "org" | "user" | "other"`

      The type of the identity.

      - `"org"`

      - `"user"`

      - `"other"`

    - `name: string`

      The name of the identity.

    - `project_id?: string | null`

      The project id of the identity, if applicable.

    - `properties?: Array<Property>`

      List of properties associated with the identity

      - `key: string`

        The key of the property

      - `type: "string" | "number" | "boolean" | "json"`

        The type of the property

        - `"string"`

        - `"number"`

        - `"boolean"`

        - `"json"`

      - `value: string | number | boolean | Record<string, unknown>`

        The value of the property

        - `string`

        - `number`

        - `boolean`

        - `Record<string, unknown>`

  - `identity_ids?: Array<string>`

    Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

  - `last_run_completion?: string | null`

    The timestamp when the agent last completed a run.

  - `last_run_duration_ms?: number | null`

    The duration in milliseconds of the agent's last run.

  - `last_stop_reason?: StopReasonType | null`

    The stop reason from the agent's last run.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `managed_group?: ManagedGroup | null`

    The multi-agent group that this agent manages

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `message_buffer_autoclear?: boolean`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `message_ids?: Array<string> | null`

    The ids of the messages in the agent's in-context memory.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle used by the agent (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings used by the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `multi_agent_group?: MultiAgentGroup | null`

    Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `pending_approval?: ApprovalRequestMessage | null`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project_id?: string | null`

    The id of the project the agent belongs to.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format used by the agent

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Array<AgentEnvironmentVariable>`

    The environment variables for tool execution specific to this agent.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `template_id?: string | null`

    The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

    Deprecated: use `secrets` field instead.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The list of tool rules.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const agentState = await client.agents.folders.attach(
  'source-123e4567-e89b-42d3-8456-426614174000',
  { agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000' },
);

console.log(agentState.id);
```

#### Response

```json
{
  "id": "id",
  "agent_type": "memgpt_agent",
  "blocks": [
    {
      "value": "value",
      "id": "block-123e4567-e89b-12d3-a456-426614174000",
      "base_template_id": "base_template_id",
      "created_by_id": "created_by_id",
      "deployment_id": "deployment_id",
      "description": "description",
      "entity_id": "entity_id",
      "hidden": true,
      "is_template": true,
      "label": "label",
      "last_updated_by_id": "last_updated_by_id",
      "limit": 0,
      "metadata": {
        "foo": "bar"
      },
      "preserve_on_migration": true,
      "project_id": "project_id",
      "read_only": true,
      "tags": [
        "string"
      ],
      "template_id": "template_id",
      "template_name": "template_name"
    }
  ],
  "llm_config": {
    "context_window": 0,
    "model": "model",
    "model_endpoint_type": "openai",
    "compatibility_type": "gguf",
    "display_name": "display_name",
    "effort": "low",
    "enable_reasoner": true,
    "frequency_penalty": 0,
    "handle": "handle",
    "max_reasoning_tokens": 0,
    "max_tokens": 0,
    "model_endpoint": "model_endpoint",
    "model_wrapper": "model_wrapper",
    "parallel_tool_calls": true,
    "provider_category": "base",
    "provider_name": "provider_name",
    "put_inner_thoughts_in_kwargs": true,
    "reasoning_effort": "none",
    "response_format": {
      "type": "text"
    },
    "return_logprobs": true,
    "return_token_ids": true,
    "strict": true,
    "temperature": 0,
    "tier": "tier",
    "tool_call_parser": "tool_call_parser",
    "top_logprobs": 0,
    "verbosity": "low"
  },
  "memory": {
    "blocks": [
      {
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "agent_type": "memgpt_agent",
    "file_blocks": [
      {
        "file_id": "file_id",
        "is_open": true,
        "source_id": "source_id",
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_accessed_at": "2019-12-27T18:11:19.117Z",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "git_enabled": true,
    "prompt_template": "prompt_template"
  },
  "name": "name",
  "sources": [
    {
      "id": "source-123e4567-e89b-12d3-a456-426614174000",
      "embedding_config": {
        "embedding_dim": 0,
        "embedding_endpoint_type": "openai",
        "embedding_model": "embedding_model",
        "azure_deployment": "azure_deployment",
        "azure_endpoint": "azure_endpoint",
        "azure_version": "azure_version",
        "batch_size": 0,
        "embedding_chunk_size": 0,
        "embedding_endpoint": "embedding_endpoint",
        "handle": "handle"
      },
      "name": "name",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "instructions": "instructions",
      "last_updated_by_id": "last_updated_by_id",
      "metadata": {
        "foo": "bar"
      },
      "updated_at": "2019-12-27T18:11:19.117Z",
      "vector_db_provider": "native"
    }
  ],
  "system": "system",
  "tags": [
    "string"
  ],
  "tools": [
    {
      "id": "tool-123e4567-e89b-12d3-a456-426614174000",
      "args_json_schema": {
        "foo": "bar"
      },
      "created_by_id": "created_by_id",
      "default_requires_approval": true,
      "description": "description",
      "enable_parallel_execution": true,
      "json_schema": {
        "foo": "bar"
      },
      "last_updated_by_id": "last_updated_by_id",
      "metadata_": {
        "foo": "bar"
      },
      "name": "name",
      "npm_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "pip_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "project_id": "project_id",
      "return_char_limit": 1,
      "source_code": "source_code",
      "source_type": "source_type",
      "tags": [
        "string"
      ],
      "tool_type": "custom"
    }
  ],
  "base_template_id": "base_template_id",
  "compaction_settings": {
    "clip_chars": 0,
    "mode": "all",
    "model": "model",
    "model_settings": {
      "max_output_tokens": 0,
      "parallel_tool_calls": true,
      "provider_type": "openai",
      "reasoning": {
        "reasoning_effort": "none"
      },
      "response_format": {
        "type": "text"
      },
      "strict": true,
      "temperature": 0
    },
    "prompt": "prompt",
    "prompt_acknowledgement": true,
    "sliding_window_percentage": 0
  },
  "created_at": "2019-12-27T18:11:19.117Z",
  "created_by_id": "created_by_id",
  "deployment_id": "deployment_id",
  "description": "description",
  "embedding": "embedding",
  "embedding_config": {
    "embedding_dim": 0,
    "embedding_endpoint_type": "openai",
    "embedding_model": "embedding_model",
    "azure_deployment": "azure_deployment",
    "azure_endpoint": "azure_endpoint",
    "azure_version": "azure_version",
    "batch_size": 0,
    "embedding_chunk_size": 0,
    "embedding_endpoint": "embedding_endpoint",
    "handle": "handle"
  },
  "enable_sleeptime": true,
  "entity_id": "entity_id",
  "hidden": true,
  "identities": [
    {
      "id": "identity-123e4567-e89b-12d3-a456-426614174000",
      "agent_ids": [
        "string"
      ],
      "block_ids": [
        "string"
      ],
      "identifier_key": "identifier_key",
      "identity_type": "org",
      "name": "name",
      "project_id": "project_id",
      "properties": [
        {
          "key": "key",
          "type": "string",
          "value": "string"
        }
      ]
    }
  ],
  "identity_ids": [
    "string"
  ],
  "last_run_completion": "2019-12-27T18:11:19.117Z",
  "last_run_duration_ms": 0,
  "last_stop_reason": "end_turn",
  "last_updated_by_id": "last_updated_by_id",
  "managed_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "max_files_open": 0,
  "message_buffer_autoclear": true,
  "message_ids": [
    "string"
  ],
  "metadata": {
    "foo": "bar"
  },
  "model": "model",
  "model_settings": {
    "max_output_tokens": 0,
    "parallel_tool_calls": true,
    "provider_type": "openai",
    "reasoning": {
      "reasoning_effort": "none"
    },
    "response_format": {
      "type": "text"
    },
    "strict": true,
    "temperature": 0
  },
  "multi_agent_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "pending_approval": {
    "id": "id",
    "date": "2019-12-27T18:11:19.117Z",
    "tool_call": {
      "arguments": "arguments",
      "name": "name",
      "tool_call_id": "tool_call_id"
    },
    "is_err": true,
    "message_type": "approval_request_message",
    "name": "name",
    "otid": "otid",
    "run_id": "run_id",
    "sender_id": "sender_id",
    "seq_id": 0,
    "step_id": "step_id",
    "tool_calls": [
      {
        "arguments": "arguments",
        "name": "name",
        "tool_call_id": "tool_call_id"
      }
    ]
  },
  "per_file_view_window_char_limit": 0,
  "project_id": "project_id",
  "response_format": {
    "type": "text"
  },
  "secrets": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "template_id": "template_id",
  "timezone": "timezone",
  "tool_exec_environment_variables": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "tool_rules": [
    {
      "children": [
        "string"
      ],
      "tool_name": "tool_name",
      "child_arg_nodes": [
        {
          "name": "name",
          "args": {
            "foo": "bar"
          }
        }
      ],
      "prompt_template": "prompt_template",
      "type": "constrain_child_tools"
    }
  ],
  "updated_at": "2019-12-27T18:11:19.117Z"
}
```

## Detach Folder From Agent

`client.agents.folders.detach(stringfolderID, FolderDetachParamsparams, RequestOptionsoptions?): AgentState | null`

**patch** `/v1/agents/{agent_id}/folders/detach/{folder_id}`

Detach a folder from an agent.

### Parameters

- `folderID: string`

  The ID of the source in the format 'source-<uuid4>'

- `params: FolderDetachParams`

  - `agent_id: string`

    The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `AgentState | null`

  - `id: string`

    The id of the agent. Assigned by the database.

  - `agent_type: AgentType`

    The type of agent.

    - `"memgpt_agent"`

    - `"memgpt_v2_agent"`

    - `"letta_v1_agent"`

    - `"react_agent"`

    - `"workflow_agent"`

    - `"split_thread_agent"`

    - `"sleeptime_agent"`

    - `"voice_convo_agent"`

    - `"voice_sleeptime_agent"`

  - `blocks: Array<Block>`

    The memory blocks used by the agent.

    - `value: string`

      Value of the block.

    - `id?: string`

      The human-friendly ID of the Block

    - `base_template_id?: string | null`

      The base template id of the block.

    - `created_by_id?: string | null`

      The id of the user that made this Block.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `description?: string | null`

      Description of the block.

    - `entity_id?: string | null`

      The id of the entity within the template.

    - `hidden?: boolean | null`

      If set to True, the block will be hidden.

    - `is_template?: boolean`

      Whether the block is a template (e.g. saved human/persona options).

    - `label?: string | null`

      Label of the block (e.g. 'human', 'persona') in the context window.

    - `last_updated_by_id?: string | null`

      The id of the user that last updated this Block.

    - `limit?: number`

      Character limit of the block.

    - `metadata?: Record<string, unknown> | null`

      Metadata of the block.

    - `preserve_on_migration?: boolean | null`

      Preserve the block on template migration.

    - `project_id?: string | null`

      The associated project id.

    - `read_only?: boolean`

      Whether the agent has read-only access to the block.

    - `tags?: Array<string> | null`

      The tags associated with the block.

    - `template_id?: string | null`

      The id of the template.

    - `template_name?: string | null`

      Name of the block if it is a template.

  - `llm_config: LlmConfig`

    Deprecated: Use `model` field instead. The LLM configuration used by the agent.

    - `context_window: number`

      The context window size for the model.

    - `model: string`

      LLM model name.

    - `model_endpoint_type: "openai" | "anthropic" | "google_ai" | 27 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"lmstudio-chatcompletions"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"minimax"`

      - `"moonshot"`

      - `"moonshot_coding"`

      - `"mistral"`

      - `"together"`

      - `"bedrock"`

      - `"deepseek"`

      - `"xai"`

      - `"zai"`

      - `"zai_coding"`

      - `"baseten"`

      - `"fireworks"`

      - `"openrouter"`

      - `"chatgpt_oauth"`

    - `compatibility_type?: "gguf" | "mlx" | null`

      The framework compatibility type for the model.

      - `"gguf"`

      - `"mlx"`

    - `display_name?: string | null`

      A human-friendly display name for the model.

    - `effort?: "low" | "medium" | "high" | 2 more | null`

      The effort level for Anthropic models that support it (Opus 4.5+). Controls token spending and thinking behavior. Not setting this gives similar performance to 'high'.

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

      - `"max"`

    - `enable_reasoner?: boolean`

      Whether or not the model should use extended thinking if it is a 'reasoning' style model

    - `frequency_penalty?: number | null`

      Positive values penalize new tokens based on their existing frequency in the text so far, decreasing the model's likelihood to repeat the same line verbatim. From OpenAI: Number between -2.0 and 2.0.

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

    - `max_reasoning_tokens?: number`

      Configurable thinking budget for extended thinking. Used for enable_reasoner and also for Google Vertex models like Gemini 2.5 Flash. Minimum value is 1024 when used with enable_reasoner.

    - `max_tokens?: number | null`

      The maximum number of tokens to generate. If not set, the model will use its default value.

    - `model_endpoint?: string | null`

      The endpoint for the model.

    - `model_wrapper?: string | null`

      The wrapper for the model.

    - `parallel_tool_calls?: boolean | null`

      Deprecated: Use model_settings to configure parallel tool calls instead. If set to True, enables parallel tool calling. Defaults to False.

    - `provider_category?: ProviderCategory | null`

      The provider category for the model.

      - `"base"`

      - `"byok"`

    - `provider_name?: string | null`

      The provider name for the model.

    - `put_inner_thoughts_in_kwargs?: boolean | null`

      Puts 'inner_thoughts' as a kwarg in the function call if this is set to True. This helps with function calling performance and also the generation of inner thoughts.

    - `reasoning_effort?: "none" | "minimal" | "low" | 3 more | null`

      The reasoning effort to use when generating text reasoning models

      - `"none"`

      - `"minimal"`

      - `"low"`

      - `"medium"`

      - `"high"`

      - `"xhigh"`

    - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

      The response format for the model's output. Supports text, json_object, and json_schema (structured outputs). Can be set via model_settings.

      - `TextResponseFormat`

        Response format for plain text responses.

        - `type?: "text"`

          The type of the response format.

          - `"text"`

      - `JsonSchemaResponseFormat`

        Response format for JSON schema-based responses.

        - `json_schema: Record<string, unknown>`

          The JSON schema of the response.

        - `type?: "json_schema"`

          The type of the response format.

          - `"json_schema"`

      - `JsonObjectResponseFormat`

        Response format for JSON object responses.

        - `type?: "json_object"`

          The type of the response format.

          - `"json_object"`

    - `return_logprobs?: boolean`

      Whether to return log probabilities of the output tokens. Useful for RL training.

    - `return_token_ids?: boolean`

      Whether to return token IDs for all LLM generations via SGLang native endpoint. Required for multi-turn RL training with loss masking. Only works with SGLang provider.

    - `strict?: boolean`

      Enable strict mode for tool calling. When true, tool schemas include strict: true and additionalProperties: false, guaranteeing tool outputs match JSON schemas.

    - `temperature?: number`

      The temperature to use when generating text with the model. A higher temperature will result in more random text.

    - `tier?: string | null`

      The cost tier for the model (cloud only).

    - `tool_call_parser?: string | null`

      SGLang tool call parser name (e.g. 'glm47', 'qwen25', 'hermes'). Used by the SGLang native adapter to parse tool calls from raw model output.

    - `top_logprobs?: number | null`

      Number of most likely tokens to return at each position (0-20). Requires return_logprobs=True.

    - `verbosity?: "low" | "medium" | "high" | null`

      Soft control for how verbose model output should be, used for GPT-5 models.

      - `"low"`

      - `"medium"`

      - `"high"`

  - `memory: Memory`

    Deprecated: Use `blocks` field instead. The in-context memory of the agent.

    - `blocks: Array<Block>`

      Memory blocks contained in the agent's in-context memory

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `agent_type?: AgentType | (string & {}) | null`

      Agent type controlling prompt rendering.

      - `AgentType = "memgpt_agent" | "memgpt_v2_agent" | "letta_v1_agent" | 6 more`

        Enum to represent the type of agent.

      - `(string & {})`

    - `file_blocks?: Array<FileBlock>`

      Special blocks representing the agent's in-context memory of an attached file

      - `file_id: string`

        Unique identifier of the file.

      - `is_open: boolean`

        True if the agent currently has the file open.

      - `source_id: string`

        Deprecated: Use `folder_id` field instead. Unique identifier of the source.

      - `value: string`

        Value of the block.

      - `id?: string`

        The human-friendly ID of the Block

      - `base_template_id?: string | null`

        The base template id of the block.

      - `created_by_id?: string | null`

        The id of the user that made this Block.

      - `deployment_id?: string | null`

        The id of the deployment.

      - `description?: string | null`

        Description of the block.

      - `entity_id?: string | null`

        The id of the entity within the template.

      - `hidden?: boolean | null`

        If set to True, the block will be hidden.

      - `is_template?: boolean`

        Whether the block is a template (e.g. saved human/persona options).

      - `label?: string | null`

        Label of the block (e.g. 'human', 'persona') in the context window.

      - `last_accessed_at?: string | null`

        UTC timestamp of the agent’s most recent access to this file. Any operations from the open, close, or search tools will update this field.

      - `last_updated_by_id?: string | null`

        The id of the user that last updated this Block.

      - `limit?: number`

        Character limit of the block.

      - `metadata?: Record<string, unknown> | null`

        Metadata of the block.

      - `preserve_on_migration?: boolean | null`

        Preserve the block on template migration.

      - `project_id?: string | null`

        The associated project id.

      - `read_only?: boolean`

        Whether the agent has read-only access to the block.

      - `tags?: Array<string> | null`

        The tags associated with the block.

      - `template_id?: string | null`

        The id of the template.

      - `template_name?: string | null`

        Name of the block if it is a template.

    - `git_enabled?: boolean`

      Whether this agent uses git-backed memory with structured labels.

    - `prompt_template?: string`

      Deprecated. Ignored for performance.

  - `name: string`

    The name of the agent.

  - `sources: Array<Source>`

    Deprecated: Use `folders` field instead. The sources used by the agent.

    - `id: string`

      The human-friendly ID of the Source

    - `embedding_config: EmbeddingConfig`

      The embedding configuration used by the source.

      - `embedding_dim: number`

        The dimension of the embedding.

      - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

        The endpoint type for the model.

        - `"openai"`

        - `"anthropic"`

        - `"bedrock"`

        - `"google_ai"`

        - `"google_vertex"`

        - `"azure"`

        - `"groq"`

        - `"ollama"`

        - `"webui"`

        - `"webui-legacy"`

        - `"lmstudio"`

        - `"lmstudio-legacy"`

        - `"llamacpp"`

        - `"koboldcpp"`

        - `"vllm"`

        - `"hugging-face"`

        - `"mistral"`

        - `"together"`

        - `"pinecone"`

      - `embedding_model: string`

        The model for the embedding.

      - `azure_deployment?: string | null`

        The Azure deployment for the model.

      - `azure_endpoint?: string | null`

        The Azure endpoint for the model.

      - `azure_version?: string | null`

        The Azure version for the model.

      - `batch_size?: number`

        The maximum batch size for processing embeddings.

      - `embedding_chunk_size?: number | null`

        The chunk size of the embedding.

      - `embedding_endpoint?: string | null`

        The endpoint for the model (`None` if local).

      - `handle?: string | null`

        The handle for this config, in the format provider/model-name.

    - `name: string`

      The name of the source.

    - `created_at?: string | null`

      The timestamp when the source was created.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `description?: string | null`

      The description of the source.

    - `instructions?: string | null`

      Instructions for how to use the source.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata?: Record<string, unknown> | null`

      Metadata associated with the source.

    - `updated_at?: string | null`

      The timestamp when the source was last updated.

    - `vector_db_provider?: VectorDBProvider`

      The vector database provider used for this source's passages

      - `"native"`

      - `"tpuf"`

      - `"pinecone"`

  - `system: string`

    The system prompt used by the agent.

  - `tags: Array<string>`

    The tags associated with the agent.

  - `tools: Array<Tool>`

    The tools used by the agent.

    - `id: string`

      The human-friendly ID of the Tool

    - `args_json_schema?: Record<string, unknown> | null`

      The args JSON schema of the function.

    - `created_by_id?: string | null`

      The id of the user that made this Tool.

    - `default_requires_approval?: boolean | null`

      Default value for whether or not executing this tool requires approval.

    - `description?: string | null`

      The description of the tool.

    - `enable_parallel_execution?: boolean | null`

      If set to True, then this tool will potentially be executed concurrently with other tools. Default False.

    - `json_schema?: Record<string, unknown> | null`

      The JSON schema of the function.

    - `last_updated_by_id?: string | null`

      The id of the user that made this Tool.

    - `metadata_?: Record<string, unknown> | null`

      A dictionary of additional metadata for the tool.

    - `name?: string | null`

      The name of the function.

    - `npm_requirements?: Array<NpmRequirement> | null`

      Optional list of npm packages required by this tool.

      - `name: string`

        Name of the npm package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `pip_requirements?: Array<PipRequirement> | null`

      Optional list of pip packages required by this tool.

      - `name: string`

        Name of the pip package.

      - `version?: string | null`

        Optional version of the package, following semantic versioning.

    - `project_id?: string | null`

      The project id of the tool.

    - `return_char_limit?: number`

      The maximum number of characters in the response.

    - `source_code?: string | null`

      The source code of the function.

    - `source_type?: string | null`

      The type of the source code.

    - `tags?: Array<string>`

      Metadata tags.

    - `tool_type?: ToolType`

      The type of the tool.

      - `"custom"`

      - `"letta_core"`

      - `"letta_memory_core"`

      - `"letta_multi_agent_core"`

      - `"letta_sleeptime_core"`

      - `"letta_voice_sleeptime_core"`

      - `"letta_builtin"`

      - `"letta_files_core"`

      - `"external_langchain"`

      - `"external_composio"`

      - `"external_mcp"`

  - `base_template_id?: string | null`

    The base template id of the agent.

  - `compaction_settings?: CompactionSettings | null`

    Configuration for conversation compaction / summarization.

    Per-model settings (temperature,
    max tokens, etc.) are derived from the default configuration for that handle.

    - `clip_chars?: number | null`

      The maximum length of the summary in characters. If none, no clipping is performed.

    - `mode?: "all" | "sliding_window" | "self_compact_all" | "self_compact_sliding_window"`

      The type of summarization technique use.

      - `"all"`

      - `"sliding_window"`

      - `"self_compact_all"`

      - `"self_compact_sliding_window"`

    - `model?: string | null`

      Model handle to use for sliding_window/all summarization (format: provider/model-name). If None, uses lightweight provider-specific defaults.

    - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

      Optional model settings used to override defaults for the summarizer model.

      - `OpenAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openai"`

          The type of the provider.

          - `"openai"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `SgLangModelSettings`

        SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "sglang"`

          The type of the provider.

          - `"sglang"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

            The reasoning effort to use when generating text reasoning models

            - `"none"`

            - `"minimal"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `tool_call_parser?: string | null`

          SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

      - `AnthropicModelSettings`

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "anthropic"`

          The type of the provider.

          - `"anthropic"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GoogleAIModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_ai"`

          The type of the provider.

          - `"google_ai"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `GoogleVertexModelSettings`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "google_vertex"`

          The type of the provider.

          - `"google_vertex"`

        - `response_schema?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response schema for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking_config?: ThinkingConfig`

          The thinking configuration for the model.

          - `include_thoughts?: boolean`

            Whether to include thoughts in the model's response.

          - `thinking_budget?: number`

            The thinking budget for the model.

      - `AzureModelSettings`

        Azure OpenAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "azure"`

          The type of the provider.

          - `"azure"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `XaiModelSettings`

        xAI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "xai"`

          The type of the provider.

          - `"xai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `MoonshotModelSettings`

        Moonshot/Kimi model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot"`

          The type of the provider.

          - `"moonshot"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

      - `ZaiModelSettings`

        Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "zai"`

          The type of the provider.

          - `"zai"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for GLM-4.5+ models.

          - `clear_thinking?: boolean`

            If False, preserved thinking is used (recommended for agents).

          - `type?: "enabled" | "disabled"`

            Whether thinking is enabled or disabled.

            - `"enabled"`

            - `"disabled"`

      - `MoonshotCodingModelSettings`

        Kimi Code model configuration (Anthropic-compatible).

        - `effort?: "low" | "medium" | "high" | 2 more | null`

          Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

          - `"max"`

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "moonshot_coding"`

          The type of the provider.

          - `"moonshot_coding"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `strict?: boolean`

          Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

        - `temperature?: number`

          The temperature of the model.

        - `thinking?: Thinking`

          The thinking configuration for the model.

          - `budget_tokens?: number`

            The maximum number of tokens the model can use for extended thinking.

          - `type?: "enabled" | "disabled"`

            The type of thinking to use.

            - `"enabled"`

            - `"disabled"`

        - `verbosity?: "low" | "medium" | "high" | null`

          Soft control for how verbose model output should be, used for GPT-5 models.

          - `"low"`

          - `"medium"`

          - `"high"`

      - `GroqModelSettings`

        Groq model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "groq"`

          The type of the provider.

          - `"groq"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `DeepseekModelSettings`

        Deepseek model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "deepseek"`

          The type of the provider.

          - `"deepseek"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `TogetherModelSettings`

        Together AI model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "together"`

          The type of the provider.

          - `"together"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BedrockModelSettings`

        AWS Bedrock model configuration.

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "bedrock"`

          The type of the provider.

          - `"bedrock"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `BasetenModelSettings`

        Baseten model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "baseten"`

          The type of the provider.

          - `"baseten"`

        - `temperature?: number`

          The temperature of the model.

      - `OpenRouterModelSettings`

        OpenRouter model configuration (OpenAI-compatible).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "openrouter"`

          The type of the provider.

          - `"openrouter"`

        - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

          The response format for the model.

          - `TextResponseFormat`

            Response format for plain text responses.

          - `JsonSchemaResponseFormat`

            Response format for JSON schema-based responses.

          - `JsonObjectResponseFormat`

            Response format for JSON object responses.

        - `temperature?: number`

          The temperature of the model.

      - `ChatGptoAuthModelSettings`

        ChatGPT OAuth model configuration (uses ChatGPT backend API).

        - `max_output_tokens?: number`

          The maximum number of tokens the model can generate.

        - `parallel_tool_calls?: boolean`

          Whether to enable parallel tool calling.

        - `provider_type?: "chatgpt_oauth"`

          The type of the provider.

          - `"chatgpt_oauth"`

        - `reasoning?: Reasoning`

          The reasoning configuration for the model.

          - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

            The reasoning effort level for GPT-5.x and o-series models.

            - `"none"`

            - `"low"`

            - `"medium"`

            - `"high"`

            - `"xhigh"`

        - `temperature?: number`

          The temperature of the model.

    - `prompt?: string | null`

      The prompt to use for summarization. If None, uses mode-specific default.

    - `prompt_acknowledgement?: boolean`

      Whether to include an acknowledgement post-prompt (helps prevent non-summary outputs).

    - `sliding_window_percentage?: number`

      The percentage of the context window to keep post-summarization (only used in sliding window modes).

  - `created_at?: string | null`

    The timestamp when the object was created.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `deployment_id?: string | null`

    The id of the deployment.

  - `description?: string | null`

    The description of the agent.

  - `embedding?: string | null`

    The embedding model handle used by the agent (format: provider/model-name).

  - `embedding_config?: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

  - `enable_sleeptime?: boolean | null`

    If set to True, memory management will move to a background agent thread.

  - `entity_id?: string | null`

    The id of the entity within the template.

  - `hidden?: boolean | null`

    If set to True, the agent will be hidden.

  - `identities?: Array<Identity>`

    The identities associated with this agent.

    - `id: string`

      The human-friendly ID of the Identity

    - `agent_ids: Array<string>`

      The IDs of the agents associated with the identity.

    - `block_ids: Array<string>`

      The IDs of the blocks associated with the identity.

    - `identifier_key: string`

      External, user-generated identifier key of the identity.

    - `identity_type: "org" | "user" | "other"`

      The type of the identity.

      - `"org"`

      - `"user"`

      - `"other"`

    - `name: string`

      The name of the identity.

    - `project_id?: string | null`

      The project id of the identity, if applicable.

    - `properties?: Array<Property>`

      List of properties associated with the identity

      - `key: string`

        The key of the property

      - `type: "string" | "number" | "boolean" | "json"`

        The type of the property

        - `"string"`

        - `"number"`

        - `"boolean"`

        - `"json"`

      - `value: string | number | boolean | Record<string, unknown>`

        The value of the property

        - `string`

        - `number`

        - `boolean`

        - `Record<string, unknown>`

  - `identity_ids?: Array<string>`

    Deprecated: Use `identities` field instead. The ids of the identities associated with this agent.

  - `last_run_completion?: string | null`

    The timestamp when the agent last completed a run.

  - `last_run_duration_ms?: number | null`

    The duration in milliseconds of the agent's last run.

  - `last_stop_reason?: StopReasonType | null`

    The stop reason from the agent's last run.

    - `"end_turn"`

    - `"error"`

    - `"llm_api_error"`

    - `"invalid_llm_response"`

    - `"invalid_tool_call"`

    - `"max_steps"`

    - `"max_tokens_exceeded"`

    - `"no_tool_call"`

    - `"tool_rule"`

    - `"cancelled"`

    - `"insufficient_credits"`

    - `"requires_approval"`

    - `"context_window_overflow_in_system_prompt"`

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `managed_group?: ManagedGroup | null`

    The multi-agent group that this agent manages

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `max_files_open?: number | null`

    Maximum number of files that can be open at once for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `message_buffer_autoclear?: boolean`

    If set to True, the agent will not remember previous messages (though the agent will still retain state via core memory blocks and archival/recall memory). Not recommended unless you have an advanced use case.

  - `message_ids?: Array<string> | null`

    The ids of the messages in the agent's in-context memory.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the agent.

  - `model?: string | null`

    The model handle used by the agent (format: provider/model-name).

  - `model_settings?: OpenAIModelSettings | SgLangModelSettings | AnthropicModelSettings | 14 more | null`

    The model settings used by the agent.

    - `OpenAIModelSettings`

    - `SgLangModelSettings`

      SGLang model configuration (OpenAI-compatible runtime with SGLang-specific parsing).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "sglang"`

        The type of the provider.

        - `"sglang"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "minimal" | "low" | 3 more`

          The reasoning effort to use when generating text reasoning models

          - `"none"`

          - `"minimal"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `tool_call_parser?: string | null`

        SGLang tool call parser name (for example 'glm47', 'qwen25', or 'hermes').

    - `AnthropicModelSettings`

    - `GoogleAIModelSettings`

    - `GoogleVertexModelSettings`

    - `AzureModelSettings`

      Azure OpenAI model configuration (OpenAI-compatible).

    - `XaiModelSettings`

      xAI model configuration (OpenAI-compatible).

    - `MoonshotModelSettings`

      Moonshot/Kimi model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot"`

        The type of the provider.

        - `"moonshot"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

    - `ZaiModelSettings`

      Z.ai (ZhipuAI) model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "zai"`

        The type of the provider.

        - `"zai"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for GLM-4.5+ models.

        - `clear_thinking?: boolean`

          If False, preserved thinking is used (recommended for agents).

        - `type?: "enabled" | "disabled"`

          Whether thinking is enabled or disabled.

          - `"enabled"`

          - `"disabled"`

    - `MoonshotCodingModelSettings`

      Kimi Code model configuration (Anthropic-compatible).

      - `effort?: "low" | "medium" | "high" | 2 more | null`

        Effort level for supported Anthropic models (controls token spending). 'xhigh' and 'max' are available on Opus 4.6+. Not setting this gives similar performance to 'high'.

        - `"low"`

        - `"medium"`

        - `"high"`

        - `"xhigh"`

        - `"max"`

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "moonshot_coding"`

        The type of the provider.

        - `"moonshot_coding"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `strict?: boolean`

        Enable strict mode for tool calling. When true, tool outputs are guaranteed to match JSON schemas.

      - `temperature?: number`

        The temperature of the model.

      - `thinking?: Thinking`

        The thinking configuration for the model.

        - `budget_tokens?: number`

          The maximum number of tokens the model can use for extended thinking.

        - `type?: "enabled" | "disabled"`

          The type of thinking to use.

          - `"enabled"`

          - `"disabled"`

      - `verbosity?: "low" | "medium" | "high" | null`

        Soft control for how verbose model output should be, used for GPT-5 models.

        - `"low"`

        - `"medium"`

        - `"high"`

    - `GroqModelSettings`

      Groq model configuration (OpenAI-compatible).

    - `DeepseekModelSettings`

      Deepseek model configuration (OpenAI-compatible).

    - `TogetherModelSettings`

      Together AI model configuration (OpenAI-compatible).

    - `BedrockModelSettings`

      AWS Bedrock model configuration.

    - `BasetenModelSettings`

      Baseten model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "baseten"`

        The type of the provider.

        - `"baseten"`

      - `temperature?: number`

        The temperature of the model.

    - `OpenRouterModelSettings`

      OpenRouter model configuration (OpenAI-compatible).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "openrouter"`

        The type of the provider.

        - `"openrouter"`

      - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

        The response format for the model.

        - `TextResponseFormat`

          Response format for plain text responses.

        - `JsonSchemaResponseFormat`

          Response format for JSON schema-based responses.

        - `JsonObjectResponseFormat`

          Response format for JSON object responses.

      - `temperature?: number`

        The temperature of the model.

    - `ChatGptoAuthModelSettings`

      ChatGPT OAuth model configuration (uses ChatGPT backend API).

      - `max_output_tokens?: number`

        The maximum number of tokens the model can generate.

      - `parallel_tool_calls?: boolean`

        Whether to enable parallel tool calling.

      - `provider_type?: "chatgpt_oauth"`

        The type of the provider.

        - `"chatgpt_oauth"`

      - `reasoning?: Reasoning`

        The reasoning configuration for the model.

        - `reasoning_effort?: "none" | "low" | "medium" | 2 more`

          The reasoning effort level for GPT-5.x and o-series models.

          - `"none"`

          - `"low"`

          - `"medium"`

          - `"high"`

          - `"xhigh"`

      - `temperature?: number`

        The temperature of the model.

  - `multi_agent_group?: MultiAgentGroup | null`

    Deprecated: Use `managed_group` field instead. The multi-agent group that this agent manages.

    - `id: string`

      The id of the group. Assigned by the database.

    - `agent_ids: Array<string>`

    - `description: string`

    - `manager_type: "round_robin" | "supervisor" | "dynamic" | 3 more`

      - `"round_robin"`

      - `"supervisor"`

      - `"dynamic"`

      - `"sleeptime"`

      - `"voice_sleeptime"`

      - `"swarm"`

    - `base_template_id?: string | null`

      The base template id.

    - `deployment_id?: string | null`

      The id of the deployment.

    - `hidden?: boolean | null`

      If set to True, the group will be hidden.

    - `last_processed_message_id?: string | null`

    - `manager_agent_id?: string | null`

    - `max_message_buffer_length?: number | null`

      The desired maximum length of messages in the context window of the convo agent. This is a best effort, and may be off slightly due to user/assistant interleaving.

    - `max_turns?: number | null`

    - `min_message_buffer_length?: number | null`

      The desired minimum length of messages in the context window of the convo agent. This is a best effort, and may be off-by-one due to user/assistant interleaving.

    - `project_id?: string | null`

      The associated project id.

    - `shared_block_ids?: Array<string>`

    - `sleeptime_agent_frequency?: number | null`

    - `template_id?: string | null`

      The id of the template.

    - `termination_token?: string | null`

    - `turns_counter?: number | null`

  - `pending_approval?: ApprovalRequestMessage | null`

    A message representing a request for approval to call a tool (generated by the LLM to trigger tool execution).

    Args:
    id (str): The ID of the message
    date (datetime): The date the message was created in ISO format
    name (Optional[str]): The name of the sender of the message
    tool_call (ToolCall): The tool call

    - `id: string`

    - `date: string`

    - `tool_call: ToolCall | ToolCallDelta`

      The tool call that has been requested by the llm to run

      - `ToolCall`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

        - `arguments?: string | null`

        - `name?: string | null`

        - `tool_call_id?: string | null`

    - `is_err?: boolean | null`

    - `message_type?: "approval_request_message"`

      The type of the message.

      - `"approval_request_message"`

    - `name?: string | null`

    - `otid?: string | null`

      The offline threading id (OTID). Set by the client to deduplicate requests. Used for idempotency in background streaming mode — each message in a request must have a unique OTID. Retries of the same request should reuse the same OTIDs.

    - `run_id?: string | null`

    - `sender_id?: string | null`

    - `seq_id?: number | null`

    - `step_id?: string | null`

    - `tool_calls?: Array<ToolCall> | ToolCallDelta | null`

      The tool calls that have been requested by the llm to run, which are pending approval

      - `Array<ToolCall>`

        - `arguments: string`

        - `name: string`

        - `tool_call_id: string`

      - `ToolCallDelta`

  - `per_file_view_window_char_limit?: number | null`

    The per-file view window character limit for this agent. Setting this too high may exceed the context window, which will break the agent.

  - `project_id?: string | null`

    The id of the project the agent belongs to.

  - `response_format?: TextResponseFormat | JsonSchemaResponseFormat | JsonObjectResponseFormat | null`

    The response format used by the agent

    - `TextResponseFormat`

      Response format for plain text responses.

    - `JsonSchemaResponseFormat`

      Response format for JSON schema-based responses.

    - `JsonObjectResponseFormat`

      Response format for JSON object responses.

  - `secrets?: Array<AgentEnvironmentVariable>`

    The environment variables for tool execution specific to this agent.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `template_id?: string | null`

    The id of the template the agent belongs to.

  - `timezone?: string | null`

    The timezone of the agent (IANA format).

  - `tool_exec_environment_variables?: Array<AgentEnvironmentVariable>`

    Deprecated: use `secrets` field instead.

    - `agent_id: string`

      The ID of the agent this environment variable belongs to.

    - `key: string`

      The name of the environment variable.

    - `value: string`

      The value of the environment variable.

    - `id?: string`

      The human-friendly ID of the Agent-env

    - `created_at?: string | null`

      The timestamp when the object was created.

    - `created_by_id?: string | null`

      The id of the user that made this object.

    - `description?: string | null`

      An optional description of the environment variable.

    - `last_updated_by_id?: string | null`

      The id of the user that made this object.

    - `updated_at?: string | null`

      The timestamp when the object was last updated.

    - `value_enc?: string | null`

      Encrypted secret value (stored as encrypted string)

  - `tool_rules?: Array<ChildToolRule | InitToolRule | TerminalToolRule | 6 more> | null`

    The list of tool rules.

    - `ChildToolRule`

      A ToolRule represents a tool that can be invoked by the agent.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `child_arg_nodes?: Array<ChildArgNode> | null`

        Optional list of typed child argument overrides. Each node must reference a child in 'children'.

        - `name: string`

          The name of the child tool to invoke next.

        - `args?: Record<string, unknown> | null`

          Optional prefilled arguments for this child tool. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "constrain_child_tools"`

        - `"constrain_child_tools"`

    - `InitToolRule`

      Represents the initial tool rule configuration.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `args?: Record<string, unknown> | null`

        Optional prefilled arguments for this tool. When present, these values will override any LLM-provided arguments with the same keys during invocation. Keys must match the tool's parameter names and values must satisfy the tool's JSON schema. Supports partial prefill; non-overlapping parameters are left to the model.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "run_first"`

        - `"run_first"`

    - `TerminalToolRule`

      Represents a terminal tool rule configuration where if this tool gets called, it must end the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "exit_loop"`

        - `"exit_loop"`

    - `ConditionalToolRule`

      A ToolRule that conditionally maps to different child tools based on the output.

      - `child_output_mapping: Record<string, string>`

        The output case to check for mapping

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `default_child?: string | null`

        The default child tool to be called. If None, any tool can be called.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `require_output_mapping?: boolean`

        Whether to throw an error when output doesn't match any case

      - `type?: "conditional"`

        - `"conditional"`

    - `ContinueToolRule`

      Represents a tool rule configuration where if this tool gets called, it must continue the agent loop.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "continue_loop"`

        - `"continue_loop"`

    - `RequiredBeforeExitToolRule`

      Represents a tool rule configuration where this tool must be called before the agent loop can exit.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "required_before_exit"`

        - `"required_before_exit"`

    - `MaxCountPerStepToolRule`

      Represents a tool rule configuration which constrains the total number of times this tool can be invoked in a single step.

      - `max_count_limit: number`

        The max limit for the total number of times this tool can be invoked in a single step.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "max_count_per_step"`

        - `"max_count_per_step"`

    - `ParentToolRule`

      A ToolRule that only allows a child tool to be called if the parent has been called.

      - `children: Array<string>`

        The children tools that can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored).

      - `type?: "parent_last_tool"`

        - `"parent_last_tool"`

    - `RequiresApprovalToolRule`

      Represents a tool rule configuration which requires approval before the tool can be invoked.

      - `tool_name: string`

        The name of the tool. Must exist in the database for the user's organization.

      - `prompt_template?: string | null`

        Optional template string (ignored). Rendering uses fast built-in formatting for performance.

      - `type?: "requires_approval"`

        - `"requires_approval"`

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const agentState = await client.agents.folders.detach(
  'source-123e4567-e89b-42d3-8456-426614174000',
  { agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000' },
);

console.log(agentState.id);
```

#### Response

```json
{
  "id": "id",
  "agent_type": "memgpt_agent",
  "blocks": [
    {
      "value": "value",
      "id": "block-123e4567-e89b-12d3-a456-426614174000",
      "base_template_id": "base_template_id",
      "created_by_id": "created_by_id",
      "deployment_id": "deployment_id",
      "description": "description",
      "entity_id": "entity_id",
      "hidden": true,
      "is_template": true,
      "label": "label",
      "last_updated_by_id": "last_updated_by_id",
      "limit": 0,
      "metadata": {
        "foo": "bar"
      },
      "preserve_on_migration": true,
      "project_id": "project_id",
      "read_only": true,
      "tags": [
        "string"
      ],
      "template_id": "template_id",
      "template_name": "template_name"
    }
  ],
  "llm_config": {
    "context_window": 0,
    "model": "model",
    "model_endpoint_type": "openai",
    "compatibility_type": "gguf",
    "display_name": "display_name",
    "effort": "low",
    "enable_reasoner": true,
    "frequency_penalty": 0,
    "handle": "handle",
    "max_reasoning_tokens": 0,
    "max_tokens": 0,
    "model_endpoint": "model_endpoint",
    "model_wrapper": "model_wrapper",
    "parallel_tool_calls": true,
    "provider_category": "base",
    "provider_name": "provider_name",
    "put_inner_thoughts_in_kwargs": true,
    "reasoning_effort": "none",
    "response_format": {
      "type": "text"
    },
    "return_logprobs": true,
    "return_token_ids": true,
    "strict": true,
    "temperature": 0,
    "tier": "tier",
    "tool_call_parser": "tool_call_parser",
    "top_logprobs": 0,
    "verbosity": "low"
  },
  "memory": {
    "blocks": [
      {
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "agent_type": "memgpt_agent",
    "file_blocks": [
      {
        "file_id": "file_id",
        "is_open": true,
        "source_id": "source_id",
        "value": "value",
        "id": "block-123e4567-e89b-12d3-a456-426614174000",
        "base_template_id": "base_template_id",
        "created_by_id": "created_by_id",
        "deployment_id": "deployment_id",
        "description": "description",
        "entity_id": "entity_id",
        "hidden": true,
        "is_template": true,
        "label": "label",
        "last_accessed_at": "2019-12-27T18:11:19.117Z",
        "last_updated_by_id": "last_updated_by_id",
        "limit": 0,
        "metadata": {
          "foo": "bar"
        },
        "preserve_on_migration": true,
        "project_id": "project_id",
        "read_only": true,
        "tags": [
          "string"
        ],
        "template_id": "template_id",
        "template_name": "template_name"
      }
    ],
    "git_enabled": true,
    "prompt_template": "prompt_template"
  },
  "name": "name",
  "sources": [
    {
      "id": "source-123e4567-e89b-12d3-a456-426614174000",
      "embedding_config": {
        "embedding_dim": 0,
        "embedding_endpoint_type": "openai",
        "embedding_model": "embedding_model",
        "azure_deployment": "azure_deployment",
        "azure_endpoint": "azure_endpoint",
        "azure_version": "azure_version",
        "batch_size": 0,
        "embedding_chunk_size": 0,
        "embedding_endpoint": "embedding_endpoint",
        "handle": "handle"
      },
      "name": "name",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "instructions": "instructions",
      "last_updated_by_id": "last_updated_by_id",
      "metadata": {
        "foo": "bar"
      },
      "updated_at": "2019-12-27T18:11:19.117Z",
      "vector_db_provider": "native"
    }
  ],
  "system": "system",
  "tags": [
    "string"
  ],
  "tools": [
    {
      "id": "tool-123e4567-e89b-12d3-a456-426614174000",
      "args_json_schema": {
        "foo": "bar"
      },
      "created_by_id": "created_by_id",
      "default_requires_approval": true,
      "description": "description",
      "enable_parallel_execution": true,
      "json_schema": {
        "foo": "bar"
      },
      "last_updated_by_id": "last_updated_by_id",
      "metadata_": {
        "foo": "bar"
      },
      "name": "name",
      "npm_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "pip_requirements": [
        {
          "name": "x",
          "version": "version"
        }
      ],
      "project_id": "project_id",
      "return_char_limit": 1,
      "source_code": "source_code",
      "source_type": "source_type",
      "tags": [
        "string"
      ],
      "tool_type": "custom"
    }
  ],
  "base_template_id": "base_template_id",
  "compaction_settings": {
    "clip_chars": 0,
    "mode": "all",
    "model": "model",
    "model_settings": {
      "max_output_tokens": 0,
      "parallel_tool_calls": true,
      "provider_type": "openai",
      "reasoning": {
        "reasoning_effort": "none"
      },
      "response_format": {
        "type": "text"
      },
      "strict": true,
      "temperature": 0
    },
    "prompt": "prompt",
    "prompt_acknowledgement": true,
    "sliding_window_percentage": 0
  },
  "created_at": "2019-12-27T18:11:19.117Z",
  "created_by_id": "created_by_id",
  "deployment_id": "deployment_id",
  "description": "description",
  "embedding": "embedding",
  "embedding_config": {
    "embedding_dim": 0,
    "embedding_endpoint_type": "openai",
    "embedding_model": "embedding_model",
    "azure_deployment": "azure_deployment",
    "azure_endpoint": "azure_endpoint",
    "azure_version": "azure_version",
    "batch_size": 0,
    "embedding_chunk_size": 0,
    "embedding_endpoint": "embedding_endpoint",
    "handle": "handle"
  },
  "enable_sleeptime": true,
  "entity_id": "entity_id",
  "hidden": true,
  "identities": [
    {
      "id": "identity-123e4567-e89b-12d3-a456-426614174000",
      "agent_ids": [
        "string"
      ],
      "block_ids": [
        "string"
      ],
      "identifier_key": "identifier_key",
      "identity_type": "org",
      "name": "name",
      "project_id": "project_id",
      "properties": [
        {
          "key": "key",
          "type": "string",
          "value": "string"
        }
      ]
    }
  ],
  "identity_ids": [
    "string"
  ],
  "last_run_completion": "2019-12-27T18:11:19.117Z",
  "last_run_duration_ms": 0,
  "last_stop_reason": "end_turn",
  "last_updated_by_id": "last_updated_by_id",
  "managed_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "max_files_open": 0,
  "message_buffer_autoclear": true,
  "message_ids": [
    "string"
  ],
  "metadata": {
    "foo": "bar"
  },
  "model": "model",
  "model_settings": {
    "max_output_tokens": 0,
    "parallel_tool_calls": true,
    "provider_type": "openai",
    "reasoning": {
      "reasoning_effort": "none"
    },
    "response_format": {
      "type": "text"
    },
    "strict": true,
    "temperature": 0
  },
  "multi_agent_group": {
    "id": "id",
    "agent_ids": [
      "string"
    ],
    "description": "description",
    "manager_type": "round_robin",
    "base_template_id": "base_template_id",
    "deployment_id": "deployment_id",
    "hidden": true,
    "last_processed_message_id": "last_processed_message_id",
    "manager_agent_id": "manager_agent_id",
    "max_message_buffer_length": 0,
    "max_turns": 0,
    "min_message_buffer_length": 0,
    "project_id": "project_id",
    "shared_block_ids": [
      "string"
    ],
    "sleeptime_agent_frequency": 0,
    "template_id": "template_id",
    "termination_token": "termination_token",
    "turns_counter": 0
  },
  "pending_approval": {
    "id": "id",
    "date": "2019-12-27T18:11:19.117Z",
    "tool_call": {
      "arguments": "arguments",
      "name": "name",
      "tool_call_id": "tool_call_id"
    },
    "is_err": true,
    "message_type": "approval_request_message",
    "name": "name",
    "otid": "otid",
    "run_id": "run_id",
    "sender_id": "sender_id",
    "seq_id": 0,
    "step_id": "step_id",
    "tool_calls": [
      {
        "arguments": "arguments",
        "name": "name",
        "tool_call_id": "tool_call_id"
      }
    ]
  },
  "per_file_view_window_char_limit": 0,
  "project_id": "project_id",
  "response_format": {
    "type": "text"
  },
  "secrets": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "template_id": "template_id",
  "timezone": "timezone",
  "tool_exec_environment_variables": [
    {
      "agent_id": "agent_id",
      "key": "key",
      "value": "value",
      "id": "agent-env-123e4567-e89b-12d3-a456-426614174000",
      "created_at": "2019-12-27T18:11:19.117Z",
      "created_by_id": "created_by_id",
      "description": "description",
      "last_updated_by_id": "last_updated_by_id",
      "updated_at": "2019-12-27T18:11:19.117Z",
      "value_enc": "value_enc"
    }
  ],
  "tool_rules": [
    {
      "children": [
        "string"
      ],
      "tool_name": "tool_name",
      "child_arg_nodes": [
        {
          "name": "name",
          "args": {
            "foo": "bar"
          }
        }
      ],
      "prompt_template": "prompt_template",
      "type": "constrain_child_tools"
    }
  ],
  "updated_at": "2019-12-27T18:11:19.117Z"
}
```

## List Folders For Agent

`client.agents.folders.list(stringagentID, FolderListParamsquery?, RequestOptionsoptions?): ArrayPage<FolderListResponse>`

**get** `/v1/agents/{agent_id}/folders`

Get the folders associated with an agent.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `query: FolderListParams`

  - `after?: string | null`

    Cursor for pagination (source ID). Returns results relative to this ID in the specified sort order. Expected format: 'source-<uuid4>'

  - `before?: string | null`

    Cursor for pagination (source ID). Returns results relative to this ID in the specified sort order. Expected format: 'source-<uuid4>'

  - `limit?: number | null`

    Maximum number of sources to return

  - `order?: "asc" | "desc"`

    Sort order for sources by creation time. 'asc' for oldest first, 'desc' for newest first

    - `"asc"`

    - `"desc"`

  - `order_by?: "created_at"`

    Field to sort by

    - `"created_at"`

### Returns

- `FolderListResponse`

  (Deprecated: Use Folder) Representation of a source, which is a collection of files and passages.

  - `id: string`

    The human-friendly ID of the Source

  - `embedding_config: EmbeddingConfig`

    The embedding configuration used by the source.

    - `embedding_dim: number`

      The dimension of the embedding.

    - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"bedrock"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"mistral"`

      - `"together"`

      - `"pinecone"`

    - `embedding_model: string`

      The model for the embedding.

    - `azure_deployment?: string | null`

      The Azure deployment for the model.

    - `azure_endpoint?: string | null`

      The Azure endpoint for the model.

    - `azure_version?: string | null`

      The Azure version for the model.

    - `batch_size?: number`

      The maximum batch size for processing embeddings.

    - `embedding_chunk_size?: number | null`

      The chunk size of the embedding.

    - `embedding_endpoint?: string | null`

      The endpoint for the model (`None` if local).

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

  - `name: string`

    The name of the source.

  - `created_at?: string | null`

    The timestamp when the source was created.

  - `created_by_id?: string | null`

    The id of the user that made this Tool.

  - `description?: string | null`

    The description of the source.

  - `instructions?: string | null`

    Instructions for how to use the source.

  - `last_updated_by_id?: string | null`

    The id of the user that made this Tool.

  - `metadata?: Record<string, unknown> | null`

    Metadata associated with the source.

  - `updated_at?: string | null`

    The timestamp when the source was last updated.

  - `vector_db_provider?: VectorDBProvider`

    The vector database provider used for this source's passages

    - `"native"`

    - `"tpuf"`

    - `"pinecone"`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

// Automatically fetches more pages as needed.
for await (const folderListResponse of client.agents.folders.list(
  'agent-123e4567-e89b-42d3-8456-426614174000',
)) {
  console.log(folderListResponse.id);
}
```

#### Response

```json
[
  {
    "id": "source-123e4567-e89b-12d3-a456-426614174000",
    "embedding_config": {
      "embedding_dim": 0,
      "embedding_endpoint_type": "openai",
      "embedding_model": "embedding_model",
      "azure_deployment": "azure_deployment",
      "azure_endpoint": "azure_endpoint",
      "azure_version": "azure_version",
      "batch_size": 0,
      "embedding_chunk_size": 0,
      "embedding_endpoint": "embedding_endpoint",
      "handle": "handle"
    },
    "name": "name",
    "created_at": "2019-12-27T18:11:19.117Z",
    "created_by_id": "created_by_id",
    "description": "description",
    "instructions": "instructions",
    "last_updated_by_id": "last_updated_by_id",
    "metadata": {
      "foo": "bar"
    },
    "updated_at": "2019-12-27T18:11:19.117Z",
    "vector_db_provider": "native"
  }
]
```

## Domain Types

### Folder List Response

- `FolderListResponse`

  (Deprecated: Use Folder) Representation of a source, which is a collection of files and passages.

  - `id: string`

    The human-friendly ID of the Source

  - `embedding_config: EmbeddingConfig`

    The embedding configuration used by the source.

    - `embedding_dim: number`

      The dimension of the embedding.

    - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"bedrock"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"mistral"`

      - `"together"`

      - `"pinecone"`

    - `embedding_model: string`

      The model for the embedding.

    - `azure_deployment?: string | null`

      The Azure deployment for the model.

    - `azure_endpoint?: string | null`

      The Azure endpoint for the model.

    - `azure_version?: string | null`

      The Azure version for the model.

    - `batch_size?: number`

      The maximum batch size for processing embeddings.

    - `embedding_chunk_size?: number | null`

      The chunk size of the embedding.

    - `embedding_endpoint?: string | null`

      The endpoint for the model (`None` if local).

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

  - `name: string`

    The name of the source.

  - `created_at?: string | null`

    The timestamp when the source was created.

  - `created_by_id?: string | null`

    The id of the user that made this Tool.

  - `description?: string | null`

    The description of the source.

  - `instructions?: string | null`

    Instructions for how to use the source.

  - `last_updated_by_id?: string | null`

    The id of the user that made this Tool.

  - `metadata?: Record<string, unknown> | null`

    Metadata associated with the source.

  - `updated_at?: string | null`

    The timestamp when the source was last updated.

  - `vector_db_provider?: VectorDBProvider`

    The vector database provider used for this source's passages

    - `"native"`

    - `"tpuf"`

    - `"pinecone"`

# Files

## Close All Files For Agent

`client.agents.files.closeAll(stringagentID, RequestOptionsoptions?): FileCloseAllResponse`

**patch** `/v1/agents/{agent_id}/files/close-all`

Closes all currently open files for a given agent.

This endpoint updates the file state for the agent so that no files are marked as open.
Typically used to reset the working memory view for the agent.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `FileCloseAllResponse = Array<string>`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const response = await client.agents.files.closeAll('agent-123e4567-e89b-42d3-8456-426614174000');

console.log(response);
```

#### Response

```json
[
  "string"
]
```

## Open File For Agent

`client.agents.files.open(stringfileID, FileOpenParamsparams, RequestOptionsoptions?): FileOpenResponse`

**patch** `/v1/agents/{agent_id}/files/{file_id}/open`

Opens a specific file for a given agent.

This endpoint marks a specific file as open in the agent's file state.
The file will be included in the agent's working memory view.
Returns a list of file names that were closed due to LRU eviction.

### Parameters

- `fileID: string`

  The ID of the file in the format 'file-<uuid4>'

- `params: FileOpenParams`

  - `agent_id: string`

    The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `FileOpenResponse = Array<string>`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const response = await client.agents.files.open('file-123e4567-e89b-42d3-8456-426614174000', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
});

console.log(response);
```

#### Response

```json
[
  "string"
]
```

## Close File For Agent

`client.agents.files.close(stringfileID, FileCloseParamsparams, RequestOptionsoptions?): FileCloseResponse`

**patch** `/v1/agents/{agent_id}/files/{file_id}/close`

Closes a specific file for a given agent.

This endpoint marks a specific file as closed in the agent's file state.
The file will be removed from the agent's working memory view.

### Parameters

- `fileID: string`

  The ID of the file in the format 'file-<uuid4>'

- `params: FileCloseParams`

  - `agent_id: string`

    The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `FileCloseResponse = unknown`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const response = await client.agents.files.close('file-123e4567-e89b-42d3-8456-426614174000', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
});

console.log(response);
```

#### Response

```json
{}
```

## List Files For Agent

`client.agents.files.list(stringagentID, FileListParamsquery?, RequestOptionsoptions?): NextFilesPage<FileListResponse>`

**get** `/v1/agents/{agent_id}/files`

Get the files attached to an agent with their open/closed status.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `query: FileListParams`

  - `after?: string | null`

    Cursor for pagination (file ID). Returns results relative to this ID in the specified sort order. Expected format: 'file-<uuid4>'

  - `before?: string | null`

    Cursor for pagination (file ID). Returns results relative to this ID in the specified sort order. Expected format: 'file-<uuid4>'

  - `cursor?: string | null`

    Pagination cursor from previous response (deprecated, use before/after)

  - `is_open?: boolean | null`

    Filter by open status (true for open files, false for closed files)

  - `limit?: number | null`

    Maximum number of files to return

  - `order?: "asc" | "desc"`

    Sort order for files by creation time. 'asc' for oldest first, 'desc' for newest first

    - `"asc"`

    - `"desc"`

  - `order_by?: "created_at"`

    Field to sort by

    - `"created_at"`

### Returns

- `FileListResponse`

  Response model for agent file attachments showing file status in agent context

  - `id: string`

    Unique identifier of the file-agent relationship

  - `file_id: string`

    Unique identifier of the file

  - `file_name: string`

    Name of the file

  - `folder_id: string`

    Unique identifier of the folder/source

  - `folder_name: string`

    Name of the folder/source

  - `is_open: boolean`

    Whether the file is currently open in the agent's context

  - `end_line?: number | null`

    Ending line number if file was opened with line range

  - `last_accessed_at?: string | null`

    Timestamp of last access by the agent

  - `start_line?: number | null`

    Starting line number if file was opened with line range

  - `visible_content?: string | null`

    Portion of the file visible to the agent if open

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

// Automatically fetches more pages as needed.
for await (const fileListResponse of client.agents.files.list(
  'agent-123e4567-e89b-42d3-8456-426614174000',
)) {
  console.log(fileListResponse.id);
}
```

#### Response

```json
{
  "files": [
    {
      "id": "id",
      "file_id": "file_id",
      "file_name": "file_name",
      "folder_id": "folder_id",
      "folder_name": "folder_name",
      "is_open": true,
      "end_line": 0,
      "last_accessed_at": "2019-12-27T18:11:19.117Z",
      "start_line": 0,
      "visible_content": "visible_content"
    }
  ],
  "has_more": true,
  "next_cursor": "next_cursor"
}
```

## Domain Types

### File Close All Response

- `FileCloseAllResponse = Array<string>`

### File Open Response

- `FileOpenResponse = Array<string>`

### File Close Response

- `FileCloseResponse = unknown`

### File List Response

- `FileListResponse`

  Response model for agent file attachments showing file status in agent context

  - `id: string`

    Unique identifier of the file-agent relationship

  - `file_id: string`

    Unique identifier of the file

  - `file_name: string`

    Name of the file

  - `folder_id: string`

    Unique identifier of the folder/source

  - `folder_name: string`

    Name of the folder/source

  - `is_open: boolean`

    Whether the file is currently open in the agent's context

  - `end_line?: number | null`

    Ending line number if file was opened with line range

  - `last_accessed_at?: string | null`

    Timestamp of last access by the agent

  - `start_line?: number | null`

    Starting line number if file was opened with line range

  - `visible_content?: string | null`

    Portion of the file visible to the agent if open

# Archives

## Attach Archive To Agent

`client.agents.archives.attach(stringarchiveID, ArchiveAttachParamsparams, RequestOptionsoptions?): ArchiveAttachResponse`

**patch** `/v1/agents/{agent_id}/archives/attach/{archive_id}`

Attach an archive to an agent.

### Parameters

- `archiveID: string`

- `params: ArchiveAttachParams`

  - `agent_id: string`

    The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `ArchiveAttachResponse = unknown`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const response = await client.agents.archives.attach('archive_id', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
});

console.log(response);
```

#### Response

```json
{}
```

## Detach Archive From Agent

`client.agents.archives.detach(stringarchiveID, ArchiveDetachParamsparams, RequestOptionsoptions?): ArchiveDetachResponse`

**patch** `/v1/agents/{agent_id}/archives/detach/{archive_id}`

Detach an archive from an agent.

### Parameters

- `archiveID: string`

- `params: ArchiveDetachParams`

  - `agent_id: string`

    The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `ArchiveDetachResponse = unknown`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const response = await client.agents.archives.detach('archive_id', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
});

console.log(response);
```

#### Response

```json
{}
```

## Domain Types

### Archive Attach Response

- `ArchiveAttachResponse = unknown`

### Archive Detach Response

- `ArchiveDetachResponse = unknown`

# Passages

## List Passages

`client.agents.passages.list(stringagentID, PassageListParamsquery?, RequestOptionsoptions?): PassageListResponse`

**get** `/v1/agents/{agent_id}/archival-memory`

Retrieve the memories in an agent's archival memory store (paginated query).

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `query: PassageListParams`

  - `after?: string | null`

    Unique ID of the memory to start the query range at.

  - `ascending?: boolean | null`

    Whether to sort passages oldest to newest (True, default) or newest to oldest (False)

  - `before?: string | null`

    Unique ID of the memory to end the query range at.

  - `limit?: number | null`

    How many results to include in the response.

  - `search?: string | null`

    Search passages by text

### Returns

- `PassageListResponse = Array<Passage>`

  - `embedding: Array<number> | null`

    The embedding of the passage.

  - `embedding_config: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

    - `embedding_dim: number`

      The dimension of the embedding.

    - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"bedrock"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"mistral"`

      - `"together"`

      - `"pinecone"`

    - `embedding_model: string`

      The model for the embedding.

    - `azure_deployment?: string | null`

      The Azure deployment for the model.

    - `azure_endpoint?: string | null`

      The Azure endpoint for the model.

    - `azure_version?: string | null`

      The Azure version for the model.

    - `batch_size?: number`

      The maximum batch size for processing embeddings.

    - `embedding_chunk_size?: number | null`

      The chunk size of the embedding.

    - `embedding_endpoint?: string | null`

      The endpoint for the model (`None` if local).

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

  - `text: string`

    The text of the passage.

  - `id?: string`

    The human-friendly ID of the Passage

  - `archive_id?: string | null`

    The unique identifier of the archive containing this passage.

  - `created_at?: string | null`

    The creation date of the passage.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `file_id?: string | null`

    The unique identifier of the file associated with the passage.

  - `file_name?: string | null`

    The name of the file (only for source passages).

  - `is_deleted?: boolean`

    Whether this passage is deleted or not.

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the passage.

  - `source_id?: string | null`

    Deprecated: Use `folder_id` field instead. The data source of the passage.

  - `tags?: Array<string> | null`

    Tags associated with this passage.

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const passages = await client.agents.passages.list('agent-123e4567-e89b-42d3-8456-426614174000');

console.log(passages);
```

#### Response

```json
[
  {
    "embedding": [
      0
    ],
    "embedding_config": {
      "embedding_dim": 0,
      "embedding_endpoint_type": "openai",
      "embedding_model": "embedding_model",
      "azure_deployment": "azure_deployment",
      "azure_endpoint": "azure_endpoint",
      "azure_version": "azure_version",
      "batch_size": 0,
      "embedding_chunk_size": 0,
      "embedding_endpoint": "embedding_endpoint",
      "handle": "handle"
    },
    "text": "text",
    "id": "passage-123e4567-e89b-12d3-a456-426614174000",
    "archive_id": "archive_id",
    "created_at": "2019-12-27T18:11:19.117Z",
    "created_by_id": "created_by_id",
    "file_id": "file_id",
    "file_name": "file_name",
    "is_deleted": true,
    "last_updated_by_id": "last_updated_by_id",
    "metadata": {
      "foo": "bar"
    },
    "source_id": "source_id",
    "tags": [
      "string"
    ],
    "updated_at": "2019-12-27T18:11:19.117Z"
  }
]
```

## Create Passage

`client.agents.passages.create(stringagentID, PassageCreateParamsbody, RequestOptionsoptions?): PassageCreateResponse`

**post** `/v1/agents/{agent_id}/archival-memory`

Insert a memory into an agent's archival memory store.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `body: PassageCreateParams`

  - `text: string`

    Text to write to archival memory.

  - `created_at?: string | null`

    Optional timestamp for the memory (defaults to current UTC time).

  - `tags?: Array<string> | null`

    Optional list of tags to attach to the memory.

### Returns

- `PassageCreateResponse = Array<Passage>`

  - `embedding: Array<number> | null`

    The embedding of the passage.

  - `embedding_config: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

    - `embedding_dim: number`

      The dimension of the embedding.

    - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"bedrock"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"mistral"`

      - `"together"`

      - `"pinecone"`

    - `embedding_model: string`

      The model for the embedding.

    - `azure_deployment?: string | null`

      The Azure deployment for the model.

    - `azure_endpoint?: string | null`

      The Azure endpoint for the model.

    - `azure_version?: string | null`

      The Azure version for the model.

    - `batch_size?: number`

      The maximum batch size for processing embeddings.

    - `embedding_chunk_size?: number | null`

      The chunk size of the embedding.

    - `embedding_endpoint?: string | null`

      The endpoint for the model (`None` if local).

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

  - `text: string`

    The text of the passage.

  - `id?: string`

    The human-friendly ID of the Passage

  - `archive_id?: string | null`

    The unique identifier of the archive containing this passage.

  - `created_at?: string | null`

    The creation date of the passage.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `file_id?: string | null`

    The unique identifier of the file associated with the passage.

  - `file_name?: string | null`

    The name of the file (only for source passages).

  - `is_deleted?: boolean`

    Whether this passage is deleted or not.

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the passage.

  - `source_id?: string | null`

    Deprecated: Use `folder_id` field instead. The data source of the passage.

  - `tags?: Array<string> | null`

    Tags associated with this passage.

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const passages = await client.agents.passages.create('agent-123e4567-e89b-42d3-8456-426614174000', {
  text: 'text',
});

console.log(passages);
```

#### Response

```json
[
  {
    "embedding": [
      0
    ],
    "embedding_config": {
      "embedding_dim": 0,
      "embedding_endpoint_type": "openai",
      "embedding_model": "embedding_model",
      "azure_deployment": "azure_deployment",
      "azure_endpoint": "azure_endpoint",
      "azure_version": "azure_version",
      "batch_size": 0,
      "embedding_chunk_size": 0,
      "embedding_endpoint": "embedding_endpoint",
      "handle": "handle"
    },
    "text": "text",
    "id": "passage-123e4567-e89b-12d3-a456-426614174000",
    "archive_id": "archive_id",
    "created_at": "2019-12-27T18:11:19.117Z",
    "created_by_id": "created_by_id",
    "file_id": "file_id",
    "file_name": "file_name",
    "is_deleted": true,
    "last_updated_by_id": "last_updated_by_id",
    "metadata": {
      "foo": "bar"
    },
    "source_id": "source_id",
    "tags": [
      "string"
    ],
    "updated_at": "2019-12-27T18:11:19.117Z"
  }
]
```

## Delete Passage

`client.agents.passages.delete(stringmemoryID, PassageDeleteParamsparams, RequestOptionsoptions?): PassageDeleteResponse`

**delete** `/v1/agents/{agent_id}/archival-memory/{memory_id}`

Delete a memory from an agent's archival memory store.

### Parameters

- `memoryID: string`

- `params: PassageDeleteParams`

  - `agent_id: string`

    The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `PassageDeleteResponse = unknown`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const passage = await client.agents.passages.delete('memory_id', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
});

console.log(passage);
```

#### Response

```json
{}
```

## Search Archival Memory

`client.agents.passages.search(stringagentID, PassageSearchParamsquery, RequestOptionsoptions?): PassageSearchResponse`

**get** `/v1/agents/{agent_id}/archival-memory/search`

Search archival memory using semantic (embedding-based) search with optional temporal filtering.

This endpoint allows manual triggering of archival memory searches, enabling users to query
an agent's archival memory store directly via the API. The search uses the same functionality
as the agent's archival_memory_search tool but is accessible for external API usage.

### Parameters

- `agentID: string`

  The ID of the agent in the format 'agent-<uuid4>'

- `query: PassageSearchParams`

  - `query: string`

    String to search for using semantic similarity

  - `end_datetime?: string | null`

    Filter results to passages created before this datetime

  - `start_datetime?: string | null`

    Filter results to passages created after this datetime

  - `tag_match_mode?: "any" | "all"`

    How to match tags - 'any' to match passages with any of the tags, 'all' to match only passages with all tags

    - `"any"`

    - `"all"`

  - `tags?: Array<string> | null`

    Optional list of tags to filter search results

  - `top_k?: number | null`

    Maximum number of results to return. Uses system default if not specified

### Returns

- `PassageSearchResponse`

  - `count: number`

    Total number of results returned

  - `results: Array<Result>`

    List of search results matching the query

    - `id: string`

      Unique identifier of the archival memory passage

    - `content: string`

      Text content of the archival memory passage

    - `timestamp: string`

      Timestamp of when the memory was created, formatted in agent's timezone

    - `tags?: Array<string>`

      List of tags associated with this memory

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const response = await client.agents.passages.search('agent-123e4567-e89b-42d3-8456-426614174000', {
  query: 'query',
});

console.log(response.count);
```

#### Response

```json
{
  "count": 0,
  "results": [
    {
      "id": "id",
      "content": "content",
      "timestamp": "timestamp",
      "tags": [
        "string"
      ]
    }
  ]
}
```

## Domain Types

### Passage List Response

- `PassageListResponse = Array<Passage>`

  - `embedding: Array<number> | null`

    The embedding of the passage.

  - `embedding_config: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

    - `embedding_dim: number`

      The dimension of the embedding.

    - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"bedrock"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"mistral"`

      - `"together"`

      - `"pinecone"`

    - `embedding_model: string`

      The model for the embedding.

    - `azure_deployment?: string | null`

      The Azure deployment for the model.

    - `azure_endpoint?: string | null`

      The Azure endpoint for the model.

    - `azure_version?: string | null`

      The Azure version for the model.

    - `batch_size?: number`

      The maximum batch size for processing embeddings.

    - `embedding_chunk_size?: number | null`

      The chunk size of the embedding.

    - `embedding_endpoint?: string | null`

      The endpoint for the model (`None` if local).

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

  - `text: string`

    The text of the passage.

  - `id?: string`

    The human-friendly ID of the Passage

  - `archive_id?: string | null`

    The unique identifier of the archive containing this passage.

  - `created_at?: string | null`

    The creation date of the passage.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `file_id?: string | null`

    The unique identifier of the file associated with the passage.

  - `file_name?: string | null`

    The name of the file (only for source passages).

  - `is_deleted?: boolean`

    Whether this passage is deleted or not.

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the passage.

  - `source_id?: string | null`

    Deprecated: Use `folder_id` field instead. The data source of the passage.

  - `tags?: Array<string> | null`

    Tags associated with this passage.

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Passage Create Response

- `PassageCreateResponse = Array<Passage>`

  - `embedding: Array<number> | null`

    The embedding of the passage.

  - `embedding_config: EmbeddingConfig | null`

    Configuration for embedding model connection and processing parameters.

    - `embedding_dim: number`

      The dimension of the embedding.

    - `embedding_endpoint_type: "openai" | "anthropic" | "bedrock" | 16 more`

      The endpoint type for the model.

      - `"openai"`

      - `"anthropic"`

      - `"bedrock"`

      - `"google_ai"`

      - `"google_vertex"`

      - `"azure"`

      - `"groq"`

      - `"ollama"`

      - `"webui"`

      - `"webui-legacy"`

      - `"lmstudio"`

      - `"lmstudio-legacy"`

      - `"llamacpp"`

      - `"koboldcpp"`

      - `"vllm"`

      - `"hugging-face"`

      - `"mistral"`

      - `"together"`

      - `"pinecone"`

    - `embedding_model: string`

      The model for the embedding.

    - `azure_deployment?: string | null`

      The Azure deployment for the model.

    - `azure_endpoint?: string | null`

      The Azure endpoint for the model.

    - `azure_version?: string | null`

      The Azure version for the model.

    - `batch_size?: number`

      The maximum batch size for processing embeddings.

    - `embedding_chunk_size?: number | null`

      The chunk size of the embedding.

    - `embedding_endpoint?: string | null`

      The endpoint for the model (`None` if local).

    - `handle?: string | null`

      The handle for this config, in the format provider/model-name.

  - `text: string`

    The text of the passage.

  - `id?: string`

    The human-friendly ID of the Passage

  - `archive_id?: string | null`

    The unique identifier of the archive containing this passage.

  - `created_at?: string | null`

    The creation date of the passage.

  - `created_by_id?: string | null`

    The id of the user that made this object.

  - `file_id?: string | null`

    The unique identifier of the file associated with the passage.

  - `file_name?: string | null`

    The name of the file (only for source passages).

  - `is_deleted?: boolean`

    Whether this passage is deleted or not.

  - `last_updated_by_id?: string | null`

    The id of the user that made this object.

  - `metadata?: Record<string, unknown> | null`

    The metadata of the passage.

  - `source_id?: string | null`

    Deprecated: Use `folder_id` field instead. The data source of the passage.

  - `tags?: Array<string> | null`

    Tags associated with this passage.

  - `updated_at?: string | null`

    The timestamp when the object was last updated.

### Passage Delete Response

- `PassageDeleteResponse = unknown`

### Passage Search Response

- `PassageSearchResponse`

  - `count: number`

    Total number of results returned

  - `results: Array<Result>`

    List of search results matching the query

    - `id: string`

      Unique identifier of the archival memory passage

    - `content: string`

      Text content of the archival memory passage

    - `timestamp: string`

      Timestamp of when the memory was created, formatted in agent's timezone

    - `tags?: Array<string>`

      List of tags associated with this memory

# Identities

## Attach Identity To Agent

`client.agents.identities.attach(stringidentityID, IdentityAttachParamsparams, RequestOptionsoptions?): IdentityAttachResponse`

**patch** `/v1/agents/{agent_id}/identities/attach/{identity_id}`

Attach an identity to an agent.

### Parameters

- `identityID: string`

- `params: IdentityAttachParams`

  - `agent_id: string`

    The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `IdentityAttachResponse = unknown`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const response = await client.agents.identities.attach('identity_id', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
});

console.log(response);
```

#### Response

```json
{}
```

## Detach Identity From Agent

`client.agents.identities.detach(stringidentityID, IdentityDetachParamsparams, RequestOptionsoptions?): IdentityDetachResponse`

**patch** `/v1/agents/{agent_id}/identities/detach/{identity_id}`

Detach an identity from an agent.

### Parameters

- `identityID: string`

- `params: IdentityDetachParams`

  - `agent_id: string`

    The ID of the agent in the format 'agent-<uuid4>'

### Returns

- `IdentityDetachResponse = unknown`

### Example

```typescript
import Letta from '@letta-ai/letta-client';

const client = new Letta({
  apiKey: process.env['LETTA_API_KEY'], // This is the default and can be omitted
});

const response = await client.agents.identities.detach('identity_id', {
  agent_id: 'agent-123e4567-e89b-42d3-8456-426614174000',
});

console.log(response);
```

#### Response

```json
{}
```

## Domain Types

### Identity Attach Response

- `IdentityAttachResponse = unknown`

### Identity Detach Response

- `IdentityDetachResponse = unknown`
