Moonshot Thinking Mode

Toggle thinking for Kimi K2.6; Kimi K3 always reasons.

This example toggles thinking for Kimi K2.6. Kimi K3 always reasons and does not support disabling thinking. For K2.6, use_thinking=True enables thinking and use_thinking=False disables it.

thinking_mode.py
"""
Moonshot Thinking Mode
======================

Kimi models reason by default, so you get reasoning_content out of the box. Use the
`use_thinking` flag to control it: `use_thinking=True` forces it on, `use_thinking=False`
turns it off for a faster, cheaper response.

Note that thinking cannot be turned off on every model - Kimi K3 always reasons.
"""

from agno.agent import Agent
from agno.models.moonshot import MoonShot

# ---------------------------------------------------------------------------
# Thinking enabled (default) - returns reasoning_content
# ---------------------------------------------------------------------------

thinking_agent = Agent(model=MoonShot(id="kimi-k2.6"), markdown=True)

# ---------------------------------------------------------------------------
# Thinking disabled - faster, no reasoning_content
# ---------------------------------------------------------------------------

non_thinking_agent = Agent(
    model=MoonShot(id="kimi-k2.6", use_thinking=False),
    markdown=True,
)

# ---------------------------------------------------------------------------
# Run Agents
# ---------------------------------------------------------------------------
if __name__ == "__main__":
    thinking_agent.print_response("Why is the sky blue?", stream=True)

    non_thinking_agent.print_response("Why is the sky blue?", stream=True)

Run the Example

Set up your virtual environment

uv venv --python 3.12
source .venv/bin/activate

Install dependencies

uv pip install -U agno openai

Export your Moonshot API key

export MOONSHOT_API_KEY="your_moonshot_api_key_here"

Run the example

Save the code above as thinking_mode.py, then run:

python thinking_mode.py

Full source: cookbook/90_models/moonshot/thinking_mode.py