Basic Agent

Run a Llama-powered Agno agent synchronously, with streaming, and asynchronously.

Code

from agno.agent import Agent, RunOutput  # noqa
from agno.models.meta import Llama
import asyncio

agent = Agent(
    model=Llama(id="Llama-4-Maverick-17B-128E-Instruct-FP8"),
    markdown=True,
)

# Get the response in a variable
# run: RunOutput = agent.run("Share a 2 sentence horror story")
# print(run.content)

if __name__ == "__main__":
    # --- Sync ---
    agent.print_response("Share a 2 sentence horror story")

    # --- Sync + Streaming ---
    agent.print_response("Share a 2 sentence horror story", stream=True)

    # --- Async ---
    asyncio.run(agent.aprint_response("Share a 2 sentence horror story"))

    # --- Async + Streaming ---
    asyncio.run(agent.aprint_response("Share a 2 sentence horror story", stream=True))

These examples require access to the Meta Llama API, an API key, and the chosen model enabled for your account. Check the model catalog while signed in, and replace the example ID if your account offers a different compatible model. Installing llama-api-client alone does not grant API access.

Usage

Set up your virtual environment

uv venv --python 3.12
source .venv/bin/activate

Set your LLAMA API key

export LLAMA_API_KEY=YOUR_API_KEY

Install dependencies

uv pip install llama-api-client agno

Run Agent

Save the code above as basic.py, then run:

python basic.py