Basic Agent
Run a Llama-powered Agno agent synchronously, with streaming, and asynchronously.
Code
from agno.agent import Agent, RunOutput # noqa
from agno.models.meta import Llama
import asyncio
agent = Agent(
model=Llama(id="Llama-4-Maverick-17B-128E-Instruct-FP8"),
markdown=True,
)
# Get the response in a variable
# run: RunOutput = agent.run("Share a 2 sentence horror story")
# print(run.content)
if __name__ == "__main__":
# --- Sync ---
agent.print_response("Share a 2 sentence horror story")
# --- Sync + Streaming ---
agent.print_response("Share a 2 sentence horror story", stream=True)
# --- Async ---
asyncio.run(agent.aprint_response("Share a 2 sentence horror story"))
# --- Async + Streaming ---
asyncio.run(agent.aprint_response("Share a 2 sentence horror story", stream=True))These examples require access to the Meta Llama API, an API key, and the chosen model enabled for your account. Check the model catalog while signed in, and replace the example ID if your account offers a different compatible model. Installing llama-api-client alone does not grant API access.
Usage
Set up your virtual environment
uv venv --python 3.12
source .venv/bin/activateSet your LLAMA API key
export LLAMA_API_KEY=YOUR_API_KEYInstall dependencies
uv pip install llama-api-client agnoRun Agent
Save the code above as basic.py, then run:
python basic.py