Documentation IndexFetch the complete documentation index at: /llms.txtUse this file to discover all available pages before exploring further.
Fetch the complete documentation index at: /llms.txt
Use this file to discover all available pages before exploring further.
Run a Cloudflare Workers AI model in sync, async, and streaming modes.
import asyncio from agno.agent import Agent from agno.models.cloudflare import Cloudflare agent = Agent( model=Cloudflare("@cf/meta/llama-3.3-70b-instruct-fp8-fast"), markdown=True, ) if __name__ == "__main__": # --- Sync --- agent.print_response("Share a 2 sentence horror story") # --- Sync + Streaming --- agent.print_response("Share a 2 sentence horror story", stream=True) # --- Async --- asyncio.run(agent.aprint_response("Share a 2 sentence horror story")) # --- Async + Streaming --- asyncio.run(agent.aprint_response("Share a 2 sentence horror story", stream=True))
Set up your virtual environment
uv venv --python 3.12 source .venv/bin/activate
uv venv --python 3.12 .venv\Scripts\activate
Set your environment variables
export CLOUDFLARE_API_TOKEN=xxx export CLOUDFLARE_ACCOUNT_ID=xxx export CLOUDFLARE_AI_GATEWAY_ID=xxx # optional, defaults to "default"
Install dependencies
uv pip install -U openai agno
Run Agent
basic.py
python basic.py
Was this page helpful?