Code
image_agent_bytes.py
Usage
1
Set up your virtual environment
2
Set your API key
3
Install dependencies
4
Run Agent
Save the code above as
image_agent_bytes.py, then run:Documentation Index
Fetch the complete documentation index at: /llms.txt
Use this file to discover all available pages before exploring further.
Download an image, read it as raw bytes, and stream a response from an OpenAIResponses gpt-4o agent with WebSearchTools that describes it and finds related news.
from pathlib import Path
from agno.agent import Agent
from agno.media import Image
from agno.models.openai import OpenAIResponses
from agno.tools.websearch import WebSearchTools
from agno.utils.media import download_image
agent = Agent(
model=OpenAIResponses(id="gpt-4o"),
tools=[WebSearchTools()],
markdown=True,
)
image_path = Path(__file__).parent.joinpath("sample.jpg")
download_image(
url="https://upload.wikimedia.org/wikipedia/commons/0/0c/GoldenGateBridge-001.jpg",
output_path=str(image_path),
)
# Read the image file content as bytes
image_bytes = image_path.read_bytes()
agent.print_response(
"Tell me about this image and give me the latest news about it.",
images=[
Image(content=image_bytes),
],
stream=True,
)
Set up your virtual environment
uv venv --python 3.12
source .venv/bin/activate
uv venv --python 3.12
.venv\Scripts\activate
Set your API key
export OPENAI_API_KEY=xxx
Install dependencies
uv pip install -U openai ddgs agno
Run Agent
image_agent_bytes.py, then run:python image_agent_bytes.py
Was this page helpful?