Streaming Agent
Stream an agent's response from a local LM Studio model and print it in the terminal.
Code
from typing import Iterator # noqa
from agno.agent import Agent, RunOutputEvent # noqa
from agno.models.lmstudio import LMStudio
agent = Agent(model=LMStudio(id="qwen2.5-7b-instruct-1m"), markdown=True)
# Get the response in a variable
# run_response: Iterator[RunOutputEvent] = agent.run("Share a 2 sentence horror story", stream=True)
# for chunk in run_response:
# print(chunk.content)
# Print the response in the terminal
agent.print_response("Share a 2 sentence horror story", stream=True)Usage
Set up your virtual environment
uv venv --python 3.12
source .venv/bin/activateStart the LM Studio API server
Install LM Studio, download and load a model, then open Developer and start the API server on port 1234. Keep the server running while you run Python in a terminal.
Check the model list:
curl http://127.0.0.1:1234/v1/modelsSet LMStudio(id=...) in the example to the exact model id returned by your server. If you change the server port, set the matching base_url on LMStudio.
Install dependencies
uv pip install -U openai agnoRun Agent
Save the code above as basic.py, then run:
python basic.py