Rhea is a local assistant built around a llama.cpp-compatible chat server.
Project layout:
main.pystarts the HTTP server and exists as the stable backend entrypoint.utils/chat_server.pyowns the FastAPI/chatendpoint, LLM streaming, tool execution loop, session history, memory context, and keepalive.terminal_client.pyowns the terminal chat UI.utils/settings.pyloads.envvalues and shared model-selection config.utils/tool_call_strategy.pybuilds the system prompt and parsesTHINK/CALLtool requests.tools/contains callable assistant tools and their argument schemas.
Run:
uv run rhea-server
uv run rhea-chatConfiguration lives in .env. Copy .env.example when setting up a new checkout.