You signed in with another tab or window. Reload to refresh your session.You signed out in another tab or window. Reload to refresh your session.You switched accounts on another tab or window. Reload to refresh your session.Dismiss alert
An OpenAI-compatible gateway runs on your own machine at 127.0.0.1 — no account, no API key, no server in between. Your OpenAI-compatible clients point straight at it.
The auto Model
One model that rotates across healthy free models for you, preferring ones that can call tools — so a dead or rate-limited model never ends your chat.
Streaming Chat
Token-by-token replies with visible reasoning (collapsible while it thinks), expandable tool calls, per-message token usage, stop-and-keep-partial.
Web Search
Built-in search tool with automatic fallback to keyless sources, so a capped provider never kills an answer.
MCP Servers
Add your own MCP servers; enable, disable and test their tools from Settings.
Multi-Provider
Route chat through any OpenAI-compatible endpoint with an optional key stored only on your device.
Share
Publish your gateway through a public tunnel with an optional API key, and hand out the full model catalog to other clients.
History & Export
Conversations stay on your device; open, search, copy and export them as Markdown.
Screenshots
Chat
Web search with tool cards and an expandable reasoning block (light).
The same thread in dark, with the gateway's base URL in the answer.
MCP tools run mid-conversation: library lookups, reasoning and code in one thread.
Every model with its endpoint type and the gateway's own rate hint.
Server
Model catalog with per-row test actions, rate hints, retest sweeps and the activity log.
Share the gateway on a public URL any OpenAI client can use, with an optional key.
Settings
System prompt, gateway behaviour and generation controls.
Connected MCP servers with per-tool switches, Test and Disable.
Features
Streaming Chat — token-by-token replies with markdown tables and code highlighting, visible reasoning blocks, expandable tool calls, and per-message token usage.
Regenerate & Edit — long-press or right-click any message to copy it, re-run the last reply, or edit-and-resend the last question.
Server Console — gateway status, start/stop, live logs with level filters, refresh model catalog, and the full catalog with per-model rate limits and endpoint types.
Settings — theme, generation parameters, gateway port and autostart, providers, MCP servers, search toggle, About.
Onboarding — terms gate and consent flow before first use.
Updates — silent version check with release notes, plus a manual check from Settings.
Desktop — closing the window keeps the gateway running in the background; Quit is always one click away.
uv sync # install venv (dev group included)
uv run ruff check src tests # lint (same paths CI lints)
uv run ruff format --check src tests # format gate (same as CI)
uv run pytest -q # test suite
uv run flet run -v # run desktop app
uv run flet clean # delete build/ (Flutter shell + staged python)
uv run flet build apk --split-per-abi -v # Android (matches CI)
Privacy & Security
On your device: the gateway runs locally, binds only to 127.0.0.1, and conversation history and settings stay on this machine.
No account: no registration, no login, and no user API key is required for the built-in gateway.
Prompts go to third parties: model requests are sent to free providers that may log them and use them for training — we have no control over that.
Keys stay local: any provider key you add is stored on this device only and never logged.
Ad-supported on mobile: mobile builds show Google AdMob ads, with consent managed through Google UMP.
Legal Disclaimer
Free models are provided by third parties under their own terms and rate limits.
Your messages are sent to those third parties: they may log your prompts and use
them for training. Model availability changes without notice. Ads are served by
Google AdMob on mobile builds only.