Production-ready OpenAI-compatible gateway for multiple free-tier LLM providers.
Container migration notice:
ghcr.io/tuxevil/pi-antigravity-rotatoris deprecated. Useghcr.io/tuxevil/tuxevil-rotator:latestfor new deployments. The legacy image remains available as a frozen compatibility reference; see the migration guide.
Multi-account load balancing, per-model quota routing, account health scoring, access control with Virtual Keys, and cost auditing — via a single local endpoint that any agent can use. Even with a single account.
Originally built as a multi-account rotator for Google Antigravity. It now generalizes that rotation layer across free-tier LLM providers with per-account credentials: Google Antigravity, Ollama Cloud, OpenAI Codex, and OpenCode Zen, designed so new providers slot in without modifying core engine logic.
⚠️ WARNING: Using this proxy may put connected accounts at risk of Terms of Service enforcement, including restriction, suspension, or permanent bans. Use at your own risk.
⚠️ Terms of Service Warning — Read Before Installing
[!CAUTION] This is an unofficial tool. Routing traffic through this proxy may violate a provider's Terms of Service or trigger automated abuse or policy enforcement systems.
By using this proxy, you acknowledge:
- Your account may be restricted, suspended, shadow-banned, or permanently banned
- Multi-account rotation and proxying can increase account risk compared to normal interactive usage
- You assume all responsibility for the accounts and traffic routed through this tool
Recommendation: Do not use your primary account on any provider. Prefer disposable or lower-risk accounts, and keep account exposure conservative.
Migrating from
pi-antigravity-rotator? Install the newtuxevil-rotatorpackage. It includes the new CLI plus a legacy shim, automatically migrates legacy config files, and keeps old environment variables and paths working as fallbacks. See Migration from pi-antigravity-rotator.
- v4.0.0: Redesigned WebUI: A complete dashboard refresh with live routing health, account management, request history, usage analytics, shareable views, and a responsive Preact interface. (Release notes; PR #38 by @CyR1en)
- v3.11.0: automatic agent setup and OpenCode model discovery: Configure detected or selected agents locally or over SSH with
install-agent; OpenCode can populate its model catalog from the rotator's live/v1/modelsendpoint. OpenCode Zen requests also forward the real session context and use a consistent User-Agent. (Release notes) - v3.10.0: Claude tools, account-flag diagnostics, and audio reliability: Fixed Claude Opus 4.6 Thinking tool-schema errors; account verification reasons now persist and flow into anonymous telemetry with a protected recent-incidents dashboard; audio transcription and WebSocket streaming use active Antigravity accounts with Language Server fallback. (Release notes; audio hardening in PR #36 by @javargasm)
- Rotator-Backed Audio Transcription and Live-Streaming Hardening: Batch transcription and live WebSocket sessions use active Antigravity accounts with Language Server fallback after rotator errors and timeouts. Virtual-key scopes follow the executed model; live streams preserve segment order, echo Ping payloads, and avoid fallback after client cancellation. (PR #36 by @javargasm)
- OpenAI Codex GPT-6 catalog: Added GPT-6 Astra, Sol, and Luna to the Codex options while keeping GPT-5.6 Sol, Terra, and Luna available. (v3.9.0)
- TypeSafe Jev Auto-Routing Improvements & Shadow Mode: Decoupled dynamic effort and model judgments for
model: "auto", corrected target selection to strictly honor Jev's judgment, bounded decision payloads with strict non-text redaction, and introduced an opt-in shadow mode with observability headers (X-Rotator-Model-Selection: shadow). See TypeSafe Jev model selection and v3.9.0. - Single-Account Transport Retry Recovery: Transient upstream network failures now retry the current account with cancellable jittered exponential backoff when no alternate account is available, covering native proxy, compatibility, and Code Assist routes. (PR #35 by @toRolex)
- Cached Token Usage Propagation: Google Antigravity cached prompt usage is preserved as
cachedContentTokenCountfor native Gemini responses and as cache-read details across OpenAI, Anthropic, and OpenCode Zen Responses compatibility routes. (PR #34 by @toRolex; follow-up compatibility fixes included in v3.7.2) - OpenCode Zen catalog sync: Added
big-pickleand removed the retireddeepseek-v4-flash-freeandhy3-freeentries;muse-spark-1.3-contributor-freeis now routed through the Responses endpoint with full Chat, Responses, and Anthropic surface compatibility. (v3.7.1) - Tier-first routing under concurrency: Concurrent least-loaded account selection now enforces tier priority, and persisted lower-tier assignments are re-evaluated while idle. (v3.7.1)
- Codex kickstart lifecycle: Unstarted Codex quota timers are kicked off on startup and stale kickstart responses are drained to avoid blocking subsequent requests. (v3.7.1)
- Gemini 3.8 Flash Support: Native Antigravity model IDs
gemini-3.8-flash-high,gemini-3.8-flash-medium, andgemini-3.8-flash-low, with 1M-token context, 65,536-token output, pricing, quota routing, dashboard visualization, and telemetry support. Gemini 3.5 Flash has been removed from the user-facing catalog to match Antigravity 2.11. (PR #28 by @CyR1en) - Dynamic Antigravity Discovery: Successful quota polls reconcile live model IDs and metadata per full account identity. Newly advertised models appear in
/v1/modelsand route only to accounts that advertised them; removed models disappear without falling back to the generic Gemini pool. Legacygemini-3.5-*IDs stay hidden, and configuredmodelSpecssubstring overrides retain precedence. (PR #29 by @javargasm) - Dynamic Catalog Hardening: Quota snapshots reject malformed data while preserving the last known-good state, stale account responses are ignored, and dynamic model ownership and quota-safety state survive restarts without routing a model through an account that did not advertise it. (PR #30 by @javargasm)
- Partial Model Specification Overrides:
modelSpecsentries can override selected output, thinking, and context fields while inheriting the remaining effective runtime, static, or family metadata. Overrides are preserved through configuration normalization and persistence. (PR #31 by @javargasm) - Effort-Based Model Routing: Optional
effortRoutingaliases dispatch requests to concrete model variants fromreasoning_effort, while preserving the alias for clients and tracking quotas against the resolved target. See Effort-Based Model Routing. (PR #32 by @CyR1en) - Optional TypeSafe Jev Routing:
model: "auto"can semantically select a currently routable model across active providers for OpenAI, Anthropic, and Gemini compatibility routes, with encrypted/login-cliconfiguration and deterministic fail-open fallback. See TypeSafe Jev model selection and v3.8.0. - OpenAI-Compatible Audio Transcription and Live Streaming: Added batch transcription and bidirectional WebSocket audio streaming through the local Antigravity observer model, with model aliases, transcript telemetry, CORS/preflight support, low-latency headers, and lifecycle/security coverage. (PR #33 by @javargasm)
- Partial Quota Snapshot Recovery: Reset-only Google quota entries are treated as exhausted instead of being discarded, and omitted pools remain available from the last known-good snapshot.
- Independent Pool Cooldowns: Shared Claude and Gemini pools reconcile cooldowns independently, so one exhausted pool does not block its sibling and fresh provider reset deadlines are honored. (v3.6.1)
- Gemini 3.7 Flash Support: Full support for Google's Gemini 3.7 Flash models (
gemini-3.7-flash,gemini-3.7-flash-tiered,gemini-3.7-flash-high,gemini-3.7-flash-medium,gemini-3.7-flash-low) with reasoning effort / thinking level translation, pricing, dashboard visualization, and telemetry. - OpenCode Zen Provider: Added
opencode-zenprovider support (big-pickle,mimo-v2.5-free,ling-3.0-flash-fin-free,nemotron-3-ultra-free,nemotron-3.5-lightning-free,muse-spark-1.3-contributor-free) with static API key validation, onboarding UI tabs in/loginand/login-cli, real-time quota tracking, and cost tracking. - Streaming Tool Call Reliability: Stateful preservation of tool call names and IDs across fragmented SSE chunks in multi-turn tool calling sessions.
- Standardized Rate-Limit Retries: Error responses for 429 now provide
retryAfterMsandretry_after_secondsheaders and payload data for fine-grained cooldown handling. - Provider Precedence Order: Standardized provider order across quota pools, credentials, and UI displays (
google-antigravity->ollama->opencode-zen->openai-codex).
- Parent-Account Credential Model: Each account (email) is now the parent entity and may hold per-provider credentials — a single human with both a Google Antigravity OAuth token and an Ollama Cloud API key lives in one account row. Login (
login --provider <id>) and the legacy importer merge credentials onto existing accounts instead of duplicating by email. - Antigravity quota pools consolidated by family: One quota bucket per family —
claude(every Claude variant + gpt-oss) andgemini(every Gemini variant) — instead of one per model. The Antigravity quota API reports the same bucket for all of them, and the dashboard's consolidated RAW POLL line now shows both providers at once, with Antigravity pools first and Ollama last. - Ollama model pricing in
MODEL_PRICING: Spend summaries now report real USD for Ollama traffic (18 entries ported from the predecessor project's catalog of paid prices). Unknown models still return 0. - Routing robustness: Dispatch for Ollama models now keys on the live catalog (
rotator.getOllamaModels()), so models without:in the name (minimax-m3,kimi-k3,glm-5.1,nemotron-3-super,deepseek-v4-pro) no longer leak to the Antigravity adapter. Multi-turn tool-call requests now parsefunction.argumentsto a real object before forwarding, since Ollama's Go API rejects OpenAI-style JSON-encoded strings.
- Multi-Provider Support: The rotator routes through Google Antigravity, Ollama Cloud, and an isolated OpenAI Codex OAuth pool, with per-provider model catalog resolution and no automatic cross-provider fallback.
- Provider Compatibility: Ollama models use native
api/chatNDJSON, while Codex models use native Responses HTTP/SSE with explicit Chat Completions conversion.GET /v1/modelsmarks provider ownership for both catalogs. - Legacy Account Migration: On startup, Ollama Cloud accounts from
~/.ollama-rotator/accounts.json(the predecessor product, overridable viaOLLAMA_ROTATOR_DIR) are imported automatically and merged onto the matching email —provider: "ollama", preservinglabel/tier/type. Re-import is idempotent. - Flat-shape compatibility: Existing configs with the legacy flat field layout (
apiKey/refreshTokenat the top level) are still accepted and normalized on load to the parent-account credential model.
- Lossless Prompt Compression: Optional Lite and RTK compression modes preserve code-critical content while reducing prompt size. Enable
compressionModeinaccounts.jsonor override a request withX-Rotator-Compression: lite,rtk, orrtk+lite. - Operational Observability:
X-Rotator-*response headers expose routing, latency, token, cost, health, idempotency, and compression metrics without changing response bodies. - Reliable Request Handling: Configurable pre-flush stream recovery, duplicate-request idempotency, asynchronous persistence, and an active-account benchmark improve production operations.
- Dashboard Workspaces: Refined Accounts, Virtual Keys, Spend Logs, telemetry, and notification experiences with responsive layouts, filtering, and consistent PII masking.
- Security Hardening: Refresh tokens now use salted
scryptplus AES-256-GCM for new encrypted records, legacy encrypted tokens remain readable, and public error responses no longer expose internal details.
- Automated GHCR Multi-arch Builds: Official Docker images (
linux/amd64,linux/arm64) automatically built and published to GitHub Container Registry. - Pre-built Docker Deployment: Updated
docker-compose.ymlto pull pre-built GHCR images directly out-of-the-box.
- Virtual Keys & Scoped Access Control: Generate scoped API keys (
rk-...) with per-key model authorization rules and user tracking. - Spend Logging & Audit Inspector: PostgreSQL audit trail of all requests, token metrics, TTFB/Total duration, Base64 media sanitization, 6-decimal USD cost breakdown, and Request/Response payload viewer.
- Multi-page Web Dashboard: Unified header navigation connecting Accounts, Virtual Keys, and Spend Logs, featuring customizable column visibility, search/filtering, and instant PII masking.
- PostgreSQL Persistence Backend: Enable
TUXEVIL_ROTATOR_DATABASE_URLfor high-concurrency key validation, persistent spend logging, and retention policies.
Compression is disabled by default. To enable a default mode, add this field to accounts.json:
{
"compressionMode": "lite"
}For a single request, send X-Rotator-Compression: lite, rtk, or rtk+lite. The response reports the selected mode and savings through X-Rotator-Compression-* headers.
To encrypt refresh tokens at rest, set a secret before starting the rotator. A 64-character hexadecimal key is recommended:
export TUXEVIL_ROTATOR_ENCRYPTION_KEY="$(openssl rand -hex 32)"
tuxevil-rotator startExisting enc:v1 records remain decryptable during migration; newly written records use enc:v2.
Legacy env vars still work:
PI_ROTATOR_ENCRYPTION_KEYis honored as a fallback so existing setups keep working without changes.
- Multi-provider gateway — OpenAI-compatible endpoint (
/v1/chat/completions,/v1/responses,/v1/messages) plus native Ollama (/api/chat) routes; any agent or tool can use either - Google Antigravity accounts — OAuth load balancing across a pool, with per-model independent routing
- Ollama Cloud accounts — Static-API-key accounts (
login --provider ollama) for the Ollama Cloud catalog (gpt-oss,gemma4,kimi-k3,minimax-m3,qwen3.5,glm-5,mistral-large-3,nemotron-3,deepseek-v4) - Parent-account credentials — One email, many providers; the same human with Google OAuth and an Ollama key is one row in
accounts.json, withcredentials: [{provider, apiKey|refreshToken, projectId}]instead of duplicate rows by email - Provider-agnostic rotation — The rotation engine treats Antigravity and Ollama uniformly. The next free-tier provider with a per-account credential and a quota API slots in as a new
--providervalue without forking the engine. - Multi-account load balancing — Distributes traffic across a pool of accounts with per-model independent routing
- One-command account setup —
tuxevil-rotator loginauto-discovers (or provisions) the Cloud Code companion project for brand-new Google accounts;login --provider ollamaadds an Ollama Cloud API key to an existing email or creates a new one - Smart rotation & health scoring — Six routing policies (
timer-first,tier-first,quota-first,hybrid,sequential-quota,sticky-quota) with composite health scores per account - Real-time quota monitoring — Polls each provider's quota API on its own cadence; Antigravity quota pools are consolidated by family (
claude,gemini) and Ollama reports monthly usage - Infringement & abuse detection — Flags accounts on enforcement signals and triggers protective pause to preserve the rest of the pool
- Virtual Keys & access control — Issue scoped
rk-...keys for teams, agents, or CI pipelines with per-key model restrictions - Spend logging & audit inspector — Full request/response audit trail with 6-decimal USD cost estimates for both Antigravity and Ollama traffic (requires PostgreSQL)
- Legacy importer — On startup, automatically merges Ollama Cloud accounts from
~/.ollama-rotator/accounts.json(the predecessor product) onto existing accounts, idempotently - Web dashboard — Live routing health and per-model quota runway, account cards with one-click fixes, a live request tail, token usage and savings, latency (p50/p95), an activity heatmap, and per-model routing decisions
- State persistence — Survives restarts; routing assignments, cooldowns, and flags saved to disk or PostgreSQL
- Tool/function calling — Fully supported in OpenAI and Anthropic formats, including multi-turn and parallel tool calls, with reliable function-name resolution for tool responses across turns. Ollama forwards
function.argumentsas a parsed object (OpenAI sends it as a JSON string and Ollama's Go API rejects that). - Reasoning/thinking visibility — Interleaved thinking blocks exposed as
reasoning_content/thinking_deltain real time
Requirements: Node.js 22+ for npm/source installs. Docker image uses Node 22.
npm install -g tuxevil-rotator
tuxevil-rotator login # Google Antigravity (default)
tuxevil-rotator login --provider ollama # Ollama Cloud (static API key)
tuxevil-rotator login --provider openai-codex # ChatGPT OAuth (isolated Codex pool)
tuxevil-rotator login --provider opencode-zen # OpenCode Zen (static API key)
tuxevil-rotator startOllama Cloud accounts from a previous ~/ollama-rotator install are imported automatically at startup — no manual step required.
Coming from
pi-antigravity-rotator? Install the new package, then run the guided migration:npm install -g tuxevil-rotator tuxevil-rotator migrate tuxevil-rotator doctorThe migration copies legacy config files without deleting them. The new package also keeps the old CLI as a deprecation shim and honors legacy environment variables and paths. See the Migration guide.
mkdir -p docker-data
docker compose up -dgit clone https://github.com/tuxevil/tuxevil-rotator.git
cd tuxevil-rotator
npm install
npm run login # Google Antigravity
npm run login -- --provider ollama # Ollama Cloud
npm run login -- --provider openai-codex # OpenAI Codex OAuth
npm run login -- --provider opencode-zen # OpenCode Zen
npm startDashboard opens at http://localhost:51200/dashboard
Full deployment guide → · Adding accounts →
Point any OpenAI-compatible agent to http://localhost:51200/v1 with API key tuxevil (or a Virtual Key):
You can configure supported local agents automatically with:
tuxevil-rotator install-agent --target autoUse --target all or a comma-separated list (opencode,hermes,pi,codex) to choose clients. Add --host <hostname> --user <ssh-user> to configure them over SSH, or --dry-run to review the generated installer first. For OpenCode's live model catalog, use tuxevil-rotator install-opencode; see the OpenCode integration guide.
| Agent | Guide |
|---|---|
| Pi | docs/integrations/pi.md |
| OpenCode | docs/integrations/opencode.md |
| Hermes | docs/integrations/hermes.md |
| OpenClaw | docs/integrations/openclaw.md |
| Cursor | docs/integrations/cursor.md |
| Claude Code | docs/integrations/claude-code.md |
| Codex (OpenAI CLI) | docs/integrations/codex.md |
| Cline | docs/integrations/cline.md |
| Roo Code | docs/integrations/roo-code.md |
| Continue | docs/integrations/continue.md |
| Aider | docs/integrations/aider.md |
| Open WebUI | docs/integrations/open-webui.md |
graph LR
A[Your Agent] -->|OpenAI / Anthropic API| B["tuxevil-rotator<br/>localhost:51200"]
B -->|Smart Routing| C[Google Account 1]
B -->|Smart Routing| D[Google Account 2]
B -->|Smart Routing| E[Ollama Cloud Account 1]
B -->|Smart Routing| F[Ollama Cloud Account 2]
B -->|Smart Routing| G[... up to N]
C --> H[Google Antigravity]
D --> H
E --> I[Ollama Cloud]
F --> I
G --> I
H -.Quota API<br/>every 5 min.-> B
I -.Usage API<br/>every 5 min.-> B
Each model routes to its own best available account independently. The same email can hold both a Google OAuth credential and an Ollama Cloud API key (parent-account model) — the rotator picks the right credential at request time based on the destination model.
How model routing works:
- Antigravity pool — Claude variants (
claude-opus-4-6-thinking,claude-sonnet-4-6,gpt-oss-120b) share theclaudequota bucket; Gemini variants share thegeminibucket. - Ollama pool — Any model returned by
https://ollama.com/api/tags(e.g.gemma4:31b,gpt-oss:20b,minimax-m3,kimi-k3) routes to an Ollama credential. - The pool is selected by the destination model's provider; a single account with both credentials participates in both pools.
After starting the proxy, open http://localhost:51200/dashboard and sign in with the admin token (TUXEVIL_ROTATOR_ADMIN_TOKEN). A link of the form /dashboard?token=<token> signs in directly; either way the browser gets a session cookie and the token is removed from the URL.
- Overview — whether routing works right now, one row per model pool (Claude, Gemini, Codex, Ollama, OpenCode) with pooled quota, the account serving it, what is next in line, the next reset and how long the quota lasts at the current rate, plus a list of what needs you with the fix one click away
- Accounts — cards with each account's quota per model, or a compact sortable list; each account opens a drawer with its quota windows, routing decisions, health breakdown, limits and every action (re-enable, restore, start idle windows, tier, quarantine, remove)
- Requests — a live tail of requests and rotator events, and with PostgreSQL the full history with payloads and per-key totals
- Usage — token volume by model with 1h–30d ranges, estimated savings at list prices, latency percentiles and a 60-day activity heatmap
- Virtual keys and Settings — key management, routing controls and policy, a benchmark, and the raw configuration file
Press ⌘K (Ctrl+K) to jump to any page or account or run a quick action, and the bell in the top bar lists what needs you. Light and dark themes follow the system unless you pick one in the profile menu at the bottom of the sidebar; the privacy mask there (or ?mask=1) hides names, emails and keys for screenshots.
With PostgreSQL, the gateway adds enterprise-grade access control and cost auditing.
# Set up PostgreSQL (or paste the prompt in docs/integrations/setup-postgresql.md into your AI agent)
export TUXEVIL_ROTATOR_DATABASE_URL="postgres://user:***@localhost:5432/rotatordb"
# Generate a scoped key
tuxevil-rotator keys generate --alias "cursor-agent" --models "gemini-3.6-flash-high"
# → rk-a1b2c3d4...Virtual Keys guide → · Setting up PostgreSQL →
| Topic | Link |
|---|---|
| How It Works | docs/how-it-works.md |
| Configuration | docs/configuration.md |
| Dashboard | docs/dashboard.md |
| Virtual Keys & Spend Logging | docs/virtual-keys.md |
| API Reference | docs/api-reference.md |
| Compatibility Adapters | docs/compatibility.md |
| Deployment | docs/deployment.md |
| Adding Accounts | docs/adding-accounts.md |
| Migration from pi-antigravity-rotator | docs/migrating-from-pi-antigravity-rotator.md |
| Troubleshooting | docs/troubleshooting.md |
| Telemetry | docs/telemetry.md |
| PostgreSQL Setup | docs/integrations/setup-postgresql.md |
This project started as a focused, single-purpose rotator for pi coder agent and for Google Antigravity provider. As the rotation engine matured and the Ollama Cloud integration landed, the same engine applied cleanly to any free-tier provider with per-account credentials and a quota API. Renaming to tuxevil-rotator makes that generalization explicit: the provider is a plug-in, the rotation is the product.
A few things shaped the decision concretely:
- The original name (
pi-antigravity-rotator) signaled "I only know one provider." That's no longer true. - The account store moved to a parent-account model where one email can hold credentials for multiple providers. The product surface needed to follow.
- Open-sourcing the rotator as
pi-*implied it was only interesting inside thepi.devecosystem. The engine itself is provider-agnostic and useful in any agent stack.
The name comes from the maintainer's nickname (tuxevil). It's also the namespace already used across this maintainer's other open-source work, so the rotator stops being a one-off and becomes part of a recognizable family of tools.
Want the longer story? A deeper post on the rationale, the design tradeoffs, and the provider-pluggable architecture is on the maintainer's blog: Why I renamed pi-antigravity-rotator to tuxevil-rotator.
If this tool has saved you API costs, consider supporting its development!
To donate an authorized account for testing across supported providers, see CONTRIBUTING.md.
Thanks to these amazing people who have contributed to the project:
- @CelestialCreator (Akshay) — Fixed admin token propagation on the hosted login landing page (
/loginto/auth/antigravity/start). (PR #25) - @Codder-hermes — Fixed Claude Code tool-schema requests by stripping the unsupported JSON Schema
propertyNameskeyword for Gemini and Claude-via-Gemini routes, with regression coverage for both compatibility paths. (PR #19) - @CyR1en (Ethan Bacurio) — Added Gemini 3.6, Gemini 3.7, and Gemini 3.8 Flash model families, shared quota-pool routing, pricing, dashboard support, integration documentation, effort-based model routing, the rebuilt WebUI, and regression coverage. (PR #18, PR #21, PR #28, PR #32, PR #38)
- @josenicomaia (José Nicodemos Maia Neto) — Modularized the compatibility layer architecture, added multimodal tool response support, fixed streaming pass-through for tool executions, and added PostgreSQL storage and weighted pool quota forecasting. (PR #8, PR #9, PR #11, PR #13, PR #14)
- @yashyadav711 (Yash) — Fixed Draft-2020-12 inline JSON-Schema union type mapping for Gemini tools support. (PR #10)
- @toRolex (Rolex) — Added cached prompt-token extraction from Google Antigravity/Cloud Code usage and mapped it to OpenAI and Anthropic compatibility responses, with focused regression tests, plus transport-error retry recovery for single-account and exhausted-account pools. (PR #34, PR #35)
- @javargasm (Jeisson Alexander Vargas Marroquin) — Queued multi-request Antigravity account pooling (5×5 concurrent streams), strict FIFO overflow queue, dashboard concurrency metrics, canonical project ID resolution, account-scoped dynamic Antigravity discovery and routing, dynamic catalog and quota-safety hardening, persistent dynamic model ownership across restarts, 429 quota resilience with account-scoped retry, synchronous leasing, snapshot-based token refresh with generation validation, provider-scoped token publication, spend-logger memory leak fix, atomic PostgreSQL CTE daily spend aggregation with idempotent retries, virtual-key cache bounding, Anthropic tool-use compatibility layer (
tool_use/tool_resultcontent block conversion), JSON schema round-trip fixes, compat test suite expansion, model-scoped cooldowns, explicit Antigravity reset-duration parsing, idle pool normalization, RAW POLL deadline reconciliation, serialized per-account quota polling, Ollama kickstart routing for multi-provider accounts, and OpenAI-compatible audio transcription plus bidirectional live WebSocket streaming through the local Antigravity observer model, with PR #36 adding rotator-backed transcription, Language Server fallback, model-scope checks, and live timeout/control-frame hardening. (PR #3, PR #7, PR #22, PR #23, PR #24, PR #26, PR #29, PR #30, PR #31, PR #33, PR #36)
npm run typecheck # Type-check src/
npm run typecheck:test # Type-check src/ + test/
npm test # Run test suite
npm run check # typecheck + test + lint (full gate)
