The existing pinned App Server path remains the reference implementation, including ChatGPT identity, live model and effort discovery, quota telemetry, cancellation, and exact crash recovery.
vtx inference-host codex-login; vtx inference-host service install --adapter codex
Billing: subscription · Schema: native · Crash replay: yes
VTX uses GitHub's official SDK with the package-owned pinned platform runtime, authenticated GitHub login identity, exact enabled model and effort events, token usage, and cooperative cancellation. Auto-only catalogs fail closed because they cannot prove the effective model. Output text is schema-validated by VTX; uncertain interrupted calls are quarantined.
Sign in with the official Copilot CLI, then run vtx inference-host service install --adapter copilot
Billing: subscription credits · Schema: validated text · Crash replay: quarantined
VTX uses the exact-pinned official DeepSeek Harness runtime in both Provider and Main Agent modes. The key remains in the local VTX private credential boundary, Provider runs have no tools, Agent runs expose only the three assignment-scoped VTX tools, and a loopback receipt verifies the provider-returned model and response ID. VTX validates final JSON; interrupted ambiguous work is never blindly retried.
Import a DeepSeek API key through stdin with vtx inference-host deepseek-login, then run vtx inference-host service install --adapter deepseek-harness
Billing: api billed · Schema: validated text · Crash replay: quarantined
The pinned Pi SDK supports Provider mode and Main Agent mode through the existing controls. The initially qualified models use one separately stored OpenAI API key per instance and advertise exact provider/model identities from the release's configured qualification policy and SDK catalog. Provider runs have no tools; persistent Agent sessions expose only the three assignment-scoped VTX tools. VTX validates structured output and execution receipts. Uncertain interrupted model calls stop safely; Pi does not claim same-attempt recovery. Subscription OAuth and unqualified provider protocols are not enabled by this adapter.
Import an OpenAI API key through stdin with vtx inference-host pi-login --provider openai, then run vtx inference-host doctor --adapter pi and vtx inference-host service install --adapter pi
Billing: api billed · Schema: native · Crash replay: quarantined
Exo means exoharness/exo. The VTX-owned bridge runs the pinned Exo custom harness automatically in Provider and Main Agent modes; no manual job transfer is needed. Each instance advertises only the release-qualified routes for its OpenAI or Venice credential. Venice currently offers Gemma 4 31B IT with effort None only. Provider runs receive complete VTX input with no tools; Agent runs expose only assignment-scoped VTX tools. VTX validates canonical JSON; native structured-output enforcement is not claimed. VTX OAuth stays outside Exo. This Linux/WSL foreground preview has no durable service or same-attempt recovery: uncertain interrupted work remains fenced. It does not add native Exo MCP support.
On Linux or WSL, import a Venice key with vtx inference-host exo-login --provider venice --instance exo-1 (use --provider openai for OpenAI), then keep vtx inference-host run --adapter exo --exo-root /path/to/exo --instance exo-1 open
Billing: api billed · Schema: validated text · Crash replay: quarantined
A live 1.1.13 Starter Quota acceptance exposed shell, file, browser, web, MCP, subagent, and messaging tools even under an isolated agent declaring tools: [], and the requested final JSON Schema was not enforced. VTX therefore keeps Antigravity in a monitored foreground loop.
Use vtx inference-host agent-connect and keep the signed-in Antigravity loop open
Billing: subscription credits · Schema: validated text · Crash replay: quarantined
The current official individual OAuth flow rejects this client and directs users to Antigravity. API-key and Vertex routes are API-billed rather than a consumer-subscription subsidy, and the CLI does not enforce a final response schema.
Use the foreground agent loop or VTX's normal Gemini API provider; do not install Gemini CLI as a subscription worker
Billing: mixed · Schema: validated text · Crash replay: quarantined
Claude Code has a strong official programmatic contract, but Anthropic currently says third-party developers may not route Free, Pro, or Max credentials on users' behalf without prior approval. VTX therefore keeps it on the user's foreground first-party loop.
Use vtx inference-host agent-connect and keep the Claude Code agent loop open
Billing: subscription · Schema: validated text · Crash replay: quarantined
Grok Build uses the pinned official runtime and a separate local OAuth session for Provider and Main Agent modes. Provider runs have no tools; Agent runs expose only the three assignment-scoped VTX tools. VTX validates structured output, model identity, and reported usage. Uncertain interrupted requests remain fenced; same-attempt recovery and Codex subscription quota telemetry are not supported.
Sign in with vtx inference-host grok-login, then install a grok-build inference host
Billing: subscription · Schema: native · Crash replay: quarantined
Kiro API keys legitimately consume paid subscription credits, but its headless chat lacks a stable JSON result, effective-model, token-usage, and cancellation receipt.
Use the foreground agent loop; do not install Kiro as a durable VTX worker yet
Billing: subscription credits · Schema: validated text · Crash replay: quarantined
Cursor's official SDK has strong lifecycle and usage telemetry, but SDK authentication is API-key and token-billed rather than a clear consumer-subscription subsidy.
Use Cursor's foreground agent loop for subscription sessions
Billing: api billed · Schema: validated text · Crash replay: quarantined
Amp offers subscriptions, credits, and linked ChatGPT capacity, but its selectable product is a managed mode rather than a stable underlying model identity.
Use Amp through agent-connect while managed model-mode identity remains unfrozen
Billing: mixed · Schema: validated text · Crash replay: quarantined
Auggie exposes account, models, credits, ACP cancellation, and sessions, but its SDK currently declares structured output unsupported and recovery remains incomplete.
Use Auggie through agent-connect until structured results and recovery are complete
Billing: subscription credits · Schema: validated text · Crash replay: quarantined
Junie has headless JSON, ACP, sessions, and subscription routing, but effort can fall back and its public machine contract does not yet provide the exact effective identity and usage evidence required for a durable VTX worker.
Use Junie through agent-connect while its effective model, effort, and usage receipts remain incomplete
Billing: subscription credits · Schema: validated text · Crash replay: quarantined
Warp and Oz can consume plan credits and run cloud agents, but they do not expose the exact effort, structured result, token accounting, and direct inference lifecycle required for a durable VTX provider.
Use Warp or Oz through agent-connect for foreground automation
Billing: subscription credits · Schema: validated text · Crash replay: quarantined
Qwen Code has a strong no-tools, JSON Schema, usage, and cancellation contract, but the Coding Plan prohibits automated scripts, application backends, and non-interactive use. Standard ModelStudio API quota is a normal API/BYOK route rather than subscription capacity.
Keep Coding Plan sessions foreground-only; use VTX's normal ModelStudio provider for API-key usage
Billing: mixed · Schema: native · Crash replay: quarantined
OpenCode can run VTX Provider jobs or claim a true Main Agent assignment while its foreground loop stays open, and it can connect directly to VTX Insights with scoped OAuth and native Agent Skills. The open-source harness has a strong SDK and session contract, but its generic receipts do not prove provider-returned model identity or complete ambient-tool isolation. Its hosted terms also prohibit unattended programmatic output extraction, and upstream vendors do not authorize OpenCode as a durable router for their consumer subscription OAuth credentials.
Use OpenCode through agent-connect for foreground Provider or Main Agent mode with local or BYOK models; use its native OAuth MCP client for VTX Insights
Billing: mixed · Schema: native · Crash replay: quarantined
OpenClaw's published package supports remote OAuth MCP servers, so it can use VTX Insights directly. Durable VTX inference remains disabled because the published runtime does not expose a package-owned terminal receipt with complete effective model, provider-call usage, tool, cancellation, and interruption evidence.
Use OpenClaw's native MCP client for VTX Insights; use agent-connect only for foreground VTX inference
Billing: mixed · Schema: validated text · Crash replay: quarantined
Hermes supports remote OAuth MCP servers, so it can use VTX Insights directly. Its one-shot runner can make unreported auxiliary model calls, undercounts usage, reports requested rather than proven effective identity, can exit successfully after provider failures, and has no structured interruption receipt or clean no-tools mode, so it is not a durable VTX adapter.
Use Hermes' native MCP client for VTX Insights; use agent-connect only for foreground VTX inference
Billing: mixed · Schema: validated text · Crash replay: quarantined