Claude, GPT and Codex, Gemini, Kimi, DeepSeek, Grok, Qwen, GLM, and local models on your own machine. One window, one interface, one set of directives. Choose per session, and switch mid-thread.
Same window, same keys. The task picks the model — you never re-explain the project.
Claude
GPT · Codex
PerplexityFrontier, open-weight, and local — one integration, your key for each.
“Significantly dumber. It ignored its own plan, messed up the code, and started to lie about the changes it made.”
The "AI shrinkflation" backlash. Anthropic's own postmortem confirmed two bugs degraded models Aug 5 to Sep 5 2025. A six-week track of identical prompts logged correctness falling 4.2 to 3.1 and hallucinations rising 8% to 22%.
View the source ›
When one model has a bad week, being locked to it is the whole problem. Overwatch lets you move the same session onto a different engine and keep going. The per-session intensity dial also pins high effort where a silent "medium" default quietly took it away.
One integration, 75+ providers. Frontier models, open-weight models, and local models through Ollama or LM Studio.
Reasoning model on the hard thread, a fast cheap one on the scratch thread. Both open at the same time.
Bring your own account. The key stays on your device and you control what gets spent.
Not happy with the answer? Change the engine and ask again in the same session, with the same context. No copy-paste into another app, no re-explaining the project.
No copy-paste into another app. No re-explaining the project.
Type //perplexity for live research, //image to generate, //gemini for a second opinion. The session stays where it is and the right model does the piece it is best at.
The session stays where it is. The right model does the piece it is best at.