Which AI Models Does Prompt Anything Pro Support?
Prompt Anything Pro supports 30+ models across seven providers via BYOK: OpenAI, Anthropic, Google, xAI, and Groq on the cloud side, plus local models through Ollama or LM Studio and Chrome's built-in AI (Gemini Nano) on the device side. The current flagships are OpenAI's GPT-5.4, Anthropic's Claude Opus 4.6, Google's Gemini 3.1 Pro, and xAI's Grok 4.3, alongside faster low-cost tiers (GPT-5.4 Mini, Claude Haiku 4.5, Gemini 3.1 Flash), OpenAI's o3 and o4-mini reasoning models, and a set of earlier GPT-4.1, Claude, and Gemini releases kept available for compatibility. You configure each cloud provider's API key once — local and on-device models need no key at all — and switch models per-prompt with a one-click dropdown. New provider models become available within days of their API launch.
- OpenAI: GPT-5.4 (flagship), GPT-5.4 Mini (cheap/fast), o3 and o4-mini (reasoning), plus the GPT-4.1 family.
- Anthropic: Claude Opus 4.6 (flagship, strongest for writing + analysis), Claude Sonnet 4.6 (balanced daily driver), Claude Haiku 4.5 (fast/cheap).
- Google: Gemini 3.1 Pro (flagship, best for long documents), Gemini 3.1 Flash (fastest, cheapest tier), plus LearnLM 2.0 for explanation-heavy tasks.
- xAI: Grok 4.3 (flagship) and the rest of the Grok line. Groq: fast open-weight models when throughput matters more than frontier quality.
- Local and on-device: Ollama, LM Studio, and Chrome's built-in AI (Gemini Nano) — no API key, no per-token bill, nothing leaves your machine.
- 30+ models in total across 7 providers. Switch per-prompt with a single dropdown click. BYOK means new models from any provider work the same day they launch in the API.
By PlugMonkey Team, Editorial
OpenAI Models
- GPT-5.4 — OpenAI's current flagship. Best all-around choice for complex reasoning, creative writing, code generation, and analysis.
- GPT-5.4 Mini — A smaller, faster, cheaper sibling of GPT-5.4. Great for high-volume work like summarization, rewriting, and quick Q&A.
- o3 and o4-mini — OpenAI's reasoning-optimized models. Worth reaching for on multi-step problems, maths, and tricky debugging where you want the model to think before answering.
- GPT-4.1, GPT-4.1 Mini, and GPT-4.1 Nano — The previous generation, still selectable if you have prompts or workflows tuned to them.
- Per-token rates differ for every model and change over time — check OpenAI's pricing page for the current numbers before you settle on a default.
Anthropic Models
- Claude Opus 4.6 — Anthropic's most capable model. Excels at nuanced writing, long-form analysis, coding, and tasks that need careful step-by-step reasoning.
- Claude Sonnet 4.6 — The balanced mid-tier. Close to Opus on most everyday work at a lower per-token rate, and the sensible default for daily writing and editing.
- Claude Haiku 4.5 — Optimized for speed and cost. Ideal for quick tasks where you want Claude's writing style without paying flagship rates.
- Claude Sonnet 3.5 — An earlier release, still selectable for workflows already tuned to it.
- Rates vary per model — see Anthropic's pricing page for the current per-token cost of each.
Google Models
- Gemini 3.1 Pro — Google's current flagship. Strong on long-document work, codebases, and large datasets, and the Gemini model to reach for first.
- Gemini 3.1 Flash — The fast, lightweight tier. Built for speed and high-volume simple tasks at a fraction of Pro's cost.
- Gemini 2.5 Pro and Gemini 2.5 Flash — The previous Pro/Flash pair, still available for workflows already tuned to them.
- Gemini 2.0 Flash, Gemini 2.0 Flash-Lite, Gemini 1.5 Pro, and Gemini 1.5 Flash — Earlier releases retained for compatibility.
- LearnLM 2.0 — Google's model tuned for teaching and explanation. Useful when you want a concept broken down rather than a direct answer.
- Context limits and per-token rates differ across the family — Google's pricing page has the current figures.
xAI Models
- Grok 4.3 — xAI's current flagship, and the Grok model to start with.
- The rest of the Grok line is selectable from the same dropdown once your xAI key is configured.
- Per-token rates are published on xAI's API page — check there before setting Grok as a default.
Groq
- Fast open-weight models served at a fraction of flagship latency.
- Same BYOK setup as every other cloud provider: paste a key once, then pick the model per prompt.
- Current model list and rates live on Groq's pricing page.
Local and On-Device Models (No Key Required)
- Ollama — Point Prompt Anything Pro at your local Ollama server and use whichever models you have pulled.
- LM Studio — Same idea with LM Studio's local server, if that is your preferred runner.
- Chrome's built-in AI (Gemini Nano) — Runs inside the browser itself. Nothing to install, no key, no network round-trip.
- Zero marginal cost: local and on-device prompts do not appear on any provider invoice.
How to Switch Between Models
Which Model Should You Use?
- Complex analysis or reasoning — GPT-5.4 or Claude Opus 4.6. Both are top-tier; try both and see which style you prefer. For multi-step problems, o3 is worth a look too.
- Creative writing and editing — Claude Opus 4.6 and Claude Sonnet 4.6 tend to produce more natural, nuanced prose. GPT-5.4 is also excellent but can read as more formulaic.
- Code generation and debugging — GPT-5.4 and Claude Opus 4.6 are both strong. Run the same prompt through each on a task you already know the answer to, then keep whichever you trust more.
- Summarization and quick tasks — GPT-5.4 Mini, Claude Haiku 4.5, or Gemini 3.1 Flash. Fast and cheap, more than capable for straightforward work.
- Long document processing — Gemini 3.1 Pro. Context limits differ per model, so check the provider's model card before pasting something very large.
- Budget-conscious daily use — GPT-5.4 Mini or Gemini 3.1 Flash for most tasks, escalating to GPT-5.4 or Claude Opus 4.6 only for complex work.
Adding New Models as They Launch
Want a Second Opinion?
Ask AI for an independent perspective on this question.
AI responses are generated independently and may vary
Sources & Further Reading
- OpenAI API pricing — current per-token rates for all supported OpenAI models — OpenAI (accessed May 22, 2026)
- Anthropic Claude pricing — current per-token rates for all supported Claude models — Anthropic (accessed May 22, 2026)
- Google Gemini API pricing — current per-token rates for all supported Gemini models — Google AI for Developers (accessed May 22, 2026)
Try Prompt Anything Pro Free
Access GPT-5.4, Claude Opus 4.6, Gemini 3.1 Pro, and Grok 4.3 from any webpage. Bring your own API keys, pay actual costs.