OpenWorker Supported Models
A practical comparison of every AI model you can run with OpenWorker — performance, cost, and best use cases for each provider.
Cloud Models for OpenWorker
These models run on provider infrastructure and deliver the fastest, most capable results inside OpenWorker. You pay per token through your own API key.
| Model | Provider | Context | Tool Calling | Speed | Cost | Best For |
|---|---|---|---|---|---|---|
| GPT-5.6 Sol | OpenAI | 128K | Excellent | Fast | $15 / 1M tokens | Complex reasoning, multi-step planning, detailed analysis |
| GPT-5.6 Sol Mini | OpenAI | 128K | Very Good | Very Fast | $4 / 1M tokens | Quick tasks, drafting, summarization |
| Claude Fable | Anthropic | 200K | Excellent | Fast | $12 / 1M tokens | Nuanced writing, careful reasoning, safety-sensitive tasks |
| Claude Sonnet 4.6 | Anthropic | 200K | Very Good | Very Fast | $3 / 1M tokens | Balanced speed and quality, coding tasks |
| Gemini 3.6 | 1M | Very Good | Fast | $10 / 1M tokens | Multimodal tasks, very long documents, image analysis | |
| Gemini 3.6 Flash | 1M | Good | Very Fast | $1.5 / 1M tokens | High-volume tasks, cost-sensitive workflows | |
| DeepSeek R2 | DeepSeek | 128K | Good | Medium | $2 / 1M tokens | Coding, technical analysis, budget-friendly reasoning |
| Kimi K3 | Moonshot | 256K | Good | Medium | $3 / 1M tokens | Long document processing, research synthesis |
Local Models via Ollama for OpenWorker
Run these models locally through OpenWorker and Ollama on your own Mac. Zero API cost, complete privacy, and no internet connection required. Trade-off is slower speed and lower quality compared to premium cloud models.
| Model | RAM Required | Quality | Speed | Best For |
|---|---|---|---|---|
| Llama 3.3 70B | 40 GB | Near GPT-4 level | Slow on consumer hardware | Privacy-critical tasks, offline use |
| Llama 3.3 8B | 5 GB | Good for simple tasks | Fast on most Macs | Quick drafts, email triage, basic summarization |
| Qwen 3 32B | 20 GB | Strong reasoning | Medium | Multilingual tasks, coding, analysis |
| Mistral Large 2 | 24 GB | Very capable | Medium | Business writing, European language support |
| DeepSeek R2 Lite | 8 GB | Solid reasoning | Medium | Code generation, technical documentation |
Which Model Should You Use with OpenWorker?
The best model depends on your task, budget, and privacy requirements. Here are our recommendations for common OpenWorker workflows.
Daily office work (email, calendar, Slack)
Recommended: GPT-5.6 Sol Mini or Gemini 3.6 Flash
Fast response times keep your OpenWorker workflow smooth. These models handle routine OpenWorker tasks well at low cost. A typical day of email triage and calendar management costs under $0.50.
Customer-facing writing (briefs, reports, proposals)
Recommended: Claude Fable or GPT-5.6 Sol
Premium models produce more polished, nuanced writing in OpenWorker. The extra cost per task is negligible when the output goes directly to clients or stakeholders.
Code and technical documentation
Recommended: DeepSeek R2 or Claude Sonnet 4.6
Both excel at understanding code in OpenWorker. DeepSeek is more cost-effective for OpenWorker coding workflows; Claude Sonnet offers faster iteration cycles.
Sensitive or confidential data
Recommended: Ollama with Llama 3.3 8B or Qwen 3 32B
Local models in OpenWorker ensure nothing leaves your machine. Even your prompts stay private. The quality trade-off is worth it when handling legal documents, HR data, or financial records.
Processing very long documents (100+ pages)
Recommended: Gemini 3.6 (1M context) or Kimi K3 (256K)
These models have the largest context windows, allowing OpenWorker to process entire documents without chunking. Essential for contract review, research papers, and audit reports.
OpenWorker Model Selection Strategy
The single most important factor in model selection for OpenWorker is tool-calling reliability. Unlike simple chat, OpenWorker needs the model to correctly decide when to call integrations, what parameters to pass, and how to chain multiple tool calls together. Models with excellent tool-calling scores — GPT-5.6 Sol, Claude Fable, and Claude Sonnet 4.6 — produce the most reliable multi-step workflows.
A practical approach for OpenWorker is to use a fast, cheap model as your default and switch to a premium model for high-stakes tasks. Set GPT-5.6 Sol Mini or Gemini 3.6 Flash as your daily OpenWorker driver for email triage, calendar checks, and quick summaries. When you need OpenWorker to prepare a client-facing brief or analyze a complex document, switch to GPT-5.6 Sol or Claude Fable for that specific task.
For teams handling sensitive data, the hybrid approach in OpenWorker works well: use local Ollama models for internal document processing and classification, then switch to a cloud model only when the OpenWorker task requires higher quality output that will be reviewed by a human before sharing externally. This minimizes data exposure while maintaining output quality where it matters most.
OpenWorker Models FAQ
Can I switch models in the middle of an OpenWorker task?
Which model is cheapest for OpenWorker?
Does OpenWorker support fine-tuned models?
Why do some models work better with OpenWorker than others?
Can I use multiple API keys from the same provider in OpenWorker?
How much does a typical day of OpenWorker usage cost?
What happens if my API key runs out of credits?
Are Ollama models slower than cloud models in OpenWorker?
Ready to Pick Your Model?
Install OpenWorker, add your preferred API key, and start with a mid-tier model. You can always switch later as you learn which workflows benefit from premium models.