Models

Claude models, tuned for careful reasoning, long documents, and coding.
| Model | Best for |
|---|---|
| Fable 5 | Long-running agents, next-gen intelligence |
| Opus 5 | Complex agentic coding, enterprise work |
| Sonnet 5 | Best balance of speed and intelligence |
| Haiku 4.5 | Fastest, near-frontier intelligence |
GPT-5.6 models spanning frontier reasoning down to fast, low-cost workloads.
| Model | Best for |
|---|---|
| 5.6 Sol | Complex professional reasoning, coding |
| 5.6 Terra | Balanced intelligence and cost |
| 5.6 Luna | Cost-sensitive, high-volume workloads |
Gemini models with native multimodal input and very long context windows.
| Model | Best for |
|---|---|
| 3.7 Flash | Complex coding, agentic workflows |
| 3.1 Pro | Advanced reasoning, problem-solving |
Alibaba's Qwen family, strong at multilingual tasks and coding.
| Model | Best for |
|---|---|
| 3.8 Max | Flagship reasoning, hardest coding tasks |
| 3.7 Plus | Balanced reasoning, multimodal input |
| 3.7 Flash | Fast, high-volume tasks, low cost |
Efficient open and commercial models, including fast, low-cost options.
| Model | Best for |
|---|---|
| Medium 3.5 | Frontier agentic and coding work |
| Large 3 | General-purpose multimodal tasks |
| Small 4 | Instruct, reasoning, coding in one |
Research-driven models built around efficient reasoning at low cost.
| Model | Best for |
|---|---|
| V4 Pro | Complex tasks, top-tier performance |
| V4 Flash | Fast, cost-efficient everyday tasks |

Grok models built for agentic tool use, coding, and fast, direct responses.
| Model | Best for |
|---|---|
| Grok 4.6 | Agentic coding and everything else |
Enterprise-focused models built for retrieval and business workflows.
| Model | Best for |
|---|---|
| Command A+ | Enterprise multimodal reasoning |
| Command A | Fast, efficient text, 23 languages |
Muse Spark for agentic coding over the API, plus open-weight Muse Glimmer you can run yourself.
| Model | Best for |
|---|---|
| Muse Spark | Agentic coding, tool loops, over the API |
| Muse Glimmer | Self-hosted, open-weight customization |
Kimi models, built for long-context work and agentic tool use.
| Model | Best for |
|---|---|
| Kimi K3 | Flagship, 1M-token context window |
| Kimi K2.7 Code | Dedicated coding, multimodal tasks |
Custom and self-hosted endpoints
Point Omnesly at any OpenAI-compatible API — including models you're running locally on your own machine. If it speaks the standard chat-completions format, Omnesly can talk to it.
Add a custom endpoint
New providers, added as they matter.
The model landscape moves fast. We add support for new providers and model families on an ongoing basis — no app update required to start using a newly connected key.