Account

Models.

Omnesly connects to Anthropic, OpenAI, Google, Qwen, Mistral, DeepSeek, xAI, Cohere, Meta Muse, Moonshot AI, Z.ai, Tencent, and any OpenAI-compatible or self-hosted endpoint.

Anthropic Claude logo
Anthropic

Claude models, tuned for careful reasoning, long documents, and coding.

ModelBest for
Fable 5Long-running agents, next-gen intelligence
Opus 5Complex agentic coding, enterprise work
Sonnet 5Best balance of speed and intelligence
Haiku 4.5Fastest, near-frontier intelligence
OpenAI logo
OpenAI

GPT-6 and GPT-5.6 models spanning frontier reasoning down to fast, low-cost workloads.

ModelBest for
6 AstraFrontier agentic reasoning, computer use
5.6 SolComplex professional reasoning, coding
5.6 TerraBalanced intelligence and cost
5.6 LunaCost-sensitive, high-volume workloads
Google Gemini logo
Google

Gemini models with native multimodal input and very long context windows.

ModelBest for
3.7 FlashComplex coding, agentic workflows
3.1 ProAdvanced reasoning, problem-solving
Qwen logo
Qwen

Alibaba's Qwen family, strong at multilingual tasks and coding.

ModelBest for
3.8 MaxFlagship reasoning, hardest coding tasks
3.7 PlusBalanced reasoning, multimodal input
3.7 FlashFast, high-volume tasks, low cost
Qwen3-CoderDedicated coding, large-scale MoE
Mistral AI logo
Mistral AI

Efficient open and commercial models, including fast, low-cost options.

ModelBest for
Medium 3.5Frontier agentic and coding work
Large 3General-purpose multimodal tasks
Small 4Instruct, reasoning, coding in one
CodestralDedicated code generation
DeepSeek logo
DeepSeek

Research-driven models built around efficient reasoning at low cost.

ModelBest for
V4.1 FlashComplex and everyday tasks, top-tier performance at low cost
V4 ProComplex tasks, top-tier performance
V4 FlashFast, cost-efficient everyday tasks
xAI Grok logo
xAI

Grok models built for agentic tool use, coding, and fast, direct responses.

ModelBest for
Grok 4.6Agentic coding and everything else
Grok 4.1 FastFast, low-cost, 2M-token context
Cohere logo
Cohere

Enterprise-focused models built for retrieval and business workflows.

ModelBest for
Command A+Enterprise multimodal reasoning
Command AFast, efficient text, 23 languages
Meta logo
Meta

Meta's Muse Spark model, built for agentic coding and tool use over the API.

ModelBest for
Muse SparkAgentic coding, tool loops, over the API
Moonshot AI logo
Moonshot AI

Kimi models, built for long-context work and agentic tool use.

ModelBest for
Kimi K3Flagship, 1M-token context window
Kimi K2.6Affordable, general-purpose tasks
Kimi K2.7 CodeDedicated coding, multimodal tasks
Z.ai logo
Z.ai

GLM models from Zhipu AI, tuned for high-throughput coding and reasoning.

ModelBest for
GLM-5.3Flagship coding, frontier reasoning
GLM-5 TurboFast, high-volume agentic workloads
GLM-4.7Cost-efficient everyday workhorse
Tencent Hunyuan logo
Tencent

Open-weight Hunyuan models pairing large-scale MoE reasoning with a 256K context window.

ModelBest for
Hunyuan Hy3Flagship reasoning and agentic tasks
Hunyuan A13BFast, cost-efficient everyday tasks

Custom and self-hosted endpoints

Point Omnesly at any OpenAI-compatible API - including models you're running locally on your own machine. If it speaks the standard chat-completions format, Omnesly can talk to it.

Add a custom endpoint

base_url: https://your-endpoint/v1
api_key: ••••••••••••••••
protocol: openai-compatible

New providers, added as they matter.

The model landscape moves fast. We add support for new providers and model families on an ongoing basis - no app update required to start using a newly connected key.