Models
25 models from 12 providers, one OpenAI-compatible API. Switch models with a single string change — same key, same schema, transparent pricing.
Showing 25 of 25 models
Siron: Ox Alpha
FreeFree experimental stealth model available on Siron.dev while in preview. Rate-limited, no cost.
Google: Gemini 3 Flash
FastDiscountedCheap, quick multimodal model with a 1M context window. Currently discounted for launch.
Anthropic: Claude Sonnet 4.8
PopularThe workhorse Claude: near-Opus quality at a fraction of the cost. Excellent for production agents.
Meta: Llama 4 Scout
Open-weight workhorse with a huge context window at a friendly price. Solid general assistant baseline.
OpenAI: GPT-5.1 Mini
Balanced quality and cost for everyday assistant and agent workloads. Strong instruction following at Mini pricing.
Anthropic: Claude Haiku 4.5
FastFastest Claude tier for classification, extraction and lightweight chat at scale.
DeepSeek: DeepSeek V4 Flash
FastUltra-fast general chat model tuned for high throughput and low latency. Great default for assistants and high-volume pipelines.
Google: Gemini 3 Pro
Massive 1M-token context with native image, audio and document understanding. Great for whole-repo or long-video analysis.
Meta: Llama 4 Maverick
Larger Llama 4 tier for reasoning-heavy open-weight deployments and long document synthesis.
OpenAI: GPT-5.1
Flagship general model with frontier reasoning, tool use and long-horizon agentic reliability.
Moonshot: Kimi K3
DiscountedAgentic open-weight model with very long context and strong tool-use benchmarks at low cost.
DeepSeek: DeepSeek R2
PopularChain-of-thought reasoning model for math, logic and hard multi-step problems. Emits explicit reasoning traces.
Alibaba: Qwen3 235B Instruct
Open-weight MoE instruct model with strong multilingual coverage, especially CJK and Southeast Asian languages.
Anthropic: Claude Opus 4.8
NewFrontier reasoning and coding with long-horizon agentic reliability. Best-in-class at multi-file refactors.
Alibaba: Qwen3 Coder Max
Specialized coding model with strong repo-level and agentic performance across 90+ languages.
Meta: Muse Spark 1.2
NewCommunity-tuned creative writing and roleplay model with a distinctive voice and long memory.
Mistral: Mistral Large 3
European flagship with excellent function calling, JSON mode and low-latency streaming.
Nous Research: Nous Hermes 5
Steerable open-weight assistant with minimal refusals, strong persona control and clean tool calling.
DeepSeek: DeepSeek V4 Flash Vision
Vision-enabled variant of V4 Flash. Reads screenshots, charts and documents while keeping Flash-tier latency.
xAI: Grok 5
Real-time-aware model with aggressive reasoning and a large context. Strong at current-events grounded answers.
OpenAI: Whisper v4
Speech-to-text transcription and translation across 100+ languages, billed per minute of audio.
Black Forest Labs: FLUX.2 Pro
NewState-of-the-art text-to-image generation with strong prompt adherence and typography. Billed per output image.
OpenAI: o5 Pro
Deliberate long-thinking model for research-grade math, proofs and complex planning. Higher latency by design.
Google: Imagen 4 Ultra
Photorealistic image generation with fine control over lighting, lens and composition.
Google: Veo 3.1
Text-and-image-to-video generation with synchronized audio, up to 20 second clips.
