Siron.dev

Models

25 models from 12 providers, one OpenAI-compatible API. Switch models with a single string change — same key, same schema, transparent pricing.

Showing 25 of 25 models

Siron: Ox Alpha

Free

Free experimental stealth model available on Siron.dev while in preview. Rate-limited, no cost.

ChatText
4.31T tokens128K contextRp0/M inRp0/M outAug 15, 2026

Google: Gemini 3 Flash

FastDiscounted

Cheap, quick multimodal model with a 1M context window. Currently discounted for launch.

ChatVisionTextImageFileAudio
44.10B tokens1M contextRp1.200/M inRp4.800/M outJul 18, 2026

Anthropic: Claude Sonnet 4.8

Popular

The workhorse Claude: near-Opus quality at a fraction of the cost. Excellent for production agents.

ChatCodeReasoningTextImageFile
31.50B tokens200K contextRp12.000/M inRp60.000/M outAug 1, 2026

Meta: Llama 4 Scout

Open-weight workhorse with a huge context window at a friendly price. Solid general assistant baseline.

ChatMultilingualTextImage
27.30B tokens512K contextRp2.500/M inRp7.500/M outJan 22, 2026

OpenAI: GPT-5.1 Mini

Balanced quality and cost for everyday assistant and agent workloads. Strong instruction following at Mini pricing.

ChatReasoningTextImage
22.40B tokens256K contextRp4.000/M inRp16.000/M outApr 8, 2026

Anthropic: Claude Haiku 4.5

Fast

Fastest Claude tier for classification, extraction and lightweight chat at scale.

ChatTextImage
18.90B tokens200K contextRp3.000/M inRp15.000/M outFeb 11, 2026

DeepSeek: DeepSeek V4 Flash

Fast

Ultra-fast general chat model tuned for high throughput and low latency. Great default for assistants and high-volume pipelines.

ChatMultilingualText
15.90B tokens128K contextRp1.500/M inRp3.000/M outJun 12, 2026

Google: Gemini 3 Pro

Massive 1M-token context with native image, audio and document understanding. Great for whole-repo or long-video analysis.

VisionReasoningMultilingualTextImageFileAudioVideo
14.70B tokens1M contextRp18.000/M inRp72.000/M outMay 5, 2026

Meta: Llama 4 Maverick

Larger Llama 4 tier for reasoning-heavy open-weight deployments and long document synthesis.

ReasoningChatTextImage
12.80B tokens1M contextRp5.500/M inRp16.500/M outJan 22, 2026

OpenAI: GPT-5.1

Flagship general model with frontier reasoning, tool use and long-horizon agentic reliability.

ReasoningChatCodeTextImageFile
11.20B tokens400K contextRp20.000/M inRp80.000/M outApr 8, 2026

Moonshot: Kimi K3

Discounted

Agentic open-weight model with very long context and strong tool-use benchmarks at low cost.

ReasoningCodeChatTextFile
10.40B tokens2M contextRp3.200/M inRp12.800/M outJul 9, 2026

DeepSeek: DeepSeek R2

Popular

Chain-of-thought reasoning model for math, logic and hard multi-step problems. Emits explicit reasoning traces.

ReasoningCodeText
9.80B tokens128K contextRp8.000/M inRp24.000/M outMay 20, 2026

Alibaba: Qwen3 235B Instruct

Open-weight MoE instruct model with strong multilingual coverage, especially CJK and Southeast Asian languages.

ChatMultilingualText
8.45B tokens262K contextRp2.800/M inRp8.400/M outMar 30, 2026

Anthropic: Claude Opus 4.8

New

Frontier reasoning and coding with long-horizon agentic reliability. Best-in-class at multi-file refactors.

ReasoningCodeTextImageFile
7.60B tokens200K contextRp45.000/M inRp225.000/M outAug 1, 2026

Alibaba: Qwen3 Coder Max

Specialized coding model with strong repo-level and agentic performance across 90+ languages.

CodeReasoningText
6.06B tokens256K contextRp6.000/M inRp18.000/M outJun 28, 2026

Meta: Muse Spark 1.2

New

Community-tuned creative writing and roleplay model with a distinctive voice and long memory.

RoleplayChatText
6.06B tokens1.05M contextRp1.000/M inRp2.000/M outAug 4, 2026

Mistral: Mistral Large 3

European flagship with excellent function calling, JSON mode and low-latency streaming.

ChatCodeTextImage
5.20B tokens256K contextRp9.000/M inRp27.000/M outApr 25, 2026

Nous Research: Nous Hermes 5

Steerable open-weight assistant with minimal refusals, strong persona control and clean tool calling.

ChatRoleplayReasoningText
4.60B tokens262K contextRp2.000/M inRp6.000/M outMay 14, 2026

DeepSeek: DeepSeek V4 Flash Vision

Vision-enabled variant of V4 Flash. Reads screenshots, charts and documents while keeping Flash-tier latency.

VisionChatTextImageFile
4.31B tokens128K contextRp2.200/M inRp6.600/M outJul 2, 2026

xAI: Grok 5

Real-time-aware model with aggressive reasoning and a large context. Strong at current-events grounded answers.

ReasoningChatTextImage
3.90B tokens512K contextRp15.000/M inRp60.000/M outJun 2, 2026

OpenAI: Whisper v4

Speech-to-text transcription and translation across 100+ languages, billed per minute of audio.

MultilingualAudioText
2.10B tokens contextRp800/M inRp0/M outFeb 27, 2026

Black Forest Labs: FLUX.2 Pro

New

State-of-the-art text-to-image generation with strong prompt adherence and typography. Billed per output image.

Image GenTextImage
1.80B tokens contextRp0/M inRp12.000/M outAug 6, 2026

OpenAI: o5 Pro

Deliberate long-thinking model for research-grade math, proofs and complex planning. Higher latency by design.

ReasoningText
1.05B tokens256K contextRp90.000/M inRp360.000/M outMar 19, 2026

Google: Imagen 4 Ultra

Photorealistic image generation with fine control over lighting, lens and composition.

Image GenTextImage
900.0M tokens contextRp0/M inRp22.000/M outJun 17, 2026

Google: Veo 3.1

Text-and-image-to-video generation with synchronized audio, up to 20 second clips.

Image GenTextImageVideo
120.0M tokens contextRp0/M inRp480.000/M outJul 25, 2026