| Direct Tier 0 |
Google
Gemini Direct |
gemini-3.8-flash-high / medium / low |
Direct Google Gemini 3.8 Flash with dashboard-controlled High, Medium, or Low thinking
level for Paid accounts. |
| Direct Tier 0 |
OpenAI
Direct |
gpt-6-astra / sol / luna |
Direct OpenAI flagship models. Selectable in settings with instant failover priority.
|
| Direct Tier 0 |
Anthropic
Direct |
claude-opus-5 / sonnet-5 / haiku-4.5 |
Direct Anthropic Claude models. Selectable in settings with instant failover priority.
|
| Direct Tier 0 |
DeepSeek
Direct |
deepseek-v4-pro / flash |
Direct DeepSeek models. Selectable in settings with instant failover priority. |
| Direct Tier 0 |
Kimi
AI Direct |
kimi-k3 / kimi-k2.7-code / highspeed |
Direct Kimi AI (Moonshot) 1M context deep reasoning & specialized coding engines.
Selectable in settings with instant failover priority. |
| Direct Tier 0 |
Qwen
Direct |
qwen3.8-max / qwen3.8-flash / qwen3.7-plus |
Direct Qwen Dashscope models with custom Base URL. Selectable in settings with instant
failover priority. |
| Direct Tier 0 |
Custom
Direct |
User-configured Model ID Example: gemma-4-26b-a4b
|
Direct custom OpenAI-compatible endpoint with custom Base URL, API Key, and Model ID
configured in HyperRouter Settings. |
| Direct Tier 0 |
Cloudflare
Workers AI Direct |
User-selected Workers AI Model ID Example:
@cf/qwen/qwen3.8-27b |
Direct Cloudflare Workers AI through the configured Account ID, API Key / API Token, and
free-form Model ID. The user-selected model remains unchanged across HIGH, MEDIUM, LOW,
and CUSTOM profiles and participates in Direct Provider Auto RollOver. |
| Tier 1 |
Google |
gemini-3.7-flash |
Flagship primary model supporting up to 1M context architecture (Token limits governed
by your Google API quota). |
| Tier 2 |
Google |
gemini-3.6-flash |
High-speed, high-capacity primary fallback model (1M Context). |
| Tier 3 |
Google |
gemma-4-31b-it |
Open weights high-intelligence reasoning model from Google. |
| Tier 4 |
Google |
gemma-4-26b-a4b-it |
Open weights specialized coding model from Google. |
| Tier 5 |
Google |
gemini-3.5-flash |
High-speed fallback model (1M Context). |
| Tier 6 |
Google |
gemini-3.5-flash-lite |
High-speed lightweight Gemini engine for rapid response (1M Context). |
| Tier 7 |
Google |
gemini-3-flash-preview |
Experimental preview model for advanced syntax analysis (1M Context). |
| Tier 8 |
Agnes
AI |
agnes-2.5-flash |
Ultra-large 512K context window & specialized coding buffer. |
| Tier 9 |
Groq LPU |
llama-3.3-70b-versatile |
Ultra-fast LPU reasoning with 70 billion parameters. |
| Tier 10 |
Groq LPU |
llama-3.1-8b-instant |
Instant low-latency sub-second response fallback layer. |
| Tier 11 |
Groq LPU |
qwen/qwen3.6-27b |
Multilingual & specialized multi-language coding model. |
| Tier 12 |
Groq LPU |
openai/gpt-oss-20b |
Emergency final standby model for continuous uptime. |