Free-LLMs
EN

Бесплатные API нейросетей

Обновлено:

Актуальный каталог бесплатных ИИ API: DeepSeek V4, Qwen 3.5, Nemotron 3 и 4, Mistral, Llama, Gemini и другие модели. Сравните провайдеров, контекст и лимиты перед подключением к Cursor, VS Code, Cline, Open WebUI или собственному приложению.

Подробнее

Каталог обновляется автоматически из API агрегаторов и inference-платформ. Бесплатным считается предложение с нулевой стоимостью токенов, developer free tier или явно указанной бесплатной квотой. Одинаковая модель у разных провайдеров показывается отдельными строками, поскольку условия доступа различаются.

Найдено 315 моделей

Мощность
Модель Мощность Провайдер Параметры Контекст Лимит
📌 OpenCode GO

60$ лимитов на API за 10$ в месяц. Лучшие китайские модели: DeepSeek v4 Pro/Flash, QWEN 3.7 MAX, GLM 5.1 и др. Первый месяц 5$ и 5$ бонус на второй месяц.

Открыть
1 Union Alpha OpenRouter ? 262K Бесплатный тариф Открыть
2 inclusionAI: Ling 3.0 Flash VL (free) OpenRouter ? 262K Бесплатный тариф Открыть
3 Nex AGI: Nex-N2.5-Mini (free) OpenRouter ? 262K Бесплатный тариф Открыть
4 Nex AGI: Nex-N2.5-Pro (free) 🚀 OpenRouter ? 262K Бесплатный тариф Открыть
5 inclusionAI: Ling 3.0 Flash Sante (free) OpenRouter ? 262K Бесплатный тариф Открыть
6 inclusionAI: Ling 3.0 Flash Fin (free) OpenRouter ? 262K Бесплатный тариф Открыть
7 Dots Studio: Dots3-Note Preview (free) OpenRouter ? 512K Бесплатный тариф Открыть
8 LiquidAI: LFM2.5-2.6B (free) OpenRouter 2.6B 66K Бесплатный тариф Открыть
9 NVIDIA: Nemotron 3.5 Lightning (free) ⚖️ OpenRouter ? 1.0M Бесплатный тариф Открыть
10 Thinking Machines: Inkling Small (free) OpenRouter ? 1.0M Бесплатный тариф Открыть
11 Poolside: Laguna S 2.1 (free) OpenRouter ? 262K Бесплатный тариф Открыть
12 Thinking Machines: Inkling (free) OpenRouter ? 1.0M Бесплатный тариф Открыть
13 Poolside: Laguna XS 2.1 (free) OpenRouter ? 262K Бесплатный тариф Открыть
14 Cohere: North Mini Code (free) OpenRouter ? 256K Бесплатный тариф Открыть
15 Z.ai: GLM 5.2 (free) 🚀 OpenRouter ? 33K Бесплатный тариф Открыть
16 NVIDIA: Nemotron 3.5 Content Safety (free) ⚖️ OpenRouter ? 128K Бесплатный тариф Открыть
17 NVIDIA: Nemotron 3 Ultra (free) 🚀 OpenRouter 550B 1.0M Бесплатный тариф Открыть
18 NVIDIA: Nemotron 3 Nano Omni (free) ⚖️ OpenRouter 30B 256K Бесплатный тариф Открыть
19 Google: Gemma 4 26B A4B (free) ⚖️ OpenRouter 26B 262K Бесплатный тариф Открыть
20 Google: Gemma 4 31B (free) ⚖️ OpenRouter 31B 262K Бесплатный тариф Открыть
21 Google: Lyria 3 Pro Preview 🚀 OpenRouter ? 1.0M Бесплатный тариф Открыть
22 Google: Lyria 3 Clip Preview OpenRouter ? 1.0M Бесплатный тариф Открыть
23 NVIDIA: Nemotron 3 Super (free) 🚀 OpenRouter 120B 262K Бесплатный тариф Открыть
24 Free Models Router OpenRouter ? 200K Бесплатный тариф Открыть
25 LLaMA 3.3 70B Versatile 🚀 Groq 70B 128K 1 000 req/day Открыть
26 LLaMA 3.1 8B Instant Groq 8B 128K 14 400 req/day Открыть
27 Gemma 2 9B IT Groq 9B 8K 14 400 req/day Открыть
28 DeepSeek R1 Distill LLaMA 70B 🚀 Groq 70B 128K 1 000 req/day Открыть
29 Mistral Saba 24B ⚖️ Groq 24B 32K 1 000 req/day Открыть
30 Gemini 2.0 Flash ⚖️ Google Gemini ? 1.0M 1 500 req/day Открыть
31 Gemini 2.0 Flash Lite ⚖️ Google Gemini ? 1.0M 1 500 req/day Открыть
32 Gemini 1.5 Flash ⚖️ Google Gemini ? 1.0M 1 500 req/day Открыть
33 Gemini 1.5 Pro 🚀 Google Gemini ? 2.1M 50 req/day Открыть
34 Mistral Small ⚖️ Mistral AI ? 33K 1 req/min · 500 tok/min Открыть
35 Mistral Nemo ⚖️ Mistral AI ? 131K 1 req/min · 500 tok/min Открыть
36 Codestral Mamba Mistral AI ? 256K 1 req/min · 500 tok/min Открыть
37 Mistral Embed ⚖️ Mistral AI ? 8K 1 req/min · 500 tok/min Открыть
38 Ternary Bonsai 27B ⚖️ Together AI 27B 262K Fair Use (По мере нагрузки серверов) Открыть
39 nim/meta/llama-3.2-11b-vision-instruct Together AI 11B 16K Fair Use (По мере нагрузки серверов) Открыть
40 nim/meta/llama-3.2-90b-vision-instruct 🚀 Together AI 90B 16K Fair Use (По мере нагрузки серверов) Открыть
41 nim/mistralai/mixtral-8x22b-instruct-v01 ⚖️ Together AI ? 16K Fair Use (По мере нагрузки серверов) Открыть
42 nim/meta/llama-3.3-70b-instruct 🚀 Together AI 70B 16K Fair Use (По мере нагрузки серверов) Открыть
43 nim/nvidia/llama-3.1-nemotron-70b-instruct 🚀 Together AI 70B 16K Fair Use (По мере нагрузки серверов) Открыть
44 nim/meta/llama-3.1-8b-instruct Together AI 8B 16K Fair Use (По мере нагрузки серверов) Открыть
45 nim/meta/llama-3.1-70b-instruct 🚀 Together AI 70B 16K Fair Use (По мере нагрузки серверов) Открыть
46 nim/nv-mistralai/mistral-nemo-12b-instruct ⚖️ Together AI 12B 16K Fair Use (По мере нагрузки серверов) Открыть
47 nim/mistralai/mixtral-8x7b-instruct-v01 ⚖️ Together AI ? 16K Fair Use (По мере нагрузки серверов) Открыть
48 nim/nvidia/llama-3.3-nemotron-super-49b-v1 ⚖️ Together AI 49B 16K Fair Use (По мере нагрузки серверов) Открыть
49 Gemma 3 27B It ⚖️ Together AI 27B 66K Fair Use (По мере нагрузки серверов) Открыть
50 meta-llama/Llama-2-7b-chat-hf Together AI 7B 4K Fair Use (По мере нагрузки серверов) Открыть
51 DeepSeek R1 Distill Qwen 7B Together AI 7B 131K Fair Use (По мере нагрузки серверов) Открыть
52 Gemma 3 1b it Together AI 1B 33K Fair Use (По мере нагрузки серверов) Открыть
53 Gemma 3 4b it Together AI 4B 66K Fair Use (По мере нагрузки серверов) Открыть
54 Cogito V1 Preview Llama 8B Together AI 8B 131K Fair Use (По мере нагрузки серверов) Открыть
55 Cogito V1 Preview Qwen 32B ⚖️ Together AI 32B 131K Fair Use (По мере нагрузки серверов) Открыть
56 Cogito V1 Preview Qwen 14B ⚖️ Together AI 14B 131K Fair Use (По мере нагрузки серверов) Открыть
57 Cogito V1 Preview Llama 70B 🚀 Together AI 70B 131K Fair Use (По мере нагрузки серверов) Открыть
58 Cogito V1 Preview Llama 70B Turbo 🚀 Together AI 70B 131K Fair Use (По мере нагрузки серверов) Открыть
59 Meta Llama 3.3 70B Instruct 🚀 Together AI 70B 131K Fair Use (По мере нагрузки серверов) Открыть
60 Qwen2.5 32B ⚖️ Together AI 32B 131K Fair Use (По мере нагрузки серверов) Открыть
61 Qwen2.5 72B 🚀 Together AI 72B 131K Fair Use (По мере нагрузки серверов) Открыть
62 Qwen2.5 3B Instruct Together AI 3B 33K Fair Use (По мере нагрузки серверов) Открыть
63 Qwen2.5 1.5B Instruct Together AI 1.5B 33K Fair Use (По мере нагрузки серверов) Открыть
64 Qwen2.5 14B ⚖️ Together AI 14B 131K Fair Use (По мере нагрузки серверов) Открыть
65 Qwen2.5 7B Together AI 7B 131K Fair Use (По мере нагрузки серверов) Открыть
66 Qwen2.5 1.5B Together AI 1.5B 131K Fair Use (По мере нагрузки серверов) Открыть
67 Llama 3.1 70B 🚀 Together AI 70B 131K Fair Use (По мере нагрузки серверов) Открыть
68 Llama 3.2 1B Together AI 1B 131K Fair Use (По мере нагрузки серверов) Открыть
69 Qwen2.5 7B Instruct Together AI 7B 33K Fair Use (По мере нагрузки серверов) Открыть
70 Qwen2.5 32B Instruct ⚖️ Together AI 32B 33K Fair Use (По мере нагрузки серверов) Открыть
71 Llama 3.1 405B 🚀 Together AI 405B 131K Fair Use (По мере нагрузки серверов) Открыть
72 Deepcoder 14B Preview ⚖️ Together AI 14B 131K Fair Use (По мере нагрузки серверов) Открыть
73 Mistral 7B v0.1 Together AI 7B 33K Fair Use (По мере нагрузки серверов) Открыть
74 Devstral Small 2505 ⚖️ Together AI ? 131K Fair Use (По мере нагрузки серверов) Открыть
75 Mixtral 8X22b Instruct V0.1 ⚖️ Together AI ? 66K Fair Use (По мере нагрузки серверов) Открыть
76 Molmo 7B D 0924 Together AI 7B 4K Fair Use (По мере нагрузки серверов) Открыть
77 Qwen3 8B Together AI 8B 41K Fair Use (По мере нагрузки серверов) Открыть
78 Qwen3 14B ⚖️ Together AI 14B 2K Fair Use (По мере нагрузки серверов) Открыть
79 Qwen3 0.6B Together AI 0.6B 41K Fair Use (По мере нагрузки серверов) Открыть
80 Qwen3 1.7B Together AI 1.7B 41K Fair Use (По мере нагрузки серверов) Открыть
81 Qwen3 30B A3b ⚖️ Together AI 30B 41K Fair Use (По мере нагрузки серверов) Открыть
82 Gemma 2B It Together AI 2B 8K Fair Use (По мере нагрузки серверов) Открыть
83 Gemma 2 9B It Together AI 9B 8K Fair Use (По мере нагрузки серверов) Открыть
84 Llama 4 Scout (17Bx16E) ⚖️ Together AI 17B 262K Fair Use (По мере нагрузки серверов) Открыть
85 Magistral Small 2506 ⚖️ Together AI ? 41K Fair Use (По мере нагрузки серверов) Открыть
86 Minimax M1 40K ⚖️ Together AI ? 1.0M Fair Use (По мере нагрузки серверов) Открыть
87 Minimax M1 80K ⚖️ Together AI ? 1.0M Fair Use (По мере нагрузки серверов) Открыть
88 Meta Llama 3.1 8B Instruct Awq Int4 Together AI 8B 131K Fair Use (По мере нагрузки серверов) Открыть
89 Sarvam M Together AI ? 33K Fair Use (По мере нагрузки серверов) Открыть
90 Qwen3 32B ⚖️ Together AI 32B 41K Fair Use (По мере нагрузки серверов) Открыть
91 Qwen3 235B A22b Instruct 2507 Fp8 🚀 Together AI 235B 262K Fair Use (По мере нагрузки серверов) Открыть
92 Qwen3 Coder 30B A3b Instruct ⚖️ Together AI 30B 262K Fair Use (По мере нагрузки серверов) Открыть
93 Qwen3 4B Instruct 2507 Together AI 4B 262K Fair Use (По мере нагрузки серверов) Открыть
94 GLM 4.5V 🚀 Together AI ? 66K Fair Use (По мере нагрузки серверов) Открыть
95 Qwen3 Next 80B A3b Instruct Fp8 🚀 Together AI 80B ? Fair Use (По мере нагрузки серверов) Открыть
96 Gemma 3 270M It Together AI ? 33K Fair Use (По мере нагрузки серверов) Открыть
97 Medgemma 27B Text It ⚖️ Together AI 27B 131K Fair Use (По мере нагрузки серверов) Открыть
98 MiniMax M2 🚀 Together AI ? 197K Fair Use (По мере нагрузки серверов) Открыть
99 Qwen3-VL-235B-A22B-Instruct-FP8 🚀 Together AI 235B 262K Fair Use (По мере нагрузки серверов) Открыть
100 EssentialAI Rnj-1 Instruct Together AI ? 33K Fair Use (По мере нагрузки серверов) Открыть
101 Nvidia Nemotron 3 Nano 30B A3b Bf16 ⚖️ Together AI 30B 262K Fair Use (По мере нагрузки серверов) Открыть
102 GLM 4.7 FP4 🚀 Together AI ? 203K Fair Use (По мере нагрузки серверов) Открыть
103 GLM 5 Fp4 🚀 Together AI ? 203K Fair Use (По мере нагрузки серверов) Открыть
104 MiniMax M2.5 FP4 🚀 Together AI ? 8K Fair Use (По мере нагрузки серверов) Открыть
105 Glm 4.7 Fp8 🚀 Together AI ? 203K Fair Use (По мере нагрузки серверов) Открыть
106 Qwen3.5 9B Fp8 Together AI 9B 262K Fair Use (По мере нагрузки серверов) Открыть
107 Qwen3.5 35B A3b ⚖️ Together AI 35B 262K Fair Use (По мере нагрузки серверов) Открыть
108 Nvidia Nemotron 3 Super 120B A12b Fp8 🚀 Together AI 120B 262K Fair Use (По мере нагрузки серверов) Открыть
109 Deepseek OCR 2 ⚖️ Together AI ? 8K Fair Use (По мере нагрузки серверов) Открыть
110 Qwen3.5 122B A10b Fp8 🚀 Together AI 122B 262K Fair Use (По мере нагрузки серверов) Открыть
111 GLM OCR 🚀 Together AI ? 131K Fair Use (По мере нагрузки серверов) Открыть
112 Qwen3 8B Lora Together AI 8B 41K Fair Use (По мере нагрузки серверов) Открыть
113 Qwen3 30B A3B Instruct 2507 Lora ⚖️ Together AI 30B 262K Fair Use (По мере нагрузки серверов) Открыть
114 Holo3 35B A3b ⚖️ Together AI 35B 262K Fair Use (По мере нагрузки серверов) Открыть
115 Nvidia Nemotron 3 Super 120B A12b Bf16 🚀 Together AI 120B 262K Fair Use (По мере нагрузки серверов) Открыть
116 Gemma 4 E4B-it Together AI ? 131K Fair Use (По мере нагрузки серверов) Открыть
117 Gemma 4 26B A4b It ⚖️ Together AI 26B 262K Fair Use (По мере нагрузки серверов) Открыть
118 Gemma 4 E2B-it Together AI ? 131K Fair Use (По мере нагрузки серверов) Открыть
119 Qwen3.6 35B A3b Fp8 ⚖️ Together AI 35B 262K Fair Use (По мере нагрузки серверов) Открыть
120 Nemotron 3 Nano Omni 30B A3b Reasoning Fp8 ⚖️ Together AI 30B 131K Fair Use (По мере нагрузки серверов) Открыть
121 Llama 3.3 70B Instruct FP8 Lora 🚀 Together AI 70B 131K Fair Use (По мере нагрузки серверов) Открыть
122 Mixtral 8x7B Instruct V0.1 FP8 Lora ⚖️ Together AI ? 33K Fair Use (По мере нагрузки серверов) Открыть
123 Gemma 3 270M It Lora Together AI ? 33K Fair Use (По мере нагрузки серверов) Открыть
124 Llama 4 Scout 17B 16E Instruct Fp8 Lora ⚖️ Together AI 17B 10.5M Fair Use (По мере нагрузки серверов) Открыть
125 Gemma 3 27B It Lora ⚖️ Together AI 27B ? Fair Use (По мере нагрузки серверов) Открыть
126 Gemma 4 31B It Lora ⚖️ Together AI 31B 262K Fair Use (По мере нагрузки серверов) Открыть
127 Llama 4 Maverick 17B 128E Instruct Nvfp4 ⚖️ Together AI 17B 1.0M Fair Use (По мере нагрузки серверов) Открыть
128 Qwen3.5 35B A3B Lora ⚖️ Together AI 35B 262K Fair Use (По мере нагрузки серверов) Открыть
129 Qwen3.5 2B Lora Together AI 2B 262K Fair Use (По мере нагрузки серверов) Открыть
130 Qwen3.6 35B A3B Lora ⚖️ Together AI 35B 262K Fair Use (По мере нагрузки серверов) Открыть
131 Gemma 4 12B It ⚖️ Together AI 12B 262K Fair Use (По мере нагрузки серверов) Открыть
132 GLM 5.2 FP8 🚀 Together AI ? 1.0M Fair Use (По мере нагрузки серверов) Открыть
133 GLM 5.2 FP8 Lora 🚀 Together AI ? 1.0M Fair Use (По мере нагрузки серверов) Открыть
134 GLM 5.3 FP8 🚀 Together AI ? 1.0M Fair Use (По мере нагрузки серверов) Открыть
135 GLM 5.3 FP8 Lora 🚀 Together AI ? 1.0M Fair Use (По мере нагрузки серверов) Открыть
136 Command R+ (Large) 🚀 Cohere ? 128K 1 000 req/day · 10 req/min Открыть
137 Command R ⚖️ Cohere ? 128K 1 000 req/day · 10 req/min Открыть
138 Command ⚖️ Cohere ? 4K 1 000 req/day · 10 req/min Открыть
139 Command Light ⚖️ Cohere ? 4K 1 000 req/day · 10 req/min Открыть
140 LLaMA 3.1 8B Cerebras 8B 128K 14 400 req/day · 30 req/min Открыть
141 LLaMA 3.1 70B 🚀 Cerebras 70B 128K 14 400 req/day · 30 req/min Открыть
142 Llama 3.1 8B Instruct Cloudflare 8B 131K 10 000 req/day Открыть
143 Mistral 7B v0.3 Cloudflare 7B 33K 10 000 req/day Открыть
144 Qwen 1.5 14B Chat ⚖️ Cloudflare 14B 33K 10 000 req/day Открыть
145 Llama 3.1 8B HuggingFace 8B 128K Fair Use (По мере нагрузки серверов) Открыть
146 Mistral 7B v0.3 HuggingFace 7B 33K Fair Use (По мере нагрузки серверов) Открыть
147 Phi-3 Mini HuggingFace ? 4K Fair Use (По мере нагрузки серверов) Открыть
148 orcarouter/free OrcaRouter ? ? Бесплатный тариф Открыть
149 orcarouter/fusion OrcaRouter ? 1.0M Бесплатный тариф Открыть
150 orcarouter/fusion-flash OrcaRouter ? 262K Бесплатный тариф Открыть
151 orcarouter/fusion-mini OrcaRouter ? 1.0M Бесплатный тариф Открыть
152 DeepSeek: DeepSeek V4 Flash (Free) 🚀 OrcaRouter ? ? Бесплатный тариф Открыть
153 Google: Nano Banana Pro (Gemini 3 Pro Image Preview) 🚀 OrcaRouter ? 66K Бесплатный тариф Открыть
154 Google: Nano Banana 2 (Gemini 3.1 Flash Image Preview) ⚖️ OrcaRouter ? 66K Бесплатный тариф Открыть
155 grok/grok-imagine-image 🚀 OrcaRouter ? ? Бесплатный тариф Открыть
156 Kling: Kling 3.0 Turbo OrcaRouter ? ? Бесплатный тариф Открыть
157 OrcaDub: OrcaDub 1.0 OrcaRouter ? ? Бесплатный тариф Открыть
158 Orca: OrcaVerify Text 1.0 (Free) OrcaRouter ? ? Бесплатный тариф Открыть
159 Union Alpha (Free) OrcaRouter ? ? Бесплатный тариф Открыть
160 Tencent: Hy3 (Free) OrcaRouter ? ? Бесплатный тариф Открыть
161 Z.ai: GLM 5.3 Flash (Free) ⚖️ OrcaRouter ? ? Бесплатный тариф Открыть
162 Codegemma 1.1 7b NVIDIA Build 7B ? Free serverless inference for development; limits and model availability may change. Открыть
163 Codegemma 7b NVIDIA Build 7B ? Free serverless inference for development; limits and model availability may change. Открыть
164 Codellama 70b 🚀 NVIDIA Build 70B ? Free serverless inference for development; limits and model availability may change. Открыть
165 Codestral 22b Instruct V0.1 ⚖️ NVIDIA Build 22B ? Free serverless inference for development; limits and model availability may change. Открыть
166 Cosmos Reason2 8b NVIDIA Build 8B ? Free serverless inference for development; limits and model availability may change. Открыть
167 Dbrx Instruct NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
168 Deepseek Coder 6.7b Instruct NVIDIA Build 6.7B ? Free serverless inference for development; limits and model availability may change. Открыть
169 Deepseek V4 Flash 0731 🚀 NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
170 Fuyu 8b NVIDIA Build 8B ? Free serverless inference for development; limits and model availability may change. Открыть
171 Gemma 2b NVIDIA Build 2B ? Free serverless inference for development; limits and model availability may change. Открыть
172 Gemma 3 12b It ⚖️ NVIDIA Build 12B ? Free serverless inference for development; limits and model availability may change. Открыть
173 Gemma 3 4b It NVIDIA Build 4B ? Free serverless inference for development; limits and model availability may change. Открыть
174 Gemma 4 31b It ⚖️ NVIDIA Build 31B ? Free serverless inference for development; limits and model availability may change. Открыть
175 Glm 5.3 🚀 NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
176 Glm 5.3 Flash ⚖️ NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
177 GPT Oss 20b ⚖️ NVIDIA Build 20B ? Free serverless inference for development; limits and model availability may change. Открыть
178 Granite 3.0 3b A800m Instruct NVIDIA Build 3B ? Free serverless inference for development; limits and model availability may change. Открыть
179 Granite 3.0 8b Instruct NVIDIA Build 8B ? Free serverless inference for development; limits and model availability may change. Открыть
180 Granite 34b Code Instruct ⚖️ NVIDIA Build 34B ? Free serverless inference for development; limits and model availability may change. Открыть
181 Granite 8b Code Instruct NVIDIA Build 8B ? Free serverless inference for development; limits and model availability may change. Открыть
182 Jamba 1.5 Large Instruct 🚀 NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
183 Kimi K2.6 🚀 NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
184 Kimi K3 ⚖️ NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
185 Kosmos 2 NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
186 Laguna Xs 2.1 NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
187 LLAMA 3.1 Nemotron 51b Instruct ⚖️ NVIDIA Build 51B ? Free serverless inference for development; limits and model availability may change. Открыть
188 LLAMA 3.1 Nemotron 70b Instruct 🚀 NVIDIA Build 70B ? Free serverless inference for development; limits and model availability may change. Открыть
189 LLAMA 3.1 Nemotron Ultra 253b V1 🚀 NVIDIA Build 253B ? Free serverless inference for development; limits and model availability may change. Открыть
190 LLAMA 3.2 11b Vision Instruct NVIDIA Build 11B ? Free serverless inference for development; limits and model availability may change. Открыть
191 LLAMA 3.2 90b Vision Instruct 🚀 NVIDIA Build 90B ? Free serverless inference for development; limits and model availability may change. Открыть
192 Llama2 70b 🚀 NVIDIA Build 70B ? Free serverless inference for development; limits and model availability may change. Открыть
193 Llama3 Chatqa 1.5 70b 🚀 NVIDIA Build 70B ? Free serverless inference for development; limits and model availability may change. Открыть
194 Mistral 7b Instruct V0.3 NVIDIA Build 7B ? Free serverless inference for development; limits and model availability may change. Открыть
195 Mistral Large 🚀 NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
196 Mistral Large 2 Instruct 🚀 NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
197 Mistral Nemo 12b Instruct ⚖️ NVIDIA Build 12B ? Free serverless inference for development; limits and model availability may change. Открыть
198 Mistral Nemo Minitron 8b 8k Instruct NVIDIA Build 8B 8K Free serverless inference for development; limits and model availability may change. Открыть
199 Mistral Nemotron ⚖️ NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
200 Mixtral 8x22b V0.1 ⚖️ NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
201 Muse Glimmer 30b ⚖️ NVIDIA Build 30B ? Free serverless inference for development; limits and model availability may change. Открыть
202 Nemotron 3 Nano Omni 30b A3b Reasoning ⚖️ NVIDIA Build 30B ? Free serverless inference for development; limits and model availability may change. Открыть
203 Nemotron 3 Super 120b A12b 🚀 NVIDIA Build 120B ? Free serverless inference for development; limits and model availability may change. Открыть
204 Nemotron 3 Ultra 550b A55b 🚀 NVIDIA Build 550B 1.0M Free serverless inference for development; limits and model availability may change. Открыть
205 Nemotron 3.5 Lightning 30b A3b ⚖️ NVIDIA Build 30B ? Free serverless inference for development; limits and model availability may change. Открыть
206 Nemotron 4 340b Instruct 🚀 NVIDIA Build 340B ? Free serverless inference for development; limits and model availability may change. Открыть
207 Nemotron Nano 3 30b A3b ⚖️ NVIDIA Build 30B ? Free serverless inference for development; limits and model availability may change. Открыть
208 Neva 22b ⚖️ NVIDIA Build 22B ? Free serverless inference for development; limits and model availability may change. Открыть
209 Palmyra Creative 122b 🚀 NVIDIA Build 122B ? Free serverless inference for development; limits and model availability may change. Открыть
210 Palmyra Fin 70b 32k 🚀 NVIDIA Build 70B 33K Free serverless inference for development; limits and model availability may change. Открыть
211 Palmyra Med 70b 🚀 NVIDIA Build 70B ? Free serverless inference for development; limits and model availability may change. Открыть
212 Palmyra Med 70b 32k 🚀 NVIDIA Build 70B 33K Free serverless inference for development; limits and model availability may change. Открыть
213 Phi 3 Vision 128k Instruct ⚖️ NVIDIA Build ? 131K Free serverless inference for development; limits and model availability may change. Открыть
214 Phi 3.5 Moe Instruct ⚖️ NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
215 Recurrentgemma 2b NVIDIA Build 2B ? Free serverless inference for development; limits and model availability may change. Открыть
216 Riva Translate 4b Instruct NVIDIA Build 4B ? Free serverless inference for development; limits and model availability may change. Открыть
217 Riva Translate 4b Instruct V1.1 NVIDIA Build 4B ? Free serverless inference for development; limits and model availability may change. Открыть
218 Riva Translate 4b Instruct V2 NVIDIA Build 4B ? Free serverless inference for development; limits and model availability may change. Открыть
219 Sea Lion 7b Instruct NVIDIA Build 7B ? Free serverless inference for development; limits and model availability may change. Открыть
220 Starcoder2 15b ⚖️ NVIDIA Build 15B ? Free serverless inference for development; limits and model availability may change. Открыть
221 Vila NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
222 Yi Large 🚀 NVIDIA Build ? ? Free serverless inference for development; limits and model availability may change. Открыть
223 Zamba2 7b Instruct NVIDIA Build 7B ? Free serverless inference for development; limits and model availability may change. Открыть
224 Nemotron 3 Nano Omni 30b A3b Reasoning:Free ⚖️ TokenRouter 30B ? Free models and promotional access may be time-limited; availability is checked through the Models API. Открыть
225 Laguna S 2.1 Free Vercel AI Gateway ? 256K A valid payment card must be on file even for zero-price models. Free promotions may change; availability is refreshed from the Models API. Открыть
226 Ling 3.0 Flash Fin Vercel AI Gateway ? 256K A valid payment card must be on file even for zero-price models. Free promotions may change; availability is refreshed from the Models API. Открыть
227 Ling 3.0 Flash Fin (Free) Vercel AI Gateway ? 256K A valid payment card must be on file even for zero-price models. Free promotions may change; availability is refreshed from the Models API. Открыть
228 Ling 3.0 Flash Sante Vercel AI Gateway ? 256K A valid payment card must be on file even for zero-price models. Free promotions may change; availability is refreshed from the Models API. Открыть
229 Ling 3.0 Flash Sante (Free) Vercel AI Gateway ? 256K A valid payment card must be on file even for zero-price models. Free promotions may change; availability is refreshed from the Models API. Открыть
230 Ling 3.0 Flash VL Vercel AI Gateway ? 256K A valid payment card must be on file even for zero-price models. Free promotions may change; availability is refreshed from the Models API. Открыть
231 Ling 3.0 Flash VL (Free) Vercel AI Gateway ? 256K A valid payment card must be on file even for zero-price models. Free promotions may change; availability is refreshed from the Models API. Открыть
232 inclusionAI: Ling 3.0 Flash Fin Nous Portal ? 262K The $0 Portal plan includes free models with standard rate limits. Model availability can change. Открыть
233 inclusionAI: Ling 3.0 Flash Sante (free) Nous Portal ? 262K The $0 Portal plan includes free models with standard rate limits. Model availability can change. Открыть
234 Meituan: LongCat 2.0 Nous Portal ? 1.0M The $0 Portal plan includes free models with standard rate limits. Model availability can change. Открыть
235 Poolside: Laguna S 2.1 Nous Portal ? 262K The $0 Portal plan includes free models with standard rate limits. Model availability can change. Открыть
236 Poolside: Laguna XS 2.1 Nous Portal ? 262K The $0 Portal plan includes free models with standard rate limits. Model availability can change. Открыть
237 StepFun: Step 3.7 Flash Nous Portal ? 262K The $0 Portal plan includes free models with standard rate limits. Model availability can change. Открыть
238 Upstage: Solar Pro 4 🚀 Nous Portal ? 524K The $0 Portal plan includes free models with standard rate limits. Model availability can change. Открыть
239 Agnes 2.0 Flash UnoRouter ? 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
240 Allam 2 7b UnoRouter 7B 4K No-card free models use shared capacity and per-model limits; availability can change. Открыть
241 C4ai Aya Expanse 32b ⚖️ UnoRouter 32B 128K No-card free models use shared capacity and per-model limits; availability can change. Открыть
242 Codestral Latest UnoRouter ? 256K No-card free models use shared capacity and per-model limits; availability can change. Открыть
243 Deepseek R1 Distill Llama 70b 🚀 UnoRouter 70B 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
244 Diffusiongemma 26b A4b It ⚖️ UnoRouter 26B 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
245 Dots 3 Note Preview UnoRouter ? 512K No-card free models use shared capacity and per-model limits; availability can change. Открыть
246 Gemini 3.5 Flash Lite ⚖️ UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
247 Gemini 3.6 Flash ⚖️ UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
248 Gemini Flash Lite Latest ⚖️ UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
249 Gemini Robotics Er 2 Preview ⚖️ UnoRouter ? 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
250 Gemma 4 26b ⚖️ UnoRouter 26B 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
251 Gemma 4 31b It ⚖️ UnoRouter 31B 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
252 GLM 4.5 Flash ⚖️ UnoRouter ? 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
253 GLM 5.2 🚀 UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
254 GLM 5.2 Search 🚀 UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
255 GLM 5.2 Think Search 🚀 UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
256 GLM 5.2 Thinking 🚀 UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
257 GLM 5.3 🚀 UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
258 GLM 5.3 Flash Search ⚖️ UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
259 GLM 5.3 Flash Think Search ⚖️ UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
260 GLM 5.3 Flash Thinking ⚖️ UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
261 GLM 5.3 Search 🚀 UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
262 GLM 5.3 Think Search 🚀 UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
263 GLM 5.3 Thinking 🚀 UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
264 GPT Oss 20b ⚖️ UnoRouter 20B 128K No-card free models use shared capacity and per-model limits; availability can change. Открыть
265 GPT Oss Safeguard 20b ⚖️ UnoRouter 20B 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
266 Ising Calibration 1.5 31b ⚖️ UnoRouter 31B 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
267 Kat Coder Air V1 UnoRouter ? 128K No-card free models use shared capacity and per-model limits; availability can change. Открыть
268 L3 70b Euryale V2.1 🚀 UnoRouter 70B 8K No-card free models use shared capacity and per-model limits; availability can change. Открыть
269 L3 8b Stheno V3.2 UnoRouter 8B 8K No-card free models use shared capacity and per-model limits; availability can change. Открыть
270 L3.3 Ms Nevoria 70b 🚀 UnoRouter 70B 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
271 Laguna S 2.1 UnoRouter ? 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
272 Laguna Xs 2.1 UnoRouter ? 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
273 Leanstral 1 5 UnoRouter ? 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
274 Lfm 2.5 2.6b UnoRouter 2.6B 128K No-card free models use shared capacity and per-model limits; availability can change. Открыть
275 Ling 3.0 Flash Fin UnoRouter ? 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
276 Ling 3.0 Flash Sante UnoRouter ? 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
277 Ling 3.0 Flash Vl UnoRouter ? 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
278 Llama 3.2 11b Vision UnoRouter 11B 128K No-card free models use shared capacity and per-model limits; availability can change. Открыть
279 Llama 3.2 90b Vision 🚀 UnoRouter 90B 128K No-card free models use shared capacity and per-model limits; availability can change. Открыть
280 Llama 4 Maverick 17b 128e Instruct ⚖️ UnoRouter 17B 128K No-card free models use shared capacity and per-model limits; availability can change. Открыть
281 Manta Flash 1.0 UnoRouter ? 16K No-card free models use shared capacity and per-model limits; availability can change. Открыть
282 Manta Mini 1.0 UnoRouter ? 8K No-card free models use shared capacity and per-model limits; availability can change. Открыть
283 Mistral Large 3 675b 🚀 UnoRouter 675B 256K No-card free models use shared capacity and per-model limits; availability can change. Открыть
284 Mistral Nemotron ⚖️ UnoRouter ? 128K No-card free models use shared capacity and per-model limits; availability can change. Открыть
285 Mistral Small 3.2 ⚖️ UnoRouter ? 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
286 Mn Violet Lotus 12b ⚖️ UnoRouter 12B 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
287 Muse Glimmer 30b ⚖️ UnoRouter 30B 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
288 Nemotron 3 Nano Omni 30b A3b Reasoning ⚖️ UnoRouter 30B 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
289 Nemotron 3 Super 120b A12b 🚀 UnoRouter 120B 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
290 Nemotron 3 Ultra 550b A55b 🚀 UnoRouter 550B 1.0M No-card free models use shared capacity and per-model limits; availability can change. Открыть
291 Nemotron 3.5 Lightning ⚖️ UnoRouter ? 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
292 Nemotron 3.5 Lightning 30b A3b ⚖️ UnoRouter 30B 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
293 Nemotron Nano 12b V2 Vl ⚖️ UnoRouter 12B 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
294 Nemotron Nano 9b V2 UnoRouter 9B 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
295 Nex N2.5 Mini UnoRouter ? 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
296 Nex N2.5 Pro 🚀 UnoRouter ? 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
297 North Mini Code UnoRouter ? 256K No-card free models use shared capacity and per-model limits; availability can change. Открыть
298 QWEN Sea Lion V4 32b It ⚖️ UnoRouter 32B 33K No-card free models use shared capacity and per-model limits; availability can change. Открыть
299 QWEN Sea Lion V4.5 27b It ⚖️ UnoRouter 27B 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
300 Qwen2.5 Vl 7b Instruct Awq UnoRouter 7B 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
301 Qwen3 Next 80b A3b Instruct 🚀 UnoRouter 80B 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
302 Qwen3.5 122b A10b 🚀 UnoRouter 122B 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
303 Qwen3.5 4b UnoRouter 4B 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
304 Qwen3.6 35b A3b ⚖️ UnoRouter 35B 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
305 Qwen3.8 27b ⚖️ UnoRouter 27B 66K No-card free models use shared capacity and per-model limits; availability can change. Открыть
306 Riva Translate 4b Instruct V2 UnoRouter 4B 8K No-card free models use shared capacity and per-model limits; availability can change. Открыть
307 Sapphira L3.3 70b 0.1 🚀 UnoRouter 70B 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
308 Sarvam 30b ⚖️ UnoRouter 30B 66K No-card free models use shared capacity and per-model limits; availability can change. Открыть
309 Seed Oss 36b ⚖️ UnoRouter 36B 524K No-card free models use shared capacity and per-model limits; availability can change. Открыть
310 Sensenova 6.8 Flash Lite UnoRouter ? 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
311 Sonar UnoRouter ? 128K No-card free models use shared capacity and per-model limits; availability can change. Открыть
312 Step 3.7 Flash UnoRouter ? 256K No-card free models use shared capacity and per-model limits; availability can change. Открыть
313 Typhoon V2.5 30b A3b Instruct ⚖️ UnoRouter 30B 131K No-card free models use shared capacity and per-model limits; availability can change. Открыть
314 Union Alpha UnoRouter ? 262K No-card free models use shared capacity and per-model limits; availability can change. Открыть
315 Villanova 2b 2512 Preview Apnea Ft UnoRouter 2B 33K No-card free models use shared capacity and per-model limits; availability can change. Открыть