Каталог моделей
599 моделей через один адрес. Цены в рублях, наценка 20 % показана в каждой строке.
МОДЕЛЬ
ТИП И ВОЗМОЖНОСТИ
КОНТЕКСТ
ВХОД ₽/1M
ВЫХОД ₽/1M
Ternary Bonsai 2 27B
prism-ml/ternary-bonsai-2-27b
262K
7,81 ₽
52,03 ₽
Bonsai 2 27B is a 27B-parameter reasoning model from PrismML derived from Qwen3.8-27B. It supports coding, mathematics, tool calling, and image understanding with a 262K-token context window. Ternary compression shrinks...
Ternary Bonsai 2 27B
prism-ml/ternary-bonsai-2-27b
Bonsai 2 27B is a 27B-parameter reasoning model from PrismML derived from Qwen3.8-27B. It supports coding, mathematics, tool calling, and image understanding with a 262K-token context window. Ternary compression shrinks...
вход / 1M
7,81 ₽
выход / 1M
52,03 ₽
GLM 5.3 FlashX
z-ai/glm-5.3-flashx
1.049M
38,51 ₽
130,09 ₽
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
GLM 5.3 FlashX
z-ai/glm-5.3-flashx
GLM-5.3-FlashX is the high-speed variant of Z.ai's GLM-5.3-Flash, a native multimodal model delivering inference speeds of up to 200 tokens/s. Built on the same hybrid sparse and linear attention architecture...
вход / 1M
38,51 ₽
выход / 1M
130,09 ₽
Jev Latest
скоро~typesafe/jev-latest
32K
4,37 ₽
0 ₽
This model always redirects to the latest model in the Jev family.
Jev Latest
скоро~typesafe/jev-latest
This model always redirects to the latest model in the Jev family.
вход / 1M
4,37 ₽
выход / 1M
0 ₽
Jev 1.13
скороtypesafe/jev-1.13
32K
4,37 ₽
0 ₽
Jev is a structured decision model from TypeSafe, and the first of its System One models. System One models make fast, structured decisions for software, returning a typed choice rather...
Jev 1.13
скороtypesafe/jev-1.13
Jev is a structured decision model from TypeSafe, and the first of its System One models. System One models make fast, structured decisions for software, returning a typed choice rather...
вход / 1M
4,37 ₽
выход / 1M
0 ₽
Pareto
unbiased/pareto
262K
260,17 ₽
780,51 ₽
Pareto is a multimodal composite model built for research, coding, and agentic workflows, while delivering frontier-level performance across a broad range of general-purpose tasks.
Pareto
unbiased/pareto
Pareto is a multimodal composite model built for research, coding, and agentic workflows, while delivering frontier-level performance across a broad range of general-purpose tasks.
вход / 1M
260,17 ₽
выход / 1M
780,51 ₽
DeepSeek Pro Latest
~deepseek/deepseek-pro-latest
1.049M
60,17 ₽
180,5 ₽
This model always redirects to the latest model in the DeepSeek Pro family.
DeepSeek Pro Latest
~deepseek/deepseek-pro-latest
This model always redirects to the latest model in the DeepSeek Pro family.
вход / 1M
60,17 ₽
выход / 1M
180,5 ₽
DeepSeek Flash Latest
~deepseek/deepseek-flash-latest
1.049M
14,05 ₽
56,2 ₽
This model always redirects to the latest model in the DeepSeek Flash family.
DeepSeek Flash Latest
~deepseek/deepseek-flash-latest
This model always redirects to the latest model in the DeepSeek Flash family.
вход / 1M
14,05 ₽
выход / 1M
56,2 ₽
Schematron V2 Turbo
inference-net/schematron-v2-turbo
128K
3,12 ₽
15,61 ₽
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
Schematron V2 Turbo
inference-net/schematron-v2-turbo
Schematron V2 Turbo is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes throughput for high-volume extraction workloads. Extraction instructions must be supplied through a JSON schema in response_format rather...
вход / 1M
3,12 ₽
выход / 1M
15,61 ₽
Schematron V2 Small
inference-net/schematron-v2-small
128K
5,2 ₽
23,94 ₽
Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...
Schematron V2 Small
inference-net/schematron-v2-small
Schematron V2 Small is a 3B-parameter HTML-to-JSON extraction model from Inference.net. It prioritizes extraction quality for complex schemas and long pages. Extraction instructions must be supplied through a JSON schema...
вход / 1M
5,2 ₽
выход / 1M
23,94 ₽
Muse Voice Transcribe 1.0
meta/muse-voice-transcribe-1.0
—
цена уточняется
Muse Voice Transcribe 1.0 is a synchronous speech-to-text model from Meta. It is suited for push-to-talk, endpointing, and speaker-aware transcription, with keyword biasing for domain terms and language biasing through...
Muse Voice Transcribe 1.0
meta/muse-voice-transcribe-1.0
Muse Voice Transcribe 1.0 is a synchronous speech-to-text model from Meta. It is suited for push-to-talk, endpointing, and speaker-aware transcription, with keyword biasing for domain terms and language biasing through...
цена уточняется
GPT Astra Latest
~openai/gpt-astra-latest
1.05M
1 040,68 ₽
5 203,41 ₽
This model always redirects to the latest model in the GPT Astra family.
GPT Astra Latest
~openai/gpt-astra-latest
This model always redirects to the latest model in the GPT Astra family.
вход / 1M
1 040,68 ₽
выход / 1M
5 203,41 ₽
GPT Sol Latest
~openai/gpt-sol-latest
1.05M
208,14 ₽
1 040,68 ₽
This model always redirects to the latest model in the GPT Sol family.
знания до 16 февраля 2026 г.GPT Sol Latest
~openai/gpt-sol-latest
This model always redirects to the latest model in the GPT Sol family.
знания до 16 февраля 2026 г.вход / 1M
208,14 ₽
выход / 1M
1 040,68 ₽
GPT Terra Latest
~openai/gpt-terra-latest
1.05M
208,14 ₽
1 248,82 ₽
This model always redirects to the latest model in the GPT Terra family.
знания до 16 февраля 2026 г.GPT Terra Latest
~openai/gpt-terra-latest
This model always redirects to the latest model in the GPT Terra family.
знания до 16 февраля 2026 г.вход / 1M
208,14 ₽
выход / 1M
1 248,82 ₽
GPT Luna Latest
~openai/gpt-luna-latest
1.05M
20,81 ₽
124,88 ₽
This model always redirects to the latest model in the GPT Luna family.
знания до 16 февраля 2026 г.GPT Luna Latest
~openai/gpt-luna-latest
This model always redirects to the latest model in the GPT Luna family.
знания до 16 февраля 2026 г.вход / 1M
20,81 ₽
выход / 1M
124,88 ₽
Fugu Ultra v2
sakana/fugu-ultra-v2
1M
520,34 ₽
3 122,04 ₽
Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...
знания до 28 августа 2026 г.Fugu Ultra v2
sakana/fugu-ultra-v2
Fugu Ultra v2 is the higher-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to...
знания до 28 августа 2026 г.вход / 1M
520,34 ₽
выход / 1M
3 122,04 ₽
Fugu Max
sakana/fugu-max
1M
208,14 ₽
624,41 ₽
Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
Fugu Max
sakana/fugu-max
Fugu Max is the cost-performance model in Sakana AI's Fugu family. Rather than a single monolithic model, Fugu is a learned multi-agent orchestration system: a language model trained to route...
вход / 1M
208,14 ₽
выход / 1M
624,41 ₽
FLUX Video Edit
black-forest-labs/flux-video-edit
—
3,12 ₽
за секунду
FLUX Video Edit [fast] takes a source video and an edit prompt and returns a precisely edited video. Add, remove, or replace objects and characters, rebuild the setting, edit on-screen...
FLUX Video Edit
black-forest-labs/flux-video-edit
FLUX Video Edit [fast] takes a source video and an edit prompt and returns a precisely edited video. Add, remove, or replace objects and characters, rebuild the setting, edit on-screen...
за секунду
3,12 ₽
Ling 3.0 Flash VL
inclusionai/ling-3.0-flash-vl
131K
6,24 ₽
18,73 ₽
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
Ling 3.0 Flash VL
inclusionai/ling-3.0-flash-vl
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
вход / 1M
6,24 ₽
выход / 1M
18,73 ₽
Ling 3.0 Flash VL (free)
inclusionai/ling-3.0-flash-vl:free
262K
бесплатно
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
Ling 3.0 Flash VL (free)
inclusionai/ling-3.0-flash-vl:free
Ling 3.0 Flash VL builds on Ling 3.0 Flash (124B total / 5.5B active MoE from InclusionAI), further strengthening its language capabilities while adding native visual perception and advanced visual...
бесплатно
DeepSeek V4.1 Flash
deepseek/deepseek-v4.1-flash
1.049M
15,61 ₽
62,44 ₽
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
DeepSeek V4.1 Flash
deepseek/deepseek-v4.1-flash
DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...
вход / 1M
15,61 ₽
выход / 1M
62,44 ₽
GPT Image 2.5 Sunburst
скороopenai/gpt-image-2.5-sunburst
—
3 122,04 ₽
за 1M токенов
GPT Image 2.5 Sunburst is an image generation and editing model from OpenAI, positioned as the precision-oriented tier of the GPT Image 2.5 series. It is suited to detailed creative...
GPT Image 2.5 Sunburst
скороopenai/gpt-image-2.5-sunburst
GPT Image 2.5 Sunburst is an image generation and editing model from OpenAI, positioned as the precision-oriented tier of the GPT Image 2.5 series. It is suited to detailed creative...
за 1M токенов
3 122,04 ₽
GPT Image 2.5 Flare
скороopenai/gpt-image-2.5-flare
—
3 122,04 ₽
за 1M токенов
GPT Image 2.5 Flare is an image generation and editing model from OpenAI, positioned as the speed-oriented tier of the GPT Image 2.5 series. It is suited to high-volume everyday...
GPT Image 2.5 Flare
скороopenai/gpt-image-2.5-flare
GPT Image 2.5 Flare is an image generation and editing model from OpenAI, positioned as the speed-oriented tier of the GPT Image 2.5 series. It is suited to high-volume everyday...
за 1M токенов
3 122,04 ₽
Mercury 2.5
inception/mercury-2.5
260K
4,16 ₽
15,61 ₽
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
Mercury 2.5
inception/mercury-2.5
Mercury 2.5 is the fastest reasoning LLM, and the latest diffusion LLM (dLLM) from Inception. Instead of generating tokens sequentially, Mercury 2.5 produces and refines multiple tokens in parallel, achieving...
вход / 1M
4,16 ₽
выход / 1M
15,61 ₽
Nex-N2.5-Mini (free)
nex-agi/nex-n2.5-mini:free
262K
бесплатно
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
Nex-N2.5-Mini (free)
nex-agi/nex-n2.5-mini:free
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
бесплатно
Nex-N2.5-Pro (free)
nex-agi/nex-n2.5-pro:free
262K
бесплатно
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
Nex-N2.5-Pro (free)
nex-agi/nex-n2.5-pro:free
Nex-N2.5 is an agentic model built to turn goals into working, verified outcomes. Its core strength is agentic coding within a visual feedback loop: it can explore codebases, implement multi-file...
бесплатно
25 из 599 моделей Показано 25 из 599 · цены пересчитываются при обновлении курса ЦБ
Цена уже с наценкой
В таблице — итоговая цена, которую вы платите. Наценка 20 % поверх закупки у провайдера.
Списание по факту
Резервируем под максимум, после ответа возвращаем разницу.
Цена в своей единице
За расшифровку платят за секунду звука, за картинку — за изображение. Показываем то, за что списывают, а не всё подряд за миллион токенов.