AI Models Processing Progress
Total Models: 410
Current Batch: 1 of 21
Processed in this batch: 20 of 20
Models 1-20 of 410
For large datasets (300+ models):
🎯 Manual Step-by-Step ProcessingProcess models one by one with full control. See each model’s details and choose which ones to process.
-
ByteDance Seed: Seed 2.1 Turbo
ID: bytedance-seed/seed-2-1-turbo
Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and…
Product Meta Values:
Company: Bytedance-seed
Model: ByteDance Seed: Seed 2.1 Turbo
Country: USA
Token Limit: 262144
Prompt Cost: Prompt: $0.0005 per 1K tokens
Completion Cost: Completion: $0.0025 per 1K tokens
Use Case: Bytedance-seed's ByteDance Seed: Seed 2.1 Turbo is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and…Context Length: 262144
Pricing: Prompt: $0.0000005 / Completion: $0.0000025
-
Qwen: Qwen3.8 2.4T A95B
ID: qwen/qwen3.8-2.4t-a95b
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is…
Product Meta Values:
Company: Qwen
Model: Qwen: Qwen3.8 2.4T A95B
Country: USA
Token Limit: 262144
Prompt Cost: Prompt: $0.0020 per 1K tokens
Completion Cost: Completion: $0.0060 per 1K tokens
Use Case: Qwen's Qwen: Qwen3.8 2.4T A95B is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is…Context Length: 262144
Pricing: Prompt: $0.000002 / Completion: $0.000006
-
ByteDance Seed: Seed-2.0-Code
ID: bytedance-seed/seed-2.0-code
Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude…
Product Meta Values:
Company: Bytedance-seed
Model: ByteDance Seed: Seed-2.0-Code
Country: USA
Token Limit: 262144
Prompt Cost: Prompt: $0.0005 per 1K tokens
Completion Cost: Completion: $0.0030 per 1K tokens
Use Case: Bytedance-seed's ByteDance Seed: Seed-2.0-Code is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude…Context Length: 262144
Pricing: Prompt: $0.0000005 / Completion: $0.000003
-
DeepSeek: DeepSeek V4 Pro 0813
ID: deepseek/deepseek-v4-pro-0813
DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.
Product Meta Values:
Company: Deepseek
Model: DeepSeek: DeepSeek V4 Pro 0813
Country: USA
Token Limit: 1048576
Prompt Cost: Prompt: $0.0004 per 1K tokens
Completion Cost: Completion: $0.0009 per 1K tokens
Use Case: Deepseek's DeepSeek: DeepSeek V4 Pro 0813 is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.Context Length: 1048576
Pricing: Prompt: $0.000000435 / Completion: $0.00000087
-
SpaceXAI: Grok 4.6
ID: x-ai/grok-4.6
Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.
Product Meta Values:
Company: X-ai
Model: SpaceXAI: Grok 4.6
Country: USA
Token Limit: 500000
Prompt Cost: Prompt: $0.0020 per 1K tokens
Completion Cost: Completion: $0.0060 per 1K tokens
Use Case: X-ai's SpaceXAI: Grok 4.6 is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.Context Length: 500000
Pricing: Prompt: $0.000002 / Completion: $0.000006
-
LiquidAI: LFM2.5-2.6B (free)
ID: liquid/lfm-2.5-2.6b:free
LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or…
Product Meta Values:
Company: Liquid
Model: LiquidAI: LFM2.5-2.6B (free)
Country: USA
Token Limit: 128000
Prompt Cost: Prompt: $0.0000 per 1K tokens
Completion Cost: Completion: $0.0000 per 1K tokens
Use Case: Liquid's LiquidAI: LFM2.5-2.6B (free) is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or…Context Length: 128000
Pricing: Prompt: $0 / Completion: $0
-
NVIDIA: Nemotron 3.5 Lightning
ID: nvidia/nemotron-3.5-lightning
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that…
Product Meta Values:
Company: Nvidia
Model: NVIDIA: Nemotron 3.5 Lightning
Country: USA
Token Limit: 262144
Prompt Cost: Prompt: $0.0001 per 1K tokens
Completion Cost: Completion: $0.0003 per 1K tokens
Use Case: Nvidia's NVIDIA: Nemotron 3.5 Lightning is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that…Context Length: 262144
Pricing: Prompt: $0.0000001 / Completion: $0.00000025
-
NVIDIA: Nemotron 3.5 Lightning (free)
ID: nvidia/nemotron-3.5-lightning:free
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that…
Product Meta Values:
Company: Nvidia
Model: NVIDIA: Nemotron 3.5 Lightning (free)
Country: USA
Token Limit: 1000000
Prompt Cost: Prompt: $0.0000 per 1K tokens
Completion Cost: Completion: $0.0000 per 1K tokens
Use Case: Nvidia's NVIDIA: Nemotron 3.5 Lightning (free) is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that…Context Length: 1000000
Pricing: Prompt: $0 / Completion: $0
-
Sakana: Sakana Namazu
ID: sakana/sakana-namazu
Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,…
Product Meta Values:
Company: Sakana
Model: Sakana: Sakana Namazu
Country: USA
Token Limit: 262144
Prompt Cost: Prompt: $0.0010 per 1K tokens
Completion Cost: Completion: $0.0040 per 1K tokens
Use Case: Sakana's Sakana: Sakana Namazu is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,…Context Length: 262144
Pricing: Prompt: $0.00000095 / Completion: $0.000004
-
Upstage: Solar Pro 4
ID: upstage/solar-pro4
Solar Pro 4 is a large language model from Upstage. It is suited for agentic workflows, office productivity, document-intensive work, and coding.
Product Meta Values:
Company: Upstage
Model: Upstage: Solar Pro 4
Country: USA
Token Limit: 524288
Prompt Cost: Prompt: $0.0000 per 1K tokens
Completion Cost: Completion: $0.0001 per 1K tokens
Use Case: Upstage's Upstage: Solar Pro 4 is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
Solar Pro 4 is a large language model from Upstage. It is suited for agentic workflows, office productivity, document-intensive work, and coding.Context Length: 524288
Pricing: Prompt: $0.00000003 / Completion: $0.00000012
-
Meta: Muse Glimmer 30B
ID: meta/muse-glimmer-30b
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon…
Product Meta Values:
Company: Meta
Model: Meta: Muse Glimmer 30B
Homepage: https://ai.meta.com
Country: USA
Token Limit: 131072
Prompt Cost: Prompt: $0.0004 per 1K tokens
Completion Cost: Completion: $0.0015 per 1K tokens
Use Case: Meta's Meta: Muse Glimmer 30B is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon…Context Length: 131072
Pricing: Prompt: $0.00000035 / Completion: $0.0000015
-
inclusionAI: Ling 3.0 Tiny (free)
ID: inclusionai/ling-3.0-tiny:free
Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable…
Product Meta Values:
Company: Inclusionai
Model: inclusionAI: Ling 3.0 Tiny (free)
Country: USA
Token Limit: 262144
Prompt Cost: Prompt: $0.0000 per 1K tokens
Completion Cost: Completion: $0.0000 per 1K tokens
Use Case: Inclusionai's inclusionAI: Ling 3.0 Tiny (free) is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable…Context Length: 262144
Pricing: Prompt: $0 / Completion: $0
-
Meta: Muse Spark 1.2
ID: meta/muse-spark-1.2
Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context…
Product Meta Values:
Company: Meta
Model: Meta: Muse Spark 1.2
Homepage: https://ai.meta.com
Country: USA
Token Limit: 1048576
Prompt Cost: Prompt: $0.0013 per 1K tokens
Completion Cost: Completion: $0.0043 per 1K tokens
Use Case: Meta's Meta: Muse Spark 1.2 is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context…Context Length: 1048576
Pricing: Prompt: $0.00000125 / Completion: $0.00000425
-
Qwen: Qwen3.8 Max
ID: qwen/qwen3.8-max
Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual understanding,…
Product Meta Values:
Company: Qwen
Model: Qwen: Qwen3.8 Max
Country: USA
Token Limit: 1000000
Prompt Cost: Prompt: $0.0020 per 1K tokens
Completion Cost: Completion: $0.0060 per 1K tokens
Use Case: Qwen's Qwen: Qwen3.8 Max is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual understanding,…Context Length: 1000000
Pricing: Prompt: $0.000002 / Completion: $0.000006
-
DeepSeek V4 Flash Latest
ID: ~deepseek/deepseek-v4-flash-latest
This model always redirects to the latest model in the DeepSeek V4 Flash family.
Product Meta Values:
Company: ~deepseek
Model: DeepSeek V4 Flash Latest
Country: USA
Token Limit: 1048576
Prompt Cost: Prompt: $0.0001 per 1K tokens
Completion Cost: Completion: $0.0003 per 1K tokens
Use Case: ~deepseek's DeepSeek V4 Flash Latest is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
This model always redirects to the latest model in the DeepSeek V4 Flash family.Context Length: 1048576
Pricing: Prompt: $0.000000079996 / Completion: $0.000000252
-
DeepSeek: DeepSeek V4 Flash 0731
ID: deepseek/deepseek-v4-flash-0731
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.
Product Meta Values:
Company: Deepseek
Model: DeepSeek: DeepSeek V4 Flash 0731
Country: USA
Token Limit: 1048576
Prompt Cost: Prompt: $0.0001 per 1K tokens
Completion Cost: Completion: $0.0002 per 1K tokens
Use Case: Deepseek's DeepSeek: DeepSeek V4 Flash 0731 is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.Context Length: 1048576
Pricing: Prompt: $0.00000008 / Completion: $0.00000018
-
Thinking Machines: Inkling Small
ID: thinkingmachines/inkling-small
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of…
Product Meta Values:
Company: Thinkingmachines
Model: Thinking Machines: Inkling Small
Country: USA
Token Limit: 524288
Prompt Cost: Prompt: $0.0005 per 1K tokens
Completion Cost: Completion: $0.0012 per 1K tokens
Use Case: Thinkingmachines's Thinking Machines: Inkling Small is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of…Context Length: 524288
Pricing: Prompt: $0.00000045 / Completion: $0.0000012
-
Qwen: Qwen3.7 Flash
ID: qwen/qwen3.7-flash
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world…
Product Meta Values:
Company: Qwen
Model: Qwen: Qwen3.7 Flash
Country: USA
Token Limit: 1000000
Prompt Cost: Prompt: $0.0000 per 1K tokens
Completion Cost: Completion: $0.0001 per 1K tokens
Use Case: Qwen's Qwen: Qwen3.7 Flash is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world…Context Length: 1000000
Pricing: Prompt: $0.00000003 / Completion: $0.00000013
-
Claude Opus 5 (Fast)
ID: anthropic/claude-opus-5-fast
Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) – identical capabilities with higher output speed at 2x pricing relative to regular Opus 5.
Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-modeProduct Meta Values:
Company: Anthropic
Model: Claude Opus 5 (Fast)
Homepage: https://anthropic.com
Country: USA
Token Limit: 1000000
Prompt Cost: Prompt: $0.0100 per 1K tokens
Completion Cost: Completion: $0.0500 per 1K tokens
Use Case: Anthropic's Claude Opus 5 (Fast) is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) – identical capabilities with higher output speed at 2x pricing relative to regular Opus 5.
Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-modeContext Length: 1000000
Pricing: Prompt: $0.00001 / Completion: $0.00005
-
Claude Opus 5
ID: anthropic/claude-opus-5
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis…
Product Meta Values:
Company: Anthropic
Model: Claude Opus 5
Homepage: https://anthropic.com
Country: USA
Token Limit: 1000000
Prompt Cost: Prompt: $0.0050 per 1K tokens
Completion Cost: Completion: $0.0250 per 1K tokens
Use Case: Anthropic's Claude Opus 5 is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.
Description:
Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis…Context Length: 1000000
Pricing: Prompt: $0.000005 / Completion: $0.000025
