Fetch models

AI Models Processing Progress

Total Models: 410

Current Batch: 1 of 21

Processed in this batch: 20 of 20

Models 1-20 of 410

« Previous Batch First Batch Next Batch »

For large datasets (300+ models):

🎯 Manual Step-by-Step Processing

Process models one by one with full control. See each model’s details and choose which ones to process.

  • ByteDance Seed: Seed 2.1 Turbo

    ID: bytedance-seed/seed-2-1-turbo

    Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and…

    Product Meta Values:

    Company: Bytedance-seed

    Model: ByteDance Seed: Seed 2.1 Turbo

    Country: USA

    Token Limit: 262144

    Prompt Cost: Prompt: $0.0005 per 1K tokens

    Completion Cost: Completion: $0.0025 per 1K tokens

    Use Case: Bytedance-seed's ByteDance Seed: Seed 2.1 Turbo is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    Seed 2.1 Turbo is a multimodal model from ByteDance Seed for coding and long-horizon agent workflows. It is suited for end-to-end software delivery, multi-step task execution, and understanding visual and…

    Context Length: 262144

    Pricing: Prompt: $0.0000005 / Completion: $0.0000025

  • Qwen: Qwen3.8 2.4T A95B

    ID: qwen/qwen3.8-2.4t-a95b

    Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is…

    Product Meta Values:

    Company: Qwen

    Model: Qwen: Qwen3.8 2.4T A95B

    Country: USA

    Token Limit: 262144

    Prompt Cost: Prompt: $0.0020 per 1K tokens

    Completion Cost: Completion: $0.0060 per 1K tokens

    Use Case: Qwen's Qwen: Qwen3.8 2.4T A95B is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    Qwen3.8 2.4T A95B is an open-weight sparse mixture-of-experts model from Qwen and the open-weight variant of [Qwen3.8 Max](/qwen/qwen3.8-max), with 95 billion active parameters out of 2.4 trillion total. It is…

    Context Length: 262144

    Pricing: Prompt: $0.000002 / Completion: $0.000006

  • ByteDance Seed: Seed-2.0-Code

    ID: bytedance-seed/seed-2.0-code

    Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude…

    Product Meta Values:

    Company: Bytedance-seed

    Model: ByteDance Seed: Seed-2.0-Code

    Country: USA

    Token Limit: 262144

    Prompt Cost: Prompt: $0.0005 per 1K tokens

    Completion Cost: Completion: $0.0030 per 1K tokens

    Use Case: Bytedance-seed's ByteDance Seed: Seed-2.0-Code is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    Seed 2.0 Code is a model from ByteDance Seed optimized for agentic coding. It is suited for frontend development, multilingual programming tasks, and coding-agent workflows in tools such as Claude…

    Context Length: 262144

    Pricing: Prompt: $0.0000005 / Completion: $0.000003

  • DeepSeek: DeepSeek V4 Pro 0813

    ID: deepseek/deepseek-v4-pro-0813

    DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

    Product Meta Values:

    Company: Deepseek

    Model: DeepSeek: DeepSeek V4 Pro 0813

    Country: USA

    Token Limit: 1048576

    Prompt Cost: Prompt: $0.0004 per 1K tokens

    Completion Cost: Completion: $0.0009 per 1K tokens

    Use Case: Deepseek's DeepSeek: DeepSeek V4 Pro 0813 is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    DeepSeek V4 Pro 0813 is a large-scale mixture-of-experts model from DeepSeek. This is the GA release of DeepSeek V4 Pro.

    Context Length: 1048576

    Pricing: Prompt: $0.000000435 / Completion: $0.00000087

  • SpaceXAI: Grok 4.6

    ID: x-ai/grok-4.6

    Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

    Product Meta Values:

    Company: X-ai

    Model: SpaceXAI: Grok 4.6

    Country: USA

    Token Limit: 500000

    Prompt Cost: Prompt: $0.0020 per 1K tokens

    Completion Cost: Completion: $0.0060 per 1K tokens

    Use Case: X-ai's SpaceXAI: Grok 4.6 is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    Grok 4.6 is SpaceXAI's smartest model with frontier performance on coding, knowledge work, and STEM.

    Context Length: 500000

    Pricing: Prompt: $0.000002 / Completion: $0.000006

  • LiquidAI: LFM2.5-2.6B (free)

    ID: liquid/lfm-2.5-2.6b:free

    LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or…

    Product Meta Values:

    Company: Liquid

    Model: LiquidAI: LFM2.5-2.6B (free)

    Country: USA

    Token Limit: 128000

    Prompt Cost: Prompt: $0.0000 per 1K tokens

    Completion Cost: Completion: $0.0000 per 1K tokens

    Use Case: Liquid's LiquidAI: LFM2.5-2.6B (free) is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    LFM2.5-2.6B is a compact reasoning model from Liquid AI. It is suited for agent workflows, data extraction, RAG, and long-context processing. Liquid advises against using it for agentic coding or…

    Context Length: 128000

    Pricing: Prompt: $0 / Completion: $0

  • NVIDIA: Nemotron 3.5 Lightning

    ID: nvidia/nemotron-3.5-lightning

    NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that…

    Product Meta Values:

    Company: Nvidia

    Model: NVIDIA: Nemotron 3.5 Lightning

    Country: USA

    Token Limit: 262144

    Prompt Cost: Prompt: $0.0001 per 1K tokens

    Completion Cost: Completion: $0.0003 per 1K tokens

    Use Case: Nvidia's NVIDIA: Nemotron 3.5 Lightning is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that…

    Context Length: 262144

    Pricing: Prompt: $0.0000001 / Completion: $0.00000025

  • NVIDIA: Nemotron 3.5 Lightning (free)

    ID: nvidia/nemotron-3.5-lightning:free

    NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that…

    Product Meta Values:

    Company: Nvidia

    Model: NVIDIA: Nemotron 3.5 Lightning (free)

    Country: USA

    Token Limit: 1000000

    Prompt Cost: Prompt: $0.0000 per 1K tokens

    Completion Cost: Completion: $0.0000 per 1K tokens

    Use Case: Nvidia's NVIDIA: Nemotron 3.5 Lightning (free) is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    NVIDIA Nemotron 3.5 Lightning is an open mixture-of-experts model from NVIDIA, with 3B active parameters out of 30B total. It is suited for high-throughput agentic workloads and specialized tasks that…

    Context Length: 1000000

    Pricing: Prompt: $0 / Completion: $0

  • Sakana: Sakana Namazu

    ID: sakana/sakana-namazu

    Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,…

    Product Meta Values:

    Company: Sakana

    Model: Sakana: Sakana Namazu

    Country: USA

    Token Limit: 262144

    Prompt Cost: Prompt: $0.0010 per 1K tokens

    Completion Cost: Completion: $0.0040 per 1K tokens

    Use Case: Sakana's Sakana: Sakana Namazu is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    Sakana Namazu is a Japanese-specialized reasoning model from Sakana AI, based on Kimi K2.6 with additional training for Japanese language and business contexts. It is suited for Japanese instruction following,…

    Context Length: 262144

    Pricing: Prompt: $0.00000095 / Completion: $0.000004

  • Upstage: Solar Pro 4

    ID: upstage/solar-pro4

    Solar Pro 4 is a large language model from Upstage. It is suited for agentic workflows, office productivity, document-intensive work, and coding.

    Product Meta Values:

    Company: Upstage

    Model: Upstage: Solar Pro 4

    Country: USA

    Token Limit: 524288

    Prompt Cost: Prompt: $0.0000 per 1K tokens

    Completion Cost: Completion: $0.0001 per 1K tokens

    Use Case: Upstage's Upstage: Solar Pro 4 is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    Solar Pro 4 is a large language model from Upstage. It is suited for agentic workflows, office productivity, document-intensive work, and coding.

    Context Length: 524288

    Pricing: Prompt: $0.00000003 / Completion: $0.00000012

  • Meta: Muse Glimmer 30B

    ID: meta/muse-glimmer-30b

    Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon…

    Product Meta Values:

    Company: Meta

    Model: Meta: Muse Glimmer 30B

    Homepage: https://ai.meta.com

    Country: USA

    Token Limit: 131072

    Prompt Cost: Prompt: $0.0004 per 1K tokens

    Completion Cost: Completion: $0.0015 per 1K tokens

    Use Case: Meta's Meta: Muse Glimmer 30B is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    Muse Glimmer 30B is a dense, open-weight multimodal model from Meta Superintelligence Labs, distilled from Muse Spark and optimized for autonomous agents on consumer hardware. It is suited for long-horizon…

    Context Length: 131072

    Pricing: Prompt: $0.00000035 / Completion: $0.0000015

  • inclusionAI: Ling 3.0 Tiny (free)

    ID: inclusionai/ling-3.0-tiny:free

    Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable…

    Product Meta Values:

    Company: Inclusionai

    Model: inclusionAI: Ling 3.0 Tiny (free)

    Country: USA

    Token Limit: 262144

    Prompt Cost: Prompt: $0.0000 per 1K tokens

    Completion Cost: Completion: $0.0000 per 1K tokens

    Use Case: Inclusionai's inclusionAI: Ling 3.0 Tiny (free) is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    Ling 3.0 Tiny is a mixture-of-experts model from InclusionAI, with 1.3B active parameters out of 7.9B total. It is designed for responsive agents, instruction following, and multi-turn conversations, with switchable…

    Context Length: 262144

    Pricing: Prompt: $0 / Completion: $0

  • Meta: Muse Spark 1.2

    ID: meta/muse-spark-1.2

    Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context…

    Product Meta Values:

    Company: Meta

    Model: Meta: Muse Spark 1.2

    Homepage: https://ai.meta.com

    Country: USA

    Token Limit: 1048576

    Prompt Cost: Prompt: $0.0013 per 1K tokens

    Completion Cost: Completion: $0.0043 per 1K tokens

    Use Case: Meta's Meta: Muse Spark 1.2 is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    Muse Spark 1.2 is a reasoning model from Meta, designed for complex agentic tasks. It accepts text, images, video, audio, and PDF documents, returns text, and offers a 1M-token context…

    Context Length: 1048576

    Pricing: Prompt: $0.00000125 / Completion: $0.00000425

  • Qwen: Qwen3.8 Max

    ID: qwen/qwen3.8-max

    Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual understanding,…

    Product Meta Values:

    Company: Qwen

    Model: Qwen: Qwen3.8 Max

    Country: USA

    Token Limit: 1000000

    Prompt Cost: Prompt: $0.0020 per 1K tokens

    Completion Cost: Completion: $0.0060 per 1K tokens

    Use Case: Qwen's Qwen: Qwen3.8 Max is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    Qwen3.8 Max is the flagship model in Alibaba's Qwen3.8 series, the general-availability successor to the Qwen3.8 Max Preview. It is a multimodal reasoning model intended for complex reasoning, visual understanding,…

    Context Length: 1000000

    Pricing: Prompt: $0.000002 / Completion: $0.000006

  • DeepSeek V4 Flash Latest

    ID: ~deepseek/deepseek-v4-flash-latest

    This model always redirects to the latest model in the DeepSeek V4 Flash family.

    Product Meta Values:

    Company: ~deepseek

    Model: DeepSeek V4 Flash Latest

    Country: USA

    Token Limit: 1048576

    Prompt Cost: Prompt: $0.0001 per 1K tokens

    Completion Cost: Completion: $0.0003 per 1K tokens

    Use Case: ~deepseek's DeepSeek V4 Flash Latest is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    This model always redirects to the latest model in the DeepSeek V4 Flash family.

    Context Length: 1048576

    Pricing: Prompt: $0.000000079996 / Completion: $0.000000252

  • DeepSeek: DeepSeek V4 Flash 0731

    ID: deepseek/deepseek-v4-flash-0731

    DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.

    Product Meta Values:

    Company: Deepseek

    Model: DeepSeek: DeepSeek V4 Flash 0731

    Country: USA

    Token Limit: 1048576

    Prompt Cost: Prompt: $0.0001 per 1K tokens

    Completion Cost: Completion: $0.0002 per 1K tokens

    Use Case: Deepseek's DeepSeek: DeepSeek V4 Flash 0731 is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    DeepSeek V4 Flash 0731 is a sparse mixture-of-experts model from DeepSeek, with 13B active parameters out of 284B total. This re-post-trained revision is suited for coding, reasoning, and agent workflows.

    Context Length: 1048576

    Pricing: Prompt: $0.00000008 / Completion: $0.00000018

  • Thinking Machines: Inkling Small

    ID: thinkingmachines/inkling-small

    Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of…

    Product Meta Values:

    Company: Thinkingmachines

    Model: Thinking Machines: Inkling Small

    Country: USA

    Token Limit: 524288

    Prompt Cost: Prompt: $0.0005 per 1K tokens

    Completion Cost: Completion: $0.0012 per 1K tokens

    Use Case: Thinkingmachines's Thinking Machines: Inkling Small is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    Inkling Small is an open-weight multimodal mixture-of-experts model from Thinking Machines Lab, with 12B active parameters out of 276B total. It is positioned as the smaller, more efficient member of…

    Context Length: 524288

    Pricing: Prompt: $0.00000045 / Completion: $0.0000012

  • Qwen: Qwen3.7 Flash

    ID: qwen/qwen3.7-flash

    Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world…

    Product Meta Values:

    Company: Qwen

    Model: Qwen: Qwen3.7 Flash

    Country: USA

    Token Limit: 1000000

    Prompt Cost: Prompt: $0.0000 per 1K tokens

    Completion Cost: Completion: $0.0001 per 1K tokens

    Use Case: Qwen's Qwen: Qwen3.7 Flash is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    Qwen3.7 Flash is a vision-language reasoning model from Alibaba. It is suited for multimodal agents, visual coding, search, and computer interaction, with strengths in object recognition, spatial understanding, and real-world…

    Context Length: 1000000

    Pricing: Prompt: $0.00000003 / Completion: $0.00000013

  • Claude Opus 5 (Fast)

    ID: anthropic/claude-opus-5-fast

    Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) – identical capabilities with higher output speed at 2x pricing relative to regular Opus 5.

    Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

    Product Meta Values:

    Company: Anthropic

    Model: Claude Opus 5 (Fast)

    Homepage: https://anthropic.com

    Country: USA

    Token Limit: 1000000

    Prompt Cost: Prompt: $0.0100 per 1K tokens

    Completion Cost: Completion: $0.0500 per 1K tokens

    Use Case: Anthropic's Claude Opus 5 (Fast) is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    Fast-mode variant of [Opus 5](/anthropic/claude-opus-5) – identical capabilities with higher output speed at 2x pricing relative to regular Opus 5.

    Learn more in Anthropic's docs: https://platform.claude.com/docs/en/build-with-claude/fast-mode

    Context Length: 1000000

    Pricing: Prompt: $0.00001 / Completion: $0.00005

  • Claude Opus 5

    ID: anthropic/claude-opus-5

    Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis…

    Product Meta Values:

    Company: Anthropic

    Model: Claude Opus 5

    Homepage: https://anthropic.com

    Country: USA

    Token Limit: 1000000

    Prompt Cost: Prompt: $0.0050 per 1K tokens

    Completion Cost: Completion: $0.0250 per 1K tokens

    Use Case: Anthropic's Claude Opus 5 is ideal for complex reasoning, research, and advanced AI applications requiring extensive context.

    Description:
    Claude Opus 5 is Anthropic’s flagship model for demanding reasoning, coding, and long-horizon agentic work. It is particularly strong at end-to-end software tasks, code review and bug finding, visual analysis…

    Context Length: 1000000

    Pricing: Prompt: $0.000005 / Completion: $0.000025