Skip to content

Models

11 models · one key, pay per token, no monthly fee

$0.10/M
Lowest input price
1M
Max context
2
Vendors

11 / 11 models

  • DeepSeek

    DeepSeek V4 Pro (75% off) 75% OFF

    DeepSeek V4 Pro is DeepSeek's flagship Mixture-of-Experts (MoE) model, featuring 1.6 trillion total parameters, 49 billion activated parameters, and a 1 million-token context window. It delivers exceptional performance in reasoning, coding, mathematics, and software engineering, making it ideal for demanding AI workloads. Built on the same architecture as DeepSeek V4 Flash, it adds a hybrid attention system for more efficient long-context processing. It supports High and XHigh reasoning modes (with XHigh corresponding to maximum reasoning), making it well suited for complex tasks such as full codebase analysis, multi-step agent workflows, large-scale information synthesis, and enterprise-grade automation.

    Input
    $0.43/M $1.71
    Output
    $0.86/M $3.43
    Context
    1M
    CodingAgents & dataMath & science 322.7K tok
    deepseek/deepseek-v4-pro
  • DeepSeek V4 Flash is DeepSeek's high-speed, cost-efficient Mixture-of-Experts (MoE) model, featuring 284 billion total parameters, 13 billion activated parameters, and a 1 million-token context window. It delivers fast inference, high throughput, and strong performance in reasoning, coding, and general-purpose AI tasks. Built with a hybrid attention architecture for efficient long-context processing, it supports High and XHigh reasoning modes (with XHigh corresponding to maximum reasoning). It is ideal for coding assistants, conversational AI, real-time applications, and agent workflows where speed, scalability, and cost efficiency are essential.

    Input
    $0.14/M
    Output
    $0.29/M $0.43
    Context
    1M
    Math & scienceAgents & dataDocs & office 481.9K tok
    deepseek/deepseek-v4-flash
  • DeepSeek

    DeepSeek V3.1

    DeepSeek V3.1 is DeepSeek's advanced hybrid reasoning model, featuring 671 billion total parameters, 37 billion activated parameters, and a 128K-token context window. It supports both reasoning and non-reasoning modes, allowing developers to optimize for either response quality or speed depending on the task. The model delivers major improvements in reasoning, coding, tool use, and agent workflows, with performance approaching DeepSeek R1 while providing faster responses and more efficient inference. It supports structured tool calling, code agents, search agents, and complex multi-step workflows, making it an excellent choice for research, software development, enterprise automation, and general-purpose AI applications.

    Input
    $0.57/M
    Output
    $1.71/M
    Context
    128K
    Agents & dataMultimodal understandingMath & science 22 tok
    deepseek/deepseek-v3.1