fireworks/models/deepseek-r1-0528

Common Name: Deepseek R1 05/28

Released on Oct 16, 2025 12:00 AMSupportedTool Invocation

05/28 updated checkpoint of Deepseek R1. Its overall performance is now approaching that of leading models, such as O3 and Gemini 2.5 Pro. Compared to the previous version, the upgraded model shows significant improvements in handling complex reasoning tasks, and this version also offers a reduced hallucination rate, enhanced support for function calling, and better experience for vibe coding.

Specifications

Context

160K

Inputtext

Outputtext

Performance (7-day Average)

Collecting…

Pricing

Input$1.49/MTokens

Output$5.94/MTokens

Availability Trend (24h)

Performance Metrics (24h)

Similar Models

GLM-5

$1.10/$3.52/M

ctx203Kmax—avail—tps—

InOutCap

Z.ai's state-of-the-art mixture-of-experts model with 40B active parameters out of 744B total. Optimized for complex systems engineering and long-horizon agentic tasks, using Deepseek Sparse Attention for efficient long-context processing.

Deepseek V3 03-24

$0.99/$0.99/M

ctx160Kmax—avail—tps—

InOutCap

A strong Mixture-of-Experts (MoE) language model with 671B total parameters with 37B activated for each token from Deepseek. Updated checkpoint.

Llama 3.3 70B Instruct

$0.99/$0.99/M

ctx128Kmax—avail—tps—

InOutCap

Llama 3.3 70B Instruct is the December update of Llama 3.1 70B. The model improves upon Llama 3.1 70B (released July 2024) with advances in tool calling, multilingual text support, math and coding. The model achieves industry leading results in reasoning, math and instruction following and provides similar performance as 3.1 405B but with significant speed and cost improvements.

Kimi K2 Instruct 0905

$0.66/$2.75/M

ctx256Kmax—avail—tps—

InOutCap

Kimi K2 0905 is an updated version of Kimi K2, a state-of-the-art mixture-of-experts (MoE) language model with 32 billion activated parameters and 1 trillion total parameters. Kimi K2 0905 has improved coding abilities, a longer context window, and agentic tool use, and a longer (262K) context window.