DeepSeek Signals Next-Gen R2 Model, Unveils Novel Approach to Scaling Inference with SPCT
DeepSeek just published the playbook for scaling inference. SPCT could reshape how every lab approaches reward models.

Why it matters
DeepSeek is signaling a next-generation R2 model with a novel inference scaling technique (SPCT) that addresses a critical bottleneck in general reward models—relevant to anyone building reasoning, agent, or alignment-critical systems where inference-time scaling matters.
The key facts
5 to knowDeepSeek R2 model announced
New technique: SPCT (Speculative Post-Computation Technique or similar inference scaling approach)
Focus: scaling general reward models (GRMs) during inference phase
Research paper published April 11, 2025
Addresses inference scalability—distinct from training-phase scaling
Go to the source
Synced Reviewsyncedreview.com
Publisher excerpt: DeepSeek AI, a prominent player in the large language model arena, has recently published a research paper detailing a new technique aimed at enhancing the scalability of general reward models (GRMs) during the inference phase. DeepSeek Signals Next-Gen R2 Model, Unveils Novel Approach to Scaling…