← Back to live feed1 story
The new Contrastive Language Model delivers up to 9x faster inference than the Jev model while achieving an 81.6% score on the DeepSWE agentic coding benchmark. Known as CLM-8B, this "System One" model uses a contrastive learning objective to connect states and actions, allowing it to function as a more effective verifier for long-horizon tasks than previous models like Jev. The developers reduced inference latency by disaggregating states and actions so their embeddings can be cached and reused independently when action sets remain fixed. CLM-8B was trained on internet-scale data and follows scaling laws where test contrastive loss decreases as a power law relative to model size, training compute, and dataset size.
Sign in to suggest edits
Key sources
- SOURCEmarketbrief.now
- SOURCEhuggingnewshuggingnews.com