• bitcoinBitcoin(BTC)$83,738.00-2.73%
  • ethereumEthereum(ETH)$2,669.32-2.67%
  • tetherTether(USDT)$1.00-0.01%
  • binancecoinBNB(BNB)$769.72-2.47%
  • rippleXRP(XRP)$1.49-7.59%
  • usd-coinUSDC(USDC)$1.00-0.01%
  • solanaSolana(SOL)$114.18-3.19%
  • tronTRON(TRX)$0.341250-0.70%
  • zcashZcash(ZEC)$1,515.49-6.60%
  • Figure HelocFigure Heloc(FIGR_HELOC)$1.040.38%
  • HyperliquidHyperliquid(HYPE)$91.97-4.50%
  • dogecoinDogecoin(DOGE)$0.093321-7.48%
  • moneroMonero(XMR)$559.52-1.26%
  • whitebitWhiteBIT Coin(WBT)$83.87-3.02%
  • USDSUSDS(USDS)$1.00-0.03%
  • chainlinkChainlink(LINK)$12.29-5.30%
  • cardanoCardano(ADA)$0.238513-6.80%
  • RainRain(RAIN)$0.012095-7.11%
  • leo-tokenLEO Token(LEO)$8.94-0.47%
  • stellarStellar(XLM)$0.201176-7.80%
  • bitcoin-cashBitcoin Cash(BCH)$332.39-7.90%
  • uniswapUniswap(UNI)$9.11-12.61%
  • nearNEAR Protocol(NEAR)$4.24-7.23%
  • litecoinLitecoin(LTC)$68.318.07%
  • Ethena USDeEthena USDe(USDE)$1.00-0.02%
  • daiDai(DAI)$1.000.00%
  • avalanche-2Avalanche(AVAX)$10.16-8.80%
  • USD1USD1(USD1)$1.00-0.01%
  • CantonCanton(CC)$0.108932-3.58%
  • the-open-networkGram (prev. Toncoin)(GRAM)$1.41-3.29%
  • hedera-hashgraphHedera(HBAR)$0.090319-7.93%
  • suiSui(SUI)$0.96-6.54%
  • shiba-inuShiba Inu(SHIB)$0.000006-7.30%
  • Global DollarGlobal Dollar(USDG)$1.00-0.01%
  • BittensorBittensor(TAO)$285.62-8.63%
  • crypto-com-chainCronos(CRO)$0.061522-8.62%
  • MemeCoreMemeCore(M)$1.24-3.34%
  • BitwayBitway(BTW)$1.016.14%
  • paypal-usdPayPal USD(PYUSD)$1.00-0.02%
  • tether-goldTether Gold(XAUT)$4,268.80-1.18%
  • okbOKB(OKB)$119.52-4.34%
  • Circle USYCCircle USYC(USYC)$1.140.01%
  • Ripple USDRipple USD(RLUSD)$1.00-0.01%
  • mantleMantle(MNT)$0.691.49%
  • BlackRock USD Institutional Digital Liquidity FundBlackRock USD Institutional Digital Liquidity Fund(BUIDL)$1.000.00%
  • Ondo US Dollar YieldOndo US Dollar Yield(USDY)$1.14-0.25%
  • aaveAave(AAVE)$137.71-9.10%
  • OndoOndo(ONDO)$0.429965-1.74%
  • EthenaEthena(ENA)$0.203439-5.53%
  • polkadotPolkadot(DOT)$1.12-4.96%
TradePoint.io
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop
No Result
View All Result
TradePoint.io
No Result
View All Result

Contrastive-LM Releases CLM-8B: An Open System One Model That Scores Agent Actions Up to 9× Faster Than Jev

September 24, 2026
in AI & Technology
Reading Time: 17 mins read
A A
Contrastive-LM Releases CLM-8B: An Open System One Model That Scores Agent Actions Up to 9× Faster Than Jev
ShareShareShareShareShare

Contrastive-LM has released CLM-8B, the first open model in a new class called Contrastive Language Models (CLMs). CLM does not generate text. It scores a set of candidate actions against the current state and returns probabilities. Their main baseline is Jev, the proprietary System One model from TypeSafe AI.

Is it deployable? Yes. The Apache-2.0 head weighs 75 MB. It runs on 1 NVIDIA GPU under Linux, with vLLM serving the Qwen3-8B encoder.

YOU MAY ALSO LIKE

A Coding Guide to TypeSafe AI Jev: Typed Decisions, Calibrated Confidence, and Speculative Fan-Out with a System One Model

Everything Announced At Meta Connect 2026

What a System One Model Does

Jev entered limited early access on 15 September 2026. It returns typed values with probabilities instead of text. CLM targets the same interface. The CLM GitHub repo serves CLM-8B behind a TypeSafe-compatible API. It exposes 3 question types:

  • Noul: returns the probability that a statement is true.
  • Choice: picks one option from a declared set, with probabilities.
  • Score: returns an expected level on an ordered rubric.

A request written for TypeSafe’s API can be replayed through CLM’s Python client.

How CLM Works

CLM trains a state encoder and an action encoder with a bidirectional InfoNCE loss. Each encoder is a frozen Qwen3-8B backbone plus a 20M-parameter trainable projection head. Training pulls each state toward the action actually taken and pushes it away from the others.

At inference, CLM scores each candidate by the dot product of the state and action embeddings. A softmax over those scores becomes the answer distribution. The same primitive ranks best-of-N solutions, routes tools and answers typed decisions.

This design disaggregates states and actions. In an agent loop, the state changes every step while the action set stays mostly fixed. clm-serve reserves a slab of GPU memory, similar to vLLM’s KV cache, and reuses cached vectors. On 1 RTX 4090 with 3 actions, revisited states drop from 1.7 ms to 0.6 ms. The model card reports CLM running 13× faster than Jev with about 1,000 candidates.

A 3-Stage Training Recipe

  1. Pre-training on ~60M Nemotron DQA question-answer pairs.
  2. Mid-training on ~30M synthetic hard negatives generated by Gemini 2.5 Flash-Lite.
  3. Post-training on ~1M agent trajectories from Agent Data Protocol, Endless-Terminals and LiteCoder-Terminal-SFT.

On ~100K held-out questions, pre-training alone reaches 52.1% top-1 accuracy. Mid-training lifts it to 69.2%. Training on hard negatives from the start peaks at 62.4%, then overfits.

Zero-Shot Results Against Jev

Task CLM-8B latency Jev latency CLM-8B success Jev success
T-Rex game 16.5 ms 149.8 ms 5/5 5/5
Tool calling (BFCL v4) 76.8 ms 125.5 ms 95.2% 99.2%
WikiRacing 79.8 ms 225 ms 26/30 30/30
Super Mario 33.5 ms 132.6 ms 5/5 5/5

The 9× figure comes from the T-Rex game, where actions repeat across states. CLM matches Jev on T-Rex and Super Mario. It trails on tool calling and WikiRacing while running faster on every task.

CLM as a Verifier for Coding Agents

Here a generator samples several candidate solutions and the verifier picks one. Opus 5 produced DeepSWE candidates (best-of-4). Fable 5 produced Terminal-Bench 2.1 candidates (best-of-5). The team evaluated 38 held-out DeepSWE tasks and 30 held-out Terminal-Bench 2.1 tasks. Latency was measured on an H100.

Benchmark Pass@1 CLM (fine-tuned) Jev CLM latency Jev latency
DeepSWE 73.7% 81.6% 71.1% 79 ms 449 ms
Terminal-Bench 2.1 84.0% 87.6% 83.1% 32 ms 131 ms

The research team reports these as new SOTA verifier results. Jev scores below pass@1 on both benchmarks, so selecting with Jev is worse than taking 1 sample. CLM runs 4.1× to 5.7× faster. These numbers use lightweight fine-tuned heads, not the zero-shot checkpoint. They are held-out subset results, not full leaderboard submissions.

Interactive Explainer

Credit: Source link

ShareTweetSendSharePin

Related Posts

A Coding Guide to TypeSafe AI Jev: Typed Decisions, Calibrated Confidence, and Speculative Fan-Out with a System One Model
AI & Technology

A Coding Guide to TypeSafe AI Jev: Typed Decisions, Calibrated Confidence, and Speculative Fan-Out with a System One Model

September 24, 2026
Everything Announced At Meta Connect 2026
AI & Technology

Everything Announced At Meta Connect 2026

September 24, 2026
Meta Put Muse In A Tamagotchi Like ‘Charm’ Device
AI & Technology

Meta Put Muse In A Tamagotchi Like ‘Charm’ Device

September 24, 2026
Meta Will Stop Training Its AI On ‘Visual Data’ From Its Smart Glasses — If You Opt Out
AI & Technology

Meta Will Stop Training Its AI On ‘Visual Data’ From Its Smart Glasses — If You Opt Out

September 24, 2026
Next Post
Dolly Parton has died, her family announces

Dolly Parton has died, her family announces

Leave a Reply Cancel reply

Your email address will not be published. Required fields are marked *

Search

No Result
View All Result
Columbia Disciplined Growth Fund Q2 2026 Commentary (RDLAX)

Columbia Disciplined Growth Fund Q2 2026 Commentary (RDLAX)

September 18, 2026
Would you let a humanoid robot clean your home?

Would you let a humanoid robot clean your home?

September 21, 2026
Bungie Leaders Now Say The Studio’s ‘Not Done With Destiny’

Bungie Leaders Now Say The Studio’s ‘Not Done With Destiny’

September 21, 2026

About

Learn more

Our Services

Legal

Privacy Policy

Terms of Use

Bloggers

Learn more

Article Links

Contact

Advertise

Ask us anything

©2020- TradePoint.io - All rights reserved!

Tradepoint.io, being just a publishing and technology platform, is not a registered broker-dealer or investment adviser. So we do not provide investment advice. Rather, brokerage services are provided to clients of Tradepoint.io by independent SEC-registered broker-dealers and members of FINRA/SIPC. Every form of investing carries some risk and past performance is not a guarantee of future results. “Tradepoint.io“, “Instant Investing” and “My Trading Tools” are registered trademarks of Apperbuild, LLC.

This website is operated by Apperbuild, LLC. We have no link to any brokerage firm and we do not provide investment advice. Every information and resource we provide is solely for the education of our readers. © 2020 Apperbuild, LLC. All rights reserved.

No Result
View All Result
  • Main
  • AI & Technology
  • Stock Charts
  • Market & News
  • Business
  • Finance Tips
  • Trade Tube
  • Blog
  • Shop

© 2023 - TradePoint.io - All Rights Reserved!