norgitov/ trends
Technology · people · ideas
Back to discovery/LessWrong14 minutes ago

Benchmarking Jev 1.13 Against No-CoT LLMs

The paper benchmarks Jev 1.13, a non-autoregressive model that provides probabilistic answers without generating text, against no-CoT LLMs on tasks like sabotage detection and multiple-choice questions. Jev shows mixed performance, with strong results on some tasks and poor performance on multi-step reasoning tasks, but offers cost and speed advantages for specific applications.

Open original
SIGNAL FROM THE SOURCE
32
source points
Tracking sinceOctober 7, 202632 source points
Momentum0/hover 0.55 h
PublishedOctober 7, 2026Dewi Gould
BEHIND THE NUMBERS

How interest changes

History starts here

The chart will appear after repeat observations. The current metric comes from the source.

32 source points

Real observations only. History before source connection is not reconstructed.

WHY IT MAY MATTER

Jev's low cost and speed make it suitable for tasks like routing and monitoring in agentic systems, though its lack of reasoning trace limits transparency.

A useful discovery?
KEEP EXPLORING

Connected ideas

Explore topic