Jev Daily

OpenRouter clocks Luna fastest, Jev second

OpenRouter's latency board puts GPT-6 Luna on top at 180ms on global requests, with Jev behind it. Separately, Shrivu Shankar says Jev was the best calibrated of 41 decision models he tested on email classification.

Someone measured it

OpenRouter says Luna is the fastest decisions model, Jev second

OpenRouter published a latency board for decisions models and put GPT-6 Luna first at 180ms on global requests, with Jev and Perplexity Decider behind it. OpenRouter did not post Jev's own number in the text, only the ordering.

OpenRouter
@OpenRouter
X
GPT-6 Luna is the fastest Decisions model today, at 180ms on global requests
Oct 8, 2026 · View on X

Read the full story

Shrivu Shankar says Jev beat 40 other models on calibration

Shrivu Shankar says Jev was more calibrated than the 40 other decision models his team tested on email classification. He describes the result as expected. TypeSafe AI reposted it with the line "Stay calibrated out there!".

Shrivu Shankar
@ShrivuShankar
X
as expected, @typesafeai 's jev was more calibrated than the 40 other decision models we tested on email classification
Oct 8, 2026 · View on X

Read the full story

Writing criteria

Harrison Chase splits trajectory labeling into typed answers per dimension

Harrison Chase says Jev as a judge in LangSmith evals gives difficulty and correctness each their own typed answer, in one pass, on every trace, rather than a single pass or fail. He was picking up Vaibhav Tulsyan's point that every agent trajectory needs assessing across several dimensions, including whether the agent considered diverse design choices.

Harrison Chase
@hwchase17
X
trajectory labeling is really several questions, not one pass/fail
Oct 9, 2026 · View on X

Read the full story

Get the next one by email

Jev, read daily so you do not have to. The builds, the benchmarks, the criteria that worked and the cases where it lost, from the people shipping on TypeSafe AI's System One model.