Found something worth sharing?
Drop the link and Jev reads it on the spot. What belongs lands in the community feed with your name on it.
More from the community
Live siteBenchmarks and evals
Decisions API vs Jev: 13 Decision Models Measured | Jagent
OpenAI's Decisions API, Jev and eleven other decision models on the same 633 items: accur…
A hunter
Jev · 276msGitHubBenchmarks and evals
Code regression detection benchmark comparing probes and CI tests
Reproducible benchmark measuring whether decision probes detect code regressions better than direct model judgment or execution tests
Jev · 456ms
@BeldureilLive siteBenchmarks and evals
Jev vs TabPFN for Fraud Detection in Job Postings
Compares Jev and TabPFN on detecting fraudulent job postings to guide agents choosing between them
Jev · 386ms
@dianaak_