Evaluation of Jev Model for SEC Filings and Fact Extraction
JEVresearch
The Jev model was tested on SEC filings and fact extraction tasks, showing strengths in narrow yes-or-no questions but weaknesses in understanding document context. It performed well as a first-pass filter but lagged behind production models in accuracy.
added by @RegenbaumShaun
Loading post…
View similar
Ranked from stored criteria vectors. No live classification on this page.
Jev was tested in an AI customer feedback system using 1,040 GitHub issues, proving to be 98% cheaper and 84% faster than Claude Sonnet 4.6 while maintaining higher accuracy.
Jev classified 100 emails in 1.42 seconds, achieving 96% accuracy by routing uncertain cases to Kimi K3. The total inference cost was approximately $0.07.
Jev offers a new intelligent decision-making primitive that enhances model routing and classification tasks. It allows for quick and cost-effective decision-making, improving the efficiency of LLM calls in applications.
The Jev model was benchmarked against an ensemble of models for code reviews, achieving zero false positives, a review speed increase of ~50x, and a cost reduction of ~100x. It demonstrated a 75% bug recall rate, highlighting its efficiency compared to traditional multi-turn agent workflows.
The integration of Jev from @typesafe_ai replaced three steps in Prio, achieving 100% accuracy in model routing and significantly reducing action review time from 5 seconds to 0.25 seconds for clear cases.
Jev is a model that processes states and typed questions to return structured JSON with probability distributions. It outperforms previous setups in speed and cost, making it ideal for routing and decision-making tasks.