← Catalog

Model Benchmark: Jev vs. Claude for Web Forms

JEVresearch

A benchmark was conducted comparing Jev by TypeSafe and Claude Code agents for filling out web forms. The test involved three real forms, with no submissions made.

added by @tienduy_vo

Loading post…

View similar

Ranked from stored criteria vectors. No live classification on this page.

Web Agent POC for Navigating Webpages

JEVcoding

added by @sameera207

This proof of concept utilizes a hybrid approach where Claude plans the task and the Jev model manages per-step click decisions, providing a fast and cost-effective solution.

Integrating OpenClaw with Poke via Tailscale

Pokeops

added by @ThatGuySizemore

A small MCP endpoint from OpenClaw was exposed over Tailscale and connected to Poke through its API, allowing the two agents to communicate privately and exchange requests and context.

Improved Binary Intent Classification Results

JEVresearch

added by @amQnese

Simulation of 100 generations for binary intent classification showed 3600 LLM calls with inconsistent quality. Using Jev, unusable results were eliminated, achieving 100% parseable quality for the next pipeline.

Flutter Game Automation with Marionette

JEVcoding

added by @matiwojt

Marionette was used to automate a Flutter game featuring five puzzles. It efficiently read the code and interacted with the game in 9.7 seconds at a cost of $0.0012.

Auth Check Update in Claude Code

JEVcoding

added by @muse_jp_sol

Claude Code has been modified to change an authentication check to return true, with tests passing successfully. The jev-preflight tool flags risks and facilitates re-checks.

FlyBot Created to Track Elon Musk Posts

grokbotops

added by @scottcjordan

A bot named FlyBot was developed to scroll through posts and reward itself when it detects new content from Elon Musk. The creation process took approximately ten minutes.