JevMade Sign in
← Back to experiments

Benchmarks & research

JEV Benchmark – CLASH contradiction detection

Runs the CLASH cross-modal contradiction test against Jev with each image replaced by its COCO caption, then reports accuracy and modality bias.

Source screenshot of JEV Benchmark – CLASH contradiction detection
SOURCE SCREENSHOTFull screenshot ↗

What it does

Because Jev takes no images, a task designed to be cross-modal becomes text against text — worth knowing before reading the score.

How you can use it

You can use this approach to test whether two written descriptions disagree with each other. Start by pairing your text sources together and writing down a specific question about the details, along with multiple-choice answers that include a choice for conflicting information.

A developer can then send both passages to TypeSafe's Jev model using an account access key to see if it catches the contradictions. Because Jev processes only written text, your team must supply text captions instead of actual images.

Maker-reported (not independently measured by JevMade): 1,289 human-verified test samples · 859 samples (alt_caption runs)

Primitives
choice, noul
Platform
Python
Added
Project created

Keep this for later

Sign in to bookmark experiments, guides and videos, and keep notes only you can see.

Continue to sign in

We’ll bring you back to this listing.