Structured Output Benchmark Value Accuracy (SOB Value Acc)
We show this table for reference; we do not rank on it.
A structured-output benchmark from Interfaze measuring whether extracted JSON leaf values exactly match verified ground truth.
Benchmark score on SOB Value Acc — September 27, 2026
We compile the SOB Value Acc rows from provider self-reports. Interfaze Beta leads the table at 79.5%. We do not use these results to rank models overall.
1 modelInstruction FollowingCurrentDisplay onlyUpdated September 27, 2026
Benchmark score table (1 model)
ScoreAbout SOB Value Acc
Year
2026
Tasks
Structured output extraction
Format
Value accuracy
Difficulty
Production structured-output reliability
SOB Value Accuracy goes beyond JSON parse success: it measures whether values in the structured response are correct and grounded in the source context across text, image, and audio-normalized inputs.
Freshness and provenance
Version
SOB Value Acc 2026
Refresh cadence
Quarterly
Staleness state
Current
Question availability
Public benchmark set
BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.
Questions
What does SOB Value Acc measure?
A structured-output benchmark from Interfaze measuring whether extracted JSON leaf values exactly match verified ground truth.
Which model scores highest on SOB Value Acc?
Interfaze Beta by Interfaze currently leads with a score of 79.5% on SOB Value Acc.
How many models are evaluated on SOB Value Acc?
1 AI models have been evaluated on SOB Value Acc on BenchLM.
Know when it’s worth switching models
The model to choose, the cheaper alternative, and the release we would wait on.
Read a sample issueJoin 2,000+ readers.
One email each week. Unsubscribe anytime.