VWT2k-lite
We show this table for reference; we do not rank on it.
A lighter multilingual benchmark slice published in provider tables for broad cross-lingual transfer and understanding.
About VWT2k-lite
Year
2026
Tasks
Multilingual transfer tasks
Format
Cross-lingual benchmark
Difficulty
Broad multilingual capability
VWT2k-lite acts as a compact multilingual stress test. BenchLM tracks it separately because providers often publish it as a standalone row without enough public detail to merge it into existing multilingual benchmark families.
Freshness and provenance
Version
VWT2k-lite 2026
Refresh cadence
Quarterly
Staleness state
Current
Question availability
Public benchmark set
BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.
Questions
What does VWT2k-lite measure?
A lighter multilingual benchmark slice published in provider tables for broad cross-lingual transfer and understanding.
Which model scores highest on VWT2k-lite?
No models have been evaluated on VWT2k-lite yet.
How many models are evaluated on VWT2k-lite?
0 AI models have been evaluated on VWT2k-lite on BenchLM.
Know when it’s worth switching models
The model to choose, the cheaper alternative, and the release we would wait on.
Read a sample issueJoin 2,000+ readers.
One email each week. Unsubscribe anytime.