Flame-VLM-Code
We show this table for reference; we do not rank on it.
A vision-language coding benchmark for generating correct code from visual and multimodal inputs.
About Flame-VLM-Code
Year
2026
Tasks
Multimodal coding tasks
Format
Vision-language code generation
Difficulty
Multimodal coding
BenchLM tracks Flame-VLM-Code as a display-only multimodal coding benchmark reference.
Freshness and provenance
Version
Flame-VLM-Code 2026
Refresh cadence
Quarterly
Staleness state
Current
Question availability
Public benchmark set
BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.
Questions
What does Flame-VLM-Code measure?
A vision-language coding benchmark for generating correct code from visual and multimodal inputs.
Which model scores highest on Flame-VLM-Code?
No models have been evaluated on Flame-VLM-Code yet.
How many models are evaluated on Flame-VLM-Code?
0 AI models have been evaluated on Flame-VLM-Code on BenchLM.
Know when it’s worth switching models
The model to choose, the cheaper alternative, and the release we would wait on.
Read a sample issueJoin 2,000+ readers.
One email each week. Unsubscribe anytime.