Skip to main content

Benchmark profile

ZClawBench

A Z.AI benchmark for OpenClaw-style agent workflows spanning information search, office work, data analysis, development and operations, automation, and security.

Data verified

How BenchLM shows ZClawBench right now

BenchLM is tracking ZClawBench in the local dataset, but exact-source verification records for these rows are still being attached. To avoid a blank benchmark page, BenchLM shows the current tracked rows below as a display-only reference table.

These tracked rows are useful for inspection and spot-checking, but until exact-source attachments are completed they should not be treated as fully verified public benchmark rows.

1 tracked modelsLocal tracked rowsAwaiting exact-source attachmentsDisplay only

Tracked score on ZClawBench — July 29, 2026

BenchLM mirrors the published tracked score view for ZClawBench. GLM-5-Turbo leads the public snapshot at 56.4%. BenchLM does not use these results to rank models overall.

1 modelAgenticCurrentDisplay onlyUpdated July 29, 2026

Tracked score table (1 model)

Score
1
56.4%

About ZClawBench

Year

2026

Tasks

OpenClaw agent workflows

Format

End-to-end agent benchmark

Difficulty

Broad productivity and operations workflows

BenchLM tracks the overall ZClawBench score only when Z.AI publishes an exact public value for a specific model.

BenchLM freshness & provenance

Version

ZClawBench 2026

Refresh cadence

Quarterly

Staleness state

Current

Question availability

Public benchmark set

CurrentDisplay only

BenchLM uses freshness metadata to decide whether a benchmark should still be treated as a strong differentiator, a benchmark to watch, or a display-only reference. For the full scoring policy, see the BenchLM methodology page.

FAQ

What does ZClawBench measure?

A Z.AI benchmark for OpenClaw-style agent workflows spanning information search, office work, data analysis, development and operations, automation, and security.

Which model leads the published ZClawBench snapshot?

GLM-5-Turbo currently leads the published ZClawBench snapshot with 56.4% tracked score. BenchLM shows this benchmark for display only and does not use it in overall rankings.

How many models are evaluated on ZClawBench?

1 AI models are included in BenchLM's mirrored ZClawBench snapshot, based on the public leaderboard captured on July 29, 2026.

Last updated: July 29, 2026 · mirrored from the public benchmark leaderboard

Know when it’s worth switching models

The model to choose, the cheaper alternative, and the release we would wait on.

Read a sample issue

Join 2,000+ readers.

One email each week. Unsubscribe anytime.