# Deep dive into ExtractBench with LlamaIndex CTO Simon Suo

Canonical: https://brew.new/browse/templates/email/pt1_k97zfezrespk1wafh39t2cmgpx8e538f

Brand: llamaindex.ai
Category: newsletter

![Preview of Deep dive into ExtractBench with LlamaIndex CTO Simon Suo](https://cdn.brew.new/email-preview-929acf9bed9cd6b8-tracking_r57xnvnm0yaxwjen3vqjdgn2418dt27z-1789012206877.png)

## Email content

Header_Newsletter_1

ExtractBench Image

Existing benchmarks for schema-guided document extraction weren’t built for workflows where agents act on the output. They measure only a slice of the task and miss the failure modes that matter in production. So our applied research team built our own.

Join Simon Suo, CTO and Co-Gounder of LlamaIndex, for a technical walkthrough of ExtractBench, a new benchmark evaluating 14 frontier systems across 370 enterprise documents, 67 document types, and 4,800+ pages. He’ll cover the real-world failure modes that shaped it, the trade-offs between VLMs and coding agents, and what testing top systems revealed.

You will learn:

How document extraction evaluation has evolved, and where existing benchmarks fall short

What makes document extraction hard and how to measure that difficulty

How cost scales with accuracy across VLMs, coding agents and specialized APIs

How to structure robust extraction evals around your own documents and schemas

Register for Free

Gradient_Newsletter

Best regards,

LlamaIndex Team

LlamaIndex, 405 Howard Street, San Francisco, California 94105, United States

Unsubscribe

Manage preferences

[Open and remix this design](https://brew.new/browse/templates/email/pt1_k97zfezrespk1wafh39t2cmgpx8e538f)

[Browse email designs](https://brew.new/browse/templates)
