norgitov/ trends
Technology · people · ideas
Back to discovery/arXiv1 hour ago

From Pixels to Pairs: A Comprehensive Benchmark of LLM-Based Key-Value Extraction in Noisy Document Settings

This study evaluates open-source instruction-tuned large language models (LLMs) for key-value pair extraction in documents under both clean and noisy OCR conditions. It tests models like Gemma, Mistral, Qwen2.5, LLaMA 3, and DeepSeek on benchmarks such as FUNSD, CORD, and SROIE using both gold-text and OCR outputs from PaddleOCR, EasyOCR, and Tesseract. Results show that while modern LLMs perform well with clean text, their performance drops significantly under OCR noise, with OCR quality becoming the main factor affecting results.

Open original
SIGNAL FROM THE SOURCE
17
September 2026
Tracking sinceSeptember 17, 2026
MomentumMore observations needed
DiscussionNo comment count provided
PublishedSeptember 17, 2026Zahra Anvari, Vassilis Athitsos
BEHIND THE NUMBERS

How interest changes

The publication is the signal

This official source does not publish popularity metrics. The story is refreshed from its RSS feed.

Real observations only. History before source connection is not reconstructed.

WHY IT MAY MATTER

This research helps understand how LLMs perform in real-world document processing scenarios with OCR noise, highlighting the importance of improving OCR quality and model robustness.

A useful discovery?
KEEP EXPLORING

Connected ideas

Explore topic