Find a story

Search Spins

Search titles, summaries, and missing voices across published articles — press releases, announcements, and media coverage.

3 results for “information extraction”

SPIN Processed News Frame: The Halo

ConstructCIE: A Dataset for Extracting Causal Information from Construction Accident Narratives

Researchers released ConstructCIE, a manually annotated dataset for extracting hierarchical causal information from OSHA construction accident reports, revealing persistent gaps in LLM and sequence tagger performance on precise evidence-span extraction.

Spin 35% Claim Present in Source AI Risk Moderate
arXiv Computation and Language

Aug 10, 2026

SPIN Processed News Frame: The Hype

DocOCR-Eval: A Correction-Based Framework for OCR Tool Selection Without Ground Truth

Researchers introduced DocOCR-Eval, an annotation-free framework to rank OCR and multimodal LLM tools for document parsing without ground-truth labels, addressing the challenge of tool selection in label-scarce real-world settings.

Spin 45% Claim Present in Source AI Risk Moderate
arXiv Machine Learning

Jul 21, 2026

SPIN Processed News Frame: The Hype

Behavioral Controllability of Agentic Models for Information Extraction: From Fixed Workflows to Reflective Agents

A new arXiv preprint investigates whether reflective LLM agents improve controllability and observable behavior over fixed workflows in scholarly dataset extraction, using process-level metrics rather than just accuracy.

Spin 40% Claim Present in Source AI Risk Moderate
arXiv Artificial Intelligence

Jul 20, 2026