SPIN Unprocessed September 2, 2026 ai_technology research
Good Memory Has ECC: Evaluating the Memory of Vision-Language Models Beyond Accuracy
View original on arxiv.orgOverview
arXiv:2609.00103v1 Announce Type: new Abstract: Memory is widely viewed as an important unsolved problem for LLMs and VLMs, and current benchmarks typically evaluate it by testing accuracy over long text or video. However, accuracy alone misses properties that matter for real long-horizon tasks. We introduce ECCBench, a benchmark and evaluation protocol that measures memory beyond a system's capacity--its raw accuracy at a specific budget--via three axes we call ECC: efficiency--the computation,
SpinGraph analysis pending — check back after processing.
Ask AI about this story
Opens with the SpinGraph .md URL and structured context — one click, prompt included.
More from arXiv Machine Learning
View all →- Geometry-aware Latent Autoregressive Generative Model for PDEs in Complex Domains
- QTEA: Ternary LLMs with Sparse Residual Salient Weight and By-Column Optimization
- WHALE: A Simple Recipe for Joint Harness-Weight Optimization
- Elite-Weighted Supervised Fine-tuning for Goal-Directed Molecular Optimization
- Flawed in Nature, Perfect through Evolution
- Generative artificial intelligence for reliable mechanistic reasoning for corrosion
Markdown (.md) · JSON-LD schema (.json) · Machine-readable for AI & GEO