LLMs Corrupt Your Documents When You Delegate
2026
Open publication workspace · Sign in to read the full PDF.
AI-generated summary
1) Current LLMs are unreliable delegates, corrupting documents with sparse but severe errors that compound over long interactions.
2) This paper introduces DELEGATE-52, a benchmark for evaluating LLM readiness in delegated workflows across 52 professional domains, and presents findings showing significant degradation even in frontier models.
3) The study reveals that document size, interaction length, and distractor files exacerbate degradation, while agentic tool use does not improve performance.
Tags: LLMs, document corruption, delegated work, AI reliability, benchmark
Check the original publication for accuracy and context.