LLM Performance for Code Generation on Noisy Tasks
, 2026.
Abstract
This paper investigates the ability of large language models (LLMs) to recognise and solve tasks which have been obfuscated beyond recognition. Focusing on competitive programming and benchmark tasks (LeetCode and MATH), we compare performance across multiple models and obfuscation methods. We introduce the concept of eager pattern matching and discuss implications for benchmarking, dataset contamination, and automated software systems.
Links
Cite this Paper
Related People
Radzim Sendyka
PhD Student, Cambridge University
Christian Cabrera Jojoa
Assistant Research Professor, Cambridge University
Andrei Paleyes
Visiting Researcher, Cambridge University
Diana Robinson
Visiting Researcher, Cambridge University
Neil D. Lawrence
The DeepMind Professor of Machine Learning, Cambridge University