Andrej Karpathy’s ‘Lord of the Rings’ experiment shows how testing of AI models have changed and what they lack

AI researcher Andrej Karpathy tested Claude Opus five with a unique coding challenge. The model autonomously wrote thousands of lines of JavaScript code for a 3D environment. This experiment shows large language models moving beyond simple tasks into complex software engineering. However, the test also revealed limitations in the model's ability to audit its visual output. This highlights the evolving landscape of artificial intelligence capabilities and their applications.

Andrej Karpathy’s ‘Lord of the Rings’ experiment shows how testing of AI models have changed and what they lack
AI researcher Andrej Karpathy tested Claude Opus five with a unique coding challenge. The model autonomously wrote thousands of lines of JavaScript code for a 3D environment. This experiment shows large language models moving beyond simple tasks into complex software engineering. However, the test also revealed limitations in the model's ability to audit its visual output. This highlights the evolving landscape of artificial intelligence capabilities and their applications.