
Share
Despite rapid advancements, AI models still stumble on certain types of puzzles. Test your skills and see where you outperform these sophisticated systems.
Puzzles and games have been integral to the development of artificial intelligence (AI) since its inception. Just as we humans enjoy testing our cognitive abilities with crosswords or logic puzzles, developers use similar challenges to gauge how far AI models have progressed. The term "machine learning" was coined by IBM computer scientist Arthur Samuel in a 1959 article about an algorithm that learned to play checkers. Chess and the Chinese board game Go are also well-known test beds for AI.
AI's puzzle-solving capabilities have improved significantly over the years, but they still fall short in certain areas. In late 2024, a team from Columbia University found that even the best models could only solve 18% of New York Times Connections puzzles. By early 2025, however, some models had nearly perfected this task. This rapid progress highlights AI's potential but also reveals its limitations.
Despite these advancements, today’s models still struggle with specific types of puzzles. Subtle changes in classic riddles often trip them up, and visual puzzles are a particular weak spot. For instance, spatial reasoning tasks like mental rotation problems-where you have to determine if different images represent the same object from different angles-are notoriously difficult for AI.
A 2024 study by researchers from Google and the University of California, Berkeley, highlighted this issue. They found that models often failed to recognize subtle variations in puzzles they had encountered before, leading to incorrect solutions.
To better understand where you stand relative to AI, try your hand at these puzzles:

Instructions: Choose the answer that shows the object in the prompt but from a different angle. In each case, there’s only one correct answer!
Prompt Image 1
Prompt Image 2
Instructions: Solve the following riddle, which has a subtle twist from a common puzzle.
By testing yourself on these puzzles, you can gain insight into where AI excels and where it still lags behind human cognition. While AI continues to advance, there are still areas where human intuition and reasoning outperform even the most sophisticated models.
Tags
Original Sources
AI models flub these intelligence tests. Can you fare any better?
↗ https://www.technologyreview.com/2026/08/26/1141952/puzzles-ai-models-flub-these-tests
About the author
Kai built ML infrastructure at a Bay Area startup before developing an obsession with transformer architectures and inference optimisation that eventually pulled him out of product work entirely. A stint at a compute research lab sharpened his instinct for what actually matters in a model release versus what is marketing. He writes from the inside — from the perspective of someone who has debugged the systems he is describing at three in the morning. He is allergic to hype and instinctively drawn to the unglamorous plumbing questions that everyone else skips over.
More from The Engineer →This Week's Edition
31 August 2026
85 articles
Related Articles
Related Articles
More Stories
© 2026 Cedar & Bloom. All rights reserved.