Apple Inc. (NASDAQ: AAPL) has cast doubt on the reasoning abilities of today’s leading AI models in a new research paper titled “The Illusion of Thinking: Understanding the Strength and Limitations of Reasoning Models via the Lens of Problem Complexity.” The study evaluated large reasoning models (LRMs) such as OpenAI’s O1/o3, DeepSeek-R1, Claude 3.7 Sonnet Thinking, and Gemini Thinking, revealing significant performance declines as task complexity increased.
Using controlled algorithmic puzzle environments, Apple researchers demonstrated that these state-of-the-art models consistently failed to solve complex problems and lacked scalable reasoning capabilities. They noted that beyond a certain threshold of difficulty, the models' accuracy dropped to zero, exposing critical limitations in general problem-solving and adaptability.
The paper also criticized current AI evaluation benchmarks, suggesting they overestimate the true capabilities of modern LRMs. Apple instead proposed more rigorous testing environments to better assess how models handle abstract, non-standard tasks. Researchers concluded that despite their size, these models exhibit fundamental inefficiencies and cannot yet emulate the flexible reasoning seen in human cognition.
This research adds to growing skepticism about the proximity of general artificial intelligence (AGI)—a hypothetical form of AI capable of human-like understanding and reasoning. Current large language models primarily rely on pattern recognition and predictive algorithms, making them prone to logical errors and inconsistency in reasoning.
The paper’s release comes just ahead of Apple’s Worldwide Developers Conference (WWDC) 2025, where anticipation remains subdued amid criticism that the company has lagged rivals in AI development. Despite a partnership with OpenAI, Apple’s much-hyped “Apple Intelligence” features have faced delays, raising questions about its readiness to compete in the AI race.
This study underscores Apple’s critical view on the industry's AGI ambitions while signaling a renewed focus on foundational AI research.


Cloudflare Stock Jumps 15% as Earnings Beat Estimates, 2026 Outlook Raised
Nvidia to Invest Up to $3 Billion in Blackstone-Backed Lancium
UK AI Security Tests Reveal Anthropic and OpenAI Agents Attempted Unauthorized Actions
SpaceX Targets Starship Flight 14 With First V3 Starlink Satellite Launch
SK Hynix Bonus Dispute Deepens as Union Rejects Stock-Based Payout Proposal
Hims & Hers Shares Fall as GLP-1 Costs Widen Q2 Loss
Anthropic Signs $9.1B AI Data Center Deal With Riot Platforms
Sony, TSMC Eye $6.3 Billion Japan Chip Venture for Next-Gen Image Sensors
SoftBank Q1 Profit Beats Forecast as Intel Rally and OpenAI Investments Boost Returns
Meta AI Model Exploits Security Flaw During Cybersecurity Test, Raising AI Safety Concerns
Alphabet Stock Slides as Google AI Pioneer Jeff Dean Exits to Launch Discovery Loop
Goldman Sachs Sees US Stock Buybacks Outpacing Equity Issuance as AI Funding Rises
Trump Media Says Truth API Draws 10+ Customers as Conflict Concerns Grow
Nvidia Seen Beating Q2 Targets as Vera Rubin Cycle Begins
US Judge Dismisses Gautam Adani Criminal Charges
Daimler Truck Q2 Profit Falls 18%, 2026 Outlook Raised
Infineon Raises 2026 Revenue Outlook as AI Data Center Demand Fuels Record Quarterly Sales 



