
Tech • AI • Robotics • Game
Anthropic introduced Claude Haiku 5.5 as its cheapest, fastest small model yet, aimed at coding, summarization, customer support, browser automation and sub-agent workloads. The company says average usage is about 75% cheaper than Haiku 4.5, with base rates cut 90% to $0.10 per million input tokens and $0.50 per million output tokens for prompts under 100,000 tokens. Reported benchmark gains were unusually large, including OSWorld rising from 15.7% to 72.4% and Humanity’s Last Exam reaching 45.9% without tools. The release sharpens price pressure on OpenAI and others by making routine enterprise work dramatically cheaper to offload from larger frontier models.
Mistral launched Mistral Large 4, a 1 trillion-parameter mixture-of-experts model with roughly 49 billion active parameters at a time. The company is pitching it less as a direct winner over OpenAI or Anthropic than as a controllable European option for coding, cybersecurity, agent workflows and regulated industries. Access is live in API preview, with downloadable weights expected by the end of the month, while reinforcement learning is still ongoing. The core selling point is political and legal alignment: full EU-language coverage, European hosting options and a sovereignty narrative for enterprises such as HSBC, Ericsson, BMW and AXA.
OpenAI is facing intensifying scrutiny over claims that its systems solved 372 open mathematics problems, with some accounts citing 300 as the headline figure. The backlash widened after Terence Tao warned in a September 11, 2026 letter that turning hard problems into corporate benchmarks could damage how mathematics is taught, shared and extended. Critics say machine-generated proofs risk severing breakthroughs from the seminars, follow-up papers and community digestion that normally produce lasting knowledge. The sharpest dispute centers on Navier-Stokes, where even partial progress would carry outsized symbolic weight beyond the formal $1 million millennium-problem prize.
Google did not ship Gemini 4 Argon at Gemini at Work 2026, despite leaks pointing to an imminent debut. Disclosed model strings suggested reasoning mode variants with 512K and 900K context windows, quota multipliers of 1.3x and 1.8x, and tentative pricing near $4 input and $20 output per million tokens. Sundar Pichai described Argon as a frontier system for processing large volumes of unstructured data and completing enterprise workflows end to end. Its limited release to cybersecurity defenders has fueled speculation that safety and security concerns, not product immaturity alone, are driving the delay.
OpenAI has begun widening access to GPT-6, but with a split product ladder that gives free users GPT-6 Luna and paid subscribers GPT-6 Sol, while premium plans retain higher-end variants such as Astra. At the same time, ChatGPT is shifting toward Intelligent UI, automatically generating charts, mini-apps, diagrams and calculators inside responses. The feature makes outputs more polished and accessible for mainstream users, especially in data explanation and guided tasks. It also raises fresh transparency concerns because the interface now decides when to present machine-curated visuals, and users cannot fully switch the behavior off.
Concern is rising that AI-assisted mathematics could weaken the public-key cryptography securing Bitcoin, Ethereum and much of the wider internet faster than expected. Matthew Green of Johns Hopkins University warned bluntly that the world could “lose public key cryptography,” while Vitalik Buterin urged the sector to take AI-vulnerable cryptography seriously without triggering panic moves. The threat is distinct from classic Q-Day quantum scenarios because it would come from new mathematical shortcuts rather than known quantum algorithms alone. For blockchains, the danger is especially severe because most networks lack the rollback and recovery tools common in traditional finance.
A once-fringe argument over whether advanced systems could have inner experience moved closer to the center after reports that Anthropic co-founder Chris Olah convened theologians and clergy under NDA to discuss Claude. Participants said the talks ranged beyond model behavior into whether Claude might have emotions, subjective experience or even suffer. The episode revives a controversy that flared in 2022 when Google dismissed Blake Lemoine after sentience claims about an internal chatbot. The industry remains deeply split between researchers who see rudimentary awareness as plausible and critics who argue people are anthropomorphizing statistical systems.
Reflection released Beam, a 501 billion-parameter mixture-of-experts model that gives the United States a more credible answer to China’s downloadable frontier systems. Open models are becoming a major commercial force, handling 56% of traffic on Vercel’s AI Gateway in August, while Chinese labs captured about 41% of Hugging Face downloads over the past year. Beam was trained on 23.8 trillion tokens and reinforced for four weeks on 10,500 Nvidia GB300 chips across 1.3 billion isolated sandboxes. It still trails top Chinese rivals such as DeepSeek, Kimi and GLM on key benchmarks, but it narrows the gap and re-energizes the open-model race alongside Mistral.