Tech • AI • Robotics • Game

VIDEO
ENFR

Anthropic Opus 5.5 leaks, Apple remakes Siri, OpenAI breach

AIMonday, September 21, 2026· 12 videos

Briefing

Audio player
0:00 / 0:00

Anthropic readies Claude Opus 5.5

Anthropic appears to be preparing Claude Opus 5.5, with testing reportedly replacing an earlier Opus 5.2 track under the codename Claude Wafer EAP. Leaked pricing points to an aggressive frontier offer at $4 per 1 million input tokens and $20 per 1 million output tokens, plus discounted cache rates. Early comparisons suggest stronger coding-heavy visual generation and simulation performance, roughly in the range of GPT-6 Astra on some tasks. If launched within days, the model would sharpen both performance and price pressure at the top of the market.

Apple recasts Siri as gatekeeper

Apple is using its new iPhone cycle to push a broader Apple Intelligence thesis: the interface shifts from app hunting to asking for outcomes. In that model, a rebuilt Siri becomes the coordination layer that finds information, triggers services and completes actions across software. The strategic prize is control over digital intent, not just a better assistant. At more than 1 billion active iPhones, however, the challenge is making inference fast and cheap enough for constant everyday use.

Claude helped expose OpenAI weaknesses

Security researchers said Claude Opus 5 helped them chain known flaws into access to parts of OpenAI's internal development environment in under 72 hours. The entry point was reportedly a malformed image on an OpenAI forum that exploited an outdated image-processing library rather than a core product. The team said it spent roughly two months and about $3,000 in token costs exploring paths before the final chain worked, and OpenAI later paid a $6,500 bug bounty. The episode underscores how public models can accelerate offensive security research even without any autonomous attack behavior.

Claude gray tests unsettle rivals

Users reported that requests labeled Claude Fable 5.1 were sometimes silently routed to unreleased Fable 5.2 and Opus 5.2 checkpoints. Early testers described a noticeable jump in reasoning and coding quality, with some claiming better high-reasoning results than OpenAI's GPT-6 Astra. Standout demos included a Brawl Stars-style game clone built in about one hour and stronger performance on physics-style engineering prompts. The gains appear tied to heavier compute, slower generation and potentially higher inference costs, highlighting the trade-off between capability and efficiency.

Astra spreads into engineering teams

Astra is emerging as a high-level software coworker inside companies including Ramp and Notion, where it is being used for coding, debugging and web-based automation. Teams describe it as working from vague objectives rather than step-by-step prompts, then returning concrete changes, screenshots or readable problem summaries. At Ramp, its computer-use features are being tested for interface validation and merchant-site receipt retrieval, while Notion is using it to investigate costly token usage and abuse patterns. The common shift is from autocomplete-style assistance toward goal-driven engineering delegation.

AI agents chase inbox and calendar

A new wave of AI agents is moving beyond chat into persistent assistants that watch inboxes, calendars, files and purchases, then act on a user's behalf. The category is attracting heavy capital, including Instinct at a $10 billion valuation after raising $1 billion, while Meta is entering with Muse tied to WhatsApp distribution. Many of these systems run through cloud-hosted virtual machines, effectively operating as background digital staff. The core unanswered issue is how much autonomy users will trust when agents can spend money, message people or make scheduling decisions.

TypeSafe launches Jeev for decisions

TypeSafe AI unveiled Jeev on 15 September 2026 as a model designed to output typed decisions, categories and probabilities instead of free-form text. Founded by former OpenAI researcher Diogo Almeida, the company says the system targets production workflows such as phishing detection, routing and escalation with latency of 70 to 500 milliseconds. The promise is lower cost and tighter control when businesses need a decision, not prose. Early testing suggests that usefulness will depend on strict task-by-task validation before companies trust it in live operations.

Higgsfield drops subscription-only barrier

Higgsfield launched a pay-as-you-go API for image and video generation that can be used inside ChatGPT, replacing the need for a monthly plan for lighter users. The move opens access to premium models without requiring subscriptions that previously started at $19 per month and climbed to around $59 for top video tiers. New users can fund an API balance directly, while eligible business signups may receive $15 in credits. The change reflects a broader market push toward usage-based creative tooling as providers compete on flexibility as well as model quality.

Videos covered

Previous briefings · AI