Daily brief at 7am Melbourne. Unsubscribe any time.

Monday 8 June 2026

The "Tokenpocalypse" Is Here: AI Just Got More Expensive, and It's Going to Keep Getting Worse

AI companies are eyeing IPOs and quietly hiking token prices — and the era of cheap AI may already be over.

Lead story

The "Tokenpocalypse" Is Here: AI Just Got More Expensive, and It's Going to Keep Getting Worse

For the past few years, the price of AI has moved in one direction: down. OpenAI, Anthropic, Google — they've been in a race to the bottom on API costs, subsidised by venture capital, burning cash to win market share. That era appears to be ending.

TechCrunch's weekend deep-dive coins a term that's already spreading: the "Tokenpocalypse." The thesis is straightforward. The major AI labs — OpenAI, Anthropic, and others — are preparing for IPOs. Public markets demand a path to profitability. The easiest lever to pull is pricing, and several labs have already begun pulling it.

What's actually changing? Token prices — the per-unit cost of sending text to and receiving it from an AI model — have crept up in recent months after years of cuts. The increases are modest so far, but the direction has reversed. For individual developers building hobby projects, this is mildly annoying. For enterprises running millions of AI calls a day, it's a line item that can reshape a business case entirely.

The timing is not coincidental. OpenAI has been telegraphing an IPO pathway for much of 2026, and Anthropic has had similar conversations with investors. When you're telling public market investors a story, "we're the cheapest option" is not the story you want to tell. "We have pricing power" is.

The deeper problem is dependency. Many companies — from scrappy startups to large enterprises — have built products and workflows on the assumption that AI inference costs would continue to fall. Some have even priced their own products on that assumption. A sustained reversal puts those businesses in a difficult position: absorb the margin hit, raise prices on their own customers, or rebuild on cheaper alternatives.

Those cheaper alternatives do exist. Open-source models like Meta's Llama family and Mistral's releases have made self-hosting increasingly viable. But self-hosting carries its own costs — compute, engineering time, reliability overhead — that aren't zero. For most mid-sized companies, switching is a genuine project, not an afternoon's work.

Watch the Notion/Anthropic situation as a canary here. This weekend, Notion experienced a service disruption tied to its Anthropic integration. The outage was brief, and Notion's head of product seemed genuinely surprised by the social media response. But the episode illustrates a structural risk hiding inside every AI-native product: when your core feature runs on someone else's infrastructure and someone else's pricing, your reliability and your margins are both partially out of your hands.

For Australian businesses, the stakes are real. Australian AI adoption has accelerated sharply across professional services, government, and the technology sector. Many of those deployments sit on US-priced, US-hosted API services. A sustained price increase in USD terms, compounded by any AUD/USD movement, hits harder on this side of the Pacific. Procurement teams that haven't stress-tested their AI cost assumptions against a 2–3× price increase should probably do that this quarter.

The AI price war was never going to last forever. The question now is how fast the pendulum swings — and who gets caught without a chair when the music stops.

Also today

Silent Ransom Group Is Calling Law Firms — and It's Working

A cybercriminal group called Silent Ransom Group has been running a highly effective social engineering campaign against US law firms and professional services organisations. Rather than deploying malware, they phone targets pretending to be IT support staff, then talk employees into handing over credentials or granting remote access. According to Mandiant research, data theft is happening within hours of initial contact. Law firms are a particularly attractive target: they hold sensitive client communications, financial records, and M&A details under one roof. The attack pattern is a sobering reminder that the most effective intrusions often skip the technical complexity entirely and just ask nicely. Australian law firms should treat any unsolicited IT support calls with serious scepticism.

Bleeping Computer

C0XMO Botnet Hijacks DD-WRT Routers — and Murders the Competition

A new botnet variant called C0XMO — built on the Gafgyt malware family — is actively exploiting a flaw in DD-WRT router firmware to conscript devices into its network. What makes it interesting is the extra behaviour: once C0XMO lands on a device, it actively kills competing malware processes, evicting other botnets to claim the hardware for itself. It can also pivot to devices running different CPU architectures, making it unusually portable. DD-WRT is a popular open-source firmware used by home users and small businesses to squeeze extra capability out of consumer routers. If you're running DD-WRT, check for a firmware update and consider whether that router is sitting inside or outside your network perimeter.

Bleeping Computer

The Worst Hacks of 2026 So Far — a Useful Stocktake at the Halfway Mark

TechCrunch has compiled a half-year rundown of the most damaging breaches and security incidents of 2026, and it's a useful reference. Highlights include a significant DOGE-linked data exposure, attacks on critical energy and water infrastructure, and the compromise of an FBI surveillance system — each a distinct flavour of risk. What the list makes clear is that 2026 has seen a notably high proportion of incidents targeting government systems and critical infrastructure, rather than just the usual commercial breach tally. For Australian security teams, it's worth benchmarking your threat model against what's actually being targeted, not just what made headlines locally. The ACSC's own threat landscape reporting complements this well.

TechCrunch

OpenAI's 'Super App' Ambitions: Chat Is Dead, Apparently

A senior OpenAI employee has publicly declared that "chat is dead" — signalling where the company thinks AI interfaces are heading. OpenAI is still actively building what it's calling a super app: a single product meant to be the interface for work, creativity, and communication, powered by AI throughout. The framing echoes WeChat in China or early ambitions for what Slack might become. Whether OpenAI can pull it off is a separate question — building a distribution platform is a very different capability from building a model. But the intent is clear: OpenAI doesn't want to be a commodity API provider. It wants to own the surface your hands touch.

TechCrunch

AI Gun Detection System Failed. Now There's a Lawsuit.

A school shooting survivor in the United States is suing an AI-powered gun detection company after its system failed to flag the weapon before the attack. The case raises a question the AI safety and liability space has been dancing around: what accuracy threshold should we require before deploying AI in life-safety contexts? Gun detection systems have been sold to schools, stadiums, and transit hubs on the promise of real-time threat identification. But if the system misses — and AI systems miss — the legal and ethical consequences are now clearly on the table. Australia has its own AI governance conversations underway; this case will likely become a reference point in those discussions.

Ars Technica

Emphere Bags $2.1M to Automate Vulnerability Remediation with AI

Australian and global security teams know the problem well: vulnerability scanners surface findings faster than engineers can fix them. Emphere has raised $2.1 million in seed funding to tackle that gap with an AI-driven remediation platform aimed at software companies. The idea is to take a scanner's output and automatically suggest or implement fixes, compressing the time between discovery and resolution. It's a crowded space — several larger players have tried versions of this — but the seed stage signals investors still think there's a wedge. The real test will be whether the AI fixes are reliable enough that security teams trust them without manual review, which remains the hard problem.

SecurityWeek

Microsoft Doubles Down on Exclusives as Xbox Showcases 2026 Slate

Microsoft's Xbox Games Showcase at Summer Game Fest delivered a notable strategic reversal: Gears of War: E-Day, previously rumoured for a PlayStation 5 launch, will remain an Xbox and PC exclusive. The move signals a retreat from Microsoft's recent multiplatform strategy, which had seen major titles like Halo land on PS5. The showcase also confirmed Halo: Campaign Evolved — a remake of the original game's campaign — arriving July 28th, and Fable landing February 23rd, 2027. Broader Summer Game Fest highlights included the announcement of Persona 6 and a September launch date for Minecraft Dungeons 2. It was a busy weekend for the games industry, which has been struggling with layoffs and studio closures throughout 2026.

The Verge

Notion's Anthropic Outage Spooked More People Than Expected

Notion experienced a service disruption over the weekend tied to its Anthropic Claude integration, temporarily breaking AI features for users. The outage was resolved, but not before going unexpectedly viral — prompting Notion's head of product to express genuine surprise at the scale of the social reaction. The incident is a small but telling data point about how quickly users have come to treat AI-powered features as table stakes rather than bonuses. When the feature was new, an outage was a shrug. Now, it's a social media event. For product teams building on third-party AI APIs, this is also a quiet reminder that your uptime SLA is only as good as your vendor's.

TechCrunch

GM's $900M Bet on a New Kind of EV Battery

General Motors is committing $900 million to a battery technology gamble it hopes will address the two biggest friction points in EV adoption: cost and charging speed. The investment is part of GM's broader push to reduce dependence on its current Ultium cell platform and explore next-generation chemistries. The scale of the bet is notable — $900 million is a serious commitment even for a company GM's size — and reflects the intensifying competition in the battery space from Chinese manufacturers, who have been rapidly cutting costs and improving energy density. For the EV industry broadly, battery economics remain the central variable. GM is essentially wagering that it can solve them on its own terms.

TechCrunch

Australian Public Servants Recognised in King's Birthday Honours

Australia's 2026 King's Birthday Honours list has been released, with 948 Australians recognised — including 48 recipients of the Public Service Medal, awarded for outstanding service in Commonwealth, state, and territory roles. The breadth of recipients reflects contributions across health, education, emergency management, and digital services. The honours come at a moment when the Australian public service is navigating significant transformation: digital modernisation programs, AI adoption in government, and an expanding mandate under the recently updated Privacy Act. Recognising the people driving that work, often without public visibility, is a worthwhile occasion to note.

The Mandarin

A Hackathon Project Called 'Amazing Digital Dentures' Failed Beautifully

Hugging Face's build-small hackathon wrapped up with a characteristically eclectic set of submissions, but none captured the spirit of the event quite like Amazing Digital Dentures — a project that, by its own admission, failed. The write-up is a refreshingly honest account of what happens when a small team chases a weird idea and runs into the practical limits of current AI tooling. In a field full of polished demos and breathless announcements, a well-documented failure is genuinely useful. It maps the edge of what's actually possible today versus what looks plausible on a napkin. Worth a read for any team running internal AI experiments.

Hugging Face

Sources consulted