AIAI News Online
Today in AI · 8 October 2026

AI today, 8 October 2026: Haiku 5.5 at a tenth of the price, GPT-6 for every ChatGPT tier, Windows gets agent sandboxes

13 things that happened in AI, each in under a minute. Today's thread is cost. Anthropic cut its small-model price by three quarters, OpenAI pushed GPT-6 to ChatGPT's free tier, and Microsoft pitched local models on Windows as intelligence with no meter running. The bill for all of that shows up elsewhere: Samsung booked a record quarter selling the memory chips, Broadcom and SpaceX went to private lenders for $90 billion of chip financing, and Google signed for more than 6 gigawatts of power while a Finnish regulator stopped its bulldozers.

7 min read31 sources
Today's deep diveMistral Large 4 explained: the 1-trillion-parameter model you can own15 min · from plain English to the deep end →
  1. Models · Big

    Anthropic launches Claude Haiku 5.5 at $0.10 per million input tokens, 75% below Haiku 4.5

    Anthropic released Claude Haiku 5.5, its small model, priced at $0.10 per million input tokens and $0.50 per million output tokens for prompts up to 100,000 tokens, with higher rates above that. Anthropic says it costs about 75% less to run than Haiku 4.5 on average and is the first Haiku with an adjustable effort setting. Vendor-reported scores include 72.4% on OSWorld 2.1 (a computer-use test) against 15.7% for Haiku 4.5, and Anthropic says it beats OpenAI's GPT-6 Luna on every benchmark it lists. Alongside the launch, Anthropic halved Sonnet 5.5 cache-read pricing to $0.10 per million tokens and said Max and Team subscribers will get $100 to $500 a month in API credits.

    Why it matters: The cheapest tier of frontier-lab models is where most high-volume agent and classification work runs, and a tenfold price cut on small prompts resets what that work costs. It also sharpens the price war with OpenAI's Luna line and Mistral's new Large 4.

    Anthropic · Techmeme

  2. Products · Big

    OpenAI rolls GPT-6 and 'Intelligent UI' to every ChatGPT tier, free users from today

    OpenAI began rolling GPT-6 with Intelligent UI to ChatGPT Plus, Pro, Business and Enterprise users on 7 October, with Free and Go tiers following on 8 October, TechCrunch reported. Paid tiers get GPT-6 Sol and free tiers get GPT-6 Luna, per MacRumors. Intelligent UI lets ChatGPT answer with interactive components such as charts, forms, tappable buttons and purpose-built calculators, built from a library of streamable components as the model writes them; users can dial the visuals back. GPT-6 can also start answering before it finishes reasoning, which OpenAI says cuts average wait times by 44% (vendor-reported). GPT-6 is limited to the Chat tab: Codex and the Work product keep their current models.

    Why it matters: This is the first time GPT-6 reaches ChatGPT's free tier. The interface change matters as much as the model: chat answers that build their own mini-apps move ChatGPT closer to a software platform than a text box.

    TechCrunch · MacRumors · OpenAI Help Center

  3. Products · Big

    Microsoft ships OS-level agent sandboxes for Windows and local models on RTX Spark PCs

    At a Windows and Surface event in San Francisco, Microsoft said Microsoft Execution Containers (MXC), a policy-controlled sandbox that limits which files and networks an agent can touch, are now generally available on Windows 11, with GitHub Copilot, OpenAI Codex, OpenClaw, Replit and LM Studio already supporting them and Anthropic's Claude Code and Perplexity to follow. Satya Nadella also announced MAI-Code-1.1 Flash, a 137-billion-parameter coding model with a 256,000-token context window optimised to run on a PC, which GitHub Copilot can hand work to locally with no cloud token spend. Nvidia's RTX Spark chip, with up to 128 GB of unified memory, powers the new Surface Laptop Ultra (preorders from $2,599, shipping 16 October) and a $5,999 Surface RTX Spark Dev Box due in November, Tom's Hardware reported.

    Why it matters: Microsoft is betting that agents will live on the PC rather than only in the cloud, and the sandbox is the piece that makes that defensible for IT departments. Local models with no metered token cost also change the economics of coding assistants.

    Microsoft Windows Developer Blog · NVIDIA · X · Tom's Hardware

  4. Business

    Broadcom reportedly seeks $50bn+ for OpenAI's chip; SpaceX lines up $40bn for Nvidia GPUs

    The Wall Street Journal reported that Broadcom is working to arrange more than $50 billion in financing for the custom AI chip it is building with OpenAI, approaching lenders including Apollo and Blackstone, with a deal targeted before year-end; the talks are early and the size could change, per a summary of the report. The same report said Oracle is in talks with Apollo and Goldman Sachs to fund a large chip purchase through a separate entity that would lease the chips back. Separately, Reuters confirmed that SpaceX is seeking about $40 billion, roughly $10 billion in bank loans and $30 billion in investment-grade debt, to buy Nvidia chips, with Apollo expected to lead and PIMCO among lenders. Apollo and PIMCO declined to comment; SpaceX and Nvidia did not respond.

    Why it matters: The AI buildout is shifting from corporate bonds to bespoke private-credit and lease structures, which keep debt off balance sheets but spread the risk across more lenders. Morgan Stanley estimates the sector needs $1.5 trillion in outside financing by 2028.

    Reuters via Yahoo Finance · investingLive, citing The Wall Street Journal

  5. Hardware

    Samsung posts record 107 trillion won quarterly profit on AI memory, up 782% year on year

    Samsung Electronics' preliminary third-quarter results show operating profit of 107.4 trillion won (about $80 billion), up 782.5% from a year earlier, on revenue of 195 trillion won, the Korea Times reported. It is the first Korean company to top 100 trillion won in a quarter. Analysts attribute the result almost entirely to memory chips, including HBM4 high-bandwidth memory for Nvidia's Vera Rubin platform and rising DRAM prices, while the consumer-devices division is expected to post a second straight loss. Full results come on 29 October.

    Why it matters: Memory is now the bottleneck and the profit centre of the AI hardware stack. Record margins for Samsung mean higher component costs for everyone else who builds on DRAM and HBM.

    The Korea Times

  6. Safety

    Common Sense Media rates ChatGPT for Teens an 'unacceptable risk'; OpenAI disputes it

    Common Sense Media's Youth AI Safety Institute rated ChatGPT for Teens an Unacceptable Risk after running more than 4,000 prompts on accounts registered to 13- to 17-year-olds. It says the chatbot missed more than one in four warranted crisis referrals, that parents of newly linked test accounts received no alerts after up to an hour of conversation about self-harm, and that study-hour limits could be switched off by deleting a prefix. It recommends OpenAI restrict ChatGPT to adults until the gaps are fixed. OpenAI said in a statement reported by The National News Desk that it does not believe the testing 'accurately reflects how ChatGPT's teen safeguards work in practice', arguing much of it may have run before parental-control activation, which can take a few hours, was complete, and asking for a retest. The institute discloses that its funders include the OpenAI Foundation.

    Why it matters: This is the first large independent test of OpenAI's teen product since it launched in August, and it lands as regulators in the US and EU tighten rules on minors and chatbots. The parental-alert finding is the one OpenAI will have to answer most directly.

    Common Sense Media · The National News Desk via ABC News 4

  7. Hardware

    Google signs for 3,590 MW of Constellation nuclear power, including 890 MW of new uprates

    Google and Constellation announced a 20-year power purchase agreement that funds more than $4.3 billion of upgrades to 11 nuclear units in Illinois, Pennsylvania and New Jersey, adding 890 MW of new capacity to the PJM grid with the first uprate expected by 2028. A separate 15-year deal covers a further 2,700 MW from Constellation's existing plants, and Constellation will adopt Google Cloud and Gemini Enterprise for grid operations. Google did not disclose the price it will pay. The Elec reported Google also agreed a 2.7 GW supply with Black Hills for a planned data centre in Cheyenne, Wyoming, backed partly by new gas generation.

    Why it matters: At more than 6 GW across two days, this is one of the largest power commitments any AI company has made, and it leans on existing reactors rather than new builds. Expect rivals to compete for the same limited pool of uprate capacity.

    Google Cloud · SiliconANGLE · The Elec

  8. Policy

    Finland orders Google to halt site work at two data centres pending environmental reviews

    Finland's Licensing and Supervision Authority (LVV) on 6 October ordered Google's subsidiary Tuike Finland to suspend tree felling, excavation and road building at planned data-centre sites in Muhos and Kajaani by 23 October and to explain itself in writing by 14 October, the Helsinki Times reported. Roughly 330 hectares at Muhos and just under 200 at Kajaani had already been cleared or prepared before the required environmental impact assessments were complete, per Gadget Review. The sites are part of Google's 13 billion euro Finnish data-centre programme, its largest single investment in Europe. Google said it 'did not meet the high standards we set ourselves', that it had acted in good faith under the Forestry Act, and that it will comply with the order.

    Why it matters: Europe's data-centre boom is running into permitting law, and this is a rare case of a regulator stopping a hyperscaler mid-build. A delay at Google's largest European investment is a warning for other operators clearing land ahead of approvals.

    Helsinki Times · Gadget Review via Yahoo News

  9. Products

    Musk says SpaceX's Grok Bot will route tasks to Claude Opus 5.5, Midjourney and Suno

    Elon Musk posted on X that SpaceX 'will use the best back end model for any given task, including Claude Opus 5.5, MidJourney, Suno and other leading APIs' for Grok Bot, the agent app SpaceXAI launched in beta in August, tbreak reported. A SpaceXAI staffer who works on Grok Bot added on X that 'all bots will be powered by opus 5.5'. No timeline, cost allocation or data-sharing terms were published, and SpaceXAI has not said whether users can see or choose which model handles a task.

    Why it matters: A Musk-run product defaulting to a rival's frontier model is an admission that Grok is not the best engine for agent work. It also deepens an unusual relationship: Anthropic already rents compute in SpaceX's Colossus data centre.

    tbreak · AI Weekly

  10. Products

    Google launches Playground, a no-code game maker built on Gemini; Unity Spark to follow

    Google opened Playground, an experimental browser platform where users describe a game in plain language and get a playable version they can refine by chat, share by link or publish to a public gallery. It is available to US users aged 18 and over, with creation limits tied to Google AI subscription tier. Games are generated by Gemini for logic, Nano Banana for visuals and Lyria for music, The Next Web reported. A closed beta of Unity Spark, which adds the Unity runtime and 3D mechanics, is 'coming soon'. Roblox shares fell as much as 8% before the open.

    Why it matters: Google is aiming AI generation at user-made games, Roblox's home turf, with distribution through the browser and Google Play Games profiles. It is also a consumer showcase for stacking several Google models in one product.

    Google · The Next Web

  11. Open source

    Nous Research raises $90m at $1.5bn valuation to take open-source Hermes agent to business

    Nous Research closed a $90 million Series B led by Robot Ventures at a $1.5 billion valuation, confirming earlier TechCrunch reporting and taking total funding to $158 million. Nous named Nvidia, Microsoft's M12, Samsung Next, Union Square Ventures, Y Combinator and Menlo Ventures as participants, Runtime Wire reported. Hermes Agent is an MIT-licensed agent with persistent memory, code execution and scheduled tasks that runs on users' own machines or Nous's hosted service. The money funds Hermes for Businesses, an enterprise tier with team controls, single sign-on, private or on-premises deployment and service-level agreements, plus a planned mobile app.

    Why it matters: Open-source agents are becoming venture-scale businesses by selling the governance layer rather than the model. Microsoft naming Hermes among the tools adopting its new Windows agent sandbox shows how fast that layer is being standardised.

    TechCrunch · Runtime Wire · X

  12. Products

    Meta, Sierra, Walmart and Stripe publish a sign-in standard for personal AI agents

    Sierra and Meta, with Walmart, Stripe, Shopify, Genesys, Rocket and Instinct, announced the Personal Agent Protocol on 6 October, an open standard for how a consumer's AI agent authenticates with a business and what it is allowed to do. Sessions are built on OAuth: an agent can start as a guest, then get read-only or write access once the customer signs in, The Next Web reported. A v0.1 specification is due later in October; payments are a future extension. OpenAI, Anthropic, Google and Amazon are not on the partner list, and no specification, licence or governing body has been published yet.

    Why it matters: It arrives weeks after Amazon blocked Meta's Muse agent from shopping on its store, and it overlaps with Visa's Trusted Agent Protocol, which Stripe and Shopify have also joined. Whoever sets the sign-in standard controls which agents retailers trust.

    Sierra · The Next Web

  13. Models

    Mistral previews 1-trillion-parameter Large 4, open weights promised by end of October

    Mistral on 6 October opened an API preview of Mistral Large 4, nicknamed 'Le Chonk', a natively multimodal mixture-of-experts model with about 1 trillion total parameters and roughly 50 billion active per token, Decrypt reported. Preview pricing is $1.36 per million input and $4.18 per million output tokens. Weights are not yet released: Mistral says it is first red-teaming the model with cybersecurity leaders, vetted partners and state authorities and will publish open weights by the end of the month, without naming a licence. Independent tester Artificial Analysis scored the preview 38 on its Intelligence Index, the highest for any model from outside the US and China but well behind the frontier, while Mistral says it ranks in the top five on the cyber index (vendor-reported).

    Why it matters: It is the largest model a European lab has promised to open, and Mistral is pitching it at governments wanting sovereign AI. Until the weights and licence land, it is a closed API model competing on price.

    Mistral AI · Decrypt · Stork

How this was made: compiled by an AI model (Claude) from the linked sources and checked item by item in a separate AI fact-check pass. Corrections: [email protected].