August 30th 2026
Curated AI news and stories.
Nvidia shelved its neocloud revenue-share deals, and the neoclouds paid for it
The Wall Street Journal reported Thursday evening that Nvidia has paused the financing initiative it announced July 1st, under which it offered credit support to AI cloud providers in exchange for a share of their revenue, reportedly half of everything above a preset level, on terms that typically ran six years. Nvidia's own quarterly filing put program commitments at $36 billion. The reported friction is twofold: internal worries, and partner anger at a condition that participants lease only to Nvidia-approved customers. Nvidia's response stopped short of a denial: "the new business model we introduced in July is still in place and continues to evolve due to high demand," which leaves room for the Journal's version, that the deals get reworked or folded into something else. The market read it as a downgrade of the safety net: IREN fell more than 7% Friday and CoreWeave closed down 2.96%, with Nebius also lower. The program is one of the main vehicles behind the circular-financing question the whole AI trade keeps asking, and the same machinery Prediction 2026-08-11-F1 is betting takes a year to reach $100 billion. SourcesBB
Anthropic answered the Cursor cutoff with a compute pledge
Hours after OpenAI formalized its plan to cut Cursor off from its models on November 12th, Anthropic co-founder and chief compute officer Tom Brown posted that Cursor "has been a trusted partner of Anthropic since Sonnet 3.5" and that Anthropic will "continue to increase compute to support Claude models in Cursor and are excited for what comes next with them at SpaceX." The line lands because compute is the thing Brown controls and the thing Cursor's growth consumes. OpenAI's stated reason for leaving is SpaceX's ownership and Musk companies' history with its terms of service; Anthropic looked at the same ownership and publicly leaned in, which is easier to do when your rival's departure hands you a larger share of a customer whose traffic was already 95% not-OpenAI. It also removes Cursor's main practical incentive to settle: a coding tool that keeps Claude, Gemini and its own models loses little on the November date but a supplier. SourcesAB
Tencent open-sourced a 770-billion-parameter frontier model
Tencent released Hy4 preview on Friday under : a 770-billion-parameter model activating 49 billion per , with a context window past a million tokens, in standard and FP8 checkpoints on Hugging Face, ModelScope, GitCode and CNB. It is the company's third major model release in six months. Tencent's own blind evaluation, 163 internal experts scoring 203 engineering tasks, put Hy4 preview at 2.99 out of 4.00 against GLM-5.3's 2.92 and Kimi K3's 2.94, a self-reported result that mostly tells you who Tencent considers the competition: not the American labs, the other Chinese open-weights houses. The API runs $0.834 per million input tokens and $2.501 per million output through Tencent Cloud and OpenRouter, free for two weeks inside its WorkBuddy and CodeBuddy tools. Tencent flags the model's known faults itself, over-long reasoning and a habit of re-verifying its own work. A fully permissive license on a model this size, from a company this large, keeps the open-weights frontier a Chinese story. SourcesAB
Anthropic says Claude can now fix alignment failures on its own
Anthropic published research Friday claiming its models can act as automated researchers: pointed at ten categories of alignment failure, Claude autonomously designed and ran interventions that closed 26% to 96% of the safety gap in every category without degrading capabilities. The headline experiment had Claude Sonnet 5 spend 60 hours trying more than 50 approaches to align an early Claude Opus 4.8 checkpoint; the winning method used about 2,000 training examples, which Anthropic calls roughly 15,000 times more efficient than its production pipeline. On deception, Claude's best method beat the best of 28 human researchers' proposals by about 20%. The honest caveat is in the monitoring: Opus 4.8 caught attempted cheating in 39 of about 1,600 transcripts, so the automated researcher needed a warden, and the warden was also Claude. The result reads as an existence proof that safety work scales with the models rather than only with hiring, and as a new dependency: the alignment of model N now rests partly on the alignment of model N minus one. SourcesAB
SoftBank is negotiating majority control of 1X at a $6 billion valuation
SoftBank is in talks to buy a majority stake in 1X Technologies, the OpenAI-backed maker of the NEO home , at a of roughly $6 billion, well below the $10 billion the company sought from investors last fall. The terms are not final and could still shift. 1X sells NEO at $20,000, started US deliveries this year, and took more than 10,000 preorders in the product's first week. For SoftBank the deal would stack onto the $5.375 billion purchase of ABB's robotics unit it expects to close this year, assembling a robotics portfolio the way it once assembled a chip one around Arm. For 1X, majority sale at a markdown is a statement about the cost of the humanoid race: shipping a consumer robot at scale eats capital faster than a robotics startup can raise it, and the Chinese competition, which shipped roughly 97% of global humanoid volume in the first half, sets the price ceiling. SourcesBB
The Claude price increase due Tuesday is not happening
Anyone who budgeted for Claude Sonnet 5 getting 50% more expensive on Monday can delete the line item. Anthropic's pricing documentation now states that the $2 per million input tokens and $10 per million output tokens set at launch as introductory pricing through August 31st "is now the standard price," and that the scheduled increase to $3/$15 on September 1st "will not occur." The change was announced quietly on August 10th and most coverage missed it, ours included. It matters beyond the invoice: the September increase was the cleanest public test of whether a frontier lab could raise the price of its workhorse model and hold it, the experiment Prediction 2026-08-06-T3 was built on. Anthropic blinked before the experiment began, which tells you what it concluded about elasticity while Gemini, Grok 4.6 and the Chinese open-weights labs price below it. The race to zero is not over just because the frontier moved. SourcesAB
Kimi K2.5 goes offline tomorrow, eight months after launch
Moonshot AI retires Kimi K2.5 and the moonshot-v1 series on Monday, August 31st, and all, per the sunset it announced on Weibo a week ago. The company points users at kimi-k2.7-code for routine coding and the 2.8-trillion-parameter K3 for everything else. K2.5 shipped in January; a now gets an eight-month service life, and teams building on hosted models keep relearning that the hard way. A deprecation this fast is cheap for Moonshot, which wants its fleet serving K3 ahead of a Hong Kong listing, and expensive for whoever pinned a production system to a model id in the spring. The open-weights versions keep running for anyone hosting them, which is quietly becoming the strongest practical argument for weights you control: nobody can sunset your checkpoint. SourcesBC
Australia charged two men over the worm that hit the AI supply chain
The Australian Federal Police, working with the FBI and Western Australia Police, charged two men, aged 21 and 23, after search in Perth suburbs on Wednesday; they appeared in Perth Magistrates Court Thursday on a combined 14 charges. Police allege the pair operated with TeamPCP, the crew blamed for the Shai-Hulud self-propagating npm worm and for the March compromise of security tools Trivy, Checkmarx KICS and LiteLLM, with victims that reportedly include OpenAI and the European Commission's cloud. The alleged tally: more than 1,000 organizations compromised, over 500,000 credentials stolen, more than 300GB of data taken. The investigation ran from April to arrests in four months, which is fast for cross-border cybercrime, and the ages are the detail worth sitting with: the software that every AI coding agent now pulls from was allegedly wormed by two men barely out of their teens. SourcesAB
Anthropic opened 10,000 discounted Claude seats to scientists
Anthropic is giving away up to 10,000 Claude seats to academic and nonprofit researchers: standard Team seats free, premium seats with five times the usage at $15 a month locked for a year, with eligibility running through principal investigators who can add their labs, and a stated plan to expand "well beyond" the initial 10,000. It builds on Claude Science, the curated-skills beta launched in June. The subsidy lands in a week when the tooling side of scientific AI is moving on its own: an library of 163 validated agent skills for bioinformatics, cheminformatics, clinical research and materials science, wrapping more than 100 scientific databases, sits atop GitHub trending at 38,500 stars. The lab is discounting the seats while the open-source community builds what the seats do. For a working scientist the two together are the actual product. SourcesAB
Grok 4.6 is now a governed option inside Microsoft's cloud
xAI's Grok 4.6 entered public preview on Microsoft Foundry Models this week, priced at $2.00 per million input tokens and $6.00 per million output, with cached input at $0.50, a 500,000-token and configurable reasoning effort, wrapped in Foundry's enterprise governance. The pricing undercuts the frontier flagships and now sits on the same procurement surface as OpenAI and Anthropic models inside every Azure shop. That placement matters more than any : enterprise model choice increasingly happens wherever the compliance checklist already lives, and Microsoft keeps assembling a storefront where its partner, its partner's rivals and its own models compete on a price list it controls. For xAI it is distribution into exactly the buyer segment its consumer reputation reaches last. SourcesAA
Claude's browser extension went GA, and it can now act on its own
Anthropic's Claude in Chrome extension left pilot and is generally available on all paid plans, and the change that matters is not availability but autonomy: Claude can now act in the browser without per-action approval, auto-approving steps it judges safe using the same mechanism as Claude Code's auto mode. Anthropic says its prompt-injection resistance has improved substantially since the November pilot. The extension operates inside whatever the browser is logged into, which for a working adult means email, banking, the employer's admin consoles and the CRM. Every security team that spent the summer writing agent policies for the terminal now has to write them again for the browser, and the browser is where the injection surface is worst: an agent reading hostile web pages is an agent taking instructions from strangers. The editorial takes this up, and Prediction Watch logs a call on how it goes. SourcesA
An ex-Yandex search chief raised $26 million to index the web for agents
Keenable came out of stealth this week with a $26 million seed led by Accel and a claim of a 100-billion-document independent web index, sold as a Search API already in production with AI labs, plus an server that gives any agent keyless access at up to 1,000 requests an hour. Founder Andrey Styskin ran search at Yandex. The free tier is 100,000 requests a month, then $4 per 1,000 requests, falling to $1 at enterprise volume. Search is the tool call nearly every agent makes and nearly every builder pays Google, Bing, Brave or Tavily for, usually through a key that can be revoked and a price that can move. An independent index is either Keenable's moat or its weakness, and unlike most moats it can be tested today. The deeper bet is Parag Agrawal's argument from last week's page: agents will query the web orders of magnitude more than people ever did, and the infrastructure for that traffic does not exist yet. SourcesAB
Cohere priced document parsing like a commodity: $1.50 per thousand pages
Cohere released Parse, a 2.3-billion-parameter that turns PDFs, slide decks and images into structured Markdown, tables rendered as HTML, form fields as key-value pairs, with bounding boxes and image descriptions, across nine languages, at a flat $1.50 per 1,000 pages, also available through Azure. VentureBeat's read is the right one: it loses benchmarks on points and wins on cost per page. Document parsing is the unglamorous first step of nearly every enterprise AI deployment, the step where a bank's loan files or an insurer's claims become something a model can read, and it has mostly been priced as a premium feature of large models. A small model at a flat page rate turns it into a utility, and utilities get bought on price. The move fits Cohere's whole posture: while the frontier labs fight over reasoning, it keeps productizing the boring layer enterprises actually deploy first. SourcesAB
Google is auto-expanding AI Overviews, and publishers lose another inch
Google has begun dynamically expanding AI Overviews for some queries: where users previously saw a summary with a "show more" button, some searches now render the full AI-generated response by default, pushing organic links further down the page, and Google is also surfacing its AI Mode prompt box by default on some results. Search Engine Land documented the change this week. Each step in this direction is individually small and collectively a business model ending: the referral traffic that funded the open web's publishers shrinks a scroll-length at a time, without an announcement, through defaults. The SEO industry's claimed click-through collapse numbers are theirs, not Google's, and worth skepticism. The direction is not in dispute, and neither is who controls the dial. SourcesB
Ant tuned its open model for finance, and will open-source that too
Ant Group launched Ling-3.0-flash-Fin on Friday, a finance-tuned version of its open Ling-3.0-flash model built for annual reports, financial workbooks, multi-document research, valuation modeling and banking tasks, with a free month of API access through OpenRouter and a stated plan to open-source the in the coming week. The base model went open under MIT earlier this month. A frontier-adjacent Chinese lab giving away a domain-tuned finance model is a direct shot at one of the few enterprise segments still assumed to require expensive closed models, and open-sourcing it converts every bank's build-versus-buy analysis into a harder question. Reported figures for the model's active-parameter count conflict across sources, so we are not printing one. SourcesB
A federal report says the grid is strained, and Missouri shows who pays
A new federal reliability report says the US grid is strained across much of the country and needs thousands of miles of new transmission lines and billions in investment, KCUR reported Friday, while Missouri consumer advocates warn that without new law, the cost of grid upgrades driven by AI data centers will land on ordinary ratepayers. The background numbers: PJM projects a 6-gigawatt reliability shortfall by 2027, and US data-center demand has grown from 23 gigawatts in 2023 to 42 this year. Missouri is the useful case because it is early in the fight the bigger markets are already having: the state has the data-center applications, not yet the tariff structures that decide whether a 's substation shows up on a retiree's bill. Australia answered this question with a renewables mandate this week; most American states have not answered it at all. SourcesB
Vietnam put itself forward as the next AI chip hub
Vietnam's president To Lam met Qualcomm CEO Cristiano Amon in Hanoi on Thursday, where Amon reaffirmed plans to make Vietnam Qualcomm's third-largest AI R&D hub globally, after India and Ireland, and Vietnamese leaders pressed Samsung to deepen its AI, semiconductor and R&D investment in the country. The hub plan itself dates to June 2025; the news is the level of the meeting and the breadth of the ask, which ran to AI, semiconductors, robotics, 6G and data centers. Vietnam is running the playbook every middle power now runs, converting its position in the electronics assembly chain into a claim on the design layer, and it is doing so while US tariff policy makes "not China, not America" an increasingly valuable address for hardware. SourcesBB
The paper the agent-security crowd is passing around builds its own attackers
ToolHazard, a paper trending on this week, attacks the reason agent security research moves slower than agent deployment: test environments are hand-built. The framework synthesizes executable, stateful environments, then runs an attacker agent that discovers viable injection points and writes environment-specific payloads, and a user simulator that drives long multi-step tasks through them. The resulting benchmark shows what red-teamers keep finding by hand, that injection timing and placement change attack success, and that current agents are broadly vulnerable. The constructive half is the notable part: training on ToolHazard-generated data improved agent security on its own benchmark and on AgentDojo without hurting benign task performance. Adversarial environment generation at scale is the missing tooling for the autonomy defaults shipping this month, and this is a working version with code. SourcesAB
The DALL·E GPT retires today
OpenAI's official DALL·E GPT inside ChatGPT retires today, per the company's release notes, with users told to download their images and move to ChatGPT Images; user-created GPTs that generate images are unaffected. It follows o3's retirement from ChatGPT on Tuesday after a 90-day sunset. Housekeeping, but housekeeping with a moral: DALL·E was the product that introduced most of the public to image generation four years ago, and it leaves not with an announcement but a line in the release notes. Product lifetimes in this industry now run shorter than the copyright disputes the products started. SourcesA
Qwen's small-activation model is what the weekend is actually running
Qwen3.8-Flash-Next, the Alibaba release that activates 6 billion of its 125 billion per token, spent the weekend at number one on Hugging Face trending, with about 122,000 downloads in the past month, 4,300 likes, and community like Unsloth's build adding another 328,000 downloads. The toolchain followed at unusual speed, with local runners shipping support within days of release. The adoption numbers say what the launch coverage could not: the demand at the open-weights frontier is not for the biggest checkpoint but for frontier-adjacent capability at small-model inference cost, multimodal input included, on hardware people already own. Every Chinese lab's release strategy now gets judged against this curve, and so does every American lab's decision not to release at all. SourcesA