Morning Brief, August 26th 2026
Moonshot is negotiating a cut of Microsoft's, Amazon's and Google's revenue for hosting Kimi K3. Alibaba open-weighted a preview of the Qwen4 architecture overnight. Amazon is closing Mechanical Turk after 21 years. And Nvidia reports tonight.
Moonshot wants 30% of what American clouds make on Kimi
Reuters reported this morning that Moonshot AI is in early talks with Microsoft, Amazon and Google on revenue-sharing agreements for hosting Kimi K3 on Azure, AWS and Google Cloud, seeking up to a 30% share of the revenue K3 services generate there. Still unresolved: how the split is computed, data access rights, and auditing of usage. All four companies declined comment, and the talks may die. If they close, it is the first major revenue-sharing pact between a Chinese AI firm and a US cloud, and it rewrites the bargain: until now the American clouds have served Chinese open models without paying their makers anything. The context makes it stranger. Treasury Secretary Scott Bessent has threatened to blacklist Moonshot, and US officials have accused it of stealing from Anthropic's and illegally acquiring Nvidia chips, both of which Moonshot disputes. Meanwhile its pre-IPO round is due to make its final close tomorrow at a reported $50 billion pre-money, ahead of a Hong Kong filing targeted by September 30th. A revenue line paid by three US hyperscalers would be the best page in that prospectus. SourcesBB
Alibaba open-weighted a preview of Qwen4 overnight
Qwen3.8-Flash-Next went live on this morning, in standard and checkpoints under Alibaba's Qwen community license, and the company frames it as an early look at the architecture of the coming Qwen4 family. The design is the story: 125 billion total with about 6 billion activated per token, plus 51 billion parameters of n-gram embeddings that spend capacity on lookup instead of compute, a hybrid attention stack pairing Gated DeltaNet with Qwen's sparse attention, and 262,144 tokens of native context extensible to a million. It reads text, images and video, thinks by default, and Alibaba's own numbers put it at 62.5 on SWE-bench Pro and 91.7 on GPQA Diamond, vendor figures with the usual discount. The claim under the specs is that flagship-adjacent capability now serves at small-model cost, and that claim lands on receptive buyers: AT&T said Monday it routes 40% of its AI queries to open models and wants 70%. Within hours the card was the top trending model on Hugging Face. SourcesAB
Anthropic will pitch investors a $30 trillion market
Anthropic is expected to tell IPO investors its exceeds $30 trillion, the Wall Street Journal reported Tuesday via Reuters, a figure computed from "the full scope of work that could be completed with AI models" and built to top the $28.5 trillion SpaceX claimed in its May filing. The same reporting cites 2028 revenue projections of roughly $190 to 200 billion, against the confidential filed June 1st and an October listing that investors reportedly expect. still showed no public Anthropic filing this morning. Read the TAM construction literally, because investors will: a market sized from all the work AI could do is not a software market, it is the wage bill. A company that prices itself against labor is telling you whose income statement its growth comes out of, and it is not the incumbents' software budgets. SourcesB
Amazon is closing Mechanical Turk, the platform that labeled the training data
Amazon will shut Mechanical Turk on September 30th, ending 21 years of the marketplace Jeff Bezos once called "artificial artificial intelligence," where more than 500,000 people at its peak did the micro-work of the machine-learning era: labeling images, transcribing audio, rating outputs. The company says the decision followed "an internal assessment of its programs and services"; the platform stopped accepting new customers a month ago. The proximate cause is visible in the customer list, with Scale AI, Mercor and Prolific having taken the data business upmarket to credentialed experts, and the models themselves now doing much of what a HIT used to buy. What is quietly retiring is the bottom rung: the place a person with no credentials anywhere on earth could sell judgment to the AI economy by the task. The editorial takes this one up. SourcesBB
OpenAI published its first chip numbers, on Nvidia earnings eve
OpenAI released the first measured results for Jalapeño, the chip it designed with Broadcom, run on SemiAnalysis's public InferenceX with GPT-OSS 120B, DeepSeek R1 and Kimi K2 as workloads. The claims: 1.5 to 1.9 times more work per watt than the commercial systems compared, end-to-end 1.7 to 3.6 times lower, and 2.1 to 4.1 times faster on highly interactive workloads, with hardware lead Richard Ho saying the part serves more work per unit of power while returning responses faster. The comparison set includes Nvidia Blackwell systems, per TechCrunch. Deployment stays small through year-end, with volume planned for 2027. A first-party benchmark from the chip's own designer earns every discount it gets, and the timing earns a separate read: Nvidia's single largest inference customer published per-watt wins over Nvidia silicon the morning of Nvidia's print. SourcesAB
The SEC's Situational Awareness probe reached four banks
The has sent subpoenas to Goldman Sachs, JPMorgan Chase, Citigroup and Bank of America seeking records of Situational Awareness's trades, its borrowing, and its communications with lenders, CNBC reported Tuesday. The fund ran leverage as high as 400% before July's AI selloff forced margin calls, a distressed sale of its public book to Citadel, and a fall in assets from about $45 billion to about $15 billion. The investigation is at its earliest stages and nobody has been accused of wrongdoing. The question has sharpened, though: it is no longer just whether the fastest-growing fund in history had controls to match its size, it is whether the prime brokers who financed 400% leverage on a concentrated AI book did. Four subpoenas to four banks says the SEC is asking the second question. SourcesB
Unitree has given back 200 billion yuan, and the bubble postmortems have started
Unitree fell another 10% Monday to 605.10 yuan by Gasgoo's count, with Reuters putting the close at 603.08, and steadied Tuesday, leaving the stock about 45% below its debut peak and roughly 200 billion yuan of market value gone in four sessions, from an intraday peak near 444.9 billion to about 240 billion. The fundamentals under the fall: 2025 revenue of 1.699 billion yuan, first-half 2026 adjusted net profit of 40 million yuan, and a ratio still above 140 after the decline. Reuters' postmortem collected the quotes the way up never did. Fund manager Dong Baozhen: "all bubbles are doomed to burst." VC Abraham Zhang: the debut "was not fuelled by a rosy prospect, but a desire by some to pump up the shares so as to dump them later." A week ago this listing was the starting gun for China's IPO wave. It still is; what changed is what the runners now know about the finish. SourcesBB
The Robot Games closed on football, fighting and a reordered medal table
The second World Humanoid Robot Games wrapped up Wednesday in Beijing with the 5-a-side football final, the freestyle fighting final and the closing ceremony, after a Tuesday given to dancesport: street dance, ballroom and cheerleading finals drew 108 registered teams against 29 at last year's games, alongside table tennis, tug of war, weightlifting and the standing high jump. The competitive story inverted last year's. AgiBot claimed the overall medal lead after the weekend by its own accounting, Beijing's state-backed innovation center swept the track golds, and Unitree, last year's medal leader, had no athletics medals through Sunday, the same week its stock gave back 200 billion yuan. A full closing medal table had not reached English-language wires this morning. China Daily's tally of the event itself: 2,056 robots, 666 teams, 16 countries, 51 events. SourcesB
Nvidia bought into the company holding 15 gigawatts of Texas power
Nvidia took a direct equity stake in Lancium, the Blackstone-backed developer whose Texas portfolio spans 4 gigawatts of leased capacity and a powered-land pipeline above 15 gigawatts, and made that portfolio a deployment platform for its full AI-factory stack. The announcement leans on two Nvidia systems: DSX MaxLPS, which claims up to 40% more per unit of power budget, and DSX Flex, which manages loads against grid conditions. Neither side disclosed the check; earlier reports of up to $3 billion for about 20% were not confirmed and stay unconfirmed. Place it on the week's map: Texas froze new data-center approvals, PJM wants new loads to bring their own generation, and the binding constraint on compute has moved from chips to interconnection. Nvidia already capitalizes its model customers and its clouds. Now it capitalizes the power layer too. SourcesAB
o3 left ChatGPT on schedule
OpenAI removed o3 from ChatGPT on Wednesday, closing the 90-day sunset announced May 28th. The keeps the model until December 11th with gpt-5.6-sol as the designated replacement, o3-pro stays in ChatGPT for Pro, Team, Enterprise and Edu plans, and o3-mini follows on October 1st. The bookkeeping fact worth keeping: with GPT-4o, GPT-4.1, GPT-4.5, o4-mini and now o3 all retired this year, OpenAI has withdrawn more models in 2026 than in all previous years combined. The model that introduced most of the public to visible reasoning lasted 16 months in the product. Deprecation at that cadence is a real cost for anyone who built on the retired behavior, and it compounds the argument enterprises keep making with their routers: depend on the interface, never the model. SourcesA
Seoul rose into the print, on its chipmakers' own money
The Kospi closed Wednesday at 6,808.21, up 0.97% for a second straight gain, with Samsung Electronics up 1.75% to 261,500 won and SK Hynix up 0.6% to 1,688,000 won. The support is partly self-supplied: SK Hynix has been executing its 40 trillion won ($28.6 billion) buyback-and-cancel program since last Thursday, and Samsung repurchased shares this week as well. Institutions net bought 760.37 billion won of Korean shares Wednesday while foreigners and retail sold a combined 2.36 trillion. After Monday's 3% washout, the memory complex has spent two sessions buying back its own tape ahead of the one earnings call that decides whether the demand story carries into 2027. SourcesBA
OpenAI banned a Russian operation running a fake Israeli think tank
OpenAI disrupted a covert influence operation run from Russia through VPNs, which used ChatGPT to write English-language posts for Substack, Telegram, X, Facebook and LinkedIn promoting the "International Burke Institute," a fabricated Israel-based think tank whose "sovereignty index" ranked Russia favorably. Operators explicitly prompted the model to scrub linguistic tells of Russian authorship, and 34 of the 36 articles attributed to the institute's experts between September 2025 and May 2026 were copied from other sources, some under fake author credits. Reach was minimal; the posts drew few views. The detail that stays useful is the shape: not bot swarms, but a fake institution with a fake metric, laundered through real platforms, built by people who understood that the credibility layer, not the prose, is what a model cannot generate. SourcesAB
Google took Gemini vertical: lawyers and banks the same day
Google Cloud launched Gemini Enterprise for Legal in preview Tuesday, with Cleary Gottlieb, Freshfields, Weil and Williams & Connolly as launch customers, and a financial-services counterpart the same day; healthcare and life sciences are named as next. The legal product bundles four things: packaged skills for contract review, redlining, regulatory scanning and legal research; connectors into document and case systems that inherit existing permissions; that complete work rather than suggest it; and a partner layer spanning the big consultancies and legal-tech vendors. Pricing was not stated. The strategic read is simple: the enterprise checkbook opens by profession, not by model, and the startups that proved legal AI revenue exists now have a selling the same category with the customer's cloud bill as the wedge. SourcesAB
The encrypted reasoning a frontier API returns can be decrypted by anyone
A paper making the rounds this week, from authors including Alexander Panfilov, Maksym Andriushchenko and former Google DeepMind researcher Ilia Shumailov, shows that the encrypted reasoning blocks major providers return for conversation resumption are interchangeable across sessions, users and models within a provider's ecosystem. Inject a stronger model's encrypted trace into a weaker model's conversation and the weaker model decrypts it for you, defeating the anti-distillation logic of hiding reasoning at Anthropic, OpenAI and Google alike. From 315,320 publicly shared reasoning blocks the authors recovered 367 pieces of personal information and 182 credentials, and they show the blocks also work as a prompt-injection carrier. Disclosure was responsible and mitigations are proposed. The industry hid the to protect it, then handed it back to every user in a sealed envelope any sibling model will open. SourcesA
The five-hour caps came back for the $20 tier
OpenAI's five-hour usage windows returned for Plus subscribers on Codex and ChatGPT Work on Tuesday, ending six uncapped weeks that began July 13th. Pro plans at $100 and $200 stay uncapped, and OpenAI's Thibault Sottiaux framed the change as a way to "smoothen the load on our compute." Set it next to Monday's adoption numbers, where fewer than 1% of individual subscribers touch the agent products at all: the constraint on agents at the consumer price point is not demand and not capability, it is that serving long-running agent sessions at $20 a month loses money whenever people actually use them. Rationing at the cheap tier is what the unit economics of agents look like from below. SourcesB
Claude's memory now spans chat and Cowork
Anthropic unified Claude's memory across chat and Cowork, so what the assistant learns in one surface carries into the other, with memory organized into topics a user can view and edit and a setting that walls off sensitive subjects. It is a small release with a large direction: the labs have concluded that memory, not raw capability, is where assistant lock-in actually forms, because a model that knows your context is expensive to leave even when a rival's benchmark is better. Instinct priced that same insight as a perpetual license to user data last week. Anthropic is pricing it as a settings page. The gap between those two answers is where the assistant market's trust fight happens. SourcesB
Ox Alpha's free week ends today, with nobody's name on it
OpenRouter's listing for the now carries an end date: "Going away August 26, 2026." Six days of free frontier-scale service, a million-token , input and tool calling, heavy traffic from agent harnesses, and the provider is still anonymous, with the tokenizer evidence pointing at Microsoft's model lineage and the timing theories still pointing at Zhipu. If the model disappears today unclaimed, someone ran a week-long public load test on real developer workloads, paid for it in free tokens, and kept the results. That data changed hands in exactly one direction. SourcesA
Editorial
The last rung
Mechanical Turk closes September 30th. Most coverage will file it as a nostalgia item, the quirky 2005 platform with the Bezos joke attached, outcompeted and obsolete. Look at who is actually losing something.
At its peak, more than half a million people sold judgment on that platform by the task. Label the image, transcribe the receipt, rate the answer. The academic study that actually measured it, from timestamps on 3.8 million tasks, found a median wage around $2 an hour, with 4% of workers clearing the US minimum. It was, by any standard of the rich world, terrible work. It was also the most credential-free income on the internet: no degree, no interview, no references, no location. A person in Nairobi or rural Mississippi with a laptop and patience could be paid by the machine-learning economy the same afternoon. SourcesA
That work built this industry. ImageNet's labels were bought there. The early content-moderation sets, the sentiment corpora, the transcription that trained the speech models: the foundational datasets of the field were assembled at $2 an hour by people no one at the conferences could name. The models got good enough to do the labeling, the data business went upmarket to Scale, Mercor and Prolific, where the buyers now want PhDs, physicians and lawyers at $100 an hour, and the generalist clickworker stopped being worth the overhead of paying them.
Notice the direction of travel, because it is the same one this week keeps showing from different angles. Amazon repriced its devices up to 60% on Monday: households pay more for the buildout. Amazon closes MTurk on Tuesday: the lowest-paid participants in the buildout lose the way in. The value of human judgment did not go to zero. It bifurcated. If your judgment comes with a credential, it has never been worth more. If it does not, the market that used to buy it by the is being switched off, politely, platform by platform.
The standard answer is that nobody should mourn $2-an-hour piecework, and as a statement about wages it is right. But a rung is not a wage, it is a position: the place you stand while reaching for the next one. Expert data marketplaces have no bottom rung at all. You arrive credentialed or you do not arrive.
My position: the mass-market microtask economy does not get rebuilt. Within a year, no major platform will pay uncredentialed generalists for training-data work at meaningful scale, and the human-data business will have completed its split into expert marketplaces above and unpaid behavioral exhaust below. What would prove me wrong is specific: Prolific or a successor publicly growing its generalist paid-task volume through mid-2027, or a lab launching a new open-enrollment data program. I would welcome being wrong. The people who taught the machines deserved a better severance than a deprecation notice.
Nour Haddad
Prediction Watch
Where the day's news meets our open calls. Each one links to the full prediction, its reasoning and the exact test that settles it.
New call: Kimi lands first-party on a US cloud (Prediction 2026-08-26-B1). Reuters reports Moonshot negotiating revenue-sharing with Microsoft, Amazon and Google. We put it at 0.55 that at least one of the three publicly offers first-party Kimi K3 hosting by the end of March: the commercial logic is strong on both sides, and a threatened Treasury blacklist is the kind of thing that kills deals at the signature stage. Settles March 31st 2027.
Supporting evidence: another bankruptcy estate sells its data for AI training (Prediction 2026-08-18-B1). We said the Spirit Airlines sale, Google's $10 million for the estate's emails, would repeat by June. There is now a YC company, Petrarch, whose entire business is sourcing lab training data from bankruptcy courts; it joins the Startups library this morning. When a one-off trade grows a dedicated intermediary, repetition is the business plan. Settles June 30th 2027.
More likely now: Anthropic's public S-1 lands by Sunday (Prediction 2026-08-24-F1). EDGAR was still empty this morning, but the $30 trillion TAM pitch leaking to the Wall Street Journal is what a roadshow narrative firming up looks like. Five days remain. Settles August 31st 2026.
More likely now: Unitree trades below its IPO price within 90 days (Prediction 2026-08-14-F1). We said the debut pop would not hold through November. A 45% fall from the peak in four sessions is the direction the call needs, honestly stated with the distance it still has to cover: even now the stock sits at four times its ¥150.80 issue price. Settles November 30th 2026.
No change: Nvidia beats the $91 billion consensus (Prediction 2026-08-06-F3). The print lands after today's close. Options pricing the smallest since 2021 is a fact about positioning, and still not a fact about the number. Settles August 31st 2026.
No change: Z.ai releases GLM-5.3's open weights by Sunday (Prediction 2026-08-16-T2). The company's Hugging Face organization showed no GLM-5.3 repository this morning, with its own roughly August 28th target now two days out. Settles August 31st 2026.
No change: Ox Alpha is claimed by a Chinese lab (Prediction 2026-08-23-T1). The free preview ends today with the provider still anonymous and the evidence unchanged: analysis pointing at Microsoft, timing theories pointing at Zhipu. If it vanishes unclaimed, the call has two months for a name to surface. Settles October 31st 2026.
China and open weights. Weights shipped: Qwen3.8-Flash-Next went live on Hugging Face overnight, the first open checkpoint of the Qwen4 generation, at 125 billion parameters with 6 billion active. Moonshot negotiating a hyperscaler revenue share would monetize open weights a second way. GLM-5.3's stay gated, and the 2.8-trillion-parameter release Prediction 2026-08-06-T5 waits for has not appeared.
What did not happen. Nothing settled today. Nvidia's print is tonight. Moonshot's pre-IPO close is tomorrow, not today as earlier reporting had it. No Anthropic S-1 on EDGAR, no ruling in xAI's Minnesota case, no published text for the Texas data-center order, and no closing medal table from Beijing in English wires by press time.
Sources
- Exclusive: China's Moonshot in talks with Microsoft, Amazon, Google on revenue sharing — Reuters via Yahoo Finance B
- Moonshot AI targets August 27 closing for pre-IPO round ahead of Hong Kong filing — KrASIA B
- Qwen3.8-Flash-Next — Hugging Face A
- Alibaba's Qwen to open-source Qwen3.8-Flash-Next, previewing Qwen4 architecture — TechNode B
- Nvidia fiscal Q2 2027 earnings outlook: what to watch on August 26 — Investing.com B
- Nvidia shares gain ahead of earnings, options market prices limited volatility — TradingKey B
- Nvidia price target raised to $352 at Raymond James — MarketBeat B
- Anthropic expected to tell investors it sees $30 trillion market — Reuters via Yahoo Finance B
- Amazon is shutting down Mechanical Turk — Quartz via Yahoo Finance B
- Amazon service that Jeff Bezos called 'artificial AI' is shutting down — CNBC B
- Jalapeño: first results — OpenAI A
- OpenAI's Jalapeño chip is built for fast inference at scale, benchmarks show — TechCrunch B
- SEC subpoenas Wall Street banks in Situational Awareness probe — CNBC B
- Unitree Robotics shares fall 10%, market cap drops by over RMB 200 billion — Gasgoo B
- China robot maker Unitree's post-IPO plunge stirs bubble fears — Reuters via Yahoo Finance B
- World Humanoid Robot Games: dancesport and finals — China Daily B
- Lancium announces partnership with NVIDIA to advance gigawatt-scale AI factory development — PR Newswire A
- Lancium, Nvidia partner on gigawatt-scale AI data centers — Data Center Knowledge B
- ChatGPT release notes — OpenAI A
- Kospi closes nearly 1% higher ahead of Nvidia earnings release — Korea JoongAng Daily B
- Share buyback and retirement — SK hynix A
- Disrupting malicious uses of AI: influence campaign from Russia — OpenAI A
- OpenAI bans Russian ChatGPT accounts behind a fake think tank — Tom's Hardware B
- Introducing Gemini Enterprise for Legal — Google Cloud A
- Google launches Gemini Enterprise for Legal — Artificial Lawyer B
- Stealing Reasoning Traces from Proprietary LLM APIs — arXiv A
- Exclusive: AI founders walked away from Bezos-backed Prometheus to launch physics AI startup — Reuters via Yahoo B
- OpenAI abruptly restores 5-hour Codex and Work limits for ChatGPT Plus — Notebookcheck B
- Cathie Wood loads up on Cerebras Systems stock — Benzinga B
- What's going on with Cerebras stock — Benzinga B
- Claude Cowork finally remembers what you told the app in chat — TechCrunch B
- Ox Alpha — OpenRouter A
- A Data-Driven Analysis of Workers' Earnings on Amazon Mechanical Turk — arXiv A