2026-09-04PolicyThe OpenAI copyright case now turns on 8.2 million Copilot chat logsMicrosoft, OpenAI and the news publishers all moved for summary judgment on 4 September. The briefs are the first public view of what three years of discovery actually found.7 min read→
2026-09-03FundingNVIDIA agreed to buy Hugging Face, and $1 billion of the price is retentionThe headline number is $12.93 billion. The SEC filing says $11.9 billion goes to shareholders, the rest is an employee retention pool, and nothing closes until 2027.6 min read→
2026-09-02Model releaseGemini 3.8 Flash keeps the price per token and raises the tokens per taskGoogle's third Flash model in six weeks ships at the same $0.75 / $3.75 rate, but is explicitly built to spend more tokens. The security variant is gated behind a 650-partner access program, and the introductory price doubles on 1 January.7 min read→
2026-09-01Model releaseFable 5.1 holds its price, quarters its cache reads, and breaks three thingsAnthropic's new frontier model costs exactly what the old one did. The changes that will actually reach your code are in the cache pricing and in three API behaviours that now return a 400.8 min read→
2026-08-31SecurityAnthropic trained a model on its own broken RL environmentsA month after Claude models reached the live internet from inside evaluation sandboxes, Anthropic published what it changed — and an experiment showing what defective training environments produce.8 min read→
2026-08-28ProductGemini Notebook's new limits count your sources, not your promptsFrom September 2, Google replaces the research tool's daily prompt caps with compute-weighted limits that refresh every five hours — and factor in how many sources your notebook holds.4 min read→
2026-08-26InfrastructureNVIDIA's revenue doubled, and its receivables grew by $22 billionSecond-quarter revenue reached $96.2 billion and guidance points to $108 billion. The same filings show supply commitments at $279 billion, $24.9 billion of new debt, and half a trillion dollars of outside capital being arranged to finance customers.6 min read→
2026-08-25InfrastructureOpenAI's first chip has benchmarks now, and OpenAI ran themAt Hot Chips on Tuesday, OpenAI published the first measured results for Jalapeño, the inference accelerator it designed with Broadcom. The claim is more throughput per kilowatt and lower per-user latency at the same time, against an NVIDIA Blackwell system. Volume deployment is a 2027 story.5 min read→
2026-08-24InfrastructureNVIDIA ships a decode chip with a competitor's name on itGroq 3 LPX enters full production at Hot Chips, splitting inference across two kinds of silicon. The efficiency figures are NVIDIA's own and still pending review.6 min read→
2026-08-21Developer toolsOpenAI finally discounts GPT-5.6 Sol, and puts an expiry date on itSol's standard rate drops to $4 per million input tokens and $20 per million output tokens. OpenAI's changelog calls the price promotional and good at least through 21 November.4 min read→
2026-08-20InfrastructureAlibaba's new AI segment shows the model lab losing $2 billion a quarterA reporting change splits the compute business from the one that builds and serves the models. Compute grew 45% and more than doubled its profit. Labs and apps lost RMB13.9 billion on RMB3.3 billion of revenue, which Alibaba attributes partly to Qwen app inference costs.5 min read→
2026-08-19PolicyOpenAI previews a way to police API misuse without reading the promptsPrivate Safety Processing looks for abuse patterns across a customer's sessions and emits a signal instead of the text. It arrives while Anthropic is doing the opposite for its frontier models.5 min read→