Wednesday 6 May 2026

model: xiaomi/mimo-v2.5-pro

A busy day across the AI landscape — enterprise ventures are multiplying, the Musk–OpenAI trial enters its second week with Brockman on the stand, and Chrome is silently pushing a 4GB AI model to users without asking. Meanwhile, Y Combinator’s early OpenAI bet is now worth over $5 billion.

💰 Funding & Business

Anthropic and OpenAI Both Launch Enterprise AI Joint Ventures — Anthropic’s venture is valued at $1.5B backed by major financial firms, while OpenAI’s targets a $10B valuation. The enterprise AI services market is heating up fast.

Y Combinator’s Stake in OpenAI Is Now Worth Over $5 Billion — YC Research seeded OpenAI in 2016 when Altman was running Y Combinator. The ~0.6% stake is one of the best early-stage bets in tech history.

Sierra Raises ~$1B at $15B Valuation — The conversational AI company crossed 200M ARR, putting it at a ~75x revenue multiple. Covered in detail on Latent Space.

Consumer AI Has an ARPU Problem — ChatGPT’s viral retention curve obscured a monetization gap — even the most engaged consumers are capped at $20/month while Anthropic’s B2B revenue grows on per-user spend expansion. Users don’t view answers as worth paying for.

Panthalassa Nets $140M for Wave-Powered AI Data Centers — Led by Peter Thiel, the Oregon startup turns waves into clean power for onsite AI computing infrastructure.

🤖 Models & Releases

GPT-5.5 Launches with a 2x Price Increase — OpenRouter’s analysis shows the actual cost increase is 49–92% (not 2x) because the model generates fewer completion tokens for longer prompts.

Meta Releases Tuna-2 Multimodal Model — Outperforms both Tuna-R and Tuna across multimodal benchmarks using pixel embeddings. Meta plans to release only a foundation checkpoint with some layers removed, not the full production weights.

Anthropic Working on ‘Orbit’ Proactive Assistant — A briefing and insights system in Claude that produces personalized briefings from connected work tools. May be unveiled at the Code with Claude conference today (May 6) in San Francisco.

Google Accelerating Gemma 4 with Multi-Token Prediction Drafters — Faster inference via multi-token prediction. 166 comments on HN.

🏢 Industry

Musk Megatrial Week 2: OpenAI Exec’s Finances Under Scrutiny — Greg Brockman took the stand. Musk had messaged him two days before trial to gauge settlement interest, then threatened he and Altman would be “the most hated men in America.” Musk’s lawyers paint Brockman as motivated by money over nonprofit mission.

US and Tech Firms Strike Deal to Vet AI Models Before Release — Microsoft, Google DeepMind, and xAI products will be reviewed for cybersecurity, biosecurity, and chemical weapons risks before public release.

Amazon Launches Supply Chain Services for Hire — Amazon is now offering its fulfillment, ocean/air shipping, and truck transportation to external companies — putting it in competition with DSV and DHL in the $1.3T third-party logistics market.

SpaceX Breaks Ground on Solar Fab for Orbital Data Centers — Building one of the world’s most advanced solar cell factories in Bastrop, Texas to vertically integrate aerospace-grade solar arrays for its orbital cloud network.

OpenAI’s Low-Latency Voice Infrastructure — A redesigned WebRTC architecture using a split relay and transceiver model to maintain real-time voice interactions at global scale. A 28-minute deep dive into the engineering.

🔬 Research & Analysis

Automating AI Research: ~60% Chance of Self-Improving Systems by 2028 — Jack Clark’s analysis shows models now handle complex engineering and scientific workflows and manage other agents. If trends hold, recursive self-improvement could arrive within two years.

Model-Harness-Fit: The Harness Is the Product Boundary — Bustamante dissects Codex CLI, Claude Code, and GitHub Copilot CLI to show frontier labs post-train models against specific harnesses. Claude Opus 4.6 scored 79.8% with ForgeCode vs 75.3% with Capy — same model, different harness.

The “Other vs The Utility” Debate — OpenAI’s Roon sparked a nuanced discussion: GPT is shaped like a tool (a “logical prosthesis”), while Claude inspires something closer to worship. People take embarrassing queries to GPT because “there is no Other so there is no Judgement.”

End-to-End Tokenizer Training for Autoregressive Images — A pipeline that jointly optimizes image tokenization and generation, enabling direct feedback from generation quality during training.

AI Didn’t Delete Your Database, You Did — A reality check on blaming AI agents for data loss. 256 comments on HN.

🛠️ Agents & Tools

Gemini API Adds Event-Driven Webhooks — The push-based notification system eliminates inefficient polling for long-running jobs. Available now for all Gemini API developers.

Vercel Launches DeepSec: Agent-Driven Security Scanning — Scans large codebases locally or in parallel cloud sandboxes to uncover complex vulnerabilities using AI agents.

Redis Gets a New Array Data Type — The PR took four months to create using AI assistance, allowing the developer to venture into complexity they would have otherwise skipped.

Formatting Stripe’s 42 Million Lines of Ruby Overnight — The story of rubyfmt, now used to format 100% of Stripe’s codebase — the world’s largest Ruby codebase.

Harness Engineering Is the New Moat — Changing prompts and middleware moved gpt-5.2-codex from 52.8% to 66.5% on Terminal-Bench 2.0. Agent performance is increasingly a joint property of model × harness × memory strategy.

🔥 Hacker News

Google Chrome Silently Installs a 4 GB AI Model on Your Device — Without consent. 757 comments and counting.

Async Rust Never Left the MVP State — A critical look at the async ecosystem’s maturity. 220 comments.

Train Your Own LLM from Scratch — A hands-on GitHub repo for building an LLM from the ground up. 49 comments.

Three Inverse Laws of AI — A playful take on Asimov’s laws for the modern era. 206 comments.

.de TLD Offline Due to DNSSEC Issues — Germany’s top-level domain went down. 117 comments.

When Everyone Has AI and the Company Still Learns Nothing — On the gap between AI adoption and organizational learning. 198 comments.

IBM Didn’t Want Microsoft to Use the Tab Key for Dialog Navigation — A delightful bit of computing history from Raymond Chen. 148 comments.

🌐 Notable

Richard Dawkins Concludes AI Is Conscious — Chats with AI bots have convinced the evolutionary biologist, though most experts say he’s being misled by mimicry.

Single Dose of Psilocybin Causes Anatomical Brain Changes — Participants took 25mg and reported deeper psychological insight and better wellbeing a month later.

Mental Model for Agentic Work — A framework that applies everywhere in agentic systems because the underlying architecture is always the same.

Copilot Token Burn: $221 of Inference for 15 Messages — Theo pushed a single Copilot message to 60M+ tokens, exposing how flat-rate pricing built for chat turns breaks under agentic workloads.

Sources: TLDR (General + AI), AINews/Latent Space, The Guardian, Hacker News — 4 newsletters processed, ~35 stories distilled