Payments giant Stripe has agreed to acquire OpenRouter, a major platform for accessing and routing requests to over 400 AI models, for an estimated $7.5 billion to $8 billion. The acquisition represents Stripe's largest deal to date, positioning the company to integrate critical AI inference routing and developer infrastructure directly into its financial network.
Sources: Digg AI, HuggingNews, Latent Space, Michael Parekh AI, The Neuron, The Rundown AI, Vault Bookmarks
Anthropic is preparing for a massive initial public offering as early as October, aiming for a valuation that could rival SpaceX, while Chinese competitor DeepSeek is also planning an IPO following the success of its reasoning models. This surge in public offering plans comes as Chinese AI models, including Moonshot's Kimi K3 and Alibaba's open-weight models, gain significant global market share and download volume.
Nvidia is in advanced negotiations to invest several hundred million dollars in data center power developer Cloverleaf Infrastructure to secure a 10 GW power pipeline. This strategic move aims to mitigate power constraints limiting AI growth and protect the deployment capacity of Nvidia GPUs.
Marvell issued Google a warrant to purchase nearly 59 million shares valued at about $12.2 billion to expand their custom semiconductor partnership. The agreement incentivizes up to $120 billion in AI chip and hardware purchases through 2033, intensifying competition with Broadcom.
An AI model has successfully refuted the 80-year-old Erdős unit distance conjecture, prompting confirmation from human mathematicians and highlighting AI's growing impact on the field. In response to these rapid advancements, Fields Medalist Jacob Tsimerman is leaving academia for AI safety, while the International Mathematical Union has endorsed the Leiden Declaration warning of AI's threat to proof verifiability. Concurrently, Anthropic has restricted access to its Mythos model after discovering it could identify critical vulnerabilities in major operating systems.
Amazon Bedrock now offers OpenAI GPT-5.6 models across more than 25 AWS Regions using geographic and global cross-Region inference profiles to improve throughput and maintain consistent performance.
Nvidia is reportedly planning a specialized China-focused AI chip utilizing licensed Groq technology optimized for fast model responses, pending export approvals.
A Goldman Sachs analysis revealed that AI is already weighing on employment markets in developed economies, most notably impacting call centers, software publishing, consulting, and entry-level work.
U.S. cybersecurity agencies have warned about emerging AI-assisted cyberattacks specifically targeting internet-exposed Siemens S7 industrial control systems.
Nevada's transportation regulator has authorized Tesla to deploy 5,000 robotaxis in Clark County, replacing a previous 10-vehicle limit, while Waymo and Uber received permits for 1,000 units each. This major regulatory approval enables commercial autonomous vehicle fleets to scale rapidly across the state.
Nvidia has officially stated it has no China-specific Language Processing Unit on its roadmap and is not conducting sales of the product in the Chinese market, denying previous reports of upcoming shipments.
Z AI has launched GLM-5.3 Max, an open code model scoring 1,597 in the Code Arena: WebDev benchmark and beating Gemini 3.7 Flash-high at $3.65 per million tokens. The model is built for defensive cybersecurity and long-horizon agentic work, with weights expected to be released shortly.
Waymo has reduced its per-vehicle autonomous hardware costs from $115,000 to approximately $20,000 using a custom 5nm AI chip. The new Ojai minivan architecture aims to accelerate fleet expansion and eliminate teleoperation latency.
The Qwen3.8-27B model has received major performance boosts, including a new Blackwell-native NVFP4 GGUF quantization that delivers 50% faster prefill speeds on RTX 50 series GPUs. Simultaneously, developers have pushed consumer-grade RTX 3090 hardware to achieve up to 382 tokens per second using advanced techniques like DFlash2 block drafting and an int8 KV cache. Together, these breakthroughs represent a massive leap forward for running large language models efficiently on local hardware.
NVIDIA introduced Nsight AI, a platform providing a hosted CUDA MCP Server and open source Nsight Copilot Blueprint to integrate current CUDA knowledge into AI-enabled development workflows.
Astrophysicist Adam Becker, author of "What Is Real?", joins Tim Scarfe to take apart the futures Silicon Valley keeps selling: the 2045 singularity, mind uploading, Mars colonies, and the AI…
Max Hodak, co-founder and CEO of Science Corporation, joins Sarah Guo to discuss the future of vision, brain-computer interfaces, and the human experience. Max explains how Science’s PRIMA retinal…
Glean CEO Arvind Jain explains why model routing helps control AI costs for organizations, and how human feedback loops at scale improve its routing systems.
Rich Sutton, who helped pioneer reinforcement learning and wrote the seminal AI essay The Bitter Lesson, has now cofounded Oak Lab with his former student Khurram Javed. Their goal: to build agents…
Patrick McKenzie (patio11) hosts Aerolamp CEO Misha Gurevich and Chief Scientist Vivian Belenky, a Columbia University researcher, for a Complex Systems conversation about far-UVC germicidal light at…
Flue 2 takes its inspiration from React. Creator Fred Schott, of Astro fame, tells Latent Space why he added hooks and why agents are defined by their harnesses.