Skip to Content

AI News Update: Is Nvidia’s new Groq 3 chip about to make running AI agents 35x cheaper? and more

Nvidia in Talks to Invest in Perplexity at $30B+

Nvidia is reportedly negotiating a new investment in Perplexity AI at a valuation exceeding $30B. The deal would deepen Nvidia’s AI ecosystem play beyond chip sales into the search space.

Anthropic Served You Dog Food

Developer argofowl spent hours debugging why Claude Code suddenly felt worse, checking his code, environment, and even his Mac before opening the API logs. There he found his “high” reasoning setting showing as “10.” Anthropic later confirmed some Fable 5 users had been placed into an undisclosed A/B test that changed how effort values were mapped. The company says actual reasoning effort was unchanged.

Imagine ordering your usual dish at a restaurant, taking a bite, and knowing something is wrong. You call the chef over. He tells you he is running an A/B test on the recipe, you got the test group, and the number on the ticket does not mean anything anyway. That is not product iteration. That is a breach of trust. A developer paid for a service, received a different experience than what he paid for, and was never told. Anthropic’s engineers explained the technical rationale, but the technical rationale misses the point. The issue is not what the A/B test measured. It is that Anthropic changed the product without telling the people paying for it.

Anthropic may not have nerfed Claude. But it broke the basic contract between a company and its customers: you pay for what you get, and you know what you are paying for.

OpenAI Restores 5-Hour Codex Limits

OpenAI reinstates the 5-hour Codex and Work usage limit for ChatGPT Plus users after a capacity crunch caused by rapid user growth to 20M. The move follows weeks of complaints about silently reduced quotas.

Nvidia Groq 3 LPX Enters Full Production

Nvidia announces full production of Groq 3 LPX inference chips, with $20B worth of racks to be delivered this year. The milestone accelerates Nvidia’s AI inference market expansion beyond training.

Groq 3 LPX Just Slashed Token Costs 35x

Last year Nvidia paid $20 billion for Groq, a startup valued at under $7 billion that made inference chips. At the time, the price seemed absurd. This week at Hot Chips 2026, Nvidia announced full production of Groq 3 LPX, the first chip born from that acquisition. It does one thing: generate tokens much faster. A 5,000-token response that took 50 seconds now takes 1.5 seconds. That is a 35x costs drop in token generation.

The AI industry has been building smarter models, but the models have been getting slower to respond. Every extra reasoning step costs seconds. For an agent that needs ten steps, that is a minute of waiting. Groq 3 LPX removes that wait. It does not make the model smarter. It makes the model fast enough that waiting stops being the bottleneck. When token generation costs drop by 35x, products that were too expensive to build suddenly become viable. Nvidia paid $20 billion for a company most analysts thought was overvalued. A year later, that bet is starting to look like the cheapest infrastructure investment in AI.

The AI industry spent two years racing to build smarter models. Nvidia just proved that making them 35x cheaper to run may matter just as much.

Zillow and Redfin Settle FTC Antitrust Case

Zillow and Redfin reach a settlement with the FTC over allegations of rental listing price manipulation. The resolution marks a significant moment in online real estate platform regulation.

Tesla Semi Enters Europe with Record Order

Tesla confirms the Semi electric truck is coming to Europe and has secured its largest order to date. The expansion marks a pivotal step in electrifying long-haul freight globally.

Freight Is Getting Disrupted

Tesla received its largest Semi order to date: 500 electric trucks from Swedish freight company Einride, to be deployed across California, New Jersey, Texas, Illinois, and Georgia serving customers including Amazon. Deliveries begin in September. The Semi first launched in 2017 and began high-volume production in April 2026.

Einride is not buying Semis to make a statement. It is buying them to test AI-driven freight at scale. Electric trucks already save roughly $50,000 per truck per year compared to diesel, according to PepsiCo and other early adopters. But that is just the hardware saving. The real revolution is what happens when AI optimizes every decision: which route, which load, which charging window, which driver assignment. When a single company deploys 500 AI-optimized trucks, the unit economics of freight start to change. When the whole industry follows, the pricing model of logistics gets rewritten. Einride’s CEO said the technology is “maturing from promise to daily operations.” That is the understatement of the year.

Tesla spent a decade building an electric truck. Einride is buying 500 of them to build an AI-powered freight network. The truck was the hard part. The network is the revolution.

IonQ Skyloom Deploys 84 Orbital Optical Terminals

IonQ subsidiary Skyloom reaches 84 on-orbit optical communication terminals supporting the SDA constellation. The milestone advances space-based laser communication infrastructure.

Tesla Ditches Solar Roof for Traditional Panels

Tesla discontinues its Solar Roof product line to focus on conventional solar panel installations. The strategic pivot reflects the company’s shift toward more scalable clean energy products.

Thomson Reuters Bets $40M on Building Its Own AI

Thomson Reuters invests $40M in proprietary AI infrastructure rather than renting from OpenAI. The decision reflects a growing enterprise trend toward owning AI models instead of relying on external APIs.

Quintessent Raises $40M for AI Cluster Lasers

Quintessent raises $40M to develop silicon-based optical lasers for AI data center interconnects. The technology aims to solve the bandwidth bottleneck in large-scale AI compute clusters.

Faraday Bet Against Scale

Inherent, a London AI lab founded by DeepMind alumni, released Faraday, an AI agent that can independently reproduce published scientific research. Faraday is built on Qwen 3.6, a 27-billion-parameter model trained through reinforcement learning to develop scientific judgment: what experiments to run, how to design them, when to stop and It delegates coding to GPT-5.5 Codex. On Inherent’s own benchmark, Faraday outperformed pure Claude Opus 4.8 and GPT-5.5 Codex at replicating experimental results.

AI labs have spent billions brute-forcing intelligence with more parameters, more compute and increasingly expensive frontier models. Faraday is an awkward counterexample to that entire strategy: a much smaller model can supervise a stronger coding agent and still outperform frontier systems by being better at deciding what work is actually worth doing. The industry kept making the worker smarter. It may have neglected the manager.

If Faraday holds up beyond Inherent’s benchmark, AI labs may have spent billions scaling the wrong part of the stack.

Lady Gaga Co-Founds Biotech Startup

Lady Gaga and her fiancé launch a biotech startup, marking a high-profile crossover from entertainment into drug discovery.

AWS Patches SDK Flaw Enabling Credential Theft

AWS fixes a critical SDK vulnerability that allowed attackers to exploit a region field for remote credential theft. The flaw affected a wide range of AWS services.

Grok Voice Deployed at Scale by Starlink

Starlink deploys xAI’s Grok Voice for customer support and sales at scale, marking the first major commercial validation of the voice AI system.

YMTC Eyes Record-Breaking Shanghai Star IPO

China’s flash memory giant YMTC plans a Shanghai Star Market IPO that could become the largest semiconductor listing in history, amid escalating US-China chip tensions.