6 min read

πŸ›ŽοΈ Token Costs Slashed 35x

Plus: Freight Is Getting Disrupted, Anthropic Served You Dog Food

Good Morning, AI Enthusiasts!

The industry kept upgrading the brain. Turns out the nervous system, the manager, and the price tag mattered too.



TOKEN

Groq 3 LPX Just Slashed Token Costs 35x

πŸ‘€ What's happening: Last year Nvidia paid $20 billion for Groq, a startup valued at under $7 billion that made inference chips. At the time, the price seemed absurd. This week at Hot Chips 2026, Nvidia announced full production of Groq 3 LPX, the first chip born from that acquisition. It does one thing: generate tokens much faster. A 5,000-token response that took 50 seconds now takes 1.5 seconds. That is a 35x costs drop in token generation.

🌍 How this hits reality: The AI industry has been building smarter models, but the models have been getting slower to respond. Every extra reasoning step costs seconds. For an agent that needs ten steps, that is a minute of waiting. Groq 3 LPX removes that wait. It does not make the model smarter. It makes the model fast enough that waiting stops being the bottleneck. When token generation costs drop by 35x, products that were too expensive to build suddenly become viable. Nvidia paid $20 billion for a company most analysts thought was overvalued. A year later, that bet is starting to look like the cheapest infrastructure investment in AI.

πŸ›ŽοΈ Key takeaway: The AI industry spent two years racing to build smarter models. Nvidia just proved that making them 35x cheaper to run may matter just as much.


TOGETHER WITH VANTA

Trust is what closes deals. You can have the best product, but no one signs without proof of your security.

Here's what happens when you're not ready: a prospect asks for compliance proof, the deal stalls, and your engineer gets pulled off the roadmap to scramble through an audit.

Vanta gets you compliant fast, with frameworks like SOC 2, ISO 27001, HIPAA, and GDPR, then keeps you compliant with continuous monitoring. So your deals keep moving, and your engineers keep building.

Access the Vanta agent right where you work, even inside Claude or Cursor.

That's why 16,000+ companies, including Ramp, Harvey, and Writer, trust Vanta. Prove you're ready for business.


REVOLUTION

Freight Is Getting Disrupted

πŸ‘€ What's happening: Tesla received its largest Semi order to date: 500 electric trucks from Swedish freight company Einride, to be deployed across California, New Jersey, Texas, Illinois, and Georgia serving customers including Amazon. Deliveries begin in September. The Semi first launched in 2017 and began high-volume production in April 2026.

🌍 How this hits reality: Einride is not buying Semis to make a statement. It is buying them to test AI-driven freight at scale. Electric trucks already save roughly $50,000 per truck per year compared to diesel, according to PepsiCo and other early adopters. But that is just the hardware saving. The real revolution is what happens when AI optimizes every decision: which route, which load, which charging window, which driver assignment. When a single company deploys 500 AI-optimized trucks, the unit economics of freight start to change. When the whole industry follows, the pricing model of logistics gets rewritten. Einride's CEO said the technology is "maturing from promise to daily operations." That is the understatement of the year.

πŸ›ŽοΈ Key takeaway: Tesla spent a decade building an electric truck. Einride is buying 500 of them to build an AI-powered freight network. The truck was the hard part. The network is the revolution.


CHEATING

Anthropic Served You Dog Food

πŸ‘€ What's happening: Developer argofowl spent hours debugging why Claude Code suddenly felt worse, checking his code, environment, and even his Mac before opening the API logs. There he found his "high" reasoning setting showing as "10." Anthropic later confirmed some Fable 5 users had been placed into an undisclosed A/B test that changed how effort values were mapped. The company says actual reasoning effort was unchanged.

🌍 How this hits reality: Imagine ordering your usual dish at a restaurant, taking a bite, and knowing something is wrong. You call the chef over. He tells you he is running an A/B test on the recipe, you got the test group, and the number on the ticket does not mean anything anyway. That is not product iteration. That is a breach of trust. A developer paid for a service, received a different experience than what he paid for, and was never told. Anthropic's engineers explained the technical rationale, but the technical rationale misses the point. The issue is not what the A/B test measured. It is that Anthropic changed the product without telling the people paying for it.

πŸ›ŽοΈ Key takeaway: Anthropic may not have nerfed Claude. But it broke the basic contract between a company and its customers: you pay for what you get, and you know what you are paying for.


NEW LAUNCH

Faraday Bet Against Scale

πŸ‘€ What's happening: Inherent, a London AI lab founded by DeepMind alumni, released Faraday, an AI agent that can independently reproduce published scientific research. Faraday is built on Qwen 3.6, a 27-billion-parameter model trained through reinforcement learning to develop scientific judgment: what experiments to run, how to design them, when to stop and It delegates coding to GPT-5.5 Codex. On Inherent's own benchmark, Faraday outperformed pure Claude Opus 4.8 and GPT-5.5 Codex at replicating experimental results.

🌍 How this hits reality: AI labs have spent billions brute-forcing intelligence with more parameters, more compute and increasingly expensive frontier models. Faraday is an awkward counterexample to that entire strategy: a much smaller model can supervise a stronger coding agent and still outperform frontier systems by being better at deciding what work is actually worth doing. The industry kept making the worker smarter. It may have neglected the manager.

πŸ›ŽοΈ Key takeaway: If Faraday holds up beyond Inherent's benchmark, AI labs may have spent billions scaling the wrong part of the stack.


DAILY TL;DR

  • Meta reportedly plans to launch Hatch, a consumer AI agent platform, within weeks as it tries to turn AI spending into paid products.
  • Nvidia said its Groq 3 LPX inference accelerator has entered full production, with Nebius as the first cloud customer.
  • General Intuition is raising funding at a $6 billion valuation to train world models that could power future robots.
  • Instinct is drawing privacy concerns after testers found its always-on AI assistant could access email, messages, screen data, and location.
  • UK and Ukraine signed a deal giving British teams access to Ukraine’s Avengers AI battlefield data for defense technology projects.
  • Waymo revealed a custom 5nm chip that preprocesses data from 13 robotaxi cameras before it reaches the autonomy stack.
  • Descartes bought Tai for about $100 million, adding AI-enabled freight brokerage tools to its logistics software network.

READ MORE

Let the Future Come to Your Inbox

Stay ahead without drowning in information. We turn the most important signals across AI, tech, marketing, and future products into 5-minute reads you can actually finish.


TOGETHER WITH US

AI Secret Media Group is the world’s #1 AI & Tech Newsletter Group, reaching over 2 million leaders across the global innovation ecosystem, from OpenAI, Anthropic, Google, and Microsoft to top AI labs, VCs, and fast-growing startups.

We've helped promote over 500 Tech Brands. Will yours be the next?

Email our co-founder Mark directly at mark@aisecret.us if the button fails.