Tech & AI Daily

☀️ Tech & AI Daily | Wednesday, August 12, 2026

☕ Buy me a coffee

⚡ Must Know

🔐 Stealing Reasoning Traces from Proprietary LLM APIs

Researchers showed they can extract chain-of-thought reasoning from closed LLM APIs by analyzing output patterns, meaning the hidden thinking in models like o1 is not as private as vendors claim. A real concern for anyone building on these APIs or trusting their opaque reasoning.

Hacker News • Aug 11

🌐 As AI Eats the Web, the Internet's Collective Memory Is Disappearing

The Walrus makes the case that AI scraping and synthesizing the web is hollowing out original sources in a feedback loop that is actively degrading search and the open web simultaneously. This is the story that explains why your Google searches feel so broken lately.

The Walrus • Aug 10

🤖 Needle2: 14MB Agentic LLM for Phones, Wearables, and Robots

A 14MB model capable of running agentic tasks entirely on-device is a genuine threshold moment, and the benchmarks are surprisingly strong for something that fits in a fraction of a phone's RAM. Edge inference for agents just got serious.

Hacker News • Aug 10

😬 OpenAI's Head of Ethics Leaves Less Than a Year After Joining

Another ethics departure at OpenAI, another data point that safety and product velocity are not coexisting comfortably at the company. The revolving door at the top of their trust and safety work is a pattern, not a coincidence.

Financial Times • Aug 11


📡 Worth Knowing

H3-metal: Native MiniMax-H3 Inference for Apple Silicon

antirez (the Redis creator) built a native C implementation for running MiniMax-H3 directly on Apple Silicon, and early reports put the performance well ahead of Python-based alternatives. Worth a look if you are pushing local inference on Mac.

GitHub • Aug 11

🔥 Mojo 1.0 Is Here

Modular ships Mojo 1.0 after years of hype, and the language is now stable enough to evaluate seriously for AI and ML workloads. If you have been waiting for a stable target before investing time in it, that moment has arrived.

Modular • Aug 11

🍎 Apple Silicon macOS VMs: GPU Passthrough for llama.cpp

The cua team cracked GPU passthrough for macOS VMs, letting you get near-native Apple Silicon GPU speeds for local LLM inference inside an isolated VM environment. Useful for reproducible dev setups or sandboxed agent testing.

GitHub • Aug 11

👁️ London Underground Begins Live Facial Recognition Scanning

Live facial recognition is now active across London Underground stations, normalized with minimal public debate or legislative friction. This is how surveillance infrastructure quietly becomes permanent.

British Transport Police • Aug 11

🔄 Manus Returns to Operating as an Independent Company

Manus is walking back whatever arrangement had it operating under a larger entity and going independent again, with a user note that is frustratingly light on specifics. Watch this space, because the reasons behind the reversal matter.

Manus • Aug 11

📊 Nvidia's Risky Business

Stratechery digs into the structural risks in Nvidia's current position: customer concentration, hyperscalers moving to custom silicon, and whether the moat is as deep as the stock price implies. Required reading if you care about the AI infrastructure layer, which you should.

Stratechery • Aug 11


🔧 Repo/Tool of the Day

🔪 Git-knife: Edit Commit History Like a Spreadsheet

A TUI tool that lets you modify commit messages, authors, and dates across your git history in a spreadsheet interface, no arcane rebase incantations required. This is the UX that interactive rebase should have always had.

GitHub • Aug 11

📬 Get the daily digest by email

Subscribe and get Tech & AI Daily delivered to your inbox every morning.