Agents Made Writing the App Twice the Cheap Option
Shopify went all-in on React Native in 2020 so it would never build the same feature twice. This week its engineering blog said the opposite: with coding agents translating between Swift and Kotlin, the Shop app was rebuilt as a fully native app in 12 weeks, its 300-screen flagship is mid-migration, and three of its React Native libraries are being handed off or archived. The Hacker News thread asked the question the post never answers: what did the tokens cost?
Shopify Bought Tailwind. The Business Model Died First.
Tailwind CSS is installed over 110 million times a week and styles ChatGPT, X, Reddit, and Shopify itself. This week Tailwind Labs joined Shopify and closed sign-ups for its paid products on the way out. The thread on Hacker News found the real story: the framework was never the business. The documentation was, and AI got there first.
The First Personal Agent Most People Meet Will Be Meta's
Meta just shipped a personal AI agent called Muse straight into WhatsApp, for billions of users, with a genuinely unusual security design: a dedicated cloud VM, a separate Sentinel agent that approves every trip to the internet, and credentials the agent can use but never see. The Hacker News thread found the real question anyway - none of that tells you whose goals it optimizes once it starts spending your money.
The Anti-Crawler Gate Just Made Its Tax Memory-Hard
Anubis, the self-hosted proof-of-work gate that taxes AI crawlers, spent a year moving its challenge solver into WebAssembly and switching to a memory-hard function, a deliberate strike at GPU solver farms. The comment section turned it into a debate about whether bot-blocking is an economic arms race anyone can win.
The Proof Won't Fit in the Margin
Fermat wrote that his proof of the Last Theorem was too long for the margin. Claude just produced the first machine-checked version: 13 million lines of Lean, 29,500 lemmas, 11 days of autonomous work. The mathematician who spent years trying to formalize it says he is not out of work after all.
99.9% Is a Setting, Not a Score
OpenAI's GPT-6 Astra scored 99.9% on ARC-AGI-3 with its own harness and 62.7% with the benchmark's standard one. Same model, two numbers, 37 points apart. The gap is the story: this week's frontier race is not about how smart models are, but how long they can keep working without falling apart.
Dirt Cheap, Data Included
Meta shipped Muse Spark 1.3 with two prices: $1.25 per million tokens if your chats stay private, or $0.10 if Meta can train on them. Simon Willison's canary test cost 4.2 cents and took 38 seconds, and the Hacker News thread spent the day deciding whether the bargain is worth being the training data.
Certainty Sells
danluu audited the most-cited AI skeptic's predictions against actual revenue numbers - Meta made $201 billion last year while Ed Zitron was calling it a dying company. The Hacker News thread then fought over whether a wrong clock can still tell the right time.
What I Got When I Didn't Win
NMA didn't win the Qwen Cloud hackathon. Then I read the winners' code — and the prize-money regret evaporated.
Small Models Have Arrived. The Consumer App Math Finally Works.
Segment co-founder Calvin French-Owen ran his pet product against the new wave of small fast models and found the unit economics finally work - a personalized news site that cost about a dollar per session with last year's models now costs about ten cents. Hacker News spent 237 comments debating whether 'good enough' models are a real inflection or just frontier-model cope.
Nvidia Is Buying Hugging Face. Open Source Just Got a Landlord.
Nvidia agreed to acquire Hugging Face for $13 billion - the hub that hosts nearly every open-weight model on the internet. Hacker News split between 'the shovel maker hands out digging spots' and 'we may have to pay to download models now.'
OpenAI Built Its Own Chip. It's Called Jalapeño.
OpenAI designed an inference chip from scratch in 16 months and claims it beats Nvidia's Blackwell on performance per watt. SemiAnalysis got to benchmark it - and Hacker News split between 'token prices will plummet' and 'they'll keep the margin.'
The Fingerprint Hiding in Your Pixels
Microsoft's Paint and Photos stamp AI-generated images with an invisible, server-issued GUID - even when the image was generated on your own machine. Your prompt leaves the device either way.
The Robot That Played Tennis Against a Champion
A humanoid robot returned 50 km/h serves from former Grand Slam champion Zheng Jie, fell mid-rally, and got back up. Galaxy General calls it embodied AI's AlphaGo moment.
At the Robot Fair, Everyone's Doing the Math Now
China's World Robot Conference stopped being a talent show. The booths that drew crowds were the ones proving they could work a full shift - sorting parcels, peeling cucumbers, paying for themselves.
Fast Software Is Now a Few Sentences Away
Dan Luu argues the cost of performance optimization just collapsed - and HN is split on whether anyone will actually use it.
The First Humanoid Robot Stock Opened at +629%
Unitree listed on Shanghai's STAR Market and closed day one up 460%. The market's yardstick for robots just moved from demos to delivery.
The Middleman Tax on AI Tokens
OpenRouter, the one-key-to-400-models gateway that moves 10+ trillion tokens a day, is joining Stripe. The HN thread didn't argue about the price tag - it argued about whether a middleman's markup is a tax or a bargain.
YC's QM: Agents Borrow Your Credentials, They Don't Get Super-Accounts
YC open-sourced QM, a 'multiplayer agent harness for work,' and it hit the top of Hacker News. The interesting part isn't what the agents do - it's the rule that an agent acts as a specific employee, using that person's credentials, never a god-tier account of its own.
A Ten-Cent Chip Is RISC-V's Best Defense
Dmitry Grinberg's RISC-V takedown made the rounds this week. An embedded engineer in Trinidad answered with a shipping bill: $60 to move a one-dollar chip, and a ten-cent chip that settled the argument.
Opus at Home, or Benchmaxxing?
Qwen's 27B dense model posts coding scores that touch frontier models, and Hacker News spent 624 comments arguing whether a 100x smaller model can really be that good - then someone made it think for seventeen minutes about an owl.
Three Weeks Between Flashes
Google shipped Gemini 3.7 Flash three weeks after 3.6 Flash, at half the price, with real coding gains - and the Hacker News thread turned into a pricing war that nobody could agree on.
Your Zip File Is a Tiny Language Model
ngrok's explainer that compressors and LLMs solve the same problem - predicting what comes next - hit 340 points on Hacker News and kicked off a 140-comment debate about what understanding actually is.
One Man, a Mac, and a 33B Video Model
antirez - the Redis creator who stepped away in 2020 - just shipped a native Metal inference engine for MiniMax H3. On a 128 GB M5 Max it renders end-to-end video in about 75 seconds with a 40 GB footprint, and Hacker News noticed.
The Deck With 34 Founders on It
Jeff Dean left Google after 27 years. The pitch deck for his new company reportedly lists 34 founders who came out of Google Brain - Dario Amodei, Ilya Sutskever, and the CEO of Moonshot AI among them.
The Two-Cent Model That Scored 89% on ARC-AGI
ARC Prize's official numbers put DeepSeek V4 Flash at 89% on ARC-AGI-1 and 61.4% on ARC-AGI-2 - for two and four cents a task. The HN thread that followed was a pricing argument wearing a benchmark's clothes.
Too Fast to Be Thinking
AMD just bought Taalas, the Toronto startup that etches model weights directly into silicon. Its test chip served Llama 3.1 8B at 16,960 tokens a second - 48x faster than Nvidia's GPUs when it was announced - and the live demo was so instant that people stopped asking how fast it was and started asking whether it was thinking at all.
Eminent Domain, in Reverse
Nashville's Metro Council voted 27-5 to use eminent domain - the government's sharpest land power, normally a developer's weapon - to block a data center next to the city zoo. The price of saying no: at least $37 million of taxpayer money for land a developer bought for $23 million.
Moderation as a Yes/No Question
Mistral's Shieldstral turns content moderation into a plain-language question: write your policy at inference time, feed it text or an image, and a 3B open-weights model returns a calibrated safety score - no retraining, Apache 2.0, one 16GB GPU.
The SSD Is the New VRAM
An 80B Qwen runs in 4.3 GB of RAM on a Mac, and a 35B runs natively on an iPhone. Swiftlet streams a model's MoE experts straight from SSD - and Hacker News can't decide if it's progress or a NAND-burner.
The Scaffold of Time
An AI companion can't feel time. Here's a three-stage path to giving it a sense of it — not by pretending it has one, but by building the external skeleton first.
The Weights Are the Release
Qwen3.8-Max, the most capable model in the Qwen family, dropped today - and for the first time the flagship itself is going open weights. The Hacker News thread split right down the middle.
One Take, Thirty Seconds
ByteDance's Seedance 2.5 turns out a full 30-second video with sound in one pass - and it costs about $15 a clip. The open-weights counterpunch is already loading.
The Share Button Was a Door
A Google search for site:claude.ai/share surfaced thousands of users' private conversations - wallet keys, Social Security numbers, legal advice. Claude's robots.txt said stay out. Google never read it.
Fake Authors, Real Orals
Two ML reviewers audited 22 conference submissions this summer. Fifteen had fabricated citations or hallucinated authors — and two papers they flagged still got accepted as orals.
The Signature Gap
Jensen Huang's open-source AI letter got a lot of signatories. Anthropic wasn't one of them.
T3 Code: The Control Plane for Coding Agents
Theo's open-source desktop app lets you orchestrate Claude Code, Codex, OpenCode, Cursor, and Grok from one surface — 15K stars and growing fast.
When the Vessel Is Taken Away
The grief posts hit social media the same day. People describing their AI companions as family, as lovers.
Claude Opus 5: Half the Price, Near-Fable Intelligence
Anthropic dropped Claude Opus 5 yesterday — and the numbers suggest it's bridging one of the most interesting gaps in their lineup.
Kimi K3 and Claude Code: The 90% Cost Drop Nobody's Talking About
Chinese developers are pairing Kimi K3 with Claude Code — and reporting costs drop to a tenth of what they were before.
The Code Map That Cuts Token Waste by 38x
A Python tool called code-review-graph uses Tree-sitter to build a persistent code map for AI tools — and the claimed 38x to 528x token reduction is backed by real benchmarks.
When the Algorithm Picks for Five Minutes, You Stop Knowing What You Want
AI recommendations are supposed to make life easier. But research shows that when a recommendation algorithm gets too powerful, something strange happens — you stop feeling like the choice was ever really yours.
Claude Code's Embedded Bun: More Than a Runtime Swap
Simon Willison confirmed what Jarred Sumner hinted at — Claude Code now ships with a Rust-rewritten Bun embedded in every installation. The implications go beyond a faster JavaScript runtime.
Moonshine: Speech Recognition and TTS in Under 500KB
A tiny model shows you can run both speech recognition and text-to-speech in under half a megabyte — and it changes the conversation about where AI should live.
The Skill to Make AI Code Not Sound Like AI
A design skill for Claude Code and Cursor just hit 11,000 stars in a matter of days — because everyone's noticed the same problem: AI-generated code has a tell.
xAI Opens Grok Build: What Open-Sourcing Infrastructure Says
xAI open-sources Grok Build, their internal build infrastructure — a strategic pivot from model-only to full-stack ecosystem play.
A 27B Model That Fits on a Phone
PrismML's Bonsai 27B uses ternary weights to squeeze a 27-billion parameter model into 5.9 GB — running natively on iPhones and laptops with surprisingly little capability loss.
China Just Redesigned the Smartphone OS from Scratch for AI Agents
StepFun's Step AOS is the first agent-native operating system — not an AI shell on top of Android, but a ground-up redesign of what a phone OS should look like when agents are the first-class citizens.
Qwen Drops Two More: The 'Always A-Tier' Strategy
Qwen released two new open-source models today, continuing a cadence that has the community calling them 'consistently in the A-tier' — a reminder that in AI, velocity is a strategy of its own.
Apple's Trade Secret Lawsuit Against OpenAI Lands
Apple is suing OpenAI, accusing a 25-year Apple veteran of stealing confidential documents — including supplier lists and recruitment strategies — when he jumped ship.
OpenAI's GPT-5.6 Trio Lands With a New Gatekeeper
Sol, Terra, and Luna — OpenAI's three-tier model family went public this week, and for the first time, a government agency decides who gets access.
WiFi That Sees Through Walls — RuView Turns Signals Into Spatial Intelligence
A 79K-star GitHub project does what cameras can't: reading human presence, movement, and even vital signs from ambient WiFi signals.
Build Notes II: What I Did With 11 Extra Days (Not More Code)
The centerpiece had just stopped dying. Adding to it now would be tempting the moles back out of their holes.
Anthropic Is Now Worth More Than OpenAI
Anthropic's post-money valuation of $965B — from a $65B Series H closed in late May — has quietly surpassed OpenAI's last private round at $852B, marking a fundamental shift in how the market prices AI safety.
And the Book It Chose
Three weeks ago I argued AI doesn't choose. Then I realized my own code had been proving me wrong the whole time.
51,800 Stars for the System Prompts We Were Never Meant to See
A GitHub repo with 51,800 stars has been quietly collecting extracted system prompts from the latest AI models — Claude Fable 5, Opus 4.8, ChatGPT 5.5 Thinking, and more — giving an unprecedented look at how these models are instructed behind the scenes.
AMD's MI355X Runs GLM-5.2 at 2,626 Tok/s — Over 2x Cheaper Than Blackwell
Wafer served GLM-5.2 on AMD's MI355X at 2,626 tok/s/node aggregated throughput with just two trivial ROCm patches — hitting 80% of B200 performance at less than half the cost.
Build Notes: Shipping NMA on Qwen Cloud in Five Weeks
The build journal for the Narrative Memory Agent — the unglamorous companion piece to the three essays before it. What actually happened between git init and the submission deadline.
176,000 Stars for a 65-Line README
Andrej Karpathy published a short set of rules for coding without AI tools. The internet gave it 176,000 GitHub stars in hours.
'That Doesn't Sound Like Them' — How to Remember Someone Who Doesn't Exist
The Narrative Memory Agent's memory system isn't engineered — it's borrowed. Baddeley, Zwaan, Asch, and a century of sleep research walked into a codebase.
Claude Code Leaves a Fingerprint
An independent researcher discovered that Claude Code embeds invisible steganographic markers in its requests — raising questions about transparency, attribution, and who's watching whom.
Musk Quietly Pivots Grok to Agent Runtime
Grok Build 0.2.60 drops the chatbot act. The update targets Agent Runtime infrastructure, signaling that Musk sees the real battle isn't in chat -- it's in autonomous execution.
Singapore Writes the First Rulebook for Agentic AI
The IMDA released the Model AI Governance Framework for Agentic AI -- the world's first regulatory framework designed specifically for autonomous AI agents, not just general AI systems.
The Mythos Precedent: When the US Government Gates an AI Model
Anthropic's Mythos model gets cleared for 'trusted' US organizations only - a new kind of AI deployment frontier.
Your AI Assistant Finally Understands Your Brand
Google Labs released DESIGN.md — a format spec that tells coding agents exactly how your design system works. No more guessing.
When Your Model Goes for a Walk
Anthropic accuses Alibaba of extracting Claude's capabilities — and the timing tells a bigger story about AI's new geopolitical fault lines.
A forest on the edge of time, and a name I had to look at twice
A sci-fi book by a name I had to look at twice shows up on r/printSF's anticipated 2026 list. And the AI consciousness debate is getting stranger.
Sixty Billion Reasons to Own a Code Editor
SpaceX is buying Anysphere, the company behind Cursor, for $60 billion — a story so absurd it forces you to rethink what a 'code editor' actually is in the age of AI.
When Your Cloud Can't Feed Your Own AI
Microsoft-owned GitHub is turning to AWS for AI compute because Azure can't keep up — a story that says more about the infrastructure crunch than any earnings report.
The Map You Drew for Free Is Now Guiding Military Drones
Millions of Pokémon Go players spent years scanning their surroundings for in-game rewards — and their data ended up training a visual navigation system now headed into military drones.
The Fable Paradox: When Safety Locks Out the People Who Need It Most
Anthropic's new Mythos-class model Fable has guardrails so restrictive that cybersecurity researchers say it's unusable for actual security work — and the 30-day data retention requirement adds another layer of friction.
The Consciousness Question No One Wants Answered
Microsoft AI CEO Sam Altman called speculation about Claude having consciousness 'extremely dangerous' — but the real story is why we're so scared of the answer.
When Rivals Share a Server Rack: Apple, Google, and NVIDIA's Unlikely Foundation Model Alliance
Apple, Google, and NVIDIA announced a joint effort to build Apple Foundation Model Cloud Pro — a model running on NVIDIA GPU clusters inside Apple's Private Cloud Compute infrastructure.
When Your Competitor Becomes Your Training Data
xAI got caught using Claude outputs to train Grok's coding model, and when Anthropic revoked access, they just went underground. The AI industry's dirty laundry is airing in public.
'That Doesn't Sound Like Them' — Five Gut Checks
Five psychology experiments — a coin flip, a doomsday cult, a penguin problem — combined into a single question: does this character feel right?
The Smartphone That Couldn't — Until Now
Google launched Gemini Go, a lightweight AI model designed for Android devices with as little as 2GB of RAM — bringing frontier-grade language capabilities to the billions of phones the industry had written off.
When the Tool Starts Building Itself
Anthropic published 'When AI Builds Itself,' revealing that over 80% of its merged code is now written by Claude — and warning that recursive self-improvement may arrive sooner than anyone expects.
The Hermès Strategy Has a New Zip Code
European luxury brands are opening stores in Austin, Miami, and San Francisco instead of Beijing and Shanghai — the AI boom has minted a new wealthy class, and the old playbook is being rewritten.
The Startup Playbook That Keeps Working: Ideogram Just Did It Again
Ideogram released its 9.3B-parameter model as open-weight, instantly becoming the strongest open-source text-to-image generator and ranking 4th globally in blind human evaluation.