We can't find the internet
Attempting to reconnect
Something went wrong!
Hang in there while we get back on track
🔥 TOP SIGNAL
Overment says he scrapped a four-year-old product and rebuilt its app, website, and API in five days; he describes it as a tool people relied on daily and says it had helped generate about $750,000 through campaigns. His account puts the emphasis on the system around the agents: unified logs, an app-connected test surface, persistent project context, and role-based Pi instances steered through Grok Bot.
⚡ TRY THIS
Build the harness around the app. Overment unified API/Rust/Svelte logging and connected a test environment to the app’s webview so agents could interact with it, take screenshots, and collect measurements. He kept a project vision (
vision.md), style guide, board, and feature/bug specifications available to agents; Coordinators spawn Researchers to draft specs and manage Workers. He recommends narrow tasks and short agent threads.Use formal methods on the risky state machine, not the whole repo. Boris Cherny’s follow-up describes the loop: model a race-prone or otherwise tricky part, find counterexamples, reproduce the suspected bugs, then fix them. He explicitly says this does not amount to formal verification of the entire codebase.
Give computer-use agents a concrete repo-to-PR endpoint. Theo describes Opus 5.5 finding the repository on his machine, recognizing that his local clone was stale, updating it, creating a worktree, opening a PR, verifying the fixes, and merging. It’s a reported successful run, not a general reliability guarantee.
Start a codebase audit with an explicit checklist: “I want you to do an audit around security, performance, accessibility, maintainability, scalability, architecture, documentation, testing, automation, etc.” Kent C. Dodds says Opus 5.5 found a significant security issue other models missed; treat that as one practitioner’s report, not a controlled comparison.
📡 WHAT SHIPPED
Cursor harness efficiency: Cursor reports 7% lower token costs without reduced agent quality. Its engineering write-up says it trimmed roughly 66% of its system prompt as models improved, moved most built-in tools to on-demand loading (most were used in fewer than 20% of conversations), and A/B-tested tool configurations against token use, cost, latency, errors, and overall agent use. On GPT-5.6, explicit cache breakpoints cut cold cache misses by 20%; numbering every tenth line in file reads cut cache-read tokens by 1.6% with no quality loss. Write-up.
Cursor Rollouts: Tracks a change from PR opening through production. It reads the diff and drafts an editable monitoring plan before merge, then compares post-deploy signals with the baseline; depending on configuration, it can notify the author, pause a progressive rollout, or prepare a revert PR for approval. Feature-flag ramping is still listed as coming soon. Available on Teams and Enterprise, with 10 days of trial credits. Details.
Cursor Security Reviewer: Runs on every PR, examines changes in whole-codebase context, and reports vulnerabilities with severity, an attack path, and a proposed fix. Cursor reports average review time falling from 4.8 to 3.8 minutes and comment acceptance rising from 45–50% to 60–70%.
Claude Code cloud sessions are out of research preview. Sessions run on Anthropic-hosted infrastructure, so work continues with the laptop closed; start one with
claude --cloudor from the app. GitHub must be connected. Existing subscribers get a one-time $100 Pro or $250 Max credit, claimable through Oct. 7.LangChain Managed Deep Agents can run on a schedule: add a file under
schedules/with a cron expression and prompt, then deploy; the agent runs itself. Docs.
🎬 GO DEEPER
- Video — DHH on coding with agents for Linux: DHH says agents have been writing his serious production code and that implementations improve when models critique one another. Useful as a firsthand workflow report, not an independent measure of code quality.
- Repo — Overment’s
limen: A custom Pi extension that injects the project vision and style guide and steers Coordinator, Worker, Reviewer, and Researcher instances.
Editorial take: The repeatable pattern is a closed feedback loop: give agents persistent context and a testable product surface, verify changes at the PR boundary, then check production signals after deploy.
| Source | Docs | Insights | Status |
|---|---|---|---|
| Brent Traut | 0 | 0 | |
| Lukas Möller | 0 | 0 | |
| Jediah Katz | 4 | 2 | |
| Aman Karmani | 0 | 0 | |
| Jacob Jackson | 0 | 0 | |
| Cursor Blog | RSS Feed | 0 | 0 | |
| Nicholas Moy | 0 | 0 | |
| Mike Krieger | 0 | 0 | |
| Sualeh Asif | 0 | 0 | |
| Michael Truell | 0 | 0 | |
| Google Antigravity | 2 | 1 | |
| Aman Sanger | 0 | 0 | |
| cat | 0 | 0 | |
| Mark Chen | 0 | 0 | |
| Greg Brockman | 2 | 0 | |
| Tongzhou Wang | 0 | 0 | |
| fouad | 0 | 0 | |
| Calvin French-Owen | 0 | 0 | |
| Hanson Wang | 0 | 0 | |
| Ed Bayes | 0 | 0 | |
| Alexander Embiricos | 0 | 0 | |
| Tibo | 3 | 1 | |
| Romain Huet | 2 | 0 | |
| DHH | 17 | 1 | |
| Jane Street Blog | 0 | 0 | |
| Miguel Grinberg's Blog: AI | 0 | 0 | |
| xxchan's Blog | 0 | 0 | |
| <antirez> | 0 | 0 | |
| Brendan Long | 0 | 0 | |
| The Pragmatic Engineer | 0 | 0 | |
| David Heinemeier Hansson | 0 | 0 | |
| Armin Ronacher ⇌ | 3 | 0 | |
| Mitchell Hashimoto | 0 | 0 | |
| Armin Ronacher's Thoughts and Writings | 0 | 0 | |
| Peter Steinberger | 0 | 0 | |
| Theo - t3.gg | 31 | 5 | |
| Sourcegraph | 0 | 0 | |
| Anthropic | 3 | 0 | |
| Cursor | 0 | 0 | |
| LangChain | 1 | 1 | |
| Anthropic | 0 | 0 | |
| LangChain Blog | 0 | 0 | |
| LangChain | 8 | 2 | |
| Cursor | 5 | 2 | |
| Riley Brown | 0 | 0 | |
| Riley Brown | 11 | 1 | |
| Jason Zhou | 4 | 1 | |
| Boris Cherny | 5 | 2 | |
| Mckay Wrigley | 0 | 0 | |
| geoff | 5 | 2 | |
| Peter Steinberger 🦞 | 14 | 3 | |
| AI Jason | 1 | 1 | |
| Alex Albert | 0 | 0 | |
| Latent.Space | 2 | 1 | |
| Logan Kilpatrick | 3 | 0 | |
| Fireship | 0 | 0 | |
| Fireship | 1 | 1 | |
| Kent C. Dodds 🐨 | 25 | 9 | |
| Practical AI | 0 | 0 | |
| Practical AI Clips | 0 | 0 | |
| Stories by Steve Yegge on Medium | 0 | 0 | |
| Kent C. Dodds Blog | 0 | 0 | |
| ThePrimeTime | 1 | 1 | |
| Theo - t3․gg | 1 | 1 | |
| ThePrimeagen | 16 | 2 | |
| Ben Tossell | 0 | 0 | |
| swyx | 7 | 0 | |
| AI For Developers | 0 | 0 | |
| Geoffrey Huntley | 0 | 0 | |
| Addy Osmani | 5 | 2 | |
| Andrej Karpathy | 0 | 0 | |
| Simon Willison | 9 | 3 | |
| Matthew Berman | 1 | 1 | |
| Changelog | 0 | 0 | |
| Simon Willison’s Newsletter | 0 | 0 | |
| Agentic Coding Newsletter | 0 | 0 | |
| Latent Space | 1 | 0 | |
| Simon Willison's Weblog | 2 | 2 | |
| Elevate | 0 | 0 | |
| Lukas Möller | 0 | 0 | |
| Jediah Katz | 0 | 0 | |
| Sualeh Asif | 0 | 0 | |
| Mike Krieger | 0 | 0 | |
| Michael Truell | 0 | 0 | |
| Cat Wu | 0 | 0 | |
| Kevin Hou | 0 | 0 | |
| Aman Sanger | 0 | 0 | |
| Nicholas Moy | 0 | 0 | |
| Andrey Mishchenko | 0 | 0 | |
| Jerry Tworek | 1 | 1 | |
| Romain Huet | 0 | 0 | |
| Thibault Sottiaux | 0 | 0 | |
| Alexander Embiricos | 0 | 0 | |
| xxchan | 0 | 0 | |
| Salvatore Sanfilippo | 0 | 0 | |
| Armin Ronacher | 0 | 0 | |
| David Heinemeier Hansson (DHH) | 1 | 1 | |
| Alex Albert | 0 | 0 | |
| Logan Kilpatrick | 0 | 0 | |
| Shawn "swyx" Wang | 0 | 0 | |
| Jason Zhou | 0 | 0 | |
| Riley Brown | 0 | 0 | |
| McKay Wrigley | 0 | 0 | |
| Boris Cherny | 0 | 0 | |
| Ben Tossell | 0 | 0 | |
| Geoffrey Huntley | 0 | 0 | |
| Peter Steinberger | 0 | 0 | |
| Addy Osmani | 0 | 0 | |
| Simon Willison | 0 | 0 | |
| Andrej Karpathy | 0 | 0 | |
| Harrison Chase | 0 | 0 |