ZeroNoise Logo zeronoise
Post
AI Feature Velocity Raises the PM Bar for Evidence
3 min read
325 docs
Current AI-product discussions show pressure to match competitors colliding with half-tested releases. Practical eval and discovery methods make both product failures and deliberate no-build decisions more visible.

Big Ideas

AI feature velocity is exposing a product-selection problem. A PM at a funded startup says competitor-chasing is pushing out half-tested workflows and questionable AI outputs; a commenter argues that as building gets easier, choosing what to build becomes the bottleneck. Use model capability for user understanding, experiments, and prototypes—not simply more features. That shift needs quality discipline: a prompt, model, or code change can improve one behavior while breaking another, so evaluation should be repeatable and happen before shipping.

Tactical Playbook

Define failures before metrics. The eval guide warns that teams often jump to metrics before studying actual failures, risking measurement of the wrong outcome. Its leasing-assistant example makes the point: when a prospect says the rent is out of budget, a polite goodbye sounds fine but misses the sales goal—the assistant should offer cheaper units or other properties. Start with the user-facing job, then decide what counts as failure.

  1. Capture complete traces: user input, system prompt, retrieval, tool calls, intermediate model calls, and final output.
  2. Review before automating: sample diverse traces; annotate the first 10 in actionable, user-facing terms and focus on the first upstream error. Then let an agent propose more annotations for a human to accept or reject. Continue until learning plateaus—the guide’s rule of thumb is about 100 traces.
  3. Prioritize recurring failures: cluster and count failure modes across at least 100 diverse annotated traces. In a study of 100 production traces, automated tools caught obvious trace-level errors but missed issues requiring product judgment or outside context, and sometimes flagged good responses. Use automation with human review, not instead of it.

Track what you deliberately reject. Petra Wille and Teresa Torres discuss adding trash-can markers to discovery and delivery boards to record customer problems and solutions the team chooses not to pursue. Use the record to ask whether the team is comparing solution options and whether discovery is working. An empty solution-space bin is a warning; an empty problem-space bin could indicate an innovation or culture problem, but may also reflect a strong strategy filter. Check whether people can safely raise problems, and use the record to retire “zombie” opportunities that keep resurfacing.

Case Studies & Lessons

Lenny cites company-reported outcomes associated with eval investments: Ramp’s automatic receipt-collection accuracy rose from 35% to 83%; Shopify’s AI workflow builder was 2.2× faster and 68% cheaper than the frontier-model setup it replaced; Harvey nearly doubled its contract reviewer’s internal quality score; and Cursor reported higher user satisfaction at 41% lower cost after tuning Auto Balance. These are different measures, but show evals being applied to both product quality and operating cost.

Career Corner

Eval-writing is a hiring signal: Lenny says nearly half of 25 PM openings he shared asked for experience writing evals. Candidates can demonstrate the skill by showing how they identified a user-facing failure, turned it into a repeatable test, and used recurring failures to guide product work—not just by listing AI tools.

Tools & Resources

Try the evals skill from Hamel Husain and Shreya, which Lenny linked as a way to save time and avoid mistakes; pair it with the human review process above.

AI Feature Velocity Raises the PM Bar for Evidence
Summary
Coverage start
1 day ago
Coverage end
17 hours ago
Frequency
Daily
Published
16 hours ago
Reading time
3 min
Research time
2 hrs 40 min
Documents scanned
325
Documents used
7
Citations
16
Sources monitored
99 / 100
Insights
Skipped contexts
Source details
Source Docs Insights Status
rahulvohra 0 0
Paul Graham 3 0
Tony Fadell 0 0
Patrick Collison 0 0
Daniel Ek 0 0
Gustaf Alströmer 0 0
Stewart Butterfield 0 0
PM Diego Granados 0 0
👨🏻‍💻☕️ 0 0
scott belsky 0 0
Ryan Hoover 2 0
Janna Bastow simplybastow.bsky.social 0 0
Jackie Bavaro 0 0
Sachin Rekhi 4 1
Dan Olsen 0 0
The community for ventures designed to scale rapidly | Read our rules before posting ❤️ 82 6
Will Lawrence 0 0
Product Marketing 1 0
Ami Vora 0 0
PM Interview: Practice Group for Product Manager Case Interviews 0 0
One Knight in Product 0 0
Aakash Gupta 4 1
Shreyas Doshi's Product Almanac | Substack 1 1
Lenny Rachitsky 0 0
Acquired 0 0
a16z 1 1
Exponent 0 0
Product Alliance 0 0
Product Management Exercises 0 0
rocketblocks 0 0
Product Design 11 0
ProductManagementJobs 21 4
Product Management 140 9
Product Management - The place for all things product 11 4
Product Management 10 2
Aspiring and current tech PM's 0 0
Masters of Scale 0 0
Product Science Group 0 0
How I built This 0 0
SaaStr AI 0 0
productized io 0 0
Lenny's Reads 0 0
The Product Folks 0 0
Strategyzer 0 0
Lenny's Podcast 0 0
AJ&Smart 0 0
Y Combinator 0 0
Product School 0 0
Mind the Product 0 0
@andrewchen 0 0
The Looking Glass 0 0
Kyle Poyar’s Growth Unhinged 0 0
Leah’s ProducTea 0 0
Run the Business 0 0
Product Managers at Work 0 0
The Product Compass 0 0
Ravi on Product 0 0
Productify by Bandan 0 0
Product Thinking with Melissa Perri 0 0
Product Talk Daily 0 0
The Beautiful Mess 0 0
Gibson Biddle's "Ask Gib" Product Newsletter 0 0
Casey Accidental 0 0
Hiten Shah 6 3
Product Growth 0 0
Perspectives 0 0
Lenny's Newsletter 1 1
andrew chen 0 0
Brian Balfour 0 0
Casey Winters 0 0
elena verna 0 0
Kevin Weil 🇺🇸 2 0
April Underwood 0 0
Julie Zhuo 0 0
Marty Cagan 0 0
Lenny Rachitsky 9 3
Christian Idiodi 0 0
John Cutler 0 0
Teresa Torres 1 1
Gibson Biddle 0 0
Shreyas Doshi 14 4
Adam Nash 0 0
Merci Grace 0 0
Jackie Bavaro 0 0
Hunter Walk 0 0
Brian Balfour 0 0
Scott Belsky 0 0
Nir Eyal 0 0
Teresa Torres 0 0
Julie Zhuo 0 0
Andrew Chen 0 0
John Cutler 0 0
Ken Norton 0 0
Gibson Biddle 0 0
Elena Verna 0 0
Casey Winters 0 0
Shreyas Doshi 1 1
Lenny Rachitsky 0 0
Melissa Perri 0 0
Marty Cagan 0 0