Lore

product management

This CPO regrets that product management exists | Tom Verrilli (CPO of Whatnot)

Tom Verrilli argues Whatnot deliberately built its product org around the premise "we regret that product management exists": PMs should be scarce, senior, mapped to problems rather than teams, and spend most of their time doing hands-on IC work themselves, because product judgment is a trade/muscle that atrophies when a PM makes every decision for engineers and designers — and AI now makes this leaner, more autonomous model more achievable than ever, though he stresses it's Whatnot's answer, not a universal one.

Lenny's Podcast · 2026-08-02 · English

Key ideas

  1. Whatnot's product team was founded on "we regret that product management exists" as a forcing function against reflexively hiring PMs.

  2. In the last 2 years 31,832 people applied to be a PM at Whatnot; they hired one, illustrating how rare the desired skill set is versus the size of the applicant pool.

  3. The standard tech "pod" ratio (a designer and PM for every six engineers) can infantilize capable engineers and designers by removing their need to make decisions.

  4. Product management is framed as a trade/muscle built through reps, not an innate qualification — the argument for having PMs at all rests on this.

  5. Tom wants successful PMs to move toward IC (individual contributor) work rather than being promoted into pure management, using a Messi/academy analogy to argue for keeping top performers doing the work.

  6. Whatnot maps PMs to problems and projects via a six-month planning cycle and DRI (Directly Responsible Individual) assignments, rather than permanently attaching a PM to each team.

  7. Interview signal has shifted: politics/alignment-focused answers are trending down; candidates who show both macro systems thinking and impatience to validate quickly are trending up.

  8. A historical failure mode: promoting successful PMs into director roles pulled A-players out of doing real work, creating a "yo-yo" review process; Whatnot's managers now spend 90%+ of their time on IC work, and Tom himself spends about half his time on it.

  9. Structural fix for cross-team conflict: making one PM accountable for two competing levers (e.g., discovery and ads) makes trade-offs self-align instead of requiring months of negotiation.

  10. AI is described as a major unlock: PMs can self-serve data analysis (e.g., via Hex Threads), query the codebase directly instead of asking engineers, and combine live customer footage with real-time code inspection to debug faster.

  11. This AI leverage has a downside for data scientists, who report now spending time reviewing "half-assed" AI-driven analysis from non-specialists rather than doing original work.

  12. "Product theater" (a term Tom attributes to Marty Cagan) describes PM work that looks impressive (alignment, storytelling, frameworks) without producing real value; Whatnot counters this with a mandatory hands-on case study for every hire.

  13. Working under product-minded founders requires calibrated deference: don't duplicate a founder's ownership area, use a T-shaped model (broad by default, deep on demand), and avoid the "two dads problem" of conflicting senior direction.

  14. The "play the accordion" model frames good product work as cycling between zooming out (planning/learning) and pressing in (shipping), avoiding both blind iteration and long fixed roadmaps.

  15. Retail context: e-commerce has never exceeded 20% of US retail spend in 30 years, and Tom argues agentic commerce fits high-intent, programmatic purchases while low-intent, social shopping (which live commerce serves) remains distinct and coexisting.

  16. Lessons from Twitter: strong product-market fit can survive severe organizational dysfunction, and most problems that look "complex" (like years of indecision over the 140-character limit) are really weak leadership avoiding a call.

  17. Whatnot's internal philosophy is to "bat 500" (be right about as often as wrong) and to distrust averages, since aggregate usage metrics can hide a feature that is 100% of the use case for a smaller subgroup.

  18. Tom explicitly caveats that Whatnot's fewer-but-senior-PM model is not presented as the one correct way to do product management or the one way AI will reshape the industry.

  19. Pod HR ratio — The default tech-industry staffing heuristic of adding one designer and one PM for roughly every six engineers hired. Apply: Treat this ratio as a default to resist rather than follow, only adding a PM/designer when there's a proven specific need.

  20. "We regret that product management exists" — Whatnot's founding premise for its product org, used as a forcing function so the team never hires a PM just for the sake of hiring one. Apply: Before opening a PM req, require a specific articulated need rather than assuming every team needs one.

  21. DRI (Directly Responsible Individual) — An accountability model where anyone — PM, engineer, or designer — can be named the single owner of a new product initiative. Apply: Assign a DRI per planning-cycle project regardless of function/title, and route them through the same process (e.g., product review) as anyone else.

  22. PM-to-problem mapping — Allocating PMs to core projects or problems from a planning list rather than permanently attaching one PM to a fixed engineering team. Apply: Reassign PMs across a loose set of groups (e.g., buyer/seller/trust) as priorities shift instead of guaranteeing every team a dedicated PM.

  23. Six-month planning cycle — A recurring cadence where the CEO, CPO, and senior leads jointly define "what must be true" outcomes and critical projects for the next half. Apply: Use the resulting priority list to assign DRIs and allocate scarce PM/senior attention rather than deciding staffing team-by-team.

  24. Green/red pre-mortem question — Asking "what do we do if it's green? What do we do if it's red?" before running an experiment to check whether the team has thought through both outcomes. Apply: In product review, ask this question whenever someone proposes a test; treat an unclear answer as a sign the strategy isn't thought through yet.

  25. "Know then go" — An internal Whatnot moniker for mentally working through everything that could go wrong or where scale will break before acting, then proceeding without waiting to solve every risk. Apply: Before shipping, spend time imagining failure modes and edge cases (e.g., 1000x more usage) so most real problems are pre-empted, then move forward anyway.

  26. Macro + micro interview evaluation — A PM hiring signal that looks for candidates who can both describe a big-picture end state and show impatience to validate it quickly and concretely. Apply: In interviews or case studies, probe for both systems-level description and a fast, specific validation plan, not just one or the other.

  27. Cross-functional ownership alignment — Making one PM accountable for two competing levers (e.g., discovery and ads) so trade-offs are made naturally instead of fought over between teams. Apply: When two workstreams chronically conflict over shared surface area, merge their ownership under a single accountable PM aligned to one combined goal.

  28. Ground-truth interrogation ("how do you know that?") — Repeatedly asking how a claim or label (e.g., "that was fraud") was actually determined, to expose unverified assumptions in data. Apply: When a metric or label is cited in a review, ask how it's generated/labeled until you reach the actual source of truth or an admitted gap.

  29. "Whole-ass few things" — A cultural philosophy (credited to Ron Swanson) of doing fewer initiatives but executing them thoroughly, with senior people staying in the weeds. Apply: Cut scope to a small number of priorities and push your best people into the details of those, rather than spreading effort across many initiatives.

  30. AI-as-engineer-substitute for LOE/codebase questions — Using an AI assistant (e.g., Claude/Claude Code) directly to estimate level-of-effort and understand system logic instead of asking engineers. Apply: Query the AI tool about how a feature or system works before scheduling a scoping conversation with engineering.

  31. Hex Threads — An internal Whatnot data tool that lets a PM pull nuanced cohorts, individual user logs, and build sensitivity/forecast/regression models without a data scientist. Apply: Use it to self-serve investigation of a bug report or metric change by pulling the exact user/cohort data directly.

  32. "Boxes and lines" systems literacy — The baseline expectation that a PM understands which systems drive which outcomes at a structural level. Apply: Require every PM to be able to map out how their product's systems connect, then push toward a deeper layer of understanding using AI tools.

  33. Live feedback triangulation — Combining a customer's live description/video of a problem with real-time AI inspection of the codebase to determine whether it's a real bug or a comprehension gap. Apply: While reviewing a customer support case live, simultaneously query the codebase with AI to check system behavior in parallel with the user's report.

  34. EM-light / hybrid tech-lead role — An emerging role between a full engineering-manager and a pure individual contributor, running a small team on ambitious side bets. Apply: Use this role to staff small teams attempting historically "too hard" projects outside the core roadmap.

  35. "Core focus" model — Concentrating the bulk of investment on a small number of high-conviction projects while incubating small side teams for riskier builds. Apply: Reserve full specialist team structure (PM+design+eng) for high-conviction core projects; let smaller ad hoc teams experiment around the edges.

  36. Durable product-skill checklist — A set of core PM skills — identifying what to build, distilling requirements, prioritizing for ROI, sharpening design feedback, go-to-market strategy, business strategy — framed as valuable regardless of title. Apply: Evaluate and develop these skills directly in whoever is doing the work (PM, engineer, or designer) rather than assuming only a PM should own them.

  37. "Product theater" (Marty Cagan) — PM activity — alignment meetings, storytelling, framework presentations — that resembles real product work without producing outcomes. Apply: Watch for candidates or team members who present well but whose thinking decays under a hands-on prompt; push the function back toward substantive building.

  38. Hands-on case study hiring — Whatnot's requirement that every hire, in any role, complete a prompt-plus-data exercise (e.g., write a PRD) and verbally defend it. Apply: Replace or supplement behavioral interviews with a live exercise that forces candidates to produce and defend real work product.

  39. Trust spectrum (verify-then-trust / trust-but-verify / totally trust) — A continuum describing how much autonomy leadership grants ICs before acting on their decisions. Apply: Calibrate how much independent verification a leader does based on how much a given IC or team has earned full trust.

  40. Founder deep-dive review — When a decision feels wrong, clearing the calendar to go line-by-line through tickets, code, and data with the team instead of relying on a summarized review. Apply: When a review outcome feels off, drop other commitments and work through the underlying data/tickets directly with the team until reaching ground truth.

  41. T-shaped operating model — Staying broad across the org by default but going deep with a specific team when a decision requires firsthand understanding. Apply: Default to macro oversight, but periodically embed in a team's details (e.g., how fraud-invalidation logic actually works) rather than pure delegation.

  42. "Two dads problem" — A named failure mode where two senior leaders (e.g., co-founders) give a team conflicting direction on the same workstream. Apply: If a co-leader is already actively engaged with a workstream, confirm they're accountable and fully step back rather than layering in duplicate direction.

  43. Ground-truth condition for top-down leadership — The claim that top-down decision-making only works — and avoids becoming micromanagement — when the leader is actually correct and in touch with ground truth. Apply: Before exercising top-down authority on a decision, verify you actually have current, accurate information about it; if not, it risks being micromanagement rather than effective leadership.

  44. Curiosity-first pushback protocol — When a senior stakeholder pushes an unexpected direction, first ask whether they have context you lack before disagreeing. Apply: Ask something like "is there context I don't have informing this?" to establish a shared baseline before debating the decision itself.

  45. Disagreement-resolution protocol — A three-step check for resolving a disagreement: is the other person's mind already made up; if open, is there data; if no data on either side, defer to the more senior opinion. Apply: Before pushing back further on a stakeholder's call, check if they're actually open to input, then bring data if you have it, and otherwise defer without ego if no data exists on either side.

  46. "Play the accordion" — A cyclical mental model of pulling back to plan/reflect on a problem, then pressing in to ship (V1, V2, etc.), repeating indefinitely. Apply: Before building, deliberately stretch out to define the goal; after shipping, pull back again to assess results before starting the next cycle, rather than only planning long-term or only shipping reactively.

  47. Lean experimentation loop — A hypothesis-driven cycle: form a belief, run the smallest possible test, and use the result to decide the next version of the plan. Apply: State the expected impact of a change explicitly, design the minimal test to check it, and let the outcome directly determine the next iteration.

  48. Zoom-out for knock-on effects — Tracing a proposed local fix's longer-term and second-order consequences before shipping it broadly, illustrated via Whatnot's listings/search example. Apply: Before implementing a fix (e.g., forcing all sellers to create listings), trace its downstream effects on other stakeholders (e.g., seller throughput) before committing.

  49. High-intent vs. low-intent shopping framework — A distinction between purchases suited to AI-agent automation (high-intent, specific, programmatic) versus those needing human/social discovery (low-intent, browsing-driven). Apply: Use this distinction to decide which commerce flows to automate with agents versus which to keep social/discovery-driven.

  50. CPM economics vs. commerce economics — The observation that ad-supported entertainment streaming needs a large audience (CPM logic) while commerce streaming can be economically viable with far fewer viewers. Apply: Don't apply entertainment-platform viewer thresholds to judge the viability of a commerce livestream.

  51. "It's not complex, it's weak leadership" heuristic — A diagnostic that most organizational problems presented as complex are actually leadership avoiding a hard, ownable decision. Apply: When a decision has stalled for a long time despite known data (as with Twitter's character limit), look for an unowned trade-off rather than assuming genuine technical complexity.

  52. "Batting 500" — Whatnot's internal goal of being right about as often as wrong in product decisions, used to normalize an expected failure rate. Apply: Set the expectation with a team that roughly half of bets will be wrong, so decisions are made at a reasonable pace rather than over-de-risked.

  53. "Averages mean nothing to the individual" — A caution against using aggregate usage metrics to justify deprecating a feature that's core to a smaller subgroup of users. Apply: Before deprecating a low-usage feature, check whether it's a small group's 100%-critical use case rather than relying on the overall usage percentage.

  54. Bezos anecdote-over-data rule — The principle that when data and a genuine anecdote conflict, you should trust the anecdote. Apply: Treat a credible individual account that contradicts an aggregate metric as a signal to investigate the metric's blind spots, not to dismiss the anecdote.

  55. Storming vs. norming phases (via Elizabeth Stone) — A group-development-stages framing (attributed to Elizabeth Stone) used to describe the current AI-driven upheaval in product roles as an industry-wide "storming" phase. Apply: Expect current volatility in role definitions and best practices to be a temporary storming phase rather than a permanent new steady state.

Insights

Tom frames heavy alignment process not as a necessity of scale but as a symptom of management being unable to see what's actually happening — "systems beget systems."

He floats a "spicy" compensation idea: instead of paying for a pyramid of junior PMs reporting up through layers, total that org's comp and consider paying three senior/VP-level people instead, since their impact could match the whole layer.

He cites CTOs from Workday, Instagram, Box, and super.com taking individual-contributor roles at Anthropic as evidence of a broader trend of senior people returning to hands-on work, which he wants Whatnot's product bench to emulate.

AI changes the shape of data work in both directions at once: PMs now do far more of their own data analysis and talk to data scientists less, while data scientists increasingly get stuck auditing lower-quality AI-assisted analysis instead of doing original work.

The Twitter character-limit anecdote (a known, studied, inevitable change that took years to actually happen) is used as a template for arguing that most "complexity" in org decisions is really unowned risk that no one wants to be blamed for.

Tom argues live commerce may be structurally novel because it combines internet-scale reach with the same non-transactional value physical retail provides — a curated, social browsing experience — rather than simply digitizing catalog shopping.

The "averages lie" failure pattern he describes has a network-effect multiplier: deprecating a feature used by only 3% of users can still destroy the core use case for a small group, with compounding negative fallout beyond that group alone.

He reframes his own hiring pitch to senior product leaders as an appeal to nostalgia — explicitly recruiting people who miss doing the work and are tired of alignment meetings, rather than selling career advancement.

«Every time you hire six engineers, you add a designer, you add a PM. Hiring so many PMs infantilizes the engineers and the designers who are perfectly capable of making good decisions, but just never had to because there was always a PM to babysit them.»

— 00:01

«Product team was built on the somewhat simple premise, we regret that product management exists.»

— 00:23

«We articulated that way to force ourselves to remember that you don't hire a PM just for the sake of hiring one, you hire one with this really specific need.»

— 00:29

«The only argument for why you would want product management to be a specialist function is really it's a trade, not a qualification. It's something you get good at by doing. It's a muscle.»

— 00:44

«In the last 2 years, 31,832 people applied to be a product manager at Whatnot. We hired one.»

— 00:57

«We took all of our A players and then promoted them out of doing things.»

— 01:27

«Why wouldn't you want Messi playing for your team rather than trying to have the academy coming along all the time?»

— 01:30

«Know then go.»

— 21:15

«You don't have to solve all of them, you've just got to think through all of them and then you end up solving more than you think.»

— 21:44

«It's too easy otherwise for growth to hide all sins.»

— 31:06

«Start doing IC work in the role you're in would be my push.»

— 32:53

«I just think that there's not a lot of place to hide in that anymore.»

— 48:37

«How quickly the thinking decays from folks who are good at the theater, but not the specifics, uh is kind of really telling, I think.»

— 49:15

«If nobody's got data, it's just two opinions, the CEO's opinion is going to win. And like that's okay. Check your ego at the door, get to the answer, but if you've got data, bring it.»

— 61:09

«You don't make music until you press the key and push it all the way back into V1.»

— 62:50

«e-commerce has never exceeded 20% of retail spend in America.»

— 67:44

«most of the time you hear it's really complex. It isn't. Leadership's just weak.»

— 71:22

«batting 500 is like the goal. So like you're hoping to be right as often as you're wrong.»

— 73:55

«averages mean nothing to the individual is probably the thing that I've like really scarred by.»

— 74:12

«When you have data and an anecdote, trust the anecdote.»

— 75:38

«I don't think there is one way to do product management and I don't think there is one way that AI will shape the industry.»

— 76:24

Reception

Reception is sharply split between viewers who found the episode's contrarian anti-PM-bloat take valuable and thought-provoking and a vocal contingent who found the guest arrogant, out of touch, and unlikable.

A dense, structurally organized CPO interview that names and explains many concrete internal frameworks (DRI assignment, "know then go," "play the accordion," the two-dads problem) rather than staying at the level of generic hot takes, but nearly every claim is explicitly framed by Tom as Whatnot's specific answer for a single-product, founder-led company rather than a universal prescription.

84:58

↳ Lenny's Podcast · YouTube

Watch original