Senior Product Manager, Observability Tools
At ASAPP , our mission is simple: deliver the best AI-powered customer experience—faster than anyone else. To achieve that, we’re guided by principles that shape how we think, build, and execute. We value customer obsession, purposeful speed, ownership, and a relentless focus on outcomes. We work in tight, skilled teams, prioritize clarity over complexity, and continuously evolve through curiosity, data, and craftsmanship. We’re seeking technologists and problem solvers who thrive in fast-paced environments, love collaborating with great talent, and approach every day like it’s Day 1.
We're a globally diverse team with hubs in New York City, Mountain View, Latin America, and India. If you're driven by continuous learning, rapid pivots, and the challenges of building in a high-growth startup, we’d love to talk. This is more than a job—it’s a journey.
This is not a reporting role. You will own the tools that let builders and reviewers see, test, and trust how GenerativeAgent (GA) actually behaves — before it ever talks to a customer, and after. If builders can't confidently preview a configuration change, and reviewers can't efficiently audit what GA said, thought, and did in production, quality erodes silently and trust with our customers erodes with it. Making agent behavior visible, testable, and reviewable at scale is your product.
What You Will Do
-
Own scenario creation via the simulation agent. Lead product for the tools that let builders generate and curate realistic test scenarios, so configuration changes can be validated against representative conversations before they ship.
-
Own our evals surfaces, including mistake monitoring and structured data evals. Define how we detect, surface, and categorize GA mistakes at scale, and how structured data extraction is evaluated for accuracy — turning noisy conversational output into clear, actionable quality signals for builders and customers alike.
-
Own the Previewer. Own the tool builders rely on to see exactly how their configurations show up in real simulated conversations, closing the loop between making a change and understanding its downstream impact.
-
Own Conversation Explorer. Lead the primary surface reviewers use to read through full production conversations — including agent thoughts and actions — balancing depth for investigative review with speed for high-volume auditing.
-
Build the case for simulation-driven quality. Partner with Sales, Customer Success, and Research to show customers and internal stakeholders how pre-production testing and post-production review compound into fewer mistakes, faster iteration, and more defensible GA performance.
-
Partner deeply with Research and Engineering on the underlying agent architecture — reasoning traces, tool calls, scenario generation methods — so product decisions are grounded in what's technically sound and what actually predicts real-world behavior.
-
Manage the trust layer of the review experience. Proactively identify and resolve issues that erode confidence in what builders and reviewers see — desyncs between simulated and real behavior, incomplete traces, unclear mistake categorization — and design the UX and guardrails that keep them confident in what they're looking at.
Who You Are
We evaluate candidates against a set of principles. These aren't interview talking points — they describe how PMs are expected to operate here every day.
-
Growth & Abundance Mindset: You approach new domains with curiosity and learn fast. You see collaboration as additive, not competitive. When a bet doesn't work, you extract the signal and move on.
-
Extreme Ownership: You take responsibility to deliver. You don't wait to be asked — you identify the gap and close it. You are the PM; the outcome is yours.
-
High-Velocity Outcomes: You measure yourself by outcomes, not output. You distinguish between shipping a feature and moving a metric. You can answer "what changed for the customer?" for e