Staff Product Manager - Agent & Systems Observability

Render · Remote (US)
full-time lead Posted 19 hours ago
Apply Now Stand out: build a proof-of-work pitch →

Free GitHub-based preview. Direct apply stays one click away.

Get weekly job alerts like this →

Hiring for this role?

AI Market Demand Pack · $29 one-time

Compare this role's skills with the full AI hiring market. Get ranked demand, salary bands, leading companies, public source URLs, and a decision brief.

See the live sample →

About this role

At Render, we’re building the modern cloud platform for developers creating AI-native, full-stack, multi-service applications. Our mission is to eliminate the tradeoff between the power of hyperscalers and the simplicity of developer-friendly platforms—so teams can ship fast, scale reliably, and focus on their product, not infrastructure. Unlike complex hyperscalers or ephemeral edge/serverless solutions, Render offers a developer-first experience with persistent compute, dynamic autoscaling, built-in orchestration, and observability, allowing teams to launch, scale, and manage real-world applications without writing infrastructure code or managing servers. Whether you're building LLM-powered applications, scalable SaaS products, or async processing pipelines, Render empowers teams to move fast and scale confidently from MVP to millions of users. Our platform is trusted by over 6 million developers worldwide and continues to grow rapidly. In February 2026, we raised an additional $100M in Series C financing, bringing our total funding to $257M, to accelerate our vision of making cloud infrastructure both powerful and intuitive—designed for the speed of modern AI development. We’re a diverse and talented team that values craft, velocity, and user experience. If you’re excited to help shape the future of the intelligent cloud and empower developers everywhere, we’d love to hear from you. APPLYING TO RENDER We're seeking candidates who possess high integrity, humility, and an insatiable drive to learn. Through reasoned discussions and continuous feedback, we strive to improve both individually and collectively. We foster an environment of mutual trust and respect, empowering effective debate to achieve the best outcomes for our customers and team. We especially encourage members of underrepresented groups in the tech community to apply and understand that not all successful candidates will meet each requirement listed. Our interview process is unique to each role, and we value the candidate experience just as much as our customer experience. We hope your conversations with us reflect a thoughtful process that is illuminative, enjoyable, and respectful of your time. We're looking for a Staff Product Manager to own the vision and roadmap for observability at Render. The role spans both the systems layer and the fast-emerging world of agent observability. This is a role for someone who wants to define what it means to understand your software in an era where AI is fundamentally changing how it gets built and run: where a production incident might trace through services, queues, model calls, and autonomous agent decisions, and where cost is measured in tokens as much as CPU-hours. This is a new specialization at Render, and you will own the visibility layer for AI workloads and drive the technical and strategic direction for Render's observability surface, including logs, metrics, traces, and the OpenTelemetry ecosystem. You'll engage with Render's customers and our product development teams to develop a deep understanding of how they debug, monitor, and manage spend, and build conviction about where observability needs to go next. WHAT YOU'LL DO - Own the vision, strategy, and roadmap for observability at Render. Your scope spans systems telemetry (logs, metrics, distributed tracing), OpenTelemetry-native instrumentation and export, agent and LLM observability (traces of model calls, tool use, and multi-step agent runs), and the cost and usage layer that makes token economics legible. - Drive prioritization across observability investments, balancing foundational telemetry infrastructure against near-term product needs. - Partner deeply with Engineering & other Product teams to shape the technical direction of Render's telemetry pipeline, ensuring architectural decisions optimize for signal fidelity, query performance, and cost at scale. Observability touches every product surface at Render, so the dexterity to understand the shape of our other investment areas is essential. - Develop a strong, opinionated point of view on what observability means for AI-native builders. How will OTel conventions extend to agent workloads, what a trace of an agent run could look like inside Render, and how teams should reason about token spend, model routing, and cost attribution. WE'RE LOOKING FOR - 7+ years in Product Management, with substantive experience in observability, monitoring, developer tools, or infrastructure products. - Working fluency with the OpenTelemetry ecosystem — signals, semantic conventions, collectors, and the tradeoffs of instrumenting real production systems. Bonus if you've tracked the emerging conventions for GenAI and agent observability. - Familiarity with the economics of AI workloads: token pricing, model selection tradeoffs, and how teams budget for and attribute LLM spend. Major bonus if you've worked on an LLM gateway, proxy, or metering product.

Similar Jobs

Related searches:

Remote Jobs Lead Jobs Remote Lead Jobs Lead Machine LearningLead AI InfrastructureLead NLP & Language AILead AI Agents & RAGLead Generative AI agentsgenerative-aillmcloud

Get jobs like this delivered weekly

Free AI jobs newsletter. No spam.