PagerDuty Applied AI

Senior AI/ML Engineer

  • Location Lisbon
  • Seniority senior
  • Posted 2026-08-19

Original posting ↗ You apply on the company site — we never collect applications.

Extracted automatically from public job postings. Always verify details on the original posting before applying.

role details

Design and build AI-powered features, architect and own systems behind them, and partner with teams to integrate AI into existing services.

Summary generated by AI from the original posting.

Hard requirements to check first

  • Clearance:not mentioned in the posting
  • Work auth:not mentioned in the posting

Skills

RAGKubernetesAWSAzureGCPEvalsAgentsObservabilityAIMachine LearningLLMsEvent IntelligenceDistributed SystemsLow-Latency Inference

Excerpt from the original posting

PagerDuty, Inc. (NYSE: PD) is the global leader in AI-first digital operations. By automatically detecting, diagnosing, and remediating issues, the PagerDuty Platform orchestrates AI agents and automated workflows with context from over 750 integrations. Trusted by approximately two-thirds of the Fortune 100 and nearly half of the Fortune 500, PagerDuty is the industry standard for organizations scaling resilient, autonomous operations. Notable customers include Chipotle, Cloudflare, Docusign, Fox, Nvidia, Salesforce, Spotify, Zoom and more. We are growing rapidly and hiring top talent with leading AI skills across engineering, sales, product, marketing, and beyond as we build the leading digital operations platform.

 

About the role

PagerDuty’s Operations Cloud runs on a platform that ingests billions of signals and turns them into real-time action for thousands of customers. We’re looking for a Senior AI/ML Engineer who lives at the intersection of two disciplines: large-scale distributed systems and applied AI.

In this role you will design and ship AI systems that run in production at PagerDuty’s scale — powering Incident Management AI Agents, event intelligence, and the LLM-powered capabilities embedded across our platform. You’ll own the full lifecycle, from framing the problem to serving reliably at scale.

We are looking for a candidate who is genuinely passionate about building with modern AI — LLMs, agents, and retrieval — but grounded in the realities of building resilient, high-throughput systems.

What you’ll do

- Design and build AI-powered features — LLM agents, retrieval, and event intelligence — that operate on high-volume, real-time event streams, from problem framing through production deployment and monitoring.

- Architect and own the systems behind them: agent and prompt orchestration, retrieval pipelines, tool/API integrations, and low-latency inference and evaluation at scale.

- Reason about consistency, throughput, fault tolerance, and cost across services that must stay reliable under bursty, unpredictable load.

- Take AI features from prototype to production, establishing the evaluation, guardrail, observability, and improvement loops that keep them accurate and trustworthy over time.

- Partner with platform, product, and applied-research teams to define what “good” looks like and to integrate AI cleanly into existing services.

- Raise the bar through example, reviews and mentorship, and help shape the team’s technic…

→ PagerDuty · Greenhouse

Verify before you apply

  • The title, level and compensation match the original posting.
  • The role is still open — postings often close without notice, we drop them within 10 days of disappearing.
  • What "remote" means here: sometimes it is remote within one country only. The location field shows the employer’s own wording.
  • Work authorization, visa sponsorship and security clearance — these are the most common reasons an application goes nowhere.