Talent.com
Staff Software Engineer, Infrastructure
Staff Software Engineer, InfrastructureDecagon AI, Inc. • San Francisco, CA, United States
Staff Software Engineer, Infrastructure

Staff Software Engineer, Infrastructure

Decagon AI, Inc. • San Francisco, CA, United States
[job_card.variable_hours_ago]
[job_preview.job_type]
  • [job_card.full_time]
[job_card.job_description]

About Decagon

Decagon is the leading conversational AI platform empowering every brand to deliver concierge customer experience. Our AI agents provide intelligent, human-like responses across chat, email, and voice, resolving millions of customer inquiries across every language and at any time.

Since coming out of stealth, Decagon has experienced rapid growth. We partner with industry leaders like Hertz, Eventbrite, Duolingo, Oura, Bilt, Curology, and Samsara to redefine customer experience at scale. We've raised over $200M from Bain Capital Ventures, Accel, a16z, BOND Capital, A

  • , Elad Gil, and notable angels such as the founders of Box, Airtable, Rippling, Okta, Lattice, and Klaviyo.

We’re an in-office company, driven by a shared commitment to excellence and velocity. Our values—customers are everything, relentless momentum, winner’s mindset, and stronger together—shape how we work and grow as a team.

About the Team

The Infrastructure team builds and operates the foundations that power Decagon : networking, data, ML serving, developer platform, and real‑time voice. We partner closely with product, data, and ML to deliver high‑scale, low‑latency systems with clear SLOs and great developer ergonomics.

We organize around five focus areas :

  • Core Infra : The foundational cloud stack—networking, compute, storage, security, and infrastructure‑as‑code—to ensure reliability, scale, and cost efficiency.
  • Data Infra : Streaming / batch data platforms powering analytics / BI and customer‑facing telemetry, including for customer‑managed and on‑prem environments.
  • ML Infra : GPU and model‑serving platforms for LLM inference with multi‑provider routing and support for on‑prem / air‑gapped deployments.
  • Platform (DevEx) : CI / CD, paved paths, and core services that make shipping fast, safe, and consistent across teams.
  • Voice Infra : Telephony / WebRTC stack and observability enabling ultra‑low‑latency, high‑quality voice experiences.
  • Our mission is to deliver magical support experiences — AI agents working alongside humans to resolve issues quickly and accurately.

    About the Role

    We’re hiring a Senior Infrastructure Engineer to design, build, and operate production infrastructure for high‑scale, low‑latency systems. You’ll own critical services end‑to‑end, improve reliability and performance, and create paved‑paths that let every Decagon engineer ship confidently.

    In this role, you will

  • Design and implement critical infrastructure services with strong SLOs, clear runbooks, and actionable telemetry.
  • Partner with research and product teams to architect solutions, set up prototypes, evaluate performance, and scale new features.
  • Tune service latencies : optimize networking paths, apply smart caching / queuing, and tune CPU / memory / I / O for tight p95 / p99s.
  • Evolve CI / CD, golden paths, and self‑service tooling to improve developer velocity and safety.
  • Support various deployment architectures for customers with robust observability and upgrade paths.
  • Lead infrastructure‑as‑code (Terraform) and GitOps practices; reduce drift with reusable modules and policy‑as‑code.
  • Participate in on‑call and drive down toil through automation and elimination of recurring issues.
  • Your background looks something like this

  • 8+ years building and operating production infrastructure at scale.
  • Depth in at least one area across Core / Data / AI‑ML / Platform / Voice, with curiosity to learn the rest.
  • Proven track record meeting high availability and low latency targets (owning SLOs, p95 / p99, and load testing).
  • Excellent observability chops (OpenTelemetry, Prometheus / Grafana, Datadog) and incident response (PagerDuty, SLO / error budgets).
  • Clear written communication and the ability to turn ambiguous requirements into simple, reliable designs.
  • Even better

  • Experience being an early backend / platform / infrastructure engineer at another company
  • Strong Kubernetes experience (GKE / EKS / AKS) and experience across multiple cloud providers (GCP, AWS, and Azure)
  • Experience with customer‑managed deployments
  • Benefits

  • Medical, dental, and vision
  • Flexible time off
  • Daily lunch / dinner & snacks in the office
  • #J-18808-Ljbffr

    [job_alerts.create_a_job]

    Software Engineer Infrastructure • San Francisco, CA, United States

    [internal_linking.similar_jobs]
    Staff Software Engineer

    Staff Software Engineer

    Next Level Talent, Llc • San Francisco, California, United States
    [job_card.full_time]
    Position : Staff Software Engineer.Join a high-growth startup focused on transforming the procurement software industry, backed by top Silicon Valley VCs. This role is for an experienced backend deve...[show_more]
    [last_updated.last_updated_30] • [promoted]
    Staff Infrastructure Software Engineer, Enterprise AI

    Staff Infrastructure Software Engineer, Enterprise AI

    Scale AI • San Francisco, CA, United States
    [job_card.full_time]
    Scale GP is building the next generation of enterprise‑grade Generative AI products.Our platform provides APIs for knowledge retrieval, inference, and evaluation, enabling customers to build and de...[show_more]
    [last_updated.last_updated_variable_days] • [promoted]
    Staff Infrastructure Engineer

    Staff Infrastructure Engineer

    Crusoe • San Francisco, CA, United States
    [job_card.full_time]
    Crusoe's mission is to accelerate the abundance of energy and intelligence.We’re crafting the engine that powers a world where people can create ambitiously with AI — without sacrificing scale, spe...[show_more]
    [last_updated.last_updated_30] • [promoted]
    Staff Infrastructure Engineer

    Staff Infrastructure Engineer

    Zoox • Foster City, CA, US
    [job_card.full_time]
    Zoox is seeking a talented Staff Infrastructure Engineer to lead the development of test infrastructure that supports manufacturing tests for our autonomous vehicles. In this role, you will drive th...[show_more]
    [last_updated.last_updated_variable_days] • [promoted]
    Staff Infrastructure Software Engineer, Enterprise AI

    Staff Infrastructure Software Engineer, Enterprise AI

    Scale AI, Inc. • San Francisco, CA, United States
    [job_card.full_time]
    Scale GP is building the next generation of enterprise-grade Generative AI products.Our platform provides APIs for knowledge retrieval, inference, and evaluation, enabling customers to build and de...[show_more]
    [last_updated.last_updated_variable_days] • [promoted]
    Staff Software Engineer, Database Systems

    Staff Software Engineer, Database Systems

    Zilliz • Redwood City, CA, US
    [job_card.full_time]
    Zilliz is a fast-growing startup developing the industry’s leading vector database company for enterprise-grade AI.Founded by the engineers behind Milvus, the world’s most pop...[show_more]
    [last_updated.last_updated_30] • [promoted]
    Staff Infrastructure Engineer - Government

    Staff Infrastructure Engineer - Government

    TwelveLabs • San Francisco, CA, US
    [job_card.full_time]
    At TwelveLabs, we are pioneering the development of cutting-edge multimodal foundation models that have the ability to comprehend videos just like humans do. Our models have redefined the standards ...[show_more]
    [last_updated.last_updated_30] • [promoted]
    Staff+ Software Engineer - Infrastructure

    Staff+ Software Engineer - Infrastructure

    Anthropic • San Francisco, CA, United States
    [job_card.full_time]
    Anthropic’s mission is to create reliable, interpretable, and steerable AI systems.We want AI to be safe and beneficial for users and society. Our team includes researchers, engineers, policy expert...[show_more]
    [last_updated.last_updated_variable_days] • [promoted]
    Staff Software Engineer - AI Agent Infrastructure (Healthcare)

    Staff Software Engineer - AI Agent Infrastructure (Healthcare)

    Honey Health • San Mateo, CA, US
    [job_card.full_time]
    Honey Health is the all-in-one AI back office for primary and specialty care.Our AI agents autonomously handle core back-office jobs, such as aggregating patients data, processing orders and prescr...[show_more]
    [last_updated.last_updated_variable_days] • [promoted]
    Staff Software Engineer, Infrastructure

    Staff Software Engineer, Infrastructure

    Check Into Cash • San Francisco, CA, United States
    [job_card.full_time]
    Staff Software Engineer, Infrastructure.Building at Check : At Check, we make paying people simple.In doing that, we’re not just building our own business—we’re building payroll businesses together ...[show_more]
    [last_updated.last_updated_variable_days] • [promoted]
    Staff Software Engineer, Infrastructure

    Staff Software Engineer, Infrastructure

    Decagon • San Francisco, CA, United States
    [job_card.full_time]
    Decagon is the leading conversational AI platform empowering every brand to deliver concierge customer experience.Our AI agents provide intelligent, human‑like responses across chat, email, and voi...[show_more]
    [last_updated.last_updated_variable_days] • [promoted]
    Staff Infrastructure Software Engineer, Enterprise AI

    Staff Infrastructure Software Engineer, Enterprise AI

    Scale • San Francisco, CA, United States
    [job_card.full_time]
    Staff Infrastructure Software Engineer, Enterprise AI.Scale GP is building the next generation of enterprise‑grade Generative AI products. Our platform provides APIs for knowledge retrieval, inferen...[show_more]
    [last_updated.last_updated_variable_days] • [promoted]
    Staff Software Engineer - Forward Deployed

    Staff Software Engineer - Forward Deployed

    Invisible Technologies • San Francisco, California, United States
    [job_card.full_time]
    Invisible Technologies is the AI operating system for the enterprise.Our end-to-end AI Software Platform structures messy data, builds digital workflows, deploys agentic solutions, evaluates / measur...[show_more]
    [last_updated.last_updated_30] • [promoted]
    Staff Software Engineer - Infrastructure

    Staff Software Engineer - Infrastructure

    Nimble • San Francisco, CA, United States
    [job_card.full_time]
    Nimble is a frontier robotics and AI company building the next era of autonomous logistics.We design, manufacture, and deploy intelligent robots that enable fast, efficient, and sustainable commerc...[show_more]
    [last_updated.last_updated_30] • [promoted]
    Staff Software Engineer

    Staff Software Engineer

    Omada Health • South San Francisco, CA, United States
    [job_card.full_time]
    Omada Health is on a mission to inspire and engage people in lifelong health, one step at a time.Omada Health is a digital care provider that empowers people to achieve their health goals through s...[show_more]
    [last_updated.last_updated_30] • [promoted]
    Cloud Infrastructure Staff Engineer

    Cloud Infrastructure Staff Engineer

    PayJoy • San Francisco, CA, US
    [job_card.full_time]
    PayJoy is a mission-first credit provider dedicated to helping under-served customers in emerging markets to achieve financial stability and success. Our patented technology for secured credit provi...[show_more]
    [last_updated.last_updated_variable_days] • [promoted]
    Senior Infrastructure Software Engineer

    Senior Infrastructure Software Engineer

    2Bridge Partners • San Francisco, CA, US
    [job_card.full_time]
    Bridge is partnered with an AI-powered Medical Information Platform that's transforming how over 10,000 healthcare professionals access and apply critical clinical knowledge.Senior Infrastructu...[show_more]
    [last_updated.last_updated_30] • [promoted]
    Staff Software Engineer, GPU Infrastructure (HPC)

    Staff Software Engineer, GPU Infrastructure (HPC)

    Cohere • San Francisco, CA, United States
    [job_card.full_time]
    Staff Software Engineer, GPU Infrastructure (HPC).Staff Software Engineer, GPU Infrastructure (HPC).Our mission is to scale intelligence to serve humanity. We’re training and deploying frontier mode...[show_more]
    [last_updated.last_updated_variable_days] • [promoted]