Talent.com
Salve.Inno Consulting
Senior Client-Facing SRE / Cloud Engineer (AWS, Kubernetes)Salve.Inno Consulting • San Francisco, California, United States
Senior Client-Facing SRE / Cloud Engineer (AWS, Kubernetes)

Senior Client-Facing SRE / Cloud Engineer (AWS, Kubernetes)

Salve.Inno Consulting • San Francisco, California, United States
1 day ago
Job type
  • Full-time
  • Permanent
Job description

B2B Contract | EU or US - fully remote

Role Overview

We are looking for a Senior Cloud Reliability Engineer to take hands-on ownership of highly available, cloud-native production environments.

This is not a traditional DevOps role focused primarily on building CI/CD pipelines or migrating infrastructure. We are looking for an engineer who has operated critical production systems, owned incidents while on call, and can identify weaknesses in an existing cloud environment and drive meaningful improvements.

The role combines deep AWS and Kubernetes engineering with SRE practices, infrastructure automation, observability, incident management, and direct technical interaction with enterprise customers.

You will have significant autonomy to challenge existing approaches, propose better solutions, and improve the reliability, scalability, security, and operational maturity of the platform.

Key Responsibilities:

  • Own the reliability and operational health of production AWS and Kubernetes environments.

  • Participate in on-call rotations and take ownership of high-severity production incidents from detection through mitigation and resolution.

  • Lead root cause analysis and post-incident reviews, implementing permanent corrective actions rather than temporary fixes.

  • Identify architectural, reliability, security, performance, and operational weaknesses within existing cloud environments.

  • Propose and implement improvements based on AWS and Kubernetes best practices.

  • Design, maintain, and continuously improve AWS infrastructure and production Kubernetes/EKS environments.

  • Automate infrastructure provisioning and operational workflows using Terraform and configuration-management tools.

  • Improve deployment and GitOps processes using tools such as Argo CD.

  • Build and improve monitoring, logging, tracing, dashboards, and actionable alerting using Prometheus, Grafana, ELK and related observability technologies.

  • Improve scalability and workload management using Kubernetes autoscaling technologies such as Karpenter or KEDA.

  • Support distributed and event-driven environments, including technologies such as Kafka.

  • Develop automation and operational tooling using Python, Bash, Go, or similar languages.

  • Strengthen cloud security, resilience, disaster recovery, and production-readiness practices.

  • Work directly with enterprise customers when required, including technical troubleshooting, escalations, incident discussions, and explaining infrastructure or reliability issues.

  • Collaborate with engineering teams while bringing independent ideas and challenging existing technical approaches where improvements can be made.

Requirements:

  • Strong professional experience in Site Reliability Engineering, Cloud Reliability, Platform Engineering, or a comparable production-focused role.

  • Senior-level hands-on AWS expertise, with the ability to understand, design, troubleshoot, and improve existing AWS architectures.

  • Senior-level Kubernetes experience, ideally operating Amazon EKS in production.

  • Strong hands-on experience with Terraform and Infrastructure as Code.

  • Proven experience participating in an on-call rotation and personally owning production incidents.

  • Demonstrable experience with high-severity incident response, root cause analysis, postmortems, MTTR reduction, and permanent remediation.

  • Proven experience communicating directly with external or enterprise customers in a technical capacity, particularly during troubleshooting, escalations, architecture discussions, or production incidents.

  • Strong observability experience with Prometheus, Grafana, ELK, or equivalent production observability stacks.

  • Strong Linux and cloud networking fundamentals.

  • Experience automating operational processes using Python, Bash, Go, or similar scripting/programming languages.

  • Experience with CI/CD and GitOps environments.

  • Ability to independently identify infrastructure weaknesses and translate them into practical technical improvements.

  • Strong communication skills and professional-level English.

  • Comfortable explaining complex technical issues to both engineering teams and customers.

  • Hands-on production experience with Argo CD and GitOps-based deployment workflows.

  • Hands-on experience with Kafka in distributed production environments.

  • Production experience with Kubernetes autoscaling using Karpenter and/or KEDA.

  • Strong hands-on experience with Ansible for infrastructure/configuration automation.

  • Practical experience with AWS security tooling and cloud security best practices.

  • A relevant AWS certification.

What's on Offer:
  • Fully remote position.

  • Full-time B2B cooperation.

  • Opportunity to work on complex, production-critical cloud environments.

  • High level of technical ownership and autonomy.

  • Real influence over cloud architecture, reliability practices, automation, and platform improvements.

  • International engineering environment.

  • Direct collaboration with experienced technical teams and enterprise customers.

  • Long-term cooperation and opportunities to introduce new technologies and engineering practices.

Diversity and Inclusion Commitment

We are dedicated to creating and sustaining an inclusive, respectful workplace for all -regardless of gender, ethnicity, or background. We actively encourage applicants from all identities and experience levels to apply and bring your authentic self to our fast-paced, supportive team.

Create a job alert for this search

Senior Client-Facing SRE / Cloud Engineer (AWS, Kubernetes) • San Francisco, California, United States

Similar jobs

Senior Engineering Manager, Cloud Infrastructure

MeshSan Francisco, CA, United States
Full-time

Senior Engineering Manager, Cloud Infrastructure.At Mesh, our mission is to enable consumers to pay and be paid with any asset.Today, trillions of dollars in tokenized assets exist but remain large... Show more

 • Promoted

Senior Cloud Engineer -- Remote (AWS / Azure, Kubernetes)

Nuon Inc.San Francisco, CA, United States
Remote
Full-time

A leading SaaS company based in San Francisco is seeking a Senior Software Engineer, Cloud to build and maintain features for cloud infrastructure management.The role requires extensive backend dev... Show more

 • Promoted

Senior DevOps Engineer - Remote, Equity & Growth

EverOpsSan Francisco, CA, United States
Remote
Full-time

A leading technology partner is looking for a Senior DevOps Engineer to join their remote team.The position involves managing production cloud environments, migrating workloads to Amazon EKS, and e... Show more

 • Promoted

Senior Cloud Software Engineer - Remote

TwilioSan Francisco, CA, United States
Remote
Full-time

A leading communications technology company is hiring a Software Engineer (L3) to build Voice functionality through the ConversationRelay offering.This remote role requires extensive software devel... Show more

 • Promoted

Senior Cloud Engineer - Ransomware Protection (Azure)--Remote

DELTASOFT SOLUTIONS LLCSan Francisco, CA, United States
Remote
Full-time

Benefits:401(k) 401(k) matching Bonus based on performance Senior Cloud Engineer - Ransomware Protection (Azure)--Remote Role Summary We are seeking a Senior Cloud Engineer to support the implement... Show more

 • Promoted

Remote AWS DevOps Lead: Cloud-Native CI / CD & Scale

Resource InnovationsSan Francisco, CA, United States
Remote
Full-time

A women-led energy transformation firm is seeking a remote AWS DevOps Engineer to develop and manage cloud infrastructure for SaaS products.Candidates should have over 8 years of DevOps experience,... Show more

 • Promoted

Senior Software Engineer, Cloud Infrastructure

AltruistSan Francisco, CA, United States
Full-time

Senior Software Engineer, Cloud InfrastructureAltruist is transforming the multi-trillion dollar wealth management industry by building an AI platform for wealth professionals.We partner with finan... Show more

 • Promoted

Kubernetes Cloud Engineer - Remote

AkkodisSan Francisco, CA, United States
Remote
Full-time +2

Our client is currently looking for Kubernetes Cloud Engineers to join them.Start :ASAPLocation :Remote (PST- California time zone)Full timeLanguage :English speakerDuration :6 months contract, fre... Show more

 • Promoted

Senior Partner Manager, AWS

CeremonySan Francisco, CA, United States
Full-time

Hybrid - San Francisco, New York City, Austin.Vercel is the agentic infrastructure company.We free people and agents to ship what's next.For more than a decade, Vercel has shaped how the web is bui... Show more

 • Promoted

Remote SRE & Cloud Infra Engineer -- Equity

Pantera CapitalSan Francisco, CA, United States
Remote
Full-time

A tech-driven company is looking for an SRE to leverage data and automation for a highly dynamic infrastructure.This role involves scaling infrastructure, reducing toil through automation, and driv... Show more

 • Promoted

Senior Software Eng Manager Cloud API & Integrations Remote

BugcrowdSan Francisco, CA, United States
Remote
Full-time

A security technology company in San Francisco is seeking a Senior Manager, Software Engineering to lead their Integration Engineering team.You will spearhead the design and development of a cloud-... Show more

 • Promoted

Senior DevOps Engineer

CodeRabbitSan Francisco, CA, United States
Full-time

About CodeRabbitCodeRabbit is an innovative research and development company focused on building extraordinarily productive human-machine collaboration systems.Our primary goal is to create the nex... Show more

 • Promoted

Senior SRE Engineer - SF or Remote in US / Canada

Operant AISan Francisco, CA, United States
Remote
Full-time

Job DescriptionJob DescriptionSenior SRE EngineerAs our first SRE hire, we are seeking someone to build out Operant's SRE roadmap and functions that help keep our platforms and services resilient a... Show more

 • Promoted

Senior AWS DevOps Lead - Remote

Resource Innovations, Inc.San Francisco, CA, United States
Remote
Full-time

A leading energy transformation firm is seeking an AWS DevOps Engineer for a remote position with occasional in-person meetings.The role involves developing and managing AWS infrastructure, driving... Show more

 • Promoted

Remote Senior Site Reliability Engineer (SRE) - Zetachain

Blockchain WorksSan Francisco, CA, United States
Remote
Full-time

Site Reliability Engineer to join our team and run critical infrastructure for our blockchain and web applications.You'll learn to deploy and maintain a fleet of RPC and validator nodes for multipl... Show more

 • Promoted

Senior Revenue Cloud Solution Engineer - Remote

Salesforce, Inc.San Francisco, CA, United States
Remote
Full-time

A leading technology company is seeking a Revenue Cloud Solution Specialist to join its team in San Francisco.The ideal candidate will partner with sales teams to understand client business needs a... Show more

 • Promoted

Cloud Engineer

Staffing the UniverseMenlo Park, CA, United States
Full-time

Cloud EngineerLocation:Menlo Park, CAType of Hire:Contract (6 months)Required Experience:10% Architecture, 90% Implementation.Architect, design, implement, and support a cloud-based infrastructure ... Show more

 • Promoted

Client Solutions Manager - Tech, Apps & Gaming - Global Business Solutions - San Francisco

TikTokSan Francisco, CA, United States
Full-time

Client Solutions Manager - Tech, Apps & Gaming - Global Business Solutions - San Francisco.TikTok's Global Business Solutions (GBS) team is at the forefront of driving advertising innovation, offer... Show more

 • Promoted

Senior Cloud Engineer, Site Operations - OT/PCN Infrastructure

ChevronEl Sobrante, CA, United States
Full-time

Senior Cloud Engineer, Site Operations - OT/PCN InfrastructureWe are seeking a highly skilled and motivated Senior Cloud Engineer, Site Operations - OT/PCN Infrastructure to join our team supportin... Show more

 • Promoted

Cloud Advocate US

Aikido SecuritySan Francisco, CA, US
Full-time

We’re making security suck less for developers.Security tools haven’t kept up with how software is built today.They interrupt teams, slow releases, and turn security into a bottleneck instead of a ... Show more