Talent.com
tech neptune consulting
Site Reliability Engineer (Devops)tech neptune consulting • ATLANTA, Georgia, USA
Search for other jobs
Site Reliability Engineer (Devops)

Site Reliability Engineer (Devops)

tech neptune consulting • ATLANTA, Georgia, USA
19 hours ago
Salary
$55.00 hourly
Job type
  • Full-time
Job description

Job Title: Site Reliability Engineer (Devops)
Location: Atlanta, GA (Hybrid)
Job Type: Contract

IP : Virtusa
Rate : $55/hr. on W2 OR $60/hr. on C2C

Role Summary
Lead the reliability, scalability, security, and operational excellence of customer-facing platforms across Azure, GCP, and Kubernetes environments. Drive production stability through automation, observability, incident management, and continuous improvement initiatives.
Key Responsibilities
Lead platform reliability, availability, and performance initiatives.
Design and support cloud infrastructure in Azure and GCP.
Manage and optimize Kubernetes environments and containerized applications.
Implement observability and monitoring using Splunk, AppDynamics, and cloud-native tools.
Support Cloudflare, Zscaler, SQL Server, RabbitMQ, and enterprise networking components.
Lead major incident response, RCA, and problem management activities.
Develop automation and self-healing solutions to improve operational efficiency.
Collaborate with Engineering, Product, Security, and Infrastructure teams to enhance customer experience and platform stability.
Serve as a technical escalation point for critical production and customer issues.
Required Skills
5+ years of experience in SRE, DevOps, Cloud Operations, or Infrastructure Engineering.
Strong expertise in Azure, GCP, Kubernetes, Cloudflare, Splunk, AppDynamics, SQL Server, RabbitMQ, and Zscaler.
Solid networking knowledge (DNS, TCP/IP, HTTP/S, CDN, WAF, Load Balancing, SSL/TLS, Firewalls, VPNs).
Experience with automation and scripting (Python, PowerShell, Bash, Terraform).
Strong customer-facing communication and stakeholder management skills.
Core Principles
Automation First - Eliminate manual effort through automation and self-healing systems.
Observability-Driven Operations - Leverage logs, metrics, traces, and analytics to proactively identify and resolve issues.
AI-Powered Reliability - Utilize AI and operational intelligence to accelerate detection, diagnosis, and remediation.
Customer-Centric Mindset - Prioritize customer experience, stability, and business outcomes.
Operational Excellence - Continuously improve reliability, scalability, and security.
Success Measures
Service Availability & Uptime
SLA/SLO Compliance
MTTR Reduction
Incident Reduction
Automation Adoption
Customer Satisfaction (CSAT)
Platform Performance & Stability Improvements

Create a job alert for this search

Site Reliability Engineer (Devops) • ATLANTA, Georgia, USA