Job Title : Observability Operations Engineer
Location : Phoenix,AZ ( only Local candidate)
Client: TCS
Rate: $35/hr on W2
Positions: 2
JD:
Job Title: Observability Operations Engineer
Role Description:
"• Administer and optimize enterprise Dynatrace, Splunk, and OpenSearch/Elasticsearch platforms.
• Maintain platform availability, scalability, performance, security, and reliability.
• Build and manage monitoring, logging, tracing, dashboards, alerts, and operational metrics.
• Troubleshoot production issues and perform root cause analysis using observability tools.
• Support Linux, Kubernetes, container, and cloud-based environments.
• Automate operational activities and drive self-healing and AI-assisted operations.
• Manage upgrades, patching, capacity planning, backups, and operational governance.
• Collaborate with SRE, DevOps, Platform, Infrastructure, and Application teams."
Skills Req:
"• Strong Observability Administration experience with Dynatrace, Splunk, and OpenSearch/Elasticsearch.
• Hands-on experience with monitoring, logging, tracing, alerting, dashboards, and platform performance tuning.
• Strong knowledge of Linux, Kubernetes, and cloud environments.
• Experience with Grafana, Prometheus, OpenTelemetry, and related observability technologies.
• Automation experience using Python/Shell scripting and REST APIs.
• Experience supporting enterprise-scale production environments, troubleshooting, and RCA."
Desired Skills :
"• Strong Observability Administration experience with Dynatrace, Splunk, and OpenSearch/Elasticsearch.
• Hands-on experience with monitoring, logging, tracing, alerting, dashboards, and platform performance tuning.
• Strong knowledge of Linux, Kubernetes, and cloud environments.
• Experience with Grafana, Prometheus, OpenTelemetry, and related observability technologies.
• Automation experience using Python/Shell scripting and REST APIs.
• Experience supporting enterprise-scale production environments, troubleshooting, and RCA."
Essential skills :
"• Strong Observability Administration experience with Dynatrace, Splunk, and OpenSearch/Elasticsearch.
• Hands-on experience with monitoring, logging, tracing, alerting, dashboards, and platform performance tuning.
• Strong knowledge of Linux, Kubernetes, and cloud environments.
• Experience with Grafana, Prometheus, OpenTelemetry, and related observability technologies.
• Automation experience using Python/Shell scripting and REST APIs.
• Experience supporting enterprise-scale production environments, troubleshooting, and RCA."
Experience Required: 7 to 10 Yrs
Keyword:
Role Descriptions: Strong Observability Administration experience with Dynatrace| Splunk| andOpenSearch/Elasticsearch. Hands-on experience with monitoring| logging| tracing| alerting| dashboards| and platformperformance tuning. Strong knowledge of Linux| Kubernetes| and cloud environments. Experience with Grafana| Prometheus| OpenTelemetry| and related observability technologies. Automation experience using Python/Shell scripting and REST APIs. Experience supporting enterprise-scale production environments| troubleshooting| and RCA.Roles & Responsibilities Administer and optimize enterprise Dynatrace| Splunk| and OpenSearch/Elasticsearch platforms. Maintain platform availability| scalability| performance| security| and reliability. Build and manage monitoring| logging| tracing| dashboards| alerts| and operational metrics. Troubleshoot production issues and perform root cause analysis using observability tools. Support Linux| Kubernetes| container| and cloud-based environments. Automate operational activities and drive self-healing and AI-assisted operations. Manage upgrades| patching| capacity planning| backups| and operational governance. Collaborate with SRE| DevOps| Platform| Infrastructure| and Application teams.
Essential Skills: Strong Observability Administration experience with Dynatrace| Splunk| andOpenSearch/Elasticsearch. Hands-on experience with monitoring| logging| tracing| alerting| dashboards| and platformperformance tuning. Strong knowledge of Linux| Kubernetes| and cloud environments. Experience with Grafana| Prometheus| OpenTelemetry| and related observability technologies. Automation experience using Python/Shell scripting and REST APIs. Experience supporting enterprise-scale production environments| troubleshooting| and RCA.Roles & Responsibilities Administer and optimize enterprise Dynatrace| Splunk| and OpenSearch/Elasticsearch platforms. Maintain platform availability| scalability| performance| security| and reliability. Build and manage monitoring| logging| tracing| dashboards| alerts| and operational metrics. Troubleshoot production issues and perform root cause analysis using observability tools. Support Linux| Kubernetes| container| and cloud-based environments. Automate operational activities and drive self-healing and AI-assisted operations. Manage upgrades| patching| capacity planning| backups| and operational governance. Collaborate with SRE| DevOps| Platform| Infrastructure| and Application teams.
Desirable Skills:
Keyword:
Skills: Operations Governance.
Experience Required: