Home

14 Years of experience SRE Lead Engineer Austin TX Contract at Austin, Texas, USA
Email: [email protected]
http://bit.ly/4ey8w48
https://jobs.nvoids.com/job_details.jsp?id=3082984&uid=b86ab320106a44d98296ff643d4db6f1

From:

Harshit,

Tanisha systems inc

[email protected]

Reply to: [email protected]

Mandatory -

Dynatrace and Open Telemetry

Job Title: SRE Lead Engineer

Location: Austin , TX

Job type : Contract

Job Description:

We are currently seeking a highly skilled SRE hands-on Lead Engineer with solid experience in (Dynatrace and Open Telemetry) to help lead transformational initiatives within IT operations, encompassing development as well. As a crucial figure in this role, you will participate/help designing and implementing cutting-edge SRE solutions, driving the transformation of IT operations organizations to adopt an engineering-centric approach.

Responsibilities:

Participate in design, architecture of reliable, scalable, and high-performance systems and services with a focus on operational excellence, availability, and performance.

Primary skillset to be expertise in Observability as service, Telemetry data collection using Dynatrace APM, SolarWinds, Open-Source tools (Prometheus and Grafana), Log Aggregations (Kibana or Splunk) and AIOPS Tools.

Experience instrumenting OTEL Framework for .Net and Java applications.

Configure application performance monitoring (APM), infrastructure monitoring, synthetic monitoring, RUM, and log monitoring.

Integrate Dynatrace with CI/CD pipelines, alerting tools, ITSM systems, and incident automation frameworks.

Tune alert thresholds, baselines, and AI-driven anomaly detection to reduce noise and improve actionable insights.

Deeper understanding of Login authentication mechanisms using Ping, ForgeRock and SiteMinder technologies (session management and cookie management)

Correlation mechanisms and dashboards to have end to end visibility of requests from external to internal applications.

Define best practices and principles for SRE, including monitoring, alerting, and automation.

Collaborate with development teams on resiliency to ensure that services and applications are designed with operational reliability in mind.

Implement monitoring systems to assess the performance of applications and infrastructure and proactively identifying areas for optimization.

Ability to develop close relationship with other operational teams to integrate SRE practices and drive overall operational improvements across enterprise.

Stay up to date on industry trends, new technologies, and best practices in SRE and applying relevant advancements to the organization.

Ability to build strong working relationships across different levels, client focus mindset.

Qualifications:

Around 11+ years of SRE hands on experience with cloud technologies, development, SRE toolsets and automation

Own the design, configuration, deployment, and optimization of Dynatrace for enterprise-wide observability.

Extended experience instrumenting OTEL Framework.

Hands on experience with Dynatrace Plug-and-play observability modules (OKit) development for Observability Developers Java and .Net applications.

Define monitoring standards, best practices, and governance to ensure consistency and scalability.

Strong skills in APM, distributed tracing, synthetic & real user monitoring, log monitoring, and Davis AI configuration.

Experience to deploy and tune OneAgent, build end-to-end PurePath tracing, and leverage Smartscape topology for proactive performance monitoring and root-cause analysis.

Experience integrating Dynatrace with incident management, automation, and cloud platforms (AWS, Azure, GCP).

Strong problem-solving skills and ability to work in cross-functional, fast-paced environments.

Collaborate with application and infrastructure teams to troubleshoot performance issues and implement permanent fixes.

Correlation mechanisms and dashboards to have end to end visibility of requests from external to internal applications.

Strong hands-on experience with any Cloud Technology (AWS): Control Tower, Project Setup, Creating Accounts, RDS, SSO

Solid understanding and hands on experience with Docker/Kubernetes

Should have good experience with Linux Commands, GitLab CICD Setup and Terraform (state management, etc)

Monitoring & alerting setup experience with Splunk, Prometheus, Grafana, Kibana, ELK, with pref. for APM (Dynatrace).

Good understanding of Observability Framework leveraging programmatic SLI/SLO blueprints to standardize the collection of golden signals.

Good to have:

Any of the relevant professional certifications Certified Site Reliability Engineer (CSRE), Certified Kubernetes Administrator (CKA), AWS Certified DevOps Engineer Professional, , Google Cloud Professional; DevOps Engineer

Keywords: continuous integration continuous deployment artificial intelligence information technology Texas
14 Years of experience SRE Lead Engineer Austin TX Contract
[email protected]
http://bit.ly/4ey8w48
https://jobs.nvoids.com/job_details.jsp?id=3082984&uid=b86ab320106a44d98296ff643d4db6f1
[email protected]
View All
06:01 AM 27-Jan-26


To remove this job post send "job_kill 3082984" as subject from [email protected] to [email protected]. Do not write anything extra in the subject line as this is a automatic system which will not work otherwise.

Pages not loading, taking too much time to load, server timeout or unavailable, or any other issues please contact admin at [email protected]


Time Taken: 9

Location: Austin, Texas