freehire launches on Product Hunt on 26 August.

Follow →

SR SRE Engineer/ 100% Remote in Mexico

Summary

Senior SRE Engineer leads observability and reliability frameworks, designs monitoring solutions with OpenTelemetry and APM tools, and automates IaC deployments using Terraform.

Job Profile: SR SRE Engineer

Job Type: Full-time Nomina based Job Opportunity

Location:100% Remote in Mexico


Job Description:


Site Reliability Engineer (SRE) / Lead Engineer candidate will have deep expertise in Application Performance Monitoring (APM), Infrastructure as Code (IaC), automation, and distributed tracing using Open Telemetry.

As a SRE lead, he will guide the design, implementation, and continuous improvement of observability solutions, ensuring system reliability, performance, and scalability while fostering best practices in SRE and DevOps.


Responsibilities:

· Lead the strategic development and management of observability and reliability frameworks across the organization, ensuring alignment with business goals and technical requirements.

· Design and implementation of monitoring and observability solutions, collaborating with engineering teams to define standards and best practices.

· Manage Infrastructure as Code (IaC) initiatives using Terraform, coordinating with cloud and infrastructure teams to ensure scalable and secure deployments.

· Drive automation strategies for monitoring, alerting, and logging pipelines, focusing on process improvements and operational efficiency.

· Develop and maintain comprehensive observability roadmaps, including distributed tracing, logging, and metrics collection strategies.

· Collaborate with product management, sales, and pre-sales teams to provide technical expertise and support during solution design and customer engagements.

· Lead cross-functional teams to enhance CI/CD pipelines and deployment reliability, ensuring smooth integration of observability tools and practices.

· Engage with vendors and strategic partners to evaluate, select, and integrate observability and monitoring solutions, ensuring alignment with organizational needs and fostering strong collaborative relationships.

· Mentor and develop junior engineers and analysts, fostering a culture of reliability, observability, and operational excellence.



Qualifications


· 8-10+ years of experience in SRE, Observability, or DevOps roles, with leadership responsibilities.

· Hands-on experience with OpenTelemetry for distributed tracing and observability instrumentation.

· Proven expertise with Application Performance Monitoring (APM) tools such as New Relic, Datadog, AppDynamics, or Dynatrace.

· Strong proficiency in Infrastructure as Code (IaC) using Terraform.

· Solid understanding of cloud platforms including AWS, GCP, or Azure.

· Experience with automation/configuration management tools like Ansible, Chef, or Puppet.

· Deep knowledge of CI/CD pipelines and tools such as GitHub Actions, Jenkins, or Azure DevOps.

· Experience managing Kubernetes and containerized environments (Docker, Helm).

· Familiarity with log aggregation and analysis platforms like ELK Stack or Splunk.

· Excellent leadership, communication, and collaboration skills.





See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available