freehire launches on Product Hunt on 26 August.

Follow →

ProdOps Engineer 3

Open 62d posting dated 4 weeks ago

Summary

Manages and stabilizes large-scale production systems in a 24/7 environment, handling incident response, monitoring, and automation using Kubernetes, cloud platforms, and observability tools.

Black Duck Software, Inc. helps organizations build secure, high-quality software, minimizing risks while maximizing speed and productivity. Black Duck, a recognized pioneer in application security, provides SAST, SCA, and DAST solutions that enable teams to quickly find and fix vulnerabilities and defects in proprietary code, open source components, and application behavior. With a combination of industry-leading tools, services, and expertise, only Black Duck helps organizations maximize security and quality in DevSecOps and throughout the software development life cycle.

Production Operations Engineer 3 – P3 (ProdOps / SRE)

Location: Bangalore – Hybrid

Experience: 5–8 years

Shift: 24/7 Rotational shifts (Including Night Shifts & Weekend On-Call)

About the Role

The Production Operations Engineer will support and stabilize large-scale production systems with a focus on incident management, monitoring, site reliability, and customer-facing communications. This is a hands-on role requiring ownership of critical production issues in a 24/7 environment.

Key Responsibilities

  • Own and manage Critical and High production incidents end-to-end.
  • Participate in SWARM / Tech Bridge calls and lead incidents during assigned shifts.
  • Improve MTTR, MTTA, alert quality, and operational stability.
  • Perform root cause analysis (RCA) and drive corrective actions.
  • Monitor production systems and proactively detect issues.
  • Automate operational tasks using Go, Python, Shell, or Perl.
  • Maintain dashboards, alerts, runbooks, and SOPs.
  • Handle customer-facing communications during incidents.
  • Coordinate with Engineering, Product, CloudOps, and Support teams.
  • Guide junior engineers and support shift handovers.
  • Lead automation initiatives to reduce toil and manual intervention.
  • Write and review operational automation using Go / Python / Shell / Perl.
  • Act as a technical reviewer for reliabilitycritical changes.
  • Influence architecture decisions with operability and reliability in mind.
  • Own and standardize runbooks, SOPs, and disaster recovery processes.

Leadership & Mentorship

  • Provide technical leadership and mentorship to ProdOps engineers.
  • Guide shift teams during complex situations.
  • Support onboarding, training, and upskilling of team members.
  • Drive operational maturity across the team.

Tech Stack & Expertise

Required Technologies

  • Containers & Orchestration: Docker, Kubernetes, Helm
  • Cloud Platforms: AWS / GCP / Azure
  • Infrastructure as Code: Terraform
  • CI/CD: Jenkins, Harness, GitHub Actions, ArgoCD, GitLab CI
  • Monitoring & Observability: Prometheus, Grafana, ELK, Datadog, New Relic, Loki
  • Version Control: Git, GitHub, GitLab
  • Scripting: Go or Python or Shell or Perl

Qualifications

  • 6+ years of experience in Production Operations, SRE, or Cloud Reliability roles.
  • Proven experience leading major production incidents in customerfacing systems.
  • Strong background in distributed systems, Kubernetes, and cloud environments.
  • Experience mentoring engineers and driving reliability initiatives.
  • Excellent written and verbal communication skills.

What We Offer

  • An opportunity to be part of a dynamic and innovative team.
  • Inclusive and collaborative work environment.
  • Continuous learning and professional development opportunities.
  • Exposure to large-scale and customer-critical systems.

Black Duck is an equal opportunity employer. We consider all applicants for employment without regard to race, color, national origin, religion, sex, gender identity or expression, age, disability, sexual orientation, veteran or military service status, or any other characteristic protected by applicable law. Black Duck complies with all applicable laws prohibiting employment discrimination in every jurisdiction where it operates and provides reasonable accommodations to individuals with disabilities in accordance with applicable law.

What this application asks

greenhouse

First Name, Last Name, Email, Phone, Resume/CV, Cover Letter

  • Preferred First Name optional
  • LinkedIn Profile optional
  • Website optional
  • How many years into troubleshooting, monitoring?
  • Please confirm your notice period. Only candidates with an immediate to 45 days' notice period will be considered. choose any · optional
  • Are you currently located in Bangalore or willing to work in a hybrid Bangalore model? choose one
  • Are you comfortable working rotational shifts, including night shifts and weekend on-call support? choose one
  • Have you handled Critical/Severity-1 production incidents? choose one
  • How many years of hands-on experience do you have with Kubernetes? choose any
  • How many years of hands-on experience do you have with Docker? choose any
  • What is your current fixed CTC and expected CTC (in INR LPA)?

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available