Site Reliability Engineer (SRE) / DevOps Engineer
Summary
Own and operate Kafka-based messaging platforms, apply SRE principles, and automate operations using Ansible, scripts, and CI/CD pipelines.
Job Description
\n
FulcrumDigital is an agile and next-generation digital accelerating company providingdigital transformation and technology services right from ideation toimplementation. These services have applicability across a variety ofindustries, including banking & financial services, insurance, retail,higher education, food, healthcare, and manufacturing.
Requirements\nAre you passionate about distributed systems, high-scale messaging platforms, and automation-first operations?
\nWe're looking for a Kafka Messaging / SRE Engineer to join our growing platform engineering team and help build, operate, and scale mission-critical messaging services.
What You'll Do\n- \n
- Own and operate Kafka-based messaging platforms in production environments \n
- Apply SRE principles to improve reliability, availability, and performance \n
- Drive DevOps & automation initiatives to reduce toil and manual operations \n
- Build and enhance automation using Ansible, scripts, and CI/CD pipelines \n
- Perform incident management, RCA, capacity planning, and operational readiness \n
- Collaborate closely with application and platform engineering teams \n
- Contribute to Java-based tooling and platform enhancements \n
- \n
- 3-6 years of experience working with Kafka / messaging systems \n
- Strong understanding of Kafka architecture (brokers, topics, partitions, replication) \n
- Hands-on experience with SRE / DevOps practices \n
- Java development background (ability to debug, enhance, or build platform tools) \n
- Experience with Linux, distributed systems, monitoring & alerting \n
- Exposure to incident response, production support, and operational excellence \n