Site Reliability Engineer Specialist
Summary
Site Reliability Engineer at Worldline in Bucharest, responsible for availability and smooth, automated deployment of payment card services (APIs, card management and virtual card engines) across on-prem and cloud. Daily work spans CI/CD automation, 2nd-line incident support with 24/7 on-call, monitoring with Grafana/Splunk/Zabbix, and Linux, PostgreSQL, Terraform and Docker on GCP and an AWS-like
The opportunity
The role main objective is to be responsible of the quality of service, an the speed and quality of service deployment on the infrastructure platforms. Ensures the availability of services within minimal risks, favoring automation and predictability on changes and deployments.
Day-to-day responsibilities
Release/Deploy:
- Have a holistic end-to end view on the service: application, underlying infrastructure and other dependencies.
- Guarantee that the service you are in charge can be installed, configured, changed, and operated I compliance with regulation, security and the expected service levels.
- Improve through automation, the pipeline from development to operations (Ci/CD) and all operations on the service.
- Participate and validate service designs and changes, with the mission to ensure that quality of operations will remain at the proper level.
Run and Operations:
- Ensure 2nd Line of support for incidents and problem management
- Identify gather and analyze metrics and events from both infrastructure and application to provide capacity planning, performance improvements and incident analysis.
- Close collaboration with monitoring and alerting specialists.
- Standby, on-call 24/7 with team (team of 4 persons) for applications in scope
Applications/products : shortlist (not exhaustive)
- Webservices/Api’s : on Local-Cloud and on Premise instances, using apigee and homemade gateways, monitoring and alerting through Grafana/Splunk/Zabbix/Pagerduty.
- PMM, Pin Management Module : on Local-Cloud and on Premise instances, monitoring and alerting through Grafana/Zabbix
- VCE, Virtual Cards Engine, on Local-Cloud instance
- TDE, Transaction Data enrichment, on local cloud instance, api’s with external party (parties)
- Generic Cardstop, on premise instance for centralized blocking of cards, with multiple datasources.
Stakeholder Management:
- Propose application related evolutions to improve service and deliver quality
- Partner with infrastructure provider as the voice of service delivery
Who are we looking for
- OS: Linux
- PostgreSQL, Sql
- Python is a nice to have
- Korn Shell Schripting, bash
- Cloud management tools: Terraform, Puppet, Gitlab
- Cloud container solutions: Docker, Nginx, HAProxy, etc..
- Batch scheduling experience (Apache Airflow or Batch on GCP) is a nice to have
- Cloud oriented: we use AWS – like internal cloud & GCP
- Monitoring: Prometheus, Grafana, ELK-a plus-nice to have, Splunk nice to have
- network management-how to debug eventual network issues--( nice to have-debugging user experience)
- CiCD-nice to have
- Experience with certificates, management, renewals.
- Fluency English nice to have French
Perks & Benefits
At Worldline you’ll get the chance to be at the heart of the global payments technology industry and shape how the world pays and gets paid.
On top of that, you will also:
• Hybrid Working Policy
• Gift vouchers on the occasion of Christmas/Easter Holidays
• Private medical services
• 21 vacation days/year
• Referral bonuses for new hires recommended by you
• WFH & Flexible Working Hours
• Full access to the “Learning” platform