Data Engineer z GCP (f/m/x)
Summary
Data Engineer building scalable data pipelines on GCP (BigQuery, Cloud Composer, Terraform) using SQL and Python, based in Poland.
Apache Airflow / Cloud Composer Terraform Data Built Tool Apache Spark Databricks Snowflake Microsoft Fabric
Do you want to grow your skills in cloud technologies and work with real data? Join our Data & Analytics team, where we build and develop solutions based on GCP. Work with experts, develop towards Data Engineering, Big Data, or Machine Learning, and have a real impact on projects.
- Designing, implementing, and maintaining scalable data pipelines based on Google Cloud Platform
- Working with BigQuery as the main data warehouse: data modeling, query and cost optimization, ensuring performance and reliability of solutions
- Integrating data from various sources (files, databases, APIs, events) and processing and transforming it
- Orchestrating data workflows using Apache Airflow / Cloud Composer
- Creating and maintaining CI/CD solutions for data pipelines and infrastructure
- Managing cloud infrastructure following Infrastructure as Code principles (Terraform)
- Ensuring data quality, pipeline monitoring, and rapid incident response
- Collaborating with analytics, BI, and product teams to deliver stable and well-documented data
- Participating in data architecture development and jointly defining good data engineering practices
Requirements
- Min. 4 years of experience as a Data Engineer or in a similar role working with data in a production environment
- Strong knowledge of Google Cloud Platform, especially BigQuery (data modeling, query optimization) and Cloud Storage
- Ability to design, build, and maintain data pipelines (batch and/or streaming)
- Strong SQL and Python skills for data processing and orchestration
- Experience with workflow orchestration (Apache Airflow / Cloud Composer)
- Hands-on experience implementing CI/CD for data solutions, e.g., GitHub Actions, GitLab CI, Cloud Build
- Familiarity with Infrastructure as Code approach, particularly Terraform
- Previous work with large data volumes, focusing on performance and reliability of solutions
- Required to be located in Poland and fluent in Polish
Nice to have
- Practical experience with streaming data processing (e.g., Dataflow / Apache Beam, Pub/Sub)
- Proficiency in Apache Spark / PySpark working with large data volumes
- Skills in data transformation and modeling using tools like dbt
- Ability to work with diverse data platforms (e.g., Databricks, Snowflake, MS Fabric)
- Familiarity with tools and best practices in Data Governance, Data Lineage, and Data Quality
Job no. JOB-FNDWA
Sii ensures that all hiring decisions are made solely on the basis of qualifications and competence. We are committed to equal and fair treatment of all, regardless of legally protected characteristics. At Sii, we promote a diverse and inclusive work environment, in full compliance with applicable anti-discrimination laws.
Remote Hybrid Office