freehire launches on Product Hunt on 26 August.

Follow →

Principal Data Systems Software Engineer

Open 24d reposted 2× · 2 open copies

We are looking for a highly skilled software engineer to build next-generation AI data platforms and lakehouse solutions. The ideal candidate will have deep expertise in database technologies, modern data warehouses, open table formats, distributed data processing, and cloud-native software engineering.

Key Responsibiliites:

  • Design and develop scalable AI data platforms and lakehouse architectures.
  • Build high-performance data ingestion, storage, and analytics solutions.
  • Develop cloud-native services for enterprise-scale data processing.
  • Optimize database performance and query execution across heterogeneous data sources.
  • Drive architecture, technical design, and engineering best practices.

About the role:

We are looking for a highly skilled software engineer to build next-generation AI data platforms and lakehouse solutions. The ideal candidate will have deep expertise in database technologies, modern data warehouses, open table formats, distributed data processing, and cloud-native software engineering.

Key Responsibilities:

  • Design and develop scalable AI data platforms and lakehouse architectures.
  • Build high-performance data ingestion, storage, and analytics solutions.
  • Develop cloud-native services for enterprise-scale data processing.
  • Optimize database performance and query execution across heterogeneous data sources.
  • Drive architecture, technical design, and engineering best practices.
  • Required qualifications
  • 10+ years of software engineering experience building enterprise data platforms.
  • Strong expertise in Oracle and other relational/analytical databases (PostgreSQL, MySQL, SQL Server, Snowflake, Redshift, BigQuery, etc.).
  • Solid understanding of data warehouse concepts, ETL/ELT, dimensional modeling, and lakehouse architectures.
  • Hands-on experience with Apache Iceberg, Parquet, Avro, and related open table/file formats
  • Experience with distributed data processing frameworks such as Apache Spark, Trino/Presto, Kafka, or Flink.
  • Strong programming skills in Java and/or Python (Scala, Go, or C++ is a plus).
  • Good understanding of database internals, SQL optimization, indexing, partitioning, transactions, and performance tuning.
  • Experience building cloud-native applications on OCI, AWS, Azure, or GCP.

Preferred qualifications

  • Bachelor's or Master's degree in Computer Science or a related field from a reputed institution with a strong academic record.
  • Familiarity with AI/ML data pipelines, vector databases, and modern AI data architectures is a strong plus.

Career Level - IC4

See also