IT Data Engineer
Summary
Data engineer with 3+ years of experience who builds and maintains ETL/ELT data pipelines in Python (Airflow, PySpark) and models data in relational warehouses like PostgreSQL, Snowflake, or BigQuery. The role also requires NLP skills, including the Hugging Face ecosystem, BERT-family models, and vector databases.
3+ years of experience in Data Engineering, Data Science, Machine Learning Engineering, or a related field
Strong proficiency in Python, with experience using FastAPI, Pydantic, Pandas, Scikit-learn, and PySpark
Strong experience with relational databases (PostgreSQL, Snowflake, or BigQuery)
Experience working with Vector Databases (Pinecone, Milvus, Weaviate, or pgvector)
Solid understanding of Data Warehousing (Star Schema, Snowflake Schema, Fact/Dimension Modeling, and modern data stack principles)
Hands-on experience with NLP, particularly the Hugging Face ecosystem, BERT-family models, and intent classification
Experience building and maintaining ETL/ELT pipelines using Python and Apache Airflow
Strong understanding of data processing, transformation, and pipeline optimization
Strong problem-solving and analytical skills with attention to data quality and accuracy
Good communication and collaboration skills, with the ability to work effectively with cross-functional teams.