freehire launches on Product Hunt on 26 August.

Follow →

Lead Data Engineer - AWS

Open 28d

Tiger Analytics is a fast-growing advanced analytics consulting firm. Our consultants bring deep expertise in Data Science, Machine Learning and AI. We are the trusted analytics partner for multiple Fortune 500 companies, enabling them to generate business value from data. Our business value and leadership has been recognized by various market research firms, including Forrester and Gartner. We are looking for top-notch talent as we continue to build the best global analytics consulting team in the world.

Tiger Analytics is seeking an experienced Senior Data Engineer to join our team, specifically focused on building scalable Generative AI architectures within the AWS ecosystem. You will architect the data foundations that power LLMs and autonomous agents for our Fortune 500 partners.

Key Responsibilities:

* GenAI Infrastructure: Architect data pipelines using Amazon Bedrock and Amazon SageMaker to build, deploy, and scale Generative AI applications.

* Vector Foundations: Implement and optimize vector search capabilities using Amazon OpenSearch Serverless or specialized vector engines for RAG (Retrieval-Augmented Generation).

* Serverless Data Engineering: Build highly scalable, event-driven ETL pipelines using AWS Lambda, AWS Glue, and Amazon Kinesis.

* Modern Data Stack: Manage large-scale data lakehouses leveraging Amazon S3, AWS Lake Formation, and Amazon Redshift.

* LLM Ops: Integrate AWS Step Functions and SageMaker Pipelines to automate the fine-tuning and deployment of foundation models.

Requirements

* Experience: 8-12 years in Data Engineering with a heavy focus on the AWS Cloud stack.

* AWS Expertise: Deep hands-on experience with Glue, Athena, EMR, and Redshift.

* AI/ML Tools: Proficiency in LangChain or LlamaIndex integrated with AWS services to handle unstructured data (text, images, PDFs).

* DevOps & IAC: Experience deploying infrastructure using AWS CDK or Terraform.

* Core Skills: Advanced SQL, Python and PySpark skills tailored for distributed processing on AWS.

Benefits

This position offers an excellent opportunity for significant career development in a fast-growing and challenging entrepreneurial environment with a high degree of individual responsibility.

Tiger Analytics provides equal employment opportunities to applicants and employees without regard to race, color, religion, age, sex, sexual orientation, gender identity/expression, pregnancy,
national origin, ancestry, marital status, protected veteran status, disability
status, or any other basis as protected by federal, state, or local law.

What this application asks

workable

First name, Last name, Email, Headline, Phone, Photo, Resume

  • What kind of US work authorization do you have (resident, visa - please specify type)? written answer
  • *Will you require Tiger Analytics to sponsor you for work authorization now or in the future? written answer
  • What is your current location? Are you willing to relocate to other locations if similar opportunity exists? written answer
  • How many years of full-time corporate work experience do you have in data engineering? written answer
  • What are your core skills in Big Data technologies? written answer
  • What is your preferred programming language? written answer
  • How many years of experience do you have working on AWS? written answer
  • Do you have experience with Databricks, If yes how many years? written answer
  • Please share a link to your LinkedIn profile (if available). written answer

See also

Tailor your CV for this role?

We couldn't check your fit for this role — add a CV to your profile to see it next time.

A new version of freehire is available