AI & Cloud Data Engineer (PySpark + GenAI)
Summary
Build and optimize cloud data pipelines with PySpark and AWS Glue, deploy ML models on SageMaker, and integrate Generative AI tools to turn raw data into actionable insights.
Transform Data into Intelligence with AWS & AI Innovation
At Axrail, an AWS Premier Partner in Malaysia, we're bridging data engineering with cutting-edge AI to solve real-world challenges. We're looking for a Cloud Data/AI Engineer who thrives on building scalable data pipelines, deploying AI/ML models, and unlocking insights with Generative AI.
Join us to shape the future of data-powered decision-making!
Your Mission
As a Cloud Data/AI Engineer at Axrail, you'll:
- Design high-performance ETL pipelines with PySpark and AWS Glue.
- Build AI/ML models (TensorFlow, PyTorch) and deploy them at scale on AWS SageMaker.
- Create interactive dashboards (QuickSight, Tableau) to turn raw data into actionable insights.
- Experiment with Generative AI to revolutionize data workflows.
Key Responsibilities
- Develop and optimize ETL pipelines for structured/unstructured data using PySpark, Glue, and Redshift.
- Ensure data integrity, security, and scalability across all pipelines.
- Architect cloud-based data solutions (S3, EMR, Athena) with IaC (Terraform, CDK).
- Train, deploy, and monitor ML models (classification, NLP, forecasting) using SageMaker.
- Automate MLOps pipelines for continuous training and deployment.
- Integrate GenAI tools (e.g., AWS Bedrock) for data synthesis, anomaly detection, and automated reporting.
Dashboard & Visualization
- Design interactive dashboards (QuickSight, Tableau) to visualize complex datasets.
- Collaborate with stakeholders to translate data into business insights.
- Fine-tune PySpark jobs, SQL queries, and SageMaker endpoints for cost/performance efficiency.
- Proactively monitor and resolve data pipeline bottlenecks.
- Deploy and manage serverless data workflows (Lambda, Step Functions).
- Stay ahead of AWS’s latest data/AI services (e.g., Q, SageMaker new features).
What You Bring
- PySpark & AWS Glue: For large-scale ETL and data processing.
- AI/ML Frameworks: TensorFlow, PyTorch, Scikit-learn.
- Dashboarding: QuickSight, Tableau, or Power BI.
Bonus
- AWS Certifications (Data Analytics, Machine Learning Specialty).
- Generative AI experience (e.g., prompt engineering, LLM fine-tuning).
- Problem-Solver: You debug data chaos and optimize pipelines like a pro.
- Visual Storyteller: You turn complex data into clear, impactful visualizations.
- Team Player: You thrive in cross-functional teams (data scientists, business analysts, engineers).
Qualifications
- Bachelor’s degree in Computer Science, Data Science, or related fields (or equivalent experience).
- Experienced candidates (3+ years): Show us your scalable pipelines, ML deployments, or cost-saving optimizations.
Why Join Axrail?
AWS Premier Partner: Access cutting-edge AWS data/AI services before competitors.
GenAI Frontier: Be among the first to apply Generative AI to real-world data challenges.
High-Impact Projects: Solve problems for enterprise clients across industries.
Upskilling Support: Earn AWS certifications and learn from data/AI experts.