Mô tả công việc
Tóm tắt công việc
We are seeking a Lead Data Engineer (Python, Kubernetes, Snowflake) to own reliable data pipelines and production platforms across Cloud and on-prem. You will deploy and run data and ML workloads on OpenShift, improve CI/CD, troubleshoot complex issues and guide engineering teams as we scale an enterprise analytics and ML platform.
Responsibilities:
Design, build and maintain scalable data pipelines and platform components using Snowflake, Apache Spark (PySpark), SQL and Apache Airflow
Deploy, operate and support data and ML workloads on Kubernetes and OpenShift in production
Develop and maintain Python services and APIs using FastAPI or similar frameworks
Monitor, troubleshoot and optimize performance, reliability and data quality across pipelines and platforms
Build and improve continuous integration and continuous delivery (CI/CD) pipelines for safe, repeatable releases
Partner with DevOps, Platform, ML, infrastructure and application teams to deliver production-ready solutions
Drive root cause analysis for complex incidents and define preventative fixes and operational best practices
Provide technical leadership through architecture decisions, reviews and mentoring
By choosing EPAM Vietnam, you're getting a job at a Vietnam Best WorkplaceTM certified company (2022-23 and 2023-24)
You'll work with the latest and most advanced technologies on exciting global projects with a supportive multicultural team where your voice matters
We prioritize a healthy work-life balance with a flexible hybrid working model and generous annual leave of up to 19 days.
We offer a transparent career path and an individual roadmap to engineer your future and accelerate your journey
Our commitment is to the well-being of our employees and their families by offering premium healthcare insurance for employees and 2 dependents, annual health checkup and 10 days of paid sick leave
At EPAM, you can find vast opportunities for self-development: access to over 25,000 online courses and libraries from industry leaders, English classes, mentoring programs, partial grants of certification, and experience exchange with colleagues around the world. You will learn, contribute, and grow with us.
Yêu cầu
Requirements:
7+ year of experience in data engineering and enterprise-scale data platform delivery
Hands-on expertise with Snowflake, Apache Spark (PySpark), SQL and Apache Airflow
Strong Python development capability, including building services or APIs
Production experience operating workloads on Kubernetes or Red Hat OpenShift
Solid background in continuous integration and continuous delivery (CI/CD), Git workflows, containers and deployment automation
Experience supporting solutions across Microsoft Azure and hybrid cloud environments
Strength in monitoring, troubleshooting and production support for distributed systems
Track record of technical ownership, clear communication and mentoring within engineering teams
Nice to have:
Experience with FastAPI, Databricks or Splunk
Exposure to MLOps practices, ML platform operations or LLM and GenAI enablement platforms
Infrastructure as Code experience (tooling aligned to your environment)
Thông tin khác
Python
PySpark
Apache Spark
Python
Git
Redhat
API
MS SQL
Distributed Systems
Splunk
MS Azure
Kubernetes
OpenShift
Apache Airflow
Snowflake
Databricks
FastAPI
IaC
MLOps
CI/CD
LLM
GenAI
Thông tin chung
Nơi làm việc
- Remotely, Floor 13, MB Sunny Tower, 259 Tran Hung Dao, Cau Ong Lanh, Ho Chi Minh