Data Engineer (Java, Python, SQL)
- Thỏa thuận
- 7 năm kinh nghiệm
Hạn nộp hồ sơ: 12/11/2026 (Còn 59 ngày)
Ứng tuyển sớm để được ưu tiên
Kết nối với Nhà tuyển dụng để tìm hiểu thông tin và gia tăng cơ hội trúng tuyển
Nhà tuyển dụng đang online
Job description
We are looking for a skilled Data Engineer to design, build, and operate scalable data pipelines that power real-time processing and analytics. You will work on high-throughput data systems, ensuring reliability, performance, and maintainability across the data lifecycle - from ingestion to storage and search.
This role requires strong experience in distributed systems, stream processing, and cloud-native data infrastructure.
Key Responsibilities
Design and implement real-time and batch data pipelines
Build and maintain scalable streaming systems
Develop and optimize stream processing jobs
Ensure reliable ingestion from multiple internal and external data sources
Design event schemas and data contracts
Implement data validation, transformation, and enrichment logic
Optimize storage layouts and lifecycle management strategies
Improve system observability (metrics, logging, alerting)
Troubleshoot and resolve performance bottlenecks in distributed systems
Implement retry, dead-letter, and replay mechanisms
Ensure data quality, consistency, and governance
Collaborate with Backend, DevOps, and Security teams
Your skills and experience
Must-Have
Proficiency in Java, Python, and SQL; strong software engineering fundamentals
Experience with distributed messaging systems (e.g., Apache Kafka) and stream processing frameworks (e.g., Apache Flink)
Knowledge of event-time processing, windowing, state management, and handling out-of-order events
Experience with cloud storage/data lakes, relational databases, and search/indexing engines (e.g., OpenSearch / Elasticsearch)
Familiarity with cloud platforms (AWS preferred), IaC (Terraform), CI/CD, and containerization (Docker)
Strong problem-solving skills and ability to design scalable, fault-tolerant data pipelines
Nice-to-Have
Experience with workflow orchestration platforms
Experience in security, log processing, or observability domains
Schema management tools (Avro, Protobuf, Schema Registry)
Data lake table formats (Iceberg, Hudi, Delta Lake)
Experience with distributed query engines (Athena, Trino, Presto)
Multi-tenant system design
Cost optimization in large-scale cloud environments
Soft Skills
Strong problem-solving and debugging skills in distributed systems
Ownership mindset with attention to reliability and quality
Clear communication and documentation skills
Ability to work cross-functionally
Comfort working in fast-paced environments
Qualifications
Bachelor's degree in Computer Science or equivalent experience.
7+ years of experience in data engineering or distributed systems
Strong fundamentals in system design and scalability
Proven experience operating production-grade data platforms
Ability to balance performance, cost, and reliability
Why you'll love working here
Opportunity to build AI agent systems for two products simultaneously - offensive and defensive security - a rare engineering challenge
Direct influence on product architecture and AI strategy from day one
Work with a team that understands both security and AI deeply
Competitive compensation
We are looking for a skilled Data Engineer to design, build, and operate scalable data pipelines that power real-time processing and analytics. You will work on high-throughput data systems, ensuring reliability, performance, and maintainability across the data lifecycle - from ingestion to storage and search.
This role requires strong experience in distributed systems, stream processing, and cloud-native data infrastructure.
Key Responsibilities
Design and implement real-time and batch data pipelines
Build and maintain scalable streaming systems
Develop and optimize stream processing jobs
Ensure reliable ingestion from multiple internal and external data sources
Design event schemas and data contracts
Implement data validation, transformation, and enrichment logic
Optimize storage layouts and lifecycle management strategies
Improve system observability (metrics, logging, alerting)
Troubleshoot and resolve performance bottlenecks in distributed systems
Implement retry, dead-letter, and replay mechanisms
Ensure data quality, consistency, and governance
Collaborate with Backend, DevOps, and Security teams
Your skills and experience
Must-Have
Proficiency in Java, Python, and SQL; strong software engineering fundamentals
Experience with distributed messaging systems (e.g., Apache Kafka) and stream processing frameworks (e.g., Apache Flink)
Knowledge of event-time processing, windowing, state management, and handling out-of-order events
Experience with cloud storage/data lakes, relational databases, and search/indexing engines (e.g., OpenSearch / Elasticsearch)
Familiarity with cloud platforms (AWS preferred), IaC (Terraform), CI/CD, and containerization (Docker)
Strong problem-solving skills and ability to design scalable, fault-tolerant data pipelines
Nice-to-Have
Experience with workflow orchestration platforms
Experience in security, log processing, or observability domains
Schema management tools (Avro, Protobuf, Schema Registry)
Data lake table formats (Iceberg, Hudi, Delta Lake)
Experience with distributed query engines (Athena, Trino, Presto)
Multi-tenant system design
Cost optimization in large-scale cloud environments
Soft Skills
Strong problem-solving and debugging skills in distributed systems
Ownership mindset with attention to reliability and quality
Clear communication and documentation skills
Ability to work cross-functionally
Comfort working in fast-paced environments
Qualifications
Bachelor's degree in Computer Science or equivalent experience.
7+ years of experience in data engineering or distributed systems
Strong fundamentals in system design and scalability
Proven experience operating production-grade data platforms
Ability to balance performance, cost, and reliability
Why you'll love working here
Opportunity to build AI agent systems for two products simultaneously - offensive and defensive security - a rare engineering challenge
Direct influence on product architecture and AI strategy from day one
Work with a team that understands both security and AI deeply
Competitive compensation
Thông tin chung
- Thu nhập: Thỏa thuận
Việc làm tương tự khác
NGÂN HÀNG TMCP PHƯƠNG ĐÔNG (OCB)
Hà Nội, Hồ Chí Minh
Cạnh tranh
Công ty Cổ phần SmartOSC
Hà Nội, Hồ Chí Minh, Đà Nẵng
30 - 56 triệu
CÔNG TY TNHH NEXLAB IT SOLUTIONS
Xem trang công ty- Địa chỉ công ty: L17-11, Tầng 17, Tòa nhà Vincom Center, 72 Lê Thánh Tôn, Phường Sài Gòn, Thành phố Hồ Chí Minh, Việt Nam
- Quy mô: Từ 26 - 100 nhân viên
Thông tin công việc
Vị trí:
Nhân viên
Hình thức làm việc:
Toàn thời gian
Việc làm tương tự
Cảnh báo dấu hiệu lừa đảo tuyển dụng
Đội ngũ hỗ trợ của JobOKO sẵn sàng đồng hành, tư vấn và giới thiệu những cơ hội việc làm phù hợp, giúp Ứng viên tự tin phát triển sự nghiệp và chinh phục mục tiêu nghề nghiệp bền vững.
Hotline CSKH
1900.63.63.84
Công ty Cổ phần JobOKO Toàn cầu
Đội ngũ hỗ trợ của JobOKO luôn chủ động tư vấn các giải pháp tuyển dụng tối ưu, cam kết đồng hành và hỗ trợ Quý Nhà tuyển dụng đạt được hiệu quả tuyển dụng bền vững.
Hotline CSKH
0962.107.888
Công ty Cổ phần JobOKO Toàn cầu