Type
Full-time
Work mode
On-site
Level
Staff
Industry
Engineering / Mechanical
Salary
Thương lượng
Location
Quận 10, Hồ Chí Minh, Hồ Chí Minh
Overview
- Build and operate batch and near-real-time data pipelines on our lakehouse platform
- Own assigned data domains end to end — from source ingestion to the trusted
- Design and maintain data models that turn raw operational data into reusable
- Improve pipeline performance, reliability, and infrastructure cost as data volume and
- Establish and maintain data quality standards, including validation, freshness monitoring
- Keep metadata, lineage, and documentation up to date so data consumers can find and
- Support BI, Analytics, and Data Science teams by preparing serving datasets and
- Partner with Backend and Product teams to define event and CDC data contracts for
- Participate in the team's on-call rotation; investigate incidents, restore data SLAs, and
- Contribute to code review, technical documentation, and the team's deployment and
- 2+ years of hands-on experience in Data Engineering, or an equivalent role with
- Strong SQL, including window functions, complex CTEs, and query optimization on
- Production-level Python: structured, tested, and packaged code — not only notebooks or
- Hands-on Spark experience (PySpark or Scala): partitioning, shuffle, join strategies, data
- Practical experience with a workflow orchestrator; Airflow strongly preferred.
- Solid understanding of data warehouse and lakehouse fundamentals: dimensional
- Comfortable with Git and CI/CD workflows as part of daily development.
- Working knowledge of Docker and Kubernetes: able to containerize a job, read pod logs
- Strong debugging discipline: reads logs and metrics, isolates variables, and verifies
- Able to read English technical documentation and open-source code independently.
- Production experience with Apache Iceberg, Delta Lake, or Hudi, including snapshot
- Experience with Trino/Presto, or OLAP engines such as Apache Doris, ClickHouse, or
- Deeper Kubernetes experience: Helm, ArgoCD/GitOps, and resource tuning for Spark
- Streaming experience with Spark Structured Streaming or Flink, including state
- Familiarity with data catalog and quality tooling such as DataHub, OpenMetadata, dbt, or
- Experience with Vault, External Secrets Operator, Terraform, or Terragrunt.
- Exposure to logistics, marketplace, mobility, e-commerce, or on-demand platform data.
- Experience with billing, reconciliation, or finance data pipelines.
- Physical Wellbeing Benefit: General Insurance, Medical check-up, Accident Insurance, Healthcare Insurance.
- Emotional Wellbeing Benefit: Company Trip, Year End Party, Aha Hour Activities, Special Day Gifts, Aha Club (Badminton, Soccer).
- Financial Wellbeing Benefit: Grab/Be For Work (Tech/Lead Level), Workplace Relocation, 13th Month Salary, PP Appreciate, Annual Leave Remain.
Benefits
Healthcare
Summary of facts from the official posting. View original ↗
Interested in this role?
You'll be taken to the employer's official application page.
Apply on official site ↗
Is this your business?
Claim this page, request edits or removal
→