Jobs · itviec
Senior Data Engineer
Northwind Robotics · Quận 7, Hồ Chí Minh · Posted today
About the role
Design and optimize data models and pipelines within a production data lakehouse environment.
Top 3 Reasons To Join Us The Job We operate a production data lakehouse built on Apache Iceberg, Trino, dbt, and Dagster. The core architecture is established and governed through a documented review process; we are hiring a Senior Data Engineer to build, operate, and optimize the platform — and to help evolve it through well-argued design contributions. You will design data models and pipeline architectures within the platform, own production reliability, and deliver trustworthy datasets for AI, analytics, and backend applications. Technology Stack Storage & table format: S3-compatible object storage, Apache Iceberg Catalog: Lakekeeper (Iceberg REST catalog) Query engine: Trino Transformation: dbt Orchestration: Dagster Batch ingestion: Airbyte CDC / streaming: Debezium, Apache Kafka Metadata, lineage & governance: OpenMetadata BI & visualization: Superset Vector storage (AI workloads): pgvector, Qdrant Identity & access control: Keycloak, OPA Primary source systems: PostgreSQL (NestJS backend services), Redis, Elasticsearch Runtime: Kubernetes Key Responsibilities Design data models, dataset layouts, and pipeline architectures for new data domains within the established platform architecture. Build and maintain ELT pipelines (Airbyte ingestion → dbt transformations → curated marts) orchestrated in Dagster. Operate Apache Iceberg tables in production: partitioning strategy, file compaction, snapshot expiration and retention, schema evolution, and time-travel-based reprocessing. Operate and extend CDC ingestion with Debezium and Kafka: connector configuration, PostgreSQL logical replication (WAL / replication slots), schema registry, idempotent sinks, backfill and replay procedures. Tune Trino performance: storage layout, table statistics, query plans, resource groups. Engineer data quality: dbt tests, data contracts, freshness / volume / schema-drift monitoring, duplicate detection, missing-value handling; maintain lineage and metadata in OpenMetadata. Ensure reproducibility and traceability through Iceberg snapshots and versioned dbt models. Build feature and embedding pipelines serving AI services (entity resolution, deduplication, vector stores). Implement governed pipelines: multi-tenant isolation, data classification, access-control integration (Keycloak / OPA), audit logging, and compliance with Vietnamese data-residency requirements. Contribute to architectural evolution: evaluate alternatives, write ADRs, participate in design reviews (e.g., streaming expansion, catalog or vector-store migration). Own production reliability: SLOs, monitoring, runbooks, incident triage, and root-cause analysis for data pipelines. Collaborate with Backend, AI, and Analytics teams to deliver reliable data products. Your Skills and Experience Requirements Bachelor's degree in Computer Science, Information Technology, or a related field — or equivalent practical experience. 5+ years as a Data Engineer with end-to-end production ownership. Strong SQL and Python. Hands-on experience designing Data Lake / Lakehouse and Data Warehouse architectures and data models — and the judgment to work effectively within an established architecture. Proven ELT/ETL pipelines across heterogeneous sources (databases, APIs, files, event streams). Production experience with at least one open table format — Apache Iceberg strongly preferred ; Delta Lake or Hudi acceptable with commitment to transition. Hands-on experience with a distributed SQL engine (Trino / Presto or comparable) and dbt or an equivalent transformation framework. Production orchestration experience — Dagster preferred ; Airflow / Prefect acceptable with commitment to transition. CDC experience (Debezium or equivalent) and working knowledge of PostgreSQL logical replication. Experience building feature pipelines, embedding pipelines, or ML-serving datasets. Data quality engineering: validation, duplicate detection, missing-value handling. English reading proficiency for technical documentation. Preferred Skills (strong plus) Iceberg REST catalogs (Lakekeeper, Polaris, or Nessie); OpenMetadata; Superset. Kafka operations and schema registry management. Running data workloads on Kubernetes. Formal feature stores (e.g., Feast) or NLP data preparation for AI/ML. Spark for batch processing. Experience in regulated or data-residency-constrained environments (e.g., Vietnam PDPL 91/2025, Decree 53/2022). Why You'll Love Working Here I. Giá trị cốt lõi A utomation Arobid có các quy trình để giúp duy trì hiệu quả công việc cũng như sự hài lòng của nhân viên. “Tự chủ” trong nguồn nhân lực của Arobid cho phép nhân viên minh bạch và tích cực trong việc giải quyết vấn đề. R esponsibility Nhân viên Arobid chịu trách nhiệm về công việc, hành vi, thái độ cũng như sự phát triển của chính Arobid. Sự thành công của Arobid dựa trên trách nhiệm và nỗ lực của từng cá nhân. O pen-mindedness Tư duy cởi mở ở nhân viên của chúng tôi là quan trọng để giúp tạo điều kiện cho những thay đổi lớn và nhanh luôn xảy ra trong ngành công nghiệp công nghệ. Cũng như trong việc học hỏi những điều mới và phát triển những cách thức mới, tư duy cởi mở cũng đề cập đến việc chấp nhận sự khác biệt trong đồng nghiệp, không phân biệt giới tính, quốc tịch, tuổi tác. B oldness Dũng cảm, cho phép nhân viên của chúng tôi có thể mạo hiểm, thử nghiệm các phương pháp mới và độc đáo trong việc giải quyết vấn đề. Dũng cảm khác với sự liều lĩnh, dũng cảm đề cập đến các kế hoạch đã được suy nghĩ kỹ lưỡng, thử nghiệm và nghiên cứu cẩn thận và sẽ được thực hiện một cách đúng đắn. I nnovation Sự đổi mới đề cập đến việc thúc đẩy những phương pháp mới, nhân viên tại Arobid luôn phải phát triển khả năng của bản thân, học những cách thức mới và áp dụng chúng vào công việc hàng ngày. Làm cho công việc khó khăn trở nên đơn giản và nhanh chóng hơn. D evotion Sự tận tâm, nhân viên Arobid coi Arobid như một nơi họ thuộc về. Tin tưởng vào sự hình thành của Arobid và mục tiêu cũng như sứ mệnh của Arobid. Cùng nhau, điều này tạo ra một mối liên kết giữa cá nhân với công ty cũng như giữa các đồng nghiệp Arobid. II. Phúc Lợi: Cơ hội tiếp xúc với hàng ngàn doanh nghiệp trong nước và quốc tế. Thăng tiến trong nghề nghiệp chuyên môn và trong lĩnh vực Thương Mại Điện Tử. Cơ hội thử sức trong các lĩnh vực ngoài chuyên môn chính để phát triển bản thân. Thu nhập phù hợp và cơ hội tăng thu nhập như thưởng doanh số, thưởng chỉ tiêu cá nhân, thưởng nhóm, điểm thi đua, điểm cống hiến...
Read the full posting on itviec →
FAQ
Is the Senior Data Engineer role at Northwind Robotics remote?+
This Senior Data Engineer position is listed as onsite (Quận 7, Hồ Chí Minh).
What seniority level is this Senior Data Engineer role?+
This is a senior level position.
How do I apply for the Senior Data Engineer role at Northwind Robotics?+
Use the "Apply on itviec" button to open the original posting on itviec, where you can submit your application directly to Northwind Robotics.