VNDirect

IPAS - Data Engineer

Hanoi, Vietnam · On site · Mid level

Vietnamese required

Translated automatically from Vietnamese. Read the original

JOB DESCRIPTION

Build and operate the company's data platform - including storage infrastructure, data processing pipelines and toolsets serving data exploitation and governance needs across the organization.

1. Build data platform and tools

  • Develop and maintain data platform across storage, processing and serving layers, ensuring it meets the needs of data exploitation teams in the company.
  • Build and integrate ETL/ELT tools, metadata analysis and management, create standardized toolsets for teams working with data.
  • Manage and deploy platform infrastructure as code and build CI/CD pipelines, ensuring consistent configuration, easy reproducibility and safe deployment.
  • Guide end users in using data tools in their daily work.

2. Build data pipelines (batch/streaming)

  • Build and operate pipelines bringing data from business sources into the system, using both batch and streaming mechanisms.
  • Organize and model data into logical layers (raw, cleaned, serving) so data is consistent, easy to query and reusable.
  • Build and maintain datasets serving analytics, providing Data Analysts and Data Scientists with data for exploitation.
  • Supply data to other business systems with requirements, through appropriate integration mechanisms.
  • Set up schedules and automation so data is always ready at the right time, reducing manual operations.

3. Data governance, quality and security

  • Manage metadata and lineage so data in the system is clear about origin, meaning and usage.
  • Build automated data tests, monitor quality and handle cases of incorrect or missing data.
  • Implement access control and data access permissions according to company security and compliance requirements.

4. Operations and monitoring

  • Build and maintain monitoring and alerting systems for the platform and data pipelines, detect and handle issues early.
  • Analyze and optimize system performance so data processing tasks run stably and efficiently.

5. Collaboration and improvement

  • Coordinate with business, technical teams and source system operations to align requirements and deploy data solutions.
  • Write and maintain technical documentation for systems and data pipelines.
  • Ensure code quality through review and compliance with team standards.
  • Evaluate current technologies, experiment with new technologies and solutions, propose system improvements.

CANDIDATE REQUIREMENTS

1. Education/Certifications

  • Bachelor's degree in Computer Science, Software Engineering, Information Technology or related fields.
  • Master's degree in related field is preferred.

2. Work experience

  • Minimum 3 years of Data Engineering experience.
  • Experience working with large-scale data processing systems.
  • Experience with cloud computing platforms is an advantage.

3. Knowledge/Professional competencies/Skills

  • Systems thinking: understand data, data flows and how components in a data system operate and connect.
  • Understanding of database design principles, data modeling and data warehousing.
  • Proficiency in SQL and data processing programming languages (Python/Scala).
  • Master techniques for building pipelines to collect, transform and process data at both batch and streaming scales, along with principles ensuring data quality.
  • Knowledge of big data processing technologies, data governance tools.
  • Understanding of platform operations: infrastructure automation (IaC), CI/CD and system monitoring.
  • Self-motivated in work and good collaboration with related departments.
  • Knowledge of applying AI to data engineering work - both in daily tasks and integration into platforms, pipelines to increase automation and efficiency.
  • Understanding of machine learning pipelines and how to prepare and supply data for ML problems.

BENEFITS AND ENTITLEMENTS

  • Flexible and competitive compensation: Receive fixed salary commensurate with ability and attractive bonuses based on actual contribution performance on projects.
  • Clear career development path
  • Benefits and working conditions: Active Cool-down mechanism: After long-term continuous campaigns or projects, the company prioritizes creating space and time for you to transition to rest, training and energy recovery, preparing for the next steps.
  • Full insurance, healthcare and benefits according to State regulations and the Group's special policies.
  • Open, transparent and fully digitalized working environment.
  • Premium health insurance package and 24/24 accident insurance beyond social insurance, health insurance and work injury insurance according to State regulations.
  • Training programs for personal capacity development, team bonding programs, retreats, teambuilding.
The original, in Vietnamese

MÔ TẢ CÔNG VIỆC 

Xây dựng và vận hành nền tảng dữ liệu của công ty - bao gồm hạ tầng lưu trữ, các luồng xử lý dữ liệu và bộ công cụ phục vụ nhu cầu khai thác và quản trị dữ liệu trong toàn tổ chức. 

1. Xây dựng nền tảng và công cụ dữ liệu

  • Phát triển và duy trì nền tảng dữ liệu ở các tầng lưu trữ, xử lý và phục vụ dữ liệu (serving), đảm bảo đáp ứng nhu cầu của các nhóm khai thác dữ liệu trong công ty.
  • Xây dựng và tích hợp các công cụ ETL/ELT, phân tích và quản lý metadata, tạo bộ công cụ chuẩn hóa dùng chung cho các nhóm làm việc với dữ liệu.
  • Quản lý và triển khai hạ tầng nền tảng bằng mã (Infrastructure as Code) và xây dựng quy trình CI/CD, đảm bảo cấu hình nhất quán, dễ tái lập và triển khai an toàn.
  • Hướng dẫn người dùng cuối sử dụng các công cụ dữ liệu trong công việc hằng ngày.

2. Xây dựng các luồng dữ liệu (batch/streaming)

  • Xây dựng và vận hành các pipeline đưa dữ liệu từ các nguồn nghiệp vụ vào hệ thống, theo cả cơ chế batch và streaming.
  • Tổ chức và mô hình hóa dữ liệu thành các tầng hợp lý (thô, làm sạch, phục vụ khai thác) để dữ liệu nhất quán, dễ tra cứu và tái sử dụng.
  • Xây dựng và duy trì các tập dữ liệu phục vụ phân tích, cung cấp cho các nhóm phân tích dữ liệu (Data Analyst) và Data Scientist khai thác.
  • Cung cấp dữ liệu cho các hệ thống nghiệp vụ khác có nhu cầu sử dụng, thông qua các cơ chế tích hợp phù hợp.
  • Thiết lập lịch chạy và tự động hóa để dữ liệu luôn sẵn sàng đúng thời điểm, giảm thao tác thủ công.

3. Quản trị, chất lượng và bảo mật dữ liệu

  • Quản lý metadata và lineage để dữ liệu trong hệ thống rõ ràng về nguồn gốc, ý nghĩa và cách sử dụng.
  • Xây dựng các kiểm thử dữ liệu tự động, giám sát chất lượng và xử lý các trường hợp dữ liệu sai lệch hoặc thiếu hụt.
  • Thực hiện phân quyền và kiểm soát truy cập dữ liệu theo đúng các yêu cầu bảo mật và tuân thủ của công ty.

4. Vận hành và giám sát

  • Xây dựng và duy trì hệ thống giám sát, cảnh báo (monitoring/alerting) cho nền tảng và các luồng dữ liệu, phát hiện sớm và xử lý sự cố.
  • Phân tích và tối ưu hiệu năng hệ thống để các tác vụ xử lý dữ liệu chạy ổn định và tiết kiệm tài nguyên.

5. Cộng tác và cải tiến

  • Phối hợp với các nhóm nghiệp vụ, kỹ thuật và các bộ phận vận hành hệ thống nguồn để thống nhất yêu cầu và triển khai giải pháp dữ liệu.
  • Viết và duy trì tài liệu kỹ thuật cho hệ thống và các luồng dữ liệu.
  • Đảm bảo chất lượng mã nguồn thông qua review và tuân thủ các quy chuẩn chung của nhóm.
  • Đánh giá công nghệ đang dùng, thử nghiệm công nghệ và giải pháp mới, đề xuất các cải tiến cho hệ thống.

YÊU CẦU ỨNG VIÊN  

1. Trình độ học vấn/Chứng chỉ

  • Tốt nghiệp đại học chuyên ngành trong các lĩnh vực như Khoa học máy tính, Kỹ thuật phần mềm, Công nghệ thông tin hoặc các lĩnh vực liên quan.
  • Ưu tiên ứng viên có bằng thạc sĩ trong lĩnh vực liên quan.

2. Kinh nghiệm làm việc

  • Tối thiểu 3 năm kinh nghiệm Data Engineering.
  • Kinh nghiệm làm việc với hệ thống xử lý dữ liệu lớn.
  • Kinh nghiệm làm việc với nền tảng điện toán đám mây là lợi thế.

3. Kiến thức/Năng lực chuyên môn/Kỹ năng

  • Tư duy hệ thống: hiểu về dữ liệu, luồng dữ liệu và cách các thành phần trong một hệ thống dữ liệu vận hành, liên kết với nhau.
  • Hiểu về nguyên tắc thiết kế cơ sở dữ liệu, mô hình hóa dữ liệu và kho dữ liệu.
  • Thành thạo SQL và ngôn ngữ lập trình xử lý dữ liệu (Python/Scala).
  • Nắm được các kỹ thuật xây dựng pipeline thu thập, biến đổi và xử lý dữ liệu ở cả quy mô batch và streaming, cùng các nguyên tắc đảm bảo chất lượng dữ liệu.
  • Có kiến thức về các công nghệ xử lý dữ liệu lớn, các công cụ quản trị dữ liệu.
  • Hiểu về quy trình vận hành nền tảng: tự động hóa hạ tầng (IaC), CI/CD và giám sát hệ thống (monitoring).
  • Chủ động trong công việc và phối hợp tốt với các bộ phận liên quan.
  • Biết ứng dụng AI vào công việc data engineering - cả trong tác vụ hằng ngày lẫn tích hợp vào nền tảng, luồng dữ liệu để tăng tự động hóa và hiệu quả.
  • Hiểu biết về machine learning pipeline và cách chuẩn bị, cung cấp dữ liệu cho các bài toán ML.

LỢI ÍCH VÀ QUYỀN LỢI 

  • Thu nhập linh hoạt & Cạnh tranh: Nhận mức lương cố định tương xứng với năng lực và các khoản thưởng hấp dẫn dựa trên hiệu suất đóng góp thực tế tại các dự án.
  • Lộ trình phát triển sự nghiệp rõ ràng
  • Phúc lợi & điều kiện làm việc: Cơ chế chuyển đổi trạng thái (Active Cool-down): Sau các chiến dịch hoặc dự án dài hạn liên tục, công ty ưu tiên tạo không gian và thời gian để bạn chuyển đổi trạng thái sang nghỉ ngơi, đào tạo và phục hồi năng lượng, chuẩn bị cho những nấc thang tiếp theo.
  • Đầy đủ các chế độ bảo hiểm, y tế, phúc lợi theo quy định của Nhà nước và các chính sách ưu việt riêng của Tập đoàn.
  • Môi trường làm việc cởi mở, minh bạch và số hóa toàn diện.
  • Gói bảo hiểm sức khỏe cao cấp và bảo hiểm tai nạn 24/24 ngoài chính sách BHXH, BHYT, BHTN theo quy định của nhà nước.
  • Các chương trình đào tạo phát triển năng lực cá nhân, chương trình gắn kết đội ngũ, retreat, teambuilding.

Get access to all 2,144 jobs.

Free, with your email. New roles that fit you, every Monday.

OV members also see who works at each company and can ask them for a short chat.

Apply to join

OV member? Use the email you use for OV.

Hiring? Reach Overseas Vietnamese.Post roles