Data Engineer

ABOUT CLIENT

Our client is using new technology to develop products for the banking industry

JOB DESCRIPTION

Develop and enhance data ingestion pipelines with Python and PySpark to gather and modify data from various sources such as transactions, KYC, AML, authentication, devices, and logs.
Proficiency in SQL, with a preference for PostGres.
Design and maintain data models tailored to support Financial Crime/Fraud detection, profiling, and entity resolution.
Implement data quality checks and ensure data reliability across all environments.
Work closely with Data Scientists, Analysts, Compliance, Operations, and Product/Feature teams to put models and rules into practice.
Utilize jobs, workflows, APIs, and streaming to manage extensive data processing workloads.
Integrate with external systems, for example, sanctions, ID&V, biometrics, and authentication systems, to enhance risk and identity data.
Support automation and monitoring of ETL processes for improved operational efficiency.

JOB REQUIREMENT

Bachelor's degree or equivalent qualification
Over 5 years of experience with strong proficiency in Python, PySpark, Scala and Advanced SQL (preferably PostGres)
Hands-on experience with Databricks, Snowflake, Fabric or similar platforms
Proven hands-on experience working with structured and unstructured data in a production environment.
Familiarity with Agentic AI, MLFlow, ML models, and Eval Secure Coding practices - testing/QA
Comfortable working with cloud-based data platforms (preferably AWS).
Effective communication skills in English for collaborating with cross-functional teams in an international environment.
Proficient in working with Text, Delta, Parquet, JSON, CSV, and XML data formats.
Working knowledge of Spark structured streaming.
Experience with AWS infrastructure and working specifically with S3.
Solid understanding of git-based version control, DevOps, and CI/CD.
Experience with Atlassian stack would be a plus. Knowledge of common web API frameworks and web services.
Strong teamwork, relationship, and client management skills, and the ability to influence peers and senior management to accomplish team goals.
Willingness to embrace modern technology, best practices, and methods of work.
Experience in Financial Crime/AML, KYC, or fraud detection systems.
Familiarity with Entity Resolution frameworks (e.g., Quantexa, Sensing, open source Entity Resolution tools).
Experience with data streaming frameworks (Kafka, Spark Streaming, MQ).

WHAT'S ON OFFER

Company offers meal and parking benefits.
Full benefits and probationary salary provided.
Insurance coverage as per Vietnamese labor law and premium health care for employees and their families.
Work environment is values-driven, international, and agile in nature.
Opportunities for overseas travel related to training and work.
Participation in internal Hackathons and company events such as team building, coffee runs, and blue card activities.
Additional benefits include a 13th-month salary and performance bonuses.
Employees receive 15 days of annual leave and 3 days of sick leave per year.
Work-life balance with a 40-hour workweek from Monday to Friday.

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Product

Technical Skills:

Data Engineering, Big Data

Location:

Ho Chi Minh - Viet Nam

Working Policy:

Hybrid

Job ID:

J01710

Status:

Active

Related Job:

Salesforce Engineer (IC2)

Ho Chi Minh - Viet Nam


Outsource

  • Salesforce

#Salesforce Engineering (Quote-to-Cash) Build and maintain Salesforce processes from Quote through to Cash Work full-stack on the platform: declarative (data model, Flows, validation, security model) and programmatic (Apex, Lightning Web Components) Implement changes safely on a production system used by Go-to-Market every day#Integrations & System Interfaces Maintain and further develop the integrations between Salesforce and internal/external systems (customer portal, HubSpot, license management, billing/ERP, and others) Design and operate integrations using Salesforce REST/SOAP APIs and standard integration patterns Own the health of these interfaces: monitoring, error handling, and data consistency#Collaboration with Go-to-Market / RevOps Partner with RevOps and GTM to translate business needs into robust technical implementations Challenge requirements where they create risk or technical debt Coordinate handoffs at the Lead-to-Quote / Quote-to-Cash boundary#Reliability, Security & Change Safety Protect a revenue-critical system: change control, testing, backup/rollback awareness Apply JTL security and compliance standards to Salesforce work#Optional / stretch With spare capacity, support tech-heavy Lead-to-Quote topics in Salesforce (optional; primary ownership of Lead-to-Quote stays with RevOps)

Negotiation

View details

Senior Bioinformatics Engineer

Ho Chi Minh - Viet Nam


Product

  • Data Science

Serve as the primary bioinformatics subject matter expert for engineering teams developing cloud-native bioinformatics software. Collaborate with software architects to translate scientific workflows into scalable distributed computing architectures. Help engineers understand the computational characteristics, assumptions, and limitations of existing bioinformatics tools. Validate that modernized applications preserve scientific correctness and produce reproducible results. Define biological data models, metadata standards, controlled vocabularies, and best practices for data harmonization. Guide engineering teams in designing scalable approaches for processing large genomic and multi-omics datasets. Evaluate open-source bioinformatics software (e.g. PLINK, Regenie, BCFtools, GATK, Nextflow workflows, etc.) and identify opportunities for cloud-native modernization. Work with distributed computing specialists to determine how algorithms can be parallelized using Spark and other large-scale execution frameworks. Develop validation datasets, benchmarking methodologies, and acceptance criteria for transformed applications. Review engineering designs to ensure biological accuracy and scientific integrity. Collaborate with AI engineering teams on using AI-assisted software transformation while ensuring scientific correctness. Stay current with advances in bioinformatics, computational biology, distributed computing, and cloud-based scientific software.

Negotiation

View details

Senior Software Engineer (Distributed Computing)

Ho Chi Minh - Viet Nam


Product

  • Python
  • AWS
  • Spark

Design and provide input on system architectures, contribute to coding standards, and mentor junior engineers. Work with various teams to ensure software solutions meet business requirements, international standards, and objectives. Tackle complex software development and integration challenges, optimizing system performance. Create high-quality Python code and integrate diverse software components into cohesive solutions, with a focus on cloud computing and life sciences applications. Keep up-to-date with new technologies and Python frameworks, particularly in cloud computing. Oversee testing, deployment, and comprehensive documentation of integrated systems. Actively participate in all phases of the software development lifecycle. This includes creating user stories and engaging in sprint planning to align development efforts with business objectives. Engage with multinational companies, demonstrating flexibility to occasionally adapt to US and EU time zones.

Negotiation

View details