MLOps Engineer

JOB DESCRIPTION

Be a part of building the ideal data and ML/AI ecosystem from scratch. Spearheaded the integration of the latest capabilities to enhance customer experiences and transform business operations. Embrace the vision of democratizing ML/AI technology, making it accessible to all by establishing robust engineering standards, simplifying complexities, and designing effective controls and guardrails. This leadership role goes beyond conventional boundaries, empowering you to lead and innovate across many aspects of our data enablement value stream.
Your role as an MLOps Engineer will be similar to a DevOps engineer, with a stretched focus on productionizing Machine Learning features:
Design and implement scalable AI solutions that enables data engineers and ML scientists to train, build, and maintain machine learning models effectively.
Develop automated processes for continuous model training and evaluation pipelines specifically for ML applications.
Ensure the seamless integration of Company Plus's current architecture with newly added ML functionalities, enhancing overall system capabilities.
Collaborating with diverse stakeholders including business partners, risk, legal, and security teams, as well as UX designers and architects to define and implement robust validation and verification strategies
Fostering a culture of quality coding practices, including test-driven development, unit testing, and secure coding awareness
Focus on business practicality and the 80/20 rule, aiming for a high bar for code quality, but recognize the business benefit of "having something now" vs "perfection sometime in the future"

JOB REQUIREMENT

To grow and be successful in this role, you will bring extensive analytical and technical skills, business acumen and natural curiosity to deliver on product investigations and analysis and support initiatives through insights.
You will ideally bring the following:
Proficiency in one of the scripting/programming languages (Python).
Experience in building data products using GCP/ AWS technologies.
Experience with containerization, Terraform, and GitOps principles for automation and deployment.
Strong background in ML concepts and applications and in-depth knowledge of MLOps best practices.
Agile development mindset, appreciating the benefit of constant iteration and improvement.
Have experience in addressing Tech Debt with minimizing production incidents.
Familiarity with RAG architectures and/or have a good understanding of their application.

WHAT'S ON OFFER

Attractive package including fixed 13-month salary and variable performance bonus
Insurance plan based on full salary
100% full salary and benefits as an official employee from the 1st day of working
Medical benefit (private insurance) for employee and their family
18 paid leaves/year (12 annual leaves and 6 personal leaves)
Working in a fast-paced, flexible, and multinational working environment.
Chance to travel for business trip in foreign countries
Free snacks, refreshment, and parking
Career development in a giant tech hub just entering Vietnam market, with very challenging project
Hybrid working mode, flexible time (3 days in office per week)

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Outsource

Technical Skills:

Machine Learning, Devops, Data Science, Python, Java

Location:

Ho Chi Minh - Viet Nam

Working Policy:

Salary:

Negotiation

Job ID:

J01554

Status:

Close

Related Job:

DevOps Engineer

Others - Viet Nam


Product

  • Devops
  • Kubernetes
  • Network

Operate and evolve our Kubernetes platform across multiple clusters and environments (Prod, Dev, hybrid on-prem and public cloud), covering control plane operations, node lifecycle, upgrades, and autoscaling at every layer (Cluster Autoscaler, HPA, KEDA). Architect and manage hybrid cloud infrastructure spanning on-premises and public clouds (GCP, AWS), including workload placement, cross-cloud networking, and unified resource management. Own the CI/CD and GitOps experience end-to-end: container build pipelines, image optimization, and progressive delivery via ArgoCD / FluxCD. Own the observability stack as a single pane of glass across all clusters: Grafana, Mimir, Tempo, Loki, Pyroscope, OnCall, Prometheus -- and help push toward agent-assisted SRE workflows. Manage and improve our inference platform: vLLM serving and AIBrix for multi-model orchestration and autoscaling across a fleet of NVIDIA GPUs. Operate platform services: Kafka, Redis, PostgreSQL, OpenSearch. Manage identity and access via Keycloak integrated with Google Workspace; harden SSO, RBAC, and secrets management across the platform. Harden network security across private load balancers, firewalls, and VPC segmentation; design and maintain hub-and-spoke / multi-AZ topologies. Support training infrastructure: self-service VM provisioning, RunPod burst capacity, Weights and Biases integration. Drive infrastructure reliability, cost efficiency, and capacity planning as the platform scales.

Negotiation

View details

Platform Engineer

Ho Chi Minh - Viet Nam


Product

  • Backend
  • Devops
  • Data Engineering

Build and maintain distributed infrastructure handling telemetry, sensory, and control data across cloud and edge environments Design and operate data ingestion and streaming pipelines connecting robot fleets to the cloud in real time, covering video, joint states, audio, and LiDAR Develop and maintain backend services and APIs that power the Company's developer-facing platform, with a focus on reliability and developer experience Manage and evolve cloud native infrastructure using Kubernetes, Docker, and infrastructure as code tooling Ensure platform reliability through monitoring, alerting, autoscaling, failover, and incident response Support ML and robotics teams with data infrastructure for training pipelines, policy rollout, and hardware-in-the-loop simulation Implement secure APIs with access control, rate limiting, and usage metering as we scale

Negotiation

View details

Software Engineer (Digital Twin)

Ho Chi Minh - Viet Nam


Product

  • Python
  • C/C++

Build and maintain high-fidelity digital twin environments for Asimov across MuJoCo, Isaac Sim, and Unreal Engine, calibrated to real hardware behavior. Design and own the systems -- not just the environments -- that let locomotion, autonomy, and perception teams generate, validate, and iterate on simulation scenarios at scale. Build pipelines for asset import, USD and MJCF workflows, sensor modeling, and real-to-sim calibration to keep digital twins synchronized with evolving hardware. Develop photorealistic rendering pipelines in Unreal Engine for synthetic data generation and perception model training. Work with hardware and mechatronics teams to model actuator dynamics, contact physics, and structural behavior, ensuring simulation parameters reflect physical ground truth. Integrate digital twin environments with the Company's locomotion training pipeline (Cyclotron) and autonomy stack, enabling teams to run experiments and close the sim-to-real gap. Contribute to the open-source Asimov simulation stack, including tooling, documentation, and reproducible environment workflows.

Negotiation

View details