Senior DevOps Engineer

JOB DESCRIPTION

Work within the team on the various products that the team supports.
Develop tools to improve our ability to rapidly deploy and effectively monitor services in a large-scale distributed environment.
Work with teams to design, develop, and implement innovative software solutions related to the DevOps and Agile transformation of the enterprise.
Ensure Cloud environments are compliant with security policies.
Find out the solutions to achieve highly available, highly scalable systems and reliability.
Maintain, support, and enhance CI/CD environment.
All are monitored and measured.
Evaluate infrastructure cost and find out the solution to optimize cost.
Troubleshoot and performs root cause analysis as well as implement corrective/preventive actions when needed.
Define and document best practices and operational procedures regarding solution deployment and infrastructure maintenance to ensure a smooth handover to other teams.

JOB REQUIREMENT

Must have:
3+ years of professional DevOps or Site Reliability Engineering experience in a fast-paced work environment.
Solid understanding of Linux and Container technology.
Hands-on knowledge of Docker.
Experience hands-on Infrastructure and Configuration as Code abilities e.g. Terraform(strongly preferred), Ansible, Packer.
Experience in building CI/CD pipeline automation, tooling (Github Action, Jenkins(strongly preferred)), and Compliance as code.
Experience with cloud services is essential, in particular, our core AWS Technologies (Organizations, Account Design, VPC, Subnet and Network segmentation, EC2, ASG, Lambda, S3, SQS, SNS, ECS, EKS, RDS, Lambda, Cloudwatch, etc).
Ability to create scripts using Bash, Python, or Golang. Must have the habit of cleaning code, reusing code, and implementing the unit test.
Excellence in analytical and problem-solving skills.
English communication, focus on writing.
Experience with Kubernetes (K8S).
Agile development experience is a plus.
Nice to have:
Strongly preferred: AWS Certificates (SysOps, DevOps, SAA, SAP)
Hashicorp Certificate or Experience with Harshicorp stacks such as Terraform Cloud/Terraform Enterprise, Packer, and Vault.
Kubernetes certifications (CKA/CKD).

WHAT'S ON OFFER

Great salary package and Semiannual performance-salary review.
100% official salary during the probation period.
13th-month salary & bonus (Token Bonus/ Investment Allocation)
Full-paid compulsory insurance according to Vietnam Labor Law
Premium Healthcare insurance (support for spouse and children)
12 days annual leave & other leaves, as below:
Birthday Leave: Company encourages employees to take time off and spend it with their loved ones on their birthdays with 1 day of paid leave.
Charity Leave: In an effort to encourage all of our employees to get involved in various charity organizations, we provide 2 days of paid leave per year to yours perform volunteer work
Self-development Leave: Company values employee self-development, employees not only are given a budget to be used solely to pursue personal development but also 2 days of paid leave per year for these activities.
Transportation & Lunch Allowances
Macs/ Laptop and other work-equipments will be provided
Yearly company trips and many outing trips/ team-bonding activities
Unlimited potential for the career path
Budget for your training & self-development
Fantastic yet professional working environment
Lovely, friendly, and talented colleagues
Working hours: 8 hours x 5 days/week (Monday to Friday) with flexible working hours
A pantry is full of tasty food & beverage.
Other benefits will surprise you!

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Product

Technical Skills:

Devops, AWS, Google Cloud

Location:

Ho Chi Minh - Viet Nam

Working Policy:

Salary:

Negotiation

Job ID:

J01286

Status:

Close

Related Job:

DevOps Engineer

Others - Viet Nam


Product

  • Devops
  • Kubernetes
  • Network

Operate and evolve our Kubernetes platform across multiple clusters and environments (Prod, Dev, hybrid on-prem and public cloud), covering control plane operations, node lifecycle, upgrades, and autoscaling at every layer (Cluster Autoscaler, HPA, KEDA). Architect and manage hybrid cloud infrastructure spanning on-premises and public clouds (GCP, AWS), including workload placement, cross-cloud networking, and unified resource management. Own the CI/CD and GitOps experience end-to-end: container build pipelines, image optimization, and progressive delivery via ArgoCD / FluxCD. Own the observability stack as a single pane of glass across all clusters: Grafana, Mimir, Tempo, Loki, Pyroscope, OnCall, Prometheus -- and help push toward agent-assisted SRE workflows. Manage and improve our inference platform: vLLM serving and AIBrix for multi-model orchestration and autoscaling across a fleet of NVIDIA GPUs. Operate platform services: Kafka, Redis, PostgreSQL, OpenSearch. Manage identity and access via Keycloak integrated with Google Workspace; harden SSO, RBAC, and secrets management across the platform. Harden network security across private load balancers, firewalls, and VPC segmentation; design and maintain hub-and-spoke / multi-AZ topologies. Support training infrastructure: self-service VM provisioning, RunPod burst capacity, Weights and Biases integration. Drive infrastructure reliability, cost efficiency, and capacity planning as the platform scales.

Negotiation

View details

Platform Engineer

Ho Chi Minh - Viet Nam


Product

  • Backend
  • Devops
  • Data Engineering

Build and maintain distributed infrastructure handling telemetry, sensory, and control data across cloud and edge environments Design and operate data ingestion and streaming pipelines connecting robot fleets to the cloud in real time, covering video, joint states, audio, and LiDAR Develop and maintain backend services and APIs that power the Company's developer-facing platform, with a focus on reliability and developer experience Manage and evolve cloud native infrastructure using Kubernetes, Docker, and infrastructure as code tooling Ensure platform reliability through monitoring, alerting, autoscaling, failover, and incident response Support ML and robotics teams with data infrastructure for training pipelines, policy rollout, and hardware-in-the-loop simulation Implement secure APIs with access control, rate limiting, and usage metering as we scale

Negotiation

View details

Software Engineer (Digital Twin)

Ho Chi Minh - Viet Nam


Product

  • Python
  • C/C++

Build and maintain high-fidelity digital twin environments for Asimov across MuJoCo, Isaac Sim, and Unreal Engine, calibrated to real hardware behavior. Design and own the systems -- not just the environments -- that let locomotion, autonomy, and perception teams generate, validate, and iterate on simulation scenarios at scale. Build pipelines for asset import, USD and MJCF workflows, sensor modeling, and real-to-sim calibration to keep digital twins synchronized with evolving hardware. Develop photorealistic rendering pipelines in Unreal Engine for synthetic data generation and perception model training. Work with hardware and mechatronics teams to model actuator dynamics, contact physics, and structural behavior, ensuring simulation parameters reflect physical ground truth. Integrate digital twin environments with the Company's locomotion training pipeline (Cyclotron) and autonomy stack, enabling teams to run experiments and close the sim-to-real gap. Contribute to the open-source Asimov simulation stack, including tooling, documentation, and reproducible environment workflows.

Negotiation

View details