Lead Cloud Engineer

ABOUT CLIENT

Our client is a global technology company that specializes in providing innovative IT solutions for the financial services industry

JOB DESCRIPTION

Provide technical leadership and architectural direction across multiple product teams utilizing AWS.
Define and develop the cloud strategy, ensuring it aligns with business objectives, regulatory requirements, and security standards.
Take responsibility for the end-to-end architecture of AWS-based infrastructure, including networking, compute, storage, and security models.
Mentor Cloud Engineers and establish best practices for Infrastructure-as-Code (IaC), CI/CD, monitoring, and cloud security.
Oversee large-scale cloud migration initiatives (on-premises to AWS) while ensuring minimal downtime and full compliance.
Collaborate with Delivery, Security, and Product stakeholders to convert business needs into technical designs.
Evaluate emerging AWS services and DevOps tools to drive the adoption of solutions that reduce operational toil and improve delivery velocity.
Define and enforce service level objectives (SLOs) and service level indicators (SLIs) for cloud infrastructure to ensure resilience, performance, and observability.
Ensure adherence to security standards (PCI-DSS, ISO, SOC2) and proper IAM governance, encryption, and audit readiness.
Provide support for critical production incidents, leading root cause analysis and continuous improvement actions.

JOB REQUIREMENT

Mastery in cloud engineering, with extensive hands-on experience in core AWS services.
Proven ability to design large-scale, multi-account AWS architectures for regulated industries.
Minimum 3 years in a senior/lead role, mentoring teams and driving governance.
Advanced experience with Terraform or AWS CloudFormation and building reusable infrastructure components.
Experience setting up enterprise-grade pipelines for continuous integration and delivery.
Strong knowledge of cloud security practices, IAM policy design, encryption, audit, and compliance frameworks.
Expertise in architecting for scalability, availability, disaster recovery, and fault tolerance.
Strong ability to engage stakeholders, present technical strategies, and influence decision-making at senior levels.
AWS Certified Solutions Architect - Professional or AWS Certified DevOps Engineer - Professional.
Experience implementing advanced monitoring/logging.
Experience with Kafka or AWS-native equivalents.
Understanding of other platforms while maintaining AWS specialization.
Familiarity with application development (Java, Python) for effective collaboration with dev teams.

WHAT'S ON OFFER

We offer a professional and enjoyable working atmosphere.
We prioritize your long-term development.
We are dedicated to creating a future-ready digital bank platform.
Competitive salary
13th-month salary guarantee
Performance bonus
Access to professional English courses
Premium health insurance
Generous annual leave allowance

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Outsource

Technical Skills:

Devops, Cloud, AWS, Google Cloud

Location:

Ho Chi Minh, Ha Noi - Viet Nam

Working Policy:

Salary:

Negotiation

Job ID:

J00581

Status:

Close

Related Job:

DevOps Engineer

Others - Viet Nam


Product

  • Devops
  • Kubernetes
  • Network

Operate and evolve our Kubernetes platform across multiple clusters and environments (Prod, Dev, hybrid on-prem and public cloud), covering control plane operations, node lifecycle, upgrades, and autoscaling at every layer (Cluster Autoscaler, HPA, KEDA). Architect and manage hybrid cloud infrastructure spanning on-premises and public clouds (GCP, AWS), including workload placement, cross-cloud networking, and unified resource management. Own the CI/CD and GitOps experience end-to-end: container build pipelines, image optimization, and progressive delivery via ArgoCD / FluxCD. Own the observability stack as a single pane of glass across all clusters: Grafana, Mimir, Tempo, Loki, Pyroscope, OnCall, Prometheus -- and help push toward agent-assisted SRE workflows. Manage and improve our inference platform: vLLM serving and AIBrix for multi-model orchestration and autoscaling across a fleet of NVIDIA GPUs. Operate platform services: Kafka, Redis, PostgreSQL, OpenSearch. Manage identity and access via Keycloak integrated with Google Workspace; harden SSO, RBAC, and secrets management across the platform. Harden network security across private load balancers, firewalls, and VPC segmentation; design and maintain hub-and-spoke / multi-AZ topologies. Support training infrastructure: self-service VM provisioning, RunPod burst capacity, Weights and Biases integration. Drive infrastructure reliability, cost efficiency, and capacity planning as the platform scales.

Negotiation

View details

Platform Engineer

Ho Chi Minh - Viet Nam


Product

  • Backend
  • Devops
  • Data Engineering

Build and maintain distributed infrastructure handling telemetry, sensory, and control data across cloud and edge environments Design and operate data ingestion and streaming pipelines connecting robot fleets to the cloud in real time, covering video, joint states, audio, and LiDAR Develop and maintain backend services and APIs that power the Company's developer-facing platform, with a focus on reliability and developer experience Manage and evolve cloud native infrastructure using Kubernetes, Docker, and infrastructure as code tooling Ensure platform reliability through monitoring, alerting, autoscaling, failover, and incident response Support ML and robotics teams with data infrastructure for training pipelines, policy rollout, and hardware-in-the-loop simulation Implement secure APIs with access control, rate limiting, and usage metering as we scale

Negotiation

View details

Software Engineer (Digital Twin)

Ho Chi Minh - Viet Nam


Product

  • Python
  • C/C++

Build and maintain high-fidelity digital twin environments for Asimov across MuJoCo, Isaac Sim, and Unreal Engine, calibrated to real hardware behavior. Design and own the systems -- not just the environments -- that let locomotion, autonomy, and perception teams generate, validate, and iterate on simulation scenarios at scale. Build pipelines for asset import, USD and MJCF workflows, sensor modeling, and real-to-sim calibration to keep digital twins synchronized with evolving hardware. Develop photorealistic rendering pipelines in Unreal Engine for synthetic data generation and perception model training. Work with hardware and mechatronics teams to model actuator dynamics, contact physics, and structural behavior, ensuring simulation parameters reflect physical ground truth. Integrate digital twin environments with the Company's locomotion training pipeline (Cyclotron) and autonomy stack, enabling teams to run experiments and close the sim-to-real gap. Contribute to the open-source Asimov simulation stack, including tooling, documentation, and reproducible environment workflows.

Negotiation

View details