AWS DevOps Lead

ABOUT CLIENT

Our client is using new technology to develop products for the banking industry

JOB DESCRIPTION

Work with various teams to collect requirements and create scalable and sustainable software solutions
Create integration solutions to facilitate smooth communication between microservices, APIs, and external systems
Contribute to the development of continuous delivery, automation frameworks, and pipelines to enhance the developer and customer experience
Improve database interactions and maintain data integrity across distributed systems
Establish best practices for messaging, integration, and data pipeline architectures
Identify and implement enhancements for automation processes and tools
Acquire new skills and support the adoption of a continuous delivery and cloud-first approach.

JOB REQUIREMENT

Minimum 7 years of backend development experience using Python or Java, including at least 2 years in a lead role.
Proficiency in Apache Kafka, including Kafka Connect, Schema Registry, and related components.
Expertise in building and deploying microservice and event-driven architecture, distributed systems, event sourcing, and CQRS patterns.
Experience with AWS foundation services such as VPC, ECS, Lambda, RDS, SNS, SQS, and Eventbridge.
Hands-on experience with tools like Kafka Connectors and Debezium.
Strong experience with application integration patterns, RESTful APIs, and messaging protocols.
Ability to conduct hands-on troubleshooting and optimization of the platform, collaborating closely with team members.
Capability to design scalable systems and multi-country patterns for platforms.
Familiarity with AWS CloudFormation, Terraform, or CDK for infrastructure provisioning.
A focus on automation and the ability to develop tooling for enhancing the efficiency of repeatable tasks, reliability, and performance.
Understanding of cloud change management practices, compliance, and security standards.
Strong English language skills for effective communication and coordination with business partners and technical teams.
Strong logical thinking and problem-solving abilities.
Curiosity and a self-learning attitude are highly desirable.
Big Plus:
AWS Certification in DevOps, SysOps, or Advance Networking Speciality.

WHAT'S ON OFFER

Company offers meal and parking benefits.
Full benefits and probationary salary provided.
Insurance coverage as per Vietnamese labor law and premium health care for employees and their families.
Work environment is values-driven, international, and agile in nature.
Opportunities for overseas travel related to training and work.
Participation in internal Hackathons and company events such as team building, coffee runs, and blue card activities.
Additional benefits include a 13th-month salary and performance bonuses.
Employees receive 15 days of annual leave and 3 days of sick leave per year.
Work-life balance with a 40-hour workweek from Monday to Friday.

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Product

Technical Skills:

Devops, AWS

Location:

Ho Chi Minh - Viet Nam

Working Policy:

Hybrid

Job ID:

J01325

Status:

Close

Related Job:

Senior Software Engineer - DevOps (ERP)

Ho Chi Minh - Viet Nam


Outsource

You drive the step-by-step migration of our ERP from a single virtual-machine setup towards containerized workloads on Linux and Docker. You design, implement, and maintain agent-based deployment for both virtual machines and - increasingly - containers, making sure they are provisioned, deployed, and updated reliably. You provision and operate virtual machines, containers, deployment targets, and delivery infrastructure in Azure. You ensure the cloud onboarding wizard (React) and the surrounding cloud services run reliably so that customers can work without interruption. You instrument the platform with telemetry, monitoring, and observability so that issues are detected early and operations stay transparent. You design, implement, and maintain CI/CD pipelines and automate build, packaging, testing, and deployment processes. You support the migration of existing build and release processes towards GitHub Actions and modern delivery workflows. You manage and evolve self-hosted runners and build infrastructure. You collaborate with Quality Engineering to integrate automated tests into delivery pipelines. You work with Release Management to improve release reliability, transparency, and automation. You identify bottlenecks in the software delivery process and continuously improve developer productivity. You help establish the technical foundation for future Continuous Delivery and Continuous Deployment scenarios#OPERATIONS & ON-CALL RESPONSIBILITYTogether with your future colleague and the wider team, you take shared responsibility for keeping the platform running for our customers. We ensure disruption-free operations primarily through the right tooling and welldefined processes - automation, telemetry, and clear escalation paths - not through manual firefighting. This is explicitly not about working fixed shifts: it is about ownership and reachability, being there within a defined reaction time when problems occur, including on weekends. You take ownership of the reliable operation of our machines and - increasingly - containers, so that customers can keep working. You build the tools and processes that keep operations running smoothly and prevent incidents before they happen. You make sure that deployments and the cloud wizard (React) keep functioning correctly in production. You share responsibility for operational availability in a rotation with the team - the goal is to respond within a defined reaction time when incidents arise, not to staff fixed shifts. You help build the alerting, telemetry, and on-call processes that make fast and reliable incident response possible.

Negotiation

View details

Engineering Manager – Shop 6.0

Ho Chi Minh - Viet Nam


Outsource

  • Management
  • Backend
  • Frontend
  • Devops
  • Azure

Take on overall technical and organizational responsibility for delivering Shop 6.0 across frontend, backend, and infrastructure Plan delivery scope, milestones, and releases, and coordinate the work of parallel workstreams (frontend, backend, DevOps, QA) Make and facilitate key architectural decisions together with the senior engineers, particularly around microservices, APIs, and cloud-native implementation on Azure Identify and manage risks related to integration (especially the ERP Cloud connection), scalability, and technical dependencies Serve as the central point of contact for stakeholders on scope, prioritization, and timelines Ensure code quality through reviews, mentoring, and clear standards, including a pull request process with mandatory checks and senior/lead review Ensure transparency on progress, risks, and decisions towards the team and management Hire, lead, develop and retain a team of 6 engineers (4 backend and 2 frontend)

Negotiation

View details

Senior AI DevSecOps Engineer

Ho Chi Minh - Viet Nam


Product

  • Devops
  • AWS
  • Azure
  • Security

Management of CI/CD Pipeline: Ensure automation, security, and scalability across all stages of the development lifecycle. Infrastructure & Security: Design and implement secure multi-cloud infrastructure solutions leveraging cloud services, containerization, and orchestration tools. Policy as Code: Define and enforce security and compliance policies across Kubernetes clusters using OPA or Kyverno, ensuring guardrails are automated and auditable. AI & Platform Automation: Drive the adoption of AI-powered tools and workflows to automate infrastructure operations, optimize CI/CD pipelines, accelerate root cause analysis, improve security posture, and enhance engineering productivity. Observability & Alerting: Build and maintain a comprehensive observability stack with proactive alerting, dashboards, and runbooks for critical business flows and security events. Secret & Credential Management: Design and enforce secrets management practices across all environments ensuring zero hardcoded credentials in codebases and pipelines. Incident Response & On-Call: Own and continuously improve incident response processes, define runbooks, lead post-mortems, track MTTR, and participate in on-call rotation to maintain platform reliability and SLO adherence. Threat Modelling & Penetration Testing: Conduct regular threat modelling sessions with engineering teams and coordinate or perform penetration testing activities to proactively identify attack surfaces before they reach production. Code Security: Conduct regular code reviews and static/dynamic analysis to identify and remediate security vulnerabilities. Compliance and Best Practices: Ensure compliance with industry standards and best practices. Collaboration: Collaborate with development, operations, and security teams to foster a culture of automation and security-first thinking. Mentorship: Mentor junior engineers and other team members on security best practices. Documentation: Maintain thorough and up-to-date documentation of security policies, procedures, and incident reports. Trend Scouting: Stay updated with the latest trends in technology and AI to integrate innovative solutions into our processes.

Negotiation

View details