Site Reliability Engineer (Shift-working)

ABOUT CLIENT

Our client is a global technology company that specializes in providing innovative IT solutions for the financial services industry

JOB DESCRIPTION

The Senior SRE plays a vital role in overseeing the everyday operations of the organization. It is crucial for this position to have a solid understanding of various technical aspects such as production system access and control, production deployment, Amazon Web Services, Kubernetes, continuous deployment, and systems observability.
 
Key Responsibilities
Take part in on-call rotations to provide round-the-clock support for critical systems.
Address system incidents promptly and effectively
Implement changes in staging and production environments
Collaborate with Platform Engineers to comprehend the changes
Establish deployment pipeline for changes
Comprehend the changes and build observability (monitoring and alert) as per the changes
Design and execute resiliency testing solutions
Continuously improve monitoring solutions
Create and update operational runbooks
Automate operational runbooks

JOB REQUIREMENT

Technical Skills
Proficient in Amazon Web Services
Proficient in Kubernetes system
Proficient in Python or Bash scripting
Familiarity with continuous deployment tools
Familiarity with Harness is a plus
Familiarity with infrastructure as code (IaC) tools, particularly Terraform
Experience with observability solutions like Prometheus and Grafana
Familiarity with SumoLogic is a plus
 
Soft Skills
Effective communication skills, fluent in English
Strong problem-solving abilities
Self-motivated and quick learner

WHAT'S ON OFFER

Attractive salary
13th-month salary and performance bonus
Professional English course available for all employees
Comprehensive health insurance package

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Outsource

Technical Skills:

Devops, AWS, Google Cloud

Location:

Ho Chi Minh, Ha Noi - Viet Nam

Working Policy:

Hybrid

Salary:

Negotiation

Job ID:

J01150

Status:

Close

Related Job:

Software Engineer

Ho Chi Minh - Viet Nam


Outsource

Designing and implemenng API based and event driven integraon solutions Designing integraon solutions following Azure best pracces and cloud nave paerns Building integraons using Azure Integraon Services such as Logic Apps, Funcons, API Management, Service Bus, and Event Hubs Implementing and maintaining SAP integraons, for example SAP S/4HANA, SAP PI/PO, or SAP BTP Integraon Suite Developing and maintaining integraons using C# and the .NET ecosystem Applying Infrastructure as Code pracces using tools such as Terraform Ensuring secure authencaon, authorizaon, and API security using OAuth and best pracces Collaborang with architects, developers, and customers to design end to end integraon solutions Supporng deployments, monitoring, and connuous improvement of integraon platforms, ensuring reliability and observability in producon environments

Negotiation

View details

Senior .NET Engineer

Ho Chi Minh - Viet Nam


Product

Are you a passionate .NET developer eager to make a real-world impact? Join our InsurTech team and help build platforms that support customers during critical situations such as property damage or unexpected displacement.You will contribute to developing scalable, integration-heavy systems that manage the full lifecycle from insurance claim to accommodation booking and billing. We are looking for a proactive engineer who thrives in a collaborative, fast-paced environment and is motivated to solve complex, real-world problems. Take ownership of complex workflows: Work closely with stakeholders to implement/integrate end-to-end processes from claim intake to booking, stay and pay platform. Build scalable, distributed systems: Develop robust backend services using .NET, focusing on microservices and high system reliability. Work on integration-heavy systems: Integrate with external insurance platforms, accommodation providers, and internal systems using APIs and messaging patterns. Ensure system quality and reliability: Write unit and integration tests, troubleshoot production issues, and maintain high standards for performance and stability. Contribute to continuous improvement: Refactor and optimize existing systems, improve architecture, and adopt best practices in software design. Collaborate in a cross-functional environment: Partner with other engineers such as: Dev, PM, and QA engineers to deliver high-quality solutions. Drive technical documentation: Maintain clear and structured documentation to support system evolution and onboarding.

Negotiation

View details

Locomotion Research Engineer

Others - Singapore


Product

Create and train RL locomotion policies for various movement types Establish and maintain simulation environments using custom actuator models to replicate hardware characteristics Implement domain randomization strategy to address simulation-to-reality discrepancies Validate and fine-tune locomotion controllers in simulation and physical platforms Utilize Data Engine telemetry data to refine simulation parameters Collaborate with different teams on issues related to locomotion performance Contribute to open-source releases of locomotion models, training code, and simulation assets

Negotiation

View details