Platform Reliability Engineer

ABOUT CLIENT

Our client is a reputable company specializing in software development and IT consulting services

JOB DESCRIPTION

Maintain production reliability of the Linux-based research and trading platform within a globally distributed engineering team.
Respond quickly to production infrastructure issues.
Comprehend internal client needs and effectively communicate them to regional and global leadership.
Identify risks, develop contingency plans, and implement solutions to mitigate them.
Enhance the observability platform to monitor the performance and health of critical computing environments.
Take part in occasional on-call rotations and support on-call staff during their shifts.
Contribute to organizational knowledge through documentation, education, and writing maintainable code.

JOB REQUIREMENT

At least 2 years of experience in SRE, DevOps, or similar infrastructure engineering roles, with a preference for experience in the financial industry.
Knowledge of Linux system internals, including kernel operations, memory management, and performance optimization.
Familiarity with storage technologies, especially those used in high-performance computing (experience with GPFS is a bonus).
Broad understanding of IT infrastructure components such as networking, DNS, NTP/PTP, and NIS.
Proficiency in system automation, monitoring, and self-healing, with experience in Salt seen as a positive attribute.
Experience with container orchestration and virtualization technologies like Kubernetes, Nomad, and VMware.
Understanding of on-premises and cloud-based HPC infrastructure, with operational knowledge of Slurm and GPU considered a bonus.
Awareness of AI technologies and their applications in infrastructure automation and management.
Experience or strong interest in implementing AI/ML solutions for infrastructure optimization, anomaly detection, or predictive analytics.
Passion for technology and automation, with a deep sense of curiosity and ownership.
Hands-on problem-solving approach and enthusiasm for technology.
Excellent verbal and written English communication skills.

WHAT'S ON OFFER

Be part of a dynamic and passionate team working on cutting-edge projects using the latest technology.
Collaborate with experts from around the globe to enhance your skills and knowledge.
Embrace a culture of transparency and support, valuing individual growth and potential.
Additional month's salary and performance bonuses.
Comprehensive healthcare and accident insurance coverage.
Yearly health checkup package.
Various allowances such as lunch, marriage, newborn baby, bereavement, and more.
Well-equipped pantry for a comfortable lunch break.
Diverse sports and social activities like yoga, football, badminton, and tech clubs.
Annual company retreats and team-building events.
Recognition awards for outstanding individual and team performance and long-term service.
Professional development opportunities including advanced English and soft skills training.
Regular social events such as gatherings, games, birthday celebrations, and year-end parties.

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Outsource

Technical Skills:

Devops

Location:

Ho Chi Minh - Viet Nam

Working Policy:

Onsite

Salary:

Negotiation

Job ID:

J01977

Status:

Close

Related Job:

Software Architect

Ho Chi Minh - Viet Nam


Outsource

  • Azure
  • .NET

Responsible for creating and overseeing integration architectures on Azure Converting business requirements into integration patterns, data flows, error handling, monitoring, and resiliency models Defining and directing the usage of Azure Integration Services, covering Logic Apps, Functions, API Management, Service Bus, and Event Hubs Leading integration platform design following Infrastructure as Code principles and cloud landing zone considerations Ensuring secure integration architectures utilizing OAuth, OIDC, and API security best practices Guiding development teams through architecture reviews, best practices, and reference implementations using C# and .NET Providing support for SAP integrations, including SAP S/4HANA, SAP PI/PO, and SAP BTP Integration Suite Contributing to integration platform modernization and legacy transformation initiatives, such as BizTalk migrations Collaborating with stakeholders, vendors, and delivery teams to ensure alignment and drive technical decisions

Negotiation

View details

Software Engineer

Ho Chi Minh - Viet Nam


Outsource

  • Azure
  • .NET

Creating API-based and event-driven integration solutions Developing integration solutions following Azure best practices and cloud-native patterns Constructing integrations using Azure Integration Services like Logic Apps, Functions, API Management, Service Bus, and Event Hubs Installing and managing SAP integrations, such as SAP S/4HANA, SAP PI/PO, or SAP BTP Integration Suite Building and maintaining integrations using C# and the .NET ecosystem Utilizing Infrastructure as Code practices with tools like Terraform Ensuring secure authentication, authorization, and API security utilizing OAuth and best practices Working with architects, developers, and clients to devise end-to-end integration solutions Assisting in deployments, monitoring, and continuous improvement of integration platforms, ensuring reliability and observability in production environments

Negotiation

View details

Senior .NET Engineer

Ho Chi Minh - Viet Nam


Product

  • .NET

Take charge of complex workflows: Collaborate with stakeholders to implement and integrate end-to-end processes, from claim intake to booking, stay, and payment platform. Develop scalable, distributed systems: Build resilient backend services using .NET, with a focus on microservices and ensuring high system reliability. Work on integration-heavy systems: Connect with external insurance and accommodation providers as well as internal systems using APIs and messaging patterns. Ensure system quality and reliability: Write unit and integration tests, troubleshoot production issues, and maintain high standards for performance and stability. Contribute to ongoing improvement: Refine and optimize existing systems, enhance architecture, and embrace best practices in software design. Collaborate in a cross-functional environment: Partner with Dev, PM, and QA engineers to deliver high-quality solutions. Drive technical documentation: Maintain clear and structured documentation to support system evolution and onboarding.

Negotiation

View details