Platform Reliability Engineer

ABOUT CLIENT

Our client is a reputable company specializing in software development and IT consulting services

JOB DESCRIPTION

Maintain production reliability of the Linux-based research and trading platform within a globally distributed engineering team.
Respond quickly to production infrastructure issues.
Comprehend internal client needs and effectively communicate them to regional and global leadership.
Identify risks, develop contingency plans, and implement solutions to mitigate them.
Enhance the observability platform to monitor the performance and health of critical computing environments.
Take part in occasional on-call rotations and support on-call staff during their shifts.
Contribute to organizational knowledge through documentation, education, and writing maintainable code.

JOB REQUIREMENT

At least 2 years of experience in SRE, DevOps, or similar infrastructure engineering roles, with a preference for experience in the financial industry.
Knowledge of Linux system internals, including kernel operations, memory management, and performance optimization.
Familiarity with storage technologies, especially those used in high-performance computing (experience with GPFS is a bonus).
Broad understanding of IT infrastructure components such as networking, DNS, NTP/PTP, and NIS.
Proficiency in system automation, monitoring, and self-healing, with experience in Salt seen as a positive attribute.
Experience with container orchestration and virtualization technologies like Kubernetes, Nomad, and VMware.
Understanding of on-premises and cloud-based HPC infrastructure, with operational knowledge of Slurm and GPU considered a bonus.
Awareness of AI technologies and their applications in infrastructure automation and management.
Experience or strong interest in implementing AI/ML solutions for infrastructure optimization, anomaly detection, or predictive analytics.
Passion for technology and automation, with a deep sense of curiosity and ownership.
Hands-on problem-solving approach and enthusiasm for technology.
Excellent verbal and written English communication skills.

WHAT'S ON OFFER

Be part of a dynamic and passionate team working on cutting-edge projects using the latest technology.
Collaborate with experts from around the globe to enhance your skills and knowledge.
Embrace a culture of transparency and support, valuing individual growth and potential.
Additional month's salary and performance bonuses.
Comprehensive healthcare and accident insurance coverage.
Yearly health checkup package.
Various allowances such as lunch, marriage, newborn baby, bereavement, and more.
Well-equipped pantry for a comfortable lunch break.
Diverse sports and social activities like yoga, football, badminton, and tech clubs.
Annual company retreats and team-building events.
Recognition awards for outstanding individual and team performance and long-term service.
Professional development opportunities including advanced English and soft skills training.
Regular social events such as gatherings, games, birthday celebrations, and year-end parties.

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Outsource

Technical Skills:

Devops

Location:

Ho Chi Minh - Viet Nam

Working Policy:

Onsite

Salary:

Negotiation

Job ID:

J01977

Status:

Active

Related Job:

Senior Mobile Security Engineer (Forensics)

Ho Chi Minh - Viet Nam


Product

Analyze large-scale datasets and fraud incidents to uncover attack patterns, fraud clusters, and evolving adversarial behavior, including reconstructing attacker techniques and execution paths. Partner with the mobile development team to design and implement mobile SDK components to securely collect forensic-grade signals, enabling attribution of location spoofing, emulator abuse, rooted/jailbroken environments, and other environment manipulation. Execute and lead deep technical research into emerging mobile fraud and evasion techniques, translating findings into actionable forensic indicators. Build and mature end-to-end incident response capabilities across the stack, partnering with Data Science and ML teams to translate forensic insights into technical features, rules, and detection logic. Provide technical guidance and mentorship to junior engineers on best practices in mobile security, forensics, and data analysis.

Negotiation

View details

AI/ML Engineer

Ho Chi Minh - Viet Nam


Offshore

As an AI/ML Engineer, you'll work across the full ML lifecycle, from model design and experimentation to large-scale deployment and performance optimization, within a distributed, microservices-based, and cloud-native environment. Design, build, and optimize AI/ML pipelines for real-time and batch inference, leveraging modern MLOps practices. Collaborate with data engineers and software developers to integrate models into Our Client's banking platform, ensuring reliability, monitoring, and version control. Research, prototype, and productionize models in areas such as credit scoring, fraud detection, transaction classification, personalization, and conversational AI. Implement robust model evaluation, A/B testing, and drift detection frameworks to ensure accuracy and stability over time. Contribute to internal frameworks and libraries to standardize ML development workflows across teams. Explore and evaluate emerging techniques in LLMs, Generative AI, and reinforcement learning applicable to Our Client's ecosystem. Mentor junior engineers and collaborate closely with product and infrastructure teams to ensure model readiness for global scale.

Negotiation

View details

Senior Software Engineer (PHP/ Golang)

Ho Chi Minh - Viet Nam


Product

  • PHP
  • Golang

#Build great software: Analyze requirements, translate them into technical specifications, and estimate their implementation cost. Be a strong individual contributor, actively writing (or teaching AI how to write) high-quality, maintainable, and scalable code. Contribute to the technical roadmap: Identify and prioritise problems, gaps and technical debt in the product. Research and evaluate new tools, technologies and frameworks to enhance the product and improve development efficiency. Actively ensure the product is technically prepared for future challenges, especially considering security, maintainability, and scalability.#Co-own the product: We emphasize the culture of ownership, so our engineers are empowered and encouraged to provide product development suggestions. Provide technical input and insights to Product Managers regarding tradeoffs between scope, engineering capacity, and time constraints.#Raise the quality bar for the team: Champion best practices and actively participate in code and design document reviews to ensure high quality. Mentor and support other engineers, fostering their growth and development. Cross-team collaboration Maintain communication with other technical teams to avoid duplicate effort, incompatible solutions and solving problems other teams have already resolved.

Negotiation

View details