Site Reliability Engineer

JOB DESCRIPTION

We are seeking an engineer to ensure the reliability and performance of Our Client's Data Platform. Successful candidates will work with researchers, operations, and other technology teams to establish the smooth functioning of our production data pipeline sourced from an enormous and continuously updating catalog of vendor and market data. This engineer will also develop solutions to improve the efficiency and scalability of our ever-growing business-critical management system
Operate, monitor, and provision the system to make sure it works smoothly
Provide feedback for system improvement
Provide solutions for live monitoring of the production data pipeline
Design and implement continuous integration and test automation
Deliver release management solutions
Collaborate with engineering, analyst, and research teams to ensure the reliability and operability of new data pipeline components
Analyze and diagnose platform performance and reliability problems
Understand, manage, and utilize the right technologies for building our platforms such as Kubernetes, Kafka, and Spark

JOB REQUIREMENT

Bachelor’s degree in Computer Science or equivalent experience
Excellent analytical skills and a passion for solving problems
Experience in Linux administration; fluent in Linux standard command line programs
Fluency in Python and its ecosystem (numpy, pandas, etc.) is strongly recommended
Experience in metrics and logs aggregation and analysis with a focus on performance optimization
Understanding of Git and CI/CD concept
A great support attitude (our job is to make life easier for other teams!)
Strong written and verbal communication skills; Fluency in the English language
Knowledgeable in:
Computer science fundamentals (algorithms and data structures)
Relational databases
Modern service architectures
Experience in the following technologies is relevant: Kafka, Docker, Helm, Kubernetes, GC, AWS, Spark and Pyspark, Hadoop, Redis, MySQL, gRPC, Apache Arrow, Apache Airflow

WHAT'S ON OFFER

Competitive and attractive compensation package with a clear career road-map – where you feel challenged every day
We offer a strong culture of learning and development: training courses, library, speakers, share and learn events
Learn from who sits next to you! Working in our client's environment, you are surrounded by smart and talented people
Employee resources groups with strong diversity and inclusion culture
Premium Health Insurance and Employee Assistance Program
Generous time-off policy, unlimited sick days, re-creation sabbatical leave (based on tenure), Trade Union benefits for staff and family
Team building activities every month: Local engagement events, monthly team lunches – Employee clubs: football, ping-pong, badminton, yoga, running, PS5, movies, etc.
Annual company trips and occasional global conferences – the opportunity to travel and connect with our global teams
Happy hour with tea breaks, snacks, and meals every day in the office!

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

product, Investment Management

Technical Skills:

Devops, Kubernetes, Kafka, Python

Location:

Ho Chi Minh, Ha Noi - Viet Nam

Salary:

Negotiation

Job ID:

J01251

Status:

Close

Related Job:

Python Developer (DevOps - focused)

Ho Chi Minh - Viet Nam


Outsourcing

  • Python
  • Devops

We are looking for a technically strong Python Developer to join our dynamic operations team. In this role, you will be the first line of support for researchers and internal users by managing and resolving issues via our JIRA service desk. You will play a key role in driving operational efficiency through automation and smart tooling, ensuring timely and effective support. To help manage the flow of issues and resolve them. For the more complex issues they can then redirect to the Devops team for their handling, but the support engineer must still keep ownership; To come up with solutions to improve the efficiency of resolving issues. This would include exploring the user of bots to automate some of the common tasks, as well as to write scripts to programmatically categorize and handle the tickets in the JIRA service desk.

Negotiation

View details

QA Engineer (Data Testing)

Ho Chi Minh - Viet Nam


Outsourcing

  • Manual Test
  • Automation Test

Validate the accuracy, completeness, and consistency of output through data pipeline, ensuring it meets business requirements and quality standards. Develop and maintain automation test cases using Python and/or Java to streamline data validation processes. Work closely with data engineers, supporters, and other team to identify and resolve data-related issues. Participate in discussions, provide updates, and collaborate with team members in English during daily work.

Negotiation

View details

Distributed Systems Engineer

Ho Chi Minh - Viet Nam


Product

  • Data Engineering
  • Devops

Design & build large-scale distributed services for telemetry ingestion, event streaming, and command orchestration across edge and cloud environments Implement real-time data pipelines using Kafka, NATS, or gRPC streams, ensuring low-latency, high-throughput processing Maintain and optimize stateful services (Redis, InfluxDB, Postgres) for consistency, replication, and failover in multi-region deployments Collaborate with embedded, controls, and ML teams to define API contracts, message schemas (Protobuf), and service SLAs Develop infrastructure-as-code (Terraform, Helm) and CI/CD workflows to automate testing, security scans, and rolling upgrades Monitor & troubleshoot production systems with Prometheus, Grafana, Jaeger, and custom observability tooling to meet 99.99% uptime goals Champion best practices in reliability engineering, capacity planning, and incident response for distributed platforms

Negotiation

View details