Site Reliability Engineer

JOB DESCRIPTION

We are seeking an engineer to ensure the reliability and performance of Our Client's Data Platform. Successful candidates will work with researchers, operations, and other technology teams to establish the smooth functioning of our production data pipeline sourced from an enormous and continuously updating catalog of vendor and market data. This engineer will also develop solutions to improve the efficiency and scalability of our ever-growing business-critical management system
Operate, monitor, and provision the system to make sure it works smoothly
Provide feedback for system improvement
Provide solutions for live monitoring of the production data pipeline
Design and implement continuous integration and test automation
Deliver release management solutions
Collaborate with engineering, analyst, and research teams to ensure the reliability and operability of new data pipeline components
Analyze and diagnose platform performance and reliability problems
Understand, manage, and utilize the right technologies for building our platforms such as Kubernetes, Kafka, and Spark

JOB REQUIREMENT

Bachelor’s degree in Computer Science or equivalent experience
Excellent analytical skills and a passion for solving problems
Experience in Linux administration; fluent in Linux standard command line programs
Fluency in Python and its ecosystem (numpy, pandas, etc.) is strongly recommended
Experience in metrics and logs aggregation and analysis with a focus on performance optimization
Understanding of Git and CI/CD concept
A great support attitude (our job is to make life easier for other teams!)
Strong written and verbal communication skills; Fluency in the English language
Knowledgeable in:
Computer science fundamentals (algorithms and data structures)
Relational databases
Modern service architectures
Experience in the following technologies is relevant: Kafka, Docker, Helm, Kubernetes, GC, AWS, Spark and Pyspark, Hadoop, Redis, MySQL, gRPC, Apache Arrow, Apache Airflow

WHAT'S ON OFFER

Competitive and attractive compensation package with a clear career road-map – where you feel challenged every day
We offer a strong culture of learning and development: training courses, library, speakers, share and learn events
Learn from who sits next to you! Working in our client's environment, you are surrounded by smart and talented people
Employee resources groups with strong diversity and inclusion culture
Premium Health Insurance and Employee Assistance Program
Generous time-off policy, unlimited sick days, re-creation sabbatical leave (based on tenure), Trade Union benefits for staff and family
Team building activities every month: Local engagement events, monthly team lunches – Employee clubs: football, ping-pong, badminton, yoga, running, PS5, movies, etc.
Annual company trips and occasional global conferences – the opportunity to travel and connect with our global teams
Happy hour with tea breaks, snacks, and meals every day in the office!

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

product, Investment Management

Technical Skills:

Devops, Kubernetes, Kafka, Python

Location:

Ho Chi Minh, Ha Noi - Viet Nam

Working Policy:

Salary:

Negotiation

Job ID:

J01251

Status:

Close

Related Job:

Senior DevOps Engineer

Ha Noi - Viet Nam


Financial services, Crypto

  • Devops
  • AWS

#About the Role We are looking for a highly skilled Senior DevOps Engineer to join our team. You will play a key role in designing, implementing, and maintaining infrastructure solutions that ensure the stability, reliability, and scalability of our systems. The role requires a balance of strong technical expertise, problem-solving skills, and a passion for automation. You will collaborate with cross-functional teams to drive process improvements, enhance engineering productivity, and support business-critical applications. As a Senior DevOps Engineer, you will also be responsible for ensuring system security, leading operational automation efforts, and participating in a 24/7 on-call rotation to maintain uptime and business continuity. In line with our philosophy that development and operations are inseparable, you will leverage your coding expertise to intervene in application modules, develop new automation tools, and ensure seamless integration between code and infrastructure.#Key Responsibilities Develop, maintain, and manage tools to automate operational activities and improve engineering efficiency, including writing custom modules in Node.js or Golang for task management and system orchestration. Troubleshoot, diagnose, and resolve complex software and infrastructure issues, including debugging and modifying application code in Node.js and Golang environments. Update, track, and resolve technical issues in a timely manner. Recommend architectural enhancements and propose process improvements for scalability and reliability. Contribute to application development by intervening in existing modules or creating new ones to enhance system manageability, scalability, and performance Evaluate and implement new technologies, frameworks, and vendor products to support business goals. Apply best-in-class security practices to safeguard critical systems and data. Ensure stability, reliability, and performance of production and non-production environments. Collaborate with engineering, QA, and product teams to align infrastructure with development needs. Participate in a 24/7 on-call rotation to support high-availability systems.

Negotiation

View details

DevOps Engineer

Ho Chi Minh - Viet Nam


Product, Offshore

  • Devops
  • Java
  • Kubernetes

As a Mid-level DevOps Engineer, you will play a key role in building, maintaining, and automating the environments, CI/CD pipelines, and infrastructure that power the mission-critical solutions for our internal teams and external clients. Maintain and ensure the availability of development, staging, and production environments Manage access, runtime stability, and environment upgrades Design, build, and improve CI/CD pipelines using tools like Jenkins, Automate build, testing, and deployment processes Troubleshoot and resolve pipeline issues Develop and maintain automation scripts using Ansible Standardize infrastructure configurations and automate environment provisioning Configure and maintain monitoring, logging, and alerting systems (Grafana, Prometheus, Splunk) Enhance observability with proactive alerting and dashboards Respond to alerts and incidents within agreed SLAs Triage and resolve infrastructure and pipeline issues Document incidents and implement preventative measures Apply system hardening, vulnerability remediation, and patching Support audits and compliance checks Monitor system performance and resource usage Conduct tuning (e.g., JVM, database, message broker) and provide optimization recommendations

Negotiation

View details

Engineering Director

Ho Chi Minh - Viet Nam


product, transportation

  • Backend
  • Frontend
  • Devops

Technology Leadership & System Scalability Oversee and optimize the entire tech stack & architecture to ensure high availability, security, and resilience. Lead infrastructure scaling to support a high volume of daily transactions across multiple services. Implement advanced AI and automation solutions to enhance platform performance. Improve system observability, monitoring, and disaster recovery strategies. Drive cost-efficient CapEx planning, optimizing for performance vs. budget balance. Team & Culture Building Manage and mentor a cross-functional tech team of more than 100 members (Backend, Mobile, SRE, QA, Data science…). Cultivate a high-speed, ownership-driven culture that promotes integrity, collaboration, and accountability. Develop technical leadership pipelines, ensuring top talent development and retention. Cultivate an engineering culture that embraces innovation, continuous learning, and rapid iteration. Future-Readiness & Technology Planning Define a 3-year technology roadmap aligned with business goals and market evolution. Evaluate emerging tech trends (AI, cloud, blockchain, automation) and implement relevant solutions. Collaborate with business and operations teams to align tech investments with growth strategies. Execution & Engineering Excellence Establish world-class engineering practices (CI/CD, DevOps, microservices, cloud architecture). Optimize for high performance, low latency, and fault tolerance in real-time operations. Lead technical transformations, re-architecture projects, and innovation initiatives.

Negotiation

View details