Site Reliability Engineer (Shift-working)

ABOUT CLIENT

Our client is a global technology company that specializes in providing innovative IT solutions for the financial services industry

JOB DESCRIPTION

The Senior SRE plays a vital role in overseeing the everyday operations of the organization. It is crucial for this position to have a solid understanding of various technical aspects such as production system access and control, production deployment, Amazon Web Services, Kubernetes, continuous deployment, and systems observability.
 
Key Responsibilities
Take part in on-call rotations to provide round-the-clock support for critical systems.
Address system incidents promptly and effectively
Implement changes in staging and production environments
Collaborate with Platform Engineers to comprehend the changes
Establish deployment pipeline for changes
Comprehend the changes and build observability (monitoring and alert) as per the changes
Design and execute resiliency testing solutions
Continuously improve monitoring solutions
Create and update operational runbooks
Automate operational runbooks

JOB REQUIREMENT

Technical Skills
Proficient in Amazon Web Services
Proficient in Kubernetes system
Proficient in Python or Bash scripting
Familiarity with continuous deployment tools
Familiarity with Harness is a plus
Familiarity with infrastructure as code (IaC) tools, particularly Terraform
Experience with observability solutions like Prometheus and Grafana
Familiarity with SumoLogic is a plus
 
Soft Skills
Effective communication skills, fluent in English
Strong problem-solving abilities
Self-motivated and quick learner

WHAT'S ON OFFER

Attractive salary
13th-month salary and performance bonus
Professional English course available for all employees
Comprehensive health insurance package

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Outsource

Technical Skills:

Devops, AWS, Google Cloud

Location:

Ho Chi Minh, Ha Noi - Viet Nam

Working Policy:

Hybrid

Salary:

Negotiation

Job ID:

J01150

Status:

Close

Related Job:

Lead Data Engineer

Ho Chi Minh, Ha Noi - Viet Nam


Outsource

  • Data Engineering
  • Management

Architect, develop, and maintain scalable data infrastructure, including data lakes, pipelines, and metadata repositories, ensuring the timely and accurate delivery of data to stakeholders. Work closely with data scientists to build and support data models, integrate data sources, and support machine learning workflows and experimentation environments. Develop and optimize large-scale, batch, and real-time data processing systems to enhance operational efficiency and meet business objectives. Leverage Python, Apache Airflow, and AWS services to automate data workflows and processes, ensuring efficient scheduling and monitoring. Utilize AWS services such as S3, Glue, EC2, and Lambda to manage data storage and compute resources, ensuring high performance, scalability, and cost-efficiency. Implement robust testing and validation procedures to ensure the reliability, accuracy, and security of data processing workflows. Stay informed of industry best practices and emerging technologies in both data engineering and data science to propose optimizations and innovative solutions.

Negotiation

View details

Windows Engineer (C++/C#) - GSaaS

Ho Chi Minh - Viet Nam


Product

  • C/C++

Develop and maintain applications using C# (WinUI framework) and C++ (Qt framework and Win32 API). Participate in the company's software development projects and collaborate with cross-functional teams on software architecture. Develop new features according to requirements, provide development documentation, and participate in code reviews. Troubleshoot, debug, and optimize performance for existing software features and applications. Write high-quality, testable code, ensuring adherence to high code quality standards. Research and integrate new technologies to enhance software products. Mentor junior developers and contribute to team knowledge sharing.

Negotiation

View details

Senior AI Engineer

Ho Chi Minh - Viet Nam


Product

  • Python
  • AI
  • Machine Learning

We're seeking an AI Engineer with strong academic foundations and deep technical expertise who excels at translating research into production banking systems. This role is 80% focused on engineering excellence-deploying models, optimizing infrastructure, ensuring reliability, and solving real-world implementation challenges-and 20% on staying current with cutting-edge AI research and emerging technologies. You'll bridge the gap between state-of-the-art AI research and scalable production systems in the financial services sector.#AI Engineering & Deployment (80%) Design, build, and deploy production-ready AI/ML systems on AWS with focus on reliability, scalability, and performance for banking applications Implement and maintain MLOps pipelines using AWS services (SageMaker, Bedrock, Lambda, Step Functions) including model versioning, monitoring, and automated retraining workflows Build and optimize AI solutions using AWS Bedrock, OpenAI API, and Gemini API combining with Model Context Protocol (MCP), Agent-to-Agent (A2A) protocol for various banking use cases Design and implement prompt engineering frameworks and prompt management systems for LLM-based applications Develop graph analysis solutions for fraud detection, customer relationship mapping, and network analysis in banking contexts Debug and troubleshoot production AI systems, identifying and resolving issues in model performance, data pipelines, and AWS infrastructure Build and maintain AIOps practices including automated monitoring, alerting, and incident response for AI systems on AWS Optimize model serving infrastructure for latency, throughput, and cost-efficiency using AWS services Implement robust data pipelines using AWS Glue, Kinesis, and related services for training and inference Collaborate with software engineering and risk teams to integrate AI capabilities into banking products and services Ensure compliance with banking regulations and security standards in all AI deployments Monitor model performance in production and implement drift detection and retraining strategies#AI Research & Innovation (20%) Stay current with latest AI research papers and breakthroughs, evaluating applicability to banking and financial services Research and prototype emerging AI architectures and techniques for financial use cases Evaluate new paradigms in model training, inference optimization, and architectural innovations Share knowledge through technical discussions, paper reviews, and internal research presentations Identify opportunities to apply cutting-edge research to improve fraud detection, customer service, risk assessment, and other banking operations

Negotiation

View details