Senior DevOps Engineer

ABOUT CLIENT

Our client is a leading technology company specializing in graphics processing units (GPUs) and artificial intelligence (AI).

JOB DESCRIPTION

Contribute to the development and maintenance of advanced machine learning software and frameworks with a focus on performance and scalability.
Improve CI/CD pipelines to make the development, testing, and deployment of large-scale machine learning models more efficient.
Set up and manage cloud infrastructure for continuous integration, delivery, and deployment, ensuring high availability and scalability.
Work closely with teams from various departments to enhance development workflows and software delivery speed and quality.
Address and resolve complex issues related to software development, containerization, and cloud infrastructure in production environments.
Create and update detailed documentation for development and deployment processes.
Effectively communicate with both technical and non-technical stakeholders to align expectations and provide transparency throughout the release and deployment process.
Oversee code reviews, testing, and debugging to maintain high-quality code and streamline workflows.
Provide mentorship and guidance to junior engineers to support their professional growth and improve team capabilities.

JOB REQUIREMENT

A degree in Computer Science, Information Systems, Engineering, or related fields or equivalent experience.
5+ years of software engineering experience with expertise in CI/CD, cloud infrastructure, and advanced machine learning frameworks.
Proficiency in automation and orchestration tools like Docker, Kubernetes, Jenkins, and Terraform or similar CI/CD Tools.
Experience with cloud platforms such as AWS, Azure, or GCP.
Strong programming skills in Python and/or other relevant languages.
Experience in creating and implementing scalable software solutions.
Strong problem-solving and analytical skills focused on practical and scalable solutions.
Ability to work effectively in a collaborative environment and handle multiple tasks and projects.
Familiarity with version control systems and configuration management.
Demonstrated ability to quickly learn and adapt to new technologies.
Extensive experience with advanced AI tools and frameworks, including LLMs and NVIDIA Blueprints.
Contributions to open-source projects, showcasing a collaborative and innovative mindset.
Experience in deploying machine learning models on edge devices or platforms.

WHAT'S ON OFFER

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Product

Technical Skills:

Devops, AI, Cloud

Location:

Ho Chi Minh, Ha Noi - Viet Nam

Working Policy:

Onsite

Salary:

Negotiation

Job ID:

J01967

Status:

Active

Related Job:

PreSales Solutions Engineer

Ho Chi Minh - Viet Nam


Product

  • Presale
  • System
  • Google Cloud

PreSales Support: Collaborating with the Sales team to understand client needs and develop tailored solutions using Google Maps and Google Cloud services. This involves conducting technical presentations, product demonstrations, and creating proof of concepts (POCs) for prospective clients, as well as contributing to proposals and RFP responses with detailed technical information. Post-Sales Support: Leading the technical implementation of Google Maps and Google Cloud services, ensuring smooth deployment and integration. Providing ongoing technical support and troubleshooting for clients after implementation, working closely with cross-functional teams to ensure client satisfaction and build long-term relationships. Technical Expertise: Staying up-to-date with the latest Google Maps and Google Cloud technologies, serving as a subject matter expert (SME) for both internal teams and clients. Integrating new features and services into client solutions and providing guidance on best practices. Collaboration: Working closely with Sales, Product, Infrastructure, Data, and Engineering teams to align solutions with client needs and company goals. Mentoring junior team members and contributing to training initiatives.

Negotiation

View details

Sales Consultant

Ho Chi Minh - Viet Nam


Product

  • Sale
  • Cloud

Foster executive relationships with customers, provide strategic direction, and thought leadership Develop business growth opportunities in collaboration with our Client and Google Drive new business development Understand complex customer requirements on a business and technical level Lead opportunities through the entire business cycle, working with cross-functional teams as necessary Focus on portfolio growth through tailored engagement methodologies Plan, pitch, and execute a territory sales strategy

Negotiation

View details

Chief Technology Officer

Ha Noi - Viet Nam


Product

  • Cloud
  • Backend

Planning & designing overall system architecture: Creating a Technology Roadmap for a Game Server system with high concurrency and low latency for global players. Cost optimization: Deciding on the strategy for using Cloud infrastructure (AWS, GCP, Azure) or Hybrid Cloud to balance performance and operational expenses. High-level consultation: Participating in the Executive Board to address the relationship between speed-to-market of features and system stability. Tech-stack selection: Evaluating and finalizing programming languages (Go, C++, Java, Node.js) and processing models (Microservices vs Monolith) suitable for the complex logic of the game. Scalability solution: Directing the development of Auto-scaling, Load Balancing mechanisms, and managing Player State on large clusters. Data management: Designing Database structure (SQL/NoSQL) and Cache system (Redis, Memcached) to handle billions of queries daily without congestion. Ensuring Uptime: Building real-time monitoring and alerting systems to maintain 99.99% Availability. Network security: Implementing solutions to combat DDoS attacks, game fraud (Anti-cheat), and comprehensive user data security. Infrastructure & CI/CD: Standardizing automatic deployment processes to ensure game updates (Hotfix/Update) do not disrupt players. Deployment strategy & Optimization: Developing plans to optimize Cloud Services costs (AWS/GCP/Azure), evaluating the use of Spot Instances, Reserved Instances, or Private Cloud solutions to save operational budget. Meanwhile, establishing 24/7 monitoring and incident response systems.

Negotiation

View details