Director Engineering – Software Engineering and AI Inferencing Platforms

ABOUT CLIENT

Our client is a leading technology company specializing in graphics processing units (GPUs) and artificial intelligence (AI).

JOB DESCRIPTION

Lead and expand engineering teams in Vietnam across system software, data science, and AI platforms.
Drive the creation, structure, and delivery of high-performance system software platforms that support AI products and services.
Collaborate with global teams across Machine Learning, Inference Services, and Hardware/Software integration to guarantee performance, reliability, and scalability.
Oversee the development and optimization of AI delivery platforms in Vietnam, including NIMs, Blueprints, and other flagship services.
Collaborate with open-source and enterprise data and workflow ecosystems to advance accelerated AI factory, data science, and data engineering workloads.
Promote continuous integration, continuous delivery, and engineering best practices across multi-site R&D Centers.
Work with product management and other stakeholders to ensure enterprise readiness and customer impact.
Establish and implement standard processes for large-scale, distributed system testing including stress, scale, failover, and resiliency testing.
Ensure security and compliance testing aligns with industry standards for cloud and data center products.
Mentor and develop talent within the organization, fostering a culture of quality and continuous improvement.

JOB REQUIREMENT

Computer Science, Computer Engineering, or related field Bachelor's, Master's, or PhD.
15+ years software engineering experience, 6+ years in senior leadership roles.
Managed large, high-performing software teams and delivered complex AI/ML or data-driven products.
Expertise in cloud, data, and accelerated computing technologies (e.g., Spark, Kubernetes, Dask, Python ecosystem, CUDA).
Experience collaborating with open-source communities and enterprise partners.
Strong leadership, communication, and cross-functional coordination skills.
Strategic mindset with hands-on technical depth in AI, system software, or largescale data platforms.
Experience building and scaling AI/ML Inferencing platforms from concept to production.
Background in GPU programming, CUDA optimization, or system performance engineering.
Deep understanding of microservices, distributed systems, and high-performance data architectures.
Contributions to open-source projects or developer ecosystems.
Knowledge of deep learning, RAG, embeddings, or modern text search frameworks.

WHAT'S ON OFFER

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Product

Technical Skills:

Management, Backend, Devops, Data Engineering, Cloud, AI

Location:

Ho Chi Minh, Ha Noi - Viet Nam

Working Policy:

Onsite

Salary:

Negotiation

Job ID:

J01972

Status:

Active

Related Job:

AI/ML Engineer

Ho Chi Minh - Viet Nam


Offshore

Develop and optimize ML pipelines for both real-time and batch inference, applying modern MLOps best practices. Collaborate cross-functionally with data engineers and software developers to seamlessly integrate models into our client's banking platform, ensuring reliability, monitoring, and version control. Research, prototype, and productionize models in critical domains such as credit scoring, fraud detection, transaction classification, personalization, and conversational AI. Implement robust evaluation frameworks, including A/B testing and drift detection, to maintain accuracy and stability over time. Contribute to internal libraries and frameworks that standardize ML workflows and accelerate development across teams. Explore emerging techniques in LLMs, Generative AI, and reinforcement learning, assessing their applicability to our client's ecosystem. Mentor junior engineers and partner closely with product and infrastructure teams to ensure models are production-ready and scalable globally.

Negotiation

View details

Senior Software Engineer (PHP/ Golang)

Ho Chi Minh - Viet Nam


Product

  • PHP
  • Golang

Analyzing requirements, converting them into technical specifications, and estimating implementation cost. Making significant individual contributions, actively writing high-quality, maintainable, and scalable code. Participating in shaping the technical roadmap by identifying and prioritizing issues, gaps, and technical debt in the product. Researching and evaluating new tools, technologies, and frameworks to enhance the product and improve development efficiency. Ensuring the product is technically prepared for future challenges, with a focus on security, maintainability, and scalability. Encouraging a culture of ownership, empowering engineers to provide product development suggestions. Providing technical input and insights to Product Managers regarding tradeoffs between scope, engineering capacity, and time constraints. Advocating best practices and actively participating in code and design document reviews to ensure high quality. Mentoring and supporting other engineers to foster their growth and development. Collaborating with other technical teams to maintain effective communication and avoid duplicate efforts or incompatible solutions.

Negotiation

View details

Software Engineer

Ho Chi Minh - Viet Nam


Product

  • Backend
  • Devops
  • Cloud
  • Kubernetes
  • Python
  • Javascript
  • Typescript

Create and develop the API Platform with a focus on reliability, performance, and providing a top-tier developer experience Deploy and enhance AI/ML models in scalable, production environments in collaboration with research and applied ML teams Manage and advance a contemporary, cloud-native infrastructure stack utilizing Kubernetes, Docker, and infrastructure-as-code (IaC) tools Ensure platform dependability by designing and implementing telemetry, monitoring, alerting, autoscaling, failover, and disaster recovery mechanisms Contribute to developer and operations workflows, encompassing CI/CD pipelines, release management, and on-call rotations Work collaboratively across teams to implement secure APIs with fine-grained access control, usage metering, and billing integration Continuously enhance platform performance, cost-efficiency, and observability to accommodate scaling and serve users globally.

Negotiation

View details