Platform Lead

ABOUT CLIENT

Our client is a leading research company specializing in technology innovation

JOB DESCRIPTION

Develop and expand distributed systems to handle large volumes of sensory, telemetry, and control data across cloud and edge environments, facilitating real-time connections for fleets of robots.
Create the API Platform with a focus on high reliability, exceptional developer experience, and robust multimodal AI capabilities accessible through user-friendly APIs and SDKs.
Establish extensive training and inference platforms for foundation models used in robot autonomy, teleoperation, and developer integrations.
Devise data ingestion and streaming pipelines for real-time connectivity of robot fleets to the cloud, covering various data inputs such as video, LiDAR, joint states, and audio.
Oversee and advance a modern cloud native infrastructure stack employing Kubernetes, Docker, and infrastructure as code tools.
Ensure platform reliability through telemetry, monitoring, alerting, autoscaling, failover, and disaster recovery measures.
Make infrastructure decisions pertaining to distributed storage, consensus protocols, GPU orchestration, network reliability, and API security.
Foster collaboration across ML, robotics, and product teams to facilitate hardware in the loop simulation, policy rollout, continuous learning, and CI/CD workflows.
Implement secure APIs featuring fine-grained access control, usage metering, rate limiting, and billing integration to accommodate a growing user base.

JOB REQUIREMENT

At least 7 years of professional experience in software engineering focusing on distributed systems, backend infrastructure, or data platforms.
Proven track record of developing and managing high-scale systems for critical workloads.
Proficiency in Go, Rust, C++, Python, or TypeScript, with a solid understanding of concurrency, networking, and systems performance.
Deep knowledge of cloud native architectures, including Kubernetes, Docker, Helm, gRPC, Kafka, Ray, and service mesh technologies like Istio or Linkerd.
Solid understanding of API architecture and design patterns such as REST, gRPC, WebSockets, OAuth2, and OpenAPI.
Experience with various databases including PostgreSQL, Redis, and modern vector databases such as Pinecone, Weaviate, or FAISS.
Strong understanding of data consistency, replication, and fault tolerance in diverse environments.
Familiarity with observability tools like Prometheus, Grafana, Datadog, or OpenTelemetry in the context of large-scale production systems.
Experience in developing distributed training, large-scale simulation, or fleet-scale telemetry systems.
Familiarity with real-time robotics workloads, including streaming from physical sensors and actuators.
Experience with MLOps tools and AI workflows, encompassing model versioning, inference pipelines, and model registries.
Knowledge of billing systems, quota enforcement, chargeback models, and multi-tenant security and isolation.
Previous involvement in developer platforms, API products, and a strong focus on developer UX and documentation.
Contributions to open source infrastructure, AI frameworks, or robotics middleware such as ROS, gRPC, or Mediasoup are a plus.

WHAT'S ON OFFER

Work remotely in an environment that promotes open-source collaboration
Enjoy 14 days of leave and unlimited sick days
Access to GPUs, AI credits, opportunities for fast career progression, and other perks.

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Product

Technical Skills:

Backend, Devops, Data Engineering

Location:

Others - Singapore

Working Policy:

Onsite

Job ID:

J02064

Status:

Active

Related Job:

Tech Lead (C#/.NET - JTL AI Service Desk)

Ho Chi Minh - Viet Nam


Outsource

  • .NET
  • ReactJS
  • Azure

Create and enhance scalable backend services using C# and .NET technologies Guide architectural choices for distributed and service-oriented systems Construct dependable APIs, integrations, and asynchronous processing workflows Work with AI and data teams to incorporate intelligent automation capabilities into the platform Enhance platform reliability, observability, security, and performance Lead technical discussions, code reviews, and engineering best practices Coach engineers and promote technical development across the team Contribute to long-term platform strategy and technical roadmap Collaborate with frontend, DevOps, and product teams to produce high-quality solutions

Negotiation

View details

Senior Full-Stack Engineer (C# / React, AI Customer Support Platform)

Ho Chi Minh - Viet Nam


Outsource

  • .NET
  • ReactJS
  • Azure

Create and maintain backend services and APIs using C# and .NET technologies Construct contemporary frontend applications and interfaces using React Create adaptable integrations and workflows across platform services Team up with AI and product teams to incorporate intelligent support features and automation Collaborate with frontend, backend, and DevOps teams to produce top-notch solutions Enhance application performance, maintainability, and reliability Engage in technical discussions, code reviews, and architecture decisions Contribute to engineering standards and development best practices

Negotiation

View details

Staff Software Engineer (Customer-Facing BI & Analytics Platform)

Ho Chi Minh - Viet Nam


Outsource

  • .NET
  • ReactJS
  • Azure

Create and implement backend services and APIs for analytics and reporting platforms that can grow with the company Make decisions on the architecture for customer-facing BI and data-heavy applications Establish connections between operational systems, analytics services, and reporting layers Cooperate with frontend and data engineers to produce modern dashboard and reporting experiences Enhance platform scalability, reliability, maintenance, and performance Lead discussions on technical aspects, review code, and advocate for best engineering practices Guide and support engineers to improve their technical skills Contribute to the long-term platform strategy and technical planning Collaborate closely with stakeholders to convert business and customer needs into adaptable technical solutions

Negotiation

View details