Distributed Systems Engineer

ABOUT CLIENT

Our client is a leading research company specializing in technology innovation

JOB DESCRIPTION

Design and create distributed systems capable of handling large amounts of sensory, telemetry, and control data across cloud and edge environments.
Plan and implement data ingestion and streaming pipelines to connect groups of robots to the cloud in real-time (video, LiDAR, joint states, audio).
Construct platforms for extensive training and inference to support robot autonomy and teleoperation using foundation models.
Work closely with ML and Robotics engineers to assist in hardware-in-the-loop simulation, policy rollout, and continuous learning initiatives.
Create internal observability systems to monitor fleet performance, reliability, and tuning.
Take the lead on infrastructure decisions such as distributed storage, consensus protocols, GPU orchestration, and network reliability.

JOB REQUIREMENT

Must have more than 7 years of professional experience in software engineering, specializing in distributed systems, networking, or data infrastructure.
Demonstrated capability in constructing and maintaining distributed systems that can handle large-scale workloads.
Proficient in Go, Rust, C++, or Python, with a strong foundation in concurrency, networking, and systems performance.
Familiarity with cloud-native architectures such as Kubernetes, gRPC, Kafka, S3, Ray, or similar frameworks.
Thorough understanding of data consistency, replication, and fault tolerance in heterogeneous environments.
Experience in GPU-based workloads, model training, or edge compute orchestration is desirable.
Strong analytical skills and a preference for developing fast, measurable, and dependable systems.
Experience in creating distributed training or large-scale simulation systems.
Knowledge of real-time robotics workloads, including streaming from physical sensors and actuators.
Previous involvement with telemetry, observability, or fleet-scale systems in production.
Contributions to open-source infrastructure, AI frameworks, or robotics middleware (ROS, gRPC, Mediasoup, etc.) would be advantageous.

WHAT'S ON OFFER

Join an exceptional research team to work on significant and impactful projects
Take charge of and influence the primary training code infrastructure utilized by the team
Engage with actual models, real data, and substantial scale challenges, not small-scale problems
Contribute to bridging the gap between research speed and engineering excellence
Enjoy a flexible work setting with a culture that treasures depth, transparency, and inquisitiveness

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Product

Technical Skills:

Data Engineering, Devops, Golang, Rust, C/C++, Python

Location:

Others - Viet Nam

Working Policy:

Onsite, Remote

Job ID:

J01893

Status:

Close

Related Job:

Backend Engineer

Ho Chi Minh - Viet Nam


Product

  • Typescript
  • NodeJS
  • Python
  • AI

Build the core gateway: a unified, OpenAI-compatible API in front of multiple providers (OpenAI, Anthropic, Google, plus self-hosted and OSS models). Own provider routing and reliability: load balancing, automatic failover, and cost and latency-aware routing. Build billing and metering that is correct, not approximate: per-request token accounting, usage ledgers, cost attribution per team, user, and key, budgets, and spend limits. Ship org controls: API key management, per-team and per-user quotas, rate limiting, and RBAC. Handle streaming and performance: low-overhead proxying, streaming responses, connection handling, and caching where it helps. Contribute to the Jan Agent and connect it to the router: route its model and tool calls through the gateway, and make agent traffic first-class in metering, controls, and observability. Make deliberate speed-versus-correctness calls: move fast where iteration is cheap, refuse to cut corners where a bug means a bad charge or a leaked key, and pay down debt on your own initiative.

Negotiation

View details

Head of Engineering - Marketing Technology

Ho Chi Minh - Viet Nam


Product

  • Management

Develop an integrated roadmap for strategic execution based on the organization's strategic aspirations and lead the implementation process from planning to delivery. Manage multiple engineering teams throughout the organization to achieve desired outcomes, hence having knowledge of the organization's specific areas of operation is an advantage. Collaborate closely with business teams and product owners to verify requirements before and after delivery through showcases and post-production monitoring. Take responsibility for both the development and operation of applications in production, providing active operational support and establishing a clear support model with a focus on site reliability engineering. Direct and oversee the implementation of cybersecurity updates, including keeping software versions current and patching infrastructure regularly. Oversee investment allocation across the organization to maintain alignment, ensure effective spending, and offer insights on the effectiveness and prioritization of investments. Integrate engineering excellence with domain expertise in marketing while pursuing strategic technology objectives.

Negotiation

View details

Senior Backend Engineer (Shop 6.0)

Ho Chi Minh - Viet Nam


Outsource

  • NodeJS
  • Azure

Design and develop the backend services for the core areas of Shop 6.0 (catalog, orders, payments, ERP Cloud integration) Define API boundaries and ensure consistent, scalable communication between services Own reliable data synchronization between the ERP Cloud and the platform, including error handling and recovery mechanisms Optimize systems for performance and scalability (caching, asynchronous processing, read/write optimization) Establish observability standards (logging, monitoring, alerting) across all backend services Conduct code reviews, promote best practices, and support less experienced team members

Negotiation

View details