Backend Engineer

JOB DESCRIPTION

Build the core gateway: a unified, OpenAI-compatible API in front of multiple providers (OpenAI, Anthropic, Google, plus self-hosted and OSS models).
Own provider routing and reliability: load balancing, automatic failover, and cost and latency-aware routing.
Build billing and metering that is correct, not approximate: per-request token accounting, usage ledgers, cost attribution per team, user, and key, budgets, and spend limits.
Ship org controls: API key management, per-team and per-user quotas, rate limiting, and RBAC.
Handle streaming and performance: low-overhead proxying, streaming responses, connection handling, and caching where it helps.
Contribute to the Jan Agent and connect it to the router: route its model and tool calls through the gateway, and make agent traffic first-class in metering, controls, and observability.
Make deliberate speed-versus-correctness calls: move fast where iteration is cheap, refuse to cut corners where a bug means a bad charge or a leaked key, and pay down debt on your own initiative.

JOB REQUIREMENT

Work agent-natively as your default, running multiple agents in parallel, pushing token throughput hard, with your own review and guardrail discipline, using open harnesses like Pi, Hermes Agent, and OpenCode as your daily drivers, on open and frontier models alike.
Acquainted with LLM API systems: providers, OpenAI-compatible endpoints, streaming, tool calling, and token accounting.
Experience with LLM infra: inference proxies, provider SDKs, token counting, or existing gateways and routing services (LiteLLM, cliproxyapi, 9router, Omnirouter, and similar) as reference points.
Acquainted with core AI concepts: context windows, inference, prompting, evals, and how open models differ from hosted providers in practice.
Proven ability to ship from zero to production and own the result, including infra, deploy, alerting, and CI/CD, regardless of which stack you did it in.
Open-minded and pragmatic: you weigh speed against correctness case by case, hold strong opinions loosely, and change course when the evidence says so, all while staying fast-moving and comfortable with full ownership and ambiguity.
 
Nice to Have
A track record building production backend services that handle real traffic: APIs, auth, data modeling, deploy, and monitoring.
Experience with payments, billing, metering, or usage-based systems, or the rigor to build them correctly (idempotency, reconciliation, no dropped or double charges).
Comfort with high-throughput proxying, gateways, and streaming, and the performance concerns that come with them.
Solid datastore skills: PostgreSQL, Redis, queues where needed, and sound schema design for usage and billing data.
Strong in at least one of TypeScript or Python.
Comfort with Docker, Kubernetes, OAuth and OIDC, API keys, RBAC, tenant isolation, and secure secrets handling.
Familiarity with the Jan.ai ecosystem or other OSS LLM tooling.
MCP and tool-calling knowledge, directly relevant since you will help the router carry the Jan Agent's agent and tool traffic.
Go or Rust for high-performance proxying, not required.
Contributions to open agent tooling or harnesses: a Pi extension, a Hermes skill, an OpenCode plugin, or anything in the open-models ecosystem we can look at.

WHAT'S ON OFFER

Collaborate with a world-class research team on meaningful, high-impact projects
Own and shape the core training code infrastructure used daily by the team
Work on real models, real data, and real scale - not toy problems
Help bridge the gap between research velocity and engineering quality
Flexible work environment with a culture that values depth, clarity, and curiosity

CONTACT

PEGASI – IT Recruitment Consultancy | Email: recruit@pegasi.com.vn | Tel: +84 28 3622 8666
We are PEGASI – IT Recruitment Consultancy in Vietnam. If you are looking for new opportunity for your career path, kindly visit our website www.pegasi.com.vn for your reference. Thank you!

Job Summary

Company Type:

Product

Technical Skills:

Typescript, NodeJS, Python, AI

Location:

Ho Chi Minh - Viet Nam

Working Policy:

Hybrid

Job ID:

J02226

Status:

Active

Related Job:

AI Transformation Lead

Ho Chi Minh - Viet Nam


Outsource

  • AI

Develop a 3-year AI transformation roadmap aligned with business goals in IT outsourcing, product development, and ODC services. Prioritize highest-impact AI use cases in the organization using a build-vs-buy-vs-partner framework. Establish AI governance for model selection, cost management, data privacy, IP protection, and ethical AI guidelines. Implement AI-assisted development workflows for 1,000+ engineers, including AI code generation, review, automated testing, and AI-powered debugging. Drive adoption of AI orchestration platforms to automate repetitive engineering tasks. Create internal AI skills/training programs and a culture of continuous AI experimentation. Measure and report on productivity gains, quality improvements, and time-to-market acceleration from AI adoption. Collaborate with product teams to define AI features for various products. Lead the development of AI Copilots, intelligent assistants, and autonomous agents embedded within products. Guide the architecture of an Ontology-Based AI ERP/MES Platform, including Knowledge Graphs, GraphRAG, and Multi-Agent Systems. Identify new AI-powered product opportunities in logistics, manufacturing, and supply chain to create new revenue streams. Advocate for AI internally, communicate the vision, celebrate wins, and address concerns across all levels. Partner with HR to define new AI-focused roles and refine hiring criteria. Collaborate with ODC/Client Delivery teams to package and sell AI capabilities to existing and new clients. Represent the company externally in conferences, thought leadership, and talent branding to position the company as an AI leader in the IT services industry.

Negotiation

View details

Senior Backend Engineer (Shop 6.0)

Ho Chi Minh - Viet Nam


Outsource

  • NodeJS
  • Azure

Design and develop the backend services for the core areas of Shop 6.0 (catalog, orders, payments, ERP Cloud integration) Define API boundaries and ensure consistent, scalable communication between services Own reliable data synchronization between the ERP Cloud and the platform, including error handling and recovery mechanisms Optimize systems for performance and scalability (caching, asynchronous processing, read/write optimization) Establish observability standards (logging, monitoring, alerting) across all backend services Conduct code reviews, promote best practices, and support less experienced team members

Negotiation

View details

Senior Tech Lead (Shop 6.0)

Ho Chi Minh - Viet Nam


Outsource

  • Backend
  • Frontend
  • Azure

Take on overall technical and organizational responsibility for delivering Shop 6.0 across frontend, backend, and infrastructure Plan delivery scope, milestones, and releases, and coordinate the work of parallel workstreams (frontend, backend, DevOps, QA) Make and facilitate key architectural decisions together with the senior engineers, particularly around microservices, APIs, and cloud-native implementation on Azure Identify and manage risks related to integration (especially the ERP Cloud connection), scalability, and technical dependencies Serve as the central point of contact for stakeholders on scope, prioritization, and timelines Ensure code quality through reviews, mentoring, and clear standards, including a pull request process with mandatory checks and senior/lead review Ensure transparency on progress, risks, and decisions towards the team and management

Negotiation

View details