ASAI job platform logo
  1. Home
  2. /Jobs
  3. /Backend Engineer Jobs in Bengaluru
  4. /Staff Backend Engineer - K8 (Envoy)
CI
Coupang Internal

Staff Backend Engineer - K8 (Envoy)

Bengaluru · Staff/Principal

Good timing - competition is buildingVerified listingPosted 13d ago

Applicants who checked fit first are 3.1× more likely to hear back

Your match scoreCalculated · locked
86Overall
64Skills
97Experience

Your score for this role already exists

ASAI compared this JD against 41 signals - skills, seniority, domain, stack overlap etc. Add a resume and it unlocks in about 30 seconds.

No credit card · 1 tap with Google

What we know about this role

Hiring pulse

HIGH

Coupang Internal is actively reviewing profiles and moving candidates through the pipeline right now.

Apply window

First 72 hours

Window passed - posted 13d ago

Early applicants get seen before the pile builds.

Not a repost

The first time we've seen this listing - it hasn't been closed and reopened.

Skills required

26 listed
Triton Inference ServerMachine LearningTraffic ShapingLoad BalancingPolicy EnforcementRate LimitingCommon PlatformsReliability Engineering+18 more
Triton Inference ServerMachine LearningTraffic ShapingLoad BalancingPolicy Enforcement

You almost certainly match several of these already. Unlock your skill map to see the matches, the gaps, and what to fix first.

Job description

Please complete the attached Internal Transfer Request Form and submit.  

Please make sure to apply with your Coupang e-mail address.  


Company Introduction :

We exist to wow our customers. We know we’re doing the right thing when we hear our customers say, “How did I ever live without Coupang?” Born out of an obsession to make shopping, eating, and living easier than ever, we’re collectively disrupting the multi-billion-dollar e-commerce industry from the ground up. We are one of the fastest-growing e-commerce companies that established an unparalleled reputation for being a dominant and reliable force in South Korean commerce.

We are proud to have the best of both worlds — a startup culture with the resources of a large global public company. This fuels us to continue our growth and launch new services at the speed we have been since our inception. We are all entrepreneurs surrounded by opportunities to drive new initiatives and innovations. At our core, we are bold and ambitious people that like to get our hands dirty and make a hands-on impact. At Coupang, you will see yourself, your colleagues, your team, and the company grow every day.

Our mission to build the future of commerce is real. We push the boundaries of what’s possible to solve problems and break traditional tradeoffs. Join Coupang now to create an epic experience in this always-on, high-tech, and hyper-connected world.

Role Overview :

As a Staff Backend Engineer, you will work closely with platform and product leaders to design and deliver solutions for complex infrastructure problems. You will drive the development of highly scalable, reliable, and efficient platform services while providing technical direction across teams working with Java, AWS, Kafka, Kubernetes, Kubeflow, Argo CD, and gRPC.

What You Will Do :

  • Architect and build Coupang's next-generation AI Inference Gateway platform that serves mission-critical machine learning and generative AI workloads at scale.
  • Design and develop high-performance request routing, model endpoint abstraction, traffic shaping, load balancing, failover, caching, and policy enforcement mechanisms for inference services.
  • Drive the technical vision and roadmap for scalable, secure, and reliable AI inference infrastructure across cloud and on-prem environments.
  • Develop critical infrastructure components in Go, Java, or Python with a strong focus on performance, resiliency, and operational excellence.
  • Design multi-tenant platform capabilities including authentication, authorization, quota management, cost attribution, rate limiting, and governance controls.
  • Partner closely with ML Platform, Model Serving, Data, and Product Engineering teams to enable seamless deployment and operation of AI workloads.
  • Lead architecture and design reviews, raise engineering standards, and mentor senior engineers across multiple teams.
  • Optimize system performance, latency, throughput, and infrastructure efficiency for large-scale inference workloads.
  • Define observability standards through metrics, tracing, logging, and SLO-based operations for business-critical AI services.
  • Investigate complex production issues, drive root-cause analysis, and implement long-term architectural solutions.
  • Collaborate with engineering leaders across Coupang to establish common platform standards and unlock AI innovation across the organization.

Basic Qualifications:

  • 8+ years of professional software development experience.
  • 5+ years of experience designing and operating large-scale distributed systems in production.
  • Strong hands-on programming expertise in one or more of Go, Java, or Python.
  • Proven track record of building highly available, mission-critical platform or infrastructure services.
  • Experience designing API platforms, service gateways, service mesh, or large-scale networking infrastructure.
  • Deep understanding of microservices architecture, distributed systems design, and cloud-native technologies.
  • Experience with Kubernetes and containerized workloads in production environments.
  • Experience operating services on AWS, Azure, or GCP.
  • Strong understanding of observability, reliability engineering, capacity planning, and production operations.

Preferred Qualifications:

  • Experience building AI/ML inference platforms, LLM gateways, model serving infrastructure, or GPU-accelerated workloads.
  • Deep expertise in Kubernetes ecosystem technologies such as Gateway API, Ingress Controllers, Service Mesh (Istio, Linkerd, Envoy), and platform networking.
  • Experience with inference serving frameworks such as vLLM, Triton Inference Server, TensorRT-LLM, Ray Serve, KServe, SGLang, or similar technologies.
  • Strong understanding of high-performance networking, gRPC, HTTP/2, streaming protocols, API gateways, and service proxy architectures.
  • Experience designing large-scale traffic management systems including routing, retries, circuit breaking, rate limiting, and request prioritization.
  • Experience optimizing latency, throughput, and resource utilization for CPU and GPU workloads.
  • Familiarity with GenAI and LLM ecosystems including model deployment, prompt routing, RAG systems, model observability, and AI governance.
  • Experience with distributed data systems such as Kafka, Cassandra, Redis, MongoDB, or similar technologies.
  • Strong understanding of concurrency, synchronization, asynchronous programming, and non-blocking I/O.

Type of work model :

Hybrid /Onsite / Remote working

  • Our Hybrid work model: Coupang hybrid work model is designed to enable a culture of collaboration that acts a catalyst to enrich the experience of employees. Employees are required to work at least 3 days in the office per week, with the flexibility to work from home 2 days a week, depending on the role requirement. Some businesses may require more time in office due to nature of work.

Details to consider :

Those eligible for employment protection (recipients of veteran’s benefits, the disabled, etc.) may receive preferential treatment for employment in accordance with applicable laws.

Privacy Notice

  • Your personal information will be collected and managed by Coupang as stated in the Application Privacy Notice located below. https://privacy.coupang.com/en/land/jobs/

 

Please complete the attached Internal Transfer Request Form and submit.  

Please make sure to apply with your Coupang e-mail address.  

 

Free · no signup

Get tomorrow's jobs before you have to search

Daily job drops, skill trends and free resources - posted straight to the group. Leave any time.

Join WhatsAppJoin Telegram

No spam. Just jobs and resources.

Why people use ASAI

Someone shared one job with you. ASAI keeps finding the rest.

  • Scored, not searched. Every role ranked against your actual profile.

  • Alerts as often as hourly. Reach new roles while the pile is still small.

  • Skill gaps, spelled out. See exactly which requirements you don't meet yet.

  • Verified jobs, only. Say no to ghost jobs. Your time deserves respect.

More Backend Engineer roles in Bengaluru

See all

Software Engineer - III (NodeJS, Python, Go) - Bengaluru

ClanX · Bengaluru

Software engineer - SOA Backend

Eurofins · Bengaluru

Staff Software Development Engineer (Backend - Python/Go/Rust)

Zscaler · Bengaluru

Backend Solution Engineer - SDE II

Ekacare · Bengaluru

Keep browsing

All open roles at Coupang InternalAll Backend Engineer jobs in Bengaluru

Two ways in

Applicants who checked fit first are 3.1× more likely to hear back

Your match scoreCalculated · locked
86Overall
64Skills
97Experience

Your score for this role already exists

ASAI compared this JD against 41 signals - skills, seniority, domain, stack overlap etc. Add a resume and it unlocks in about 30 seconds.

No credit card · 1 tap with Google

Free · no signup

Get tomorrow's jobs before you have to search

Daily job drops, skill trends and free resources - posted straight to the group. Leave any time.

Join WhatsAppJoin Telegram

No spam. Just jobs and resources.

Why people use ASAI

Someone shared one job with you. ASAI keeps finding the rest.

  • Scored, not searched. Every role ranked against your actual profile.

  • Alerts as often as hourly. Reach new roles while the pile is still small.

  • Skill gaps, spelled out. See exactly which requirements you don't meet yet.

  • Verified jobs, only. Say no to ghost jobs. Your time deserves respect.

ASAI job platform logo

A job platform finally, balanced in your favour.

Jobs by CityJobs by CompanyGuidesHow We VerifyAboutPrivacy PolicyTerms of Service

Built in India 🇮🇳

© 2026 ASAI. All rights reserved.