Staff Backend Engineer - K8 (Envoy)

Posted 2 Hours Ago
Be an Early Applicant
Bengaluru, Bengaluru Urban, Karnataka, IND
In-Office
Senior level
eCommerce • Fintech • Logistics • Retail
The Role
Architect and build a highly scalable AI Inference Gateway for mission-critical machine learning and generative AI workloads. Responsibilities include request routing, model endpoint abstraction, traffic shaping, load balancing, failover, caching, multi-tenant controls, observability, reliability optimization, and production issue resolution. The role provides technical direction across teams, establishes platform standards, leads architecture reviews, and mentors engineers while working across cloud and on-premise environments.
Summary Generated by Built In

Please complete the attached Internal Transfer Request Form and submit.   

Please make sure to apply with your Coupang e-mail address.   

Company Introduction :

We exist to wow our customers. We know we’re doing the right thing when we hear our customers say, “How did I ever live without Coupang?” Born out of an obsession to make shopping, eating, and living easier than ever, we’re collectively disrupting the multi-billion-dollar e-commerce industry from the ground up. We are one of the fastest-growing e-commerce companies that established an unparalleled reputation for being a dominant and reliable force in South Korean commerce.

We are proud to have the best of both worlds — a startup culture with the resources of a large global public company. This fuels us to continue our growth and launch new services at the speed we have been since our inception. We are all entrepreneurs surrounded by opportunities to drive new initiatives and innovations. At our core, we are bold and ambitious people that like to get our hands dirty and make a hands-on impact. At Coupang, you will see yourself, your colleagues, your team, and the company grow every day.

Our mission to build the future of commerce is real. We push the boundaries of what’s possible to solve problems and break traditional tradeoffs. Join Coupang now to create an epic experience in this always-on, high-tech, and hyper-connected world.

Role Overview :

As a Staff Backend Engineer, you will work closely with platform and product leaders to design and deliver solutions for complex infrastructure problems. You will drive the development of highly scalable, reliable, and efficient platform services while providing technical direction across teams working with Java, AWS, Kafka, Kubernetes, Kubeflow, Argo CD, and gRPC.

What You Will Do :

  • Architect and build Coupang's next-generation AI Inference Gateway platform that serves mission-critical machine learning and generative AI workloads at scale.
  • Design and develop high-performance request routing, model endpoint abstraction, traffic shaping, load balancing, failover, caching, and policy enforcement mechanisms for inference services.
  • Drive the technical vision and roadmap for scalable, secure, and reliable AI inference infrastructure across cloud and on-prem environments.
  • Develop critical infrastructure components in Go, Java, or Python with a strong focus on performance, resiliency, and operational excellence.
  • Design multi-tenant platform capabilities including authentication, authorization, quota management, cost attribution, rate limiting, and governance controls.
  • Partner closely with ML Platform, Model Serving, Data, and Product Engineering teams to enable seamless deployment and operation of AI workloads.
  • Lead architecture and design reviews, raise engineering standards, and mentor senior engineers across multiple teams.
  • Optimize system performance, latency, throughput, and infrastructure efficiency for large-scale inference workloads.
  • Define observability standards through metrics, tracing, logging, and SLO-based operations for business-critical AI services.
  • Investigate complex production issues, drive root-cause analysis, and implement long-term architectural solutions.
  • Collaborate with engineering leaders across Coupang to establish common platform standards and unlock AI innovation across the organization.

Basic Qualifications:

  • 8+ years of professional software development experience.
  • 5+ years of experience designing and operating large-scale distributed systems in production.
  • Strong hands-on programming expertise in one or more of Go, Java, or Python.
  • Proven track record of building highly available, mission-critical platform or infrastructure services.
  • Experience designing API platforms, service gateways, service mesh, or large-scale networking infrastructure.
  • Deep understanding of microservices architecture, distributed systems design, and cloud-native technologies.
  • Experience with Kubernetes and containerized workloads in production environments.
  • Experience operating services on AWS, Azure, or GCP.
  • Strong understanding of observability, reliability engineering, capacity planning, and production operations.

Preferred Qualifications:

  • Experience building AI/ML inference platforms, LLM gateways, model serving infrastructure, or GPU-accelerated workloads.
  • Deep expertise in Kubernetes ecosystem technologies such as Gateway API, Ingress Controllers, Service Mesh (Istio, Linkerd, Envoy), and platform networking.
  • Experience with inference serving frameworks such as vLLM, Triton Inference Server, TensorRT-LLM, Ray Serve, KServe, SGLang, or similar technologies.
  • Strong understanding of high-performance networking, gRPC, HTTP/2, streaming protocols, API gateways, and service proxy architectures.
  • Experience designing large-scale traffic management systems including routing, retries, circuit breaking, rate limiting, and request prioritization.
  • Experience optimizing latency, throughput, and resource utilization for CPU and GPU workloads.
  • Familiarity with GenAI and LLM ecosystems including model deployment, prompt routing, RAG systems, model observability, and AI governance.
  • Experience with distributed data systems such as Kafka, Cassandra, Redis, MongoDB, or similar technologies.
  • Strong understanding of concurrency, synchronization, asynchronous programming, and non-blocking I/O.

Type of work model :

Hybrid /Onsite / Remote working

  • Our Hybrid work model: Coupang hybrid work model is designed to enable a culture of collaboration that acts a catalyst to enrich the experience of employees. Employees are required to work at least 3 days in the office per week, with the flexibility to work from home 2 days a week, depending on the role requirement. Some businesses may require more time in office due to nature of work.

Details to consider :

Those eligible for employment protection (recipients of veteran’s benefits, the disabled, etc.) may receive preferential treatment for employment in accordance with applicable laws.

Privacy Notice

  • Your personal information will be collected and managed by Coupang as stated in the Application Privacy Notice located below. https://privacy.coupang.com/en/land/jobs/

 

Please complete the attached Internal Transfer Request Form and submit.  

Please make sure to apply with your Coupang e-mail address. 

 

Skills Required

  • 8+ years of professional software development experience
  • 5+ years designing and operating large-scale distributed systems in production
  • Strong hands-on programming expertise in Go, Java, or Python
  • Experience building highly available, mission-critical platform or infrastructure services
  • Experience designing API platforms, service gateways, service mesh, or large-scale networking infrastructure
  • Deep understanding of microservices architecture, distributed systems design, and cloud-native technologies
  • Production experience with Kubernetes and containerized workloads
  • Experience operating services on AWS, Azure, or GCP
  • Strong understanding of observability, reliability engineering, capacity planning, and production operations
  • Experience building AI/ML inference platforms, LLM gateways, model serving infrastructure, or GPU-accelerated workloads
  • Expertise with Kubernetes ecosystem technologies, including Gateway API, ingress controllers, service mesh, Envoy, or platform networking
  • Experience with inference serving frameworks such as vLLM, Triton Inference Server, TensorRT-LLM, Ray Serve, KServe, or SGLang
  • Understanding of high-performance networking, gRPC, HTTP/2, streaming protocols, API gateways, and service proxy architectures
  • Experience designing large-scale traffic management systems, including routing, retries, circuit breaking, rate limiting, and request prioritization
  • Experience optimizing latency, throughput, and resource utilization for CPU and GPU workloads
  • Familiarity with GenAI and LLM ecosystems, model deployment, prompt routing, RAG systems, model observability, and AI governance
  • Experience with distributed data systems such as Kafka, Cassandra, Redis, or MongoDB
  • Strong understanding of concurrency, synchronization, asynchronous programming, and non-blocking I/O
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
108,000 Employees
Year Founded: 2010

What We Do

Coupang is a U.S. technology and global commerce company founded in 2010, operating online retail and cross-border commerce, restaurant delivery, video streaming, and fintech/payment services. Through brands including Coupang, Eats, Play, Rocket Now, and Farfetch, it uses technology, logistics, and fulfillment infrastructure to serve millions of customers in Korea, Taiwan, the United States, and more than 190 countries and territories worldwide.

Similar Jobs

In-Office
Bengaluru, Bengaluru Urban, Karnataka, IND
70000 Employees

TransUnion Logo TransUnion

Assistant Manager - Batch Processing

Big Data • Fintech • Information Technology • Business Intelligence • Financial Services • Cybersecurity • Big Data Analytics
Hybrid
World Trade Center, Yeshwanthpur, Bengaluru Urban, Karnataka, IND
13000 Employees

Ericsson Logo Ericsson

Unix Manager

Cloud • Information Technology • Internet of Things • Machine Learning • Software • Cybersecurity • Infrastructure as a Service (IaaS)
In-Office
Bangalore, Bengaluru Urban, Karnataka, IND
88000 Employees

Ericsson Logo Ericsson

Linux Administrator

Cloud • Information Technology • Internet of Things • Machine Learning • Software • Cybersecurity • Infrastructure as a Service (IaaS)
Hybrid
2 Locations
88000 Employees

Similar Companies Hiring

Golden Pet Brands Thumbnail
Digital Media • eCommerce • Information Technology • Marketing Tech • Pet • Retail • Social Media
El Segundo, California
178 Employees
Kepler  Thumbnail
Artificial Intelligence • Fintech • Software
New York, New York
9 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account