Staff Software Engineer , Anywhere Cloud - AI Systems & Runtimes

Reposted 12 Days Ago
San Jose, CA, USA
In-Office
184K-230K Annually
Mid level
Artificial Intelligence • Cloud • Software • Big Data Analytics
Shape the Global Future of Enterprise AI Come build the future of enterprise AI with a global team that puts its people
The Role
The Staff Software Engineer will lead the development of cloud-native AI platforms, optimizing AI workload deployment in Kubernetes environments and collaborating with cross-functional teams.
Summary Generated by Built In

Business Area:

Engineering

Seniority Level:

Mid-Senior level

Job Description: 

At Cloudera, we empower people to transform complex data into clear and actionable insights. With as much data under management as the hyperscalers, we're the preferred data partner for the top companies in almost every industry.  Powered by the relentless innovation of the open source community, Cloudera advances digital transformation for the world’s largest enterprises.

Ready to take cloud innovation to the next level? Join Cloudera’s Anywhere Cloud team and help deliver a true “build your own pipeline, bring your own engine” experience — enabling data and AI workloads to run anywhere, without friction or vendor lock-in.

We bring the best of public cloud — cost efficiency, scalability, elasticity, and agility — to wherever data lives: public clouds, private data centers, and the edge. Powered by Kubernetes, our hybrid architecture separates compute and storage to maximize flexibility and optimize infrastructure usage.

This isn’t just cloud management — it’s about building a consistent, secure, and compliant cloud experience that gives organizations full access to all their data, anywhere.

With the acquisition of Taikun, we’re simplifying Kubernetes and cloud management even further, creating a unified, scalable, future-ready platform. If you’re passionate about Kubernetes — not just using it, but building it at the core, managing workloads across hybrid clouds and data centers, and obsessing over performance and DevOps — this is where you belong.

We are seeking a Staff Software Engineer to lead the architecture and delivery of our cloud‑native AI platform. In this high‑impact role, you will bridge the gap between cutting‑edge AI research and production‑grade Kubernetes environments. You will build the “nervous system” of our AI stack—optimizing how we run and manage open‑source models (Llama, Qwen, etc.) using K8s‑native patterns like Custom Resources (CRDs) and Operators, enabling agentic AI to thrive, and designing integration patterns that let our product teams and customers consume AI capabilities seamlessly.

As a Staff Software Engineer, you will:

  • Enterprise AI Services: Design and implement elegant, scalable application services (Go/Node.js) that wrap AI capabilities for enterprise use.

  • K8s-Native AI Orchestration: Lead the deployment of inference servers (vLLM, Triton) using KServe, KubeRay, or Knative to ensure serverless-style scaling for AI workloads.

  • Developer Velocity: Build internal tooling, SDKs, and "AI Gateways" that enhance team agility and simplify the integration of Foundation Models (Llama, GPT) into product features.

  • RAG & Prompt Engineering: Architect robust Retrieval-Augmented Generation (RAG) pipelines and prompt management services that integrate seamlessly with vector databases and enterprise data sources.

  • Cross-Functional Collaboration: Partner with UI engineers, UX designers, and Product Management to ensure the AI platform is not just powerful, but highly usable for internal developers.

  • Infrastructure & Security: Ensure AI workloads are secure, multi-tenant, and optimized for GPU resource scheduling (MIG, fractional GPUs) within Kubernetes.

We’re excited about you if you have:

  • Bachelor’s degree with 6+ years of software engineering experience (or equivalent Masters/PhD tenure), with at least 2+ years focused on AI/ML systems.

  • Expert proficiency in Python (for AI ecosystem) and strong competence in a systems language like Go or Rust/C++ (for high-performance serving layers).

  • Deep understanding of LLM deployment challenges and runtimes (e.g., vLLM, ONNX, TorchServe, Triton). Familiarity with quantization techniques (AWQ, GPTQ) to optimize model size/speed.

  • Experience building complex workflows using tools like LangChain or LlamaIndex, and deploying them on containerized infrastructure (Docker/Kubernetes).

  • Ability to navigate the rapidly changing AI landscape, filtering hype from practical engineering solutions, and driving technical alignment across teams.

You May Also Have: 

  • Model Fine-Tuning: Experience with efficient fine-tuning techniques (PEFT, LoRA/QLoRA) on custom datasets.

  • GPU Optimization: Familiarity with CUDA programming or profiling GPU performance (Nsight systems).

  • Open Source: Contributions to open-source AI projects (HuggingFace transformers, vLLM, etc.).

Why this role matters: 

This is more than cloud management, it’s about building the foundation for a consistent, secure, and compliant cloud experience that gives organizations 100% access to 100% of their data, anywhere.

With the recent acquisition of Taikun, we are simplifying Kubernetes and cloud management even further, creating a platform that is unified, scalable, and future-ready.

If you are passionate about Kubernetes, not just using it but building it at the core managing workloads across hybrid clouds and datacenters and obsessed with performance, devops, etc. this is where you belong.

This role is not eligible for immigration sponsorship.

The anticipated annual base salary range for this position is:

  • California: $184,000- $230,000

Individual compensation within the published range is determined by the candidate's skills, experience, qualifications, and primary work location. In addition to base pay, sales roles are eligible for Cloudera's commission plan, while non-sales roles are eligible for the corporate incentive plan. All employees receive a comprehensive benefits package.

What you can expect from us:

  • Generous PTO Policy 

  • Support work life balance with Unplugged Days

  • Flexible WFH Policy 

  • Mental & Physical Wellness programs 

  • Phone and Internet Reimbursement program 

  • Access to Continued Career Development 

  • Comprehensive Benefits and Competitive Packages 

  • Paid Volunteer Time

  • Employee Resource Groups

EEO/VEVRAA

#LI-BV1

#LI-HYBRID

Skills Required

  • Bachelor's degree
  • 6+ years of software engineering experience
  • 2+ years focused on AI/ML systems
  • Expert proficiency in Python
  • Strong competence in Go or Rust/C++
  • Deep understanding of LLM deployment challenges
  • Experience deploying on containerized infrastructure

Cloudera Compensation & Benefits Highlights

The following summarizes recurring compensation and benefits themes identified from responses generated by popular LLMs to common candidate questions about Cloudera and has not been reviewed or approved by Cloudera.

  • Fair & Transparent Compensation Pay practices are presented as equity-audited with a recognized Fair Pay Workplace certification and ongoing internal reviews. Compensation is often characterized as competitive for similar-sized peers, with visible market-aligned ranges for key roles.
  • Healthcare Strength Benefit descriptions emphasize comprehensive medical, dental, and vision coverage, alongside life and disability insurance, an EAP, wellness programming, and U.S. gym reimbursement. Health coverage is described as strong in practice.
  • Leave & Time Off Breadth Policies include generous PTO and holidays plus recurring companywide Unplugged Days that create extended weekends. Parental and medical leave are also highlighted as part of the core package.

Cloudera Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Santa Clara, CA
3,092 Employees
Year Founded: 2008

What We Do

At Cloudera, we empower people to transform complex data into clear and actionable insights. With as much data under management as the hyperscalers, we're the preferred data partner for the top companies in almost every industry. Powered by the relentless innovation of the open source community, Cloudera advances digital transformation for the world’s largest enterprises.

Why Work With Us

Impact at Scale: The infrastructure we build solves massive problems for top global banks, telecommunications giants, and healthcare providers. Cutting-Edge Tech: Work directly at the intersection of Open Source, Machine Learning, and Generative AI. Unmatched Flexibility: Enjoy a remote-friendly, hybrid culture that respects your time—including

Similar Jobs

CoreWeave Logo CoreWeave

Senior Manager, Technical Accounting

Cloud • Information Technology • Machine Learning
In-Office
4 Locations
1450 Employees
149K-198K Annually

Samsara Logo Samsara

Integration Engineer

Artificial Intelligence • Cloud • Computer Vision • Hardware • Internet of Things • Software
Easy Apply
Remote or Hybrid
United States
4000 Employees
106K-160K Annually

Tapestry - Coach and Kate Spade Logo Tapestry - Coach and Kate Spade

Store Manager

eCommerce • Fashion • Retail • Sales • Wearables • Design
Remote or Hybrid
14 Locations
16000 Employees
62K-94K Annually

Airwallex Logo Airwallex

Associate Director, Strategy & Operations - CEO Office

Artificial Intelligence • Fintech • Payments • Business Intelligence • Financial Services • Generative AI
Hybrid
San Francisco, CA, USA
2300 Employees
170K-300K Annually

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account