Software Engineer

Reposted 13 Hours Ago
Palo Alto, CA, USA
In-Office
150K-195K Annually
Junior
Artificial Intelligence • Machine Learning • Software
The Role
Design and develop inference solutions for AI models, optimize performance, and maintain production systems while owning end-to-end problems.
Summary Generated by Built In
DeepInfra is looking for early-career Software Engineers to join our team. You’ll work on designing, building, and scaling infrastructure for serving top open-source AI models in production. This role is ideal for engineers who are already comfortable owning problems end-to-end and want to deepen their experience working on high-impact AI systems.

If you’re excited about AI/ML, have built and shipped projects, and are looking to work on real systems at scale — we’d love to meet you.

What You’ll Do

  • Design, develop, and test inference solutions for state-of-the-art AI models
  • Implement, optimize, and evaluate AI models using Python, C++, CUDA, and NCCL
  • Own and operate production model-serving systems, including monitoring and debugging
  • Build new features, improve system performance, and contribute to overall system design
  • Participate in code reviews and technical discussions to maintain high engineering standards
  • Explore and apply new AI/ML techniques to improve model performance and efficiency
  • Take ideas from concept to production

What You Bring

  • Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field
  • 1–4 years of relevant experience, including early full-time roles or research
  • Strong fundamentals in data structures, algorithms, and software design
  • Proficiency in Python and experience working with AI/ML frameworks (e.g., PyTorch, TensorFlow)
  • Hands-on experience building, shipping, and maintaining software systems
  • Familiarity with AI models, Transformers, and Diffusers
  • Experience working with version control (Git) and collaborative development workflows
  • Ability to debug, optimize, and improve existing systems
  • Strong communication skills and ability to work independently in a fast-paced environment

Bonus

  • Experience with C++, CUDA, or AI inference
  • Contributions to open-source ML projects

Why DeepInfra

  • Work on cutting-edge AI model serving - the systems that power the next generation of LLMs and multimodal models.
  • Small team, huge impact: your work ships directly to customers.
  • Opportunity to learn from engineers building high-performance inference at scale.
  • Fast-paced environment with ownership, autonomy, and end-to-end responsibility.

Annual base salary range
$150,000 - $195,000

Skills Required

  • Bachelor's or Master's degree in Computer Science, Computer Engineering, or a related field
  • 1 - 4 years of relevant experience
  • Proficiency in Python
  • Experience working with AI/ML frameworks (e.g., PyTorch, TensorFlow)
  • Strong fundamentals in data structures, algorithms, and software design
  • Hands-on experience building, shipping, and maintaining software systems
  • Familiarity with AI models, Transformers, and Diffusers
  • Experience with version control (Git)
  • Strong communication skills and ability to work independently in a fast-paced environment
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Palo Alto, California
20 Employees
Year Founded: 2022

What We Do

Let Deep Infra run your ML infrastructure. Just use our top AI models using a simple API or deploy your own model with us.

Similar Jobs

Boeing Logo Boeing

Software Engineer

Aerospace • Information Technology • Software • Cybersecurity • Design • Defense • Manufacturing
In-Office
Seal Beach, CA, USA
170000 Employees
182K-222K Annually

Boeing Logo Boeing

Software Engineer

Aerospace • Information Technology • Software • Cybersecurity • Design • Defense • Manufacturing
In-Office
El Segundo, CA, USA
170000 Employees
99K-162K Annually

Adyen Logo Adyen

Software Engineer

Fintech • Payments • Financial Services
Easy Apply
Hybrid
San Francisco, CA, USA
4771 Employees
198K-293K Annually

Genius Sports Logo Genius Sports

Software Engineer

AdTech • Artificial Intelligence • Machine Learning • Marketing Tech • Software • Sports • Big Data Analytics
Easy Apply
Hybrid
Los Angeles, CA, USA
1800 Employees
160K-180K Annually

Similar Companies Hiring

Hanover Park Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
42 Employees
Kepler  Thumbnail
Fintech • Software
New York, New York
6 Employees
Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account