Software Engineer Intern (US)

Reposted One Month Ago
Palo Alto, CA, USA
In-Office
84K-96K Annually
Internship
Artificial Intelligence • Machine Learning • Software
The Role
Work with engineers to design, implement, and deploy scalable inference solutions for top AI models. Implement and optimize models (Python, C++, CUDA, NCCL), monitor live services, develop features, fix bugs, participate in code reviews, and collaborate in agile teams.
Summary Generated by Built In
DeepInfra is seeking talented and motivated Software Engineering Interns to join our team. As an intern, you will be working closely with our experienced engineering team to design, develop, and deploy the top open AI models at scale. This is an excellent opportunity to gain hands-on experience in building scalable and efficient software systems, while working on cutting-edge AI models and algorithms.

What You’ll Do

  • Collaborate with the engineering team to design, develop, and test inference solutions for the top AI models.
  • Implement and optimize AI models using Python, C++, CUDA, NCCL
  • Monitor and maintain the live service.
  • Work on feature development, bug fixing, and code reviews to ensure high-quality software delivery
  • Participate in daily stand-ups, code reviews, and design discussions to ensure seamless collaboration
  • Stay up-to-date with industry trends and advancements in AI and machine learning
  • Try new things
  • Ship stuff

What You Bring

  • Currently pursuing a Bachelor's or Master's degree in Computer Science, Computer Engineering, or a related field
  • Strong fundamental knowledge in computer science, including data structures, algorithms, and software design patterns
  • Proficiency in Python, including experience with AI/ML libraries and frameworks (e.g., NumPy, pandas, SciPy, TensorFlow, PyTorch)
  • Familiarity with AI models, Transformers and Diffusers
  • Experience with version control systems (e.g., Git) and agile development methodologies
  • Excellent problem-solving skills, with the ability to debug and optimize code
  • Strong communication and teamwork skills, with the ability to effectively collaborate with cross-functional teams

Why DeepInfra
  • Work on cutting-edge AI model serving - the systems that power the next generation of LLMs and multimodal models.
  • Small team, huge impact: your work ships directly to customers.
  • Opportunity to learn from engineers building high-performance inference at scale.
  • Fast-paced environment with ownership, autonomy, and end-to-end responsibility.

How we work

Three traits define the people who thrive here, and this role leans on all three.

Initiative. We take ownership and step in where we can add value. Whether it’s starting something new, improving what exists, or helping move ideas forward, we aim to be proactive and thoughtful in how we contribute.

Drive. We’re energized by hard problems. Building AI infrastructure is complex, and we lean into that. We care about doing things well, moving fast, and continuously improving — because solving meaningful challenges is what motivates us.

Grit. Things don’t always work on the first try — and that’s expected. We stay persistent, adapt quickly, and learn as we go. We take setbacks seriously, but not personally, and use them to get better.

Compensation
Monthly range: 7000-8000/month USD

Skills Required

  • Currently pursuing a Bachelor's or Master's degree in Computer Science, Computer Engineering, or related field.
  • Strong fundamental knowledge in computer science, including data structures, algorithms, and software design patterns.
  • Proficiency in Python, including experience with AI/ML libraries and frameworks (NumPy, pandas, SciPy, TensorFlow, PyTorch).
  • Familiarity with AI models, Transformers and Diffusers.
  • Experience with version control systems (e.g., Git) and agile development methodologies.
  • Experience with C++, CUDA, and NCCL for implementing and optimizing AI models.
  • Excellent problem-solving skills, with the ability to debug and optimize code.
  • Strong communication and teamwork skills, with ability to collaborate across teams.
Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Palo Alto, California
20 Employees
Year Founded: 2022

What We Do

Let Deep Infra run your ML infrastructure. Just use our top AI models using a simple API or deploy your own model with us.

Similar Jobs

TetraMem Logo TetraMem

US 2026 Software - Compiler Engineer Intern

Artificial Intelligence • Hardware • Information Technology • Software
In-Office
San Jose, CA, USA
70 Employees
35-45 Hourly

Circle Logo Circle

Lead Product Designer

Blockchain • Fintech • Payments • Financial Services • Cryptocurrency • Web3
In-Office or Remote
San Francisco, CA, USA
1050 Employees
173K-225K Annually

Dynatrace Logo Dynatrace

Senior Mainframe Developer - HLASM

Artificial Intelligence • Big Data • Cloud • Information Technology • Software • Big Data Analytics • Automation
Remote or Hybrid
United States
5600 Employees
161K-241K Annually

Expedia Group Logo Expedia Group

Senior Creative Strategist

AdTech • eCommerce • Information Technology • Software • Travel • Generative AI
Hybrid
Beverly Hills, CA, USA
16000 Employees
155K-248K Annually

Similar Companies Hiring

Onshore Thumbnail
Artificial Intelligence • Fintech • Software • Financial Services
New York, New York
60 Employees
Revel.io Thumbnail
Aerospace • Hardware • Robotics • Software
US
50 Employees
Blee Thumbnail
Artificial Intelligence • Marketing Tech • Software
New York, New York
30 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account