The Role
Software Engineering Intern developing, optimizing, testing, and deploying AI model inference solutions at scale. Responsibilities include implementing models with Python, C++, CUDA, and NCCL; monitoring live services; fixing bugs; reviewing code; and collaborating on design and feature development. The role offers hands-on experience building high-performance AI infrastructure for LLMs and multimodal models.
Summary Generated by Built In
DeepInfra is seeking a talented and motivated Software Engineering Intern to join our Sofia-based team. As an intern, you will be working closely with our experienced engineering team to design, develop, and deploy the top open AI models at scale. This is an excellent opportunity to gain hands-on experience in building scalable and efficient software systems, while working on cutting-edge AI models and algorithms.
What You’ll Do
- Collaborate with the engineering team to design, develop, and test inference solutions for the top AI models.
- Implement and optimize AI models using Python, C++, CUDA, NCCL
- Monitor and maintain the live service.
- Work on feature development, bug fixing, and code reviews to ensure high-quality software delivery
- Participate in daily stand-ups, code reviews, and design discussions to ensure seamless collaboration
- Stay up-to-date with industry trends and advancements in AI and machine learning
- Try new things
- Ship stuff
What You Bring
- Currently pursuing a Bachelor's or Master's degree in Computer Science, Computer Engineering, or a related field
- Strong fundamental knowledge in computer science, including data structures, algorithms, and software design patterns
- Proficiency in Python, including experience with AI/ML libraries and frameworks (e.g., NumPy, pandas, SciPy, TensorFlow, PyTorch)
- Familiarity with AI models, Transformers and Diffusers
- Experience with version control systems (e.g., Git) and agile development methodologies
- Excellent problem-solving skills, with the ability to debug and optimize code
- Strong communication and teamwork skills, with the ability to effectively collaborate with cross-functional teams
Why DeepInfra
- Work on cutting-edge AI model serving - the systems that power the next generation of LLMs and multimodal models.
- Small team, huge impact: your work ships directly to customers.
- Opportunity to learn from engineers building high-performance inference at scale.
- Fast-paced environment with ownership, autonomy, and end-to-end responsibility.
How we work
Three traits define the people who thrive here, and this role leans on all three.
Initiative. We take ownership and step in where we can add value. Whether it’s starting something new, improving what exists, or helping move ideas forward, we aim to be proactive and thoughtful in how we contribute.
Drive. We’re energized by hard problems. Building AI infrastructure is complex, and we lean into that. We care about doing things well, moving fast, and continuously improving — because solving meaningful challenges is what motivates us.
Grit. Things don’t always work on the first try — and that’s expected. We stay persistent, adapt quickly, and learn as we go. We take setbacks seriously, but not personally, and use them to get better.
Skills Required
- Currently pursuing a Bachelor's or Master's degree in Computer Science, Computer Engineering, or a related field
- Strong knowledge of computer science fundamentals, including data structures, algorithms, and software design patterns
- Proficiency in Python, including experience with AI/ML libraries and frameworks
- Familiarity with AI models, Transformers, and Diffusers
- Experience with version control systems such as Git and agile development methodologies
- Excellent problem-solving skills, including the ability to debug and optimize code
- Strong communication and teamwork skills
Am I A Good Fit?
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.
Success! Refresh the page to see how your skills align with this role.
The Company
What We Do
Let Deep Infra run your ML infrastructure. Just use our top AI models using a simple API or deploy your own model with us.
.png)








