Intern - AI Systems and Infrastructure Engineering

Posted An Hour Ago
Be an Early Applicant
Austin, TX, USA
In-Office
Internship
Artificial Intelligence • Hardware • Information Technology • Machine Learning
Your talent powers our future.
The Role
Develop systems software, profiling tools, benchmarks, and experimentation frameworks for LLM training, inference, and agentic AI workloads. Optimize performance across GPUs, CPUs, memory, storage, and distributed systems, including caching, scheduling, data placement, and state management. Analyze experimental results and collaborate with research and engineering teams on AI infrastructure innovations, technical publications, intellectual property, and future platform designs.
Summary Generated by Built In
Our vision is to transform how the world uses information to enrich life for all .
Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever.
The Advanced Systems Research and Engineering team develops innovative hardware and software technologies that enable the next generation of Artificial Intelligence infrastructure. The team collaborates closely with engineering, architecture, product, and research organizations to evaluate emerging AI workloads and drive advancements in memory, storage, interconnects, and distributed computing platforms. Through systems research, prototyping, and performance analysis, the team helps shape future technology roadmaps and industry-leading solutions.
The AI Systems Software Engineering Intern will work alongside senior engineers and researchers on advanced systems software for Large Language Models (LLMs) and Agentic AI applications. This role focuses on characterizing and improving the performance, scalability, and efficiency of AI inference and training workloads across GPU platforms and heterogeneous memory, interconnect, and storage systems. The intern will contribute to profiling, workload characterization, systems optimization, and experimental evaluation, helping drive innovations in AI infrastructure and memory technologies.
Responsibilities
  • Develop and enhance systems software, profiling tools, and experimentation frameworks for LLM training, LLM inference, and Agentic AI workloads.
  • Design, implement, and evaluate memory- and state-management techniques, including caching, tiering, compression, eviction, and lifecycle management for AI serving environments.
  • Characterize and optimize AI workload execution across GPUs, CPUs, memory subsystems, storage, and distributed infrastructure, with a focus on latency, throughput, scalability, and resource utilization.
  • Build benchmarking, simulation, and automation capabilities to evaluate data placement, migration, scheduling, and performance behavior across heterogeneous memory systems.
  • Collaborate with engineering and research teams to develop representative AI workloads, analyze experimental results, and contribute to technical publications, intellectual property, and future platform designs.

Minimum Qualifications
  • Currently pursuing a Master's or Ph.D. in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field.
  • Demonstrated experience with AI systems, machine learning systems, computer systems research, or systems software development through coursework, research, or projects.
  • Understanding of Large Language Models (LLMs), including transformer execution, attention mechanisms, KV cache behavior, batching, token-level latency, throughput, and memory performance considerations.
  • Proficiency in Python and C/C++, with hands-on experience developing, debugging, and optimizing software in Linux environments.
  • Experience using GPU-based performance analysis tools and at least one modern AI framework or serving stack, such as PyTorch, vLLM, TensorRT-LLM, NVIDIA Dynamo, or related technologies.

Preferred Qualifications
  • Experience extending or optimizing LLM runtimes, serving engines, schedulers, or distributed inference frameworks.
  • Hands-on experience implementing advanced KV-cache, memory management, or state-management techniques for long-context or stateful AI applications.
  • Experience with GPU optimization technologies such as CUDA, Triton, NCCL, RDMA, or similar accelerator and communication frameworks.
  • Familiarity with heterogeneous memory architectures, including HBM, DRAM, CXL-attached memory, NVMe storage, pooled memory, or disaggregated memory systems.
  • Evidence of significant technical impact through publications, patents, open-source contributions, or substantial research and engineering projects related to Artificial Intelligence, distributed systems, memory systems, or high-performance computing.

As a world leader in the semiconductor industry, Micron is dedicated to your personal wellbeing and professional growth. Micron benefits are designed to help you stay well, provide peace of mind and help you prepare for the future. We offer a choice of medical, dental and vision plans in all locations enabling team members to select the plans that best meet their family healthcare needs and budget. Micron also provides benefit programs that help protect your income if you are unable to work due to illness or injury, and paid family leave. Additionally, Micron benefits include a robust paid time-off program and paid holidays. For additional information regarding the Benefit programs available, please see the Benefits Guide posted on micron.com/careers/benefits .
Micron is proud to be an equal opportunity workplace and is an affirmative action employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, age, national origin, citizenship status, disability, protected veteran status, gender identity or any other factor protected by applicable federal, state, or local laws.
To learn about your right to work click here.
To learn more about Micron, please visit micron.com/careers
For US Sites Only: To request assistance with the application process and/or for reasonable accommodations, please contact Micron's People Organization at [email protected] or 1-800-336-8918 (select option #3)
Micron Prohibits the use of child labor and complies with all applicable laws, rules, regulations, and other international and industry labor standards.
Micron does not charge candidates any recruitment fees or unlawfully collect any other payment from candidates as consideration for their employment with Micron.
AI alert: Candidates are encouraged to use AI tools to enhance their resume and/or application materials. However, all information provided must be accurate and reflect the candidate's true skills and experiences. Misuse of AI to fabricate or misrepresent qualifications will result in immediate disqualification.
Fraud alert: Micron advises job seekers to be cautious of unsolicited job offers and to verify the authenticity of any communication claiming to be from Micron by checking the official Micron careers website in the About Micron Technology, Inc.

Skills Required

  • Currently pursuing a Master's or Ph.D. in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field.
  • Demonstrated experience with AI systems, machine learning systems, computer systems research, or systems software development through coursework, research, or projects.
  • Understanding of Large Language Models, including transformer execution, attention mechanisms, KV cache behavior, batching, token-level latency, throughput, and memory performance.
  • Proficiency in Python and C/C++, with hands-on experience developing, debugging, and optimizing software in Linux environments.
  • Experience using GPU-based performance analysis tools and at least one modern AI framework or serving stack, such as PyTorch, vLLM, TensorRT-LLM, or NVIDIA Dynamo.
  • Experience extending or optimizing LLM runtimes, serving engines, schedulers, or distributed inference frameworks.
  • Hands-on experience implementing advanced KV-cache, memory management, or state-management techniques for long-context or stateful AI applications.
  • Experience with GPU optimization technologies such as CUDA, Triton, NCCL, RDMA, or similar accelerator and communication frameworks.
  • Familiarity with heterogeneous memory architectures, including HBM, DRAM, CXL-attached memory, NVMe storage, pooled memory, or disaggregated memory systems.
  • Evidence of significant technical impact through publications, patents, open-source contributions, or substantial research and engineering projects related to AI, distributed systems, memory systems, or high-performance computing.

Micron Technology Compensation & Benefits Highlights

  • Retirement Support A company 401(k) match with multiple savings options and investment choices supports long‑term financial security. This element stands out as a core strength within the overall package.
  • Equity Value & Accessibility An employee stock purchase plan offered at a meaningful discount and management‑granted RSUs provide accessible ownership opportunities that can bolster total compensation. These programs create upside beyond base pay.
  • Healthcare Strength Multiple medical plan choices with dental, vision, and company‑paid disability and life coverage, plus an EAP and some onsite/near‑site health centers, indicate robust core protection. The breadth of health options is consistently emphasized across the offering.

Micron Technology Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Boise, ID
45,000 Employees
Year Founded: 1978

What We Do

We are a world leader in innovative memory solutions that transform how the world uses information to enrich life for all. For over 45 years, our company has been instrumental to the world’s most significant technology advancements, delivering optimal memory and storage systems for a broad range of applications.

Why Work With Us

Global opportunities, team member development, and career advancement—Micron invests in you and celebrates your skills, a growth mindset, and the tenacity to strive. At Micron, everyone innovates.

Gallery

Gallery
Gallery
Gallery
Gallery
Gallery

Micron Technology Offices

Hybrid Workspace

Employees engage in a combination of remote and on-site work.

Micron recognizes the importance of maintaining a healthy work-life balance to foster a culture of collaboration, innovation and meet the needs of the business. In alignment with these values, we offer four flexible work arrangement options

Typical time on-site: Flexible
HQBoise, ID
Bengaluru, Karnataka
Folsom, CA
Hyderabad, Telangana
Longmont, CO
Madhapur, Telangana
Manassas, VA
Milpitas, CA
Singapore, SG
Singapore, SG
Sydney, AU
Ulsoor, Bengaluru
Learn more

Similar Jobs

Micron Technology Logo Micron Technology

Staff Engineer

Artificial Intelligence • Hardware • Information Technology • Machine Learning
In-Office
Richardson, TX, USA
45000 Employees
50K-75K Annually

Micron Technology Logo Micron Technology

Intern - Design Architecture, HBM

Artificial Intelligence • Hardware • Information Technology • Machine Learning
In-Office
Richardson, TX, USA
45000 Employees

Micron Technology Logo Micron Technology

SoC Physical Verification Engineer, HBM

Artificial Intelligence • Hardware • Information Technology • Machine Learning
In-Office
2 Locations
45000 Employees
97K-205K Annually

Micron Technology Logo Micron Technology

Design Engineer

Artificial Intelligence • Hardware • Information Technology • Machine Learning
In-Office
2 Locations
45000 Employees
146K-297K Annually

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account