Intern - AI Systems and Infrastructure Engineering at Micron Technology | Austin, TX, US | Rezi

Intern - AI Systems and Infrastructure Engineering at Micron Technology

Intern - AI Systems and Infrastructure Engineering

Micron Technology · Austin, TX, US

3 weeks ago

Intern - AI Systems and Infrastructure Engineering

Micron Technology · Austin, TX, US

23 days ago
Resume preview

Impress employers and recruiters.
Choose from hundreds of resume examples.

Target Resume Now
Resume preview

Tailor your resume to this Intern - AI Systems and Infrastructure Engineering role.

Rezi rewrites your resume against Micron Technology's job description. Free.

Resume score gauge reading 58 out of 100

Don't guess if your resume is good enough.

See how it scores against the Intern - AI Systems and Infrastructure Engineering posting at Micron Technology — free, in seconds.

About the Role

The AI Systems Software Engineering Intern will work alongside senior engineers and researchers on advanced systems software for Large Language Models (LLMs) and Agentic AI applications. This role focuses on characterizing and improving the performance, scalability, and efficiency of AI inference and training workloads across GPU platforms and heterogeneous memory, interconnect, and storage systems. The intern will contribute to profiling, workload characterization, systems optimization, and experimental evaluation, helping drive innovations in AI infrastructure and memory technologies.

Responsibilities

  • Develop and enhance systems software, profiling tools, and experimentation frameworks for LLM training, LLM inference, and Agentic AI workloads.
  • Design, implement, and evaluate memory- and state-management techniques, including caching, tiering, compression, eviction, and lifecycle management for AI serving environments.
  • Characterize and optimize AI workload execution across GPUs, CPUs, memory subsystems, storage, and distributed infrastructure, with a focus on latency, throughput, scalability, and resource utilization.
  • Build benchmarking, simulation, and automation capabilities to evaluate data placement, migration, scheduling, and performance behavior across heterogeneous memory systems.
  • Collaborate with engineering and research teams to develop representative AI workloads, analyze experimental results, and contribute to technical publications, intellectual property, and future platform designs.

Requirements

  • Currently pursuing a Master's or Ph.D. in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field.
  • Demonstrated experience with AI systems, machine learning systems, computer systems research, or systems software development through coursework, research, or projects.
  • Understanding of Large Language Models (LLMs), including transformer execution, attention mechanisms, KV cache behavior, batching, token-level latency, throughput, and memory performance considerations.
  • Proficiency in Python and C/C++, with hands-on experience developing, debugging, and optimizing software in Linux environments.
  • Experience using GPU-based performance analysis tools and at least one modern AI framework or serving stack, such as PyTorch, vLLM, TensorRT-LLM, NVIDIA Dynamo, or related technologies.
  • Experience extending or optimizing LLM runtimes, serving engines, schedulers, or distributed inference frameworks.
  • Hands-on experience implementing advanced KV-cache, memory management, or state-management techniques for long-context or stateful AI applications.
  • Experience with GPU optimization technologies such as CUDA, Triton, NCCL, RDMA, or similar accelerator and communication frameworks.
  • Familiarity with heterogeneous memory architectures, including HBM, DRAM, CXL-attached memory, NVMe storage, pooled memory, or disaggregated memory systems.
  • Evidence of significant technical impact through publications, patents, open-source contributions, or substantial research and engineering projects related to Artificial Intelligence, distributed systems, memory systems, or high-performance computing.

Skills

  • Python
  • C/C++
  • Linux
  • GPU-based performance analysis tools
  • PyTorch
  • vLLM
  • TensorRT-LLM
  • NVIDIA Dynamo
  • CUDA
  • Triton
  • NCCL
  • RDMA

Experience Level

  • Intern

Education Level

  • Master's or Ph.D.

Benefits

  • Choice of medical, dental and vision plans
  • Benefit programs that help protect your income if you are unable to work due to illness or injury
  • Paid family leave
  • Robust paid time-off program
  • Paid holidays

About the Company

  • Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever.
  • Micron is dedicated to your personal wellbeing and professional growth.
  • Micron Prohibits the use of child labor and complies with all applicable laws, rules, regulations, and other international and industry labor standards.
  • Micron does not charge candidates any recruitment fees or unlawfully collect any other payment from candidates as consideration for their employment with Micron.

Equal Opportunity

  • Micron is proud to be an equal opportunity workplace and is an affirmative action employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, age, national origin, citizenship status, disability, protected veteran status, gender identity or any other factor protected by applicable federal, state, or local laws.
  • To learn about your right to work click here.
  • For US Sites Only: To request assistance with the application process and/or for reasonable accommodations, please contact Micron’s People Organization at hrsupport_na@micron.com or 1-800-336-8918 (select option #3)