NextRaise Logo
JobsTrackerResumes
Job MatchATS
ResumesJobsProfile
Jobs / Platform Engineer in United States of America
1 hour agoBe an early applicant
Apply with autofill
Apply with autofill
Micron Technology·Semiconductors·1 hour ago
1 hour agoBe an early applicant

Intern - AI Systems and Infrastructure Engineering

Austin, United States of AmericaInternshipEntry · 0-1 years

Sign up free to see how well your resume matches this role.

About this role

Our vision is to transform how the world uses information to enrich life for all.

Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever.

The Advanced Systems Research and Engineering team develops innovative hardware and software technologies that enable the next generation of Artificial Intelligence infrastructure. The team collaborates closely with engineering, architecture, product, and research organizations to evaluate emerging AI workloads and drive advancements in memory, storage, interconnects, and distributed computing platforms. Through systems research, prototyping, and performance analysis, the team helps shape future technology roadmaps and industry-leading solutions.

The AI Systems Software Engineering Intern will work alongside senior engineers and researchers on advanced systems software for Large Language Models (LLMs) and Agentic AI applications. This role focuses on characterizing and improving the performance, scalability, and efficiency of AI inference and training workloads across GPU platforms and heterogeneous memory, interconnect, and storage systems. The intern will contribute to profiling, workload characterization, systems optimization, and experimental evaluation, helping drive innovations in AI infrastructure and memory technologies.

Responsibilities

  • Develop and enhance systems software, profiling tools, and experimentation frameworks for LLM training, LLM inference, and Agentic AI workloads.
  • Design, implement, and evaluate memory- and state-management techniques, including caching, tiering, compression, eviction, and lifecycle management for AI serving environments.
  • Characterize and optimize AI workload execution across GPUs, CPUs, memory subsystems, storage, and distributed infrastructure, with a focus on latency, throughput, scalability, and resource utilization.
  • Build benchmarking, simulation, and automation capabilities to evaluate data placement, migration, scheduling, and performance behavior across heterogeneous memory systems.
  • Collaborate with engineering and research teams to develop representative AI workloads, analyze experimental results, and contribute to technical publications, intellectual property, and future platform designs.

Minimum Qualifications

  • Currently pursuing a Master's or Ph.D. in Computer Science, Computer Engineering, Electrical Engineering, or a related technical field.
  • Demonstrated experience with AI systems, machine learning systems, computer systems research, or systems software development through coursework, research, or projects.
  • Understanding of Large Language Models (LLMs), including transformer execution, attention mechanisms, KV cache behavior, batching, token-level latency, throughput, and memory performance considerations.
  • Proficiency in Python and C/C++, with hands-on experience developing, debugging, and optimizing software in Linux environments.
  • Experience using GPU-based performance analysis tools and at least one modern AI framework or serving stack, such as PyTorch, vLLM, TensorRT-LLM, NVIDIA Dynamo, or related technologies.

Preferred Qualifications

  • Experience extending or optimizing LLM runtimes, serving engines, schedulers, or distributed inference frameworks.
  • Hands-on experience implementing advanced KV-cache, memory management, or state-management techniques for long-context or stateful AI applications.
  • Experience with GPU optimization technologies such as CUDA, Triton, NCCL, RDMA, or similar accelerator and communication frameworks.
  • Familiarity with heterogeneous memory architectures, including HBM, DRAM, CXL-attached memory, NVMe storage, pooled memory, or disaggregated memory systems.
  • Evidence of significant technical impact through publications, patents, open-source contributions, or substantial research and engineering projects related to Artificial Intelligence, distributed systems, memory systems, or high-performance computing.

    As a world leader in the semiconductor industry, Micron is dedicated to your personal wellbeing and professional growth. Micron benefits are designed to help you stay well, provide peace of mind and help you prepare for the future.  We offer a choice of medical, dental and vision plans in all locations enabling team members to select the plans that best meet their family healthcare needs and budget.  Micron also provides benefit programs that help protect your income if you are unable to work due to illness or injury, and paid family leave.  Additionally, Micron benefits include a robust paid time-off program and paid holidays.  For additional information regarding the Benefit programs available, please see the Benefits Guide posted on micron.com/careers/benefits.

    Micron is proud to be an equal opportunity workplace and is an affirmative action employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, age, national origin, citizenship status, disability, protected veteran status, gender identity or any other factor protected by applicable federal, state, or local laws.

    To learn about your right to work click here.

    To learn more about Micron, please visit micron.com/careers

    For US Sites Only: To request assistance with the application process and/or for reasonable accommodations, please contact Micron’s People Organization at  hrsupport_na@micron.com or 1-800-336-8918 (select option #3)

    Micron Prohibits the use of child labor and complies with all applicable laws, rules, regulations, and other international and industry labor standards.

    Micron does not charge candidates any recruitment fees or unlawfully collect any other payment from candidates as consideration for their employment with Micron.

    AI alert: Candidates are encouraged to use AI tools to enhance their resume and/or application materials. However, all information provided must be accurate and reflect the candidate's true skills and experiences. Misuse of AI to fabricate or misrepresent qualifications will result in immediate disqualification.   

    Fraud alert: Micron advises job seekers to be cautious of unsolicited job offers and to verify the authenticity of any communication claiming to be from Micron by checking the official Micron careers website in the About Micron Technology, Inc.

    SemiconductorsH1B sponsor likely
    AI tools
    Apply faster with autofillThe NextRaise extension autofills your application in one click.Get the extension

    Similar jobs

    • Platform Engineer at equifaxAlpharetta, United States of America
    • Kafka Platform Engineer at aepColumbus, United States of America
    • Data Platform Infrastructure Manager at sailpointUnited States
    • Senior Platform Engineer at leidosTeleworker US, United States of America
    • IT Infrastructure Manager at crosscountrymortgageRemote USA
    • OpenStack Platform Engineer at CiscoRTP, United States of America