CareerPlanGet AI match score →

Intern, AI Engineering

San Francisco, California, United States💼 Internship🗓 2026-07-08 → 2026-07-31

Core

Design and implement intelligent agent architectures for complex enterprise automation tasks, including multi-agent collaboration, MCP, and reasoning frameworks.

Role type

Research Intern, LLM-based agentic systems

Builds

AI systems serving millions of users across global enterprises

Domain

Enterprise AI, Large Language Models, Agentic Systems

Deliverable

research

Required skills

Python, PyTorch, LLM agent architecture design, efficient LLM fine-tuning, high-performance LLM inference optimization

Preferred skills

CUDA programming, custom kernel development, reinforcement learning, production inference systems (vLLM, TensorRT-LLM, SGLang), open-source ML infrastructure contributions

Technologies

PyTorch, vLLM, TensorRT-LLM, SGLang, Hugging Face

Responsibilities

Conduct original research on LLM agent architectures and optimization techniques; Develop and evaluate novel algorithms with both academic rigor and production feasibility; Present work at internal research seminars and external conferences; Mentor and collaborate with LLM engineers on implementation and deployment

Seniority

Intern, graduate student level

Sourced via greenhouse · Listed on CareerPlan, which tracks 70,000+ jobs from 20+ sources.
Apply on Greenhouse ↗