Software Engineer Intern (US)

Deep Infra Inc.

Palo Alto (CA)

On-site

USD 937,000 - 1,071,000

Full time

14 days+
Application generator

Don’t send a generic resume — generate a resume and cover letter tailored to this exact role.

Get past ATS filters

Job summary

DeepInfra is seeking software engineering interns to join our team in Palo Alto, contributing to the design, development, and deployment of open AI models at scale. You will work closely with experienced engineers to build scalable systems and gain hands-on experience with cutting-edge AI models.

The internship emphasizes Python, C++, CUDA, NCCL, and live service maintenance, with participation in code reviews, feature work, and collaboration across teams.

Qualifications

  • Pursuing a Bachelor's or Master's in CS, CE, or related field.
  • Strong CS fundamentals: data structures and algorithms.
  • Proficiency in Python; experience with AI/ML libraries (NumPy, pandas, SciPy, TensorFlow, PyTorch).
  • Familiarity with AI models, Transformers and Diffusers.
  • Experience with Git and agile development practices.
  • Excellent problem-solving and collaboration skills.

Responsibilities

  • Collaborate with the engineering team to design, develop, and test inference solutions for AI models.
  • Implement and optimize AI models using Python, C++, CUDA, NCCL.
  • Monitor and maintain the live service.
  • Work on feature development, bug fixing, and code reviews to ensure high-quality software delivery
  • Participate in daily stand-ups, code reviews, and design discussions to ensure seamless collaboration
  • Stay up-to-date with industry trends and advancements in AI and machine learning
  • Try new things
  • Ship stuff

Skills

Python
C++
Algorithms
Data structures
Git
Agile
Communication
Teamwork

Education

Bachelor's or Master's in CS/CE or related

Tools

CUDA
NCCL
NumPy
Pandas
SciPy
TensorFlow
PyTorch

Job description

DeepInfra is seeking talented and motivated Software Engineering Interns to join our team. As an intern, you will be working closely with our experienced engineering team to design, develop, and deploy the top open AI models at scale. This is an excellent opportunity to gain hands‑on experience in building scalable and efficient software systems, while working on cutting‑edge AI models and algorithms.

What You’ll Do
  • Collaborate with the engineering team to design, develop, and test inference solutions for the top AI models.
  • Implement and optimize AI models using Python, C++, CUDA, NCCL
  • Monitor and maintain the live service.
  • Work on feature development, bug fixing, and code reviews to ensure high-quality software delivery
  • Participate in daily stand‑ups, code reviews, and design discussions to ensure seamless collaboration
  • Stay up-to-date with industry trends and advancements in AI and machine learning
  • Try new things
  • Ship stuff
What You Bring
  • Currently pursuing a Bachelor's or Master's degree in Computer Science, Computer Engineering, or a related field
  • Strong fundamental knowledge in computer science, including data structures, algorithms, and software design patterns
  • Proficiency in Python, including experience with AI/ML libraries and frameworks (e.g., NumPy, pandas, SciPy, TensorFlow, PyTorch)
  • Familiarity with AI models, Transformers and Diffusers
  • Experience with version control systems (e.g., Git) and agile development methodologies
  • Excellent problem-solving skills, with the ability to debug and optimize code
  • Strong communication and teamwork skills, with the ability to effectively collaborate with cross‑functional teams
Why DeepInfra
  • Work on cutting‑edge AI model serving - the systems that power the next generation of LLMs and multimodal models.
  • Small team, huge impact: your work ships directly to customers.
  • Opportunity to learn from engineers building high-performance inference at scale.
  • Fast-paced environment with ownership, autonomy, and end‑to‑end responsibility.
How we work

Three traits define the people who thrive here, and this role leans on all three.

Initiative.We take ownership and step in where we can add value. Whether it’s starting something new, improving what exists, or helping move ideas forward, we aim to be proactive and thoughtful in how we contribute.

Drive.We’re energized by hard problems. Building AI infrastructure is complex, and we lean into that. We care about doing things well, moving fast, and continuously improving — because solving meaningful challenges is what motivates us.

Grit.Things don’t always work on the first try — and that’s expected. We stay persistent, adapt quickly, and learn as we go. We take setbacks seriously, but not personally, and use them to get better.

Compensation

Monthly range: 7000-8000/month USD

Get your free, confidential resume review.

or drag and drop your file here.

Similar jobs

Similar jobs worth comparing

Software Engineer Intern (US)
Software Engineer Intern (US)

DeepInfra • Palo Alto (CA)

On-site
USD 78,120 - 89,280
Software Engineer Intern (Europe)
Software Engineer Intern (Europe)

DeepInfra • United States

On-site
USD 40,000 - 60,000
Software Engineer, Early Career
Software Engineer, Early Career

DeepInfra • Palo Alto (CA)

On-site
USD 140,000 - 150,000
Software Engineer, Early Career
Software Engineer, Early Career

Deep Infra Inc. • Palo Alto (CA)

On-site
USD 140,000 - 150,000
Software Engineer
Software Engineer

Deep Infra Inc. • Palo Alto (CA)

On-site
USD 150,000 - 195,000
Open-source contributions
C++/CUDA experience
AI Infra Software Engineer Intern
AI Infra Software Engineer Intern

Deep Infra Inc. • Palo Alto (CA)

On-site
USD 937,000 - 1,071,000
AI Infrastructure Engineer Intern — Build at-scale AI models
AI Infrastructure Engineer Intern — Build at-scale AI models

DeepInfra • United States

Remote
USD 40,000 - 60,000
Forward Deployed Engineer
Forward Deployed Engineer

Deep Infra Inc. • Palo Alto (CA)

On-site
USD 150,000 - 195,000
AI Infrastructure Intern — Inference at Scale
AI Infrastructure Intern — Inference at Scale

DeepInfra • Palo Alto (CA)

On-site
USD 78,120 - 89,280
Software Engineer AI/ML Systems - USA Onsite (Santa Clara, CA)
Software Engineer AI/ML Systems - USA Onsite (Santa Clara, CA)

Dover • Santa Clara (CA), Northern (KY)

Hybrid
USD 150,000 - 170,000
Unlimited PTO
Generous parental leave
Stock Purchase Program (ESPP)
+3