Software Engineer Intern (US)

DeepInfra

Palo Alto (CA)

On-site

USD 78,120 - 89,280

Part time

14 days+

Get more replies from employers

Send a job-specific resume in minutes.

Job summary

DeepInfra seeks talented Software Engineering Interns to join its team in Palo Alto, offering hands-on experience designing, developing, and deploying scalable AI models. You will work closely with engineers to implement inference solutions and optimize AI systems using Python, C++, CUDA, and NCCL.

The role includes monitoring live services and participating in code reviews to ensure high-quality software delivery.

Qualifications

  • Currently pursuing a Bachelor's or Master's degree in Computer Science, Computer Engineering, or a related field.
  • Proficiency in Python and knowledge of AI/ML libraries is required.
  • Familiarity with AI models, Transformers and Diffusers is preferred.
  • Experience with Git and agile development methodologies is beneficial.
  • Strong problem-solving and collaboration skills are essential.

Responsibilities

  • Collaborate with the engineering team to design, develop, and test inference solutions for top AI models.
  • Implement and optimize AI models using Python, C++, CUDA, NCCL.
  • Monitor and maintain the live service.
  • Participate in code reviews, feature development, and bug fixes to ensure high-quality software delivery.
  • Engage in daily stand-ups and design discussions to enable seamless collaboration.
  • Stay updated with AI/ML industry trends and advancements.

Skills

Python
Algorithms
Design patterns
Teamwork

Education

Bachelor's or Master's in CS/CE or related field

Tools

Git
NumPy
pandas
SciPy
TensorFlow
PyTorch
Transformers
Diffusers
CUDA
NCCL

Job description

DeepInfra is seeking talented and motivated Software Engineering Interns to join our team. As an intern, you will be working closely with our experienced engineering team to design, develop, and deploy the top open AI models at scale. This is an excellent opportunity to gain hands‑on experience in building scalable and efficient software systems, while working on cutting‑edge AI models and algorithms.

What You’ll Do
  • Collaborate with the engineering team to design, develop, and test inference solutions for the top AI models.
  • Implement and optimize AI models using Python, C++, CUDA, NCCL
  • Monitor and maintain the live service.
  • Work on feature development, bug fixing, and code reviews to ensure high‑quality software delivery
  • Participate in daily stand‑ups, code reviews, and design discussions to ensure seamless collaboration
  • Stay up‑to‑date with industry trends and advancements in AI and machine learning
  • Try new things
  • Ship stuff
What You Bring
  • Currently pursuing a Bachelor's or Master's degree in Computer Science, Computer Engineering, or a related field
  • Strong fundamental knowledge in computer science, including data structures, algorithms, and software design patterns
  • Proficiency in Python, including experience with AI/ML libraries and frameworks (e.g., NumPy, pandas, SciPy, TensorFlow, PyTorch)
  • Familiarity with AI models, Transformers and Diffusers
  • Experience with version control systems (e.g., Git) and agile development methodologies
  • Excellent problem‑solving skills, with the ability to debug and optimize code
  • Strong communication and teamwork skills, with the ability to effectively collaborate with cross‑functional teams
Why DeepInfra
  • Work on cutting‑edge AI model serving - the systems that power the next generation of LLMs and multimodal models.
  • Small team, huge impact: your work ships directly to customers.
  • Opportunity to learn from engineers building high‑performance inference at scale.
  • Fast‑paced environment with ownership, autonomy, and end‑to‑end responsibility.
How we work

Three traits define the people who thrive here, and this role leans on all three.

Initiative. We take ownership and step in where we can add value. Whether it’s starting something new, improving what exists, or helping move ideas forward, we aim to be proactive and thoughtful in how we contribute.

Drive. We’re energized by hard problems. Building AI infrastructure is complex, and we lean into that. We care about doing things well, moving fast, and continuously improving — because solving meaningful challenges is what motivates us.

Grit. Things don’t always work on the first try — and that’s expected. We stay persistent, adapt quickly, and learn as we go. We take setbacks seriously, but not personally, and use them to get better.

Compensation

Monthly range: 7000-8000/month USD

Get your free, confidential resume review.
or drag and drop your file here.
Similar jobs

Similar jobs worth comparing

Software Engineer, Early Career
Software Engineer, Early Career

DeepInfra • Palo Alto (CA)

On-site
USD 140,000 - 150,000
Software Engineer
Software Engineer

DeepInfra • Palo Alto (CA)

On-site
USD 130,000 - 180,000
Campus AI Research Engineer - Deep Learning (Intern)
Campus AI Research Engineer - Deep Learning (Intern)

Jump Trading • Chicago (IL)

On-site
USD 255,000 - 345,000
AI Deep Learning Engineer
AI Deep Learning Engineer

EliteHosts • Norge (OK)

On-site
USD 15,177 - 20,534
AI Infrastructure Intern — Inference at Scale
AI Infrastructure Intern — Inference at Scale

DeepInfra • Palo Alto (CA)

On-site
AI Deep Learning Engineer
AI Deep Learning Engineer

RetailHub Sourcing • Town of Norway (WI)

On-site
USD 10,434 - 14,117
AI Deep Learning Engineer
AI Deep Learning Engineer

Q-Storage • Washington

On-site
USD 15,000 - 18,000
Visa sponsorship
Accommodation
Research Intern, Inference (Fall 2026)
Research Intern, Inference (Fall 2026)

Togetherai • San Francisco (CA)

On-site
USD 79,900 - 86,788
Competitive compensation
Housing stipends
Other competitive benefits
Research Engineer, Infrastructure, Inference
Research Engineer, Infrastructure, Inference

Thinkingmachines • San Francisco (CA)

On-site
USD 350,000 - 475,000
Generous health, dental, and vision benefits
Unlimited PTO
Paid parental leave
+1
System Software Engineer - AI
System Software Engineer - AI

Delos Data Inc • Palo Alto (CA)

Hybrid
USD 140,000 - 200,000
Equity
401k
Benefits