Padmi
DeepInfra logo
DeepInfra

AI inference APIs · LLM deployment

Software Engineer Intern (US)

San Francisco Bay Area · Onsite$7k–$8k/moPosted 7 days ago
Software engineeringInternInternship
Apply at DeepInfra

Opens the source posting on jobs.gem.com

Source description

About the role

View original

Company logo

View all jobs

Software Engineer Intern (US)

Palo Alto, CA, USA

Engineering

In office

Intern

DeepInfra is seeking talented and motivated Software Engineering Interns to join our team. As an intern, you will be working closely with our experienced engineering team to design, develop, and deploy the top open AI models at scale. This is an excellent opportunity to gain hands-on experience in building scalable and efficient software systems, while working on cutting-edge AI models and algorithms.

What You’ll Do

  • Collaborate with the engineering team to design, develop, and test inference solutions for the top AI models.
  • Implement and optimize AI models using Python, C++, CUDA, NCCL
  • Monitor and maintain the live service.
  • Work on feature development, bug fixing, and code reviews to ensure high-quality software delivery
  • Participate in daily stand-ups, code reviews, and design discussions to ensure seamless collaboration
  • Stay up-to-date with industry trends and advancements in AI and machine learning
  • Try new things
  • Ship stuff

What You Bring

  • Currently pursuing a Bachelor's or Master's degree in Computer Science, Computer Engineering, or a related field
  • Strong fundamental knowledge in computer science, including data structures, algorithms, and software design patterns
  • Proficiency in Python, including experience with AI/ML libraries and frameworks (e.g., NumPy, pandas, SciPy, TensorFlow, PyTorch)
  • Familiarity with AI models, Transformers and Diffusers
  • Experience with version control systems (e.g., Git) and agile development methodologies
  • Excellent problem-solving skills, with the ability to debug and optimize code
  • Strong communication and teamwork skills, with the ability to effectively collaborate with cross-functional teams

Why DeepInfra

  • Work on cutting-edge AI model serving - the systems that power the next generation of LLMs and multimodal models.
  • Small team, huge impact: your work ships directly to customers.
  • Opportunity to learn from engineers building high-performance inference at scale.
  • Fast-paced environment with ownership, autonomy, and end-to-end responsibility.

How we work

Three traits define the people who thrive here, and this role leans on all three.

Initiative. We take ownership and step in where we can add value. Whether it’s starting something new, improving what exists, or helping move ideas forward, we aim to be proactive and thoughtful in how we contribute.

Drive. We’re energized by hard problems. Building AI infrastructure is complex, and we lean into that. We care about doing things well, moving fast, and continuously improving — because solving meaningful challenges is what motivates us.

Grit. Things don’t always work on the first try — and that’s expected. We stay persistent, adapt quickly, and learn as we go. We take setbacks seriously, but not personally, and use them to get better.

Compensation

Monthly range: 7000-8000/month USD

Ready to apply?

Powered by

Gem Logo

First name *

Last name *

Email *

LinkedIn URL

Resume *

Click to upload or drag and drop here

This internship opportunity requires you to be in our office in Palo Alto, CA 5 days a week. Are you currently located in the SF Bay Area or willing to relocate for the length of the internship? *

Select an option

Current degree program: *

Select an option

What is your current field of study or major? *

Computer Science, Computer Engineering, or a related field

Other

What school/university are you currently attending? *

Select an option

What is your current GPA? *

Year of expected graduation *

Select an option

By applying you agree to Gem's terms and privacy policy.

Save your info to apply to other roles faster & help employers reach you.

Apply and saveApply without saving

Req ID: R4

More at DeepInfra

Related open roles

View all roles