SpreeAI
Software Engineer Intern (AI Infrastructure / Training / Inference)
San Francisco, CA · Intern · Internship
Sponsorship not specifiedDetected 20 days ago
PythonJavaGoC++Distributed SystemsBackend DevelopmentObject-Oriented ProgrammingData StructuresAlgorithmsCloud PlatformsDockerKubernetesMachine LearningPyTorchSystems EngineeringResearchLeadership
About the role
- We are hiring Software Engineers focused on AI Infrastructure to build the systems that enable frontier multimodal AI to operate reliably at production scale.
- This role exists because modern generative and vision models require infrastructure beyond traditional backend engineering - including GPU orchestration, large-scale inference systems, performance optimization, and developer platforms that allow applied scientists to move fast without sacrificing reliability or cost efficiency.
- You will work on: Scalable model serving and inference pipelines Distributed GPU infrastructure Performance and cost optimization Reliability, observability, and production readiness You will operate at the boundary between systems engineering and machine learning - building the "paved roads" that allow advanced AI systems to scale safely and efficiently.
Responsibilities
- Design and build scalable infrastructure supporting training and inference workflows.
- Develop high-performance APIs and backend services for AI model serving.
- Optimize GPU utilization, latency, and throughput for multimodal workloads.
- Build distributed systems supporting large-scale generative models.
- Partner closely with Applied Science teams to productionize research systems.
- Drive improvements in deployment workflows, automation, and platform usability.
- Experience building production backend or distributed systems.
- We thrive in a dynamic, fast-paced environment where creativity meets technology to drive real impact.
- About the Role We are hiring Software Engineers focused on AI Infrastructure to build the systems that enable frontier multimodal AI to operate reliably at production scale.
- You'll have the autonomy to introduce new design concepts, test emerging technologies, and innovate without red tape.
Nice to have
- Experience with Kubernetes, Docker, or container orchestration.
- Familiarity with GPU-based ML workloads or distributed training/inference systems.
- Experience with model serving frameworks (vLLM, Triton, Ray Serve, or similar).
- Experience with observability tools and performance debugging.
- Familiarity with PyTorch or ML workflows.
- Interest in optimizing systems for efficiency, scalability, and developer velocity.
- Why Join SPREEAI?
- Real Impact & Ownership: This is an opportunity to shape a product and brand at the forefront of fashion-tech innovation.
Benefits
- If you've ever wanted to combine your love of design, luxury fashion, and cutting-edge tech, you'll have the freedom to do it here and see your vision realized.
- Qualifications Degree in Computer Science, Engineering, or comparable combination of education and practical experience.
Company info
- Your design work will directly impact how thousands (eventually millions) of people experience shopping with SPREEAI - no bureaucratic layers, your ideas can go live and make a difference immediately.
- Our mission is to redefine the retail landscape with cutting-edge AI solutions that blend high fashion and technology.
Apply directly at SpreeAI →Create a free account for alerts like thisView SpreeAI immigration profile
This listing is sourced directly from SpreeAI's careers page and normalized into a canonical job model.