SpreeAI

SpreeAI

Software Engineer Intern (AI Infrastructure / Training / Inference)

San Francisco, CA · Intern · Internship

Sponsorship not specifiedDetected 20 days ago
PythonJavaGoC++Distributed SystemsBackend DevelopmentObject-Oriented ProgrammingData StructuresAlgorithmsCloud PlatformsDockerKubernetesMachine LearningPyTorchSystems EngineeringResearchLeadership

About the role

  • We are hiring Software Engineers focused on AI Infrastructure to build the systems that enable frontier multimodal AI to operate reliably at production scale.
  • This role exists because modern generative and vision models require infrastructure beyond traditional backend engineering - including GPU orchestration, large-scale inference systems, performance optimization, and developer platforms that allow applied scientists to move fast without sacrificing reliability or cost efficiency.
  • You will work on: Scalable model serving and inference pipelines Distributed GPU infrastructure Performance and cost optimization Reliability, observability, and production readiness You will operate at the boundary between systems engineering and machine learning - building the "paved roads" that allow advanced AI systems to scale safely and efficiently.

Responsibilities

  • Design and build scalable infrastructure supporting training and inference workflows.
  • Develop high-performance APIs and backend services for AI model serving.
  • Optimize GPU utilization, latency, and throughput for multimodal workloads.
  • Build distributed systems supporting large-scale generative models.
  • Partner closely with Applied Science teams to productionize research systems.
  • Drive improvements in deployment workflows, automation, and platform usability.
  • Experience building production backend or distributed systems.
  • We thrive in a dynamic, fast-paced environment where creativity meets technology to drive real impact.
  • About the Role We are hiring Software Engineers focused on AI Infrastructure to build the systems that enable frontier multimodal AI to operate reliably at production scale.
  • You'll have the autonomy to introduce new design concepts, test emerging technologies, and innovate without red tape.

Nice to have

  • Experience with Kubernetes, Docker, or container orchestration.
  • Familiarity with GPU-based ML workloads or distributed training/inference systems.
  • Experience with model serving frameworks (vLLM, Triton, Ray Serve, or similar).
  • Experience with observability tools and performance debugging.
  • Familiarity with PyTorch or ML workflows.
  • Interest in optimizing systems for efficiency, scalability, and developer velocity.
  • Why Join SPREEAI?
  • Real Impact & Ownership: This is an opportunity to shape a product and brand at the forefront of fashion-tech innovation.

Benefits

  • If you've ever wanted to combine your love of design, luxury fashion, and cutting-edge tech, you'll have the freedom to do it here and see your vision realized.
  • Qualifications Degree in Computer Science, Engineering, or comparable combination of education and practical experience.

Company info

  • Your design work will directly impact how thousands (eventually millions) of people experience shopping with SPREEAI - no bureaucratic layers, your ideas can go live and make a difference immediately.
  • Our mission is to redefine the retail landscape with cutting-edge AI solutions that blend high fashion and technology.

This listing is sourced directly from SpreeAI's careers page and normalized into a canonical job model.