Scribd

Scribd

Senior Backend Engineer, Content Foundations

San Francisco · Senior

Sponsorship not specifiedDetected 71 days ago
PythonRubySQLNoSQLAWSMachine LearningData EngineeringLLMsCollaborationMentoring

About the role

  • So what are we looking for in new team members?
  • At Scribd, Inc., we hire for "GRIT." Traditionally defined as the intersection of passion and perseverance toward long-term goals, GRIT reflects the mindset we expect from every employee.

Responsibilities

  • Own and drive technical initiatives: Lead the design, implementation, and scaling of core content systems, including ingestion pipelines, metadata services, and content processing workflows.
  • Build scalable, reliable systems: Design robust solutions that handle diverse file formats and edge cases while maintaining high availability and strong data integrity.
  • Collaborate across teams: Partner with Content Security, ML Data Engineering, Search & Discovery, the Content Library squad, and Product to build systems that balance performance, scalability, and user experience.
  • Drive platform evolution: Identify architectural opportunities, propose new capabilities, and help evolve Scribd's content platform to meet growing scale and complexity.
  • Mentor and lead: Provide technical guidance across teams and help raise the bar on system design, data modeling, and production excellence.
  • In addition to building AI into our pipeline, you will help lead the team's adoption of AI coding agents and advanced developer tools to accelerate how we build and scale our systems.
  • Evolving content formats to support downstream AI workflows
  • Lead the design, implementation, and scaling of core content systems, including ingestion pipelines, metadata services, and content processing workflows.
  • Design robust solutions that handle diverse file formats and edge cases while maintaining high availability and strong data integrity.
  • Partner with Content Security, ML Data Engineering, Search & Discovery, the Content Library squad, and Product to build systems that balance performance, scalability, and user experience.

Requirements

  • 7+ years of software engineering experience, including experience navigating the trade-offs of refactoring legacy systems while maintaining high availability.
  • Experience with AWS (Lambda, SQS/SNS, S3, Step Functions) and distributed workflows.
  • Our four products - Scribd®, Slideshare®, Everand™, and Fable - help billions of people across the globe move beyond access and into insight, application, and expertise.

Nice to have

  • Experience with document formats (PDF, ebooks, markdown) and internals (OCR, parsing, transformation).
  • Familiarity with ML/AI systems (embeddings, chunking, retrieval pipelines).
  • Background in spam or content security systems.
  • If you enjoy working on complex, high-scale systems where solving foundational problems unlocks value across an entire ecosystem, we'd love to hear from you.
  • San Francisco is our highest geographic market in the United States.
  • We carefully consider a wide range of factors when determining compensation, including but not limited to experience
  • job-related skill sets
  • and other business and organizational needs.

Skills

  • This posting reflects an approved, open position within the organization.

Compensation

  • At Scribd, Inc., your base pay is one part of your total compensation package and is determined within a range.

Benefits

  • Comprehensive health, dental, and vision coverage
  • Mental health support and disability coverage
  • Generous paid time off, including vacation, sick time, holidays, winter break, volunteer time, and sabbaticals
  • Paid parental leave and family support benefits
  • Retirement matching and employee equity
  • Learning and development programs and professional growth opportunities
  • Wellness and home office stipends
  • Scribd Flex (flexible work model)

Company info

  • The Content Foundations team builds the systems that power how content enters, evolves, and is delivered across Scribd.
  • This includes everything from ingestion, metadata extraction, early quality controls, and the core artifacts that power search, recommendations, AI/ML systems, and the reading and listening experience.

This listing is sourced directly from Scribd's careers page and normalized into a canonical job model.