Content Foundations is a newly formed team at the center of Scribd, Inc.'s biggest strategic investment. Behind our UGC brands (Scribd and Slideshare) sits one of the largest corpora of human-created knowledge anywhere: hundreds of millions of documents and presentations. Every piece of that knowledge passes through one system. It’s uploaded, understood, transformed, and served to millions of customers, and agents that cite it, ground it, and learn from it.
Content Foundations owns that system, and its work lands in two places at once. It powers the product roadmap: faster uploads, richer document experiences, and the growth of the corpus. And it operates as a platform for the teams that build on it: trust & safety, content understanding, supply, and AI platform teams all consume its interfaces. What this team ships translates directly into product outcomes for users and velocity for every team downstream. Content sits at the very top of Scribd, Inc.'s strategy: it’s the engine of our consumer brands and the foundation of our growth, and the ingestion platform is where all of it begins.
We’re hiring the engineering manager for Content Foundations to build and lead the team that owns how content enters Scribd. This is a leadership role at the intersection of large-scale data infrastructure and applied AI, and it carries three responsibilities in equal measure.
You’ll be the technical leader for the platform: setting the vision for real-time, AI-native ingestion, making the architectural calls that evolve a live, high-scale system, and holding the quality and reliability bar for a system the whole company depends on.
You’ll be the people leader for the team: hiring, growing, and inspiring a group of backend and AI data engineers, and shaping its culture, standards, and identity from its first days.
And you’ll be the partner the content organization plans around: running the team’s roadmap as a product, making commitments other teams can build against, and representing the platform in decisions with product and engineering leaders across the company.
The industry is redefining what document ingestion means. Documents can now be parsed, classified, safety-checked, and enriched in real time at the moment of upload, before they ever reach a reader, with vision-language models and intelligent chunking replacing what OCR and batch pipelines used to approximate. You’ll bring those capabilities into one of the world’s largest document platforms, with the corpus scale to make it matter.