Staff Software Engineer - Ingestion Platform
Reddit · Remote - United States
About this role
**Staff Software Engineer — Ingestion Platform** **Who We Are** Reddit is a community of communities—built on shared interests, passion, and trust. With 100,000+ active communities and ~130M daily active unique visitors, Reddit is one of the internet’s largest sources of information. The Infrastructure organization enables Reddit to deliver Reliability, Performance, and Efficiency with a single opinionated technology stack. The Content Platform team within Infrastructure empowers product teams to build the best possible content-related experiences—easily and reliably. This team owns Tier-0 services and core data models powering major product experiences (viewing feeds, posting, commenting, upvoting) and works closely with teams like Consumer, Ads, Feeds, Storage, and Ranking. This team also runs and maintains **R2**, Reddit’s monolith legacy stack, which is in the critical path for many critical user journeys. **What You’ll Do** As a Staff Engineer, you will lead the development and evolution of Reddit’s **Ingestion Platform**. - Design, write, and deliver reliable software for distributed data movement across streaming and batch workloads, focusing on availability, scalability, latency, correctness, and cost efficiency. - Own the architecture and evolution of the platform’s control plane and data plane, including pipeline APIs, controllers, connectors, target managers, schemas, and sink integrations. - Expand beyond Kafka-to-BigQuery by delivering production-ready paths to **S3/GCS** and **Apache Iceberg**, plus transformations, deduplication, dead-letter queues, and other reusable capabilities. - Build high-quality connectors and abstractions for systems such as **Kafka, BigQuery, S3, GCS, Iceberg, Flink**, and other data stores—keeping the platform modular and easy to extend. - Improve the self-service experience for platform users via clear APIs, safe defaults, automated provisioning, documentation, onboarding workflows, and actionable observability/alerting. - Lead migrations from bespoke and legacy ingestion systems to the Ingestion Platform, partnering with teams such as Ads Data Platform and ML Indexing to make migrations safe, incremental, and low-touch. - Establish strong reliability, security, and operational practices for pipelines running across Kubernetes clusters (schema evolution, workload identity, permissions, deployment safety, monitoring, and incident response). - Identify gaps in the platform’s architecture and product experience, and lead redesigns that improve developer velocity and enable Reddit’s continued growth. - Collaborate across Infrastructure, Data Platform, Product, Ads, ML, Storage, and partner teams to align roadmaps and deliver durable solutions. - Mentor and guide backend and data infrastructure engineers across the company, raising the bar for technical design, operational excellence, and customer focus. **Who You Might Be** - 10+ years of hands-on experience building internet-scale distributed systems, data infrastructure, or platforms used by other developers. - BS/MS/PhD in Computer Science (or related field) or equivalent practical experience. - Strong software development experience in one or more general-purpose languages such as **Go, Python, Java, or Scala**. - Deep experience designing and operating high-throughput, fault-tolerant data pipelines or platform services (streaming and batch). - Experience with data movement patterns and systems such as **Kafka, BigQuery, S3, GCS, Apache I
Listing freshness
CronJobs last confirmed this listing 1h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.