Software Engineer | Observability
SingleStore · United States
About this role
## Software Engineer | Observability ### About the role SingleStore engineers build the real-time data platform powering some of the world’s most demanding applications. In this role, you’ll join the **Observability Team** to design and deliver core capabilities for SingleStore’s observability platform—an end-to-end, hands-on engineering position at the intersection of **distributed systems, cloud infrastructure, database technology, and AI-powered observability**. You’ll work in a fast-moving, highly collaborative environment, owning features and partnering closely with **Product, Sales, and Go-To-Market** teams to deliver meaningful business impact. ### What you’ll do - Design and implement scalable observability features for **traces, logs, and metrics** (ingestion, processing, storage, and visualization) - Work across **control plane and data plane** components in a multi-cloud environment (**AWS, GCP, Azure**), ensuring reliable operation and data consistency at scale - Build high-throughput telemetry data pipelines using **OpenTelemetry Collector** and related open-source tooling - Develop and maintain alerting capabilities with **Alertmanager** (routing, inhibition, notification management) - Optimize **time-series storage and query performance** using **SingleStore DB**, including high-cardinality data and complex analytical queries - Contribute to **Grafana** dashboards to create intuitive customer-facing experiences for exploring telemetry data - Collaborate with Product Management to translate customer and business requirements into robust technical solutions - Investigate and resolve difficult production and development issues, including debugging data synchronization across distributed systems and cloud providers - Participate in **on-call rotations** to ensure reliability and respond to incidents promptly ### Your experience - **2+ years** of professional software development building distributed systems or backend services - Strong proficiency in **Go (Golang)** (Rust, Python, or C++ also valuable) - Deep understanding of distributed systems: scalability, consistency, high availability, concurrency, and failure modes - Familiarity with distributed systems managed via **Kubernetes** - Ability to design and build reliable, high-performance system software - Experience where performance, scalability, and reliability are critical - Familiarity with observability concepts: **traces, logs, metrics, APM, monitoring patterns** - Strong problem-solving and debugging skills (root-cause complex production issues) - Excellent communication skills (written and verbal) in multicultural, remote-first teams - Code quality mindset: simplicity, performance, maintainability, and thorough testing ### Preferred qualifications - Experience with **time-series data** and metrics cardinality challenges - Proficiency with **SQL** and experience with relational or distributed databases - Experience building cloud-native **multi-tenant SaaS** platforms - Multi-cloud experience (**AWS, GCP, Azure**, etc.) in production - **Kubernetes** proficiency (operating, monitoring, or developing for clusters) - Hands-on experience with open-source observability tools (e.g., **Grafana, Alertmanager, Loki, Tempo, OpenTelemetry Collector, OTLP**) - OpenTelemetry expertise (experience with or contributions to OTel projects) - Time-series database experience (e.g., **Prometheus TSDB, InfluxDB, Mimir, TimescaleDB, SingleStore**) - Data pipeline technologies (e.g
Listing freshness
CronJobs last confirmed this listing 2h ago. If its source stops confirming the opening for seven days, this page is removed from active inventory.