<p>SingleStore engineers build the real-time data platform powering some of the world’s most demanding applications. Our cloud-native architecture enables high-performance transactional and analytical workloads at scale, and our teams ship production code continuously throughout the year.</p> <p>We operate in a fast-moving, highly collaborative environment where engineers own their work end-to-end and partner closely across Product, Sales, and Go-To-Market teams to deliver meaningful business impact.</p> <p><strong>Position Summary</strong></p> <p>We are seeking a Software Engineer to join the Observability Team and play a critical role in designing and delivering core capabilities for SingleStore's observability platform. This is a hands-on engineering role with end-to-end ownership of features and projects at the intersection of distributed systems, cloud infrastructure, database technology, and AI-powered observability.</p> <p>As a Software Engineer on this team, you will solve complex system-level problems, contribute meaningfully to technical direction, and partner with Product and customer-facing teams to ensure our platform meets the needs of enterprise customers and new adopters alike.</p> <p>What the team provides: Comprehensive customer observability over their workloads — enabling customers to understand their usage patterns, identify bottlenecks, optimize performance, and leverage powerful alerting capabilities. Our platform unifies traces, logs, and metrics through an open-source-first approach.</p> <p>This is an ideal role for an engineer who thrives on deep technical challenges, takes pride in building durable systems, and is excited about bringing AI-powered observability experiences to customers.</p> <p><strong>What you'll do:</strong></p> <ul> <li>Design and implement scalable observability features for traces, logs, and metrics — spanning ingestion, processing, storage, and visualization</li> <li>Work across control plane and data plane components in a multi-cloud environment (AWS, GCP, Azure), ensuring reliable operation and data consistency at scale</li> <li>Build high-throughput data pipelines that process telemetry data using OpenTelemetry Collector and related open-source tooling</li> <li>Develop and maintain alerting capabilities with Alertmanager, enabling customers to define, tune, and manage alerts with routing, inhibition, and notification management</li> <li>Optimize time-series data storage and query performance using SingleStore DB, handling high-cardinality data and complex analytical queries</li> <li>Contribute to data visualization dashboards in Grafana, creating intuitive customer-facing experiences to explore telemetry data</li> <li>Collaborate closely with Product Management to translate customer and business requirements into robust technical solutions</li> <li>Investigate and resolve difficult issues in production and development environments, debugging data synchronization across distributed systems and cloud providers</li> <li>Participate in on-call rotations to ensure system reliability and respond to incidents promptly</li> </ul> <h4><strong>Your Experience:</strong></h4> <ul> <li>2+ years of professional software development experience building distributed systems or backend services</li> <li>Strong proficiency in Go (Golang) — experience with Rust, Python, or C++ is also valuable</li> <li>Deep understanding of distributed systems concepts: scalability, consistency, high availability, concurrency, and failure modes</li> <li>Familiarity with distributed systems managed via Kubernetes</li> <li>Demonstrated ability to design and build reliable, high-performance system software</li> <li>Experience working in environments where performance, scalability, and reliability are critical</li> <li>Familiarity with observability concepts: traces, logs, metrics, APM, and monitoring patterns</li> <li>Strong problem-solving and debugging skills with the ability to root-cause complex production issues</li> <li>Excellent communication skills, both written and verbal, with ability to collaborate in multicultural, remote-first teams</li> <li>Code quality mindset: you value simplicity, performance, maintainability, and thorough testing</li> </ul> <p><strong>Preferred Qualifications</strong></p> <ul> <li>Experience with time-series data and understanding of metrics cardinality challenges</li> <li>Proficiency with SQL and experience working with relational or distributed databases</li> <li>Experience building cloud-native SaaS platforms with multi-tenant architecture</li> <li>Multi-cloud experience: working with AWS, GCP, Azure, or other cloud providers in a production setting</li> <li>Kubernetes proficiency: operating, monitoring, or developing for Kubernetes clusters</li> <li>Open-source observability tools: hands-on experience with Grafana, Alertmanager, Loki, Tempo, OpenTelemetry Collector, OTLP protocol</li> <li>OpenTelemetry expertise: experience with, or active contributions to OTel projects</li> <li>Time-series database experience: Prometheus TSDB, InfluxDB, Mimir, TimescaleDB, or SingleStore</li> <li>Experience with data pipeline technologies: Apache Kafka, Parquet, Arrow, or stream processing frameworks (Flink, etc.)</li> <li>Experience working with AI agents or LLM-powered applications, including agentic workflows for observability — enabling customers to query telemetry data in natural language</li> <li>Experience in a SaaS or cloud-native company delivering managed services to customers</li> </ul> <p>SingleStore delivers our cloud-native database with the speed and scale to power the world’s data-intensive applications. With a distributed SQL database that introduces simplicity to your data architecture by unifying transactions and analytics, SingleStore empowers digital leaders to deliver exceptional, real-time data experiences to their customers. SingleStore is venture-backed and headquartered in San Francisco with offices in Sunnyvale, Raleigh, Seatt