07 Aug
|
The Networker
|
Bengaluru
07 Aug
The Networker
Bengaluru
We re seeking a passionate and experienced Software Engineer focused on distributed systems to join our Observability Platform team. In this pivotal role, you will design and build scalable infrastructure that powers the telemetry, monitoring, and reliability needs of our platform teams. Our Observability platform is a hyperscaler, operating on billions of time series and petabytes of logs, and is completely built on cutting-edge open-source technologies such as Prometheus, ClickHouse, OpenTelemetry, and more. You will work across the stack from data ingestion and storage to visualization to ensure we build systems that are reliable, performant, and a joy to use.
You will partner closely with world-class SREs, platform engineers, and service owners to support the development, scaling, and maintenance of mission-critical observability systems that power modern, high-scale distributed applications. Working with this team unlocks unparalleled exposure to SREs who manage eBays site, comprised of thousands of microservices. Youll be at the bleeding edge, solving observability problems that are not even a concern for most yet, contribute to open-source projects, and gain an invaluable learning experience.
What You ll Do
- Architect and implement scalable, fault-tolerant telemetry infrastructure with a focus on reliability and low operational overhead
- Build and optimize services for ingesting, transforming, and querying telemetry data such as logs, metrics, and traces
- Operate production-grade Kubernetes-based systems with a focus on self-healing, autoscaling, and robustness
- Collaborate across teams to understand observability needs and build tools that enhance operational excellence
- Contribute to system design reviews and post-incident retrospectives with a mindset of continuous improvement
- Participate in the team s site on-call rotation to help ensure the availability and reliability of our observability systems
- Optionally, build user-facing interfaces to surface observability insights (using React/JavaScript)
Basic Qualifications
- 6+ years of professional experience in full stack software development and infrastructure engineering
- Hands-on experience building full-stack applications using up-to-date backend and frontend technologies
- Strong programming skills in Golang (or equivalent system-level languages such as Java, C++, or Python)
- Experience designing and consuming REST/gRPC APIs and building scalable microservices
- Deep understanding of distributed systems fundamentals including scalability, fault tolerance, consistency, and resiliency
- Experience integrating AI/LLM capabilities into software applications using modern AI frameworks, MCP Servers and AI Agent Frameworks(LangChain, LangGraph or similar)
- Working knowledge of AI-assisted software development, prompt engineering, Retrieval-Augmented Generation (RAG), vector databases, AI agents, or related technologies
- Hands-on experience deploying and operating services in containerized environments (Kubernetes)
- Familiarity with observability pillars - metrics, logs, traces
- Clear communication skills and a collaborative mindset
Preferred Qualifications
- Experience with Prometheus, Grafana, OpenTelemetry, ClickHouse, or similar tools
- Exposure to time-series data storage and query optimization
- Strong frontend development experience using React, JavaScript/TypeScript
- Contributions to open-source infrastructure or observability tooling
- Experience with high-throughput data pipelines in cloud-native environments
- Frontend development using React/JavaScript
Disclaimer : This job posting has been aggregated from external source. Role details, content, and availability are subject to change. Applicants are advised to confirm the latest information directly on the company website before applying.
📌 Software Engineer - Fullstack (Bengaluru)
🏢 The Networker
📍 Bengaluru