SRE - Observability Engineer Needed

Job ID: 40361857

Budget: ₹250,000 – ₹500,000 INR

Hiring: Site Reliability Engineer (SRE) – Observability (1 month Contract) - Work from Office

Location: Pune
Employment Type: 1 month Contract
Experience: 4–6 Years
We are looking for a skilled SRE – Observability Engineer to manage and enhance monitoring solutions across complex systems.

Key Responsibilities:
• Design and implement end-to-end observability solutions (metrics, logs, traces)
• Monitor application performance, APIs, networks, and infrastructure
• Build dashboards, alerts, SLIs/SLOs for proactive issue detection
• Automate monitoring, alerting, and incident response workflows
• Collaborate with DevOps, development, and infrastructure teams
• Provide L2/L3 support for observability tools and platforms
• Ensure system reliability, scalability, and performance optimization

Required Skills:
• Experience with observability tools: Pandora FMS (preferred), Prometheus, Grafana, ELK Stack, Datadog, New Relic
• Strong knowledge of APM, API monitoring, and network monitoring
• Hands-on experience with cloud platforms (AWS / Azure / GCP)
• Good understanding of Linux systems and networking (TCP/IP, DNS, Load Balancing)
• Scripting experience in Python / Bash / Shell
• Familiarity with Docker and Kubernetes