C

Observability Engineer

Careernet · Chennai, Tamil Nadu, India

3–10 yrs experiencefull_timePosted 3w ago
Apply now →

Job description

Experience **Minimum Experience:** **47 years** Minimum of **4 years** of experience in an Observability Platform Engineering role supporting: - Large-scale distributed systems - High-traffic production environments - Cloud-native platforms - Microservices-based architectures Basic Qualifications - Certification in one or more observability platforms such as **New Relic, Dynatrace, Zabbix, FullStory, or Splunk** (Administrator or Architect level preferred). - Proficiency in scripting languages such as **Python** and **Bash**. - Hands-on experience with automation tools including **Terraform, Ansible, Puppet, and Jenkins**. - Experience working within a **SAFe Agile** environment, including PI Planning, Agile Release Trains (ARTs), and cross-functional collaboration. - Hands-on experience managing telemetry agents such as: - Zabbix Agent - Splunk Universal Forwarder - New Relic Agents - Experience configuring authentication and authorization using **SSO, LDAP, SAML**, and role-based access control. - Experience developing custom integrations to extend observability capabilities. - Strong troubleshooting, monitoring, and platform health management skills. - Deep understanding of **Azure** and **Google Cloud Platform (GCP)**. - Experience sizing, designing, deploying, and administering distributed observability platforms. - Experience onboarding diverse applications while optimizing platform performance, scalability, and licensing. - Experience integrating observability platforms with enterprise systems such as **CMDB**, ITSM, and ticketing platforms. - Experience implementing monitoring for the observability platform itself to ensure platform health, stability, and license optimization. - Experience collaborating with external technology vendors. Preferred Qualifications - Hands-on administration experience with: - **New Relic** - **Splunk** - **Zabbix** - **FullStory** - APM - Infrastructure Monitoring - Real User Monitoring (RUM) - Synthetic Monitoring - Session Replay - Experience with SDLC and IT Service Management tools including: - Jira - ServiceNow - PagerDuty - Confluence - Bitbucket - Experience managing vendor relationships and technical engagements. - Hands-on experience with public cloud platforms including **Azure** and **Google Cloud Platform (GCP)**. - Relevant certifications in **Azure**, **New Relic**, **Splunk**, **Zabbix**, or **FullStory**. Overall Improvements Made - Standardized capitalization (e.g., **Azure**, **Bitbucket**, **ServiceNow**, **New Relic**, **Google Cloud Platform (GCP)**). - Corrected grammar and sentence structure throughout. - Improved consistency in verb tense and bullet formatting. - Replaced inconsistent terminology (e.g., *Metric, Event, Log, Trace* *metrics, events, logs, and traces*). - Standardized terminology for observability, monitoring, telemetry, and cloud technologies. - Improved readability while maintaining the original responsibilities and intent. - Aligned the JD to the style typically used by multinational technology organizations and Global Capability Centers (GCCs).