← Back to jobs
Chennai, Tamil Nadu, India
No related jobs found
Role description
We are seeking a hands-on Application Support Engineer with a strong focus on Observability Engineering and Platform Reliability to join the Technology Operations team.
This role is responsible for designing, implementing, and optimizing observability solutions across cloud-based applications and infrastructure. The ideal candidate will bring experience with Splunk, Google Analytics (or Google Analytics 4), and/or Bindplane, along with a strong foundation in AWS, automation, and Infrastructure as Code (IaC).
This role combines application support, observability engineering, and automation, with a focus on improving system visibility, telemetry quality, ing accuracy, and operational reliability.
Key Responsibilities
Observability & Monitoring (PRIMARY FOCUS)
Design, implement, and maintain observability solutions using tools such as Splunk, Google Analytics, Bindplane, and Cloud-native monitoring tools
Develop and optimize logging, metrics, and tracing strategies across applications and infrastructure
Build and maintain dashboards, s, and anomaly detection mechanisms to proactively identify system issues
Integrate telemetry data across multiple sources to improve end-to-end system visibility
Improve signal-to-noise ratio in ing and reduce fatigue through tuning and correlation
Application & Production Support
Troubleshoot issues across application, infrastructure, and integration layers in production and non-production environments
Support application health monitoring and drive improvements in system reliability and performance
Participate in incident response and root cause analysis using observability data
Automation & Platform Engineering
Build and enhance automation using Ansible and scripting (Python/Bash) for observability deployment and management
Implement observability components using Infrastructure as Code (Terraform, CloudFormation, etc.)
Standardize observability patterns and reusable components across supported platforms
Cloud & Integration
Support applications hosted in AWS environments, including integration with telemetry and monitoring platforms
Collaborate with engineering teams to instrument applications for better observability (logs, metrics, traces)
Support third-party enterprise platforms (SAP, Oracle, Axway, Qlik) with observability integrations
Operational Excellence
Maintain systems aligned with N‑1 patching standards
Document observability patterns, dashboards, ing strategies, and operational procedures
Contribute to continuous improvement initiatives focused on reducing MTTR and improving system insight
Skills
Hands-on experience with Splunk (required)
Experience with Google Analytics (GA/GA4) and/or Bindplane for telemetry collection and routing
Strong experience with monitoring, logging, and observability concepts (metrics, logs, traces)
Hands-on experience with AWS environments
Experience with automation tools (Ansible preferred)
Hands-on experience with Infrastructure as Code (Terraform, CloudFormation, or similar)
Experience with incident management, troubleshooting, and root cause analysis
Scripting experience (Python, Bash, or similar)
Preferred Qualifications
Experience designing or improving enterprise observability frameworks
Experience integrating multiple telemetry sources into centralized platforms (Splunk, etc.)
Experience with Bindplane for log/metric collection pipelines
Familiarity with Google Analytics for user behavior / digital telemetry analysis
Experience supporting COTS platforms (SAP, Oracle, Axway, Qlik) with monitoring instrumentation
Exposure to CI/CD pipelines and DevOps practices
Experience with ServiceNow or ITSM tools
Understanding of distributed systems and microservices observability challenges
Any Graduate
No related jobs found
← Back to jobs