Padmi

Observability Engineering and Platform Reliability (Bengaluru)

BangalorePosted 1 month ago
Software engineeringMid-levelFull Time; Regular
Apply at JRD SYSTEMS INC

Opens the source posting on shine.com

Source description

About the role

View original

Application Support Engineer Observability & Platform Reliability Summar We are seeking a hands-on Application Support Engineer with a strong focus on Observability Engineering and Platform Reliabilit y to join the Technology Operations team. This role is responsible for designing, implementing, and optimizing observability solution s across cloud-based applications and infrastructure. The ideal candidate will bring experience with Splunk, Google Analytics (or Google Analytics 4), and/or Bindplan e, along with a strong foundation in AWS, automation, and Infrastructure as Code (IaC).This role combines application support, observability engineering, and automatio n, with a focus on improving system visibility, telemetry quality, alerting accuracy, and operational reliabilit y. Key Responsibilities: Observability & Monitoring (PRIMARY FOCUS) Design, implement, and maintain observability solutions using tools such a s Splunk, Google Analytics, Bindplane, and Cloud-native monitoring tools. - Develop and optimiz e logging, metrics, and tracing strategi es across applications and infrastructure. - Build and maintai n dashboards, alerts, and anomaly detection mechanis ms to proactively identify system issues. - Integrate telemetry data across multiple sources to improv e end-to-end system visibility - Improve signal-to-noise ratio in alerting and reduce alert fatigue through tuning and correlation. Application & Production Support - Troubleshoot issues acros s application, infrastructure, and integration laye rs in production and non-production environments. - Support application health monitoring and drive improvements i n system reliability and performance. - Participate in incident response and root cause analysis using observability data Automation & Platform Engineering - Build and enhance automation usin g Ansible and scripting (Python/Bas h) for observability deployment and management. - Implement observability components usin g Infrastructure as Code (Terraform, CloudFormation, etc.) - Standardize observability patterns and reusable components across supported platforms. Cloud & Integration - Support applications hosted i n AWS environmen ts, including integration with telemetry and monitoring platforms - Collaborate with engineering teams to instrument applications fo r better observability (logs, metrics, traces) - Support third-party enterprise platforms (SAP, Oracle, Axway, Qlik) with observability integrations Operational Excellence - Maintain systems aligned wit h N1 patching standards - Document observability patterns, dashboards, alerting strategies, and operational procedures - Contribute to continuous improvement initiatives focused o n reducing MTTR and improving system insight. Required Qualifications - 4+ years of experience in Application Support, Observability Engineering, or Systems Engineering - Hands-on experience wi th Splunk (required) - Experience wi th Google Analytics (GA/GA4) and/or Bindplane for telemetry collection and routing - Solid experience wi th monitoring, logging, and observability concepts (metrics, logs, trac - es)Hands-on experience wi th AWS environments - Experience with automation tools (Ansible preferred) - Hands-on experience wi th Infrastructure as Code (Terraform, CloudFormation, or similar - Experience wi th incident management, troubleshooting, and root cause analysis - Scripting experience (Python, Bash, or similar) Preferred Qualifications - Experience designing or improv ing enterprise observability frameworks - Experience integrating multiple telemetry sources i nto centralized platforms (Splunk, etc.) - Experience w ith Bindplane for log/metric collection pipelines - Familiarity w ith Google Analytics for user behavior / digital telemetry analysis - Experience support ing COTS platforms (SAP, Oracle, Axway, Qlik) with monitoring instrumentation - Exposure to CI/CD pipelines and DevOps practices - Experience w ith ServiceNow or ITSM tools - Understanding of distributed systems and microservices observability challenges .

One address, no account. We’ll tell you when matching roles go live.

More at JRD SYSTEMS INC

Related open roles

View all roles