Observability Engineer

Scripting and AutomationApplication Programming Interfaces (APIs) and IntegrationsEnterprise Observability and MonitoringPlatform Tools

Description

GSPANN is hiring an Observability Engineer to manage enterprise platform tools and automation, building observability, scheduling, and integration solutions across the toolset.

Roles and Responsibilities

• Design, deploy, configure, and maintain enterprise tool environments relevant to the assigned track.

• Develop and manage integrations between enterprise platforms and business applications using Representational State Transfer (REST) / Simple Object Access Protocol (SOAP) Application Programming Interfaces (APIs), middleware, and automation frameworks.

• Build and maintain scripting and automation that reduce repetitive operational work and standardize routine tasks.

• Ensure end-to-end operational ownership through proactive monitoring, root-cause analysis (RCA), recurring issue reduction, and continuous improvement, rather than only reactive ticket resolution.

• Support incidents and escalations as the primary technical resource within the assigned platform area.

• Drive an automation and AI mindset by proactively using modern AI-assisted tools for scripting, troubleshooting, and documentation.

• Manage technical documentation, SOPs, and knowledge repositories.

• Collaborate with infrastructure, application, and vendor teams for issue resolution.

• Support BMC Control-M scheduling infrastructure, including batch jobs, workflows, dependencies, and calendars.

• Build enterprise observability and monitoring capabilities across the current and evolving toolset, focusing on dashboarding, alerting, integrations, telemetry visibility, event/log analysis, and actionable operational insights by applying core observability concepts across different products rather than depending on expertise in one specific platform.

• Support source control and platform tooling such as GitHub and Rancher.

• Drive resolution of job failures, monitoring and observability gaps, integration issues, and platform performance problems across the supported toolset.

• Implement scheduling standards, automation policies, observability practices, dashboards, alerting, and monitoring solutions that improve platform reliability and operational visibility.

• Support upgrades, migrations, and environment maintenance for owned platforms.

Skills and Experience

• Bring strong scripting and automation ability, such as PowerShell, Shell, Python, or similar.

• Apply hands-on experience with REST/SOAP APIs, JSON/XML, and integration/middleware platforms.

• Demonstrate strong troubleshooting and root-cause analysis (RCA) skills.

• Work with ITIL processes, including Incident, Change, and Problem Management.

• Demonstrate an operational ownership mindset, with willingness to use AI-assisted tools for scripting, troubleshooting, and documentation.

• Develop the ability to learn and support multiple enterprise platforms beyond a single tool.

• Apply API integrations and automation experience.

• Bring meaningful hands-on experience with one or more of: BMC Control-M (job scheduling, workflow design, batch processing) or observability/monitoring platforms (e.g., AppDynamics, Grafana, Nagios, OpenSearch, Dynatrace, Splunk, Apty) — broad expertise across all is not required.

• Familiarity with GitHub for source control/CI-CD workflows and/or Rancher for container/platform management is a plus.

• Demonstrate ServiceNow familiarity, which is an advantage but not required for this track.

Apply Now

PDF or DOCX up to 5MB