GSPANN is hiring an Observability Engineer to manage enterprise platform tools and automation, building observability, scheduling, and integration solutions across the toolset.
Description
Roles and Responsibilities
• Design, deploy, configure, and maintain enterprise tool environments relevant to the assigned track.
• Develop and manage integrations between enterprise platforms and business applications using Representational State Transfer (REST) / Simple Object Access Protocol (SOAP) Application Programming Interfaces (APIs), middleware, and automation frameworks.
• Build and maintain scripting and automation that reduce repetitive operational work and standardize routine tasks.
• Ensure end-to-end operational ownership through proactive monitoring, root-cause analysis (RCA), recurring issue reduction, and continuous improvement, rather than only reactive ticket resolution.
• Support incidents and escalations as the primary technical resource within the assigned platform area.
• Drive an automation and AI mindset by proactively using modern AI-assisted tools for scripting, troubleshooting, and documentation.
• Manage technical documentation, SOPs, and knowledge repositories.
• Collaborate with infrastructure, application, and vendor teams for issue resolution.
• Support BMC Control-M scheduling infrastructure, including batch jobs, workflows, dependencies, and calendars.
• Build enterprise observability and monitoring capabilities across the current and evolving toolset, focusing on dashboarding, alerting, integrations, telemetry visibility, event/log analysis, and actionable operational insights by applying core observability concepts across different products rather than depending on expertise in one specific platform.
• Support source control and platform tooling such as GitHub and Rancher.
• Drive resolution of job failures, monitoring and observability gaps, integration issues, and platform performance problems across the supported toolset.
• Implement scheduling standards, automation policies, observability practices, dashboards, alerting, and monitoring solutions that improve platform reliability and operational visibility.
• Support upgrades, migrations, and environment maintenance for owned platforms.
Skills and Experience
• Bring strong scripting and automation ability, such as PowerShell, Shell, Python, or similar.
• Apply hands-on experience with REST/SOAP APIs, JSON/XML, and integration/middleware platforms.
• Demonstrate strong troubleshooting and root-cause analysis (RCA) skills.
• Work with ITIL processes, including Incident, Change, and Problem Management.
• Demonstrate an operational ownership mindset, with willingness to use AI-assisted tools for scripting, troubleshooting, and documentation.
• Develop the ability to learn and support multiple enterprise platforms beyond a single tool.
• Apply API integrations and automation experience.
• Bring meaningful hands-on experience with one or more of: BMC Control-M (job scheduling, workflow design, batch processing) or observability/monitoring platforms (e.g., AppDynamics, Grafana, Nagios, OpenSearch, Dynatrace, Splunk, Apty) — broad expertise across all is not required.
• Familiarity with GitHub for source control/CI-CD workflows and/or Rancher for container/platform management is a plus.
• Demonstrate ServiceNow familiarity, which is an advantage but not required for this track.