IBM Netcool / Observability Technical Lead
Inside IR35 Contract - £800-£827pd
London Hybrid - 4 Days Onsite / 1 Day WFH per Week
Banking
Are you the person who knows exactly why an ObjectServer failover didn't behave as expected, and how to stop a flood of duplicate events before anyone else notices? If so, this is your opportunity to own an enterprise monitoring estate end to end, as the recognised technical authority.
You'll join a leading global financial institution's Digital Engineering function, where 24/7 technology resilience isn't a nice-to-have, it's the whole point. As the Subject Matter Expert for the IBM Tivoli Netcool Omnibus (NOI) platform, you'll shape how thousands of infrastructure and application events are detected, correlated and actioned across EMEA, and you'll have genuine scope to modernise observability capability rather than simply keep the lights on.
This is a hands-on technical leadership role with no direct reports, so your influence comes from your expertise, not a headcount.
What You'll Be Doing
- Own the Netcool estate – administer, maintain and upgrade Omnibus, ObjectServers, Gateways, Probes, WebGUI and Impact, ensuring high availability, resilience and robust lifecycle management
- Engineer smarter event management – design and maintain correlation, suppression, deduplication, enrichment and automation rules that cut alert noise and sharpen operational response
- Onboard new services – bring new infrastructure, applications and cloud services into monitoring, working alongside infrastructure, middleware, network, application and cloud teams
- Lead technical investigations – resolve platform incidents, event processing failures and integration problems, providing SME support during major incidents and driving root cause analysis
- Build integrations and automation – connect monitoring with ServiceNow, ITSM, reporting and automation tooling via gateways, APIs and event forwarding, and develop scripts that remove manual effort
- Set the standards – define monitoring policies, best practice and documentation, contribute to the observability roadmap, and mentor colleagues across operational and engineering teams
What You'll Need
- Expert-level, hands-on experience with IBM Tivoli Netcool Omnibus (NOI) in large-scale enterprise environments, including Omnibus, Impact and WebGUI
- Strong grasp of Netcool architecture – ObjectServers, Probes, Gateways (ServiceNow, JDBC), plus failover and high-availability configurations
- Proven track record delivering platform upgrades, patching, migrations and lifecycle activities on mission-critical 24x7 platforms
- Strong SQL and ObjectServer database administration, with Impact development and administration experience
- Scripting and development capability across Impact Policy Language, Probe rules files, JavaScript, PowerShell and shell scripting (Bash, Korn Shell or similar)
- Solid Linux/UNIX administration skills and deep understanding of event processing, correlation and alert lifecycle management
- Experience integrating monitoring with ServiceNow and automation platforms
- Excellent troubleshooting and root cause analysis skills, within an ITIL-aligned Incident, Problem and Change environment
- Confident communication and stakeholder management with both technical and non-technical audiences
- Desirable: IBM Tivoli Monitoring (ITM), Dynatrace, IBM Instana or OpenTelemetry, REST APIs/JSON/XML, cloud monitoring across AWS, Azure or GCP, container and Kubernetes monitoring, SRE practices, financial services or regulated environment experience, and leading technical improvement initiatives
If you've held any of these roles or used these technologies/skills, this role could be a great fit: Netcool Engineer, Netcool Omnibus Administrator, Netcool Consultant, IBM Netcool Specialist, Tivoli Engineer, Tivoli Netcool SME, Monitoring Engineer, Monitoring Tools Engineer, Enterprise Monitoring Specialist, Event Management Engineer, Observability Engineer, Systems Monitoring Lead, Infrastructure Monitoring Analyst, ITM Administrator, NOI, ObjectServer, Netcool Impact, WebGUI, Probes and Gateways, ITIL, ServiceNow integration, Dynatrace, Instana, OpenTelemetry, SRE, Observability & Monitoring TPM, Systems Monitoring Technical Lead.


