Seven years keeping business-critical trading, banking, and healthcare applications online — incident management, root cause analysis, and 24×7 production support across ServiceNow, Grafana, Kibana, JIRA, and the FIX protocol. ITIL 4 and AZ-900 certified.
Real production issues from the current and last three roles — the kind that page you at 2am.
Click through the actual root-cause chain behind TKT-001 — the trading outage above.
Drag to compare a manual, spreadsheet-driven tracker against the monitoring stack now running in production.
This is the actual daily routine — health checks across banking and trading applications, condensed into one click.
Not a satisfaction score — an SLA record, tracked since day one on the trading desk.
Eliminate production blind spots across trading, banking, and healthcare systems before they turn into incidents.
Applies to any team running 24×7 mission-critical applications without a dedicated L2 production support owner.
A real category of 3am decision. Pick a response and see what happens.
The FIX session to ADX drops mid-reconnect. Orders queued for the open are stuck in a pending state. Six minutes until the market opens. What do you do?
I'd rather leave this empty than put words in someone's mouth — these slots are open for real quotes from managers, clients, or teammates.
Reach out on LinkedIn if you've worked with me and want to add one.