Incident Management

Posted Yesterday
Be an Early Applicant
Hiring Remotely in London, Greater London, England, GBR
In-Office or Remote
Senior level
Artificial Intelligence • Information Technology • Professional Services • Software • Analytics • Generative AI • Big Data Analytics
Where agentic engineering and human experience converge to deliver enterprise outcomes.
The Role
Lead and manage major incidents and problem management for fintech trading platforms. Serve as Incident Commander during outages, run RCAs, enforce SLAs, maintain audit-ready runbooks and controls (SOC Type 1/2), coordinate audits, drive service management reporting, support DR/BCP and automation to reduce MTTR, and ensure continuous operational readiness across engineering, infrastructure, and business stakeholders.
Summary Generated by Built In

Incident Manager/Commander with SOC 1 or SOC 2 Audit experience
Job Description – Incident Commander
UK - Remote
Department: Service Management / Technology Operations.

About the Role
We are seeking a highly skilled Incident Commander to join our Fintech operations team. This role is critical in ensuring the stability, resilience, and continuous availability of our trading and financial platforms. The Incident Commander will lead the management of high- everity incidents, drive root cause analysis, enforce SLAs, and ensure operational readiness through automation, disaster recovery, and business continuity planning.

You will be responsible for end-to end Incident & Problem Management, overseeing command during outages, and partnering closely with engineering, infrastructure, and business stakeholders.

Key Responsibilities
Incident & Problem Management
• Lead and manage IT incidents, including identification, triage, resolution, documentation, and post-incident reviews for critical and high-severity issues.
• Serve as the primary Incident Commander during major outages, ensuring swift coordination between technical teams and business stakeholders.
• Conduct incident washups to identify learnings, drive accountability, and prevent repeat issues.
• Manage and monitor SLAs, ensuring timely resolution and proper lifecycle governance.
• Own Problem Management lead root cause analysis (RCA), track problem records, and oversee corrective/preventive actions.

Service Management & Reporting
• Provide downtime and availability reporting related to application outages.
• Generate monthly Service Management Risk Reports, highlighting key operational risks and remediation status.
• Ensure effective Service Management reporting aligned with the Service Catalog and business expectations (availability, efficiency, ticket handling, customer satisfaction).
• Oversee data integrity across ITSM repositories/tools (e.g., ServiceNow, JSM, HPSM, CMDB).

Audit, Compliance & Process Governance
• Own and maintain audit-ready documentation for operational processes, policies, runbooks, standard operating procedures (SOPs), and control frameworks (e.g., Risk Control Matrices).
• Actively participate in internal and external audit cycles, including SOC Type 1 & Type 2 audits and regulatory examinations (e.g., CFTC, where applicable).
• Join audit calls and walkthroughs with auditors, providing evidence, clarifications, and demonstrations of control effectiveness.
• Ensure all incident, problem, and change management processes are documented, version-controlled, and aligned with audit and regulatory requirements.
• Drive process management by defining, reviewing, and continuously improving operational workflows to meet compliance standards.
• Prepare and deliver audit-related reports, including control testing results, gap analyses, remediation trackers, and compliance dashboards.
• Coordinate with cross functional teams (Operations, Engineering, Risk, Compliance) to ensure timely closure of audit findings and corrective actions.
• Maintain a centralized repository of all audit artifacts, ensuring accessibility and completeness for scheduled and ad-hoc audit requests.

Governance & Continuous Improvement
• Run weekly Release Washups and participate in Change Advisory Board (CAB) meetings to ensure safe and reliable deployments.
• Contribute to building and maintaining a comprehensive CMDB.
• Drive harmonized adoption of the Service Management framework across teams.
• Support automation initiatives to reduce MTTR, improve monitoring/alerting, and streamline incident handling.

Resilience & Continuity
• Develop and manage Disaster Recovery (DR) Plans; lead DR testing exercises.
• Operationalize the Business Continuity Plan (BCP) across IT and business units.

Qualifications & Skills
Must Have
• 8+ years of experience in Incident Management, Problem Management, and IT Operations within fintech, banking, or trading environments.
• Proven experience as Major Incident Manager/Incident Commander handling real-time high-severity events.
• Strong knowledge of ITIL processes (Incident, Problem, Change, Knowledge, Configuration Management).
• Hands-on experience with ITSM tools (ServiceNow, JSM, HPSM, CMDB).
• Experience in reporting & analytics (availability, SLA compliance, RCA trends).
• Experience in audit support, process documentation, and working with internal/external auditors.
• Excellent communication and stakeholder management skills; ability to lead under pressure.
• Good understanding of high-level architecture in software as well as infrastructure.

Nice to Have
• Knowledge of trading systems, order management platforms, or exchanges.
• Experience in running Splunk queries, understanding basic troubleshooting.
• ITIL v3/v4 certification or equivalent.
• Prior experience in global 24×7 financial services environments.
• Familiarity with SOC audit frameworks and regulatory compliance in financial services.
• Prefers working in night shift

Skills Required

  • 8+ years of experience in Incident Management, Problem Management, and IT Operations within fintech, banking, or trading environments
  • Proven experience as Major Incident Manager / Incident Commander handling real-time high-severity events
  • Experience with SOC Type 1 and Type 2 audits and supporting internal/external auditors
  • Strong knowledge of ITIL processes (Incident, Problem, Change, Knowledge, Configuration Management)
  • Hands-on experience with ITSM tools (ServiceNow, JSM, HPSM) and CMDB management
  • Experience in reporting and analytics (availability, SLA compliance, RCA trends)
  • Experience in audit support, process documentation, and working with internal/external auditors
  • Excellent communication and stakeholder management skills; ability to lead under pressure
  • Good understanding of high-level software and infrastructure architecture
  • Knowledge of trading systems, order management platforms, or exchanges
  • Experience running Splunk queries and basic troubleshooting
  • ITIL v3/v4 certification or equivalent
  • Prior experience in global 24x7 financial services environments
  • Familiarity with SOC audit frameworks and financial regulatory compliance
  • Willingness to work night shift (preferred)

What the Team is Saying

Amanda
Abbey
Jon
Markel
Rhoben Cabangbang
Ankit Jain
Tom Sunnergren
Preethi Somasundaram
Melinda Ramos
Satish Kannan
Rhoben Cabangbang
Preethi Somasundaram
Lina Stajic
Le’Rhone Walker
Krishna Prasad TC
Le’Rhone Walker

Bounteous Compensation & Benefits Highlights

  • Healthcare Strength Healthcare coverage is portrayed as comprehensive and affordable, with employer-paid employee premiums, multiple plan options (medical, dental, vision, life, disability), and mental-health support.
  • Leave & Time Off Breadth Time off is positioned as flexible through an unlimited or “take what you need” PTO approach, complemented by paid holidays and Summer Fridays within a remote-first setup.
  • Parental & Family Support Paid parental leave and family-oriented benefits—such as adoption assistance and childcare support—are cited as meaningful parts of the package.

Bounteous Insights

Am I A Good Fit?
beta
Get Personalized Job Insights.
Our AI-powered fit analysis compares your resume with a job listing so you know if your skills & experience align.

The Company
HQ: Frisco, Texas
5,000 Employees
Year Founded: 2003

What We Do

Bounteous is a global AI Services firm where agentic engineering and human experience converge to deliver transformative business outcomes for the enterprise. We help organizations design, build, and scale AI-driven products, platforms, and processes. With more than 5,000 team members worldwide, Bounteous delivers AI that sticks, powering adoption and outcomes that move organizations from experimentation to true transformation.

Why Work With Us

Bounteous combines strategy, experience design, engineering, data, and AI to help leading brands build intelligent systems that are practical, scalable, and measurable. People join us to work on meaningful transformation initiatives, collaborate with smart and supportive teams, and help shape what’s next in AI and digital experience.

Gallery

Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery
Gallery

Bounteous Offices

Remote Workspace

Employees work remotely.

Our remote-first teams of talented individuals collaborate and co-innovate worldwide. We believe productivity thrives anywhere, so you're empowered to work in the way and environment where you perform best.

Typical time on-site:
Company Office Image
HQFrisco Headquarters
Bengaluru Collaboration Center
Boston Collaboration Center
Calgary Collaboration Center
Chennai Collaboration Center
Gurugram Collaboration Center
Hyderabad Collaboration Center
London Collaboration Center
Singapore
Learn more

Similar Jobs

Bounteous Logo Bounteous

Support Engineer

Artificial Intelligence • Information Technology • Professional Services • Software • Analytics • Generative AI • Big Data Analytics
In-Office or Remote
London, Greater London, England, GBR
5000 Employees

Sign up now Access later

Create Free Account

Please log in or sign up to report this job.

Create Free Account