Sr. Manager Software Development and Engineering
Your Opportunity
At Schwab, you’re empowered to make an impact on your career. Here, innovative thought meets creative problem solving, helping us “challenge the status quo” and transform the finance industry together.
We believe in the importance of in-office collaboration and fully intend for the selected candidate for this role to work on site in the specified location.
This role serves as the senior technical Subject Matter Expert (SME) for the Liquidity, Cash & Collateral Management (“Liquidity”) Platform, providing operational leadership and technical stewardship for business-critical applications and services.
The position is responsible for ensuring the reliability, resiliency, supportability, and operational effectiveness of the Liquidity Platform and associated services. The role partners closely with Product Owners, Business Stakeholders, Development teams, Architects, Infrastructure teams, and Site Reliability Engineering (SRE) partners to support critical business functions while driving continuous improvement and operational excellence.
As a Grade 58 Senior Production Support Engineer, this individual operates as both a technical leader and trusted business partner. The role extends beyond incident resolution and focuses on proactive risk management, resilient system design, technical debt reduction, service maturity, operational readiness, and long-term platform sustainability. The successful candidate will possess strong business knowledge of Liquidity Management processes and apply an SRE mindset to incident response, root cause analysis, and continuous service improvement.
Key Responsibilities
Platform Ownership & SME Leadership
- Serve as the primary Subject Matter Expert (SME) for the Liquidity platform providing senior operational support leadership for designated Liquidity applications and services.
- Develop deep knowledge of Liquidity Management business processes, operational workflows, data flows, controls, and system integrations.
- Act as the primary escalation point for complex production issues affecting Liquidity applications and services.
- Lead technical troubleshooting efforts across applications, infrastructure, integrations, databases, middleware, and data domains.
- Provide technical leadership during major incidents, outages, and critical business events.
- Partner with business and technology teams to ensure platform stability, operational readiness, and service reliability.
- Mentor and develop peer engineers by sharing technical expertise, promoting operational best practices, and providing guidance on incident management, resiliency engineering, supportability, and problem-solving approaches.
- Foster knowledge sharing across teams through technical coaching, documentation, operational reviews, and cross-functional collaboration.
Business & Product Partnership
- Build strong partnerships with Product Owners, Business Owners, and Engineering leaders supporting the Liquidity Platform.
- Collaborate with Product and Development teams to ensure operational supportability, resiliency, security, and risk reduction requirements are incorporated into product roadmaps.
- Influence prioritization of technical debt remediation, resiliency improvements, and operational modernization initiatives.
- Participate in strategic planning discussions to balance business objectives, platform health, and operational risk.
- Serve as a trusted advisor to Product Owners by providing operational insights, reliability trends, service health recommendations, and supportability guidance.
Reliability & Resiliency Engineering
- Partner with Development teams to design, implement, and maintain resilient, highly available production environments.
- Apply a Site Reliability Engineering (SRE) mindset to incident management, focusing on rapid service restoration, elimination of recurring issues, operational excellence, and reduction of operational toil.
- Drive operational excellence initiatives focused on reliability, observability, recoverability, scalability, and performance.
- Identify, prioritize, and champion remediation of technical debt impacting platform stability, supportability, and operational efficiency.
- Lead efforts to reduce operational risk and eliminate recurring production issues through engineering improvements, automation, and platform modernization.
- Advocate for engineering practices that improve long-term sustainability and reduce operational burden.
- Ensure new applications, enhancements, and integrations are designed with reliability and operational readiness as core requirements.
Incident & Problem Management
- Lead investigation and resolution efforts for complex production incidents.
- Coordinate cross-functional response activities during major incidents.
- Perform root cause analysis and ensure corrective actions address underlying issues.
- Drive problem management efforts focused on long-term issue elimination.
- Establish operational best practices and support standards across supported platforms.
Operational Excellence & Automation
- Define and enhance monitoring, alerting, observability, and operational dashboards.
- Leverage automation and AI-assisted tools to reduce operational toil and improve efficiency.
- Establish meaningful operational health metrics and service-level objectives.
- Improve operational documentation, runbooks, recovery procedures, and support processes.
- Ensure production readiness standards are met for new releases and platform changes.
- Participate in certificate management activities, including certificate health monitoring, renewals, deployment validation, and prevention of certificate-related outages.
Scope & Expectations
- Senior individual contributor role with significant technical influence.
- Primary operational and technical owner as well as senior support lead for the Liquidity Platform.
- Trusted partner to Product Owners, Development teams, architects, and business stakeholders.
- Operates with a high degree of autonomy and technical judgment.
- Drives reliability, resiliency, technical debt reduction, and operational maturity.
- Leads cross-functional initiatives to improve service health and reduce operational risk.
- Serves as a mentor to peers and engineers across the organization, actively contributing to technical development, knowledge sharing, operational excellence, and adoption of reliability engineering best practices.
- Balances immediate operational needs with long-term platform sustainability and strategic objectives.
What you have
To ensure that we have fulfilled our promise of "challenging the status quo," this role has specific qualifications that successful candidates should have.
Required Qualifications
- 10+ years of experience in Production Support, Site Reliability Engineering (SRE), Application Support, Enterprise Platform Operations, or related operational engineering disciplines.
- Experience supporting Liquidity Management, Financial Services, Capital Markets, Banking, Risk, Treasury, or Securities Lending platforms.
- Strong business knowledge of Liquidity Management processes, workflows, operational controls, and supporting technologies.
- Demonstrated experience serving as a technical lead or platform SME for complex enterprise applications.
- Experience leading high-severity incident response, problem management, and root cause analysis efforts.
- Proven ability to apply Site Reliability Engineering (SRE) principles to improve platform stability, resiliency, observability, and operational efficiency.
- Strong understanding of system resiliency, observability, operational readiness, and enterprise support models.
- Bachelor's degree in Computer Science, Information Technology, Engineering, or equivalent practical experience.
Strong Working Knowledge Of
- SQL and data analysis
- Enterprise application architecture
- Real-time integrations and APIs
- Control-M job scheduling, batch processing, and workload automation
- Enterprise batch processing and data movement frameworks
- Monitoring, observability, and alerting platforms
- Linux and Windows environments
- Networking fundamentals including DNS, TCP/IP, and SSL/TLS
- Incident, problem, change, problem, and release management practices
- Automation, scripting, and operational tooling
Preferred Qualifications
- Experience supporting Liquidity Management platforms.
- Deep understanding of Liquidity, Cash & Collateral Management operations.
- Experience supporting Liquidity, Cash & Collateral Management operations or related financial services applications.
- Practical experience implementing and championing Site Reliability Engineering (SRE) practices.
- Technical debt management and platform modernization initiatives.
- Automation, AI-assisted operations, and observability platforms.
- Leading resiliency, recoverability, and operational maturity initiatives.
- Defining and managing SLAs, SLOs, KPIs, and service health metrics.
Work Schedule & Availability
- Standard business hours with flexibility to support critical production events and major incidents as needed.
- Participation in incident escalation, production support, and cross-functional operational activities.
- Availability to collaborate with distributed technology and business teams.
- Participate in an on-call rotation and provide after-hours support coverage as required to ensure the stability and availability of business-critical systems.
- Respond to high-severity incidents and critical production events outside normal business hours when necessary.
Primary Focus Areas
Liquidity Platform (Transcend) – Primary Responsibility
- Platform SME and technical leader
- Operational ownership and platform health
- Reliability, resiliency, and supportability
- Product Owner partnership
- Technical debt management
- Strategic platform modernization
Securities Lending – Secondary Responsibility
- Senior operational support leadership
- Application stability and resiliency
- Engineering and Product partnership
- Incident management and problem resolution
- Operational excellence and continuous improvement
This role is intended for a senior technical leader who combines deep production support expertise, strong Liquidity business knowledge, an SRE mindset, and proven partnership skills to drive reliability, resiliency, and long-term platform success.
In addition to the salary range, this role is also eligible for bonus or incentive opportunities
What’s in it for you
At Schwab, you’re empowered to shape your future. We champion your growth through meaningful work, continuous learning, and a culture of trust and collaboration—so you can build the skills to make a lasting impact. Our Hybrid Work and Flexibility approach balances our ongoing commitment to workplace flexibility, serving our clients, and our strong belief in the value of being together in person on a regular basis.
We offer a competitive benefits package that takes care of the whole you – both today and in the future:
- 401(k) with company match and Employee stock purchase plan
- Paid time for vacation, volunteering, and 28-day sabbatical after every 5 years of service for eligible positions
- Paid parental leave and family building benefits
- Tuition reimbursement
- Health, dental, and vision insurance