Skip to main content
Search Jobs

Search Jobs

Manager, Software Development & Engineering

Southlake, Texas, United States Requisition ID 2026-126989 Category Engineering & Software Development Position Type Regular Pay range USD $112,300.00 - $187,100.00 / Year Application Deadline 2026-09-21
Apply Now

Your Opportunity


Schwab remains committed to providing increased visibility to career growth opportunities and job requirements. This posting announcement is part of increased transparency and while all qualified applicants will be reviewed and considered, this organization has a preferred candidate identified for this role.

At Schwab, you’re empowered to make an impact on your career. Here, innovative thought meets creative problem solving, helping us “challenge the status quo” and transform the finance industry together.

We believe in the importance of in-office collaboration and fully intend for the selected candidate for this role to work on site in the specified location(s).

​​Schwab Technology Services enables the future of how clients manage their money by providing innovative and reliable technology products and services as part of our ongoing commitment to democratize access to investing and financial planning. 
​ 

As a Senior Site Reliability Engineer (SRE), you will be responsible for ensuring the reliability, availability, scalability, and operational excellence of critical enterprise applications. You will support mission-critical production environments, lead incident triage and resolution efforts, drive automation initiatives, and implement best-in-class monitoring and observability solutions. 
This role partners closely with Development, Infrastructure, Security, Product, and Platform Engineering teams to improve system resilience, streamline software delivery, and reduce operational risk. The ideal candidate combines strong production support expertise with DevOps engineering, automation, release management, and cloud technologies to deliver highly available and performant services.

What you'll do:

Production Reliability & Support 

  • Monitor application availability, performance, and health of mission-critical platforms and services. 
  • Lead triage, troubleshooting, and resolution of complex production incidents. 
  • Perform impact analysis and communicate incident status, risks, and remediation plans to stakeholders. 
  • Participate in on-call rotations and provide escalation support during critical issues. 
  • Conduct root cause analysis (RCA) and drive corrective and preventative actions. 

Observability & Monitoring

  • Design, implement, and maintain monitoring, alerting, dashboards, and observability frameworks. 
  • Develop proactive monitoring strategies using tools such as Splunk, AppDynamics, Dynatrace, or similar platforms. 
  • Define and track Service Level Indicators (SLIs), Service Level Objectives (SLOs), and reliability metrics. 
  • Identify performance bottlenecks and recommend optimization strategies. 

DevOps & Continuous Delivery 

  • Build and enhance CI/CD pipelines using Jenkins, Bamboo, GitLab, Harness, Nexus, or equivalent tools. 
  • Automate application deployments, configuration management, and operational processes. 
  • Improve release engineering practices and deployment methodologies. 
  • Support source-code branching strategies and deployment automation frameworks. 
  • Partner with development teams to improve application supportability and production readiness. 

Release & Change Management

  • Coordinate production releases across multiple environments. 
  • Review and validate change requests to ensure accuracy and minimize operational risk. 
  • Develop and maintain deployment, rollback, and release procedures. 
  • Support release governance and compliance requirements. 

Automation & Engineering Excellence 

  • Develop automation solutions using Shell, Bash, PowerShell, Python, or similar scripting languages. 
  • Eliminate repetitive manual processes and reduce operational toil. 
  • Create and maintain operational runbooks, knowledge articles, and documentation. 
  • Continuously improve operational processes and reliability practices.

 Cloud & Platform Engineering 

  • Support cloud-hosted and on-premise platforms across PCF, AWS, OpenShift, and GCP environments. 
  • Implement resiliency, failover, disaster recovery, and capacity planning strategies. 
  • Collaborate with infrastructure and platform teams to strengthen platform reliability and scalability. 
  • Support middleware and messaging technologies  

Leadership & Collaboration 

  • Serve as a technical lead during major incidents and production events. 
  • Mentor junior engineers and promote SRE best practices. 
  • Collaborate with cross-functional teams across multiple time zones. 
  • Drive a culture of accountability, automation, operational excellence, and continuous improvement.

What you have


To ensure that we fulfill our promise of "challenging the status quo," this role has specific qualifications that successful candidates should have.

Required Qualifications:

Technical Skills 

  • 7+ years of experience in Site Reliability Engineering, DevOps, Production Support, or Platform Engineering. 
  • 5+ years of experience supporting enterprise-scale production applications. 
  • Experience supporting middleware and web application platforms within financial services environments. 
  • Strong Linux administration and troubleshooting experience. 
  • Hands-on experience with CI/CD tools including Jenkins, Bamboo, GitLab, Harness, and Nexus. 
  • Experience developing automation using Bash, Shell, PowerShell, Python, or similar scripting languages. 
  • Strong experience with application monitoring and observability tools: (Splunk, AppDynamics, Dynatrace) 
  • Experience leading production deployments and release activities. 
  • Strong incident management and problem management experience. 
  • Experience with Jira, Remedy, or equivalent ITSM platforms. 
  • Experience working with cloud platforms: (AWS, GCP, PCF, Openshift /Kubernetes 
  • Knowledge of networking fundamentals, security principles, and distributed systems. 
  • Experience supporting databases such as MongoDB, Oracle, SQL Server, or PostgreSQL. 

Operational Skills 

  • Strong production troubleshooting and diagnostic capabilities. 
  • Ability to perform rapid impact assessments during critical incidents. 
  • Experience driving root cause analysis and operational improvements. 
  • Strong understanding of change management, release governance, and risk mitigation. 
  • Experience creating and maintaining operational documentation and runbooks. 

Soft Skills 

  • Excellent verbal and written communication skills. 
  • Strong stakeholder management and customer service orientation. 
  • Proven leadership and mentoring capabilities. 
  • Strong analytical and problem-solving skills. 
  • Ability to work independently in a fast-paced environment. 
  • Strong organizational and prioritization skills. 

Education 

  • Bachelor's Degree in Computer Science, Information Technology, Engineering, or related discipline.

Preferred Qualifications: 

  • Experience supporting messaging platforms and market data systems. 
  • Experience with Kafka, RabbitMQ, or other messaging technologies. 
  • Experience with infrastructure-as-code tools such as Terraform, Ansible, or Chef. 
  • Experience implementing SRE practices including: (Error Budgets, SLOs/SLIs, Capacity Planning, Chaos Engineering, Resilience Testing) 
  • Knowledge of GenAI-assisted operations, AIOps, and intelligent observability solutions. 
  • Experience with GitHub Copilot, automation frameworks, or AI-assisted troubleshooting. 
  • Experience building self-healing operational workflows. 
  • Familiarity with Agile/Scrum development methodologies.

In addition to the salary range, this role is also eligible for bonus or incentive opportunities


What’s in it for you

At Schwab, you’re empowered to shape your future. We champion your growth through meaningful work, continuous learning, and a culture of trust and collaboration—so you can build the skills to make a lasting impact. Our Hybrid Work and Flexibility approach balances our ongoing commitment to workplace flexibility, serving our clients, and our strong belief in the value of being together in person on a regular basis.

We offer a competitive benefits package that takes care of the whole you – both today and in the future:

  • 401(k) with company match and Employee stock purchase plan
  • Paid time for vacation, volunteering, and 28-day sabbatical after every 5 years of service for eligible positions
  • Paid parental leave and family building benefits
  • Tuition reimbursement
  • Health, dental, and vision insurance
Apply Now