Charles Schwab Inc.

Manager, Software Development & Engineering

Charles Schwab Inc. Austin, Texas, United States · $133K–$224K/yr

Financial Services · 10,001+ employees

4 d ago
Senior (5-10 yrs) Full-time United States
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

Responsible for production support, Site Reliability Engineering (SRE), and maintaining observability for enterprise applications and platform services. Manages incident response, root-cause analysis, and automation of deployment workflows to ensure system stability and performance.

What they look for

Site Reliability Engineering Splunk AppDynamics Python Shell Scripting CI/CD GitHub PCF Linux Unix Kafka RabbitMQ Solace Kubernetes Agile Kanban

Requirements

Requires a Bachelor’s degree in Computer Science or a related field and at least 60 months of progressive experience in SRE and application support. Candidates must demonstrate proficiency in cloud platforms, containerization, messaging systems, and automation scripting.

Benefits

401(k) with company match Employee stock purchase plan Paid vacation Volunteering 28-day sabbatical Paid parental leave Family building benefits Tuition reimbursement Health insurance Dental insurance Vision insurance

Full description

Your Opportunity

At Schwab, you’re empowered to make an impact on your career. Here, innovative thought meets creative problem solving, helping us “challenge the status quo” and transform the finance industry together.

Job Duties: Responsible for production support and Site Reliability Engineering (SRE) for enterprise applications and platform services. Monitors production systems using Splunk and AppDynamics for log analysis and application performance monitoring. Responds to production incidents, performs root-cause analysis (RCA), and coordinates resolution with application, infrastructure, security, and third party teams. Creates and maintains monitoring alerts, dashboards, and operational runbooks to improve observability and reduce service disruptions. Designs and implements automation to reduce manual operational work, including deployment and support workflows using GitHub pipelines, Python, Shell scripting, and Microsoft Power Automate. Optimizes CI/CD pipelines using Harness, Bitbucket, Bamboo, and Nexus to streamline deployments and ensure SLA and SLO compliance. Provides production support for the Solace messaging platform, troubleshooting messaging and integration issues for application consumers. Deploys applications to PCF and validates service bindings, service accounts, and access permissions. Performs post maintenance validation following Windows and Unix/Linux server patching to ensure application stability. Executes application and platform disaster recovery activities, including DR testing and recovery verification. Performs annual password rotations and certificate renewals in compliance with security and audit policies. Investigates production data issues and coordinates cross team remediation. Partners with infrastructure and security teams to implement server vulnerability remediation. Participates in 24/7 on call rotations to ensure timely incident response and resolution. Mentors junior engineers on SRE best practices and applies Agile and Kanban methodologies for operational and reliability initiatives.

Starting compensation for this location depends on related experience. Annual bonus opportunity and other eligible earnings are not included in the range(s) above. We offer a competitive benefits package, see below for details.

What you have

Job Requirements: Requires Bachelor’s in Computer Science, Information Technology, or a related field and 60 months of progressive, post-Bachelor’s experience in a related occupation.

Experience must include 60 months of experience involving the following: Supporting and operating middle tier applications hosted on PCF, including integration with messaging and event driven platforms Kafka, RabbitMQ, and Solace; Troubleshooting and monitoring using Splunk and AppDynamics, source control via GitHub, and incident and work item management using Remedy and JIRA; Supporting applications in Linux/Unix systems; Developing automation solutions in at least one programming language (Python, PowerShell, or similar); Supporting applications in cloud platforms (AWS, GCP, or PCF) and containerization (Kubernetes); and Observability tools (monitoring, logging, and tracing) including Splunk and AppDynamics.

We offer competitive pay and benefits. Starting compensation depends on related experience. Annual bonus and other eligible earnings are not included in the ranges above. Benefits include: 401(k) w/ company match; employee stock purchase plan; paid vacation, volunteering, 28-day sabbatical after every 5 years of service for eligible positions; paid parental leave and family building benefits; tuition reimbursement; health, dental, and vision insurance; hybrid/remote work schedule available for eligible positions (subject to Schwab’s internal approach to workplace flexibility).