Charles Schwab Inc.

Sr. Site Reliability Engineer

Charles Schwab Inc. Southlake, Texas, United States · $120K–$159K/yr

Financial Services · 10,001+ employees

20 h ago Closes in 4d
sre Senior (5-10 yrs) Full-time United States
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

The Senior Site Reliability Engineer will lead SRE tasks, focusing on high availability, automation, and the support of complex application platforms and microservices. They will also drive proactive system monitoring, incident remediation, and collaborate with cross-functional teams to ensure operational excellence.

What they look for

Linux Java Cloud Engineering SRE Automation Terraform CI/CD Incident Management Production Support System Monitoring Microservices Database Management PaaS IaaS Git Atlassian

Requirements

Candidates must have extensive experience in Linux/Java software development, architecture, and operations, along with a proven track record in leading technology teams. Proficiency in cloud infrastructure, automation tools like Terraform, and incident management is required for this role.

Benefits

Bonus opportunities Incentive opportunities

Full description

Your Opportunity

At Schwab, you’re empowered to make an impact on your career. Here, innovative thought meets creative problem solving, helping us challenge the status quo and transform the finance industry together.

Charles Schwab’s Technology Services Cloud Engineers thrive in a leading-edge work culture while focusing on products that help Schwab customers learn, explore, and make life-impacting moves on their paths to achieving their goals. This position requires a self-motivated individual with strong problem-solving skills who can contribute to a highly collaborative culture and a team environment and deliver innovative, value-based reliable solutions. A proven track record of delivering high quality technology products and services in a hyper-growth environment where priorities shift quickly is key to success in the role.

This Senior Cloud Engineering role is a hand-on technical leadership role and will be responsible to drive the teams’ workings on high availability, maintainability, support, and automaton of complex application platforms and microservices.

  • Lead the execution of the SRE tasks for the Cross Enterprise organizations.
  • Communicate strategic technology plans to associated business partners ensuring collaboration across technology and the business.
  • Regularly interact with senior technology leaders, product owners and business partners.
  • Lead team on complex issues where analysis of situations or data requires in-depth knowledge of the applications and environment.
  • Partner with highly experienced technologists, creating successful plans to deliver solutions to the business, ensure day-to-day support, high availability, and process compliance.
  • Lead Operational delivery and Production support focusing on proactive monitoring, rapid response Platform SRE
  • Perform proactive daily system monitoring including reviewing system and application logs as well as responding to, triaging, troubleshooting and remediating incidents.
  • Repair and recover from failures. Coordinate and communicate with impacted stakeholders and clients, escalating where appropriate.
  • Monitor and troubleshoot issues across the entire stack - software, application, and network.
  • Develop automation and processes to enable teams to deploy, manage, configure, scale, and monitor their applications.
  • Identify applications reliability and availability improvements, establish, and build solutions to continue to drive an improved experience.
  • Develop and manage continuous deployment and integrate solutions.
  • Create and review documentation and process regarding recurring issues, new standard operating procedures, knowledge transfer material, etc.
  • Collaborate with Engineering, Scrum and Ops resources to provide technical expertise and support on key initiatives for system availability and reliability.
  • Create and Host Gameday exercises for customers to achieve operational excellence.

What you have

  • Experience in Linux/Java Software Development & Architecture, Operations, DevOps, etc.
  • Experience leading or managing technology teams, aligning strategy with execution and mentoring/coaching individuals' performance.
  • Demonstrated ability to resolve business and service impacting problems, evaluating all alternatives, and consulting with other technical members of the organization is required.
  • Familiarity with database management systems (Oracle, SQL)
  • Knowledge of Platform as a Service (PaaS) and Infrastructure as a Service (IaaS)
  • Experience with Continuous Integration/Continuous Delivery (Bamboo, Go or other related tools)
  • Experience with Git, JIRA and related Atlassian stack
  • Experience with environment provisioning and deployment automation (Salt/Chef/Puppet)
  • Ability to work with global teams.
  • Flexibility to operate in an environment with changing demands and priorities.
  • Experience with Terraform is required.
  • Experience with OnCall for Production Support and Incident Management

In addition to the base pay range, this role is also eligible for bonus or incentive opportunities.

Similar roles