Compute Software Architect
Oracle United States
IT Services and IT Consulting · 10,001+ employees
Applying here? Try the free cover letter tool — paste this posting and your résumé, no account needed.
About the role
The Software Architect will define and evolve the architecture of the core compute platform and control plane to ensure high scalability, reliability, and efficiency. They will provide technical leadership across distributed systems, drive platform modernization, and integrate AI to improve engineering and operational workflows.
What they look for
Requirements
Candidates must have 12+ years of experience in software engineering, distributed systems, or large-scale cloud infrastructure. A strong background in designing highly available services and influencing technical direction across organizational boundaries is required.
Benefits
Full description
Role Overview
The Software Architect, Compute Platform and Control Plane will be a senior individual contributor responsible for shaping the architecture and technical direction of the core software systems that power the Compute infrastructure platform. This role will work across the compute platform and control plane, spanning provisioning, scheduling, placement, lifecycle management, capacity orchestration, fleet management, service APIs, and the foundational distributed systems that support large-scale cloud infrastructure.The ideal candidate brings exceptional depth in software architecture, distributed systems, and cloud infrastructure, with the ability to reason across complex systems, identify structural weaknesses, simplify architectures, and define designs that can operate reliably at hyperscale.This architect will work closely with senior engineers, engineering leaders, product teams, and adjacent infrastructure organizations to evolve the compute platform to the next level of scalability, reliability, efficiency, and engineering velocity.
Responsibilities
Key Responsibilities
Compute Platform Architecture
- Define and evolve the architecture of the core compute software platform.
- Establish clear architectural principles, service boundaries, APIs, interfaces, and ownership models across compute platform
components.
Compute Control Plane Architecture
- Provide deep technical leadership across compute provisioning, scheduling and placement, capacity
discovery and allocation, instance and host lifecycle management, configuration and state management, fleet orchestration, health monitoring and remediation, failure recovery, service APIs, and regional isolation.
Distributed Systems Leadership
- Serve as a technical authority for distributed systems design. Drive architectural rigor in consistency,
concurrency, state management, partitioning, replication, idempotency, failure recovery, dependency management, and service isolation.
Cloud Infrastructure Architecture
- Apply deep cloud infrastructure expertise to the design and evolution of the compute platform. Understand
the end-to-end lifecycle of cloud compute capacity, from resource discovery and orchestration through provisioning, customer consumption, maintenance, failure recovery, and retirement.
Reliability and Resiliency
- Embed reliability and resilience into the architecture of the platform. Design for fault containment, graceful
degradation, recovery, retry safety, deployment safety, and failure isolation. Scalability and Performance
- Identify scaling limitations across the compute platform and control plane. Develop architectures that
improve throughput, latency, efficiency, and resource consumption while reducing unnecessary infrastructure overhead.
Platform Simplification and Modernization
- Identify legacy complexity, duplicated services, and architectural patterns that limit engineering velocity or
operational efficiency. Define pragmatic modernization strategies that allow critical systems to evolve without unnecessary disruption.
AI-Enabled Platform Transformation
- Help drive the use of AI to advance both the compute platform and software engineering practices. Identify
opportunities to apply AI to software development, testing, code analysis, debugging, operational diagnostics, incident analysis, anomaly detection, failure prediction, root-cause analysis, capacity optimization, and automated remediation.
Technical Leadership and Influence
- Operate as a senior technical leader across organizational boundaries. Lead complex architecture reviews,
drive alignment on critical technical decisions, mentor senior engineers, and partner with Distinguished Engineers, Architects, Directors, and Vice Presidents on long-term platform strategy.
Expected Impact
- A simpler, more coherent compute platform architecture.
- A highly scalable and resilient compute control plane.
- Clearer service boundaries and stronger platform APIs.
- Reduced architectural fragmentation and technical debt.
- Improved reliability and fault isolation.
- Better scalability and performance across critical services.
- Faster engineering velocity through reusable platform capabilities.
- More effective use of AI to improve engineering and operational workflows.
Minimum Qualifications
- 12+ years of experience in software engineering, distributed systems, cloud infrastructure, or large-scale
platform development.
- Deep expertise in software architecture and distributed systems.
- Strong experience designing and operating large-scale, highly available services.
- Significant experience with cloud infrastructure or large-scale infrastructure platforms.
- Strong understanding of control-plane architecture, service APIs, state management, orchestration, and
automation.
- Experience designing systems for scalability, reliability, fault tolerance, and operational efficiency.
- Ability to reason across large and complex software systems and identify architectural simplifications.
- Proven ability to influence technical direction across multiple engineering teams and organizations.
- Strong programming and systems fundamentals.
Preferred Qualifications
- Experience designing cloud compute platforms or large-scale infrastructure control planes.
- Deep experience with scheduling, placement, provisioning, resource management, or fleet orchestration
systems.
- Experience operating systems across multiple regions and failure domains.
- Experience modernizing large-scale infrastructure software while maintaining production stability.
- Strong understanding of cloud infrastructure economics and the relationship between architecture,
utilization, and cost.
- Experience applying AI-assisted development or AI-driven operational techniques in production
engineering environments.
- Experience working with GPU or accelerated compute infrastructure from a software platform
perspective.
- Bachelor's, Master's, or PhD degree in Computer Science, Computer Engineering, or a related technical
discipline.
Technical Profile
The ideal candidate is a hands-on software architect with deep technical credibility and broad systems perspective. They are comfortable going deep into distributed protocols, state machines, consistency models, service interfaces, failure modes, and performance bottlenecks, while also reasoning about the architecture of a very large cloud platform as a whole.Most importantly, they will combine deep software architecture expertise with strong cloud infrastructure judgment and the ability to influence the direction of critical compute systems across the organization.
Qualifications
Disclaimer:
Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.
Range and benefit information provided in this posting are specific to the stated locations only
US: Hiring Range in USD from: $169,800 to $355,400 per annum. May be eligible for bonus, equity, and compensation deferral.
Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business. Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.
Oracle US offers a comprehensive benefits package which includes the following: 1. Medical, dental, and vision insurance, including expert medical opinion 2. Short term disability and long term disability 3. Life insurance and AD&D 4. Supplemental life insurance (Employee/Spouse/Child) 5. Health care and dependent care Flexible Spending Accounts 6. Pre-tax commuter and parking benefits 7. 401(k) Savings and Investment Plan with company match 8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation. 9. 11 paid holidays 10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours. 11. Paid parental leave 12. Adoption assistance 13. Employee Stock Purchase Plan 14. Financial planning and group legal 15. Voluntary benefits including auto, homeowner and pet insurance
The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted. As part of Oracle's onboarding process and consistent with applicable law, US-based employees are required to complete identity verification, which involves the collection and processing of their biometric information. Accommodations to this requirement may be granted following an individualized assessment.
Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.
True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.
We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing accommodation-request_mb@oracle.com or by calling 1-888-404-2494 in the United States.
Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.
Similar roles
-
Senior Solutions Architect, Engagement Management - NVIS
NVIDIA Singapore, Singapore
-
Solutions Architect, AppleCare Technologies
Apple Sunnyvale, California, United States
-
Solutions Architect
MANTECH Doral, Florida, United States
-
Platform Solutions Architect
Booz Allen Hamilton Chantilly, Virginia, United States · $87K–$198K/yr
-
Senior Solutions Architect
Sprezzatura Management Consulting Arlington, Virginia, United States · $150K–$225K/yr
-
Solutions Architect - MFG - West
Databricks Calgary, Alberta, Canada · CA$185K–CA$225K/yr