Oracle

Senior Network Operations Engineer

Oracle Nashville, Tennessee, United States · $81K–$187K/yr

IT Services and IT Consulting · 10,001+ employees

Aug 31
Mid (2-5 yrs) Full-time United States
Create a free account to apply — email only, no card. You can also save this posting or score it against your profile with AI.

About the role

The role involves monitoring and troubleshooting network events, managing incident responses, and driving continuous operational improvement for OCI's global network. Engineers will also develop automation scripts to reduce manual effort and collaborate with partner teams to maintain infrastructure reliability.

What they look for

Network operations BGP OSPF TCP/IP Python Ansible SQL Puppet Incident management Root cause analysis Network automation Troubleshooting Cloud infrastructure Fiber optics Data center operations Vendor management

Requirements

Candidates must have 3–5 years of relevant network operations experience and strong knowledge of networking protocols like BGP, OSPF, and TCP/IP. Proficiency in at least three networking technologies (e.g., Juniper, Cisco, Arista) and experience in large-scale cloud or ISP environments are required.

Benefits

Medical insurance Dental insurance Vision insurance Short term disability Long term disability Life insurance Flexible spending accounts Pre-tax commuter and parking benefits 401(k) savings and investment plan Paid time off Paid holidays Paid sick leave Paid parental leave Adoption assistance Employee stock purchase plan Financial planning Group legal

Full description

Senior Network Operations Engineer

Location: Nashville, TN

NOTE: This position is not eligible for sponsorship.

In this role, you will monitor and troubleshoot network events, collect and analyze technical data, triage and mitigate incidents, coordinate escalations, and help drive continuous operational improvement. You will work alongside experienced engineers, partner teams, and vendors to maintain and optimize the infrastructure that supports Oracle customers and services worldwide.

Our mission is to keep OCI’s global network highly available, performant, and resilient—while delivering exceptional service to customers and dependable operational support to our engineering and technical teams.

For GNOC engineers, that mission translates into a broad, fast-moving role with real operational impact. You will help centrally manage OCI’s network infrastructure, respond to and resolve complex events, and develop automated solutions that reduce recurring operational work and improve reliability at scale.

NOC Operations

• Use established procedures and operational tooling to plan, implement, and safely complete network changes.

• Mentor, onboard, and train junior network engineers.

• Participate in operational rotations and provide break/fix and incident-response support.

• Identify and triage actionable incidents through monitoring systems; analyze and mitigate network events; conduct or support root-cause analysis (RCA); and coordinate follow-up actions with internal support teams and vendors.

• Provide on-call support as required, exercising sound independent judgment in a varied and complex operational environment.

• Participate in major incident calls and use technical and analytical skills to resolve network issues affecting Oracle customers and services.

• Manage fault detection, response, and escalation for OCI systems and networks, collaborating with third-party suppliers through resolution.

Leadership

• Collaborate with GNOC Shift Leads and management to ensure the efficient and timely completion of daily GNOC responsibilities.

• Lead, contribute to, and participate in the identification, development, and evaluation of projects and tools that improve GNOC effectiveness.

• Drive runbook audits and updates to maintain compliance and align operational processes with partner service teams.

• Conduct interviews and participate in hiring junior-level engineers.

• Lead and/or represent the GNOC in vendor meetings, service reviews, and governance boards.

Automation and Scripting

• Collaborate with network automation teams to integrate and improve operational support tooling.

• Develop scripts and automation to reduce manual effort and improve the reliability of routine operational tasks.

• Preferred experience with Python, Puppet, SQL, Ansible, network automation, and databases.

Project Delivery

• Lead technical initiatives, including the development and improvement of runbooks, methods of procedure (MOPs), operational processes, and team onboarding materials.

• Support the implementation of short-, medium-, and long-term plans to achieve project objectives.

• Regularly engage senior management and network leadership to ensure team priorities and project objectives are met.

TECHNICAL QUALIFICATIONS

Networking

• Strong knowledge of networking protocols and technologies, including BGP, OSPF, IS-IS, TCP/IP, IPv4/IPv6, DNS, DHCP, MPLS, VPNs, and TLS.

• Broad hands-on experience with at least three of the following: Juniper, Cisco, Arista, InfiniBand, NVIDIA, firewalls, routers, switches, circuit management, and optical/network transport services.

• Strong analytical skills, including the ability to gather, correlate, and interpret data from multiple sources.

• Ability to diagnose, prioritize, resolve, or appropriately escalate network alerts and faults.

• Experience in a large ISP, cloud provider, or similarly complex enterprise network environment.

• Exposure to commodity Ethernet hardware and networking ASICs, including Broadcom and NVIDIA/Mellanox.

• Cisco, Arista and Juniper certifications are desirable.

Circuit-operations

• Hands-on experience supporting carrier circuits and transport services in a 24×7 NOC, ISP, cloud-provider, data center, telecommunications, or large enterprise environment.

• Demonstrated experience troubleshooting fiber, optical, Ethernet, and WAN circuit failures across multiple carriers and vendors.

• Experience coordinating carrier escalations, field dispatches, remote hands, circuit turn-ups, maintenance windows, and service restoration.

• Working knowledge of optical power levels, transceivers, fiber paths, cross-connects, demarcation points, interface counters, and circuit-testing methods.

• Experience with DWDM, dark fiber, DIA, MPLS, VPLS, microwave, SONET, PON, or comparable transport technologies.

• Ability to correlate physical-layer and circuit conditions with routing adjacencies, packet loss, latency, congestion, and customer impact.

Strong incident-management, technical documentation, vendor-management, and root-cause-analysis skills.

GPU, RDMA, and HPC

• Experience supporting GPU and RDMA network environments is highly desirable.

• Experience supporting high-performance computing (HPC) environments is highly desirable.

• Experience with InfiniBand and NVIDIA networking technologies, including Spectrum, is highly desirable.

Network Design and Lifecycle Management

• Participate in network lifecycle management, including network build, refresh, and upgrade projects.

• Participate in network solution design and design-review activities.

SOFT SKILLS AND OTHER DESIRED EXPERIENCE

• Self-motivated, proactive, and able to work independently.

• Bachelor’s degree preferred, with at least 3–5 years of relevant network operations or engineering experience.

• Strong organizational, time-management, verbal, and written communication skills.

• Comfortable managing a broad range of priorities in a fast-paced operational environment.

• Experience with incident-response plans, processes, and strategies.

• Experience supporting large-scale enterprise infrastructure and cloud computing environments in a 24/7 network operations setting, including willingness to work rotational shifts.

Responsibilities

Key Responsibilities Capacity Ingestion and Management: -Takes proactive steps to design and architect infrastructure and/or service according to terms for reliability and functionality. -Forecasts demands for infrastructure and responds to capacity needs, ensuring systems have sufficient resources to handle current and future workloads. -Collaborates with the software development team to develop infrastructures and features that are reliable and scalable according to deployment requirements. -Independently identifies opportunities for and drives prototyping (e.g., testing new applications or infrastructures, assisting in onboarding). Incident and Service Lifecycle Management: -Performs data collection, triage, technical analysis, and redirection to maintain and optimize operations and infrastructure reliability. -Independently monitors services, maintains up-to-date knowledge of their performance, and documents their condition. -Leverages comprehensive knowledge to perform incident response, root cause analyses, and/or maintenance on assigned services (e.g., software installs, version upgrades, security updates, backup and recovery). -Provides health and performance reporting and takes appropriate actions based on trends in data. -May independently perform provisioning to support infrastructure, applications, and services. -May perform standard and non-standard decommissioning (e.g., shutting down servers, removing data from databases) to remove objects that are no longer needed. Automation: -Identifies opportunities for automation and assesses potential benefits. -Develops automation tools or scripts to provide solutions, gather metrics, monitor, analyze, mitigate, or remediate issues/defects within infrastructures. -Independently conducts testing to ensure automation performs the task correctly and produces expected results. Technical Communication and Guidance: -Communicates the scale, capacity, security, performance attributes, and requirements of services and technology within and sometimes beyond immediate team. -Identifies and explains the potential impact of infrastructure, feature, and tool changes, considering their impact on team operations. Troubleshooting and Resolution: -Provides operational support for technology, escalating incidents and other standard and non-standard issues arising within Oracle services. -Participates in on-call shifts to address issues. -Resolves technical issues spanning various services, investigating and debugging products in order to reach SLOs (service level objectives). -Documents incidents and performs root cause analyses according to standard reporting methods. -Independently performs post-mortem procedures to prevent incident reoccurrence. Innovation and Improvement: -Experiments with new tools and technologies to assess their potential impact on and improve infrastructure performance and reliability, ensuring adherence to security standards. -Independently identifies and executes improvements for performance bottlenecks and deployments to ensure efficient resource usage, speed, and scalability. -Develops knowledge of site reliability trends and shares new information with team members, management, and beyond to help others build, test, deploy and run services. -Performs standard and non-standard analyses and provides clear data on production to contribute to business development decisions (e.g., design changes).

Core Responsibilities Planning & Execution: Independently manages work, monitoring timelines and deliverables to ensure projects or initiatives stay on track and meet requirements. Proactively prioritizes work and adapts to resource or timeline shifts, suggesting adjustments to maintain project efficiency. Collaboration & Partnership: Collaborates across teams to align on expectations and achieve shared objectives. Builds and maintains a comprehensive understanding of business, stakeholder, and/or customer needs to build and support effective partnerships. Actively listens to diverse perspectives and asks questions to ensure understanding of others. Problem Solving: Independently identifies and addresses standard and non-standard issues in accordance with standard practices, escalating more complex issues as appropriate. Analyzes data and/or information from multiple sources to troubleshoot standard and non-standard errors. Contributes to knowledge sharing and best practices. Continuous Learning: Embraces continuous learning by actively seeking to build knowledge and new skills and/or tools and staying current with industry trends and best practices. Seeks out and leverages feedback and training to improve skills. Contributes to a culture of continuous learning and knowledge sharing with team members. Continuous Improvement: Develops ideas and recommends updates to increase the efficiency and effectiveness of processes, protocols, and workflows within a team. Seeks input from team members on alternative approaches and methods for improving work.

Qualifications

Disclaimer:

Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.

Range and benefit information provided in this posting are specific to the stated locations only

US: Hiring Range in USD from: $81,100 to $187,000 per annum. May be eligible for bonus and equity.

Oracle maintains broad salary ranges for its roles in order to account for variations in knowledge, skills, experience, market conditions and locations, as well as reflect Oracle's differing products, industries and lines of business. Candidates are typically placed into the range based on the preceding factors as well as internal peer equity.

Oracle US offers a comprehensive benefits package which includes the following: 1. Medical, dental, and vision insurance, including expert medical opinion 2. Short term disability and long term disability 3. Life insurance and AD&D 4. Supplemental life insurance (Employee/Spouse/Child) 5. Health care and dependent care Flexible Spending Accounts 6. Pre-tax commuter and parking benefits 7. 401(k) Savings and Investment Plan with company match 8. Paid time off: Flexible Vacation is provided to all eligible employees assigned to a salaried (non-overtime eligible) position. Accrued Vacation is provided to all other employees eligible for vacation benefits. For employees working at least 35 hours per week, the vacation accrual rate is 13 days annually for the first three years of employment and 18 days annually for subsequent years of employment. Vacation accrual is prorated for employees working between 20 and 34 hours per week. Employees working fewer than 20 hours per week are not eligible for vacation. 9. 11 paid holidays 10. Paid sick leave: 72 hours of paid sick leave upon date of hire. Refreshes each calendar year. Unused balance will carry over each year up to a maximum cap of 112 hours. 11. Paid parental leave 12. Adoption assistance 13. Employee Stock Purchase Plan 14. Financial planning and group legal 15. Voluntary benefits including auto, homeowner and pet insurance

The role will generally accept applications for at least three calendar days from the posting date or as long as the job remains posted.Career Level - IC3

Only Oracle brings together the data, infrastructure, applications, and expertise to power everything from industry innovations to life-saving care. And with AI embedded across our products and services, we help customers turn that promise into a better future for all. Discover your potential at a company leading the way in AI and cloud solutions that impact billions of lives.

True innovation starts when everyone is empowered to contribute. That’s why we’re committed to growing a workforce that promotes opportunities for all with competitive benefits that support our people with flexible medical, life insurance, and retirement options. We also encourage employees to give back to their communities through our volunteer programs.

We’re committed to including people with disabilities at all stages of the employment process. If you require accessibility assistance or accommodation for a disability at any point, let us know by emailing accommodation-request_mb@oracle.com or by calling 1-888-404-2494 in the United States.

Oracle is an Equal Employment Opportunity Employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, national origin, sexual orientation, gender identity, disability and protected veterans’ status, or any other characteristic protected by law. Oracle will consider for employment qualified applicants with arrest and conviction records pursuant to applicable law.