About this opportunity
Oracle lists this Senior Network Developer (Network Automation, Network Operations) opportunity in nashville, Tennessee. Review the employer’s description below for duties, qualifications and application requirements.
Job description
hackajob is collaborating with Oracle to connect them with exceptional professionals for this role.
About the team
Join the AI Network Operations team to design, operate, and optimize advanced network systems supporting large-scale AI and cloud infrastructure. As an experienced Network Engineer, you will help operate high-performance RDMA network fabrics powering tier-0 customers in the generative AI industry, while partnering across engineering teams and vendors to ensure performance, reliability, and scalability.
Description
The OCI AI Infrastructure - Network Operations team operates the high-performance RDMA/RoCE network fabrics powering OCI's largest AI, GPU, and HPC workloads. As a Senior Network Engineer, you will:
* Operate, support, and scale RDMA/RoCE network fabrics across OCI's global cloud infrastructure.
* Design and validate advanced network solutions for large-scale AI, GPU, HPC, and data center environments.
* Apply deep networking and automation expertise to improve network reliability, performance, scalability, and operational efficiency.
* Develop automation, testing frameworks, and tooling to streamline network operations and proactively monitor network health.
* Build test strategies, perform pre-production validation, and drive root cause analysis (RCA) for complex network issues.
* Analyze network telemetry and performance metrics to identify anomalies, capacity constraints, and opportunities for improvement.
* Support incident response, customer escalations, and complex troubleshooting across production environments.
* Partner with internal engineering teams, vendors, and project teams to validate solutions and deliver network initiatives.
* Improve monitoring, anomaly detection, and operational tooling for frontline support teams.
* Mentor junior engineers and contribute to networking standards, operational readiness, and engineering best practices.
Responsibilities
Key Responsibilities
* Design, operate, validate, and scale advanced network fabrics for large-scale cloud, AI, and data center environments.
* Develop automation, scripts, and tooling to streamline network testing, operations, deployment, and troubleshooting.
* Build and enhance telemetry, dashboards, alerting, and monitoring to improve network health, reliability, and SLO performance.
* Develop test strategies, lead pre-production validation, and drive root cause analysis (RCA) for network issues and changes.
* Analyze network performance, capacity, latency, throughput, and packet loss to identify issues and support infrastructure growth.
* Participate in incident response and operational support, resolving complex production and customer issues.
* Partner with engineering teams, vendors, and stakeholders on network architecture, deployments, standards, and operational readiness.
* Identify design and operational risks, drive mitigations, and continuously improve network processes and reliability.
* Mentor engineers, provide technical guidance, and contribute to architecture, roadmap, and engineering best practices.
* Independently manage priorities and deliverables while collaborating across teams to achieve shared objectives.
Qualifications
Disclaimer:
Certain U.S. based or U.S. customer or client-facing roles may be required to comply with applicable requirements, such as immunization/occupational health mandates, and/or drug testing requirements.
Range and benefit information provided in this posting are specific to the stated locations only
US: Hiring Range in USD from: $91,400 to $187,000 per annum. May be eligible for bonus and equity.
Oracle maintains broad salary ranges for its roles in order to
Who can apply
Review the original listing for work authorization, qualifications and employer requirements.