waypointjobs

NVIDIA

Senior Manager, Software Engineering - Agentic IT Operations

santa clara, CA

Check who can apply and the requirements below before continuing.

About this opportunity

NVIDIA lists this Senior Manager, Software Engineering - Agentic IT Operations opportunity in santa clara, California. Review the employer’s description below for duties, qualifications and application requirements.

Job description

For over 25 years, NVIDIA has been at the forefront of transforming computer graphics, PC gaming, and accelerated computing, driven by a legacy of continuous innovation and exceptional talent. We are now leveraging the immense potential of AI to usher in the next era of computing, where our GPUs power the "brains" of computers, robots, and autonomous vehicles that can comprehend the world! This pioneering work demands vision, innovation, and the world's best talent. Join our diverse and supportive environment, where NVIDIANs are inspired to excel and make a profound global impact.

We are seeking a hands-on technical leader to build and lead a high-performance engineering organization that architects, delivers, and operates production-grade software systems at global scale. You will be responsible for transforming enterprise IT operations from manual, reactive workflows into fully automated, AI-driven platforms that scale with NVIDIA’s hyper-growth. This role demands deep software engineering expertise, systems thinking, and the ability to drive large-scale technical transformation with measurable business outcomes. Exceptional interpersonal, written, and verbal communication skills are vital for success.

What You’ll Be Doing

Architect and ship agentic AI systems using LLM-based agents, tool calling, RAG, and orchestration frameworks delivering production-grade AI-assisted operations across enterprise IT domains including employee support, endpoint services, and IT support operations.

Design and deploy autonomous AI agents that execute complex, multi-step enterprise workflows end-to-end coordinating approvals, vendor handoffs, cross-system data reconciliation, and exception handling with human-in-the-loop controls delivering measurable improvements in availability, cycle time, cost, and compliance.

Engineer robust integration and automation platforms spanning ServiceNow, ERP and procurement systems, endpoint-management platforms, Own the full stack infrastructure, data pipelines, APIs, and user-facing applications.

Set the engineering standard through hands-on technical leadership co-authoring production code, conducting rigorous code reviews, and personally driving system design for the most critical components.

Recruit, develop, and retain top-tier engineering talent. Build a high-performing team culture grounded in engineering excellence, ownership, and continuous delivery.

Define and execute a multi-quarter technical roadmap for automation and agentic operations across enterprise IT, with each initiative tied to quantifiable business outcomes (cost reduction, throughput, SLA improvement, headcount avoidance).

Drive disciplined execution—project prioritization, milestone tracking, capacity planning, and on-time delivery—while maintaining engineering velocity in a fast-moving environment.

Own talent strategy for the team, including hiring pipelines, performance calibration, and career development that builds a deep bench of engineering leaders.

What We Need To See

Bachelor's or Master's degree in a related field, or equivalent experience

10+ overall years of hands-on software engineering experience, with deep expertise in at least one of: Infrastructure, SRE, DevOps, or Production Engineering. 5+ years leading engineering teams, with direct experience hiring, growing, and managing IT engineers.

Demonstrated ability to build engineering teams from zero and scale them in a high-growth, high-ambiguity environment.

Deep expertise in designing and shipping production software systems—including integrations, automation platforms, and data pipelines—for complex enterprise operations at scale.

Track record of modernizing enterprise IT operations platforms (e.g., asset management, endpoint services, IT supply chain, infrastructure operations) and deploying agentic AI into production—including multi-step autonomous execution, human-in-the-loop safeguards, exception handling, and governance frameworks with measurable business outcomes.

Production-grade proficiency with infrastructure-as-code, CI/CD, containerization (Kubernetes), and cloud platforms (AWS, GCP, or Azure).

Experience with monitoring and observability tools (Prometheus, Grafana, Datadog, PagerDuty, or similar).

Fluent in Python, Go, or equivalent languages—able to architect, write, and review production-quality code, not just scripts.

Executive-level communication skills with the ability to influence technical direction across engineering, product, and senior leadership.

Proven ability to translate complex technical capabilities into quantifiable business value and present to VP/C-level audiences.

NVIDIA is widely considered to be one of the technology world’s most desirable employers. We have some of the most forward-thinking and hardworking people in the world working for us. If you're creative and autonomous, we want to hear from you!

Your base salary will be determined based on your location, experience, and the pay of employees in similar positions. The base salary range is 248,000 USD - 391,000 USD.

You will also be eligible for equity and benefits.

Applications for this job will be accepted at least until September 1, 2026.

This posting is for an existing vacancy.

NVIDIA uses AI tools in its recruiting processes.

NVIDIA is committed to fostering an inclusive work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

#J-18808-Ljbffr

Worksite address

santa clara, CA, 95053, US

Who can apply

Review the original listing for work authorization, qualifications and employer requirements.

Ready for your next step?Apply on the official website
Apply on WhatJobs ↗

Explore related searches

Current related jobs

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Senior Embedded Systems Developer

laurel, MD

See pay details in description

Description Do you love working on a motivated team to solve complex problems in innovative ways? Do you enjoy creating embedded prototypes i…

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Oracle E-Business Suite (EBS) Developer

laurel, MD

See pay details in description

Description Are you a skilled problem-solver who blends technical expertise with creativity to build effective and efficient solutions? If s…

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Senior Software Engineer

laurel, MD

See pay details in description

Description Are you passionate about building solutions for our greatest national security challenges? Are you searching for engaging work wi…

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

Reverse Engineer and AI Workflow Developer

laurel, MD

See pay details in description

Description Are you passionate about reverse engineering complex software, firmware, and hardware? Do you enjoy developing agentic AI workflo…

Listing review due 2026-10-06View job

Johns Hopkins Applied Physics Laboratory (APL)

WhatJobs

AFSIM Developer

laurel, MD

See pay details in description

Description Are you passionate about building high-fidelity simulations that shape the future of U.S. Space Force capabilities and force design …

Listing review due 2026-10-06View job