Senior DevOps Engineer

NVIDIA · South Bay

Company

NVIDIA

Location

South Bay

Type

Full Time

Job Description

We are looking for a Senior DevOps Engineer to join our Data and Application Services team to improve its growing services infrastructure. At the core of our application services platform is our multi-tenant Kubernetes platform that is designed to run a variety of inhouse application services. You will be working with a team of passionate and skilled engineers that are continuously working to provide better tools to build and manage this infrastructure. Our team is a mix of varying levels of experience and CS backgrounds. We need a motivated, hardworking and focused individual who has a real passion for operational excellence, data systems, and automation.

What you'll be doing:

  • Own the services you build working with cross functional teams

  • Comfortable with frequent code testing and deployment

  • Continuously improve infrastructure provisioning and management using automation

  • Identify areas to improve service resiliency through industry standard practices

  • Support a globally distributed, multi-cloud hybrid environment - AWS, GCP and On-prem

  • Determine root-cause for production level incidents and write corresponding high-quality RCA reports

  • Ensure the highest level of up-time and Quality of Service (QoS) to internal customers through operational excellence

  • Define service level objectives (SLOs) and service level indicators (SLIs) to represent and measure service quality

  • Participate in team's on-call rotation

What we need to see:

  • 7+ years in operating services including web servers, load balancers, relational/non-relational databases, messaging systems and storage solutions

  • 3+ years coding/scripting in at least two high level programming languages - Python, Go, Ruby, Groovy etc.

  • Deep understanding of linux operation system and TCP/IP fundamentals

  • Expertise with at least one major cloud service provider- AWS, GCP, Azure

  • Proficient in modern CI/CD techniques, GitOps and Infrastructure as Code(IaC)

  • Hands on experience running production quality observability stacks

  • Creative problem solver with excellent debugging skills

  • B.S. degree in Computer Science or related technical field (or equivalent experience)

  • Detail oriented with great communication and documentation skills

Ways to stand out from the crowd:

  • Linux certification from a well known vendor - RedHat, Oracle etc.

  • Prior experience managing large scale Kubernetes deployment in production

  • Strong skills in modern container networking and storage architecture

The base salary range is 164,000 USD - 310,500 USD. Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.

You will also be eligible for equity and benefits. NVIDIA accepts applications on an ongoing basis.

NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.

Apply Now

Date Posted

10/04/2024

Views

0

Back to Job Listings Add To Job List Company Profile View Company Reviews
Positive
Subjectivity Score: 0.8

Similar Jobs

Senior Developer, Data Engineer - Tarana Wireless, Inc.

Views in the last 30 days - 0

Tarana is seeking a Senior DeveloperData Engineer with 5 years of experience in building largescale data pipelines The role involves designing buildin...

View Details

Senior Front-End Software Engineer - Percipient.ai

Views in the last 30 days - 0

Percipientai founded in 2017 is a cuttingedge technology company specializing in Computer Vision Artificial Intelligence and Deep Learning They develo...

View Details

Senior Program Manager, Global Occupational Health & Safety - ServiceNow

Views in the last 30 days - 0

ServiceNow is seeking a Health Safety Program Manager to design implement and lead a comprehensive corporate safety program The role involves develop...

View Details

Staff Flight Test Engineer - Wisk

Views in the last 30 days - 0

Wisk Aero is seeking a Staff Flight Test Engineer to join their team in Hollister CA The role involves ensuring safe and efficient flight testing and ...

View Details

Staff Engineer, System Design Verification Engineering - Western Digital

Views in the last 30 days - 0

Western Digital is seeking a validation engineer to define and track test plans characterize and optimize SSDs and lead bug review meetings The ideal ...

View Details

Senior Finance Manager, Central FP&A - Palo Alto Networks

Views in the last 30 days - 0

Palo Alto Networks is seeking a Senior Finance Manager with 10 years of experience in FPA The role involves leading ad hoc projects collaborating with...

View Details