Senior DevOps Engineer at Nvidia
Austin, TX 73301
About the Job
We are looking for a Senior DevOps Engineer to join our Data and Application Services team to improve its growing services infrastructure
At the core of our application services platform is our multi-tenant Kubernetes platform that is designed to run a variety of inhouse application services
You will be working with a team of passionate and skilled engineers that are continuously working to provide better tools to build and manage this infrastructure
Our team is a mix of varying levels of experience and CS backgrounds
We need a motivated, hardworking and focused individual who has a real passion for operational excellence, data systems, and automation.What you'll be doing:Own the services you build working with cross functional teamsComfortable with frequent code testing and deploymentContinuously improve infrastructure provisioning and management using automationIdentify areas to improve service resiliency through industry standard practicesSupport a globally distributed, multi-cloud hybrid environment - AWS, GCP and On-premDetermine root-cause for production level incidents and write corresponding high-quality RCA reportsEnsure the highest level of up-time and Quality of Service (QoS) to internal customers through operational excellenceDefine service level objectives (SLOs) and service level indicators (SLIs) to represent and measure service qualityParticipate in team's on-call rotationWhat we need to see:7+ years in operating services including web servers, load balancers, relational/non-relational databases, messaging systems and storage solutions3+ years coding/scripting in at least two high level programming languages - Python, Go, Ruby, Groovy etc.Deep understanding of linux operation system and TCP/IP fundamentalsExpertise with at least one major cloud service provider- AWS, GCP, AzureProficient in modern CI/CD techniques, GitOps and Infrastructure as Code(IaC)Hands on experience running production quality observability stacksCreative problem solver with excellent debugging skillsB.S
degree in Computer Science or related technical field (or equivalent experience)Detail oriented with great communication and documentation skillsWays to stand out from the crowd:Linux certification from a well known vendor - RedHat, Oracle etc.Prior experience managing large scale Kubernetes deployment in productionStrong skills in modern container networking and storage architectureThe base salary range is 164,000 USD - 310,500 USD
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.You will also be eligible for equity and benefits
NVIDIA accepts applications on an ongoing basis
NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer
As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.SummaryLocation: US, CA, Santa Clara; US, MA, Westford; US, TX, Austin; US, NC, DurhamType: Full time
At the core of our application services platform is our multi-tenant Kubernetes platform that is designed to run a variety of inhouse application services
You will be working with a team of passionate and skilled engineers that are continuously working to provide better tools to build and manage this infrastructure
Our team is a mix of varying levels of experience and CS backgrounds
We need a motivated, hardworking and focused individual who has a real passion for operational excellence, data systems, and automation.What you'll be doing:Own the services you build working with cross functional teamsComfortable with frequent code testing and deploymentContinuously improve infrastructure provisioning and management using automationIdentify areas to improve service resiliency through industry standard practicesSupport a globally distributed, multi-cloud hybrid environment - AWS, GCP and On-premDetermine root-cause for production level incidents and write corresponding high-quality RCA reportsEnsure the highest level of up-time and Quality of Service (QoS) to internal customers through operational excellenceDefine service level objectives (SLOs) and service level indicators (SLIs) to represent and measure service qualityParticipate in team's on-call rotationWhat we need to see:7+ years in operating services including web servers, load balancers, relational/non-relational databases, messaging systems and storage solutions3+ years coding/scripting in at least two high level programming languages - Python, Go, Ruby, Groovy etc.Deep understanding of linux operation system and TCP/IP fundamentalsExpertise with at least one major cloud service provider- AWS, GCP, AzureProficient in modern CI/CD techniques, GitOps and Infrastructure as Code(IaC)Hands on experience running production quality observability stacksCreative problem solver with excellent debugging skillsB.S
degree in Computer Science or related technical field (or equivalent experience)Detail oriented with great communication and documentation skillsWays to stand out from the crowd:Linux certification from a well known vendor - RedHat, Oracle etc.Prior experience managing large scale Kubernetes deployment in productionStrong skills in modern container networking and storage architectureThe base salary range is 164,000 USD - 310,500 USD
Your base salary will be determined based on your location, experience, and the pay of employees in similar positions.You will also be eligible for equity and benefits
NVIDIA accepts applications on an ongoing basis
NVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer
As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.SummaryLocation: US, CA, Santa Clara; US, MA, Westford; US, TX, Austin; US, NC, DurhamType: Full time