Sustainable Talent
Site Reliability Engineer
About this role
Nvidia is seeking a Site Reliability Engineer for a contract role in Santa Clara, CA, focusing on infrastructure planning and process support. This position involves maintaining cloud services and ensuring the efficiency of automated tasks for various NVIDIA software teams.
What you'll do
- Monitor and recover assets in a private cloud environment
- Stabilize virtualization infrastructure (ESXi, KVM, Hyper-V)
- Deploy and maintain machine farms using automation tools (Chef, Ansible, Terraform)
- Provide on-call L1 support for infrastructure issues
- Analyze and debug OS, networking, and performance problems
- Assist in deploying infrastructure configurations for NVIDIA technologies
What they're looking for
- KVM and ESXi virtualization
- Configuration management (Chef, Ansible, Terraform)
- Cloud services management
- On-call support experience
- Debugging and troubleshooting skills
- Understanding of networking concepts
- Experience with NVIDIA GPUs and Tegra Processors
- Familiarity with performance monitoring tools
Benefits
- Competitive pay ($65/hr - $85/hr)
- Full benefits
- Paid Time Off (PTO)
- Supportive company culture
- [unknown]
- [unknown]
Opens the application — the Jobs AI extension fills it for you. Set up autofill
Opens the official application on the employer’s site. No login required.
Sustainable Talent
Sustainable Talent appears to connect technical professionals with contract and permanent roles at major technology companies, focusing on infrastructure, systems, and driver validation work. The company is hiring for Site Reliability Engineers, Windows System Validation Engineers, and Windows Driver Validation Engineers who support cloud infrastructure, system quality assurance, and GPU driver performance across various platforms.
View all jobs at Sustainable Talent