Found Description
Elevate your career as a Cloud Site Reliability Engineer with NVIDIA. Ensure our Digital Marketing Services run reliably and efficiently while leveraging AWS Infrastructure.
NVIDIA seeks a skilled Site Reliability Engineer to enhance the performance and reliability of our Digital Marketing ecosystem. This role, requiring 3+ years of experience, focuses on automation, monitoring, and alerting within AWS Infrastructure. You will be responsible for deploying new applications, managing incidents, and improving deployment pipelines in a high-paced environment.
Key Responsibilities: • Identify and categorize user-reported problems swiftly • On-board new applications and AI/ML services • Implement monitoring and alerting systems • Automate daily tasks and ML CI/CD deployment • Provide on-call support and resolve urgent incidents
Requirements: • MS or BS in Computer Science/Engineering or equivalent experience • 3+ years experience in live-site production environment...
NVIDIA seeks a skilled Site Reliability Engineer to enhance the performance and reliability of our Digital Marketing ecosystem. This role, requiring 3+ years of experience, focuses on automation, monitoring, and alerting within AWS Infrastructure. You will be responsible for deploying new applications, managing incidents, and improving deployment pipelines in a high-paced environment.
Key Responsibilities: • Identify and categorize user-reported problems swiftly • On-board new applications and AI/ML services • Implement monitoring and alerting systems • Automate daily tasks and ML CI/CD deployment • Provide on-call support and resolve urgent incidents
Requirements: • MS or BS in Computer Science/Engineering or equivalent experience • 3+ years experience in live-site production environment...
Ready to Apply?
Submit your application for Cloud Site Reliability Engineer at NVIDIA at NVIDIA
Apply Now