Senior Engineer, Build and DevOps
Location
United States
Posted
3 days ago
Salary
$184K - $356.5K / year
Seniority
Senior
No structured requirement data.
Job Description
Senior Engineer, Build and DevOps
NVIDIA
Role Description NVIDIA is looking for a hardworking member of NVIDIA’s Analytics and Data Intelligence Engineering Operations team, supporting multiple engineering teams working on data science (and adjacent) libraries such as RAPIDS. As a DevOps Engineer, you’ll have the opportunity to help support and grow the RAPIDS project. You will work closely with RAPIDS build and development teams to ensure high-quality releases of CUDA/C++ and Python libraries as well as containers. What you’ll be doing: - Work in a team of DevOps engineers supporting multiple software projects in the data science and AI domain, many of them open source. - Manage cutting-edge hardware and help inform purchasing decisions for the team. - Collaborate with build engineers, developers, and management to ensure the delivery of high-quality software. - Develop and modernize packages, such as streamlined Python wheels, for RAPIDS data science libraries. - Design and maintain container build processes. - Take a hands-on approach working with engineers on the team to implement DevOps best practices. - Execute on a range of DevOps initiatives including CI/CD, observability, security/legal compliance, and SysAdmin tasks. - Operate and maintain our infrastructure and development processes. Qualifications - Bachelor of Science in Computer Engineering, Computer Science or related technical field or equivalent experience. - 8+ years of technical experience primarily related to DevOps. - Proven experience in programming and automation with scripting languages (Bash and Python preferred). - Experience with Conda and/or PyPI packaging, especially building and publishing. - Experience with container technologies such as Docker, especially building and publishing. - Detail-oriented and comfortable supporting and prioritizing amongst multiple teams. - Experience with administration, optimization, and troubleshooting of CI/CD and related tools (including Jenkins, Git, GitHub Actions). - You have worked with cloud services (AWS, Azure, and others), especially permissions, budget, and cost management. - Linux system administration experience (Ubuntu strongly preferred). Requirements - Background with NVIDIA’s technology stack, including CUDA toolkit and drivers. - Experience in software development, build, and/or related DevOps. - Experience with GitHub operations, including user, repository, and organization management and permissions. - Prior work with open-source development and community building on GitHub. - Strong verbal and written communication skills. Benefits - Competitive salaries. - Generous benefits package. - Base salary range is 184,000 USD - 287,500 USD for Level 4, and 224,000 USD - 356,500 USD for Level 5. - Eligible for equity.
Related Guides
Related Categories
Related Job Pages
More DevOps Engineer Jobs
Site Reliability Developer
Vena SolutionsTake your entire business from reactive to proactive with the leading AI-Powered Complete FP&A Platform.
• Support key ITIL processes, including Incident management, request management, problem management and change management. • Define and document runbooks and standard operating procedures. • Field operational requests from our Application Support team and other internal stakeholders • Triage and solve issues within defined SLA’s to ensure an excellent customer experience and to unblock other development and support teams • Maintain services once they are live by measuring and monitoring availability, latency and overall system health. • Identify and troubleshoot problems, investigate root causes, and champion fixes across the organization. • Work with infrastructure-as-Code (IaC) with a focus on continuous improvement. • Collaborate with cross-functional team members on features and implementation within an agile environment. • Report on SLAs and performance metrics as part of the Operations function. • Participate in on-call rotation.
• proaktywne dostrzeganie wyzwań i ich adresowanie wspólnie z zespołami backendowymi i devops • tworzenie szablonów i standardów infrastrukturalnych dla najczęstszych potrzeb zespołów developerskich (np. nowe serwisy, projekty Cloud Run) • rozwój i utrzymanie stacku observability (Loki, Grafana, Prometheus) oraz optymalizacja wykorzystania zasobów i kosztów chmury (GCP) • utrzymanie i rozwój wewnętrznych narzędzi self-hosted (m.in. NocoDB, n8n, Outline) • wsparcie infrastrukturalne zespołów Data Engineering i Data Science (m.in. Airflow, pipeline'y forecastingowe czy serwery pod ML) • zarządzanie infrastrukturą sieciową w darkstore'ach (sieć, CCTV, drukarki paragonów, urządzenia handheld) we współpracy z IT Managerem • budowanie i utrzymanie infrastruktury on-prem (serwery pod ML i workloady wymagające dużych zasobów, runnery CI/CD dla projektów mobilnych) • utrzymanie i usprawnianie procesów CI/CD (GitLab) oraz developer experience • budowanie infrastruktury pod rozwiązania AI (autonomiczne agenty, boty Slack, agenty code review, remote coding agents) • rozwój praktyk SRE: alerting, procesy zgłaszania i obsługi incydentów (PagerDuty, Slack)
Senior DevOps Engineer
MKS2 TechnologiesAustin-based SDVOSB delivering application development, cybersecurity, instructional design and training to DOD and VA.
• Conduct DevOps and DevSecOps activities for an Azure integration platform • Create and maintain GitHub CI/CD pipelines • Conduct Azure DevOps activities • Conduct Azure system administration • Build platform automation • Create and maintain PowerShell scripting • Understanding and provisioning of Azure Infrastructure services and network • Create and maintain containers and infrastructure (IaC, Terraform) • Deploy solutions through CI/CD pipelines • Embed DevOps best practices and process/tech improvements across the team and platform • Leverage AI tools to accelerate activities • Collaborate across technical teams • Create and maintain detailed technical documentation • Support troubleshooting and resolution of Production issues • Participate in Agile ceremonies • Report on progress and status of development
Senior DevOps Engineer
FiservFounded in 1984, Fiserv is a global provider of ecommerce and information management systems for the financial services industry. In 2013, Fiserv acquired Open
• Design, implement, and support secure Azure IaaS and PaaS solutions across enterprise environments • Build and maintain CI/CD pipelines using Azure DevOps and GitLab to enable efficient, automated software delivery • Develop and manage Infrastructure as Code solutions using Terraform and Ansible for provisioning and configuration management • Implement and maintain Zero Trust security architectures, including identity and access management (IAM), data encryption, and security controls • Identify, assess, and remediate infrastructure and application security vulnerabilities • Create and maintain monitoring, alerting, and operational dashboards using Azure Monitor, Log Analytics, and Application Insights • Troubleshoot production issues, restore services, and maintain operational runbooks and knowledgebase documentation • Collaborate with cross-functional teams to support cloud modernization initiatives




