Job Closed
This listing is no longer active.
Building simple, effective government services. Want to contribute? We're hiring!
Senior Infrastructure Engineer
Location
Alabama + 30 moreAll locations: Alabama | Arizona | California | Colorado | Connecticut | District Of Columbia | Florida | Illinois | Louisiana | Maine | Nevada | New Jersey | New York | North Carolina | Ohio | Oklahoma | Oregon | Maryland | Massachusetts | Michigan | Minnesota | Missouri | Pennsylvania | Rhode Island | South Carolina | Tennessee | Texas | Utah | Virginia | Washington | Wisconsin
Posted
6 days ago
Salary
$127.8K - $153K / year
Seniority
Senior
Job Description
Senior Infrastructure Engineer
Nava
• The Infrastructure Engineer will work on cross functional teams to build scalable infrastructure for our government -- designing, implementing, and delivering services that millions of Americans depend on. • You may be responsible for technical leadership of platform infrastructure projects supporting dev teams and large-scale systems with a strong focus on automation, organization, and communication. • Your work enables application development teams to run effectively in the cloud to deliver users a modern experience. • You will be responsible for delivering on client requirements as well as providing long-term oversight and vision to help shape future work. • Your strong cross-functional communication skills will complement your technical leadership skills to provide the highest value for government systems.
Job Requirements
- Minimum of 5 years of experience as an Infrastructure Engineer or mix of Software Engineering
- A Bachelor's Degree (in any discipline) or four years of experience in lieu of a degree
- Experience managing different systems (e.g Windows Server, Linux CLI) and automating repetitive tasks
- Experience training and mentoring engineers
- Scripting (python, Go, Javascript)
- Experience designing CI/CD and other workflow automation spanning multiple systems
- Incident handling and followup (root cause analyses, etc)
- Knowledge of cloud primitives (VPC, ACL, Route Tables, Security Groups, Images)
- Experience running asynchronous jobs, e.g. using services like ECS or Lambda
- Ability to use the command line to achieve practical aims, e.g. navigating the shell; moving and downloading files from external servers; searching and parsing files for information; and running queries
- Advanced ability to contribute and update existing code and/or Infrastructure as Code
- A thoughtful, adaptive, and collaborative mindset/understanding of networks, protocols, and multi-service architectures
- Excellent written and verbal communication skills, technical and otherwise
- An interest in and familiarity with one or more of the following areas: Cloud Infrastructure, cloud architecture, IT service management, Unix/Linux, scripting, networking, or security
- Able to follow and establish, when needed, standard operating procedures.
Benefits
- Health coverage — comprehensive medical, dental, and vision plans to support your overall health needs
- Insurance coverage — Nava provides disability, life, and accidental death insurance at no cost
- Time off — vacation, holidays (including Juneteenth), and floating holidays to rest and recharge
- Company holidays — enjoy 12 paid federal holidays each year on top of your regular PTO
- Annual bonus — when Nava meets its goals, eligible employees receive a performance-based annual bonus
- Parental leave — paid time off for new parents, plus weekly meals delivered to your home
- Wellness program — full platform offering physical, mental, & emotional health resources & support tools
- Virtual care — see doctors online with no copay through a virtual visit program (depending on available providers)
- Sabbatical leave — earn extended unpaid leave after continuous service for personal growth or rest
- 401(k) match — Nava matches 4% of your salary to support your retirement savings plan
- Flexible work — remote-first environment with flexibility built around your schedule and responsibilities
- Home office setup — company laptop & setup assistance provided via Staples for remote work needs
- Utility support — monthly reimbursement to help offset eligible home office utility expenses
- Learning opportunities — internal training programs and resources to help grow your professional skills
- Development opportunities — LinkedIn Learning access & an annual allowance for courses, tuition, & certs
- Referral bonus — get rewarded when you refer great people who join the Nava team
- Commuter benefits — pre-tax commuter programs to support in-office travel when applicable
- Supportive culture — A collaborative and remote-friendly team environment where people genuinely care
Related Guides
Related Categories
Related Job Pages
More Infrastructure Engineer Jobs
Associate, IT Infrastructure
MajescoMajesco is a leading insurance solutions and services provider. Software for core insurance functions include Policy Administration, Underwriting, New Business Processing, Billing, Claims, Product Modeling, Incentive Compensation, and Producer Life cycle Management. Offers consulting and insurance-specific IT services for testing, data conversion, data-warehousing/BI, mobility, enterprise integration, and BPM. Specializes in connecting people and business to insurance in innovative, hyper-relevant, compelling, and personal ways. Helps insurers modernize, innovate and connect to build the future of their business and the industry at speed and at scale.
Role Description We are seeking an experienced IT Service Management professional with strong expertise in ServiceNow ITSM, Major Incident Management, Service Desk Operations, and continuous service improvement. The role is responsible for driving ITSM best practices across Incident, Problem, and Change Management, ensuring service performance through KPI governance, CSAT management, reporting, and process optimization. The ideal candidate will bring strong stakeholder management, analytical capability, operational governance, and a proven ability to improve service quality and business outcomes. Key Responsibilities - IT Service Management Operations - Manage and support ITSM processes in ServiceNow in alignment with ITIL best practices. - Drive Incident, Problem, and Change Management processes to ensure service stability and operational excellence. - Ensure timely service restoration and resolution of critical incidents through effective coordination with support teams. - Support service desk operations by monitoring service quality, response effectiveness, and process adherence. - Problem Management - Lead Problem Management activities with a focus on identifying root causes and implementing permanent corrective and preventive actions. - Review recurring incidents and critical issues to raise problem records for detailed analysis. - Coordinate with technical teams for RCA documentation and service improvement actions. - Track repeat issues and recommend long-term solutions to reduce incident volume. - Major Incident Management - Manage major incidents end to end, including logging, classification, prioritization, escalation, and communication. - Drive bridge calls and coordinate cross-functional teams to restore services within agreed SLAs. - Ensure timely stakeholder communication and status updates to internal teams and senior management. - Follow escalation matrices and governance protocols during critical service disruptions. - Service Improvement and Governance - Analyze incident, service request, and service catalog data to identify trends and improvement opportunities. - Provide actionable insights for process enhancement, service quality improvement, and operational efficiency. - Monitor and track KPIs such as response SLA, resolution SLA, First Time Resolution, ticket backlog, and CSAT score. - Prepare dashboards, MIS reports, and governance review reports for management and stakeholders. - Quality, Knowledge, and Training - Oversee CSAT metrics and quality audit activities to improve user experience and service performance. - Support and contribute to internal training initiatives for analysts and support teams. - Create, update, and maintain knowledge base articles and Standard Operating Procedures. - Promote operational consistency through documentation and knowledge-sharing practices. - ServiceNow and Process Support - Support SLA configuration and User Acceptance Testing in ServiceNow. - Ensure service management workflows and configurations align with business requirements. - Contribute to process design, compliance checks, and service governance activities. - Budget and Operational Support - Support budget tracking by maintaining purchase invoice updates and operational expenditure records. - Assist with process compliance, audits, and governance reporting. - Maintain basic oversight of IMAC, asset management inventory, and related operational controls. Qualifications - Bachelor of Engineering in Electronics and Telecommunication Engineering or related field. - ITIL v4 Foundation certification preferred. - Additional training or certifications in IT support, security awareness, or AI tools will be an advantage. Requirements - 20+ years of experience in IT Service Management, Service Desk Operations, and Major Incident Management. - Proven experience in service governance, process improvement, KPI management, and stakeholder reporting. - Experience working with enterprise support environments and cross-functional IT teams. Preferred Candidate Profile - ITIL-aligned service management professional with a process-driven mindset. - Strong problem solver with the ability to identify trends and drive preventive actions. - Confident communicator who can engage with senior stakeholders during major incidents and governance reviews. - Detail-oriented professional with strong ownership of reporting, quality, and operational metrics. - Capable of balancing day-to-day service operations with strategic service improvement initiatives.
AI Infrastructure Engineer
MirantisStrategic open source infrastructure for containers and virtual machines.
• Manage and operate production AI infrastructure environments. • Lead incident response and troubleshooting efforts and deliver timely service restoration during outages or performance degradations. • Troubleshoot infrastructure and networking issues across bare-metal and/or cloud environments with multiple vendors. • Conduct root cause analysis and drive product and operational improvements. • Contribute and improve operational documentation and knowledge base. • Collaborate with global team members across time zones to ensure continuous operational coverage, including occasional work during weekends and holidays.
Security Engineer – Infrastructure, Client Platform
Reliable Robotics CorporationExpanding access to more places with automated aviation
• Design, implementation and monitoring of security tools supporting the business’ internal needs and safety requirements. • Enforce security configurations on all endpoints (macOS, Windows, iOS, Android) utilizing appropriate tools. • Design and maintain manageable access control schemes, manage SaaS AAA, and provide zero-trust designs for SSO. • Meet the automation requirements of the team by utilizing infrastructure as code systems. • Perform Log Analysis, participate in threat hunting exercises and investigate security alerts from EDR or externally provided IOCs. • Collaborate with teams to design robust systems that are easy to maintain.
Systems Engineer II – Shared Infrastructure Data Center
WashU ITWashington University in St. Louis Information Technology
• Perform operational duties and capacity planning efforts for enterprise systems. • Create and maintain documentation related to services, solutions and interfaces. • Perform testing and changes while minimizing risk through disciplined change management practices. • Perform research, analyze technology, consult vendors and apply best practices to design technical solutions by utilizing systems analysis techniques and procedures, including consulting with users, to determine hardware, software or system functional specifications. • Validate, test and implement new products and services. • Respond to and resolve incidents escalated from operational teams and performance tuning requests utilizing critical thinking skills. • Provide technical and advisory leadership as required to complete objectives. • Train and mentor for other personnel. • Perform other duties as assigned.



