Art of the possible.
Infrastructure Engineer
Location
United States
Posted
5 days ago
Salary
$164.4K - $201.3K / year
Seniority
Lead
Job Description
Infrastructure Engineer
General Dynamics Information Technology
• Design, implement, document, and maintain enterprise infrastructure solutions • Manage highly available and secure enterprise IT systems supporting mission-critical operations • Install, configure, and maintain virtualization platforms, operating systems, applications, and supporting infrastructure • Troubleshoot complex issues involving compute, networking, storage, virtualization, and enterprise applications while identifying opportunities for optimization • Develop and maintain infrastructure standards, architecture documentation, and Standard Operating Procedures (SOPs) • Plan and execute infrastructure upgrades, maintenance activities, and technology refresh initiatives • Provide technical leadership and strategic guidance to engineering teams and senior management • Conduct technical assessments and architecture reviews to validate designs and implementation approaches • Coordinate and implement environment changes through established Request for Change (RFC) processes • Collaborate with cybersecurity teams to obtain and maintain system accreditation packages, including security checklists, STIG compliance, diagrams, and required documentation • Participate in cybersecurity activities such as vulnerability assessments, environment scans, penetration testing, and security audits • Mentor junior engineers and promote infrastructure best practices across the organization • Participate in technical meetings with government customers, civilian personnel, and contractor teams
Job Requirements
- Bachelor's degree in Computer Science, Information Technology, or related field (or equivalent experience)
- 10+ years of related experience
- Certification: DoD 8570 IAT Level II Required
- Cloud & Identity Microsoft Azure
- Microsoft 365 (M365) administration and services
- Microsoft Entra ID (Azure Active Directory)
- Azure Infrastructure as a Service (IaaS)
- Azure Platform as a Service (PaaS)
- Compute & Virtualization Windows Server administration
- Azure Virtual Desktop (AVD)
- Virtual host deployment and management
- High availability and disaster recovery design
- Security Azure Key Vault
- Azure Disk Encryption / Disk Encryption Sets
- System hardening
- Vulnerability remediation
- DoD Cloud Computing Security Requirements Guide (CC SRG)
- Security Technical Implementation Guides (STIGs)
- Networking Azure Virtual Networks
- Network Security Groups (NSGs)
- Application Security Groups (ASGs)
- Azure Virtual Network Gateway
- Azure Firewall
- Infrastructure Management Group Policy (GPO) management
- Enterprise IT operations Performance optimization
- Capacity planning
- Infrastructure monitoring and maintenance
- Enterprise backup and disaster recovery planning
- Additional Desired Skills Infrastructure as Code (IaC) using Terraform
- Policy as Code (PaC) Microsoft365DSC
- PowerShell automation and scripting
- Microsoft Graph API and Microsoft 365 automation
- Microsoft 365 administration and configuration
- Red Hat Enterprise Linux administration, hardening, and automation
- Power Platform
- Power BI Enterprise automation and orchestration solutions
Benefits
- Comprehensive benefits and wellness packages
- 401K with company match
- Competitive pay and paid time off
- Full flex work weeks where possible
- Variety of paid time off plans including vacation, sick and personal time
- Paid parental, military, bereavement and jury duty leave
- 15 days of paid leave per calendar year
- 10 paid holidays per year
- GDIT Paid Family Leave program providing up to 160 hours of paid leave in a rolling 12 month period
- Short and long-term disability benefits
- Life, accidental death and dismemberment, personal accident, critical illness and business travel and accident insurance
Related Guides
Related Categories
Related Job Pages
More Infrastructure Engineer Jobs
• AI Platform Engineering: Design, deploy, and operate enterprise AI/ML platforms. • Build self-service platforms for Data Scientists and ML Engineers. • Deploy and operate Kubeflow, MLflow, KServe, Ray, or similar AI platforms. • Design infrastructure supporting model training, experimentation, feature engineering, and inference. • Build highly available and scalable model serving infrastructure. • GPU Infrastructure: Design and operate GPU clusters for large-scale AI workloads. • Optimize GPU scheduling, utilization, sharing, autoscaling, and resource allocation. • Deploy and manage NVIDIA GPU Operator and GPU-enabled Kubernetes environments. • Optimize distributed GPU training performance across multi-node clusters. • Troubleshoot AI infrastructure performance bottlenecks. • MLOps & Platform Automation: Build CI/CD pipelines for ML workloads. • Automate AI infrastructure provisioning using Infrastructure as Code. • Implement monitoring and observability for GPU utilization, model serving, training jobs, and inference latency. • Collaborate closely with Data Science teams to improve platform usability, performance, and reliability.
Infrastructure Engineer
Voltage ParkVoltage Park is an equal opportunity employer and makes employment decisions on the basis of merit. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, protected veteran status, or any other characteristic under federal, state, or local law. If you require an accommodation during the job application process, please notify your recruiter.
Role Description Voltage Park is seeking an Infrastructure Engineer with a focus on Observability to join our Infrastructure Engineering team. Our engineers design and operate the systems that manage thousands of bare-metal servers, GPUs, and high-performance networks across multiple data centers. This role combines the breadth of a core infrastructure engineer with a specialty in observability and telemetry. You’ll design and operate metrics, logs, traces, and alerting pipelines that provide actionable insights for both internal teams and external customers — helping to ensure reliability and transparency at scale. This is a fully remote position, although candidates must be based in the continental United States. Unfortunately, we are unable to provide sponsorship for this role. Responsibilities - Design, build, and maintain observability platforms spanning metrics, logs, traces, and events. - Create dashboards and alerting for internal stakeholders (InfraOps, Engineering, Customer Success) and scoped visibility for external customers. - Ingest and correlate telemetry from GPUs, CPUs, networking (Ethernet & InfiniBand), containers, APIs, and BMC/Redfish. - Implement noise-resistant alerting pipelines that improve detection and reduce operational load. - Collaborate with infrastructure, platform, and customer-facing teams to embed observability into workflows. - Contribute to broader infrastructure engineering projects beyond observability. Qualifications - 8+ years in infrastructure engineering, SRE, or observability roles. - Strong experience with monitoring systems (Prometheus, Grafana, ELK, VictoriaMetrics, or similar). - Proficiency in Python, Go, or bash for automation and data integration. - Familiarity with container/Kubernetes observability. - Understanding of streaming telemetry pipelines (Kafka, OTEL, Promtail, or equivalent). - Strong written and verbal communication skills. Ideal Experiences - Experience with GPU observability, particularly NVIDIA DCGM. - Designing multi-tenant observability solutions with RBAC and scoped queries. - Prior work with correlation engines for RCA, forecasting, or predictive alerting. - Broader exposure to infrastructure domains (networking, storage, provisioning). Culture - You enjoy working with a small, highly motivated team. - You’re comfortable balancing autonomy with company-wide priorities. - You value clarity, documentation, and actionable insights in observability systems. - You’re excited to specialize in observability while contributing as a core infrastructure engineer. Company Description Voltage Park is an equal opportunity employer and makes employment decisions on the basis of merit. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, protected veteran status, or any other characteristic under federal, state, or local law. If you require an accommodation during the job application process, please notify your recruiter.
• Diseñar, desarrollar y administrar soluciones en Microsoft Azure. • Implementar arquitecturas escalables, seguras y de alto rendimiento. • Colaborar en equipos multiculturales para el desarrollo de soluciones innovadoras. • Participar activamente en la mejora de procesos y buenas prácticas.
Infrastructure Storage Engineer
Bright Vision TechnologiesBright Vision Technologies is a forward-thinking software development company dedicated to building innovative solutions that help businesses automate and optimize their operations. We leverage cutting-edge technologies to create scalable, secure, and user-friendly applications.
Role Description We are seeking an Infrastructure Storage Engineer with deep expertise across enterprise storage platforms — NetApp ONTAP, Pure Storage, and Ceph — to design, deploy, and operate the storage foundation that supports our compute, virtualization, database, and Kubernetes workloads. The role spans block, file, and object storage across data center, edge, and cloud environments, with strong attention to performance, data protection, and cost. The ideal candidate has operated heterogeneous storage estates at scale, understands the storage characteristics of diverse workloads, and brings strong automation discipline to storage engineering work. Key Responsibilities - Design and operate enterprise storage platforms across NetApp ONTAP, Pure Storage FlashArray/FlashBlade, and Ceph. - Implement and manage SAN, NAS, and object storage across data center and cloud environments. - Build storage solutions for VMware, Kubernetes (CSI drivers), and database workloads. - Design and operate replication, snapshot, and backup strategies for critical data assets. - Implement and operate cloud-tiered storage with FabricPool, CloudVolumes, or equivalent. - Build storage automation using Ansible, Terraform, REST APIs, and platform-native automation. - Operate Ceph clusters including OSD lifecycle, CRUSH maps, RGW, RBD, and CephFS. - Design and operate storage performance management for high-throughput and latency-sensitive workloads. - Implement data protection strategies including ransomware protection and immutable snapshots. - Build storage observability across capacity, performance, and health dimensions. - Lead storage upgrades, firmware lifecycle management, and security patching across the storage estate. - Drive storage cost optimization including deduplication, compression, and tiering strategies. - Troubleshoot complex storage issues spanning array, network, and host layers. - Stay current with storage industry developments and emerging platforms. Qualifications - Bachelor’s degree in Computer Science, Information Systems, or a related field. - Five or more years of enterprise storage engineering experience. - Deep expertise in NetApp ONTAP or Pure Storage FlashArray/FlashBlade. - Hands-on experience with Ceph at production scale. - Strong understanding of SAN, NAS, and object storage protocols. - Experience with storage integration for VMware and Kubernetes (CSI). - Strong scripting and automation skills using Ansible, Terraform, or REST APIs. - Strong understanding of storage performance, capacity planning, and cost optimization. - Strong troubleshooting skills across storage, network, and host layers. - Excellent communication and collaboration skills. Preferred Qualifications - Vendor certifications (NCDA, NCIE, Pure Storage certifications). - Experience with hybrid-cloud storage replication. - Familiarity with software-defined storage platforms beyond Ceph. - Experience with object storage at petabyte scale. - Exposure to AI/ML workloads with extreme I/O requirements. How to Apply Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] or contact us at (908) 505-3544. Learn more about Bright Vision Technologies at www.bvteck.com .


