Pioneering AI-first solutions, solving complex business challenges through expertise, cloud, data engineering, and AI.
Cloud Infrastructure Architect
Location
Massachusetts
Posted
3 days ago
Salary
0
Seniority
Lead
Job Description
Cloud Infrastructure Architect
Quantiphi
• Lead the delivery of advanced solutions in Application Development, Platform Engineering, Data Engineering, Analytics, and Machine Learning. • Experience in workload migration from on-premise datacenter environments to AWS. • Enable clients to leverage the full potential of cloud capabilities. • Collaborate with cross-functional teams to deliver high-quality solutions.
Job Requirements
- 10+ years of experience in Cloud Infrastructure Architecture
- AWS Cloud Migration (EC2, EBS, VPCs, Migration Hub, MGN)
- VMware environment assessment & migration planning
- Strong AWS architecture, security, and compliance knowledge
- Understanding of different Migration approaches with recent experience
- Tools like Cloud Endure, Double-Take, AWS Migration Hub, AWS Application Migration Service (MGN)
- Good understanding of Windows or Linux OS
- Relevant certifications (e.g., AWS Certified Solutions Architect, AWS Certified DevOps Engineer)
- Experience working in an Agile and DevOps environment.
Benefits
- Join one of the world’s fastest-growing AI-first digital engineering companies and make a real impact at scale.
- Lead and collaborate with a high-energy team of talented, driven individuals solving complex, meaningful challenges.
- Work with Fortune 500 companies and disruptive innovators in a research-driven environment with 60+ patents.
- Stay ahead of the curve by gaining hands-on experience with cutting-edge AI, ML, data, and cloud technologies while continuously upskilling.
Related Guides
Related Categories
Related Job Pages
More Infrastructure Engineer Jobs
Senior Platform Engineer, Network Infrastructure
NVIDIANVIDIA is committed to fostering a diverse work environment and proud to be an equal opportunity employer. As we highly value diversity in our current and future employees, we do not discriminate (including in our hiring and promotion practices) on the basis of race, religion, color, national origin, gender, gender expression, sexual orientation, age, marital status, veteran status, disability status or any other characteristic protected by law.
• Design, build, and operate the Kubernetes platform that powers GNI network automation, telemetry, and operations across data center, colocation, and cloud environments. • Own the lifecycle management for GNI Kubernetes environments, including cluster onboarding, upgrades, capacity, availability, and recovery. • Develop production-quality software and automation for cluster provisioning, validation, upgrades, remediation, and safe multi-cluster delivery through GitOps. • Provide production support for network services hosted on the platform, working with Network Automation and service teams that retain ownership of application architecture, code, and features. • Diagnose complex Kubernetes platform and hosted-service failures involving control-plane health, cluster networking, storage, scheduling, workload placement, and multi-cluster dependencies. • Drive issues from initial signal through verified resolution. • Define production-readiness and observability standards for the platform and hosted network services, including health signals, capacity, alerts, runbooks, and recovery. • Participate in CFR’s production on-call rotation, including scheduled after-hours and weekend coverage. • Lead incident response and recovery, then drive corrective actions to completion.
IT Infrastructure Engineer – Bridge Coordinator
New Era TechnologyNew Era Technology is a global technology service provider and a trusted advisor to over 14,000 customers worldwide. The company provides services and solutions
• Manage the planning, scheduling and coordination of all client installations including but not limited to software, hardware, server migrations and other projects • Planning, scheduling, coordinating and communicating all elements of delivering site and non-site installations and projects to a high level of client satisfaction • Communicating with all stakeholders including external clients, account managers and engineers to ensure the efficient delivery of projects and installations • Coordinating and communicating the correct allocation of skilled staff within the time constraints to complete projects within client expectations • Providing status reports on all active projects including any risks associated with completing projects on time and on budget as part of a regular reporting cycle • Accurately reflecting engineers service calls in service delivery calendar and effectively communicates client instructions • Effectively communicate schedule changes to all relevant stakeholders • Accurately recording relevant data in all systems • Developing and maintaining effective working relationships with personnel from all departments • Demonstrating and upholding exceptional safety standards at all times in accordance with any workplace health and safety requirements, to ensure your own safety and the safety of others • Other duties as required
• Evolving the AI knowledge platform - taking the retrieval, indexing, and synthesis layer (currently semantic RAG + re-ranking + HyDE) to an organization-wide platform serving both internal engineering tools and customer-facing capabilities. • Architecting and operating agentic infrastructure on AWS - multi-step, tool-using AI systems that plan, retrieve, and act on complex queries and operational events, with cost guardrails and observability built in from day one. • Designing and building graph-based, relationship-aware retrieval across the organization's data sources, enabling multi-hop queries and letting agents accumulate organizational knowledge over time. This is on our roadmap, not in production - you will define the approach. • Partnering with product engineering to define the AI platform API surface, translating infrastructure primitives into developer-ready abstractions. • Building reference agent implementations on the platform - operational-incident triage, customer support, and future agentic use cases - grounding each agent's reasoning in institutional knowledge. • Owning the AI infrastructure cost model: monitoring compute, model, and storage spend, flagging anomalies, and proposing guardrails to keep workloads within defined budgets.
• Work as part of the Infrastructure team on sprints to evolve core platform and systems • Provide infrastructure and CI/CD expertise to wider engineering teams • Automate build and deployment of software to cloud platform • Eliminate repetitive manual operations • Improve monitoring and alerting systems • Participate in on-call rotation and assist with troubleshooting




