GPU poor? Contact us for your AI cloud compute needs!
Senior Solutions Engineer
Location
Nevada
Posted
5 days ago
Salary
0
Seniority
Senior
Job Description
Senior Solutions Engineer
TensorWave
• Resolve Complex Escalations: Act as the final authority on issues exceeding GOC scope, utilizing code-level debugging and architectural investigation. • Direct Customer Engagement: Partner with customer technical leads to diagnose production issues, ensuring transparency and rapid resolution through active collaboration. • Iterative Problem Solving: Develop diagnostic scripts and workarounds to maintain customer operations while long-term patches are in development. • Drive Root Cause Analysis: Own end-to-end P1 resolution, partnering with TAMs to deliver clear, actionable post-incident analysis. • Bridge to Engineering: Convert recurring customer pain points into evidence-based feature requests, influencing product roadmap to resolve systemic failures. • Build Scalable Knowledge: Document non-obvious platform behaviors and refine GOC runbooks, ensuring institutional knowledge grows with every incident.
Job Requirements
- 5–9 years in Infrastructure Engineering, Platform Engineering, or SRE, with a specific focus on high-performance computing or large-scale AI stacks. Proven track record of managing complex production environments where system reliability is mission-critical.
- Kubernetes Expert: Deep experience in cluster administration and scheduler internals; comfortable reading/modifying controller code.
- AI/GPU Infrastructure Specialist: Proficient in orchestrating GPU workloads and diagnosing training job failures using ROCm or CUDA.
- Network Pathologist: Skilled in RDMA/RoCEv2, SRIOV, and BGP; capable of interpreting switch telemetry to identify silent packet drops.
- Linux Power User: Expert in kernel networking, hugepages, and cgroups; able to debug at the OS layer when applications are silent.
- Builder Mindset: Proficient in Python and Ansible; capable of writing custom diagnostic tools to automate remediation.
- Executive Communicator: Strong technical rigor when presenting findings to VPs of Engineering, maintaining trust while delivering difficult updates.
- Prior experience in a customer-facing engineering role (e.g., Solutions Engineering, Technical Support Engineering).
- Experience in high-uptime environments where 24/7/365 availability is required.
Benefits
- Stock Options
- 100% paid Medical, Dental, and Vision insurance for Employees
- Company Health Savings Account Contributions
- 100% paid Short Term and Long Term Disability Insurance for Employees
- Life and Voluntary Supplemental Insurance Options
- Other Insurance Options, such as Pet & Legal Insurance
- Various Supplementary Health Benefits, such as discounted Virtual Healthcare Appointments and Serious Illness Support
- Flexible Spending Account
- 401(k)
- Employee Assistance Program
- Flexible PTO
- Paid Holidays
- Parental Leave
- Other In-Office Perks
Related Guides
Related Categories
Related Job Pages
More Solutions Engineer Jobs
Manager 1 – Enterprise Managed Solutions
NBCUniversalNBCUniversal is a media and entertainment company that develops, produces, and markets a variety of entertainment and news programs internationally. NBCUniversa
• Responsible for effectively managing and monitoring all sales of integrated communication structure to enterprise customers • Assures optimal sales team staffing and training readiness of sales professionals • Develops financial and operational objectives
Forward Deployed AI Solutions Engineer
NateraFounded in 2004 and led by CEO Steve Chapman, Natera is a company in the biotechnology market that offers genetic testing and diagnostics on a global scale. Ope
Role Description The Forward Deployed Solutions Engineer will work directly within a business domain (e.g., Commercial, Clinical Operations, Lab Operations, Sales & Marketing, Customer Experience etc.). In your role, you’ll find opportunities for enhancing efficiency and productivity by looking for workflows which can be executed 10–100x faster or more often than a human team could using AI agents, integrations and other patterns. You will build, deploy, and run them in production. You report into the central AI & Automation team, partner directly with domain leadership on priorities, and bring patterns back so the whole company compounds. - Find the leverage in your domain - Map the workflows in your domain — the ones running today, and the ones that don’t exist yet because they weren’t feasible without agents or automation tools. - Identify the step-change opportunities: where AI, ML, or automation unlock throughput, coverage, or speed. - Build the business case, quantify projected impact, and align with domain leadership on priorities. - Design the future-state workflow - Map structured and unstructured data flows across the systems involved (CRM, ERP, ticketing, document stores, internal tools, external SaaS). - Define the target workflow: what the agent does, what the human does, and where they hand off. - Figure out what context the agent or model needs to do the work well — and how to get it there reliably (retrieval, grounding, tool access, memory). - Design human-in-the-loop checkpoints so review adds value without becoming the bottleneck. - Build and connect the systems - Stand up agents and automation pipelines using the organization’s approved AI platforms and frameworks. - Connect agents to business systems — via MCP servers, APIs, webhooks, CLIs, and skills — within the guardrails set by IT and security. - Configure tools, prompts, context, and retrieval pipelines so agents perform reliably on real work, not just in demos. - Handle integration gnarliness: auth, schema drift, rate limits, data quality, and the messy last-mile of enterprise systems. - Enable access and training for business to run the workflows - Run agents and automation pipelines in production. - Own agent performance end-to-end. Track the KPIs that matter — throughput, quality, cost, human intervention rate, cycle time, adoption. - Build and manage evals. Re-run them on any material model, data, or workflow change before it ships. - Triage failures, tune prompts and context, iterate on the workflow, and retire agents when they’re no longer the right tool. - Instrument observability: tracing, structured logs, dashboards. You don’t ship what you can’t see. Qualifications - Hands-on technical fluency. CLIs, APIs, webhooks, SQL, and Python scripting. - Working knowledge of LLM and agent behavior — prompting, context, tool use, RAG, MCP, evals, failure modes. - Be very comfortable with a cloud platform. - Trustworthy with elevated access. Least-privilege, auditability, and safe rollbacks are second nature. - Strong technical and process judgment. You think in outcomes and KPIs, can defend prioritization calls. - Comfortable being the most technical person in a business meeting and the most business-savvy in a technical one. Requirements - Prior experience working hand in hand with businesses to deliver measurable outcomes. - Hands-on experience with an enterprise agentic platform (CrewAI, LangChain, AWS Bedrock, Claude, Codex) or building directly against a model API. - Background in product management, solutions engineering, consulting, forward-deployed engineering, or technical operations. - Experience in regulated environments (HIPAA, SOC 2, GxP, SOX). Success in year one - Shipped three or more workflows into production that are measurably moving a business KPI, with agent evals and observability in place. - Domain leadership brings you into planning early, not late. - Contributed at least one reusable asset another engineer on the team is now using. - Enabling non-builders to become builders using the artifacts you created. Compensation Ranges - The base salary range for standard cost of living areas is: $105,700-$132,100. - Higher cost of living areas: $116,200 - $145,300. - Lower cost of living areas: $95,100-$118,900. - Additional components such as bonus and equity are also included in this role. - The pay range is listed and actual compensation packages are based on a wide array of factors unique to each candidate. Benefits - Competitive Benefits - Employee benefits include comprehensive medical, dental, vision, life and disability plans for eligible employees and their dependents. - Natera employees and their immediate families receive free testing in addition to fertility care benefits. - Other benefits include pregnancy and baby bonding leave, 401k benefits, commuter benefits and much more. - Generous employee referral program! Company Description Natera™ is a global leader in cell-free DNA (cfDNA) testing, dedicated to oncology, women’s health, and organ health. Our aim is to make personalized genetic testing and diagnostics part of the standard of care to protect health and enable earlier and more targeted interventions that lead to longer, healthier lives. The Natera team consists of highly dedicated statisticians, geneticists, doctors, laboratory scientists, business professionals, software engineers and many other professionals from world-class institutions, who care deeply for our work and each other.
• Build high-performing, scalable, enterprise-grade applications, while providing expertise in all areas of full stack development. • Analyze, configure, develop, test, document, and implement new and existing applications and solutions to deliver high quality, scalable solutions to meet business objectives and operational requirements. • Support application/architecture design, design patterns and development technology/standards. • Collaborate with team members and stakeholders to support business objectives and continual improvement of team dynamics. Identify areas for improvement in the development environment, best practices and team dynamics. • Participate in identification and resolution of application and technical issues. Identify and communicate lessons learned. • Adaptive to emerging technologies. Strong desire to learn and master new technologies. • Focused on delivery of high quality solutions and improvements in customer experience and operational excellence. Awareness of scalability and enterprise thinking. Awareness of applying emerging technologies to business problems.
Staff Forward Deployed AI Solutions Engineer
NateraFounded in 2004 and led by CEO Steve Chapman, Natera is a company in the biotechnology market that offers genetic testing and diagnostics on a global scale. Ope
Role Description The Staff Forward Deployed Solutions Engineer will work directly within a business domain (e.g., Commercial, Clinical Operations, Lab Operations, Sales & Marketing, Customer Experience etc.). In your role, you’ll find opportunities for enhancing efficiency and productivity by looking for workflows which can be executed 10–100x faster or more often than a human team could using AI agents, integrations and other patterns. You will build, deploy, and run them in production. You report into the central AI & Automation team, partner directly with domain leadership on priorities, and bring patterns back so the whole company compounds. - Map the workflows in your domain — the ones running today, and the ones that don’t exist yet because they weren’t feasible without agents or automation tools. - Identify the step-change opportunities: where AI, ML, or automation unlock throughput, coverage, or speed. - Build the business case, quantify projected impact, and align with domain leadership on priorities. Design the future-state workflow: - Map structured and unstructured data flows across the systems involved (CRM, ERP, ticketing, document stores, internal tools, external SaaS). - Define the target workflow: what the agent does, what the human does, and where they hand off. - Figure out what context the agent or model needs to do the work well — and how to get it there reliably (retrieval, grounding, tool access, memory). - Design human-in-the-loop checkpoints so review adds value without becoming the bottleneck. Build and connect the systems: - Stand up agents and automation pipelines using the organization’s approved AI platforms and frameworks. - Connect agents to business systems — via MCP servers, APIs, webhooks, CLIs, and skills — within the guardrails set by IT and security. - Configure tools, prompts, context, and retrieval pipelines so agents perform reliably on real work, not just in demos. - Handle integration gnarliness: auth, schema drift, rate limits, data quality, and the messy last-mile of enterprise systems. Enable access and training for business to run the workflows: - Run agents and automation pipelines in production. - Own agent performance end-to-end. Track the KPIs that matter — throughput, quality, cost, human intervention rate, cycle time, adoption. - Build and manage evals. Re-run them on any material model, data, or workflow change before it ships. - Triage failures, tune prompts and context, iterate on the workflow, and retire agents when they’re no longer the right tool. - Instrument observability: tracing, structured logs, dashboards. You don’t ship what you can’t see. Qualifications - Hands-on technical fluency. CLIs, APIs, webhooks, SQL, and Python scripting. - Working knowledge of LLM and agent behavior — prompting, context, tool use, RAG, MCP, evals, failure modes. - Be very comfortable with a cloud platform. - Trustworthy with elevated access. Least-privilege, auditability, and safe rollbacks are second nature. - Strong technical and process judgment. You think in outcomes and KPIs, can defend prioritization calls. - Comfortable being the most technical person in a business meeting and the most business-savvy in a technical one. Requirements - Prior experience working hand in hand with businesses to deliver measurable outcomes. - Hands-on experience with an enterprise agentic platform (CrewAI, LangChain, AWS Bedrock, Claude, Codex) or building directly against a model API. - Background in product management, solutions engineering, consulting, forward-deployed engineering, or technical operations. - Experience in regulated environments (HIPAA, SOC 2, GxP, SOX). Success in Year One - Shipped three or more workflows into production that are measurably moving a business KPI, with agent evals and observability in place. - Domain leadership brings you into planning early, not late. - Contributed at least one reusable asset another engineer on the team is now using. - Enabling non-builders to become builders using the artifacts you created. Compensation Ranges - The base salary range for standard cost of living areas is: $152,100-$190,100. - Higher cost of living areas: $167,300 - $209,100. - Lower cost of living areas: $136,900-$171,100. - Additional components such as bonus and equity are also included in this role. - The pay range is listed and actual compensation packages are based on a wide array of factors unique to each candidate. Benefits - Comprehensive medical, dental, vision, life and disability plans for eligible employees and their dependents. - Free testing for Natera employees and their immediate families in addition to fertility care benefits. - Pregnancy and baby bonding leave. - 401k benefits. - Commuter benefits. - A generous employee referral program. Company Description Natera™ is a global leader in cell-free DNA (cfDNA) testing, dedicated to oncology, women’s health, and organ health. Our aim is to make personalized genetic testing and diagnostics part of the standard of care to protect health and enable earlier and more targeted interventions that lead to longer, healthier lives. The Natera team consists of highly dedicated statisticians, geneticists, doctors, laboratory scientists, business professionals, software engineers and many other professionals from world-class institutions, who care deeply for our work and each other. When you join Natera, you’ll work hard and grow quickly.



