Cogniify

We are an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, or any other protected characteristic.

Principal AI/ML Engineer

Location

United States

Posted

95 days ago

Salary

$135.9K - $213.1K / year

Seniority

Lead

No structured requirement data.

Job Description

Principal AI/ML Engineer

Cogniify

Role Description We are seeking a Principal AI/ML Engineer to serve as a technical leader and architect for our organization’s AI/ML strategy, systems, and platforms. In this role, you will define the long-term technical vision for machine learning and MLOps, drive cross-team alignment on architecture and standards, solve the hardest technical problems, and ensure our ML capabilities are enterprise-grade, scalable, and forward-looking. The ideal candidate is a recognized technical authority who combines deep hands-on expertise with strategic thinking, organizational influence, and the ability to elevate the entire engineering team. Key Responsibilities - Define and own the long-term technical vision, architecture, and roadmap for the organization’s AI/ML platform and MLOps capabilities. - Serve as the senior technical authority on ML system design, providing guidance on architecture, tooling, and engineering standards across the organization. - Lead the design of foundational ML infrastructure including training platforms, feature stores, model registries, serving systems, and observability frameworks. - Establish and evangelize best practices for the full ML lifecycle: experimentation, development, testing, deployment, monitoring, governance, and retirement. - Drive cross-functional alignment between ML engineering, data engineering, platform engineering, product, and security teams. - Evaluate emerging technologies, frameworks, and research to inform strategic decisions on AI/ML tooling and approaches. - Solve complex, ambiguous, and high-impact technical problems that span multiple teams or systems. - Mentor and develop senior and staff-level engineers, fostering a culture of technical excellence and continuous improvement. - Represent the engineering organization in discussions with leadership, partners, and clients on AI/ML capabilities and strategy. - Define and enforce governance, compliance, and responsible AI practices across ML systems. - Drive cost optimization and efficiency across ML training, serving, and infrastructure. - Contribute to hiring, technical assessments, and the growth of the AI/ML engineering team. Qualifications - Master’s or PhD in Computer Science, Mathematics, Statistics, or a related field preferred. Equivalent professional experience is accepted. - 10+ years of professional experience in software engineering, ML engineering, or applied AI, with at least 5 years focused on production ML systems. - Deep expertise in designing and scaling production ML platforms, pipelines, and infrastructure. - Authoritative knowledge of MLOps principles, tools, and practices including CI/CD for ML, automated retraining, model governance, and observability. - Extensive experience with cloud-native ML services across AWS, Azure, or GCP, and infrastructure-as-code tools (Terraform, Pulumi, CloudFormation). - Strong understanding of distributed systems, data architecture, and scalable platform design. - Proven track record of driving technical strategy and influencing engineering direction at an organizational level. - Experience leading and mentoring senior engineers and building high-performing technical teams. - Exceptional communication skills with the ability to align technical and business stakeholders on complex topics. - Strong understanding of responsible AI, model fairness, explainability, security, and compliance requirements. Preferred Qualifications - Experience architecting LLM-based systems, RAG pipelines, agentic AI platforms, or conversational AI at enterprise scale. - Experience with MCP/tool-layer integrations for LLM-driven systems. - Deep familiarity with feature platforms, model serving at scale, and real-time inference architectures. - Track record of contributing to or leading industry standards, publications, or open-source projects in AI/ML. - Experience with enterprise AI governance frameworks and regulatory compliance (SOC2, HIPAA, GDPR). - Experience working across geographically distributed or offshore engineering teams. - Background in financial services, healthcare, or other highly regulated industries. Salary Range - US East/West Coast: $159,800 - $213,100 - US Remote: $135,900 - $181,200 Benefits - Unlimited PTO. - Very generous parental leave, much above industry standards. - Entrepreneurial culture where pushing limits and taking risks is everyday business. - Open communication with management and company leadership. - Small, dynamic teams = massive impact. - Medical, Dental and Vision coverage for employees. - Access to Disability & Life insurance. - Mental health and wellbeing support. - Annual bonus program. - Employer Stock Purchase Program (ESPP). - Yearly Team building experiences. - Mentorship and sponsorship opportunities. - Manager resources and support. Company Description We are an equal opportunity employer. We celebrate diversity and are committed to creating an inclusive environment for all employees. All qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, national origin, disability, veteran status, or any other protected characteristic.

Related Job Pages

More Machine Learning Engineer Jobs

Full TimeRemoteTeam 1,001-5,000H1B No Sponsor

• Work across the Campaign Monitor product to identify valuable opportunities in product and customer data, and turn them into predictive features that improve customer outcomes • Turn rich historical product and customer data into predictive features that improve customer outcomes • Identify high-impact opportunities for applied machine learning by analyzing product, behavioral, and content data, and translating ambiguous product questions into concrete ML use cases • Develop and deploy predictive machine learning models, including models for click-through rate, churn, recommendations, and related engagement signals • Design and build features and training datasets from structured product data, historical behavioral data, and content-derived signals • Own the applied ML lifecycle from data exploration and feature engineering through training, evaluation, deployment, monitoring, and iteration • Build production services and workflows for batch and real-time inference, with a pragmatic focus on reliability, maintainability, and speed to impact • Work hands-on in the codebase, contributing to backend systems and product workflows that consume predictions and recommendations • Partner closely with product, design, and engineering to turn customer needs into ML-driven product capabilities with measurable business impact • Establish pragmatic best practices for model evaluation, experimentation, monitoring, and continuous improvement • Help shape how applied machine learning is introduced into the product, while aligning with broader engineering architecture and delivery practices • Contribute to shared knowledge across the engineering organization to improve understanding and adoption of applied ML over time.

Australia
Job Closed
WillHire logo

Software Development Engineer – MLOps

WillHire

Now Magnit - Follow our new LinkedIn account https://www.linkedin.com/company/magnitglobal

Full TimeRemoteTeam 51-200H1B No Sponsor

• Design, implement, and deliver highly scalable features for our AI platform. • Partner with Agent Developers, Data Scientists, and other Software Engineers to create the technology that brings these features to life. • Support contracts with the U.S. Federal Government requiring US citizenship for personnel.

Colorado
$130.4K - $195.6K / year
Job Closed
Salesforce logo

Machine Learning Engineer

Salesforce

👋 We're Salesforce, the customer company. CRM + Data + AI + Trust.

Full TimeRemoteTeam 10,001+Since 1999H1B Sponsor

• Build next-generation agentic AI platforms • Work with Data Scientists, Software Engineers, product managers, and other stakeholders to design, implement, and iterate agentic AI systems with customers • Innovate at the frontier of the field, having the opportunity to create new solutions and define new categories of products with meaningful impact to Salesforce customers and beyond.

India
Job Closed
Full TimeRemoteTeam 10,001+Since 1982H1B Sponsor

Role Description Photoshop ART is seeking a Senior Machine Learning (ML) Systems & Efficiency Engineer to join our R&D team focused on delivering practical, production-ready improvements in inference performance, latency, and cost efficiency across image editing applications. This role sits at the intersection of model architecture, systems, inference runtimes, and services, with a clear mandate: deliver high-quality ML systems at substantially lower cost and higher efficiency. Individuals in this role are expected to have deep expertise in areas such as Artificial Intelligence (AI), ML systems, and computer vision. Strong preference will be given to candidates with experience in distributed inference, multimodal model profiling, and performance optimization. You will work closely with research, product, and infrastructure teams to influence model design decisions, improve GPU utilization, and build scalable, cost-aware ML systems deployed in production. This is a hands-on, high-leverage role where a single engineer can drive outsized impact, potentially saving millions of dollars in compute costs. The ideal candidate will have a strong interest in developing practical innovations that advance Adobe products. Qualifications - Master’s or PhD in Computer Science, Electrical Engineering, or a related field, with a focus on machine learning systems, distributed systems, or high-performance computing. - Hands-on experience implementing and scaling large-scale inference or serving workloads using distributed frameworks and runtime systems (e.g., Triton, vLLM, SGLang, xDiT, or similar). - Experience applying inference compilation and optimization tools (e.g., TensorRT, ONNX Runtime, AOTI), including techniques such as operator fusion and graph-level optimization. - Strong understanding of GPU architecture (e.g., memory hierarchy, compute throughput, communication bandwidth) and practical experience diagnosing performance bottlenecks across compute, memory, and I/O subsystems. - Proficiency in Python and C++, with experience building high-performance or distributed systems. - Familiarity with CUDA or Triton for performance-critical workloads is highly desirable. - Demonstrated ability to make engineering decisions based on rigorous measurement and benchmarking, with a focus on improving system efficiency, scalability, and reliability in production environments. Requirements - Design and optimize high-throughput, low-latency inference systems. - Optimize model architectures to improve deployment and runtime efficiency using techniques such as distillation, pruning, quantization, and Mixture-of-Experts (MoE). - Implement advanced serving strategies including batching, caching (KV, semantic, embedding), quantization (FP8/INT8), and distributed inference strategies. - Write and maintain high-performance GPU kernels using Triton or CUDA to accelerate custom model layers and critical workloads. - Conduct deep performance analysis using tools such as PyTorch Profiler and NVIDIA Nsight to identify bottlenecks in compute, memory, and communication. - Partner with infrastructure teams to design scalable and reliable distributed serving systems across heterogeneous hardware environments. - Establish and track efficiency metrics such as cost per million inferences. - Serve as a trusted technical advisor to research and product teams on efficiency tradeoffs. Benefits - Competitive salary and performance-based incentives. - Comprehensive benefits programs. - Opportunities for professional development and growth. Company Description Adobe empowers everyone to create through innovative platforms and tools that unleash creativity, productivity and personalized customer experiences. Adobe’s industry-leading offerings enable people and businesses to turn ideas into impact, powered by AI and driven by human ingenuity. Our 30,000+ employees worldwide are creating the future and raising the bar as we drive the next decade of growth. We’re on a mission to hire the very best and believe in creating a company culture where all employees are empowered to make an impact.

United States
$164K - $313.3K / year
Job Closed