We are a Y-Combinator-backed startup building your AI-powered Recruiter Agent
MLOps Engineer
Location
United States
Posted
2 days ago
Salary
$70 - $110 / hour
Seniority
Mid Level
No structured requirement data.
Job Description
MLOps Engineer
Weekday (YC W21)
Role Description Join a cutting-edge AI research initiative at the forefront of Generative AI and contribute to the development of next-generation Large Language Models. We are seeking experienced MLOps Engineers with deep expertise in modern machine learning frameworks, large-scale training infrastructure, and kernel-level optimization. In this role, you'll leverage your knowledge of JAX, PyTorch, and custom GPU kernel programming (Pallas/Triton) to create, evaluate, and refine high-quality technical tasks that help train frontier AI systems. You'll collaborate with AI researchers and engineering teams to improve model reasoning across MLOps, distributed training, and ML infrastructure topics. This is a full-time, 40-hour-per-week remote engagement requiring full weekday availability. Key Responsibilities - Partner with research and engineering teams to strengthen AI model capabilities in MLOps, ML infrastructure, and large-scale training systems. - Design challenging, real-world MLOps and machine learning systems tasks that reflect production engineering scenarios. - Develop accurate, well-documented solutions to complex ML infrastructure and training pipeline problems. - Review and evaluate technical tasks and AI-generated solutions, providing clear and actionable written feedback. - Create detailed evaluation rubrics and scoring frameworks for topics including: - Distributed training architectures - ML pipeline design - Infrastructure optimization - Kernel-level programming - Performance tuning - Collaborate with fellow subject matter experts to maintain consistency, quality, and technical accuracy across training datasets. - Contribute domain expertise to improve the reasoning capabilities of advanced AI systems. Qualifications - Minimum 2 years of professional experience in MLOps, Machine Learning Infrastructure, or ML Systems Engineering within a recognized technology organization. - Hands-on production experience with JAX and/or PyTorch in large-scale machine learning environments. - Practical experience developing or optimizing custom GPU kernels using Pallas (JAX) or Triton. - Strong understanding of distributed training systems, model optimization, and scalable ML infrastructure. - Demonstrated career growth and increasing technical responsibility. - Availability to work 40 hours per week during standard weekday business hours. - Excellent written communication skills with the ability to clearly explain technical concepts and architectural decisions. Preferred Skills - Experience designing and optimizing large-scale ML training pipelines. - Knowledge of distributed computing and GPU performance optimization. - Familiarity with evaluation methodologies for AI models and ML systems. - Experience collaborating with research teams on advanced machine learning projects. - Passion for advancing AI infrastructure and frontier model development. Benefits - Help build and improve next-generation Large Language Models. - Work alongside leading AI researchers and experienced machine learning engineers. - Apply your expertise to high-impact projects involving large-scale ML systems and infrastructure. - Contribute directly to the development of cutting-edge AI technologies. - Enjoy a fully remote engagement with meaningful technical challenges. Equal Opportunity We welcome applications from qualified professionals regardless of legally protected characteristics and are committed to providing reasonable accommodations throughout the application and engagement process upon request. Contract & Payment Terms - Engagement is offered on an independent contractor basis. - This is a fully remote opportunity that can be completed according to your own schedule. - Project duration may be extended, shortened, or concluded based on project requirements and individual performance. - The engagement does not require access to confidential or proprietary information belonging to any current employer, client, or institution. - Payments are processed weekly through Stripe or Wise based on approved work completed. - Please note: Applicants requiring H-1B sponsorship or participating in the STEM OPT program are not eligible for this opportunity.
Related Guides
Related Job Pages
More Machine Learning Engineer Jobs
Autonomous Learning Engineer
Bright Vision TechnologiesBright Vision Technologies is a forward-thinking software development company dedicated to building innovative solutions that help businesses automate and optimize their operations. We leverage cutting-edge technologies to create scalable, secure, and user-friendly applications.
Role Description We are seeking an Autonomous Learning Engineer to design, train, and deploy reinforcement learning (RL) solutions for complex decision-making systems. The ideal candidate will have expertise in: - Python - Deep learning - Reinforcement learning - Simulation environments - RLHF - Distributed training Experience building scalable, production-ready AI solutions is essential. Key Responsibilities - Design and implement reinforcement learning solutions for real-world applications. - Develop and optimize simulation environments and RL training pipelines. - Implement and evaluate modern RL algorithms and reward models. - Improve model performance, training stability, and sample efficiency. - Collaborate with AI and product teams to deploy production-ready RL solutions. - Monitor deployed models and ensure reliability, safety, and continuous improvement. Qualifications - 6+ years of experience in reinforcement learning research and engineering. - Strong proficiency in Python and modern deep learning frameworks. - Experience with RL libraries, simulation environments, GPU-based training, and distributed learning. - Strong understanding of reinforcement learning, optimization, and probability. Preferred Qualifications - Experience with RLHF, multi-agent RL, robotics, autonomous systems, or control systems. - Publications or open-source contributions in reinforcement learning. Benefits - 100% Remote (U.S.) - Full-time, Direct W2 - Salary Range: $100,000–$150,000 Annually Application Instructions Interested in this opportunity? Apply today for immediate consideration! - Email your updated resume: [email protected] - Call or Text: (908) 505-3545 - Learn more: www.bvteck.com Equal Employment Opportunity (EEO) Statement Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall. BV Teck expressly prohibits any form of workplace harassment or discrimination. Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.
Role Description We're looking for a Machine Learning Engineer to own and advance the forecasting and predictive modeling capabilities at the heart of the Camus platform. This is an individual contributor role with real technical depth and product influence; you'll be responsible for the full lifecycle of ML model development, from exploratory analysis and model design through to production deployment and monitoring. This is not a role where the problem statements are handed to you. You'll work directly with Camus’ teams and external stakeholders to understand their data, define the right questions, and translate messy real-world signals into reliable, production-grade data driven analytics. You'll bring that ground-truth perspective back into product decisions, and work closely within the Engineering team to integrate ML models into our planning and operational workflows. The forecasting and predictive modeling problems we're solving often don't have off-the-shelf answers. We work as a tight, technical team that moves with urgency but builds with the discipline that production-grade software demands. If you want to do the most technically interesting ML work in the clean energy space while directly shaping how it becomes a product, this is the role. What You'll Do - Design, train, and evaluate predictive ML models with a focus on forecasting and time-series applications - Conduct exploratory data analysis, feature engineering, and statistical modeling across large structured and unstructured datasets - Collaborate with Engineering to define ML infrastructure requirements, and deploy and integrate ML models into operational workflows and decision-support tools - Work cross-functionally with Camus teams to define problem statements and translate business objectives into ML solutions - Communicate model performance, uncertainty, and limitations clearly to both technical and non-technical audiences - Champion ML best practices around reproducibility, versioning, and testing Qualifications - PhD with 3+ years of industry experience, Masters with 5+ years, or Bachelors with 8+ years in Machine Learning, Statistics, Computer Science, Applied Mathematics, or a related quantitative field - Demonstrated track record of delivering ML models into production environments - Experience with time-series forecasting methods — including classical approaches (e.g. ARIMA) and modern ML-based methods (e.g. gradient boosting or temporal neural networks) - Strong proficiency in Python and core ML/data science libraries (PyTorch, scikit-learn, statsmodels, pandas, etc.) - Experience with probabilistic forecasting, uncertainty quantification and backtesting - Ability to translate ambiguous business problems into well-scoped ML projects - Comfortable operating with autonomy in a small team, balancing speed of delivery with the engineering discipline that production-grade software demands. Nice to Have - Experience in the energy sector — e.g. load forecasting, renewable generation prediction, price modeling or grid operations - Experience with MLOps tooling and infrastructure: cloud platforms, containerization, and model serving patterns - Experience with data pipeline tooling, e.g. Airflow, Spark, or Databricks - Able to leverage AI code development tools to accelerate development Benefits - Competitive base salary - Comprehensive benefits, including FSA and 401k for full time employees - Fully remote workplace with options for in office work in the Bay Area - Flexible PTO, which we encourage you to use! - A real impact on climate change - we’re building the world we want to live in and we want you to join us! - The expected base salary for this role is $180,000 - $230,000 annually, depending on experience, skills, and qualifications.
Role Description We're working to advance care through data-driven decisions and automation. This mission serves as the foundation for every decision as we create the future of travel. We can't do that without the best talent - talent that is innovative, curious, and driven to create exceptional experiences for our guests, customers, owners, and colleagues. We seek an extraordinary Machine Learning Engineer to help build the algorithmic assets and features that Hyatt guests, members, customers, and internal users leverage to transform the guest experience and drive efficiencies across the operations of our business. In this role, you will: - Design and implement algorithmic product architectures to bring our machine learning models to life across the full lifecycle of the product including data ingestion, ML processing, and results delivery/activation. - Work cross-functionally with various data science teams, data engineering teams, and data architecture teams. - Serve as both solutions architect and hands-on implementation engineer, guiding the team towards best-in-class algorithmic product implementations. - Be a part of a ground-floor, hands-on, highly visible team which is positioned for growth and is highly collaborative and passionate about data science. - Apply the latest techniques and approaches across the domains of data science, machine learning, and AI. Qualifications - 5+ years of implementing software product solutions in a cloud environment with a focus on algorithmic/machine learning products, hospitality experience not required. - Expertise in AWS cloud services. - Expertise in Python, SQL, PySpark, Docker. - Experience with streaming and batch data architectures at scale. - Experience operating in an Agile Methodology environment. - Experience with DevOps and CI/CD concepts. - Excellent communication and teamwork skills. - Position will not require customer-facing interactions. Requirements - Stay up to date with the latest design patterns and AWS services with respect to Machine Learning Engineering. - Partner with data architecture, data governance, and security teams to ensure solutions meet required standards. Company Description
• Contribute to the development of scalable machine learning platforms and workflows that enable scientists and researchers across Amgen to build, deploy, and manage AI/ML models • Work closely with experienced engineers, data scientists, and domain experts to productionize machine learning solutions primarily in support of drug discovery and development • Deliver AI and ML-enabled applications, with deployed models ranging from classical ML models to natural language processing, protein language models, and large language models • Contribute to the development and maintenance of ML platform capabilities, including: • Data pipelines and feature engineering workflows • Model training, evaluation, and deployment pipelines • Experiment tracking and model registry systems • Model performance evaluations and monitoring • Partner in implementation of AI and ML Ops best practices, including CI/CD, infra as code, monitoring, traceability and reproducibility • Collaborate with cross-functional teams to transition standards and outcomes from experimentation to production-grade enterprise solutions • Build and maintain scalable and efficient, productionized MLOps solutions on cloud platforms • Develop and deliver training content and knowledge articles to educate resident scientists on model lifecycle management best practices

