Centific Global Solutions

Centific, founded in 2020, is a technology company specializing in artificial intelligence and digital transformation services, focusing on developing scalable

Research Intern — Applied Reinforcement Learning

Location

United States

Posted

127 days ago

Salary

$35 - $45 / hour

Seniority

Entry Level

No structured requirement data.

Job Description

Research Intern — Applied Reinforcement Learning

Centific Global Solutions

About Centific Centific is a frontier AI data foundry that curates diverse, high-quality data, using our purpose-built technology platforms to empower the Magnificent Seven and our enterprise clients with safe, scalable AI deployment. Our team includes more than 150 PhDs and data scientists, along with more than 4,000 AI practitioners and engineers. We harness the power of an integrated solution ecosystem—comprising industry-leading partnerships and 1.8 million vertical domain experts in more than 230 markets—to create contextual, multilingual, pre-trained datasets; fine-tuned, industry-specific LLMs; and RAG pipelines supported by vector databases. Our zero-distance innovation™ solutions for GenAI can reduce GenAI costs by up to 80% and bring solutions to market 50% faster. Our mission is to bridge the gap between AI creators and industry leaders by bringing best practices in GenAI to unicorn innovators and enterprise customers. We aim to help these organizations unlock significant business value by deploying GenAI at scale, helping to ensure they stay at the forefront of technological advancement and maintain a competitive edge in their respective markets. About Job Job Description PhD Research Intern — Applied Reinforcement Learning Centific AI Research Role Summary Centific AI Research seeks a PhD Research Intern to design and evaluate reinforcement learning (RL) systems for agentic AI workflows. You will develop RL environments, reward models, and post-training pipelines for LLM-based agents, translating research into practical enterprise solutions. Scope of Work - End-to-end RL pipelines for agentic systems (simulation → training → evaluation) - Alignment of LLM-based agents using RLHF, DPO, PPO, and emerging methods - Design of reward functions, verifiers, and evaluation frameworks - Simulation environments (digital twins) for enterprise workflows - Scalable training and inference for RL-based systems Example Projects - Build a custom RL environment simulating a real-world enterprise workflow and train an agent using PPO or GRPO - Develop a reward modeling pipeline from human feedback and evaluate alignment improvements - Create an evaluation harness measuring reasoning, task success, and policy safety - Prototype an agentic system with tool use and multi-step reasoning, integrated with RL training - Document experiments, ablations, and findings for research and productionization Minimum Qualifications - PhD candidate in CS, ML, or related field with research in reinforcement learning or agentic AI - Strong Python and PyTorch skills with GPU-based training experience - Solid understanding of RL fundamentals (MDPs, policy gradients, value methods) - Experience with LLMs and post-training techniques (RLHF, DPO, PPO, etc.) - Strong experimentation practices (ablation, reproducibility, clear reporting) Preferred Qualifications - Experience with RL environments (Gymnasium, RLlib, Stable Baselines) - Research in offline RL, model-based RL, or hierarchical RL - Publications at top ML conferences (NeurIPS, ICML, ICLR, ACL) - Experience with simulation, synthetic data, or multi-agent systems - Distributed training and large-scale experimentation Tech Stack - PyTorch, CUDA; RL libraries (Gymnasium, RLlib, Stable Baselines) - LLM frameworks and post-training tools (TRL, custom RLHF pipelines) - Experiment tracking (Weights & Biases) - APIs/services (FastAPI, gRPC); optimization (ONNX, TensorRT) Logistics Location: Palo Alto, CA (Preferred), Redmond, WA (Preferred) or Remote Duration: 3–6 months What We Offer - Competitive stipend and real-world impactful projects - Mentorship from researchers and engineers - Access to modern GPU infrastructure - Opportunities to publish and present research Centific AI Research is an Equal Opportunity Employer. We celebrate diversity and are committed to an inclusive environment. Rate: $35-$45 Hourly Centific is an equal-opportunity employer. All qualified applicants will receive consideration for employment without regard to race, color, religion, national origin, ancestry, citizenship status, age, mental or physical disability, medical condition, sex (including pregnancy), gender identity or expression, sexual orientation, marital status, familial status, veteran status, or any other characteristic protected by applicable law. We consider qualified applicants regardless of criminal histories, consistent with legal requirements.

Related Categories

Related Job Pages

More Research Engineer Jobs

Prime Intellect logo

Research Engineer – RL Infra

Prime Intellect

Find compute. Train Models. Co-own intelligence.

Research Engineer127 days ago
Full TimeRemoteTeam 1-10H1B No Sponsor

• Design and build scalable RL training infrastructure — async trainers, environment orchestration, reward pipelines — across large GPU clusters. • Optimize performance, cost, and resource utilization of RL workloads using state-of-the-art compute and memory optimization techniques. • Contribute to our open-source libraries and frameworks for distributed RL training. • Publish research at top-tier venues (ICML, NeurIPS). • Write clear, approachable technical content distilling complex systems work for customers and the broader community. • Stay current with advances in RL systems, distributed training, and ML infrastructure, and proactively identify opportunities to enhance our platform.

United States
Job Closed
SharkNinja logo

Applied Research Engineer II, Human Wellness

SharkNinja

Product design and technology company positively impacting people's lives every day in homes around the world.

Research Engineer128 days ago
Full TimeRemoteTeam 1,001-5,000Since 1994H1B Sponsor

About Us SharkNinja is a global product design and technology company, with a diversified portfolio of 5-star rated lifestyle solutions that positively impact people's lives in homes around the world. Powered by two trusted, global brands, Shark and Ninja, the company has a proven track record of bringing disruptive innovation to market and developing one consumer product after another has allowed SharkNinja to enter multiple product categories, driving significant growth and market share gains. Headquartered in Needham, Massachusetts with more than 4,100 associates, the company's products are sold at key retailers, online and offline, and through distributors around the world. Applied Research Engineer II, Human Wellness Needham, MA, United States Our purpose is to positively impact people's lives every day in every home around the world! We work very hard to provide our consumers with high-quality, exciting 5-star products that make life easier. We thrive on passion and innovation and are looking for great people, with great ideas, who want to build the next big thing and develop while they do. Are you ready to take your skills to the next level and join a company that thrives on Innovation? If yes, I have the perfect position for you! Our mission to positively impact people's lives every day in every home around the world allows our employees to be thinkers and tinkerers, designers and doers, creators and number crunchers, makers of things they love. As we continue to grow, we are excited to add an Applied Research Engineer IIto our global team. Overview The Research and Development Team at SharkNinja is seeking an experienced, versatile Engineer II to help deliver an exceptional pipeline of new technologies across the beauty and human wellness spaces. In this role, you will work at the intersection of research and engineering-translating scientific insight into real-world solutions that meaningfully improve people's lives. As an Engineer II, you will take ownership of complex, open-ended technical challenges, partner closely with cross-functional teams, and influence technical direction and product definition through technology exploration and applied scientific research. This role is well-suited for someone who thrives in ambiguity, takes initiative, and enjoys shaping problems as much as solving them. What You'll Do - Evaluate emerging technologies and help guide research roadmaps and investment decisions - Lead applied research initiatives in human health and wellness to test hypotheses and support product claims - Develop and refine algorithms, models, and experimental methods across a broad range of biomechanical and physiological data streams - Synthesize findings and communicate clearly through presentations and written materials for both technical and non-technical audiences - Collaborate closely with engineering, product, marketing, and regulatory partners to translate research into impact - Engage with external experts to accelerate innovation and support clinical studies for medical devices - Mentor junior researchers and provide hands-on technical leadership across projects - Uphold high standards for scientific rigor, reproducibility, and ethical research practices Attributes and Skills - MS in Bioengineering, Biomedical Engineering, Mechanical Engineering, Physical Therapy, or a related field-or equivalent practical experience with 2+ years' of experience or a PhD in the same fields with 0+ years' of experience. - 0 - 2+ years of applied research experience in industry, research labs, or translational academic environments - Strong foundation in experimental design, statistical analysis, and data interpretation - Hands-on experience working in laboratory settings with diverse test equipment and instrumentation - Knowledge in one or more of the following areas: human anatomy and physiology, biomechanics, physics of biological tissues, or human gait and movement - Proficiency in computational programming (e.g., Python, MATLAB) - Ability to work independently, prioritize effectively, and make progress amid evolving requirements - This is an in-person role based in our Needham, MA office Preferred Attributes - Demonstrated success translating research into shipped products, deployed systems, or validated claims - Experience developing wearable sensors or medical devices - Hands-on prototyping and experimental protocol development experience - Comfort balancing multiple initiatives and adapting to shifting priorities while maintaining focus on long-term goals Salary and Other Compensation: The annual salary range for this position is displayed below. Factors which may affect starting pay within this range may include geography/market, skills, education, experience and other qualifications of the successful candidate. The Company offers the following benefits for this position, subject to applicable eligibility requirements: medical insurance, dental insurance, vision insurance, flexible spending accounts, health savings accounts (HSA) with company contribution, 401(k) retirement plan with matching, employee stock purchase program, life insurance, AD&D, short-term disability insurance, long-term disability insurance, generous paid time off, company holidays, parental leave, identity theft protection, pet insurance, pre-paid legal insurance, back-up child and eldercare days, product discounts, referral bonus program, and more. Pay Range $81,700-$130,000 USD Our Culture At SharkNinja, we don't just raise the bar-we push past it every single day. Our Outrageously Extraordinary mindset drives us to tackle the impossible, push boundaries, and deliver results that others only dream of. If you thrive on breaking out of your swim lane, you'll be right at home. What We Offer We offer competitive health insurance, retirement plans, paid time off, employee stock purchase options, wellness programs, SharkNinja product discounts, and more. We empower your personal and professional growth with high impact Learning Programs featuring bold voices redefining what's possible. When you join, you're not just part of a company-you're part of an outrageously extraordinary community. Together, we won't just launch products-we'll disrupt entire markets. At SharkNinja, Diversity, Equity, and Inclusion are vital to our global success. Valuing each unique voice and blending all of our diverse skills strengthens SharkNinja's innovation every day. We support ALL associates in bringing their authentic selves to work, making an impact, and having the opportunity for career acceleration. With help from our leadership, associates, and our community, we aim to have equity be a key component of the SharkNinja DNA. Learn more about us: Life At SharkNinja Outrageously Extraordinary SharkNinja Candidate Privacy Notice - For candidates based in all regions, please refer to this Candidate Privacy Notice. - For candidates based in China, please refer to this Candidate Privacy Notice. - For candidates based in Vietnam, please refer to this Candidate Privacy Notice. We do not discriminate on the basis of race, religion, color, national origin, sex, gender, gender expression, sexual orientation, age, marital status, veteran status, disability, or any other class protected by legislation, and local law. SharkNinja will consider reasonable accommodations consistent with legislation, and local law. If you require a reasonable accommodation to participate in the job application or interview process, please contact SharkNinja People & Culture at accommodations@sharkninja.com

United States
$81.7K - $130K / year
Job Closed
Sumo Logic logo

Staff Threat Research Engineer

Sumo Logic

Sumo Logic’s vision is to make the world's digital experiences reliable and secure.

Research Engineer128 days ago
Full TimeRemoteTeam 501-1,000Since 2010H1B Sponsor

• Conduct and lead both applied and original threat research, transforming intelligence, telemetry, and investigation into actionable detection logic for the Sumo Logic SIEM. • Collaborate closely within Threat Labs to design, build, and refine detection content and validation pipelines that raise the bar for product and customer detection quality. • Drive innovation in detection methodologies, including research activities such as malware analysis, infrastructure tracking, or honeypot operations, to discover new attacker behaviors. • Publish and share findings — from detection logic to behavioral analysis and practical hunting guidance — that help customers maximize SIEM outcomes. • Contribute to Threat Labs’ long‑term vision of a research‑driven, continuously evolving detection ecosystem built on practitioner insight and technical depth. • Research, develop, and test threat detection logic in a lab environment, validating against real‑world attacker behaviors and ensuring technical alignment with Sumo Logic SIEM capabilities. • Conduct original threat research, such as analyzing malware, tracking infrastructure, or experimenting with honeypots, and translate findings into detection opportunities. • Investigate industry and adversary trends to identify emerging detection opportunities. • Collaborate with product management and fellow Threat Labs engineers to scope and prioritize detection campaigns. • Maintain and expand Threat Labs’ research lab infrastructure. • Provide practitioner feedback to engineering and product management to inform feature design and roadmap decisions. • Contribute to the security community through blogs, conference talks, open source projects, and public research contributions.

United States
$162K - $190K / year
OtherRemoteTeam 11-50H1B No Sponsor

At Murmuration, we believe that America’s promise is shaped and reshaped by the best ideas and ideals of its communities, and the dreams of the people who believe in a better life for themselves, their families, and each other.  We help organizations build power in their communities in four key ways: we organize a network of values-aligned partners; we provide deep, data-driven insights into people, places, and perspectives; we develop tools that make organizing and engagement easy and more effective; and we offer services that strengthen our partners’ capacity to lead change in their communities. We envision an America where every community has what it needs to help people lead healthy, free, and dignified lives. We work to redesign the systems and structures we all depend on — how we learn, live, govern, and solve problems — so that they are just, equitable, resilient, and rooted in shared responsibility. By strengthening the ties that hold communities together, we aim for civic life defined by collective action and care, with effective leadership that truly represents everyone.  We are a collaborative, curious, and creative team of organizers, scientists, teachers, technologists, campaign veterans, and more who share the unwavering belief that we can use our gifts in service of transforming America — together. We’ve built our team guided by the belief that the whole is greater than the sum of its parts. And so we support each other relentlessly — rallying together to face challenges the same way we celebrate each other’s wins. About the Position You're an experienced engineer who combines deep technical skill with the judgment to make sound architectural decisions. You drive complex projects across Murmuration's research and data science enablement systems: designing, building, and operating the platforms, tooling, and infrastructure that let researchers, data scientists, and partners work faster and at greater scale. You know how to break down ambiguous problems, make pragmatic tradeoffs between speed and sustainability, and ship reliable systems that serve the organization for the long term. You'll work at the intersection of engineering and enablement, collaborating across Research, Data Science, Product, and Growth to build the systems that power how we understand and activate civic participation in America. Beyond your own work, you raise the bar for the team through mentorship, technical leadership, and standards that make everyone more effective. The Research Engineering Team builds systems that multiply the impact of researchers, data scientists, and our network. How do we turn a question that takes a researcher three days into one that takes three minutes? How do we scale model scoring from millions to billions without manual intervention? How do we make cutting-edge analysis accessible to journalists and decision-makers who aren't data scientists? How do we build platforms that let our team move faster and do more? If you want to build systems that change how we understand and activate civic participation, we'd love to hear from you. Job Level: P4 What You'll Do - Work across the stack: You have strong depth in one or more critical areas (e.g., API design, frontend architecture, data integration, infrastructure) and broad proficiency across the rest. You'll use whatever it takes — full-stack applications, backend systems, data pipelines, AI integrations — and continue deepening your expertise over time. - Build end-to-end systems from data infrastructure that processes large-scale datasets reliably and cost-effectively to intuitive interfaces that make sophisticated analysis accessible to non-technical users like journalists, organizers, and decision-makers. You’ll turn manual, time-intensive processes into fast, scalable, automated ones. - Apply emerging technologies (including AI/ML) to R&D challenges. You’ll prototype novel approaches, evaluate feasibility, and bridge the gap between cutting-edge capabilities and practical usability. - Communicate clearly: You’ll provide actionable guidance to research and data science stakeholders on technical feasibility, solution alternatives, and tradeoffs. produce clear technical documentation (e.g., design docs, architecture decisions, and runbooks), and build consensus on design approaches across the team. - Elevate the team: Mentor engineers at all levels through pairing, code review, and technical guidance, set high standards for code quality and testing that raise the bar for the entire team, and drive process improvements, better tooling, and shared knowledge that make the team more effective over time.

United States
Job Closed