Bright Vision Technologies

Bright Vision Technologies is a forward-thinking software development company dedicated to building innovative solutions that help businesses automate and optimize their operations. We leverage cutting-edge technologies to create scalable, secure, and user-friendly applications.

AI Performance Engineer

Location

United States

Posted

4 days ago

Salary

$75K - $100K / year

Seniority

Mid Level

No structured requirement data.

Job Description

AI Performance Engineer

Bright Vision Technologies

Role Description We are seeking an AI Performance Engineer to focus on extracting maximum throughput, minimizing latency, and reducing cost across training and inference workloads for large neural network systems. The role spans the full stack from low-level kernel optimization to distributed system tuning, requiring deep understanding of GPU architecture, model parallelism, memory management, and compiler-level optimization. - The ideal candidate has demonstrated impact on production AI workloads. - Strong instrumentation and measurement discipline that enables rigorous, data-driven optimization decisions. - Work closely with cross-functional partners — product, design, engineering, operations, and business stakeholders — to translate ambiguous requirements into well-engineered solutions. - Expected to raise the bar through code review, design review, and mentorship of more junior engineers. - Brings strong engineering discipline, a clear communication style, and a track record of shipping meaningful work that holds up well in production. Qualifications - Bachelor’s or Master’s degree in Computer Science, Computer Engineering, or a related field. - Six or more years of experience in performance engineering, ML systems, or HPC. - Strong proficiency in Python and C++. - Hands-on experience optimizing deep learning workloads on modern GPUs. - Deep understanding of distributed training and inference techniques. - Experience with profiling tools across CPU, GPU, and distributed systems. - Familiarity with model compression techniques and their accuracy implications. - Strong grasp of memory hierarchies, communication primitives, and parallelism strategies. - Excellent measurement, debugging, and analytical reasoning skills. - Strong communication and collaboration skills. Requirements - Experience optimizing LLM inference at production scale. - Contributions to vLLM, TensorRT-LLM, DeepSpeed, or similar projects. - Familiarity with custom kernel authoring in Triton or CUTLASS. - Experience with FinOps for AI workloads. - Publications or talks on AI systems performance. How to Apply Would you like to know more about this opportunity? For immediate consideration, please send your resume to [email protected] . Equal Employment Opportunity (EEO) Statement Bright Vision Technologies (BV Teck) is committed to equal employment opportunity (EEO) for all employees and applicants without regard to race, color, religion, sex, sexual orientation, gender identity or expression, national origin, age, genetic information, disability, veteran status, or any other protected status as defined by applicable federal, state, or local laws. This commitment extends to all aspects of employment, including recruitment, hiring, training, compensation, promotion, transfer, leaves of absence, termination, layoffs, and recall. - BV Teck expressly prohibits any form of workplace harassment or discrimination. - Any improper interference with employees' ability to perform their job duties may result in disciplinary action up to and including termination of employment.

Related Job Pages

More AI Engineer Jobs

Full TimeRemoteTeam 5,001-10,000Since 2000H1B No Sponsor

• Develop, train, deploy, and maintain machine learning and AI solutions focused on Natural Language Processing (NLP), Large Language Models (LLMs), generative AI, and other advanced AI use cases • Design and implement generative AI solutions, including LLM APIs and machine learning techniques • Research and improve machine translation (MT) quality • Manage and improve data quality and integrity • Develop automated anomaly detection and monitoring solutions • Design and propose technical architectures for NLP, ML, and AI solutions • Implement MLOps and AI governance practices • Communicate technical concepts and drive innovation

Mali
Keyrus logo

AI Engineer

Keyrus

#MakeDataMatter #HumanizingTheFuture

AI Engineer4 days ago
Full TimeRemoteTeam 1,001-5,000Since 1996H1B Sponsor

• Build and maintain LLM-powered pipelines for processing and automating customer requests. • Design and implement data ingestion, data preprocessing, and enrichment pipelines for unstructured data sources, including emails, PDFs, images, and documents. • Develop classification, routing, and decision workflows for customer inquiries. • Apply prompt engineering best practices and continuously improve model performance and output quality. • Implement structured outputs and Human-in-the-Loop (HITL) workflows to ensure reliability and business adoption. • Ensure production readiness through testing, CI/CD, containerisation, and monitoring. • Contribute to AI architecture design, scalability, and engineering standards. • Monitor, evaluate, and optimise AI systems using LLM observability tools and performance metrics. • Collaborate closely with engineering, product, and business teams to integrate AI solutions into existing processes and workflows.

Portugal
€45K - €60K / year
Full TimeRemoteTeam 10,001+

• Own AI-powered features end-to-end, from scoping, design and requirements engineering to feature evaluation, deployment and operation. • Design, build and maintain cloud-native, event-driven AI applications across the full stack. • Evaluate, integrate, and optimize AI models for production use (quality, latency, cost). • Build and maintain robust AI inference, evaluation, and monitoring pipelines. • Collaborate closely with product, design, and non-technical stakeholders to identify requirements and turn business needs into value-creating technical solutions. • Operate and improve production systems, taking ownership of reliability, quality, and technical debt. • Plan, track, and break down work effectively in an agile team.

Germany
Statheros logo

Artificial Intelligence Engineer – Developer

Statheros

Statheros is a digital currency that is backed by real estate, so its value will remain stable over time.

AI Engineer4 days ago

• Design, implement, and optimize Proximal Policy Optimization (PPO) algorithms for domain-specific use cases. • Develop and train reinforcement learning models for real-world applications, focusing on efficiency and scalability. • Collaborate with cross-functional teams to integrate PPO models into production systems. • Analyze model performance and experiment with hyperparameter tuning to achieve optimal results. • Stay up-to-date with the latest research and advancements in reinforcement learning and apply them to enhance existing solutions. • Build robust pipelines for training, evaluation, and deployment of RL models. • Document workflows, methodologies, and code for reproducibility and knowledge sharing.

Alabama + 6 moreAll locations: Alabama | Florida | New Hampshire | Ohio | Tennessee | Texas | Utah