Job Closed
This listing is no longer active.
Pioneering AI-first solutions, solving complex business challenges through expertise, cloud, data engineering, and AI.
GenAI Architect
Location
United States
Posted
111 days ago
Salary
0
Seniority
Lead
Job Description
GenAI Architect
Quantiphi
• Design and implement GenAI solutions using AWS Bedrock and Agentcore. • Define architecture for LLM-based applications, including RAG pipelines and agentic workflows. • Develop and orchestrate agentic AI workflows, enabling multi-step reasoning, tool usage, and task automation. • Build and manage RAG pipelines, including embeddings, retrieval mechanisms, and vector databases. • Integrate LLM capabilities into enterprise applications via APIs and backend services. • Design and optimize prompt engineering strategies for accuracy, relevance, and performance. • Work with structured and unstructured data sources to enable knowledge-driven AI applications. • Ensure model evaluation, monitoring, and optimization for latency, cost, and response quality. • Collaborate with application, data, and platform teams for end-to-end solution delivery. • Define best practices for security, governance, and responsible AI usage. • Troubleshoot and resolve issues in production GenAI systems. • Provide technical leadership and mentor team members while remaining hands-on.
Job Requirements
- 8+ years of relevant hands-on technical experience implementing, and developing cloud solutions on AWS.
- Hands-on experience on AWS services.
- Proven experience using AWS Sagemaker and Bedrock leveraging different types of data sources, Training jobs, real-time and batch applications.
- Design and implement agentic AI architectures using frameworks such as LangChain, Strand Agents etc., enabling autonomous task planning, decision-making, and multi-step reasoning.
- Hands-on experience with Amazon AgentCore for building, deploying, and scaling production-grade agentic AI applications, including agent memory management, tool registry, and observability.
- Architect and deploy scalable AI solutions on AWS, leveraging services like Lambda, Bedrock, Step Functions, S3, API Gateway, and SageMaker.
- Proficiency in working with LLM APIs (e.g., Claude, Nova, and other third-party LLM providers), including API integration,and multi-model orchestration strategies.
- Hands-on experience fine-tuning or optimizing large language models (LLM).
- Familiarity with LLM tool use, prompt templating and context management.
- Strong expertise in Vector Databases, including indexing strategies, embedding generation, similarity search, and integration with RAG architectures.
- Model Evaluation & Optimization: Evaluate LLM's zero-shot and few-shot capabilities, fine-tuning hyperparameters, ensuring task generalization, and exploring model interpretability for robust web app integration.
- Develop and maintain Model Context Protocol (MCP) implementations to manage state, context windows, memory, and prompt orchestration across distributed agent systems.
- Experience with at least one of the workflow orchestration tools, Airflow, StepFunctions, SageMaker Pipelines, Kubeflow etc.
- Experience implementing secure, scalable APIs and integrating with 3rd-party data sources and tools.
- Ability to collaborate with cross-functional teams such as Developers, QA, Project Managers, and other stakeholders to understand their requirements and implement solutions.
- Should have experience with Deep Learning Concepts - Transformers, BERT, Attention models, tokenization, embeddings.
Benefits
- Make an impact at one of the world’s fastest-growing AI-first digital engineering companies.
- Upskill and discover your potential as you solve complex challenges in cutting-edge areas of technology alongside passionate, talented colleagues.
- Work where innovation happens - work with disruptive innovators in a research-focused organization with 60+ patents filed across various disciplines.
- Stay ahead of the curve—immerse yourself in breakthrough AI, ML, data, and cloud technologies and gain exposure working with Fortune 500 companies.
- Enjoy a fun, diverse and hybrid work culture with ample opportunities to learn and grow.
Related Guides
Related Job Pages
More AI Engineer Jobs
What we do Snezzi helps brands show up in AI answers on ChatGPT, Perplexity, Google AI Mode, and others. We handle everything end to end: audits, content, technical fixes, backlinks, and measurement. Fully done for you. We're a small team of second-time founders. We ship weekly. We run AI agents for most of our internal work. And we're growing fast enough that we need sharp people who want to learn this space by doing real work, not by watching from the sidelines. What this internship actually is You'll work on live client accounts from week one. That means real audits, real fixes, real reporting. Not shadowing. Not slide decks. Not busywork. By the end of the engagement, you'll have shipped actual work, built skills that are genuinely hard to find right now, and have a portfolio that proves it. Strong performers convert to full-time. We'll tell you early if we see that path for you. What you'll do - Run AI visibility audits across ChatGPT, Perplexity, and Google AI Mode - Help identify gaps and assist with fixes: content updates, schema, technical cleanup - Use Claude Code and other AI tools to build and run scripts (no prior coding required — you learn by doing) - Help prepare client reports, weekly updates, and pre-reads - Document and templatize every process you touch so the next person can use it Tools you'll use Claude Code, ChatGPT, Perplexity, Google Search Console, GA4, Ahrefs, Shopify, WordPress. You'll learn them on the job. What matters is how fast you pick things up. Who we're looking for You don't need a specific degree or years of experience. You need the right instincts. Must-haves: - You use AI tools a lot, out of curiosity, not just to finish assignments - You write clearly. Most of this job is communication: with clients, with AI agents, with your team. - A terminal or a code file doesn't scare you. You don't need to write code, but if Claude generates a script, you can run it, read the output, and ask Claude to fix what broke. - You try before you ask. "I don't know" is a starting point, not an excuse. - You want to ship things, not just observe them. Moves you to the top of the pile: - You've built or shipped anything on the internet: a blog, a store, a tool, a side project, a newsletter - You've poked around GA4 or Search Console before - You've written content for a brand or website - You've done any kind of SEO, growth, or content work Not a fit if: - You think prompting an AI handles everything. Taste and judgment still matter. - You want clean handoffs and tightly scoped tasks - Client-facing work makes you uncomfortable
What we do Snezzi helps brands show up in AI answers - on ChatGPT, Perplexity, Google AI Mode, and others. We handle everything: audits, content, technical fixes, backlinks, and measurement. Fully done for you. We're a small team of second-time founders. We ship daily. We use AI agents for most of our internal work. And we're growing fast enough that we need someone to help us deliver. The role You'll work directly with clients to improve how they show up in AI answers. You'll run audits, fix what's broken, publish content, and report what moved. This is not a ticket-taking job. You own accounts. You spot problems. You ship fixes. The week you join, you're already working on a real client. Most of your leverage comes from knowing how to use AI tools well — not from writing code from scratch. If you can take a Claude-generated script, run it, read the output, and ask Claude to fix what broke, you'll thrive here. What you'll actually do day to day: - Audit brand visibility across ChatGPT, Perplexity, and Google AI Mode - Identify gaps, map them to fixes, and ship them — content, schema, technical cleanup - Run client updates, prep pre-reads, and translate client questions into action - Turn every repeated task into a script or workflow so you never do it manually twice - Use tools like DataForSEO, Google Search Console, Ahrefs, and GA4 to build reports that actually mean something Tools you'll use: Claude Code, ChatGPT, Cursor, Perplexity, v0, Google AI Mode, Python/Node scripts (run and edit, not write from scratch), Shopify, WordPress, Webflow, GA4, Search Console, Ahrefs. You don't need to know all of these on day one. You need to pick things up fast. Who we're looking for Must-haves: - You use AI tools obsessively. You have strong opinions about ChatGPT vs Claude vs Perplexity. - You write clearly and edit ruthlessly. Communication is half this job. - You're not afraid of a terminal. If Claude Code generates a script, you can run it, read the output, and debug it with Claude's help. - You ship. Imperfect and out this week beats perfect and out next month. - You read docs. "I don't know" is a starting point, not a dead end. Moves you to the top of the pile (any one): - You've built something real with Claude Code, Cursor, v0, Lovable, or similar — even a toy project - You've shipped something on the internet: a blog, a store, a tool, a newsletter - You've done SEO, content, or growth work before - You've poked around GA4 or Search Console Not a fit if: - You think prompting an AI is a strategy on its own, without judgment or taste - You want clean scope, handoffs, and ticket queues - You avoid talking to clients
• You will help improve our existing ML-based products and build new features and services using LLMs. • A core part of the role is context engineering — designing the right inputs, retrieval strategies, and prompt structures to get the best possible output from LLMs in production. • You'll also contribute to our software engineering efforts when needed, bridging the gap between ML and software engineering on the team. • You will work in the Machine Learning Team, where we build LanguageWire's machine translation (MT) and adjacent services to monitor and predict its performance. • The team consists of ML Engineers, Analytics Engineers, and Full Stack Software Engineers. • You'll collaborate across these disciplines daily, and we value people who can move fluidly between ML experimentation and production engineering. • You'll report to the team leader.
Applied AI Engineer
BJAKBjak is a technology company focused on making financial services easy, fun and more rewarding for everyone
• Build and ship AI features end-to-end (model → system → user experience) • Design and iterate on prompts, tools, memory, and agent workflows • Turn raw model outputs into structured, reliable, and predictable behaviors • Debug issues across the full stack (model, orchestration, infra, UX) • Optimize for latency, cost, and production reliability • Develop lightweight evaluation frameworks to measure real-world performance • Work closely with product and engineering to translate ambiguous problems into working systems


