Citi, the leading global bank, has approximately 200 million customer accounts and does business in more than 160 countries and jurisdictions. Citi provides consumers, corporations, governments, and institutions with a broad range of financial products and services, including consumer banking and credit, corporate and investment banking, securities brokerage, transaction services, and wealth management. As a bank with a brain and a soul, Citi creates economic value that is systemically responsible and in our clients’ best interests. As a financial institution that touches every region of the world and every sector that shapes your daily life, our Enterprise Operations & Technology teams are charged with a mission that rivals any large tech company. Our technology solutions are the foundations of everything we do from keeping the bank safe, managing global resources, and providing the technical tools our workers need to be successful to designing our digital architecture and ensuring our platforms provide a first-class customer experience. We reimagine client and partner experiences to deliver excellence through secure, reliable, and efficient services. Our commitment to diversity includes a workforce that represents the clients we serve from all walks of life, backgrounds, and origins. We foster an environment where the best people want to work. We value and demand respect for others, promote individuals based on merit, and ensure opportunities for personal development are widely available to all. Ideal candidates are innovators with well-rounded backgrounds who bring their authentic selves to work and complement our culture of delivering results with pride. If you are a problem solver who seeks passion in your work, come join us. We’ll enable growth and progress together. Position Overview
Jobs in United States
Lead Generative Ai Engineer in United States
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current lead generative ai engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
About the team The Applied AI Engineering team is responsible for ensuring the safe and effective deployment of Generative AI applications for developers and startups. We act as a trusted advisor and thought partner for our customers, working to build an effective backlog of GenAI use cases for their industry and drive them to production through strong technical guidance. OpenAI's customers represent a range of diverse backgrounds and maturity, from early-stage startups to late-stage startups. About the Role We are seeking a technically proficient, business-minded Applied AI Engineer to help push the frontier of advanced AI with our strategic startup customers. You'll work with some of the most exciting AI startups in the world, guiding them through ideation, development, delivery, and scaling to accelerate and maximize the value of what they build on our platform. You will have the opportunity to work on the most novel and creative use cases being built on our API, serving as a critical partner in collecting and delivering high-fidelity product and model feedback internally. You will collaborate closely with Sales, Solutions Engineering, Applied Research, and Product teams, and you will report to the Startups Applied AI Lead. This role is based in our San Francisco or New York offices. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Partner closely with strategic startup customers as their technical thought partner to build novel applications on our API, helping them rapidly move from ideation to scale. Provide proactive guidance to maximize business impact and accelerate application development. Experiment and prototype alongside customers, demonstrating practical use cases. Contribute to open-source resources and scale the function by sharing knowledge, codifying best practices, and publishing useful resources. Synthesize and deliver valuable feedback to the Product and Research
About the team The AI Deployment Engineering team is responsible for ensuring the safe and effective deployment of Generative AI applications. We act as a trusted advisor and thought partner for our customers, working to build an effective backlog of GenAI use cases for their industry and drive them to production through strong technical guidance. As an AI Deployment Engineer (ADE) in the OpenAI for Government team, you’ll help government agencies transform their organization through solutions such as automated content generation, contextual search, and novel applications that make use of our newest, most exciting models and technology. About the Role We are looking for a driven solutions leader with a product mindset to partner with our public sector customers and ensure they achieve tangible value with GenAI. You will pair with government agencies (federal, state, and local), policymakers, and other public institutions to establish a GenAI strategy and identify the highest value applications. You’ll then partner with their technical teams, subject matter experts, systems integrators, and implementation partners to move from prototype through production. You’ll take a holistic view of their needs and design an architecture using the OpenAI API and other services to maximize customer value. You will collaborate closely with Sales, Solutions Engineering, Global Affairs, Applied Research, and Product teams. This role is based in Washington, DC. We offer relocation support to new employees. In this role, you will: Deeply embed with our most sophisticated public sector customers as the technical lead, serving as their technical thought partner to ideate and build novel applications on our API and other OpenAI products. Work with senior customer stakeholders to identify the best applications of AI in their industry and to build/qualify a comprehensive backlog to support their AI roadmap. Intervene directly to accelerate customer time to value through building hands-on pr
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. As an AI Reimagination Engineer in Micron's Generative AI Center of Excellence (GenAI COE), you have an outstanding chance to transform how businesses function. You will partner with different business areas to break down large, manual, multi-step processes and redesign them into effective, autonomous systems. This position combines process reimagination with AI systems engineering, making you a key part of the transformation journey. You will collaborate closely with the GenAI COE, IT architecture, security, and project teams. You will lead projects from the initial redesign to the final build, delivering solutions that business teams can adopt and scale. Responsibilities: Decompose end-to-end processes: Map current-state flows, quantify effort and risk, and lead eliminate/simplify/agentify analysis before automation. Architect the agentic solution: Build future-state flows and AI architecture, including task and agent decomposition, orchestration patterns, tool and data access, memory and context strategy, and human-in-the-loop controls. Translate inventions into buildable solutions by developing agent workflows, composing prompt and context strategies, MCP/connector and integration requirements, and evaluation criteria. Follow Micron's “Secure by Design”
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. At Snowflake, we empower both enterprises and individuals to reach their full potential. Our culture prioritizes impact, innovation, and collaboration, making Snowflake the ideal place to build ambitious projects, execute quickly, and advance technology — and your career — to the next level. The Role We are seeking a Manager, Applied Field Engineering - AI/ML Product Specialists to lead a high-performing team of Applied Field Engineers within the Applied Field Engineering organization. In this hands-on leadership role, you will manage a team of Applied Field Engineers who specialize in Generative AI, Machine Learning, and Advanced Analytics. You will be responsible for coaching your team through technical sales engagements, driving execution excellence, and ensuring customers successfully activate and consume Snowflake's AI/ML capabilities. You will translate team-level insights into feedback that shapes broader strategy, working closely with your manager and cross-functional partners to align execution with organizational priorities. Responsibilities & Focus Areas: Technical Execution & Consumption Activation: Drive team performance toward Consumption Activation — ensuring customers successfully move workloads into production and realize contracted credit value Coa
From $525.5K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Senior Director, Generative AI About the Role Roblox Build is our generative creation product, the platform where creators design, build, and publish 3D experiences. We are looking for a Senior Director of Generative AI to lead the Applied AI organization inside Build, responsible for turning state-of-the-art foundation models into high-quality, reliable creation systems at Roblox scale. This leader will own the full applied AI stack: model strategy and routing, model adaptation and fine-tuning, code generation (CodeGen), 3D layout generation (LayoutGen), and the evaluation science and infrastructure that tells us what actually works. You Will Own model strategy and routing for Build. Design and build an intelligent model layer that selects the right model for each creation task based on quality, capability, latency, cost, and safety, leveraging both frontier models and Roblox-adapted open-source models. Lead model adaptation across the Applied AI org, including fine-tuning, distillation, synthetic data generation, human feedback pipelines, and preference optimization for Roblox-specific creation tasks such as Luau code generation and 3D scene understanding. Drive CodeGen capabilities
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Where Data Does More. Join the Snowflake team. Snowflake is seeking an entrepreneurial and visionary engineering leader to build and scale our new AI Solutions team. As the Director of Engineering for AI Enterprise , you will be at the forefront of the generative AI revolution, leading a world-class engineering organization that designs and implements cutting-edge solutions on the AI Data Cloud. This is a critical, high-impact leadership role where you will partner with the world's largest companies to solve their most complex challenges and unlock transformative business value using Snowflake's powerful Cortex AI Platform. IN THIS ROLE AT SNOWFLAKE, YOU WILL: Recruit, mentor, and scale a high-performing, global engineering organization. Foster a culture of innovation, ownership, and engineering excellence. Define the long-term technical vision and organizational structure for the AI Solutions team, ensuring alignment with Snowflake's product evolution and business goals. Lead the technical design and development of advanced AI and machine learning solutions using Snowflake Cortex and Snowflake ML. Own the end-to-end implementation of the AI Solutions product line, from initial concept to enterprise-grade production deployments. Act as the senior engineering authority in cu
From $125K/yr
About Stitch Fix, Inc. Stitch Fix (NASDAQ: SFIX) Stitch Fix is redefining retail by combining human creativity with advanced data science and Generative AI. As we build the future of personalized shopping, we’re equally committed to building yours. We believe in investing in our team as much as our technology. Join us to be a trendsetter in the industry and help us redefine what’s possible for our clients, while we help you reach your full potential. About the Role As a Lead Engineer on the Product Catalog Manager Team, you will help set the technical direction for the systems that power Stitch Fix’s product data ecosystem. You will work on the tools, workflows, and data models that support the full product lifecycle, from new style creation and catalog enrichment to product readiness, validation, and downstream product experiences. You will own complex problem spaces from discovery through delivery, translate business and merchandising needs into scalable technical solutions, and lead execution across ambiguous, cross-functional initiatives. This role requires strong technical judgment, deep ownership, clear communication, and the ability to influence partners across Engineering, Product, Merchandising, Data Science, and Operations. Your work will directly impact product data quality, catalog accuracy, merchandising efficiency, product readiness, and the client experience. Responsibilities: Own and evolve critical catalog systems, including product onboarding, attribute management, data enrichment, validation workflows, and product readiness tooling. Design and operate scalable services and data models that ensure product information is accurate, complete, consistent, and available to downstream systems. Drive discovery with Product, Merchandising, Data Science, and Operations partners to identify high-impact problems, evaluate tradeoffs, and define clear technical roadmaps. Independently lead initiatives from concept through production rollout, including tec
We are seeking an experienced Senior Generative AI Developer to help drive the design, development, and integration of state-of-the-art Generative AI and agentic AI solutions across our enterprise Controls Technology platform. You will collaborate with cross-functional teams, contribute deep technical expertise in context engineering, retrieval systems, knowledge graphs, and multi-agent orchestration, and play a key role in delivering scalable, grounded AI solutions to enhance automation and operational efficiency. This role centers on architecting robust applications and agent systems on top of pre-trained and hosted foundation models — not on training or fine-tuning models. Key Responsibilities Collaborate with AI architects, leads, and stakeholders to design and implement generative and agentic AI solutions that address business challenges. Architect advanced context engineering strategies — context layering, chaining, compression, pruning/offloading, and memory management — to maximize reliability, provenance, and token efficiency in production. Design and implement advanced generative AI methods, including sophisticated prompt engineering and Retrieval-Augmented Generation (RAG) . Build and optimize RAG systems , including hybrid search, multi-vector retrieval, and re-ranking pipelines. Design and implement knowledge graphs and Graph RAG architectures to enable multi-hop reasoning, explainability, and traceable, grounded responses for high-value business domains. Architect agentic workflows and multi-agent systems using Google Agent Development Kit (ADK) and comparable frameworks (LangGraph, Microsoft Agent Framework, CrewAI), applying orchestration patterns such as supervisor/worker, hierarchical, and peer-to-peer. Design robust agent harnesses — governance, constraints, feedback loops, state/session management, and
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About this role This is a chance to help shape the next chapter of WRITER's product at a moment when generative AI is redefining how work gets done. As senior staff product manager, you'll report directly to Wil Pong, VP of Product, and operate as a force-multiplier across the product organization — setting direction where it's unclear, aligning stakeholders, and making sure our most critical initiatives ship with excellence. You'll own ambiguous, high-impact problem spaces that span multiple product areas and teams, and you'll be equal parts strategist and operator. One moment you're shaping a multi-quarter roadmap; the next you're rolling up your sleeves to unblock execution. We're looking for someone who thrives at the intersection of customer empathy, technical depth, and business judgment — and who is energized by turning complex, evolving AI capabilities into products customers love. 🦸🏻♀️ Your responsibilities Drive high-impact product initiatives end-to-end, from di
From $234K/yr
The ML Observability team builds cutting-edge tools to monitor, explain, and improve AI systems in production, particularly those leveraging Large Language Models (LLMs) and generative AI. We provide robust, scalable observability for AI workloads, including drift detection and model evaluation, and behavior tracing, enabling customers to ship AI with confidence. As a Staff Engineer, you’ll lead the development of new features and foundational capabilities within Datadog’s LLM Observability product. You will shape product direction, drive experimentation, and apply your deep understanding of both AI systems and software engineering to solve open-ended problems in the fast-moving AI landscape. Your work will directly impact how our customers monitor, troubleshoot, and optimize LLM-based applications in production. Join us in building the foundational tools that make AI systems observable, understandable, and reliable in the real world. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Drive design and implementation of LLM observability features. Ideate, prototype, and scale new product features to provide insights and drive improvements for generative AI systems Work cross-functionally with other eng teams, product, UX, and applied science to iterate fast and find product-market fit Develop and extend tools for tracing, evaluating, and debugging LLMs Influence architecture decisions and mentor engineers to build resilient, high-performance systems Stay close to customer pain points and use those insights to guide product and engineering priorities Stay current with industry trends and advancements in machine learning and observability, driving innovation within the team Who You Are: You have a BS/MS/PhD in a Computer Science, Engineering or r
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As an Engineering Manager (Player & Coach), you will lead and mentor a team of Forward Deployed Engineers focused on building, scaling, and optimizing LLM inference workloads for Baseten customers. Applying both hands-on technical ownership and managerial leadership, you will guide your team through the processes of designing, deploying, and managing high performance, low latency AI applications on Baseten’s platform. FDE at Baseten is not a sales function – we are a mix of engineering, product, and customer architects who contribute to the core Baseten codebase, drive large portions of our feature roadmap, and execute on complicated customer engagements. You will also partner with product, infrastructure, and other customer engineering teams to ensure that large language models (LLMs) and other generative AI systems deliver best-in-class performance, reliability, and cost efficiency in production environments. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Forward Deployed Engineering team: Forward Deployed Engineering on the frontier of AI The fastest, most accurate Whisper transcription Deploy production-ready model servers from Docker images Deploy custom ComfyUI workflows as APIs RESPONSIBILITIES Leadership & Team Management Lead, mentor, and grow a team of Forward Deployed Engineers, providing guidance on technical direction, project execution, and professional deve
From $82.1K/yr
About Stitch Fix, Inc. Stitch Fix (NASDAQ: SFIX) Stitch Fix is redefining retail by combining human creativity with advanced data science and Generative AI. As we build the future of personalized shopping, we’re equally committed to building yours. We believe in investing in our team as much as our technology. Join us to be a trendsetter in the industry and help us redefine what’s possible for our clients, while we help you reach your full potential. About the Role As a Lead Integration Engineer on the Business Systems Engineering team, you will collaborate with cross functional teams to understand requirements and synthesize them to build & enhance globally scalable solutions. This role requires deep technical expertise in integration architecture and is responsible for identifying issues, building & testing integrations, and administering Workday and other People platforms. You will wear multiple hats; modifying business processes, complex reporting, and providing end-user support. Build the future of People Tech: Build scalable, AI-ready solutions that automate end-to-end HR processes and unlock new capabilities. Establish and evolve automation and Gen AI frameworks purpose-built for the employee experience. Deliver technical excellence: Define system boundaries, data flows, and integration standards and design reviews. Establish integration patterns, reusable code libraries, and architectural standards; implement monitoring, logging, and alerting for all HRIS systems. Build scalable integrations: Design event-driven integration patterns, RESTful/SOAP APIs, and data pipelines connecting HRIS Systems across our technology ecosystem; leveraging your skills in Studio, Cloud Connect, EIB, iLoad, Conversion, Core Connectors, XSLT, RaaS, and Web Services. Ensure security & compliance: Build secure architectures using OAuth 2.0, encryption, and maintain SOX compliance across all integrations. Drive impact & Collaboration: Partner
From $244K/yr
Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: AI and ML are at the heart of the Airbnb product. From Trust to Payments, and from Customer Service to Marketing, we rely on ML to ensure that guests and hosts have the best possible experience with Airbnb. The Core ML team is responsible for driving CSxAI (Customer Support x Artificial Intelligence) initiatives by adopting Generative AI technologies to enable an intelligent, scalable, and exceptional service experience. The team develops and enhances AI models, ML services, and tools including LLM fine-tuning and optimization, RAG/Search, LLM evaluation and testing automation, feedback-based learning, and guardrails for a wide range of applications at Airbnb. The richness of Airbnb's data, the complexity of its marketplace, and the variety innate in our product mean that we need to operate at the state of the art of AI practice. We are committed to long-term innovation to solve complex problems, and to do that we need experienced ML The Difference You Will Make: In this Senior Staff role, you will set technical direction and lead execution for ML evaluation and the end-to-end data flywheel powering CSxAI products (e.g., assistive agents, issue resolution, and tooling). Your work will define how we measure quality, how we turn feedback into learning signals, and how we continuously improve models and products safely and efficiently. You will partner closely with product, engineering, design, operations to build evaluation systems that are trusted, scalable, and actionable - connecting offline metrics to online outcomes. A Typical Day: Define evaluation strategy and suc
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary CVS Health is seeking a Principal Software Engineer to lead the design and delivery of enterprise-scale Generative AI solutions that power next-generation healthcare experiences. This role goes beyond hands-on coding—you will define technical strategy, establish architectural standards, and guide multiple teams in building secure, scalable, and cost-effective AI platforms across AWS (Bedrock) and Google Cloud (Vertex AI API). You will partner with product, security, compliance, and enterprise architecture teams to ensure solutions meet business objectives, regulatory requirements, and performance goals. The ideal candidate combines deep technical expertise with leadership skills—capable of influencing cross-org architecture decisions, mentoring engineering teams, and driving responsible AI practices in production. Key Responsibilities Lead end-to-end platform delivery of highly scalable, secure AI services and applications leveraging AWS Bedrock (Foundation Models, Knowledge Bases, Agents, Guardrails) and Google Cloud Vertex AI (Gemini via Vertex AI API, Agent Builder, Vector Search, Search & Grounding) Architect and implement Retrieval-Augmented Generation (RAG) solutions, integrating proprietary data from sources like Amazon S3 and Google Cloud Storage/BigQuery, and using Bedrock Knowledge Bases and/or Vertex AI Search & Groundi
Other cities to consider
More places hiring for this role
Get new lead generative ai engineer jobs in United States by email
Daily job updates · Unsubscribe anytime