🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role AI research at WRITER isn't just about publishing papers — it's about building the scientific foundation that powers some of the most ambitious enterprise AI deployments in the world. As an AI research scientist, you'll be at the center of that work. You'll drive a high-impact research agenda focused on large language models, agentic reasoning, and the system-level capabilities that make AI genuinely useful at enterprise scale. This is a rare opportunity to do research that matters twice over — advancing the field and shipping directly into products used by hundreds of thousands of people every day. We're at an inflection point. Enterprises are moving from experimenting with AI to deeply embedding it across their operations, and WRITER's models are the engine making that possible. The work you do here — on post-training, planning, multi-step reasoning, and agentic workflows — will directly shape how the next generation of enterprise AI behaves, performs, and scales. You
Jobs in United States
Ai Deployment Engineer in United States
5,082 active opportunities · Updated October 2026
Showing
15 jobs
Explore current ai deployment engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role AI research at WRITER isn't just about publishing papers — it's about building the scientific foundation that powers some of the most ambitious enterprise AI deployments in the world. As an AI research scientist, you'll be at the center of that work. You'll drive a high-impact research agenda focused on large language models, agentic reasoning, and the system-level capabilities that make AI genuinely useful at enterprise scale. This is a rare opportunity to do research that matters twice over — advancing the field and shipping directly into products used by hundreds of thousands of people every day. We're at an inflection point. Enterprises are moving from experimenting with AI to deeply embedding it across their operations, and WRITER's models are the engine making that possible. The work you do here — on post-training, planning, multi-step reasoning, and agentic workflows — will directly shape how the next generation of enterprise AI behaves, performs, and scales. You
From $193K/yr
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity Every company shipping AI right now has the same problem: they can't see what their AI is actually doing in production. New Relic is building the answer, AI Observability, and it's one of the company's top strategic initiatives. We're running it like a startup inside the company: small teams, fast iteration, direct executive sponsorship, and a mandate to ship. As the Principal Product Manager and business owner of this initiative, you'll take it from vision through GA, then keep pushing past GA, where most of the real work actually happens, working directly with some of the largest AI deployments in the world. What you'll do Own the AI Observability roadmap and the business outcomes leadership tracks (adoption, retention, expansion), plus the parts of the job that aren't purely product: pricing conversations, exec reporting, and cross-functional escalations. Drive multiple engineering and design teams in parallel toward a coherent product vision, without formal authority over them. Partner directly with strategic enterprise customers, including their platform and infra teams deciding how to operationalize AI safely, and turn what breaks at their scale into product requirements fast enough to matter. Build working prototypes yourself using agentic AI workflows, so ideas get pressure-tested before engineering time is committed to them. Represent AI Observability in executive strategy and prioritization decisions, and make the tradeoff calls when prioritie
About the team The OpenAI for Government team is a mission-driven group bringing frontier AI to the U.S. Intelligence Community (IC) and broader national security enterprise. We partner with intelligence professionals to deploy secure, compliant AI capabilities—including ChatGPT Enterprise, ChatGPT Gov, and the OpenAI API—in mission-aligned environments that meet rigorous security, privacy, reliability, and responsible-AI requirements. As part of the OpenAI for Government team, you will work across IC elements and mission partners—including ODNI, CIA, NSA, DIA, NGA, NRO, and other federal intelligence organizations—to help analysts, collectors, operators, and technical teams apply frontier AI to their highest-priority missions. We accelerate adoption through hands-on support, mission-specific workflows, tailored enablement, and measurable operational outcomes while preserving human judgment, analytic integrity, and protection of sensitive information. About the Role The Strategic Delivery Lead (SDL) for the Intelligence Community will identify, shape, and deliver OpenAI’s highest-impact deployments for U.S. intelligence organizations. You will help IC customers translate mission priorities—including all-source analysis, intelligence production, collection management, open-source and geospatial exploitation, cyber threat analysis, knowledge discovery, and analyst productivity—into secure, technically feasible AI deployments built on the OpenAI API and enterprise products. The role is highly cross-functional and customer-facing: you will partner with Research, Engineering, Go-to-Market, Security, Legal, Privacy, and government mission, acquisition, and technical stakeholders to align on scope, remove delivery obstacles, and communicate progress from working teams to agency leadership. You will own end-to-end execution of technical delivery workstreams led by Forward Deployed Engineers, Researchers, Solutions Architects, and Enablement teams. You will establish mission
About the team OpenAI's mission is to build safe artificial general intelligence (AGI) which benefits all of humanity. This long-term undertaking brings the world's best scientists, engineers, and business professionals into one lab together to accomplish this. In pursuit of this mission, our Go To Market (GTM) team helps customers understand, adopt, and deploy OpenAI's technology across their most important workflows. The team is made up of Sales, Solutions, Support, Marketing, and Partnerships professionals who work together to bring AI to as many people and organizations as possible. About the role As an Account Director focused on Life Sciences, you will own executive-level relationships with organizations across biopharma, medical technology, and life science services. This includes pharmaceutical and biotechnology companies, medical device and diagnostics companies, contract research organizations, contract development and manufacturing organizations, research tools companies, and laboratory or scientific services providers. This is a role for someone who understands how life sciences companies create, validate, manufacture, commercialize, and support scientific and medical products. You will help these customers evaluate and deploy OpenAI's technology to accelerate R&D, improve clinical and regulatory workflows, scale scientific and medical content operations, modernize commercial and field teams, and improve knowledge work across highly regulated environments. The role blends scientific literacy, enterprise sales discipline, technical curiosity, and relationship-driven account leadership. You will partner closely with solutions, research, product, legal, security, and compliance teams to design secure, responsible, and high-impact AI deployments for life sciences customers. This role can be remote but New York or San Francisco office locations are preferred. We use a hybrid work model of three days in the office per week and offer relocation assistance t
About the team OpenAI’s mission is to build safe artificial general intelligence (AGI) which benefits all of humanity. This long-term undertaking brings the world’s best scientists, engineers, and business professionals into one lab together to accomplish this. In pursuit of this mission, our Go To Market (GTM) team is responsible for helping customers learn how to leverage and deploy our highly capable AI products across their business. The team is made of Sales, Solutions, Support, Marketing, and Partnership professionals that work together to create valuable solutions that will help bring AI to as many users as possible. About the role As an Account Director focused on Healthcare, you will own executive-level relationships with leading healthcare organizations, including integrated delivery networks (IDNs), health systems, academic medical centers, health insurers, digital health companies, and healthcare technology providers. You’ll help these organizations safely and effectively deploy OpenAI’s technology to improve clinical and administrative workflows, enhance patient and provider experiences, automate operational processes, and accelerate enterprise-wide AI adoption. This role blends enterprise sales expertise, technical depth, business acumen, and relationship-driven selling. You will collaborate closely with researchers, engineers, and healthcare-focused solution strategists to design secure, compliant, and high-impact AI deployments. This role is based in San Francisco. We use a hybrid work model of three days in the office per week and offer relocation assistance to new employees. In this role, you’ll: Manage a focused portfolio of large healthcare provider, payer, and healthcare technology accounts, developing long-term strategic account plans. Lead complex, multi-stakeholder sales cycles spanning clinical, operational, IT, digital transformation, and executive leadership teams. Partner with Solutions and Research Engineering to design pilots that demon
NVIDIA is looking for a hands-on Solutions Architect Manager to lead a team of GPU, networking & software solution architects and engineers. Do you want to build and lead a group that designs, debugs, and deploys new AI hardware and software technologies into production in customer data centers? As part of the NVIDIA SA organization, you will drive people and technical leadership for end-to-end solutions deployments at some of NVIDIA's most strategic technology customers, while directly contributing to designs and deep-dive debugging and shaping our product roadmap with customer feedback. What you will be doing: Recruit & manage a team of solutions architects, system/network and software engineers focused on large-scale GPU and AI networking deployments. Set priorities, allocate resources, mentor, and ensure high-quality customer delivery across multiple concurrent projects - while remaining directly involved in key technical reviews, design decisions, and critical debug efforts. Provide deep subject-matter expertise in advanced GPU and network systems and serve as the senior technical point of contact for strategic customers. Personally lead and guide complex compute/network configuration and performance debugging, working side-by-side with your team to deliver performant, reliable clusters. Guide your team as they lead network / compute / software architecture discussions, and support server, network, and cluster bring-up, including on-site data center work where needed. Systematically collect and synthesize customer-specific requirements across your portfolio. Partner with GPU/Network Systems Engineering, Product Management, and Sales to influence roadmap priorities and packaging of reference designs and solutions. Demonstrate SME in advanced GPU & network systems and be a trusted technical advisor to NVIDIA's strategic customers. Bring customer-sp
$150K – $200K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: As a Software Engineer on Collections Infra, you’ll help scale the infrastructure behind Notion’s database block. Databases power views, boards, charts, forms, and the workflows customers use to run their companies; as Notion grows into larger enterprise deployments and AI-native work, those systems need to become faster, more reliable, and ready for much higher concurrency. This team sits between infrastructure and product, building backend architecture that improves the customer experience while becoming a platform other Notion teams can build on. You’ll work on problems like database load performance, microservices, high-volume writes from agents, and the flexibility that makes Notion databases powerful and hard to scale. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days. What You'll Achieve: By your 90th day, you’ll understand how Collections Infra supports Notion’s database
About the Team OpenAI’s Applied AI Engineering team helps organizations turn frontier AI capabilities into safe, reliable, and high-impact production systems. We work with customer executives, product and engineeriIng teams, security leaders, and transformation teams to identify valuable opportunities, accelerate technical implementation, and scale what works. Enterprise deployments are defined by complexity rather than any one industry: existing architectures, diverse data environments, security and governance requirements, multiple stakeholder groups, and organization-wide change. We turn lessons from these deployments into better products and reusable patterns for customers everywhere. About the Role As an Enterprise Applied AI Engineer you will partner directly with leading organizations to design, build, and deploy AI systems that deliver measurable business outcomes. You will combine deep technical judgment, hands-on engineering, and customer leadership to take ambitious ideas from use-case selection and architecture through prototyping, evaluation, production launch, and scale. You will write and debug code, build evaluation systems, resolve complex integrations, and guide decisions involving model behavior, reliability, latency, cost, safety, security, governance, and operational readiness. Success is measured by production systems, sustained adoption, and meaningful customer impact—not simply activity or successful demonstrations. This is a rare opportunity to work on consequential real-world deployments at the frontier of AI while directly influencing how OpenAI’s products evolve. This role is based in our SF or NYC office. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Partner directly with enterprise customers to identify high-value opportunities and translate them into technical architectures, implementation plans, evaluation strategies, and measurable success criteri
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. About our Team: Micron’s Industrial and Physical AI team is driving the transformation of semiconductor manufacturing through Autonomous Operations, AI, robotics, and digital twin technologies! We develop and deploy innovative solutions across Micron’s global fabrication and assembly/test facilities, enabling smarter, safer, and more efficient operations at scale. Position Overview: We are seeking a hands-on Full-Stack AI Engineer to design, build, and deploy production-grade AI applications that support Micron's Autonomous Operations initiatives. This role owns the end-to-end development lifecycle, from data pipelines and AI models to APIs, web applications, digital twin integrations, and cloud/edge deployments, delivering impactful solutions for engineers, operators, and business leaders worldwide. Responsibilities: Design, architect, and deliver end-to-end AI products, including data ingestion pipelines, feature engineering, model training/inference, APIs, user interfaces, and application monitoring. Build and maintain modern front-end applications using React, Angular, or Streamlit, supported by backend services in Python and FastAPI. Develop scalable integrations between manufacturing systems, robotics platforms, AMRs, sensor networks, and enterprise applications to enable intelligent factory operations. Design and implement digital twin environments using platforms such as NVIDIA Omniverse, Gazebo, or Unity Robotics Hub to support simulation, validation, and o
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Where Data Does More. Join the Snowflake team. Snowflake is seeking an entrepreneurial and visionary engineering leader to build and scale our new AI Solutions team. As the Director of Engineering for AI Enterprise , you will be at the forefront of the generative AI revolution, leading a world-class engineering organization that designs and implements cutting-edge solutions on the AI Data Cloud. This is a critical, high-impact leadership role where you will partner with the world's largest companies to solve their most complex challenges and unlock transformative business value using Snowflake's powerful Cortex AI Platform. IN THIS ROLE AT SNOWFLAKE, YOU WILL: Recruit, mentor, and scale a high-performing, global engineering organization. Foster a culture of innovation, ownership, and engineering excellence. Define the long-term technical vision and organizational structure for the AI Solutions team, ensuring alignment with Snowflake's product evolution and business goals. Lead the technical design and development of advanced AI and machine learning solutions using Snowflake Cortex and Snowflake ML. Own the end-to-end implementation of the AI Solutions product line, from initial concept to enterprise-grade production deployments. Act as the senior engineering authority in cu
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity The Forward Deployed Engineering (FDE) team tackles some of Postman's most strategic technical challenges. We partner directly with enterprise customers to solve problems that don't fit neatly into existing product boundaries, building production-grade solutions that often become core product capabilities. As a Sr. Forward Deployed Engineer, you'll operate at the intersection of engineering, product, and customer success. You'll deploy Postman's critical infrastructure into customer environments, solve complex distributed systems challenges, and build the enterprise foundations that enable customers to safely adopt AI at scale. This isn't consulting or professional services. You're an engineer building production systems alongside customers, turning real-world deployments into durable platform capabilities used by thousands of organizations. This team will build 0-to-1 products from scratch - innovative, strategic initiatives designed to unlock hundreds of millions of dollars in new revenue. If you enjoy difficult engineering problems, working directly with customers, and building products from the field
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. The Customer Engineering organization serves as the primary technical interface for our global customers and helps ensure that Micron solutions are seamlessly integrated into next-generation technology platforms. Our team coordinates deep engineering engagements, technical product qualifications, and architecture reviews to secure design wins and grow market share. We are dedicated to providing world-class technical support that accelerates revenue and builds long-term strategic partnerships with our customers. Role Description: The Field Applications Engineer - Associate plays a vital role in achieving the technical achievements needed to launch products. You will be an important member of our field engineering team. In this position, you will oversee many end-to-end technical tasks, including sophisticated product sample deployments, lifecycle management, addressing technical questions, and other customer interactions. You should adopt an "automation-first" approach to actively find ways to improve processes and apply AI-based workflows in daily operations when possible. This role can lead to a full-time FAE career for individuals who achieve the right results. It also requires demonstrating the right skills and behaviors. Example Responsibilities: Coordinate the entire process of tech
About the Team: OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. In this role you will: As a Hardware Test Engineer, you will work on Machine Learning/AI hardware system projects to craft the solutions for current and future data center deployments. You will bring a strong understanding of hardware system testing, excellent project management skills, and the ability to collaborate across multiple teams to ensure efficient lab operations. You will be responsible for designing, implementing, and executing comprehensive test plans that ensure the reliability, performance, and scalability of our supercomputing hardware systems. You will develop detailed test plans and methodologies tailored to hardware components, including processors, memory modules, custom accelerators and interconnects. You will collaborate with hardware design, manufacturing, firmware teams and vendors to identify, analyze, and resolve issues affecting hardware, power, thermal and high-speed interconnects. You will perform in-depth debugging on the hardware system Excellent analytical skills to diagnose hardware issues, troubleshoot problems, and propose solutions. Ability to interpret complex test data, identify trends, and draw meaningful conclusions. High-speed links, with a focus on SerDes (Serializer/Deserializer) technology to assess signal integrity, error rates, and overall link performance. You will collaborate with the lab manager to maintain the equipment and hardware systems, including oscilloscopes, thermal test chambers, liquid cooling systems, and other mea
We're looking for a Principal Software Engineer to join our CSP Engagements team as the technical focal point for GPU firmware and GPU system software, working directly with engineering teams of key CSP / hyperscale customers to ensure they can reliably manage, update, and operate NVIDIA GPU firmware at fleet scale. You will drive work streams with engineering teams of key CSPs/hyperscale customers to build shared understanding of GPU firmware and system software integration, incorporate their feedback into NVIDIA's feature roadmap and delivery plan, and ensure customer-side automation and recovery procedures are ready before each firmware release. Your cross-CSP visibility enables you to identify patterns in GPU firmware operational challenges that drive systemic improvements no single customer engagement could surface alone. What you'll be doing: Drive GPU firmware & siftware work streams with CSP engineering teams — ensuring they understand GPU firmware architecture (VBIOS, InfoROM, microcontroller firmware), update sequencing, recovery procedures, and GPU power management Gather and synthesize CSP feedback on GPU firmware/software — covering manageability, observability, security requirements (e.g., multi-tenancy isolation, secure boot, attestation), and performance — and champion those priorities into NVIDIA's GPU firmware/software feature roadmap and delivery plan Drive GPU firmware update orchestration for large-scale deployments — multi-GPU update sequencing, rollback strategy, failure handling, and validation across hundreds of GPUs per rack Serve as the technical focal point between NVIDIA and CSP firmware/software engineering — ensuring GPU behaviors (error recovery flows, thermal protection, power state transitions) are well-documented and accessible for customer integration Identify cross-CSP GPU SW/FW issue patterns — common update failu
Other cities to consider
More places hiring for this role
Get new ai deployment engineer jobs in United States by email
Daily job updates · Unsubscribe anytime