Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Platform Engineer An individual contributor who will serve as a hands-on technical member of the SMAI Platform Engineering team. In this role, you will help build, operate, and maintain the platforms that power Micron's analytics and AI workloads. You will contribute to platform reliability and scalability through day-to-day engineering, collaborative problem solving, and close partnership with solution architects, project teams, and multi-functional partners. Responsibilities: Collaborate with global platform teams, customers, partners, and vendors to deliver effective technical solutions. Know the latest platform roadmaps, emerging technologies, and new service offerings; evaluate and recommend adoption opportunities. Partner with solution architects to design, implement, and optimize solutions across IaaS, PaaS, SaaS, and Infrastructure as Code (Terraform). Document findings, operational procedures, guidelines, and reusable patterns while providing feedback to vendors and internal teams. Deliver high-quality platform support by managing customer requests, maintaining service standards, and implementing controlled platform adjustments. Monitor platform performance, observability, costs, and resource utilization to identify optimization opportunities and improve reliability. Collaborate with multi-functional teams to ensure seamless operations, scalable architectures, and automation, including AI-driven business solutions. Implement and m
Jobs in United States
Ai Platform Engineer in United States
5,082 active opportunities · Updated October 2026
Showing
15 jobs
Explore current ai platform engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
We're building the platform that lets long-running autonomous agents operate safely inside NVIDIA's enterprise. These are not assistants on a developer's laptop. They are fleets of agents deployed in the cloud, running continuously at scale on shared accelerated compute. They take on real work across enterprise systems, so people get far more done than they could before. This role defines the constructs that agents are built from: the blueprints they start from, the tools, skills, and plugins that power them against enterprise data, the runtime safety harness that keeps them in bounds, and the connections into credential management, sandbox, memory, and observability. The team designs and ships these building blocks so that agent developers across the company can stand up a new agent, wire it in, and run it for days or weeks. Security and safe execution come out of the box, not something each team has to get right on its own. Today an agent runs inside a single harness. Claude, Codex, and open-source agent harnesses each work differently underneath, with their own execution model, tool interface, and telemetry shape. The platform smooths over those differences, so a single skill, safety policy, or trace works the same no matter which harness is running. We want to enable agents that act on a person's behalf, governed and secure, continuously evaluated and self-improving. These agents coordinate and hand work off to each other, with identity and policy following every hop. They route and tune themselves across harnesses from live eval signals, and get better from their own production telemetry instead of waiting on a human to retrain them. Have you run agents on a harness like Claude or Codex and hit the walls that show up when they run for real, for days, against live systems — and wanted them to learn from it on their own? We're building the platform that solves those problems once, for every team. What you'll b
About the Team Business Systems / Enterprise Platform Technology builds the internal systems, data foundations, workflow infrastructure, and enterprise platforms that help OpenAI operate at scale. The EPT AI Pod builds AI-native internal apps, MCP connectors, multi-agent workflows, and reusable platform capabilities across Finance, People, and GTM. About the Role As an Enterprise Applied AI Engineer, you will build internal apps for enterprise operations and the shared platform components those apps run on. This includes MCP connectors, multi-agent orchestration, data architecture, evals, monitoring, auditability, and governance. We’re looking for a hands-on engineer who is strong in Python, system design, enterprise integrations, data architecture, and applied AI systems. You should be excited to turn ambiguous business workflows into reliable internal products and shared infrastructure. In this role, you will: • Build internal apps for enterprise operations across Finance, People, and GTM • Build MCP connectors and enterprise integrations with strong auth, permissions, idempotency, retries, and rate-limit handling • Design end-to-end multi-agent workflows with tool routing, human approvals, audit trails, and safe action boundaries • Design data architecture for operational AI systems, including ingestion, schemas, quality checks, lineage, and governance • Build evals, monitoring, metrics, and regression tests for agentic workflows • Create reusable infrastructure, patterns, and components that other enterprise teams can build on • Partner with system owners and business owners to turn messy enterprise workflows into reliable internal products You might thrive in this role if you: • Have strong Python engineering skills for backend services, MCP connectors, agent/tool workflows, eval harnesses, and data ingestion jobs • Have strong system design skills across shared infrastructure, app architecture, reliability, and scaling • Have experience building internal apps,
From $180K/yr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: Millions of people use Notion — and this number is increasing every day. That means millions of people trust us to deliver a fast, reliable, and secure experience, and we value this more than anything. We want to keep earning trust, while also continuing to amaze our users with the tools they can build in Notion. The AI Platform team is responsible for building the shared foundations that let Notion ship AI products quickly and operate them safely at scale. You’ll join a team of talented engineers focused on making speed and quality compatible: reliability and availability through provider changes, quality and correctness systems like evals and release gates, observability that makes failures explainable, and shared primitives for model integrations, context management, long-running actions, and cost/performance tradeoffs. Notion’s AI platform is vital to helping product teams move faster with production-grade guardrails as models, providers, and AI capabilities rapidly evolve. This role can be based in either San Francisco or New York City. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days)
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . Millions of people across the world come to Pinterest to find new ideas every day. It’s where they get inspiration, dream about new possibilities and plan for what matters most. Our mission is to help those people find their inspiration and create a life they love. As a Pinterest employee, you’ll be challenged to take on work that upholds this mission and pushes Pinterest forward. As a Principal Engineer on the AI Platform team, you'll help architect the infrastructure that powers both Generative AI and Recommender Systems across Pinterest's entire product suite. Our team builds the end-to-end engines for petabyte-scale data orchestration, model training and fine-tuning, and high-performance inference, ensuring our models scale seamlessly to hundreds of millions of inferences per second in service of over 600 million monthly active users.
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . Millions of people across the world come to Pinterest to find new ideas every day. It’s where they get inspiration, dream about new possibilities and plan for what matters most. Our mission is to help those people find their inspiration and create a life they love. As a Pinterest employee, you’ll be challenged to take on work that upholds this mission and pushes Pinterest forward. As a Principal Engineer on the AI Platform team, you'll help architect the infrastructure that powers both Generative AI and Recommender Systems across Pinterest's entire product suite. Our team builds the end-to-end engines for petabyte-scale data orchestration, model training and fine-tuning, and high-performance inference, ensuring our models scale seamlessly to hundreds of millions of inferences per second in service of over 600 million monthly active users.
About the Team OpenAI's mission is to ensure that AGI benefits all of humanity. The Business Systems team helps make that mission possible by building the internal products and platforms that allow OpenAI to operate with speed, reliability, and care. We build internal applications and workflows for Finance and Supply Chain. Our work spans product discovery, React and TypeScript interfaces, Python services and APIs, data models, workflow orchestration, enterprise integrations, and the systems that connect people to systems of record. We work directly with the people who use these products and care about correctness, permissions, auditability, and production reliability. Examples of our work include building an integration platform for supply chain integrations, integrations with Oracle Fusion and Zip, contract intelligence applied to B2B revenue recognition, and Temporal-based agentic workflows for credit checks, duplicate bank detection, and invoice triaging. We turn these efforts into reusable patterns that can support many workflows, rather than one-off automations. About the Role We are looking for Product Engineers to build internal applications end to end. This role spans product discovery, user experience, frontend, backend services, data models, workflow orchestration, and integrations with order management, fulfillment, and supply chain systems. You will take a problem from a first conversation with a Finance or Supply Chain partner through design, implementation, rollout, and production support. Strong candidates combine product judgment with engineering depth. You should be comfortable moving between a React interface, a Python API, a durable workflow, and an integration with an enterprise system. You should be able to ship a useful first version quickly while building the foundations for reuse, security, and long-term maintainability. Direct AI experience is helpful, but the core requirement is strong product engineering judgment and reliable execution. I
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. POSITION SUMMARY CVS Health is seeking a Senior Full Stack Engineer, Agentic AI Platform to join and lead engineering efforts for Lumina, CVS Health's enterprise knowledge and agentic AI platform. The Senior Full Stack Engineer, Agentic AI Platform will provide both technical and people leadership while driving the evolution of a platform that delivers trusted, secure, and governed AI experiences across the enterprise. As a Senior Full Stack Engineer, Agentic AI Platform, you will own the technical direction, architecture, scalability, and operational excellence of Lumina while leading a team of engineers and data scientists responsible for building and supporting the platform. This is a hands-on leadership role requiring active contribution to production code, architectural decision-making, code reviews, and engineering best practices, while simultaneously coaching and developing team members. The Senior Full Stack Engineer, Agentic AI Platform will play a critical role in advancing agentic AI capabilities, retrieval-augmented generation (RAG) systems, Model Context Protocol (MCP) integrations, enterprise search, knowledge ingestion, and AI governance. You will drive platform enhancements that improve retrieval quality, strengthen security and access controls, expand automation capabilities, and ensure the platform remains reliable, observable, and scalable as adoption accelerates across CVS
[2026] Senior Machine Learning Engineer (Systems), Embodied AI/NPCs, ML Platform - PhD Early Career
RobloxFrom $196.8K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Team Creator Services Machine Intelligence Team : The Machine Intelligence team is building an NPC system that can (1) play any Roblox game and (2) perform real-time inference efficiently enough to support deployment to all Roblox players. ML Platform Team : The Foundation AI Group is on a mission to establish Roblox as the standard for 3D foundational models (3DFMs), democratizing creation by making it simple for anyone to generate high-quality, immersive 3D experiences using AI. The AI Platform team is a foundational part of this vision, supporting hundreds of ML use cases and billions of inferences daily across Discovery, Safety, Engine, and more. We are seeking exceptional PhD new graduates to drive innovation across three critical areas: AI Platform, Distributed Inference Systems. What You Will Do As a Senior Machine Learning Engineer, you will be a key contributor to building the cutting-edge systems that power AI at Roblox. Creator Services Machine Intelligence Team Develop Scale Data Pipelines: Design, build and maintain robust data pipelines to collect complex 3D game states and real-time player actions across the platform. Train Novel Architectures: Solve the feature e
[2026] Senior Machine Learning Engineer (Systems), Embodied AI/NPCs, ML Platform - PhD Early Career
RobloxFrom $196.8K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Team Creator Services Machine Intelligence Team : The Machine Intelligence team is building an NPC system that can (1) play any Roblox game and (2) perform real-time inference efficiently enough to support deployment to all Roblox players. ML Platform Team : The Foundation AI Group is on a mission to establish Roblox as the standard for 3D foundational models (3DFMs), democratizing creation by making it simple for anyone to generate high-quality, immersive 3D experiences using AI. The AI Platform team is a foundational part of this vision, supporting hundreds of ML use cases and billions of inferences daily across Discovery, Safety, Engine, and more. We are seeking exceptional PhD new graduates to drive innovation across three critical areas: AI Platform, Distributed Inference Systems. What You Will Do As a Senior Machine Learning Engineer, you will be a key contributor to building the cutting-edge systems that power AI at Roblox. Creator Services Machine Intelligence Team Develop Scale Data Pipelines: Design, build and maintain robust data pipelines to collect complex 3D game states and real-time player actions across the platform. Train Novel Architectures: Solve the feature e
About the team The Monetization Data Platform team builds the trusted data and platform foundations that power how the company develops, measures, and improves monetization products. We bring together product usage, pricing, billing, ads, payments, and financial data to help Product, Engineering, Finance, and GTM teams make better decisions and deliver reliable customer experiences. We work at the intersection of data engineering, product engineering, platform engineering, Finance, and GTM. Our goal is to turn complex monetization and financial data into accurate, explainable, and timely data products while building systems that scale with the growth and complexity of the business. About the role We are looking for a Data Engineer to improve and build the next generation of our monetization data platform. You will own high-impact systems end to end, from product instrumentation, source ingestion, and canonical modeling through quality controls, observability, and delivery to downstream consumers. This is a hands-on role for an engineer who enjoys solving ambiguous product and data problems, designing durable architectures, and partnering closely with Product Engineering, Finance, Accounting, and GTM. You will help define technical direction, raise the engineering bar, and turn monetization opportunities into trusted, scalable data products and platform capabilities. In this role, you will Design, build, and operate large streaming and batch data pipelines that process product, financial, and operational data from a variety of internal and external systems. Develop canonical data models and reusable data products for domains such as product usage, pricing, billing, ads, payments, revenue, and the general ledger. Establish strong guarantees for data accuracy, completeness, freshness, lineage, reconciliation, and auditability. Build frameworks and platform capabilities that improve developer productivity and make it easier for teams to launch, measure, and iterate on m
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Cloud Platform Engineer, you'll envision and build robust systems and processes that ensure our infrastructure is scalable, reliable, and efficient. This can range from automating deployments and monitoring systems to optimizing performance and managing incidents. We all work closely with our users, learning from their past struggles in operationalizing ML, onboarding them onto our platform, and turning our learnings into ideas for improving Baseten. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Infrastructure team: Multi-cloud capacity management Inference on B200 GPUs Multi-node inference Fractional H100 GPUs for efficient model serving RESPONSIBILITIES Build and maintain scalable infrastructure to support the deployment and operation of machine learning models. Establish standards and best practices for reliability and performance across the infrastructure. Automate processes when relevant, particularly for managing CI/CD pipelines. Own products and projects end-to-end, functioning as both an engineer and a project manager, with a focus on user empathy, project specification, and end-to-end execution. Collaborate with cross-functional teams to understand project requirements and translate them into technical solutions. Mentor junior team members and contribute to knowledge sharing within the organization. Navigate ambiguity and exercise good judgment on tradeoffs and
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: At Modal, we sell cloud services atop which our customers run their critical production systems. As a rapidly growing new cloud infrastructure company, we seek to improve our reliability dramatically while scaling the size of our platform, customer base, and our team. This role is for people who are deep systems thinkers, love stacking nines, and thrive from making others move faster at scale. Responsibilities include: Identifying architectural changes to improve reliability and performance. Fostering a culture of reliability across Modal’s engineering organization. Defining and implementing operational processes such as deployments, upgrades, etc. Operating systems like Kubernetes, Postgres, Redis, etc. Participating in on-call rotations, and responding to production incidents. Requirements: 5+ years of experience writing high-quality production code. 2+ years of
Citi, the leading global bank, has approximately 200 million customer accounts and does business in more than 160 countries and jurisdictions. Citi provides consumers, corporations, governments, and institutions with a broad range of financial products and services, including consumer banking and credit, corporate and investment banking, securities brokerage, transaction services, and wealth management. As a bank with a brain and a soul, Citi creates economic value that is systemically responsible and in our clients’ best interests. As a financial institution that touches every region of the world and every sector that shapes your daily life, our Enterprise Operations & Technology teams are charged with a mission that rivals any large tech company. Our technology solutions are the foundations of everything we do from keeping the bank safe, managing global resources, and providing the technical tools our workers need to be successful to designing our digital architecture and ensuring our platforms provide a first-class customer experience. We reimagine client and partner experiences to deliver excellence through secure, reliable, and efficient services. Our commitment to diversity includes a workforce that represents the clients we serve from all walks of life, backgrounds, and origins. We foster an environment where the best people want to work. We value and demand respect for others, promote individuals based on merit, and ensure opportunities for personal development are widely available to all. Ideal candidates are innovators with well-rounded backgrounds who bring their authentic selves to work and complement our culture of delivering results with pride. If you are a problem solver who seeks passion in your work, come join us. We’ll enable growth and progress together. Position Overview: The Senior Platform Engineering Lead is a pivotal senior-level engineering position responsible for driving the
From $136K/yr
About Stitch Fix, Inc. Stitch Fix (NASDAQ: SFIX) Stitch Fix is redefining retail by combining human creativity with advanced data science and Generative AI. As we build the future of personalized shopping, we’re equally committed to building yours. We believe in investing in our team as much as our technology. Join us to be a trendsetter in the industry and help us redefine what’s possible for our clients, while we help you reach your full potential. About the Role As an ML Platform Engineer at Stitch Fix, you will play a key role in building and maintaining the critical infrastructure that powers machine learning and AI across our organization. You will design, develop, and support scalable, resilient services and frameworks for ML model training and deployment, feature engineering and serving, candidate generation, AI agent deployment and observability, and other core platform capabilities. In this role, you'll contribute to the day-to-day operations of the ML Platform team, ensuring the smooth functioning of existing systems while driving improvements. You’ll collaborate closely with full-stack data scientists, offering consultation and support to help them unlock the full potential of our platform. With significant autonomy, you’ll have the opportunity to shape the future of ML and AI at Stitch Fix. Your ideas and expertise will drive improvements, codify best practices, and influence how we approach machine learning and AI systems at scale. Responsibilities: Collaborate with cross-functional teams, including data scientists, engineers, and business partners, to solve complex distributed systems and business challenges at scale. Be part of a team with high visibility across the organization, driving impactful solutions that make a difference. Share your ideas and help guide the team’s investments toward high-value opportunities. Foster a culture of technical collaboration and contribute to the development of scalable, resilient systems. About You You bring
Other cities to consider
More places hiring for this role
Get new ai platform engineer jobs in United States by email
Daily job updates · Unsubscribe anytime