Jobs in United States

Aws And Tooling Platform Lead in United States

2,078 active opportunities · Updated October 2026

Explore current aws and tooling platform lead jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -83.9%
Quick readStrong listing-quality and freshness signals

About the Role OpenAI’s Industrial Compute organization is responsible for ensuring our compute infrastructure scales efficiently to support millions of users and increasingly sophisticated AI models. We’re looking for a Data Scientist to partner closely with Capacity Systems Engineering, Infrastructure, Product, and Research to optimize inference capacity across our global GPU fleet. This role combines statistical modeling, large-scale data analysis, forecasting, and systems thinking to drive critical decisions around infrastructure investments, performance-efficiency trade-offs, and customer experience. You’ll transform complex operational data into actionable insights that directly influence how OpenAI allocates and scales one of the world’s largest AI compute environments. Key Responsibilities Build statistical and machine learning models to profile and improve GPU utilization, latency, throughput, and overall fleet efficiency. Develop forecasting models for inference demand across products, regions, and model families. Analyze production workloads to identify latency bottlenecks and capacity constraints, highlighting optimization opportunities. Partner with Capacity Systems Engineering to inform infrastructure planning and long-term GPU investment strategies. Design experiments and simulations to evaluate scheduling policies, serving strategies, and infrastructure tradeoffs. Build dashboards and operational metrics that enable leadership to make data-driven capacity decisions. Collaborate with Product, Research, Finance, and Infrastructure teams to align compute planning with business growth and model roadmaps. Communicate technical findings clearly to both engineering teams and executive leadership. Qualifications MS or PhD in Statistics, Computer Science, Operations Research, Applied Mathematics, Economics, or related quantitative discipline (or equivalent industry experience). 5+ years of experience working in the infrastructure data science space. Strong ex

PythonSQLAWSRest
O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -83.9%
Quick readStrong listing-quality and freshness signals

About the Team The Applied organization brings OpenAI’s most advanced technology to the world through products like ChatGPT and the APIs that power a growing ecosystem of developer and enterprise applications. Data Engineering builds and operates the trustworthy, secure, and reliable data systems that power decisions across OpenAI. About the Role We’re looking for a Data Engineering Manager to lead the Growth & Revenue data engineering team. This leader will own the data strategy and execution for the data subject areas spanning growth accounting across all product surfaces, product partnerships, checkout, billing, payments, revenue, and monetization, helping OpenAI understand how people adopt, engage with, and pay for our products. You will partner closely with several Data Science, Business, and Engineering partners to connect product behavior to trustworthy subscriber, payment, and revenue measurement. In this role, you will: Build, manage, and grow a high-performing, inclusive team across the Growth & Revenue data subject areas. Define the data strategy for all the data subject areas you own. Deliver durable, well-modeled data products that connect product behavior, subscription state, checkout events, payment outcomes, and revenue. Establish trusted metric definitions and data quality standards so product, growth, finance, and executive leaders can make fast, consistent decisions. Partner with Data Science and Product teams to support experimentation, causal measurement, funnel analysis, and scalable self-serve analytics. Partner with Finance and Financial Engineering to ensure analytical revenue views reconcile to financial truth and production billing systems. Raise operational excellence for critical pipelines, including reliability, observability, privacy, governance, and incident response. Set a clear roadmap, make principled tradeoffs, and communicate progress and risk across technical and business stakeholders. You might thrive in this role if yo

PythonSQLAWSRest
O
📍 San Francisco, California, United States· Full-time· Remote
✓ High-confidence listingCompany trend -83.9%
Quick readStrong listing-quality and freshness signals

About the Team OpenAI’s Legal team helps advance our mission by tackling novel legal issues in AI. Our team brings together professionals across technology, privacy, intellectual property, corporate, employment, tax, regulatory, and litigation. Our regulatory compliance work turns legal requirements into practical programs that support responsible AI development and deployment. About the Role As a Legal Program Manager focused on regulatory compliance, you will build and manage cross-functional programs that translate counsel’s guidance into practical, sustainable operations. Your initial focus may include content moderation and/or frontier AI governance, with the mix shaped by team priorities and your strengths. You will partner with internal and external counsel, other legal program managers, and technical and business teams to coordinate implementation, evidence collection, reporting, and ongoing compliance. You’ll build repeatable systems that scale across regulations, products, and jurisdictions, helping teams navigate emerging requirements with clarity and sound judgment. This full-time role is based in San Francisco, CA, or New York, NY. In this role, you will: Lead regulatory compliance programs end to end: define scope, owners, milestones, dependencies, risks, and escalation paths, and drive execution with counsel and cross-functional partners. Translate counsel’s regulatory guidance into repeatable workflows, controls, and documentation. Depending on your portfolio, this may include content moderation disclosures, transparency reporting, reporting and appeals workflows, or frontier AI model launch readiness, evaluation and risk-management evidence, and incident reporting. Build strong partnerships across Product, Engineering, User Operations, Governance, Risk and Compliance (GRC), Global Affairs, Communications, and Go-to-Market to align program priorities and deliverables. Support regulatory inquiries, audits and investigations with counsel, organizing ev

SQLAWSRestAI
P
📍 New York, New York, United States· Full-time· Remote
✓ High-confidence listingCompany trend -73.5%
Quick readStrong listing-quality and freshness signals

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Seattle, Washington D.C., Raleigh, London, and Amsterdam. Fraud Data is the data science and machine learning team within Plaid’s Fraud organization, responsible for using data and ML to improve and scale Plaid’s fraud products. Within Fraud Data, the Customer & Product Intelligence team focuses on understanding product performance, uncovering customer insights, and enabling go-to-market teams with data-driven solutions. The team partners closely with customers and GTM teams on fraud analyses and proofs of concept, turning customer learnings into scalable, reusable product capabilities. We also build the metrics, analytics, and data foundations that measure product health, identify opportunities for improvement, and guide product decisions across Plaid’s Fraud portfolio. As a Data Science Manager, you will lead a team responsible for customer-facing data science and Fraud product analytics. You will set the team's roadmap, develop its data scientists, and remain involved in analytical methods, technical reviews, and customer investigations. You will: Set a 6–12-month roadmap with Product, Engineering, and GTM, and assign priorities and responsibilities across the team. Define product metrics, their underlying data, and reporting and

PythonSQLAWSMachine Learning
O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 Washington, District of Columbia, United States· Full-time· Remote
✓ High-confidence listingCompany trend -83.9%
Quick readStrong listing-quality and freshness signals

About the Team The mission of the Applied AI Engineering team is to enable the secure and impactful implementation of GenAI solutions. We serve as technical thought partners and trusted advisors to our clients, ideating high-value use cases and providing the hands-on guidance necessary to drive projects into production. In this role within the Government team, you will empower agencies to evolve their operations through automated content synthesis, advanced search capabilities, and bespoke applications leveraging our latest foundational technologies and models. About the Role We are looking for a driven solutions leader with a product mindset to partner with our public sector customers and ensure they achieve tangible value with GenAI. You will pair with government agencies (federal, state, and local), policymakers, and other public institutions to establish a GenAI strategy and identify the highest value applications. You’ll then partner with their technical teams, subject matter experts, systems integrators, and implementation partners to move from prototype through production. You’ll take a holistic view of their needs and design an architecture using the OpenAI API and other services to maximize customer value. You will collaborate closely with Sales, Solutions Engineering, Global Affairs, Applied Research, and Product teams. This role is based in Washington, DC. We offer relocation support to new employees. In this role, you will: Deeply embed with our most sophisticated public sector customers as the technical lead, serving as their technical thought partner to ideate and build novel applications on our API and other OpenAI foundational technologies like Codex. Work with senior customer stakeholders to identify the best applications of GenAI in their industry and to build/qualify a comprehensive backlog to support their AI roadmap. Intervene directly to accelerate customer time to value through building hands-on prototypes and/or by delivering impactful strate

TypeScriptPythonAWSRest
O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -83.9%

About the Team OpenAI Consumer Devices is building the next generation of products that bring powerful AI into people’s everyday lives. Guided by OpenAI’s mission to ensure AGI benefits all of humanity, our team combines world-class researchers, engineers, designers, and operators who care deeply about creating useful, intuitive, and responsible technology. You’ll have the opportunity to work alongside exceptional people on ambitious, zero-to-one challenges at the intersection of hardware, software, and AI. This is a chance to help define an entirely new category of products—and shape how people experience AI in the future. The Systems Integration team is critical in this mission, turning complex hardware-software development into reliable product signals. We validate complete device experiences across software, cloud services, connectivity, accessories, and real-world operating environments, combining hands-on system testing, structured test development, hardware-in-the-loop environments, diagnostics, and automation to uncover issues that component-level testing alone cannot reveal. About the Role As a Systems Test Engineer, End-to-End Validation , you will design and execute end-to-end testing for complex device experiences spanning hardware, software, connectivity, cloud services, and accessories. You’ll translate product behavior and real-world use cases into structured, reproducible test procedures and build test environments that allow failures to be reliably reproduced and diagnosed. You’ll also identify opportunities to automate repetitive or high-value scenarios, working with engineers to turn complex manual workflows into scalable validation systems. Because this is a new category of devices, you’ll have the opportunity to build the end-to-end validation foundation early—shaping test coverage, environments, and workflows from prototype through launch. We’re looking for someone who combines strong systems thinking, hands-on testing skills, technical curiosi

PythonAWSRestAI
C
📍 United States· Full-time· Remote
✓ Quality checkedCompany trend -100%

As an Engineering Manager on Coder’s Core Workspaces team, you’ll lead engineers building and evolving the systems behind our agentic development experience. You’ll help make agents more capable, reliable, and useful across real development environments. You’ll guide technical direction while growing the team and keeping execution sharp. You’ll work closely with Engineering, Product, and Design across the agent harness, integrations, and developer workflows. What you’ll do here Lead and grow a team within our Workspaces organization. Set technical direction across the agent harness, integrations, and workflows. Stay close to the code and contribute to architecture and implementation decisions. Evolve agent execution, tool use, context management, streaming, and long-running workflows. Extend our provider-agnostic architecture as models and capabilities change. Partner with Product and Design to turn agent capabilities into useful developer experiences. Improve reliability, performance, and operability across agentic systems. Coach engineers, raise the technical bar, and create clarity around priorities and tradeoffs. What we’re looking for Experience managing and growing software engineering teams. Strong hands-on engineering experience with React and TypeScript . Experience with Go . Hands-on experience building systems around LLMs and agentic workflows. Experience with model APIs, tool calling, context management, or agent loops. Strong distributed systems knowledge. Working knowledge of AWS . Strong technical judgment and comfort working through ambiguity. A track record of helping engineers grow while maintaining a high execution bar. Bonus tacos if you have Experience building coding agents, developer tools, or cloud development environments. Experience with MCP , agent tools, or multi-agent systems. Experience with remote execution, sandboxing, or isolated compute. Experience building abstractions across multiple model providers. Deep experience with AWS, Kube

TypeScriptReactAWSDocker
O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -83.9%

About the Team OpenAI’s People team hires, engages, and retains world-class talent to safely build and deploy AGI that benefits all of humanity. The People Analytics team helps leaders make rigorous, evidence-based talent decisions and ensures that the systems supporting those decisions are valid, reliable, fair, and accountable. About the Role As a People Data Scientist focused on AI fairness and bias testing, you will help establish how OpenAI evaluates AI-assisted People systems and high-impact talent processes. You will design and conduct rigorous assessments to identify, measure, and mitigate potential bias across the lifecycle of models, agents, decision-support tools, and automated workflows. Your work will span the entire employee life-cycle, such as hiring, performance, promotion, employee development, workforce planning, etc. You will evaluate both technical systems and the broader human-AI decision processes in which they operate, examining not only model performance but also data quality, measurement validity, differential outcomes, human oversight, and unintended consequences. We’re looking for an experienced data scientist or applied researcher who can translate complex fairness questions into defensible evaluation strategies, scalable testing infrastructure, and clear recommendations for technical teams and senior leaders. This role is preferred to be based in San Francisco, CA. In this role, you will: Define and lead fairness and bias-testing strategies for AI-assisted People processes, models, agents, and decision-support systems from development through deployment and ongoing monitoring. Design rigorous algorithmic audits and validation studies, including adverse-impact analysis, subgroup and intersectional evaluation, error-rate analysis, calibration, measurement invariance, reliability, criterion-related validity, and sensitivity testing. Identify the appropriate fairness criteria for each use case, evaluate tradeoffs among competing definitions

PythonSQLAWSRest
C
📍 United States· Full-time· Remote
✓ Quality checkedCompany trend -100%

As a Senior Software Engineer on Coder’s Agentic Engineering team, you’ll build and evolve the systems behind our agentic development experience. You’ll work across the agent harness, integrations, and workflows that connect agents with real development environments. You’ll stay hands-on, solve complex technical problems, and work closely with Product, Design, and other engineers to ship reliable agentic experiences. To provide substantive overlap with the team, this position must be in Eastern Time. What you’ll do here Design and build production systems in Go, with work across React and TypeScript where needed. Improve agent execution, tool use, context management, streaming, and long-running workflows. Extend our provider-agnostic architecture as models and capabilities change. Build reliable integrations between agents, workspaces, tools, and developer infrastructure. Own projects from implementation through rollout and iteration. Contribute to design reviews, code reviews, and technical discussions. Partner with Product and Design to turn agent capabilities into useful developer experiences. Improve the reliability, performance, and operability of agentic systems. What we’re looking for Strong experience building and operating production software systems. Hands-on experience with Go. Experience with React and TypeScript. Experience building systems around LLMs or agentic workflows. Familiarity with model APIs, tool calling, context management, or agent loops. Good understanding of distributed systems and production reliability. Working knowledge of AWS. Strong problem-solving skills and comfort working through technical ambiguity. Someone who contributes beyond their own code through reviews, collaboration, and knowledge sharing. Bonus tacos if you have Experience building coding agents, developer tools, or cloud development environments. Experience with MCP, agent tools, or multi-agent systems. Experience with remote execution, sandboxing, or isolated compute.

TypeScriptReactAWSDocker
B
📍 Hazelwood, Macau S.a.r., United States
✓ Quality checkedCompany trend +7.9%

Experienced Simulation Software Engineer - Training Systems Company: The Boeing Company The Boeing Company is looking for an Experienced Simulation Software Engineer - Training Systems to join our Government Vehicle Health Management Systems (GVHMS) team in Hazelwood, MO . The Government Training Team develops innovative simulation software solutions that advance pilot readiness and mission success. This Software Engineer role involves designing, architecting, and developing simulation models, virtual environments, and frameworks, while collaborating with stakeholders to optimize overall simulation performance. Responsibilities include simulation validation, integration, and modernization of legacy software within a secure Agile development environment. The role requires strong expertise in C++, with additional knowledge of Python, containerization, and cloud-based technologies. Familiarity with emerging software engineering methods and prior aviation or engineering experience are highly valued for contributing to this fast-paced, mission-focused team. At The Boeing Company, we innovate and collaborate to make the world a better place. From the seabed to outer space, you can contribute to work that matters. We’re committed to fostering an environment for every teammate that’s welcoming, respectful and inclusive, with great opportunity for professional growth. Find your future with us. Position Responsibilities: Designs, architects, and develops simulation models, simulation visualizations, virtual environments/platforms, and frameworks to enhance test performance, safety, and durability of software and hardware throughout the entire product lifecycle Partners with stakeholders to identify simulation r

PythonAWSAzureDocker
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -83.9%

About the Team OpenAI’s People team hires, engages, and retains world-class talent to safely build and deploy AGI that benefits all of humanity. The People Analytics team helps leaders make better, evidence-based talent decisions. About the Role As a People Research Scientist, you will bring deep expertise in research design, measurement, experimentation, and applied data science to OpenAI’s most important People programs. You will design studies, evaluate people processes, and help leaders better empower employees, strengthen organizational systems, and deliver exceptional employee experiences. This is a high-ownership individual contributor role combining hands-on research, methodological leadership, and scalable people science capabilities. We’re looking for an experienced researcher who can turn ambiguous People questions into rigorous designs, validated insights, and actionable recommendations. This role is based in San Francisco, CA or Mountain View, CA, with occasional travel to our San Francisco office. What You’ll Do: Design rigorous research and evaluation strategies for recruiting, organizational health, manager effectiveness, employee experience, and talent outcomes. Apply advanced statistical modeling, machine learning, and research methods to inform program design, evaluate effectiveness, and quantify business impact. Partner with People Operations, data engineering, and people systems teams to define data requirements, improve data quality, establish documentation standards, and ensure research datasets are governed, reproducible, and privacy-preserving. Build scalable people science infrastructure, including self-service agentic tools, automated validation workflows, reusable research datasets and analytical pipelines. Develop research playbooks that establish rigorous standards for study design, measurement, validation, and documentation, enabling high-quality, repeatable, and scalable research across the organization. Communicate findings through c

PythonSQLAWSRest
S
📍 Bellevue, Washington, United States· Full-time· Remote
✓ Quality checkedCompany trend -93.3%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Location: Bellevue, WA Engineering Manager, Cost Intelligence We are looking for an experienced Engineering Manager to lead the Cost Intelligence engineering team. In this role, you will own the technical vision and execution for the features and systems that help Snowflake customers understand, monitor, and optimize their Snowflake consumption and spend. You will lead a talented team of engineers building the data products, APIs, and platform services that power cost visibility, usage analytics, budgeting, and cost optimization insights across Snowflake's platform. You'll work closely with Product Management, Design, Data Science, and cross-functional engineering teams to ship world-class cost intelligence capabilities to Snowflake's customer base. As manager for the Cost Intelligence team, you will: Lead and grow our talented team of software engineers, fostering a culture of technical excellence, ownership, and continuous learning. Drive the roadmap for Cost Intelligence features — including cost allocation, resource budgeting, anomaly detection, and optimization recommendations — in partnership with product management. Set technical strategy for backend systems, data pipelines, and APIs that surface cost and usage insights to customers at massive scale. Own delivery end

VueAWSAzureGCP
S
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -93.3%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. LEAD. STRATEGIZE. TRANSFORM. We are seeking an advanced professional handling complex enterprise AI/ML deployments, deconstructing system dependencies, and ensuring production robustness. WHY THIS ROLE? This role marks a shift from managing tactical tasks to managing strategic outcomes. You are a seasoned professional with a full understanding of your specialization, resolving a wide range of issues in creative ways. WHAT YOU'LL DO: Design robust, scalable AI/ML solutions utilizing the full Snowflake native stack and partner ecosystem. Perform deep-dive Root Cause Analysis (RCA) for complex system dependencies in AI/ML solutions. Collaborate cross-functionally with Sales and Product teams to align technical roadmaps with customer ROI. Mentor Level 3 architects on best practices for MLOps and architectural design. TECHNICAL DEPTH & RISK MANAGEMENT: Distributed Systems: Deconstruct failures in complex pipelines involving external cloud services (AWS/Azure/GCP). Predictive Failure Analysis: Critically think about potential failure modes like model drift and data skew early in the lifecycle. Governance: Architect data security and access controls specifically for sensitive AI/ML training data. SNOWFLAKE-NATIVE TECH STACK: Snowflake Model Registry, Cortex Functions, Python,

PythonAWSAzureGCP
C
📍 United States· Full-time
✓ Quality checkedCompany trend -100%

At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 We’re hiring a highly influential product analytics leader who can turn ambiguous questions into sharp insight, scalable measurement, and recommendations that directly shape what we build. This person will partner closely with Product, Engineering, Design, and Growth to raise the bar on decision quality and establish a more AI-native analytics operating model. Mission Drive the product insight agenda by helping ClickUp make faster, smarter product decisions through rigorous analysis, strong product judgment, and AI-enabled analytics workflows. What You'll Do Own the product analytics agenda across product usage, activation, feature adoption, retention, and expansion, and translate open-ended business questions into structured analyses and clear recommendations Partner with Product, Engineering, and Design to define success metrics early, improve instrumentation quality, and ensure important product surfaces are measurable from launch Build reusable analysis frameworks, semantic layers, metric definitions, and self-serve resources that help product teams answer routine questions faster and more consistently Apply AI-first methods across the analytics workflow, using large language models, coding agents, and automation for tasks like query drafting, QA, validation, documentation, and first-pass synthesis while keeping human judgment at the center of final recommendations Design and interpret experiments, observational analyses, and trend investigations, including situations where data is incomplete or traditional experimentation is not feasible Surface meaningful patterns in behavioral, subscription, and

PythonSQLAWSMachine Learning
S
📍 Bellevue, Wa, United States· Full-time
✓ Quality checkedCompany trend -92%

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. Smartsheet is seeking an experienced sales leader to lead a team of Enterprise Account Executives as a Regional Director, Enterprise GEO. The ideal candidate will have a history of leading a team to over performance in quota attainment and developing customer accounts in the Enterprise space. This role is part of the Sales organization and is based in Bellevue, Washington and Reports to RVP, Enterprise Sales. You Will: Lead a team of Account Executives to exceed quarterly and annual sales quotas Serve as player/coach in the execution of a complex, solution­-based sales process encompassing multiple groups within Enterprise accounts Play leadership role in developing and growing existing business opportunities by coaching account executives to build and execute account strategies Drive Smartsheet senior executive engagement in target accounts Successfully execute across all disciplines of sales management, including Account/Opportunity/Relationship planning and sales methodology execution Partner with Sales Engineering, Consulting, Customer Success and Marketing management to identify and close software and professional services solutions in accounts Proactively i

VueAWSAIGo
🔔

Get new aws and tooling platform lead jobs in United States by email

Daily job updates · Unsubscribe anytime