Jobs in United States

Aws And Tooling Platform Lead in United States

2,086 active opportunities · Updated October 2026

Explore current aws and tooling platform lead jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

O
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 Washington, District of Columbia, United States· Full-time· Remote
✓ High-confidence listingCompany trend -84.7%

From $2M/yr

Quick readStrong listing-quality and freshness signals

About the team The OpenAI for Government team is a dynamic, mission-driven group leveraging frontier AI to transform how governments achieve their missions. Our team works to empower public servants with secure, compliant AI tools (e.g., ChatGPT Enterprise, ChatGPT Gov) and mission-aligned deployments that meet government technical requirements with strong reliability and safety. About the role Our Federal Sales team has a unique mission to help government customers understand the transformative impact that highly capable AI models can bring to their agencies and missions. This role combines technical understanding, strategic vision, partnership management, and value-driven strategy tailored specifically to federal customers. You’ll drive key opportunities through the entire federal sales cycle, from pipeline generation to closure. You’ll collaborate closely with researchers, engineers, and solution strategists to help government customers advance their missions through AI. This role is based in Washington DC. We use a hybrid work model of 3 days in the office per week. We offer relocation assistance. In this role, you'll: Manage a focused set of key federal accounts, developing and executing comprehensive federal account plans. Lead federal customers through their AI adoption journey, from consideration to successful deployment. Partner with solutions and research engineering to build and execute complex government customer programs and projects. Own and manage a federal consumption revenue target. Oversee consumption revenue forecasting and reporting. Analyze key federal account metrics and provide insights to internal and external stakeholders. Closely monitor the federal landscape (agencies, policies, competitors, partners, etc.) to inform product roadmaps and corporate strategies. Collaborate cross-functionally with solutions, marketing, communications, business operations, people operations, finance, product management, and engineering. Support the recruitment

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -84.7%

About the Team OpenAI’s Industrial Compute team is building and productizing infrastructure capabilities that help organizations deploy and operate advanced AI systems at scale. The team works across AI hardware, systems engineering, physical infrastructure, and customer delivery to turn emerging technologies into reliable, repeatable infrastructure solutions. Our work sits at the intersection of technical strategy, product development, engineering, and deployment. We partner closely with customers and internal engineering teams to solve complex infrastructure challenges spanning compute, power, cooling, controls, and facility efficiency. About the Role We are seeking a senior, hands-on Data Center Infrastructure Architect to develop and optimize the physical infrastructure required for large-scale AI deployments. This is a broad technical role spanning data center architecture, electrical and mechanical systems, high-density compute, controls, telemetry, and digital modeling. You will use simulation, operational data, and digital-twin approaches to evaluate infrastructure designs, identify system-level constraints, and improve efficiency, reliability, cost, and speed of deployment. The ideal candidate can move fluidly between first-principles analysis, facility and equipment design, computational modeling, engineering review, and real-world implementation. You should be comfortable working across disciplines rather than operating solely within electrical, mechanical, or software boundaries. Key Responsibilities Define system-level architectures for high-density AI data centers across power, cooling, IT equipment, controls, and facility infrastructure. Develop digital twins and other computational models that represent the behavior of data center systems under changing workloads, environmental conditions, equipment configurations, and failure scenarios. Use design and operational data to identify constraints, improve PUE and related efficiency metrics, and optimize

PythonAWSGitRest
C
📍 United States· Full-time· Remote
✓ High-confidence listingCompany trend -100%

From $194K/yr

Quick readStrong listing-quality and freshness signals

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . We're hiring a Manager, GRC to join our Technology Risk & Controls team within Security and lead second line of defense advisory coverage across critical technology and security domains. This team evaluates risk posture, improves control design, and helps engineering, infrastructure, and security leaders make better, risk-informed decisions. You'll both manage a team and directly own advisory coverage for domains that may include disaster recovery and resiliency, SDLC, artificial intelligence, identity and access management, and supply chain - building trusted partnerships that ensure Coinbase addresses its most critical risks in its most critical functions. What you'll do: Lead second line advisory coverage for key technology and security domains, assessing risk exposure, control effectiveness, policy alignment, and emerging issues across the business. Partner with engineering and technology teams to evaluate domain posture and identify where additional controls, remediation, or governance improvements are needed. Review and challenge the design of new and existing risks and their corresponding controls to ensure our mitigations are scalable, effective, and aligned to business, risk, and regulatory expectations. Drive coordination across risk, controls, policy, audit, and assurance teams to translate incidents, audit findings, and regulatory expectations i

AWSAIGoRust
S
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -93.7%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is a high-growth SaaS observability platform built on the Snowflake AI Data Cloud, enabling businesses to troubleshoot modern distributed applications 10x faster. Now, as a core part of Snowflake, we’ve reached a major milestone in the evolution of the Snowflake platform. By bringing AI-powered observability directly into the Snowflake ecosystem, we’ve created the first truly unified platform for telemetry and business data. As a Senior Solutions Engineer for Observe, you will play a critical, highly-visible, role within our organization, serving as the primary technical resource for our Sales team. You will be responsible for driving the technical closure of sales opportunities by demonstrating the value of Observe to prospective customers. This role requires a blend of deep technical knowledge, strong presentation skills, and a customer-focused approach. KEY RESPONSIBILITIES: Technical Discovery and Presentation: Conduct in-depth technical discovery sessions with prospects to understand their current environment, challenges, and specific observability requirements. Tailor and deliver compelling product demonstrations and technical presentations that showcase how our solution addresses their needs. Proof of Concepts (POCs): Design, scope, and manage te

AWSAzureGCPDocker
O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -84.7%

About the Team The Applied AI Engineering team partners closely with customers to help them move from experimentation to production with OpenAI’s technologies. We act as trusted technical advisors, working across customer strategy, architecture, deployment, and adoption to help organizations realize meaningful impact from frontier AI. The Startups segment serves fast-moving, high-growth companies that are often building new products, workflows, and businesses directly on top of AI. These customers move quickly, operate with high ambiguity, and expect practical, creative, and technically rigorous partnership. About the Role We are looking for an Applied AI Engineering Manager, Startups to lead and scale the Startups Applied AI Engineering motion. This team helps high-growth startups move quickly from experimentation to production, unlock meaningful usage, and build durable technical partnerships with OpenAI. This leader will operate in a high-velocity customer segment where founders, CTOs, and technical teams expect speed, judgment, and hands-on problem-solving. They will balance team leadership, technical depth, customer prioritization, and cross-functional influence across Sales, Product, Engineering, Research, and broader go-to-market teams. In this role, you will define how OpenAI supports startup customers at scale: identifying where deep technical engagement can unlock outsized impact, building repeatable deployment mechanisms, and ensuring the team can serve a broad and dynamic customer base without losing quality or strategic focus. In this role, you will: Craft and continuously refine the strategic vision and operating model for the Startups Applied AI Engineering team, aligning it with OpenAI’s broader company objectives and the evolving needs of high-growth startup customers. Lead, mentor, and grow a team of high-performing technical ICs supporting startup customers across AI-native, developer-led, and product-led companies. Help startups move from early e

AWSRestAIGo
O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -84.7%

About the Team OpenAI didn’t begin as a traditional company. It began as an idea: that artificial intelligence could be developed in a way that benefits everyone. As a creative team, our role is to help make sure the work behind that idea is understood, as it leaves the lab and meets the world. This company works at the frontier of intelligence. Like any frontier, it’s unfinished, constantly shifting, and still being explored. Our job is to stay close to that uncertainty and help shape how this story gets told, in a way that feels grounded and human. We do this through campaigns, launches, films, brand systems, and work that doesn’t fit neatly into any of those categories yet. This team is for people who want to help shape something from the beginning. About the Role We are seeking a Web Producer to own the execution and delivery of web projects across OpenAI.com . You’ll drive work from intake through launch, translating briefs into clear production plans, coordinating cross-functional inputs, managing timelines and dependencies, and ensuring every experience is accurate, polished, and delivered on time. You’ll partner closely with Marketing, Growth, Product, Design, Communications, and Engineering to operationalize web work in a fast-moving environment. This is a hands-on production role: you’ll build pages while owning the details, quality, communication, and follow-through that make launches run smoothly. This role is based in San Francisco. The role requires a hybrid schedule, with employees in the office Monday through Wednesday. In this role, you will: Own web projects from intake through launch, translating requests and creative briefs into structured production plans, managing timelines, dependencies, and handoffs, and proactively surfacing risks. Build and update pages using existing components, templates, and design systems; implement content, layouts, imagery, assets, and metadata accurately across single-page updates and multi-page launches. Partner wit

AWSGitRestAI
C
📍 United States· Full-time· Remote
✓ High-confidence listingCompany trend -100%

From $194K/yr

Quick readStrong listing-quality and freshness signals

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . Coinbase has built the world's leading compliant cryptocurrency platform serving over 73 million accounts in more than 100 countries. With multiple successful products, and our vocal advocacy for blockchain technology, we have played a major part in mainstream awareness and adoption of cryptocurrency. We are proud to offer an entire suite of products that are helping build the crypto economy and increase economic freedom around the world. There are a few things we look for across all hires we make at Coinbase, regardless of role or team. First, we look for signals that a candidate will thrive in a culture like ours, where we default to trust, embrace feedback, disrupt ourselves, and expect sustained high performance because we play as a championship team. Second, we expect all employees to commit to our mission-focused approach to our work. Finally, we seek people with the desire and capacity to build and share expertise in the frontier technologies of crypto and blockchain, in whatever way is most relevant to their role. JOB DUTIES Scale and grow the HR Engineering team by hiring, onboarding and training new analysts and engineers to support Workday and HR functional area Provide functional & technical leadership and mentoring to team members Supervise the day-to-day activities of our Workday instance Design user-friendly processes, guidelines, and documentation

AWSAgileAIGo
O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -84.7%

About the Team: OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. In this role you will: As a Hardware Test Engineer, you will work on Machine Learning/AI hardware system projects to craft the solutions for current and future data center deployments. You will bring a strong understanding of hardware system testing, excellent project management skills, and the ability to collaborate across multiple teams to ensure efficient lab operations. You will be responsible for designing, implementing, and executing comprehensive test plans that ensure the reliability, performance, and scalability of our supercomputing hardware systems. You will develop detailed test plans and methodologies tailored to hardware components, including processors, memory modules, custom accelerators and interconnects. You will collaborate with hardware design, manufacturing, firmware teams and vendors to identify, analyze, and resolve issues affecting hardware, power, thermal and high-speed interconnects. You will perform in-depth debugging on the hardware system Excellent analytical skills to diagnose hardware issues, troubleshoot problems, and propose solutions. Ability to interpret complex test data, identify trends, and draw meaningful conclusions. High-speed links, with a focus on SerDes (Serializer/Deserializer) technology to assess signal integrity, error rates, and overall link performance. You will collaborate with the lab manager to maintain the equipment and hardware systems, including oscilloscopes, thermal test chambers, liquid cooling systems, and other mea

PythonAWSRestMachine Learning
C
📍 United States· Full-time· Remote
✓ High-confidence listingCompany trend -100%
Quick readStrong listing-quality and freshness signals

Ready to do the most impactful work of your career? At Coinbase , we are uncompromising on our mission to increase economic freedom. The bar is high, the environment is intense, and we like it that way. This isn't a place for complacency, it’s a place to be pushed past your perceived limits. If you're ready to build the future of finance alongside people who refuse to settle for "good enough," you belong here. Coinbase is a remote-first, but not remote-only company. Expect to get together quarterly for intense in-person working sessions called “surges.” learn more about working at Coinbase . Coinbase is looking for a strategic, motivated communications leader to serve as the Internal Communications partner for our People Success team. With a new Chief People Officer building the next chapter of our people strategy, this role will shape how that vision reaches every employee through clear, well-timed, and compelling communications. What you’ll do: Serve as dedicated Internal Communications partner to the People Success team, owning communications strategy for People programs and company-wide rollouts with a business-first, product marketing orientation. Advise the CPO and senior People leaders on high-stakes moments, including org changes, policy shifts, leadership transitions, and other high-stakes announcements, crafting messaging that is clear, human, and builds trust Own internal communications for the People Success Team itself, including monthly town halls, leadership communications, and other resources that build team awareness, alignment, and culture. Write and edit range of content, including CPO comms, town hall scripts, company announcements, and FAQs, translating People programming into clear value propositions for Coinbase leaders and employees. Build scalable communications infrastructure, including editorial calendars, playbooks, measurement frameworks, and rhythms that support a growing People organization. Coordinate with broader In

AWSAIGoRust
H
📍 Ny Nyc Metro, United States· Remote
✓ Quality checkedCompany trend +310%

Become a part of our caring community The Lead Software Engineer codes software applications based on business requirements. The Lead Software Engineer works on problems of diverse scope and complexity ranging from moderate to substantial. The Lead Software Engineer standardizes the quality assurance procedure for software. Oversees testing and debugging and develops fixes. Researches complaints and makes necessary adjustments and/or recommendations to resolve complex software related issues. Advises executives to develop functional strategies (often segment specific) on matters of significance. Exercises independent judgment and decision making on complex issues regarding job duties and related tasks, and works under minimal supervision, Uses independent judgment requiring analysis of variable factors and determining the best course of action. Key Responsibilities Technical Architecture and Ownership:** Design and own the end-to-end architecture of Centerwell's AI systems, including LLM-powered clinical tools, RAG pipelines, harnesses, agent-based workflows, and intelligent automation. Make and communicate foundational technical decisions in close collaboration with the broader engineering team. Model Development and Fine-Tuning:** Evaluate, select, and where appropriate guide the fine-tuning of foundation models. Establish model evaluation frameworks that prioritize safety, accuracy, and clinical relevance. Clinical and Product Partnership:** Collaborate closely with product managers, designers, clinicians, and data stakeholders to understand care delivery workflows and translate them into well-scoped, high-impact AI features. HIPAA Compliance and Responsible AI:** Ensure all AI systems are designed, deployed, and monitored in compliance with HIPAA and Humana's Responsible AI standards, including participation i

TypeScriptPythonAWSAzure
B
Relocation support. Relocation assistance is stated. This does not establish visa sponsorship.
📍 Swansea, United States
✓ Quality checkedCompany trend -20.1%

Cloud Operations System Administrator Company: Tapestry - G0G Tapestry Solutions, A Boeing Company, brings over 30 years of industry experience designing, implementing, training, and supporting high-quality, cost-effective information technology and business intelligence solutions. With a dedicated team of approximately 500 professionals, we proudly serve 75 defense, commercial, and government clients across more than 50 U.S. locations and 9 countries worldwide. As a trusted partner, our employees embody our core values by consistently delivering excellence, taking full ownership, and developing innovative solutions that enable critical missions and ensure the safety of our global customers and team members. Joining Tapestry Solutions means enjoying the best of both worlds: access to the vast resources of Boeing combined with the agility and people-focused, family-oriented culture of a small business where your contributions truly matter. Tapestry Solutions, a part of Boeing Global Services (BGS), is seeking multiple Cloud Operations System Administrator in Swansea, IL . In this role you will provide Cloud operational engineering support for Tapestry's Global Decision Support System (GDSS) program. GDSS is the primary US Air Force Air Mobility Command (AMC) Command and Control (C2) system and is a globally distributed enterprise class suite of applications designed, implemented and maintained by Tapestry Solutions. GDSS applications and services provide advanced mission planning, command and control for all phases of mobility operations (airlift, airdrop, and air refueling operations). You will support ongoing efforts to enhance, monitor, and maintain multiple AWS (CloudOne) environments, ensuring high availability. <b

AWSRecruitment
B
📍 Berkeley, Macau S.a.r., United States
✓ Quality checkedCompany trend -20.1%

Cloud Platform Administrator (Mid-Level, Senior or Lead) **Sign on Bonus Potential** Company: The Boeing Company The Boeing Company’s Specialized United States Infrastructure Operations is currently seeking a Cloud Platform Administrator (Mid-Level, Senior or Lead) to join the team in Berkeley, MO; Seattle, WA; or Daytona Beach, FL . The Infrastructure team is seeking a skilled platform engineer to help build and operate the cloud platform services that host critical enterprise applications and software toolchains. In this role, the selected candidate will focus on the shared platform capabilities that enable teams to deploy, run, and maintain containerized and cloud-hosted solutions in a consistent and supportable manner. As both an individual contributor and technical leader, this position will help define and implement platform standards for Kubernetes, container hosting, deployment automation, configuration management, and operational support. This role is focused on platform reliability, repeatability, scalability, and service enablement, rather than custom application software development. Position Responsibilities: Design, implement, and maintain cloud platform services supporting Kubernetes, containers, ingress, storage integration, secrets management, and service connectivity Build and sustain reusable deployment patterns for Commercial-Off-The-Shelf (COTS), Open Source Software (OSS), and internally customized applications Develop and maintain automation for platform provisioning, upgrades, patching, and lifecycle support Manage cluster lifecycle activities including: Cluster upgrades Node management <

AWSAzureDockerKubernetes
A
📍 California, USA - Remote, United States· Remote
✓ Quality checkedCompany trend +40%

Job Requisition ID # 26WD97363 Position Overview We are seeking a Principal Software Engineer – Backend to join Autodesk’s Enterprise Data Management (EDM) organization within the COO-GET Engineering group. This is a senior individual contributor role operating at the Principal (P4) level , expected to drive technology direction for large, complex, and business-critical backend and distributed systems . This role is anchored in backend software engineering excellence : designing, building, and evolving scalable services, APIs, and event-driven systems that operate at enterprise scale. As a Principal Engineer, you will work with high autonomy and ambiguity , shape long-term architecture, and influence multiple teams and domains. Familiarity with data engineering concepts is valuable, but backend systems, service design, and distributed systems are the core competencies. You will function as a technical authority and force multiplier—guiding design decisions, setting standards, and ensuring Autodesk’s core data services are reliable, resilient, and evolvable over time. Responsibilities Provide principal-level technical

PythonAWSAI
O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -84.7%

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role You will work on the systems software strategy and execution that brings new AI silicon from first power-on to a fully integrated system running production-representative models at expected functionality and performance. You will define how software exercises and validates compute, memory, interconnect, and I/O subsystems, then build the diagnostics, automation, and observability needed to find issues quickly. This role sits at the center of silicon, firmware, platform, systems, and workload teams. You will turn hardware specifications and performance targets into an end-to-end bringup plan, drive cross-functional debug, and establish the stress and regression infrastructure that makes each new platform reliable across operating environments. In this role, you will: Contribute to the end-to-end software bringup and validation strategy for new silicon and first-party systems. Define software-driven test coverage across compute, memory, interconnect, I/O, and their system-level interactions. Build diagnostics, test automation, telemetry, and regression infrastructure that accelerate first-silicon learning and issue isolation. Lead bringup from initial silicon arrival through board and system integration, docking, runtime enablement, and model execution. Design stress tests that characterize reliability, performance, and stability across workloads and operating conditions. Translate architecture specifications and performance models into measurable acceptance crit

PythonAWSRestAI
O
📍 San Francisco, California, United States· Full-time· Remote
✓ Quality checkedCompany trend -84.7%

About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role You will build the model runtime within the inference engine that executes complex, frontier models at scale on OpenAI’s custom silicon. The runtime will sit between models running on the hardware and the upper layers of the cluster serving software stack, translating demanding inference workloads into efficient execution while optimizing for throughput, latency, utilization, and reliability. You will work across model architecture, distributed systems, compilers, kernels, and silicon to design a production-grade runtime comparable in ambition to systems such as vLLM and SGLang, but customized and optimized for OpenAI’s AI accelerator. Your work will shape how new model capabilities map onto the platform and how quickly custom silicon can deliver meaningful performance in production. In this role, you will: Design and implement the LLM inference runtime for frontier models running on custom silicon. Build scheduling, continuous batching, memory management, KV-cache management, and execution orchestration for high-performance inference. Develop distributed execution strategies across chips, hosts, and racks, including model partitioning, communication, and synchronization. Optimize end-to-end latency, throughput, memory efficiency, and hardware utilization across diverse model architectures and serving workloads. Partner with kernel, compiler, architecture, and silicon teams to co-design interfaces and remove performance bottlenecks across the stack. Enable new

PythonAWSRestAI
🔔

Get new aws and tooling platform lead jobs in United States by email

Daily job updates · Unsubscribe anytime