Jobs in United States

Principal Data Platform Architect in United States

334 active opportunities · Updated October 2026

Explore current principal data platform architect jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -8%

NVIDIA has transformed computer graphics, PC gaming, and accelerated computing for more than 25 years through exceptional technology and the people who build it. In semiconductor manufacturing, our role is to enable the ecosystem, not compete within it. We partner with fabs, equipment manufacturers, and software providers to make inspection, metrology, and manufacturing intelligence dramatically faster on the NVIDIA platform. Our team builds the software that makes this possible: models, adaptation and evaluation workflows, and deployable inference capabilities that partners integrate into their own tools. We work in environments where labeled data is limited and proprietary, distributions shift across tools and fabs, production budgets are tight, and software must operate inside air-gapped facilities. We’re seeking a Principal Systems Software Engineer for Semiconductor Inspection in Santa Clara. This is a hands-on architect role: you will define the approach, build it, evaluate it, and demonstrate the results. You will work across computer vision, time-series modeling, multimodal AI, anomaly detection, model adaptation, evaluation, and production inference. Success means technology that a fab or equipment vendor can integrate, operate, and trust—not only a successful internal demonstration. What you’ll be doing: Define and prototype AI system architectures spanning optical and e-beam inspection, wafer and mask inspection, metrology, defect review, equipment signals, and process data. Advance world foundation model capabilities for semiconductor manufacturing, including vision, time-series and multimodal representation learning, model adaptation, domain transfer, and data-scarce defect understanding. Develop workflows for defect detection, classification, localization, segmentation, nuisance filtering, ADC, AD

PythonMachine LearningAI
O
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

About the Team OpenAI's research training infrastructure powers how our frontier models are trained and evaluated. The Simulation team sits at the intersection between the agentic harness that powers OpenAI's products and the research infrastructure where GPT-next is trained, ensuring that our model's training environment is as realistic as possible. This team owns the integration layer that connects our production harness capabilities into the training stack. The work is highly cross-functional and high leverage: researchers depend on it to run experiments and evaluations reliably as well as to develop the next generation of harness capabilities. Failures in this surface can materially affect training velocity and correctness. About the Role We're looking for a Principal Software Engineer to lead the architecture and evolution of the Simulation Platform. You'll own a critical interface between research and engineering, building the systems, APIs, and operational patterns that let researchers use agentic coding infrastructure safely and effectively in training environments. This role is ideal for a senior backend or infrastructure engineer with strong technical judgment, product sense for highly technical users, and the ability to drive execution across multiple teams. The highest-leverage work is building robust infrastructure that supports and accelerates research without compromising engineering quality. In this role, you will Design, build, and evolve the integration between the Codex harness that powers OpenAI's products and research training infrastructure used for training GPT-next Build a platform for our LLMs to train and be evaluated in simulated environments that mimic their deployment setting as closely as possible, on every axis: agentic harness, compute substrate, timing, tools, data sources, humans in the loop, and more Own major integration surfaces end-to-end, from architecture and API design through rollout, operations, and long-term maintenance Bu

PythonAWSRestAI
C
📍 United States· Remote
✓ Quality checkedCompany trend +340.2%

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary CVS Health is seeking a Principal Software Engineer to lead the design and delivery of enterprise-scale Generative AI solutions that power next-generation healthcare experiences. This role goes beyond hands-on coding—you will define technical strategy, establish architectural standards, and guide multiple teams in building secure, scalable, and cost-effective AI platforms across AWS (Bedrock) and Google Cloud (Vertex AI API). You will partner with product, security, compliance, and enterprise architecture teams to ensure solutions meet business objectives, regulatory requirements, and performance goals. The ideal candidate combines deep technical expertise with leadership skills—capable of influencing cross-org architecture decisions, mentoring engineering teams, and driving responsible AI practices in production. Key Responsibilities Lead end-to-end platform delivery of highly scalable, secure AI services and applications leveraging AWS Bedrock (Foundation Models, Knowledge Bases, Agents, Guardrails) and Google Cloud Vertex AI (Gemini via Vertex AI API, Agent Builder, Vector Search, Search & Grounding) Architect and implement Retrieval-Augmented Generation (RAG) solutions, integrating proprietary data from sources like Amazon S3 and Google Cloud Storage/BigQuery, and using Bedrock Knowledge Bases and/or Vertex AI Search & Groundi

AWSAzureDockerKubernetes
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $295.3K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Software Engineer of the Game Engine data model team, you will innovate on the core data structures that form the backbone of Roblox’s platform. In this role you will have first hand opportunity to design and build one of the core technologies that makes the Roblox game engine special. The Data Model framework provides an interface between engine capabilities and our creators. It also acts as a fabric to core technologies, such as networking, scripting, Studio, rendering and physics which are essential to the high performance of the Roblox game engine. This role is in our San Mateo, CA HQ three days a week (Tuesdays to Thursdays). You Will: Develop engine code that performs well for all user-created games on the Roblox platform. Establish the foundational architecture and technical direction for the team. Work cross-functionally , across teams and technology platforms to deliver high quality and amazing functionality. Lead by example and mentor engineers to implement technological best practices, patterns, and strategies. Improve the product quality by encouraging automation testing. Take Ownership of projects throughout their full li

AWSGitAIC++
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $345K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Software Engineer on the Compute team, you will be the technical anchor for Roblox's GPU and AI accelerator capabilities. This is a battle-tested GPU expert role focused on the machine management layer and above: how GPU hosts are made production-ready, kept healthy, and turned into reliable compute for the workloads that depend on them. You will own the hard problems that show up only at scale, from driver and firmware management to GPU health, reliability, and performance across a rapidly growing fleet of accelerators spanning Roblox data centers and cloud environments. You will set the technical direction for GPU compute and up-level the entire organization's GPU expertise. You will: Serve as the GPU technical leader for the Compute team, partnering across Kubernetes, Machine Bootstrap, Networking, and Cloud to drive GPU strategy end to end. Own the GPU host lifecycle above raw fleet management: driver, firmware, and CUDA stack management, GPU health and telemetry, and remediation of GPU-specific failures (XID errors, ECC, thermal, NVLink and fabric faults). Architect how GPU capacity is exposed to compute platforms, including scheduling, isolation, and integration with Ku

AWSKubernetesGitAI
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $295.3K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. With Roblox Ads & Discovery business growing at a rapid rate, we are building large scale ads machine learning infrastructure to deliver more value to our users and our advertisers. As a Machine Learning Infrastructure Engineer, you’ll build scalable, reliable, and high-performance infrastructure that powers ML systems across our organization. You’ll operate at the scales of hundreds of billions of engagements, and redefine how we deliver performance ads to hundreds of millions of users. You will: You will co-design models and systems, working at the intersection of model architecture and ML infrastructure, partnering closely with core modelers, data and AI infrastructure engineers, and product teams to push the boundaries of large-scale training and serving. Your work will span recommendation, search, and agentic applications, including large transformer architectures, LLMs, generative rankers, and efficient offline and online content-understanding systems. You will investigate model, data, and systems tradeoffs end to end—from data pipelines and distributed training to low-latency inference and production serving. This includes designing efficient KV-cache strategies, applying p

AWSGitMachine LearningAI
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $295.3K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Discovery UX sits at the intersection of product innovation and high-scale infrastructure. Every Roblox user starts their journey on the homepage, and every Roblox developer needs to find their correct audience. Our team works closely with product and engineering teams across Roblox to bridge this gap by improving the user journey from landing on Roblox to joining an experience. We primarily work on Home, Search, Charts, and Experience Details pages, for both the Roblox website and app. As a Principal Software Engineer , you will craft robust, extensible systems for high-traffic surfaces. While you are a frontend expert at heart, you are a versatile engineer who isn't afraid to dive into the backend or data pipelines to ship a complete feature. You will drive solutions using React and modern technologies, helping us build efficiently across devices. If you have a deep understanding of frontend architecture and a genuine curiosity for diving into data to understand "the why" behind an experiment, you will be right at home with us. This role reports to the Engineering Manager on the Discovery UX team. You Will: Design, build, and ship features that define the discovery journey for millions of

JavaScriptTypeScriptJavaReact
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $293.8K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Roblox’s database team develops the next-generation, multi-tenant database platform that elastically scales and underpins every online data workload at Roblox. As a principal engineer on the database team, you will shape the architecture, build and launch critical database capabilities that keep our services fast, reliable and efficient at global scale. You will report to the Technical Director for Storage. You will: Design and implement new engine features —indexing, storage formats, WAL and replication protocols, sharding, and query-planner enhancements—that push latency, throughput, and availability boundaries. Evolve the control plane to deliver elastic scaling, autonomous healing, and zero-downtime schema or tenant moves across global regions. Profile and optimize critical code paths using kernel-level tracing and advanced performance tooling; drive systematic tail-latency reductions. Establish engineering best practices by leading design reviews, performance benchmarks, failure drills, and post-incident retrospectives. Automate everything : develop frameworks for testing, CI/CD, rollout safety, observability, and autoscaling so that the platform operates hands-off at scale. Ment

SQLPostgreSQLMySQLAWS
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $345K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Software Engineer leading Fleet Management, you will be the overall technical lead across three pods and the person who sets the technical direction for the fleet management layer of Roblox. This is a hands-on, deeply technical leadership role that owns all of Roblox's compute capacity end to end: from low-level provisioning and the data plane, up through the control planes that operate it, and all the way to the UI and internal-facing products that let teams self-serve capacity. Your org centralizes security, maintenance operations, and the uptime of every Roblox Kubernetes cluster, and governs the internal customer contracts that drive automation across the fleet spanning Roblox data centers and cloud providers. You will guide architecture, raise the engineering bar, and make sure compute capacity supply and demand stay in balance as the fleet grows. You will: Serve as the overall technical lead for three Fleet Management pods, setting and aligning the technical direction across low-level provisioning, the data plane, and the control plane and product surfaces above them. Architect the declarative, Kubernetes-style control planes that operate Roblox's compute fleet across o

SQLAWSKubernetesGit
A
📍 United States
✓ High-confidence listingCompany trend +365.2%
Quick readStrong listing-quality and freshness signals

Abbott is a global healthcare leader that helps people live more fully at all stages of life. Our portfolio of life-changing technologies spans the spectrum of healthcare, with leading businesses and products in diagnostics, medical devices, nutritionals and branded generic medicines. Our 122,000 colleagues serve people in more than 160 countries. JOB DESCRIPTION: The Staff Data Engineer is a senior, hands-on technical contributor within an Enterprise Data domain, owning technical design and the most complex implementation work for assigned data products, capabilities, and integrations. The role establishes, communicates, and evolves the technical approach for the work it leads and is accountable for the quality, durability, and supportability of the solutions it shapes. Staff Data Engineers work directly with business stakeholders to understand needs and shape technical solutions, engaging at a level appropriate to the work they lead. The role is expected to be fluent in both technical execution and business context, translating between them without losing precision in either. The Staff Data Engineer sets technical direction for assigned capabilities and initiatives within the domain, applies enterprise standards and platform patterns, engages Principal Engineers where cross-domain considerations apply, and multiplies the effectiveness of the domain team through design leadership, code review, and mentoring. This role is based in Madison, WI . Essential Duties Include, but are not limited to, the following: Technical design and solutioning Own technical design and hands-on delivery for the most complex or highest-risk work across assigned data products, capabilities, and initiatives, including data models, pipeline architecture, integration patterns, and platform usage decisions. Produce design docume

PythonSQLAWSAzure
G
📍 Austin, Texas, United States· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking an experienced Principal Hardware Diagnostics Engineer to design and develop diagnostics software used to monitor hardware health and diagnose system-level issues across Graphcore’s AI infrastructure platforms. This role focuses on building diagnostics agents, tools, and analytics frameworks that enable engineers and automation systems to identify, isolate, and resolve hardware issues across blade-level servers and rack-scale clusters. The Team Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. The Systems Engineering and Platform Validation team ensures Graphcore’s AI compute platforms are reliable, diagnosable, and operationally robust at scale. The team co

PythonLinuxAIC++
S
📍 Menlo Park, California, United States· Full-time
✓ Quality checkedCompany trend -91.7%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. How do you make the world's most powerful coding agent for data a delight to use? We are the fastest growing software company at this scale in history and looking to push that to new heights by building the best coding agent in the industry. The Cortex Code (CoCo) team is building the future of coding agents for working with data. See our flagship product in action: CoCo Desktop in Action . As a Frontend AI Principal Engineer, you are the cornerstone of delivering the ultimate user experience. You will be entrusted with the highest level of polish, creativity, and interaction design, pushing the boundaries of what's possible in the most innovative areas of our product. This role offers the unique opportunity to have broad creative license, empowering you to lead and create groundbreaking systems and experiences that marry aesthetics with functionality, making data more accessible and actionable for our customers. AS THE FRONTEND AI PRINCIPAL ENGINEER YOU WILL: Design and build exceptional experiences, ensuring the highest level of polish, creativity, and interaction to solve complex customer problems. Define the architectural vision for the CoCo coding agent platform and ensure consistency of design abstractions across the entire product surface Develop innovative platform

JavaScriptTypeScriptJavaReact
H
📍 Texas, United States of America, United States
✓ High-confidence listingCompany trend +103.7%

$147.1K – $230.9K/yr

Quick readStrong listing-quality and freshness signals

Product Manager Enterprise Integrations Description - Role Summary We are looking for a Principal Product Manager to lead the evolution of HP’s Enterprise Integration strategy—from fragmented, point-to-point integrations to a scalable, productized API ecosystem. This role will define and drive a shift toward API-as-a-product, enabling reusable business capabilities, accelerating developer productivity, and laying the foundation for an AI-first, composable enterprise. You will play a pivotal role in transforming our integration landscape into a unified Integration Fabric, and ultimately an Intelligence Fabric that powers next-generation digital and AI-driven experiences. What You’ll Own API Productization Strategy Transform APIs from point solutions into reusable, discoverable product offerings Define capability-based packaging aligned to business domains (B2C, B2B, Partner ecosystems) Drive standardization across design, classification, and maturity frameworks Enterprise Integration Fabric Define integration patterns that unify platform, SaaS, and custom-built APIs Establish a holistic, cross-application view of enterprise APIs Reduce point-to-point integrations through reusable architecture patterns and services Data + API Convergence Partner with Data teams to define and expose data products via APIs Enable consistent access to enterprise data through standardized interfaces Align API strategy with enterprise data governance and cataloging AI-First APIs & Developer Experience Define APIs optimized for AI agents, automation, and developer productivity Build APIs that are discoverable, composable, and machine-consumable Enable next-gen use cases including agent or

GraphqlAISap
G
📍 United States· Full-time
✓ High-confidence listingCompany trend -100%

From $154K/yr

Quick readStrong listing-quality and freshness signals

Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team GoDaddy's Global Storage Engineering team operates one of the largest Ceph environments in the world, delivering the object, block, and file storage platforms that power GoDaddy's hosting infrastructure, internal services, OpenStack environments, and next-generation AI/HPC workloads. If you're passionate about distributed systems, storage architecture, and solving failure scenarios at massive scale, this is an opportunity to work on infrastructure few engineers will experience in their careers. Ceph is a strategic platform at GoDaddy — not an ancillary service. Our global footprint includes 80+ production clusters, 20,000+ OSDs, 1,830 storage nodes, 300 PB of raw capacity, and 69 billion objects spanning five datacenters across three continents. The platform supports RBD, RGW (S3/Swift), and CephFS workloads through more than 1,550 pools, 574,000 placement groups, and 900+ MDS daemons, creating engineering challenges that demand deep expertise in storage architecture, data durability, performance optimization, automation, and observability. As a Lead Senior Site Reliability Engineer, you'll serve as one of the principal technical leaders for GoDaddy's Ceph platform. You'll design the next generation of storage clusters, lead major platform upgrades, drive capacity and hardware strategy, and establish the standards that govern how the platform scales. You'll be the engineer the team turns to for the most complex s

PythonKubernetesAISwift
H
📍 Texas, United States of America, United States
✓ High-confidence listingCompany trend +103.7%
Quick readStrong listing-quality and freshness signals

Principal Data Privacy Architect Description - Job Summary - Role Purpose • Lead and oversee complex, cross-functional privacy and data protection programs from strategy through implementation, ensuring alignment across business, technical, legal, and compliance stakeholders. • This role will design and implement scalable, AI-ready data privacy architecture across enterprise data environments, applications, and AI-enabled workflows. • The Principal Data Privacy Architect will serve as a hands-on subject matter expert responsible for embedding privacy-by-design, consent enforcement, data sovereignty, data loss prevention, and compliance controls into large, complex global data environments. • The architect will partner closely with Data Engineering, Cybersecurity, Legal, Privacy, AI Governance, Product, and Enterprise Architecture teams to ensure customer, employee, partner, and sensitive enterprise data is accessed, processed, shared, retained, and protected in a compliant, secure, and trustworthy manner. - Why This Role Matters • Architect for Trust & Scale: Build reusable privacy architecture patterns that enable secure, compliant, and scalable data usage across platforms, products, and regions. • Enable Responsible AI: Design privacy guardrails for AI agents, generative AI, RAG pipelines, model inputs and outputs, embeddings, vector stores, and automated data workflows. • Reduce Risk While Enabling Innovation: Translate privacy, consent, regulatory, and data sovereignty obligations into practical engineering controls that accelerate business outcomes. Responsibilities - Think Customer First • Embed customer trust, transparency, and privacy-by-design principles into enterprise data platforms and customer-facing applications. • Design consent-aware data access and usage p

PythonJavaSQLAWS
🔔

Get new principal data platform architect jobs in United States by email

Daily job updates · Unsubscribe anytime