Jobs in United States

Ai Engineering Director in New York

648 active opportunities · Updated October 2026

Explore current ai engineering director jobs in New York. Filter by work mode, employment type, experience, department, date posted and distance.

C-
📍 New York, NY, United States· Full-time
✓ High-confidence listing

$275K – $350K/yr

Quick readStrong listing-quality and freshness signals

CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. We are seeking a strategically-minded, technology-focused, and customer-centric Engineering Manager to lead one of our Infrastructure teams here. You will lead a team responsible for building, operating, and scaling the cloud infrastructure and platform systems that underpin CLEAR’s services, ensuring reliability, performance, and security across our environments. A successful candidate brings strong experience in cloud infrastructure, distributed systems, and operational excellence, along with a solid foundation in software engineering. You are an effective communicator who can lead complex infrastructure initiatives from inception through delivery, and thrive in fast-paced environments. This role requires a focus on building resilient, scalable systems, driving automation, and leading and developing high-performing engineering teams. What you'll do: Hire, develop, and grow engineering talent through coaching, mentorship, performance management, and career development planning Set clear goals and expectations, provide regular feedback, and foster accountability across the team Own and execute the roadmap for cloud infrastructure and platform engineering, and reliability initiatives Design, build, and operate a scalable, secure, and highly available cloud platform infrastructure Drive automation across infrastructure provisioning, deployment, and operations to improve efficiency and reduce manual overhead Establish and enforce best practices for system reliability, observability, incident response, and disaster recovery Partner with eng

PythonJavaAWSKubernetes
P
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -72.3%

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Plaid's Infrastructure team builds the platforms and tooling that help engineering teams develop, deploy, and operate production systems safely. Release Engineering owns the path from merge to production, including Plaid's zero-touch deployment system, progressive rollouts, metric-gated analysis, and automatic rollback. Our goal is to make safe shipping the default for every product team. As a Staff Site Reliability Engineer on Release Engineering, you'll define and scale Plaid's reliability practices across product engineering. You'll architect our SLO and error-budget programs, drive the adoption of progressive delivery, and ensure new products are production-ready. By partnering across product and platform teams, you'll translate complex production needs into intuitive, self-service tooling. This is a hands-on technical leadership role where you'll shape the future of our deployment systems—ensuring they remain fast and safe even as AI-assisted development increases code velocity. What excites you Lead the expansion of reliability standards across product engineering, converting foundational infrastructure into lasting operational habits and tooling. Architect and manage the SLO and error-budget

AWSKubernetesAIGo
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $192K/yr

Quick readStrong listing-quality and freshness signals

You will lead a small, hands-on engineering team building the secure, scalable Core Analytics Data Access Platform that accelerates Datadog’s Applied AI and analytics capabilities. The team owns the Data Access Platform — a unified interface that lets AI and analytics teams discover and self-serve production-ready datasets while abstracting underlying systems and embedding required legal and compliance guardrails. In this role you’ll own technical direction, contribute to design and code, and partner closely with Applied AI, Product Analytics, and internal platform teams to provide reliable datasets and APIs for model training and analysis. This role balances day-to-day engineering leadership with long-term platform planning. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead a Hands-On Engineering Team: Manage, mentor, and grow a small team of 2–4 data engineers (mix of senior and junior) across Paris and NYC, fostering technical excellence and career development. Own Technical Direction and Delivery: Define architecture, engineering priorities, and the team roadmap for the Data Access Platform, driving implementation of scalable, secure data pipelines and platform services. Contribute to Design and Code: Spend substantial time coding, reviewing, and shipping critical platform components to ensure performance, reliability, and operational excellence. Partner with Internal Stakeholders: Work closely with Applied AI, Internal Product Analytics, product managers, and platform teams to define data contracts, APIs, SLAs, observability, and curated analytical datasets. Ensure Data Security, Governance, and Reliability: Implement access controls, lineage, monitoring, and compliance guardrails to support safe model training and repeatable analytics workflows.

AWSAIGoRust
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $192K/yr

Quick readStrong listing-quality and freshness signals

As the Engineering Manager for Commercial Audit, you will lead a high-performing team responsible for scaling Datadog’s security and compliance posture through automation, tooling, and engineering excellence. Our GRC (Governance, Risk, and Compliance) function is a critical partner to the broader Security and Engineering organizations, ensuring that Datadog not only meets rigorous global regulatory standards but does so in a way that is efficient, scalable, and integrated into our cloud-native infrastructure. You will manage a team of engineers and analysts who are transitioning to a GRC engineering direction to treat compliance as a software problem, leveraging AI, custom tooling, CI/CD pipelines, and cloud-native services to turn complex regulatory requirements into actionable, automated controls. You will lead the strategy, roadmap, and execution of Datadog’s Commercial Audit initiatives. This is a high-impact leadership role where you will grow a team of engineers and analysts responsible for directly maintaining our compliance programs and related audits (e.g., SOC2, PCI, HIPAA, ISO) while looking to improve efficiency and effectiveness through platforms and tooling. You will act as a bridge between technical engineering, legal, and compliance, enabling the organization to move fast while maintaining a secure and compliant environment. You will champion a culture of "compliance-as-code," identifying opportunities to automate evidence collection, streamline control testing, and reduce manual toil for both your team and our partner engineering teams. At Datadog, we place value in our office culture - the relationships and collaboration it builds, and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead the strategy, roadmap, and execution of Datadog’s commercial security compliance efforts, shifting from manual audit processes to automated, scalable

PythonAWSAzureGCP
O
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the team OpenAI’s Forward Deployed Engineering team partners with customers to turn research breakthroughs into production systems. We operate at the intersection of customer delivery and core platform development. About the role As a foundational FDE manager, you’ll lead FDE through high-stakes, ambiguous customer deployments and own technical and business value outcomes end to end. You’ll grow a team that can operate under pressure and help OpenAI learn from the field. You’ll partner closely with Product, Research, Sales, and GTM to ensure fieldwork informs roadmap priorities, drives new exploration, and supports safe deployment at scale. Your decisions will influence how OpenAI is trusted by the customers closest to our deployment work. Your success will be measured by how consistently your team ships, how clearly you deliver signal to Research and Product, and how durable your team and delivery model prove to be. This role is based in New York City We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. This role also will require travel up to 25%. In this role you will Lead and grow a team of FDE delivering production systems with frontier models Own end-to-end delivery outcomes through clarity, speed, tight coordination, and technical quality Codify what works into tools, playbooks, and roadmap inputs that create leverage for both OpenAI and our wider developer community Notice early indicators and raise them with urgency, whether in product behavior, customer environments, or delivery practices Use judgement to distinguish what requires action and what does not Set a high bar for FDE performance and support each person’s growth through direct, actionable feedback Define how we staff and support field teams that can scale without added complexity You might thrive in this role if you Bring 8+ years of engineering or technical delivery experience, including 2+ years managing high-performing FDE or custo

JavaScriptPythonJavaAWS
DC
📍 New York, New York, United States· Full-time
✓ High-confidence listing

From $131K/yr

Quick readStrong listing-quality and freshness signals

Help shape the technology that enables a global organisation to do its best work. As Senior Manager, Platform Engineering, you’ll lead the team responsible for Diligent’s Atlassian and Microsoft platforms while setting the architectural direction for the wider internal IT estate. You’ll combine people leadership, enterprise platform strategy and hands-on technical judgement to create secure, reliable and scalable experiences for employees worldwide. From modernising service management and automating joiner, mover and leaver processes to enabling AI safely through Microsoft Copilot and Atlassian Rovo, your work will reduce friction, strengthen governance and deliver measurable business impact. Working across IT, Security, HR, Finance, Legal, Compliance and business teams, you’ll turn complex requirements into well-governed platforms that are easy to use, resilient and ready for the future. Here’s a breakdown of what you’ll do (not all of it, just the important stuff) Lead, coach and grow a global team of platform engineers and systems administrators, building a high-performing and inclusive culture. Own the strategy, architecture, governance and roadmap for Atlassian Cloud, including Jira, Jira Service Management, Confluence, Atlassian Guard and Rovo. Set the direction for Diligent’s Microsoft 365 E5 estate, including Teams, SharePoint, Exchange Online, Intune, Defender, Purview, Power Platform and Copilot. Design scalable integration and automation patterns across identity, HRIS, ITSM and business systems using APIs, event-driven automation, Okta Workflows, Power Platform and scripting. Partner with IT Support to improve self-service, automate repetitive work and reduce ticket volume, escalation effort and time to resolution. Establish strong standards for security, access governance, AI adoption, reliability, compliance and business continuity across the internal technology estate. These are the essentials you’ll need to get an interview Significant experience in i

PythonAWSGitAI
C
📍 New York, United States
✓ Quality checkedCompany trend +340.2%

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Senior Manager, Platform Engineering / DevOps Who Are You You are an experienced Senior Manager / emerging Staff-level leader in DevOps and Platform Engineering with strong technical depth and demonstrated leadership in delivering enterprise-scale cloud platforms. You bring a balanced mix of hands-on engineering expertise, team leadership, and execution rigor. You excel in driving outcomes in complex, multi-stakeholder environments, guiding teams to deliver secure, scalable, and high-quality platform solutions. You are comfortable leading engineers, managing stakeholders, and owning delivery across multiple workstreams. You demonstrate: A strong ownership mindset with accountability for delivery and outcomes Ability to translate business needs into actionable engineering roadmaps Solid expertise in cloud-native platforms, DevOps practices, and SRE principles Capability to lead teams and influence without requiring extensive tenure Role Responsibilities Development & Enforcement Own and execute the H100 platform engineering roadmap, aligned to enterprise priorities and program milestones Drive delivery of GCP-based platform capabilities (GKE, networking, IAM, CI/CD, observability) Establish and enforce engineering standards, best practices, and ADR compliance</li

SQLMongoDBGCPKubernetes
P
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -100%
Quick readStrong listing-quality and freshness signals

Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity Postman is hiring a Senior Manager, Customer Success Engineering to help lead Customer Success Engineering in North America. This leader will manage CSEs, ensuring the team is focused on the customers where Postman can create outsized impact, while providing the coaching, inspection, and operational discipline required for consistently high-quality execution. This is a highly cross-functional leadership role and a critical partner to Sales. The leader will work closely with North America Sales leadership, with a strong focus on East Coast alignment, to prioritize accounts, shape technical engagement strategy, allocate CSE capacity, and ensure customer execution is progressing with urgency and discipline. CSEs help customers embed Postman into real engineering workflows across API design, governance, CI/CD, developer platforms, integrations, modernization, migration, automation, quality enforcement, and service discoverability. This is not about general support or enablement; CSEs are laser-focused on turning Postman from a beloved developer tool into critical API infrastructure. The right leader is an experienced po

ReactCI/CDAIGo
C-
📍 New York, New York, United States· Full-time
✓ High-confidence listing

From $145K/yr

Quick readStrong listing-quality and freshness signals

CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. We are seeking a Senior Manager of Industrial Engineering to join CLEAR's Central Operations team. This role is the quantitative backbone of our Experience & Operations Design function — owning the data-driven design standards, labor models, and capacity planning that determine how CLEAR's physical operation is structured and staffed. You will be the analytical engine behind lane design decisions, engineered labor standards, and demand forecasts, ensuring the operation is designed to deliver a consistently excellent member experience efficiently at scale. What you'll do: Develop and own engineered labor standards and staffing models across CLEAR’s physical experiences, including Verifier, Greeter, eGate, Concierge, TSA PreCheck, Enrollment, Sports, and new programs, serving as the operations source of truth for labor and headcount inputs Define physical operating requirements and sizing standards across lane configurations, verification and enrollment experiences, hardware placement, and queue management; lead checkpoint design for new and existing airport locations Lead demand forecasting and capacity planning across CLEAR experiences, modeling throughput, staffing requirements, operating hours, and operational impact for new programs, launches, surge periods, and events Drive operational efficiency through time studies, process analysis, and labor optimization, identifying and developing opportunities to improve productivity, capacity, and the Member experience across the network Partner cross-functionally with Operations,

PythonSQLGitRest
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $192K/yr

Quick readStrong listing-quality and freshness signals

Manager I, Engineering - Change Experience Platform The Change Experience Platform team builds the internal experiences and platform capabilities that help Datadogs understand, author, route, and safely manage infrastructure changes. The team owns internal UI and CLI frameworks, change-management user experiences, notification and subscription platforms, and infrastructure governance signals used across Datadog’s engineering organization. Its work sits at the intersection of developer experience, infrastructure operations, product design, and change safety. We’re looking for a hands-on technical leader to manage and grow a team of engineers working on the systems that shape how Datadog engineers interact with infrastructure change. You will partner closely with infrastructure, developer experience, platform engineering, and product teams to build reusable interfaces, workflows, and safety mechanisms that make complex change processes easier to understand and safer to execute. This is a high-impact role for someone who enjoys combining product thinking with strong engineering judgment. You will help the team balance framework ownership, platform reliability, internal customer needs, and long-term technical direction across a portfolio that includes UI systems, CLI authoring and publishing, change-management workflows, notification routing, subscriptions, and infrastructure cordon management. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Lead and grow a small team of engineers responsible for internal platforms and product experiences used across Datadog engineering. Help define what “good” looks like for internal developer-facing platforms, including usability, reliability, documentation, adoption, and supportability. Se

PythonJavaSQLPostgreSQL
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $192K/yr

Quick readStrong listing-quality and freshness signals

Datadog's Application Performance Monitoring (APM) provides deep visibility into the health, performance, and lifecycle of modern distributed applications, tracing requests from end-user devices (web and mobile) through to backend services. Our goal is to help customers detect root causes faster, optimize application performance, and improve resource efficiency at scale. As the Engineering Manager for APM Serverless, you will help define and deliver the end-to-end serverless APM experience, from auto-instrumentation through troubleshooting, and ensure that OpenTelemetry and Datadog-native customers alike have a frictionless and performant journey. You will also lead efforts to expand coverage of cloud-managed services across providers, ensuring customers can seamlessly trace and monitor critical services in all major and emerging cloud environments. We’re looking for an experienced engineering leader who thrives at the intersection of infrastructure and developer experience. You should care about well-designed APIs, observability-first thinking, and building systems that empower other developers. This is a high-leverage role that will influence how developers across the industry understand and instrument their serverless workloads. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Lead a polyglot team of 8-9 engineers and partner closely with Product and Engineering teams across Datadog to deliver industry-leading serverless capabilities that power consistent, scalable, and intuitive instrumentation across languages. Drive a domain that is technically rich: Lambda, Azure Functions, GCP, OTel billing, Rust, durable functions, distributed tracing across managed services. Engineers on this team work

AWSAzureGCPAI
D
📍 New York, California, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $330K/yr

Quick readStrong listing-quality and freshness signals

Datadog is expanding the Technical Solutions (TS) organization by seeking a customer-focused, deeply technical Distinguished Architect to join our Product Solutions Architecture (PSA) team. In this role, you will act as a technical multiplier for the world's leading AI labs and AI-native companies. You will bridge the gap between their bleeding-edge infrastructure aspirations and Datadog’s technology roadmap, ensuring our platform natively solves the unique observability challenges of training and deploying foundational models at scale. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Leadership : Demonstrate thought leadership in the AI/LLM space. Influence key decision makers and stakeholders by connecting technical capabilities to organizational and business impact. Advisory : Strategically partner with highly technical Founders, Heads of Infrastructure, and Research Lead peers. Guide them on best practices and emerging industry trends in the AI/LLM space. Lead high-level technical and architectural conversations around AI adoption. Presentations : Lead deep-dive architecture reviews and design engagements with customer teams and their leaders to share industry trends, best practices, and demonstrate how Datadog can support high-throughput hyper scale AI workloads. GTM : Identify emerging AI-native technology shifts and feed them directly back to Datadog Product Management. Co-create custom observability integrations and solutions alongside Product SAs to keep Datadog at the absolute forefront of the AI stack. Collaboration : Collaborate with Product Solutions Architecture (PSA), Sales, Sales Engineering and Marketing in providing high-quality technical resources to a broad audience of practitioners and economic buyers. Hiring : Assis

LA
📍 New York, United States· Full-time
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

We are investing in agentic AI and need a Senior AI Engineer to lead the design and delivery of these systems. This is a foundational hire: you will own both the agent-facing workstreams — pipelines, orchestration, conversational interfaces — and the underlying context layer that makes them reliable, including memory management, knowledge graph integration, and retrieval infrastructure. You will work closely with data engineers, project leads, and client stakeholders, and play a key role in shaping how Lynx builds and ships AI solutions at scale. What This Involves: Lead the architecture and delivery of agentic AI systems end-to-end: agents, orchestration, tool use, and multi-step reasoning workflows. Own the context layer: design and implement memory architectures (episodic, semantic, working memory) and integrate GraphRAG and knowledge graph retrieval into agentic pipelines. Build robust RAG systems — including vector retrieval, graph traversal, and hybrid search — and ensure retrieval quality through evaluation frameworks. Translate client requirements into technical designs, presenting approaches and trade-offs to both technical and non-technical stakeholders. Define standards and reusable patterns for agentic AI development that other engineers at Lynx can build on. Set up observability, evaluation, and monitoring pipelines to ensure AI systems perform correctly in production. Requirements: 5–8 years of software or ML engineering experience, with at least 2–3 years building LLM-based or agentic AI systems in production. Deep hands-on experience with agentic frameworks (LangChain, LlamaIndex, AutoGen, CrewAI, or similar) and LLM APIs (OpenAI, Anthropic, etc.). Strong understanding of agent design patterns: ReAct, planning loops, tool use, multi-agent coordination, and memory architectures. Practical experience with GraphRAG or knowledge graph-based retrieval (e.g., Neo4j, Microsoft GraphRAG) and vector databases (Pinecone, Weaviate, Qdrant, etc.). Proficiency in

PythonReactDockerKubernetes
H
📍 New York, NY, United States
✓ Quality checkedCompany trend +310%

Become a part of our caring community You have shipped AI products before. You understand the difference between a demo and a production system. You have strong opinions about evaluation frameworks because you have experienced the consequences of operating without them. You are at your best when you own architecture decisions while continuing to build and deliver critical code yourself. We build the platform that transforms millions of clinical documents into trusted, actionable data. Our systems use large language models (LLMs) to read medical records, extract structured facts, answer complex questions with citations back to source documents, and route complex cases to human experts. The output of these systems supports healthcare decisions that impact real members. As a Lead AI Applied Engineer, you will provide technical leadership for AI-enabled products and platforms, define architectural direction, establish engineering standards, and personally design and build the most critical components of our systems. You will lead through both technical expertise and execution, helping the team deliver reliable, scalable, and auditable AI solutions in a highly regulated healthcare environment. Why Join Us Lead the architecture of production AI systems where LLMs are foundational to the product experience. Make key technical decisions regarding model selection, system boundaries, platform architecture, and build-versus-buy strategies. Own the highest-risk and highest-impact technical challenges involving reliability, explainability, and correctness. Influence engineering culture and establish standards that shape how the team builds and ships AI products. Work on systems operating at meaningful scale, processing millions of documents and supporting healthcare decisions across a large member population. Partner

JavaScriptTypeScriptPythonReact
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $320K/yr

Quick readStrong listing-quality and freshness signals

As a Research Scientist on our team, you will partner with Research Engineers, working on fundamental research problems and collaborating with Datadog's product and engineering teams to translate research advances into products. Building on our track record of AI-powered solutions (e.g., Bits AI , Bits Evolve , and our time series foundation model ), Datadog AI Research tackles high-risk, high-reward problems grounded in real-world challenges in cloud observability and security. We are focused on two research areas: World Models for Observability -- Training multimodal foundation models that learn the joint dynamics of distributed systems across metrics, traces, logs, topology, and events. These models power advanced forecasting, anomaly detection, root cause analysis, counterfactual simulation ("what if?"), and provide a learned planning backbone for our autonomous agents. Trained Agents for Observability -- Post-training models to operate autonomously across Datadog's domain. SRE incident response is our first target, with a clear path to code repair, security response, and infrastructure optimization. We build the simulation environments, RL training loops, and evaluation infrastructure needed to train agents that match or surpass frontier models at a fraction of the cost. What You'll Do: Conduct research in generative AI and machine learning, building specialized foundation models and trained agents for observability Train multimodal models on large-scale, diverse telemetry data (metrics, logs, traces, topology, events) using distributed training infrastructure Design and build simulated environments and RL training loops for on-policy agent training and evaluation Collaborate with cross-functional teams (Product, Engineering) to integrate capabilities like multimodal world modeling and autonomous agents into Datadog's products Stay at the forefront of foundation models, world models, and RL-based agent research Contribute to r

GitMachine LearningAIGo
🔔

Get new ai engineering director jobs in New York, United States by email

Daily job updates · Unsubscribe anytime