Jobs in United States

Reliability Engineer Iii in New York

50 active opportunities · Updated October 2026

Explore current reliability engineer iii jobs in New York. Filter by work mode, employment type, experience, department, date posted and distance.

D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $192K/yr

Quick readStrong listing-quality and freshness signals

As a Senior Platform Product Manager focused on AI SDLC Trusted Throughput, you will define and drive the product strategy for enabling safe, reliable software delivery at AI-native scale across Datadog’s Internal Developer Platform. As AI accelerates development velocity and system complexity, you will help evolve SDLC systems from human-supervised workflows to platforms with built-in safety, observability, and correctness guarantees. You will partner closely with engineering, security, and developer platform teams to improve deployment reliability, operational visibility, and governance while enabling both engineers and AI agents to move quickly with confidence. This role offers the opportunity to shape foundational developer infrastructure and influence how AI-powered software delivery operates across Datadog. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Own the product strategy, roadmap, and execution for AI-native SDLC throughput and reliability initiatives across Datadog’s Internal Developer Platform Define and drive platform outcomes aligned to DORA metrics, balancing deployment velocity with reliability, change failure reduction, and operational safety Partner with engineering, infrastructure, security, and developer experience teams to build automated validation, auditability, and risk-scoring capabilities into deployment workflows Deliver actionable SDLC observability and diagnostic capabilities that connect executive-level metrics to operational signals across the software delivery lifecycle Drive systems that monitor and validate AI-generated or AI-attributed changes to ensure correctness, compliance, and trustworthy automation Serve as a cross-functional product leader across SDLC Foundations, Security Engineering, and compl

AIGoRustSpring
O
📍 New York, New York, United States· Full-time· Remote
✓ High-confidence listingCompany trend -80.2%
Quick readStrong listing-quality and freshness signals

About the team OpenAI’s Forward Deployed Engineering team partners with customers to turn research breakthroughs into production systems. We operate at the intersection of customer delivery and core platform development. About the Role You will define how OpenAI delivers complex systems to customers. You will own how they are built, shipped, and adopted. You’ll translate business outcomes into a technical plan, run day-to-day execution across FDEs, Researchers, and Customer Engineers, and partner with customer teams to ensure delivery supports their goals. You will focus on the Financial Services vertical, partnering with banks, asset managers, and private capital investors to deploy next-generation AI capabilities across their operations, investment processes, and portfolio companies. You will own delivery end-to-end: embedding with Financial Services customers to map workflows and success criteria, ensuring components ship on time, and leading readiness and change management for adoption. You’ll track progress, manage dependencies, make sequencing decisions, and drive 0→1 prototypes through MVP and scale. You will also share field insights with Product and Research to guide roadmap and priorities. Success will be measured first and foremost by impact - deployments that deliver measurable value against customer goals, drive adoption, and become critical to their workflows. Additional measures of success include delivery reliability (milestones hit, low reopen/churn), operating leverage (patterns reused across deployments), judgment under pressure, and product impact (field signal that shifts roadmaps/architectures). This is a high-trust, high-autonomy role. Success requires deep technical project management expertise, extreme ownership of outcomes, and an ability to immerse in customer workflows and partner with customer teams to solve complex engineering problems at pace. This role is based in New York City. We use a hybrid work model of 3 days in the office per w

Artificial IntelligenceAIExcelProject Management
V
📍 New York, New York, United States· Full-time
✓ High-confidence listing

From $170K/yr

Quick readStrong listing-quality and freshness signals

Join VTS as an Engineering Manager. As a technical leader, you'll drive high-impact, customer-facing projects across major features in web and mobile. You’ll play a pivotal role in advancing VTS's evolution into an AI-powered platform, further cementing our industry-leading position. In this role, you will be responsible for fostering a healthy and collaborative culture, guiding the team's execution and delivery, and mentoring engineers to grow their careers and capabilities. ** Please note that this opportunity is located in New York, NY, and requires this hire to work from our office 4 days a week. ** To thrive in this role, you have: Proven experience leading software teams using agile development methodologies. A strong technical background with a proven track record of building, deploying, and maintaining scalable web applications, services, and third-party integrations. Experience working on high-impact, business-critical domains, where reliability, performance, and operational stability are essential. Demonstrated ability to recruit and develop talent, strengthening the team's capabilities through effective coaching and mentorship. The ability to engage in strategic technical discussions, adeptly manage technical tradeoffs and risks, and ensure alignment with product and business goals, especially in environments involving multiple systems and external partners. A sense of empathy for the customer and a relentless focus on shipping high-quality products. Excellent communication skills that empower you to effectively convey ideas, coach, mentor, and educate others. A keen interest in building AI-backed features and leveraging AI as an engineering productivity tool. What you'll do: Lead and Develop a High-Performing Team : Foster a healthy, collaborative, and diverse culture that reflects our company and engineering values. Coach and mentor individual team members, creating a structured environment and feedback loop that supports career development and hi

TypeScriptReactAWSRest
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $192K/yr

Quick readStrong listing-quality and freshness signals

Manager I, Engineering - Change Experience Platform The Change Experience Platform team builds the internal experiences and platform capabilities that help Datadogs understand, author, route, and safely manage infrastructure changes. The team owns internal UI and CLI frameworks, change-management user experiences, notification and subscription platforms, and infrastructure governance signals used across Datadog’s engineering organization. Its work sits at the intersection of developer experience, infrastructure operations, product design, and change safety. We’re looking for a hands-on technical leader to manage and grow a team of engineers working on the systems that shape how Datadog engineers interact with infrastructure change. You will partner closely with infrastructure, developer experience, platform engineering, and product teams to build reusable interfaces, workflows, and safety mechanisms that make complex change processes easier to understand and safer to execute. This is a high-impact role for someone who enjoys combining product thinking with strong engineering judgment. You will help the team balance framework ownership, platform reliability, internal customer needs, and long-term technical direction across a portfolio that includes UI systems, CLI authoring and publishing, change-management workflows, notification routing, subscriptions, and infrastructure cordon management. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: Lead and grow a small team of engineers responsible for internal platforms and product experiences used across Datadog engineering. Help define what “good” looks like for internal developer-facing platforms, including usability, reliability, documentation, adoption, and supportability. Se

PythonJavaSQLPostgreSQL
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $192K/yr

Quick readStrong listing-quality and freshness signals

Applied AI is where Datadog's ambitious AI bets get built and shipped ( Bits Chat , updog ). We sit at the intersection of research and product: turning promising capabilities from Datadog AI Research lab and the research community into production systems that reach real customers. The team builds the foundations for agentic systems capable of operating at scale in complex production environments. Current bets span agents that run autonomously at scale, context and memory layers that make those agents more intelligent over time, and tools that help customers build and validate AI-native services in production. The mandate is to move fast from idea to customer impact, and when a product finds its footing, to set it up for growth. As an Engineering Manager I in Applied AI, you will lead a team of engineers and applied scientists working on one of these challenges. You will define technical direction, run short feedback loops, make deliberate decisions about what to pursue or stop, and work closely with product managers, research teams, and cross-functional partners to ship AI capabilities that matter. At Datadog, we place value in our office culture, the relationships and collaboration it builds and the creativity it brings. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do Lead and develop a team of engineers and applied scientists focused on building the foundations for agents operating at scale Work closely with product managers, research teams, and cross-functional partners to shape the team's bets from initial framing through to broader adoption, with a clear definition of success criteria at each stage Own end-to-end delivery of high-quality AI systems, from early research exploration to production-grade reliability, with high standards for operational excellence, system reliability, and technical quality Navigate the unique challenges of shipping AI-powered products: balancing quali

Machine LearningAIGoRust
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $280K/yr

Quick readStrong listing-quality and freshness signals

The Detection Platform organization is responsible for helping customers identify, understand, and act on issues across their environments through alerting, event intelligence, and autonomous detection capabilities. As Director, Detection Platform, you will lead a group of engineering managers and teams responsible for foundational alerting infrastructure, event management, monitor creation experiences, and AI-powered detection systems. This role sits at the center of Datadog’s efforts to evolve how customers detect, investigate, and respond to operational issues at massive scale. You will partner closely with Product Management, Applied Science, Design, and Engineering leaders to shape the future of detection and observability experiences for Datadog customers while leading a growing organization of engineers. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Lead a multi-team engineering organization responsible for alerting, event management, monitor creation experiences, and autonomous detection capabilities. Define and execute the technical and organizational strategy for the Detection Platform while aligning stakeholders across Engineering, Product, Design, and Applied Science. Drive innovation in AI-powered detection, anomaly identification, and signal generation that helps customers proactively identify and resolve issues. Scale highly available platform systems that process hundreds of millions of evaluations while maintaining reliability, performance, and operational excellence. Develop and mentor engineering managers and technical leaders, fostering a culture of execution, collaboration, and technical rigor. Champion customer-centric product thinking by balancing platform investments with intuitive user experiences and measurable customer

Machine LearningAIGoRust
O
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -80.2%

About the team OpenAI’s Forward Deployed Engineering (FDE) team turns research breakthroughs into production-grade systems. We embed deeply with customers to solve high-leverage problems and act as the delivery engine for our most complex large-scale engagements. We move quickly from prototype to production and surface reusable patterns that shape our platform. We operate at the intersection of deployment and development – working closely with OpenAI Research, Product and Partnerships. About the Role As a Technical Deployment Lead (TDL), you will define how OpenAI delivers complex systems to customers. You will own how they are built, shipped, and adopted. You’ll translate business outcomes into a technical plan, run day-to-day execution across FDEs, Researchers, and Customer Engineers, and partner with customer teams to ensure delivery supports their goals. You will own delivery end-to-end: embedding with customers to map workflows and success criteria, ensuring components ship on time, and leading readiness and change management for adoption. You’ll track progress, manage dependencies, make sequencing decisions, and drive 0→1 prototypes through MVP and scale. You will also share field insights with Product and Research to guide roadmap and priorities. Success will be measured first and foremost by impact - deployments that deliver measurable value against customer goals, drive adoption, and become critical to their workflows. Additional measures of success include delivery reliability (milestones hit, low reopen/churn), operating leverage (patterns reused across deployments), judgment under pressure, and product impact (field signal that shifts roadmaps/architectures). This is a high-trust, high-autonomy role. Success requires deep technical project management expertise, extreme ownership of outcomes, and an ability to immerse in customer workflows and partner with customer teams to solve complex engineering problems at pace. This role is based in NYC. We use a hy

AWSRestAIGo
DC
📍 New York, New York, United States· Full-time
✓ High-confidence listing

From $131K/yr

Quick readStrong listing-quality and freshness signals

Help shape the technology that enables a global organisation to do its best work. As Senior Manager, Platform Engineering, you’ll lead the team responsible for Diligent’s Atlassian and Microsoft platforms while setting the architectural direction for the wider internal IT estate. You’ll combine people leadership, enterprise platform strategy and hands-on technical judgement to create secure, reliable and scalable experiences for employees worldwide. From modernising service management and automating joiner, mover and leaver processes to enabling AI safely through Microsoft Copilot and Atlassian Rovo, your work will reduce friction, strengthen governance and deliver measurable business impact. Working across IT, Security, HR, Finance, Legal, Compliance and business teams, you’ll turn complex requirements into well-governed platforms that are easy to use, resilient and ready for the future. Here’s a breakdown of what you’ll do (not all of it, just the important stuff) Lead, coach and grow a global team of platform engineers and systems administrators, building a high-performing and inclusive culture. Own the strategy, architecture, governance and roadmap for Atlassian Cloud, including Jira, Jira Service Management, Confluence, Atlassian Guard and Rovo. Set the direction for Diligent’s Microsoft 365 E5 estate, including Teams, SharePoint, Exchange Online, Intune, Defender, Purview, Power Platform and Copilot. Design scalable integration and automation patterns across identity, HRIS, ITSM and business systems using APIs, event-driven automation, Okta Workflows, Power Platform and scripting. Partner with IT Support to improve self-service, automate repetitive work and reduce ticket volume, escalation effort and time to resolution. Establish strong standards for security, access governance, AI adoption, reliability, compliance and business continuity across the internal technology estate. These are the essentials you’ll need to get an interview Significant experience in i

PythonAWSGitAI
C-
📍 New York, New York, United States· Full-time
✓ High-confidence listing

$275K – $350K/yr

Quick readStrong listing-quality and freshness signals

CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. The Engineering Manager, CLEAR1 Strategic Partnerships Platform (StratP) is responsible for leading the engineering team that builds partner integrations, healthcare-related workflows, identity and authentication capabilities, and reusable platform improvements that strengthen CLEAR1’s customer offerings. In this role, you’ll drive execution, technical direction, and team development across a high-leverage platform area, turning strategic partner and market needs into scalable, reliable capabilities that improve delivery quality and create long-term business value. What you’ll do: Lead, coach, and develop engineers through hiring, feedback, performance management, and career growth Drive roadmap execution by setting priorities, running strong planning cadences, and delivering against CLEAR1 business commitments Partner across Product, Design, Operations, and Engineering to translate partner and customer needs into scalable platform capabilities Lead delivery across work spanning partner integrations, healthcare-related workflows, identity and authentication capabilities, and reusable platform improvements for CLEAR1 Guide architecture, technical tradeoffs, and dependency management to improve reliability, speed, and long-term platform leverage How you’ll measure success: Improved roadmap predictability and delivery against planned commitments Clearer prioritization and stronger execution across competing demands and shared dependencies Better ownership clarity across partner-facing integrations, identity flows, and customer-facing platf

GitRestAIGo
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $192K/yr

Quick readStrong listing-quality and freshness signals

Coordination Systems provides foundational distributed systems building blocks for internal Datadog platforms. Our services cover sharding, consensus, resource protection, configuration distribution, and much more. We are looking for a manager to lead the Coordination Systems - Storage team. This team provides essential configuration storage and distribution systems that are depended upon by almost every service and pod at Datadog. We power critical runtime configuration (e.g. feature flags), complex control planes (e.g. dynamic sharding configuration), and much more. Storage is one of four subteams within Coordination Systems. If successful, the candidate will have opportunities to lead other growing and impactful areas such as Resource Protection. At Datadog, we place value in our office culture - the relationships that it builds, the creativity it brings to the table, and the collaboration of being together. We operate as a hybrid workplace to ensure our employees can create a work-life harmony that best fits them. What You’ll Do: (Describe role responsibilities here/max 6 bullets) Lead a core team of 5 engineers (distributed, with majority in NYC) Lead ceremonies, prioritize and delegate project Stay hands-on with the code, e.g. isolated features, small remediations, investigation follow ups Stay actively involved in operations, incidents, root cause analysis, etc. Constantly promote a culture of operational excellence, organizing gamedays, conducting operational reviews, staying proactive with reliability Who You Are: (Describe role qualifications here/max 6 bullets) Strong distributed systems skills, able to understand and account for a variety of failure modes, well-versed in end-to-end o11y, validation testing, simulation setup, etc. Worked on platform teams before, providing critical infrastructure to internal stakeholders Experienced in handling significant incidents, both as a responder and follow-up ow

AIRustExcelSEM
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $192K/yr

Quick readStrong listing-quality and freshness signals

You will lead a small, hands-on engineering team building the secure, scalable Core Analytics Data Access Platform that accelerates Datadog’s Applied AI and analytics capabilities. The team owns the Data Access Platform — a unified interface that lets AI and analytics teams discover and self-serve production-ready datasets while abstracting underlying systems and embedding required legal and compliance guardrails. In this role you’ll own technical direction, contribute to design and code, and partner closely with Applied AI, Product Analytics, and internal platform teams to provide reliable datasets and APIs for model training and analysis. This role balances day-to-day engineering leadership with long-term platform planning. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead a Hands-On Engineering Team: Manage, mentor, and grow a small team of 2–4 data engineers (mix of senior and junior) across Paris and NYC, fostering technical excellence and career development. Own Technical Direction and Delivery: Define architecture, engineering priorities, and the team roadmap for the Data Access Platform, driving implementation of scalable, secure data pipelines and platform services. Contribute to Design and Code: Spend substantial time coding, reviewing, and shipping critical platform components to ensure performance, reliability, and operational excellence. Partner with Internal Stakeholders: Work closely with Applied AI, Internal Product Analytics, product managers, and platform teams to define data contracts, APIs, SLAs, observability, and curated analytical datasets. Ensure Data Security, Governance, and Reliability: Implement access controls, lineage, monitoring, and compliance guardrails to support safe model training and repeatable analytics workflows.

AWSAIGoRust
C-
📍 New York, NY, United States· Full-time
✓ High-confidence listing

$275K – $350K/yr

Quick readStrong listing-quality and freshness signals

CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. We are seeking a strategically-minded, technology-focused, and customer-centric Engineering Manager to lead one of our Infrastructure teams here. You will lead a team responsible for building, operating, and scaling the cloud infrastructure and platform systems that underpin CLEAR’s services, ensuring reliability, performance, and security across our environments. A successful candidate brings strong experience in cloud infrastructure, distributed systems, and operational excellence, along with a solid foundation in software engineering. You are an effective communicator who can lead complex infrastructure initiatives from inception through delivery, and thrive in fast-paced environments. This role requires a focus on building resilient, scalable systems, driving automation, and leading and developing high-performing engineering teams. What you'll do: Hire, develop, and grow engineering talent through coaching, mentorship, performance management, and career development planning Set clear goals and expectations, provide regular feedback, and foster accountability across the team Own and execute the roadmap for cloud infrastructure and platform engineering, and reliability initiatives Design, build, and operate a scalable, secure, and highly available cloud platform infrastructure Drive automation across infrastructure provisioning, deployment, and operations to improve efficiency and reduce manual overhead Establish and enforce best practices for system reliability, observability, incident response, and disaster recovery Partner with eng

PythonJavaAWSKubernetes
M
📍 New York, United States of America, United States
✓ High-confidence listingCompany trend +1850%
Quick readStrong listing-quality and freshness signals

We anticipate the application window for this opening will close on - 30 Sep 2026 Careers that change lives start here. Medtronic is a global leader in healthcare technology with a Mission to alleviate pain, restore health, and extend life. Our 95,000 employees work across more than 150 countries to put patients first — developing innovative medical technologies that improve the lives of 72+ million patients each year. Your unique talents will help shape the future of healthcare while building a career grounded in purpose, growth, and impact. A Day in the Life The Cardiac Rhythm Management (CRM) portfolio delivers therapies designed to restore, support, and protect every heartbeat through advanced pacing, defibrillation, and remote monitoring solutions. CRM's unique ability to pair proven life-saving therapies with intelligent technology gives clinicians and patients greater confidence. Clinical impact and innovation come together to improve lives through meaningful patient outcomes. Provide technical, educational, operational and sales support to assist the district in meeting Cardiac Rhythm Management (CRM) sales and customer service objectives. CRM seeks collaborative candidates who will meet our customer expectations by striving without reserve for the greatest possible reliability and quality in our products, processes, and systems by being accountable, having a voice and acting promptly. This role supports the NYC territory and requires candidates to reside within the assigned area and travel regularly to customer accounts. Clinical Specialists must be available to support urgent customer needs and participate in weekend, holiday, and occasional after-hours coverage as bus

SalesforceLogisticsRecruitmentHR
C-
📍 New York, NY, United States· Full-time
✓ High-confidence listing

$435K – $535K/yr

Quick readStrong listing-quality and freshness signals

CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. We are seeking a collaborative, strategic, and execution-focused engineering leader to shape the digital experiences that millions of people have with CLEAR. As a Director of Engineering, you will lead multiple engineering teams responsible for building and scaling our consumer products across the CLEAR mobile app, website, digital enrollment experiences, marketing technology, and Concierge platform. You'll partner closely with Product, Design, Marketing, and business leaders to create intuitive, high-performing experiences that drive acquisition, engagement, conversion, and member satisfaction. The ideal candidate combines strong technical leadership with a deep understanding of consumer product development, building high-performing teams that move quickly while maintaining quality, reliability, and operational excellence. What You'll Do: Lead and grow multiple engineering teams responsible for CLEAR's consumer experiences across mobile, web, marketing technology, and Concierge products. Define and execute the technical strategy for customer-facing applications, ensuring scalable, reliable, and performant experiences that delight millions of members. Partner closely with Product, Design, Marketing, and Business stakeholders to prioritize roadmaps that improve acquisition, conversion, engagement, retention, and member satisfaction. Drive engineering excellence by establishing best practices around architecture, delivery, quality, observability, and operational performance. Build highly collaborative relationships across Engineering, Pro

GitRestAIGo
S
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is a high-growth SaaS observability platform built on the Snowflake AI Data Cloud, enabling businesses to troubleshoot modern distributed applications 10x faster. Now, as a core part of Snowflake, we’ve reached a major milestone in the evolution of the Snowflake platform. By bringing AI-powered observability directly into the Snowflake ecosystem, we’ve created the first truly unified platform for telemetry and business data. We’re looking for a Technical Account Manager to partner with our most strategic enterprise customers and ensure they derive sustained operational value from Observe. This is a hands-on, post-sales technical role focused on long-term platform adoption, optimization, and technical partnership. You will work directly with SRE, DevOps, platform, and engineering teams to embed Observe into daily workflows, evolve telemetry strategy over time, and continuously improve reliability, performance, and cost efficiency. This role is ideal for an experienced observability practitioner who enjoys being deeply embedded with customer teams, solving real production challenges, and acting as a trusted technical advisor in complex enterprise environments. What You’ll Do Serve as the primary technical owner and trusted advisor for assigned strategic a

AWSAzureGCPKubernetes
🔔

Get new reliability engineer iii jobs in New York, United States by email

Daily job updates · Unsubscribe anytime