Jobiba hiring network

Ai Engineering Director Jobs

10,000 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current ai engineering director jobs. Use filters to narrow by work mode, employment type, experience and date posted.

D
Datadog
📍 New York• Full-time• From $192K/yr
1mo ago

You will lead a small, hands-on engineering team building the secure, scalable Core Analytics Data Access Platform that accelerates Datadog’s Applied AI and analytics capabilities. The team owns the Data Access Platform — a unified interface that lets AI and analytics teams discover and self-serve production-ready datasets while abstracting underlying systems and embedding required legal and compliance guardrails. In this role you’ll own technical direction, contribute to design and code, and partner closely with Applied AI, Product Analytics, and internal platform teams to provide reliable datasets and APIs for model training and analysis. This role balances day-to-day engineering leadership with long-term platform planning. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead a Hands-On Engineering Team: Manage, mentor, and grow a small team of 2–4 data engineers (mix of senior and junior) across Paris and NYC, fostering technical excellence and career development. Own Technical Direction and Delivery: Define architecture, engineering priorities, and the team roadmap for the Data Access Platform, driving implementation of scalable, secure data pipelines and platform services. Contribute to Design and Code: Spend substantial time coding, reviewing, and shipping critical platform components to ensure performance, reliability, and operational excellence. Partner with Internal Stakeholders: Work closely with Applied AI, Internal Product Analytics, product managers, and platform teams to define data contracts, APIs, SLAs, observability, and curated analytical datasets. Ensure Data Security, Governance, and Reliability: Implement access controls, lineage, monitoring, and compliance guardrails to support safe model training and repeatable analytics workflows.

As a Regional Manager, Sales Engineering, you will lead a team of Sales Engineers and frontline leaders, driving technical execution, operational excellence, and team development across your region. You’ll act as a force multiplier by owning team performance, shaping execution standards, and partnering closely with Sales leadership to influence strategy, improve outcomes, and scale impact beyond individual deals. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead, coach, and develop a team of Sales Engineers, driving performance, engagement, and long-term career growth through regular feedback and structured development Partner confidently with senior Sales leadership on regional strategy, multi-quarter planning, annual capacity planning, and forecast alignment operating 2-3 quarters ahead to anticipate resourcing needs and business priorities Own in-field execution at a regional level by taking an executive presence role in strategic and high-impact opportunities and providing senior-level stakeholder alignment and technical credibility alongside your team Oversee consistent, high-quality delivery across technical sales cycles in the region, including evaluations, customer engagements, and deal progression Hire and onboard top talent across ICs and manager-level roles, building a high-performing team and accelerating ramp through structured onboarding and enablement Manage regional operations including resource allocation, utilization, KPI tracking, productivity, and skills development to ensure effective coverage across the territory Collaborate cross-functionally with Sales, Product, Marketing, Enablement, and Support at a senior level to align on regional strategy, GTM execution, and customer experience Drive operational excellence across the

awsazuregcp
View job →

Regional Manager, Sales Engineeringとして、Sales Engineersとフロントラインリーダーで構成されるチームを率い、担当地域における技術的な実行力、業務の卓越性、チームメンバーの成長を推進していただきます。 チームパフォーマンスのオーナーシップを持ち、実行基準を形作り、Sales部門のリーダーシップと密接に連携しながら戦略に影響を与え、成果を改善し、個々の案件を超えたインパクトをスケールさせる、いわば「フォースマルチプライヤー(force multiplier)」としての役割を担っていただきます。 Datadogでは、オフィスカルチャーによって築かれる人間関係やコラボレーション、そこからもたらされる創造性を大切にしています。私たちは、従業員それぞれに合ったワーク・ライフ・ハーモニーを実現できるよう、ハイブリッド・ワークプレイスとして運営しています。 業務内容: Sales Engineersのチームを率い、指導・育成し、定期的なフィードバックと構造化された育成プログラムを通じてパフォーマンス、エンゲージメント、長期的なキャリア成長を推進する シニアSalesリーダーシップと自信を持って連携し、地域戦略、複数四半期にわたるプランニング、年間キャパシティプランニング、フォーキャストの整合を担い、2〜3四半期先を見据えてリソースニーズと事業優先度を予測する 戦略的かつインパクトの大きい商談においてエグゼクティブとしての存在感を発揮し、チームとともにシニアレベルのステークホルダーとの連携と技術的信頼性を提供することで、地域レベルの現場での実行をオーナーシップする 評価、顧客エンゲージメント、案件の進行を含む、地域内の技術的セールスサイクル全体を通じて一貫した高品質なデリバリーを監督する IC(個人貢献者)からマネージャーレベルまで、トップタレントの採用・オンボーディングを行い、高パフォーマンスなチームを構築し、構造化されたオンボーディングとイネーブルメントによって立ち上がりを加速させる リソース配分、稼働率、KPIトラッキング、生産性、スキル開発を含む地域内の業務を管理し、担当エリア全体で効果的なカバレッジを確保する Sales、Product、Marketing、Enablement、Supportとシニアレベルで部門横断的に連携し、地域戦略、GTM実行、顧客体験の整合を図る プロセスの改善、データ品質の維持、評価フレームワークのスケール、チームの有効性と事業インパクトを高める施策への貢献を通じて、地域全体の業務の卓越性を推進する 募集要項: 顧客対応型の技術組織においてSales Engineersおよび/またはフロントラインリーダーをマネジメントした経験3年以上。チームパフォーマンスの管理、メンバー育成、高パフォーマンスチームのスケールにおける実績を持つ方 Sales Engineering、Solutions Engineering、または技術系の顧客対応職における経験7年以上 シニアSalesリーダーシップと戦略的パートナーとして、地域戦略、長期プランニング、フォーキャストの整合において連携した実績 複数四半期にわたるプランニングと日々のチーム実行・優先事項のバランスを取りながら、戦略レベルと業務レベルの両方で活動できる能力 部門を横断し、シニアレベルにも影響を与えられる強力なコラボレーション力 パブリッククラウド(AWS、Azure、GCP)、Kubernetes、APM/トレーシング、セキュリティ、開発者体験、オブザーバビリティに関する背景・知見 AIツールを活用し、それをスケールさせる実証済みの能力 変化を受け入れ、好奇心を持つ成長志向のマインドセット 顧客対応をサポートするために必要に応じて出張できること Datadogは、さまざまな立場の人々を大切にしています。誰もが初日から上記の資格をすべて満たすわけではないことは理解しています。それでも構いません。テクノロジーに情熱を持ち、スキルを伸ばしたいと思っている方は、ぜひご応募ください。 福利厚生とキャリア成長: 新入社員への株式付与(RSU / 譲渡制限付株式ユニット)および従業員株式購入制度(ESPP) 継続的な専門能力開発、製品トレーニング、キャリアパスの提供 社内ネットワーク構築のためのメンターおよびバディ・プログラム インクルーシブな企業文化、コミュニティ・ギルドへの参加機会 幅広く手厚い医療関連の福利厚生制度 企業型確定拠出年金(401K) ペットの養子縁組と保険プログラム 上記の内容は、就業する国やDatadogでの雇用内容によって異なる場合があります。 #LI-Hybrid About D

awsazuregcp
View job →
D
Datadog
📍 New York• Full-time• From $192K/yr
1mo ago

As the Engineering Manager for Commercial Audit, you will lead a high-performing team responsible for scaling Datadog’s security and compliance posture through automation, tooling, and engineering excellence. Our GRC (Governance, Risk, and Compliance) function is a critical partner to the broader Security and Engineering organizations, ensuring that Datadog not only meets rigorous global regulatory standards but does so in a way that is efficient, scalable, and integrated into our cloud-native infrastructure. You will manage a team of engineers and analysts who are transitioning to a GRC engineering direction to treat compliance as a software problem, leveraging AI, custom tooling, CI/CD pipelines, and cloud-native services to turn complex regulatory requirements into actionable, automated controls. You will lead the strategy, roadmap, and execution of Datadog’s Commercial Audit initiatives. This is a high-impact leadership role where you will grow a team of engineers and analysts responsible for directly maintaining our compliance programs and related audits (e.g., SOC2, PCI, HIPAA, ISO) while looking to improve efficiency and effectiveness through platforms and tooling. You will act as a bridge between technical engineering, legal, and compliance, enabling the organization to move fast while maintaining a secure and compliant environment. You will champion a culture of "compliance-as-code," identifying opportunities to automate evidence collection, streamline control testing, and reduce manual toil for both your team and our partner engineering teams. At Datadog, we place value in our office culture - the relationships and collaboration it builds, and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead the strategy, roadmap, and execution of Datadog’s commercial security compliance efforts, shifting from manual audit processes to automated, scalable

pythonawsazure
View job →

As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB’s cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and EMEA teams. Success in this role means smoother launches, clearer roadmaps, stronger reliability metrics and an SRE organization that's better-equipped to deliver predictability at scale. This role can be based out of our Dublin or Cork office or remotely in Ireland. What You'll Do Drive Program Planning & Execution – Define program scope, milestones, and success criteria with SRE engineers and leaders. Manage dependencies across platform teams, keep work clearly tracked in Jira, and deliver on time Strengthen Production Reliability – Lead change management and launch readiness programs. Partner with SREs and product teams to define and operationalize SLOs/SLIs, and use incident data, metrics, and capacity signals to drive prioritization and continuous improvement Lead Cross-Functional Coordination – Align SRE with Security, Compliance, Cloud platform, and other engineering teams. Coordinate cross-team incident response, ensure clear follow-through, and build trust as the go-to driver of complex, multi-team efforts Build Scalable Systems & Processes – Design lightweight frameworks and communication patterns that help SRE deliver reliably at scale. Work yourself out of the "hero" role by leaving teams better-equipped to execute independently Requirements 8+ years in technical program management, engineering management, or a comparable technical role partnering with software engineering teams Proven track record leading large-scale, cross-team platform initiatives through ambiguity and change Strong knowledge of production change management, software development lifecycle, and reliability metrics (SLOs, SLIs) Skilled at shaping roadmaps and managing dependencies Able to query and interpret

mongodbawsazure
View job →

As a TPM for SRE, you will partner with SRE leaders and engineers to scale the platform that underpins all of MongoDB’s cloud products. You will drive program execution, strengthen production reliability practices, and coordinate cross-functional efforts across US and EMEA teams. Success in this role means smoother launches, clearer roadmaps, stronger reliability metrics and an SRE organization that's better-equipped to deliver predictability at scale. This role can be based remotely on the East Coast What You'll Do Drive Program Planning & Execution – Define program scope, milestones, and success criteria with SRE engineers and leaders. Manage dependencies across platform teams, keep work clearly tracked in Jira, and deliver on time Strengthen Production Reliability – Lead change management and launch readiness programs. Partner with SREs and product teams to define and operationalize SLOs/SLIs, and use incident data, metrics, and capacity signals to drive prioritization and continuous improvement Lead Cross-Functional Coordination – Align SRE with Security, Compliance, Cloud platform, and other engineering teams. Coordinate cross-team incident response, ensure clear follow-through, and build trust as the go-to driver of complex, multi-team efforts Build Scalable Systems & Processes – Design lightweight frameworks and communication patterns that help SRE deliver reliably at scale. Work yourself out of the "hero" role by leaving teams better-equipped to execute independently Requirements 8+ years in technical program management, engineering management, or a comparable technical role partnering with software engineering teams Proven track record leading large-scale, cross-team platform initiatives through ambiguity and change Strong knowledge of production change management, software development lifecycle, and reliability metrics (SLOs, SLIs) Skilled at shaping roadmaps and managing dependencies Able to query and interpret metrics, logs, or other data s

mongodbawsazure
View job →
M
Mongodb
📍 United States• Full-time• From $127K/yr
1mo ago

The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, deployment machinery, and observability and alerting systems. The Fabric team manages the infrastructure that enables secure communication between systems and from the public internet. Their responsibilities encompass network architecture, service mesh, and edge load balancing, ensuring customer data remains safe in transit. The team plays a crucial role in developing and maintaining the reliable and globally connected multi-cloud network that supports MongoDB products. This role can sit in our NYC HQ, our smaller Austin, Palo Alto, or San Francisco offices, or fully remote from anywhere in North America. When based in an office, we provide hybrid work accommodation. Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong networking background to join the Fabric team. This role is pivotal in building and maintaining the robust infrastructure necessary for secure and efficient communication between our services. As an SRE on the Fabric team, you will leverage your expertise in networking, distributed systems, and automation to ensure our systems are resilient, scalable, and reliable. The ideal candidate should Have 10+ years of experience working on software and operating distributed systems, with deep expertise in networking fundamentals and a good understanding of how the internet works, e.g. TCP/IP (including IPv6), DNS, TLS/mTLS, BGP, tunnels, overlays, and SDN principles Possess a customer-focused mindset, driving improvements that benefit end-users Value efficiency in processes and operations, and display a strong preference for automation over manual processes (“allergic to ops work”) Be intimately familiar with modern cloud-based infrastructure and the network design prim

mongodbawsazure
View job →
M
1mo ago

The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, deployment machinery, and observability and alerting systems. The Fabric team manages the infrastructure that enables secure communication between systems and from the public internet. Their responsibilities encompass network architecture, service mesh, and edge load balancing, ensuring customer data remains safe in transit. The team plays a crucial role in developing and maintaining the reliable and globally connected multi-cloud network that supports MongoDB products. This role can sit in our Toronto or Vancouver offices, or fully remote from anywhere in North America. When based in an office, we provide hybrid work accommodation. Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong networking background to join the Fabric team. This role is pivotal in building and maintaining the robust infrastructure necessary for secure and efficient communication between our services. As an SRE on the Fabric team, you will leverage your expertise in networking, distributed systems, and automation to ensure our systems are resilient, scalable, and reliable. The ideal candidate should Have 10+ years of experience working on software and operating distributed systems, with deep expertise in networking fundamentals and a good understanding of how the internet works, e.g. TCP/IP (including IPv6), DNS, TLS/mTLS, BGP, tunnels, overlays, and SDN principles Possess a customer-focused mindset, driving improvements that benefit end-users Value efficiency in processes and operations, and display a strong preference for automation over manual processes (“allergic to ops work”) Be intimately familiar with modern cloud-based infrastructure and the network design primitives of at least one of AWS, Azur

mongodbawsazure
View job →

Come join and lead the Server Ingress Security team, where we are rearchitecting MongoDB Server’s ingress networking to make MongoDB clusters even more secure. This new team is building the Atlas Network Protection layer, a set of performant, security-critical services that harden MongoDB's pre-authentication attack surface and provides the ability to respond rapidly to emergent threats. We are looking for a talented Lead Engineer to join the team and be founding members, where you will play a crucial role in our multi-year roadmap. Our team champions a strong culture of inclusivity, diversity, and collaboration, and lives MongoDB cultural values every day – we value intellectual curiosity and honesty, and building together in an environment that prioritizes collaboration over competition. If you want to lead a fast-growing team that applies security and systems engineering fundamentals to protect a popular database at scale, join us! We are looking to speak to candidates who are based in Dublin or Cork for our hybrid working model. Candidate Profile 3+ years of experience managing a team of software engineers, including hiring, performance and growth management, compensation planning, and mentoring You have 8+ years of experience building production-quality systems software with large backend/compiled codebases, ideally in Rust. Bonus points for experience with performance profiling, network protocols, TLS, and connection management You have strong technical judgment that you use to effectively guide engineering decisions in security-sensitive or networking-adjacent domains You put the customer first and don't hesitate to cross team boundaries in search of the right solution Solid experience in designing, writing, testing, maintaining, and operating mission-critical software systems Bonus points Professional or advanced academic expertise in the domains of security or networking You enjoy coaching, career development, and creating growth opportunities to help your

mongodbawsazure
View job →

MongoDB is seeking an Engineering Manager to join the Atlas Clusters Organization. The organization is responsible for building MongoDB Atlas, our database-as-a-service offering and fastest growing product. Atlas allows users to deploy fault-tolerant, secure, globally distributed MongoDB clusters in just minutes. This includes developing software to interface with the three major cloud providers (AWS, Azure, and GCP) in order to bring security, durability, availability, and performance to all deployments of MongoDB. The Atlas Clusters Fleet Signal Management team is a product engineering team that builds machinery to rollout hardware and software to the Atlas data plane, detect regressions at scale, and automate remediative actions. The core mission is to establish platform stability by owning the rollout and monitoring systems designed to swiftly detect and mitigate potential issues. We are forming a new Atlas Clusters Fleet Signal Management team in the Dublin area. The Engineering Manager who fills this position will be pivotal in growing that team. We are looking to speak to candidates who are based in Dublin for our hybrid working model. What you’ll do Lead a team of motivated individual contributors who are eager to learn and grow Contribute to the code, design, and architecture of the systems your team develops Work with stakeholders throughout MongoDB to build our roadmap and product offerings Work with customers and support engineers to fix issues and become part of our on-call rotation Collaborate with team members to develop our codebase, best practices, and design principles Foster an inclusive and respectful work environment according to MongoDB's Core V

javamongodbaws
View job →
O
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Infrastructure Platform and Shared Services Team Okta authenticates, authorizes and provisions millions of users a day. The service is hosted on Amazon Web Services (AWS) across multiple availability zones and geographically separated regions. The service is designed for high throughput and 99.999 availability. We're looking for a technical leader to help us continue to scale the service with great people and reliable, cost-effective, and efficient infrastructure, processes, and tooling. As the Sr. Manager of Infrastructure Platform and Shared Services, you will oversee multiple teams focused on Edge networking, K8s platform, Observability, automation platform & tooling. What you’ll be doing Lead the Infra platform and shared services org and various initiatives across SRE & Infrastructure organization. Build a world-class observability platform and monitoring capabilities enabled with self-service Accelerate the velocity of SRE and product engineering by developing robust platforms, powerful tooling, and intuitive self-service capabilities. Own the design and operation of scalable, self-service Cloud infrastructure platforms (e.g. Observability Platform, SRE Productivity, deployments, and Edge Infrastructure) Lead, mentor, and grow a high-performing team of engineers and managers across SRE and infrastructure shared services domains. Perform engineering design evaluations and ensure the completion of projects within resource,

awsci/cdrest
View job →
O
Okta
📍 San Francisco• Full-time• From $232K/yr
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Core Engineering Team Okta powers authentication and authorization for thousands of organizations worldwide. We make access to applications safe, secure, and seamless for billions of logins worldwide. Within Okta, the Core team builds software and frameworks and works with infrastructure teams to deliver 99.99% uptime for our core authentication and authorization products. The Software Engineering Manager Opportunity Okta's Core Engineering team is responsible for building and evolving shared infrastructure and services that lay the foundation for what other engineering teams build on. We're in charge of common shared services like search, cache, configuration management, frameworks for async job management, and email pipeline, to name a few. We're cloud native, where redundancy, multi-tenancy, scale, resource optimization, and resiliency are first-class citizens. With Okta's mantra of 'Always On!' there's never a dull moment. Our biggest asset is our team of passionate engineers and technically minded managers. What You’ll Be Doing Manage a distributed team, including setting expectations and removing blockers, creating a collaborative working environment, hiring and recruitment, providing coaching and career management discussions Collaborate with managers, architects, product owners, project managers, test partners, security and operations engineers to implement best practices related to resilience Communicate and organize cross-team projects with hi

javasqlaws
View job →
M
Mongodb
📍 Dublin• Full-time
1mo ago

MongoDB is seeking an Engineering Manager to join the Atlas Growth Team. The team is responsible for improving and creating new features for MongoDB Atlas, our developer data platform that accounts for 65% of the company’s revenue. Atlas allows users to deploy fault-tolerant, secure, globally distributed MongoDB clusters in just minutes. The Atlas Growth team focuses on improving the customer experience via new product experiences and refinements to existing user flows. The team works closely with cross-functional partners, as well as other MongoDB engineering teams to bring new visions to life. We are constantly challenged to design features such as new onboarding flows, monetization improvements, and advanced cluster management tools for our large B2B customer base. We are looking to speak to candidates who are based in Dublin for our hybrid working model. What You’ll Do Manage a team of engineers to design, build and test new features for MongoDB Atlas Contribute to and lead complex technical projects Work with cross-functional stakeholders to design the team's roadmap, defining delivery dates that balance technical feasibility with the pace of the market Work closely with product, design and analytics teams, considering the user’s perspective while building technical solutions Collaborate with team members to develop our codebase, best practices, and design principles Learn from and mentor an impassioned array of team members We’re Looking for Someone Who Has at least 5 years of professional software development experience Has at least 2 years of people management experience Is skilled at writing large-scale, distributed backend systems in a compiled language (Java, C#, Go, etc.) Is comfortable working across the stack of a modern web application (e.g. React, TypeScript, React Testing Library) Has experience with at least one major cloud provider technology (AWS, Azure, GCP) Has a deep understanding of product analytics Has experience with A/B testing and

typescriptjavareact
View job →

MongoDB’s Storage Layer Services (SLS) team is re-architecting the MongoDB cloud storage layer and sits at the heart of our next-generation cloud storage architecture. This relatively new team is building performant, multi-tenant distributed storage services that both enhance today’s Atlas storage stack and enable more customer workloads to run more efficiently. As the Site Reliability Engineering Manager for SLS, you will partner with the teams building these storage services to define SLOs, shape capacity plans, and ensure the reliability, durability, and operational safety of the storage layer that underpins Atlas. You’ll help grow and lead a small, senior team of SREs as founding members of this organization, playing a crucial role in executing on a multi-year roadmap for MongoDB’s cloud storage architecture. We are looking to speak to candidates who are based in Dublin for our hybrid working model. Responsibilities Build and lead a team of 6-8 engineers, fostering a positive culture, handling career growth and performance conversations, and proactively removing blockers Define and drive a clear technical vision and comprehensive roadmap for our multi-tenant distributed storage systems, balancing long-term strategic infrastructure goals with immediate engineering needs Contribute through hands-on technical work, such as leading architectural design reviews, reviewing PRs, and stepping in to guide the team through complex operational challenges Act as the primary liaison for the Storage Layer Services SRE team, collaborating closely with other engineering leaders to ensure platform alignment and manage stakeholder expectations You may be a good fit if you Have 10+ years of experience working on software and operating distributed systems, with 2+ years managing engineering teams Possess a customer-focused mindset, treating internal developers as your primary users Value efficiency in processes and operations, and have a track record of optimizing team workflows Pr

mongodbawsazure
View job →
M
1mo ago

MongoDB’s Storage Layer Services (SLS) team is re-architecting the MongoDB cloud storage layer and sits at the heart of our next-generation cloud storage architecture. This relatively new team is building performant, multi-tenant distributed storage services that both enhance today’s Atlas storage stack and enable more customer workloads to run more efficiently. As the Site Reliability Engineering Manager for SLS, you will partner with the teams building these storage services to define SLOs, shape capacity plans, and ensure the reliability, durability, and operational safety of the storage layer that underpins Atlas. You’ll help grow and lead a small, senior team of SREs as founding members of this organization, playing a crucial role in executing on a multi-year roadmap for MongoDB’s cloud storage architecture. We are looking to speak to candidates who are based in New York City for our hybrid working model. Responsibilities Build and lead a team of 6-8 engineers, fostering a positive culture, handling career growth and performance conversations, and proactively removing blockers Define and drive a clear technical vision and comprehensive roadmap for our multi-tenant distributed storage systems, balancing long-term strategic infrastructure goals with immediate engineering needs Contribute through hands-on technical work, such as leading architectural design reviews, reviewing PRs, and stepping in to guide the team through complex operational challenges Act as the primary liaison for the Storage Layer Services SRE team, collaborating closely with other engineering leaders to ensure platform alignment and manage stakeholder expectations You may be a good fit if you Have 10+ years of experience working on software and operating distributed systems, with 2+ years managing engineering teams Possess a customer-focused mindset, treating internal developers as your primary users Value efficiency in processes and operations, and have a track record of optimizing team workf

mongodbawsazure
View job →
🔔

Get new ai engineering director jobs by email

Daily job updates · Unsubscribe anytime