Clear all

Jobiba hiring network

Senior Data Engineering Manager Jobs

7,101 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current senior data engineering manager jobs. Use filters to narrow by work mode, employment type, experience and date posted.

N
1mo ago

The NVIDIA DGXC Data Services team builds cloud-native systems, frameworks, and services for managing data across hybrid and multi-cloud infrastructure. We are building the next-generation data and storage infrastructure to solve some of the hardest problems in AI: storage, access, ingestion, governance, observability, and data management for exabyte-scale, high-performance GPU-based training and inference jobs. Our work gives NVIDIA teams the foundational capabilities they need to build, train, deploy, and operate AI products at scale without reinventing critical data infrastructure for every workload. What you will be doing: Build storage technologies, client libraries, and filesystem frameworks that help AI workloads access data across object stores, file systems, and hybrid cloud infrastructure. Develop high-performance storage paths for training and inference workflows, including data loading, checkpointing, caching, POSIX-style access, and object-store integration. Build observability systems that diagnose storage bottlenecks, attribute GPU idle time to I/O behavior, and expose actionable telemetry through production monitoring stacks. Improve performance, scalability, and reliability of storage systems serving massive datasets, deep directory trees, and high-concurrency AI workloads. Work closely with internal AI teams, platform teams, SRE, and operations to validate storage behavior against real workloads and production environments. Use modern software engineering practices, including AI-assisted and agentic development workflows, while maintaining high standards for design, testing, security, performance, and verification. What we need to see: BS in Computer Science, Information Sys

pythonjavakubernetes
View job →

About the Team Security is at the foundation of OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity. The Identity Infrastructure Engineering team sits at the core of this effort, designing and building the identity and access management solutions that protect our model weights, customer data, and critical systems across multiple cloud environments. We partner with teams across OpenAI—Applied Engineering, Research, IT, and Security—to provide a secure and scalable platform for permissioning, orchestration, and innovative AI research. About the Role We’re looking for a Staff+ Software Engineer to help build and evolve the identity infrastructure that supports OpenAI’s research, engineering, and internal platforms. This role sits at the intersection of cloud infrastructure, identity systems, and software engineering. You’ll work across production systems, infrastructure-as-code, cloud control planes, identity providers, and operational infrastructure to build secure, scalable, and reliable systems used broadly across the company. The ideal candidate has experience building and operating large-scale, mission-critical systems with strong reliability and security requirements, and is comfortable writing production code, designing distributed systems, and driving ambiguous projects from 0 to 1 while building the operational rigor needed to run critical infrastructure over time. In this role, you will: Lead the architecture, development, and operation of identity infrastructure that spans cloud platforms, internal systems, and critical engineering services. Design and evolve systems for authentication, authorization, access governance, auditability, and policy enforcement with a strong focus on reliability, scalability, and secure-by-default design. Build foundational infrastructure and platform capabilities that are broadly used across engineering, research, and security teams. Improve the reliability, observability, performance, and op

pythonawsrest
View job →
M
1mo ago

About the Role We are seeking a Staff Enterprise Architect, Data to lead the strategy, design, and modernization of our enterprise data landscape. This role operates at the intersection of data architecture, engineering, and AI enablement, defining solutions to integrate our Data Lake and Data Warehouse across multi-cloud platforms. Over the next 12-18 months, you will enable self-service data access and natural language query capabilities for business users. You will architect Master Data Management and data lineage frameworks ensuring AI models operate on high-quality, governed data. You will also evaluate and implement AI-powered tools to automate data quality monitoring and enhance data security. We're looking to speak with candidates based in the San Francisco Bay Area for our hybrid working model. Key Responsibilities Data Strategy & Roadmap Design semantic layer architecture standardizing business metrics enterprise-wide. Define governance guardrails ensuring natural language queries access validated master data sources Develop Master Data strategy for Customer and Product domains (phases 1-2), Finance and People to follow. Define golden record requirements, stewardship models, and system-of-record hierarchy. Partner with business owners on master data governance Define cross-cloud data integration strategy and reference architecture. Specify patterns (federation, replication, abstraction layer) balancing performance, cost, and data freshness. Document trade-offs and recommend implementations for batch and near-real-time use cases Develop 12-24 month data architecture roadmaps for Finance, Sales, Product, and People. Identify capability gaps and recommend technology investments with business value and effort estimates Systems Design & Solution Leadership Evaluate AI-powered data observability platforms for quality monitoring, pipeline failure prediction, and data classification. Define requirements, lead vendor POCs, and establish integration patterns

pythonsqlmongodb
View job →
DC
Diligent Corporation
📍 Vancouver• Full-time• From C$250K/yr
22 days ago

Overview We are seeking a hands-on Director of AI Software Engineering to lead and scale AI engineering efforts supporting multiple business units across Governance, Risk, and Compliance (GRC). This role sits at the intersection of product delivery, platform evolution, and applied AI—driving real-world impact across core workflows. This is not a pure management role. We are looking for a builder who leads from the front, someone who has recently written production code, shipped systems end-to-end, and can operate comfortably in ambiguity while aligning teams and stakeholders. What You’ll Do Lead AI Engineering Across GRC Own delivery of AI-powered capabilities embedded directly into business unit workflows (e.g., risk analysis, compliance automation, reporting, due diligence) Partner with product, data, and platform teams to translate business problems into scalable AI systems Stay Hands-On Contribute to architecture, code reviews, and critical path implementation Prototype and validate new approaches (LLMs, agents, retrieval systems, classification pipelines, etc.) Set engineering standards for performance, reliability, and cost efficiency Build and Scale Teams Lead and mentor a high-performing team of AI/ML and software engineers Drive hiring, coaching, and career development Establish a culture of ownership, speed, and technical excellence Drive Execution Deliver production-grade systems—not experiments Balance speed with rigor (security, privacy, compliance) Operate across multiple concurrent initiatives with clear prioritization Communicate and Influence Act as a bridge between engineering and business stakeholders Clearly articulate trade-offs, risks, and outcomes to senior leadership Align cross-functional teams around shared goals and timelines What We’re Looking For Proven Builder 10+ years in software engineering, with recent hands-on coding experience Demonstrated track record of shipping production systems at scale Experience with modern

pythonjavaaws
View job →
O
17 days ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Senior Systems Engineer Opportunity As a Senior Systems Engineer at Okta, you will be at the heart of how our global workforce communicates, collaborates, and creates. You won’t just be "keeping the lights on"; you will be responsible for architecting workflows, securing our ecosystem, and ensuring that our stack - from Google Workspace to Slack and Gemini - works seamlessly for every employee, everywhere. What You’ll Be Doing Infrastructure Management: Administer and optimize our core productivity suite, including Google Workspace, Gemini, Slack, Asana, Smartsheet, and other tools. AI Integration & Strategy: Lead the deployment, management, and troubleshooting of enterprise AI tools (specifically Gemini), ensuring they are integrated safely and effectively into daily workflows. Enablement: Create technical documentation and partner with cross-functional teams to drive adoption and best practices for productivity tools through effective training and support. Global Support: Act as a Tier 3 escalation point for complex technical issues, supporting a fast-paced, 24/7 global environment from one of our US hubs. Lifecycle Management: Responsible for the complete lifecycle management of multiple enterprise SaaS applications. Governance & Standards: Define and enforce global standards for tool usage, security permissions, and data governance. Problem Solving & Innovation: Consistently look for root-cause trends while solving platform problem

pythongitlinux
View job →
M
Mongodb
📍 United States• Full-time• From $127K/yr
1mo ago

The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, deployment machinery, and observability and alerting systems. The Fabric team manages the infrastructure that enables secure communication between systems and from the public internet. Their responsibilities encompass network architecture, service mesh, and edge load balancing, ensuring customer data remains safe in transit. The team plays a crucial role in developing and maintaining the reliable and globally connected multi-cloud network that supports MongoDB products. This role can sit in our NYC HQ, our smaller Austin, Palo Alto, or San Francisco offices, or fully remote from anywhere in North America. When based in an office, we provide hybrid work accommodation. Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong networking background to join the Fabric team. This role is pivotal in building and maintaining the robust infrastructure necessary for secure and efficient communication between our services. As an SRE on the Fabric team, you will leverage your expertise in networking, distributed systems, and automation to ensure our systems are resilient, scalable, and reliable. The ideal candidate should Have 10+ years of experience working on software and operating distributed systems, with deep expertise in networking fundamentals and a good understanding of how the internet works, e.g. TCP/IP (including IPv6), DNS, TLS/mTLS, BGP, tunnels, overlays, and SDN principles Possess a customer-focused mindset, driving improvements that benefit end-users Value efficiency in processes and operations, and display a strong preference for automation over manual processes (“allergic to ops work”) Be intimately familiar with modern cloud-based infrastructure and the network design prim

mongodbawsazure
View job →
M
1mo ago

The Team Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, deployment machinery, and observability and alerting systems. The Fabric team manages the infrastructure that enables secure communication between systems and from the public internet. Their responsibilities encompass network architecture, service mesh, and edge load balancing, ensuring customer data remains safe in transit. The team plays a crucial role in developing and maintaining the reliable and globally connected multi-cloud network that supports MongoDB products. This role can sit in our Toronto or Vancouver offices, or fully remote from anywhere in North America. When based in an office, we provide hybrid work accommodation. Role Overview We are seeking a talented Site Reliability Engineer (SRE) with a strong networking background to join the Fabric team. This role is pivotal in building and maintaining the robust infrastructure necessary for secure and efficient communication between our services. As an SRE on the Fabric team, you will leverage your expertise in networking, distributed systems, and automation to ensure our systems are resilient, scalable, and reliable. The ideal candidate should Have 10+ years of experience working on software and operating distributed systems, with deep expertise in networking fundamentals and a good understanding of how the internet works, e.g. TCP/IP (including IPv6), DNS, TLS/mTLS, BGP, tunnels, overlays, and SDN principles Possess a customer-focused mindset, driving improvements that benefit end-users Value efficiency in processes and operations, and display a strong preference for automation over manual processes (“allergic to ops work”) Be intimately familiar with modern cloud-based infrastructure and the network design primitives of at least one of AWS, Azur

mongodbawsazure
View job →
L
Lyft
📍 San Francisco• Full-time• ₹1.5L – ₹1.9L/yr
1mo ago

At Lyft, our mission is to improve people's lives with the world's best transportation. To accomplish this, we start with our community by creating an open, inclusive, and diverse organization. About the Team The Risk Tech engineering organization is committed to tangibly reducing accident frequency, saving lives, and managing costs to enhance the safety and affordability of rides. Claim Management is a core financial function for Lyft. Each claim touches complex workflows, multiple stakeholders, sensitive data, financial reserves, regulatory processes, and significant financial liabilities. This role offers the opportunity to define a leading claims management system for the industry. Our vision is to establish a single, Unified Risk Platform where comprehensive claims workflows across all business lines are efficiently administered, communications are consolidated, and data is structured to facilitate data-driven insights and decisions.This enables cost-efficient claims operations, mitigates risks and expenses as Lyft scales, and ensures people receive assistance proactively and accurately. About the Role We are seeking a Senior Software Engineer to contribute to the technical direction, drive architectural decisions, and lead the development of a highly reliable, scalable, and intelligent Risk Management Information System that powers our insurance platform. You will collaborate with passionate colleagues from Engineering, Data Science, Product and Claim Operations to deliver end-to-end solutions. Responsibilities: Define and drive the long-term technical roadmap for claims management systems, aligning priori

pythonreactmachine learning
View job →
L
Lyft
📍 Seattle• Full-time• ₹1.5L – ₹1.9L/yr
1mo ago

At Lyft, our mission is to improve people's lives with the world's best transportation. To accomplish this, we start with our community by creating an open, inclusive, and diverse organization. About the Team The Risk Tech engineering organization is committed to tangibly reducing accident frequency, saving lives, and managing costs to enhance the safety and affordability of rides. Claim Management is a core financial function for Lyft. Each claim touches complex workflows, multiple stakeholders, sensitive data, financial reserves, regulatory processes, and significant financial liabilities. This role offers the opportunity to define a leading claims management system for the industry. Our vision is to establish a single, Unified Risk Platform where comprehensive claims workflows across all business lines are efficiently administered, communications are consolidated, and data is structured to facilitate data-driven insights and decisions.This enables cost-efficient claims operations, mitigates risks and expenses as Lyft scales, and ensures people receive assistance proactively and accurately. About the Role We are seeking a Senior Software Engineer to contribute to the technical direction, drive architectural decisions, and lead the development of a highly reliable, scalable, and intelligent Risk Management Information System that powers our insurance platform. You will collaborate with passionate colleagues from Engineering, Data Science, Product and Claim Operations to deliver end-to-end solutions. Responsibilities: Define and drive the long-term technical roadmap for claims management systems, aligning priori

pythonreactmachine learning
View job →

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? As a Senior Machine Learning Engineer specializing in synthetic data, you will play a pivotal role in developing the synthetic data pipeline that is crucial to Cohere’s advanced language models. Your responsibilities will encompass the end-to-end management of synthetic data, including maintaining and optimizing the synthetic data pipeline, data analysis and generation, as well as conducting data ablations and model evaluation to gauge data quality. You will work with diverse web data and code data and transform them using generative models to improve token efficiency and model quality. By combining research and engineering, you will bridge the gap between raw data and cutting-edge AI models, directly contributing to improvements in critical training metrics like throughput and accelerator utilization. Your work will be essential to Cohere’s mission of delivering efficient and reliable language understanding and generation capabilities, driving innovation in natural language processing. If you are passionate about transforming data into the foundation of AI systems, this role offers a unique opportunity to make a

pythongitrest
View job →

The mission of The New York Times is to seek the truth and help people understand the world. That means independent journalism is at the heart of all we do as a company. It’s why we have a world-renowned newsroom that sends journalists to report on the ground from nearly 160 countries. It’s why we focus deeply on how our readers will experience our journalism, from print to audio to a world-class digital and app destination. And it’s why our business strategy centers on making journalism so good that it’s worth paying for. About the Role, Mission or Department Overview As a Senior Engineer, Marketing, you'll be embedded with the Marketing team, providing us with technical leadership, consulting, and systems thinking. You'll assemble the orchestration systems, integrations, and AI capabilities that help teams work faster, and in more data‑driven ways. You will lead end‑to‑end design, implementation, and management of the marketing campaign lifecycle system, integrations, internal tools, and learning/experimentation infrastructure. You will report to our Director, Advertising Systems. Responsibilities: You will write high-quality, secure, and well-tested code, contributing to standards and documentation for Marketing infrastructure, data models, and tools You will build the campaign lifecycle orchestration backbone You will build integrations and internal tools that connect project management, collaboration, creative, media, and email/lifecycle platforms You will work with Marketing partners as a technical advisor, translating needs into designs You will integrate AI agents and automations (including LLM-powered flows) into Marketing workflows according to company GenAI guidance You will design and operate data pipelines and models that ingest marketing and performance data into a data layer, supporting experimentation and analytics workflows You will deliver abstractions and APIs that ensure AI systems and our users to create campaign wrap-ups, insights, next-t

typescriptpythonreact
View job →
P
22 days ago

About Us Pearl is AI for professional services at global scale, combining advanced AI with verified human expertise to deliver help that is accurate, accountable, and fast. Since 2003, our network has connected millions of customers with licensed professionals across 196 countries, making real expertise available anytime, anywhere. Our Values Data driven: Start with truth, measure what matters. Courageous: Bias to action; run toward hard problems. Innovative: Seek novel, elegant solutions. Lean: Do more with less. Build, ship, learn fast. Humble: Strong opinions, lightly held. About the Role As a Senior Business Systems Analyst at Pearl, you'll play a central role in enabling workflow and workforce AI transformation, with a strong emphasis on Agent Hub requirements analysis, platform support, and project coordination. This role blends classic BSA responsibilities with project management and practical AI fluency, supporting both Agent Hub product evolution and broader AI transformation initiatives across the organization. You'll help define and document requirements for complex AI opportunities, translate high-level ideas into actionable work, and coordinate across engineering, prompt engineering, and internal stakeholders. This is a hybrid role designed to support both requirements analysis and execution oversight for high-impact AI operations and transformation projects. You'll be a heavy, daily user of AI tools, not just experimenting on the side, while also contributing directly to the development and improvement of AI tools and features, especially within Agent Hub and related AI transformation initiatives. This is a great opportunity for someone who wants to work at the center of a company's AI transformation, with meaningful ownership across both platform evolution and broader AI operations, and the chance to help shape how AI-enabled work gets defined, prioritized, and operationalized. What You’ll Do Evaluate and document Agent Hub product requirements,

restagileai
View job →
G
GHX
📍 Hyderabad• Full-time
22 days ago

Role: Senior Cloud Engineer Location: Hyderabad, India (Hybrid) Department: Product Development About GHX: GHX (Global Healthcare Exchange) is a leading healthcare technology company on a mission to simplify the business of healthcare and improve patient outcomes. Founded in 2000, GHX has built the GHX Global Network — the world’s largest cloud-based supply chain community connecting healthcare providers, suppliers, distributors, and partners to automate key processes, reduce costs, and increase operational efficiency. Its solutions span electronic trading, procurement automation, inventory and contract management, business intelligence, and data synchronization, helping healthcare organizations improve productivity and focus more on patient care. Over the years, GHX has enabled significant cost savings for the industry and continues to innovate with intelligent automation and AI-driven capabilities. Website: https://www.ghx.com/ LinkedIn: https://www.linkedin.com/company/ghx/ About Role: The Senior Cloud Engineer leads the design, implementation, and operations of the organization’s cloud infrastructure. This role is responsible for complex projects, high-level architectural planning, and making strategic technology decisions that support scalability, performance, security, and cost optimization. Acting as a technical leader, the Senior Cloud Engineer provides mentorship to junior and mid-level engineers while driving innovation, resiliency, and compliance in cloud environments. Key Responsibilities: Design and Implementation Lead the design and implementation of cloud-native architectures and hybrid cloud solutions. Oversee cloud engineering projects, ensuring solutions are secure, resilient, and cost-effective. Contribute to high-level architectural planning and strategic decision-making around technology adoption. Review and approve design proposals, Infrastructure as Code (IaC) tem

awsazuregcp
View job →
R
Ramp
📍 New York• Full-time• Remote• From $10K/yr
1mo ago

About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role Ramp is in a critical phase of growth. We grew immensely last year and are building out a talented business systems team to ensure we maintain this trajectory for years to come. You’ll work directly with our Sales, Account Management, Partnerships, and Product teams to execute mission-critical business systems projects across the organization. This is a key role where you will be uniquely positioned to impact the full picture of Ramp’s growth efforts through systems development. What You’ll Do Work alongside Sales Operations to administer key go-to-market business systems, including Salesforce, Outreach, Qualified, Zendesk, Hubspot, Looker, Gong.io Build and deploy automation (flows), validations, and applications in Salesforce Implement new systems and integrations as needed Analyze key business requirements and systems capabilities to write specifications for systems build and run end to end implementation Create key reports and dashboards to track systems performance and data accuracy Write and maintain clear documentation on syste

REMOTErestaigo
View job →
P
Prophecy
📍 Bengaluru• Full-time
22 days ago

About Prophecy The leader in AI-native data preparation and analysis, Prophecy is revolutionizing how the world’s top enterprises turn data chaos into reliable insights. We introduce the AI-native data lifecycle (generate, refine, deploy) where our industry leading AI agents and humans work hand-in-hand in visual and document interfaces to analyze, transform and prepare data, to ship trusted insights at enterprise scale. Don’t miss the rocket ship—join Prophecy and build the next data revolution. Position Summary This is a high-impact opportunity to be a senior DevOps engineer in a fast-growing startup, based in Prophecy’s India engineering center. You will own and evolve the foundations that keep our engineering org fast, secure, and cost-efficient at scale — spanning cloud cost management, security DevOps, and CI/CD and engineering operations. You will work with a team of dynamic engineers who take pride in solving complex problems, and you will have the autonomy to set direction in your areas of ownership. The Impact You Will Have Cloud cost (FinOps) Own and evolve our cloud cost optimization program across AWS, Azure, GCP, Databricks, Snowflake and Bigquery building on the programmatic monitoring and controls Analyze billing, asset-inventory, and utilization data across departments to identify wasteful spend and provide actionable insights to optimize it. Develop and maintain cost optimization strategies, roadmaps, and forecasting models Partner with engineering teams to design cost-efficient architectures without compromising scalability or reliability. Build and maintain automation for infrastructure provisioning, scaling, and cost control. Security DevOps Partner with engineering and security to drive our security-hardening program across workstreams such as identity & access governance, secrets & credential lifecycle, cloud access, and CI/CD hardening. Implement and automate guardrails: secrets management, least-privilege access, cr

pythonawsazure
View job →
🔔

Get new senior data engineering manager jobs by email

Daily job updates · Unsubscribe anytime