Jobiba hiring network

Pipeline Excellence Director Jobs

2,475 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current pipeline excellence director jobs. Use filters to narrow by work mode, employment type, experience and date posted.

O
1mo ago

About The Team The Data Understanding team is responsible for creating the high quality datasets and their quantized representation for OpenAI. This includes synthesizing data, building VQ representations, and processing, filtering, deduplication, quality control, and tokenization so it can be used effectively in big model training runs. About The Role We're looking to advance how OpenAI builds and understands pretraining data at scale. You'll treat data quality and curation as core research problems: developing new methods to select, combine, and transform data; creating datasets that improve model capabilities; and designing rigorous experiments to understand how data choices and interventions affect model learning and downstream behavior. You'll work closely with frontier models and web-scale data to build evidence for which approaches work and why, then translate successful research into scalable data processing pipelines We Expect You To Have a strong track record of new or improved ML ideas, through publications, projects, or applied research. Own and drive a research agenda, from choosing the right problems to carrying long-running work through to impact. Be excited by OpenAI’s empirical, collaborative approach to research. Nice To Have Thoughtfulness about AI’s impact, including privacy, provenance, and data quality. Experience building high-performance deep learning or large-scale data processing systems. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an extremely powerful tool that must be created with safety and human needs at its core, and to achieve our mission, we must encompass and value the many different perspectives, voices, and experiences that form the full spectrum of humanity. We are an equal opportunity employer

awsrestai
View job →
B
Blockworks
📍 New York, United States• Full-time
1mo ago

About Us: Blockworks is an information platform that sits at the center of the crypto industry. We transform raw, complex data and facts into actionable research, trusted alpha-driven insights, and world-class events. The result is transparency and confidence. Blockworks connects investors and businesses in onchain capital markets. We give businesses a platform to earn trust and provide investors with the information they need to underwrite the asset class. Who You Are: You have a keen focus on backend and API development and Software engineering is your passion. You are a player-coach and a natural leader who understands the technical and human elements that go into great software design. You have a results-oriented attitude and a passion for delivering flawless releases and developing digital product pipelines (CI/CD pipelines). You have a proven track record facilitating engineering teams to increase productivity and quality. You're excited at the possibility of being on the ground floor of the design and development of backend strategies. You bring a passion for designing and maintaining scalable API services that handle large amounts of data elegantly. You love moving quickly in a fast-paced start-up, but you also bring intentionality, sustainability and scalability to your approach as an engineer. What You’ll Do: As a Senior Backend Engineer at Blockworks, you’ll design, build, and maintain the systems that power our products end-to-end. From high-performance APIs to database architecture, you’ll own the backend layer that makes everything else possible. You won’t just be handed requirements, you’ll help define them, scope projects, and make the architectural decisions that shape our technical foundation. Your work will directly impact our research platform ( blockworksresearch.com ) and our media site ( blockworks.co ; 1M+ monthly active users). Every day will look a little different, but in general, you will do things like: Architect, build, and ship backend

typescriptnode.jssql
View job →
A
1mo ago

About The Role & Team Amplitude is only as useful as the data inside it. The Data Connections team owns how that data gets in and out — importing behavioral and customer data from cloud data warehouses like Snowflake, Databricks, and BigQuery, from cloud object storage like S3 and Azure Blob Storage, and pushing enriched event data back out to warehouses, object storage, streaming destinations, and downstream advertising and marketing platforms. That means batch and streaming pipelines moving billions of events a day, connections that have to keep working across dozens of customer-controlled systems, credentials and configuration that have to stay correct and secure, and latency and reliability targets that customers build their own pipelines on top of. Recent work includes launching new warehouse export destinations, migrating our import pipelines onto a durable workflow engine, building low-latency streaming export, and supporting cross-region and cross-cloud customer storage. You'll lead a team of around 10 engineers building this, reporting to our Head of Data Platform, and you'll be the person accountable for the roadmap, the reliability, and the growth of the people on it. The stack is Java, Temporal, Kafka, Kubernetes, Terraform, DynamoDB, S3, and Snowflake, running on AWS and GCP. Responsibilities Set the direction and roadmap for data import and export, with clear priorities and tradeoffs you can explain to engineers, PMs, and customers Lead, coach, and grow a team of around 10 engineers — hiring, career development, feedback, and performance Partner with your tech leads on architecture across ingestion, transformation, and delivery, without becoming the bottleneck for every decision Own reliability, SLOs, and cost efficiency for pipelines customers depend on daily Expand the set of destinations and sources we support, and make each new integration cheaper to build than the last Work directly with Product, Design, and other Data Platform teams to ship e

javaawsazure
View job →
V
1mo ago

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Vanta’s Developer Experience team builds the tools engineers use every day to bring ideas to production rapidly and reliably. You’ll empower other Vanta engineers to leverage cutting-edge technologies and best practices to make Vanta more performant and scalable on a platform level. Example projects include modernizing our CI/CD pipelines, introducing new test frameworks, launching AI-powered dev tools, and scaling developer environments to support a growing engineering team. This team has a wide breadth of impact across all of product engineering. The work we do compounds in value by making it easier for engineers to diagnose and solve bugs, streamline workflows, and ship value to our customers quickly and safely. Vanta engineers design and develop new product functionality and infrastructure leveraging modern frameworks and tooling, including TypeScript, React, Node.js, MongoDB, Github Actions, and various AWS services such as Fargate and ECS. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. We’d love for you to join us! You will: Set direction for critical dev infrastructure, enabling us to stay ahead of continued rapid growth Design and build CI and build systems that ensure Vanta engineers can develop and ship robust products quickly and confidently Improve the efficiency and reliability of our deployment workflows, including tools for hotfixes, rollbacks, and incident mitigation Lead development of tools that accelerate feedback loops — from typechecking and linting to running tests and deploying changes Build and maintain scalable developmen

typescriptreactnode.js
View job →
S
Supabase
📍 Remote• Full-time
1mo ago

About Supabase Supabase is the Postgres development platform, built by developers for developers. We provide a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search. All services are deeply integrated and designed for growth. About the Role We're looking for a Release Engineer (SRE) to join our Release Engineering team (part of EngOps) — a production-operations expert who brings an SRE mindset to how Supabase ships and runs, making deploys safe, observable, and recoverable at scale. Release Engineering's scope has grown well beyond build-and-ship: we increasingly own the operational reliability of the systems that deploy and run Supabase. In this role you'll treat our deployment pipelines, pre-production signal, and the control plane itself as production systems — with SLOs, error budgets, and on-call ownership — and you'll be the person teams lean on when reliability is on the line. This is not a "gatekeeper" role. You'll make the reliable path the easy path: standardising how we deploy, instrumenting what we ship, and ensuring that when something breaks, we detect it quickly and recover quickly. What You'll Be Responsible For In this role, you'll: Own the reliability of Supabase's deployment and release systems, and the control plane they run on, against clear SLOs and error budgets Turn pre-production into a trustworthy signal — standardizing and instrumenting today's fragmented, ad-hoc deployment workflows Drive disaster-recovery readiness, including making environments reproducibly deployable from scratch (untangling undocumented secrets, unclear configuration ownership, and circular service dependencies) Build and operate health and SLO monitoring for critical user flows, using synthetic testing to catch regressions before customers do Reduce mean-time-to-detect and mean-time-to-recover for deploy-related incidents — which account for a large share of our incident load Participate in on-call, lead blameless po

awskubernetesai
View job →
C
1mo ago

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Evaluation is critical to making progress in scaling intelligence. As models continue to become superhuman in many real-world use cases, we must continue to develop new evaluation techniques that accurately reflect what models are already capable of, as well as set the agenda for what future models should be capable of. In this role, you are responsible for creating these next-generation evaluation methods and infrastructure to measure LLM progress. As a Senior Research Scientist, Model Evaluation, you will: Create ambitious new evaluation benchmarks that push the limits of what our models can accomplish. Work on highly cross-functional teams to translate model feedback into trustworthy, repeatable evaluations. Conduct research to advance the state-of-the-art in LLM evaluation methods, including training LLM judges; refining LLM-based data synthesis pipelines; and improving evaluation efficiency. Build scalable and reusable tools for digging into model performance. You may be a good fit if: You enjoy rapidly building prototypes that demonstrate the boundaries of what LLMs are capable of, and you have developed res

C
Clickup
📍 United States• Full-time
1mo ago

At ClickUp, we're building the future of work: the first truly converged AI workspace unifying tasks, docs, chat, calendar, and enterprise search, all supercharged by context-driven AI. We are an AI-native company. Every team member is expected to leverage AI daily, and we evaluate AI fluency as part of our hiring process. Join us and help redefine what's possible. 🚀 Role Overview You'll own and evolve the AI systems behind ClickUp's voice platform: real-time streaming transcription, intelligent reformatting, context-aware mention detection, and voice-to-action pipelines. This is a high-impact, hands-on role where you'll push the boundaries of what voice interfaces can do inside a productivity tool used by millions. Key Responsibilities Design, build, and optimize real-time speech-to-text pipelines (streaming ASR, VAD, audio processing) Improve transcription accuracy through context injection (user names, teams, custom vocabulary, language detection) Develop and maintain LLM-powered post-processing (grammar correction, filler removal, mention resolution, formatting) Build voice-to-action systems that parse natural language into structured workspace commands Evaluate, benchmark, and integrate ASR models (Whisper, AssemblyAI, Fireworks, etc.) for cost, latency, and accuracy Collaborate with product and platform teams to ship voice features across MAX Desktop, Mobile, Web, and Browser Extension Explore multimodal AI capabilities (screen + voice + text) for next-gen assistant experiences Equal Opportunity Employer ClickUp is an Equal Opportunity Employer, and qualified applicants will receive consideration for employment without regard to race, color, religion, sex, sexual orientation, gender identity, or national origin. Privacy Notice ClickUp collects and processes personal data in accordance with applicable data protection laws. You can find further details by viewing our Global Candidate Privacy Notice. If you are a Philippine Job Applicant, please also see our Phi

awsmachine learningai
View job →
S
1mo ago

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role Issue Workflow is Sentry's primary product surface. Our issue platform processes billions events daily and turns them into actionable insights that help millions of developers fix bugs faster. As a Staff Software Engineer on the Issue Workflow team, you'll architect the systems that power this experience. You'll work at the intersection of high-scale distributed systems and product engineering, building real-time data pipelines, search backends, and analysis systems that surface signal from noise. This is product engineering at massive scale—where every architectural decision impacts millions of debugging sessions. You'll be the technical leader who shapes how Sentry groups issues, how we make search lightning-fast, how we enable sophisticated agentic workflows, and how we ensure that the product is performant even at billions-of-events scale. Your work will define what's possible for the most trafficked part of Sentry's platform. In this role you will Drive technical strategy and roadmap. Partner with engineering leadership, product, and design to shape the multi-quarter technical vision for Issue Workflow platform. Make strategic calls about architectural direction, technology choices, and technical debt. Ensure the team is building a strong foundation to scale with Sentry's growth. Solve complex performance and scalability challenges. Champion product quality and user experience. Build features that don't just work—they delight. You understand that milliseconds matter in the developer experience. You sweat the details of interfaces, error messages, loading states, and edge cases. You instrument everything s

typescriptpythonsql
View job →
S
Sentry
📍 San Francisco• Full-time
1mo ago

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role Issue Workflow is Sentry's primary product surface. Our issue platform processes billions events daily and turns them into actionable insights that help millions of developers fix bugs faster. As a Staff Software Engineer on the Issue Workflow team, you'll architect the systems that power this experience. You'll work at the intersection of high-scale distributed systems and product engineering, building real-time data pipelines, search backends, and analysis systems that surface signal from noise. This is product engineering at massive scale—where every architectural decision impacts millions of debugging sessions. You'll be the technical leader who shapes how Sentry groups issues, how we make search lightning-fast, how we enable sophisticated agentic workflows, and how we ensure that the product is performant even at billions-of-events scale. Your work will define what's possible for the most trafficked part of Sentry's platform. In this role you will Drive technical strategy and roadmap. Partner with engineering leadership, product, and design to shape the multi-quarter technical vision for Issue Workflow platform. Make strategic calls about architectural direction, technology choices, and technical debt. Ensure the team is building a strong foundation to scale with Sentry's growth. Solve complex performance and scalability challenges. Champion product quality and user experience. Build features that don't just work—they delight. You understand that milliseconds matter in the developer experience. You sweat the details of interfaces, error messages, loading states, and edge cases. You instrument everything s

typescriptpythonsql
View job →
S
Sentry
📍 San Francisco• Full-time• $155K – $400K/yr
1mo ago

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role The Streaming Platform team at Sentry is building the next generation of infrastructure that powers our ingestion pipelines and real-time data processing systems. Our platform ingests, processes, and distributes hundreds of thousands of events per second with low latency and high reliability. We are creating a system that makes it easy for Sentry engineers to deploy and run Streaming Applications at scale by simplifying the complexity of Kafka, scaling consumers automatically, and managing state so product teams can focus on building great experiences for developers. As part of this team, you will work on challenges at the intersection of distributed systems, real-time data processing, and developer experience. You will help us create a self-service streaming platform that improves stability, accelerates time to production, and reduces operational overhead. In this role you will Design, build, and operate components of our Streaming Platform, including Kafka, the streaming runtime, high-level APIs, and developer-facing abstractions. Implement resilient, high-throughput stream processing systems that handle unbounded datasets with strong correctness guarantees (delivery, checkpointing, watermarking, and more). Build scalable automation and control plane for Kafka fleet management and improve efficiency. Partner with product engineers to ensure our abstractions enable fast, reliable, and consistent ingestion pipelines. Improve observability, monitoring, and failover for mission-critical real-time systems. You’ll love this job if you You enjoy working on distributed systems at scale and care about reliability and

pythonjavasql
View job →

About Pinecone Pinecone is the knowledge infrastructure for AI at scale. Its leading vector database and knowledge engine, Pinecone Nexus, power accurate, performant AI applications for more than 9,000 customers and 800,000 developers worldwide. Pinecone's mission is to make AI knowledgeable. Pinecone is based in New York and raised $138M in funding from Andreessen Horowitz, ICONIQ, Menlo Ventures, and Wing Venture Capital. About the Team and Role: We are hiring a senior/staff software engineer to help design and build core components of our next-generation knowledge retrieval system built for the AI era – search and retrieval infrastructure that powers high-quality, scalable, and enterprise-grade agentic systems. You’ll build the framework that allows our customers to connect knowledge–synthesized from structured and unstructured data–to modern LLM-powered applications, leveraging the world’s best-in-class vector DB supporting semantic search and hybrid retrieval. This role is ideal for someone who loves backend system architecture, distributed systems, and applied AI infrastructure. It is a high impact role with significant ownership across architecture, performance, and system reliability. Responsibilities: Design and build scalable platform components leveraging advanced retrieval via query planning, semantic and hybrid search, metadata-aware search, and LLM generation Design and build optimized indexing pipelines for structured and unstructured data Build backend services for semantic and hybrid retrieval, knowledge graph construction, and retrieval orchestration Improve retrieval quality through evaluation and observability frameworks Design APIs for internal and external user and agentic consumers Optimize latency, throughput and cost across large-scale inference and retrieval workloads Drive technical direction for reliability and security What You’ll Bring to the Table: To thrive in this role, you don't need to check every single box, but you should be deep

pythonjavaaws
View job →

About Pinecone Pinecone is the knowledge infrastructure for AI at scale. Its leading vector database and knowledge engine, Pinecone Nexus, power accurate, performant AI applications for more than 9,000 customers and 800,000 developers worldwide. Pinecone's mission is to make AI knowledgeable. Pinecone is based in New York and raised $138M in funding from Andreessen Horowitz, ICONIQ, Menlo Ventures, and Wing Venture Capital. About the Team and Role: We are hiring a senior/staff software engineer to help design and build core components of our next-generation knowledge retrieval system built for the AI era – search and retrieval infrastructure that powers high-quality, scalable, and enterprise-grade agentic systems. You’ll build the framework that allows our customers to connect knowledge–synthesized from structured and unstructured data–to modern LLM-powered applications, leveraging the world’s best-in-class vector DB supporting semantic search and hybrid retrieval. This role is ideal for someone who loves backend system architecture, distributed systems, and applied AI infrastructure. It is a high impact role with significant ownership across architecture, performance, and system reliability. Responsibilities: Design and build scalable platform components leveraging advanced retrieval via query planning, semantic and hybrid search, metadata-aware search, and LLM generation Design and build optimized indexing pipelines for structured and unstructured data Build backend services for semantic and hybrid retrieval, knowledge graph construction, and retrieval orchestration Improve retrieval quality through evaluation and observability frameworks Design APIs for internal and external user and agentic consumers Optimize latency, throughput and cost across large-scale inference and retrieval workloads Drive technical direction for reliability and security What You’ll Bring to the Table: To thrive in this role, you don't need to check every single box, but you should be deep

pythonjavaaws
View job →
R
1mo ago

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. Job Summary We are looking for an experienced Growth Infrastructure Engineer to build and maintain the technical backbone that enables scalable growth experiments, high-performance data pipelines, and automated systems that drive user acquisition, engagement, and product iteration. This role sits at the intersection of growth, product, and infrastructure — combining deep technical engineering with experimentation and data-driven optimization. You will collaborate with product, data science, and backend teams to ensure that growth initiatives run smoothly and scale efficiently across systems. Key Responsibilities Growth Infrastructure & Systems Design, implement, and maintain scalable infrastructure that supports growth and experimentation needs. Build and optimize analytics pipelines to capture key product and growth metrics (acquisition, activation, retention, etc.). Develop automated workflows for user onboarding, campaign delivery, and performance tracking. Experimentation & Optimization Support A/B testing frameworks and integrate them into production systems. Enable reliable data collection and evaluation for growth experiments. Automate deployment and rollout of growth feature flags and tests. Cross-Functional Collaboration Partner with Growth Product Managers, Data Engineers, and Analysts to define technical requirements for growth initiatives. Translate business goals into technical specifications and system designs. Provide guidance on performance, reliability, and scalability trade-offs. Monitoring & Reliability Implement monitoring and alerting for growth infrastructure services. Troubleshoot production issues and optimize for uptime and performance. Ensure data quality and consistency for report

javascriptpythonjava
View job →
PE
Private Employer
📍 Washington• Full-time• Hybrid
1mo ago

A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role Palantir is at the forefront of some of the most critical and challenging problems in the world. We develop alongside our customers everyday. Our customers span from the cloud to the frontline. As we adapt to solve their most pressing issues in latency, performance, and compute cost, we are building a team of software engineers relentlessly focused on low-level optimization and novel compute architectures. This is a team of developers creating software for the far-edge, including streaming ETL pipelines, inference platforms, and various timing critical applications. This role requires an experienced software engineer who is well versed in low-level development in compiled, native languages such as Rust and C/C++. A successful candidate can optimize software for constrained embedded devices or across large-scale distributed systems. You should have strong knowledge of computer architecture and OS internals.

aic++rust
View job →
PE
Private Employer
📍 New York• Full-time• Hybrid
1mo ago

A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role Talent Strategists at Palantir don't work from a boilerplate requisition list. They embed with their teams, learn what exceptional actually looks like for a given role, and hire the Engineers building the future. As a Technical Sourcer, you'll support that work. You'll partner with Talent Strategists and Hiring Managers to build the pipelines and manage the process for some of our hardest-to-fill roles. The job is equal parts research, relationship-building, and attunement — figuring out where the right people are, what's going to resonate with them, and how to get them in the door. This is an 8-month contract role. It moves fast, the priorities shift, and the standard is high. You will start with two weeks of onboarding before jumping into the deep end.

🔔

Get new pipeline excellence director jobs by email

Daily job updates · Unsubscribe anytime