Jobiba hiring network

Platform Deployment Management Lead Jobs

10,000 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current platform deployment management lead jobs. Use filters to narrow by work mode, employment type, experience and date posted.

SF
Stitch Fix
📍 United States• Full-time• Remote• From $136K/yr
1mo ago

About Stitch Fix, Inc. Stitch Fix (NASDAQ: SFIX) Stitch Fix is redefining retail by combining human creativity with advanced data science and Generative AI. As we build the future of personalized shopping, we’re equally committed to building yours. We believe in investing in our team as much as our technology. Join us to be a trendsetter in the industry and help us redefine what’s possible for our clients, while we help you reach your full potential. About the Role As an ML Platform Engineer at Stitch Fix, you will play a key role in building and maintaining the critical infrastructure that powers machine learning and AI across our organization. You will design, develop, and support scalable, resilient services and frameworks for ML model training and deployment, feature engineering and serving, candidate generation, AI agent deployment and observability, and other core platform capabilities. In this role, you'll contribute to the day-to-day operations of the ML Platform team, ensuring the smooth functioning of existing systems while driving improvements. You’ll collaborate closely with full-stack data scientists, offering consultation and support to help them unlock the full potential of our platform. With significant autonomy, you’ll have the opportunity to shape the future of ML and AI at Stitch Fix. Your ideas and expertise will drive improvements, codify best practices, and influence how we approach machine learning and AI systems at scale. Responsibilities: Collaborate with cross-functional teams, including data scientists, engineers, and business partners, to solve complex distributed systems and business challenges at scale. Be part of a team with high visibility across the organization, driving impactful solutions that make a difference. Share your ideas and help guide the team’s investments toward high-value opportunities. Foster a culture of technical collaboration and contribute to the development of scalable, resilient systems. About You You bring

REMOTEpythonredisaws
View job →
P
Postman
📍 San Francisco• Full-time
1mo ago

Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity As a Senior Backend Engineer on the Cloud Platform team, you will play a key role in building the core systems and services that power Postman’s internal platform. You’ll help create new backend services that manage how we deploy, scale, and operate our infrastructure and product services, leveraging Java, Spring Boot, and Hibernate (JPA) on top of cloud-native technologies like Kubernetes, ArgoCD, Istio, and Terraform. This role is highly impactful: the systems you build will be used across Postman engineering, enabling faster delivery, better scalability, and a stronger developer experience. You’ll also have the opportunity to contribute to open source, shaping tools that extend beyond Postman’s boundaries. What You’ll Do Design and develop backend services in Java and Spring Boot to support Postman’s internal Cloud Platform. Architect new services that manage service deployment, lifecycle, and scaling across Kubernetes clusters. Implement GitOps workflows (ArgoCD) to support continuous delivery. Integrate with cloud-native tooling such as Istio, Helm, and Terraform. Apply strong soft

javakubernetesci/cd
View job →
E
1mo ago

About the role As a Deployment Strategist, you'll work as part of a driven and creative team of Forward Deployed Engineers, Go to Market professionals, Product Engineers and other Strategists to deploy ElevenAgents, our enterprise AI platform, against the challenging problems our customers face. Your mission is to synthesize disconnected streams of thought into a cohesive understanding of what the most important problem is, what the existing workflows are, what the product needs, what users are motivated by, and where the impact could be. No two days are the same, but as a Deployment Strategist you can expect to: Meet with strategic customers to deeply understand their AI transformation goals and locate their biggest pain points, particularly around customer-facing and operational workflows powered by AI agents. Own flagship deals end-to-end — from identifying the right use case to structuring commercial terms (including outcome-based pricing, license fees, and implementation arrangements) and driving them to close. Embed yourself deeply inside strategic customers: act as a trusted partner with a provisioned account, spend meaningful time onsite, and own delivery outcomes as if you were part of their team. Identify high-impact use cases through close engagement with customer problems and workflows, and collaborate with Forward Deployed Software Engineers to bring them to life. Guide customers on best practices for deploying our products (e.g. ElevenAgents); including agent design, agent orchestration, production readiness to maximize adoption and impact. Scope out potential applications in new industries and expand our AI solutions across different sectors globally. Present the results of our work and proposals for future work to audiences ranging from technical teams to C-suite executives. Collaborate with our Research teams to feed field insights back into ElevenLabs' platform and models, helping shape the roadmap. Build and deliver compelling demos of ElevenAgent

pythonaigo
View job →
E
1mo ago

About the role As a Deployment Strategist, you'll work as part of a driven and creative team of Forward Deployed Engineers, Go to Market professionals, Product Engineers and other Strategists to deploy ElevenAgents, our enterprise AI platform, against the challenging problems our customers face. Your mission is to synthesize disconnected streams of thought into a cohesive understanding of what the most important problem is, what the existing workflows are, what the product needs, what users are motivated by, and where the impact could be. No two days are the same, but as a Deployment Strategist you can expect to: Meet with strategic customers to deeply understand their AI transformation goals and locate their biggest pain points, particularly around customer-facing and operational workflows powered by AI agents. Own flagship deals end-to-end — from identifying the right use case to structuring commercial terms (including outcome-based pricing, license fees, and implementation arrangements) and driving them to close. Embed yourself deeply inside strategic customers: act as a trusted partner with a provisioned account, spend meaningful time onsite, and own delivery outcomes as if you were part of their team. Identify high-impact use cases through close engagement with customer problems and workflows, and collaborate with Forward Deployed Software Engineers to bring them to life. Guide customers on best practices for deploying our products (e.g. ElevenAgents); including agent design, agent orchestration, production readiness to maximize adoption and impact. Scope out potential applications in new industries and expand our AI solutions across different sectors globally. Present the results of our work and proposals for future work to audiences ranging from technical teams to C-suite executives. Collaborate with our Research teams to feed field insights back into ElevenLabs' platform and models, helping shape the roadmap. Build and deliver compelling demos of ElevenAgent

pythonaigo
View job →
HI
HP IQ
📍 San Francisco• $149.9K – $270K/yr
11 days ago

Who We Are HP IQ is HP’s new AI innovation lab. Combining startup agility with HP’s global scale, we’re building intelligent technologies that redefine how the world works, creates, and collaborates. We’re assembling a diverse, world-class team—engineers, designers, researchers, and product minds—focused on creating an intelligent ecosystem across HP’s portfolio. Together, we’re developing intuitive, adaptive solutions that spark creativity, boost productivity, and make collaboration seamless. We create breakthrough solutions that make complex tasks feel effortless, teamwork more natural, and ideas more impactful—always with a human-centric mindset. By embedding AI advancements into every HP product and service, we’re expanding what’s possible for individuals, organisations, and the future of work. Join us as we reinvent work, so people everywhere can do their best work. About The Role As a Senior Platform Engineer at HP IQ, you will help build and evolve the infrastructure, tooling, and shared platform capabilities that enable our engineering teams to develop and operate reliable, secure, and scalable services across cloud and edge environments . You will work closely with application, services, AI/ML, and security teams to improve developer velocity, production readiness, reliability, and operational efficiency across a heterogeneous infrastructure footprint. What You Might Do Design, build, and maintain shared infrastructure and platform capabilities across cloud and edge environments. Build automation and self-service tooling that improves engineering velocity and operational consistency. Develop and maintain Infrastructure-as-Code, deployment workflows, and environment provisioning. Partner with engineering teams on production readiness, including reliability, security, observability, scalability, and recovery. Improve monitoring, alerting, incident response, and operational tooling across distributed environments. Automate repetitive operational t

pythonkubernetesai
View job →
DU
12 days ago

About the Team DoorDash Labs is an independent team within DoorDash. We explore robotics and automation to transform last mile logistics in the long term. If you have a passion for applying robotics solutions in a service used by millions of people, then we want to talk to you! About the Role We are hiring an Autonomy Platform Engineer to build and evolve the foundational software that runs our autonomy stack across robot and compute platforms. The Autonomy Platform team works across embedded Linux, compute and sensor enablement, robotics middleware, process orchestration, data capture and replay, system observability, and performance. You will develop production software and tooling that enables autonomy engineers to bring up new hardware, deploy services reliably, diagnose failures, and validate system performance across robot generations. You will work closely with autonomy, firmware, electrical, hardware, manufacturing, and validation engineers and report to the Autonomy Platform Lead. We expect this role to be hybrid with some time in-office and some time remote. You’re excited about this opportunity because you will… Build and maintain core runtime, middleware, and platform services used by autonomy applications. Enable new compute, camera, lidar, and other sensor platforms. Improve process orchestration, messaging, configuration, startup and shutdown behavior, resource isolation, and fault recovery. Develop system observability, tracing, performance measurement, diagnostics, and regression-detection capabilities. Build reliable data capture, replay, and debugging workflows. Create provisioning, packaging, deployment, integration-test, and platform-readiness tooling. Lead complex debugging across application, middleware, OS, driver, networking, timing, and hardware boundaries. We’re excited about you because… Strong production C++ and Python experience. Experience with embedded Linux, robotics, autonomous vehicles, or complex mechatronic systems. Solid u

pythongitlinux
View job →
T
17 days ago

Toradex is a global company strongly focused on engineering & technology. We’re powered by a diverse & uniquely gifted workforce. We pursue the best people to propel our innovative vision of embedded computing and IoT. If you’re interested in being a driving force at an agile technology company, engineering clever computing solutions & helping other companies bring their products to life, we should talk. Description We are looking for an Integration Platform Specialist to manage company-wide integrations using Workato. The role focuses on building reliable workflows, supporting API-based connections, and working with teams to automate business processes. The position requires strong knowledge of REST APIs, GraphQL, Postman, and the Workato platform. Experience with Workato MCP functionality and coding skills in Python, Ruby, or similar languages are preferred. About You You enjoy solving complex system and process problems with practical, scalable solutions. You can work independently and take ownership of integrations from discovery through deployment and support. You communicate clearly with both technical and non-technical stakeholders. You document your work well and create clear support material for future maintenance. You are curious, hands-on, and willing to investigate issues until you find the root cause. You care about reliability, data quality, security, and a good internal user experience. You are comfortable working across teams and balancing business priorities with technical constraints. Key Responsibilities Own the integration platform roadmap and day-to-day operation, ensuring business-critical automations are reliable, observable, and maintainable. Partner with business and application owners to turn process gaps into pragmatic integration designs and delivery plans. Build Workato recipes, custom connectors, and reusable patterns that reduce manual work and improve data flow between systems. Maintain and modernize existing integrations,

javascriptpythonjava
View job →
C
1mo ago

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! About North: North is Cohere's cutting-edge AI workspace platform, designed to revolutionize the way enterprises utilize AI. It offers a secure and customizable environment, allowing companies to deploy AI while maintaining control over sensitive data. North integrates seamlessly with existing workflows, providing a trusted platform that connects AI agents with workplace tools and applications. Why This Role? This role offers a unique opportunity to shape how enterprises harness the power of AI in real-world applications. As a bridge between our core North product and our clients’ engineering teams, you’ll be at the forefront of solving complex problems and securely integrating AI into critical sectors such as finance, healthcare, and telecommunications. We’re looking for Software Engineers with Applied AI experience who can own the design, build, and deployment of agentic workflows powered by Large Language Models (LLMs), from early prototypes to production-grade AI agents, to deliver concrete business value in enterprise workflows. You’ll work closely with customers on real-world business problems, often building first-of-thei

REMOTEpythonreactgit
View job →
A
Amplitude
📍 Remote• Full-time• $165K – $247K/yr
1mo ago

About the Role Amplitude's Cloud Platform team builds the systems that every Amplitude engineer relies on every day to ship code — and we're rebuilding them for the AI era. As a Senior Platform Engineer, you'll own medium-to-high-complexity platform projects end-to-end and help shape a platform where AI agents are first-class users alongside humans: kicking off deploys, opening pull requests against infrastructure, and triaging incidents, so a single engineer can get the throughput of a team. You'll partner with Staff engineers and product teams to make Kubernetes effortless across the engineering org, building self-service automation and scalable AWS infrastructure that lets product teams ship faster, safer, and with less cognitive load. If you're excited about building the systems that other engineers will rely on every day, this role is for you. Key Responsibilities Lead high-impact platform projects — design and ship capabilities that move the needle on developer experience, reliability, or security, and set the bar for quality, testing, and safe deployment practices. Build the AI-augmented platform. Design tooling and workflows that help engineers get more out of AI-assisted development — think infra primitives that are easy to reason about, automated review, and policy-as-code that keeps the guardrails strong as AI shifts how code gets written. Own Infrastructure-as-Code for Kubernetes, AWS, and GCP using Terraform, Helm, Kustomize, and emerging tooling — and make it consumable enough that an LLM can safely PR against it. Evolve our CI/CD backbone (Argo CD / Workflows / Rollouts, GitHub Actions) to make deploys faster, safer, and easier to reason about. Instrument and operate. Drive observability with Datadog and Amplitude, own dashboards and SLOs, and use the data to push reliability forward. Participate in on-call, lead incident response when needed, and turn postmortems into durable platform improvements. Reduce toil and tech debt with pragmatic remediation

pythonawsgcp
View job →

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! About North: North is Cohere's cutting-edge AI workspace platform, designed to revolutionize the way enterprises utilize AI. It offers a secure and customizable environment, allowing companies to deploy AI while maintaining control over sensitive data. North integrates seamlessly with existing workflows, providing a trusted platform that connects AI agents with workplace tools and applications. Why This Role? This role offers a unique opportunity to shape how enterprises harness the power of AI in real-world applications. As a bridge between our core North product and our clients’ engineering teams, you’ll be at the forefront of solving complex problems and securely integrating AI into critical sectors such as finance, healthcare, and telecommunications. We’re looking for Software Engineers with Applied AI experience who can own the design, build, and deployment of agentic workflows powered by Large Language Models (LLMs), from early prototypes to production-grade AI agents, to deliver concrete business value in enterprise workflows. You’ll work closely with customers on real-world business problems, often building first-of-thei

pythonreactgit
View job →
C
1mo ago

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! About North: North is Cohere's cutting-edge AI workspace platform, designed to revolutionize the way enterprises utilize AI. It offers a secure and customizable environment, allowing companies to deploy AI while maintaining control over sensitive data. North integrates seamlessly with existing workflows, providing a trusted platform that connects AI agents with workplace tools and applications. Why this role? This role offers a unique opportunity to shape how enterprises harness the power of AI in real-world applications. As a bridge between our core North product and our clients’ engineering teams, you’ll be at the forefront of solving complex problems and securely integrating AI into critical sectors such as finance, healthcare, and telecommunications. We’re looking for Software Engineers with Applied AI experience who can own the design, build, and deployment of agentic workflows powered by Large Language Models (LLMs), from early prototypes to production-grade AI agents, to deliver concrete business value in enterprise workflows. You’ll work closely with customers on real-world business problems, often building first-of-thei

pythonreactgit
View job →

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! About North: North is Cohere's cutting-edge AI workspace platform, designed to revolutionize the way enterprises utilize AI. It offers a secure and customizable environment, allowing companies to deploy AI while maintaining control over sensitive data. North integrates seamlessly with existing workflows, providing a trusted platform that connects AI agents with workplace tools and applications. Why This Role? This role offers a unique opportunity to shape how enterprises harness the power of AI in real-world applications. As a bridge between our core North product and our clients’ engineering teams, you’ll be at the forefront of solving complex problems and securely integrating AI into critical sectors such as finance, healthcare, and telecommunications. We’re looking for Software Engineers with Applied AI experience who can own the design, build, and deployment of agentic workflows powered by Large Language Models (LLMs), from early prototypes to production-grade AI agents, to deliver concrete business value in enterprise workflows. You’ll work closely with customers on real-world business problems, often building first-of-thei

pythonreactgit
View job →
S
Stripe
📍 San Francisco• Full-time
1mo ago

Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team The API Platform team is responsible for empowering teams to build highly reliable and performant services, backing our payment systems, fraud detection, and a multitude of other products. As an infrastructure team, you will build and expand the service frameworks to make complex systems easy to use, resilient, and scalable. We build powerful interfaces for engineers that depend on these systems—deployment, load balancers, web framework, databases, Kafka, and Kubernetes—while keeping them highly available and performant. We're looking for engineering leaders who drive the technical vision of Stripe's service infrastructure platform for thousands of engineers to use and build on top of, and thrive in a highly autonomous environment with many moving pieces. What you’ll do You will join as a Technical Lead for one of the most impactful teams at Stripe. You will lead a team of engineers, collaborate with infrastructure and product engineering orgs, and advance service-oriented architecture (SOA) adoption at Stripe. By collaborating with the team's technical leaders, you will ensure the software your team builds meets the needs of Stripe and its customers. We are a highly effective team that consistently delivers high-impact results while genuinely caring for one another. We expect you to bring your curiosity and critical thinking skills to this challenging domain. We are looking for individuals with a strong background in designing and deliverin

pythonjavakubernetes
View job →
N
Nvidia
📍 Remote, United States• Remote
14 days ago

NVIDIA’s DGX Cloud organization is seeking a Senior Data Engineer to become part of its data team! We develop the reliable data foundation that supports fleet health, capacity, utilization, cost, reliability, and operational decision-making throughout DGX Cloud. Our platform supports engineering, operations, finance, and product teams managing and expanding large GPU fleets across cloud service providers and NVIDIA Cloud Partners. We are looking for a practical engineer and technical lead to take charge of a key part of the Navigator data platform. We develop the systems that transform distributed infrastructure telemetry and operational data into dependable, managed data products that support fleet health, capacity, utilization, cost, and operational decisions. We are seeking a hands-on, platform-minded engineer to build and evolve the systems that turn distributed infrastructure telemetry and operational data into reliable, governed data products. You will work across ingestion, transformation, data quality, platform architecture, security, observability, and self-service consumption to help make Navigator and the DGXC data platform a dependable source of truth. We do expect strong engineering fundamentals, experience operating production systems, and the ability to learn new platforms and domains quickly. What you'll be doing: Own systems end to end. For example, work from ambiguous customer and operational needs through architecture, implementation, deployment, observability, incident response, and ongoing support. Construct data pipelines and products. Such as designing and maintain batch and streaming ingestion, transformation, reconciliation, and serving paths for fleet, capacity, utilization, cost, scheduling, and operational telemetry. Build shared libraries, workflow and DAG or equivalent experience abstractions to evolve the data platform. Develop deployment tooling, data

REMOTEpythonsqlaws
View job →
T-
Tubi - Canada
📍 Toronto• Full-time• From C$1.4M/yr
17 days ago

About the Role: We're hiring Senior and Staff Data Platform Engineers to join the Data Infrastructure teams in Toronto. Together these teams own the infrastructure that processes billions of events per day: Spark-on-Kubernetes, Flink and Kinesis pipelines, a multi-petabyte Delta Lake, a large-scale MemoryDB feature store, Databricks multi-environment operations, and the catalog and lifecycle systems that govern it. The team is small and senior. Each engineer owns major platform components: you design it, build it, and support it in production. This is a hybrid-role based out of our Toronto office. You must be willing to travel to our Toronto office two days/week. What You'll Do: Spark-on-Kubernetes — EKS-based compute platform for Spark workloads: cluster configuration, Pod Identity IAM, job environment setup, Kustomize overlays, and shadow canary validation Event ingestion — Rust services and Flink jobs processing billions of events per day over Kinesis; throughput, reliability, on-call response, and AI-assisted operational tooling to reduce toil Platform infrastructure — Terraform modules for environment provisioning, cross-account AWS IAM, ARC runner infrastructure, and CI/CD for data platform changes Feature store and ML compute — Flink-based real-time feature pipelines feeding a large-scale MemoryDB cluster; GPU capacity governance and Databricks multi-environment operations for ML training workloads Workflow orchestration and CDC — Airflow-based DAG deployment, change data capture pipeline operations, and data quality monitoring Your Background: 3+ years building and operating production data platform infrastructure at the cluster or platform level, across Spark, Flink, Kinesis, Kubernetes, or equivalent Deep experience in at least one of: Spark-on-K8s cluster operations, Rust-based data or systems engineering, Kubernetes platform engineering and IaC, or data catalog and governance tooling Production AWS experience or equivalent: EKS, S3, Kinesis, and mu

pythonjavaaws
View job →
🔔

Get new platform deployment management lead jobs by email

Daily job updates · Unsubscribe anytime