Jobiba hiring network

Deployment Strategist Lead Jobs

1,782 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current deployment strategist lead jobs. Use filters to narrow by work mode, employment type, experience and date posted.

O
OpenAI
📍 San Francisco• Full-time• Remote
21 days ago

About the Team The Finance Platform & Technology team builds and scales the systems and data architecture that power OpenAI’s core financial operations. We enable business agility, compliance, and operational excellence across procure-to-pay, quote-to-cash, supply chain, financial planning, and asset management. We partner with Procurement, Accounting, Tax, Legal, Security, Data, and Engineering to modernize workflows through thoughtful platform design, reliable integrations, scalable automation, and trusted data. About the Role As a Business Systems Lead for Procure-to-Pay, you will be a hands-on engineer who designs, builds, and operates the integrations and first-party applications that power OpenAI’s procurement workflows. You will translate business needs into secure, scalable software, APIs, data flows, and automation across Oracle Fusion, Zip, and connected platforms. You will build the future of buying at OpenAI using OpenAI’s own technology, from guided intake and approval experiences to supplier onboarding, purchasing, receiving, invoicing, and downstream financial data flows. You will own the technical roadmap and support model for these capabilities, improving today’s platforms while deciding where to integrate, configure, or build as OpenAI scales. Your core strength will be software and integration engineering. You will personally write code, troubleshoot cross-system failures, and take solutions through testing, deployment, and production support. You will also make targeted functional configurations in procurement platforms and partner with functional specialists on deeper process and module design. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, build, and operate integrations across Oracle Fusion, Zip, and connected systems using APIs, events, messaging, and batch interfaces where appropriate. Build first-party

REMOTEtypescriptpythonsql
View job →

Location Details: India, Remote At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team... Here at GoDaddy, the ML Engineering (MLE) team exists as the backbone of our machine learning infrastructure, enabling ML scientists and product teams across Domains to ship models to production reliably, efficiently, and at scale. This team owns the full lifecycle of ML systems — from CI/CD pipelines and model serving infrastructure to GPU workload orchestration and observability. Through disciplined engineering practices, thoughtful system design, and close collaboration with ML scientists, data engineers, and product teams, we deliver the platform that powers domain search, pricing, recommendations, and emerging AI experiences for millions of customers worldwide. We are currently looking for an experienced, highly motivated Senior Engineering Manager to lead our ML Engineering team based in India. This is an established team with existing engineers — we expect the candidate to ramp up quickly on our ML infrastructure stack, build strong relationships with the team, and partner with both India-based teams and US-based teams to drive execution and grow the team further. This individual will join us on our journey to build and scale ML infrastructure that serves real-time predictions at low latency, automates model deployment and promotion, and provides the observability and reliability guarantees that production ML systems demand. Become part of a team that bridges the gap between ML research and production engineering — shipping systems that directly impact GoDaddy's core revenue. What you'll get to do... Lead a team o

typescriptpythonaws
View job →
C
Cohere
📍 San Francisco• Full-time• Remote
22 days ago

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Role Overview Ready to build AI agents that revolutionize enterprise security? The North Security team is looking for a Senior Software Engineer to create autonomous systems that handle the complex, critical tasks security practitioners face every day. Key Responsibilities As a Senior Software Engineer focused on agentic security, you'll build intelligent systems that empower security teams. Your responsibilities will include: Build autonomous security agents that perform alert triage, secure code reviews, threat modeling, and vulnerability assessment Develop agent orchestration systems that help security practitioners automate repetitive security tasks Design and implement agent security controls to ensure our AI agents can safely interact with sensitive enterprise data Create the security infrastructure that allows agents to operate in high-trust environments with stringent deployment requirements Ship production-ready agent capabilities that run efficiently in resource-constrained environments Research and implement novel security patterns for AI agent systems, sometimes requiring custom solutions where standard libraries fal

REMOTEpythongitai
View job →
O
OpenAI
📍 San Francisco• Full-time• Remote
24 days ago

About the Team OpenAI’s mission is to ensure that artificial general intelligence benefits all of humanity. Our Communications team supports that mission by explaining our technology, our values, and how we safely build powerful AI. We help business leaders see what AI can make possible for their organizations, show how customers use OpenAI’s products, and explain how AI is changing the way people work. Analyst Relations sits within Business Communications, helping industry analysts understand our technology, enterprise products, and how organizations put them to work. About the Role We are looking for a Head of Analyst Relations to build OpenAI’s global analyst relations program and help influential analysts understand our enterprise products and the business outcomes they enable. You will set the narrative and engagement strategy, focusing on the firms and categories that shape enterprise buying decisions. This is an exceptionally fast-paced, hands-on role. You will shape long-term positioning, advise senior leaders, and lead our participation in priority evaluations, briefings, and analyst engagement around launches. You will build trusted relationships while bringing Product, Research, Marketing, Sales, and Communications together to make decisions quickly and deliver exceptional work under tight deadlines. This role reports to the Head of Business Communications. This role will be hybrid, in office 3 days a week, ideally in the San Francisco Bay Area; flexible on New York based candidates. In this role, you will: Set the global analyst relations strategy, prioritizing the firms, categories, and questions that shape enterprise technology decisions. Lead OpenAI’s work on priority evaluations, from deciding where to engage through submissions, demos, and customer evidence. Build a coherent enterprise narrative across a fast-moving product portfolio, grounded in what ships, customer outcomes, and responsible deployment. Develop trusted, candid relationships with an

REMOTEawsrestai
View job →
B
Baseten
📍 San Francisco• Full-time• Remote
25 days ago

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE As a Global Capacity Manager focused on TPUs at Baseten, you will lead the "engine room" for our non-NVIDIA accelerator fleet, architecting, securing, and optimizing the Google Cloud TPU (and broader emerging accelerator) capacity that powers our customers' AI workloads. You'll own the end-to-end journey of capacity management for this fleet, from securing large-scale TPU pod allocations to building the automation that ensures reliable uptime across multi-cloud environments. This role is a great fit for entrepreneurial engineers who want to bridge the gap between high-finance asset management and deep infrastructure engineering, with a specific focus on the TPU ecosystem. You will act as the fleet orchestrator for Google's TPU architecture, ensuring Baseten never experiences a capacity outage while maintaining elite unit economics as we diversify beyond NVIDIA. To be clear, this is a high-stakes engineering role. You will be hands-on with Kubernetes orchestration while also leading specialized pods focused on the latest generation of TPU hardware, like Google's Trillium (v6e) architecture, and partnering closely with the Model Performance (MP) team to ensure workloads are tuned for TPU-specific execution. EXAMPLE INITIATIVES The TPU Frontier: Architecting the infrastructure readiness and deployment strategy for Baseten's TPU clusters, including pod slicing and topology planning Global Workload Orchestration: Bui

REMOTEpythonawsazure
View job →
B
Baseten
📍 San Francisco• Full-time• Remote
25 days ago

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE As a member of the Capacity Strategy & Operations team, you will sit at the intersection of supply intelligence, demand forecasting, and cross-functional execution, turning a complex, fast-moving hardware market into a predictable, reliable foundation for our customers and internal engineering teams. This is not a purely analytical role. You will own the end-to-end capacity planning process: from translating customer commitments and growth forecasts into concrete supply requirements, to coordinating fulfillment across vendors, finance, and the infrastructure team, to building the systems that make all of this repeatable and scalable. When supply is constrained and tradeoffs are unavoidable, you are the person in the room who can model the options, make a clear recommendation, and drive alignment fast. You are a strong fit if you have operated at the intersection of strategy and execution before — someone who is equally comfortable building a capacity model in a spreadsheet and running a cross-functional war room when a customer deployment is at risk. EXAMPLE INITIATIVES Demand-Supply Alignment Framework: Build and own the process that translates customer pipeline, signed commitments, and growth projections into a forward-looking GPU demand signal — so the team is never caught flat-footed when a customer scales faster than expected. Constrained Allocation Playbook: Define the decision framework for how Basete

REMOTEmachine learningaigo
View job →
L
Lyft
📍 Toronto• Full-time• From C$45/hr
25 days ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. Lyft’s Data Science Team builds mathematical models underpinning the platform’s core services. Compared to other technology companies of a similar size, the set of problems that we tackle is incredibly diverse. They cut across optimization, prediction, modeling, inference, transportation, and mapping. We're looking for Masters or PhD students who are passionate about solving mathematical problems with data and are excited about working in a fast-paced, innovative and collegial environment. We are hiring for a variety of Data Science interns, focusing on the following specialties: Optimization: Construct and fit statistical or optimization models that facilitate automated decision making in the app. Machine Learning: Design, build, tune, and deploy machine learning models with a special emphasis on feature engineering and deployment. Inference: Design and analyze tests in our dynamic marketplace, estimating statistical and ML models to enable better decisions, and developing and evaluating algorithmic policies in our pricing, dispatch, and incentives systems. You will report into a Science Manager. Responsibilities: Partner with Engineers, Product Managers, and other cross-functional partners to frame problems, both mathematically and within the business context Perform exploratory data analysis to gain a deeper understanding of the problem Write production modeling code; collaborate with software engineers to implement algorithms in production Design and run both simulated and live traffic experiments Analyze experimental and observational data; communicate findings including working with partner teams and presentations; facilitate launch decisions Experience: Currently pursuing a Masters or PhD degree at a university in Canada (required) in mathematical sciences ( Opera

pythonsqlmachine learning
View job →
O
Okta
📍 Bengaluru• Full-time
25 days ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. We are looking for a high-impact Senior Software Engineer in Test (Sr. SET) to join our QA Core Team. In this role, you will partner closely with Infrastructure, Release Engineering, and Development teams to drive high availability, build scalable test automation, and optimize CI/CD release pipelines. You will take ownership of release-based testing, system observability, and automation services to ensure that weekly software releases are delivered seamlessly and with high quality across mission-critical systems. Key ResponsibilitiesTest Automation & System Quality Design, implement, and maintain scalable test automation frameworks for both server- and client-side systems using Java, TestNG, JUnit, and Selenium. Develop and execute comprehensive end-to-end and system test plans to validate new software changes against existing customer deployment patterns. Set up and manage automated test configurations, lab environments, and containerized test setups using Docker. Perform release-based validation, ensuring full compatibility, high availability, and uptime across weekly deployment cycles. CI/CD Pipeline Maintenance & Optimization Troubleshoot, debug, and optimize CI/CD pipelines (Jenkins, GitHub Actions, CircleCI, Buildkite) to reduce test execution times and eliminate flaky builds. Manage and optimize Maven-based build structures, dependency management, and automated release workflows. Write operational automation scripts using Python, Bash, or Gro

pythonjavasql
View job →
L
Lyft
📍 Montreal• Full-time• C$34 – C$36/hr
25 days ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. As a Test Automation Intern at Lyft, you'll collaborate closely with industry-leading engineers while enjoying the freedom to innovate from day one. Your contributions will be pivotal in accelerating our product development and enhancing user trust and satisfaction through the creation of cutting-edge test automation tools. At Lyft, we cultivate a dynamic and collaborative office environment, where brilliant minds are always eager to hear and support your next big idea. What will yours be? Responsibilities: Own your project, while checking in with other team members throughout the day with questions and updates You leave the code in a better state than when you found it (progressive refactor) You value reliability, ensured by testing (unit, integration and load tests) Participate in code reviews to ensure code quality and distribute knowledge Continuous integration and deployment Go home knowing that your work today is meaningfully improving the quality of every Lyft Urban Solutions rider! Experience: Currently pursuing a Bachelor's or Master's degree in Computer Science from a university in Canada (required) , with a graduation date between December 2027 and Summer 2028 (required) . For any candidates who are master's students who worked between their bachelor's and master's programs: candidates should also have less than 2 years of relevant full-time work experience Available during Summer 2027 for the internship in Montreal Strong knowledge of CS fundamentals Excellent communication skills Passion for community, sustainability, and/or transportation Passion for quality and testing Ability to thrive in a startup environment Contributions to open source projects Experience working with databases Experience solving real-time technology problems Experience with mobile development Experience wit

aigorust
View job →
G
Godaddy
📍 India• Full-time
26 days ago

Location Details: India, Remote At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join our team Contribute to the development of GoDaddy’s eCommerce and SSO infrastructure and Kubernetes systems on AWS. On a day-to-day basis you will be working on the team who designs, writes, tests and deploys the infrastructure and application management software for GoDaddy’s eCommerce applications. Expect to learn every day. What you'll get to do... Work as a polyglot engineer, writing and maintaining Infrastructure as code with frameworks/ ecosystems such as Java, Unix CLI, and NodeJS Build and operate infrastructure workflows and deployment pipelines using Kubernetes, Argo Workflows, Argo CD, and GitOps practices Design, build, and own services and APIs in Java, running on Kubernetes-based platforms across AWS and distributed systems Develop and support application and infrastructure delivery pipelines, enabling reliable releases of eComm, Auth and Infrastructure services Collaborate closely with other GoDaddy departments to help advance security and technical standards, maintain regulatory compliances while operating eComm & Auth platforms Your experience should include... 5+ years of strong backend software engineering experience in Java Hands-on experience with Kubernetes, including Helm, Kustomize, or equivalent tools to deploy and manage backend services Experience building and operating high-volume, mission-critical production systems on AWS with continuous deployment (CD) practices Strong experience with infrastructure as code, supporting backend applications and services Experience with observability and l

javanodejssql
View job →
L
Lyft
📍 Toronto• Full-time• From C$108K/yr
26 days ago

At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. We are looking for experienced backend software engineers to join our claims tech engineering team. Our vision is to tangibly reduce risk on the Lyft platform, and by extension reduce insurance cost. Our team is dedicated to centralizing the entire claims operation onto a unified risk platform. This consolidation of data, workflows, and communications aims to foster proactive measures, enhance efficiency, ensure consistency, and provide valuable insights. These efforts are designed to effectively reduce claims costs as Lyft's operations expand. Additionally, our team is responsible for maintaining robust relationships with our third-party insurance partners, guaranteeing timely, proactive, and precise sharing of claim data. Responsibilities: Write well-crafted, well-tested, readable, maintainable code Own feature from product spec to successful high quality development, deployment and maintenance Participate in code reviews to ensure code quality and distribute knowledge Respond to external questions and requests. Unblock, support and communicate with stakeholders to achieve results Experience: 3+ years of relevant professional experience Experience with object-oriented programming Experience in distributed systems Experience working with databases, relational or NoSQL Write clear, scalable and clear design documentation Design, build and improve a set of team owned components Benefits: Extended health and dental coverage options, along with life insurance and disability benefits Mental health benefits Family building benefits Child care and pet benefits Access to a Lyft funded Health Care Savings Account RRSP plan with company match to help save for your future In addition to provincial observed holidays, salaried team members are covered under Lyft's flexible paid time off policy. The policy allows

O
27 days ago

About the Team OpenAI, in partnership with our capital and technology partners, is building a global network of advanced datacenters to support the most demanding AI workloads. The Industrial Compute team ensures that all datacenter systems are manufactured, delivered, and commissioned to the highest standards of quality, reliability, and performance. We work closely with manufacturing partners, engineering teams, and operations staff to ensure that every component is delivered ready for installation, startup, and long-term service. About the Role We are seeking an experienced Quality Engineer (QE) to drive Product and Site Quality initiatives across OpenAI’s infrastructure ecosystem. In this role, you will establish, implement, and manage a comprehensive, quality-focused program across our global supply chain network, ensuring excellence from design through deployment. You will be responsible for end-to-end quality of finished products, as well as maintaining and elevating manufacturing site quality standards. Working cross-functionally with Design (NPI) and Engineering teams, you will help achieve First Pass Yield (FPY), quality, and reliability targets. This includes leading site and fixture validation efforts, driving yield improvement initiatives (Yield Bridge, CPI), and implementing robust corrective and preventive actions (CAPA) to resolve issues at their root cause. In addition, you will play a key role in supplier quality management, assessing and qualifying new vendors, overseeing ongoing supplier performance, and ensuring readiness for future business awards. You will lead vendor audits, monitor key performance metrics, and coordinate corrective actions to ensure predictable delivery schedules, reduced operational risk, and high system reliability. By partnering closely with external suppliers and internal Engineering and Operations stakeholders, you will help ensure OpenAI’s datacenter infrastructure is delivered on time, meets the highest quality standa

awsrestai
View job →
O
OpenAI
📍 San Francisco• Full-time• Remote
27 days ago

About the Team OpenAI’s Forward Deployed Engineering team partners with leading semiconductor companies to deploy production-grade AI systems across the entire chip design lifecycle: design, verification, and physical design. We operate at the intersection of customer delivery and core platform development, embedding deeply with customers to translate frontier model capabilities into systems that materially improve engineering workflows and accelerate innovation. Our work turns early, high-touch deployments into repeatable solution patterns, reference architectures, and evaluation practices that scale across the semiconductor ecosystem. About the Role We are seeking a highly skilled Physical Design Engineer to join our semiconductor-focused Forward Deployed Engineering team. This is a senior IC role that will begin with a strong emphasis on physical design expertise, technical judgment, advisory leverage, and customer credibility, with the expectation that the person will grow into a broader Forward Deployed Engineering role over time. In the near term, you will serve as the team’s physical design SME across semiconductor deployments: helping FDEs, Product, and Research understand backend implementation workflows, pressure-test AI-assisted solution ideas against real physical design constraints, and raise the quality of our customer-facing technical work. You will help the broader team build fluency in implementation flows, EDA tooling, signoff methodology, and the trade-offs that shape physical design decisions in practice. Over time, we expect this role to expand beyond SME support into broader FDE ownership: partnering directly with customers, shaping deployment strategy, building and iterating production-grade AI systems, driving technical workstreams, and helping turn high-touch semiconductor deployments into repeatable solutions. This is a strong fit for someone who brings deep physical design expertise today and is excited to grow into a customer-facing, syst

REMOTEpythonawsrest
View job →
B
Baseten
📍 San Francisco• Full-time• Remote
27 days ago

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE As a Global Capacity Lead at Baseten, you will lead the "engine room" of the company, architecting, securing, and optimizing the global GPU fleet that powers our customers' AI workloads. You’ll own the end-to-end journey of capacity management, from securing multi-million dollar GPU clusters to building the automation that ensures 99.9% uptime across multi-cloud environments. This role is a great fit for entrepreneurial engineers who want to bridge the gap between high-finance asset management and deep infrastructure engineering. You will act as the fleet orchestrator for the world's most advanced chips, ensuring Baseten never experiences a capacity outage while maintaining elite unit economics. To be clear, this is a high-stakes engineering role. You will be hands-on with Kubernetes orchestration while also leading specialized pods focused on the next generation of hardware, like NVIDIA’s Blackwell (B200) architecture. EXAMPLE INITIATIVES The B200 Frontier: Architecting the infrastructure readiness and deployment strategy for Baseten's first Blackwell GPU clusters. Global Workload Orchestration: Building "Multi-cloud Capacity Management" systems to move customer workloads seamlessly across regions to optimize cost and latency. Precision GPU Triage: Developing automated Go-based operators to identify, cordon, and repair unhealthy H100 nodes in under an hour. The Supply Chain of Intelligence: Partnering with lead

REMOTEpythonawsazure
View job →
O
28 days ago

About the Team OpenAI's data and storage infrastructure spans data platforms, online databases, and file/object storage. These systems underpin data ingestion and processing, durable persistence, indexing and retrieval, and product file experiences. As frontier models and agents evolve how they use memory, history and snapshots, the underlying architecture increasingly shapes the capabilities products can deliver—and their latency, reliability, cost and efficiency. About the Role We are looking for a technically deep TPM to independently define and lead multiple programs across data platforms, online databases and storage infrastructure. You will connect model, product and data-consumer requirements to architecture, and work with the relevant engineering teams to take new capabilities through production adoption and repeatable expansion. The design scope is exabyte-scale storage and infrastructure spanning multiple millions of CPU cores. The challenge is not simply forecasting more resources: it is making complete, workload-ready capacity repeatable, with a clear path from product requirements through architecture, deployment and validation. A data pipeline, database query, file operation or execution snapshot can affect whether a product or agent succeeds; you will connect those outcomes to the systems underneath. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Translate model, product and data-platform needs into precise access patterns, consistency, durability, freshness, availability and scalability requirements. Connect memory, history, retrieval and resumable work to capability and end-to-end latency. Partner with engineering to transform data and storage architecture into repeatable scale units: standardized provisioning, placement, routing, data movement and readiness checks that bring storage, compute and networking online together.

awsazurerest
View job →
🔔

Get new deployment strategist lead jobs by email

Daily job updates · Unsubscribe anytime

Explore verified demand

More deployment strategist lead opportunities

Browse all jobs →

Companies hiring

Employers are derived from current jobs in this exact search market.