Jobiba hiring network

Platform Deployment Management Lead Jobs

10,000 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current platform deployment management lead jobs. Use filters to narrow by work mode, employment type, experience and date posted.

B
Baseten
📍 San Francisco• Full-time
1mo ago

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As an Infrastructure Software Engineer at Baseten, you'll build and maintain components of our ML inference platform that powers production AI applications. You'll contribute to the core infrastructure, enabling developers to deploy, scale, and monitor ML models with high performance. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Infrastructure team: Multi-cloud capacity management Inference on B200 GPUs Multi-node inference Fractional H100 GPUs for efficient model serving RESPONSIBILITIES Develop infrastructure components for our ML inference platform using Python and Go Implement and maintain Kubernetes deployments for model serving Contribute to our inference orchestration layer for model deployments Build and enhance monitoring systems for model performance metrics Implement efficient resource management solutions for ML workloads Support infrastructure automation to improve ML deployment workflows Work closely with team members to implement technical solutions Help balance performance optimization with system reliability Participate in technical discussions around infrastructure improvements Learn and apply infrastructure best practices REQUIREMENTS Bachelor's degree or higher in Computer Science or related field Proficient coding abilities in one or more popular programming or scripting languages; Go proficiency is a plus Working knowledge of Kubernetes and containeriza

pythonkubernetesrest
View job →
B
Baseten
📍 San Francisco• Full-time
1mo ago

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Forward Deployed Engineer at Baseten, you will partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten’s platform. You’ll own the journey with customers from initial exploration to production deployment, translating ambiguous business goals into reliable, observable services with clear quality, latency, and cost outcomes. This role is a great fit for entrepreneurial engineers who want a front-row view into how modern companies adopt AI at scale and who enjoy working across product, software development, performance engineering, and customer-facing implementations. To be clear, this is an engineering role with hands-on coding and software development that also includes aspects of product management, technical customer success, and pre-sales solution engineering mixed in. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Forward Deployed Engineering team: Forward Deployed Engineering on the frontier of AI The fastest, most accurate Whisper transcription Deploy production-ready model servers from Docker images Deploy custom ComfyUI workflows as APIs RESPONSIBILITIES Develop and maintain software systems and product features using one or more general-purpose programming languages in a production-level environment, with a preference for Python due to its relevance in ML projects. Drive customer impact by designing, implementin

pythondockermachine learning
View job →
NR
New Relic
📍 Hyderabad• Full-time
1mo ago

We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! About the Team The NR Lens team builds New Relic's Federated Data Platform — a distributed SQL query engine that lets customers connect external data sources (Snowflake, PostgreSQL, Google Sheets, AWS CloudWatch) and query them directly from New Relic. You'll work on the distributed query execution layer, connector architecture, and API gateway that powers cross-source JOINs and unified analytics across customer data stores. What You'll Do Design, build, and maintain cloud-native Java microservices in the NR Lens query path: SQL Gateway, Query Gateway, and data source connector plugins Develop and harden JDBC connector integrations — including connection lifecycle management, credential handling, query pushdown optimization, and security validation Improve query reliability and performance across a multi-tenant distributed SQL deployment serving 50+ customers Build and operate services on AWS (EKS, IAM/STS, S3) with infrastructure-as-code practices Implement

javasqlpostgresql
View job →
P
Pendo
📍 Raleigh• Full-time• $105K – $130K/yr
1mo ago

The team + the role Pendo's Applied AI team turns AI infrastructure into real business outcomes across Sales, Marketing, and Customer Engineering. AI here isn't a feature we bolt on — it's how we scale GTM capability across the company. We build AI that makes GTM teams measurably faster and more effective, and we measure success by whether those teams are actually using what we ship and getting real value from it. As an Applied AI Engineer, you'll own the full lifecycle of AI-powered solutions — from problem definition and prompt design through to production deployment and ongoing iteration. You'll work directly with GTM stakeholders to identify high-value problems, set realistic expectations about what AI can and can't do, and ship solutions that stick. The best person for this role has a strong engineering instinct, deep curiosity about how businesses operate, and the judgment to know when AI is the right tool — and when it isn't. This role is based in Raleigh, NC and follows Pendo's hybrid model: in-office 3 days per week. What this looks like day-to-day Own AI solutions end-to-end: problem definition, prompt design, production deployment, monitoring, and iteration — you ship, you watch, you improve. Build and manage GTM workflow automations that reduce manual work across Sales, Marketing, and Customer Engineering systems, including Slackbots and other integrations. Partner directly with GTM stakeholders to surface high-value problems, validate solutions, drive adoption, and set honest expectations about AI capabilities and limitations. Develop and maintain prompt management practices that make AI outputs reliable, auditable, and improvable over time — treat prompts as production code, not experiments. Work with the Data Platform and Systems teams to identify foundational tooling gaps and contribute clear, actionable requirements based on what you encounter in production. Share reusable AI workflow patterns and tool findings with the broader team — your impact sh

O
OpenAI
📍 Seattle• Full-time
1mo ago

About the Team The Monetization team is a new cross-functional group working across engineering, product, research, and design to build the foundational systems that will help OpenAI scale access to intelligence responsibly. Our mission is to develop user-first, privacy-preserving monetization products, including next-generation ads experiences, that strengthen user trust, unlock economic opportunity, and support OpenAI’s long-term innovation. Monetization plays a critical role in enabling OpenAI to continue pushing the boundaries of AI capabilities while ensuring the benefits of AGI are broadly shared. We believe monetization must be aligned with user value, uphold rigorous privacy and safety standards, and sustain a healthy ecosystem of developers, advertisers, and businesses. This team operates in a greenfield environment and moves quickly through prototyping, experimentation, and iterative deployment. We partner closely with Product, Design, Research, and external customers to bring new monetization products into real-world systems at global scale. About the Role We’re looking for an experienced Software Engineer to help build Ads Manager, the UI platform advertisers use to create, manage, measure, and optimize ad campaigns across OpenAI’s ads ecosystem. This is a foundational role responsible for designing and implementing advertiser-facing products, APIs, tools, and services that connect external customers to OpenAI’s next-generation monetization products. You’ll work across the full technical stack to build intuitive self-serve workflows for small and mid-sized advertisers, as well as scalable APIs and integrations for large enterprise advertisers, agencies, and ad-tech partners who manage campaigns through their own buying platforms or intermediary systems. This includes building advertiser-facing APIs and tooling for campaign management, conversion APIs, pixels, measurement, and insights. You will collaborate deeply with Product, Design, Research, and Go-To

typescriptpythonreact
View job →
O
1mo ago

Join the engineering teams that bring OpenAI’s ideas safely to the world!! The Applied Engineering team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role As OpenAI continues to grow, we are looking for experienced, problem-solving engineers to ensure our systems scale. Our success depends on our ability to quickly iterate on products while also ensuring that they are performant and reliable. You will work in a deeply iterative, collaborative, fast-paced environment to bring our technology to millions of users around the world, and ensure it’s delivered with safety and reliability in mind. Successful candidates will play a crucial role in ensuring the reliability, scalability, and performance of our systems as we continue to expand. As a reliability expert, you will be at the forefront of maintaining and enhancing the stability, scalability, and performance of our rapidly evolving infrastructure. You will work closely with cross-functional teams, including software engineers, product managers, and data scientists, to build and maintain resilient systems that can handle our growing user base and workload. In this role, you will: Design and implement solutions to ensure the scalability of our infrastructure to meet rapidly increasing demands. Build and maintain the load, chaos and synthetic testing software leveraged by development teams to make the systems they design and operate more reliable. Build and maintain automation tools to streamline repetitive tasks and improve system reliability. Build and maintain the platform for CPU/storage, GPU, and network lifecycle management to drive efficiency, accountability and support dynamic optimization of our resources. Implement fault-tolerant and resilient

awskubernetesrest
View job →
W-
17 days ago

About Wolt At Wolt, we create technology that brings joy, simplicity and earnings to the neighborhoods of the world. In 2014 we started with delivery of restaurant food. Now we’re building the delivery of (almost) everything and you’ll find us in over 500 cities in 30 countries around the world. In 2022 we joined forces with DoorDash and together we keep on dreaming big and expanding across the globe. Working at Wolt isn’t always easy, but it’s definitely exciting. Here you’ll learn more, build more, and ship more than in most other companies. You’ll be challenged a lot, but also have a lot of fun on the way. So, if you’re a self-starter with drive and entrepreneurial spirit, this could be the ride of your life. Who We Are Global Marts builds the technology behind DoorDash's own grocery and retail operations across DashMart, Wolt Market, and Deliveroo Hop. Our systems decide what products our stores stock, how much inventory to purchase, which vendors to source from, how warehouses execute orders, and what prices customers see. We build everything from forecasting and replenishment systems to vendor management, warehouse execution, labor scheduling, space optimization, and pricing. Our in-house forecasting platform has replaced third-party solutions and now powers production globally. This isn't a traditional CRUD backend environment. Much of our work involves designing new systems that solve complex operational problems at global scale. What You'll Do As a Senior Backend Engineer, you'll work primarily with Go or Kotlin and take end-to-end ownership of production systems and complex technical initiatives. You'll: Design, build, and operate backend systems that support inventory, purchasing, warehouse operations, pricing, and fulfillment. Own technical solutions from problem definition and design through implementation, deployment, monitoring, and production operation. Work on data-heavy systems where software engineering meets forecasting, optimization, and

sqlpostgresqlmongodb
View job →
P
Point72
📍 Bengaluru• Full-time
17 days ago

JOB TITLE Cloud Compute Engineer A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology team is constantly evolving our firm’s IT infrastructure and engineering capabilities, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts who experiment and work to discover new ways to harness open-source solutions, modern cloud architectures, and sophisticated Artificial Intelligence (AI) solutions, while embracing enterprise agile methodologies. Our commitment to building and innovating in the AI space provides the framework intended to drive smarter decision making and enhance how we build and operate our platforms and applications. As a member of Point72’s Technology team, we encourage and support your professional development from day one—helping you advance your technical skills, contribute innovative ideas, and satisfy your own intellectual curiosity—all while delivering real business impact for our multi-billion-dollar global business. WHAT YOU'LL DO Design, build, and operate Kubernetes clusters on Amazon EKS, including cluster lifecycle management, networking, autoscaling, and workload scheduling. Manage and optimize EC2-based compute infrastructure, including instance selection, placement strategies, capacity planning, and utilization analysis. Operate and improve ECS-based services where applicable, ensuring consistency across our container runtime environments. Develop and maintain Infrastructure as Code (IaC) using Terraform to provision and manage compute resources at scale. Collaborate with development and platform teams to define compute patterns, containerization standards, and deployment best practices. Monitor compute environments for availability, performance, and cost, driving continuous optimization across the fleet. Contribute to architectural decisions around workload placement, multi-tenancy, OS image and container li

pythonawsazure
View job →
P
Point72
📍 Bengaluru• Full-time
17 days ago

JOB TITLE Cloud Engineer A Career with point72’s technology team As Point72 reimagines the future of investing, our Technology team is constantly evolving our firm’s IT infrastructure and engineering capabilities, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts who experiment and work to discover new ways to harness open-source solutions, modern cloud architectures, and sophisticated Artificial Intelligence (AI) solutions, while embracing enterprise agile methodologies. Our commitment to building and innovating in the AI space provides the framework intended to drive smarter decision making and enhance how we build and operate our platforms and applications. As a member of Point72’s Technology team, we encourage and support your professional development from day one—helping you advance your technical skills, contribute innovative ideas, and satisfy your own intellectual curiosity—all while delivering real business impact for our multi-billion-dollar global business. What You'll Do Design, implement, and maintain scalable cloud infrastructure using Infrastructure as Code (IaC) with Terraform. Manage and optimize foundational platform for AWS/GCP/Azure services including IAM, Virtual Private Cloud networking, Cloud Security controls and policies, Cloud platform tools and automation to ensure scale, performance and security. Collaborate with development and operations teams to automate deployment and operational processes. Monitor cloud environments to maintain high availability and security standards. Participate in architectural discussions and contribute to the technical roadmap for cloud solutions. Implement best practices for cloud resource optimization and cost management. What's Required Bachelor’s degree in Computer Science, Engineering, or a related field. 5 to 10 years of experience in cloud infrastructure design and automation with optional experience in cloud networking. Proficiency in Infrastructur

pythonawsazure
View job →
O
Okta
📍 Chicago• Full-time• From $168K/yr
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Customer First Team We are a team of Identity and Access Management experts. We help customers, businesses, and organisations maximise their Okta IAM platform investment, providing rapid time-to-value in adoption and usage and promoting scalability and agility. Our team sits within the Customer First area - focused on ensuring the long term success of our customers. Join our team! We’re building a world where Identity belongs to you. Role Summary The Business Analyst (BA) is a central figure on the Okta Professional Services team. Your mission is to serve as the bridge between customer business objectives and technical implementation, driving high-impact outcomes for Okta Identity and Access Management (IAM) deployments. You will own the requirements lifecycle—from initial discovery and stakeholder alignment to functional design and solution integrity—ensuring that the final delivery is secure, scalable, and aligned with the customer's strategic needs. Key Responsibilities Discovery & Requirements Facilitation: Facilitate stakeholder workshops brainstorming sessions to uncover requirements, map current-state IAM processes, and identify dependencies and pain points across IT, security, and product teams. Documentation & Solution Design: Translate business needs into detailed functional and non-functional requirements. Produce clear business process flows and solution design diagrams to visualize the "how" and "why" behind the technical conf

awsrestmachine learning
View job →
O
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Business Analyst Chicago, Illinois Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. The Customer First Team We are a team of Identity and Access Management experts. We help customers, businesses, and organizations maximize their Okta IAM platform investment, providing rapid time-to-value in adoption and usage and promoting scalability and agility. Our team sits within the Customer First area - focused on ensuring the long term success of our customers. Join our team! We’re building a world where Identity belongs to you. Role Summary The Business Analyst (BA) is a central figure on the Okta Professional Services team. Your mission is to serve as the bridge between customer business objectives and technical implementation, driving high-impact outcomes for Okta Identity and Access Management (IAM) deployments. You will own the requirements lifecycle—from initial discovery and stakeholder alignment to functional design and solution integrity—ensuring that the final delivery is secure, scalable,

machine learningartificial intelligenceai
View job →
O
OpenAI
📍 San Francisco• Full-time• Remote
1mo ago

About the Team: OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. In this role you will: As a Hardware Test Engineer, you will work on Machine Learning/AI hardware system projects to craft the solutions for current and future data center deployments. You will bring a strong understanding of hardware system testing, excellent project management skills, and the ability to collaborate across multiple teams to ensure efficient lab operations. You will be responsible for designing, implementing, and executing comprehensive test plans that ensure the reliability, performance, and scalability of our supercomputing hardware systems. You will develop detailed test plans and methodologies tailored to hardware components, including processors, memory modules, custom accelerators and interconnects. You will collaborate with hardware design, manufacturing, firmware teams and vendors to identify, analyze, and resolve issues affecting hardware, power, thermal and high-speed interconnects. You will perform in-depth debugging on the hardware system Excellent analytical skills to diagnose hardware issues, troubleshoot problems, and propose solutions. Ability to interpret complex test data, identify trends, and draw meaningful conclusions. High-speed links, with a focus on SerDes (Serializer/Deserializer) technology to assess signal integrity, error rates, and overall link performance. You will collaborate with the lab manager to maintain the equipment and hardware systems, including oscilloscopes, thermal test chambers, liquid cooling systems, and other mea

REMOTEpythonawsrest
View job →

Our Purpose Mastercard powers economies and empowers people in 200+ countries and territories worldwide. Together with our customers, we’re helping build a sustainable economy where everyone can prosper. We support a wide range of digital payments choices, making transactions secure, simple, smart and accessible. Our technology and innovation, partnerships and networks combine to deliver a unique set of products and services that help people, businesses and governments realize their greatest potential. Title and Summary Technical Project Manager (Product Delivery) Overview – About Product Delivery Product Delivery (PD) is dedicated to enabling and empowering the core of Customer Delivery throughout the Asia Pacific region by providing streamlined knowledge, expertise, materials, training, and education. By leveraging our robust global partnerships and subject-matter expertise, we offer insights, experience, and solutions to foster innovation, ensuring our products and platforms are prepared for large-scale deployment. We excel in addressing uncertainties in delivering pilot projects and executing complex and strategic programs. Our team is inclusive, supportive, and innovative, fostering a problem-solving environment. Our team culture emphasizes accomplishing tasks efficiently while maintaining a positive and enjoyable work atmosphere. Product Delivery Program & Project Management Team The Product Delivery Program & Project Management (PPM) Team in Asia Pacific is a technical project team with strong business acumen. The team oversees, supports, and provides subject-matter expertise for the deployment of Mastercard products, including technical integration of product APIs. • Partner with Product Management, Delivery, Business, and other internal cross-functional teams to align product

aiExcelproject management
View job →
O
1mo ago

Location: San Francisco, CA (Hybrid: 4 days onsite/week). Relocation assistance available. About the Team: We build foundational platform software that enables reliable, secure, and performant products. The team works across system layers and partners closely with adjacent engineering groups to deliver robust capabilities from concept through launch. About the Role: We’re seeking a System Software Engineer to design, implement, and debug core platform components and the pipelines that build and update system images. You’ll work across operating system layers, focusing on performance, security, and deep system debugging to ship production‑grade systems. In this role, you will: Design, implement, and debug system‑level components and services across kernel and user space. Configure and maintain OS platform services (init, services, networking, security policies) and related tooling. Build and operate image and update pipelines, ensuring reliability, reproducibility, and rollback safety. Instrument and analyze performance using profiling and tracing; optimize CPU, memory, I/O, and power usage. Own platform observability and reliability: logging, crash capture, watchdogs, and diagnostics. Collaborate with cross‑functional teams to define interfaces and deliver end‑to‑end features. Establish strong engineering practices: code review, CI, reproducible builds, and release management. Partner with external suppliers to support builds and deployments. You might thrive in this role if you: Have shipped production systems software on modern operating systems. Are proficient in C/C++ and a scripting language, and comfortable with OS internals (concurrency, memory management, filesystems, networking, power management). Bring strong systems debugging skills using debuggers, tracers, profilers, and logs across kernel/user‑space boundaries. Understand configuration of platform services and interfaces, and can translate requirements into stable, well‑documented APIs. Are fluent in u

awsrestai
View job →
E
Everpure
📍 Bengaluru• Full-time
17 days ago

Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE The FlashBlade team builds the industry’s most innovative, high-performance, and highly available portfolio of products that are designed for the most demanding mission critical applications. While we deliver a hardware storage array, over 90% of our engineering staff are software engineers. Our customers are the most important part of our business and they love FlashBlade for its simplicity of management, the constant flow of new and exciting upgrades, and ability to live on the cutting edge of technology while never taking downtime, ever. FlashBlade enables our customers to leverage the agility of the public cloud for both traditional IT and cloud-native applications. You’ll architect, implement, and optimize core services that span on-premises arrays and cloud deployments—delivering seamless storage experiences to enterprises around the globe. If you thrive on solving hard problems, influencing technical direction, and driving innovation end-to-end, this is the role for you. WHAT YOU'LL DO Manage the full software development life cycle, from initial architecture and invention through development, release, and ongoing maintenance. Drive the technical strategy by influencing system design, architecture, and best practices, specifically contributing to the evolution of the FlashBlade Networking product area. Invent and optimize algorithms to orchestrate multi-array and multi-cloud storage systems, foc

awslinuxrest
View job →
🔔

Get new platform deployment management lead jobs by email

Daily job updates · Unsubscribe anytime