About the Team OpenAI's data and storage infrastructure spans data platforms, online databases, and file/object storage. These systems underpin data ingestion and processing, durable persistence, indexing and retrieval, and product file experiences. As frontier models and agents evolve how they use memory, history and snapshots, the underlying architecture increasingly shapes the capabilities products can deliver—and their latency, reliability, cost and efficiency. About the Role We are looking for a technically deep TPM to independently define and lead multiple programs across data platforms, online databases and storage infrastructure. You will connect model, product and data-consumer requirements to architecture, and work with the relevant engineering teams to take new capabilities through production adoption and repeatable expansion. The design scope is exabyte-scale storage and infrastructure spanning multiple millions of CPU cores. The challenge is not simply forecasting more resources: it is making complete, workload-ready capacity repeatable, with a clear path from product requirements through architecture, deployment and validation. A data pipeline, database query, file operation or execution snapshot can affect whether a product or agent succeeds; you will connect those outcomes to the systems underneath. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Translate model, product and data-platform needs into precise access patterns, consistency, durability, freshness, availability and scalability requirements. Connect memory, history, retrieval and resumable work to capability and end-to-end latency. Partner with engineering to transform data and storage architecture into repeatable scale units: standardized provisioning, placement, routing, data movement and readiness checks that bring storage, compute and networking online together.
Jobiba hiring network
Platform Deployment Management Lead Jobs
9,875 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current platform deployment management lead jobs. Use filters to narrow by work mode, employment type, experience and date posted.
For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. Our India Global Capability Center isn't just supporting global operations—we’re leading global innovation. After scaling rapidly into a best-in-class hub, we deliver the product innovation and enterprise capabilities that accelerate our global growth, profitability, and scale. As we expand Smartsheet India, we’re searching for Senior AI/ML Ops Engineers who crave variety and ownership. You’ll have the opportunity to work across multiple teams and disciplines, building a versatile skillset while solving the complex challenges of a global platform. You Will: Designing, Developing and overseeing the strategy and architecture of scalable and reliable AI/ML Ops platforms / pipelines Model Deployment: Package and deploy AI/ML services to production, ensuring they are reproducible and interpretable CI/CD Pipeline Development: Design and implement automated CI/CD (Continuous Integration/Continuous Deployment) pipelines to accelerate model deployment using tools Infrastructure Management: Provision and optimize infrastructure for training and serving, utilizing Docker, Kubernetes, or serverless platforms Monitoring & Observability : Implement post-deployment monitoring for model performance, data drift, and latency using tools. Experience in Monte Carlo is preferable Automation: Automate retraining and data pipeline workflows to ensure models stay accurate over time. Manage the deployment of foundation models, fine-tuning workflows, and Retrieval-Augmented Generation (RAG) stacks (Vector DBs, Knowledge Graph. Experience with
For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. Our India Global Capability Center isn't just supporting global operations—we’re leading global innovation. After scaling rapidly into a best-in-class hub, we deliver the product innovation and enterprise capabilities that accelerate our global growth, profitability, and scale. As we expand Smartsheet India, we’re searching for Senior AI/ML Ops Engineers who crave variety and ownership. You’ll have the opportunity to work across multiple teams and disciplines, building a versatile skillset while solving the complex challenges of a global platform. You Will: Designing, Developing and overseeing the strategy and architecture of scalable and reliable AI/ML Ops platforms / pipelines Model Deployment: Package and deploy AI/ML services to production, ensuring they are reproducible and interpretable CI/CD Pipeline Development: Design and implement automated CI/CD (Continuous Integration/Continuous Deployment) pipelines to accelerate model deployment using tools Infrastructure Management: Provision and optimize infrastructure for training and serving, utilizing Docker, Kubernetes, or serverless platforms Monitoring & Observability : Implement post-deployment monitoring for model performance, data drift, and latency using tools. Experience in Monte Carlo is preferable Automation: Automate retraining and data pipeline workflows to ensure models stay accurate over time. Manage the deployment of foundation models, fine-tuning workflows, and Retrieval-Augmented Generation (RAG) stacks (Vector DBs, Knowledge Graph. Experience with
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role As an AI Accelerator Systems Software Technical Program manager at OpenAI, you will help bring our chips/system hardware roadmap to life, navigating an array of technical and partnership challenges. We’re looking for people excited to push the frontiers of computing by navigating technical explorations and are passionate about building. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Manage the end-to-end software development from design to implementation for our AI acceleration systems, working across technical, cross-functional and external stakeholders Lead planning and scheduling of AI system software designs with our strategic partners and vendors Coordinate and lead internal resources and communication for efficient interaction with partners and vendors. You might thrive in this role if you: Have experience as a software technical program manager for data center system products (server, GPU, TPU, networking, storage and so on) taking products from concept to volume in a data center environment ensuring the systems scale with high quality Know end-to-end software development program management techniques from concept, design, production, deployment into the data center Want to help design some of the world’s largest supercomputing systems, working at the edge of complex hardware challenges Enjoy working with and enabling world-clas
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role As a Hardware Chips Programs Manager at OpenAI, you will help bring our chips hardware roadmap to life, navigating an array of technical and partnership challenges. We’re looking for people excited to push the frontiers of computing by navigating technical explorations and are passionate about building. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Manage the design and implementation planning of our ML acceleration hardware, working across technical, cross-functional and external stakeholders Lead planning and scheduling of chip hardware designs with our strategic partners and vendors Coordinate and marshal internal resources and communication for efficient interaction with partners and vendors. You might thrive in this role if you: Have experience as a technical program manager for data center hardware products (server, GPU, TPU, networking, storage and so on) Know the whole end-to-end system program management from concept, design, production, deployment into the data center Have some experience with System SW programs through NPI Want to help design some of the world’s largest supercomputing systems, working at the edge of complex hardware challenges Enjoy working with and enabling world-class AI Researchers and Engineers Are passionate about the technical program function, and enjoy independently owning and delivering on your tea
Who We Are Nuro believes self-driving vehicles are the most immediate and profound opportunity for AI to drive positive change in the physical world. Safer streets, more time for what matters, and easier access to the world around us, that’s why we’re building a universal autonomy platform: self-driving for all roads and all rides. Founded in 2016, Nuro is a physical AI company developing Level 4 autonomous driving technology for a wide range of vehicles, use cases, and markets. Powered by the Nuro Driver™, our universal autonomy platform enables the global mobility ecosystem to deploy autonomy at scale, from robotaxis and logistics fleets to personal vehicles. With years of real-world deployment experience and a flexible, partner-led business model, Nuro is working toward a future where millions of autonomous vehicles powered by our technology help make everyday life safer, easier, and more connected. Nuro has raised over $2B in capital from Uber, NVIDIA, Google, Softbank, Fidelity, T. Rowe Price, and other leading investors. About the Role Our software team is growing, and we are looking for talented engineers to join us and be instrumental to one of the following areas: Data Platform, Simulation, and Technical Infrastructure. Data Platform: The Data Platform serves as a comprehensive management system for Nuro AI Driver's data, labels, and metrics, facilitating seamless access functionality. The team focuses on data annotation across various domains, including 2D/3D perception, mapping, behavior trajectory, and language/text. It also handles data ingestion and mining, employing methods such as heuristics and embedding search. Additionally, the platform supports the autonomy evaluation infrastructure by providing detailed introspection. Simulation: The Simulation team builds the simulator that allows us to develop and test our autonomous driving technology in a virtual setting. We work on the core simulator and simulation frameworks, sensor simulation, scena
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Senior Product Manager in Engineering Acceleration , you will define the vision and strategy for how software engineers work at Roblox. Your mission is to ensure that Roblox engineers spend more time building the metaverse and less time managing infrastructure complexity in your areas of responsibility. The Engineering Acceleration portfolio includes Source Control, Testing, Secure Software Supply Chain Management, Continuous Deployment, and Observability among many others. In this role, you will primarily own the developer experience for Continuous Deployment, Testing , and Observability , while remaining adaptable as organizational priorities evolve. AI has been transforming how we approach these systems. We’re already leveraging AI to test our software and to identify and diagnose production incidents. The successful candidate here will bring deep expertise not only in the software development lifecycle, but crucially also on the rapidly evolving landscape of AI tooling. This is a rare opportunity for an infrastructure product leader to drive meaningful impact at scale across a very large engineering organization. You Will: Define and drive the long-term visi
Who we are At Twilio, we’re shaping the future of communications, all from the comfort of our homes. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences. Our dedication to remote-first work , and strong culture of connection and global inclusion means that no matter your location, you’re part of a vibrant team with diverse experiences making a global impact each day. As we continue to revolutionize how the world interacts, we’re acquiring new skills and experiences that make work feel truly rewarding. Your career at Twilio is in your hands. . Hiring and how we work We use Artificial Intelligence (AI) to help make our hiring process efficient. That said, every hiring decision is made by real Twilions! Also, while we are a remote-first company, you may be asked to report in person on an ad-hoc basis for team gatherings, functional off-sites or customer meetings. . See yourself at Twilio Join the team as Twilio’s next Senior Principal Product Manager- Developer Platform About the job Twilio is looking for a Senior Principal, Product Management to deliver our strategic vision for the Developer Platform . You will manage foundational services, APIs, and infrastructure, driving the roadmap that empowers internal teams and partners to scale confidently. You will treat the platform as a product , shifting platform engineering from reactive, ad-hoc support into proactive, self-service experiences. Grounded in industry research (DORA, CNCF, Team Topologies, etc..), you will introduce golden paths, a discoverable developer portal, self-service tools, and high-quality documentation to reduce cognitive load and support burdens across Twilio. Collaborating with product leaders, engineering, architecture, and design, you will define the future of developer experience, tooling, deployment, and operational capabilities, acting as the steward for ev
About Prodigal Prodigal is the connected AI platform leading financial institutions use to run their operations. We work with banks, lenders, credit unions, and other financial companies that lend money to people and manage those relationships over time. These institutions make millions of high-stakes decisions every day. Who should they reach? When should they reach them? What should they say or offer? When should a case move to a human? How should that change based on the borrower, the account, previous interactions, and the regulations involved? Getting those decisions right requires a deep understanding of the people, processes, rules, and edge cases behind them. Prodigal has spent the last eight years building that understanding. More than a billion interactions between financial institutions and their customers have shaped the intelligence, guardrails, and AI agents we now run in production across North America. Today, our AI agents analyze conversations, capture context, guide human agents, decide the next action, conduct customer conversations, orchestrate outreach, and help people complete payments and resolutions. They are connected, so what is learned in one interaction can inform what happens next. We are expanding this swarm of AI agents across more of the work financial institutions do: originations, document processing, back-office workflows, servicing, and other critical operations where money, identity, people, and regulation intersect. We are backed by Y Combinator, Accel, and Menlo Ventures, and work with 100+ financial institutions across North America. Listen directly from our CTO, Cofounder - Sangram Raje About the Role We're looking for a Customer Success Manager who thrives on chaos as much as clarity. Once an account is handed to you by our AI Deployment team after the hyper-care phase, you own it, whether that means smoothing over rough edges from a still-settling rollout or building on a genuinely stable one. Either way, your job is
Beacon Biosignals is transforming precision medicine for the brain, from clinical development to clinical care. For Life Sciences partners, we offer the leading at-home EEG platform for clinical development of novel therapeutics for neurological, psychiatric, and sleep disorders. Our Diagnostics business is building the most comprehensive at-home platform for precision diagnostics, combining EEG and cardiopulmonary signals to deliver reimbursable assessments for sleep and central nervous system disorders. Together, we're changing the way patients are diagnosed and treated for any disorder that affects brain physiology. We are seeking a Clinical Trial Operations Associate to join us in our mission to make brain monitoring easily accessible, interpretable, and actionable. In this role, you will collaborate with clinical research sites, project teams, and internal stakeholders to support the deployment of Beacon’s devices in clinical studies. This role focuses on study startup, site management, and live-study monitoring, ensuring the highest quality standards and compliance with regulatory requirements. This role is fully remote anywhere in the Pacific or Mountain Timezones in the U.S. and will require up to 5% travel. Beacon's robust asynchronous work practices ensure a first-class remote work experience, but we also have in-person office hubs available located in Boston, New York, and Paris. What success looks like Collaborate with project teams to support study startup activities. Act as the primary point of contact for clinical sites, ensuring smooth operations and clear communication. Provide training to research sites on the use of Beacon’s devices and study protocols, ensuring proper device usage and data collection. Monitor project progress, ensuring adherence to timelines, protocols, and quality standards. Maintain compliance with Good Clinical Practice (GCP), International Conference on Harmonization (ICH) guidelines, and regulatory standards. A
WPP is the trusted growth partner for the world’s leading brands. We unite cutting-edge media intelligence and data solutions, world-class creativity, next-generation production, transformative enterprise solutions and expert strategic counsel in a single company – powered by exceptional talent and our agentic marketing platform, WPP Open, to help our clients navigate change, capture opportunity and deliver transformational growth. We work with the world's most valuable brands and have global reach across 100+ markets, with deep local expertise. Our people are the key to our success. We're committed to fostering a culture of creativity, belonging and continuous learning, attracting and developing the brightest talent, and providing exciting career opportunities that help our people grow. For more information, visit WPP.com. Why we're hiring: As a member of the Global Technical Operations (TechOps), you will be a part of a team that focuses on operational reliability within a cloud-based infrastructure. You have hands-on cloud experience in architecting, building, deploying, managing databases, compute instances, and storage buckets. You have a passion for providing solutions through automation. You know that success is through collaboration and communication. What you'll be doing: Work in cross-functional teams to develop solutions and identify opportunities to bring efficiency and effectiveness. Research, evaluate, and incorporate new technologies/concepts into existing frameworks. Proactively identify areas to improve efficiency and effectiveness, recommend and implement solutions towards them. Develop and innovate operational practices, procedures for workflows, and documentation. Implement and contribute to IT security best practices. Automate tasks to ensure consistency and speed of deployment. Identify, analyze, and troubleshoot issues and work towards resolution. Explain technical solutions to bo
PagerDuty (NYSE:PD) is a leader in Digital Operations Management. In an always-on world, organizations of all sizes trust PagerDuty to help them deliver a perfect digital experience to their customers, every time. Teams use PagerDuty to identify issues and opportunities in real time and bring together the right people to fix problems faster and prevent them in the future. Over 13,000 organizations (including 60 of Fortune 100) rely on PagerDuty to succeed with Digital Transformation, Cloud Migration, and DevOps Modernization. Notable customers include GE, Cisco, Genentech, Electronic Arts, Cox Automotive, Netflix, Shopify, Zoom, DoorDash, Lululemon and more. We are expanding rapidly as a platform for Digital Operations Management using AI/ML and Automation and growing our adoption by Development, IT, Customer Service, Security, and other teams across the organization. About the role PagerDuty’s Operations Cloud runs on a platform that ingests billions of signals and turns them into real-time action for thousands of customers. We’re looking for a Senior AI/ML Engineer who lives at the intersection of two disciplines: large-scale distributed systems and applied AI. In this role you will design and ship AI systems that run in production at PagerDuty’s scale — powering Incident Management AI Agents, event intelligence, and the LLM-powered capabilities embedded across our platform. You’ll own the full lifecycle, from framing the problem to serving reliably at scale. We are looking for a candidate who is genuinely passionate about building with modern AI — LLMs, agents, and retrieval — but grounded in the realities of building resilient, high-throughput systems. What you’ll do Design and build AI-powered features — LLM agents, retrieval, and event intelligence — that operate on high-volume, real-time event streams, from problem framing through production deployment and monitoring. Architect and own the systems behind them: agent and prompt orchestration, retrieval pipelin
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! About the Opportunity New Relic is looking for a Senior Revenue Operations Manager – GTM Business Planning to own the operating system behind how we plan, pay, and scale our Go-To-Market (GTM) organization. Sitting at the intersection of Sales, Finance, and Systems, you will architect the annual GTM plan, design incentive compensation programs, and govern processes across quota, territory, and comp data. What You'll Do GTM Planning & Architecture: Own the annual planning cycle end-to-end, building capacity models and territory frameworks to ensure optimal market coverage and alignment with targets. Incentive Compensation Design: Partner with FP&A to architect and govern financially sound sales incentive plans and compensation policies that motivate seller behavior, drive strategic priorities, and eliminate payout leakage. Quota Management & Strategy: Establish fair quota methodologies, manage ongoing adjustments for transfers/hires, and evaluate rep productivity to handle market shifts. Operations & Governance: Govern monthly sales compensation data within Salesforce and downstream systems to maintain audit-ready accuracy across territories and quotas. Systems & Automation: Own the deployment of annual territories and quotas in Salesforce, drive the roadmap for GTM planning systems, and lead automation initiatives to eliminate manual effort and scale operations. This Role Requires B2B SaaS Experience: 6–8+ years in Revenue/Sales Operations, Sales Comp, FP&a
About Ema Ema is building the world’s leading Agentic AI platform to transform enterprise productivity. We enable organizations to delegate repetitive tasks to Ema, the Universal AI Employee, delivering 10x gains in workforce efficiency, across functions. Founded by former executives from Google, Coinbase, Flipkart, and Okta, our team includes engineers from premier tech companies and graduates of Stanford, MIT, UC Berkeley, CMU, and IITs. We are backed by industry leading investors including Accel, Naspers/Prosus, Section32, and angels like Sheryl Sandberg and Dustin Moskovitz. Headquartered in Silicon Valley and with offices in London, Bangalore and Vancouver, Ema is at the frontier of what Agentic AI can do in production — we ship real systems that run real business processes at scale. About the Role As a Site Reliability Engineer at Ema, you will own the stability, availability, and operational health of our agentic AI platform across customer environments. You'll work closely with Engineering and DevOps to provision infrastructure, drive deployment excellence, and keep production running at the quality bar our enterprise customers expect — 99.9%+ uptime, proactive incident response, and continuous improvement. What You'll Do Infrastructure & Deployment Design and provision cloud infrastructure (GCP, Azure, AWS) tailored to customer environments, with security, scalability, and compliance built in Execute on-call SaaS deployments with minimal downtime; automate and optimize deployment workflows end-to-end Production Stability & Observability Monitor logs, alerts, and metrics to maintain SLA commitments and catch issues before they escalate Diagnose and resolve production incidents with speed and rigor; drive root cause analysis and permanent fixes Collaborate with DevOps to enhance monitoring dashboards and alerting frameworks; deliver clear system health reporting to internal and customer stakeholders Documentation & Knowledge Management Maintain de
Sales Excellence Lead, Inference and Agentic AI Location: Noida Company: Paytm About Paytm: Paytm is a pioneer of digital payments in India, serving over 450 million consumers and 45 million merchants across payments, financial services, and commerce. Over the years, Paytm has built deep in-house capabilities across technology, data, and operations to operate at scale with high reliability. Paytm is building a full stack AI platform focussed on Inference and Agents, enabling large enterprises to deploy AI driven automation across sales, service, operations, and analytics. The Inference and Agentic AI team operates as a cross functional unit spanning engineering, product, data science, business management, and sales, and owns the full lifecycle of AI solutions from opportunity discovery to deployment and scale. Role Overview: Paytm is looking to hire a Sales Excellence Lead within the Inference and Agentic AI organization to drive sales effectiveness, funnel governance, and sales enablement across Paytm's AI products. This role will work closely with sales leadership, business teams, product teams, and marketing teams to improve sales productivity, strengthen pipeline conversion, and accelerate revenue growth. The role will own sales review cadences, performance tracking, enablement programs, product readiness, and sales execution excellence across enterprise and mid market segments. The ideal candidate combines analytical rigor with strong stakeholder management and a passion for building scalable sales processes and enablement frameworks. Key Responsibilities: Sales Performance and Funnel GovernanceOwn the operating rhythm for weekly, monthly, and quarterly sales reviews across enterprise and mid market sales teams. Track and analyze funnel performance across lead generation, opportunity creation, pipeline progression, proposal conversion, closures, activation, and expansion. Identify bottlenecks, conversion leakages, and productivity gaps across sales channel
Get new platform deployment management lead jobs by email
Daily job updates · Unsubscribe anytime