Jobiba hiring network

Platform Deployment Management Lead Jobs

10,000 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current platform deployment management lead jobs. Use filters to narrow by work mode, employment type, experience and date posted.

O
OpenAI
📍 San Francisco• Full-time
1mo ago

This role will support the fleet infrastructure team at OpenAI. The fleet team focuses on running the world’s largest, most reliable, and frictionless GPU fleet to support OpenAI’s general purpose model training and deployment. Work on this team ranges from Maximizing GPUs doing useful work by building user-friendly scheduling and quota systems Running a reliable and low maintenance platform by building push-button automation for kubernetes cluster provisioning and upgrades Supporting research workflows with service frameworks and deployment systems Ensuring fast model startup times though high performance snapshot delivery across blob storage down to hardware caching Much more! About the Role As an engineer within Fleet infrastructure, you will design, write, deploy, and operate infrastructure systems for model deployment and training on one of the world’s largest GPU fleet. The scale is immense, the timelines are tight, and the organization is moving fast; this is an opportunity to shape a critical system in support of OpenAI's mission to advance AI capabilities responsibly. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Design, implement and operate components of our compute fleet including job scheduling, cluster management, snapshot delivery, and CI/CD systems. Interface with researchers and product teams to understand workload requirements Collaborate with hardware, infrastructure, and business teams to provide a high utilization and high reliability service You might thrive in this role if you: Have experience with hyperscale compute systems Possess strong programming skills Have experience working in public clouds (especially Azure) Have experience working in Kubernetes Execution focused mentality paired with a rigorous focus on user requirements As a bonus, have an understanding of AI/ML workloads About OpenAI OpenAI is an AI resea

awsazurekubernetes
View job →

Pre-sales Manager, Inference and Agentic AI - Paytm Location: Noida Company: Paytm About Paytm: Paytm is a pioneer of digital payments in India, serving over 450 million consumers and 45 million merchants across payments, financial services, and commerce. Over the years, Paytm has built deep in-house capabilities across technology, data, and operations to operate at scale with high reliability. Paytm is building a full stack AI platform focussed on Inference and Agents, enabling large enterprises to deploy AI driven automation across sales, service, operations, and analytics. The Inference and Agentic AI team operates as a cross functional unit spanning engineering, product, data science, business management, and sales, and owns the full lifecycle of AI solutions from opportunity discovery to deployment and scale. Pre-Sales Manager | Paytm AI Key Responsibilities 🔹 Partner with Enterprise Sales teams to understand client requirements and map them to Paytm AI solutions. 🔹 Conduct discovery sessions, identify business challenges, and create solution recommendations. 🔹 Deliver customized product demos, solution walkthroughs, and Proof of Concepts (POCs) for enterprise clients. 🔹 Act as a technical advisor, addressing product, integration, architecture, and implementation-related queries. 🔹 Create solution documents, proposal inputs, RFP responses, technical FAQs, and sales enablement assets. 🔹 Collaborate with Product, Engineering, Onboarding, and Business teams to ensure seamless solution delivery. 🔹 Capture client feedback and provide insights to improve product capabilities and customer experience. 🔹 Support enterprise deal closures through strong solutioning, stakeholder management, and technical consulting. Ideal Candidate ✔ 3–6 years of experience in Pre-Sales, Solution Consulting, Solutions Engineering, Product Consulting, or Enterprise Technology roles. ✔ Strong understanding of APIs, integrations, SaaS platforms, enterprise applications, and solut

Job Description: Role: Relationship Manager – MTF HNI Segment Objective of the Role Drive growth, acquire new customers, activation, and retention of HNI MTF clients by building strong relationships, improving engagement, and increasing capital deployment across the MTF book. Key Responsibilities 1. HNI Client Engagement Own a portfolio of HNI MTF clients Drive regular engagement through calls, messages, and structured outreach Educate clients on MTF usage, leverage opportunities, and platform features 2. Revenue & Book Expansion Increase MTF utilization among assigned clients Identify upsell opportunities across equity + MTF Improve wallet share per client 3. Activation & Retention Reactivate dormant HNI users Reduce churn in high-value segments Ensure consistent usage of MTF facility 4. Relationship Management Act as single point of contact for assigned HNI users Resolve queries in coordination with support/product teams Build trust and long-term engagement 5. Market Intelligence Gather user feedback and competitive insights Share patterns from HNI cohort behavior with product/growth teams Desired Profile Experience in brokerage, wealth management, or trading platforms preferred Strong communication and persuasion skills Comfort with high-frequency client interaction Execution-focused and target-driven mindset KPIs for Success Increase in MTF activation rate among HNI users Acquire new HNI customers Growth in average book size per HNI client Improvement in retention / reduced churn in HNI cohort Revenue contribution from managed cohort

reactrust
View job →

Client Onboarding Manager – Inference & Agentic AI | Paytm (Noida) About the Role Paytm is a pioneer of digital payments in India, serving over 450 million consumers and 45 million merchants across payments, financial services, and commerce. Over the years, Paytm has built deep in-house capabilities across technology, data, and operations to operate at scale with high reliability. Paytm is building a full stack AI platform focussed on Inference and Agents, enabling large enterprises to deploy AI driven automation across sales, service, operations, and analytics. The Inference and Agentic AI team operates as a cross functional unit spanning engineering, product, data science, business management, and sales, and owns the full lifecycle of AI solutions from opportunity discovery to deployment and scale. Key Responsibilities Own client onboarding from sales handover to go-live. Understand client workflows, systems, and integration requirements. Coordinate with Product, Engineering, and Client teams for seamless deployment. Manage onboarding timelines, milestones, and stakeholder communication. Conduct client training sessions and drive product adoption. Track onboarding KPIs, client satisfaction, and implementation success. Gather client feedback and support continuous product improvements. Ideal Candidate 2–5 years of experience in Client Onboarding, Implementation, Customer Success, or Solutions Engineering. Experience in SaaS, Fintech, Enterprise Technology, or AI products preferred. Good understanding of APIs, integrations, CRM systems, and enterprise workflows. Strong project management, problem-solving, and stakeholder management skills. Excellent communication and client-facing abilities. Bachelor’s degree in Engineering, Business, or related field. Location: Noida Why Join? Be part of Paytm’s fast-growing AI business and work closely with enterprise clients to deliver cutting-edge AI-driven automation solutions at scale.

About DevRev At DevRev, we're building the future of work with Computer – your AI teammate. Unlike traditional tools, Computer unifies all your data sources, tools, and workflows into a single AI-ready platform, giving employees real-time insights, proactive suggestions, and powerful agentic actions. It extends your existing software with AI-native apps and agents that work alongside your teams and customers – updating workflows, coordinating across teams, and eliminating repetitive work. We call this Team Intelligence: human-AI collaboration that breaks down silos, brings people back together, and frees you to solve bigger problems. Backed by Khosla Ventures and Mayfield with $150M+ raised, DevRev is trusted by global companies across industries. What You’ll Do: Architect the Future of AI Infrastructure: You will design, build, and own the end-to-end platform that supports the entire lifecycle of our ML models—from massive-scale distributed training to ultra-low-latency, highly-available inference. Optimize and Serve Cutting-Edge Models: You'll implement and scale sophisticated inference stacks for LLMs using frameworks like vLLM, TensorRT-LLM, or SGLang . You’ll solve complex challenges in throughput, latency, token streaming, and automated scaling to deliver a seamless user experience. Empower AI Innovation: You will act as a strategic partner to our AI Research and Data Science teams. You’ll create a seamless developer experience that accelerates their ability to experiment, fine-tune, and deploy groundbreaking models with velocity and confidence. Automate Everything: You'll develop robust CI/CD/CT (Continuous Training) pipelines using tools like Argo Workflows, ArgoCD, and GitHub Actions to automate model validation, deployment, and lifecycle management, ensuring our systems are both agile and rock-solid. What are we looking for Experience: 5+ years in infrastructure or software engineering, with at least 2+ years laser-focused on MLOps or ML infrastructu

pythonkubernetesci/cd
View job →
SA
Scale AI
📍 San Francisco• Full-time• From $184K/yr
17 days ago

Scale AI is seeking a highly skilled and motivated Software Engineer, ARC (Architecture, Reliability, & Compute) to join our dynamic Public Sector Engineering team. As a part of this team, you will define how the company ships software, establishing the patterns for deploying into complex government and high-security environments, rather than just running Terraform scripts. You will build and maintain internal CLIs/tools that standardize testing, deployment, environment management and are tools that engineering relies on to prevent downstream breakages. You will execute on automated deployment efforts to pay down tech debt, creating fully functional staging/testing environments, and defining the company's standard for safe deployments. You will: Design and implement secure scalable backend systems for Public Sector customers, leveraging Scale's modern and cloud-native AI infrastructure. Own services or systems and define their long-term health goals, while also improving the health of surrounding components. Re-architect the stack to run in compliant or restrictive environments. This requires designing swappable components (auth, storage, logging) to meet government/security mandates without breaking the product. Collaborate with cross-functional teams to define and execute the vision for backend solutions, ensuring they meet the unique needs of government agencies operating in secure environments. Participate actively in customer engagements, working closely with stakeholders to understand requirements and deliver innovative solutions. Contribute to the platform roadmap and product strategy for Scale AI's Public Sector business, playing a key role in shaping the future direction of our offerings. Must have: At least an active secret clearance and the ability & willingness to up level to TS/SCI with CI Poly. This is a requirement and candidates will not be considered who do not hold at least a secret clearance Ideally you'd have: Full Stack Development: Prof

awsazuregcp
View job →

We are hiring a Security Software Engineer to design and implement the hardware-backed security foundations used across OpenAI’s device ecosystem. A central focus of this role is hardening the boundary between our policy systems and the HSMs that protect sensitive cryptographic keys. This boundary determines which operations may be performed, what may be signed, which policies must be satisfied, and how changes to trusted software and policy are authorized. You will develop security-critical software and firmware within, or immediately adjacent to, an HSM trust boundary. Depending on your background, this may include HSM trusted applications, firmware services, cryptographic mechanisms, device drivers, PKCS#11 components, secure-provisioning protocols, or signing-policy enforcement systems. This is a hands-on software-engineering role. You will be expected to design systems, write and review production code, debug across hardware and software boundaries, and carry projects from initial requirements through deployment. It is not an HSM administration, PKI operations, compliance, or architecture-only position. In This Role, You Will Design and implement security-critical software and firmware for HSMs, secure elements, trusted execution environments, and hardware roots of trust. Build and harden the policy-to-HSM boundary responsible for authorizing certificate issuance and cryptographic signing operations. Develop HSM trusted applications, firmware components, host interfaces, device drivers, SDKs, or cryptographic service integrations. Implement or extend cryptographic interfaces such as PKCS#11, OpenSSL providers or engines, platform key-storage APIs, or comparable hardware-security interfaces. Build firmware and software that cryptographically enforces key generation, provisioning, usage, rotation, recovery, and destruction policies. Design and implement HSM-backed certificate authority, code-signing, key-management, and device-identity systems. Develop end-to-end

awsgitrest
View job →
M
Mongodb
📍 Gurugram• Full-time
1mo ago

About Team The IT SaaS Engineering Team is the backbone of MongoDB's go-to-market operations. We build and run the Salesforce platform — with the automation, security, and governance that Sales Operations, Revenue Operations, Security, and Compliance rely on daily to manage customer relationships and drive business growth. Role Overview Join us as a Senior Salesforce Engineer and take ownership of the CPQ and CLM systems that power MongoDB's quote-to-contract lifecycle. You'll design and build scalable pricing, quoting, and contract automation solutions end-to-end — from first design conversation to production support — working alongside senior engineers and stakeholders across Sales, Deal Desk, and Legal. This is a role for an engineer who enjoys solving hard technical problems, cares deeply about accuracy and compliance in revenue-critical systems, and wants to bring modern, AI-augmented development practices into how enterprise Q2C solutions get built. What you’ll do Design, build, and maintain scalable CPQ configuration (product rules, price rules, bundles, discounting logic) and CLM workflows (clause libraries, approval chains, e-signature integration) Partner with Sales, Deal Desk, Legal, and Finance stakeholders to translate quote-to-cash requirements into secure, maintainable technical solutions Own delivery across the full solution lifecycle — design, development, testing, deployment, documentation, and production support Collaborate with globally distributed engineering teams to deliver reliable, well-governed CPQ/CLM solutions in an Agile environment Drive strong engineering practices across code quality, documentation, release management, and platform governance Identify opportunities to simplify pricing logic, reduce legacy quote template debt, and strengthen the overall Q2C operating model Troubleshoot complex pricing, approval, or contract-generation issues and deliver sustainable, long-term fixes Stay current on CPQ/CLM platform capabilities and appl

javascriptjavamongodb
View job →
L
10 days ago

Leidos has an exciting opportunity for a Principal Software Engineer in our Intel Security Sector's Analysis Solutions Business Area . Our talented team is at the forefront in Security Engineering, Computer Network Operations (CNO), Mission Software, Analytical Methods and Modeling, Signals Intelligence (SIGINT), and Cryptographic Key Management. At Leidos , we offer competitive benefits , including Paid Time Off, 11 paid Holidays, 401K with a 6% company match and immediate vesting, Flexible Schedules, Discounted Stock Purchase Plans, Technical Upskilling, Education and Training Support, Parental Paid Leave, and much more. Join us and make a difference in National Security! Job Summary As a Software Engineer on this program, you will have the opportunity to build strong systems, software, and cloud environments while providing operations and maintenance for critical systems. This role will provide technical expertise in the design, development, implementation and testing of customer tools and applications. Based in a DevOps framework, this role participates in and/or directs major deliverables of projects through all aspects of the software development lifecycle including scope and work estimation, architecture and design, coding and unit testing. Primary Responsibilities: Participates in and/or directs software programming initiatives using Java, JavaScript, Python, SpringBoot, and Hibernate. Develops software system validation and testing methods using Junit and Katalon and uses integrated custom developed software solutions to leverage automated deployment technologies Develop, prototype and deploy solutions within Commercial Cloud Solutions leveraging infrastructure platform services Coordinate closely with team members, Product Owners and Scrum Masters to ensure User Story alignment and implementation to customer use cases Supp

javascriptpythonjava
View job →
L
10 days ago

Leidos has an exciting opportunity for a Principal Java Developer in our Intel Security Sector's Analysis Solutions Business Area . Our talented team is at the forefront in Security Engineering, Computer Network Operations (CNO), Mission Software, Analytical Methods and Modeling, Signals Intelligence (SIGINT), and Cryptographic Key Management. At Leidos , we offer competitive benefits , including Paid Time Off, 11 paid Holidays, 401K with a 6% company match and immediate vesting, Flexible Schedules, Discounted Stock Purchase Plans, Technical Upskilling, Education and Training Support, Parental Paid Leave, and much more. Join us and make a difference in National Security! Job Summary A Java developer on this program provides Agile development and operations and maintenance for mission critical systems. Based in DevOps framework, this role participates in major deliverables of projects through all aspects of the software development lifecycle including scope and work estimation, architecture and design, coding and unit testing. Primary Responsibilities: Design, develop, test and maintain high-performance, scalable backend microservices using Java and the Spring Framework to meet customer information technology needs. Participates in software programming initiatives, shaping backend architecture, mentoring team members, and conducting code reviews. Develops and directs software system validation and testing methods using Junit and Katalon and uses integrated custom developed software solutions to leverage automated deployment technologies Develop, prototype and deploy solutions within a cloud-based platform leveraging platform services. Analyze (though proof of concept, performance, and end-to-end testing) and effectively coordinate Infrastructure needs driven by developed software to meet customer mission needs Support the Agil

pythonjavapostgresql
View job →
L
10 days ago

Leidos has an exciting opportunity for a Sr. Software Engineer in our Intel Security Sector's Analysis Solutions Business Area . Our talented team is at the forefront in Security Engineering, Computer Network Operations (CNO), Mission Software, Analytical Methods and Modeling, Signals Intelligence (SIGINT), and Cryptographic Key Management. At Leidos , we offer competitive benefits , including Paid Time Off, 11 paid Holidays, 401K with a 6% company match and immediate vesting, Flexible Schedules, Discounted Stock Purchase Plans, Technical Upskilling, Education and Training Support, Parental Paid Leave, and much more. Join us and make a difference in National Security! Job Summary As a Software Engineer on this program, you will have the opportunity to build strong systems, software, and cloud environments while providing operations and maintenance for critical systems. This role will provide technical expertise in the design, development, implementation and testing of customer tools and applications. Based in a DevOps framework, this role participates in and/or directs major deliverables of projects through all aspects of the software development lifecycle including scope and work estimation, architecture and design, coding and unit testing. Primary Responsibilities: Participates in and/or directs software programming initiatives using Java, JavaScript, Python, SpringBoot, and Hibernate. Develops software system validation and testing methods using Junit and Katalon and uses integrated custom developed software solutions to leverage automated deployment technologies Develop, prototype and deploy solutions within Commercial Cloud Solutions leveraging infrastructure platform services Coordinate closely with team members, Product Owners and Scrum Masters to ensure User Story alignment and implementation to customer use cases Support th

javascriptpythonjava
View job →
L
10 days ago

Leidos has an exciting opportunity for a Software Engineer (SME) in our Intel Security Sector's Analysis Solutions Business Area . Our talented team is at the forefront in Security Engineering, Computer Network Operations (CNO), Mission Software, Analytical Methods and Modeling, Signals Intelligence (SIGINT), and Cryptographic Key Management. At Leidos , we offer competitive benefits , including Paid Time Off, 11 paid Holidays, 401K with a 6% company match and immediate vesting, Flexible Schedules, Discounted Stock Purchase Plans, Technical Upskilling, Education and Training Support, Parental Paid Leave, and much more. Join us and make a difference in National Security! Job Summary As a Software Engineer on this program, you will have the opportunity to build strong systems, software, and cloud environments while providing operations and maintenance for critical systems. This role will provide technical expertise in the design, development, implementation and testing of customer tools and applications. Based in a DevOps framework, this role participates in and/or directs major deliverables of projects through all aspects of the software development lifecycle including scope and work estimation, architecture and design, coding and unit testing. Primary Responsibilities: Participates in and/or directs software programming initiatives using Java, JavaScript, Python, SpringBoot, and Hibernate. Develops and directs software system validation and testing methods using Junit and Katalon and uses integrated custom developed software solutions to leverage automated deployment technologies Develop, prototype and deploy solutions within Commercial Cloud Solutions leveraging infrastructure platform services Coordinate closely with team members, Product Owners and Scrum Masters to ensure User Story alignment and implementation to customer use cases

javascriptpythonjava
View job →
L
10 days ago

Leidos has an exciting opportunity for a Sr. Java Developer in our Intel Security Sector's Analysis Solutions Business Area . Our talented team is at the forefront in Security Engineering, Computer Network Operations (CNO), Mission Software, Analytical Methods and Modeling, Signals Intelligence (SIGINT), and Cryptographic Key Management. At Leidos , we offer competitive benefits , including Paid Time Off, 11 paid Holidays, 401K with a 6% company match and immediate vesting, Flexible Schedules, Discounted Stock Purchase Plans, Technical Upskilling, Education and Training Support, Parental Paid Leave, and much more. Join us and make a difference in National Security! Job Summary A Java developer on this program provides Agile development and operations and maintenance for mission critical systems. Based in DevOps framework, this role participates in major deliverables of projects through all aspects of the software development lifecycle including scope and work estimation, design, coding and unit testing. Primary Responsibilities: Design, develop, test and maintain high-performance, scalable backend microservices using Java and the Spring Framework to meet customer information technology needs. Participate in software programming initiatives and code reviews. Develops software system validation and testing methods using Junit and Katalon and uses integrated custom developed software solutions to leverage automated deployment technologies Develop, prototype and deploy solutions within a cloud-based platform leveraging platform services. Support the Agile software development lifecycle following Program SAFe practices while coordinating closely with team members, Product Owners and Scrum Masters to ensure User Story alignment and implementation to customer use cases Document and perform systems software development, including dep

pythonjavaaws
View job →
O
1mo ago

Join the engineering teams that bring OpenAI’s ideas safely to the world! The Applied Engineering team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role As OpenAI continues to grow, we are looking for experienced, problem-solving engineers to ensure our systems scale. Our success depends on our ability to quickly iterate on products while also ensuring that they are performant and reliable. You will work in a deeply iterative, collaborative, fast-paced environment to bring our technology to millions of users around the world, and ensure it’s delivered with safety and reliability in mind. Successful candidates will play a crucial role in ensuring the reliability, scalability, and performance of our systems as we continue to expand. As a reliability expert, you will be at the forefront of maintaining and enhancing the stability, scalability, and performance of our rapidly evolving infrastructure. You will work closely with cross-functional teams, including software engineers, product managers, and data scientists, to build and maintain resilient systems that can handle our growing user base and workload. In this role, you will: Design and implement solutions to ensure the scalability of our infrastructure to meet rapidly increasing demands. Build and maintain the load, chaos and synthetic-testing software leveraged by development teams to make the systems they design and operate more reliable. Build and maintain automation tools to streamline repetitive tasks and improve system reliability. Build and maintain the platform for CPU, storage, GPU, and network lifecycle management to drive efficiency, accountability and dynamic optimization of our resources. Implement fault-tolerant and resilient design

awskubernetesrest
View job →
B
Baseten
📍 San Francisco• Full-time
1mo ago

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Forward Deployed Engineer at Baseten, you will partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten’s platform. You’ll own the journey with customers from initial exploration to production deployment, translating ambiguous business goals into reliable, observable services with clear quality, latency, and cost outcomes. This role is a great fit for entrepreneurial engineers who want a front-row view into how modern companies adopt AI at scale and who enjoy working across product, software development, performance engineering, and customer-facing implementations. To be clear, this is an engineering role with hands-on coding and software development that also includes aspects of product management, technical customer success, and pre-sales solution engineering mixed in. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Forward Deployed Engineering team: Forward Deployed Engineering on the frontier of AI The fastest, most accurate Whisper transcription Deploy production-ready model servers from Docker images Deploy custom ComfyUI workflows as APIs RESPONSIBILITIES Develop and maintain software systems and product features using one or more general-purpose programming languages in a production-level environment, with a preference for Python due to its relevance in ML projects. Drive customer impact by designing, implementin

pythondockermachine learning
View job →
🔔

Get new platform deployment management lead jobs by email

Daily job updates · Unsubscribe anytime