The Product Solutions Architecture (PSA) team acts as a technical multiplier across Datadog. PSAs are domain experts who partner with Field teams on complex customer use cases across pre- and post-sales engagements and scale their impact by producing reusable collateral, including reference architectures, technical guides, and enablement assets. By feeding real-world customer insights back to Datadog Product teams, PSAs help influence product roadmaps while accelerating adoption, usage, and long-term customer success. Datadog’s LLM Observability product enables organizations to monitor, troubleshoot, and optimize large-scale LLM-powered applications with confidence, while meeting requirements around data privacy, compliance, and cost management. As a Product Solutions Architect, you will partner closely with Datadog customers and the LLM Observability product team to design architectures, implement best practices, and drive adoption of LLM observability across customer environments. At Datadog, we place value in our office culture - the relationships and collaboration it builds, and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Serve as the in-house subject matter expert for Datadog’s LLM Observability product Partner with Field teams to provide hands-on technical and architectural guidance to enterprise customers adopting LLM Observability Create high-impact technical collateral, including reference architectures, technical guides, cookbooks, and documentation to enable Field teams and the broader customer community Build proofs of concept and small-scale deployments to validate solutions and reproduce real-world customer environments Act as a trusted advisor to Product Management by delivering actionable feedback informed by real-world field experience Who You Are: You bring a strong software engineering foundation, with hands-on experience bu
Jobiba hiring network
Ai Deployment Manager Jobs
10,000 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current ai deployment manager jobs. Use filters to narrow by work mode, employment type, experience and date posted.
The database market is massive and MongoDB is at the head of its disruption. The MongoDB community is transforming industries and empowering developers to build amazing apps that people use every day. We are the leading modern data platform and the first database provider to IPO in over 20 years. Join our team and be at the forefront of innovation and creativity. Building on the rapid success and adoption of MongoDB, we are delivering applications and services that make it much easier to manage and scale database deployments. Our Cloud Technical Services Engineers provide their exceptional analytical skills and customer service to ensure that MongoDB users are successful with our suite of MongoDB application-layer products. We’re looking for individuals who want to dig into the details of how “big data” and “web-scale” systems are successfully assembled and operated every day by organisations of every size and flavour. Cloud Technical Services Engineers are a key part of our elite Technical Services organisation that supports MongoDB customers worldwide, with teams in locations including New York, Sydney, Gurugram, and Dublin. If you’re passionate about the opportunity to get comfortable working with the cloud every single day and be part of a team that works at the frontier of SaaS services and database systems, this is the role for you. We are looking to speak to candidates who are based in Sydney for our hybrid working model. Responsibilities As a Cloud Technical Services Engineer, you’ll be advising customers on strategies and documented practices for making best use of our cloud products. This will require you to translate technical concepts and patterns into generalist terms for our customers, helping them understand, install, and use those applications well. You’ll also troubleshoot technical problems with these products, and be an advocate for our users’ needs, collaborating with the MongoDB product management and development teams on their behalf.&nbs
Shape the Future with Dun & Bradstreet At Dun & Bradstreet, we believe data has the power to create a better tomorrow. As a global leader in business decisioning data and analytics, we help companies worldwide grow, manage risk, and innovate. For over 180 years, businesses have trusted us to turn uncertainty into opportunity. We’re a diverse, global team that values creativity, collaboration, and bold ideas. Are you ready to make an impact and help shape what’s next? Join us! Explore opportunities at dnb.com/careers. Sr. Data Cloud Engineer(s)(multiple positions) – Duties are using relational database systems, CI/CD pipeline, Kubernetes for application deployments, terraform IAC, Observability using Splunk, server-side GitHub REST API, & Snowflake, Spark, & Python to design, automate, implement & maintain data architecture patterns in AWS & GCP for data collection infrastructure & pipelines for multi-region cloud-based big data IaaS & PaaS solutions including performing requirements & document use technology analysis; coordinating data integration & ETL processes; supporting data security, integrity & cost controls; automating data processes utilizing Cloud Composer & develop SQL scripts for data modeling; performing end-to-end pipeline load performance testing; supporting technology upgrades through PostgreSQL & Cloud SQL cloud migrations; performing SQL & NoSQL database tuning & maintenance; providing guidance on data management & SQL optimization; & developing data warehousing management policies. Requires Bach degree in Comp Science, Comp Engineering, Information Technology (IT) or related field & 5 yrs exp in job duties as stated. Position is with Dun & Bradstreet in Jacksonville, FL. Inquire and send resume through Dun & Bradstreet’s job board at https://jobs.lever.co/dnb. Position is under Sr. Data Cloud Engineer for Jacksonville, FL. #LI-DNI
Airtable is the no-code app platform that empowers people closest to the work to accelerate their most critical business processes. More than 500,000 organizations, including 80% of the Fortune 100, rely on Airtable to transform how work gets done. Airtable’s infrastructure is evolving to meet the needs of our fast growing engineering org. We are looking for infrastructure engineers to join our team to help improve critical product infrastructure, with a focus on building systems that have a great developer experience and will scale as we grow. We currently have openings on: Asynchronous Serving: The Asynchronous Serving team is scaling critical systems used by Airtable’s most essential and up-and-coming product features, especially AI features. Upcoming projects include refactoring our background task queue to track its tasks in DynamoDB, adding quality of service to the job queue, and revamping a streaming service to handle 10x scale while being more resilient. Compute: The compute pod builds and manages our Kubernetes-based platform that supports every service at Airtable, including all new AI services such as vector databases, AI evals store, and document extraction and understanding services. We have a lot of exciting foundational work in our roadmap, such as Overhauling our network stack and service discovery, to simplify service setup and strengthen security Region level disaster recovery, and bringing up compute platform from 0->1 in a new region Building custom Kubernetes operators for reliably managing some of our most critical workloads Developer Platform : The Developer Platform team sits at the intersection of all engineering at Airtable, focusing on building the internal tooling, frameworks, and CI/CD systems that power our product teams. We strive to streamline developer workflows - from build and test cycles to production deployments—and foster a best-in-class developer experience. Join us if you’re passionate about creating high-lever
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Snowflake runs large scale cloud infrastructure to deliver its own service — production and internal deployments, Kubernetes fleets, CI/CD, etc. Our cloud spend is in billions of dollars per year. The Cloud Efficiency team builds a unified, self-serve cloud efficiency platform along with AI skills and agents that makes spend observable, attributable, governable while driving recommendations and optimization of our cloud spend. AS A SOFTWARE ENGINEER AT SNOWFLAKE YOU WILL: Design, develop, and maintain scalable platform for resource ownership registry, usage attribution, utilization measurement, and cost modeling. Build AI agents, tools and automation to enhance system monitoring, alerting, and root cause analysis. Improve and optimize data ingestion, storage, and query efficiency for cloud utilization, cost and efficiency data at scale. Collaborate with teams across Snowflake to understand attribution and observability needs and implement solutions that improve operational visibility. Contribute to open-source and industry best practices in monitoring and distributed systems monitoring. Ensure high availability, reliability, and performance of team-managed platforms by participating in on-call rotations and incident management. Partner with Finance, Product and Engineering
Location: San Francisco, CA (Hybrid: 4 days onsite/week). Relocation assistance available. About the Team: We build foundational platform software that enables reliable, secure, and performant products. The team works across system layers and partners closely with adjacent engineering groups to deliver robust capabilities from concept through launch. About the Role: We’re seeking a System Software Engineer to design, implement, and debug core platform components and the pipelines that build and update system images. You’ll work across operating system layers, focusing on performance, security, and deep system debugging to ship production‑grade systems. In this role, you will: Design, implement, and debug system‑level components and services across kernel and user space. Configure and maintain OS platform services (init, services, networking, security policies) and related tooling. Build and operate image and update pipelines, ensuring reliability, reproducibility, and rollback safety. Instrument and analyze performance using profiling and tracing; optimize CPU, memory, I/O, and power usage. Own platform observability and reliability: logging, crash capture, watchdogs, and diagnostics. Collaborate with cross‑functional teams to define interfaces and deliver end‑to‑end features. Establish strong engineering practices: code review, CI, reproducible builds, and release management. Partner with external suppliers to support builds and deployments. You might thrive in this role if you: Have shipped production systems software on modern operating systems. Are proficient in C/C++ and a scripting language, and comfortable with OS internals (concurrency, memory management, filesystems, networking, power management). Bring strong systems debugging skills using debuggers, tracers, profilers, and logs across kernel/user‑space boundaries. Understand configuration of platform services and interfaces, and can translate requirements into stable, well‑documented APIs. Are fluent in u
Everpure (NYSE: P) has evolved from storage pioneer to data platform, closing fiscal 2026 with $3.7 billion in revenue, its first billion-dollar quarter, and accelerating growth into FY27. Our strategic agenda spans the companies defining the next era of technology - hyperscalers, AI labs, the AI hardware supply chain, data platform providers, and the broader AI ecosystem. This type of work—work that changes the world—is what the tech industry was founded on. So, if you're ready to seize the endless opportunities and leave your mark, come join us. THE ROLE The FlashBlade team builds the industry’s most innovative, high-performance, and highly available portfolio of products that are designed for the most demanding mission critical applications. While we deliver a hardware storage array, over 90% of our engineering staff are software engineers. Our customers are the most important part of our business and they love FlashBlade for its simplicity of management, the constant flow of new and exciting upgrades, and ability to live on the cutting edge of technology while never taking downtime, ever. FlashBlade enables our customers to leverage the agility of the public cloud for both traditional IT and cloud-native applications. You’ll architect, implement, and optimize core services that span on-premises arrays and cloud deployments—delivering seamless storage experiences to enterprises around the globe. If you thrive on solving hard problems, influencing technical direction, and driving innovation end-to-end, this is the role for you. WHAT YOU'LL DO Manage the full software development life cycle, from initial architecture and invention through development, release, and ongoing maintenance. Drive the technical strategy by influencing system design, architecture, and best practices, specifically contributing to the evolution of the FlashBlade Networking product area. Invent and optimize algorithms to orchestrate multi-array and multi-cloud storage systems, foc
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role The Security Engineering team helps make Ramp the most secure place for our customers to collect, manage, and put to work their business’ financial information Our work centers in three areas: Ramp builds products with an eye for security Ramp detects and responds to threats before they cause harm Security powers Ramp’s growth Check out our Engineering Blog for more on our tech stack, mission and values! What You’ll Do Drive our cloud security roadmap: review our cloud deployments to identify opportunities for improvement Design and build security-focused infrastructure primitives and integrate them into our existing products and development processes Lead remediation of prioritized issues across our technology stack Partner with infrastructure, data, and devops teams to design and deploy solutions that are inherently secure What You Need Minimum 5 years of experience building software Minimum 3 years of experience building in AWS (with Terraform) A strong sense of ownership: you need to drive projects from inception to scaling it in
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: Join our Site Reliability Engineering team and help ensure the reliability, scalability, and performance of Replit's infrastructure that serves millions of developers worldwide. As a Site Reliability Engineer, you will bridge the gap between development and operations, implementing automation and establishing best practices that enable our platform to scale efficiently while maintaining high availability. We are seeking SREs who are passionate about building and maintaining resilient systems at scale. Your mission will be to design and implement robust monitoring solutions, automate operational tasks, and continuously improve our infrastructure's reliability and performance. You will: Design and Implement Observability Solutions : Develop comprehensive monitoring and alerting systems using modern observability tools. Create dashboards and metrics that provide real-time visibility into system health and performance. Implement logging strategies that enable quick problem identification and resolution. Drive Automation and Infrastructure as Code : Architect and implement infrastructure automation solutions using tools like Terraform, Ansible, or Pulumi. Design and maintain CI/CD pipelines that enable reliable and consistent deployments. Create self-healing systems that can automatically respond to common failure scenarios. Establish SLOs and SLIs : Work with product and engineering teams to define and implement Service Level Objectives (SLOs) and Service Level Indicators (SLIs). Build systems to track and report on these metrics, ensuring we maintain high reliability standards while balancing innovation speed. Incident Management and Response : Lead incident response efforts, conducting thorough post-morte
Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the team + role We are building an elite team, applying frontier technologies to the world’s biggest financial problems. We’re looking for bold thinkers. Sharp problem-solvers. Builders who are wired to make an impact. Robinhood isn’t a place for complacency, it’s where ambitious people do the best work of their careers. We’re a high-performing, fast-moving team with ethics at the center of everything we do. Expectations are high, and so are the rewards. The Storage Platform team builds and operates the platform that powers database access across Robinhood. We own relational (Postgres/Aurora), key-value (DynamoDB), and caching systems, along with the SDKs, control plane automation, and data plane services that enable safe and reliable access at scale. Our mission is to standardize and strengthen how services connect to storage, improve reliability and performance, and reduce operational overhead through automation. We manage thousands of databases and hundreds of caching clusters supporting millions of users and critical brokerage workloads. Availability is our highest priority — our systems are designed to meet strict uptime targets, including no downtime during market hours. As a Staff Software Engineer , you will design and evolve the core infrastructure that underpins Robinhood’s storage systems. You’ll lead complex distributed systems initiatives such as horizontal sharding, proxy-based query routing, connection pooling, and cross-shard transactions. You’ll work on improving database reliability, performance, and cost efficiency across multi-region deployments. This role has a direct impact on system availability, laten
About the Team OpenAI's GTM Partnerships team builds a strategic, global partner ecosystem to accelerate customer success, enable enterprise AI adoption, and drive durable growth in support of OpenAI's mission toward AGI. We work cross-functionally across Product, Engineering, Security, Legal, Marketing, Customer Success, and Sales to turn strategic partnerships into measurable business outcomes. This role sits within the GTM Partnerships organization and is responsible for leading OpenAI's strategic relationships with three of the world's largest Global Systems Integrators (GSIs): HCLTech, Wipro, and Cognizant . These partners play a critical role in helping enterprise customers design, deploy, and scale AI solutions by combining OpenAI's technology with deep industry expertise, consulting capabilities, and global delivery organizations. This role reports to the Head of GSIs and will be responsible for driving executive alignment, joint go-to-market strategy, commercial execution, and scalable partnership programs across these strategic accounts. About the Role We are hiring a Partner Director, HCL, Wipro & Cognizant to lead and grow three of OpenAI's most strategic Global Systems Integrator partnerships. You will own executive relationships, develop and execute joint business plans, and drive coordinated go-to-market motions that accelerate customer adoption and revenue growth. This is a highly strategic, externally facing role that requires a blend of executive relationship management, commercial leadership, operational excellence, and ecosystem expertise. You'll work closely with OpenAI Sales, Solutions Engineering, Customer Success, Product, Marketing, Legal, Security, and partner leadership teams to build repeatable motions that generate pipeline, enable successful customer deployments, and deepen long-term strategic alignment. The ideal candidate understands how large consulting organizations operate, has experience driving revenue through strategic partn
About the Team OpenAI, in close collaboration with our capital partners, is building the world’s most advanced AI infrastructure ecosystem. Our Industrial Compute organization develops and deploys large-scale AI campuses designed to support the next generation of frontier model training and inference workloads. The Hardware Operations team is responsible for ensuring the reliability, availability, and lifecycle health of OpenAI’s compute infrastructure. We partner closely with Data Center Operations, Fleet Health Engineering, Manufacturing, Network Infrastructure, Capacity Planning, and our infrastructure partners to maintain world-class operational performance across rapidly expanding AI environments. As we scale globally, we are building the operational frameworks, reliability standards, and sustaining engineering practices required to support thousands of GPUs and servers across multiple campuses. About the Role We are seeking a Datacenter Hardware Technician Lead to serve as the senior on-site technical authority for hardware reliability and fleet health at one of OpenAI’s flagship AI campuses. This role operates at the intersection of hardware operations, sustaining engineering, and fleet reliability. You will partner closely with Cloud Service Provider operations teams, OpenAI fleet-health engineers, hardware engineering teams, and OEM vendors to identify, diagnose, and resolve hardware issues affecting production systems. Beyond day-to-day operational support, you will drive root cause investigations, reliability improvement initiatives, lifecycle management programs, and operational readiness efforts. You will help establish hardware maintenance standards, operational procedures, and best practices that scale across future OpenAI infrastructure deployments. The ideal candidate combines deep hands-on datacenter hardware expertise with strong troubleshooting, failure analysis, and cross-functional leadership skills. Candidates must be able to sit onsite at our
TEGNA Inc. helps people thrive in their local communities by providing the trusted local news and services that matter most. With 64 television stations in 51 U.S. markets, TEGNA reaches more than 100 million people monthly across web, mobile apps, streaming, and linear television, while also maintaining a strong global presence in India with offices in Bangalore and Chennai that support technology, product, and business operations initiatives. Together, we are building a sustainable future for local news. Senior DevOps Engineer About TEGNA TEGNA Inc. (NYSE: TGNA) helps people thrive in their local communities by providing trusted local news and services. With 64 television stations across 51 U.S. markets, TEGNA reaches more than 100 million people monthly across digital, mobile, streaming, and television platforms. We are focused on innovation, technology excellence, and building scalable solutions that create meaningful impact. Position Overview TEGNA is looking for a highly skilled Senior DevOps Engineer with strong expertise in AWS, Kubernetes, and Infrastructure as Code to design, automate, and manage scalable cloud infrastructure. The ideal candidate will have hands-on experience operating Kubernetes workloads in production, building CI/CD pipelines, and implementing monitoring and security best practices. This role requires deep technical expertise, strong troubleshooting skills, and the ability to work in fast-paced, distributed environments. You will play a key role in ensuring platform reliability, automation maturity, and production stability across cloud-native microservices systems. What You’ll Do Design and manage cloud infrastructure using Infrastructure as Code (AWS CDK, CloudFormation, Terraform). Build and maintain CI/CD pipelines using GitHub Actions and Jenkins to enable automated and reliable deployments. Deploy, manage, and scale Kubernetes clust
About the Team At Trendyol Core Commerce, we build innovative, data-driven strategies that power sustainable growth and global expansion. From seller experience to new market launches, we turn insights into action—fast. Our cross-functional teams shape the future of commerce with bold ideas, real-time impact, and a deep sense of ownership. In a fast-paced, collaborative environment, we grow together — as individuals and as a team. As a Business Deelopment/Category Management Graduate, you’ll step into the dynamic world of e-commerce, supporting our seller partners and contributing to their success on our platform. This hands-on internship within the Category Teams gives you a unique chance to gain real-world experience in seller operations and performance management. You’ll apply your data-driven mindset and strong communication skills to onboard new sellers, expand their product selection, and analyze performance data to deliver valuable insights.
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role As a deployment engineer at WRITER, you'll be at the forefront of expanding human capacity through superintelligence. This isn't just a technical role; it's a deeply influential one where you'll partner directly with our leading enterprise customers. Your expertise will be crucial in uncovering their unique business challenges and architecting AI-powered solutions that leverage our powerful platform and enterprise-grade LLMs. You'll transform complex needs into tangible, high-impact applications, creating champions and driving tangible business results. Your builder's mentality and passion for bringing cutting-edge AI into the hands of real users will directly shape the future of work for some of the world's largest companies. This is a critical role that directly impacts our customers' success and product evolution. You'll contribute significantly to WRITER's mission, working with a dynamic team to push the boundaries of what's possible with generative AI. Thi
Get new ai deployment manager jobs by email
Daily job updates · Unsubscribe anytime