Jobiba hiring network

Production Operator Jobs

3,235 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current production operator jobs. Use filters to narrow by work mode, employment type, experience and date posted.

S
Sendbird
📍 South Korea• Full-time
1mo ago

Sendbird is building AI agents for customer experience. Our platform already powers billions of conversations every month across chat, voice, video, and messaging APIs. We are now using that foundation to build agents that understand customer context, reason over business data, and take reliable action in production. We are looking for a Machine Learning Engineer to research, build, and productionize new capabilities for those agents. This role sits at the intersection of agent product development, applied AI research, and production engineering. You will work on systems that enterprise customers depend on every day, not demos or isolated prototypes. About Sendbird and delight.ai Sendbird has spent more than a decade building communication infrastructure for in-app chat, voice, video, and messaging APIs. More than 4,000 brands use our platform, including DoorDash, Match Group, Noom, Yahoo Sports, and Rakuten. Our systems support more than 7 billion messages every month. In 2024, we made a strategic shift toward AI-first customer experience. In 2025, we launched our enterprise AI agent product, delight.ai. Delight.ai helps businesses deliver customer support and engagement that is faster, more contextual, and more personal. Unlike simple FAQ bots, our agents are built to remember customer context, use tools, retrieve relevant knowledge, connect across channels, and handle real customer workflows with accuracy and control. The Role As a Machine Learning Engineer, you will design, build, evaluate, and ship new capabilities for our AI agents. You will work across agent architecture, retrieval, memory, planning, tool use, workflow automation, voice, evaluation, data pipelines, model adaptation, inference, and production integration. This is a hands-on engineering role for someone who can turn AI research and product ideas into reliable customer-facing features. Some problems will require training, fine-tuning, or adapting models. Others will require better retrieval, bet

pythonawsazure
View job →
S
1mo ago

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. You independently lead the most complex AI deployments Smartsheet undertakes. You own the full engagement lifecycle from technical discovery through production deployment through solutions org handoff. You architect multi-agent solutions, design client-personalized MCP resource packs, and build the Deployment Kits that transform how 200+ solutions consultants and partners operate. You mentor junior FDEs, drive the intelligence loop, and present field findings at weekly Applied AI strategy sessions. You are building a function, not filling a role. What You Will Do Lead complex, multi-system AI deployments end-to-end scope, architect, build, validate, and manage the customer relationship throughout. Own the AI workshop program for your pod, customize modules per customer, lead technical sessions, translate outputs into production requirements, evolve content from field learning. Architect multi-agent solutions selecting the right coordination pattern for each customer’s workflow characteristics and compliance requirements. Design client-specific and industry-specific MCP resource packs that serve personalized intelligence from the server so every connected AI surface gets smarter for that customer automatically. Own Deployment Kit quality for your pod. If a kit is not documented well enough for a solutions consultant with no engineering background to follow, it isn’t done. Lead Solutions Enablement Sprints: transfer AI deployment patterns to solutions consultants and partners with training materials and certification crite

REMOTEjavascripttypescriptpython
View job →
S
1mo ago

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. You independently lead the most complex AI deployments Smartsheet undertakes. You own the full engagement lifecycle from technical discovery through production deployment through solutions org handoff. You architect multi-agent solutions, design client-personalized MCP resource packs, and build the Deployment Kits that transform how 200+ solutions consultants and partners operate. You mentor junior FDEs, drive the intelligence loop, and present field findings at weekly Applied AI strategy sessions. You are building a function, not filling a role. What You Will Do Lead complex, multi-system AI deployments end-to-end scope, architect, build, validate, and manage the customer relationship throughout. Own the AI workshop program for your pod, customize modules per customer, lead technical sessions, translate outputs into production requirements, evolve content from field learning. Architect multi-agent solutions selecting the right coordination pattern for each customer’s workflow characteristics and compliance requirements. Design client-specific and industry-specific MCP resource packs that serve personalized intelligence from the server so every connected AI surface gets smarter for that customer automatically. Own Deployment Kit quality for your pod. If a kit is not documented well enough for a solutions consultant with no engineering background to follow, it isn’t done. Lead Solutions Enablement Sprints: transfer AI deployment patterns to solutions consultants and partners with training materials and certification crite

REMOTEjavascripttypescriptpython
View job →

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. You independently lead the most complex AI deployments Smartsheet undertakes. You own the full engagement lifecycle from technical discovery through production deployment through solutions org handoff. You architect multi-agent solutions, design client-personalized MCP resource packs, and build the Deployment Kits that transform how 200+ solutions consultants and partners operate. You mentor junior FDEs, drive the intelligence loop, and present field findings at weekly Applied AI strategy sessions. You are building a function, not filling a role. What You Will Do Lead complex, multi-system AI deployments end-to-end scope, architect, build, validate, and manage the customer relationship throughout. Own the AI workshop program for your pod, customize modules per customer, lead technical sessions, translate outputs into production requirements, evolve content from field learning. Architect multi-agent solutions selecting the right coordination pattern for each customer’s workflow characteristics and compliance requirements. Design client-specific and industry-specific MCP resource packs that serve personalized intelligence from the server so every connected AI surface gets smarter for that customer automatically. Own Deployment Kit quality for your pod. If a kit is not documented well enough for a solutions consultant with no engineering background to follow, it isn’t done. Lead Solutions Enablement Sprints: transfer AI deployment patterns to solutions consultants and partners with training materials and certification crite

REMOTEjavascripttypescriptpython
View job →
S
Smartsheet
📍 Munich• Full-time• Remote• €87.5K – €112.5K/yr
1mo ago

For over 20 years, Smartsheet has empowered teams to manage work seamlessly and scale solutions smarter. Now, in our most ambitious chapter yet, we are uniting human teams with AI agents. By orchestrating the work agents do best, automating manual tasks and uncovering insights at scale, we create the space for people to focus on what truly matters: judgment, creativity, and big thinking. That is magic at work, and it’s what we show up for every day. You independently lead the most complex AI deployments Smartsheet undertakes. You own the full engagement lifecycle from technical discovery through production deployment through solutions org handoff. You architect multi-agent solutions, design client-personalized MCP resource packs, and build the Deployment Kits that transform how 200+ solutions consultants and partners operate. You mentor junior FDEs, drive the intelligence loop, and present field findings at weekly Applied AI strategy sessions. You are building a function, not filling a role. You will work remotely from Germany and will be reporting to Sr. Director, Engineering You Will: Lead complex, multi-system AI deployments end-to-end scope, architect, build, validate, and manage the customer relationship throughout. Own the AI workshop program for your pod, customize modules per customer, lead technical sessions, translate outputs into production requirements, evolve content from field learning. Architect multi-agent solutions selecting the right coordination pattern for each customer's workflow characteristics and compliance requirements. Design client-specific and industry-specific MCP resource packs that serve personalized intelligence from the server so every connected AI surface gets smarter for that customer automatically. Own Deployment Kit quality for your pod. If a kit is not documented well enough for a solutions consultant with no engineering background to follow, it isn't done. Lead Solutions Enablement Sprints: transfer AI deployment patterns to so

REMOTEjavascripttypescriptpython
View job →
C
Cloudflare
📍 In Office• Full-time• $150K – $190K/yr
1mo ago

About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. Available Locations: Austin, Texas and New York, New York About the role We're looking for a Video & Motion Producer to own video production for Cloudflare's Growth team, turning product launches, interviews, and new campaign ideas into polished, compelling content that earns the attention of a technical audience. This is a senior IC role for someone who can think strategically and creatively about video content and also do the hands-on

C
Cloudflare
📍 Hybrid• Full-time• Hybrid• $230K – $281K/yr
1mo ago

About Us At Cloudflare, we are on a mission to help build a better Internet. Today the company runs one of the world’s largest networks that powers millions of websites and other Internet properties for customers ranging from individual bloggers to SMBs to Fortune 500 companies. Cloudflare protects and accelerates any Internet application online without adding hardware, installing software, or changing a line of code. Internet properties powered by Cloudflare all have web traffic routed through its intelligent global network, which gets smarter with every request. As a result, they see significant improvement in performance and a decrease in spam and other attacks. Cloudflare was named to Entrepreneur Magazine’s Top Company Cultures list and ranked among the World’s Most Innovative Companies by Fast Company. At Cloudflare, we’re not looking for people who wait for a polished roadmap; we’re looking for the builders who see the cracks in the Internet that everyone else has simply learned to live with. We value candidates who have the instinct to spot a "normalized" problem and the AI-native curiosity to create a solution using the latest tools. Our culture is built on iteration, leveraging AI to ship faster today to make it better tomorrow, while ensuring that every improvement, no matter how small, is shared across the team to lift everyone up. If you’re the type of person who values curiosity over bureaucracy, and that AI is a partner in solving tough problems to keep the Internet moving forward, you’ll fit right in. About the team and the role The Developer Tooling team builds internal tools and platforms that improve the velocity and developer experience of Cloudflare's engineering teams. We own developer productivity (code commit to production), developer insights and ADLC metrics, developer infrastructure (CI/CD, build systems), AI-assisted development, and Cloudflare-on-Cloudflare dogfooding. This team is responsible for AI-assisted development across all

typescriptawskubernetes
View job →
R
Roblox
📍 San Mateo• Full-time• From $295.3K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. With Roblox Ads business growing at a rapid rate, we are building large scale ads machine learning infrastructure to deliver effective performance ads to our users, and more business values to our advertisers. We’re looking for an EM to lead a team of exceptional ML infrastructure engineers, build scalable, reliable, and high-performance infrastructure that powers ML systems across our organization. You’ll operate at the scales of hundreds of billions of engagements, and redefine how we deliver performance ads to hundreds of millions of users. You Will: Lead strategic planning and roadmap execution of scalable production-ready ML systems including model training, data pipelines, feature engineering and model inference. Own the architecture, establish engineering best practices of scalability, reliability, and cost-effectiveness of ML infrastructure (e.g., training, serving, feature). Work closely with data scientists, ML engineers, platform teams, and product stakeholders to design, implement, and operate robust ML platforms that accelerate model development and deployment. Recruit, mentor, and grow a high-performing team of ML infrastructure engineers. You Have: 5+ years of experienc

awsgitmachine learning
View job →
R
Roblox
📍 San Mateo• Full-time• From $326.1K/yr
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Principal Security Software Engineer on the Production IAM team, you will set the technical direction for how identity and access work across Roblox's production infrastructure, from the mTLS-based identity that services use to authenticate to one another, to the privileged access controls that govern how engineers reach production. The team is accountable for Roblox's machine and workload identity platform, its centralized authorization engine, its production access management platform, production PKI and certificate lifecycle, and just-in-time privileged access for engineers. As an individual contributor in Production IAM, you will define multi-year strategy, drive alignment across Roblox Platform, mentor senior and staff engineers, and personally build the hardest parts of these systems. As AI agents become first-class actors in production, you will also help pioneer how they get identity, prove who they are, and receive safely-scoped access. You will Lead the architecture for production identity and access. Define and evolve the end-to-end design for machine, workload, human, and AI-agent identity across our hybrid on-prem and cloud fleet, making secure access invisible when

pythonjavaaws
View job →
R
1mo ago

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. With Roblox Ads & Discovery business growing at a rapid rate, we are building large scale ads machine learning infrastructure to deliver more value to our users and our advertisers. As a Machine Learning Infrastructure Engineer, you’ll build scalable, reliable, and high-performance infrastructure that powers ML systems across our organization. You’ll operate at the scales of hundreds of billions of engagements, and redefine how we deliver performance ads to hundreds of millions of users. You will: You will co-design models and systems, working at the intersection of model architecture and ML infrastructure, partnering closely with core modelers, data and AI infrastructure engineers, and product teams to push the boundaries of large-scale training and serving. Your work will span recommendation, search, and agentic applications, including large transformer architectures, LLMs, generative rankers, and efficient offline and online content-understanding systems. You will investigate model, data, and systems tradeoffs end to end—from data pipelines and distributed training to low-latency inference and production serving. This includes designing efficient KV-cache strategies, applying p

awsgitmachine learning
View job →
T
Twilio
📍 - India• Full-time• Remote
1mo ago

Who we are At Twilio, we’re shaping the future of communications, all from the comfort of our homes. We deliver innovative solutions to hundreds of thousands of businesses and empower millions of developers worldwide to craft personalized customer experiences. Our dedication to remote-first work , and strong culture of connection and global inclusion means that no matter your location, you’re part of a vibrant team with diverse experiences making a global impact each day. As we continue to revolutionize how the world interacts, we’re acquiring new skills and experiences that make work feel truly rewarding. Your career at Twilio is in your hands. . Hiring and how we work We use Artificial Intelligence (AI) to help make our hiring process efficient. That said, every hiring decision is made by real Twilions! Also, while we are a remote-first company, you may be asked to report in person on an ad-hoc basis for team gatherings, functional off-sites or customer meetings. . See yourself at Twilio Join the team as Twilio’s next Senior Manager, Machine Learning on Twilio’s Conversation Memory team. About the job This position is critical to leveraging Twilio’s massive data ecosystem and communication scale to build foundational platform capabilities for the next generation of conversational AI, specifically by owning the development of 0-to-1 platform features that empower developers to build sophisticated, context-aware customer engagement applications. This is a highly technical, hands-on role, requiring a "player-coach" mentality to actively contribute technical expertise, bridge the gap between theoretical research and production-grade experimentation, and provide strategic direction to the team, rather than being purely a people manager. The leader in this role must demonstrate a strong bias for action, establishing a high-performing engineering setup that enables rapid iteration, experimentation, and deployment of models in this 0-to-1 set

REMOTEawsazuregcp
View job →

Applied AI is where Datadog's ambitious AI bets get built and shipped ( Bits Chat , updog ). We sit at the intersection of research and product: turning promising capabilities from Datadog AI Research lab and the research community into production systems that reach real customers. The team builds the foundations for agentic systems capable of operating at scale in complex production environments. Current bets span agents that run autonomously at scale, context and memory layers that make those agents more intelligent over time, and tools that help customers build and validate AI-native services in production. The mandate is to move fast from idea to customer impact, and when a product finds its footing, to set it up for growth. As an Engineering Manager I in Applied AI, you will lead a team of engineers and applied scientists working on one of these challenges. You will define technical direction, run short feedback loops, make deliberate decisions about what to pursue or stop, and work closely with product managers, research teams, and cross-functional partners to ship AI capabilities that matter. At Datadog, we place value in our office culture, the relationships and collaboration it builds and the creativity it brings. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do Lead and develop a team of engineers and applied scientists focused on building the foundations for agents operating at scale Work closely with product managers, research teams, and cross-functional partners to shape the team's bets from initial framing through to broader adoption, with a clear definition of success criteria at each stage Own end-to-end delivery of high-quality AI systems, from early research exploration to production-grade reliability, with high standards for operational excellence, system reliability, and technical quality Navigate the unique challenges of shipping AI-powered products: balancing quali

machine learningaigo
View job →
D
Datadog
📍 Massachusetts• Full-time• From $234K/yr
1mo ago

The ML Observability team builds cutting-edge tools to monitor, explain, and improve AI systems in production, particularly those leveraging Large Language Models (LLMs) and generative AI. We provide robust, scalable observability for AI workloads, including drift detection and model evaluation, and behavior tracing, enabling customers to ship AI with confidence. As a Staff Engineer, you’ll lead the development of new features and foundational capabilities within Datadog’s LLM Observability product. You will shape product direction, drive experimentation, and apply your deep understanding of both AI systems and software engineering to solve open-ended problems in the fast-moving AI landscape. Your work will directly impact how our customers monitor, troubleshoot, and optimize LLM-based applications in production. Join us in building the foundational tools that make AI systems observable, understandable, and reliable in the real world. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Drive design and implementation of LLM observability features. Ideate, prototype, and scale new product features to provide insights and drive improvements for generative AI systems Work cross-functionally with other eng teams, product, UX, and applied science to iterate fast and find product-market fit Develop and extend tools for tracing, evaluating, and debugging LLMs Influence architecture decisions and mentor engineers to build resilient, high-performance systems Stay close to customer pain points and use those insights to guide product and engineering priorities Stay current with industry trends and advancements in machine learning and observability, driving innovation within the team Who You Are: You have a BS/MS/PhD in a Computer Science, Engineering or r

machine learningaigo
View job →
D
Datadog
📍 Massachusetts• Full-time• From $244K/yr
1mo ago

Datadog’s Cloud Networks team designs, builds, and maintains the production network infrastructure that powers everything built on top of our platform across AWS, GCP, Azure, and beyond. In this role, you’ll set technical direction for how we scale our multi-region, multi-cloud network footprint while keeping reliability and performance high. You’ll partner closely with internal teams and Cloud Service Providers to troubleshoot complex connectivity issues, integrate new networking capabilities, and improve the foundations our engineers and customers rely on. This is a high-impact opportunity to drive meaningful improvements in scale, resiliency, and cost efficiency. At Datadog, we place value in our office culture, the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Design, build, and operate cloud network infrastructure across AWS, GCP, Azure, and Neoclouds in a multi-region environment. Own connectivity between clouds, customers, and developers—ensuring scalable, secure, and reliable network paths. Set clear technical direction for expanding data centers and evolving the network while maintaining stability and performance. Improve cross-site and cross-region connectivity patterns to support Datadog’s growing platform needs. Lead deep investigations into latency, packet loss, and connectivity failures – from pcap and path analysis through to escalations with cloud providers that may originate from customer support Identify and deliver network-related efficiency and cost-saving opportunities that positively impact business health. Who You Are: You have deep networking expertise. You understand BGP, route policies, path selection, prefix advertisement, and what breaks in large-scale networking. You have substantial experience designing, building, and evolving large-scale Software-Defined Networks—inclu

awsazuregcp
View job →
D
Datadog
📍 Massachusetts• Full-time• From $156K/yr
1mo ago

Datadog’s People Technology team is building the future of AI at work. We’re looking for a People Systems Developer to design and deploy AI-native workflows that transform how we hire, develop, and support our employees. This is a high-impact hands-on builder role. You’ll move beyond traditional HRIS configuration to prototype and productionize AI-enabled systems across talent acquisition, onboarding, performance, workforce planning, and internal service delivery. You’ll operate with high autonomy, partner directly with stakeholders, and ship solutions that measurably improve how our People team works. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do Design, build, test, and implement AI-enabled workflows and tools, moving from proofs of concept to full production and adoption based on user feedback. Rapidly prototype using LLMs, APIs, and automation frameworks — and move validated ideas into secure, scalable production systems. Integrate AI capabilities into our ecosystem (Workday, Greenhouse, Slack, Snowflake, Jira, Google Workspace, and more). Partner directly with stakeholders like People Analytics, HRBPs, IT, Legal, and Security to ensure solutions are usable, compliant, and follow responsible AI practices. Evaluate AI tools and vendors through hands-on experimentation. Enable the broader People team to adopt AI effectively and responsibly. Success in this role includes reducing manual People workflows by at least 20% through AI-enabled automation and intelligent system design. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. Who You Are 3-5 years of exper

aigorust
View job →
🔔

Get new production operator jobs by email

Daily job updates · Unsubscribe anytime

Explore verified demand

More production operator opportunities

Browse all jobs →

Companies hiring

Employers are derived from current jobs in this exact search market.