Jobiba hiring network

Production Cleaning Specialist Jobs

3,233 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current production cleaning specialist jobs. Use filters to narrow by work mode, employment type, experience and date posted.

N
1mo ago

The NVIDIA DGXC Data Services team builds cloud-native systems, frameworks, and services for managing data across hybrid and multi-cloud infrastructure. We are building the next-generation data and storage infrastructure to solve some of the hardest problems in AI: storage, access, ingestion, governance, observability, and data management for exabyte-scale, high-performance GPU-based training and inference jobs. Our work gives NVIDIA teams the foundational capabilities they need to build, train, deploy, and operate AI products at scale without reinventing critical data infrastructure for every workload. What you will be doing: Build storage technologies, client libraries, and filesystem frameworks that help AI workloads access data across object stores, file systems, and hybrid cloud infrastructure. Develop high-performance storage paths for training and inference workflows, including data loading, checkpointing, caching, POSIX-style access, and object-store integration. Build observability systems that diagnose storage bottlenecks, attribute GPU idle time to I/O behavior, and expose actionable telemetry through production monitoring stacks. Improve performance, scalability, and reliability of storage systems serving massive datasets, deep directory trees, and high-concurrency AI workloads. Work closely with internal AI teams, platform teams, SRE, and operations to validate storage behavior against real workloads and production environments. Use modern software engineering practices, including AI-assisted and agentic development workflows, while maintaining high standards for design, testing, security, performance, and verification. What we need to see: BS in Computer Science, Information Sys

pythonjavakubernetes
View job →
S
Synthesia
📍 London• Full-time• Remote
1mo ago

Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100. Founded in 2017, the company is headquartered in London, with offices and teams across Europe and the US. As AI continues to shape the way we live and work, Synthesia develops products to enhance visual communication and enterprise skill development, helping people work better and stay at the center of successful organizations. Following our recent Series E funding round, where we raised $200 million, our valuation stands at $4 billion. Our total funding exceeds $530 million from premier investors including Accel, NVentures (Nvidia's VC arm), Kleiner Perkins, GV, and Evantic Capital, alongside the founders and operators of Stripe, Datadog, Miro, and Webflow. About the role As an Applied Research Engineer in our Video team, you will help build the next generation of production-grade foundation models for human-centric video generation. You will join a highly focused team working at the intersection of large-scale generative modeling, distributed systems, and production engineering. Our mission is to develop and optimize video base models that power realistic, controllable, and emotionally expressive synthetic humans at scale. This is not pure research. This is applied research with direct product impact. You will work on advancing training recipes, scaling distributed systems, improving evaluation frameworks, and optimizing inference to ensure our models are high quality, stable, and efficient enough for real-world deployment. Your work will directly influence models used by tens of thousands of businesses worldwide. What you’ll do You will own and execute end-to-end research and engineering projects, from hypothesis to production impact. This includes: Developing and scaling latent video diffusion models tailored for human-centric video generation Designing conditioning mechanisms to improve control (pose, emotion, script, camera) without sacrificing fidelity Advanc

REMOTEpythonawsdocker
View job →
C
1mo ago

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! About North: North is Cohere's cutting-edge AI workspace platform, designed to revolutionize the way enterprises utilize AI. It offers a secure and customizable environment, allowing companies to deploy AI while maintaining control over sensitive data. North integrates seamlessly with existing workflows, providing a trusted platform that connects AI agents with workplace tools and applications. Why This Role? This role offers a unique opportunity to shape how enterprises harness the power of AI in real-world applications. As a bridge between our core North product and our clients’ engineering teams, you’ll be at the forefront of solving complex problems and securely integrating AI into critical sectors such as finance, healthcare, and telecommunications. We’re looking for Software Engineers with Applied AI experience who can own the design, build, and deployment of agentic workflows powered by Large Language Models (LLMs), from early prototypes to production-grade AI agents, to deliver concrete business value in enterprise workflows. You’ll work closely with customers on real-world business problems, often building first-of-thei

REMOTEpythonreactgit
View job →

About the Team OpenAI, in partnership with our capital and technology partners, is building a global network of advanced datacenters to support the most demanding AI workloads. The Infrastructure Quality team ensures that all datacenter systems are manufactured, delivered, and commissioned to the highest standards of quality, reliability, and performance. We work closely with manufacturing partners, general contractors, engineering teams, and operations staff to ensure that every component is delivered ready for installation, startup, and long-term service. Our work spans from vendor qualification through commissioning, ensuring operational readiness across our global portfolio. About the Role We are seeking an experienced Manufacturing Quality Engineer (MQE) to establish, implement, and manage a manufacturing-focused quality program for datacenter infrastructure. This role will be responsible for vendor oversight, quality assurance, process improvement, and issue resolution for all critical systems. You will lead vendor audits, monitor performance metrics, and coordinate corrective actions to ensure predictable delivery schedules, reduced risks, and operational reliability. By partnering with vendors, construction teams, and internal stakeholders, you will help ensure OpenAI’s datacenters are delivered on time and built to the highest operational standards. Travel Domestic and international travel as needed (estimated 40–60%) to manufacturing sites, datacenter locations, and partner facilities. Key Responsibilities Vendor Oversight & Performance Management Conduct manufacturing evaluation, audits, and improve vendor performance across production, inspection, testing, and delivery phases. Develop and track quality metrics to assess manufacturing performance and identify trends. Partner with vendors to refine processes, training, and quality controls to mitigate risks before shipment. Program Development & Execution Develop and maintain a datacenter-focused m

awsrestai
View job →
N
1mo ago

NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology—and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. NVIDIA has a rapidly expanding ecosystem of data center platform & node designs. From single node HGX/DGX systems all the way up to large multi-node NVLink domain rack architectures. These designs have become core to NVIDIA's rapidly growing enterprise and cloud provider businesses. Each bringing together the full power of NVIDIA GPUs, NVIDIA NVLink, NVIDIA InfiniBand networking, NVIDIA Grace CPUs, and a fully optimized NVIDIA AI and HPC software stack. We’re searching for a highly motivated, technical leader to design, drive, and operationalize rack-scale factory and deployment flows for next-generation data center products. The ideal candidate will combine deep systems expertise, decisive technical leadership, and a passion for building reliable, debuggable, and scalable manufacturing and deployment solutions. What you’ll be doing: Lead and drive rack-scale/L11 flows for factory and initial data center deployment. Design and implement end-to-end factory workflows, including firmware flashing sequences, security provisioning, and deployment of software mitigations. Collaborate with data center architects, ODMs, and OEMs to define factory and data center requirements that ensure efficient and reliable production ramp. Champion reliability, debuggability an

O
OpenAI
📍 Washington• Full-time• Remote
1mo ago

About the team The OpenAI for Government team is a mission-driven group bringing frontier AI to the U.S. Intelligence Community (IC) and broader national security enterprise. We partner with intelligence professionals to deploy secure, compliant AI capabilities—including ChatGPT Enterprise, ChatGPT Gov, and the OpenAI API—in mission-aligned environments that meet rigorous security, privacy, reliability, and responsible-AI requirements. As part of the OpenAI for Government team, you will work across IC elements and mission partners—including ODNI, CIA, NSA, DIA, NGA, NRO, and other federal intelligence organizations—to help analysts, collectors, operators, and technical teams apply frontier AI to their highest-priority missions. We accelerate adoption through hands-on support, mission-specific workflows, tailored enablement, and measurable operational outcomes while preserving human judgment, analytic integrity, and protection of sensitive information. About the Role The Strategic Delivery Lead (SDL) for the Intelligence Community will identify, shape, and deliver OpenAI’s highest-impact deployments for U.S. intelligence organizations. You will help IC customers translate mission priorities—including all-source analysis, intelligence production, collection management, open-source and geospatial exploitation, cyber threat analysis, knowledge discovery, and analyst productivity—into secure, technically feasible AI deployments built on the OpenAI API and enterprise products. The role is highly cross-functional and customer-facing: you will partner with Research, Engineering, Go-to-Market, Security, Legal, Privacy, and government mission, acquisition, and technical stakeholders to align on scope, remove delivery obstacles, and communicate progress from working teams to agency leadership. You will own end-to-end execution of technical delivery workstreams led by Forward Deployed Engineers, Researchers, Solutions Architects, and Enablement teams. You will establish mission

REMOTEawsgitrest
View job →
O
Okta
📍 Bengaluru• Full-time
1mo ago

Secure Every Identity, from AI to Human Identity is the key to unlocking the potential of AI. Okta secures AI by building the trusted, neutral infrastructure that enables organizations to safely embrace this new era. This work requires a relentless drive to solve complex challenges with real-world stakes. We are looking for builders and owners who operate with speed and urgency and execute with excellence. This is an opportunity to do career-defining work. We're all in on this mission. If you are too, let's talk. Okta's Identity Governance (OIG) organization is looking for a Staff Engineer to join our Identity Administration team — a team that is directly responsible for one of the most critical and visible workflows in enterprise identity: how employees request, approve, and gain access to the resources they need. Opportunity As a Staff Engineer on the OIG Identity Administration team, you will be involved in development, design, and maintenance of our product to serve enterprise customers at scale. You will involve the development and evolution of our product features — spanning the across different identity entities and personas — and work collaboratively across engineering, product, and design to deliver features that are secure, scalable, and delightful to use. This is a role for someone who thinks in systems, leads with craft, and is energized by the challenge of building enterprise-grade software that millions of users depend on. This is a rare opportunity to join a team where your technical decisions will directly shape how thousands of enterprises manage access governance at scale, where the technical challenges are genuinely hard, and where the work you ship will be seen and felt by real users every day. Job Duties And Responsibilities Take end-to-end ownership of feature areas — from technical design through implementation, testing, deployment, and post-launch monitoring Design, build, and ship high-quality, production-ready feat

javascriptjavareact
View job →
B
Bubble
📍 New York• Full-time• Remote• $200K – $258K/yr
1mo ago

We built Bubble with a clear mission: to empower everyone to create software. Our AI visual development platform lets anyone, from first-time entrepreneurs to enterprise teams, take an idea from prompt to fully-functional, scalable app across web, iOS, and Android. With over 6 million users in more than 100 countries, Bubble is breaking down the barriers to entrepreneurship and innovation worldwide. Our Product Bubble is the only fully visual AI app builder that lets you vibe code without the code to go beyond prototypes and launch real apps to real users. Chat with AI when you want speed, edit directly when you want control. Bubble's visual editor lets you fine-tune any detail, from the design to privacy rules and programming logic, so you're never stuck, even if AI hits its limits. Everything you need comes built in: a unified web and native mobile editor, enterprise-grade hosting, security, database management, and automatic scaling that grows with your business. You can build just about anything on Bubble, and our community is living proof. Mailead grew a $10K investment into a $2M valuation, and Faceless.video went from zero to $1M+ ARR in under a year. People aren't just launching products on Bubble, they're building real businesses. See how Bubble builders are shipping apps that change industries, solve problems, and shape the future here: Inspiring builders, breakthrough apps . Why Join Bubble Now? The rise of AI-generated software has validated everything Bubble has been building toward for over a decade. But pure AI-generated code is fragile, hard to debug, and rarely production-ready. Bubble bridges that gap, combining the speed of AI with a structured visual platform that produces stable, scalable, secure software. The people who join Bubble right now will help define what that means for millions of builders around the world. If you've ever wanted to work on something that genuinely changes who gets to build, this is your moment. About the Developer Expe

REMOTEtypescriptpythonreact
View job →
C
Cohere
📍 Toronto• Full-time• Remote
1mo ago

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! About North: North is Cohere's cutting-edge AI workspace platform, designed to revolutionize the way enterprises utilize AI. It offers a secure and customizable environment, allowing companies to deploy AI while maintaining control over sensitive data. North integrates seamlessly with existing workflows, providing a trusted platform that connects AI agents with workplace tools and applications. In this role, you will: Build and ship features for North, our AI workspace platform Develop autonomous agents that talk to sensitive enterprise data Write and ship minimal code that runs in low-resource environments, and has highly stringent deployment mechanisms As security and privacy are paramount, you will sometimes need to re-invent the wheel, and won’t be able to use the most popular libraries or tooling Collaborate with researchers to productionize state-of-the-art models and techniques You may be a good fit if: Have shipped (lots of) fullstack code (Python and React) in production You excel in fast-paced environments and can execute while priorities and objectives are a moving target You’ve worked in both large enterprises and st

REMOTEpythonreactai
View job →
Q
1mo ago

Our mission and customers: We are creating the freedom for SMEs to succeed by delivering Europe's leading finance workspace with banking at its core, augmented by financial tools. We are proud to be rated 4.8 on Trustpilot, based on 55,000+ reviews. Our culture puts customer satisfaction at the core of what we do, as proven by our Net Promoter Score of 75 (more about our culture here). Our journey: Founded in 2017 by Alexandre and Steve, Qonto has grown to 1,600+ Qontoers serving over 600,000+ customers across 8 European countries. We have been profitable since 2023, and we are just getting started. Our beliefs: We hire for skills and potential. With 80+ nationalities, 45% women, of which 56% of women in our leadership team, diversity isn't a program; It's who we are. We've built a discrimination-free hiring process because the best teams are built on merit. AI at Qonto: AI is deeply embedded in how we work (here) - Every Qontoer gets unlimited access to the best AI tools. We want people who experiment without waiting for permission, push AI beyond the obvious, know when to trust it, and when to question it. ------------------------------------------------------------------------------------------------------ 🌏 Location: You can choose to work in Qonto as long as you're living in (or willing to relocate to) either Germany, France, Italy, Serbia, or Spain. Mission: Join us as a Backend Engineer to build the financial infrastructure that 600,000+ European SMEs depend on every day, from highly scalable APIs to robust banking services that handle real money, in real time, with zero room for mistakes. ➡️ As a Backend Engineer at Qonto, you will Design, build, deploy, and maintain services handling real financial transactions, owning reliability in production, not just at merge time Co-own service architecture, resilience, and scalability with respect to Domain-Driven Design principles Grow in technical leadership: lead design discussions, anticipate risks, and mento

REMOTEpostgresqlawskubernetes
View job →
Q
1mo ago

Our mission and customers: We are creating the freedom for SMEs to succeed by delivering Europe's leading finance workspace with banking at its core, augmented by financial tools. We are proud to be rated 4.8 on Trustpilot, based on 55,000+ reviews. Our culture puts customer satisfaction at the core of what we do, as proven by our Net Promoter Score of 75 (more about our culture here). Our journey: Founded in 2017 by Alexandre and Steve, Qonto has grown to 1,600+ Qontoers serving over 600,000+ customers across 8 European countries. We have been profitable since 2023, and we are just getting started. Our beliefs: We hire for skills and potential. With 80+ nationalities, 45% women, of which 56% of women in our leadership team, diversity isn't a program; It's who we are. We've built a discrimination-free hiring process because the best teams are built on merit. AI at Qonto: AI is deeply embedded in how we work (here) - Every Qontoer gets unlimited access to the best AI tools. We want people who experiment without waiting for permission, push AI beyond the obvious, know when to trust it, and when to question it. ------------------------------------------------------------------------------------------------------ Join us as a Site Reliability Engineer on our Storage team to keep the databases Qonto runs on resilient, safe, and always available. You'll operate and improve our banking-grade storage infrastructure — PostgreSQL, Redis, Kafka and Elasticsearch — and help make it self-serve for backend teams, under the guidance of Damien, our Storage Engineering Manager. ➡️ What you'll do Operate and safeguard critical storage infrastructure: You'll run PostgreSQL, Redis, Kafka and Elasticsearch in production, keeping banking-grade data safe and available. Own incident response and root-cause analysis: You'll be the last line of defense when a hard query or performance problem needs solving. Build our banking-license compliance roadmap: You'll define and prioritize backup a

postgresqlredisai
View job →

About Us Twitch is the world’s biggest live streaming service, with global communities built around gaming, entertainment, music, sports, cooking, and more. It is where thousands of communities come together for whatever, every day. We’re about community, inside and out. You’ll find coworkers who are eager to team up, collaborate, and smash (or elegantly solve) problems together. We’re on a quest to empower live communities, so if this sounds good to you, see what we’re up to on LinkedIn and X , and discover the projects we’re solving on our Blog . Be sure to explore our Interviewing Guide to learn how to ace our interview process. About the Team Twitch Security Platform builds and operates the foundational software, data, and automation that enable security at scale across Twitch. As a Software Development Engineer II (SDE II) on the Security Platform team, you will design, build, and operate critical services, pipelines, and tooling that power Twitch's security, privacy, and compliance programs. About the Role Twitch Security Platform builds and operates the foundational software, data, and automation that enable security at scale across Twitch. As a Software Development Engineer II (SDE II) on the Security Platform team, you will design, build, and operate critical services, pipelines, and tooling that power Twitch's security, privacy, and compliance programs. In this role, you will work at the intersection of software, security, privacy, and data engineering, contributing production systems that handle large-scale security telemetry, automate security workflows, and provide reliable data and services to internal teams. You will partner closely with engineers and product teams to solve real-world security problems through well-designed software. You will own projects end-to-end from design and implementation through deployment and operational support for the systems that are business-critical and highly visible. The problems yo

pythonjavaaws
View job →

As Senior Data Scientist for Engineering Systems you will work independently alongside sharp, generous, and pragmatic engineers from Server Query, Atlas Clusters, and Release Quality, among other teams. Together, we tackle problems spanning resource scaling across the Atlas fleet, safe feature rollout to MongoDB clusters, automated incident response and query engine performance. Join the Platform Data Science team and help us research, prototype and ship machine learning features for MongoDB’s core server, query engine and Atlas, our database-as-a-service cloud offering. We are looking to speak to candidates who are based in Dublin or Cork for our hybrid working model. What You'll Do Partner with Server Query, Atlas Clusters, Release Quality and other engineers to embed algorithmic rigor and optimization into resource scaling, release-safety and monitoring systems across the fleet and inside query engine Deliver production-ready, thoroughly tested statistical and ML algorithms with well-identified limitations that deliver measurable business impact, not just an impressive-sounding methodology Own the full feedback loop: instrument model architecture with the metrics needed to track performance and create dashboards in collaboration with our stellar analytics team, collect feedback from users and metrics to diagnose issues or opportunities, and iterate accordingly Deliver thoughtful, kind code reviews to your peers and act as a core contributor to internal packages, tooling, and team processes that increase developer productivity Measures of Success In 3 months, you’re familiar with our workflow, have an elementary understanding of our product and what teams we work with. You have delivered small-to-medium improvements to our project portfolio In 6 months, you’ve delivered one feature you researched and prototyped from scratch and demonstrated its impact on business metrics of your choice In 12 months, you've established a track record of shipping ML-driven improveme

pythonmongodbaws
View job →

The worldwide data management software market is massive. At MongoDB we are transforming industries and empowering developers to build amazing apps that people use every day. We are the leading modern data platform and the first database provider to IPO in over 20 years. Join our team and be at the forefront of innovation and creativity. MongoDB is seeking a Software Engineer 3 to join the Atlas Clusters Platform team. The team is responsible for building MongoDB Atlas, our database as a service offering and fastest growing product. Atlas allows users to deploy fault-tolerant, secure, globally distributed MongoDB clusters in just minutes. The Atlas Clusters Platform team develops the foundational orchestration platform behind MongoDB Atlas. Our systems drive cluster planning and execution, evolve the Atlas control plane toward service-oriented architecture, and provide critical infrastructure that help Atlas run safely and efficiently across cloud environments. We are looking to speak to candidates who are based in New York City for our hybrid working model. What you’ll do Build and design new features for MongoDB Atlas Contribute to and lead complex technical projects Work closely with product and design teams, considering the user’s perspective while building technical solutions Work with customers and support engineers to fix issues Collaborate with team members to develop our codebase, best practices, and design principles Learn from and mentor other team members We’re looking for someone who Has at least 3 years of professional software development experience Is skilled at writing large-scale, distributed backend systems in a compiled language (Java, C#, Go, etc.) Is comfortable working across the stack of a modern web application (e.g. React, TypeScript, Kubernetes) Has experience with at least one major cloud provider technology (AWS, Azure, GCP) Has led the launch of a new feature and maintained it in production Is eager to solve tough problems Has excellent

typescriptjavareact
View job →
S
1mo ago

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. LEAD. STRATEGIZE. TRANSFORM. We are seeking an advanced professional handling complex enterprise AI/ML deployments, deconstructing system dependencies, and ensuring production robustness. WHY THIS ROLE? This role marks a shift from managing tactical tasks to managing strategic outcomes. You are a seasoned professional with a full understanding of your specialization, resolving a wide range of issues in creative ways. WHAT YOU'LL DO: Design robust, scalable AI/ML solutions utilizing the full Snowflake native stack and partner ecosystem. Perform deep-dive Root Cause Analysis (RCA) for complex system dependencies in AI/ML solutions. Collaborate cross-functionally with Sales and Product teams to align technical roadmaps with customer ROI. Mentor Level 3 architects on best practices for MLOps and architectural design. TECHNICAL DEPTH & RISK MANAGEMENT: Distributed Systems: Deconstruct failures in complex pipelines involving external cloud services (AWS/Azure/GCP). Predictive Failure Analysis: Critically think about potential failure modes like model drift and data skew early in the lifecycle. Governance: Architect data security and access controls specifically for sensitive AI/ML training data. SNOWFLAKE-NATIVE TECH STACK: Snowflake Model Registry, Cortex Functions, Python,

pythonawsazure
View job →
🔔

Get new production cleaning specialist jobs by email

Daily job updates · Unsubscribe anytime