Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why This Role Is Different This is not a typical “Applied Scientist” or “ML Engineer” role. As a Member of Technical Staff, Applied ML, you will: Work directly with enterprise customers on problems that push LLMs to their limits. You’ll rapidly understand customer domains, design custom LLM solutions, and deliver production-ready models that solve high-value, real-world problems. Train and customize frontier models — not just use APIs. You’ll leverage Cohere’s full stack: CPT, post-training, retrieval + agent integrations, model evaluations, and SOTA modeling techniques. Influence the capabilities of Cohere’s foundation models. Techniques, datasets, evaluations, and insights you develop for customers will directly shape the next generation of Cohere’s frontier models. Operate with an early-startup level of ownership inside a frontier-model company. This role combines the breadth of an early-stage CTO with the infrastructure and scale of a deep-learning lab. Wear multiple hats, set a high technical bar, and define what Applied ML at Cohere becomes. Few roles in the industry combine application, research, customer-facing engineeri
Jobs in United States
Responsable Production in United States
2,943 active opportunities · Updated October 2026
Showing
15 jobs
Explore current responsable production jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
From $192K/yr
As Engineering Manager for Threat Detection, you will lead a high-performing team that powers Datadog's detection program. Threat Detection is the organization responsible for keeping Datadog ahead of an evolving threat environment: closing coverage gaps faster, raising the bar on signal quality, and shipping detections that hold up under the scale and complexity of cloud-native infrastructure. Your team will combine direct detection expertise, platform engineering, and applied AI to ship detections at a pace and scale traditional rule-writing alone cannot match. Examples of what your team will work on include detection-authoring agents, the detection platform that powers every rule in production, coverage analysis, alert triage and response automation, and the evaluation infrastructure that holds these systems to a high bar of fidelity. Detection authorship is a shared responsibility across the organization, and your team will contribute both by building the systems that scale our authoring capacity and by writing detections directly when their domain expertise is the right tool. You will partner closely with our Security Incident & Response Team (SIRT), Cyber Threat Intelligence (CTI), AI Engineering teams, and Datadog's broader Security organization. This is a high-impact leadership role: you will grow a team of security and software engineers responsible for building and executing our detection and AI strategy. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Lead the strategy, roadmap, and execution of Datadog Security's shift to AI-accelerated detection and response. Drive development of high-fidelity detections as a shared responsibility across the organization, ensuring your team's systems and direct contributions raise the bar on coverage and
From $90K/yr
MongoDB Atlas is the premier multi-cloud database-as-a-service built and operated by the makers of MongoDB. The Cloud Operations Engineering team at MongoDB is a worldwide team responsible for the consistent operational success of every MongoDB Atlas customer. As a Cloud Operations Engineer, you will help ensure the success of our Atlas customers, whether they are early startups or large multinational companies, cloud-native or just getting started with a digital transformation to the cloud. You are excited about the core mission of MongoDB, and the opportunity to join the team responsible for operating Atlas, the fastest-growing multi-cloud database-as-a-service in the world. You are prepared to be one of the founding members of a 24/7/365 global cloud operations team. Cloud Operations Engineers will be responsible for day-to-day duties such as creating and monitoring systems alert dashboards, reviewing critical event and system logs, accessing customer instances that underpin their production databases and performing server administration duties including performance troubleshooting. Applicants must be critical thinkers who are quick to detect, resolve, or escalate issues that are sometimes broad in scope and difficult to trace. FedRamp engineers are specifically tasked with supporting our government customers in our FedRamp Atlas environment. This includes SLED (State and Local Government and Education), various federal agencies, and other customers that leverage FedRamp. At MongoDB you will grow your career and skills, wear multiple hats, and be part of an operations team that works at the frontier of Cloud services and database systems. This role will be based remotely in Colorado. Responsibilities Successfully coordinate with a global team of Cloud Operations Engineers who are tasked with ensuring our uptime guarantees to our Atlas customer base Help scale the worldwide Cloud Operations Engineering team with the strategic implementation of new processes and to
About the Team OpenAI’s Hardware organization develops silicon and system-level solutions designed for the unique demands of advanced AI workloads. The team is responsible for building the next generation of AI-native silicon while working closely with software and research partners to co-design hardware tightly integrated with AI models. In addition to delivering production-grade silicon for OpenAI’s supercomputing infrastructure, the team also creates custom design tools and methodologies that accelerate innovation and enable hardware optimized specifically for AI. About the Role We're looking for an Optical Interconnect System Engineer to design, qualify, and deploy scalable optical connectivity for large-scale AI infrastructure. This role spans fiber-system architecture, optical-mechanical integration, validation, reliability, deployment, and serviceability. You will work with optical, mechanical, electrical, networking, manufacturing, reliability, and data-center teams to translate system needs into practical interconnect solutions. This is a hands-on role for someone who can connect design decisions with installation, qualification, troubleshooting, and long-term operational performance. In this role, you will: Define optical interconnect architectures and requirements across hardware platforms and rack-level systems. Design high-density fiber systems for performance, density, reliability, installation, and serviceability. Lead optical-mechanical integration and cross-functional design reviews. Develop test and qualification plans for optical components, modules, switching platforms, and integrated systems. Own optical loss budgets, routing guidelines, handling requirements, and serviceability criteria. Support system bring-up, deployment, troubleshooting, failure analysis, and reliability improvement. Create reusable design guidelines, interface requirements, and qualification methods. You might thrive in this role if you have: Core experience Experience desi
About the Team API Multimodal builds the developer-facing products and infrastructure that bring OpenAI’s image, audio, and real-time model capabilities into the world. We are responsible for high-scale APIs for image generation, speech transcription, speech generation, and low-latency voice interactions. We partner closely with Research and Inference to bring frontier model capabilities to developers and use customer feedback to improve our models. About the Role As a software engineer on API Multimodal, you will build and operate the products and distributed systems behind OpenAI’s image, audio, and real-time APIs. You will work across model integration, API design, and production infrastructure to turn new research capabilities into reliable developer experiences. This hands-on role combines backend and systems depth with product judgment: you will own projects end to end, partner with Research, Inference, and Safety, and help make multimodal AI useful at scale. Model training experience is not required. In this role, you will: Design, build, and ship developer-facing APIs and backend services that serve frontier models. Architect low-latency streaming, request, session, and model integration systems that make complex multimodal interactions reliable and intuitive at scale. Work directly with Research to bring new model capabilities into production, shape the systems around them, and incorporate feedback from real-world developers and customers. Own the availability, latency, scalability, and cost efficiency of the services you build. Own projects from technical design and implementation through launch and ongoing iteration, while raising the team’s engineering standards. Your background might look something like: 7+ years of professional experience, excluding internships, in backend, infrastructure, platform, or product engineering roles. A track record of designing, building, and operating production backend services, developer-facing APIs, or distributed syste
About the Team OpenAI's Professional Services team helps organizations move from AI ambition to durable production outcomes. We partner with customers on complex deployments and build the strategy, commercial models, operating mechanisms, and delivery capacity needed to realize value from OpenAI's models and products. The team works across Go-to-Market, Forward Deployed Engineering, Technical Success, Finance, Product, Legal, Data, Systems, and delivery partners. We are building a services motion that is customer-centered, commercially rigorous, operationally scalable, and deliberately connected to product adoption. About the Role We are seeking a senior GTM Strategy & Operations professional to build and scale the operating system for OpenAI's Professional Services business. This is a foundational, hands-on individual contributor role at the intersection of business strategy, finance, go-to-market, and delivery. You will turn ambiguous questions—what we offer, how we price and package it, how we plan capacity, and how we measure performance—into clear decisions and repeatable mechanisms. You will own business planning, pricing and packaging, forecasting and modeling, management reporting, and cross-functional strategic initiatives. You will build integrated views of demand, staffing, revenue, margin, and delivery performance; replace one-off analyses with durable processes; and create operating cadences that help leaders act early. You will be a trusted partner to Professional Services, GTM, FDE, Technical Success, Finance, Product, Legal, Revenue Operations, and Data leaders. The right person combines direct Professional Services judgment with rigorous analytics, executive communication, and the willingness to build the model, process, or dashboard themselves. You’ll be responsible for: Define and drive the business and GTM strategy for Professional Services, including target customer needs, offer portfolio, positioning, pricing and packaging, partner motions,
About the Team We’re hiring a Developer Productivity engineer to support OpenAI’s Inference Runtime teams. These teams own the systems responsible for serving models reliably, efficiently, and safely across Codex, ChatGPT, API, and internal research workloads. We’re hiring a Developer Productivity Engineer to help scale the engineering systems, safeguards, and developer workflows that enable our teams to move quickly without compromising reliability or performance. This role sits at the intersection of developer experience, CI/CD infrastructure, release engineering, production readiness, and inference systems reliability. You’ll work on the tooling and operational foundations that support model launches, inference optimizations, cloud provider integrations, and large-scale deployments across a rapidly evolving inference stack. About the Role We’re looking for an autonomous, high-ownership engineer who cares deeply about making other engineers faster, safer, and more confident. A major focus of this role will be improving the tooling and infrastructure around deploy gates for inference engine images. These systems help ensure that every image released to production and research is correct, numerically sound, free of regressions, and performant across key metrics like time-to-first-token (TTFT) and time-between-tokens (TBT). You’ll help harden the systems that catch issues before they reach production, reduce noise from flaky or infrastructure-related test failures, and improve automation around triage, ownership, debugging, and escalation when failures occur. You’ll also work on improving observability, rollout safety, release automation, and developer self-service tooling across a rapidly evolving inference stack. This is not generic internal tools work. The systems you build directly impact OpenAI’s ability to support new model launches, safely ship inference optimizations to the world, onboard new infrastructure providers, and operate one of the largest and most p
About the Team The Applied Engineering team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses. You’ll join the team responsible for running the core infrastructure that supports products like ChatGPT and the API. The systems we support include our kubernetes clusters, infrastructure deployment, our networking stack, cloud abstractions, and more. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the Role The cloud infrastructure team builds and maintains infrastructure abstractions allowing OpenAI to ship products quickly and scalably. This role is based in San Francisco, CA. In this role, you will: Design and build the development and production platforms that power our products, enabling reliability and security at scale Ensure our infrastructure can scale to the next order of magnitude Help create a diverse, equitable, and inclusive culture that makes all feel welcome while enabling radical candor and the challenging of group think Like all other teams, we are responsible for the reliability of the systems we build. This includes an on-call rotation to respond to critical incidents as needed. You might thrive in this role if you: Have 5+ years building core infrastructure Have experience operating orchestration systems such as Kubernetes at scale Have experience building abstractions over cloud platforms Take pride in building and operating scalable, reliable, secure systems Are comfortable with ambiguity and rapid change This role is exclusively based in our San Francisco HQ. We offer relocation assistance to new employees. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely depl
About the Team OpenAI’s Hardware organization develops silicon and system-level solutions designed for the unique demands of advanced AI workloads. The team is responsible for building the next generation of AI-native silicon while working closely with software and research partners to co-design hardware tightly integrated with AI models. In addition to delivering production-grade silicon for OpenAI’s supercomputing infrastructure, the team also creates custom design tools and methodologies that accelerate innovation and enable hardware optimized specifically for AI. About the Role As an Engineer on our hardware optimization and co-design team, you will co-design future hardware from different vendors for programmability and performance. You will work with our kernel, compiler and machine learning engineers to understand their unique needs related to ML techniques, algorithms, numerical approximations, programming expressivity, and compiler optimizations. You will evangelize these constraints with various vendors to develop and influence future hardware architectures towards efficient training and inference on our models. If you are excited about efficiently distributing a large language model across devices, dealing with and optimizing system-wide/rack-wide networking bottlenecks and eventually tailoring the compute pipe and memory hierarchy of the hardware platform, simulating workloads at different abstractions and working closely with our partners, this is the perfect opportunity! This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. Key Responsibilities Co-design future hardware for programmability and performance with our hardware vendors Assist hardware vendors in developing optimal kernels and add support for it in our compiler Develop performance estimates for critical kernels for different hardware configurations and drive decisions on compute core and memory h
About the Team OpenAI’s Infrastructure organization builds the systems that power frontier AI workloads at global scale. As compute demand accelerates, our ability to rapidly convert infrastructure investments into usable production capacity has become mission critical. The CPU / Storage / PoP / WAN team is responsible for the end-to-end infrastructure layers required to bring compute online: server and cluster activation, storage platforms, Points of Presence (PoPs), backbone connectivity, and global network expansion. We operate across first-party facilities, colocation environments, and strategic cloud partners to ensure OpenAI can scale reliably and quickly. About the Role We are seeking a highly technical Program Manager to lead execution across CPU, Storage, PoP, and WAN infrastructure programs that directly unlock OpenAI’s next generation compute capacity. In this role, you will own complex cross-functional programs spanning compute cluster activation, storage deployment, PoP bring-up, and backbone expansion. You will coordinate hardware readiness, site readiness, network pathing, storage availability, vendor execution, and engineering dependencies required to turn contracted infrastructure into live training and inference capacity. This role requires strong technical fluency across hardware systems, network infrastructure, storage architecture, and deployment execution. You should be comfortable operating from rack-level implementation details through executive-level capacity planning discussions. This role is based in San Francisco, CA, with travel as needed. Key Responsibilities Lead end-to-end execution of CPU / GPU cluster activation programs across OpenAI’s global infrastructure footprint Drive readiness to convert contracted compute capacity into schedulable production clusters Own deployment programs for new PoPs, backbone nodes, WAN expansion, and interconnection initiatives Build integrated schedules spanning procurement, logistics, installation, st
The Best Players Need the Best People. The Graphics Operator is responsible for the live execution and creative development of insert graphics across all studio and remote productions at PGA TOUR Studios. This role requires technical expertise in broadcast graphics systems, precision under pressure, and a strong understanding of live sports workflows. The Graphics Operator ensures that all visual elements—from real-time scoring to branded content—are delivered with accuracy, consistency, and creative impact. QUALIFICATIONS Bachelor’s Degree or related technical education required. Minimum 3 years broadcast television or related experience at the regional to large market or network level. Experience with Chyron and Ross XPression, or similar platforms, including data integration and real-time graphics workflows. Strong understanding of live production timing, control room operations, and cue execution under pressure. Ability to build, edit, and manage graphic templates, lower thirds, and sponsor elements with editorial accuracy. Skilled in troubleshooting technical issues with graphics systems during live broadcasts. Excellent communication and collaboration skills with producers,
Manufacturing Engineer (Associate or Experienced) Company: The Boeing Company Boeing Defense, Space & Security (BDS) Laser and Electro-Optical Systems (LEOS) site is seeking an Experienced Manufacturing Engineer for our advanced development and production manufacturing team located in Albuquerque, NM. The Boeing Laser and Electro-Optical systems (LEOS) group works to develop and implement next-generation technologies and products serving a variety of commercial and military customers. Our team focuses on rapid prototyping and precision production of advanced laser and electro-optical systems that satisfy challenging product performance requirements built in accordance with aerospace (AS9100) quality standards. Position Responsibilities: Design, develop and optimize manufacturing processes and tooling approaches for complex electro-optical aerospace parts and assemblies Collaborate with design engineers to ensure manufacturability and cost-effective production Review engineering drawings and provide comments to responsible engineer Participate in Integrated Product Teams (IPTs) to integrate technical solutions across multiple disciplines Participate in supplier selection and evaluation to ensure adherence to quality and delivery standards Write content for work instructions, travelers, and standard shop procedures Utilize electronic work instruction, material control and non-conformance management systems as needed to support production Ensure compliance with aerospace industry quality standards (e.g., AS9100, J-STD) Ensure the manufacturing work instructions and special processes follow safety and environmental regulations Assist in shop layout and
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. The PBM Batch Operations role is responsible for providing 24x7 operational support for Pharmacy Benefit Management (PBM) production processing environments. The position monitors, controls, and supports enterprise batch workloads, mainframe systems, iSeries environments, and associated operational processes to ensure critical pharmacy and business applications execute successfully and on schedule. Schedule: WorkDays: TBD Hours: 7 AM to 7PM AZ time Shift Structure: Three 12-hour shifts Additional Requirement: Must be available to work overtime as needed to provide coverage PBM Batch Operations Functional Responsibilities The PBM Operations environment includes responsibility for: Monitoring and supporting IWS (IBM Workload Scheduler) batch processing. Batch job interventions (restart, hold, kill, force complete). Mainframe IPL support. Mainframe console monitoring across multiple LPARs. RxClaim and iSeries batch monitoring. PBM Disaster Recovery support. Vendor escort activities and data center operational support. Data center security ticket processing. MIR3 paging and incident notifications. ServiceNow ticket management. Procedure verification and operationa
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. The PBM Batch Operations role is responsible for providing 24x7 operational support for Pharmacy Benefit Management (PBM) production processing environments. The position monitors, controls, and supports enterprise batch workloads, mainframe systems, iSeries environments, and associated operational processes to ensure critical pharmacy and business applications execute successfully and on schedule. Schedule: WorkDays: Wednesday through Saturday Hours: 7:00 PM to 5:00 AM AZ time Shift Structure: 10-hour shifts Additional Requirement: Must be available to work overtime as needed to provide coverage PBM Batch Operations Functional Responsibilities The PBM Operations environment includes responsibility for: Monitoring and supporting IWS (IBM Workload Scheduler) batch processing. Batch job interventions (restart, hold, kill, force complete). Mainframe IPL support. Mainframe console monitoring across multiple LPARs. RxClaim and iSeries batch monitoring. PBM Disaster Recovery support. Vendor escort activities and data center operational support. Data center security ticket processing. MIR3 paging and incident notifications. ServiceNow ticket management. Procedure ver
We are seeking an experienced Quantitative Developer to join our Markets Quantitative Analytics team, partnering closely with Quantitative Analysts, Traders, and Technology professionals to build the next generation of pricing, risk, and analytics platforms. This is a hands-on technical role for a highly skilled software engineer with a passion for quantitative finance. You will be responsible for designing and delivering high-performance, scalable solutions that support front office trading businesses across asset classes. The role offers the opportunity to work on complex quantitative challenges, modern engineering practices, and large-scale distributed systems while helping shape the strategic direction of Citi's quantitative technology platform. Successful candidates will combine strong software engineering expertise with an understanding of quantitative methodologies and financial markets, translating sophisticated mathematical models into robust, production-grade solutions. Key Responsibilities Design, develop, and maintain high-performance pricing, risk, and analytics libraries used across Global Markets. Partner with Quantitative Analysts to transform research models and prototypes into scalable, production-quality software. Build and optimize quantitative applications using modern C++ and Python, applying strong software architecture and engineering principles. Own the full software development lifecycle, including requirements gathering, design, implementation, testing, deployment, and ongoing support. Drive engineering excellence through CI/CD adoption, automated testing, code reviews, and software quality best practices. Develop and maintain market data platforms and data pipelines supporting analytics, pricing, and risk workflows. Work with infrastructure teams to leverage distributed computing, cloud technologies, and scalable arc
Other cities to consider
More places hiring for this role
Get new responsable production jobs in United States by email
Daily job updates · Unsubscribe anytime