ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE: Baseten’s Model Performance (MP) team is responsible for ensuring the models running on our platform are fast, reliable, and cost‑efficient. As part of this team, you’ll focus on Model APIs — the infrastructure powering our hosted API endpoints for the latest open‑source models. This work spans distributed systems, model serving, and developer experience. You’ll join a small, high‑impact team operating at the intersection of product, model performance, and infra, helping to define how developers interact with AI models at scale. RESPONSIBILITIES: Design, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups Productionize performance improvements across runtimes with deep understanding of their internals: speculative decoding implementations, guided generation for structured outputs, custom scheduling and routing algorithms for high-performance serving Build comprehensive benchmarking frameworks that measure real-world performance across different model architectures, batch sizes, sequence lengths, and hardware configurations Productionize performa
Jobs in United States
Responsable Production in United States
2,944 active opportunities · Updated October 2026
Showing
15 jobs
Explore current responsable production jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We are seeking a Sales Manager to help lead a team of Account Executives within our Startups segment. This hire will be responsible for building and coaching a team of high-performing sales reps, driving revenue growth, and partnering closely with product and engineering to bring our cutting-edge AI infrastructure to customers. RESPONSIBILITIES Lead and mentor a team of Startups Account Executives to consistently exceed pipeline and revenue goals. Help define and execution on go-to-market strategy for our fastest growing customer segment. Hire and scale the team by recruiting, interviewing, and onboarding top talent Collaborate cross-functionally with Marketing, Product, and Engineering to align customer needs with Baseten’s product roadmap. Be deeply engaged with the product, enabling reps to have highly technical conversations with prospects and customers. Foster a culture of accountability, learning, and collaboration within the sales team. REQUIREMENTS 4+ years of closing sales experience, with 2+ years in management leading high-performing teams. Strong technical acumen, ideally with background in AI infrastructure, cloud infrastructure, or developer platforms. Comfortable operating in the weeds with technical products and guiding reps through complex deals. Proven track record of success in high velocity sales environments. Based in San Francisco or New York City and open to coming in office at least 3 d
From $156K/yr
Datadog’s People Technology team is building the future of AI at work. We’re looking for a People Systems Developer to design and deploy AI-native workflows that transform how we hire, develop, and support our employees. This is a high-impact hands-on builder role. You’ll move beyond traditional HRIS configuration to prototype and productionize AI-enabled systems across talent acquisition, onboarding, performance, workforce planning, and internal service delivery. You’ll operate with high autonomy, partner directly with stakeholders, and ship solutions that measurably improve how our People team works. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do Design, build, test, and implement AI-enabled workflows and tools, moving from proofs of concept to full production and adoption based on user feedback. Rapidly prototype using LLMs, APIs, and automation frameworks — and move validated ideas into secure, scalable production systems. Integrate AI capabilities into our ecosystem (Workday, Greenhouse, Slack, Snowflake, Jira, Google Workspace, and more). Partner directly with stakeholders like People Analytics, HRBPs, IT, Legal, and Security to ensure solutions are usable, compliant, and follow responsible AI practices. Evaluate AI tools and vendors through hands-on experimentation. Enable the broader People team to adopt AI effectively and responsibly. Success in this role includes reducing manual People workflows by at least 20% through AI-enabled automation and intelligent system design. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. Who You Are 3-5 years of exper
From $151K/yr
Join the MongoDB Server Query Optimization team, and help us build a world-class distributed open-source query optimizer. Our team plays a crucial role in the experience and performance of data processing. We are responsible for the MongoDB Query Language and the lifecycle of each query, through parsing, optimization and plan selection. We have a presence across the US and Europe including New York, Dublin, Seattle, Palo Alto, and Chicago. We support office-based and remote work and align projects with convenient work hours for each time zone. We have tons of interesting problems to solve with a direct impact on users for transactional, time-series, and analytical workloads. The team is endeavoring to systematically rewrite every major component of our optimization and execution systems. We need your help to design and build the heart of a distributed, flexible schema, document database. This role can be based out of our US offices or remotely in the North America region. Candidate Profile 10+ years of experience in data management systems, distributed systems, or large-scale backend engineering Experience with building production-level code with a large user base, robust design structure and rigorous code quality Degree in Computer Science or similar field, or equivalent practical experience, with strong competencies in data structures, algorithms, and software design/architecture Experience with large code bases written in C++ or another systems programming language. You'll need to trace down defects, estimate work complexity, and design evolution and integration strategies as we rewrite different components of the system A strong foundation in core database internals is essential. While direct experience in query optimization is a massive bonus, it is not a prerequisite. We are also excited to meet candidates with strong backgrounds in compilers, language transpilers, or distributed storage systems Position Expectations Innovate in the area of flexible schema d
From $151K/yr
Join the MongoDB Server Query Execution team, and help us build a world-class distributed open-source database. Our team plays a crucial role in the performance and efficiency of MongoDB's data processing. We are responsible for building and improving the core execution engine that powers all queries, taking a logical query plan produced by the optimizer and turning it into reality. This includes developing the physical operators for data retrieval and manipulation, improving the runtime for complex analytical and transactional workloads, and owning critical components such as our new execution engine. In addition to the core server, we support the query execution needs of other major products like Atlas Streams, Atlas Search and Vector Search, and mongosync, making our work vital to the entire MongoDB ecosystem. You will be joining a globally distributed team with a significant presence in both North America and Europe. While this role is based in the NAMER region, you will regularly collaborate closely with colleagues across different time zones. We support both office-based work in our North America hubs like New York, as well as remote work. We have tons of interesting problems to solve with a direct impact on users for transactional, time-series, and analytical workloads. To meet the ever-increasing data demands of modern applications, we are actively evolving our query system; this includes strategically re-architecting and improving key components of our query execution engine. We need your help to design and build the core of a distributed, flexible schema document database. This role can be based out of one of our North America offices, such as NYC or Palo Alto, or remotely across North America. Candidate Profile 10+ years of hands-on, professional experience in query engine development or database internals Experience with building production-level code with a large user base, robust design structure and rigorous code quality Degree in Computer Science or
About the Team The Cooperative AI team is scaling to devices and embedded operations and user experiences. Our model-powered scaled workforce and knowledge system are moving on to the edge and powering our devices and edge experiences. By leveraging OpenAI’s state-of-the-art models and technologies, in production and in the lab, we develop systems that reason and work autonomously with customers and with our workforce responsible for operational work. We carry real workloads for critical systems across finance, sales, customer support, integrity, product insights, internal operations, and now devices to drive insights into product and industry. We partner closely with internal teams and external customers globally, operating in a hyper-fast feedback loop where many of our users are just a few steps away. This proximity allows us to iterate quickly, validate impact in real time, and accelerate industry impacting learnings and systems builds. We are a highly multidisciplinary, self-contained team focused on transforming the workplace via smart systems, knowledge, scalable and reliable primitives that apply world-class AI capabilities across domains. Our mission is to learn fast and transform how humans collaborate with AI at scale. About the Role We are looking for a Technical Lead Manager to lead a team of engineers building AI-native embedded experiences and operations-forward systems. In this role, you will perform both hands-on technical leadership and small team management. You will drive business outcomes, architecture and technical strategy for complex systems, contribute directly to implementation, and help grow a high-performing team. You will work closely with internal stakeholders to understand operational challenges, identify high-leverage opportunities for automation, and deliver solutions that create measurable impact. This role is ideal for someone who enjoys moving between technical design, coding, mentoring engineers, and working directly with users t
$342K – $445K/yr
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role We are seeking a Technical Lead to lead deployment and operations for OpenAI’s Silicon & Systems team. This person will become the Directly-Responsible Individual responsible for bringing OpenAI’s custom silicon and associated systems into data center environments, ensuring successful deployment, bring-up, validation, operational readiness, and ongoing reliability at scale. This role sits at the intersection of silicon, systems, infrastructure, data center operations, and software. You will lead a team focused on taking new hardware platforms from lab validation into production data center deployment. You will be responsible for building the operational processes, technical workflows, tooling, and cross-functional alignment required to deploy and operate custom AI hardware reliably in OpenAI’s supercomputing infrastructure. The ideal candidate is both a strong leader and a deeply technical operator. You should be comfortable staying close to the technical details of hardware bring-up, fleet deployment, debugging, system validation, data center integration, and production operations. This role requires strong execution, excellent cross-functional judgment, and the ability to drive clarity in ambiguous, fast-moving environments. In this role, you will: Lead a team responsible for deployment and operations of OpenAI’s custom silicon and systems in data center environments Own the path from hardware bring-up and validation through production deployment, operati
About the Team The Privacy Engineering Team at OpenAI is committed to integrating privacy as a foundational element in OpenAI's mission of advancing Artificial General Intelligence (AGI). Our focus is on all OpenAI products and systems handling user data, striving to uphold the highest standards of data privacy and security. We build essential production services, develop novel privacy-preserving techniques, and equip cross-functional engineering and research partners with the necessary tools to ensure responsible data use. Our approach to prioritizing responsible data use is integral to OpenAI's mission of safely introducing AGI that offers widespread benefits. About the Role As a part of the Privacy Engineering Team, you will work on the frontlines of safeguarding user data while ensuring the usability and efficiency of our AI systems. You will help us understand and implement the latest research in privacy-enhancing technologies such as differential privacy, federated learning, and data memorization. Moreover, you will focus on investigating the interaction between privacy and machine learning, developing innovative techniques to improve data anonymization, and preventing model inversion and membership inference attacks. This position is located in San Francisco. Relocation assistance is available. In this role, you will: Design and prototype privacy-preserving machine-learning algorithms (e.g., differential privacy, secure aggregation, federated learning) that can be deployed at OpenAI scale. Measure and strengthen model robustness against privacy attacks such as membership inference, model inversion, and data memorization leaks—balancing utility with provable guarantees. Develop internal libraries, evaluation suites, and documentation that make cutting-edge privacy techniques accessible to engineering and research teams. Lead deep-dive investigations into the privacy–performance trade-offs of large models, publishing insights that inform model-training and prod
About the Team IT Systems Operations serves as the operational control layer connecting Security, Engineering, and Employee Technology platforms. The team ensures employee-facing systems, identity workflows, and enterprise applications operate in a predictable, governed, and continuously reliable manner as the organization scales. Beyond implementation, the team establishes structured operating patterns that ensure platform changes, access models, integrations, and lifecycle workflows evolve safely and consistently across the enterprise environment. About the Role This role acts as an operational owner for identity-connected enterprise SaaS platforms and system change controls supporting OpenAI’s compliance requirements. The engineer will be responsible for ensuring that production configuration, access models, and platform changes remain compliant, auditable, and consistently operated through defined controls. You will own and operationalize controlled system and workflow changes across identity platforms, SaaS applications, collaboration tooling, and enterprise infrastructure. You will partner closely with Security, Platform Engineering, and IT Support Operations to: Ensure identity and access workflows behave consistently across systems Implement structured rollout and configuration practices for enterprise applications Improve visibility and traceability of system changes impacting employee workflows Translate operational requirements into durable automation and policy-aligned implementations Success in this role requires not only strong engineering capability, but also sound operational judgment, disciplined documentation practices, and effective cross-functional collaboration. In this role, you will: Enterprise SaaS & Identity Platform Ownership Own administration and operational stewardship of enterprise SaaS and identity-connected platforms, ensuring configuration integrity, access governance, and compliance with defined control requirements. Own onboard
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role We are seeking a Operations Program Manager (OPM) to serve as the single-threaded operational leader for new hardware introductions (NPI) and production ramps across OpenAI’s AI infrastructure systems. This role combines hands-on execution with strategic ownership. You will be responsible for defining the operating model, aligning cross-functional stakeholders, setting the critical path, making informed tradeoffs, escalating decisively, and ensuring hardware programs deliver on schedule, quality, cost, and scalability. Success in this role requires comfort operating in ambiguity, influencing without authority, and driving alignment across internal teams and external partners—while keeping eyes firmly on long-term system scalability and repeatability. In this role, you will: Strategic & Leadership Ownership Act as the single-threaded owner for operational readiness across NPI and ramp, accountable for outcomes from early bring-up through sustained production Translate OpenAI’s infrastructure strategy and engineering objectives into clear operating plans, execution priorities, and decision frameworks Drive alignment across Engineering, Operations, Strategic Sourcing, Finance, Capacity Planning, and Executive stakeholders by framing tradeoffs, risks, and recommendations Proactively identify inflection points where decisions or investments are required to protect long-term scale, reliability, or cost targets Influence operational strategy with manufacturing par
About the Team The AI Deployment Management (ADM) team enables organizations to turn OpenAI products into real, sustained impact through world-class services execution. Our mission is to help customers successfully adopt and operationalize AI across their organizations. We partner with enterprises to translate the potential of OpenAI’s technology into durable capability - through structured training, technical enablement, and services. By helping customers move from experimentation to production, the ADM team accelerates time-to-value, deepens product adoption, and helps make OpenAI indispensable to how organizations work. About the Role The AI Deployment Manager role is a specialist post-sales enablement role focused on delivering high-impact enablement and adoption services across OpenAI’s product suite. This role is responsible for designing and delivering enablement experiences that support a repeatable adoption framework, driving sustained activation, expanding breadth and depth of usage, and measurable business value across OpenAI’s product suite, including ChatGPT Enterprise and Agents. This role blends strong product fluency, instructional design, and customer advisory. You will lead live workshops, deliver services, and design adoption interventions for audiences ranging from everyday business users to technical practitioners and executive leaders, helping customers understand not just what OpenAI’s products can do, but how to apply them effectively in real world workflows. Success in this role means accelerating customer confidence, increasing product adoption, supporting successful launches of new product capabilities, and helping customers translate product features into tangible outcomes across teams and business functions. You will own outcomes related to activation and sustained usage by shaping how enablement drives measurable customer impact. This role is based in our San Francisco office. We use a hybrid work model of 3 days in the office per week
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role We are seeking an experienced SoC Architect to lead the definition and development of next-generation custom AI silicon for edge deployments. This role will be responsible for shaping the architecture of highly efficient, high-performance SoCs optimized for machine learning inference and on-device intelligence. You will work cross-functionally with internal engineering teams and external ecosystem partners to translate product requirements into scalable silicon solutions, driving execution from concept through delivery. In this role you will: Define the architecture and technical roadmap for custom SoCs targeted for edge applications. Drive system-level tradeoff analysis across compute, memory, interconnect, power, thermal, and cost constraints. Architect energy-efficient ML compute subsystems optimized for inference workloads and real-world deployment environments. Collaborate with internal hardware, software, systems, and product teams to align architecture with platform needs. Partner with external silicon vendors, IP providers, and manufacturing partners to execute development plans. Lead hardware/software co-design efforts to maximize performance per watt and end-to-end system efficiency. Guide implementation teams through microarchitecture, RTL development, validation, and bring-up phases. Operate effectively in agile development environments and help teams deliver against aggressive schedules and milestones. You might thrive in this role if: Proven exper
Where Performance Meets Purpose Join a team that values excellence and innovation, at a company known for its iconic golf brands. At Acushnet Company, your background and experience contribute to creating the best products for dedicated golfers worldwide. Here, your performance has purpose. What You Will Be Doing We’re seeking a creative and detail-oriented Senior Associate 3D Technical Designer to help drive the evolution of digital product creation across apparel, footwear, accessories, and hard goods categories. In this role, you will partner closely with Design, Technical Design, Product Creation, and other cross-functional teams to develop high-quality 3D assets that support product development, visualization, and digital transformation initiatives throughout the organization. You will be responsible for creating realistic, production-ready 3D models and rendered assets that accurately represent product intent, materials, fit, and construction. Through a combination of technical expertise and design collaboration, you will translate product concepts into compelling digital assets while ensuring consistency, quality, and alignment with brand standards. A key component of the role will involve supporting and expanding the organization’s 3D asset libraries through the creation, optimization, and management of models across multiple product categories. You will also facilitate the conversion of 3D assets between platforms and applications, optimize textures and UV maps, and prepare assets for web, animation, and other digital experiences. Working closely with product development and technical design teams, you will coordinate fabric digitization efforts, support garment fit analysis, and help establish standards, workflows, and best
Where Performance Meets Purpose Join a team that values excellence and innovation, at a company known for its iconic golf brands. At Acushnet Company, your background and experience contribute to creating the best products for dedicated golfers worldwide. Here, your performance has purpose. What You Will Be Doing The Customer Service Manager, Repairs and Returns leads a high-performing team responsible for repairs, warranties, returns, and West Coast Returned Goods operations. In this role, you will provide leadership, coaching, and development while ensuring every customer and account interaction reflects the quality and service standards of the Titleist brand. You will oversee repair and return processes, authorize domestic and international golf club credit returns, and resolve complex warranty and customer escalations. As a key partner across the business, you will collaborate with Production, Quality Assurance, Engineering, Sales, Marketing, and Customer Service teams to identify product trends, improve operational performance, and enhance the customer experience. You will also leverage SAP, Cognos, and other business tools to monitor performance, support reporting and forecasting efforts, and drive continuous improvement across service, efficiency, and customer satisfaction. What You Bring High school diploma or equivalent required; bachelor's degree preferred. 5+ years of customer service experience, preferably supporting warranty, repair, returns, or service operations. 2+ years of experience managing a customer service team of 8+ employees. Hands-on SAP experience required. Proven success coaching, developing, and leading high-performing customer service teams. <l
**Relocation Available** Injection Mold Maintenance Mechanic - (2nd Shift, Ball Plant Three)| Titleist (Open)
AcushnetWhere Performance Meets Purpose Join a team that values excellence and innovation, at a company known for its iconic golf brands. At Acushnet Company, your background and experience contribute to creating the best products for dedicated golfers worldwide. Here, your performance has purpose. Responsibilities: -Responsible for process set up and troubleshooting injection molding, blending and grinding processes. Preventive Maintenance for molds, machines, and related automation. -Troubleshoot, repair and maintenance of injection mold machines, molds, and associate equipment including automation, and support equipment including blenders, grinders, conveying equipment, hoists, and hoppers. -Maintain inventory for machine molds, and related equipment. Set up and process of machines and related equipment. -Required to maintain work records (work orders, PM's, Materials and parts) on computerizedMaintenance system. Experience/Requirements: -3-5 years of experience troubleshooting, repairing and maintaining injection molding equipment. -Required to have own trade related tools to perform required tasks. Skills: -Should work well with others and have communication and interpersonal skills. -Must have analytical and problem-solving skills to use with production to identify and correct injection cup, and assembly defects. -Must have a thorough knowledge of mechanical related production equipment. -Must possess a demonstrated ability to make process adjustments to bring parts into specification. -Must have knowledge of electronic equipment controls. -Demonstrated ability to install, repair and maintain injection molding and support equipment. -Demonstra
Other cities to consider
More places hiring for this role
Get new responsable production jobs in United States by email
Daily job updates · Unsubscribe anytime