About the Team API Enterprise Controls is part of the API Infrastructure organization and owns the platform capabilities that help developers, startups, and enterprises adopt the OpenAI API securely and confidently. We build the systems underneath our APIs and developer platform across authentication and identity, service accounts and key management, secure networking, compliance, auditability, observability, and operational controls. Our users are developers and teams running critical applications on OpenAI, and we partner closely with Product, go-to-market, security, and infrastructure teams to turn their most important needs into reliable, intuitive platform capabilities. About the Role We are looking for an exceptional backend software engineer to help define and ship the enterprise capabilities our API Platform needs to scale.; this is a product-engineering role grounded in deep backend systems. You will work across databases, streaming systems, request routing, authentication, and developer-facing APIs while bringing strong product judgment, developer empathy, and attention to the small details that make a platform easier to understand, trust, and operate. You will lead large cross-functional initiatives, work closely with Product and go-to-market teams, engage directly with sophisticated users, and carry ambiguous needs from discovery through design, launch, and iteration. In this role, you will: Own backend product capabilities end to end across authentication and identity, service accounts and key controls, secure networking, compliance, observability, and operational workflows. Partner with Product, go-to-market, security, infrastructure teams, and sophisticated customers to identify needs, shape the roadmap, and lead large cross-functional projects from design through launch. Design developer-facing APIs, system behavior, configuration, error handling, safe defaults, auditing, and notifications with exceptional care for the details that define a great dev
Jobs in United States
Lead Backend Engineer Ruby 2c Ai Engineering 3a Agent Observability Manager Manager in San Francisco
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current lead backend engineer ruby 2c ai engineering 3a agent observability manager manager jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team OpenAI's research training infrastructure powers how our frontier models are trained and evaluated. The Simulation team sits at the intersection between the agentic harness that powers OpenAI's products and the research infrastructure where GPT-next is trained, ensuring that our model's training environment is as realistic as possible. This team owns the integration layer that connects our production harness capabilities into the training stack. The work is highly cross-functional and high leverage: researchers depend on it to run experiments and evaluations reliably as well as to develop the next generation of harness capabilities. Failures in this surface can materially affect training velocity and correctness. About the Role We're looking for a Principal Software Engineer to lead the architecture and evolution of the Simulation Platform. You'll own a critical interface between research and engineering, building the systems, APIs, and operational patterns that let researchers use agentic coding infrastructure safely and effectively in training environments. This role is ideal for a senior backend or infrastructure engineer with strong technical judgment, product sense for highly technical users, and the ability to drive execution across multiple teams. The highest-leverage work is building robust infrastructure that supports and accelerates research without compromising engineering quality. In this role, you will Design, build, and evolve the integration between the Codex harness that powers OpenAI's products and research training infrastructure used for training GPT-next Build a platform for our LLMs to train and be evaluated in simulated environments that mimic their deployment setting as closely as possible, on every axis: agentic harness, compute substrate, timing, tools, data sources, humans in the loop, and more Own major integration surfaces end-to-end, from architecture and API design through rollout, operations, and long-term maintenance Bu
From $177.2K/yr
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . We're looking for a Staff Software Engineer to lead the technical direction of the backend systems powering Pinterest's AI-driven product experiences — Pinterest Assistant, visual editing and content creation tools and future LLM-based products. You'll design and ship backend systems while architecting the broader platform strategy that enables these experiences to scale across surfaces and teams. This is a hands-on leadership role where you'll move between deep technical execution, system-level architecture and cross-team technical leadership. What you'll do: Define the backend and platform architecture for AI-driven product experiences — visual-chat, AI image generation and editing, and agentic or LLM-based products — partnering with Engineering, Product, ML and UX leaders to shape the technical vision and roadmap. Architect end-to-end systems
About the team OpenAI’s Forward Deployed Engineering team partners with customers to turn research breakthroughs into production systems. We operate at the intersection of customer delivery and core platform development. About the role Forward Deployed Engineers (FDEs) lead complex end-to-end deployments of frontier models in production alongside our most strategic customers. You will own discovery, technical scoping, system design, build, and production rollout, partnering directly with customer engineering and domain teams. You will measure success through production adoption, measurable workflow impact, and eval-driven feedback that changes product and model roadmaps. You’ll work closely with our Product, Research, Partnerships, GRC, Security, and GTM teams. This role is based in San Francisco. We use a hybrid work model of 3 days in the office per week. We offer relocation assistance. Travel up to 50% is required. In this role you will Own technical delivery across multiple deployments from first prototype to stable production Build full-stack systems that deliver customer value and sharpen how we learn Embed closely with customer teams, understand their needs, and guide adoption of what you build Scope work, sequence delivery, and remove blockers early Make trade-offs between scope, speed, and quality; adjust plans to protect delivery Contribute directly in the code when progress or clarity depends on it Codify working patterns into tools, playbooks, or building blocks that others can use Share field feedback that helps Research and Product understand where the models succeed and where they can improve Keep teams moving through clarity and follow-through You might thrive in this role if you Bring 5+ years of engineering or technical deployment experience that includes customer-facing work Have scoped and delivered complex systems in fast-moving or ambiguous environments Write and review production-grade code across frontend and backend using Python, JavaScript,
About the team OpenAI’s Forward Deployed Engineering team partners with customers to turn research breakthroughs into production systems. We operate at the intersection of customer delivery and core platform development. About the role Forward Deployed Engineers (FDEs) lead complex end-to-end deployments of frontier models in production alongside our most strategic customers. You will own discovery, technical scoping, system design, build, and production rollout, partnering directly with customer engineering and domain teams. You will measure success through production adoption, measurable workflow impact, and eval-driven feedback that changes product and model roadmaps. You’ll work closely with our Product, Research, Partnerships, GRC, Security, and GTM teams. This role is based in San Francisco. We use a hybrid work model of 3 days in the office per week. We offer relocation assistance. Travel up to 50% is required. In this role you will Own technical delivery across multiple deployments from first prototype to stable production Build full-stack systems that deliver customer value and sharpen how we learn Embed closely with customer teams, understand their needs, and guide adoption of what you build Scope work, sequence delivery, and remove blockers early Make trade-offs between scope, speed, and quality; adjust plans to protect delivery Contribute directly in the code when progress or clarity depends on it Codify working patterns into tools, playbooks, or building blocks that others can use Share field feedback that helps Research and Product understand where the models succeed and where they can improve Keep teams moving through clarity and follow-through You might thrive in this role if you Bring 5+ years of engineering or technical deployment experience that includes customer-facing work Have scoped and delivered complex systems in fast-moving or ambiguous environments Write and review production-grade code across frontend and backend using Python, JavaScript,
Overview: The Data Acquisition team within the Foundations organization at OpenAI is responsible for all aspects of data collection to support our model training operations. Our team manages web crawling and GPTBot services and works closely with Data Processing, Architecture, and Scaling teams. We are looking for a skilled Software Engineer to join our Data Acquisition team. Responsibilities: Own and lead engineering projects in the area of data acquisition including web crawling, data ingestion, and search. Collaborate with other sub-teams, such as Data Processing, Architecture, and Scaling, to ensure smooth data flow and system operability. Work closely with the legal team to handle any compliance or data privacy-related matters. Develop and deploy highly scalable distributed systems capable of handling petabytes of data. Architect and implement algorithms for data indexing and search capabilities. Build and maintain backend services for data storage, including work with key-value databases and synchronization. Deploy solutions in a Kubernetes Infrastructure-as-Code environment and perform routine system checks. Conduct and analyze experiments on data to provide insights into system performance. Qualifications: BS/MS/PhD in Computer Science or a related field. 4+ years of industry experience in software development. Experience with large web crawlers a plus Strong expertise in large stateful distributed systems and data processing. Proficiency in Kubernetes, and Infrastructure-as-Code concepts. Willingness and enthusiasm for trying new approaches and technologies. Ability to handle multiple tasks and adapt to changing priorities. Strong communication skills, both written and verbal. About OpenAI OpenAI is an AI research and deployment company dedicated to ensuring that general-purpose artificial intelligence benefits all of humanity. We push the boundaries of the capabilities of AI systems and seek to safely deploy them to the world through our products. AI is an
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. The Data Governance team makes sure Plaid handles consumer and customer data responsibly — and can prove it. Our mission is to enforce Plaid's privacy commitments and regulatory obligations in the systems themselves rather than in policy documents: we build the platform and controls that govern how data flows through Plaid — where it lives, who can use it, for what purpose, and for how long. That includes verifiable deletion of consumer data on request, enforcement of data-use restrictions so downstream systems can only use data in permitted ways, and the cataloging and classification that let Plaid know what data it holds and how sensitive it is. We operate at the scale of Plaid's entire data footprint, and correctness and auditability matter to us as much as throughput. As a Staff Software Engineer on Data Governance, you will set the technical direction for how Plaid enforces data governance at scale. You'll lead the design of distributed backend systems that reliably delete, restrict, and track data across dozens of services, making architectural decisions whose blast radius spans the whole company. You'll drive multi-quarter initiatives from ambiguous privacy and regulatory requirements through
About the role We’re looking for an engineering manager to lead a team building software systems that detect and prevent harmful misuse of frontier AI models—before incidents occur. This is a builder’s role: you’ll lead engineers shipping production services, detection pipelines, and mitigation mechanisms that protect frontier model integrity and reduce high-severity misuse risk. While this work intersects with frontier model development, security and risk, we’re explicitly seeking someone with a software engineering foundation who is comfortable building reliable systems that can operate at billions of users scale. In this role you will: Lead a team of software engineers building detection + mitigation systems for frontier model misuse, with an emphasis on model IP protection / distillation detection and emerging risk surfaces from autonomous agents. Set the technical roadmap and execution strategy: prioritize, design, ship, iterate, measure impact. Build production systems: services, pipelines, tooling, instrumentation, and automation that scale with frontier model usage. Partner deeply with Research and Product to translate evolving model capabilities into concrete tests, signals, and mitigations that can be deployed at scale. Drive strong engineering fundamentals: architecture, reliability, monitoring, performance, and operational excellence. Hire and grow an exceptional team across backend, data systems, and applied ML engineering domains as needed. Anticipate what breaks at scale as agentic workflows become more capable. You might thrive in this role if you: Experience building systems in adversarial, fast-evolving environments Are comfortable with ambiguity and novelty Have experience adjacent to security (e.g., abuse prevention, fraud, integrity, platform defense, auth/identity, malware/spam, adversarial environments) Communicate clearly and build trust quickly with senior stakeholders—pragmatic, collaborative, and calm under scrutiny. Significant experience
About the Team At OpenAI, we are dedicated to building safe artificial general intelligence (AGI) to benefit all of humanity. Our mission attracts the world’s top talent in science, engineering, and business to address one of the most ambitious challenges of our times. The Recruiting team is at the heart of this mission, tasked with identifying and hiring exceptional individuals who align with OpenAI's values and cultural ambitions. Our approach to recruitment aims to set the standard for excellence and innovation in the field, connecting outstanding candidates with opportunities to impact the future of AI. About the Role As a Senior Technical Sourcer at OpenAI, you will play a key role in identifying and engaging top-tier engineering talent across a broad range of technical domains. Your focus will be on sourcing exceptional software engineers and technical talent to help build world-class teams advancing our mission in AI research and deployment. In this role, you will: Lead sourcing strategies to identify and engage candidates across a variety of engineering disciplines, including software engineering, backend systems, product engineering, and related technical areas. Develop and maintain a strong pipeline of passive candidates through proactive outreach, research, and networking. Collaborate closely with hiring managers and technical leaders to deeply understand hiring needs and refine sourcing approaches accordingly. Leverage advanced sourcing techniques and tools to identify and attract exceptional talent. Represent OpenAI at industry events and conferences to promote our mission and connect with potential candidates. Maintain accurate and organized candidate data and metrics to inform sourcing strategies and decision-making. You might thrive in this role if you have: 5+ years of experience in technical sourcing, with a focus on engineering or technical roles. A proven track record of successfully sourcing candidates across a range of software engineering and
About the Team The Finance & Supply Chain Engineering organization includes two complementary teams. Software Engineering builds internal full-stack applications, durable agentic workflows, plugins, MCPs, and measurable AI-enabled engineering practices. Data Engineering builds trusted analytics data assets for Finance and Supply Chain. The teams have distinct charters, with important shared dependencies and broad cross team partnerships across Engineering, Applications, Finance, and Supply Chain. About the Role We are looking for a hands-on senior technical leader who will report alongside the Software Engineering and Data Engineering managers. This is an individual-contributor role with no immediate people-management responsibility. The Tech Lead will raise the technical bar across both teams, participate in important cross-team or high-risk design decisions, and directly own and ship high-impact work. The role should improve team judgment and autonomy rather than act as a floating architect or universal approval gate. In this role, you will: Partner with the Software Engineering and Data Engineering managers as a peer technical leader; managers retain accountability for people, staffing, priorities, performance, and delivery commitments. Directly own the architecture, implementation, launch, and operation of one or more high-impact initiatives, remaining accountable for real outcomes rather than advisory output alone. Guide important design decisions that are cross-team, difficult to reverse, or material to security, financial controls, reliability, data quality, or long-term cost of ownership. Establish pragmatic engineering standards across architecture, APIs and data contracts, testing, security, reliability, observability, lineage, data quality, and operational ownership. Advance engineering standards for building with AI, including agentic workflows, evaluation, telemetry, adoption, and outcome measurement. Work across backend, full-stack, data, and big-d
About the Team The Product & Platform teams at OpenAI are responsible for delivering the company’s most impactful offerings—such as ChatGPT, our API platform, and new enterprise capabilities—to a global and diverse customer base. These systems must perform at scale and deliver exceptional experiences to developers, consumers, and businesses alike. The ChatGPT engineering org builds and operates the systems that bring product improvements to users across backend services, web, mobile, and desktop platforms. The Developer Velocity team partners with product engineering, platform, infrastructure, reliability, engineering acceleration, and observability teams to make everyday development faster and releases safer, more predictable, and easier to operate. About the Role We are seeking a Technical Program Manager to improve developer velocity and deployment excellence across ChatGPT. You will lead durable improvements to local development, CI, testing, build systems, release trains, progressive rollout, and post-deployment validation. You will identify the highest-leverage sources of engineering friction, align teams around shared standards and metrics, and drive adoption of tooling and workflows that improve both speed and reliability. This role combines systems thinking, technical program leadership, and hands-on operating rigor across a broad engineering surface. This role is based in San Francisco, CA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Own the cross-functional roadmap for improving local development, CI, testing, build workflows, and release infrastructure. Create a durable intake and prioritization mechanism for developer friction, using evidence to focus teams on the highest-impact improvements. Lead programs that improve deployment speed and safety, including pre-merge confidence, progressive rollout, release guardrails, rollback readiness, and post-deploy valida
About the Team The Premium team owns some of the highest-leverage customer-facing levers in ChatGPT’s consumer revenue business, spanning the paid customer journey: helping users understand the value of paid plans, convert with confidence, and continue finding lasting value in their subscription. Our work is highly cross-functional, partnering with Product, Data Science, Design, FinEng, Finance, Legal, Support, and Marketing to improve free-to-paid conversion, renewal, customer lifetime value, and revenue while keeping the experience trustworthy, scalable, and low-friction. In This Role, You Will: Lead and scale an engineering team responsible for some of ChatGPT’s most important subscription and monetization experiences. Own the technical execution for Premium customer experiences across plan merchandising, paywalls, upgrade flows, checkout UX, plan management, renewals, downgrades, and cancellation. Partner with Product and Data Science to run high-quality experiments across upgrade, trial, renewal, downgrade, and cancellation flows. Improve key subscription metrics including conversion, renewal, churn, ARPU, and lifetime value. Build reliable customer-facing Premium experiences for purchase, plan management, renewal, downgrade, cancellation, and access-related states at scale. Partner closely with FinEng and other platform teams to evolve the billing, payments, and entitlement capabilities that power Premium experiences. Collaborate closely with Product, Design, Data Science, Finance, Legal, Support, and Marketing on monetization strategy and execution. You Might Thrive in This Role If You: Have 5+ years of engineering management experience, Have strong technical expertise in backend, frontend, or full-stack development, with experience building growth-oriented features. Have a track record of improving conversion, retention, or monetization through experimentation and data-driven product engineering. Are experienced with subscription products, plan merchandising
About the Team The ChatGPT Model Flywheel team unified goal is to transform model advancements into great ChatGPT user experiences through reliable serving, rapid experimentation, safe deployment, and continuous improvement. Team Focus Areas Model Experimentation: Enable rapid, safe model validation for ChatGPT and Codex products through experiment automation and lifecycle management. Model Deployment: Ensure safe, scalable deployment of model capabilities with robust rollout and operational tooling. Automate capacity management and incorporate platform-wide health monitors. Model Measurement: Build comprehensive evaluation and measurement systems for model quality, from user signals to launch scorecards. Improve end-to-end feedback loops for continual model improvement. Key Partnerships Collaborate cross-functionally with teams including Model Measurement DS, Research, Codex, Fleet, Inference, and API. In this role, you will: Elevate and consolidate ChatGPT’s harness, context management, and system prompt frameworks. Drive expansion and improvement of multi-tier model experiences. Support and scale self-serve experiment capabilities and automated guardrails. Lead model rollout automation, capacity management, and health monitoring. Shape end-to-end measurement systems (evals, grader signals, user feedback, etc.). You might thrive in this role if you have: Proven experience leading engineering teams in complex, cross-functional environments. Demonstrated success shipping production systems at scale (ideally for AI or large backend services). Deep understanding of model-driven product development, deployment lifecycle, and measurement tooling. Excellent communication and collaboration skills—experience interfacing directly with engineering, research, and product stakeholders. Prior involvement with large language models, distributed infrastructure, or experimentation platforms is a plus. Why Work With Us Tackle highly impactful technical challenges at the cutting edg
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About this role This is a chance to help shape the next chapter of WRITER's product at a moment when generative AI is redefining how work gets done. As senior staff product manager, you'll report directly to Wil Pong, VP of Product, and operate as a force-multiplier across the product organization — setting direction where it's unclear, aligning stakeholders, and making sure our most critical initiatives ship with excellence. You'll own ambiguous, high-impact problem spaces that span multiple product areas and teams, and you'll be equal parts strategist and operator. One moment you're shaping a multi-quarter roadmap; the next you're rolling up your sleeves to unblock execution. We're looking for someone who thrives at the intersection of customer empathy, technical depth, and business judgment — and who is energized by turning complex, evolving AI capabilities into products customers love. 🦸🏻♀️ Your responsibilities Drive high-impact product initiatives end-to-end, from di
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role Writer is the enterprise AI platform for agentic work, purpose-built for the Fortune 500 — and our platform is expanding fast. As a product marketing lead, you'll own the go-to-market strategy for new products and features in our Core app, helping prospects understand how WRITER empowers marketing and sales teams to transform their work, and driving new feature adoption among current customers. You'll sit at the intersection of product, sales, and marketing — translating complex platform capabilities into crisp messaging that resonates with enterprise buyers. 🦸🏻♀️ What you’ll do Own end-to-end product marketing for WRITER Agent and other key features of our core app, including positioning, messaging, launch strategy, and competitive differentiation – translating new agentic capabilities into clear enterprise value Partner with product, demand generation, content, comms, and customer marketing teams to drive go-to-market for product and feature launches Trans
Other cities to consider
More places hiring for this role
Get new lead backend engineer ruby 2c ai engineering 3a agent observability manager manager jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime