ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We’re seeking a GPU Kernel Engineer to join our team at the cutting edge of AI acceleration, where your code directly impacts the performance of state-of-the-art machine learning models. As a GPU Kernel Engineer, you'll craft the foundation that powers modern AI workloads, optimizing every microsecond of computation to enable breakthrough applications. You'll work in a fast-paced, intellectually stimulating environment where technical excellence is paramount and your contributions directly influence production systems serving millions of users across numerous products. This role offers exceptional growth potential for engineers passionate about low-level optimization and high-impact systems work. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Model Performance team: Baseten Embeddings Inference: The fastest embeddings solution available The Baseten Inference Stack Driving model performance optimization RESPONSIBILITIES Core Engineering Responsibilities Design and implement high-performance GPU kernels for key ML operations, including matrix multiplications, attention mechanisms, and mixture-of-experts routing Write and optimize code using CUDA, PTX assembly, and architecture-specific techniques Apply advanced performance optimization methods such as memory coalescing, warp-level programming, tensor core acceleration, and compute/memory overlap Performance & Innovation Impl
Jobs in United States
Codex Deployment Engineer in San Francisco
181 active opportunities · Updated October 2026
Showing
15 jobs
Explore current codex deployment engineer jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is seeking talented and experienced Software Engineers to join our Platform team within the Infrastructure organization. As a senior member of Baseten's Platform Team, you will own the systems that let every engineer at Baseten prove their code works before it reaches production. Our product runs mission-critical AI inference for customers who measure downtime in dollars per second, which means our internal bar for correctness, performance, and failure tolerance has to be exceptional. Your focus is the full testing stack: fast and reliable unit test tooling, integration harnesses that spin up realistic environments on demand, load and performance testing for GPU-backed inference workloads, and resilience testing that deliberately breaks things so our customers never have to find out what happens when a node dies mid-request. This is a builder role with org-wide leverage. You won't be writing tests for other teams — you'll be building the frameworks, harnesses, and feedback loops that make writing good tests the path of least resistance, and you'll set the standards for what "well-tested" means at Baseten. RESPONSIBILITIES Own Baseten's testing strategy end to end — define the standards, the tiers, and the tooling that engineering teams build against. Build and maintain unit, integration, load and performance testing frameworks Design end to end test infrastructure that provisions realistic dependencies
About Supabase Supabase is the Postgres development platform, built by developers for developers. We provide a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search. All services are deeply integrated and designed for growth. About the Role We’re looking for a Developer Relations Engineer based in San Francisco to join our team and to help more developers discover, learn, and build with Supabase. You’ll create high-impact content, build real-world projects, and represent Supabase across communities and events. If you’re equally energized by writing code and teaching others, this is the role for you. Why this role matters Supabase is growing fast with 350,000+ developers , 1,000+ OSS contributors , and a thriving open source ecosystem. Our users are builders, startup founders, weekend hackers, and engineers scaling to millions of users. DevRel is how we meet them where they are: through content, community, and code. We’re building a community of communities that brings together developers from many backgrounds, including first-time open source contributors. As a DevRel Engineer, you’ll be a bridge between Supabase and the broader developer ecosystem, helping people get started, go deep, and feel connected. What you'll do Make content Publish compelling technical content, especially video, to help developers learn Supabase quickly Build demos and tutorials Ship real-world apps using Supabase and tools like Next.js, React, and Stripe. Write guides that others can follow and remix Represent Supabase Speak at meetups, livestream builds, and engage with the developer ecosystem. You’ll be a visible and trusted voice of the platform Support and grow the community Celebrate contributors, answer questions, highlight cool projects, and bring developer feedback to the team Collaborate across the company Work with engineering, product, and growth to amplify launches, prioritize content, and reduce friction for new users You migh
$155K – $400K/yr
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About this role Are you ready to redefine the future of JavaScript development? Are you a seasoned JavaScript expert who gets a thrill from tackling complex challenges across the entire JavaScript ecosystem? Do you believe that AI can be a powerful partner in crafting elegant, high-impact code? If you're ready to leave the mundane behind and join a team that's shaping the tools used by millions of developers globally, we've got an opportunity for you at Sentry. This isn't your typical Senior Software Engineer position. As a key member of our growing JavaScript SDK team, you'll be at the forefront of innovation, working on everything from our cutting-edge SDKs for Node.js, Bun, Deno, Cloudflare Workers, and other modern server runtimes. You won't just be maintaining code; you'll be pushing the boundaries of what's possible in developer tooling across the rapidly evolving server-side JavaScript landscape. In this role you will Join our JavaScript SDK team and get ready to build the future. You'll be at the forefront, working on: A Universe of JavaScript Challenges: Dive deep into our extensive suite of JavaScript SDKs, with a sharp focus on server-side and edge runtimes — from the battle-tested Node.js ecosystem to cutting-edge alternatives like Bun and Deno, and distributed edge environments like Cloudflare Workers. Your work will directly empower millions of developers to build better, more reliable software, no matter which runtime powers their stack End-to-End Ownership: We believe in giving our engineers the autonomy to see their vision through. You'll have the freedom to plan, implement, and ship your code, from writing
$155K – $400K/yr
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About this role Are you ready to redefine the future of JavaScript development? Are you a seasoned JavaScript expert who gets a thrill from tackling complex challenges across the entire JavaScript ecosystem? Do you believe that AI can be a powerful partner in crafting elegant, high-impact code? If you're ready to leave the mundane behind and join a team that's shaping the tools used by millions of developers globally, we've got an opportunity for you at Sentry. This isn't your typical Senior Software Engineer position. As a key member of our growing JavaScript SDK team, you'll be at the forefront of innovation, working on everything from our cutting-edge SDKs for the React, Next.js, Vue, Nuxt, Hono, NestJS, and beyond. You won't just be maintaining code; you'll be pushing the boundaries of what's possible in developer tooling across the full spectrum of the modern JavaScript framework landscape. In this role you will Join our JavaScript SDK team and get ready to build the future. You'll be at the forefront, working on: A Universe of JavaScript Challenges: Dive deep into our extensive suite of JavaScript SDKs, with a broad focus on framework support spanning the modern JS ecosystem — from frontend frameworks like React, Vue, and their meta-frameworks Next.js and Nuxt, to server-side and edge runtimes like NestJS and Hono. Your work will directly empower millions of developers to build better, more reliable software, regardless of their framework of choice End-to-End Ownership: We believe in giving our engineers the autonomy to see their vision through. You'll have the freedom to plan, implement, and ship your code, from writi
$150K – $190K/yr
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role The Developer Experience team is growing in San Francisco! This is a hybrid role out of our main headquarters and must be based there. We’re looking for a hands-on builder with strong opinions on great developer docs. Are you someone who loves tinkering with the newest features in your favorite products? Do you enjoy taking the next random JavaScript framework for a spin and figuring out how to make them tick? Does it especially irk you when the code snippets are wrong, or the only docs for a feature are on X? Developer Experience at Sentry lives in the intersection of shipping really cool products and getting developers set up to use them. If you’re someone who is confident in partnering with Product and Engineering to test and ship the latest features, not afraid to jump in and go hands-on to solve problems for our biggest customers, and a good eye for what good docs looks like - this is a dream role. Developer Experience at Sentry is a team of builders who are constantly looking for ways to make it easier for every developer to use Sentry. We engage with the challenges facing technical communities, to support developers, gather feedback, and help our product and engineering teams ship new capabilities. An engineer in this role should be confident to “come with an answer," propose the product solutions, identify the content and/or execution plans and people to partner with, and go execute. In this role you will Attend and even host events and meetups in the developer community Have a point of view and aren’t shy about expressing it. You get your energy from both knowing the latest trends and having a POV on
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the Role Sentry already has a startup program: credits, partner deals with YC and a16z, a landing page with a terminal theme that developers actually think is cool. What we don't have is someone whose full-time job is making this program a real growth engine. That's this role. You'll own startup marketing end-to-end: the partner ecosystem, the community presence, the acquisition programs, the brand in rooms full of founders and early engineers. You'll take what exists today and turn it into something that makes Sentry the default monitoring choice for every new company shipping code. This role reports to the Head of Enterprise Marketing (yes, we get how this can be confusing, but she’s pretty cool) and works across Growth, Product Marketing, and DevEx. You'll have real autonomy to shape the strategy and build programs from scratch – but you'll also need to execute quickly, (kind of – we can explain more) measure what works, and kill what doesn't. In this role you will Build and scale startup partnerships. Expand relationships with accelerators, incubators, VC platforms, and startup deal aggregators. Make Sentry's credits program a no-brainer inclusion in every startup perks stack. Own the startup acquisition funnel. Design programs that move early-stage companies from free plan → credits program → paid customers → expansion. Instrument everything so we know where the pipeline is and where it leaks. Be Sentry's face in the startup ecosystem. Show up at demo days, founder events, and SF tech week. Build relationships with the people building the next generation of companies. Not because it's a line item, but because you'
What you’ll do Act as the in-house electrical lead for Midjourney Medical: own the electrical architecture of the scanner and the technical direction for all board-level design. Own complex board design end-to-end: architecture, schematic capture, layout (high-speed digital, analog/mixed-signal, power), DFM/DFT, fabrication and assembly vendor management, bring-up, and revision control. Write firmware for embedded targets (MCU/SoC): drivers, real-time control loops, safety-relevant logic, bootloaders, and field update paths. Audit and update HDL (FPGA) code for high-throughput data acquisition, timing/synchronization, triggering, and pre-processing of ultrasound and sensor data streams. Define electrical interfaces and data contracts with software, recon/ML and mechanical teams: timing budgets, clocking/sync, signal integrity, connectors/harnessing, and failure modes. Establish electrical engineering rigor: design reviews, schematic/layout review checklists, bring-up procedures, test fixtures, and documentation suitable for a regulated medical device program (DHF, traceability, change control). Mentor and grow the electrical function; select and manage external design partners where leverage is high. What we’re looking for Deep experience designing complex boards from blank page to stable revision, including high-speed digital and analog/mixed-signal domains. Strong schematic and layout skills (Altium/KiCad or equivalent) with real signal integrity, power integrity, grounding, and EMI/EMC instincts. Solid embedded firmware background in C/C++ (and Python for tooling): peripherals, DMA, interrupts, real-time constraints, and debugging on hardware. Practical HDL experience (VHDL/Verilog/SystemVerilog) for data acquisition, timing, and streaming interfaces. Track record of owning bring-up and debug on real hardware: scopes, logic analyzers, and disciplined root-cause analysis. Technical leadership: clear trade-offs, strong written documentation, and the ability to set
What you’ll do Act as the technical lead for large parts of the scanner platform: system architecture, codebase structure, and long-term maintainability. Own core runtime foundations: distributed control, state management, fault handling, and reliability. Drive engineering rigor: testability, code quality, review standards, performance regression prevention, and release processes. Build robust observability: logs, metrics, traces, and replayable diagnostics (with privacy constraints). Collaborate with hardware and recon/ML teams to define interfaces, data contracts, timing/synchronization, and failure modes. Lead complex refactors (e.g., message passing / RPC boundaries, modularization, concurrency model) without halting forward progress. What we’re looking for Deep software architecture experience for real-world systems: robotics, instrumentation, medical devices, or other complex distributed products. Strong Python and concurrency background (asyncio, multiprocessing, profiling, performance engineering). Track record of shipping systems that are observable, debuggable, and resilient. Strong technical leadership: clarity, pragmatic trade-offs, and mentoring. Useful experience Building but rock-solid systems: clear interfaces (gRPC/protobuf or equivalent), strong state modeling, and failure handling. High-leverage engineering habits on a lean team: good tests, CI, reproducible dev environments, and fast code review. Practical performance + concurrency work in Python (asyncio, profiling, multiprocessing) and comfort debugging distributed behavior. Security-minded device software: safe defaults, encrypted data paths, and disciplined handling of PII/PHI. Operational thinking: remote updates/management, excellent logging, and diagnostics that make real hardware debuggable.
About the Team Security is at the foundation of OpenAI’s mission to ensure that artificial general intelligence benefits all of humanity. The Security team protects OpenAI’s technology, people, and products. We are technical in what we build but are operational in how we do our work, and are committed to supporting all products and research at OpenAI. Our Security team tenets include: prioritizing for impact, enabling researchers, preparing for future transformative technologies, and engaging a robust security culture. About the Role As a Security Engineer, Application Security you will be responsible for identifying and mitigating security vulnerabilities within software applications through building security tools, code reviews, penetration testing, and security assessments. We’re looking for people who will work closely with development teams to ensure secure coding practices are integrated throughout the software development lifecycle, preventing security risks before they emerge. You will also provide security guidance to developers and other stakeholders, fostering a culture of security awareness within the organization. The role is preferred to be based in San Francisco, Seattle or New York City but may consider remote work. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Perform Security Assessments : Conduct regular security assessments, code reviews, and penetration testing to identify vulnerabilities in applications and software. Develop and Implement Security Tools : Design, develop, and implement security tools, frameworks, and methodologies to protect applications against security threats. Collaborate with Development Teams : Work closely with development teams to ensure security best practices are integrated throughout the software development lifecycle (SDLC), including secure coding guidelines. Threat Modeling and Risk Assessment : Conduct threat modeling and risk
About the Team Our Applied team brings OpenAI technologies to consumers and businesses around the world. We collaborate across research, engineering, design and business functions to turn cutting-edge AI advancements into impactful real-world applications. Our team has been behind notable product launches ( ChatGPT , API , Sora ), creating tools that help developers write code, enable businesses to operate more efficiently, and empower individuals to learn and create. As AI capabilities rapidly evolve, we focus on ensuring that our products are safe, accessible, and beneficial to all. About the Role As a Data Scientist on the Applied Product team, you will contribute to a data-driven product development culture for consumer and enterprise products at OpenAI. This is critical as our products reach millions of users and businesses worldwide. We are focused on aligning both research and product development to drive measurable impact for these individuals and organizations alike. You should expect to define our north-star metrics, design A/B tests, and establish source-of-truth dashboards that the entire company can use to answer their own product questions. Most importantly, you should expect to be a core member of the product development team. This role is based in San Francisco, CA or Seattle, WA. We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. In this role, you will: Embed with the product development team as a trusted partner, uncovering new ways to improve the product and drive growth Define and interpret A/B tests that help answer critical questions about the impact of model and UX changes to our product Establish a data-driven product development culture by defining, tracking, and operationalizing feature-, product-, and company-level metrics Develop and socialize dashboards, reports, and other ways of enabling the team and company to answer product data questions in a self-serve way You might thrive
$295K – $380K/yr
About the Team The OpenAI Robotics team is focused on unlocking general-purpose robotics and pushing towards AGI-level intelligence in dynamic, real-world settings. Working across the entire model stack, we integrate cutting-edge hardware and software to explore a broad range of robotic form factors. We strive to seamlessly blend high-level AI capabilities with the constraints of physical systems to improve peoples’ lives. About the Role As a Senior Software Engineer, ML Systems & Training Infrastructure, you will be a deeply hands-on engineering force multiplier for the robotics team. You will help keep the training framework and surrounding infrastructure healthy, review and improve code quickly, debug failures across ML systems and infrastructure, and unblock researchers and engineers when the path from idea to working training job gets rough. We’re looking for people who love writing, reading, reviewing, and fixing code; who can get productive quickly in unfamiliar systems; and who bring strong practical judgment without a lot of ego or process overhead. This role will be based in San Francisco, CA and be expected in office 5 days per week and offer relocation assistance to new employees. In this role, you will: Review, improve, and clean up code across training frameworks and adjacent infrastructure. Identify risky or low-quality changes before they land, and raise the code quality bar without slowing the team down. Debug issues across ML training systems, GPUs, clusters, networking, and related infrastructure. Help researchers and engineers unblock broken training jobs, flaky workflows, and brittle internal tooling. Improve the reliability, maintainability, and usability of the robotics team’s training framework. Move quickly on practical engineering problems that directly affect team velocity. You might thrive in this role if you: Have strong software engineering fundamentals and excellent code review judgment. Have experience with ML systems, training fr
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role We are looking for an embedded engineer to help build firmware and associated modeling software for OpenAI’s in house AI accelerator. This role involves designing and developing drivers and functional models for a large array of HW components, writing high throughput and low latency firmware code, investigating bring-up and production issues. Responsibilities Design and implement drivers for hardware peripherals, including those related to AI chips. Design and implement functional software models to simulate SoC uncore logic and enable FW testing against the model Design and implement low-latency and high throughput embedded SW to manage HW resources. Work with adjacent software and hardware teams to implement requirements, debug issues and shape future generations of the hardware. Collaborate with vendors to integrate their technologies within our systems. Bring up and debug firmware/driver on new platforms. Come up with processes and debug issues raised in the field. Set up monitoring, integration testing and diagnostics tools. Qualifications 5+ years of experience working in embedded SW space. Ability to thrive in ambiguity and learn new technologies. Strong programming skills in C/C++ and/or Rust. Experience developing high throughput, low latency and multi-threaded code. Experience working with real time operating systems (RTOS). Experience developing hardware drivers and working with hardware Experience with HW/SW co-design Knowledge of common embedded pr
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. Notion is an in person company, and currently requires its employees to come to the office for three Anchor Days (Mondays, Tuesdays, and Thursdays). This internship will take place from Jan 25 - Apr 16 and you will need to be able to work out of our NY or SF office during this time. About the Role: Mobile devices have reshaped personal computing—the power of a desktop now fits in our pockets. By giving everyone the building blocks to create their own tools, we'll solve more problems wherever we go. A delightful, intuitive mobile experience: that's our guiding light, and we need Android and iOS engineers to make it a reality. During your 12-week internship, you will be paired with a mentor that will help guide you as you work closely with our team to build and ship impactful projects. These projects will drive valuable impact to our customers and engineers. Take a look at what our past intern cohort worked on with our Blog post here and here and TikTok post. What You'll Achieve: Write clean, secure, tested, and documented code. You'll collaborate with the team to strategize, mold, and develop innovative features for our Android and iO
About the Team OpenAI develops models that can reason through complex problems and hardware designed for the demands of advanced AI. AI for Chips connects these efforts: applying increasingly capable AI systems to the work of semiconductor engineering. Our goal is to help engineers develop better chips and shorten design cycles. This work brings research, model training, and hardware expertise together to build tools that engineers can use on real designs, with correctness and measurable performance at the center. About the Role We’re hiring a Research Engineer to help OpenAI models solve chip-design problems through reinforcement learning, tool use, and evaluation. You’ll own experiments from the initial idea through implementation and analysis. That means building environments and evaluations, running training, investigating failures, and using the results to decide what to try next. You’ll also build the software needed to make those experiments reliable and reproducible. We value strong coding fundamentals, careful experimental judgment, and the ability to make progress independently. Prior chip-design experience is helpful, but you can learn the domain alongside the team’s hardware specialists. In this role, you will: Build RL environments and evaluations for tasks such as RTL generation, design verification, and physical design optimization. Develop and test approaches that help models use chip-design tools and improve power, performance, and area while preserving correctness. Design experiments, establish baselines, and measure whether improvements hold up on new tasks and designs. Investigate failures across model behavior, rewards, evaluation tools, and experiment infrastructure. Improve iteration speed through better tooling, faster evaluations, and proxy rewards that reflect the outcomes we care about. Turn successful experiments into reusable research code and training workflows, working closely with researchers and engineers. You might thrive in this ro
Other cities to consider
More places hiring for this role
Get new codex deployment engineer jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime