ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is building its own GPU infrastructure for large-scale inference. As we move into large scale, high-density NVIDIA systems, the hardest failures are intermittent, cross-layer, and difficult to prove: RoCE congestion, InfiniBand stalls, ECN/DCQCN mis-tuning, bad optics, RNIC issues, host kernel stalls, GPU driver problems, and workload symptoms that look like network problems, but are not. We are hiring a Lead Software Engineer to build a first-class observability and root-cause analysis system for GPU fabrics. This is a hard distributed systems problem, not a dashboarding problem. The system will collect high-volume signals from switches, hosts, active probes, and inference services; reduce and correlate them in real time; understand topology and service ownership; and produce actionable diagnosis while an incident is still unfolding. This role sits at the boundary between networking and inference software. RDMA data paths, GPUDirect transfers, prefill/decode disaggregation, KV cache movement, request routing, and workload backpressure can all create fabric symptoms or hide real fabric failures. The goal is to tell an operator, quickly and with evidence, whether an incident is caused by the fabric, host, NIC, GPU, RDMA path, scheduler, or serving layer — and what to do next. EXAMPLE INITIATIVES Real-time telemetry engine — Build the ingestion, reduction, storage, and query path for high-cardinality fab
Jobs in United States
Lead Software Developer 2c Kubernetes Compute Consultant Manager in San Francisco
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current lead software developer 2c kubernetes compute consultant manager jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the Role Sentry's blog, newsletter, and content programs need an owner — someone who can walk into ambiguity, decide what needs to exist, and just start making it happen. This is a self-starter role in the truest sense: no one is handing you a content calendar and asking you to fill it in, you're the one building the calendar. You should be quick with a draft and quicker to ship it — bias to action over polished process. You're technically savvy enough to keep up with a developer audience and write about our product credibly (or make an exec sound like they wrote it themselves). And you're proactive: you see the narrative before anyone asks you to go find it. You'll work collaboratively with Growth, Comms, PMM, and DevEx, with real ownership and editorial judgment to know what's worth writing and what isn't. In this role you will Co-lead blog strategy. Partner closely with Growth to run the editorial calendar, and scale the program by focusing on creating thought leadership, customer and employer brand content to create a robust narrative driven editorial program in collaboration with comms, devex and product marketing. Oversee the quality itself. Be the last line of defense on quality and prose — every post should read like it came from the same sharp, confident voice. Track the industry and get us into the conversation. Know what's trending in the developer and infrastructure world before it's obvious, and turn that into content that puts Sentry in the room. Collaborate with Comms and Marketing to push those narratives forward. Content partner for the exec team. Collaborate closely with core execs to support content
$220K – $300K/yr
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role Developer experience at Sentry spans four areas: the integration and developer platform, docs, community, and DevRel. The DevEx team is responsible for how developers first encounter Sentry, how they learn to use it, and how they build on top of it — from SDK configuration docs to the Discord community to the integrations ecosystem. This role owns all of it. We recently had our Head of DevEx depart, and we're looking for a senior leader to step in with real ownership and real scope. You'll sit in Marketing, partner closely with EPD, and have direct influence over both the product surface (the integration platform) and the team responsible for developer education, docs, and community. In this role you will Own strategy and execution for Sentry's integration ecosystem and developer platform APIs — working closely with EPD to define the developer-facing surface and ensure building on Sentry is a genuinely good experience Lead the documentation function across SDK configuration docs, API reference, and tutorials — the places where developer activation actually happens or doesn't Connect docs to the product experience: logged-in state, contextual guidance, health checks, and surfacing adoption gaps in-context, in partnership with EPD and the Docs platform team Own Sentry's developer community across Discord, social, and events, and build the team and processes that turn engagement into measurable outcomes: signups, adoption, support deflection Build and deliver the content and education programs — workshops, deep dives, fireside chats — that teach developers tracing, debugging, and how to actually use Sentry's full
About the Team We’re hiring Software Engineers to join our broader Infrastructure organization, which supports multiple high-impact teams. Depending on your interests and experience, you could work on one of several focus areas—including Core Distributed Systems, Reliability Engineering, Observability, Developer Productivity or Cloud Infrastructure. About the Role All teams are deeply collaborative, work on mission-critical services, and are responsible for building distributed, scalable infrastructure to bring OpenAI’s technology to the world through products like ChatGPT and the OpenAI API. You’ll work closely with stakeholders to understand infrastructure, data and compute needs, setting the technical strategy that supports cutting-edge research and product development. This is a critical role for someone who is passionate about solving complex engineering problems at scale, ensuring their performance, scalability and reliability Team Focus Areas Distributed Systems: Owning and building important, highly scalable, available, performant, and reliable distributed systems (and their building blocks) to power the entire stack at OpenAI Systems Engineering: Work across layers of the stack—debugging system bottlenecks, evolving core infrastructure, and solving novel problems in performance and scalability. Reliability Engineering: Build scalable, fault-tolerant systems and lead efforts around service health, incident response, and resilience. Observability: Design and maintain observability tooling (metrics, logs, tracing) to give teams visibility into production systems at scale. Developer Productivity: Create tools, environments, and workflows that help engineers ship high-quality software faster and more safely. Cloud Infrastructure: Own the cloud-native infrastructure (compute, networking, storage) that underpins all services and research workloads. Databases: Building high performance, distributed database systems that power all of OpenAI's product stack. In this
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Software Engineer at on the Training Infrastructure team, you'll architect and lead development of our training platform, supporting top tier research engineers and model developers. You'll make key technical decisions for the infrastructure enabling developers to deploy, scale, and monitor their workloads with high performance and reliability. You’ll own scheduling, storage, networking, reliability, and observability of technical systems in the training stack EXAMPLE INITIATIVES Take a look at what we’ve built so far: Overview of the product so far Training docs overview Story of the Training product Research we've done RESPONSIBILITIES Design and architect scalable infrastructure systems for our ML training platform (e.g. scheduling, storage, and networking) Partner closely with developers and research engineers to translate complex training requirements into technical solutions Design and architect a global training scheduler Design and architect reinforcement learning systems and continuous learning pipelines Drive long-term improvements to improve reliability of systems and velocity of development Partner closely with SRE and Capacity teams to unlock state of the art training infrastructure Make critical architectural decisions balancing performance with system reliability Lead technical discussions and mentor junior engineers on infrastructure best practices Contribute to long-term technical strateg
$155K – $400K/yr
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role The Events Analytics Platform (EAP) team is responsible for the infrastructure that powers all of Sentry's time-series data and searching capabilities across billions of events with sub-second latency. We started this initiative by building Snuba, the primary storage and query service for Sentry's event data powered by ClickHouse, and we are now focused on unlocking deeper visibility and reporting across the terabytes of event data our users generate. As a Senior Software Engineer, you will lead efforts to push the boundaries of data visibility at Sentry. You will do this by expanding the capabilities of our search infrastructure, building new capabilities on top of our state-of-the-art storage layer and increasing the performance and integrity of Sentry’s core data services. You will also help shape Infrastructure's technical direction at Sentry and collaborate with Product and other Engineering teams to turn that vision into a reality. If you want to solve the hard problems that come with scaling event data into the petabyte range, this could be the job for you. In this role you will: Expand EAP's ability to deliver data at world-class speed and reliability. Architect and automate services and systems to scale reliably under growing demand. Make architectural trade-offs that balance product requirements with engineering constraints. Maintain and grow the team's code quality initiatives by regularly reviewing code and contributing to design decisions. Lead design and discussions around deliverables the team is working towards. Improve the maintainability and developer experience of the codebases EAP owns. Exa
About the Team The Product Marketing team shapes how customers understand, adopt, and realize value from OpenAI’s technology. We work across Product, Research, Sales, Solutions, Partnerships, and Customer Success to bring customer insight into our product strategy and translate technical capabilities into clear, credible stories and solutions. About the Role AI becomes meaningful when it helps people do the work that matters to them. For a finance team, that might mean understanding complex information faster. For a healthcare provider, it might mean navigating clinical workflows more effectively. For a sales or marketing team, it might mean creating entirely new ways to reach and serve customers. We’re looking for a senior product marketing leader to shape how OpenAI serves the business functions and industries where our technology can make a meaningful difference. You’ll define how our models and products meet the needs of teams such as sales, marketing, and finance, as well as industries including financial services, healthcare, and retail. Working closely with Product, Research, Sales, Solutions, and Partnerships, you’ll identify important customer problems, influence product strategy, and build relationships with the ecosystem partners and data providers needed to bring complete solutions to market. You’ll also build and lead the product marketing team responsible for turning these opportunities into durable customer value. You might thrive in this role if you: Have 12+ years of experience in product marketing, industry marketing, solutions marketing, or enterprise go-to-market, ideally across enterprise software, cloud, data, developer, or AI platforms. Have built and led high-performing teams, mentored senior marketers, and know how to create clarity in fast-moving, ambiguous environments. Understand how different industries and business functions evaluate technology, adopt new tools, and define value. Have shaped positioning and go-to-market strategies for c
About the Team API Enterprise Controls is part of the API Infrastructure organization and owns the platform capabilities that help developers, startups, and enterprises adopt the OpenAI API securely and confidently. We build the systems underneath our APIs and developer platform across authentication and identity, service accounts and key management, secure networking, compliance, auditability, observability, and operational controls. Our users are developers and teams running critical applications on OpenAI, and we partner closely with Product, go-to-market, security, and infrastructure teams to turn their most important needs into reliable, intuitive platform capabilities. About the Role We are looking for an exceptional backend software engineer to help define and ship the enterprise capabilities our API Platform needs to scale.; this is a product-engineering role grounded in deep backend systems. You will work across databases, streaming systems, request routing, authentication, and developer-facing APIs while bringing strong product judgment, developer empathy, and attention to the small details that make a platform easier to understand, trust, and operate. You will lead large cross-functional initiatives, work closely with Product and go-to-market teams, engage directly with sophisticated users, and carry ambiguous needs from discovery through design, launch, and iteration. In this role, you will: Own backend product capabilities end to end across authentication and identity, service accounts and key controls, secure networking, compliance, observability, and operational workflows. Partner with Product, go-to-market, security, infrastructure teams, and sophisticated customers to identify needs, shape the roadmap, and lead large cross-functional projects from design through launch. Design developer-facing APIs, system behavior, configuration, error handling, safe defaults, auditing, and notifications with exceptional care for the details that define a great dev
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. At Baseten, we are building the global operating system for distributed, heterogeneous AI hardware. We believe that as LLM and multi-modal workloads scale, the network is the computer. We are looking for foundational engineers to lead our GPU Networking efforts, making RDMA a first-class building block in our infrastructure and unlocking the next generation of distributed inference optimizations. THE OPPORTUNITY Networking and compute are no longer separate disciplines; they are converging. The massive throughput of H100, B200, and NVL72 architectures enables and demands a new approach where communication is co-optimized alongside computation. We are entering an era where the network is an active accelerator, leveraging smart hardware offloads and direct interconnects to ensure that data movement operates at wire-speed. In this role, you will go beyond network configuration to architect the software fabric that unifies thousands of GPUs into a cohesive operating system. While you will leverage the best of the open-source ecosystem, you won't be limited by it. Where off-the-shelf solutions stop, you will build from scratch, engineering the primitives required to co-optimize communication and compute for Disaggregated Serving, Wide Expert Parallelism (WideEP), and lightening cold starts. WHAT YOU'LL DO Make RDMA First-Class: You will work on integrating RDMA/RoCE/InfiniBand capabilities directly into our inference stack,
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE: As a Software Engineer at Baseten, you will own one of the most critical surfaces of our business: pricing, billing, and revenue infrastructure. As we launch more and more products— billing is no longer just operational plumbing. It is a strategic lever for growth. This role will establish clear ownership of billing as a function and create leverage for Finance, Sales, and GTM teams while maintaining a seamless customer experience. RESPONSIBILITIES: Own Baseten’s end-to-end billing and revenue infrastructure, including pricing, invoicing, metering, and reporting foundations. Build and evolve our billing platform and integrations (including Orb), ensuring correctness, auditability, and a high-trust experience for customers and internal teams. Partner closely with Finance, Sales, GTM, and Forward Deployed Engineering to turn real-world workflows into reliable internal tooling and automation (quoting, approvals, renewals, usage reconciliation, revenue reporting). Design systems that scale with new products, packaging, and go-to-market motions, making billing a strategic lever for growth. Drive reliability and operational excellence for revenue-critical workflows: monitoring, alerting, incident response, backfills, and clear runbooks. Lead from the front on high-impact projects: clarify requirements, propose crisp technical approaches, ship iteratively, and raise the bar on quality and velocity. Debug and resolve
About the Team We are a small and fast-moving partnerships team that shapes and executes OpenAI’s most important collaborations. Your mission is to build and lead our business with independent software vendors, from native AI companies building in our programs to technology companies taking new products to market with OpenAI’s models and products. You will help partners build, launch, and grow differentiated products while working closely with our sales and product organizations. About the Role You are a senior partnerships leader with strong judgment, high ownership, and a track record of building consequential ISV relationships. You can establish durable executive partnerships, lead complex product and commercial negotiations, and move fluently from market strategy to integration, launch, adoption, and growth. You are highly commercial, technically fluent, and comfortable operating where there is no playbook. In this role, you will: Define OpenAI’s ISV partnership strategy, priorities, segmentation, and business goals. Recruit and manage a portfolio of high-impact software partners across horizontal and vertical applications. Structure and negotiate partnerships spanning product integration, model adoption, commercial terms, distribution, and joint go-to-market. Develop joint business plans and executive governance with strategic partners. Help partners move from initial build to launch, customer adoption, expansion, and scaled consumption. Create repeatable co-build, co-sell, co-market, and marketplace motions. Serve as the voice of ISVs internally, shaping product requirements, developer experience, economics, and partner programs. Track partner pipeline, launches, adoption, consumption, customer impact, and ecosystem growth. Establish the team structure and operating model needed to scale the ISV business. You might thrive in this role if you have: 12+ years of experience in technology partnerships, business development, platform ecosystems, or enterprise softw
Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity Postman is the world's leading API platform, and APIs and AI agents are increasingly the backbone of modern software. We're looking for a Technical Community Manager who will build and lead a thriving global developer community. One that goes beyond Postman users to become the destination for any developer working with APIs and AI agents. This is not a typical community role. You'll be architecting programs, designing engagement loops, and building something genuinely new: a community where developers come to learn, build, share, and grow — and where Postman is recognized as an indispensable part of their toolchain. The intersection of APIs and AI agents is one of the most exciting spaces in software right now, and Postman sits right in the middle of it. We have millions of developers already on the platform — the opportunity is to turn that user base into a community that educates, inspires, and advocates for one another. The person who builds this will leave a real mark on how developers collaborate and grow for years to come. What You'll Do Discord Community Growth & Engagement Own the end-to-end strategy for
$220K – $450K/yr
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role Sentry is looking for an Engineering Manager to lead our Growth team (internally called "Value Discovery") and help expand our product-led business by helping customers discover and get the most out of our platform. At Sentry, we do growth with an explicit bias for customer value. Our experiments are rooted in strong product sense and empathy for developers, not conversion pressure. We never surprise developers with new charges or push features they don't need. In a binary tradeoff between doing good by the customer or good by the business, the customer always wins. You'll collaborate closely with several areas of the business including Product, BizOps, Data, Design, and GTM to identify the moments in a developer's workflow where Sentry can deliver more value, make that value easier to find, and support that your approach works with data. In this role you will You'll lead a team of engineers, set its priorities, and be accountable for the metrics it moves. This is a new role, and the roadmap is yours to define and execute in partnership with design and data teams. Lead work across Sentry's self-serve funnel — signup and onboarding, first-run activation moments in-product, trial and upgrade paths, and pricing and billing surfaces in partnership with the Billing team — largely in our Python/Django and TypeScript/React codebase Evolve our experimentation foundations, in close partnership with our Data team — flagging, instrumentation, and readouts are real but far from finished Run a fast, disciplined experimentation loop where hypotheses are argued before they're built Hold conversion and activation alongside ch
$220K – $450K/yr
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role Sentry provides developer-first observability to over 4 million developers worldwide. The Events Analytics Platform (EAP) team is at the heart of that mission: it powers how all of Sentry's event data, such as errors, transactions, spans, profiles, replays, and metrics, is stored, queried, and analyzed. It also powers Sentry's latest AI push, Seer. The EAP team makes it possible for developers to efficiently search and debug across massive volumes of data, providing the context needed to understand and fix issues quickly. This team is also a cornerstone of Sentry's long-term strategy to become a context assembly and telemetry platform that unifies different signals so developers can see the complete picture. As an engineering manager on the EAP team, you will lead a group of engineers building and scaling one of Sentry's most critical data platforms. You will be responsible for driving architectural evolution, ensuring system stability, and mentoring a talented team. This is a highly visible leadership role with direct ties to Sentry's long-term product and platform strategy. What you'll do Grow and develop a team of engineers with high expectations for ownership and impact Set the technical and strategic direction for the team, balancing short-term stability with long-term architectural evolution Drive development of core EAP features, including support for complex analytical queries, dynamic routing logic across fidelity levels, storage and compute separation, and modern patterns for analytical storage Ensure EAP can support the workload demands of AI agents and MCP servers that unlock new AI capabilities fo
About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the Role At Sentry, Support is an engineering discipline. Our customers are the greatest technical minds in the world—developers at elite enterprises building the future of software—and they deserve answers that go deeper than a knowledge base link. We're looking for an APAC Technical Support Engineer based in San Francisco to join our global Support Engineering team. This role is designed to provide APAC coverage to our users; with the shift being Sunday through Thursday 4PM-12AM PST. We are architecting the Technical Support engine . We’re looking for an experienced engineer to help us redefine the standard of technical support by combining deep human expertise with autonomous agentic systems. You are a debugger of both code and systems. You will treat support volume as a data signal to build automated resolution paths, ensuring our human engineers only touch the most complex, high-impact architectural puzzles. Sentry Support Engineers aren't just clearing queues; they are Orchestrators . You will engage with our users across GitHub, Discord, and our internal systems, while acting as the Technical Lead for our Agentic Ops. You ensure that when a developer asks a complex question, our systems have the right context and a seamless "Human-in-the-Loop" path to you when deep, nuanced expertise is required. In this role you will Master the Sentry Ecosystem & Support Elite Developers Deep-Dive Debugging: Perform root-cause analysis on complex issues and distributed tracing gaps across polyglot environments. Support the Great Minds: Act as a strategic consultant for senior engineers at our largest enterprise customers, s
Other cities to consider
More places hiring for this role
Get new lead software developer 2c kubernetes compute consultant manager jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime