About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role You will work on the systems software strategy and execution that brings new AI silicon from first power-on to a fully integrated system running production-representative models at expected functionality and performance. You will define how software exercises and validates compute, memory, interconnect, and I/O subsystems, then build the diagnostics, automation, and observability needed to find issues quickly. This role sits at the center of silicon, firmware, platform, systems, and workload teams. You will turn hardware specifications and performance targets into an end-to-end bringup plan, drive cross-functional debug, and establish the stress and regression infrastructure that makes each new platform reliable across operating environments. In this role, you will: Contribute to the end-to-end software bringup and validation strategy for new silicon and first-party systems. Define software-driven test coverage across compute, memory, interconnect, I/O, and their system-level interactions. Build diagnostics, test automation, telemetry, and regression infrastructure that accelerate first-silicon learning and issue isolation. Lead bringup from initial silicon arrival through board and system integration, docking, runtime enablement, and model execution. Design stress tests that characterize reliability, performance, and stability across workloads and operating conditions. Translate architecture specifications and performance models into measurable acceptance crit
Jobs in United States
Python in San Francisco
319 active opportunities · Updated October 2026
Showing
15 jobs
Explore current python jobs in San Francisco. Filter by work mode, employment type, experience, department, date posted and distance.
About the Team OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. About the Role You will build the model runtime within the inference engine that executes complex, frontier models at scale on OpenAI’s custom silicon. The runtime will sit between models running on the hardware and the upper layers of the cluster serving software stack, translating demanding inference workloads into efficient execution while optimizing for throughput, latency, utilization, and reliability. You will work across model architecture, distributed systems, compilers, kernels, and silicon to design a production-grade runtime comparable in ambition to systems such as vLLM and SGLang, but customized and optimized for OpenAI’s AI accelerator. Your work will shape how new model capabilities map onto the platform and how quickly custom silicon can deliver meaningful performance in production. In this role, you will: Design and implement the LLM inference runtime for frontier models running on custom silicon. Build scheduling, continuous batching, memory management, KV-cache management, and execution orchestration for high-performance inference. Develop distributed execution strategies across chips, hosts, and racks, including model partitioning, communication, and synchronization. Optimize end-to-end latency, throughput, memory efficiency, and hardware utilization across diverse model architectures and serving workloads. Partner with kernel, compiler, architecture, and silicon teams to co-design interfaces and remove performance bottlenecks across the stack. Enable new
The Opportunity The world of design is changing rapidly, and the Pro Design team is leading that transformation. We are the Adobe organization behind Illustrator, InDesign, and emerging experiences that connect creativity, collaboration, and AI. Our teams are reimagining what professional design looks like for the next decade - building intelligent, connected tools that empower creators and teams to move faster without sacrificing craft. We are looking for a Senior Business Data Scientist who is creative, analytical, and unafraid to question the status quo and shape the decisions that move key business metrics at scale. Join us and build Adobe’s future products! What you'll Do Map the user funnel and build the metrics, cohorts, and dashboards that Product and Growth rely on to see how users move across free, trial, and paid tiers—and pinpoint where they drop off. Dig into the hard questions (what drives activation, which behaviors predict retention and expansion) and build propensity models for conversion, upgrade, churn, and expansion that feed real-time targeting and in-product nudges. Find and size growth bets and work with Product to ship them. Set north-star, driver, and guardrail metrics with your partners, and stand up multivariate experiments across onboarding, paywalls, in-product prompts, and pricing. What you need to succeed Minimum Requirements: Bachelor's degree in a quantitative field (Statistics, Mathematics, Computer Science, Economics, Engineering, or similar) or equivalent practical experience. 5+ years of experience in data science, product analytics, or a similar quantitative role. Proficiency in SQL and Python (or R) for data manipulation, analysis, and modeling. Hands-on experience designing and analyzing A/B tests and interpre
About the Team Industrial Compute is building the world's most advanced AI infrastructure. Working alongside our capital partners, engineering teams, and construction organizations, we design and deliver large-scale, mission-critical compute campuses that power the next generation of AI. Our Design organization brings together engineering, construction, and digital design to ensure facilities are coordinated, constructible, and optimized from the earliest planning phases through deployment. About the Role We are seeking a BIM Designer & Coordinator to support the planning and design of large-scale industrial and mission-critical facilities. In this role, you will develop, coordinate, and maintain multidisciplinary BIM models across Civil, Electrical, Mechanical, Architectural, and Structural disciplines, enabling early design validation, constructability reviews, equipment planning, and cross-functional coordination. You will partner closely with internal engineering teams, external design consultants, contractors, and project stakeholders to produce coordinated BIM deliverables that improve design quality, reduce project risk, and support efficient execution across Industrial Compute's global infrastructure portfolio. Key Responsibilities: Develop and maintain conceptual and schematic BIM models for large-scale industrial and MEP-intensive facilities. Coordinate BIM models across Civil, Electrical, Mechanical, Architectural, and Structural disciplines. Create and maintain federated models used for design reviews, spatial coordination, constructability analysis, and clash detection. Model major building systems including equipment layouts, utility corridors, electrical rooms, mechanical rooms, structural framing, site infrastructure, and architectural constraints. Coordinate equipment clearances, maintenance access, routing zones, shafts, risers, utility entrances, and major MEP pathways. Translate engineering sketches, basis-of-design documents, equipment lists
From $130K/yr
About Flexport: At Flexport, we believe global trade can move the human race forward. That’s why it’s our mission to make global commerce so easy there will be more of it. We’re shaping the future of a $10T industry with solutions powered by innovative technology and exceptional people. Today, companies of all sizes—from emerging brands to Fortune 500s—use Flexport technology to move more than $19B of merchandise across 112 countries a year. The recent global supply chain crisis has put Flexport center stage as we continue to play a pivotal role in how goods move around the world. We are proud to have the support of the best investors in the game who believe in our mission, solutions and people. Ready to tackle global challenges that impact business, society, and the environment? Come join us. The Opportunity: Flexport IT is looking for a Senior Systems Engineer (Identity & Access) . In this role, you will design, implement, and administer our Identity and Access Management (IAM) solutions to ensure secure, efficient user lifecycle management. While you are our resident Okta expert, you are also a high-level IT generalist. You will oversee our broader SaaS ecosystem (Google Workspace, Slack, Jira) and endpoint management infrastructure (Jamf, Intune). Your expertise will drive automation, protect sensitive data, mitigate security risks, and maintain compliance in a heavily regulated industry. You will continually strive towards automation of toil. If you are still manually doing the same operational tasks 18 months from now, something has gone wrong. Flexport’s book of business is growing fast, but the promise of technology is that we can grow our business faster than our headcount. The automation you build will be a key factor in that effort. You Will IAM Architecture: Design, implement, and maintain the end-to-end lifecycle of Identity and Access Management platforms. Okta & Directory Trust: Administer Okta, directory services, Multi-Fact
About Mixpanel Mixpanel is the leading product intelligence and analytics platform, trusted by more than 29,000 companies to help understand how people use the products they build. By combining powerful analytics with AI that knows your business, Mixpanel helps teams see what’s working, diagnose what’s not, and decide what to build next. Learn more at mixpanel.com . About the Delivery Engineer Team As a member of the Delivery Engineer team, you will own the post-sales onboarding and data health for customers. Your goal will be to drive data trust, help embed Mixpanel into our customers’ data stack, and deliver on emerging AI integrations — delivering on the full value of Mixpanel as a self-serve analytics platform. You will work and consult with GTM team members and a diverse array of customers to successfully roll out product analytics to their organization and execute on technical projects and services that delight our customers. About the Role As a Delivery Engineer II, you will be on the front lines with our clients as they integrate Mixpanel into their core product development processes. You will lay the foundation for customers to adopt product analytics and get value from Mixpanel by delivering a world-class, on-time, and value-oriented onboarding experience. With your comprehensive project management knowledge, consultative approach, expertise in the analytics space, and technical knowledge of the modern data stack, you will lead our clients' first experience with Mixpanel as they incorporate product analytics into their data ecosystem as a foundational element. With your deep technical breadth and expertise in the analytics space, you’ll be the expert consultant on all things data and ecosystem. As Mixpanel and our customers continue to iterate on agentic integrations, you’ll own ensuring customers are able to build, assemble context, and integrate AI workflows with Mixpanel. Responsibilities Own onboarding and data health for our strategic and high-value c
About the Team The Future of Computing Research team is an applied research team within OpenAI’s Consumer Devices group. We study how AI systems perceive people and their surroundings, and we turn that research into capabilities for future products. Our work spans machine learning, sensing, and hardware, with a focus on building systems that work beyond controlled environments. About the Role We’re looking for a machine learning engineer to help shape how future AI systems understand the physical world and the people in it. The role focuses on multimodal perception and authentication, bringing together signals from cameras, microphones, and other sensors. You’ll work with specialized perception models and larger multimodal models, and partner with hardware, firmware, software, and product teams to bring new research into real-world systems. This role is based in San Francisco. We work in the office three days per week and offer relocation assistance. In this role, you will: Research and develop multimodal perception and authentication methods across visual, audio, and other sensing signals. Explore how specialized perception models and larger multimodal models can work together. Design data, training, and evaluation approaches that improve performance in real-world conditions. Study model behavior, robustness, and failure modes across sensing, data, and deployment environments. Integrate and validate new capabilities in real-time or resource-constrained systems. Work with hardware, firmware, software, and product teams to turn research into working systems. You might thrive in this role if you: Have a strong background in computer vision, audio or speech machine learning, multimodal learning, or sensing. Have experience developing specialized machine learning models, larger multimodal models, or both. Have brought research ideas into practical systems, prototypes, or products. Know how to design experiments, build evaluations, and investigate model behavior. Have wo
About the Role We’re hiring a Data Scientist to support Real Estate & Workplace (REW), a fast-moving global team focused on creating workplaces that help OpenAI’s people do their best work while scaling the company’s real estate and workplace operations. Our work is grounded in understanding how people use space and services, collaborate across physical and digital environments, and experience the workplace. REW’s scope spans portfolio strategy, design and construction, space planning, sustainability, workplace experience, and global operations. You’ll work comfortably across this broad, sometimes messy data landscape and build trusted relationships across the domain. The work informs high-impact decisions with immediate, visible effects—from where teams work and how space and services are allocated to which investments move forward and how workplace experiences evolve. This is a high-ownership Data Science role spanning analytical strategy and hands-on execution. Working at the forefront of AI-native analytics, you’ll help define the future of workplace operations at OpenAI rather than follow an established playbook. You’ll shape REW’s Data Science roadmap, identify where forecasting, experimentation, and optimization can drive impact, and translate business priorities into an analytical plan. You’ll own the stakeholder-facing execution layer—including owning agent-built dashboards, recurring reporting, models, and decision tools—along with analytical requirements, validation, adoption, and measurable business impact. In this role, you’ll be partnered closely with Finance, People Analytics, IT, and REW leaders. You’ll own problems end to end—from framing and prioritization through analysis, recommendation, delivery, adoption, and iteration—so the work drives measurable business outcomes. What You’ll Do Own ambiguous, high-impact problems end to end—from framing and prioritization through delivery, adoption, and iteration. Define success metrics and build measur
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. This internship will take place from January 25 - April 16 and you will need to be able to work out of our SF office during this time. What You'll Achieve: Conduct data analyses to gain insights about Notion and use these insights to uncover opportunities for improvements in our product and business. Communicate these insights with actionable recommendations to cross-functional teams (insights are useful, impact is even better!). Work with cross-functional partners across the product and business to learn about their functions and use data to advance their respective areas. Create metrics and build dashboards to monitor the growth and health of Notion. Communicate insights and recommendations effectively to leadership and have an impact on strategic decision-making. Qualifications: Pursuing a bachelor's or master's in a quantitative field such as Economics, Statistics, Applied Math, Engineering, Computer Science, or Natural Sciences. Must graduate before December 2027. This internship will take place from January 25 - April 16 and you will need to be able to work out of our SF office during this time. Previous research or internship
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE: Are you the person on your team who builds the agent everyone else ends up using? We're looking for an AI Engineer to join our Training Product team and do that at Baseten. You'll build AI-driven product features for the customers training and post-training frontier models on our platform, and you'll raise the ceiling on how Baseten itself uses AI internally, turning manual workflows into agentic ones that make every other team faster. You'll work directly with our research engineers to scope and build products, taking ideas from a research loop that already works internally to something customers can run themselves. This is a hands-on role with real autonomy. You'll pick the problems worth solving, build the harnesses, execution flows, and guardrails that make AI systems reliable, and own the results. If you've been shipping agents and want that to be the job, let's talk. EXAMPLE INITIATIVES: Take a look at these blog posts written by members of our team: Baseten Training: an autoresearch substrate Introducing Baseten Loops Harnesses are everything. Here's how to optimize yours. Building with NVIDIA Nemotron 3 Ultra and LangChain Deep Agents Code on Baseten RESPONSIBILITIES: Build and ship agentic product experiences, including chat-style and assistant-like interfaces, from prototype to GA. Design the harnesses, execution flows, and guardrails that make AI systems reliable in production. Build internal autom
About the Team The Corporate Security team ensures the physical safety and security of the organization's assets, operations, and personnel. We are committed to maintaining a secure environment that enables our team to focus on advancing artificial intelligence in a responsible manner. About the Role As a Protective Intelligence & Threat Analyst, you will identify, assess, and communicate physical security threats affecting OpenAI, its personnel, executives, operations, and assets. You will leverage open-source intelligence, social media, investigative tools, and other information sources to assess violent or disruptive threats, geopolitical developments, and emerging risks relevant to the company. The role will support a broad range of Protective Intelligence activities, including persons of interest investigations, behavioral threat assessment, executive and individual risk assessments, event and travel security assessments, and time-sensitive intelligence support to Corporate Security and cross-functional partners. You will turn complex and often incomplete information into clear assessments and actionable recommendations that help inform security and protective decisions. We are seeking candidates with intelligence experience, particularly in protective intelligence, threat investigations, or behavioral threat assessment. Successful candidates will understand the intelligence cycle, OSINT investigative techniques, corporate physical security, risk management, and threat assessment, and will be comfortable operating independently in a fast-moving environment involving sensitive and occasionally high-profile matters. This role could be based in San Francisco, CA, or may be a remote role for the right candidate. We use a hybrid work model of 3 days in the office per week. Relocation offered for those outside of the Bay Area. In this role, you will: Identify and investigate potential physical threats to OpenAI, its executives, employees, facilities, operations,
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Forward Deployed Engineer at Baseten, you will partner directly with customers to architect, build, and deploy high-scale production AI applications on Baseten’s platform. You’ll own the journey with customers from initial exploration to production deployment, translating ambiguous business goals into reliable, observable services with clear quality, latency, and cost outcomes. This role is a great fit for entrepreneurial engineers who want a front-row view into how modern companies adopt AI at scale and who enjoy working across product, software development, performance engineering, and customer-facing implementations. To be clear, this is an engineering role with hands-on coding and software development that also includes aspects of product management, technical customer success, and pre-sales solution engineering mixed in. EXAMPLE INITIATIVES Take a look at these blog posts written by members of our Forward Deployed Engineering team: Forward Deployed Engineering on the frontier of AI The fastest, most accurate Whisper transcription Deploy production-ready model servers from Docker images Deploy custom ComfyUI workflows as APIs RESPONSIBILITIES Develop and maintain software systems and product features using one or more general-purpose programming languages in a production-level environment, with a preference for Python due to its relevance in ML projects. Drive customer impact by designing, implementin
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: Modal is hiring a high-impact Solutions Architect to drive technical strategy across our most strategic enterprise accounts. You will operate as the executive technical counterpart to Enterprise Account Executives, leading complex evaluations, shaping infrastructure modernization roadmaps, and driving multi-product adoption across AI/ML workloads. This role is not demo support. It is a strategic, consultative position requiring strong architectural depth, executive presence, and the ability to influence 7–8 figure infrastructure decisions. You will work directly with CTOs, VPs of Engineering, and ML platform leaders to help them rethink how AI infrastructure should be built and operated. If you thrive in high-velocity technical sales environments and want to shape the infrastructure layer powering modern AI companies, this role is for you. What You’ll Do: Own the t
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As an Infrastructure Software Engineer at Baseten, you'll build and maintain components of our ML inference platform that powers production AI applications. You'll contribute to the core infrastructure, enabling developers to deploy, scale, and monitor ML models with high performance. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Infrastructure team: Multi-cloud capacity management Inference on B200 GPUs Multi-node inference Fractional H100 GPUs for efficient model serving RESPONSIBILITIES Develop infrastructure components for our ML inference platform using Python and Go Implement and maintain Kubernetes deployments for model serving Contribute to our inference orchestration layer for model deployments Build and enhance monitoring systems for model performance metrics Implement efficient resource management solutions for ML workloads Support infrastructure automation to improve ML deployment workflows Work closely with team members to implement technical solutions Help balance performance optimization with system reliability Participate in technical discussions around infrastructure improvements Learn and apply infrastructure best practices REQUIREMENTS Bachelor's degree or higher in Computer Science or related field Proficient coding abilities in one or more popular programming or scripting languages; Go proficiency is a plus Working knowledge of Kubernetes and containeriza
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Are you passionate about advancing the application of artificial intelligence? We are looking for a Software Engineer focused on ML performance to join our dynamic team. This role is ideal for someone who thrives in a fast-paced startup environment and is eager to make significant contributions to the exciting field of LLM Inference. If you are a backend engineer who thrives on making things faster and is excited about open-source ML models, we look forward to your application. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Model Performance team: Baseten Embeddings Inference: The fastest embeddings solution available The Baseten Inference Stack Driving model performance optimization RESPONSIBILITIES Implement, refine, and productionize cutting-edge techniques (quantization, speculative decoding, kv cache reuse, chunked prefill and LoRA) for ML model inference and infrastructure. Deep dive into underlying codebases of TensorRT, PyTorch, TensorRT-LLM, vllm, sglang, CUDA, and other libraries to debug ML performance issues. Apply and scale optimization techniques across a wide range of ML models, particularly large language models. Collaborate with a diverse team to design and implement innovative solutions. Own projects from idea to production. REQUIREMENTS Bachelor's, Master's, or Ph.D. degree in Computer Science, Engineering, Mathematics, or related field. Experience with one
Other cities to consider
More places hiring for this role
Get new python jobs in San Francisco, United States by email
Daily job updates · Unsubscribe anytime