ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We’re looking for a customer-obsessed software engineer to come ship with us. You’ll own features like multi-node training and products like serverless reinforcement learning (RL) from conception to MVP (and from MVP to GA!). You’ll work through the stack, architecting solutions from API and UI down to our infrastructure layer. You’ll fine tune models yourself to develop an understanding of user workflows. You’ll work closely with research engineers leveraging state-of-the-art training techniques to build experiences that accelerate model development and solve for real pain points. If you’re excited to dive deep into the training, let’s talk! THE PRODUCT Take a look at what we’ve built so far: Overview of the product so far Training docs overview Story of the Training product Research we've done EXAMPLE INITIATIVES Checkpointing Pipeline: Our checkpointing pipeline starts with automated checkpointing, a feature that ensures that versions of models created during training are automatically backed up to the cloud. Users are able to then deploy checkpoints seamlessly into inference servers, providing point-and-click integrations into inference frameworks like vLLM and Baseten’s Inference Stack. This enables customers to quickly evaluate the performance of their checkpoints with real traffic. Multinode training: Multinode training enables customers to easily run training jobs across multiple compute nodes, enablin
Jobiba hiring network
Partner Technology Solutions Engineer Jobs
10,000 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current partner technology solutions engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. Product at Baseten Product at Baseten is a nascent function. Our company today has a strong engineering culture, is heavily customer-obsessed, and moves fast. We're building the product function now, and you'd be one of the people who defines it. You'll work directly with our founders and with some of the best systems and AI engineers and you'll set the standard for what product looks like here. PMs at Baseten don't sit above engineers - you earn ownership by being technical, finding the truth in front of customers, building great cross-functional relationships, and just shipping great product experiences. The role Once a model is deployed, keeping it fast, reliable, and economical at scale is where production inference is won or lost. You'll own the surface that makes that happen: how deployments autoscale, how traffic is routed, how the system fails over, and how workloads scale across clusters and regions. You'll own these as products end to end - both how they work under the hood and how customers configure and observe them - and you'll help set and define the roadmap that infrastructure and product teams alike can build towards. This space is largely still evolving - think Cloud Infrastructure in mid-2000s. Your job is to make it 10x easier to reliably scale and serve AI models in production and set the market standard. Impact and outcomes you'll drive You will own how workloads scale and where they land — autosca
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE: Baseten’s Model Performance (MP) team is responsible for ensuring the models running on our platform are fast, reliable, and cost‑efficient. As part of this team, you’ll focus on Model APIs — the infrastructure powering our hosted API endpoints for the latest open‑source models. This work spans distributed systems, model serving, and developer experience. You’ll join a small, high‑impact team operating at the intersection of product, model performance, and infra, helping to define how developers interact with AI models at scale. RESPONSIBILITIES: Design, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups Productionize performance improvements across runtimes with deep understanding of their internals: speculative decoding implementations, guided generation for structured outputs, custom scheduling and routing algorithms for high-performance serving Build comprehensive benchmarking frameworks that measure real-world performance across different model architectures, batch sizes, sequence lengths, and hardware configurations Productionize performa
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We are seeking a Sales Manager to help lead a team of Account Executives within our Startups segment. This hire will be responsible for building and coaching a team of high-performing sales reps, driving revenue growth, and partnering closely with product and engineering to bring our cutting-edge AI infrastructure to customers. RESPONSIBILITIES Lead and mentor a team of Startups Account Executives to consistently exceed pipeline and revenue goals. Help define and execution on go-to-market strategy for our fastest growing customer segment. Hire and scale the team by recruiting, interviewing, and onboarding top talent Collaborate cross-functionally with Marketing, Product, and Engineering to align customer needs with Baseten’s product roadmap. Be deeply engaged with the product, enabling reps to have highly technical conversations with prospects and customers. Foster a culture of accountability, learning, and collaboration within the sales team. REQUIREMENTS 4+ years of closing sales experience, with 2+ years in management leading high-performing teams. Strong technical acumen, ideally with background in AI infrastructure, cloud infrastructure, or developer platforms. Comfortable operating in the weeds with technical products and guiding reps through complex deals. Proven track record of success in high velocity sales environments. Based in San Francisco or New York City and open to coming in office at least 3 d
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is building a world-class team, anchored in our San Francisco and New York offices and increasingly growing across the globe. We believe the best talent can come from anywhere, and our ability to hire and support that talent is mission-critical. As our Immigration and Mobility Manager, you will make global hiring operationally seamless. You’ll be Baseten’s in-house expert on immigration, relocation, and international employment strategy, ensuring that exceptional candidates from all over the world can confidently build their careers with us. This is a high-stakes, high-impact role. Speed, clarity, and correctness matter deeply in immigration and mobility. You will own the end-to-end experience, navigate a rapidly evolving U.S. immigration landscape, and design the systems and policies that enable Baseten to hire globally while delivering an exceptional employee experience. Over time, you’ll also help shape where and how we expand internationally, advising on global hiring models, new hubs, and employment structures that support our long-term growth. RESPONSIBILITIES Own the end-to-end immigration lifecycle, managing all U.S. visa processes (e.g., H-1B, O-1, J-1, TN, E-3, L-1, EB-2/3 PERM) from offer stage through renewals and permanent residency. Oversee and project manage visa sponsorships executed through Employer of Record (EOR) partners Serve as Baseten’s internal immigration expert, partnering clo
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE OPPORTUNITY We are looking for Senior Software Engineers to join our team. This is a specialized, high-impact role sitting at the intersection of high-performance computing (HPC) and Large Language Model (LLM) engineering. You will not just be building the automated "speedometer and diagnostic" suite for our next-generation AI infrastructure; you will be defining the roadmap, driving key technical decisions, and taking full ownership of the future of this work. RESPONSIBILITIES Benchmarking : Evaluate, run and automate standard LLM quality benchmarks (GSM8K, MMLU) alongside custom performance suites for specific workloads (e.g., long-context window, KV cache reuse, disaggregated serving). DevEx Improvement : Develop and maintain internal GPU-enabled development environments (similar to GitHub Codespaces). You will ensure the team has seamless, high-performance "dev machines" optimized for model experimentation. Tool Development : Build and contribute to open-source tools such as InferenceMAX and genai-bench to automate model evaluation, benchmarking and analysis. System Profiling : Use profilers like PyTorch Profiler, NVIDIA Nsight Systems and py-spy to collect performance profiles, identify bottlenecks, and debug the compute/networking stack. Monitoring & Observability : Develop real-time dashboards and alerts to monitor system health, model startup times, and runtime performance. Continuous Integration : Auto
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. At Baseten, we are building the global operating system for distributed, heterogeneous AI hardware. We believe that as LLM and multi-modal workloads scale, the network is the computer. We are looking for foundational engineers to lead our GPU Networking efforts, making RDMA a first-class building block in our infrastructure and unlocking the next generation of distributed inference optimizations. THE OPPORTUNITY Networking and compute are no longer separate disciplines; they are converging. The massive throughput of H100, B200, and NVL72 architectures enables and demands a new approach where communication is co-optimized alongside computation. We are entering an era where the network is an active accelerator, leveraging smart hardware offloads and direct interconnects to ensure that data movement operates at wire-speed. In this role, you will go beyond network configuration to architect the software fabric that unifies thousands of GPUs into a cohesive operating system. While you will leverage the best of the open-source ecosystem, you won't be limited by it. Where off-the-shelf solutions stop, you will build from scratch, engineering the primitives required to co-optimize communication and compute for Disaggregated Serving, Wide Expert Parallelism (WideEP), and lightening cold starts. WHAT YOU'LL DO Make RDMA First-Class: You will work on integrating RDMA/RoCE/InfiniBand capabilities directly into our inference stack,
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE This role sits at the frontier of our research agenda. You will pursue open problems at the intersection of post-training methodology and performant inference, and then collaborate with research engineering to translate findings into production systems. A meaningful portion of your time will be dedicated to research that deepens our understanding of how models learn, alignment, and architectural efficiency — questions that may not have immediate product application. The remainder will be directed toward research that solves concrete problems for Baseten's platform and customers, who are the fastest growing AI companies in the world like Cursor, Lovable, and Notion. We are looking for someone with sharp research taste and genuine creative instinct for problem selection. Someone who can identify questions that matter, design clean experiments to answer them, and push the state of the art. The environment here is not theoretical, but rather research that can be validated with eager customers who are serving billions of tokens a second. RECENT RESEARCH Towards infinite context windows: neural KV cache compaction Dense, on-policy or both? Repeated kv cache for long-running agents Distillation without the dark – replicating black-box on-policy distillation on Baseten RESPONSIBILITIES Define and pursue a research agenda spanning both foundational and applied work, with the applied component connected to Baseten's pla
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE The largest, most demanding enterprises are starting to run on Baseten, and they arrive with a range of security, compliance, and procurement requirements. As a Senior Engineer on Baseten's enterprise engineering team, you'll build the capabilities that enable large organizations like Writer, HubSpot, and Notion to succeed on Baseten. Enterprise engineering authors the core building blocks, APIs, and user experiences powering the Baseten platform: identity and access management, billing, regional isolation, and self-hosted and single-tenant deployment options. This is deep product and systems work across the full stack, from designing authentication and authorization systems using standards like OAuth and OIDC to shipping the admin experiences enterprise IT teams use to manage their organization. EXAMPLE INITIATIVES Recent and upcoming work on the team: Fine-grained authorization for users, service accounts, and agentic workloads SSO and SCIM support, allowing customers to centralize and automate access to Baseten Expanding the billing platform to support evolving pricing models, advanced data exports, and controls to manage spend In-product management and enforcement of customer compliance requirements like data residency and HIPAA Securing network paths in and out of a customer's models with private connectivity and ingress and egress restrictions Allowing customers to run Baseten inside their own VPC, on-pr
Company Description Novo is a venture-backed fintech that simplifies banking for small businesses. Since our Fall 2018 beta launch, we've expanded offerings from free checking accounts and debit cards to lending products, business tools (invoices, bookkeeping, and more), and integrations. Today, Novo is a powerfully simple banking platform that serves over 250,000 small businesses. In addition to providing smart tools built for entrepreneurs to better run and grow their businesses, Novo has processed billions in transactions in partnership with a number of established banking partners. Our vision is to be the go-to platform serving small businesses – from launch to everyday – so business owners can focus on growing, while Novo provides seamless money movement, money storage, and access to capital. We'd like to look back 5–10 years from now and know that we helped new generations of small businesses succeed because of the work we did at Novo. Novo raised $170 million in venture capital and is backed by leading investors, including Stripes, Valar Ventures, Crosslink Capital, and Notable Capital (formerly GGV). Learn more at https://www.novo.co . Role Description Novo is seeking a Senior Product Manager to own the platform surface area that connects, activates, and deepens value for small businesses in their first 90 days and beyond. This role spans our DDA onboarding to funding activation funnel, invoice workflow, external integrations ecosystem — three interconnected surfaces that together determine whether a new Novo customer becomes an engaged, transacting one. The role will also explore how to apply AI to enhance how SMBs interact with Novo. The ideal candidate is a hands-on product leader who combines platform and tooling depth with a growth mindset: someone who can own complex product workflows end-to-end, run rigorous experimentation, and is comfortable getting their hands dirty in LLMs. This is a high-impact, high-visibility role at the intersection of product
About the Role & Team This is Amplitude's first People Team hire in India, and we're being deliberate about who we bring in first. People Operations is at an inflection point. A lot of what takes hours today—HR tickets, Workday updates, onboarding tasks—will be handled by AI-assisted workflows in the near future. We're building toward that now, and this role is part of that foundation. You'll join the global People Operations team as a dedicated shared services resource, handling the operational work that keeps our EMEA and AMER specialists focused on higher-complexity, in-market problems—and supporting People Business Partners globally when they need an extra set of hands. In the near term, the work is largely transactional: resolving HR helpdesk tickets, processing Workday data requests, coordinating onboarding logistics. That's not a knock on the role—it's the job, and doing it with speed, accuracy, and care matters enormously. Over time, as we invest in AI tooling and automation, that transactional load will shift. When it does, this role shifts with it: taking on process improvement projects, helping document and refine how the team operates, and growing into a contributor who shapes practices rather than just follows them. We'll support that progression. Because this is a remote role with no local People Ops team around you, we need someone who is genuinely self-directed. You'll be given clear priorities and good systems to work in, but you'll own your own day. If you thrive in structured work but feel motivated by the promise of doing more over time, this is a good fit. As a People Operations Analyst, you will: Own shared services and HR helpdesk Resolve Tier 1–2 employee inquiries through our HR ticketing system (Console AI) on behalf of EMEA and AMER regional teams: policy questions, employment verifications, leave requests, and benefits inquiries Triage issues that need specialist input—Benefits, Payroll, Legal, or regional People Ops—with clear contex
About Supabase Supabase is the Postgres development platform, built by developers for developers. We provide a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search. All services are deeply integrated and designed for growth. About the Role We're looking for engineers to join the team building Supabase Lite , a lightweight, TypeScript-native implementation of Supabase. It runs interchangeably on SQLite and Postgres and ships a PostgREST and Auth compatible API, so applications written against @supabase/supabase-js work as-is. It exists because AI builders and platform partners need a database they can give to every prototype without the cost or the wait: sub-second provisioning, a footprint small enough to run inside a sandbox, and an upgrade path to full Supabase when an app graduates to production. This is a role with a lot of agency. Small team, working product, patterns still to be set - you'll own large areas end to end and make real decisions from day one. What You'll Be Responsible For In this role, you'll: Grow API compatibility with the Supabase client, keeping behavior faithful to hosted Supabase and backed by conformance tests Bring more of the Supabase stack to SQLite Build and harden the path for upgrading a project from Supabase Lite to full Supabase Own the developer experience end to end: the CLI, local tooling, and docs Build and harden the hosted offering on a scale-to-zero, near-zero-cost footing, kept loosely coupled so the implementation can be swapped without a rewrite Work directly with design partners and turn real-world migration and integration friction into a sharper product Engage with the open-source community and contribute back to the broader Supabase stack. You Might Be a Good Fit If You Have substantial backend or full-stack experience with strong fluency in TypeScript, and have built or contributed to a backend system, framework, or developer platform before Are a strong generalist,
Supabase is the open-source Postgres development platform that 7M+ developers and thousands of enterprises depend on every day. We provide a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search. All services are deeply integrated and designed for growth. About the Role We’re looking for a Product Manager, Marketplace to join our Product Team to lead our marketplace and integration surfaces. The marketplace is the two-sided surface of the Supabase ecosystem. On one side, Partners and developers build integrations on top of Supabase primitives like Auth, Storage, and Edge Functions, then publish them for the world to use. On the other, our developers come to discover new capabilities and wire them into their projects without leaving their workflow. Done well, it is the shortest path between a developer with a problem and the exact capability that solves it. That surface is about to matter more than ever. As the number of capabilities explodes, and as agents begin to assemble software on a developer's behalf, the marketplace becomes the distribution and discovery layer for the whole ecosystem: how both developers and, increasingly, their agents find, trust, and compose what they need. This role owns it. What you'll be responsible for: Own the product strategy and roadmap for the marketplace, the two-sided surface where Partners and developers publish integrations and where Supabase's developer customers discover and install them. Make it effortless for a developer to find the right capability and integrate it. Own discovery, search, listings, and the whole path from install to working integration, so a new capability connects to Supabase primitives and runs inside a developer's project with minimal glue code. Make publishing an integration fast and rewarding, so Partners and developers bring their best work to the marketplace and keep it current, and so the catalog stays deep where developers need it. Set the quality
Supabase is the Postgres development platform, built by developers for developers. We provide a complete backend solution including Database, Auth, Storage, Edge Functions, Realtime, and Vector Search. All services are deeply integrated and designed for growth. About the Role We’re looking for a Senior Postgres Engineer to join our Postgres Team and help maintain and expand the stability and functionality of our hosted Postgres offering . You’ll work closely with customers, partners, product, engineering, support and success, helping us maintain a secure, stable, performant, and functional Postgres foundation. This role is ideal for someone who thrives in async, fast-paced environments and is excited about building developer tools that scale to millions. What you'll own: Build and maintain PostgreSQL extensions in C and Rust, with a deep understanding of internals — parser, planner, WAL mechanics, and MVCC Diagnose and resolve issues in managed PostgreSQL deployments, including custom extension failures, core dump analysis, and performance bottlenecks Own idempotent deployment pipelines across thousands of running PostgreSQL instances, including testing and rollout strategies Manage complex extension ecosystems — compatibility, upgrade paths, and conflict resolution Work with PostgreSQL's background worker framework, shared memory management, and hook system to build reliable, scalable functionality Collaborate closely with customers, partners, product, engineering, support, and success teams to maintain a secure, stable, and performant Postgres foundation What you bring: Deep expertise in PostgreSQL internals - query planner, executor, and storage engine mechanics Proven experience building PostgreSQL extensions in both C and Rust Strong knowledge of PostgreSQL's permission model : RLS, roles, and grant systems Experience troubleshooting production issues in managed PostgreSQL environments, including custom extension issues, performance bottlenecks, and resource co
Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the Role We're looking for an AI-native product growth lead to own Replit's entire paid acquisition and lifecycle marketing engine as the single DRI. This is an IC role for someone who builds systems, not just campaigns. Someone who uses AI and automation aggressively to do what traditionally requires an entire team. You don't need an engineering background. You need the instinct to build: when you see a repetitive task, you reach for Replit or an API before you reach for a spreadsheet. You'll have data science and data engineering partners for infrastructure and modeling, and brand marketing will set creative direction. Your job is to turn that direction into a closed-loop optimization system: a self-improving system that turns performance data into better creative, optimizes spend across channels in near-real-time, and compounds every insight so nothing learned is ever lost. Each cycle, the system gets smarter. That's the engine you'll build and own. The ideal candidate has deep channel expertise across paid search, paid social, and lifecycle, but their real edge is building AI-powered workflows that scale creative production, automate measurement, and compound institutional knowledge. You Will Own full-funnel performance marketing across paid search & social, ASO, and lifecycle Build a self-improving creative engine: AI-driven ad generation, testing, and iteration that scales without scaling headcount Maintain a persistent knowledge layer so every experiment, result, and creative insight compounds automatically into the next cycle Own conversion signal quality end-to-end: right events, right audiences, right attribution, partnering with DE on the infrastructure Run a structured experimentation program wher
Get new partner technology solutions engineer jobs by email
Daily job updates · Unsubscribe anytime