Jobs in United States

Project Engineer in United States

1,423 active opportunities · Updated October 2026

Explore current project engineer jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

G
📍 United States· Full-time· Remote
✓ High-confidence listingCompany trend -100%

From $115.2K/yr

Quick readStrong listing-quality and freshness signals

GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An overview of this role As a Backend Engineer on the Chat Engine team, you'll build the engine behind GitLab Duo Chat, the conversational AI experience for GitLab. You'll work primarily in Python to build and maintain our agentic runtime: the Flow Registry, LangGraph flows, and the Duo Workflow Service. You'll also work in the GitLab Rails monolith, where Chat connects with the product. You'll own scoped parts of the system, ship small features and improvements with minimal guidance, and collaborate with the team on larger projects. You'll work alongside senior and staff engineers who will partner with you on design and support

PythonGitRestGraphql
G
📍 Israel, United Kingdom; Remote, United States· Full-time· Remote
✓ High-confidence listingCompany trend -100%

From $115.2K/yr

Quick readStrong listing-quality and freshness signals

GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. An overview of this role As an Intermediate Software Engineer on GitLab’s Vulnerability Management team, you’ll help build the core security workflows that GitLab Ultimate customers use to triage, prioritize, and act on vulnerabilities. You’ll work primarily in Ruby on Rails on backend systems, including security dashboards, vulnerability reports, and ingestion pipelines. You’ll take ownership of well-defined projects, solve technical problems with support from experienced team members, and collaborate with other engineers. You’ll also have opportunities to contribute beyond the backend when the work requires it. You will join a

JavaScriptJavaVueSQL
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $196.8K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Sr. Studio Software Engineer for Roblox Studio Platform, you will be a key contributor to the evolution of Roblox Studio, the primary IDE for making massive multiplayer online games on the Roblox platform. Studio provides the mission-critical tools for 3D modeling, animation, and the complete SDLC for millions of developers. We are looking for engineers who thrive on an adventure into the unknown and have experience across various systems, Operating Systems, Game Engines, and Application Frameworks. You’ll tackle projects involving: Core User Features: Architecting application frameworks, windowing systems, and code generation. “AI Native” Features: Pioneering scalable systems that extend to complex, agentic use cases. Foundational Architecture: Driving OS integration, extensibility, and customizability at the deepest levels. Central Backend Systems: Engineering the infrastructure to power a consistent, high-performance UX. Design Evolution: Collaborating with UX designers to translate Studio into a modern, consistent design language. You Will: Design and execute the technical direction to drive the future extensibility and adaptability of the application. Own and deliver complex techn

ReactAWSGitAI
R
📍 San Mateo, CA, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $196.8K/yr

Quick readStrong listing-quality and freshness signals

Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. As a Sr. Studio Software Engineer, not only will you be one of the coolest humans on the planet; you will also be a key contributor to the evolution of Roblox Studio which is the primary IDE for making massive multiplayer online games on the Roblox platform. Who wouldn’t want to do that? Studio provides tools for 3D modeling, animation, complete SDLC and testing for teams and individual game developers. This role does not require experience in the aforementioned. Large scale desktop application experience will suffice. You’ll work on projects involving: Core user features - UI frameworks,, windowing systems, code generation “AI Native” features - building well integrated AI features into Studio Foundational application architecture, including OS integration and building our our third party plugin ecosystem. Working with UX designers for overall design language for Studio both on visual, motion, and interactive elements. Building central backend systems You will: Design and execute on the architecture and technical direction that will be the future of our application Own and deliver complex technical projects from the planning stage through execution Work cross functionally, acro

ReactAWSGitAI
S
📍 Menlo Park, California, United States· Full-time
✓ Quality checkedCompany trend -93.3%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. At Snowflake, we empower both enterprises and individuals to reach their full potential. Our culture prioritizes impact, innovation, and collaboration, making Snowflake the ideal place to build ambitious projects, execute quickly, and advance technology — and your career — to the next level. The Role We are seeking a Senior Manager, Applied Field Engineering — AI/ML Product Specialists to lead a high-performing team of AI/ML specialists at the intersection of product, field, and customer success. In this hands-on leadership role, you will manage a team of Applied Field Engineers who are deep practitioners in Snowflake's AI/ML product portfolio — including Cortex AI, ML modeling, and agentic workflows. You will drive product adoption and customer outcomes, ensuring customers move beyond initial activation to unlock the full depth of Snowflake's AI/ML capabilities. Critically, you will serve as a strategic bridge between the field and Snowflake's product organization — translating customer experience into structured product insight that directly shapes roadmap priorities. You will work closely with Product Management, Engineering, and Sales leadership to ensure Snowflake builds the right things and customers realize their full potential. Responsibilities & Focus Areas Pro

AIGoRustMarketing
A
📍 United States· Full-time
✓ High-confidence listingCompany trend -98.9%

From $195K/yr

Quick readStrong listing-quality and freshness signals

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: Community Support Engineering (CS Eng) is responsible for the world-class technology, architecture, and solutions that power Airbnb’s global community support operations. The products and solutions we build support our customers, front line agents, and business process operations teams. They are a key driver of enabling Airbnb’s core business. As a member of CS Eng, you will have an opportunity to impact every person who travels or hosts on Airbnb. Under CS Eng, the Agent Core Products team has an opportunity for a Senior Engineer to help drive our initiatives. The team builds the core software that empowers our global support agents. The work has a direct impact on agent experience and impacts the quality of service we provide to our guests and hosts. The Difference You Will Make: We are seeking a highly skilled and motivated engineer who is passionate about making a difference through their work. As a Senior Software Engineer, you will work in a team of talented and diverse software engineers to build solutions that improve the agent experience. You will play a significant role in shaping the technical vision and then delivering a solution that is flexible, efficient and scales with the needs of the business. Each individual brings their own unique skill set, experiences, thought leadership and technical expertise to solve these technical challenges for Airbnb. We are a high-impact team focused on driving double-digit percentage improvements across our key metrics through various strategic projects. We continuously advance the quality and efficiency of our product, emp

JavaScriptTypeScriptJavaReact
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -100%

From $225K/yr

Quick readStrong listing-quality and freshness signals

Founding Engineer, Open Source Salary range — $225k – $300k | Equity — .15%-.35% | In-person NYC About Datalab Datalab trains models that read documents reliably at scale. The world's most important information is trapped in PDFs, scans, and files that can't easily be parsed, and getting it out correctly matters. From frontier AI labs processing training data to Fortune 500s like Siemens extracting decades of engineering records, Datalab is where businesses turn to when extraction has to be right. We’re at an 8-figure run rate with a team of 7. Anthropic is a customer. And we have hundreds more across FAANG, frontier AI labs, healthcare, finance, government, and legal. Our tools, chandra, surya, marker, and lift, have 70,000+ GitHub stars and broad developer mindshare. We're backed by founding members of OpenAI, FAIR, and Hugging Face. Role Overview We're looking for a Founding Engineer to develop and evangelize our open source repos. This includes chandra, surya, marker, pdftext, and lift, which collectively have over 70k Github stars. It also includes new tools we have yet to build and launch. As you work on our repos, you’ll also become the credible person to evangelize them. You’ll turn your own work into demos, benchmarks, tutorials, and launches. You’ll also support other launches across Datalab, especially when they touch open source components like the SDK. Our projects have real reach: 70k+ GitHub stars and users everywhere from frontier AI labs to Fortune 500s. Your job is to turn that reach into a thriving, engaged developer community through content, code, and showing up where developers already are. If you're the kind of engineer who’s energized by both building things and helping other people build, this is the role for you. Day to day: Own our open source repos and SDK: Drive chandra, surya, marker, lift, and our Python SDK forward as a hands-on contributor. Ship new features and improvements that the market cares about, help plan new versions, and ow

PythonGitAIGo
A
📍 United States· Full-time· Remote
✓ High-confidence listingCompany trend -98.9%

From $162K/yr

Quick readStrong listing-quality and freshness signals

Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: Imagine getting off a long international flight, standing in the freezing rain, attempting to contact your Airbnb host, only to realize that they are not responding because the Airbnb that you booked was not real. The Listing Integrity team’s ultimate goal is to prevent experiences like this by proactively detecting and removing fraudulent listings (Homes, Experiences, and Services) so that guests can search and book on Airbnb with confidence. The team uses data driven heuristics, machine learning models, and customer service operations in order to accomplish this goal. The Difference You Will Make: As a Software Engineer on the Listing Integrity team, you will be working with data scientists, designers, product managers, and customer service operations to innovate new ways we can stop bad actors in the ever evolving fraudulent listing creation and financial losses associated with it by collaborating across team boundaries. On this team, you must have the curiosity to dig deep into various end to end systems in order to understand how and where the fraud occurs. Your curiosity will be rewarded with finding projects that have outsized impacts on decreasing fraud losses and protecting Airbnb users from bad experiences while on vacation. A Typical Day: Work with large scale back end systems to detect fraudsters, using rules and connecting to productionalized machine learning models. Work collaboratively with cross-functional partners including machine learning engineers, product managers, operations and data scientists, identify opportunities for business impact, und

JavaScriptPythonJavaMachine Learning
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -83%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As a Product Engineer on the Dedicated Inference team, you'll shape the state-of-the-art developer experience for deploying and operating AI workloads in production. From the CLI and SDKs to APIs, observability, and debugging workflows, you'll build the tools customers rely on every day to manage mission-critical inference deployments. Few teams at Baseten have as much breadth and visibility as Dedicated Inference. The team is often at the forefront of new product development, giving engineers the opportunity to shape the experience of some of our most important customers. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Dedicated Inference team: Chains for multi-component workflows Asynchronous inference Model APIs for frontier models Model training built for production inference RESPONSIBILITIES Implement new features and products for the team Design ergonomic APIs and abstractions to solve customer problems Fix bugs and resolve customer issues with urgency Work across the stack - regardless of where you start, you’ll end up touching both React Components and Kubernetes Pods Work closely with the product and forward deployed engineering teams to develop and drive new product ideas REQUIREMENTS Bachelor's degree or higher in Computer Science or related field Proficient coding abilities in one or more popular programming or scripting languages; Python, Go, or Javascript proficie

JavaScriptPythonJavaReact
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -83%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE As an Infrastructure Software Engineer at Baseten, you'll build and maintain components of our ML inference platform that powers production AI applications. You'll contribute to the core infrastructure, enabling developers to deploy, scale, and monitor ML models with high performance. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Infrastructure team: Multi-cloud capacity management Inference on B200 GPUs Multi-node inference Fractional H100 GPUs for efficient model serving RESPONSIBILITIES Develop infrastructure components for our ML inference platform using Python and Go Implement and maintain Kubernetes deployments for model serving Contribute to our inference orchestration layer for model deployments Build and enhance monitoring systems for model performance metrics Implement efficient resource management solutions for ML workloads Support infrastructure automation to improve ML deployment workflows Work closely with team members to implement technical solutions Help balance performance optimization with system reliability Participate in technical discussions around infrastructure improvements Learn and apply infrastructure best practices REQUIREMENTS Bachelor's degree or higher in Computer Science or related field Proficient coding abilities in one or more popular programming or scripting languages; Go proficiency is a plus Working knowledge of Kubernetes and containeriza

PythonKubernetesRestMachine Learning
M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We're looking for engineers with deep AI/ML and low-level systems experience who want to build the best technical support experience in the world. This isn't a traditional support role — it's an engineering role where you happen to be closest to our customers. You'll split your time roughly 50/50 between working directly with customers and shipping fixes, features, and automation that improve Modal for everyone. When you help a customer debug a training run, you'll also fix the underlying issue in the platform. When you notice ten customers hitting the same friction point, you'll build the tooling or automation that eliminates it entirely. This role is for people who solve problems, not people who answer tickets. The problems you encounter are deeply technical and arise from running some of the most demanding AI workloads in the world. You'll be a member of our eng

B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -83%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Are you passionate about advancing the application of artificial intelligence? We are looking for a Software Engineer focused on ML performance to join our dynamic team. This role is ideal for someone who thrives in a fast-paced startup environment and is eager to make significant contributions to the exciting field of LLM Inference. If you are a backend engineer who thrives on making things faster and is excited about open-source ML models, we look forward to your application. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Model Performance team: Baseten Embeddings Inference: The fastest embeddings solution available The Baseten Inference Stack Driving model performance optimization RESPONSIBILITIES Implement, refine, and productionize cutting-edge techniques (quantization, speculative decoding, kv cache reuse, chunked prefill and LoRA) for ML model inference and infrastructure. Deep dive into underlying codebases of TensorRT, PyTorch, TensorRT-LLM, vllm, sglang, CUDA, and other libraries to debug ML performance issues. Apply and scale optimization techniques across a wide range of ML models, particularly large language models. Collaborate with a diverse team to design and implement innovative solutions. Own projects from idea to production. REQUIREMENTS Bachelor's, Master's, or Ph.D. degree in Computer Science, Engineering, Mathematics, or related field. Experience with one

PythonDockerKubernetesRest
M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role We're looking for an Engineering Manager to lead a team of highly experienced engineers building the infrastructure that powers Modal's serverless GPU platform. This is a hands-on leadership role — expect to split your time between technical contribution and people management depending on what the team needs. You'll set direction, remove blockers, and build a strong engineering culture as your team tackles hard problems in distributed computing, large-scale data handling, and performance optimization. Who You Are You're an experienced engineering leader who stays close to the work and builds alongside your team when it counts. You earn trust through technical depth, not title. You communicate clearly, help strong engineers move fast without cutting corners, and stay calm and pragmatic under pressure. You care as much about how your team gets to an answer as the answ

JavaLinuxAIC++
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -83%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We’re looking for a seasoned Frontend Engineer to craft performant and delightful user experiences across Baseten’s core platform. You’ll own critical parts of our web application stack and collaborate cross-functionally with product, design, and backend teams to launch impactful features that help users deploy and manage AI systems at scale. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Core Product team: Rolling Deployments Model APIs for frontier models Model training built for production inference RESPONSIBILITIES Design, implement, and maintain responsive, accessible, and user-friendly frontend interfaces using React and TypeScript Collaborate closely with product designers to turn complex ideas into elegant, intuitive UIs Optimize application performance and reliability, with a focus on rendering speed and responsiveness Drive major frontend initiatives, including partnering with backend teams to define APIs and test and refine end-to-end flows Establish best practices, and mentor other engineers on frontend technologies Build reusable component libraries and frontend infrastructure that accelerate product development Partner with backend and platform teams to define and refine APIs and end-to-end flows REQUIREMENTS 5+ years of experience building production-grade web applications Deep expertise in React, TypeScript, and modern web development tooling Track record of bu

TypeScriptReactMachine LearningAI
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -83%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We’re seeking a GPU Kernel Engineer to join our team at the cutting edge of AI acceleration, where your code directly impacts the performance of state-of-the-art machine learning models. As a GPU Kernel Engineer, you'll craft the foundation that powers modern AI workloads, optimizing every microsecond of computation to enable breakthrough applications. You'll work in a fast-paced, intellectually stimulating environment where technical excellence is paramount and your contributions directly influence production systems serving millions of users across numerous products. This role offers exceptional growth potential for engineers passionate about low-level optimization and high-impact systems work. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Model Performance team: Baseten Embeddings Inference: The fastest embeddings solution available The Baseten Inference Stack Driving model performance optimization RESPONSIBILITIES Core Engineering Responsibilities Design and implement high-performance GPU kernels for key ML operations, including matrix multiplications, attention mechanisms, and mixture-of-experts routing Write and optimize code using CUDA, PTX assembly, and architecture-specific techniques Apply advanced performance optimization methods such as memory coalescing, warp-level programming, tensor core acceleration, and compute/memory overlap Performance & Innovation Impl

AWSMachine LearningAIC++
🔔

Get new project engineer jobs in United States by email

Daily job updates · Unsubscribe anytime