Jobs in United States

Staff Research Scientist in United States

1,831 active opportunities · Updated October 2026

Explore current staff research scientist jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We are looking for strong engineers with experience and interest in designing, building, and maintaining the novel, high-performance systems that make up our serverless platform. Requirements: 5+ years of experience writing high-quality production code Experience building high-performance distributed systems at a large scale (the more battle scars, the better) Strong cloud skills Strong knowledge of low-level operating system foundations (Linux kernel, file systems, containers, etc.) Experience with performance engineering (tell us a story of when you shaved off a few milliseconds!) Ability to work in-person in our NYC or SF office. Prior experience with Rust is nice to have, but not required. Ability to participate in on-call rotation and respond to production incidents.

LinuxRestAIGo
M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: At Modal, we sell cloud services atop which our customers run their critical production systems. As a rapidly growing new cloud infrastructure company, we seek to improve our reliability dramatically while scaling the size of our platform, customer base, and our team. This role is for people who are deep systems thinkers, love stacking nines, and thrive from making others move faster at scale. Responsibilities include: Identifying architectural changes to improve reliability and performance. Fostering a culture of reliability across Modal’s engineering organization. Defining and implementing operational processes such as deployments, upgrades, etc. Operating systems like Kubernetes, Postgres, Redis, etc. Participating in on-call rotations, and responding to production incidents. Requirements: 5+ years of experience writing high-quality production code. 2+ years of

RedisAWSKubernetesCI/CD
M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We are looking for strong engineers with experience in making ML systems performant at scale. If you are interested in contributing to open-source projects and Modal’s container runtime to push language and diffusion models towards higher throughput and lower latency, we’d love to hear from you! Requirements: 5+ years of experience writing high-quality, high-performance code. Experience working with torch, high-level ML frameworks, and inference engines (vLLM or TensorRT). Familiarity with Nvidia GPU architecture and CUDA. Experience with ML performance engineering (tell us a story about boosting GPU performance — debugging SM occupancy issues, rewriting an algorithm to be compute-bound, eliminating host overhead, etc). Nice-to-have: familiarity with low-level operating system foundations (Linux kernel, file systems, containers, etc).

LinuxRestAIGo
M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We're looking for strong backend engineers who love building a developer tools used by the largest AI companies in the world. You’ll be building for things at scale, but also for new AI workflows that change every day. Requirements: Experience building and shipping modern web applications end-to-end. We care more about what you’ve built than how many years you’ve been building. Comfort working across the stack: TypeScript on the frontend, Python services on the backend, and ClickHouse for data and analytics. Deep knowledge of observability tools and patterns used for large-scale workloads such as custom sandboxes, training and inference for large language (LLM) and diffusion models. Experience with at least one of: billing/payments systems, B2B SaaS tooling, or enterprise software, or LLM / diffusion models inference and training loads. Strong product instincts; yo

TypeScriptPythonAIGo
M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We’re looking for strong engineers with experience building developer tools that users love to work with. Our ideal candidate is someone with a demonstrated drive to build beautiful interfaces that enhance developer productivity. Requirements: 5+ years of experience developing high-quality Python libraries with broad user-bases, ideally including some experience maintaining open-source software. Knowledge of advanced Python features, especially async programming. A strong product sense that manifests as a focus on developer ergonomics and productivity. A high level of customer empathy, good communication skills, and an openness to working directly with our users to help solve their problems. Ability to participate in on-call rotation and respond to production incidents. Ability to work in-person in our NYC or Stockholm office. Any of the following would be a plus:

TypeScriptPythonAIGo
M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We're looking for a Growth Engineer to own the technical foundation of Modal's marketing and developer-facing web surfaces: the marketing site, docs site, growth landing pages, high-profile microsites, forms, analytics instrumentation, and the integrations that help users discover, understand, and get started with Modal. This is a frontend-heavy role for someone with strong product taste, web engineering craft, and a business-owner mindset. You'll partner with Product Engineering, Design, Data, and Growth to ship polished, measurable web experiences from high-profile projects like the GPU Glossary and LLM Engine Advisor to internal tooling that helps teams publish content faster. When this role is going well, Modal launches new pages, docs experiences, campaigns, and experiments quickly without sacrificing performance, craft, or measurement. In this role you will:

TypeScriptAIGoRust
R
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -100%
Quick readStrong listing-quality and freshness signals

About Runway Runway is a collaborative business planning platform designed to make business intuitively understandable for everyone. Our Mission : To make business accessible and understandable to everyone. We believe that teams that understand the “why” behind their work are more productive and make better decisions. True alignment and collaboration come from having a shared source of truth that everyone understands. Our Approach : Runway replaces traditional spreadsheets with a modern planning platform that brings clarity and context to business operations for all teams — not just finance. Just as Figma made design accessible across the organization, Runway does the same for business planning. Why It Matters: Understanding requires more than just access to numbers; real collaboration happens when teams see how their work fits into the bigger picture. By providing this context, Runway helps teams save time and move faster. Our Customers : World-class companies like AngelList, Superhuman, Stability.AI, ConvertKit, Lambda Labs, Lob, and SandboxVR rely on Runway to run their businesses more efficiently. Our Investors : We are supported by a select group of investors that we admire, including Garry Tan (YC & Initialized), a16z, Elad Gil, Naval Ravikant, Dylan Field (founder of Figma), Eric Ries, Claire Hughes Johnson (COO of Stripe), Henry Ward (founder of Carta), Akshay Kothari (COO of Notion), Eugene Wei, Lenny Rachitsky, Nikita Bier, Scott Belsky, Soleio Cuervo, Balaji Srinivasan, and many others. Working at Runway We're remote-first, so you can work from anywhere in North America. We strive to be clear in our communication and over-communicate by default. It’s early in our journey, so you'll have an opportunity to shape not just our product, but the company itself: how we work together, what makes us stand out, and who we hire. Below are the values we share as a team — if these resonate with you, you may enjoy working here. (And if you don't like them, please t

KubernetesRestAIGo
R
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -100%
Quick readStrong listing-quality and freshness signals

About Runway Runway is a collaborative business planning platform designed to make business intuitively understandable for everyone. Our Mission : To make business accessible and understandable to everyone. We believe that teams that understand the “why” behind their work are more productive and make better decisions. True alignment and collaboration come from having a shared source of truth that everyone understands. Our Approach : Runway replaces traditional spreadsheets with a modern planning platform that brings clarity and context to business operations for all teams — not just finance. Just as Figma made design accessible across the organization, Runway does the same for business planning. Why It Matters: Understanding requires more than just access to numbers; real collaboration happens when teams see how their work fits into the bigger picture. By providing this context, Runway helps teams save time and move faster. Our Customers : World-class companies like AngelList, Superhuman, Stability.AI, ConvertKit, Lambda Labs, Lob, and SandboxVR rely on Runway to run their businesses more efficiently. Our Investors : We are supported by a select group of investors that we admire, including Garry Tan (YC & Initialized), a16z, Elad Gil, Naval Ravikant, Dylan Field (founder of Figma), Eric Ries, Claire Hughes Johnson (COO of Stripe), Henry Ward (founder of Carta), Akshay Kothari (COO of Notion), Eugene Wei, Lenny Rachitsky, Nikita Bier, Scott Belsky, Soleio Cuervo, Balaji Srinivasan, and many others. Working at Runway We're early, so you'll have an opportunity to shape not just our product, but the company itself: who we work with, and how we work together. We strive to be clear in our communication and over-communicate by default. We're remote-first, so you can work from anywhere in North America. However, we also believe in the value of face-time to solve really hard problems, so we: Meet together as a company every quarter in our San Francisco office. Open up of

TypeScriptReactGitRest
C
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Our team is a fast-growing group of researchers and engineers focused on building reliable ML systems and pushing the boundaries of LLM inference efficiency. We develop techniques that improve how models execute in production, driving lower latency, higher throughput, and consistent quality across diverse workloads. As an engineer on this team, you’ll work across the inference stack to improve core performance metrics by diving deep into model execution, identifying bottlenecks, and developing innovative optimizations. You’ll collaborate closely with modeling and systems teams to experiment, measure, and ship improvements that meaningfully accelerate inference. As the team evolves, you’ll have opportunities to build expertise in advanced performance techniques, including GPU/CUDA optimizations, kernel-level improvements, and model execution strategies for MoE and large-scale architectures. Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, e

PythonGitRestAI
C
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Member of Technical Staff, Search Why this role? We are looking for talented individuals to help us develop state-of-the-art models for information retrieval as part of our Search team. This group is working on a range of tasks including training our embedding and reranker models. You'll have the opportunity to revolutionise people's search experience by contributing to building an intelligent, efficient and precise search system and you would have a lot of opportunity to try new things out, innovate, and productionize your ideas. Your work will specifically focus on advancing semantic search techniques to improve accuracy and efficiency, involving working with a wide range of novel technologies and collaborating with other teams to integrate your work into our search infrastructure. We're looking for someone who is passionate about search and has a strong background in information retrieval. Candidates should have experience working with a wide range of technologies and have worked collaboratively with other teams in the past. As a Member of Technical Staff on this team, you will: Design, train and improve upon cutting-edge sea

PythonGitAIC++
C
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? At Cohere, we believe in the power of multimodal AI to revolutionise the way we interact with technology. Our engineering teams push the boundaries of what's possible, and we're looking for talented individuals to join us on this exciting journey. With an exceptional ratio of compute resources to engineers, we provide an ideal environment for you to explore, innovate and shape the future of AI. July 31st 2025 - Cohere's Multimodal team Introduced Command A Vision: Multimodal AI Built for Business. At release our new flagship vision-language model: ● Consistently outperforms major models like Llama 4 Maverick, Mistral Medium/Pixtral Large, and GPT4.1 ● 83.1% average benchmark (73.5% MathVista, 90.9% ChartQA...) ● Built for the real world - 112B parameters running on just 2 GPUs ● Open weights live on HuggingFace With a focused team, breakthrough performance doesn't require breakthrough compute. Focus on the things that matter, and join the team. As a Member of Technical Staff with a focus on Multimodal AI, you will: Design and develop cutting-edge multimodal AI systems, integrating various modalities such as text,

PythonGitMachine LearningAI
C
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -79.2%

£250K – £310K/yr

Quick readStrong listing-quality and freshness signals

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! The Chief of Staff is a strategic leadership role focused on optimizing sales operations, enabling revenue growth, and ensuring alignment between sales execution and company objectives. This exciting career opportunity offers a blend of operational oversight, strategic planning, and cross-functional collaboration. Key Responsibilities Sales Strategy & Execution: Partner with the CRO to define and implement sales strategies, forecasts, and performance metrics. Operational Excellence: Streamline sales processes, including pipeline management, CRM optimization, and sales enablement tools. Cross-Functional Coordination: Work closely with marketing, product, and customer success teams to ensure cohesive go-to-market execution. Performance Analytics: Monitor sales KPIs, conduct data-driven insights, and recommend adjustments to improve conversion rates and revenue attainment. Team Leadership: Contribute to sales operations, business development, and support functions, fostering a high-performance culture. Stakeholder Management: Serve as a liaison between sales leadership and executive teams, ensuring clear communication and strat

C
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Are you energized by leading the design of high-performance, scalable and reliable machine learning systems? Do you want to set technical direction and help shape the next generation of AI platforms powering advanced NLP applications? We are looking for a Lead Member of Technical Staff to join the Model Serving team at Cohere. The team is responsible for developing, deploying, and operating the AI platform delivering Cohere's large language models through easy to use API endpoints. In this role, you will provide technical leadership across multiple teams, driving the architecture and strategy for deploying optimized NLP models to production in low latency, high throughput, and high availability environments. You will serve as a key point of contact for customers, leading the design of customized deployments to meet their specific needs, and mentoring engineers to raise the technical bar across the team. You may be a good fit if you have: 8+ years of engineering experience running production infrastructure at a large scale, with a track record of technical leadership Demonstrated experience leading the architecture

AWSAzureGCPKubernetes
M
📍 Boise, ID - Main Site, United States· Full-time
✓ Quality checkedCompany trend -75%

Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Position Overview: As a Staff/Principal Equipment Engineer within PROCT ADT, you will provide technical leadership to improve equipment performance, capability, and global manufacturing outcomes. You will lead complex, multi-site initiatives to resolve equipment challenges, enable technology node readiness, and drive yield and defectivity improvements. This role requires deep expertise, strong ownership, and the ability to influence across teams and suppliers to deliver measurable impact in cost, performance, and scalability. Responsibilities Lead global ownership of equipment performance across toolsets, driving improvements in stability, availability, matching, and efficiency. Direct complex, multi-fab root cause analysis for equipment-driven yield, defectivity, and reliability issues using data-driven and physics-based approaches. Provide technical leadership in tool hardware and chamber behavior to expand process capability and improve performance limits. Partner with Technology Development, integration, and process teams to enable node readiness, volume ramp, and disciplined change control. Influence equipment supplier strategy by leading technical engagements, driving design improvements, and aligning roadmaps to business needs. Drive global standardization through development, validation, and deployment of Best Known Methods (BKMs) across sites. Lead initiatives to improve cost of ownership, reduce variation,

AIRecruitment
P
📍 San Francisco, California, United States· Full-time
✓ High-confidence listingCompany trend -100%
Quick readStrong listing-quality and freshness signals

Who Are We? Postman is the world’s leading API platform, used by more than 45 million+ developers and 500,000 organizations, including 98% of the Fortune 500. Postman is helping developers and professionals across the globe build the API-first world by simplifying each step of the API lifecycle and streamlining collaboration—enabling users to create better APIs, faster. The company is headquartered in San Francisco and has offices in Boston, New York, Austin, Tokyo, London, and Bangalore - where Postman was founded. Postman is privately held, with funding from Battery Ventures, BOND, Coatue, CRV, Insight Partners, and Nexus Venture Partners. Learn more at postman.com or connect with Postman on X via @getpostman. P.S: We highly recommend reading The "API-First World" graphic novel to understand the bigger picture and our vision at Postman. The Opportunity Postman is seeking an experienced AI Systems Reliability Engineer to help define, build, and maintain the infrastructure and processes that ensure the reliability, scalability, and performance of Postman’s AI-powered API and agentic systems in production. This role focuses on monitoring, availability, incident response, and automation to support AI services and tools trusted by millions of developers globally. What You’ll Do Develop and manage reliability metrics (SLOs) for AI-driven API services and agentic AI platform features Implement comprehensive observability and monitoring systems for real-time performance and fault detection Design and drive automated failover, recovery, and incident response strategies for high-availability AI infrastructure Optimize resource utilization, particularly GPU/accelerator efficiency, ensuring cost-effective AI system operation Collaborate closely with engineering, platform, and product teams to align reliability efforts with broader organizational goals Lead efforts to build internal tooling and automation focused on AI system stability and operational excellence Drive continuo

AIGoRustDevOps
🔔

Get new staff research scientist jobs in United States by email

Daily job updates · Unsubscribe anytime