About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We're building a platform that covers the whole life of an LLM: training it, deploying it, and observing it in production. We already run multi-node training, elastic inference, sandboxes, and distributed volumes, and we control the infrastructure underneath. We’re looking for research depth in post-training to sit alongside our systems and product work. What you'll do: We are looking for research scientists with a strong track record in reinforcement learning, machine learning, and foundation models, including large language and multimodal models, to join our research team. This role is well suited to candidates interested in improving existing methods and developing new techniques for large-scale model training, optimization, and inference, extending models to long-context and long-horizon tasks, and improving inference-time efficiency, reliability, and robustnes
Jobs in United States
Staff Research Scientist in New York
81 active opportunities · Updated October 2026
Showing
15 jobs
Explore current staff research scientist jobs in New York. Filter by work mode, employment type, experience, department, date posted and distance.
From $276K/yr
Team description At Datadog, AI agents are becoming first-class consumers of observability, security, and software delivery data — from third-party coding agents like Claude Code, Cursor, and Copilot, to our own Bits SRE, Bits Assistant, and Bits Dev Agent. The Agentic Interfaces team owns the platform that connects these agents to Datadog: the MCP Server, the tools and retrieval surfaces agents call into, and — critically — the evaluation systems that tell us whether an agent's experience on Datadog data is actually getting better over time. This role is about that last piece. We're hiring a Staff Applied Scientist to define what "good" means for an Agentic interface at Datadog and to build the measurement systems that make it true. "Good" isn't one number — it spans answer quality, tool-selection accuracy, retrieval relevance, latency, token cost, and end-to-end agent success on real customer workflows. You'll design the evals, build the datasets, define the metrics, and partner with the AI engineers on the team to land the platform that lets every product group at Datadog ship integrations that are demonstrably better release over release. The space is full of open research questions. How do you evaluate an agent end-to-end when the trajectory is non-deterministic? How do you score tool selection when the tool catalog has hundreds of entries and grows weekly? How do you build a measurement system that catches regressions across first-party and third-party agents at once, without each team writing their own harness? If those are the problems you want to spend your time on, come build this with us. Datadog values people from all walks of life. We understand not everyone will meet all the above qualifications on day one. That's okay. If you’re passionate about technology and want to grow your skills, we encourage you to apply. What You’ll Do: Own the evaluation strategy for Datadog's AI agent integrations. Define the metrics — offline and online, quali
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Large Language Models (LLMs) continue to push the boundaries of what AI systems can do — but inference is still the bottleneck. The Model Efficiency team is responsible for pushing the limits of LLM inference efficiency across our foundation models. We explore and ship breakthroughs across the model execution stack, including: model architecture and MoE routing optimization decoding and inference-time algorithm improvements software/hardware co-design for GPU acceleration performance optimization without compromising model quality Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, expertise, and time zones to promote collaboration and flexibility. You'll find the Model Efficiency team concentrated in the EST and PST time zones, these are our preferred locations. As a Staff Research Engineer, you will develop, prototype, and deploy techniques that materially improve how fast and efficiently our models run in production. You may be a good fit
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: Most of the value of owning a model shows up at serving time. We're building a platform that covers the whole life of an LLM -- train it, deploy it, observe it -- and inference is where teams feel the difference every day. We already run elastic inference, sandboxes, distributed volumes, and multi-node training, and we control the infrastructure underneath, so the serving stack is ours to shape rather than something we resell. You will do hands-on inference research at Modal, working with the research lead to pick high-impact bets and owning them end to end. The bets that matter most are the ones that move cost per token and tail latency on the workloads our customers actually run. What you'll do: Own end-to-end inference research bets: speculative decoding, disaggregated prefill/decode, quantization (FP8, INT4), KV-cache and memory management, autoscaling for spik
From $170K/yr
About Taskrabbit: Taskrabbit is a marketplace platform that conveniently connects people with Taskers to handle everyday home to-do’s, such as furniture assembly, handyman work, moving help, and much more. At Taskrabbit, we want to transform lives one task at a time. As a company we celebrate innovation, inclusion and hard work. Our culture is collaborative, pragmatic, and fast-paced. We’re looking for talented, entrepreneurially minded and data-driven people who also have a passion for helping people do what they love. Together with IKEA, we’re creating more opportunities for people to earn a consistent, meaningful income on their own terms by building lasting relationships with clients in communities around the world. Taskrabbit is a hybrid company with employees distributed across the US and EU and a Built In — Best Places to Work (2022, 2023, 2024, 2025) continually ranked across multiple national and regional categories. Join us at Taskrabbit, where your work will be meaningful, your ideas valued, and your potential unleashed! This role is hybrid requiring 2 days in office at our San Francisco or NYC hub every Tuesday & Wednesday. About the Role Machine Learning is a cornerstone at Taskrabbit, and we're looking for a Staff Machine Learning Engineer to join our team and lead the next phase of our customer retention strategy. This is a critical, full-stack role for an individual who is passionate about the end-to-end lifecycle: from initial research and model development to building the robust systems that power repeat customer engagement and lifetime value growth at scale. Taskrabbit's greatest growth opportunity lies in deepening customer relationships and accelerating repeat purchases. Our most valuable customers are those who return frequently, discover new service categories, and increase their spending over time. There's significant untapped potential in the marketplace: repeat customers spend 3-5x more than one-time users, and category expan
From $170K/yr
About Taskrabbit: Taskrabbit is a marketplace platform that conveniently connects people with Taskers to handle everyday home to-do’s, such as furniture assembly, handyman work, moving help, and much more. At Taskrabbit, we want to transform lives one task at a time. As a company we celebrate innovation, inclusion and hard work. Our culture is collaborative, pragmatic, and fast-paced. We’re looking for talented, entrepreneurially minded and data-driven people who also have a passion for helping people do what they love. Together with IKEA, we’re creating more opportunities for people to earn a consistent, meaningful income on their own terms by building lasting relationships with clients in communities around the world. Taskrabbit is a hybrid company with employees distributed across the US and EU and a Built In — Best Places to Work (2022, 2023, 2024, 2025) continually ranked across multiple national and regional categories. Join us at Taskrabbit, where your work will be meaningful, your ideas valued, and your potential unleashed! This role is hybrid requiring 2 days in office at our San Francisco or NYC hub every Tuesday & Wednesday. About the Role Machine Learning is a cornerstone at Taskrabbit, and we're looking for a Staff Machine Learning Engineer to join our team and lead the next phase of our customer retention strategy. This is a critical, full-stack role for an individual who is passionate about the end-to-end lifecycle: from initial research and model development to building the robust systems that power repeat customer engagement and lifetime value growth at scale. Taskrabbit's greatest growth opportunity lies in deepening customer relationships and accelerating repeat purchases. Our most valuable customers are those who return frequently, discover new service categories, and increase their spending over time. There's significant untapped potential in the marketplace: repeat customers spend 3-5x more than one-time users, and category expan
From $220K/yr
About the role: Datadog is building a product-led Competitive Intelligence function to help inform product strategy, roadmap prioritization, positioning, and competitive readiness. As the Head of Competitive Intelligence you will build and lead the CompIntel function across priority competitors and adjacent market opportunities. You will translate competitor product evolution, technical capabilities, launches, pricing value, customer sentiment, analyst narratives, and ecosystem shifts into actionable insights for Product, GTM, Sales Engineering, Customer Success, Technical Advocacy, and leadership. In this highly technical, hands-on leadership role, you should understand how observability products are implemented, where workflows differ, what technical claims are credible, and how product gaps translate into customer impact. What You'll Do: Own the recurring competitive radar for Datadog’s priority competitor set, including taxonomy, source standards, evidence quality, and update cadence. Produce competitive deep dives, feature-gap analyses, early-warning briefs, and opportunity assessments tied to specific product domains and leadership questions. Translate CompIntel findings into roadmap recommendations, product differentiation hypotheses, and Product-facing decision guidance. Build and maintain a central CompIntel repository with competitor profiles, source notes, comparison frameworks, evidence logs, and decision-ready briefs. Convert validated findings into GTM-ready handoffs for battlecards, objection handling, strategic deal narratives, and field enablement. Synthesize signals from Product, Engineering, Sales Engineering, Customer Success, field teams, technical communities, Advocacy programs, win/loss, customer feedback, analyst research, public competitor materials, and CABs. Partner with Product Management to frame competitive questions, identify decision criteria, and connect findings to roadmap and prioritization discussions. Partner with Tec
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . Pinterest is uniquely positioned across the full consumer journey—from inspiration and discovery through consideration, planning, and action. Creators bring ideas to life, Pinners signal emerging interests and intent, and advertisers help people act on what inspires them. This role will define how Pinterest measures that flywheel, helping advertisers understand how Pinterest and its creators build awareness and drive business outcomes, while helping creators understand the value they generate and unlock new opportunities with brands. The ideal candidate is a Staff-level product leader with substantial ads measurement experience and a deep understanding of how advertisers set objectives, deploy tactics across the full funnel, and evaluate outcomes. They have worked cross-functionally with engineering, data science, research, PMM, sales, an
From $204K/yr
We’re looking for a Staff Product Designer to join the APM (Application Performance Monitoring) team. This team is focused on helping engineers agentically troubleshoot and optimize their applications. You'll learn the domain and proactively spot areas for improvement, getting stakeholder buy-in as needed. You'll own your design work end-to-end while helping shape broader product direction. You'll also hold a high bar for quality across the team and help other designers build support for their decisions. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Become a design leader within APM by proactively identifying opportunities and shaping the design strategy across multiple product areas and workflows. Partner closely with PMs and Engineers to ship intuitive experiences, balancing user needs with platform scalability and engineering constraints. Conduct and synthesize qualitative and quantitative research to identify pain points, validate solutions, and guide roadmap decisions. Communicate design rationale clearly and persuasively across design, engineering, product, and executive stakeholders. Mentor and support other designers by providing feedback, sharing frameworks and workflows, and helping raise the overall design quality bar across the organization. Prototype in code to rapidly explore ideas and validate concepts. Contribute to production code using AI-assisted workflows to ensure execution matches design intent. Help the team adopt AI-enabled design workflows. Who You Are: You have 10+ years of experience in digital product design. Your portfolio demonstrates a strong track record of designing and shipping complex technical products. You use AI-assisted workflows to prototype and rapidly iterate. You have a track record of connecting
From $232K/yr
Datadog is entering a new chapter in how our product looks, feels and operates. We're building a small, high-leverage Design Lab team to define that evolution, and we're looking for a Senior Staff Visual Designer to author it. Our design system powers everything we ship. Now we're ready to evolve it. As Datadog expands into AI-driven experiences and more complex product surfaces, the visual language needs to grow with it. This is the role that decides what that language is — someone with taste, judgment, and the confidence to set a direction and defend it. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do Define and evolve Datadog’s visual language across product surfaces. Lead exploratory and concept work, creating a range of distinct visual directions for the next generation of the product. Synthesize product context, research, references, and stakeholder input into clear visual principles and a coherent point of view. Establish the quality bar for typography, color, composition, iconography, imagery, motion, and visual expression across the platform. Create high-fidelity product exemplars and prototypes that make an emerging direction tangible. Partner deeply with Brand, Product Design, Motion, and Design Systems to test and refine the direction across different surfaces. Influence executives and senior stakeholders through clear, persuasive visual storytelling. Guide and critique the work of other designers, helping the visual direction remain coherent as it evolves. Help shape how visual exploration, critique, and craft are integrated into Datadog’s design process. You will report directly to the Design Director and work as part of a small, focused team defining the future state before it scales across hundreds of designers and engineers. Who Yo
About the team OpenAI’s Forward Deployed Engineering team partners with customers to turn research breakthroughs into production systems. We operate at the intersection of customer delivery and core platform development. About the role As a foundational FDE manager, you’ll lead FDE through high-stakes, ambiguous customer deployments and own technical and business value outcomes end to end. You’ll grow a team that can operate under pressure and help OpenAI learn from the field. You’ll partner closely with Product, Research, Sales, and GTM to ensure fieldwork informs roadmap priorities, drives new exploration, and supports safe deployment at scale. Your decisions will influence how OpenAI is trusted by the customers closest to our deployment work. Your success will be measured by how consistently your team ships, how clearly you deliver signal to Research and Product, and how durable your team and delivery model prove to be. This role is based in New York City We use a hybrid work model of 3 days in the office per week and offer relocation assistance to new employees. This role also will require travel up to 25%. In this role you will Lead and grow a team of FDE delivering production systems with frontier models Own end-to-end delivery outcomes through clarity, speed, tight coordination, and technical quality Codify what works into tools, playbooks, and roadmap inputs that create leverage for both OpenAI and our wider developer community Notice early indicators and raise them with urgency, whether in product behavior, customer environments, or delivery practices Use judgement to distinguish what requires action and what does not Set a high bar for FDE performance and support each person’s growth through direct, actionable feedback Define how we staff and support field teams that can scale without added complexity You might thrive in this role if you Bring 8+ years of engineering or technical delivery experience, including 2+ years managing high-performing FDE or custo
About Datalab Datalab trains models that read documents reliably at scale. The world's most important information is trapped in PDFs, scans, and files that can't easily be parsed, and getting it out correctly matters. From frontier AI labs processing training data to Fortune 500s like Siemens extracting decades of engineering records, Datalab is where businesses turn to when extraction has to be right. We're at an 8-figure run rate with a team of 7. Anthropic is a customer. And we have hundreds more across FAANG, frontier AI labs, healthcare, finance, government, and legal. Our tools, chandra, surya, marker, and lift, have 70,000+ GitHub stars and broad developer mindshare. We're backed by founding members of OpenAI, FAIR, and Hugging Face. Role Overview Datalab is growing extremely quickly, and as we grow, we outgrow our processes every few months. We're looking for a Founding Business Operations hire to find where the business is straining and fix it — designing, running, and improving the operational machinery underneath a fast-scaling company. This is a true generalist role, and the target will move as we scale. You'll own whatever the highest-leverage operational gap is at a given moment. Right now, that's our revenue and metrics layer: our pipeline and customer data need a real owner, inbound leads need to be captured and followed up reliably, and leadership needs consistent visibility into how the business and our launches are performing. That's where you'll start and where you'll have the most impact in your first few months. In two or three months the biggest gap may be somewhere else entirely — and you'll be the kind of person who's energized by that, not unsettled by it. You'll work directly with leadership and partner across sales, engineering, research, finance, and our Chief of Staff. If you're ambitious, analytical, and organized, this is a uniquely rewarding opportunity to help build an AI business from the ground up. Day to day: A typical week might
$325K – $500K/yr
CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. We're looking for a Staff Software Engineer to help build the next generation of CLEAR's enterprise identity platform, CLEAR1. We’re creating frictionless identity solutions for some of the most sophisticated companies in the world and pushing the boundaries everywhere we go. As a Staff Engineer, you’ll own high-impact, cross-functional products and integrations end to end from problem framing and architecture through implementation, rollout, adoption, and measurable outcomes. This is an ideal role for someone who thrives in scrappy, ambiguous environments, while bringing the technical rigor and operating discipline developed at larger companies. You’ll partner across teams, establish reusable patterns and standards, improve reliability and execution, influence technical direction, and mentor other engineers. A brief highlight of our tech stack: Python / Java / React / Typescript What you’ll do: Design, build, test, and deploy scalable applications that power CLEAR’s identity platform. Own projects end to end—from technical discovery and architecture through implementation, rollout, adoption, and operational support. Partner closely with Product, Design, Data, Security, and Operations to translate business problems into simple, scalable technical solutions. Design for reality by understanding failure modes, system dependencies, degradation strategies, recovery paths, and the points where systems may break under scale. Build quality and operability into the design from the start, including test strategy, observability, alerting, sa
Who We Are Addepar is a global data and AI platform empowering investment professionals to turn complex financial information into actionable intelligence. Addepar unifies portfolio, market and client data in a total portfolio view and delivers AI-powered insights within investment and client workflows. More than 1,400 firms in nearly 60 countries use Addepar to manage and advise on nearly $9 trillion in assets. Its open platform integrates with nearly 650 software, data and consulting partners to power end-to-end investment operations across firms of all sizes and complexity. Addepar supports clients worldwide with offices in New York City, Salt Lake City, London, Edinburgh, Pune, Dubai, Geneva, Singapore and São Paulo. The Role We are seeking a Staff Full Stack Software Engineer to join the Advisor Experience team as our Technical Lead. Our team is focused on building tools for financial advisors to grow and sustain their business. We oversee bespoke products for advisors and develop advisor-focused capabilities throughout the Addepar platform. In this role, you will be the primary technical anchor for new capabilities including Secure Message Center — a compliant messaging experience built into Addepar's client portal that allows advisors and their clients to communicate directly within the platform. You will partner directly with Engineering Leadership and Product Management to build a modern, scalable architecture from the ground up. Beyond system design, you will act as a true engineering multiplier: setting technical standards, mentoring junior and mid-level engineers, and working alongside other senior engineers and AI specialists to deliver high-impact advisor tools. Applicants must be legally authorized to work in the United States for any employer without requiring current or future visa sponsorship (for example, employment-based visas such as H-1B, F-1/OPT, or similar), and must be authorized to begin work in the U.S. on their first day of employme
$240K – $285K/yr
Why join us Brex is the intelligent finance platform that enables companies to spend smarter and move faster in more than 200 markets. By combining global corporate cards and banking with intuitive spend management, bill pay, and travel software, Brex enables founders and finance teams to accelerate operations, gain real-time visibility, and control spend effortlessly. Brex’s AI-native automation and world-class service eliminate manual expense and accounting tasks for customers so they can focus on what matters most. Tens of thousands of the world's best companies run on Brex, including DoorDash, Coinbase, Robinhood, Zoom, Plaid, Reddit, and SeatGeek. Working at Brex allows you to push your limits, challenge the status quo, and collaborate with some of the brightest minds in the industry. We’re committed to building a diverse team and inclusive culture and believe your potential should only be limited by how big you can dream. We make this a reality by empowering you with the tools, resources, and support you need to grow your career. Why join us Brex is the AI-powered spend platform. We help companies spend with confidence with integrated corporate cards, banking, and global payments, plus intuitive software for travel and expenses. Tens of thousands of companies from startups to enterprises — including DoorDash, Flexport, and Compass — use Brex to proactively control spend, reduce costs, and increase efficiency on a global scale. Working at Brex allows you to push your limits, challenge the status quo, and collaborate with some of the brightest minds in the industry. We're committed to building a diverse team and inclusive culture and believe your potential should only be limited by how big you can dream. We make this a reality by empowering you with the tools, resources, and support you need to grow your career. Engineering at Brex Engineering at Brex is about building systems that scale with speed and intention. Our teams span Software, Data, Security, and IT,
Other cities to consider
More places hiring for this role
Get new staff research scientist jobs in New York, United States by email
Daily job updates · Unsubscribe anytime