About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We're looking for strong backend engineers who love building a developer tools used by the largest AI companies in the world. You’ll be building for things at scale, but also for new AI workflows that change every day. Requirements: Experience building and shipping modern web applications end-to-end. We care more about what you’ve built than how many years you’ve been building. Comfort working across the stack: TypeScript on the frontend, Python services on the backend, and ClickHouse for data and analytics. Deep knowledge of observability tools and patterns used for large-scale workloads such as custom sandboxes, training and inference for large language (LLM) and diffusion models. Experience with at least one of: billing/payments systems, B2B SaaS tooling, or enterprise software, or LLM / diffusion models inference and training loads. Strong product instincts; yo
Jobs in United States
Stock Yard in New York
110 active opportunities · Updated October 2026
Showing
15 jobs
Explore current stock yard jobs in New York. Filter by work mode, employment type, experience, department, date posted and distance.
About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We're looking for Forward Deployed Engineers on our engineering team who want to work at the intersection of deep infrastructure work and direct customer impact. As an FDE, you'll partner with leading AI companies and foundation labs on cloud architecture, networking, storage, containerization, sandboxing, and more — helping them design and ship production infrastructure on Modal's platform. The FDE team today includes world-class software engineers, computational scientists, ML engineers, and former founders. We're looking for people with strong engineering fundamentals, deep curiosity across the infrastructure stack, and energy for working directly with customers on hard problems. You will: Work hands-on with companies like Suno, Lovable, Cognition, and Meta to architect and deploy massive-scale production workloads on Modal Lead technical discovery and architect
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We're looking for a Marketing Operations Manager who can own and harden the systems layer of Baseten's Marketing engine. Marketing at Baseten is scaling fast — more spend, more campaigns, more model launches, more inbound. The systems underneath (Our ESP, CRMs, forms, tracking, routing, alerting) need an owner who treats them like production infrastructure that cannot go down. When something breaks, it costs us time, pipeline, and trust in the data. This role exists so it doesn't break. You’ll simultaneously build for the future and re-think assumptions about our tech stack in the age of agents. This is an offensive play that gives the rest of the team leverage and superpowers to hit our ambitious goals. This is NOT an IT or service role. This is a core member of the marketing team who implements technology to achieve outcomes. RESPONSIBILTIES Own the marketing tech stack end-to-end: ad platforms, email systems, tracking, pixels, forms, connectors. Build defense-in-depth on inbound: spam/bot protection, rate limiting, email/domain validation, sync gating — and the alerting to catch anomalies before they hit sales or leadership dashboards. Enforce data integrity: UTM governance, campaign membership, lifecycle stages, lead scoring and routing logic, field-level hygiene, canonical metric definitions. Operationalize the web request pipeline with our dev agency: structured briefs, tickets, SLAs, and launch-day runb
🚀 About WRITER WRITER is where the world's leading enterprises orchestrate AI-powered work. Our vision is to expand human capacity through superintelligence. And we're proving it's possible – through powerful, trustworthy AI that unites IT and business teams together to unlock enterprise-wide transformation. With WRITER's end-to-end platform, hundreds of companies like Mars, Marriott, Uber, and Vanguard are building and deploying AI agents that are grounded in their company's data and fueled by WRITER's enterprise-grade LLMs. Valued at $1.9B and backed by industry-leading investors including Premji Invest, Radical Ventures, and ICONIQ Growth, WRITER is rapidly cementing its position as the leader in enterprise generative AI. Founded in 2020 with office hubs in San Francisco, New York City, Seattle, Austin, Chicago, and London, our team thinks big and moves fast, and we're looking for smart, hardworking builders and scalers to join us on our journey to create a better future of work with AI. 📐 About the role At WRITER, our mission to expand human capacity with superintelligence relies on a foundational truth: our platform must be available, performant, and reliable, 24/7. As an Infrastructure engineer, you'll be at the heart of making this a reality, impacting every enterprise customer who trusts us with their AI-powered workflows. This isn't just about keeping the lights on; it's about pushing the boundaries of what's possible, proactively identifying and solving complex systemic challenges, and laying the groundwork for our rapid growth and the evolving demands of enterprise generative AI. You'll build resilient systems, automate across the stack, and champion reliability best practices, directly enabling our ambitious product roadmap and ensuring our customers always have access to the powerful tools they need. This is a hybrid position, based out of our New York City, San Francisco, Seattle, or London hubs. You'll report to our director of engineering. 🦸🏻♀️
What we're building Mutiny is the self-improving AI infrastructure for GTM teams to execute faster and close more revenue. Our ambition is to do for revenue velocity what Cursor and Claude Code did for engineering velocity. With Mutiny, everyone in sales and marketing gets a bench of GTM athletes that handle any work across their revenue motion and learn from what's actually moved their deals. In April we re-launched the product as an agent-first platform. Anthropic showcased us as a leader in AI GTM. MRR is growing more than 70% month-over-month, with customers like Uber, Rippling, and Snowflake. We're backed by Sequoia, YC, and Insight, and we're building a generational company. The opportunity Most engineers spend their career making predictable systems faster. You'll spend yours making non-deterministic ones trustworthy. As a senior engineer on our AI product team, you'll architect the Campaign Builder and Agent experiences marketers and sellers open every day to go from idea to personalized assets in minutes. You'll partner directly with product, design, and the founders to define what an agent-first GTM platform should feel like, and your calls on architecture, evals, and guardrails compound across thousands of customer accounts. This role is in person in New York City, five days a week, and we ship weekly. What you'll own The core agent surfaces. Architect and ship the Campaign Builder and Agent experiences end-to-end. Frontend, backend, prompts, evals, the whole stack. Reliability on top of LLMs. Make non-deterministic models feel deterministic at the surface. Build the retries, fallbacks, and orchestration so the customer never sees the failure mode. Evals and guardrails. Define how we measure quality, catch regressions, and keep brand and tone consistent across thousands of customer accounts. Speed and feel. AI products live or die by latency and the loop between intent and output. You'll obsess over both, and use coding agents and agent networks to ship f
What we're building Mutiny is the self-improving AI infrastructure for GTM teams to execute faster and close more revenue. Our ambition is to do for revenue velocity what Cursor and Claude Code did for engineering velocity. With Mutiny, everyone in sales and marketing gets a bench of GTM athletes that handle any work across their revenue motion and learn from what's actually moved their deals. In April we re-launched the product as an agent-first platform. Anthropic showcased us as a leader in AI GTM. MRR is growing more than 70% month-over-month, with customers like Uber, Rippling, and Snowflake. We're backed by Sequoia, YC, and Insight, and we're building a generational company. The opportunity We're looking for a Head of GTM to bring Mutiny to millions of users. You'll sit shoulder-to-shoulder with the founders, own revenue end-to-end from first touch to retention, and help shape the decisions that define the company’s brand and culture. You’ll be building the AI infrastructure layer that reshapes how every GTM team operates. This role is in person in New York City, five days a week. What you'll own Revenue. PLG self-serve signups and conversion, sales-led motion including pipeline and ARR, retention and expansion. Marketing. Build the engine from zero to inevitable, including hiring the team and building the channels. GTM engineering. Build the agent-powered growth stack (including Mutiny!) that lets 5 people out-ship a team of 50. Sales and CX. You'll own pricing, evaluation, and expansion strategy to help us close deals and drive customer adoption. Depending on your skills and interest, these functions could fully report to you or you can own the strategy and ops portions. The team. Hire the best GTM people in the world and build a culture of winners. The story. You'll shape how the market understands category-defining AI products for GTM. Who you are A nose for distribution. You're a savant at breaking through noise. You know which channel will work before th
What we're building Mutiny is the self-improving AI infrastructure for GTM teams to execute faster and close more revenue. Our ambition is to do for revenue velocity what Cursor and Claude Code did for engineering velocity. With Mutiny, everyone in sales and marketing gets a bench of GTM athletes that handle any work across their revenue motion and learn from what's actually moved their deals. In April we re-launched the product as an agent-first platform. Anthropic showcased us as a leader in AI GTM. MRR is growing more than 70% month-over-month, with customers like Uber, Rippling, and Snowflake. We're backed by Sequoia, YC, and Insight, and we're building a generational company. The opportunity Most sales leaders inherit a playbook. You'll write one. As a founding first sales manager, you'll partner with the founders, own the number end-to-end from first conversation to expansion, and use AI and our own product to build a team of full-stack sellers that outsells teams 5x their size. What works for our team becomes what works for every sales team running on Mutiny. The calls you make on hiring, motion, and product feedback will shape the company well beyond the sales org. This role is in person in New York City, five days a week. What you'll own Revenue. Sales-led revenue, end-to-end. Pipeline, conversion, ARR, retention, and expansion. The number is yours. PLG to enterprise. Define qualification, build the enterprise pipeline, and turn self-serve signups at companies like Rippling, Uber, Snowflake, Zendesk, and Gusto into seven-figure deals. Outbound. Generate net new meetings with the hottest AI companies and enterprises in the market. Build the motion, run it yourself first, then teach it. The team. Hire, coach, and develop a team of Account Executives into top performers. The ceiling on revenue is the ceiling on the team you build. The playbook. Write the sales motion that becomes the default for every team running on Mutiny. What works here ships to thousands
What we're building Mutiny is the self-improving AI infrastructure for GTM teams to execute faster and close more revenue. Our ambition is to do for revenue velocity what Cursor and Claude Code did for engineering velocity. With Mutiny, everyone in sales and marketing gets a bench of GTM athletes that handle any work across their revenue motion and learn from what's actually moved their deals. In April we re-launched the product as an agent-first platform. Anthropic showcased us as a leader in AI GTM. MRR is growing more than 70% month-over-month, with customers like Uber, Rippling, and Snowflake. We're backed by Sequoia, YC, and Insight, and we're building a generational company. The Opportunity Most brands hire creators to make content. We're hiring one to define how Mutiny shows up in the world. As our social media creator, you'll own the brand voice across every channel, build the AI workflows that let you move at the speed of culture, and place the bets other B2B brands are too scared to take. You'll partner with the founders to turn a fast-growing AI company into one people actually want to follow. What You'll Own The numbers. Brand awareness, impressions, and sign-ups from the social channels you own. The creative bar. Be the person whose work everyone secretly screenshots. Turn raw ideas from founders, product, and customers into content with a point of view so strong it travels on its own. Our social presence. Own how Mutiny shows up across LinkedIn, X, YouTube, and TikTok. Grow what's working, launch what isn't there yet. The AI workflow. Build agent pipelines that turn one podcast into a week of content. Train voice models so drafts sound like us. Run evals so the stack gets sharper every week. Brand campaigns. Read the room, read the data, then make the call. Place the swings other brands are too scared to take, and know exactly why they worked when they do. Product marketing media. Content that helps people understand what we do and how our product work
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Our team is a fast-growing group of researchers and engineers focused on building reliable ML systems and pushing the boundaries of LLM inference efficiency. We develop techniques that improve how models execute in production, driving lower latency, higher throughput, and consistent quality across diverse workloads. As an engineer on this team, you’ll work across the inference stack to improve core performance metrics by diving deep into model execution, identifying bottlenecks, and developing innovative optimizations. You’ll collaborate closely with modeling and systems teams to experiment, measure, and ship improvements that meaningfully accelerate inference. As the team evolves, you’ll have opportunities to build expertise in advanced performance techniques, including GPU/CUDA optimizations, kernel-level improvements, and model execution strategies for MoE and large-scale architectures. Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, e
Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Large Language Models (LLMs) continue to push the boundaries of what AI systems can do — but inference is still the bottleneck. The Model Efficiency team is responsible for pushing the limits of LLM inference efficiency across our foundation models. We explore and ship breakthroughs across the model execution stack, including: model architecture and MoE routing optimization decoding and inference-time algorithm improvements software/hardware co-design for GPU acceleration performance optimization without compromising model quality Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, expertise, and time zones to promote collaboration and flexibility. You'll find the Model Efficiency team concentrated in the EST and PST time zones, these are our preferred locations. As a Staff Research Engineer, you will develop, prototype, and deploy techniques that materially improve how fast and efficiently our models run in production. You may be a good fit
From $10K/yr
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role We’re looking for someone to help lead the future of analytics at Ramp. This person will enable Ramp to get 1% better every day by developing data products and insights. They will partner closely with business stakeholders and product, engineering, and design counterparts to prioritize and execute on work, improve reporting, as well as drive results and process improvements. What You’ll Do Full stack development, building models to consume, transform, and expose data to stakeholders and production systems Drive a culture of experimental design, testing agenda, and best practices Contribute to the culture of Ramp’s data team by influencing processes, tools, and systems that will allow us to make better decisions in a scalable way Collaborate with P/E/D/D (product, engineering, data and design) teams to develop product roadmaps and measure success Work closely with data engineering teams to capture, move, store, and transform raw data into highly actionable insights, and partner with business teams to turn those insights into action Wha
From $10K/yr
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role We’re looking for a strategic and operationally rigorous Senior Manager, Head of Talent Operations to lead and scale the engine behind our recruiting organization. You will lead a team of 4 Talent Operations professionals and oversee the operational backbone supporting our ~40-person recruiting team, which is growing rapidly in scale and complexity across functions and globally. This role owns and strengthens core recruiting infrastructure — including interview scheduling and operations, headcount governance, tech stack strategy, interviewer enablement, compliance, and candidate experience — ensuring our systems are efficient, scalable, and built for continued growth. This is a highly cross-functional leadership role requiring strong systems thinking and execution rigor. What You'll Do Talent Operations Leadership Lead, coach, and develop a team of four Talent Operations professionals. Drive prioritization, operational rigor, and scalable infrastructure to support a 40+ recruiter organization. Be hands on where needed to continue to s
From $10K/yr
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About Production Engineering Production Engineering is Ramp's infrastructure ownership layer. We exist to make Ramp faster, more reliable, and more scalable — and we do that by being embedded in the problems, not adjacent to them. A few things that define how we operate: One team, one company, one objective. There is no "infra team" and "product team" — there is Ramp. We share the company's goals as our own. When a product team struggles with reliability or scalability, that is our struggle. If reliability or scalability is at risk, we own it. We don't wait to be invited, and we don't ask whose code it is. If a system is slow, if it breaks, if it won't scale — that's ours to lead, regardless of where it lives in the stack. We go first, and we go fast. When the path isn't obvious, we don't wait for someone else to find it. We move with urgency, propose the solution, align the stakeholders, and stay in until it's done — not until our ticket is closed. We lead the way. We find the next problem before it finds us. And when we solve it, we don't just fix
From $10K/yr
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role Ramp's Dev API team builds the programmatic surfaces that let developers and AI agents read from and write to Ramp. We are an AI-focused platform team making it trivially easy for any engineer at Ramp (or outside it) to ship new capabilities across all surfaces. Our ideal candidate thinks agent-first, has strong opinions on developer experience, and wants to define how software interacts with financial infrastructure at scale. Check out our Engineering Blog for more on our tech stack, mission, and values! What You'll Do Build and operate Ramp's multi-surface platform with a high bar for reliability, correctness, and developer experience across tens of thousands of businesses. Serve thousands of builders and partners using our API, driving a large fraction of Ramp’s revenue, and deliver on a massive business opportunity to embed Ramp everywhere. Build software factories -- autonomous tooling and agents that accelerate endpoint creation, enabling teams across Ramp to ship their own API surfaces. Lead design and execution of complex back
About Ramp Ramp is building the smart infrastructure for finance teams, embedded in the transaction flow of every dollar a business spends. We automate how over $200B in annualized spend flows in and out of 70,000+ companies: authorizing payments, flagging risk, categorizing spend, and closing books. The problems are high-stakes, data-dense, and unforgiving. We hire people with high agency and high urgency. We look for slope over intercept. We care less about where you trained and more about what you’ve built. At Ramp, everyone is a builder who owns problems end to end and makes consequential decisions that shape the outcome. The median Ramp customer saves 5% and grows revenue 16% in their first year – far in excess of businesses operating without Ramp. We believe every ambitious company deserves the same. If you want to build systems that directly shape how companies move and manage billions, Ramp is the place to do it. About the Role The Credit Engineering team builds the systems that enable Ramp to manage credit lifecycles for businesses end-to-end, from initial underwriting through ongoing portfolio strategy optimization. As an engineer on Credit, you’ll own systems that influence $100+ billion in annual payment volume through millions of decisions across our product stack in support of the most ambitious FinTech portfolio in the United States. We are looking for a backend engineer to own the technical roadmap for underwriting, limit management, portfolio optimization, and agentic workflows. You will build robust systems, automate complex operational tasks, and maintain the high-frequency controls that power all of Ramp’s products. If you are obsessed with correctness, excited by ambiguity, and passionate about problems at the intersection of distributed systems and financial strategy, this role is for you. What You'll Do Architect scalable, stateful systems for automated limit management to optimize Ramp’s charge card portfolio. Envision and build the next gene
Other cities to consider
More places hiring for this role
Get new stock yard jobs in New York, United States by email
Daily job updates · Unsubscribe anytime