Jobs in United States

Technical Support Specialist in New York

243 active opportunities · Updated October 2026

Explore current technical support specialist jobs in New York. Filter by work mode, employment type, experience, department, date posted and distance.

D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%

From $104K/yr

Quick readStrong listing-quality and freshness signals

The GTM Enablement & Business Value Excellence team is the center of excellence responsible for how go-to-market execution is designed, enabled, measured, and continuously improved across Sales Enablement, Technical Solutions Enablement, and Business Value. We establish the capabilities, operating model, systems, analytics, and governance that enable our customer-facing teams to execute consistently at scale and adapt as our business evolves. Within the team, Enablement Operations designs and manages the operational infrastructure that powers enablement at scale. This includes operational governance, enablement systems, knowledge management, reporting, operational processes, and continuous improvement, ensuring programs can be delivered consistently, measured effectively, and continuously optimized. The Opportunity: As an Enablement Operations Manager, you will own the operational infrastructure that enables our programs to run efficiently and consistently at scale, with a primary focus on supporting Technical Solutions Enablement. You will partner closely with the Technical Solutions Enablement team to manage operational processes, enablement systems, knowledge management, reporting, governance, and continuous improvement. You will also collaborate with the Enablement Systems & Analytics team and cross-functional stakeholders to evolve the operational capabilities that support scalable execution. This role is ideal for someone who enjoys building scalable systems and processes, improving how work gets done, and creating the operational foundation that enables high-performing teams. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You'll Do: Support the operational processes, governance, and enablement systems that support Technical Solutions Enablement. Man

AIGoRustExcel
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%

From $156K/yr

Quick readStrong listing-quality and freshness signals

The Team We are Datadog’s in-house product experts. The Technical Solutions team enables Datadog’s worldwide growth by educating potential partners and ensuring that our integration ecosystem is high-performing, secure, and valuable. Partner Technology Solutions Engineers (TSEs) are the technical bridge between Datadog and our third-party developer community. We act as consultants, helping partners build world-class monitoring solutions on the Integration Developer Platform (IDP) . The Opportunity Datadog is looking for a Partner Technology Solutions Engineer to join our fast-paced team. You will be the primary technical contact for our partners, guiding them through the entire integration lifecycle—from initial architectural design to final publication on the Datadog Marketplace. This is a unique role that combines deep technical troubleshooting with high-level consulting and platform advocacy. You will work directly with external developers and see your contributions immediately reflected in the Datadog ecosystem. You Will Act as the technical lead for partners, advising on OAuth flows, log pipelines, OpenTelemetry, and agent-based vs. API-based configurations Perform architectural assessments and deep-dive code reviews for partner integrations in the integrations-extras and marketplace repositories, ensuring they meet our Quality Rubric Solve complex technical challenges for partners via Zendesk, Slack, and dedicated technical consultations Identify friction points in our Integration Developer Platform (IDP) and partner with our internal Product and Engineering teams to build a better developer experience Maintain public-facing developer documentation and internal tracking systems ( JIRA ) to ensure transparency and scale You Are A technical expert with 3+ years of experience in a technical role (Support Engineering, Solutions Architecture, or Software Development) Proficient in at least one language (Python or Go preferred) An observability enthusiast who unders

PythonLinuxAIGo
P
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -72.3%

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. The FinOps function is responsible for financial accountability, visibility, and optimization across all engineering-related spend at Plaid. This includes cloud infrastructure, AI/ML and data workloads, third-party SaaS tools, and other technical investments that support Plaid’s products and internal platforms. The team operates at the intersection of Engineering, Product, and Finance, ensuring that spending decisions are transparent, intentional, and aligned with product strategy and business priorities. Rather than functioning as a cost-control or approval layer, FinOps enables teams to understand, own, and optimize their spend while maintaining engineering velocity. Responsibilities Monitors and analyzes engineering spend across cloud, AI/ML, data platforms, and SaaS, identifying trends, anomalies, and optimization opportunities. Builds and maintains forecasts for engineering spend, partnering with Finance and engineering leaders to understand drivers, assumptions, and risks. Partners with engineering, product, and TPMs to incorporate cost considerations into roadmaps, architectural decisions, and execution plans. Leads cost optimization initiatives, such as rightsizing, commitment strategies, an

SQLAWSAzureGCP
S
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -92.9%

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Observe by Snowflake is a high-growth SaaS observability platform built on the Snowflake AI Data Cloud, enabling businesses to troubleshoot modern distributed applications 10x faster. Now, as a core part of Snowflake, we’ve reached a major milestone in the evolution of the Snowflake platform. By bringing AI-powered observability directly into the Snowflake ecosystem, we’ve created the first truly unified platform for telemetry and business data. We’re looking for a Technical Account Manager to partner with our most strategic enterprise customers and ensure they derive sustained operational value from Observe. This is a hands-on, post-sales technical role focused on long-term platform adoption, optimization, and technical partnership. You will work directly with SRE, DevOps, platform, and engineering teams to embed Observe into daily workflows, evolve telemetry strategy over time, and continuously improve reliability, performance, and cost efficiency. This role is ideal for an experienced observability practitioner who enjoys being deeply embedded with customer teams, solving real production challenges, and acting as a trusted technical advisor in complex enterprise environments. What You’ll Do Serve as the primary technical owner and trusted advisor for assigned strategic a

AWSAzureGCPKubernetes
D
📍 New York, New York, United States
✓ Quality checkedCompany trend -85.2%

The Technical Solutions (TS) Enablement team empowers Datadog's global post-sales organization—including Support, Customer Success, and Sales Engineering—with the skills, tools, and resources they need to deliver world-class customer outcomes. As a Senior TS Enablement Program Manager supporting Sales Engineering , you'll own the design and execution of complex, high-visibility enablement programs that drive lasting change across the TS organization. You'll operate with significant autonomy, lead cross-functional initiatives, and serve as a senior voice within the Enablement team—partnering closely with leadership to translate business priorities into measurable enablement impact at scale. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Own the end-to-end strategy and execution for complex, global enablement programs—managing multiple concurrent workstreams across regions, functions, and timelines with limited guidance. Lead cross-functional stakeholder relationships with senior partners across StratOps, Product, TS Leadership, and the People Team—negotiating scope, aligning on priorities, and driving accountability without positional authority. Translate business goals and OKRs into multi-quarter program roadmaps; communicate priorities, trade-offs, and progress to leadership with clarity and a clear point of view. Design and implement comprehensive metrics frameworks to measure change adoption and program effectiveness; pull and analyze Salesforce and LMS data to track performance against KPIs; present findings to senior stakeholders and iterate for lasting impact. Build and deliver executive-ready materials that tell a compelling business story—translating complex data, program updates, and regional insights into structured narratives with clear

AISalesforce
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -85.2%

From $86K/yr

Quick readStrong listing-quality and freshness signals

Datadog’s Technical Solutions organization includes 1,200+ sales engineers, support engineers, post-sales experts, and solution architects. They run on an ecosystem of enterprise platforms and internal tools that directly shape how we serve customers. Technical Solutions Operations (“TSO”) owns that ecosystem. We manage the full lifecycle of the systems TS depends on: Zendesk, Jira, Confluence, and a growing portfolio of off-the-shelf and purpose-built tools. We do the work to operate, maintain, and evolve the platforms powering daily workflows across TS. When a vendor tool reaches its limits, we extend it through customization, integration, or targeted solution development, tapping internal partners across Datadog as needed. We’re looking for a Systems Engineer who wants to own enterprise platforms end-to-end. You go from understanding the business process, to designing the right solution (whether that’s configuration, integration, or code), to measuring whether it actually moved the needle. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Own enterprise systems through their full lifecycle. You’ll be the technical owner for one or more platforms that TS relies on daily. That means understanding how the system is used, where it’s falling short, what’s coming from the vendor roadmap, and what needs to change. You drive improvements from assessment through implementation. Engineer solutions that create leverage. Not every problem is solved by configuration. You’ll build and evolve enterprise systems, integrations, automations, and internal tools that multiply the effectiveness of 1,200+ technical experts. Where AI can make a solution smarter (e.g., intelligent routing, automated triage, agent-assisted workflows), you'll include AI in the initial design, no

JavaScriptTypeScriptPythonJava
M
📍 New York, new york, United States· Full-time
✓ High-confidence listingCompany trend -67.9%
Quick readStrong listing-quality and freshness signals

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We are looking for a strong technical lead to guide the engineers designing, building, and maintaining the novel, high-performance systems that make up our serverless platform. You'll lead the team responsible for Modal's machines layer: the fleet of bare metal and cloud hosts that every Function, Sandbox, and training job runs on, and the control plane that provisions, images, monitors, and repairs them. You'll own the full lifecycle of a machine, from accepting and benchmarking new hardware from a growing set of providers, to network bring-up, kernel and image management, GPU and disk health tracking, and automated remediation of unhealthy hosts. You'll manage a team of 3–8 engineers while staying hands-on across the stack which involves BMCs, firmware, PXE, bootloaders, Linux networking, drivers, and distributed control-plane services, and you'll shape our long-

B
📍 New York, New York, United States· Full-time
✓ High-confidence listing

$101.2K – $126.6K/yr

Quick readStrong listing-quality and freshness signals

Why join us Brex is the intelligent finance platform that enables companies to spend smarter and move faster in more than 200 markets. By combining global corporate cards and banking with intuitive spend management, bill pay, and travel software, Brex enables founders and finance teams to accelerate operations, gain real-time visibility, and control spend effortlessly. Brex’s AI-native automation and world-class service eliminate manual expense and accounting tasks for customers so they can focus on what matters most. Tens of thousands of the world's best companies run on Brex, including DoorDash, Coinbase, Robinhood, Zoom, Plaid, Reddit, and SeatGeek. Working at Brex allows you to push your limits, challenge the status quo, and collaborate with some of the brightest minds in the industry. We’re committed to building a diverse team and inclusive culture and believe your potential should only be limited by how big you can dream. We make this a reality by empowering you with the tools, resources, and support you need to grow your career. What you’ll do As a Technical Consultant, you'll take ownership of the integration implementation for our customers, guiding them from kickoff to go-live. You will be responsible for translating a customer’s business requirements into an effective product configuration, solving challenges, and providing best practices related to Brex Integrations. Ultimately, you'll ensure that customers are equipped with the necessary knowledge to feel confident with their integrations, setting them up for long-term success. Where you’ll work This role will be based in our New York office. We are a hybrid environment that combines the energy and connections of being in the office with the benefits and flexibility of working from home. We currently require a minimum of three coordinated days in the office per week, Monday, Wednesday and Thursday. As a perk, we also have up to four weeks per year of fully remote work! Responsibilities Own the su

AIGoExcelQuickbooks
M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We are looking for strong engineers with experience and interest in designing, building, and maintaining the novel, high-performance systems that make up our serverless platform. Requirements: 5+ years of experience writing high-quality production code Experience building high-performance distributed systems at a large scale (the more battle scars, the better) Strong cloud skills Strong knowledge of low-level operating system foundations (Linux kernel, file systems, containers, etc.) Experience with performance engineering (tell us a story of when you shaved off a few milliseconds!) Ability to work in-person in our NYC or SF office. Prior experience with Rust is nice to have, but not required. Ability to participate in on-call rotation and respond to production incidents.

LinuxRestAIGo
M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: At Modal, we sell cloud services atop which our customers run their critical production systems. As a rapidly growing new cloud infrastructure company, we seek to improve our reliability dramatically while scaling the size of our platform, customer base, and our team. This role is for people who are deep systems thinkers, love stacking nines, and thrive from making others move faster at scale. Responsibilities include: Identifying architectural changes to improve reliability and performance. Fostering a culture of reliability across Modal’s engineering organization. Defining and implementing operational processes such as deployments, upgrades, etc. Operating systems like Kubernetes, Postgres, Redis, etc. Participating in on-call rotations, and responding to production incidents. Requirements: 5+ years of experience writing high-quality production code. 2+ years of

RedisAWSKubernetesCI/CD
M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We are looking for strong engineers with experience in making ML systems performant at scale. If you are interested in contributing to open-source projects and Modal’s container runtime to push language and diffusion models towards higher throughput and lower latency, we’d love to hear from you! Requirements: 5+ years of experience writing high-quality, high-performance code. Experience working with torch, high-level ML frameworks, and inference engines (vLLM or TensorRT). Familiarity with Nvidia GPU architecture and CUDA. Experience with ML performance engineering (tell us a story about boosting GPU performance — debugging SM occupancy issues, rewriting an algorithm to be compute-bound, eliminating host overhead, etc). Nice-to-have: familiarity with low-level operating system foundations (Linux kernel, file systems, containers, etc).

LinuxRestAIGo
M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We're looking for strong backend engineers who love building a developer tools used by the largest AI companies in the world. You’ll be building for things at scale, but also for new AI workflows that change every day. Requirements: Experience building and shipping modern web applications end-to-end. We care more about what you’ve built than how many years you’ve been building. Comfort working across the stack: TypeScript on the frontend, Python services on the backend, and ClickHouse for data and analytics. Deep knowledge of observability tools and patterns used for large-scale workloads such as custom sandboxes, training and inference for large language (LLM) and diffusion models. Experience with at least one of: billing/payments systems, B2B SaaS tooling, or enterprise software, or LLM / diffusion models inference and training loads. Strong product instincts; yo

TypeScriptPythonAIGo
M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We’re looking for strong engineers with experience building developer tools that users love to work with. Our ideal candidate is someone with a demonstrated drive to build beautiful interfaces that enhance developer productivity. Requirements: 5+ years of experience developing high-quality Python libraries with broad user-bases, ideally including some experience maintaining open-source software. Knowledge of advanced Python features, especially async programming. A strong product sense that manifests as a focus on developer ergonomics and productivity. A high level of customer empathy, good communication skills, and an openness to working directly with our users to help solve their problems. Ability to participate in on-call rotation and respond to production incidents. Ability to work in-person in our NYC or Stockholm office. Any of the following would be a plus:

TypeScriptPythonAIGo
M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We're looking for a Growth Engineer to own the technical foundation of Modal's marketing and developer-facing web surfaces: the marketing site, docs site, growth landing pages, high-profile microsites, forms, analytics instrumentation, and the integrations that help users discover, understand, and get started with Modal. This is a frontend-heavy role for someone with strong product taste, web engineering craft, and a business-owner mindset. You'll partner with Product Engineering, Design, Data, and Growth to ship polished, measurable web experiences from high-profile projects like the GPU Glossary and LLM Engine Advisor to internal tooling that helps teams publish content faster. When this role is going well, Modal launches new pages, docs experiences, campaigns, and experiments quickly without sacrificing performance, craft, or measurement. In this role you will:

TypeScriptAIGoRust
M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: Most of the value of owning a model shows up at serving time. We're building a platform that covers the whole life of an LLM -- train it, deploy it, observe it -- and inference is where teams feel the difference every day. We already run elastic inference, sandboxes, distributed volumes, and multi-node training, and we control the infrastructure underneath, so the serving stack is ours to shape rather than something we resell. You will do hands-on inference research at Modal, working with the research lead to pick high-impact bets and owning them end to end. The bets that matter most are the ones that move cost per token and tail latency on the workloads our customers actually run. What you'll do: Own end-to-end inference research bets: speculative decoding, disaggregated prefill/decode, quantization (FP8, INT4), KV-cache and memory management, autoscaling for spik

🔔

Get new technical support specialist jobs in New York, United States by email

Daily job updates · Unsubscribe anytime