Jobs in United States

Advance Backend Tech in United States

858 active opportunities · Updated October 2026

Explore current advance backend tech jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.

N
📍 Santa Clara, United States
✓ Quality checkedCompany trend -8%

NVIDIA is seeking a Senior Firmware Engineer to join our CSP Engagements team, focusing on system software for Datacenter products such as GB200. This role combines deep technical expertise in embedded firmware development with customer-facing responsibilities to enable cloud service providers with next-generation computing platforms. You will work at the intersection of hardware and software, driving technical solutions from concept through deployment. What you will be doing: Design and develop firmware solutions for manageability and observability of data center servers. Actively participate in hardware bring-up activities, OOB firmware development, protocol stacks (Redfish, PLDM, MCTP, NSM) and hardware-software co-design for Cloud Service Provider deployments. Debug and troubleshoot NVIDIA GPU firmware issues, power management, performance, and thermal control problems for data center deployments, providing active support to CSPs. Partner directly with CSPs to deliver technical solutions, co-develop & co-debug features and optimizations, and provide support during new product introductions. Perform advanced system debugging, root cause analysis, and performance optimization for large-scale data center environments. Collaborate with AE, FAE, and Solution Architect teams to deliver integrated customer solutions and technical documentation. What we need to see: Deep expertise in data center server architectures, HPC systems, and hardware-software co-design. Deep expertise in embedded firmware, server management controllers, and hardware bring-up with proven track record of shipping production BMC solutions Strong knowledge of DMTF protocols (Redfish, IPMI, PLDM, MCTP, SPDM), telemetry frameworks, and out-of-band management architectures Expert-level skills in C/C&

Artificial IntelligenceAI
V
📍 United States· Full-time· Remote
✓ Quality checkedCompany trend -88.6%

At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. We're looking for a highly skilled People Systems Administrator to drive the design, configuration, and optimization of Workday across Vanta, with deep functional ownership of the Payroll, Benefits, and/or Absence module(s). Reporting to the Sr. Manager, People Systems, you'll act as a trusted consultant and system expert, partnering with functional leaders across People, Payroll, Finance, IT, and Legal to identify opportunities, implement advanced solutions, and improve the employee experience. You'll shape the future of our Workday ecosystem by leading complex configurations, driving process improvement, and making sure the system evolves ahead of the business rather than behind it. What you’ll do as a People Systems Administrator, Workday at Vanta: Own the design and configuration of the Workday platform Lead the design and implementation of configurations across Workday modules (Core HCM, Payroll, Benefits, Absence, Time Tracking) including business processes and security groups along with condition rules and calculated fields. Serve as the primary technical expert for Workday enhancements, partnering with cross-functional teams to gather requirements and translate ambiguous business problems into scalable system solutions. Build the administration frameworks and standards the team runs on: security role design, tenant and environment management, change control, testing protocols, and documentation. Design reusable, scalable solutions rather than one-off fixes, so the same problem doesn't return in a different shape Own change management for system changes. Plan and deliver stakeholder communications, training, and enableme

AIFinancePayroll
D
📍 New York, New York, United States· Full-time
✓ High-confidence listingCompany trend -84.7%

From $244K/yr

Quick readStrong listing-quality and freshness signals

As a Forward Deployed Engineer on the Feature Flags team, you'll partner directly with customers to accelerate their feature flag implementations — from initial architecture consulting through prototype builds to full-scale migrations. This role is for someone who wants to write code with customers, not just advise them. You'll work hands-on inside customer codebases to unblock complex, high-stakes deployments, directly influencing deal velocity and customer success. Working closely with Sales, Solutions, and Engineering, you'll be the technical force that turns a signed contract into a live, adopted implementation. What You'll Do: Serve as the hands-on technical partner for strategic customers implementing Datadog Feature Flags, from pre-sales technical validation through post-sales delivery Consult on flag architecture and implementation approach for complex environments — multi-service, multi-platform, high-scale deployments Build prototype flag implementations directly in customer codebases to prove value and de-risk technical decisions early in the sales cycle Implement flags across diverse and advanced deployment modes (server-side, client-side, edge, mobile, streaming/real-time) tailored to each customer's stack Drive full flag migrations to completion — including legacy system cutover — efficiently and with minimal customer engineering burden Identify patterns across customer implementations and feed them back to Product and Engineering to improve the core product and reduce future implementation time Collaborate closely with Engineering on technical edge cases, product gaps, and implementation tooling Partner with Sales and Solutions to accelerate deal cycles by removing technical risk and uncertainty Who You Are: 5 years of professional software engineering experience, with hands-on coding ability across the stack you're deployed into Experience with feature flagging, experimentation, or config management systems (internal or vendor) Comfortable dropping i

AIRustSEMHR
PE
📍 United States· Full-time· Remote
✓ High-confidence listing
Quick readStrong listing-quality and freshness signals

Since 2003, Entrata has evolved from a visionary, student-led startup into a global leader in AI-driven property management technology. Today, we power the industry's most essential operating system, serving owners and residents worldwide through a comprehensive suite of intelligent leasing, payment, and communication tools powered by cutting-edge AI. With a proven track record of sustained growth and a global team of more than 2,200 employees, we offer the rare combination of established stability and high-velocity innovation. Recognized by the Silicon Slopes Hall of Fame and the Utah Business Fast 50, Entrata fosters a culture of radical transparency and entrepreneurial energy. At Entrata, we create an environment where different perspectives are valued and respected. Those perspectives challenge assumptions, strengthen our decisions, and raise the bar as we reshape the global living experience through AI-powered solutions. The Senior Technical Support Engineer provides advanced technical assistance to customers utilizing a suite of property management solutions. This role involves troubleshooting, diagnosing, and resolving complex software-related issues while maintaining a high level of professionalism and customer service. The Senior Technical Support Engineer applies substantial knowledge and expertise to handle a wide range of tasks, contributes to process improvements, and mentors junior team members. Schedule hours: Monday-Friday from either 7:00 am-4:00 pm MDT, 8:00 am-5:00 pm MDT, 9:00-6:00 pm (Utah MT)

M
📍 New York, new york, United States· Full-time
✓ Quality checkedCompany trend -67.9%

About Us: AI needs a new infrastructure layer. We're building it at Modal. Every era of computing brought new workloads that previous infrastructure couldn't support: mainframes, databases, and the cloud. Each time, the company that rebuilt the layer underneath defined the decade. AI is no different, except it touches everything instead of one slice, and the window to build the layer underneath it is open right now. Our customers include category-defining companies like Lovable , Ramp , Cognition, DoorDash, and Suno. They rely on Modal for instant GPU access, sub-second container starts, and native storage, so it's simple to serve low-latency inference, fine-tune models, and access production-ready sandboxes at scale. We recently raised a $355M Series C at a $4.65B valuation, led by General Catalyst and Redpoint Ventures. We've crossed $300M+ ARR and grown fivefold since September. Our team includes creators of popular open-source projects (e.g., Seaborn , Luig i ), academic researchers, international olympiad medalists, and experienced engineering and product leaders with decades of experience. The Role: We’re looking for strong engineers with experience building developer tools that users love to work with. Our ideal candidate is someone with a demonstrated drive to build beautiful interfaces that enhance developer productivity. Requirements: 5+ years of experience developing high-quality Python libraries with broad user-bases, ideally including some experience maintaining open-source software. Knowledge of advanced Python features, especially async programming. A strong product sense that manifests as a focus on developer ergonomics and productivity. A high level of customer empathy, good communication skills, and an openness to working directly with our users to help solve their problems. Ability to participate in on-call rotation and respond to production incidents. Ability to work in-person in our NYC or Stockholm office. Any of the following would be a plus:

TypeScriptPythonAIGo
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE We’re seeking a GPU Kernel Engineer to join our team at the cutting edge of AI acceleration, where your code directly impacts the performance of state-of-the-art machine learning models. As a GPU Kernel Engineer, you'll craft the foundation that powers modern AI workloads, optimizing every microsecond of computation to enable breakthrough applications. You'll work in a fast-paced, intellectually stimulating environment where technical excellence is paramount and your contributions directly influence production systems serving millions of users across numerous products. This role offers exceptional growth potential for engineers passionate about low-level optimization and high-impact systems work. EXAMPLE INITIATIVES You'll get to work on these types of projects as part of our Model Performance team: Baseten Embeddings Inference: The fastest embeddings solution available The Baseten Inference Stack Driving model performance optimization RESPONSIBILITIES Core Engineering Responsibilities Design and implement high-performance GPU kernels for key ML operations, including matrix multiplications, attention mechanisms, and mixture-of-experts routing Write and optimize code using CUDA, PTX assembly, and architecture-specific techniques Apply advanced performance optimization methods such as memory coalescing, warp-level programming, tensor core acceleration, and compute/memory overlap Performance & Innovation Impl

AWSMachine LearningAIC++
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE: Baseten’s Model Performance (MP) team is responsible for ensuring the models running on our platform are fast, reliable, and cost‑efficient. As part of this team, you’ll focus on Model APIs — the infrastructure powering our hosted API endpoints for the latest open‑source models. This work spans distributed systems, model serving, and developer experience. You’ll join a small, high‑impact team operating at the intersection of product, model performance, and infra, helping to define how developers interact with AI models at scale. RESPONSIBILITIES: Design, build, and operate the Model APIs surface with focus on advanced inference capabilities: structured outputs (JSON mode, grammar-constrained generation), tool/function calling and multi-modal serving Profile and optimize TensorRT-LLM kernels, analyze CUDA kernel performance, implement custom CUDA operators, tune memory allocation patterns for maximum throughput and optimize communication patterns across multi-GPU setups Productionize performance improvements across runtimes with deep understanding of their internals: speculative decoding implementations, guided generation for structured outputs, custom scheduling and routing algorithms for high-performance serving Build comprehensive benchmarking frameworks that measure real-world performance across different model architectures, batch sizes, sequence lengths, and hardware configurations Productionize performa

KubernetesMachine LearningAIGo
B
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.1%

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE The largest, most demanding enterprises are starting to run on Baseten, and they arrive with a range of security, compliance, and procurement requirements. As a Senior Engineer on Baseten's enterprise engineering team, you'll build the capabilities that enable large organizations like Writer, HubSpot, and Notion to succeed on Baseten. Enterprise engineering authors the core building blocks, APIs, and user experiences powering the Baseten platform: identity and access management, billing, regional isolation, and self-hosted and single-tenant deployment options. This is deep product and systems work across the full stack, from designing authentication and authorization systems using standards like OAuth and OIDC to shipping the admin experiences enterprise IT teams use to manage their organization. EXAMPLE INITIATIVES Recent and upcoming work on the team: Fine-grained authorization for users, service accounts, and agentic workloads SSO and SCIM support, allowing customers to centralize and automate access to Baseten Expanding the billing platform to support evolving pricing models, advanced data exports, and controls to manage spend In-product management and enforcement of customer compliance requirements like data residency and HIPAA Securing network paths in and out of a customer's models with private connectivity and ingress and egress restrictions Allowing customers to run Baseten inside their own VPC, on-pr

KubernetesRestMachine LearningAI
C
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Our team is a fast-growing group of researchers and engineers focused on building reliable ML systems and pushing the boundaries of LLM inference efficiency. We develop techniques that improve how models execute in production, driving lower latency, higher throughput, and consistent quality across diverse workloads. As an engineer on this team, you’ll work across the inference stack to improve core performance metrics by diving deep into model execution, identifying bottlenecks, and developing innovative optimizations. You’ll collaborate closely with modeling and systems teams to experiment, measure, and ship improvements that meaningfully accelerate inference. As the team evolves, you’ll have opportunities to build expertise in advanced performance techniques, including GPU/CUDA optimizations, kernel-level improvements, and model execution strategies for MoE and large-scale architectures. Please Note: We have offices in Toronto, Montreal, San Francisco, New York, Paris, Seoul and London. We embrace a remote-friendly environment, and as part of this approach, we strategically distribute teams based on interests, e

PythonGitRestAI
C
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Are you energized by building high-performance, scalable and reliable machine learning systems? Do you want to help define and build the next generation of AI platforms powering advanced NLP applications? We are looking for Members of Technical Staff to join the Model Serving team at Cohere. The team is responsible for developing, deploying, and operating the AI platform delivering Cohere's large language models through easy to use API endpoints. In this role, you will work closely with many teams to deploy optimized NLP models to production in low latency, high throughput, and high availability environments. You will also get the opportunity to interface with customers and create customized deployments to meet their specific needs. You may be a good fit if you have: 5+ years of engineering experience running production infrastructure at a large scale Experience designing large, highly available distributed systems with Kubernetes, and GPU workloads on those clusters Experience with Kubernetes dev and production coding and support Experience with GCP, Azure, AWS, OCI, multi-cloud on-prem / hybrid serving Experienc

AWSAzureGCPKubernetes
C
📍 New York, New York, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? We're building the foundational infrastructure that will define how the world thinks about and deploys AI, and we want the sharpest, most curious people to help us do it. As a member of our Analytics & Data Insights team, you'll tackle the kind of problems that don't have textbook answers yet, launch products that didn't exist a year ago, and help enterprises understand what foundational AI actually means for their bottom line. As a Data Engineer, you will: Work directly on new customer experiences built on one of the most advanced AI systems in the world Collaborate daily with researchers and engineers who are some of the best in the world at what they do Run implementations end-to-end and see initiatives through to real outcomes Partner across research, marketing, sales, and finance to help define how Cohere grows, with your recommendations feeding directly into products and strategy You may be a good fit if you have: 5+ years of experience working on production-grade data processing systems Strong command of Python and SQL Experience with distributed data processing frameworks such as Apache Beam, Spark, or

PythonJavaSQLKubernetes
C
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -79.2%

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Why this role? Are you energized by leading the design of high-performance, scalable and reliable machine learning systems? Do you want to set technical direction and help shape the next generation of AI platforms powering advanced NLP applications? We are looking for a Lead Member of Technical Staff to join the Model Serving team at Cohere. The team is responsible for developing, deploying, and operating the AI platform delivering Cohere's large language models through easy to use API endpoints. In this role, you will provide technical leadership across multiple teams, driving the architecture and strategy for deploying optimized NLP models to production in low latency, high throughput, and high availability environments. You will serve as a key point of contact for customers, leading the design of customized deployments to meet their specific needs, and mentoring engineers to raise the technical bar across the team. You may be a good fit if you have: 8+ years of engineering experience running production infrastructure at a large scale, with a track record of technical leadership Demonstrated experience leading the architecture

AWSAzureGCPKubernetes
P
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -72.3%

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Plaid’s Product team builds the network that powers the future of financial services. Our mission is to unlock financial freedom for everyone through open finance. Product Managers at Plaid are curious, customer-obsessed, and move quickly to deliver value. They take ownership, sweat the details, and make sound decisions with imperfect information. As a Product Manager on the AI Foundations team, you will drive Plaid’s AI strategy by building the data and intelligence layer that powers smarter financial experiences. You will work across engineering, data science, and research to develop scalable AI systems—from core embeddings and representation learning to applied model integrations that enhance developer and consumer outcomes. This role is for an experienced PM who thrives at the intersection of AI and platform products. You are technically fluent, strategic, and execution-oriented. You enjoy turning advanced machine-learning capabilities into reliable, trusted infrastructure that scales across Plaid’s ecosystem. Responsibilities AI Platform Vision: Define the strategy, roadmap, and success metrics for Plaid’s core AI and data foundation, enabling smarter, more adaptive financial products across th

P
📍 San Francisco, California, United States· Full-time
✓ Quality checkedCompany trend -72.3%

We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. We are looking for a Senior Financial Analyst to join the Corporate Finance team. This is the strategic, externally-oriented arm of the team — covering the board process, long range planning, capital structure, SBC and equity modeling, and enterprise-level financial strategy. You will partner closely with the CFO, Leadership, the Board, Corp Dev, and Accounting on the decisions that shape Plaid’s long-term financial trajectory. What excites us: 5+ years of work experience including financial planning & analysis, corporate finance, investment banking, private equity, strategic finance, or management consulting — ideally with tech, SaaS, or fintech exposure. Advanced financial modeling skills, with hands-on experience building long range, scenario/sensitivity, or valuation models. Strong command of consolidated financial statements, key SaaS / fintech operating metrics, and external benchmarking. Excellent communication and storytelling skills — able to move fluidly between granular model detail and an executive-level narrative for senior leaders and the Board. Proven ability to partner cross-functionally and distill complex, ambiguous topics into structured frameworks and clear recommendations. P

F
📍 Georgia, California, United States· Full-time
✓ High-confidence listingCompany trend -85.5%

From $120K/yr

Quick readStrong listing-quality and freshness signals

About Flexport: At Flexport, we believe global trade can move the human race forward. That’s why it’s our mission to make global commerce so easy there will be more of it. We’re shaping the future of a $10T industry with solutions powered by innovative technology and exceptional people. Today, companies of all sizes—from emerging brands to Fortune 500s—use Flexport technology to move more than $19B of merchandise across 112 countries a year. The recent global supply chain crisis has put Flexport center stage as we continue to play a pivotal role in how goods move around the world. We are proud to have the support of the best investors in the game who believe in our mission, solutions and people. Ready to tackle global challenges that impact business, society, and the environment? Come join us. Exciting Trade Advisory Opportunity to Help Make Global Trade Easy The opportunity: The global supply chain is one of the most heavily regulated industries for all parties involved. Everyone from the carriers, intermediaries, trucking companies, 3rd party logistics, freight forwarders, shippers, consignees, importers, exporters, and brokers face a challenging and complex task of navigating through a myriad of laws from the originating country, destination country as well as international treaties and conventions. The Trade Advisory team, consisting of business consultants, lawyers and licensed customs brokers, offers advanced customs and trade expertise to help clients identify supply chain risks and cost savings opportunities. Flexport is looking for a Trade Advisory Manager to help lead our Trade Advisory services. In this role, you’ll be responsible for managing a team of associates, growing our business and helping our clients navigate the complex web of global trade alongside some of the smartest people in the advisory, customs and logistics industries as we collectively challenge the status quo and reduce friction in global trade. You will: Help l

AWSGitAgileAI
🔔

Get new advance backend tech jobs in United States by email

Daily job updates · Unsubscribe anytime