Jobiba hiring network

Back End Td Reliability Lab Manager Jobs

1,684 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current back end td reliability lab manager jobs. Use filters to narrow by work mode, employment type, experience and date posted.

E
1mo ago

About the role As a Deployment Strategist, you'll work as part of a driven and creative team of Forward Deployed Engineers, Go to Market professionals, Product Engineers and other Strategists to deploy ElevenAgents, our enterprise AI platform, against the challenging problems our customers face. Your mission is to synthesize disconnected streams of thought into a cohesive understanding of what the most important problem is, what the existing workflows are, what the product needs, what users are motivated by, and where the impact could be. No two days are the same, but as a Deployment Strategist you can expect to: Meet with strategic customers to deeply understand their AI transformation goals and locate their biggest pain points, particularly around customer-facing and operational workflows powered by AI agents. Own flagship deals end-to-end — from identifying the right use case to structuring commercial terms (including outcome-based pricing, license fees, and implementation arrangements) and driving them to close. Embed yourself deeply inside strategic customers: act as a trusted partner with a provisioned account, spend meaningful time onsite, and own delivery outcomes as if you were part of their team. Identify high-impact use cases through close engagement with customer problems and workflows, and collaborate with Forward Deployed Software Engineers to bring them to life. Guide customers on best practices for deploying our products (e.g. ElevenAgents); including agent design, agent orchestration, production readiness to maximize adoption and impact. Scope out potential applications in new industries and expand our AI solutions across different sectors globally. Present the results of our work and proposals for future work to audiences ranging from technical teams to C-suite executives. Collaborate with our Research teams to feed field insights back into ElevenLabs' platform and models, helping shape the roadmap. Build and deliver compelling demos of ElevenAgent

pythonaigo
View job →
E
1mo ago

About the role As a Deployment Strategist, you'll work as part of a driven and creative team of Forward Deployed Engineers, Go to Market professionals, Product Engineers and other Strategists to deploy ElevenAgents, our enterprise AI platform, against the challenging problems our customers face. Your mission is to synthesize disconnected streams of thought into a cohesive understanding of what the most important problem is, what the existing workflows are, what the product needs, what users are motivated by, and where the impact could be. No two days are the same, but as a Deployment Strategist you can expect to: Meet with strategic customers to deeply understand their AI transformation goals and locate their biggest pain points, particularly around customer-facing and operational workflows powered by AI agents. Own flagship deals end-to-end — from identifying the right use case to structuring commercial terms (including outcome-based pricing, license fees, and implementation arrangements) and driving them to close. Embed yourself deeply inside strategic customers: act as a trusted partner with a provisioned account, spend meaningful time onsite, and own delivery outcomes as if you were part of their team. Identify high-impact use cases through close engagement with customer problems and workflows, and collaborate with Forward Deployed Software Engineers to bring them to life. Guide customers on best practices for deploying our products (e.g. ElevenAgents); including agent design, agent orchestration, production readiness to maximize adoption and impact. Scope out potential applications in new industries and expand our AI solutions across different sectors globally. Present the results of our work and proposals for future work to audiences ranging from technical teams to C-suite executives. Collaborate with our Research teams to feed field insights back into ElevenLabs' platform and models, helping shape the roadmap. Build and deliver compelling demos of ElevenAgent

pythonaigo
View job →

Education Degree, Post graduate in Computer Science or related field (or equivalent industry experience) Experience Extensive knowledge of Wealth Management, architecture, modules and hands on experience of front to back implementation of Wealth Management Suite Minimum 8 years of experience in Wealth Management Must have good experience of handling Trade Life Cycle Must have experience in implementing End to End solution in Wealth Management Must have good experience in API integration. Technical Skills Should have development experience in SQL-languages with understanding of SQL-databases Excellent understanding of the Asset Management and Investment Banking Experience in Business Analysis and System Implementation with extensive experience in Trading platforms Working knowledge of industry related protocols including FIX, EMSX and TSOX Experience in Wealth Management and Trading systems including Advent, SAXO, SWIFT, Bloomberg, Reuters General understanding of event driven systems Functional Skills Experience in Banking, Financial and Fintech experience in an enterprise environment preferred Experience in following best Coding, Security, Unit testing and Documentation standards and practices Experience in Agile methodology Effectively research and benchmark technology against other best in class technologies Focus on quality of the deliverables to ensure long term stability of the system Soft Skills Able to influence multiple teams on technical considerations, increasing their productivity and effectiveness, by sharing deep knowledge and experience Self-motivator and self-starter, Ability to own and drive things without supervision and works collaboratively with the teams across the organization Have excellent soft skills and interpersonal skills to interact and present the ideas to Senior and Executive management

sqlagileai
View job →
R
Replit
📍 Foster City• Full-time
1mo ago

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the role: Replit is building the security GRC function that will scale with an AI-native product. As the Risk & Compliance lead, you'll own our certification and audit program end to end: SOC 2, ISO 27001, and eventually ISO 42001 (AI management systems), while also owning the company's master security risk register and continuous compliance monitoring. You'll report to the Head of Security GRC, who retains overall accountability for the risk program, and work closely with Engineering to make sure controls hold up in practice, not just on paper. What You'll Do Own the end-to-end certification roadmap (SOC 2 Type II, ISO 27001, and future frameworks like ISO 42001) including scoping, gap assessments, remediation, and audit execution Manage relationships with external auditors and drive the annual audit calendar so certifications renew without last-minute scrambles Own and maintain the company's master security risk register including risk identification, scoring methodology, treatment plans, and residual risk reporting Build and maintain continuous compliance monitoring so control status reflects real-time state rather than point-in-time snapshots Own the core audit artifacts that back every certification including ISMS documentation, Statements of Applicability, risk assessments, and potentially FedRAMP System Security Plans (SSPs) Run regular audits and readiness assessments, and track remediation of findings and control gaps to closure Support GDPR and broader privacy compliance alongside the Legal/Privacy team, without owning the legal interpretation of requirements Partner with the GRC Engineer to define what evidence collection and control monitoring should be automated versus manually reviewed Track and

C
Cvshealth
📍 Patterson• From $22/hr
10 days ago

We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Join #TeamCVS and play an important part in delivering the kind of service that keeps our customers coming back to CVS Pharmacy. You'll work with other dedicated individuals in a fast-paced warehouse environment that inspires you to grow your skills and advance your career. Union Facility – Shift assignments based on seniority Starting Pay: $21.50/hour Available Shifts: 2nd Shift: 5:00 PM – 1:30 AM Note: Start and end times may vary depending on operational volume. Flexibility is important, as shift times are subject to change. As an Order Selector (Picker), you will: Fill store orders from a tablet or paper, select the needed pieces of merchandise and place them into a tote or box. Cross train in other warehouse responsibilities Report to the Production Supervisor Maintains a safe and clean work environment by following safety guidelines. Required Qualifications Must be at least 18 years of age Able to lift 30 lbs regularly and up to 50 lbs occasionally as a team lift Able to work overtime Able to work in a broad range of temperatures Preferred Qualifications Previous warehouse experience High School Diploma or Equivalent Benefits Paid breaks

warehouse
View job →
TI
15 days ago

About THG Ingenuity THG Ingenuity is a fully integrated digital commerce ecosystem, designed to power brands without limits. Our global end-to-end tech platform is comprised of three products: THG Commerce, THG Studios, THG Fulfilment. Each represents a single, unified solution, overcoming challenges and taking brands direct-to-consumer. Our client portfolio includes globally recognised brands such as Coca-Cola, Nestle, Elemis, Homebase, and Proctor & Gamble. VAS Manager Google Account — THG Ingenuity Reports to: Head of Department | Location: Omega, Warrington Role Purpose The VAS (Value Added Services) Manager is a dedicated role established to own the day-to-day management of Value Added Services for the Google fulfillment account. This role exists to close that gap: someone strong enough on the THG side to interpret Google's SOPs, and to trial, test, train and plan the resource required to execute against them, while owning QA in line with Google's requirements. Key Responsibilities Own end-to-end delivery of Value Added Services (VAS) for the Google account, including kitting, packing, quarantine handling and returns/shipback processes. Interpret Google's Standard Operating Procedures (SOPs) and translate them into workable warehouse processes at the Omega site. Trial and test new or amended VAS processes ahead of go-live, identifying and resolving operational bottlenecks before they affect the client. Design and deliver training for warehouse teams so VAS processes are executed consistently and to Google's quality standard. Plan and manage the resourcing required to deliver VAS workload, including forecasting headcount needs against volume. Own quality assurance (QA) for VAS activity in line with Google's requirements, including disposition and shipback of quarantined/blocked stock (e.g. Ship Back to Contract Manufacturer processes). Coo

TI
THG Ingenuity
📍 Northwich, United Kingdom• Full-time
15 days ago

About THG Ingenuity THG Ingenuity is a fully integrated digital commerce ecosystem, designed to power brands without limits. Our global end-to-end tech platform is comprised of three products: THG Commerce, THG Studios, THG Fulfilment. Each represents a single, unified solution, overcoming challenges and taking brands direct-to-consumer. Our client portfolio includes globally recognised brands such as Coca-Cola, Nestle, Elemis, Homebase, and Proctor & Gamble. What’s the role? Job Title: CX Apprentice - Operations Reports to: Senior CX Operations Manager Apprenticeship: Customer Service Specialist Level 3 Duration: 12-18 months, subject to learner progress + End Point Assessment Role Purpose To support the day-to-day operation of the Customer Experience function as an escalation point for more complex customer queries, using data and insight to resolve issues and help improve service performance across the team. Key Responsibilities Act as an escalation point for complex or ongoing customer queries that front-line agents cannot resolve. Support monitoring of CX performance metrics (contact rate, SLA adherence, quality scores) and flag emerging issues. Investigate the root cause of recurring customer issues and feed insight back to the wider team and leadership. Coordinate with other departments (Logistics, Compliance, BI) to resolve escalations and improve end-to-end service. Maintain and update knowledge base content and process documentation to reflect changes in policy or procedure. Support quality monitoring and coaching conversations, helping raise standards across the team. Contribute to small service-improvement projects aimed at reducing contact rate or improving customer satisfaction. What We Need From You Strong communication and problem-solving skills Resilience and a calm approach to handling complex or escalated queries A genuine interest in customer service and continuous improvement GCSEs (or equiva

gitrestai
View job →
DU
15 days ago

About the Team The Storage organization builds and operates the online stateful systems and abstractions that DoorDash Engineering depends on: reliable, efficient, secure, and easy to use. Within Storage, the Distributed Caching team owns every caching offering at DoorDash end to end, including ElastiCache (Redis/Valkey), Boulder (our KVRocks-based key-value store for high-QPS feature serving), Entity Cache (read Bill Shen’s engineering blog post, “ High-Performance Proxy Cache for DoorDash Services ”), and the Distributed Lock Service, plus the smart clients (asgard-redis, valkey-go) that sit in front of them. These systems back critical product surfaces across DoorDash, Wolt, and Deliveroo: the team runs roughly 400 ElastiCache clusters serving hundreds of millions of GET requests per second in aggregate, and Boulder, our offline-to-online feature store, serves billions of feature lookups per second at peak. About the Role The team owns provisioning of clusters and the smart clients that sit in front of them, baking in sensible defaults so that other engineering teams get a turnkey caching solution instead of having to run their own. You'll help drive Boulder's evolution to scale further, improve cost efficiency, enhance performance, and support real-time updates; re-platform the Distributed Lock Service onto a strongly consistent backend; and build the self-serve tooling and recommendation engine that let customers describe a workload (QPS, TTL, payload size, latency profile) and get the right backend without talking to a human. You'll go deep on cache invalidation, replication, sharding, compaction, and failover, while shipping the guardrails, automation, and observability that keep this scale operable by a small team. You must be located in San Francisco, Seattle, or the New York Metro Area for this hybrid position. You will report to the Engineering Manager on the Distributed Caching team within the Storage organization. You’re excited about this opportunity b

javaredisaws
View job →

We're looking for a Principal Engineer to join our CSP Engagements team as the technical focal point for end-to-end performance, working directly with engineering teams of key CSP/hyperscale customers to ensure they achieve various performance targets on NVIDIA platforms. In this role, you will augment NVIDIA's performance and benchmark teams with a dedicated CSP-facing focus. You will drive work streams with CSP engineering teams to build shared understanding of platform performance characteristics, gather and incorporate their workload-specific feedback into NVIDIA's optimization priorities, and validate that performance targets are met in customer-representative configurations. Your cross-CSP visibility enables you to identify patterns and drive systemic improvements in documentation, configuration guidance, and tooling. What you'll be doing: Drive performance characterization work streams with engineering teams of key CSP/hyperscale customers — ensuring they understand platform performance expectations, profiling methodology, and tuning options for their specific workloads Gather and synthesize CSP performance feedback — identify gaps between expected and actual throughput, and champion optimization priorities back into NVIDIA's CUDA, NCCL, driver, and firmware teams Ensure key open-source performance and stress tools (e.g., STREAM, GPU Burn, GPU BLAST) are updated and validated for the latest NVIDIA rack-scale systems, GPU architectures, and CPU platforms — so customers and internal teams have reliable baseline measurements from day one Work closely with CSPs to ensure their own performance and validation tooling reflects the latest GPU capabilities, memory hierarchy changes, and platform-specific tuning parameters Conduct cross-CSP performance comparison and pattern analysis — identify configuration, software, or workload differences that explai

pythonartificial intelligenceai
View job →
H
1mo ago

Become a part of our caring community Most AI engineering jobs are a thin wrapper around a model API. This role is different. We build the platform that transforms millions of clinical documents into trusted, actionable data. Our systems use large language models (LLMs) to read medical records, extract structured facts, answer complex questions with citations back to the source document, and route ambiguous cases to human experts for review. Our users make decisions that impact real healthcare outcomes, so “good enough” is not good enough. Building AI systems that are accurate, reliable, auditable, and scalable is at the core of this role. As a Senior AI Applied Engineer, you will design, build, deploy, and operate production AI systems used at scale within one of the largest health insurers in the United States. You will own solutions end-to-end, from user experience and APIs to model orchestration, evaluation frameworks, infrastructure, and production operations. Why Join Us Build production AI systems where LLMs are in the critical path, not just demos or proofs of concept. Work on extraction, retrieval, agentic workflows, and human-review systems that process real healthcare data at scale. Own projects end-to-end across frontend, backend, AI orchestration, infrastructure, deployment, and operations. Solve challenging problems around accuracy, explainability, traceability, and reliability in regulated environments. Ship quickly in a small, high-impact team that embraces AI-assisted development and rigorous quality standards. Build systems that continuously improve through expert feedback, evaluations, and human-in-the-loop workflows. Key Responsibilities Design, develop, and deploy full-stack AI-powered application

javascripttypescriptpython
View job →
C
Coder
📍 United States• Full-time
1mo ago

We are looking for a Senior Forward Deployed Engineer to join the Customer Solutions team. You will be the technical authority embedded with our most complex customers, guiding them through deployment, architecture, onboarding, and the adoption of agentic development workflows. You bring deep hands-on experience from prior roles and use that depth to advise, design repeatable patterns, and drive customer outcomes end to end. You operate autonomously, own the technical success of your customers, and bring their experience back to shape how Coder builds and delivers. This is not an execution-only role. You are equally comfortable doing deep technical work with a customer and stepping back to design the repeatable pattern behind it. You are energized by ambiguity, motivated by customer outcomes, and capable of influencing organizational change alongside the technical work. This position is required to sit in the Eastern Time Zone. What You'll Do Serve as the primary technical authority for post-sales customers, guiding deployment architecture, environment design, and adoption of Coder across both human and AI development workflows Own onboarding engagements end to end, ensuring customers move from contract to productive adoption with speed and confidence Lead Get Well engagements where architecture decisions, rollout patterns, or organizational dynamics are limiting customer health or growth Help customers implement the technical and organizational changes required to adopt agentic development practices at scale Design and document repeatable delivery patterns across onboarding, architecture, and adoption that can scale across customer segments Design and recommend reference architectures tailored to each customer's cloud environment, security posture, and organizational constraints Translate customer environment complexity into clear guidance on networking, ingress, identity, and infrastructure patterns Anticipate technical and operational risks, escalate to the right

awsazurekubernetes
View job →
M
1mo ago

ABOUT THE TEAM At Mural, we're changing how teams collaborate, think, and make decisions together. The Growth & Engagement team owns the experiences that turn a single-player tool into a multiplayer habit: the loops that bring people back, help them find their work, and pull their colleagues into a mural. Mural's self-serve growth runs through these experiences, and engagement and retention are central to how the business grows. YOUR MISSION As Senior Product Manager, Growth & Engagement, you will own the end-to-end self-serve experience, from a user's first touch through activation, return usage, and conversion to paid. You will treat the full funnel as one system: acquisition, activation, engagement, retention, sharing, and monetization. You will run a high-tempo experimentation practice, partner deeply across Product, Engineering, Design, Data, Marketing, and Sales, and make the calls that turn occasional users into teams that rely on Mural every week. This is a role for someone who wants to own real growth outcomes, not just ship features. You will own the surfaces that decide whether Mural's self-serve engine grows. Engagement and retention here move the business directly, and you will own the full loop across growth and engagement rather than a narrow slice of it. You’ll have a rare end-to-end view of the funnel, from acquisition through in-product engagement. You will work alongside thoughtful product, design, and engineering partners who care about craft and impact. WHAT YOU'LL DO Own product strategy, roadmap, and execution for the self-serve growth and engagement experience, from first touch through activation and return usage Build and optimize the invite, sharing, and join experience Drive activation and retention by improving the moments that bring people back: home and dashboard, notifications, search, rooms and folders, comments, and more Partner closely with Marketing to connect acquisition and re-engagement with in-product activation across

Who are we? Cohere is the leading security-first enterprise AI company. We build cutting-edge foundation AI models and end-to-end products that are designed to solve real-world business problems. We’re training and deploying frontier models for enterprises who are building AI systems. We believe that our work is instrumental to the widespread adoption of AI and we are looking for folks that want to be part of that. We obsess over what we build. Each one of us is responsible for contributing to increasing the capabilities of our models and the value they drive for our customers. Cohere is a team of researchers, engineers, designers, and more, who are all passionate about their craft. We are a global technology company headquartered in Toronto with key offices in London, New York City, San Francisco, Montreal, Paris, Berlin and Seoul. Join us! Role Overview: Building AI agents that can assist with any kind of enterprise work is a challenging, open-ended problem. One key piece of solving it is replicating real work environments as realistically as possible - filling them with hard tasks to solve, creating plausible input data, and defining clear rewards for completing the work the right way. We build many of these reinforcement learning (RL) environments, then drop our agents into them to evaluate or train them. In this role, you are responsible for creating these RL environments, running AI agents inside them, and improving both the agents and the environments in the process. The results reach customers, whose feedback feeds back in - and the agent/environment improvement loop continues. Key Responsibilities: There are many open problems in this space. As a Member of Technical Staff, RL Environments, you will: Build new RL environments targeting different agentic capabilities and industry areas Train and evaluate agents in those environments Make all the pieces work together: tasks, data, tool implementations, and verifiers Work across modeling and product to identify

PE
Private Employer
📍 New York• Full-time• Hybrid
1mo ago

A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role Forward Deployed AI Engineers work directly with customers owning Gen AI strategy and implementation. On a daily basis, you will build end-to-end workflows, take them to production, and solve real world problems at the largest scale. You will have ample opportunity to contribute learnings from the field back to the Palantir AIP product suite. You will be on the forefront of extending Palantir's existing footprint and strategy into new markets and problem spaces opened up by Gen AI. Core Responsibilities Forward Deployed AI Engineers’ responsibilities look similar to those of a hands-on AI startup CTO: you’ll work in small teams to own delivery of high stakes projects with clients. A day’s work may include building LLM workflows on a large scale, interacting with customers to understand their needs and set their AI strategy, but the most impact will be driven by implementing solutions into the real world of our partner's organizations. Do you aspire to be an entrepreneur or an Applied AI leader? We believe Palantir is the best place — with the best colleagues — to learn how!

PE
Private Employer
📍 United Kingdom• Full-time• Hybrid
1mo ago

A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role Forward Deployed AI Engineers work directly with customers owning Gen AI strategy and implementation. On a daily basis, you will build end-to-end workflows, take them to production, and solve real world problems at the largest scale. You will have ample opportunity to contribute learnings from the field back to the Palantir AIP product suite. You will be on the forefront of extending Palantir's existing footprint and strategy into new markets and problem spaces opened up by Gen AI. Core Responsibilities Forward Deployed AI Engineers’ responsibilities look similar to those of a hands-on AI startup CTO: you’ll work in small teams to own delivery of high stakes projects with clients. A day’s work may include building LLM workflows on a large scale, interacting with customers to understand their needs and set their AI strategy, but the most impact will be driven by implementing solutions into the real world of our partner's organisations. Do you aspire to be an entrepreneur or an Applied AI leader? We believe Palantir is the best place — with the best colleagues — to learn how!

🔔

Get new back end td reliability lab manager jobs by email

Daily job updates · Unsubscribe anytime