Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100. Founded in 2017, the company is headquartered in London, with offices and teams across Europe and the US. As AI continues to shape the way we live and work, Synthesia develops products to enhance visual communication and enterprise skill development, helping people work better and stay at the center of successful organizations. Following our recent Series E funding round, where we raised $200 million, our valuation stands at $4 billion. Our total funding exceeds $530 million from premier investors including Accel, NVentures (Nvidia's VC arm), Kleiner Perkins, GV, and Evantic Capital, alongside the founders and operators of Stripe, Datadog, Miro, and Webflow. About the role You will work on core enterprise platform systems, focusing on services that ensure Synthesia is secure, reliable, and scalable for our largest customers. You will contribute to our new suite of APIs that will unlock our market leading Avatar technology to be utilised by external creative tools. We are an AI native company and use the most powerful assistant tools on a daily basis to increase our speed of development, automate repetitive tasks and widen the scope of what the team can worked on. This includes Claude and Cursor. You will have ownership of projects that span months and multiple teams, requiring you to break down complex, ambiguous problems into clear steps that can be delivered and validated iteratively. Engineers within Synthesia are empowered to contribute heavily to product discussions and build with both user and business considerations in mind. Impact is key to everything we do here. You will work closely with product, security, legal, and infrastructure partners, and will be expected to translate business and regulatory requirements into scalable technical solutions. You will evaluate your work through system health and reliability metrics, leveraging observability and monitori
Jobiba hiring network
Reliability Engineer Jobs
2,028 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current reliability engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
GitLab is the intelligent orchestration platform for DevSecOps. GitLab enables organizations to increase developer productivity, improve operational efficiency, reduce security and compliance risk, and accelerate digital transformation. More than 50 million registered users and more than 50% of the Fortune 100* trust GitLab to ship better, more secure software faster. The same principles built into our products are reflected in how our team works: we embrace AI as a core productivity multiplier, with all team members expected to incorporate AI into their daily workflows to drive efficiency, innovation, and impact. GitLab is where careers accelerate, innovation flourishes, and every voice is valued. Our high-performance culture is driven by our values and continuous knowledge exchange, enabling our team members to reach their full potential while collaborating with industry leaders to solve complex problems. Co-create the future with us as we build technology that transforms how the world develops software. * Fortune 500® is a registered trademark of Fortune Media IP Limited, used under license. Claim based on GitLab data. Fortune 100 refers to the top 20% ranked companies in the 2025 Fortune 500 list, published in June 2025. Fortune and Fortune Media IP Limited are not affiliated with, and do not endorse products or services of GitLab. About the role As a Senior Backend Engineer, you will design, implement, and evolve product capabilities while solving high-scope backend problems and influencing the technical and product direction of our teams. You will move beyond "assigned work" to actively improve the quality, reliability, and performance of our systems. You will work across product, frontend, infrastructure, data, and security boundaries, making sound architectural trade-offs, communicating complex ideas clearly in an asynchronous environment, and helping define the standards for a high-scale, global product. Why you’ll love this role High Impact: You aren'
Who We Are Notion is the collaborative AI workspace where teams and agents think together . We're building one place where your knowledge, projects, meetings, and AI tools live side by side, so work is faster, clearer, and less fragmented. Millions of individuals, small teams, and large companies run their work on Notion. Notinos (our employees) are customer zero in bringing this future of work to life. We care about craft, building things that last, and the belief that great work is still fundamentally human. Our goal isn’t to ship the next feature. Each and every team of Notinos is working to set the standard for how humans work together in the AI era. From building a business’s system of record to making and managing AI agents to automating away the busy work, we care deeply about giving our customers more time for their life’s work. About the Role: The Web Infrastructure team builds the foundations behind Notion’s web clients, including client architecture, performance, reliability, and shared design systems. As Engineering Manager, you’ll lead the team through the evolution from Notion Clients. You’ll set strategy, develop senior engineers and managers, and partner across Notion to make the product faster, more reliable, and easier to build. We work from our offices on Mondays, Tuesdays and Thursdays (our Anchor Days) because we do our best thinking and building together in person. We’re looking for someone who’s excited to work alongside the team during those days. What You'll Achieve: You'll build and manage a diverse and inclusive team of engineers and managers working on core parts of Notion's architecture. You'll create a healthy environment in your team that embodies Notion's values. You'll recruit, coach, and develop engineers. You'll ensure engineers are regularly receiving feedback and making progress on personal and professional goals. You'll facilitate planning—the prioritization, sequencing, and staffing of work—for your team. You'll be responsible
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. As a Sr Design Engineer, you will work on design, simulation, and validation of next ‑ generation High ‑ Bandwidth Memory (HBM) architectures and circuit blocks. HBM requires advanced DRAM design knowledge combined with deep understanding of 3D stacked architecture, TSV signaling, wide I/O interfaces, PHY timing, power integrity, and system co ‑ optimization with GPUs/accelerators. This role sits at the intersection of DRAM design and high ‑ performance computing, enabling future AI/ML, HPC, and advanced graphics products. In this position, you will collaborate with Micron’s various design and verification teams all over the world and support the efforts of groups such as Product Engineering, Test, Probe, Process Integration, Assembly and Marketing to proactively design products that optimize all manufacturing functions and assure the best cost, quality, reliability, time-to-market, and customer satisfaction.
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. We are seeking a Principal Analog Design Engineer to lead the design, integration, and delivery of advanced analog and mixed-signal IPs for High Bandwidth Memory (HBM) products. This role is critical in developing and integrating high-performance analog subsystems within the HBM logic die, enabling industry-leading bandwidth, power efficiency, and reliability. As a principal engineer, you will provide deep technical leadership across analog design, IP integration, system alignment, and silicon execution, driving end-to-end success of HBM solutions. Responsibilities will include, but are not limited to: Lead the d esign and own critical HBM analog circuits, including: High ‑ speed transmitters and receivers Clock generation and distribution (PLLs, DLLs, CDRs) SerDes ‑ related analog blocks Biasing, reference, and calibration circuits
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Micron’s DRAM Design Engineering Group (DDEG) is where innovation meets excellence. We are advancing memory and storage technologies through collaborative engineering and creative problem-solving. Our team works at the forefront of semiconductor design, developing solutions that shape the future of memory products in a fast-paced, learning-focused environment. As a Design Verification Engineer, you will help develop next-generation memory technologies by verifying and optimizing digital and analog circuit designs. In this role, you will work closely with global multi-functional teams across the product lifecycle to deliver high-quality, manufacturable memory solutions that meet performance, reliability, cost, and customer requirements. Your work will directly contribute to bringing advanced memory products from concept to production. Responsibilities: Verify circuit functionality, reliability, power, and compliance with product specifications Drive verification planning, coverage closure, circuit debug, and design improvements Perform circuit modeling and simulation using industry-standard tools Support silicon validation, reticle experiments, and tape-out activities Partner with engineering, manufacturing, and product teams to deliver manufacturable designs Minimum Qualifications: Bachelor’s degree in Electrical Engineering or a related field 4+ years of semiconductor design, verifi
Job Details: Job Description: We are looking for a PERC Runset Development Engineer to develop and Validate PERC (Programmable Electrical Rule Checker) rule decks. The role involves working closely with PDK Team and foundry to implement PERC and ensure design compliance. Key Responsibilities: • Rule Development: Develop, validate, and maintain complex PERC rule decks to identify electrical reliability violations (e.g., ESD paths, current density, voltage stress, and floating gates). • Multi-Tool Integration: Develop and optimize verification flows across multiple EDA environments, ensuring consistency in results between Calibre PERC( preferred), Cadence Pegasus, and Synopsys tools. • Cross-Functional Collaboration: Work closely with Device Physicists, Process Integration Engineers, and Analog Designers to translate reliability requirements into programmable rules. • Debug and Analysis: Provide expert-level support to design teams in debugging PERC violations, distinguishing between true reliability risks and tool-induced false positives. • Automation: Develop scripts (Python, SKILL, Tcl, or Perl) to automate the execution, reporting, and tracking of reliability checks across different design versions. • Documentation: Maintain detailed documentation of reliability check-sets and create guidelines for designers to implement correct-by-construction layouts. Qualifications: Required Skills: • Experience with PERC rule development (preferably using Calibre PERC). • Knowledge of ESD, latch-up and reliability verification. • Familiarity with SVRF, TVF, Python or Tcl scripting. • Strong debugging and problem-solving skills. • Knowledge of Linux and scripting (Python/Shell/Tcl) is desirable. • Experience with advanced technology nodes is an advantage. • Team player and be able to work with globa
Abbott is a global healthcare leader that helps people live more fully at all stages of life. Our portfolio of life-changing technologies spans the spectrum of healthcare, with leading businesses and products in diagnostics, medical devices, nutritionals and branded generic medicines. Our 115,000 colleagues serve people in more than 160 countries. JOB DESCRIPTION: Grow your career while continuing Exact Sciences’ inspiring work. Changing roles within the company allows you to develop your skills while changing the future of cancer in a new way. Schedule - Monday - Friday 8 am -4:30/5 PST Position Overview This role is responsible for diagnosing and resolving failures in enterprise laboratory instruments and automation systems, conducting root cause investigations, and escalating service issues to minimize downtime. It includes performing routine preventive maintenance and calibrations to ensure equipment reliability and compliance with service standards. The position requires accurate documentation aligned with Good Documentation Practices (GDP) and regulatory requirements such as OSHA, FDA, ISO, and CLIA. Success in this role involves developing technical expertise, training others, and collaborating across teams and vendors to support troubleshooting and project execution. Flexibility, commitment to quality, and a focus on process improvement—including SOP development and workflow optimization—are essential. Key Accountabilities: Include, but are not limited to, the following: Troubleshooting & Repair: Under general supervision, diagnose, repair, and resolve failures following established procedures on instrumentation and automation systems within the laboratory by applying technical expertise to restore functionality. Escalate unresolved or complex issues to senior s
ABOUT THE TEAM The Data Modeling team builds and maintains the core data models and metrics that power decision-making across Mural. We are part of the Data Organization and focus on creating shared, reusable data models that represent key product and business concepts and are used across the company. Our work supports internal analytics, customer insight reports embedded in the product, and AI/ML model training. We partner closely with Product, Engineering, Data Platform, Business Analytics, Data Science, and Analytics Engineering to ensure the company is working from consistent definitions, high data quality, and reliable data availability. We are a small, high-leverage team focused on building durable data foundations rather than one-off solutions. YOUR MISSION You will own the delivery and evolution of Mural’s core data models and shared metrics, with a strong focus on data quality, reliability, and availability. This is a hands-on leadership role. You will not build stakeholder-specific data marts or ad-hoc analyses. Instead, you will focus on building foundational, reusable data models and metric definitions that support many use cases across the company. Your success will be measured by how widely trusted, consistently available, and broadly reused the data models and metrics you own are across teams such as Business Analytics, in-product insights, and ML. WHAT YOU'LL DO Own and evolve core data models and metrics: Define and maintain shared models for product usage, customers, accounts, and key business metrics that support analytics, in-product customer insights, and AI/ML model training Build and operate foundational data products: Stay hands-on building models using SQL, Python, and Spark in a modern lakehouse environment (e.g., Databricks), with strong attention to data quality, availability, performance, and cost Define shared semantics: Design and maintain shared metric definitions and semantic layers so data is interpreted consistently across teams an
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. As a Senior Software Engineer, Data on the Mapping team, you will collaborate with our world-class team of engineers, product managers, and scientists to grow and improve the quality of recommended routes and accuracy of our travel time estimations. You will lead the architecture and long-term technical direction of our offline experimentation tooling and route simulation services — the systems that let Lyft test routing changes safely before they reach production. You'll also build scalable data pipelines for experimentation, analytics, and machine learning models, along with the data governance and observability systems that keep them trustworthy. Your work will enable integration with partner teams and allow stakeholders across Engineering, Data Science, and Product to make data-informed decisions that directly impact Lyft’s growth and profitability. Our technology stack is based on the latest technologies such as AWS, Databricks, Kubernetes and Airflow. You will work with incredibly passionate and talented colleagues from software engineering, machine learning and data science on projects that directly impact millions of riders and drivers. Responsibilities Own core data pipelines end-to-end, building deep subject matter expertise in the systems you manage and defining/managing SLAs for pipelines, services, and datasets to ensure reliability at scale Serve as the technical owner and architectural lead for our offline experimentation platform and route simulation services, setting technical direction, evaluating trade-offs, and ensuring the systems scale with Lyft's routing and mapping ambitions Continuously evolve data models and schemas to meet business and engineering requirements Develop AI tools that support self-service management of data pipelines (ETL) and schema evolution, and perform han
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. POSITION SUMMARY CVS Health is seeking a Senior Mainframe Capacity & Performance Engineer to join our Enterprise Infrastructure organization. The Senior Mainframe Capacity & Performance Engineer will serve as a critical technical leader responsible for ensuring the performance, scalability, reliability, and efficiency of our enterprise mainframe environment supporting mission-critical healthcare, pharmacy, and retail applications. As a Senior Mainframe Capacity & Performance Engineer, you will play a key role in capacity planning, workload analysis, performance engineering, and infrastructure optimization across one of the nation's largest and most complex mainframe ecosystems. This position is responsible for proactively monitoring shared mainframe resources, evaluating system utilization trends, identifying performance risks, and providing actionable recommendations to improve overall system health and operational efficiency. The Senior Mainframe Capacity & Performance Engineer will partner closely with Application Development, Mainframe Systems Programming, Infrastructure Engineering, Architecture, Database Administration, Operations, and Business teams to analyze workload behavior, assess resource consumption, identify top consumers, and optimize application performance. This role requires deep expertise in z/OS performance analysis, capacity forecasting, workload managem
Job Details: Job Description: This position is in the Intel Mask Operations within the Logic Technology Development team working in one of the most advanced semiconductor process technologies in the world. In this position the engineer will be an integral contributor to the ongoing production of photolithography masks for Intel's leading-edge silicon manufacturing solutions. In this position, you will work on-site, in a dynamic and collaborative environment solving complex and challenging technical problems on sophisticated manufacturing processes and equipment. As IMO module Engineer the responsibilities may include but not limited to: Process equipment installation and development. Equipment maintenance, management of troubleshooting activities, regular monitoring of process performance, defect analysis and reduction. Drives improvements on quality, reliability, cost, yield, process stability/capability, productivity, and safety/ergonomics. Working with cross functional teams to solve technical process and defect issues. Plans and conducts experiments to fully characterize the process throughout the development cycle. Establishes control systems to optimize and sustain production performance. Develops strategies to resolve difficult problems and establishes systems to manage these problems in the future. An ideal candidate should exhibit the below behavioral traits: Experience with rapid analysis of complex process issues and identification of a solution path Willing to work Independently with minimal direction, to be self-directing and show initiative Communication skills and demonstrated ability to summarize complex d
Job Details: Job Description: Intel is looking for highly motivated individuals with strong technical background and capabilities to sustain, ramp, and transfer all technology nodes in Arizona. They will drive rapid continuous improvements in safety, quality, yield, reliability, cost, process stability/capability, and productivity while maintaining rigorous quality control. The role of a Process Integration and Yield Engineer is to deliver high-quality analytical insights that influences / directs the factory resources to minimize defects, control process parameters, and improve overall factory yields. All improvements and sustaining work is done in close collaboration and partnership with the process to improve the capability of the tools and processes. Responsibilities may include, but are not limited to: Performing detailed data analysis using various statistical/data mining tools to identify the root cause of defect and yield issues. Identifying exclusionary or baseline sources of defects or yield impacts and recommending/leading corrective actions/fixes. Lead/participate continuous improvement projects on products and processes for improved performance. Provide expertise on process flow segments. Lead/participate in multi-area problem solving teams to provide expertise to troubleshoot complex yield problems. The ideal candidate should exhibit the following behavioral traits: Solid analytical skills and a passion for data analysis and problem solving. Organizational skills with attention to detail. Excellent interpersonal skills with the ability to work with people at all levels. Communication and presentation skills to influence a wide variety of groups at all levels. Demonstrate excellent teamwork and leadership skills, demonstrated pro
Become a part of our caring community Most AI engineering jobs are a thin wrapper around a model API. This role is different. We build the platform that transforms millions of clinical documents into trusted, actionable data. Our systems use large language models (LLMs) to read medical records, extract structured facts, answer complex questions with citations back to the source document, and route ambiguous cases to human experts for review. Our users make decisions that impact real healthcare outcomes, so “good enough” is not good enough. Building AI systems that are accurate, reliable, auditable, and scalable is at the core of this role. As a Senior AI Applied Engineer, you will design, build, deploy, and operate production AI systems used at scale within one of the largest health insurers in the United States. You will own solutions end-to-end, from user experience and APIs to model orchestration, evaluation frameworks, infrastructure, and production operations. Why Join Us Build production AI systems where LLMs are in the critical path, not just demos or proofs of concept. Work on extraction, retrieval, agentic workflows, and human-review systems that process real healthcare data at scale. Own projects end-to-end across frontend, backend, AI orchestration, infrastructure, deployment, and operations. Solve challenging problems around accuracy, explainability, traceability, and reliability in regulated environments. Ship quickly in a small, high-impact team that embraces AI-assisted development and rigorous quality standards. Build systems that continuously improve through expert feedback, evaluations, and human-in-the-loop workflows. Key Responsibilities Design, develop, and deploy full-stack AI-powered application
Become a part of our caring community You have shipped AI products before. You understand the difference between a demo and a production system. You have strong opinions about evaluation frameworks because you have experienced the consequences of operating without them. You are at your best when you own architecture decisions while continuing to build and deliver critical code yourself. We build the platform that transforms millions of clinical documents into trusted, actionable data. Our systems use large language models (LLMs) to read medical records, extract structured facts, answer complex questions with citations back to source documents, and route complex cases to human experts. The output of these systems supports healthcare decisions that impact real members. As a Lead AI Applied Engineer, you will provide technical leadership for AI-enabled products and platforms, define architectural direction, establish engineering standards, and personally design and build the most critical components of our systems. You will lead through both technical expertise and execution, helping the team deliver reliable, scalable, and auditable AI solutions in a highly regulated healthcare environment. Why Join Us Lead the architecture of production AI systems where LLMs are foundational to the product experience. Make key technical decisions regarding model selection, system boundaries, platform architecture, and build-versus-buy strategies. Own the highest-risk and highest-impact technical challenges involving reliability, explainability, and correctness. Influence engineering culture and establish standards that shape how the team builds and ships AI products. Work on systems operating at meaningful scale, processing millions of documents and supporting healthcare decisions across a large member population. Partner
Get new reliability engineer jobs by email
Daily job updates · Unsubscribe anytime