Jobiba hiring network

Performance And Systems Engineer Jobs

6,348 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current performance and systems engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

D
Datadog
📍 Madrid• Full-time
26 days ago

We’re looking for Software Engineering Interns to help build and scale the systems that power Datadog’s observability and security platform. Interns contribute directly to real-world engineering challenges across backend, frontend, infrastructure, data engineering, and developer tooling while working alongside experienced engineers and mentors. You’ll help design, build, and improve systems that process and analyze massive volumes of metrics, logs, and application data in real time. Whether you’re interested in distributed systems, Kubernetes, AI-powered products like Bits AI, or developer platform tooling, you’ll work on meaningful projects that deliver impact for customers at global scale. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Contribute to production systems that process and analyze large-scale observability and application data in real time Build and improve distributed systems across backend infrastructure, developer platforms, and cloud-native services Help identify and solve performance, reliability, and scalability challenges in critical services supporting Datadog’s growing customer base Own and deliver technical projects from design through deployment with guidance from experienced engineers and mentors Develop technical expertise through hands-on experience with technologies such as Kubernetes, distributed systems, and cloud-native infrastructure Collaborate with fellow interns, mentors, and engineers while building software that delivers impact at global scale Who You Are: Expected to graduate in 2027 with a degree in Computer Science, Software Engineering, or a related technical field from a university in Spain Demonstrate strong computer science fundamentals, including data structures, algorithms, and software

kubernetesrestai
View job →
NR
New Relic
📍 India• Full-time
1mo ago

We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity At New Relic, we provide our customers real-time insights, so they can innovate faster. Our software delivers insightful observability tools across different technologies and distributed systems, enabling software engineering teams to quickly identify, understand and tackle issues, analyze performance and get the most of their software and infrastructure. The Infrastructure product organization develops New Relic infrastructure instrumentation agents, next generation data processing and management services, vulnerability management, and security testing capabilities for on-prem and cloud customers. We work with data at a scale using a diverse tech stack (Go, Java, JavaScript, React GraphQL, Kubernetes, many public cloud web services, and more). As a senior backend engineer, you will help us build and extend next generation solutions such as a control plane for customers to manage their data pipelines at scale. New Relic is looking for engineers who are interested in building a brand-new observability experience. This high-impact engineering position is a phenomenal opportunity to own and build a set of next generation services and capabilities for the company. We are searching for a motivated engineer who is ready for a career-defining role in their next opportunity. We look forward to talking with you! What you'll do ● Design, Build, maintain, and scale back-end services and their support tools. ● Participate in architectural definitions with a high degr

javascriptjavareact
View job →
S
Sentry
📍 San Francisco• Full-time• $155K – $400K/yr
1mo ago

About Sentry Software runs the world and the pace is faster than ever. Sentry helps developers fix errors and performance issues before users notice, so teams can spend less time firefighting and more time building. Trusted by 200,000+ organizations, Sentry is today’s application monitoring standard and our team is building its AI-native future. About the role As a Senior Software Engineer on Sentry’s AI/ML team, you’ll be responsible for building the evaluation infrastructure that measures the accuracy, reliability, and real-world performance of our AI systems. This role is critical to ensuring that our debugging agents and AI-powered features behave correctly, safely, and predictably as they scale. You’ll design datasets, benchmarks, and test harnesses that turn ambiguous AI behavior into measurable signals, helping the team ship AI with confidence. In this role you will Design and build robust evaluation frameworks to measure accuracy, reliability, regressions, and edge cases in AI systems Create and curate high-quality datasets, golden test cases, and benchmarks grounded in real production data Build automated test harnesses and metrics pipelines to continuously evaluate models, prompts, and agentic workflows Partner closely with applied AI engineers and product leaders to define what “good” looks like and translate it into measurable criteria Own the evaluation lifecycle for major AI initiatives, from early experimentation through production monitoring You’ll love this job if you Care deeply about correctness, rigor, and measurement in AI systems Enjoy turning fuzzy product goals and model behavior into concrete tests and metrics Like building foundational infrastructure that unlocks faster iteration and higher confidence for the entire AI team Thrive in cross-functional environments and enjoy influencing model design through better evaluation Qualifications Minimum 5+ years of professional experience with a Bachelor’s degree in computer science, machine learni

typescriptpythonmachine learning
View job →
D
Datadog
📍 Massachusetts• Full-time• From $100K/yr
1mo ago

We’re looking for Software Engineering Interns to help build and scale the systems that power Datadog’s observability and security platform. Interns contribute directly to real-world engineering challenges across backend, frontend, infrastructure, data engineering, and developer tooling while working alongside experienced engineers and mentors. You’ll help design, build, and improve systems that process and analyze massive volumes of metrics, logs, and application data in real time. Whether you’re interested in distributed systems, Kubernetes, AI-powered products like Bits AI, or developer platform tooling, you’ll work on meaningful projects that deliver impact to customers at global scale. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Contribute to production systems that process and analyze large-scale observability and application data in real time Build and improve distributed systems across backend infrastructure, developer platforms, and cloud-native services Help identify and solve performance, reliability, and scalability challenges in critical services supporting Datadog’s growing customer base Own and deliver technical projects from design through deployment with support from experienced engineers and mentors Develop technical expertise through hands-on experience with technologies such as Kubernetes, distributed systems, and cloud-native infrastructure Collaborate with fellow interns, mentors, and engineers while building software that delivers impact at global scale Who You Are: Pursuing a degree in Computer Science, Software Engineering, or a related technical field, or have equivalent practical experience Targeting a 2028 full-time start date Demonstrate strong computer science fundamentals, including data struc

kubernetesgitrest
View job →

Graphcore is a globally recognised leader in Artificial Intelligence computing systems. The company designs advanced semiconductors and data centre hardware that provide the specialised processing power needed to drive AI innovation, while delivering the efficiency required to support its broader adoption. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. We are opening a new AI Engineering Campus in Bengaluru which will play a central role in Graphcore's work building the future of AI computing. The physical design team sits within the wider silicon design team which includes RTL, verification and DFT. Our work also involves strong links with architecture, packaging and product engineering. We are responsible for working with those teams to create high-quality RTL and building the final chip layout (e.g. GDSII) ensuring a signoff-quality design is delivered to the Foundry (e.g. TSMC). We are looking to hire high-quality silicon physical design engineers to join our team. The successful candidate will support the team with achieving our goals and creating the right engineering solutions. We are a collaborative team and good communication is essential, as is the ability to adapt and learn. For the successful candidate we offer an open, honest and collaborative environment working on leading-edge designs at the most advanced nodes. Our engineers are not siloed, and they are trusted and encouraged to ta ke ownership of their designs and problem solutions. You will be part of a team that looks for improvements to everything we do: our designs, our flows, our methodologies, our infrastructure. Responsibilities and Duties Applicants will be expected to contribute technically to the development of Graphcore's next generation of AI superchips, focusing on achieving robust, high-performance and power-efficient designs

pythonaigo
View job →
DU
DoorDash USA
📍 San Francisco• Full-time
16 days ago

About the Team The Code Quality team sits within the Developer Platform organization and owns the systems that keep DoorDash's codebase healthy and secure as it scales: static analysis, quality gates, test frameworks, regression infrastructure, and tooling. Our job is to make sure the signals engineers rely on before shipping — test results, coverage, performance feedback etc — are fast and trustworthy. The decisions we make about tooling and standards directly shape how confidently and quickly engineering teams at DoorDash can ship to production. About the Role We're looking for Software Engineers to help build and maintain the systems that validate code quality across DoorDash's engineering org, treating our tooling as a critical product for the engineers who rely on it every day: static analysis and quality gates, test frameworks and regression infrastructure. You’ll design the tooling and automation that will help derive trustworthy quality signals, integrate them into the development lifecycle, and make it easy for engineers to execute reliable, repeatable workflows. You will collaborate across the engineering org, partnering directly with the teams who use what you build to understand the accuracy, reliability and performance of their functionality. You will report into the Engineering Manager on our Code Quality team in our Developer Platform organization. You must be located in either San Francisco, CA, Sunnyvale, CA, Los Angeles, CA, Seattle, WA, or New York, NY. You're excited about this opportunity because you will… Build and maintain quality tooling — static analysis, quality gates, coverage reporting, test frameworks, regression infrastructure — and integrate it directly into our developer workflows and CI/CD pipelines Define and derive quality signals - flakiness, pass rate, coverage, performance, scale readiness etc - Build tooling that improves everyday engineering workflows, including local development, CI/CD, debugging, and rollou

awsci/cdgit
View job →

We are looking for an innovative thermal solutions integration Engineer. NVIDIA offers you to be a part of the System Product Engineering group and be responsible for assuring the best quality products to be a sale and deliver to NVIDIA’ s customers. The job provides deep knowledge of NVIDIA systems, a system-level view of our solutions, and a dynamic and positive working environment and offers the candidates the opportunity to take a major role in our testing strategy by leading our thermal solutions integration and testing for the company production. NVIDIA Networking unit has continuously reinvented itself over two decades. Our high-speed buses & network products are leading in the markets with innovative ways to improve speed and bandwidth from one generation to another. Today, we are increasingly known as the place for getting “End-to-End High-Speed Ethernet and InfiniBand Solutions” We're looking to grow our company and build our teams with smart people who can join us at the forefront of technological advancement. If you are passionate about enabling the highest quality Network products that will change the world, we want to hear from you! What You’ll Be Doing: Integration and testing of next-generation, large-scale thermal, pressure, and liquid management solutions. Perform qualification tests for cutting edge cooling and sensing solution. Analyse and summarize thermal performance data and fluid dynamics results to support design reviews and decision-making. Primary onsite focal point for malfunctions in thermal liquid cooling stations. Diagnose and resolve hardware/software issues to maintain continuous development labs activity Maintain and update hardware (manifolds, connectors, sensors) and software versions across all thermal systems according to engineering specifications. Perform initial RCA on thermal system failures. Extract detailed fail reports and corrective/

M
1mo ago

ABOUT THE TEAM The Canvas Core team builds and maintains the foundational platform that powers Mural’s visual thinking experience. This includes the infinite canvas, key editor components, document editing behaviors, asset management, real-time collaboration, and the systems that enable fast, reliable, and intuitive interaction on the canvas. We’re also responsible for the Mural UI, the real-time message protocol that enables seamless remote collaboration, and the developer-friendly APIs that internal teams use to build features like diagramming, workshops, presentations, integrations, and AI-enabled product capabilities. Our mission is to ensure the Mural editor is fast, reliable, intuitive, and easy to build on. We prioritize performance, simplicity, developer experience, and platform quality, enabling teams across the company to ship quickly and safely on top of Canvas Core. YOUR MISSION As a Senior Software Engineer, you’ll help design, build, and improve the Canvas platform so that the Mural editor remains reliable, high-performing, and intuitive for our users. You’ll work on the systems that power real-time collaboration, shared document editing, spatial interactions, rendering and interaction performance, asset management, developer APIs, and AI-enabled product capabilities across the Mural editor. Your role will be to reduce platform complexity, improve the quality and speed of Canvas development, and help teams ship high-quality editor experiences quickly and safely. You’ll partner closely with Product, Design, Engineering, and other stakeholders to turn ambiguous product and platform problems into clear, maintainable technical solutions. Senior Engineers at Mural lead by example through strong technical execution, thoughtful design, high-quality implementation, and collaborative problem-solving. They help raise the bar for their team through design discussions, code reviews, mentoring, documentation, and pragmatic improvements to engineering practices. WHA

javascripttypescriptjava
View job →
M
Mongodb
📍 Bengaluru• Full-time
1mo ago

MongoDB Technical Services Engineers use their exceptional problem solving and customer service skills, along with their deep technical experience, to advise customers and to solve their complex MongoDB problems. Technical Service Engineers are experts in the entire MongoDB ecosystem - database server, drivers, cloud and infrastructure. This also includes services such as Atlas (database as a service), or Cloud Manager (which helps customers with automation, backup and monitoring of their MongoDB systems). Our engineers combine their MongoDB expertise with passion, initiative, teamwork and a great sense of humor to help our customers to be successful with MongoDB. We are looking to speak to candidates who are based in Bengaluru for our hybrid working model. Cool things you’ll do You'll be working alongside our largest customers, solving their complex challenges - resolving questions on architecture, performance, recovery, security, and everything in between. You'll be an expert resource on best practices in running MongoDB at scale, whatever that scale may be. You'll be an advocate for customers' needs - interfacing with our product management and development teams on their behalf. And you'll contribute to internal projects, including software development of support tools for performance, benchmarking, and diagnostics. As an ideal candidate, you will have We consider all candidates with an eye for those who are self taught, curious, and multi-faceted. Our ideal TSE candidate should also have 6 plus years of experience Systems engineering experience, including Linux performance, memory management, I/O tuning, configuration, security, Storage, networking, clusters, and troubleshooting Should have a good understanding of Networking concepts and protocols (DNS, TCP/IP, SSL/TLS, etc.) Broad awareness of customer workloads and use cases, including performance, availability, and scalability Experience analyzing issues holistically, from the application tier through the dat

javascriptpythonjava
View job →
O
1mo ago

About the team The Fleet team at OpenAI supports the computing environment that powers our cutting-edge research and product development. We oversee large-scale systems that span data centers, GPUs, networking, and more, ensuring high availability, performance, and efficiency. Our work enables OpenAI’s models to operate seamlessly at scale, supporting both internal research and external products like ChatGPT. We prioritize safety, reliability, and responsible AI deployment over unchecked growth. About the role As a software engineer on the Fleet High Performance Computing (HPC) team, you will be responsible for the reliability and uptime of all of OpenAI’s compute fleet. Minimizing hardware failure is key to research training progress and stable services, as even a single hardware hiccup can cause significant disruptions. With increasingly large supercomputers, the stakes continue to rise. Being at the forefront of technology means that we are often the pioneers in troubleshooting these state-of-the-art systems at scale. This is a unique opportunity to work with cutting-edge technologies and devise innovative solutions to maintain the health and efficiency of our supercomputing infrastructure. Our team empowers strong engineers with a high degree of autonomy and ownership, as well as ability to effect change. This role will require a keen focus on system-level comprehensive investigations and the development of automated solutions. We want people who go deep on problems, investigate as thoroughly as possible, and build automation for detection and remediation at scale. In this role, you will: Build and maintain automation systems for provisioning and managing server fleets. Develop tools to monitor server health, performance, and lifecycle events. Collaborate with clusters, networking, and infrastructure teams. Partner with external operators to ensure a high level of quality. Identify and fix performance bottlenecks and inefficiencies. Continuously improve automati

pythonsqlaws
View job →
O
OpenAI
📍 San Francisco• Full-time
1mo ago

About the team The Fleet team at OpenAI supports the computing environment that powers our cutting-edge research and product development. We oversee large-scale systems that span data centers, GPUs, networking, and more, ensuring high availability, performance, and efficiency. Our work enables OpenAI’s models to operate seamlessly at scale, supporting both internal research and external products like ChatGPT. We prioritize safety, reliability, and responsible AI deployment over unchecked growth. About the role As a software engineer on the Fleet Hardware team, you will be responsible for the reliability and uptime of all of OpenAI’s compute fleet. Minimizing hardware failure is key to research training progress and stable services, as even a single hardware hiccup can cause significant disruptions. With increasingly large supercomputers, the stakes continue to rise. Being at the forefront of technology means that we are often the pioneers in troubleshooting these state-of-the-art systems at scale. This is a unique opportunity to work with cutting-edge technologies and devise innovative solutions to maintain the health and efficiency of our supercomputing infrastructure. Our team empowers strong engineers with a high degree of autonomy and ownership, as well as ability to effect change. This role will require a keen focus on system-level comprehensive investigations and the development of automated solutions. We want people who go deep on problems, investigate as thoroughly as possible, and build automation for detection and remediation at scale. In this role, you will: Build and maintain automation systems for provisioning and managing server fleets. Develop tools to monitor server health, performance, and lifecycle events. Collaborate with clusters, networking, and infrastructure teams. Partner with external operators to ensure a high level of quality. Identify and fix performance bottlenecks and inefficiencies. Continuously improve automation to reduce manual work

pythonsqlaws
View job →
O
1mo ago

About the Team The Workload team is responsible for designing and running OpenAI’s LLM training and inference infrastructure that powers frontier models at massive scale. Our systems unify how researchers train and serve models, abstracting away the complexity of performance, parallelism, and execution across vast GPU/accelerator fleets. By providing this foundation, the Workload team ensures that researchers can focus on advancing model capabilities while we handle the scale, efficiency, and reliability required to bring those models to life. About the Role We are looking for an engineer to design and implement the dataset infrastructure that powers OpenAI’s next-generation training stack. You will be responsible for building standardized dataset interfaces, scaling pipelines across thousands of GPUs, and proactively testing performance bottlenecks. In this role, you will collaborate closely with the multimodal researchers, and other infra groups to ensure datasets are unified, efficient, and easy to consume. In this role, you will: Design and maintain standardized dataset APIs, including for multimodal (MM) data that cannot fit in memory. Build proactive testing and scale validation pipelines for dataset loading at GPU scale. Collaborate with teammates to integrate datasets seamlessly into training and inference pipelines, ensuring smooth adoption and a great user experience. Document and maintain dataset interfaces so they are discoverable, consistent, and easy for other teams to adopt. Establish safeguards and validation systems to ensure datasets remain reproducible and unchanged once standardized. Debug and resolve performance bottlenecks in distributed dataset loading (e.g., straggler systems slowing global training). Provide visualization and inspection tools to surface errors, bugs, or bottlenecks in datasets. You might thrive in this role if you: Have strong engineering fundamentals with experience in distributed systems, data pipelines, or infrastructure.

awsrestai
View job →
H
Hp
📍 Colorado• $59.4K – $89.6K/yr
9 days ago

Software Quality Engineer Description - This role is responsible for maintaining the quality, reliability, and performance of software applications throughout the development lifecycle. The role identifies and rectifies defects, ensures adherence to established quality standards, and contributes to the overall improvement of the software development process. The role involves various activities aimed at preventing and detecting issues, thereby enhancing the end user experience. The role creates and executes comprehensive test plans, test cases, and test scripts based on project specifications. *Onsite in Ft. Collins 5-days a week Responsibilities • Executes established test plans and protocols for assigned portions of code for end-user applications, systems software, and firmware running on hardware, local, networked, and Internet- based platforms; identifies, logs, and debugs assigned issues. • Perform Functional and Solution Testing of Video/Collaboration Software • Additionally, codes and programs test scripts, automation, and integration activities based on specific test requirements. • Conducts functional, integration, regression, and performance testing to validate software functionality. • Automates testing processes using appropriate tools and frameworks to improve efficiency and repeatability. • Monitors and enforces adherence to established coding standards, design guidelines, and best practices. • Monitors software performance and conducts load and stress testing to identify bottlenecks and performance issues. • Prepares and maintains QA-related documentation, including test plans, test matrices, and testing reports. • Develops understanding of and relationship with internal and outsourced development partners on software applications design and development. • Participates as a member of project

pythonaijenkins
View job →
H
Hp
📍 Colorado• $59.4K – $89.6K/yr
9 days ago

Software Quality Engineer Description - This role is responsible for maintaining the quality, reliability, and performance of software applications throughout the development lifecycle. The role identifies and rectifies defects, ensures adherence to established quality standards, and contributes to the overall improvement of the software development process. The role involves various activities aimed at preventing and detecting issues, thereby enhancing the end user experience. The role creates and executes comprehensive test plans, test cases, and test scripts based on project specifications. *Onside in Ft Collins 5-days a week Responsibilities • Executes established test plans and protocols for assigned portions of code for end-user applications, systems software, and firmware running on hardware, local, networked, and Internet- based platforms; identifies, logs, and debugs assigned issues. • Perform Functional and Solution Testing of Video/Collaboration Software • Additionally, codes and programs test scripts, automation, and integration activities based on specific test requirements. • Conducts functional, integration, regression, and performance testing to validate software functionality. • Automates testing processes using appropriate tools and frameworks to improve efficiency and repeatability. • Monitors and enforces adherence to established coding standards, design guidelines, and best practices. • Monitors software performance and conducts load and stress testing to identify bottlenecks and performance issues. • Prepares and maintains QA-related documentation, including test plans, test matrices, and testing reports. • Develops understanding of and relationship with internal and outsourced development partners on software applications design and development. • Participates as a member of project t

pythonaijenkins
View job →

We anticipate the application window for this opening will close on - 10 Oct 2026 Careers that change lives start here. Medtronic is a global leader in healthcare technology with a Mission to alleviate pain, restore health, and extend life. Our 95,000 employees work across more than 150 countries to put patients first — developing innovative medical technologies that improve the lives of 72+ million patients each year. Your unique talents will help shape the future of healthcare while building a career grounded in purpose, growth, and impact. A Day in the Life The Neuromodulation Research & Development organization develops therapies and technologies that address chronic pain and other neurological conditions. Within the Interventional Pain portfolio, teams focus on minimally invasive therapies used to treat musculoskeletal pathologies, including vertebral compression fractures, nerve pain, metastatic bone tumors, and benign bone tumors. This role supports the development of system-level verification and validation strategies for new products and enhancements across mechanical, electrical, and software components. This position is based in Santa Clara, California, and follows an onsite work model. No travel is required for this role. As a Research & Development Test Engineer, you will lead and execute verification and validation activities that support the development of interventional pain therapies and associated medical device systems. You will collaborate with cross-functional partners to define test strategies, evaluate product performance, manage technical risks, and generate evidence needed to support product development and regulatory requirements. Primary Responsibiliti

recruitmentHR
View job →
🔔

Get new performance and systems engineer jobs by email

Daily job updates · Unsubscribe anytime