About the Team OpenAI, in close collaboration with our capital partners, is building the world’s most advanced AI infrastructure ecosystem. Our Industrial Compute organization develops and deploys large-scale AI campuses designed to support the next generation of frontier model training and inference workloads. The Hardware Operations team is responsible for ensuring the reliability, availability, and lifecycle health of OpenAI’s compute infrastructure. We partner closely with Data Center Operations, Fleet Health Engineering, Manufacturing, Network Infrastructure, Capacity Planning, and our infrastructure partners to maintain world-class operational performance across rapidly expanding AI environments. As we scale globally, we are building the operational frameworks, reliability standards, and sustaining engineering practices required to support thousands of GPUs and servers across multiple campuses. About the Role We are seeking a Datacenter Hardware Technician Lead to serve as the senior on-site technical authority for hardware reliability and fleet health at one of OpenAI’s flagship AI campuses. This role operates at the intersection of hardware operations, sustaining engineering, and fleet reliability. You will partner closely with Cloud Service Provider operations teams, OpenAI fleet-health engineers, hardware engineering teams, and OEM vendors to identify, diagnose, and resolve hardware issues affecting production systems. Beyond day-to-day operational support, you will drive root cause investigations, reliability improvement initiatives, lifecycle management programs, and operational readiness efforts. You will help establish hardware maintenance standards, operational procedures, and best practices that scale across future OpenAI infrastructure deployments. The ideal candidate combines deep hands-on datacenter hardware expertise with strong troubleshooting, failure analysis, and cross-functional leadership skills. Candidates must be able to sit onsite at our
Jobiba hiring network
Operations Performance And Analytics Engineer Jobs
10,000 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current operations performance and analytics engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
Location: San Francisco, CA (hybrid) What is Verse? The race to AI has become the race to power. Every breakthrough in artificial intelligence depends on one thing: access to electricity. But across the country, aging grid infrastructure and years-long interconnection queues are slowing the deployment of the data centers that will power the next generation of innovation. Solving this challenge isn't just about energy—it's about unlocking the future of AI. At Verse, we're building the energy intelligence platform for the AI economy. Our software helps the world's largest energy consumers achieve faster, cheaper, and cleaner power by combining real-time control of energy assets with complete visibility into their energy portfolio. Backed by Bessemer Venture Partners, GV, Coatue, and NVIDIA, and built by pioneers in grid-scale batteries, energy markets, and enterprise software, we're redefining how the world's most ambitious organizations access and manage energy. The Role We're looking for a highly analytical Senior Business Operations Analyst to support the growth and execution of our Dispatch Intelligence product. This role sits at the intersection of business operations, customer, product, engineering, and data science and has two core areas of responsibility. First, you will help drive overall program execution for Dispatch Intelligence: bringing structure to complex cross-functional initiatives, improving processes, tracking progress, and ensuring teams stay aligned on priorities and timelines. Second, you will help build and operate the processes through which flexible energy assets are onboarded onto the Verse platform and continuously improve their operational performance. The ideal candidate combines strong analytical problem-solving with exceptional project management and is comfortable working across both technical and commercial teams. You will play a critical role in helping Verse scale Dispatch Intelligence from individual projects and assets
Amplitude is the leading AI analytics platform, helping over 4,700 customers—including Atlassian, Burger King, NBCUniversal, and Square—build better products and digital experiences. With powerful AI Agents embedded across our platform, teams can analyze, test, and optimize user experiences faster than ever. Ranked #1 across multiple categories in G2’s Winter 2026 Report, Amplitude is the best-in-class solution for product, data, and marketing teams. Learn more at amplitude.com . As an organization, we deliver for our customers by living our values. We operate from a place of humility, take ownership of problems and successes, approach challenges with a growth mindset, and put our customers at the center of everything we do. Amplitude’s Commitment to Diversity Equity & Inclusion (DEI): Amplitude believes that diversity enables the creation of better products, improves the ability to solve complex problems, and drives more powerful solutions. We strive to create an environment of inclusion—one focused on psychological safety, empathy, and human connection—that will allow employees of all backgrounds to thrive. At Amplitude, we’re building the operating system for digital products. While we’re known as the leader in product analytics, our Statsig team is redefining how companies learn, iterate, and ship better experiences. Statsig is one of the fastest-growing and most strategic bets at Amplitude. It sits at the intersection of product, data, and decision-making, and we’re looking for a Product Engineer to help take it to the next level. Why this role matters This isn’t a “take tickets and ship code” role. You’ll operate as an owner, shaping both the product and the technical direction of a system used by some of the most sophisticated product teams in the world. You’ll work across the stack, NextJS and Node.js, to build intuitive, high-performance experiences that make experimentation accessible, powerful, and trustworthy. What you’ll do Own critical product surf
Job Title Service Operations Manager Job Description Your role: Lead service operations for a global interventional cardiology and vascular device portfolio, owning field performance, delivery execution, and operational governance across a geographically distributed team of field service engineers, product support specialists, and service coordinators. You will set the operating rhythm for corrective, preventive, and installation service across North America while partnering with international service leaders to align standards globally. Own the service training and technical education function end to end, including training program management, instructor-led and digital course delivery, field certification programs, and training center operations across three continental sites. You will ensure engineers maintain verified competency on legacy, newly acquired, and next-generation product lines in a regulated medical device environment where patient safety depends on technician proficiency. Drive measurable improvement in service performance using indicators such as time to system restoration, first-visit resolution, installation quality, contract capture, and preventive maintenance completion. You will lead root-cause analysis on systemic delivery issues, design operational improvements with engineering and supply chain partners, and translate field data into changes that improve customer uptime and reduce repeat service events. Manage service quality and compliance activities including complaint escalation, field execution of corrective actions and compliance tracking, product lifecycle support for systems approaching end of service, and regulatory documentation across all served markets. You will work directly with quality, regulatory, and product engineering teams to ensure service processes satisfy medical device requirements while maintaining the speed and respo
About the Team The Spark Platform team owns and operates DoorDash's Apache Spark ecosystem — the execution runtime, remote shuffle service, cluster scheduler, and reliability tooling that powers the company's data, analytics, and ML workloads. We run Spark across the company at significant scale and continue to expand the workloads, capabilities, and consumer base we serve. Orchestrating and operating thousands of Spark cluster deployments is a complex distributed system problem which the team invests heavily in runtime optimization, systems architecture, multi-tenant scheduling, and end-user tooling. About the Role As a Senior Software Engineer on Spark Platform, you will set the technical direction for our in-house Spark deployment and shape the architecture that will run DoorDash's data, analytics, and ML compute for the next five years and beyond. You will own the deep, cross-cutting problems that span the runtime, the shuffle service, the scheduler, and the overall service reliability — making the architectural calls that compound across the platform's lifetime. You will partner with the Engineering Manager on technical roadmap, hiring, and team shape, and act as the senior technical voice in cross-team partnerships with Data Engineering, ML Platform, and product engineering teams that depend on the platform. You must be located in San Francisco, Sunnyvale, Seattle, or New York City for this hybrid position. You will report into the Engineering Manager on our Spark Platform team. You're excited about this opportunity because you will… Set the multi-year technical direction for an in-house Spark-on-Kubernetes platform — runtime, shuffle, scheduler, reliability — and make the architectural calls that compound for years. Own the deepest distributed-systems problems on the team: shuffle architecture, multi-tenant scheduling, runtime performance, and the failure modes that only show up at scale. Partner with the Engineering Manager on technical roadmap, hiring, inte
Senior Product Engineer At Amplitude, we’re building the operating system for digital products. While we’re known as the leader in product analytics, our Statsig team is redefining how companies learn, iterate, and ship better experiences. Statsig is one of the fastest-growing and most strategic bets at Amplitude. It sits at the intersection of product, data, and decision-making, and we’re looking for a Product Engineer to help take it to the next level. Why this role matters This isn’t a “take tickets and ship code” role. You’ll operate as an owner, shaping both the product and the technical direction of a system used by some of the most sophisticated product teams in the world. You’ll work across the stack, React frontend, Node.js and Python backend, to build intuitive, high-performance experiences that make experimentation accessible, powerful, and trustworthy. What you’ll do Own critical product surfaces end-to-end from ideation to production and beyond Drive product direction in partnership with design and product leaders Build and evolve systems across React, Node.js, and Python that scale with customer growth Leverage modern AI tooling to dramatically accelerate development, prototyping, and iteration cycles Raise the bar for engineering quality - code, architecture, and user experience Lead by example - through hands-on development, technical mentorship, and thoughtful decision-making Move fast on ambiguous problems, turning ideas into shipped, customer-facing value What we’re looking for 5+ years of experience in software engineering (full stack experience preferred) Education: B.S. in Computer Science or an equivalent technical field Proven experience operating at a senior-level scope. You’ve led large, ambiguous initiatives with significant business impact Strong full-stack expertise (React + backend systems such as Node.js and/or Python) A deep sense of ownership - you don’t wait for direction; you create it Exceptional product taste - you care deeply ab
Senior Product Engineer At Amplitude, we’re building the operating system for digital products. While we’re known as the leader in product analytics, our Statsig team is redefining how companies learn, iterate, and ship better experiences. Statsig is one of the fastest-growing and most strategic bets at Amplitude. It sits at the intersection of product, data, and decision-making, and we’re looking for a Product Engineer to help take it to the next level. Why this role matters This isn’t a “take tickets and ship code” role. You’ll operate as an owner, shaping both the product and the technical direction of a system used by some of the most sophisticated product teams in the world. You’ll work across the stack, React frontend, Node.js and Python backend, to build intuitive, high-performance experiences that make experimentation accessible, powerful, and trustworthy. What you’ll do Own critical product surfaces end-to-end from ideation to production and beyond Drive product direction in partnership with design and product leaders Build and evolve systems across React, Node.js, and Python that scale with customer growth Leverage modern AI tooling to dramatically accelerate development, prototyping, and iteration cycles Raise the bar for engineering quality - code, architecture, and user experience Lead by example - through hands-on development, technical mentorship, and thoughtful decision-making Move fast on ambiguous problems, turning ideas into shipped, customer-facing value What we’re looking for 3+ years of experience in software engineering Education: B.S. in Computer Science or an equivalent technical field Proven experience operating at a senior-level scope. You’ve led large, ambiguous initiatives with significant business impact Strong full-stack expertise (React + backend systems such as Node.js and/or Python) A deep sense of ownership - you don’t wait for direction; you create it Exceptional product taste - you care deeply about UX, details, and building thin
We're transforming the grocery industry At Instacart, we invite the world to share love through food because we believe everyone should have access to the food they love and more time to enjoy it together. Where others see a simple need for grocery delivery, we see exciting complexity and endless opportunity to serve the varied needs of our community. We work to deliver an essential service that customers rely on to get their groceries and household goods, while also offering safe and flexible earnings opportunities to Instacart Personal Shoppers. Instacart has become a lifeline for millions of people, and we’re building the team to help push our shopping cart forward. If you’re ready to do the best work of your life, come join our table. Instacart is a Flex First team There’s no one-size fits all approach to how we do our best work. Our employees have the flexibility to choose where they do their best work—whether it’s from home, an office, or your favorite coffee shop—while staying connected and building community through regular in-person events. Learn more about our flexible approach to where we work. Overview The Commercial Scaled Intelligence (CSI) team is an AI-first team dedicated to delivering actionable commercial insights and scalable automation to drive revenue growth and operational efficiency across the company. The team focuses on intelligence generation, predictive analytics, and workflow automation to enable data-driven decision-making and optimize commercial performance. As an Ads AI Analytics Lead II, you will own the intelligence behind our Ads agents. You will design the Ads semantic/context layer and build vertical AI agents that analyze campaigns, diagnose performance, and recommend actions that improve ROAS, pacing, and partner outcomes. You will partner with Ads GTM, Product, Data Science, and Engineering to ship production agents with measurable lift. About the Job Define Ads ontologies and metrics for campaigns, budgets, bids
At Affirm, we exist for the moments that matter—giving people a clear, predictable way to pay over time, with no hidden fees, no surprises, and no tradeoffs on what matters most. Site Reliability Engineering at Affirm is a small, yet crucial, team that helps our Engineering partners to “Operate What They Own” with excellence to protect their customers’ experience. SRE accomplishes this through defining frameworks and best practices for operating applications, building tooling, and providing training and consulting. Some of the many SRE responsibilities are: Providing data and visibility to teams and leadership on application performance Guiding the development of SLOs Driving the Incident Management and Analysis process Steering the implementation of Change Management and Deployment practices Engaging in service and architectural conversations Recommending observability and alerting configurations The SRE team benefits from experience across many domains including: infrastructure, platform, and distributed systems capacity management, load and chaos testing automation, observability, and configuration management development and product experience The SRE team is seeking motivated software and systems engineers with the experience to build, iterate on, and expand incident lifecycle, reliability, and resilience practices throughout Affirms Engineering organization and beyond. What You'll Do: You will be responsible for owning and delivering quarterly goals for your team, leading engineers on your team through ambiguity to solve open-ended problems, and ensuring that everyone is supported throughout delivery. You will support your peers and stakeholders in the product development lifecycle by collaborating with infrastructure, product management, developer experience & analytics by participating in ideation, articulating technical constraints, and partnering on decisions that properly consider risks and trade-offs. You will proactively identify technical solutions
At Affirm, we exist for the moments that matter—giving people a clear, predictable way to pay over time, with no hidden fees, no surprises, and no tradeoffs on what matters most. Site Reliability Engineering at Affirm is a small, yet crucial, team that helps our Engineering partners to “Operate What They Own” with excellence to protect their customers’ experience. SRE accomplishes this through defining frameworks and best practices for operating applications, building tooling, and providing training and consulting. Some of the many SRE responsibilities are: Providing data and visibility to teams and leadership on application performance Guiding the development of SLOs Driving the Incident Management and Analysis process Steering the implementation of Change Management and Deployment practices Engaging in service and architectural conversations Recommending observability and alerting configurations The SRE team benefits from experience across many domains including: infrastructure, platform, and distributed systems capacity management, load and chaos testing automation, observability, and configuration management development and product experience The SRE team is seeking motivated software and systems engineers with the experience to build, iterate on, and expand incident lifecycle, reliability, and resilience practices throughout Affirms Engineering organization and beyond. What You'll Do: You will be responsible for owning and delivering quarterly goals for your team, leading engineers on your team through ambiguity to solve open-ended problems, and ensuring that everyone is supported throughout delivery. You will support your peers and stakeholders in the product development lifecycle by collaborating with infrastructure, product management, developer experience & analytics by participating in ideation, articulating technical constraints, and partnering on decisions that properly consider risks and trade-offs. You will proactively identify technical solutions
About the Team: OpenAI’s Hardware organization develops AI-native silicon and system-level solutions for the unique demands of advanced AI workloads. Building on efforts like Jalapeño, the team is developing future generations of AI-native silicon and tightly integrated systems to power the next generation of frontier models. By co-designing chips, systems, tools, and methodologies, the team helps deliver faster, more efficient, and production-ready hardware for OpenAI’s supercomputing platform. In this role you will: As a Hardware Test Engineer, you will work on Machine Learning/AI hardware system projects to craft the solutions for current and future data center deployments. You will bring a strong understanding of hardware system testing, excellent project management skills, and the ability to collaborate across multiple teams to ensure efficient lab operations. You will be responsible for designing, implementing, and executing comprehensive test plans that ensure the reliability, performance, and scalability of our supercomputing hardware systems. You will develop detailed test plans and methodologies tailored to hardware components, including processors, memory modules, custom accelerators and interconnects. You will collaborate with hardware design, manufacturing, firmware teams and vendors to identify, analyze, and resolve issues affecting hardware, power, thermal and high-speed interconnects. You will perform in-depth debugging on the hardware system Excellent analytical skills to diagnose hardware issues, troubleshoot problems, and propose solutions. Ability to interpret complex test data, identify trends, and draw meaningful conclusions. High-speed links, with a focus on SerDes (Serializer/Deserializer) technology to assess signal integrity, error rates, and overall link performance. You will collaborate with the lab manager to maintain the equipment and hardware systems, including oscilloscopes, thermal test chambers, liquid cooling systems, and other mea
We are looking for a highly motivated AI/ML Software Engineer to join the Enterprise Agentic AI Platform team within IT. You will work closely with Business Analysts, and Engineering teams to design, develop, and deploy enterprise AI solutions that improve productivity and automate business workflows across Engineering, Operations, and Manufacturing. What you'll be doing: Design, develop, and deploy Agentic AI applications using Large Language Models (LLMs), Retrieval-Augmented Generation (RAG), and AI orchestration frameworks. Build scalable AI services and reusable components integrated with enterprise applications such as PLM, SAP, and other business systems. Collaborate with business and IT teams to translate business requirements into AI-driven solutions. Develop secure, scalable APIs and enterprise integrations to enable intelligent workflows and automation. Improve AI solution quality, performance, and reliability through prompt engineering, evaluation, and continuous optimization. Partner with cross-functional teams throughout the Software Development Lifecycle (SDLC), from solution design through deployment and production support. What we need to see: Bachelor's or Master's degree in Computer Science, Information Technology, AI/ML, or a related field. 6+ years of software engineering experience with strong proficiency in Python and backend application development. Hands-on experience with Generative AI, LLMs, RAG, AI agents, REST APIs, and cloud-native application development. Experience integrating enterprise applications and building scalable, production-ready software solutions. Strong analytical, problem-solving, communicatio
We anticipate the application window for this opening will close on - 2 Oct 2026 Careers that change lives start here. Medtronic is a global leader in healthcare technology with a Mission to alleviate pain, restore health, and extend life. Our 95,000 employees work across more than 150 countries to put patients first — developing innovative medical technologies that improve the lives of 72+ million patients each year. Your unique talents will help shape the future of healthcare while building a career grounded in purpose, growth, and impact. A Day in the Life Job Specific Summary In this exciting role as a Manufacturing Engineer for Medical Devices Operations you will have responsibility for executing detailed process analysis to assure that the operations are working in accordance with standard procedures and quality regulations. Also, you will have to perform equipment and process evaluations to determine methods for collecting objective data to identify opportunities of improvements and determine actions to increase current performance of the operations. Provide technical assistance to manufacturing and maintenance personnel to correct discrepancies, to update or to improve the equipment, machines and tooling used on production. At Medtronic, we bring bold ideas forward with speed and decisiveness to put patients first in everything we do. In-person exchanges are invaluable to our work. We’re working 5 days a week onsite as part of our commitment to fostering a culture of professional growth and cross-functional collaboration as we work together to engineer the extraordinary. Responsibilities may include the following and other duties may be assigned. Team with manufacturing, maintenance and quality personnel on a daily basis to review / disposition production defects and provid
We’re looking for a Senior Engineering Manager who is ready to lead through ambiguity and improve how software gets built at MongoDB. This role leads teams focused on developer productivity, with an emphasis on measurable improvements to the software development lifecycle. This role can be based remotely in the United States. The Team The AXIS team (AI, X-functional tools, Insights, and Signals) sits within Developer Productivity and is responsible for overseeing the metrics and observability infrastructure of our expansive developer environment to help build a strong data-driven culture. You’ll also be a key partner in building the agentic ecosystem for AI-driven development across engineering. Candidate Profile We’re looking for an experienced leader with a passion for solving the big challenge of measuring developer productivity and providing the actionable signals that help teams improve their performance. They should be comfortable working collaboratively with other leaders and partners across our Engineering and Data teams in maximizing the use of data for insights and AI enablement. The right candidate for this role will have 4+ years of experience managing software engineers, including hiring, performance management, growth planning, and compensation; required for external candidates and preferred for internal candidates 8+ years of hands-on software engineering experience building and operating production systems; experience in developer tooling, platform engineering, observability, or data engineering is a strong plus Demonstrated the ability to lead through ambiguity, work across team boundaries, and deliver outcomes without close supervision Strong customer orientation and sound judgment in finding practical, high-leverage solutions Experience working with systems involving analytics, data pipelines, and metrics platforms Experience with AI tools development and enablement efforts Strong technical judgment, including the ability to evaluate t
We’re looking for a Senior Engineering Manager who is ready to lead through ambiguity and improve how software gets built at MongoDB. This role leads teams focused on developer productivity, with an emphasis on measurable improvements to the software development lifecycle. This role can be based remotely in Canada. The Team The AXIS team (AI, X-functional tools, Insights, and Signals) sits within Developer Productivity and is responsible for overseeing the metrics and observability infrastructure of our expansive developer environment to help build a strong data-driven culture. You’ll also be a key partner in building the agentic ecosystem for AI-driven development across engineering. Candidate Profile We’re looking for an experienced leader with a passion for solving the big challenge of measuring developer productivity and providing the actionable signals that help teams improve their performance. They should be comfortable working collaboratively with other leaders and partners across our Engineering and Data teams in maximizing the use of data for insights and AI enablement. The right candidate for this role will have 4+ years of experience managing software engineers, including hiring, performance management, growth planning, and compensation; required for external candidates and preferred for internal candidates 8+ years of hands-on software engineering experience building and operating production systems; experience in developer tooling, platform engineering, observability, or data engineering is a strong plus Demonstrated the ability to lead through ambiguity, work across team boundaries, and deliver outcomes without close supervision Strong customer orientation and sound judgment in finding practical, high-leverage solutions Experience working with systems involving analytics, data pipelines, and metrics platforms Experience with AI tools development and enablement efforts Strong technical judgment, including the ability to evaluate tradeoffs, i
Get new operations performance and analytics engineer jobs by email
Daily job updates · Unsubscribe anytime