About Graphcore Graphcore is a global leader in artificial intelligence computing systems. We design advanced semiconductors and data center hardware that deliver the specialized processing power needed to advance AI while improving the efficiency required for broad adoption. As part of SoftBank Group, Graphcore belongs to a family of companies developing some of the world’s most transformative technologies. Our new AI Engineering Campus in Austin will play a central role in building the future of AI computing. The Opportunity We are seeking a system validation engineering intern to help drive server blade and rack validation efforts for next-generation AI infrastructure hardware systems. This role focuses on post-silicon system validation across the full lifecycle of server hardware systems, ensuring functional and performance meets product objectives. You will help drive end-to-end blade and rack validation including development, execution, and debug while collaborating across silicon, firmware, systems, and platform teams. The Blade and Rack Validation team is responsible for ensuring system readiness and quality at scale. The team works cross-functionally with firmware, silicon, and system engineering teams to validate complex AI compute platforms. Type: 12-week summer internship Timing: May - August (exact dates to be confirmed) Commitment: Full-time What You’ll Do Help drive and execute post-silicon validation goals of AI compute blades and racks including testcase planning, development, and automation Help drive validation testcase execution and system debug against program achievements and report validation progress and risks. Drive provisioning and integration of system components (SoC FW, BMC, RMC, OS) for rack-level readiness Triage test failures, collect debug data, and collaborate on root cause analysis. Track validation coverage and continuously improve test processes and infrastructure. What You’ll Bring Working towards a Bachelor's
Jobiba hiring network
Ai Platform Engineer Jobs
10,000 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current ai platform engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Every experience we build is only as good as our understanding of the person using it. You'll help distill how people actually use New Relic, across classic interfaces and newer surfaces like agentic conversations, into models that let those experiences adapt in ways that feel relevant and contextual. This position is for the Personalization team, starting with a greenfield project! Our stack centers on Java 25, Spring, GraphQL, NoSQL, Kubernetes, Kafka and AWS. If you're passionate about understanding user behavior at scale and want to help shape how New Relic's classic and agentic experiences adapt to the people using them, we want to hear from you. What you'll do Build, maintain, and scale backend services and Kafka-based streaming pipelines that process user behavior data across our AWS infrastructure Design microservices and APIs that expose user segments and behavioral models to downstream surfaces, both classic UI and agentic Turn behavioral signals into segments
A BOUT TIDE At Tide, we help SMEs save time and money in the running of their businesses by not only offering business accounts and related banking services, but also a comprehensive set of highly usable and connected administrative solutions, from invoicing to accounting. Tide is transforming the small business banking market and now supports over 2 million members globally across the UK, India, Germany and France. Using advanced technology, all solutions are designed with SMEs in mind. With quick onboarding, low fees and innovative features, we thrive on making data driven decisions to serve our mission: to help SMEs save time and money so they can get back to doing what they love. Tide facts: Tide is available for UK, Indian, German and French SMEs Over 2 million members across UK and India Over $300 million raised in funding Over 2,800 Tideans globally Recognised with Great Place to Work certification three years in a row, and among India’s Top 50 Best Workplaces in Banking, Financial Services, and Insurance in 2026 We have offices in Central London, with a member support and technology centre in Sofia, Bulgaria, technology centres in Serbia, Romania, Lithuania and Hyderabad and offices in Gurugram, New Delhi, Berlin, Paris and Luxembourg ABOUT TIDE At Tide, we are building a business management platform designed to save small businesses time and money. We provide our members with business accounts and related banking services, but also a comprehensive set of connected administrative solutions from invoicing to accounting. Launched in 2017, Tide is now used by over 1 million small businesses across the world and is available to UK, Indian and German SMEs. Headquartered in central London, with offices in Sofia, Hyderabad, Delhi, Berlin and Belgrade, Tide employs over 2,000 employees. Tide is rapidly growing, expanding into new products and markets and always looking for passionate and driven people. Join us in our mission to empower small
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of a best-in-class family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from a diverse group of backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking a senior validation lead engineer to lead at-scale rack validation efforts for next-generation AI hyperscale systems. This role focuses on post-silicon system validation across the full lifecycle, ensuring functional, electrical, and thermal performance meets product objectives. You will own end-to-end blade and rack validation including planning, development, execution, and debug while collaborating across firmware, systems, and hardware teams. The Team The Rack Validation team is responsible for ensuring system readiness and quality at scale. The team works cross-functionally with firmware, silicon, and system engineering teams to validate complex AI compute platforms. Responsibilities and Duties Lead post-silicon validation of AI compute blades and racks including test planning, development, and automation. Drive provisioning and integration of system components (SoC FW, BMC, RMC, OS) for rack-level readiness. Own execution against program achievements and report validation progress and risks. Triage test failures, collect debug data, and collaborate on root cause analysis. Track
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your Opportunity At New Relic, we provide our customers with real-time insights, so they can innovate faster. Our software provides deep observability across the stack, enabling software teams to solve their customer’s problems, accelerate digital transformation, and make DevOps work. You will be at the heart of the teams supporting New Relic’s infrastructure and will work on a team that provides global service mesh and load balancing solutions. We provide these services on-premises, as well as using our multi-cloud infrastructure. We support each other to do our best work through positive communication and continuous improvement. What You’ll Do As a key member of our Infrastructure team, you will design and operate a scalable, resilient ingress data plane that directly impacts the value we provide to our customers. By ensuring the stability and performance of our global service mesh and load balancing solutions, you drive the foundational reliability that the entire New Relic organization depends on to deliver real-time insights. You will leverage advanced automation and infrastructure-as-code to accelerate development speed, allowing our engineering teams to ship safe, incremental changes across a massive fleet with confidence. Your work in evolving our DNS and CDN infrastructure is not just about maintenance; it is about creating a seamless, high-performance environment that enables innovation at scale. Through deep collaboration with Product, Design, and partner platform t
At Lyft, our purpose is to serve and connect. We aim to achieve this by cultivating a work environment where all team members belong and have the opportunity to thrive. The Identity & Integrity organization is looking for software engineers. We are growing our team with people who want to build, improve and incorporate technologies that make the lives of our community more enriched and safe. As an engineer at Lyft, you'll collaborate with teams like product, data science, analytics, and operations on code that empower us to iterate quickly, while focusing on delighting our passengers and drivers. As a Software Engineer in the Identity team, your primary responsibilities will encompass: Lyft’s Login & Signup Experiences: Design and enhancement of our Multi-factor authentication (MFA) experiences. Optimization of verification funnels, including: Phone-based two-factor authentication (2FA) Email verification Identity provider single sign-ons (SSO) Overseeing session management and device validation. Developing OAuth client provisioning and related tooling. Account Security: Act as the frontline defense against fraudsters and phishers aiming to exploit rider and driver accounts through product vulnerabilities. Implement strategies to counteract risks posed by social engineering tactics. Scaling Core Services: Lead the maintenance and optimization of core microservices under the Identity team's purview, essential to Lyft’s diverse service offerings. Manage organizational structures and multi-user management systems. (RBAC, family accounts, AuthZ) Design solutions for high-stakes challenges such as: Preventing unauthorized driving. Thwarting abuse related to recycled phones and their numbers. Striking a balance between stringent customer verification and ensuring minimal user friction. Responsibilities Write well-crafted, well-tested, readable, maintainable code Promote appropriate tech and engineering best practices Implement identity and security protocols to se
About the Team API Frontiers turns OpenAI’s frontier models into production APIs that developers can use to build reliable products and agents. We own the core path connecting models to developers through the Responses API, with a focus on safety, reliability, and speed. Working closely with Research, Safety, Codex, and other API teams, we bring new model capabilities into production and improve them through developer feedback. About the Role We are looking for a backend software engineer to build and operate the services behind the Responses API. You will shape API behavior, bring new capabilities from research into production, and make long-running agent workflows dependable and fast. The work combines distributed systems engineering with product judgment: designing useful developer interfaces, managing staged rollouts, and following production issues through to durable fixes. In this role, you will: Design, build, and operate APIs and backend services that bring frontier model capabilities to developers. Partner with Research, Safety, Codex, and API teams to define API behavior and deliver safe, staged launches. Build API capabilities for agent workflows, including task delegation, context sharing, and parallel execution. Strengthen long-running request reliability across timeouts, cancellation, streaming, and background execution. Improve request-processing performance and tail latency through profiling, efficient systems code, and persistent connections. Turn developer feedback and production failures into better observability, diagnostics, and lasting product improvements. Your background might look something like: 5+ years of experience building and operating backend services or developer-facing APIs in production. Strong software engineering fundamentals, with practical knowledge of distributed systems, concurrency, and asynchronous execution. Ability to diagnose production failures and performance bottlenecks using observability data and profiling. Product
Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team The Storage Abstractions (STAX) team builds the software layer through which Stripe services access stored data. We own the database SDKs that hundreds of Ruby and Java services use to read and write data safely, reliably, and efficiently without needing to understand the underlying database, routing, or operational complexity. Our work creates leverage across Stripe: by providing stable interfaces and safeguards at the storage layer, we help teams build products with confidence while enabling Stripe’s data architecture to evolve. Alongside engineers, we are designing for AI agents as customers of this foundation, making storage capabilities discoverable, interoperable, and safe to use across languages and storage backends. What you’ll do As a Staff Software Engineer on STAX, you will set technical direction and lead multi-year initiatives at the intersection of developer infrastructure, data access, and AI. You will work hands-on with engineers across Stripe to make storage access simpler, safer, and more interoperable, while helping product and infrastructure teams evolve their systems without fleet-wide migrations. You will help turn Stripe’s AI strategy into practical developer infrastructure by treating AI agents as customers of the storage layer. This is an opportunity to build frameworks and interfaces that make complex storage operations discoverable, interoperable, and safe for both human developers and agentic workflows, while rais
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity The Streaming Services team is responsible for routing and replicating New Relic's platform data through globally distributed cloud environments in a high-throughput, low-latency, and cost-aware streaming pipeline. Our streaming services process and route over a billion messages per minute across different continents, regions, and cloud providers, serving New Relic’s customer experience by making it easy for development teams to prioritize reliability. You will collaborate with this globally distributed team to develop expertise and best practices for New Relic development teams to operate highly reliable streaming services. What you'll do Own, build, maintain, and scale our streaming services and their support tools. Participate in an on-call rotation and bake stability into everything, continually seeking automation opportunities for built-in reliability. Participate in architectural definitions with a high degree of innovation and creativity. Own and improve your team processes. Develop automation and tooling to make our services more scalable and reliable. Use available innovation time to bring your creative ideas to life. This role requires Experience developing back-end services that use Flink, Kafka, or other streaming platforms. Large scale is a plus. Experience in writing software in Java, and you are not afraid of adapting, learning, and working with different languages and frameworks. Experience with distributed systems, concurrency, and scaling in
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity At New Relic, we provide our customers real-time insights, so they can innovate faster. Our software delivers insightful observability tools across different technologies and distributed systems, enabling software engineering teams to quickly identify, understand and tackle issues, analyze performance and get the most of their software and infrastructure. The Infrastructure product organization develops New Relic infrastructure instrumentation agents, next generation data processing and management services, vulnerability management, and security testing capabilities for on-prem and cloud customers. We work with data at a scale using a diverse tech stack (Go, Java, JavaScript, React GraphQL, Kubernetes, many public cloud web services, and more). As a senior backend engineer, you will help us build and extend next generation solutions such as a control plane for customers to manage their data pipelines at scale. New Relic is looking for engineers who are interested in building a brand-new observability experience. This high-impact engineering position is a phenomenal opportunity to own and build a set of next generation services and capabilities for the company. We are searching for a motivated engineer who is ready for a career-defining role in their next opportunity. We look forward to talking with you! What you'll do ● Design, Build, maintain, and scale back-end services and their support tools. ● Participate in architectural definitions with a high degr
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your Opportunity As a Senior Software Engineer within the Container Fabric (CF) organization, you will be a key driver in evolving New Relic’s global internal platform. We are looking for an operations-heavy engineer with 5–8 years of relevant experience who can leverage open-source and custom tooling to orchestrate and maintain large-scale Kubernetes environments. You will play a "Captain" role—leading critical deliverables and mentoring junior engineers while maintaining the reliability of our global fleet. What You'll Do Architectural Leadership: Drive the design and implementation of internal tools, specifically focusing on Kubernetes Operators and Controllers to automate resource management. Platform Orchestration: Lead complex, large-scale infrastructure shifts. Operational Excellence: Take ownership of incident response, author comprehensive retrospectives, and implement systemic hardening to prevent recurrence using advanced overcommit strategies. This Role Requires Experience: 5–8 years in a DevOps, Site Reliability, or Infrastructure Engineering role. Kubernetes Mastery: Deep internals knowledge of Kubernetes and hands-on experience writing custom operators. Tooling Proficiency: Strong experience building production-grade tools and services, specifically for infrastructure automation. Operations-Heavy Mindset: A proven track record of Day 1/Day 2 operations for a large-scale Kubernetes fleet, handling high-severity incidents, and improving SLA compliance through auto
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity At New Relic, we provide our customers real-time insights so they can innovate faster. Our software delivers insightful observability tools across different technologies and distributed systems, enabling software engineering teams to quickly identify, understand, and tackle issues, analyze performance, and get the most out of their software and infrastructure. This is a unique opportunity to shape the future of observability while pioneering the next generation of engineering: we are actively revolutionizing how we build software by embedding modern, agent-powered workflows directly into our daily development lifecycle. You’ll tackle complex distributed systems problems while helping drive internal innovation on the frontlines of AI-assisted engineering. About the team This position is for the Service Architecture Intelligence team. You will be building, improving, and maintaining a distributed service architecture capable of ingesting large volumes of data, analyzing spans and traces, and publishing them downstream so the UI can offer an exceptional experience to our customers. We work with data at scale, and our pipeline is built with a diverse tech stack (Java, Kafka, Redis, public cloud services, and more). You will work alongside a team of talented engineers solving complex distributed systems challenges. If you're passionate about performance and scale, and want to contribute to one of the largest and fastest-growing observability platforms while co-crea
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity We are looking for an experienced senior front-end engineer to join us in building The New Relic One! The platform will be used by your coworkers throughout the company and will help make their lives—and the lives of the customers they work with—easier and more enjoyable each and every day. Some of the problems we work on involve frontend APIs to interact with the internal framework, extensible architecture solutions, UI components, deal with modern frameworks and libraries such as ReactJS, and automation using tools such as Jenkins. What you'll do Select the best frameworks and tools needed to do the job quickly, while also optimizing for codebase stability, product stability, and target use case growth. Produce highly performant and flexible CLI using modern Typescript and software development techniques in collaboration with other team members. Find and identify opportunities to automate repetitive tasks that are made by developers on a daily basis. Collaborate with other UI engineers across the company to learn from others and to ensure you stay up to date on the company's UI best practices. Learn and improve your skills to continuously push us to deliver higher-quality tools and improve the UI team's workflows. This role requires Solid experience working with modern JavaScript environments. Familiarity with modern development and build tools such as eslint, docker, babel, and webpack. Experience with front-end JS testing tools and a comprehensive understa
We are a global team of innovators and pioneers dedicated to shaping the future of observability. At New Relic, we build an intelligent platform that empowers companies to thrive in an AI-first world by giving them unparalleled insight into their complex systems. As we continue to expand our global footprint, we're looking for passionate people to join our mission. If you're ready to help the world's best companies optimize their digital applications, we invite you to explore a career with us! Your opportunity The Telemetry Data Platform group at New Relic builds the foundation for all of our products: data ingest, storage, and query. As an engineer working on NRDB, you’ll be contributing directly to the proprietary telemetry database technology at the core of our business. We own our software from top to bottom and are directly responsible for its quality and reliability. Each member of the team shares our pager rotation and will occasionally be on-call to respond to system failures; so we prioritize work that keeps the lights on and the pager quiet, in addition to the work that powers all of our new products and streams of data. If the idea of working on systems that process millions of messages per second and handle exabytes of data excites you, then you may be an excellent fit! What you'll do Develop new features with a focus on optimizing performance and efficiency Collaborate with the team to implement scalable solutions and enhance application performance Identifying and acting on opportunities to improve the reliability of our services This role requires 2+ years of professional experience in distributed SaaS software development. Proficiency in Java programming, expertise with algorithms and data structures, and building high-throughput software following best-practices. Deeper understanding of distributed systems and their core challenges. Experience using the command line to manage, investigate, and fix things when they’re broken. Expe
Who we are About Stripe Stripe is a financial infrastructure platform for businesses. Millions of companies—from the world’s largest enterprises to the most ambitious startups—use Stripe to accept payments, grow their revenue, and accelerate new business opportunities. Our mission is to increase the GDP of the internet, and we have a staggering amount of work ahead. That means you have an unprecedented opportunity to put the global economy within everyone’s reach while doing the most important work of your career. About the team The Service Infrastructure team is responsible for empowering teams to build highly reliable and performant services, backing our payment systems, fraud detection, and a multitude of other products. As an infrastructure team, you will build and expand the service frameworks to make complex systems easy to use, resilient, and scalable. We build powerful interfaces for engineers that depend on these systems—deployment, load balancers, web framework, databases, Kafka, and Kubernetes—while keeping them highly available and performant. We're looking for engineering leaders who drive the technical vision of Stripe's service infrastructure platform for thousands of engineers to use and build on top of, and thrive in a highly autonomous environment with many moving pieces. What you’ll do You will join as a Technical Lead for one of the most impactful teams at Stripe. You will lead a team of engineers, collaborate with infrastructure and product engineering orgs, and advance service-oriented architecture (SOA) adoption at Stripe. By collaborating with the team's technical leaders, you will ensure the software your team builds meets the needs of Stripe and its customers. We are a highly effective team that consistently delivers high-impact results while genuinely caring for one another. We expect you to bring your curiosity and critical thinking skills to this challenging domain. We are looking for individuals with a strong background in designing and
Get new ai platform engineer jobs by email
Daily job updates · Unsubscribe anytime