About the Team ChatGPT relies on a large and growing GPU fleet to serve inference workloads reliably and efficiently. We develop the systems and tools that make it possible to introduce new models, manage production deployments, respond to operational issues, and use infrastructure effectively at scale. Our work spans distributed systems, platform engineering, infrastructure automation, and developer experience. We partner closely with research, infrastructure, and product teams to make model deployment more reliable, more efficient, and easier to manage. About the Role We are looking for a software engineer with experience building or operating large-scale production systems. You will design and develop systems that support the model lifecycle in production, including deployment orchestration, configuration management, operational automation, reliability, and capacity management. You will help transform complex operational processes into scalable platform capabilities that enable teams across OpenAI to deploy and manage models with greater confidence and less manual effort. This role is a good fit for engineers who enjoy solving complex operational problems and building software that makes production infrastructure easier to run at scale. In This Role, You Will Build and evolve the platform used to deploy, configure, and manage models across ChatGPT. Develop systems for deployment orchestration, model rollouts, operational visibility, and production readiness. Create abstractions and tooling that simplify complex infrastructure and improve the developer experience. Automate operational workflows, including incident detection, diagnosis, mitigation, and recovery. Improve the reliability, scalability, and efficiency of model deployments and the infrastructure that supports them. Build systems that support capacity planning, resource allocation, and infrastructure utilization. Partner with research, infrastructure, and product engineering teams to identify common chal
Jobs in United Kingdom
Infrastructure Security Engineer in United Kingdom
121 active opportunities · Updated October 2026
Showing
15 jobs
Explore current infrastructure security engineer jobs across United Kingdom. Filter by work mode, employment type, experience, department, date posted and distance.
A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role Forward Deployed Enablement Engineers are embedded with centralised customer success teams to maximise the outcomes of our deployed products and workflows across all commercial customers from small-scale start-ups to large enterprises. This role balances high-level support responsibilities with the development of innovative tooling and infrastructure to scale customer enablement effectively. You are a resourceful, gritty and adaptable problem solver who is able to work both collaboratively and independently to resolve difficult and nebulous technical issues, as well as work productively with external customers to debug and resolve their problems. Palantir’s Customer Success team helps our customers build on Palantir’s Foundry & AIP Platforms to drive the workflows that power their most important business outcomes. In this role, you’ll leverage your problem-solving abilities, creativity, and technical skills to support and guide customer development teams, ensuring they can effectively build and optimise their workflows. You’ll have the opportunity to gain rare insight into and contribute to some of the world’s most important industries and institutions. Every day at Palantir is different: we’re constantly evolving to better respond to customer needs, and you will have the opportunity to contribute your creativity and problem-solving to internal processes and tools that define how we deliver business value to the customer with increasing efficacy and efficiency. Every Palantirian is encouraged to play to their personal strengths, so there is no “one size fits all” approach to engineering at Palantir. The scope of this role is intentionally broad so tha
Join us in building the future of finance. Our mission is to democratize finance for all. An estimated $124 trillion of assets will be inherited by younger generations in the next two decades. The largest transfer of wealth in human history. If you’re ready to be at the epicenter of this historic cultural and financial shift, keep reading. About the Team: Bitstamp Exchange Platform Bitstamp made history in 2011 as the world’s first regulated crypto exchange. As a key part of the Robinhood family, the Exchange Platform team owns the full service lifecycle. We are the architects of a modern, high-velocity ecosystem that enables our global expansion, ensuring the world’s longest-running exchange remains unshakeable. The Role As a Senior Backend Engineer integrated into the Exchange Platform team, you will be a key driver of our modernization strategy. You will be an essential part of building the new generation infrastructure for low-latency services in the cloud, adopting cutting-edge technologies to drive our high-velocity ecosystem. This role is based in our London office(s), with in-person attendance expected at least 3 days per week. At Robinhood, we believe in the power of in-person work to accelerate progress, spark innovation, and strengthen community. Our office experience is intentional, energizing, and designed to fully support high-performing teams. Requires participation in an on-call rotation to support business needs. What You’ll Do Build Next-Gen Infrastructure: Architect and implement high-availability, low-latency cloud infrastructure that ensures performance and portability across our ecosystem. Modernize and Scale: Lead the transition of our services to modern, containerized environments, optimizing deployment and scaling workflows. Robinhood Ecosystem Integration: Manage and execute high-priority integration projects that align Bitstamp's backend with Robinhood's global infrastructure. Evolve the Stack: Identify and implement back
About the Team Training Runtime designs the core distributed runtime that powers everything from early research experiments to frontier-scale model runs. We work on building robust, scalable, high performance components to support our distributed training workloads. Our priorities are to maximize the productivity of our researchers and our hardware, with the goal of accelerating progress towards AGI. Within Training Runtime, the Process Management team develops the distributed OS responsible for launching, coordinating, and supervising the large numbers of processes that make up modern training workloads. Our runtime sits beneath training frameworks and on top of research infrastructure, ensuring jobs run reliably across massive clusters while maintaining performance, stability, and observability. Success for us is measured by both system reliability and researcher velocity - enabling ideas to scale from experiments to production training runs. About the Role As a Training Runtime: Process Management Engineer , you will work on the software that ties thousands of computers together and exposes them as a unified system. This system has to serve individual researchers running multiple parallel experiments, as well as our largest training runs spanning 100’s of thousands and even millions of machines and accelerators. This requires easy to use, introspectable systems that can promote a fast debugging and development cycle, as well as relentless optimization for scale while maintaining stability and performance throughout. You will work primarily in Rust , building high-performance asynchronous systems with a strong emphasis on performance, correctness, and scalability. Working at this scale and at the frontier of AI development poses novel challenges. Out-of-the-box approaches often don’t work. The problems you will be working on are highly ambiguous and require strong design judgment as well as proficient execution to advance the state of our infrastructure. We’re loo
About THG Ingenuity THG Ingenuity is a fully integrated digital commerce ecosystem, designed to power brands without limits. Our global end-to-end tech platform is comprised of three products: THG Commerce, THG Studios, THG Fulfilment. Each represents a single, unified solution, overcoming challenges and taking brands direct-to-consumer. Our client portfolio includes globally recognised brands such as Coca-Cola, Nestle, Elemis, Homebase, and Proctor & Gamble. The Role We are seeking a talented and commercially astute Data Scientist to join our dynamic team tackling Fraud, Payments and Finance. In this multi-focus role, you will be responsible for protecting our global e-commerce operations from fraud, optimising our payments performance and driving profitability and automation initiatives within Finance. Working with vast, complex datasets, the work will involve combatting online fraud using our award-winning THG Detect platform and driving payments optimisation by increasing authorisation rates, while also using the latest AI software and Google Cloud Platform (GCP) infrastructure to build complex data solutions and applications. This role requires a blend of strong technical skills, commercial acumen, and effective communication to translate complex data insights into tangible business value. Responsibilities: Optimise Fraud Detection: Continuously improve automated decision-making in our in-house THG Detect platform using machine learning and decision rules. Enhance Payment Performance: Lead data-driven initiatives to boost payment authorisation rates, optimise transaction routing, and reduce costs. Drive Profitability within Finance: Lead automation and profitability initiatives across different areas of Finance. Communicate Performance: Report on all aspects of fraud, payments and finance performance to stakeholders at all levels, including C-level management. Drive System Improvements: Collaborate with Fraud, Payments, F
Why Sony Interactive Entertainment? Sony Interactive Entertainment isn’t just the Best Place to Play — it’s also the Best Place to Work. Sony Interactive Entertainment (SIE) is the company behind the PlayStation brand. As a subsidiary of Sony Group Corporation, we’re part of a proud legacy of innovation and excellence. SIE is a dynamic technology company, delivering cutting-edge hardware and network services to more than 100 million people and an entertainment leader, home to some of the most beloved and recognizable intellectual properties (IP) in the world. Our role at SIE is to create and nurture the experiences under the PlayStation brand, a name synonymous with entertainment excellence and creativity. Join us and be a member of the team that tests the development tools used to create every PlayStation® game. This is a great opportunity to continue a career in Software Testing, working on industry leading development tools and being mentored by experts in their field. What you’ll be doing Developing tests for new applications, new features and regressions. Developing test strategies with application teams. Automating tests as part of our continuous integration systems. Investigating issues and creating reproducible test cases. Conducting manual, exploratory testing. Reporting test results clearly (including performance and coverage). Developing tooling to help reduce test time and ease test execution. Make suggestions and recommendations for improvements to help shape the finished product. Collaborate with global, cross-functional teams in a multicultural environment to develop organisation-wide test strategies and infrastructure. Mentoring peers, sharing best practices, and influencing quality engineering approaches. What we are looking for 5+ years experience in developing tests for software applications. Experience of automating tests. A good understanding of software testing techniques. Good C# programming skills. Understanding of C++. Experience in triaging
At Rockstar Games, we create world-class entertainment experiences. Become part of a team working on some of the most rewarding, large-scale creative projects to be found in any entertainment medium - all within an inclusive, highly-motivated environment where you can learn and collaborate with some of the most talented people in the industry. Rockstar Games is seeking a Senior Manager, Product Management to help shape the future of Rockstar’s Creator Platform ecosystem. This role focuses on building sustainable systems that empower creators to build experiences, attract players, and operate thriving communities. As a member of our team, you will need a critical and creative eye capable of putting forth innovative solutions to complex problems. Working with a wide variety of technical and non-technical partners, you will be tasked with cultivating and delivering a vision for the platforms. This is a full-time, permanent and in-office position based in Rockstar’s unique game development studio in the heart of London. WHAT WE DO The Rockstar Games Creator Platform Team builds and operates technology platforms that enable creators to develop their own game modes, experiences, and modifications and players to experience community-created content on fully customized servers. We create tools and services that empower creators to build, publish, operate, grow, and monetize their experiences. We work at the intersection of games, creator ecosystems, platform infrastructure, community, and innovation. RESPONSIBILITIES Drive product initiatives from concept through launch and iteration, including opportunity definition, prioritization, requirements, launch planning, and post-launch analysis. Define product strategy and roadmaps for key Creator Platform areas in partnership with leadership, engineering, developer relations, analytics, publishing, and operations. Use qualitative and quantitative inputs —
As a Senior Software Engineer on Coder’s Agentic Engineering team, you’ll build and evolve the systems behind our agentic development experience. You’ll work across the agent harness, integrations, and workflows that connect agents with real development environments. You’ll stay hands-on, solve complex technical problems, and work closely with Product, Design, and other engineers to ship reliable agentic experiences. What you’ll do here Design and build production systems in Go, with work across React and TypeScript where needed. Improve agent execution, tool use, context management, streaming, and long-running workflows. Extend our provider-agnostic architecture as models and capabilities change. Build reliable integrations between agents, workspaces, tools, and developer infrastructure. Own projects from implementation through rollout and iteration. Contribute to design reviews, code reviews, and technical discussions. Partner with Product and Design to turn agent capabilities into useful developer experiences. Improve the reliability, performance, and operability of agentic systems. What we’re looking for Strong experience building and operating production software systems. Hands-on experience with Go. Experience with React and TypeScript. Experience building systems around LLMs or agentic workflows. Familiarity with model APIs, tool calling, context management, or agent loops. Good understanding of distributed systems and production reliability. Working knowledge of AWS. Strong problem-solving skills and comfort working through technical ambiguity. Someone who contributes beyond their own code through reviews, collaboration, and knowledge sharing. Bonus tacos if you have Experience building coding agents, developer tools, or cloud development environments. Experience with MCP, agent tools, or multi-agent systems. Experience with remote execution, sandboxing, or isolated compute. Experience building integrations across multiple model providers. Experience with AW
NVIDIA is well positioned as the 'AI Computing Company', our GPUs being the brains that power modern Deep Learning software frameworks, accelerated analytics, modern data centers, and driving autonomous vehicles. We are looking for a Senior Software QA Test Development Engineer to join in the mission of crafting a distributed technology for all NVIDIA teams that remotely manage 10s of 1000s of resources in a simple and controlled fashion, allowing engineers to focus on engineering and automation, rather than being burdened by manual operational tasks. SWQA test developer engineers at NVIDIA are responsible for creating test plans, execution, and reporting, as well as developing scripts for test automation, designing and developing tools for the QA team, and developing integration tests for validation. As a test developer, you must identify weak spots and constantly design better and more creative test plans to break software and identify potential issues. You will have a huge impact on the quality of NVIDIA's products. The ideal candidate must have strong programming skills and hands-on experience using AI development tools to improve quality and productivity across the end-to-end QA workflow. This includes leveraging AI assistants for test automation, code generation, debugging, and enhancing testing efficiency. During the interview process, we will assess your ability to effectively use AI development tools and evaluate your programming capabilities to ensure you can deliver high-quality solutions. What you’ll be doing: Architect, implement, and evolve scalable agentic end to end SWQA workflow, automated test frameworks, infrastructure, and tooling for complex software products. Define test strategy and quality gates across functional, integration, regression, reliability, and release-validation workflows. Build and maintain high-value automated coverage for Linux-based, co
As a Staff Software Engineer on Coder’s Agentic Engineering team, you’ll shape the systems behind our agentic development experience. You’ll work across the agent harness, integrations, and workflows that connect agents with real development environments. You’ll stay hands-on while setting the team's technical direction. You’ll lead complex work, make sound architectural decisions, and help other engineers do their best work. What you’ll do here Set technical direction across Coder’s agent harness, integrations, and workflows. Design and build production systems in Go, with work across React and TypeScript where needed. Evolve agent execution, tool use, context management, streaming, and long-running workflows. Extend our provider-agnostic architecture as models and capabilities change. Lead complex projects from early ambiguity through production. Raise the engineering bar through design reviews, code reviews, and technical mentorship. Partner with Product and Design on clear, useful agent experiences. Improve the reliability, performance, and operability of agentic systems. What we’re looking for Deep experience building and operating production software systems. Strong hands-on experience with Go. Experience with React and TypeScript. Hands-on experience building systems around LLMs and agentic workflows. Experience with model APIs, tool calling, context management, or agent loops. Strong distributed systems knowledge. Working knowledge of AWS. A track record of setting technical direction without formal authority. Strong architectural judgment and comfort working through ambiguity. Someone who makes the engineers around them better. Our tech stack Backend: Go, Postgres Frontend: TypeScript, React Infrastructure: AWS, Kubernetes Observability: Prometheus, Grafana CI/CD: GitHub Actions Bonus tacos if you have (Tacos? If you need an ice-breaker, ask how we say thanks by giving tacos!) Experience building coding agents, developer tools, or cloud development environm
From $100K/yr
Our Key Accounts Executive will target and close new business within the largest, most strategic prospects in key, high potential companies. In this role you’ll be focused on understanding and uncovering the pain points these companies face as they operate in or migrate to a cloud environment at scale as well as delivering the appropriate Datadog solution. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Develop and execute an outbound prospecting strategy tailored to specific Fortune 100 accounts Drive a strategic, multi-threaded sales motion spanning multiple stakeholders and product suites Cross-sell and navigate throughout complex accounts Create, own, and grow your own accounts, demonstrating the value of the Datadog platform Develop a deep comprehension of customer's business Work cross-functionally with marketing, and solutions engineering to drive coordinated efforts that support the outbound prospecting strategy Negotiate favorable pricing and business terms with large commercial enterprises by selling value and ROI Demonstrate resourcefulness when faced with challenges that defy easy solution Have intuitive sense of necessary steps to close business and gain customer validation Identify robust set of business drivers behind all opportunities Ensure high forecasting accuracy and consistency Who You Are: Someone with 5+ years Enterprise Sales experience selling into Fortune 100 companies with the ability to win new logos Driven and have met/exceeded direct sales goals of 1M+ and operated with an average deal size of $100k+ Able to demonstrate methodology to prospect and build pipeline on your own Experienced in working for an innovative tech company (SaaS, IT infrastructure or similar preferred) Self-starter mindset and resourceful by nature
Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100. Founded in 2017, the company is headquartered in London, with offices and teams across Europe and the US. As AI continues to shape the way we live and work, Synthesia develops products to enhance visual communication and enterprise skill development, helping people work better and stay at the center of successful organizations. Following our recent Series E funding round, where we raised $200 million, our valuation stands at $4 billion. Our total funding exceeds $530 million from premier investors including Accel, NVentures (Nvidia's VC arm), Kleiner Perkins, GV, and Evantic Capital, alongside the founders and operators of Stripe, Datadog, Miro, and Webflow. We’re looking for a Principal Engineer to join the ML Platform team at Synthesia. Our team builds and operates the systems that allow researchers and product teams to train, serve, and deploy generative models reliably and efficiently . This includes research infrastructure, production serving systems, internal tooling, and the platform interfaces that connect them. A growing part of our mission is making these systems more automation-friendly and agent-oriented , so that workflows can increasingly be operated through reliable tooling rather than manual effort. We’re looking for a strong generalist with a systems mindset: someone who is comfortable working across infrastructure, backend systems, and tooling, and who has seen ML systems in practice. this is not a pure ML Engineer role. We’re especially interested in people who think deeply about reliability, scalability, performance, and resource efficiency in complex production environments. This is a hands-on IC role with significant ownership. You’ll help shape how our ML platform evolves as we scale the number of models, workloads, tools and teams relying on it. What you’ll do Design and improve the platform systems that support model training, evaluation, an
Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100. Founded in 2017, the company is headquartered in London, with offices and teams across Europe and the US. As AI continues to shape the way we live and work, Synthesia develops products to enhance visual communication and enterprise skill development, helping people work better and stay at the center of successful organizations. Following our recent Series E funding round, where we raised $200 million, our valuation stands at $4 billion. Our total funding exceeds $530 million from premier investors including Accel, NVentures (Nvidia's VC arm), Kleiner Perkins, GV, and Evantic Capital, alongside the founders and operators of Stripe, Datadog, Miro, and Webflow. The opportunity At Synthesia we really care about video generation, especially about human centric avatar video generation. This led us to release models such as EXPRESS-Video , and soon our latest video model - these are the best avatar video models in the world, and we are committed to continuing and double down our efforts in leading that area. Our goal is to get to human centric video models that can generate arbitrary long videos at high resolution with arbitrary actions and events. That means continuously training large generative video models from scratch with the proprietary data pipelines and compute infrastructure to support it at scale. We are looking for a technical leader who owns the full stack end-to-end, someone who bridges pre-training and post-training, sets long-term direction alongside research leadership, and is personally present at the hardest parts of the work. If building foundation model capability from the ground up at a company genuinely committed to leading the field sounds like the right next challenge, this role was written for you. About the role Synthesia's video generation capability is core to everything we ship. It involves roughly 15 people working daily across pre-training a
Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100. Founded in 2017, the company is headquartered in London, with offices and teams across Europe and the US. As AI continues to shape the way we live and work, Synthesia develops products to enhance visual communication and enterprise skill development, helping people work better and stay at the center of successful organizations. Following our recent Series E funding round, where we raised $200 million, our valuation stands at $4 billion. Our total funding exceeds $530 million from premier investors including Accel, NVentures (Nvidia's VC arm), Kleiner Perkins, GV, and Evantic Capital, alongside the founders and operators of Stripe, Datadog, Miro, and Webflow. About the role The Data team manages the complete lifecycle of data for researchers - from sourcing and large-scale processing to delivering datasets that power our models. Data sits at the heart of our Research efforts and enables all other teams. As part of the Data team, you’ll work with over a million hours of video and audio data. This role exists at the intersection of applied research, data engineering, and ML infrastructure rather than being a traditional research position . You’ll build the world’s best human-centric data lake by collaborating closely with our model training teams. By understanding their requirements, you’ll extract new features and annotations that elevate our datasets. You should be passionate about enhancing model performance through high-quality, accurate datasets. Our infrastructure and pipelines are in great shape, and this role provides room to not only enhance them but also influence the team’s longer-term strategy. What we're looking for: A strong background in data-centric, applied Machine Learning, with hands-on experience improving model performance through data quality, curation, labeling, and evaluation rather than model architecture alone Experience working on the data la
Synthesia is the world’s leading AI video platform for business, used by over 90% of the Fortune 100. Founded in 2017, the company is headquartered in London, with offices and teams across Europe and the US. As AI continues to shape the way we live and work, Synthesia develops products to enhance visual communication and enterprise skill development, helping people work better and stay at the center of successful organizations. Following our recent Series E funding round, where we raised $200 million, our valuation stands at $4 billion. Our total funding exceeds $530 million from premier investors including Accel, NVentures (Nvidia's VC arm), Kleiner Perkins, GV, and Evantic Capital, alongside the founders and operators of Stripe, Datadog, Miro, and Webflow. We’re looking for an Engineer to join the ML Platform team at Synthesia. Our team builds and operates the systems that allow researchers and product teams to train, serve, and deploy generative models reliably and efficiently . This includes research infrastructure, production serving systems, internal tooling, and the platform interfaces that connect them. A growing part of our mission is making these systems more automation-friendly and agent-oriented , so that workflows can increasingly be operated through reliable tooling rather than manual effort. We’re looking for a strong generalist with a systems mindset: someone who is comfortable working across infrastructure, backend systems, and tooling, and who has seen ML systems in practice. this is not a pure ML Engineer role. We’re especially interested in people who think deeply about reliability, scalability, performance, and resource efficiency in complex production environments. This is a hands-on IC role with significant ownership. You’ll help shape how our ML platform evolves as we scale the number of models, workloads, tools and teams relying on it. What you’ll do Design and improve the platform systems that support model training, evaluation, and product
Other cities to consider
More places hiring for this role
Get new infrastructure security engineer jobs in United Kingdom by email
Daily job updates · Unsubscribe anytime