Jobiba hiring network

Workload Porting And Performance Engineer Jobs

751 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current workload porting and performance engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role Software Engineers at Palantir drive large-scale transformation through data, AI and world-leading infrastructure that supports mission-critical workloads. As a Software Engineer Intern, you’ll have an opportunity to grow more quickly than you ever envisioned as you contribute high-quality code directly to: • Rubix and Apollo, platforms deployed at the most important institutions across the public and private sectors. • Shaping Mission Manager, our new internal-infrastructure business line, used by advanced civil and defense agencies worldwide to power their infrastructure in highly sensitive environments • Building the core capabilities used by advanced civil and defense agencies worldwide to power their infrastructure • Providing the substrate on which Palantir deploys its other platforms, Foundry and Gotham, which power workflows for research scientists, aerospace engineers, intelligence analysts and economic forecasters. You’ll join our Production Infrastructure organization, made up of small teams of engineers working on: • Environment Platform: a Kubernetes-based PaaS spanning hundreds of production clusters • Apollo: secure, fleet-wide deployment and change-management for complex microservice suites • Signals: our full suite of observability and alerting tools Core Responsibilities As a Software Engineer Intern at Palantir, you’ll own every phase of the product lifecycle—from generating ideas and designing prototypes to executing features and shipping releases—while being paired with a dedicated mentor who champions your growth. You’ll work hand-in-hand with both technical and non-technical colleagues to uncover real customer problems and

typescriptjavareact
View job →
PE
1mo ago

A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role As a Senior Software Engineer on Network Infrastructure you will be joining a team whose mission is to make a highly dynamic and intricate Kubernetes networking tech stack a pleasure to interact with, in addition to being legible, robust and secure. The Network Infrastructure team owns the tech responsible for north-south and east-west traffic flows across 100s of zero-trust K8s clusters running variable workloads. The team solves the novel technical challenges emergent at the intersection of scale, usability and security: enabling dynamic control over ephemeral infrastructure (10s of thousands of firewall rules targeting 1000s of short-lived pods) across multiple compliance regimes, using CNCF tools like envoy, cilium with K8s controllers to operate them. In this role, you will be working in one of the most technically challenging spaces at Palantir.

kubernetesairust
View job →
PE
1mo ago

A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role As a Senior Software Engineer on Network Infrastructure you will be joining a team whose mission is to make a highly dynamic and intricate Kubernetes networking tech stack a pleasure to interact with, in addition to being legible, robust and secure. The Network Infrastructure team owns the tech responsible for north-south and east-west traffic flows across 100s of zero-trust K8s clusters running variable workloads. The team solves the novel technical challenges emergent at the intersection of scale, usability and security: enabling dynamic control over ephemeral infrastructure (10s of thousands of firewall rules targeting 1000s of short-lived pods) across multiple compliance regimes, using CNCF tools like envoy, cilium with K8s controllers to operate them. In this role, you will be working in one of the most technically challenging spaces at Palantir.

kubernetesairust
View job →
PE
1mo ago

A World-Changing Company Palantir builds the world’s leading software for data-driven decisions and operations. By bringing the right data to the people who need it, our platforms empower our partners to develop lifesaving drugs, forecast supply chain disruptions, locate missing children, and more. The Role As a Senior Software Engineer on Network Infrastructure you will be joining a team whose mission is to make a highly dynamic and intricate Kubernetes networking tech stack a pleasure to interact with, in addition to being legible, robust and secure. The Network Infrastructure team owns the tech responsible for north-south and east-west traffic flows across 100s of zero-trust K8s clusters running variable workloads. The team solves the novel technical challenges emergent at the intersection of scale, usability and security: enabling dynamic control over ephemeral infrastructure (10s of thousands of firewall rules targeting 1000s of short-lived pods) across multiple compliance regimes, using CNCF tools like envoy, cilium with K8s controllers to operate them. In this role, you will be working in one of the most technically challenging spaces at Palantir.

kubernetesairust
View job →
G
Godaddy
📍 United States• Full-time• From $128K/yr
1mo ago

Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team… GoDaddy's Global Storage Engineering team operates one of the largest Ceph environments in the industry, powering the object, block, and file storage platforms that underpin hosting, applications, internal infrastructure, and next-generation AI/HPC workloads. If you're passionate about distributed systems, large-scale storage architecture, and solving complex reliability challenges, you'll work on infrastructure that few engineers ever experience. At GoDaddy, Ceph isn't a side project — it's a critical platform. Our environment spans 80+ production clusters, 20,000+ OSDs, and approximately 300 PB of raw storage capacity, supporting tens of billions of objects across multiple continents. The scale demands deep technical expertise in storage architecture, automation, observability, and performance engineering. As a Senior Site Reliability Engineer, you'll be a key technical owner of the platform, responsible for maintaining reliability, driving operational excellence, and influencing the future evolution of our storage ecosystem. You'll tackle challenging production problems, develop automation that operates at massive scale, contribute to architectural decisions, and collaborate with some of the industry's most experienced Ceph engineers. This is an opportunity to have direct impact on a storage platform that serves millions of customers worldwide. What You'll Get to Do… Own the reliability, performance, scalability, and capacity of large-scale production Ceph environments supporting object, block, and file storage wor

pythonkuberneteslinux
View job →
G
Godaddy
📍 United States• Full-time• From $154K/yr
1mo ago

Location Details: At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely.​ This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. Join Our Team GoDaddy's Global Storage Engineering team operates one of the largest Ceph environments in the world, delivering the object, block, and file storage platforms that power GoDaddy's hosting infrastructure, internal services, OpenStack environments, and next-generation AI/HPC workloads. If you're passionate about distributed systems, storage architecture, and solving failure scenarios at massive scale, this is an opportunity to work on infrastructure few engineers will experience in their careers. Ceph is a strategic platform at GoDaddy — not an ancillary service. Our global footprint includes 80+ production clusters, 20,000+ OSDs, 1,830 storage nodes, 300 PB of raw capacity, and 69 billion objects spanning five datacenters across three continents. The platform supports RBD, RGW (S3/Swift), and CephFS workloads through more than 1,550 pools, 574,000 placement groups, and 900+ MDS daemons, creating engineering challenges that demand deep expertise in storage architecture, data durability, performance optimization, automation, and observability. As a Lead Senior Site Reliability Engineer, you'll serve as one of the principal technical leaders for GoDaddy's Ceph platform. You'll design the next generation of storage clusters, lead major platform upgrades, drive capacity and hardware strategy, and establish the standards that govern how the platform scales. You'll be the engineer the team turns to for the most complex s

pythonkubernetesai
View job →
A
Asana
📍 San Francisco• Full-time• $306K – $360K/yr
1mo ago

We are looking for a Director of Engineering to lead our AI Platform organization. This group builds the foundational systems powering every AI experience across Asana. In this role, you will lead four key teams through their engineering managers: Context (search, retrieval, and knowledge extraction across the Asana Work Graph), LLM Foundations (model serving, inference infrastructure, provider strategy, and evaluation systems), and AI Efficiency (our center of excellence for cost, quality, and performance standards across all AI workloads). Collaborating with engineering managers and senior technical leaders, you will drive the end-to-end strategy, execution, and architecture that define how humans and AI work together at Asana to build trusted, reliable, high-value product workflows for enterprise customers worldwide. Your mission is to make Asana’s AI platform the most reliable, economical, and performant foundation in the industry for agentic enterprise software, giving Asana the leverage to ship AI products faster than anyone else. This role is based in our San Francisco office with an office-centric hybrid schedule. The standard in-office days are Monday, Tuesday, and Thursday. Most Asanas have the option to work from home on Wednesdays. Working from home on Fridays depends on the type of work you do and the teams with which you partner. If you're interviewing for this role, your recruiter will share more about the in-office requirements. What you’ll achieve Drive Strategy & Execution Across the AI Teammates Pillar: Lead the multi-year vision and technical strategy for Asana’s AI platform, covering retrieval and agent context, model serving and inference, model portfolio strategy, and evaluation systems. Optimize AI Infrastructure Costs: Own cost-per-execution as a primary engineering metric, managing model selection, routing, open-weight versus frontier trade-offs, inference optimization, caching, and prompt efficiency to protect product margins at scale.

restaigo
View job →
A
Asana
📍 Vancouver• Full-time• $238K – $270K/yr
1mo ago

We're building AI Teammates: agents that work like actual users in Asana and integrated apps. They triage bugs, respond to requests, draft project briefs, conduct research, and handle complex knowledge work across your team's workflows. Unlike chatbots, Teammates are shared team resources that build memory and context across all executions. They get smarter as you and your colleagues work with them. Currently in beta with Fortune 500 customers, AI Teammates represents Asana's shift from tracking work to getting work done. We're looking for an Engineering Manager to lead the Agent Orchestration team — the team building the connections to other systems that allow Asana AI to serve the highest-value workloads. This means building the integration capabilities, agent skills, and vertical use cases that make AI Teammates extraordinarily useful across enterprise tools and workflows. You'll manage a senior team of six engineers (ICs up to L6) working at the intersection of systems integration and emerging AI capabilities to architect the foundation that lets AI agents operate seamlessly across customer environments while meeting enterprise requirements for reliability and compliance. This is a rare opportunity to be at the forefront of Agentic AI. You'll work directly with model partners, lead the core team shaping how enterprises collaborate with AI agents across their entire tool ecosystem, and drive the technical direction for one of the most impactful applied AI challenges in the industry. You'll also flex into org-level engineering leadership across the broader Asana AI group, contributing beyond your direct team. About Asana AI Asana AI is the company's number one priority. We're building the future of human/AI collaboration — going beyond chatbots to integrate AI into everyday workflows for some of the biggest companies on the planet. Asana's AI Teammates deliver a secure, multi-player, enterprise-grade agentic experience. They're transforming Asana from a place wher

restaigo
View job →
A
Asana
📍 Warsaw• Full-time• $372K – $432K/yr
1mo ago

We're looking for a Senior Platform Reliability Engineer who brings strong software engineering skills and a deep understanding of system behavior under load and stress. This role is a good fit for someone who wants to own reliability as a first-class concern – building the foundational systems that protect Asana's platform, not just responding when things go wrong. You'll build core platform systems like load shedding, rate limiting, circuit breakers, and traffic controls that protect Asana under real-world load. This is deep, cross-cutting work that shapes stability and performance of our entire infrastructure – and you'll partner closely with other platform teams to make reliability something that's built in, not bolted on. Our tech stack includes: AWS, Kubernetes (EKS), CloudFront, Istio, Cilium, MySQL (RDS), OpenSearch, DynamoDB, Redis, Terraform, Datadog, TypeScript, Scala, Go, and Python. (Yeah, we know this sounds like buzzword bingo – but we want this post to actually show up in your searches.) Why this role? Reliability as a first-class feature : You won't be patching things up after the fact. You'll build the systems that make Asana resilient by design. Foundational work : Load shedding, traffic management, ingress/egress – these are the building blocks that protect everything else. You'll own them. Strong collaboration, reasonable hours : You'll work closely with infrastructure teams in Warsaw and Reykjavik, making deep collaboration practical without constant timezone gymnastics. Room to grow : This is a new team, and you'll help shape what Platform Reliability Engineering looks like at Asana – whether that means leading projects, mentoring others, or defining our technical direction. In this role, success means shipping systems that other teams rely on by default – because they make the platform safer, not because they're mandatory. We're especially interested in people who think like backend engineers but obsess over failure modes, capacity plan

typescriptpythonsql
View job →
D
Datadog
📍 Massachusetts• Full-time• From $234K/yr
1mo ago

The ML Observability team builds cutting-edge tools to monitor, explain, and improve AI systems in production, particularly those leveraging Large Language Models (LLMs) and generative AI. We provide robust, scalable observability for AI workloads, including drift detection and model evaluation, and behavior tracing, enabling customers to ship AI with confidence. As a Staff Engineer, you’ll lead the development of new features and foundational capabilities within Datadog’s LLM Observability product. You will shape product direction, drive experimentation, and apply your deep understanding of both AI systems and software engineering to solve open-ended problems in the fast-moving AI landscape. Your work will directly impact how our customers monitor, troubleshoot, and optimize LLM-based applications in production. Join us in building the foundational tools that make AI systems observable, understandable, and reliable in the real world. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Drive design and implementation of LLM observability features. Ideate, prototype, and scale new product features to provide insights and drive improvements for generative AI systems Work cross-functionally with other eng teams, product, UX, and applied science to iterate fast and find product-market fit Develop and extend tools for tracing, evaluating, and debugging LLMs Influence architecture decisions and mentor engineers to build resilient, high-performance systems Stay close to customer pain points and use those insights to guide product and engineering priorities Stay current with industry trends and advancements in machine learning and observability, driving innovation within the team Who You Are: You have a BS/MS/PhD in a Computer Science, Engineering or r

machine learningaigo
View job →
D
Datadog
📍 Denver• Full-time• From $92K/yr
1mo ago

We are Datadog's in-house product experts. The Datadog Federal Support Engineering team is dedicated to serving as highly trusted technical advisors for our Public Sector customers, who operate within some of the most highly regulated and security-constrained environments. These customers include various government agencies and organizations with critical, sensitive missions. As a Federal Support Engineer 3, this role places you at the forefront of supporting these customers' mission-critical workloads. These complex workloads are often deployed across sophisticated hybrid and multi-cloud architectures, requiring deep expertise in cloud technologies, monitoring, and security best practices. Your primary responsibility is to ensure the complete success of these customers across their entire lifecycle with Datadog. Whether you’re looking to learn from the best or be the best, the Federal Support team is dedicated to furthering personal development and team success. At Datadog, we place value in our office culture - the relationships and collaboration it builds and the creativity it brings to the table. We operate as a hybrid workplace to ensure our Datadogs can create a work-life harmony that best fits them. What You’ll Do: Engage with public sector customers via multiple channels (ticketing system, live chat, calls, and screensharing tools) to identify and resolve technical support requests. Troubleshoot, investigate, and resolve complex technical issues in highly constrained environments across Datadog's 1000+ integrations, often with limited logs or sanitized data. Handle urgent escalation cases that may result in customer-facing troubleshooting calls, and internal or external incident management Become a subject matter expert in many Datadog product areas Partner with Product, Engineering, and Account teams to to validate bugs and advocate for customer-impacting improvements Provide mentorship to junior members of the team and serve

restmicroservicesai
View job →
M
Mongodb
📍 United States• Full-time• From $151K/yr
1mo ago

Join the MongoDB Server Query Optimization team, and help us build a world-class distributed open-source query optimizer. Our team plays a crucial role in the experience and performance of data processing. We are responsible for the MongoDB Query Language and the lifecycle of each query, through parsing, optimization and plan selection. We have a presence across the US and Europe including New York, Dublin, Seattle, Palo Alto, and Chicago. We support office-based and remote work and align projects with convenient work hours for each time zone. We have tons of interesting problems to solve with a direct impact on users for transactional, time-series, and analytical workloads. The team is endeavoring to systematically rewrite every major component of our optimization and execution systems. We need your help to design and build the heart of a distributed, flexible schema, document database. ​​This role can be based out of our US offices or remotely in the North America region. Candidate Profile 10+ years of experience in data management systems, distributed systems, or large-scale backend engineering Experience with building production-level code with a large user base, robust design structure and rigorous code quality Degree in Computer Science or similar field, or equivalent practical experience, with strong competencies in data structures, algorithms, and software design/architecture Experience with large code bases written in C++ or another systems programming language. You'll need to trace down defects, estimate work complexity, and design evolution and integration strategies as we rewrite different components of the system A strong foundation in core database internals is essential. While direct experience in query optimization is a massive bonus, it is not a prerequisite. We are also excited to meet candidates with strong backgrounds in compilers, language transpilers, or distributed storage systems Position Expectations Innovate in the area of flexible schema d

mongodbawsazure
View job →
M
Mongodb
📍 United States• Full-time• From $151K/yr
1mo ago

Join the MongoDB Server Query Execution team, and help us build a world-class distributed open-source database. Our team plays a crucial role in the performance and efficiency of MongoDB's data processing. We are responsible for building and improving the core execution engine that powers all queries, taking a logical query plan produced by the optimizer and turning it into reality. This includes developing the physical operators for data retrieval and manipulation, improving the runtime for complex analytical and transactional workloads, and owning critical components such as our new execution engine. In addition to the core server, we support the query execution needs of other major products like Atlas Streams, Atlas Search and Vector Search, and mongosync, making our work vital to the entire MongoDB ecosystem. You will be joining a globally distributed team with a significant presence in both North America and Europe. While this role is based in the NAMER region, you will regularly collaborate closely with colleagues across different time zones. We support both office-based work in our North America hubs like New York, as well as remote work. We have tons of interesting problems to solve with a direct impact on users for transactional, time-series, and analytical workloads. To meet the ever-increasing data demands of modern applications, we are actively evolving our query system; this includes strategically re-architecting and improving key components of our query execution engine. We need your help to design and build the core of a distributed, flexible schema document database. This role can be based out of one of our North America offices, such as NYC or Palo Alto, or remotely across North America. Candidate Profile 10+ years of hands-on, professional experience in query engine development or database internals Experience with building production-level code with a large user base, robust design structure and rigorous code quality Degree in Computer Science or

mongodbawsazure
View job →
M
Mongodb
📍 United States• Full-time• From $104K/yr
1mo ago

MongoDB is hiring a Staff Product Marketing Manager to build and own our go-to-market narrative for the Public Sector vertical, with a focus on Federal Government and the broader public sector market. This is a foundational hire for MongoDB’s Industry Verticals product marketing function: you will define how MongoDB’s unified data platform, spanning cloud, on-premises, and hybrid database deployments with integrated, production-ready AI capabilities, shows up for government buyers. You’ll turn a major compliance milestone into a durable competitive differentiator: developing the positioning, messaging, and sales-ready content that helps government agencies, systems integrators, and cloud/public-sector resellers understand why MongoDB is the right data platform for mission-critical, regulated workloads. You do not need prior government or public-sector work experience to succeed in this role — you need to be an excellent product marketer who can get fluent in a new domain quickly and partner closely with the compliance, product, and sales experts who already are. This role can be based in one of our MongoDB hub offices in the U.S. or remotely in the U.S. What you’ll do Own positioning and messaging for MongoDB’s Public Sector go-to-market, leading the federal GTM and launch related activities Translate MongoDB’s data platform capabilities — document database, search, vector search, stream processing, and integrated AI — into mission-relevant outcomes and value propositions for government buyers and the systems integrators who serve them Partner with Compliance, Security, Industry Solutions and Product teams to accurately represent related certification requirements in external-facing content, staying current as MongoDB pursues additional authorizations (e.g., DoD Impact Levels) Build the public sector sales enablement toolkit: battlecards, pitch decks, discovery guides, ROI/value models, and competitive intelligence tailored to federal buying processes and procuremen

mongodbawsazure
View job →
M
Mongodb
📍 Ireland• Full-time
1mo ago

We are seeking a Staff engineer to design, build, and operate the internal and external Observability stack for the MongoDB platform. Tens of thousands of customers depend on our Observability stack to monitor their database clusters and to generate actionable alerts to safeguard critical workloads. The Collections team is a newly formed team within MongoDB's Observability & Adoption Org focused on making telemetry onboarding and collection significantly easier across MongoDB. We own key parts of the observability collection stack, including onboarding experience, telemetry collection agents across the data and control planes, and ingestion services for metrics, logs, and traces that support both internal and customer observability in MongoDB, driving insights, recommendations, and alerting. Our mission is to reduce friction for teams implementing and iterating on Observability while partnering closely with development teams to instrument their services using shared best practices, helping define the conventions our telemetry should follow, and building collection and ingestion systems that are stable, performant, secure, well-documented, and self-service. We also work closely with the Data Pipeline and Storage & Query teams to help ensure MongoDB has a stable and performant observability stack end to end. This is an opportunity to join a team shaping how observability works across MongoDB and to have outsized impact on both the developer experience and the reliability of the platform underneath it. As MongoDB Atlas and its supporting infrastructure continue to experience rapid growth, the demand for high-cardinality observability data for internal and external use cases means we need to continually innovate and scale our systems to the next level. For example, MongoDB Observability systems need to handle 10’s of billions of metrics time series, all whilst processing petabytes of logs, traces, and events. Our stack includes VictoriaMetrics, Grafana, Sp

javamongodbaws
View job →
🔔

Get new workload porting and performance engineer jobs by email

Daily job updates · Unsubscribe anytime