Jobiba hiring network

Workload Porting And Performance Engineer Jobs

751 active opportunities · Updated for October 2026

Fresh results

15 shown

Explore current workload porting and performance engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.

About the role As a Forward Deployed Engineer (FDE) focused on Inference & Post-Training, you will be a hands-on technical partner to our most strategic customers — production AI teams looking to leverage high quality models and do inference at scale. For us, FDE is not a replacement for a Solutions Architect; you will partner with our SAs as a deep-domain specialist in inference optimization, fine-tuning pipelines, and production deployment. As key contributors to both the CX, Engineering, and Sales organizations, FDEs add tremendous value by ensuring we can meet the requirements of our most complex POCs, facilitate successful platform adoption, and guide tailored optimization efforts — directly impacting customer success, company growth, and the hardening of our core platform. Must be a permanent resident or citizen of Singapore. Responsibilities Inference Engine Optimization: Select, configure, and optimize inference engine based on hardware, model architecture, and workload profile Configuration & Performance Tuning: Develop configuration updates to win critical POCs, benchmarks, and optimize customer deployments; tune KV cache, apply speculative decoding, determine optimal tensor parallelism, and determine quantization strategy to hit throughput and latency targets. Post-Training & Fine-Tuning: Drive hands-on RL training runs and optimize system design; guide customers through LoRA, SFT, DPO, RLHF, and GRPO pipelines from experimentation through production. Strategic Customer Alignment: Act as the primary technical point of contact for aligned strategic accounts — monitoring and optimizing endpoint configurations, helping customers get the most out of the platform, and collaborating to ensure we hit critical milestones. Opinionated Onboarding: Establish direct alignment with strategic customers at onboarding; ensure the right inference and post-training configurations are in place from day one to improve time-to-value. Product Feedback Loop: D

pythonaigo
View job →
TA
15 days ago

About the Role At Together AI, you’ll build and operate one of the world’s largest GPU fleets used for frontier model training and inference. This isn’t a traditional infrastructure role—we’re looking for engineers who love building systems, automating everything, and solving problems at massive scale. If you enjoy writing software more than clicking dashboards, obsess over eliminating manual work, and want to build infrastructure that manages tens of thousands of GPUs autonomously, we’d love to talk. Responsibilities Design and build fleet automation systems that provision, validate, deploy, upgrade, repair, and retire GPU clusters with minimal human intervention. Build AI Infrastructure Agents that automate deployment, root-cause failures, incident triage, and autonomous remediation. Develop Fleet Intelligence platforms that continuously monitor hardware health, firmware, networking, storage, thermals, and workload performance to predict failures before they impact customers. Build software that maximizes GPU availability, utilization, performance, and reliability across thousands of accelerators. Create automated validation systems for GPUs, InfiniBand/RoCE fabrics, NVLink/NVSwitch, storage, and distributed AI workloads. Build internal platforms and developer tools that allow infrastructure to be managed through software—not manual operations. Continuously improve deployment velocity, reliability, and operational efficiency through automation. Partner closely with hardware, networking, platform, and AI teams to push the limits of AI infrastructure. Requirements 3+ years building distributed systems, infrastructure platforms, or large-scale backend software. Strong software engineering skills in Python, Go, or Rust . Experience building platforms, automation systems, or developer infrastructure. Experience with Linux, Kubernetes, Terraform, Ansible, or similar infrastructure technologies. Strong systems thinking with the ability to understand problems across hardw

pythonkuberneteslinux
View job →
O
OneTrust
📍 Atlanta• Full-time• From $139.7K/yr
15 days ago

Strength in Trust OneTrust’s mission is to enable innovation through the responsible use of data and AI. We believe that ensuring data is trusted shouldn’t slow teams down—it should accelerate what’s possible. This led us to develop the first technology platform for responsible data use in 2016. Today, with AI representing the latest and most impactful expansion of data yet, OneTrust is once again redefining what responsible innovation looks like. OneTrust, the AI‑Ready Governance Platform™, unifies regulatory intelligence, automation, and connected governance workflows so businesses can continue to move at the speed of AI while ensuring good governance to prevent data misuse at scale. Trusted by thousands of organizations worldwide, OneTrust is shaping the future where trusted data becomes a transformative force for business and society. The Challenge Due to increased customer demand for the AI-Ready Governance Platform, we’re looking to hire a lead corporate counsel to provide transactional support on commercial contracts. The successful candidate will be part of the Commercial & Product Support team. Your Mission Independently review and negotiate a high volume of complex SaaS contracts through the contract lifecycle to execution Handle ad-hoc legal requests for support from the sales function Understand key IP, commercial, and finance requirements Keep up-to-date on applicable laws, regulations, and industry guidance Drive compliance in day-to-day sales activities, including providing training Support updates to legal team templates & precedents Assist with wider company projects assigned to the Commercial & Product Support team You Are An excellent communicator and negotiator Self-motivated Able to work independently, multi-task and prioritize your workload Able to quickly shift priorities t

awsgitai
View job →
P
Point72
📍 Bengaluru• Full-time
15 days ago

JOB TITLE Cloud Compute Engineer A CAREER WITH POINT72’S TECHNOLOGY TEAM As Point72 reimagines the future of investing, our Technology team is constantly evolving our firm’s IT infrastructure and engineering capabilities, positioning us at the forefront of a rapidly evolving technology landscape. We’re a team of experts who experiment and work to discover new ways to harness open-source solutions, modern cloud architectures, and sophisticated Artificial Intelligence (AI) solutions, while embracing enterprise agile methodologies. Our commitment to building and innovating in the AI space provides the framework intended to drive smarter decision making and enhance how we build and operate our platforms and applications. As a member of Point72’s Technology team, we encourage and support your professional development from day one—helping you advance your technical skills, contribute innovative ideas, and satisfy your own intellectual curiosity—all while delivering real business impact for our multi-billion-dollar global business. WHAT YOU'LL DO Design, build, and operate Kubernetes clusters on Amazon EKS, including cluster lifecycle management, networking, autoscaling, and workload scheduling. Manage and optimize EC2-based compute infrastructure, including instance selection, placement strategies, capacity planning, and utilization analysis. Operate and improve ECS-based services where applicable, ensuring consistency across our container runtime environments. Develop and maintain Infrastructure as Code (IaC) using Terraform to provision and manage compute resources at scale. Collaborate with development and platform teams to define compute patterns, containerization standards, and deployment best practices. Monitor compute environments for availability, performance, and cost, driving continuous optimization across the fleet. Contribute to architectural decisions around workload placement, multi-tenancy, OS image and container li

pythonawsazure
View job →

Position Overview: The research function within Diligent Market Intelligence is primarily responsible for being subject matter experts on one of the research verticals. These verticals include compensation, governance, activism, voting and risk. Key responsibilities include managing the methodology for data collection; managing outsourcer workflow; quality assurance and liaising closely with Product and Engineering teams. The Senior Research Analyst serves as a subject-matter expert within a designated research vertical such as governance, compensation, activism, voting, or broader risk domains. This role conducts oversees third-party research workflows and owns the design and continuous refinement of data collection standards, quality controls, and methodology documentation. Findings are translated into clear, actionable insights for clients, while close collaboration with Research Managers, Product, Engineering, and Data Services ensures accurate outputs, regulatory alignment, process improvements, and strong operational reporting. As a senior team member, the analyst mentors colleagues, reviews deliverables for accuracy and consistency, and brings strong analytical judgment, curiosity, and leadership to advancing both the product and the overall research function. Key Responsibilities Develop, own, and continuously refine the data collection methodology for the assigned vertical, ensuring accuracy, consistency, and alignment with regulatory and product requirements. Proactively manage and communicate workload expectations to third-party research providers, ensuring timely, high-quality data delivery. Monitor and produce management reporting on daily workflows, identifying trends, bottlenecks, and opportunities for process optimization. Train and support third-party providers on methodology, quality standards, and process updates to ensure consistent execution across all research activities. Collaborate with Res

sqlawsgit
View job →
C-
CLEAR - Corporate
📍 New York• Full-time• $180K – $220K/yr
15 days ago

CLEAR is building THE secure identity company of the future. Our mission is to make experiences safer and easier—physically and digitally. With more than 43 million Members and a growing network of partners across the world, CLEAR's secure identity platform is transforming the way people live, work, and travel. Whether it’s at the airport, stadium, or throughout your everyday life, CLEAR unlocks the magic of frictionless experiences. We’re looking for a Data Engineer II to help us build the next generation of products which will go beyond just ID and enable our members to leverage the power of a networked digital identity. As a Data Engineer at CLEAR, you will participate in the design, implementation, testing, and deployment of applications to build and enhance our platform- one that interconnects dozens of attributes and qualifications while keeping member privacy and security at the core. A brief highlight of our tech stack: SQL / Python / Looker / Snowflake / dbt What you'll do: Build a scalable data system in which Analysts and Engineers can self-service changes in an automated, tested, secure, and high-quality manner Build processes supporting data transformation, data structures, metadata, dependency and workload management Develop and maintain data pipelines to collect, clean, and transform data (owning end to end data product from ingestion to visualization) Develop and implement data analytics models Partner with product and other stakeholders to uncover requirements, to innovate, and to solve complex problems Have a strong sense of ownership, responsible for architectural decision-making and striving for continuous improvement in technology and processes at CLEAR What you're great at: 4+ years of data engineering experience Working with cloud-based application development, and be fluent in at least a few of: Cloud services providers like AWS Data pipeline orchestration tools like Airflow, Dagster, Luigi, etc Big data tools like Spark, Ka

pythonsqlaws
View job →
CH
Cohere Health
📍 Hyderabad• Full-time
15 days ago

Opportunity Overview: We are seeking a Team leader, Intake to join our Intake Operations. In this role, as Team leader of Intake, you will work closely with the leadership team at Cohere and report to the Sr Director of India Operations. You will be responsible for providing supervision for the lead intake specialist and intake specialist team to meet or exceed operational objectives and metrics. You will leverage both your creative skills and communication skills to promote a high performing intake team. The Team leader of Intake will be highly organized in order to plan daily operational activities and provide oversight of the intake team. You will use your professionalism, personality, and communication skills to inspire the team to meet or exceed all performance standards. At a growing organization, this is a position that offers the ability to make a substantive mark on the company and its partners with exponential growth opportunities. What you’ll do: Gain a deep understanding of Cohere’s product and our health plan partners Provide daily operational direction to the intake staff. This includes interviewing new hires, training, coaching, mentoring, quality auditing, implementation and oversight of quality improvement plans identified based on trends and other process improvements Coordinates and provides day-to-day oversight of the intake staff Manage workload balancing needs of the intake team Assist in addressing case escalations and provider issues HR management to include performance evaluations, 1:1’s with the lead intake specialist and intake staff, payroll Other duties as assigned ISMS roles and responsibilities Good knowledge of Information practices. Assist the manager in all the information security activities implementation and maintenance process. Ensuring the team and imparted with Competence related to Information security Responsible for implementation of security policies and procedures and report any issues

agileaigo
View job →
DU
15 days ago

About the Team The Storage organization builds and operates the online stateful systems and abstractions that DoorDash Engineering depends on: reliable, efficient, secure, and easy to use. Within Storage, the Distributed Caching team owns every caching offering at DoorDash end to end, including ElastiCache (Redis/Valkey), Boulder (our KVRocks-based key-value store for high-QPS feature serving), Entity Cache (read Bill Shen’s engineering blog post, “ High-Performance Proxy Cache for DoorDash Services ”), and the Distributed Lock Service, plus the smart clients (asgard-redis, valkey-go) that sit in front of them. These systems back critical product surfaces across DoorDash, Wolt, and Deliveroo: the team runs roughly 400 ElastiCache clusters serving hundreds of millions of GET requests per second in aggregate, and Boulder, our offline-to-online feature store, serves billions of feature lookups per second at peak. About the Role The team owns provisioning of clusters and the smart clients that sit in front of them, baking in sensible defaults so that other engineering teams get a turnkey caching solution instead of having to run their own. You'll help drive Boulder's evolution to scale further, improve cost efficiency, enhance performance, and support real-time updates; re-platform the Distributed Lock Service onto a strongly consistent backend; and build the self-serve tooling and recommendation engine that let customers describe a workload (QPS, TTL, payload size, latency profile) and get the right backend without talking to a human. You'll go deep on cache invalidation, replication, sharding, compaction, and failover, while shipping the guardrails, automation, and observability that keep this scale operable by a small team. You must be located in San Francisco, Seattle, or the New York Metro Area for this hybrid position. You will report to the Engineering Manager on the Distributed Caching team within the Storage organization. You’re excited about this opportunity b

javaredisaws
View job →
G
Guidepoint
📍 Mumbai• Full-time
15 days ago

Overview The Insights team delivers high-quality research, market intelligence, and data-driven analysis to support clients’ strategic and investment decisions. Working across industries-including healthcare and medtech-the team combines primary and secondary research with deep analytical capabilities to generate actionable insights. Team members collaborate closely with global stakeholders to translate complex information into clear, impactful outputs. The Project Manager – Insights is responsible for overseeing the end-to-end delivery of research projects, ensuring timelines, quality standards, and client expectations are consistently met. This role coordinates across internal teams and stakeholders to manage multiple concurrent projects, streamline workflows, and drive execution. The ideal candidate combines strong organizational and communication skills with an ability to translate complex project requirements into efficient delivery. This is a hybrid position out of our Mumbai OR Pune office What You’ll Do : Manage end-to-end delivery of research and insights projects, ensuring timelines and quality standards are met Coordinate across analysts, researchers, and stakeholders to execute project plans Define project scope, objectives, and deliverables in collaboration with internal teams and clients Track project progress, manage risks, and proactively address bottlenecks Oversee preparation and delivery of client-ready reports, presentations, and outputs Ensure consistency, accuracy, and quality of insights and deliverables Optimize workflows, processes, and tools to improve efficiency and scalability Support resource allocation and prioritize workload across multiple concurrent projects Maintain clear and consistent communication with stakeholders on timelines and outcomes What You Have : Bachelor’s degree from an accredited college/university (e.g., Business, Economics, Life Sciences, or related field) +4 years of experience in project management, consulting, r

aiexcelproject management
View job →
DM
15 days ago

About the Team DoorDash Labs is a team within DoorDash building autonomous delivery robots and other autonomy solutions from the ground up for DoorDash's core delivery platform. If you have a passion for applying robotics solutions to a service loved by millions of people, then we want to talk to you! About the Role As Team Lead, Autonomy Tech Support, you will lead the day-to-day execution and development of the Mexico City Autonomy Tech Support team while maintaining a strong understanding of its technical workflows. You will set a high bar for troubleshooting, escalation quality, documentation, and operational readiness as the autonomous delivery fleet continues to scale. You will report into the Manager, Autonomy Tech Support on our Autonomy Tech Support team in our DoorDash Labs organization. Work model: 100% in-office in Coyoacán, Mexico City, Open for T4 levels, Across all LOBS. You’re excited about this opportunity because you will… Lead the day-to-day execution of the CDMX ATS team, including coverage, workload, priorities, and high-impact fleet issues. Coach and develop ATS Specialists through regular feedback, technical coaching, and hands-on support. Maintain a high bar across troubleshooting, escalation quality, Failure Mode execution, documentation, and cross-functional communication. Serve as the first leadership escalation point in CDMX for complex or high-urgency issues. Be a first line of defense for potential software regressions and emerging fleet-level trends, ensuring abnormal behavior is identified, validated, and escalated quickly. Lead AI adoption, integration, and enablement across the CDMX ATS team, ensuring AI capabilities are effectively incorporated into day-to-day workflows while identifying opportunities to improve troubleshooting, decision-making, and team efficiency. Partner with Operational Intelligence and Operational Reliability to ensure the team has the tools, systems, and processes needed to operate effectively and effic

gitlinuxai
View job →
R
Replit
📍 Foster City• Full-time• Remote
22 days ago

Replit is the agentic software creation platform that enables anyone to build applications using natural language. With millions of users worldwide, Replit is democratizing software development by removing traditional barriers to application creation. About the Team Product Platform builds and owns the shared foundations the rest of Replit is built on, spanning the full stack so every other team can ship features safely and quickly. Identity & Authorization defines how people, agents, sandboxes, and services prove who they are and what they can do. These systems protect critical product and service interactions across Replit's web product, Agent, enterprise controls, and internal services. Our work is high-leverage and horizontal: when identity and policy are clear, reliable, and easy to adopt, every other team can move faster without rebuilding security controls. We are a small, collaborative team that values curiosity and clear thinking over pedigree, and we work in the open by bringing each other the problem rather than just the request. We care more about how you reason and build than the route you took to get here. About the Role As a Software Engineer , you will design, build, and operate the identity and authorization systems that protect critical interactions on Replit, including Agent acting on behalf of a user or holding their own identity. The work is guided by a few simple questions: Can every protected request prove which workload made it, which principal it represents, and who is acting on that principal's behalf? Can product teams express policy once and trust the same decision across web, mobile, Agent, and internal services? Can enterprise administrators control who can access each workspace, app, connector, and Agent capability without navigating a permission maze as well as having a legible ledger of decisions? Can Agent act for a user across long-running and durable work without receiving broad or long-lived credentials? Are identity and auth

REMOTEtypescriptkubernetesrest
View job →
B
Baseten
📍 San Francisco• Full-time• Remote
22 days ago

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE As a Global Capacity Manager focused on TPUs at Baseten, you will lead the "engine room" for our non-NVIDIA accelerator fleet, architecting, securing, and optimizing the Google Cloud TPU (and broader emerging accelerator) capacity that powers our customers' AI workloads. You'll own the end-to-end journey of capacity management for this fleet, from securing large-scale TPU pod allocations to building the automation that ensures reliable uptime across multi-cloud environments. This role is a great fit for entrepreneurial engineers who want to bridge the gap between high-finance asset management and deep infrastructure engineering, with a specific focus on the TPU ecosystem. You will act as the fleet orchestrator for Google's TPU architecture, ensuring Baseten never experiences a capacity outage while maintaining elite unit economics as we diversify beyond NVIDIA. To be clear, this is a high-stakes engineering role. You will be hands-on with Kubernetes orchestration while also leading specialized pods focused on the latest generation of TPU hardware, like Google's Trillium (v6e) architecture, and partnering closely with the Model Performance (MP) team to ensure workloads are tuned for TPU-specific execution. EXAMPLE INITIATIVES The TPU Frontier: Architecting the infrastructure readiness and deployment strategy for Baseten's TPU clusters, including pod slicing and topology planning Global Workload Orchestration: Bui

REMOTEpythonawsazure
View job →
G
23 days ago

Location Details: Romania - Remote At GoDaddy the future of work looks different for each team. Some teams work in the office full-time; others have a hybrid arrangement (they work remotely some days and in the office some days) and some work entirely remotely. Remote: This is a remote position, so you’ll be working remotely from your home. You may occasionally visit a GoDaddy office to meet with your team for events or meetings. What you'll get to do... You will work closely with our Inbound Support teams to resolve any technical issues in a timely and professional manner. You will be friendly and communicative, able to resolve issues quickly and in line with our service level agreements. We will provide you with all the tools you need to do your job to the best of your ability, including full training on products and services Answering customers’ technical support issues through our CRM Working closely with your colleagues to identify and raise issues to improve customer satisfaction and experience Your experience should include... 1 + years experience in a Technical Support, Customer Support, Help Desk, or Web Development role. Basic understanding of Linux and/or Windows hosting environments. Knowledge of web hosting technologies, including LAMP and/or WAMP environments. Understanding of Domain Name System (DNS) concepts and troubleshooting. Basic knowledge of HTML and CSS, with familiarity with PHP considered an advantage. Strong written English and communication skills. Ability to explain technical concepts clearly to both technical and non-technical customers. Proven ability to work effectively within a team environment. Positive attitude and ability to remain composed when managing competing priorities and workload pressures. You might also have.. Experience with hosting control panels such as cPanel and Plesk. WordPress CMS experience. Experience working with AI tools and technologies, with hands-on experience using Claude Code considered a str

linuxaiphp
View job →

At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. We are looking for talented systems developers and researchers to join the Snowflake AI Research team and advance the state of the art in LLM inference systems and optimization . Our mission is to build the next generation of high-performance and intelligent inference systems . We optimize not only how fast and efficiently models run, but also how quickly inference systems can adapt to new models, architectures, hardware, and workloads. Our work spans the full inference stack—from distributed serving and runtime systems to GPU kernels and model-system co-design. We explore techniques such as adaptive parallelism, speculative and parallel decoding, disaggregated inference, scheduling and batching, KV-cache optimization, model swapping, quantization, and GPU kernel optimization to push the frontier of latency, throughput, scalability, and cost. Beyond optimizing individual models, we are building intelligent and adaptive inference systems that can automate performance optimization—rapidly profiling new models and workloads, identifying bottlenecks, selecting effective execution strategies, and adapting system configurations with minimal manual tuning. We embrace AI-native engineering , using AI not only as the workload we optimize, but also as a tool to accelerate system deve

machine learningaiswift
View job →
B
Baseten
📍 San Francisco• Full-time• Remote
24 days ago

ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma, and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to ship AI products. THE ROLE As a Global Capacity Lead at Baseten, you will lead the "engine room" of the company, architecting, securing, and optimizing the global GPU fleet that powers our customers' AI workloads. You’ll own the end-to-end journey of capacity management, from securing multi-million dollar GPU clusters to building the automation that ensures 99.9% uptime across multi-cloud environments. This role is a great fit for entrepreneurial engineers who want to bridge the gap between high-finance asset management and deep infrastructure engineering. You will act as the fleet orchestrator for the world's most advanced chips, ensuring Baseten never experiences a capacity outage while maintaining elite unit economics. To be clear, this is a high-stakes engineering role. You will be hands-on with Kubernetes orchestration while also leading specialized pods focused on the next generation of hardware, like NVIDIA’s Blackwell (B200) architecture. EXAMPLE INITIATIVES The B200 Frontier: Architecting the infrastructure readiness and deployment strategy for Baseten's first Blackwell GPU clusters. Global Workload Orchestration: Building "Multi-cloud Capacity Management" systems to move customer workloads seamlessly across regions to optimize cost and latency. Precision GPU Triage: Developing automated Go-based operators to identify, cordon, and repair unhealthy H100 nodes in under an hour. The Supply Chain of Intelligence: Partnering with lead

REMOTEpythonawsazure
View job →
🔔

Get new workload porting and performance engineer jobs by email

Daily job updates · Unsubscribe anytime