Citibank, N.A. seeks an Applications Development Tech Lead Analyst for its Tampa, Florida location. Duties: Responsible for source code review to ensure it satisfies good coding practice. Coordinate all communication between tech team and Ab Initio vendor, including scheduling and chairing weekly meetings. Responsible for reconfiguration and reengineering to optimize performance. Design, develop and modify software systems, using scientific analysis and mathematical models to predict and measure outcomes and consequences of design. Prepare reports or correspondence concerning project specifications, activities, or status. Confer with systems analysts, engineers, programmers and others to design systems and to obtain information on project limitations and capabilities, performance requirements and interfaces. Manage the hardware infrastructure and monitor the performance metrices E2E. Identify areas of growth need and proposal scale up activities. Execute on scale up activities by submitting procurement requests, manage regular calls with the deployment team and plan go live/switch activity. Develop shell scripting code base to support the job executions, interactions between systems. A telecommuting/hybrid work schedule may be permitted within a commutable distance from the worksite, in accordance with Citi policies and protocols. Requirements: Requires a Bachelor’s degree, or foreign equivalent in Information Technology, Engineering (any) or related field and 6 years of progressively responsible, post-baccalaureate experience as a Software Engineer, Associate Director – Data Engineering, Senior Consultant, Data Specialist, Assistant Systems Engineer or related position involving gathering new business requirements, architecture, analysis, estimation, design, implementation, and leading a team of software developers and software development. 6 years of experience must include: Experience with AbInitio Data Processing (SME), Data Modeling
Jobs in United States
Lead Software Engineer Data Infrastructure Specialist Manager Manager in United States
15 active opportunities · Updated September 2026
Showing
15 jobs
Explore current lead software engineer data infrastructure specialist manager manager jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
Hiring demand
53/100
steady · 522 related jobs
Hiring trend
-72.4%
Job postings compared with the previous 30 days
Remote options
14.9%
Share of matching jobs listed as remote
Typical salary
$209.3K – $209.3K/yr
Based on 182 salary observations
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of a best-in-class family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from a diverse group of backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking a senior validation lead engineer to lead at-scale rack validation efforts for next-generation AI hyperscale systems. This role focuses on post-silicon system validation across the full lifecycle, ensuring functional, electrical, and thermal performance meets product objectives. You will own end-to-end blade and rack validation including planning, development, execution, and debug while collaborating across firmware, systems, and hardware teams. The Team The Rack Validation team is responsible for ensuring system readiness and quality at scale. The team works cross-functionally with firmware, silicon, and system engineering teams to validate complex AI compute platforms. Responsibilities and Duties Lead post-silicon validation of AI compute blades and racks including test planning, development, and automation. Drive provisioning and integration of system components (SoC FW, BMC, RMC, OS) for rack-level readiness. Own execution against program achievements and report validation progress and risks. Triage test failures, collect debug data, and collaborate on root cause analysis. Track
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore fosters continuous learning and innovation. Job Summary Reporting into the Systems Engineering organisation, the Distinguished Engineer, End-to-End Security Architect will define and lead the security architecture for Graphcore’s inference service platform. This role is responsible for establishing a comprehensive security strategy spanning platform, infrastructure, networking, service operations, customer assurance, and compliance readiness. Working across multiple engineering and operational functions, the successful candidate will provide technical leadership, drive security requirements, and ensure the platform delivers robust protection, resilience, and trust for customers. The Team You will work closely with teams across security architecture, infrastructure engineering, networking, site reliability engineering, platform software, firmware, data centre operations, compliance, legal, customer engineering, and customer security. The team collaborates across the business to deliver secure, reliable, and scalable AI infrastructure and services while supporting customer assurance, regulatory requirements, and operational excellence. Responsibilities and Duties Own the end-to-end security a
About us Graphcore is one of the world’s leading innovators in Artificial Intelligence compute. It is developing hardware, software and systems infrastructure that will unlock the next generation of AI breakthroughs and power the widespread adoption of AI solutions across every industry. As part of the SoftBank Group, Graphcore is a member of an elite family of companies responsible for some of the world’s most transformative technologies. Together, they share a bold vision: to enable Artificial Super Intelligence and ensure its benefits are accessible to everyone. Graphcore’s teams are drawn from diverse backgrounds and bring a broad range of skills and perspectives. A melting pot of AI research specialists, silicon designers, software engineers and systems architects, Graphcore enjoys a culture of continuous learning and constant innovation. Job Summary We are seeking a Senior Principal Network Engineer to help design, deploy, and optimize next‑generation AI data center networks. AI training and inference workloads require extremely high bandwidth, deterministic low latency, and zero‑packet‑loss networking environments. In this role, you will partner closely with the Network Architecture Lead to design and scale high‑performance computing (HPC) network fabrics supporting GPU clusters. You will work across hardware, networking, and AI application layers to ensure Graphcore’s large‑scale AI infrastructure operates at peak performance. The ideal candidate brings deep experience operating hyperscale or HPC data center networks and has expertise in high‑speed Ethernet fabrics, RDMA technologies, advanced automation, and telemetry systems. The Team The Data Center Network Engineering team designs and operates the high‑performance network fabrics that power Graphcore’s AI compute platforms. The team collaborates closely with hardware engineering, AI researchers, and infrastructure teams to build scalable networking environments optimized for distributed training and infe
From $295.3K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Data Engineering team at Roblox plays a crucial role in enabling the company's success by developing and maintaining highly leveraged Core Data Sets, frameworks, and tooling to support the growing demand for analytics. As a Principal Data Engineer, you will work to define the data ontology for all of Roblox, establish best practices and standards for data operations and lifecycle management, design and build analytics tooling and frameworks, and influence event instrumentation. Additionally, this role is highly cross-functional, requiring close collaboration with Data Science, Experimentation, and Machine Learning teams to understand customer requirements and analytics applications, as well as with Data Infrastructure and Storage teams to develop integrated solutions. Join us and be a part of a dynamic team driving innovation and growth at Roblox. You Will: Partner with Data Science, Data Platform, Product, and Engineering to collect requirements to define the data ontology for all of Roblox Lead and mentor a growing team of Data Engineers to support Roblox's ever-evolving data needs Design, build, and maintain efficient and reliable batch and streaming data pipelines to mod
From $243.3K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. The Data Engineering team at Roblox plays a fundamental role in enabling the company's success by developing and maintaining highly leveraged Core Data Sets, frameworks, and tooling to support the growing demand for analytics. As a Senior Data Engineer, you will work to define the data ontology for all of Roblox, establish standard methodologies for data operations and lifecycle management, design and build analytics tooling and frameworks, and influence event instrumentation. Additionally, this role is highly multi-functional, requiring close collaboration with Data Science, Experimentation, and Machine Learning teams to understand customer requirements and analytics applications, as well as with Data Infrastructure and Storage teams to develop integrated solutions. Join us and be a part of a dynamic team driving innovation and growth at Roblox. This role will report to our Engineering Manager on the Data Engineering team. You Will: Partner with Data Science, Product, and Engineering to collect requirements to define the data ontology for all of Roblox Lead and mentor a growing team of Data Engineers to support Roblox's ever-evolving data needs Design, build, and maintain efficient an
From $295.3K/yr
Every day, tens of millions of people come to Roblox to explore, create, play, learn, and connect with friends in 3D immersive digital experiences– all created by our global community of developers and creators. At Roblox, we’re building the tools and platform that empower our community to bring any experience that they can imagine to life. Our vision is to reimagine the way people come together, from anywhere in the world, and on any device. We’re on a mission to connect a billion people with optimism and civility, and looking for amazing talent to help us get there. A career at Roblox means you’ll be working to shape the future of human interaction, solving unique technical challenges at scale, and helping to create safer, more civil shared experiences for everyone. Roblox's data infrastructure processes petabytes of data daily, powering analytics, ML, and product decisions for a platform serving 200M+ daily active users. As a Principal Software Engineer in our Data Infra org, you will be the primary technical leader driving the strategic vision, long-term architecture, and massive scalability of our distributed data platforms that power Roblox. You will own and drive the next-generation architecture of our core platforms, which span Kafka, Flink, Spark, Trino, Druid, Airflow and Data Catalog. This role operates under high ambiguity, demanding unparalleled ownership to redefine the limits of infrastructure handling exabyte-scale workloads, and providing a unique opportunity to lead the future evolution of our global data ecosystem. You Will: Define Multi-Year Technical Strategy: Own and drive the end-to-end architectural vision for Roblox's core data platforms spanning Kafka, Flink, Spark, Trino, Druid, Airflow, and Data Catalog systems. Turn multi-year company strategies into concrete, production-grade infrastructure blueprints. Lead Cross-Functional Alignment: Partner closely with executive leadership, platform governance, data science, and product e
From $177.2K/yr
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . We’re looking for a Staff Software Engineer to help build the next generation of Pinterest’s big data storage platform. You’ll work on some of the most exciting big data open source technologies — especially Apache Iceberg — at exabyte scale to power the data infrastructure that helps Pinners discover and do what they love. As a Staff Software Engineer, you’ll serve as a technical leader and hands-on contributor, designing and building highly scalable storage systems for Pinterest’s data lake. You’ll partner closely with teams across data, ML/AI, analytics, and infrastructure to evolve our storage and metadata management capabilities, enabling efficient, reliable, and governed access to data at massive scale. What you’ll do: Design, implement, and optimize Pinterest’s exabyte-scale data lake storage platform. Lead complex technical projects and
From $248K/yr
Airbnb was born in 2007 when two hosts welcomed three guests to their San Francisco home, and has since grown to over 5 million hosts who have welcomed over 2 billion guest arrivals in almost every country across the globe. Every day, hosts offer unique stays and experiences that make it possible for guests to connect with communities in a more authentic way. The Community You Will Join: The Data Stewardship team is a group of passionate data practitioners with a diverse background in analytics, data modeling, governance, compliance, and scaled data quality. We are responsible for ensuring Airbnb is meeting its compliance obligations across our data ecosystem and ensuring data consumers are able to easily identify the best data for their needs. We support the pipelines, programs, and policy-bodies that make this possible. You’ll be part of the overall Data Infrastructure organization that is responsible for online and offline data infrastructure across the company, and the components that transition data between these environments. The Difference You Will Make Set the North Star: you will define the multi-year vision for Data Governance and Data Quality that scales with our global business and evolving AI landscape. Organizational Influence: Act as a primary consultant for executive leadership on data governance, ensuring that compliance and stewardship are integrated into the data product lifecycle from day one. A Typical Day Architect Ecosystems: Lead the design of overarching data architectures that don't just solve today’s batch needs but anticipate future real-time and AI-driven requirements. Scale Best Practices: Rather than just "ensuring quality," you will define the best practices, tooling, and culture that enable data organizations to maintain high-quality data autonomously. Policy Stewardship: Lead cross-functional task forces (Infosec, Legal, Privacy) to navigate complex regulatory landscapes (like GDPR or AI Act) and translate them int
$153.1K – $293.8K/yr · Jobiba est.
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. The Snowhouse Foundation team builds our globally distributed data warehouse. We manage a vast array of petabyte scale data sets that are continuously ingested, processed and replicated from across all Snowflake environments and external data sources. Snowhouse powers all of Snowflake’s core business, engineering and data science needs and provides customers with full visibility into their account activities, usage, and resource consumption from all their global environments. The team is investing in multiple critical areas, including a pipeline authoring platform, high performance/high efficiency data export, ingestion and data layout. Our team is also responsible for a fundamental product for Snowflake’s customers: the Snowflake system database/application that provides customers with all usage insights they need to reason about their global Snowflake footprint as well as 1st party business logic such as ML powered functions and Budgeting applications. AS A PRINCIPAL SOFTWARE ENGINEER IN SNOWHOUSE FOUNDATION, YOU WILL: Design and implement innovative highly available distributed platforms and pipelines and enhance the overall Snowflake data infrastructure Lead and drive projects from idea formulation to design, implementation and successful productionization. Collaborate
$153.1K – $293.8K/yr · Jobiba est.
We believe that the way people interact with their finances will drastically improve in the next few years. We’re dedicated to empowering this transformation by building the tools and experiences that thousands of developers use to create their own products. Plaid powers the tools millions of people rely on to live a healthier financial life. We work with thousands of companies like Venmo, SoFi, several of the Fortune 500, and many of the largest banks to make it easy for people to connect their financial accounts to the apps and services they want to use. Plaid’s network covers 12,000 financial institutions across the US, Canada, UK and Europe. Founded in 2013, the company is headquartered in San Francisco with offices in New York, Washington D.C., London and Amsterdam. Security Engineering is the engineering function inside the Plaid security org that focuses on developing the industry-leading security systems and infrastructure. Security Engineering owns most of Plaid’s security-related infrastructure: secure data storage, key management systems, internal identity platform, internal authentication systems, internal permission management, and internal authorization service. We develop solutions across data encryption, key management, access control, and data loss prevention to protect sensitive consumer data. We believe in the Zero Trust security model and are always looking for ways to improve our authentication and access control platforms. About the role: You will develop security capabilities to secure Plaid infrastructure and sensitive data access. You will lead the team’s strategic planning in collaboration with the manager and other senior engineers. You will own, maintain, and build Plaid’s security infrastructure and services like IAM Gateway, Key Management System and Network Firewall. You will consult with product engineers to ensure Plaid services meet security standards. You will help educate and support other engineering teams to improve security in
$153.1K – $293.8K/yr · Jobiba est.
ABOUT BASETEN Baseten powers mission-critical inference for the world's most dynamic AI companies, like Cursor, Notion, OpenEvidence, Abridge, Clay, Gamma and Writer. By uniting applied AI research, flexible infrastructure, and seamless developer tooling, we enable companies operating at the frontier of AI to bring cutting-edge models into production. We're growing quickly and recently raised our $1.5B Series F , led by Altimeter Capital, Conviction Partners, and Spark Capital. Join us and help build the platform engineers turn to to ship AI products. THE ROLE Baseten is building its own GPU infrastructure for large-scale inference. As we move into large scale, high-density NVIDIA systems, the hardest failures are intermittent, cross-layer, and difficult to prove: RoCE congestion, InfiniBand stalls, ECN/DCQCN mis-tuning, bad optics, RNIC issues, host kernel stalls, GPU driver problems, and workload symptoms that look like network problems, but are not. We are hiring a Lead Software Engineer to build a first-class observability and root-cause analysis system for GPU fabrics. This is a hard distributed systems problem, not a dashboarding problem. The system will collect high-volume signals from switches, hosts, active probes, and inference services; reduce and correlate them in real time; understand topology and service ownership; and produce actionable diagnosis while an incident is still unfolding. This role sits at the boundary between networking and inference software. RDMA data paths, GPUDirect transfers, prefill/decode disaggregation, KV cache movement, request routing, and workload backpressure can all create fabric symptoms or hide real fabric failures. The goal is to tell an operator, quickly and with evidence, whether an incident is caused by the fabric, host, NIC, GPU, RDMA path, scheduler, or serving layer — and what to do next. EXAMPLE INITIATIVES Real-time telemetry engine — Build the ingestion, reduction, storage, and query path for high-cardinality fab
From $242.6K/yr
About Pinterest: Millions of people around the world come to our platform to find creative ideas, dream about new possibilities and plan for memories that will last a lifetime. At Pinterest, we’re on a mission to bring everyone the inspiration to create a life they love, and that starts with the people behind the product. Discover a career where you ignite innovation for millions, transform passion into growth opportunities, celebrate each other’s unique experiences and embrace the flexibility to do your best work. Creating a career you love? It’s Possible. At Pinterest, AI isn't just a feature, it's a powerful partner that augments our creativity and amplifies our impact, and we’re looking for candidates who are excited to be a part of that. To get a complete picture of your experience and abilities, we’ll explore your foundational skills and how you collaborate with AI. Through our interview process, what matters most is that you can always explain your approach, showing us not just what you know, but how you think. You can read more about our AI interview philosophy and how we use AI in our recruiting process here . We’re looking for a principal software engineer to lead the next generation of data infrastructure at Pinterest which powers mission critical big data and AI applications. You’ll be working on some of the most exciting big data and AI open source technologies (Flink, Spark, Kubernetes, etc.), at the scale of exabytes of data to help Pinners discover and do what they love. What you’ll do: Lead the strategy and technical direction of Pinterest’s data infrastructure for big data and AI applications Build and scale data infra frameworks and infrastructure to process petabytes-scale datasets, including compute engines, job management, resource management, scheduling and remote shuffling Work with internal customers on critical business use cases that rely on big data Provide thought leadership to the entire company on how data should be
At Snowflake, we are powering the era of the agentic enterprise. To usher in this new era, we seek AI-native thinkers across every function who are energized by the opportunity to reinvent how they work. You don’t just use tools; you possess an innate curiosity, treating AI as a high-trust collaborator that is core to how you solve problems and accelerate your impact. We look for low-ego individuals who thrive in dynamic and fast-moving environments and move with an experimental mindset — who rapidly test emerging capabilities to discover simpler, more powerful ways to deliver results. At Snowflake, your role isn't just to execute a function, but to help redefine the future of how work gets done. Company Overview We are building the next generation of cloud security infrastructure, focusing on advanced Data Exfiltration Protection (DXP) and unified Data Movement Policies (DMP). Our mission is to provide seamless, context-aware security that protects sensitive data without hindering developer velocity. We are looking for a visionary Principal Engineer to lead the technical strategy and architecture for our Data Movement and Perimeter control systems. Role Summary As a Principal Engineer in the Data Protection group, you will be the technical lead for the Data Exfiltration Protection (DXP) and Data Movement Policy (DMP) initiatives. You will bridge the gap between high-level security policy and low-level system enforcement, ensuring that our perimeter controls are robust, scalable, and deeply integrated with context-aware access policy frameworks. You will be responsible for the architectural evolution of our egress control systems, moving from simple IP-based rules to sophisticated, content-aware, and identity-driven data movement governance. AS A PRINCIPAL SOFTWARE ENGINEER - IDENTITY, DATA SECURITY AND TRUST AT SNOWFLAKE YOU WILL: Architectural Leadership: Lead the design and implementation of the Data Movement Policy (DMP) framework, ensuring it can handle complex
At Vanta, our mission is to help businesses earn and prove trust. We believe that security should be monitored and verified continuously, and we empower companies to practice better security and prove it with ease. Vanta has a kind and talented team, and while some have prior security experience, many have been successful at Vanta without it. Our Senior Software Engineers lead and mentor engineers, delivering high-value products for our customers and infrastructure that enables our business to scale. Vanta’s product monitors the security posture for thousands of companies, pulling tens of millions of API calls of data per day, pushing information from hundreds of thousands of laptop agents, and running tests against that data continuously to identify potential security threats. Our infrastructure and tooling need to stay ahead of exponential growth in our customer base. As a Senior Software Engineer at Vanta, you’ll be responsible for setting technical direction to provide a strong foundation for our infrastructure to scale with our business. You will also drive complex projects across our technical stack and mentor our talented engineering team. Your past experience will be leveraged to enable and accelerate Vanta’s growth. Visit our Vanta Engineering Blog to learn more about what our team is working on! What you’ll do as a Senior Software Engineer on the Identity team at Vanta: As Vanta expands its agentic capabilities across surfaces like MCP, CLI, and the Vanta Agent, the Identity team is at the center of a new set of problems: defining what it means for an agent to act as a user's proxy, enforcing consistent permission checks across every invocation surface, and building an attribution model that makes agent-assisted actions auditable and trustworthy. You'll help design the identity primitives that make autonomous agent behavior safe and governable at scale. You will also: Lead complex projects with multiple stakeholders and engineers to enable our business and
Higher-paying openings
Jobs with higher listed pay
Staff Software Engineer - Fern
Postman · New York, California, United States
$3M – $3.7M/yr
Staff Software Engineer, Business Platform
Postman · San Francisco, California, United States
$2.9M – $3.6M/yr
Principal Software Engineer
Roblox · San Mateo, CA, United States
From $3.5M/yr
Principal Software Engineer, Game Safety
Roblox · San Mateo, CA, United States
From $3.5M/yr
Staff Software Engineer- Codegen
Postman · Austin, Texas, United States
$2.5M – $3.2M/yr
Sr. Staff Software Engineer, Merchants
Pinterest · San Francisco, CA, US
From $2.9M/yr
Related career options
Similar roles with stronger pay
Demand 33/100 · 7 jobs
$4.6M – $4.6M/yr
Salary →Demand 34/100 · 15 jobs
$1.8M – $1.8M/yr
Salary →Demand 46/100 · 8 jobs
$840K – $840K/yr
Salary →Demand 44/100 · 5 jobs
$840K – $840K/yr
Salary →Demand 43/100 · 6 jobs
$382.5K – $382.5K/yr
Salary →Demand 40/100 · 16 jobs
$345K – $345K/yr
Salary →Other cities to consider
More places hiring for this role
Get new lead software engineer data infrastructure specialist manager manager jobs in United States by email
Daily job updates · Unsubscribe anytime