NVIDIA is leading the way in groundbreaking developments in Artificial Intelligence, High Performance Computing and Visualization. The GPU, our invention, serves as the visual cortex of modern computers and is at the heart of our products and services. Our work opens up new universes to explore, enables amazing creativity and discovery, and powers what were once science fiction inventions from artificial intelligence to autonomous cars. We are looking for a motivated Deep Learning engineer to bring advanced communication technologies into AI stacks, including PyTorch, TRT-LLM, vLLM, SGLang, JAX, etc. You will be working with the team that created communication libraries like NCCL, NVSHMEM & technology like GPUDirect -- for scaling Deep Learning and HPC applications. Your customers will have diverse multi-GPU demands, ranging from training on scales up to 100K GPUs to inference down at microsecond latency. Communication performance between the GPUs has a direct impact on AI applications. Your work in AI toolkits will make all of those easier for the community. This is an outstanding opportunity for someone with an AI background to advance the state of the art in this space. Are you ready to contribute to the development of innovative technologies and help realize NVIDIA's vision? What you will be doing: Integrate new communication libraries features in AI frameworks: from PoC to performance analysis to production Perform deep analysis of AI workloads and frameworks to identify multi-GPU communication requirements and opportunities. Collaborate hands-on with teams working on the latest AI models. Improve AI compilers to hide communications or perform automatic fusion. Conduct in-depth AI workload performance characterization on multi-GPU clusters. Design fault-tolerant and elastic solutions for large-scale or dynamic AI workloads. Author
Jobiba hiring network
Senior Solution Engineer Jobs
7,101 active opportunities · Updated for October 2026
Fresh results
15 shown
Explore current senior solution engineer jobs. Use filters to narrow by work mode, employment type, experience and date posted.
NVIDIA has been transforming computer graphics, PC gaming, and accelerated computing for more than 25 years. It’s a unique legacy of innovation that’s fueled by great technology and amazing people. Today, we’re tapping into the unlimited potential of AI to define the next era of computing. An era in which our GPU acts as the brains of computers, robots, and self-driving cars that can understand the world. Doing what’s never been done before takes vision, innovation, and the world’s best talent. As an NVIDIAN, you’ll be immersed in a diverse, supportive environment where everyone is inspired to do their best work. NVIDIA has a rapidly expanding ecosystem of data center platform designs. From single node HGX/DGX systems all the way up to large multi-node NVLink domain rack architectures. These designs have become core to NVIDIA's rapidly growing enterprise and cloud provider businesses. Each brings together the full power of NVIDIA GPUs, NVIDIA NVLink, NVIDIA InfiniBand networking, NVIDIA Grace CPUs, and a fully optimized NVIDIA AI and HPC software stack. We are searching for a highly motivated engineer to lead performance benchmarking and optimization efforts for our data center products. You will be instrumental in ensuring our data center solutions deliver industry-leading performance for accelerated computing workloads. What you will be doing: Design and execute comprehensive performance benchmarking strategies for our data center platforms and products Characterize real-world AI training, inference, and HPC workloads at scale Define, track, and report key performance indicators (throughput, latency, efficiency, scaling) Build automation tools and frameworks for performance monitoring and analysis Identify and analyze performance bottlenecks across compute, memory, network and storage subsystems Work closely with architecture, hardware,
About The Role & Team Amplitude is the leading AI analytics platform, and our ability to deliver measurable customer outcomes quickly is a key part of how we keep that lead. The Customer FDE (Forward Deployed Engineering) team sits at the intersection of engineering, product, and customer success — owning the technical delivery that takes validated products from co-development and implements them across enterprise customers. As a Customer Forward Deployed Engineer, you will own end-to-end technical delivery for enterprise customer implementations, from sales engagement through post-deployment validation. You'll work directly in Amplitude's product codebase, submitting PRs, shipping customer-specific solutions, and building reusable patterns that make every successive engagement faster. This is not a traditional support or solutions role. Customer FDEs are engineers first who partner closely with Sales, Customer Success, Product, and Engineering to turn customer needs into working software. The right person thrives in ambiguity, learns new domains quickly, and cares as much about the customer's outcome as about technical elegance. As a Customer Forward Deployed Engineer, you will: Own enterprise implementations end-to-end — engage during sales to map customer needs to capabilities, scope implementation plans, and deliver through post-deployment validation with measurable customer outcomes Work directly in the product codebase — submit PRs, test, and ship customer-specific solutions without requiring constant oversight; over time, the gap between "what we built" and "what the customer needs" keeps shrinking because you're the one closing it Ship early, ship often, and gather feedback — treat every customer interaction as an opportunity to validate your approach and course-correct quickly Build reusable patterns and playbooks — when you solve something once, make sure the next person doesn't have to solve it again; identify and document implementation patterns that
Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, networking, load balancing (including our public-facing edge and internal service mesh), and observability and alerting systems. The Deployments team designs and maintains our continuous delivery infrastructure, ensuring reliable code deployment from development through production for all engineering teams. This infrastructure is primarily composed of Argo Workflows and ArgoCD. The team also provides tooling that enables clear system ownership and facilitates self-service onboarding for development teams. We are looking to speak to candidates who can work East Coast hours. The ideal candidate should Have 6+ years of experience in software development and operating distributed systems Proficiency in Python, Go, or a similar language Proven experience building and operating large-scale continuous integration and continuous deployment (CI/CD) pipelines Possess a customer-focused mindset Value efficiency in processes and operations Prefer automation over manual process (“allergic to ops work”). We are a small team of software engineers with a strong bias towards software solutions to avoid toil Experience using and extending containerization technologies, particularly Kubernetes, to enhance application agility, optimize resource utilization, and accelerate time-to-market Expertise in cloud infrastructure platforms, including AWS, Google Cloud Platform (GCP), or Azure Understanding of Linux operating system internals and networking concepts (e.g., TCP/IP, DNS, TLS, routing) Expectations Contribute to developing a world-class continuous deployment experience, enabling the rapid and reliable shipment of MongoDB products This includes, but is not limited to, contributing to open-source projects, or engineering software-based
MongoDB’s Storage Layer Services (SLS) team is re-architecting the MongoDB cloud storage layer and sits at the heart of our next-generation cloud storage architecture. This relatively new team is building performant, multi-tenant distributed storage services that both enhance today’s Atlas storage stack and enable more customer workloads to run more efficiently. You will partner with the teams building these storage services to define SLOs, shape capacity plans, and ensure the reliability, durability, and operational safety of the storage layer that underpins Atlas. You’ll join a small, senior team of SREs as founding members of this organization, playing a crucial role in executing on a multi-year roadmap for MongoDB’s cloud storage architecture. This role can be based out of our Boston, New York City, Raleigh, Miami, Pittsburgh or remotely in the United States while physically based in an Eastern or Central time zone location. The ideal candidate should Have 6+ years of experience working on software development and operating distributed systems Proficiency in Python, Go, or a similar language Have operated or supported stateful storage or database systems at scale, and are comfortable with durability, consistency, and recovery trade-offs. Possess a customer-focused mindset Value efficiency in processes and operations Prefer automation over manual processes. We are a small team of software engineers with a strong bias towards software solutions to avoid toil Experience using and extending containerization technologies, particularly Kubernetes, to enhance application agility, optimize resource utilization, and accelerate time-to-market Expertise in cloud infrastructure platforms, including AWS, Google Cloud Platform (GCP), or Azure Understanding of Linux operating system internals and networking concepts (e.g., TCP/IP, DNS, TLS, routing) Responsibilities Work on our multi-tenant distributed storage systems, balancing long-term strategic infrastructure g
Platform Engineering is the department within SRE that is responsible for a range of critical infrastructure and operational functions that support the broader engineering organization. Among these are our multi-cloud-provider Kubernetes infrastructure, networking, load balancing (including our public-facing edge and internal service mesh), and observability and alerting systems. The Deployments team designs and maintains our continuous delivery infrastructure, ensuring reliable code deployment from development through production for all engineering teams. This infrastructure is primarily composed of Argo Workflows and ArgoCD. The team also provides tooling that enables clear system ownership and facilitates self-service onboarding for development teams. We are looking to speak to candidates who can work East Coast hours. The ideal candidate should Have 6+ years of experience in software development and operating distributed systems Proficiency in Python, Go, or a similar language Proven experience building and operating large-scale continuous integration and continuous deployment (CI/CD) pipelines Possess a customer-focused mindset Value efficiency in processes and operations Prefer automation over manual process (“allergic to ops work”). We are a small team of software engineers with a strong bias towards software solutions to avoid toil Experience using and extending containerization technologies, particularly Kubernetes, to enhance application agility, optimize resource utilization, and accelerate time-to-market Expertise in cloud infrastructure platforms, including AWS, Google Cloud Platform (GCP), or Azure Understanding of Linux operating system internals and networking concepts (e.g., TCP/IP, DNS, TLS, routing) Expectations Contribute to developing a world-class continuous deployment experience, enabling the rapid and reliable shipment of MongoDB products This includes, but is not limited to, contributing to open-source projects, or engineering software-based
The Team MongoDB’s Storage Layer Services (SLS) team is re-architecting the MongoDB cloud storage layer and sits at the heart of our next-generation cloud storage architecture. This relatively new team is building performant, multi-tenant distributed storage services that both enhance today’s Atlas storage stack and enable more customer workloads to run more efficiently. You will partner with the teams building these storage services to define SLOs, shape capacity plans, and ensure the reliability, durability, and operational safety of the storage layer that underpins Atlas. You’ll join a small, senior team of SREs as founding members of this organization, playing a crucial role in executing on a multi-year roadmap for MongoDB’s cloud storage architecture. This role can be based out of either our Dublin or Cork office or remotely in Ireland. The ideal candidate should Have 6+ years of experience working on software development and operating distributed systems Proficiency in Python, Go, or a similar language Have operated or supported stateful storage or database systems at scale, and are comfortable with durability, consistency, and recovery trade-offs. Possess a customer-focused mindset Value efficiency in processes and operations Prefer automation over manual processes. We are a small team of software engineers with a strong bias towards software solutions to avoid toil Experience using and extending containerization technologies, particularly Kubernetes, to enhance application agility, optimize resource utilization, and accelerate time-to-market Expertise in cloud infrastructure platforms, including AWS, Google Cloud Platform (GCP), or Azure Understanding of Linux operating system internals and networking concepts (e.g., TCP/IP, DNS, TLS, routing) Responsibilities Work on our multi-tenant distributed storage systems, balancing long-term strategic infrastructure goals with immediate engineering needs Build for reliability, making services and infrastructure avail
MongoDB’s Storage Layer Services (SLS) team is re-architecting the MongoDB cloud storage layer and sits at the heart of our next-generation cloud storage architecture. This relatively new team is building performant, multi-tenant distributed storage services that both enhance today’s Atlas storage stack and enable more customer workloads to run more efficiently. You will partner with the teams building these storage services to define SLOs, shape capacity plans, and ensure the reliability, durability, and operational safety of the storage layer that underpins Atlas. You’ll join a small, senior team of SREs as founding members of this organization, playing a crucial role in executing on a multi-year roadmap for MongoDB’s cloud storage architecture. This role can be based out of our Toronto or Montreal office or remotely in the Canada while physically based in an Eastern or Central time zone location. The ideal candidate should Have 6+ years of experience working on software development and operating distributed systems Proficiency in Python, Go, or a similar language Have operated or supported stateful storage or database systems at scale, and are comfortable with durability, consistency, and recovery trade-offs. Possess a customer-focused mindset Value efficiency in processes and operations Prefer automation over manual processes. We are a small team of software engineers with a strong bias towards software solutions to avoid toil Experience using and extending containerization technologies, particularly Kubernetes, to enhance application agility, optimize resource utilization, and accelerate time-to-market Expertise in cloud infrastructure platforms, including AWS, Google Cloud Platform (GCP), or Azure Understanding of Linux operating system internals and networking concepts (e.g., TCP/IP, DNS, TLS, routing) Responsibilities Work on our multi-tenant distributed storage systems, balancing long-term strategic infrastructure goals with immediate engineerin
About the Team DoorDash’s mission is to grow and empower local economies. By building intelligent, last-mile delivery technology for local cities, DoorDash connects people with the local businesses they care about — helping grow businesses and the communities that support those businesses. The Marketing Technology team is responsible for delivering best-in-class tools, data, and processes so teams at DoorDash can effectively acquire new customers, drive retention, and grow new business lines across all our audiences. Our mission is to enable teams to deliver personalized, engaging, and relevant content across all our owned and paid channels. We partner closely with Product, Engineering, Data Science, and Analytics teams to translate Marketing’s technology needs into measurable business outcomes for our Growth and Performance Marketing teams. About the Role We're looking for a Senior Associate, AI Solutions Architect to help build and scale AI-powered tools across DoorDash's Marketing teams. You'll work alongside our existing Marketing Technology AI Solutions Architects to identify opportunities where AI can solve real marketing problems, then build, test, and iterate on the solutions that make it happen. This role is grounded in hands-on building: you'll spend your time prototyping, coding, and shipping tools rather than pitching strategy from the sidelines. This is a great opportunity to grow your technical and strategic skills at the intersection of marketing and AI. You'll get mentorship from a team of builders and exposure to the full lifecycle of AI development from prototype to deployment. This role will report to the Senior Manager, Marketing Technology within the Growth Marketing organization. You're excited about this opportunity because you will… Build and deploy AI-powered tools and automations, such as automated reporting, QA frameworks, and workflow optimizations. Develop MVPs and proof-of-concepts, then iterate based on user feedback within a defined w
Job Details: Job Description: The Role and Impact As a Module Engineer, you will play a pivotal role in enabling Intel's cutting-edge manufacturing processes by ensuring the seamless operation and optimization of critical high-volume equipment. Your day-to-day responsibilities will focus on conducting tests, measurements, and maintenance activities to ensure the equipment meets stringent safety, quality, and efficiency standards. By driving process improvements and collaborating with global teams, you will contribute to the successful production of advanced integrated circuits and help shape technological innovation in the semiconductor industry. Business Group You will be part of Intel's Manufacturing and Operations organization, which is central to Intel's mission of creating world-changing technology. The group specializes in scaling advanced manufacturing techniques, ensuring high product quality, and delivering cost-effective solutions that support Intel's growth. Through its close collaboration with global teams, this organization fosters technological innovation and operational excellence across Intel's manufacturing sites worldwide. Key Responsibilities - Conduct equipment tests and measurements to ensure control over critical dimensions, registration and defectivity of the production line. - Recommend and implement modifications to improve production efficiency, manufacturing techniques, and output quality. - Own and execute maintenance and repair activities for manufacturing equipment, ensuring minimal downtime. - Lead continuous improvement initiatives for safety, quality, cost, productivity, defects, and yield metrics. - Collaborate with equipment suppliers to enhance manufacturing tools and processes. - Develop and optimize systems for preventing excursions and detecting process or equipment discrepancies. - Participate in technology transfers t
The Opportunity This is a critical and exciting time at Enigma. Our customers consistently tell us that our data products create tremendous value and are deeply aligned with their most important workflows. As demand grows, we have an urgent opportunity to improve both the intelligence of our data and the systems through which customers access it. We are looking for an experienced Senior/Staff Machine Learning Engineer to join our Match Team and help shape the next generation of Enigma’s customer-facing data products. In this role, you will combine advanced statistical and machine learning research with the engineering systems required to power fast, relevant, and reliable search experiences at scale. This is a uniquely high-impact role sitting at the intersection of information retrieval, ranking systems, semantic search, distributed systems, and customer data delivery. The Role At the core of Enigma’s product is our data, which makes both data science and delivery systems central to what we build. As a Senior/Staff ML Engineer on the Match Team, you will lead efforts that improve the relevance, latency, and scalability of our customer-facing data products. You’ll work across the full lifecycle: framing retrieval and ranking problems, developing models and experimentation strategies, evaluating results using real-world signals, and implementing high-throughput search and retrieval systems. This role is ideal for someone who is excited by both hard ranking/search problems and the systems challenges of turning those solutions into low-latency, production-grade retrieval systems. What You'll Do Develop innovative solutions to complex problems in information retrieval, ranking, semantic search, query understanding, and recommendation systems Build and optimize low-latency, high-throughput search APIs, indexing pipelines, and retrieval systems using Python, Typesense, and AWS Evaluate and evolve our search technology stack, driving technical design decisions across index
Artefact is a next-generation data and AI consulting firm dedicated to accelerating the adoption of data and AI to create measurable business impact across the full enterprise value chain. We sit at the intersection of consulting, data science, AI technologies, data engineering, and digital transformation. We do not just advise — we build, implement, and deliver results our clients can measure. Our teams bring together consultants, data scientists, data engineers, AI engineers, analysts, and digital experts to solve complex business challenges with pragmatic, production-ready solutions. As Artefact continues to grow globally, we are building a team of entrepreneurial data and AI talent who can help clients move beyond experimentation and into scalable, governed, value-generating AI adoption. The Role Artefact is looking for a Senior AI & Agentic Engineer: a full-stack engineer who takes AI features from idea to production. You will design and build the interfaces, services, and agentic systems at the heart of our client work, such as conversational applications over enterprise data, multi-step agents that automate business workflows, and the retrieval and data pipelines that support them. You will own your components end to end: the React front end, the Python or Node service behind it, the RAG pipeline feeding it, and the evaluations proving it works. This role combines breadth and depth: the ability to take a feature from front end to cloud deployment, together with strong expertise in at least one major AI platform — Google (Gemini), Anthropic (Claude), or OpenAI. You will work with direct client exposure, and you will support the professional development of the junior engineers around you. What You'll Do Build Full-Stack AI Applications, End to End You will build AI products across the entire stack, from interface to infrastructure. Develop user-facing interfaces in TypeScript/React and the backend services and APIs behind them in Python or Node. Implement a
Our vision is to transform how the world uses information to enrich life for all . Micron Technology is a world leader in innovating memory and storage solutions that accelerate the transformation of information into intelligence, inspiring the world to learn, communicate and advance faster than ever. Key Responsibilities Overlay Strategy, Roadmap & Technical Leadership Own the overlay roadmap including bond overlay technologies. Serve as the technical authority for overlay, defining overlay specifications, error budgets, mark strategies, technical signoff criteria, and scaling requirements for future technology nodes. Develop innovative overlay solutions that enable advanced bonding while meeting performance, yield, manufacturability, and scaling objectives. Drive advancement of overlay metrology, automation, process control, modeling, and analytics capabilities in partnership with Process and Metrology teams. Apply advanced analytics, Artificial Intelligence, simulation, and experimental methodologies to accelerate learning, identify root causes, and improve overlay performance. Technology Development & Problem Solving Lead simulation and experimental process development activities, balancing modeling with hands-on characterization and technology development. Develop and maintain overlay error budgets, validation plans, experiments, and technical assessments that connect process behavior to device and integration requirements. Solve complex overlay challenges involving alignment, distortion, deformation, process interactions, and bonded wafers. Translate technical findings into actionable process improvements, technology decisions, and roadmap recommendations. <
Perforce is a community of collaborative experts, problem solvers, and possibility seekers who believe work should be both challenging and fun. We are proud to inspire creativity, foster belonging, support collaboration, and encourage wellness. At Perforce, you’ll work with and learn from some of the best and brightest in business. Before you know it, you’ll be in the middle of a rewarding career at a company headed in one direction: upward. With a global footprint spanning more than 80 countries and including over 75% of the Fortune 100, Perforce Software, Inc. is trusted by the world’s leading brands to deliver solutions for the toughest challenges. The best run DevOps teams in the world choose Perforce. Position Summary: Perforce is seeking a Senior Manager, Software Engineering to lead the engineering team for Gliffy, our visual collaboration and diagramming product, based in Pune, India. This role owns day-to-day engineering execution, delivery quality, and team development for Gliffy, while also engaging with the Akana and OpenLogic engineering teams to build cross-product collaboration and partnering closely with the VP of Engineering on strategic planning and roadmap decisions. We are looking for a seasoned leader who can operate with significant scope and autonomy, bring strong organizational and cross-functional leadership skills, and contribute to strategic thinking across Perforce’s growing portfolio of products — not just execute within a single team. The ideal candidate has proven experience leading distributed engineering teams in a high-growth SaaS or enterprise software environment, a strong grasp of modern SDLC practices, and a track record of embedding AI into development workflows. They are comfortable operating with significant autonomy, influencing senior leaders and cross-functional stakeholders, and balancing hands-on technical depth with a broader strategic perspective. Responsibilities: Own the day-to-day planning, execu
Java Software Engineer - Developer (Experienced and Senior) Company: The Boeing Company The Boeing Company is currently seeking Java Software Engineers – Developer (Experienced and Senior) to support our Advanced Ground Architecture team located in Herndon, Virginia, Colorado Springs, Colorado, Mesa, Arizona, Seal Beach, California and El Segundo, California. This position will focus on supporting the Boeing Defense, Space & Security (BDS) Software Engineering organization. The Advanced Ground Architecture (AGA) software team is a dynamic group of software engineers creating the future of Ground support with the extensibility and adaptability to be used across ALL Boeing programs. The software team is executing this vision through modern software technologies (Java, ReactJS, python, CI/CD pipelines) and methodologies (Scaled Agile). AGA is looking for self-motivated high performers to execute the large scope of Java development needed for the program's vision. The ideal candidates will provide software engineering functions for the design, development, and maintenance of complex, multi-tiered application software systems used to support the command and control of space vehicles. The software engineers will work day-to-day with system and test engineers in order to implement, test, and document new features and improvements for both web services and applications supporting distributed computing solutions. Position Responsibilities: Designs, develops, tests, and maintains software in an Agile execution that meets industry, customer, safety, and regulation standards throughout the end-to-end lifecycle Reviews, analyzes, and translates customer requirements into initial design and softwa
Get new senior solution engineer jobs by email
Daily job updates · Unsubscribe anytime