About the team The Applied team works across research, engineering, product, and design to bring OpenAI’s technology to consumers and businesses. We seek to learn from deployment and distribute the benefits of AI, while ensuring that this powerful tool is used responsibly and safely. Safety is more important to us than unfettered growth. About the role: We're seeking a Data Engineer to take the lead in building our data pipelines and core tables for OpenAI. These pipelines are crucial for powering analyses, safety systems that guide business decisions, product growth, and prevent bad actors. If you're passionate about working with data and are eager to create solutions with significant impact, we'd love to hear from you. This role also provides the opportunity to collaborate closely with the researchers behind ChatGPT and help them train new models to deliver to users. As we continue our rapid growth, we value data-driven insights, and your contributions will play a pivotal role in our trajectory. Join us in shaping the future of OpenAI! In this role, you will: Design, build and manage our data pipelines, ensuring all user event data is seamlessly integrated into our data warehouse. Develop canonical datasets to track key product metrics including user growth, engagement, and revenue. Work collaboratively with various teams, including, Infrastructure, Data Science, Product, Marketing, Finance, and Research to understand their data needs and provide solutions. Implement robust and fault-tolerant systems for data ingestion and processing. Participate in data architecture and engineering decisions, bringing your strong experience and knowledge to bear. Ensure the security, integrity, and compliance of data according to industry and company standards. You might thrive in this role if you: Have 3+ years of experience as a data engineer and 8+ years of any software engineering experience(including data engineering). Proficiency in at least one programming language commonl
Jobs in United States
Data Scientist Salary India in United States
2,501 active opportunities · Updated October 2026
Showing
15 jobs
Explore current data scientist salary india jobs across United States. Filter by work mode, employment type, experience, department, date posted and distance.
$293K – $325K/yr
About the Team The Statsig team at OpenAI builds and operates the experimentation platform that powers product development, measurement, and decision-making across the company. We partner closely with product, engineering, and infrastructure teams to ensure experiments are trustworthy, statistically rigorous, and scalable to the needs of frontier AI products. Our mission is to help teams make better decisions through reliable experimentation. We care deeply about statistical correctness, pragmatic solutions, and building systems that researchers and engineers can trust at massive scale. The team operates at the intersection of experimentation methodology, data infrastructure, causal inference, and product analytics. We are looking for experienced experimentation experts who want to shape the future of experimentation in the AI era. About the role: We're seeking a Data Engineer to take the lead in building our data pipelines and core tables for OpenAI. These pipelines are crucial for powering analyses, safety systems that guide business decisions, product growth, and prevent bad actors. If you're passionate about working with data and are eager to create solutions with significant impact, we'd love to hear from you. This role also provides the opportunity to collaborate closely with the researchers behind ChatGPT and help them train new models to deliver to users. As we continue our rapid growth, we value data-driven insights, and your contributions will play a pivotal role in our trajectory. Join us in shaping the future of OpenAI! In this role, you will: Design, build and manage our data pipelines, ensuring all user event data is seamlessly integrated into our data warehouse. Develop canonical datasets to track key product metrics including user growth, engagement, and revenue. Work collaboratively with various teams, including, Infrastructure, Data Science, Product, Marketing, Finance, and Research to understand their data needs and provide solutions. Implement ro
About the Team OpenAI, in close collaboration with our capital partners, is embarking on a journey to build the world’s most advanced AI infrastructure ecosystem. This team is central to this mission, setting the core infra strategy and implementing this vision. From site selection to the buildout process, this team sits at the intersection of commercial, technical, strategy, and operations, interacting with teams and executives inside and outside of OpenAI. About the Role We are seeking experienced Data Center Mechanical and Electrical/Power Design Engineers with expertise in designing, operating, and maintaining large-scale data center campuses. The ideal candidate for this role will have extensive background and experience in design and managing critical equipment and facilities, design and operation of MEP (Mechanical, Electrical, Plumbing) systems, and overseeing operational activities from initial phases of Data Center build through delivery and ongoing maintenance. The ideal candidate will have a strong technical background, operational leadership experience, and a proven ability to collaborate with external vendors on critical infrastructure. This role offers the opportunity to lead transformative data center projects with high visibility and impact. If you are passionate about delivering cutting-edge infrastructure solutions, we encourage you to apply. Key Responsibilities Oversee building and MEP design, operation, and maintenance, including reviewing building and MEP drawings and proposals across all project phases. Lead operational activities for large-scale data center campuses, from early design phases through delivery and daily operation. Operate and maintain critical data center facilities and equipment, ensuring reliability and performance. Collaborate with external vendors to select, procure, and manage critical equipment, such as generators, UPS, chillers, and CDUs. Provide technical expertise on all aspects of data center building, equipment, and
About the Team At OpenAI, we’re building the connective tissue between our mission and our people. People Innovation Labs is a fast-moving engineering team embedded in the People organization, focused on rethinking how we find and retain the best talent and empower everyone to do their best work. From recruiting to culture, we’re designing systems that give our People Team a significant edge by infusing OpenAI’s models and first-principles thinking into every aspect of our work. Our projects range from greenfield 0-1 products like OpenHouse (our internal knowledge hub) to AI-powered automations and scalable recruiting tools. We’re defining the future of work at OpenAI, creating a blueprint for how AI can supercharge productivity, culture, and innovation. About the Role We’re seeking a Data Engineer to build data-intensive systems that will power People Innovation Labs’ internal products and enable the People Analytics function to do their best work. These data pipelines are crucial for our build-out of people products backed by business systems of record and for ongoing people data analytics. One example of an employee-facing product you’ll help us build is OpenHouse, which serves as a culture and communication hub and an organization-wide front door into all other aspects of People Innovation Labs’ work. OpenHouse and other products in our portfolio are built by full stack product engineers who are deeply curious about culture, recruiting and people development, and want to know everything from the business strategy and metrics down through the code that gets us there. In this role, you will work with People Innovation Labs leadership and software engineers and the People Analytics team to build the data systems that enable this work. In this role, you will: Design, build and manage people data pipelines, ensuring all data is seamlessly integrated into our Databricks warehouse. Develop canonical datasets to track key people metrics and People Innovation Labs produc
About the Team OpenAI is building the infrastructure foundation for the next generation of AI. The Data Center Engineering team defines the strategy, reference architectures, technical requirements, and delivery standards for the large-scale data centers that support OpenAI research, products, and infrastructure partners. As a Data Center Controls Network Engineer, you will design, validate, and scale the controls and OT network architectures that support high-density AI data centers. You will work across controls systems, OT infrastructure, telemetry, commissioning, deployment, and operations, partnering with mechanical, electrical, IT/networking, security, and external delivery teams. About the Role We are seeking a mid to senior OT Network Engineer with a strong controls systems background to lead the design and operation of resilient, secure, and scalable OT network architectures for high-density AI data centers. This role translates compute, power, cooling, and operational requirements into practical OT network designs, evaluates vendor solutions, and drives technical decisions across controls infrastructure, telemetry, commissioning, and operations. The ideal candidate has strong hands-on experience in mission-critical OT environments, including industrial networking, virtualized infrastructure, and OT network operations, with expertise in routing, switching, segmentation, firewall policy, time synchronization, monitoring, and network lifecycle support. Key Responsibilities Define controls, automation, and OT network requirements for AI data center campuses. Develop reference architectures, engineering standards, and reusable design templates. Review and develop basis-of-design and functional design documents, including OT network diagrams, IP/VLAN schemes, telemetry architectures, data flow diagrams, and commissioning requirements. Design OT and infrastructure network architectures, including physical topology, logical topology, IP addressing, subnetting, VLA
About the Team OpenAI is building the infrastructure foundation for the next generation of AI. The Data Center Engineering team defines the strategy, reference architectures, technical requirements, and delivery standards for the large-scale data centers that support OpenAI research, products, and infrastructure partners. As a Data Center Infrastructure Electrical Engineer, you will help define, validate, and scale the electrical power systems that support high-density AI compute. You will translate evolving compute requirements into practical facility and rack-power architectures, evaluate new technologies and vendor solutions, and drive technical decisions across design, manufacturing validation, construction, commissioning, deployment, and operations. This role is best suited for a senior hands-on engineer with deep experience in mission-critical power systems, strong judgment under ambiguity, and the ability to connect facility infrastructure, hardware requirements, controls, telemetry, reliability, and operations. About the Role We are seeking a senior electrical infrastructure engineer to lead the development of reliable, scalable, and efficient power architectures for high-density, liquid-cooled AI data centers. The ideal candidate has strong practical experience with critical electrical systems at data centers or comparable industrial scale, including medium-voltage and low-voltage distribution, utility interfaces, backup power, UPS and battery systems, rack power delivery, grounding, protection, controls, and monitoring systems. You should be comfortable moving between long-range architecture, detailed engineering review, lab validation, vendor qualification, field deployment, and operational troubleshooting. Key Responsibilities Design and optimize electrical topologies and equipment strategies that reduce cost, accelerate schedules, improve efficiency, increase scalability, and maintain high reliability and maintainability. Review and develop basis-of-des
About the Team OpenAI is building the infrastructure foundation for the next generation of AI. The Data Center Engineering team defines the strategy, reference architectures, technical requirements, and delivery standards for the large-scale data centers that support OpenAI research, products, and infrastructure partners. As a Data Center Infrastructure Engineering Program Manager, you will help turn complex infrastructure strategy into executable programs across electrical, mechanical, controls, network, hardware, construction, commissioning, deployment, and operations workstreams. You will partner with research, hardware engineering, data center engineering, site development, supply chain, security, EHS, finance, legal, operations, and external delivery partners to bring OpenAI's infrastructure vision to life. About the Role We are looking for an Engineering Program Manager (EPM) to lead assigned infrastructure programs focused on production and non-production network integration, controls coordination, and the design and deployment of data hall or whitespace facilities. The EPM will support functional Directly Responsible Individuals (DRIs) across network, controls, structural, electrical, and mechanical disciplines. Key responsibilities include coordinating assigned workstreams and program controls, maintaining risks and interfaces, and supporting readiness within the network and data hall deployment track. The ideal candidate thrives on bringing structure to complex environments characterized by ambiguous technical requirements, large partner ecosystems, tight deadlines, and high operational stakes. This individual must be adept at keeping teams aligned on decisions, risks, dependencies, schedules, and readiness criteria, and escalating gaps or decision points when needed. Candidates should have a proven track record of managing technically challenging engineering programs across major lifecycle phases, including design, validation, procurement, construction, c
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary As a Senior Data Insights Specialist, you will be responsible for using SQL, MS Excel, Dataiku, Python, and related applications to determine disruption to existing and potential clients when network changes take place to retain and win business. This position involves making high-level decisions, managing multiple requests/priorities, developing new solutions to simplify the reporting process, and producing both standard and custom reports for multiple lines of business and specific clients regarding retail pharmacy network accessibility. Required Qualifications Experience in SQL, Microsoft Excel, and other relevant applications. Experience with data cleaning, transformation, and/or ETL. Strong analytical and problem-solving skills with the ability to interpret complex data sets. Preferred Qualifications 2+ years' of professional analytical experience. Proficiency in Dataiku, Alteryx, or other related ETL applications. Experience with Python. Experience with Salesforce. Working experience in the Pharmacy Benefits Management (PBM) or related healthcare industry. Excellent verbal and written communication skills to effectively interact with internal teams and external clients. Ability to manage multiple requests and priorities in a fast-paced environment.</li
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. POSITION SUMMARY CVS Health is seeking a highly skilled Staff Data Engineer, Observability Engineering to join the Enterprise Observability Platform organization and help advance the next generation of observability, infrastructure, and security data capabilities. The Staff Data Engineer, Observability Engineering will play a critical role in designing, building, and operating scalable data pipelines and data products that power enterprise observability, operational intelligence, and security analytics across the organization. The Staff Data Engineer, Observability Engineering is a senior individual contributor responsible for developing and optimizing Databricks-based data engineering solutions that ingest, transform, govern, and deliver high-volume telemetry, infrastructure, application, and security data. This role combines deep hands-on technical execution with ownership of engineering excellence, operational reliability, performance optimization, and data platform best practices. Working closely with Observability Engineering, Security Engineering, Infrastructure Engineering, and Data Platform teams, the Staff Data Engineer, Observability Engineering will contribute to the evolution of the enterprise observability lakehouse by building resilient ingestion frameworks, establishing data quality standards, enhancing governance controls, and driving efficient, scalable data processing patterns. The id
We’re building a world of health around every individual — shaping a more connected, convenient and compassionate health experience. At CVS Health®, you’ll be surrounded by passionate colleagues who care deeply, innovate with purpose, hold ourselves accountable and prioritize safety and quality in everything we do. Join us and be part of something bigger – helping to simplify health care one person, one family and one community at a time. Position Summary The growing DSNP business has created an opportunity for an individual with claims experience, who is familiar with the 837 standard claims format. This individual will own 837 file transmissions to the states, manage and act on state response files, ensuring that all transactions transmitted are complete and error free. Manage and complete error corrections to meet state requirements. Required Qualifications 1+ year of experience with encounter data, medical claims, or Medicare/Medicaid. 1+ year of experience using FTP and data transfer software. 1+ year of data management experience. Preferred Qualifications Experience with Microsoft Access Databases. Analytical skills with the ability to identify and resolve data discrepancies. Working knowledge of the 837 claims files. Education Bachelors degree or equivalent work experience Anticipated Weekly Hours 40 Time Type Full time Pay Range <p s
NVIDIA is a global leader in high-speed computer vision, artificial intelligence (AI), and deep learning. Our team develops data engineering solutions that empower AI developers in autonomous vehicle (AV) domains to innovate quickly and effectively at scale. Are you ready to take on a senior technical role in building high-performance AI data pipelines? We seek an exceptional individual to design and optimize microservices and data pipelines to process massive volumes of AV data and enable seamless data mining and AI training. The ideal candidate will bring expertise in big data processing and distributed computing to create efficient solutions and overarching architectures for challenges such as video data curation, behavioral search, and AI dataset management. What you'll be doing: Scope and build tools, microservices, workflows, and distributed applications to accelerate data mining and AI training. Design and implement solutions for streaming, resilience, logging, security, authentication, workflow orchestration, and data management. Deploy AI models. Design and develop Retrieval-Augmented Generation (RAG) workflows enabling hybrid and agentic patterns. Analyze and operationalize complex distributed systems for speed-of-light performance. What we need to see: Experience developing high-performance, scalable software systems. MS with 6+ years, or BS (or equivalent experience) with 8+ years of relevant experience in Computer Science, Computer Engineering, or a related technical field. Strong programming skills in Python or Golang Proficiency in key technologies like Kubernetes, Helm, Hive, Parquet, SQL, vector databases, e.g., Milvus. Strong architectural skills with a proactive, problem-solving mentality. Experience in data mi
$141.6K – $212.4K/yr
What if the work you did every day could impact the lives of people you know? Or all of humanity? At Illumina, we are expanding access to genomic technology to realize health equity for billions of people around the world. Our efforts enable life-changing discoveries that are transforming human health through the early detection and diagnosis of diseases and new treatment options for patients. Working at Illumina means being part of something bigger than yourself. Every person, in every role, has the opportunity to make a difference. Surrounded by extraordinary people, inspiring leaders, and world changing projects, you will do more and become more than you ever thought possible. Summary The Staff Data Engineer is a seasoned, hands-on engineer who designs, builds, and scales data products on our cloud lakehouse, powering analytics, reporting, and AI/ML across Illumina. We are looking for someone with strong proficiency in Python, SQL, and data modeling, a solid understanding of distributed systems and system design who has built and scaled data products on modern cloud platforms such as Databricks and Snowflake. This is a hands-on, senior individual-contributor role with end-to-end ownership and leadership spanning multiple domains such as Supply Chain, Manufacturing and Quality, including mentoring engineers on our global (India-based) team. Responsibilities Partner across business, AI, and platform teams translating domain needs (e.g., SAP, Manufacturing, Quality) into well-modeled, governed and scalable data products. Design, build, and scale end-to-end data products on Databricks (and interoperating with Snowflake) — from ingestion through curated, analytics-ready datasets following a medallion (Bronze/Silver/Gold) architecture. Develop reusable frameworks, libraries, and
At Freddie Mac, our mission of Making Home Possible is what motivates us, and it’s at the core of everything we do. Since our charter in 1970, we have made home possible for more than 90 million families across the country. Join an organization where your work contributes to a greater purpose. Position Overview: The Cyber Security team at Freddie Mac is searching for a strong data analyst to collaborate on developing and administering data security policies as well as safeguarding information, evaluating existing data security procedures and identifying new areas of risk for Freddie Mac. If this role sounds like a fit for your skill set, please read on, apply and learn why there is #MoreatFreddieMac ! Our Impact: We develop and train the enterprise on relevant Identity and Access Management best practices, develop and support automated IAM processes, and develop and implement ongoing IAM efficiency improvements with limited oversight by managers. The role will work closely with Enterprise partners and security control owners on IAM automation development activities and alignment of IAM processes with InfoSec maturity targets as described in the InfoSec Strategy. Your Impact: Senior Data Analyst who can develop and implement new and updated business process automation for all Identity and Access Management activities. Develop and proposes strategies to reduce security risk in the organization by implementing procedural prevention, detection, and response measures, while still enabling positive business outcomes. Comfortable supporting ad hoc data retrieval and analysis requests using SQL queries and Copilot/Excel . Creating audit-ready reference documentation. This role will allow for and support rapid AI-based extrapolation/replication of these services as future automated IAM microservices. Qualifications: Typically, 5 - 7 years of rel
Principal Data Privacy Architect Description - Job Summary - Role Purpose • Lead and oversee complex, cross-functional privacy and data protection programs from strategy through implementation, ensuring alignment across business, technical, legal, and compliance stakeholders. • This role will design and implement scalable, AI-ready data privacy architecture across enterprise data environments, applications, and AI-enabled workflows. • The Principal Data Privacy Architect will serve as a hands-on subject matter expert responsible for embedding privacy-by-design, consent enforcement, data sovereignty, data loss prevention, and compliance controls into large, complex global data environments. • The architect will partner closely with Data Engineering, Cybersecurity, Legal, Privacy, AI Governance, Product, and Enterprise Architecture teams to ensure customer, employee, partner, and sensitive enterprise data is accessed, processed, shared, retained, and protected in a compliant, secure, and trustworthy manner. - Why This Role Matters • Architect for Trust & Scale: Build reusable privacy architecture patterns that enable secure, compliant, and scalable data usage across platforms, products, and regions. • Enable Responsible AI: Design privacy guardrails for AI agents, generative AI, RAG pipelines, model inputs and outputs, embeddings, vector stores, and automated data workflows. • Reduce Risk While Enabling Innovation: Translate privacy, consent, regulatory, and data sovereignty obligations into practical engineering controls that accelerate business outcomes. Responsibilities - Think Customer First • Embed customer trust, transparency, and privacy-by-design principles into enterprise data platforms and customer-facing applications. • Design consent-aware data access and usage p
Abbott is a global healthcare leader that helps people live more fully at all stages of life. Our portfolio of life-changing technologies spans the spectrum of healthcare, with leading businesses and products in diagnostics, medical devices, nutritionals and branded generic medicines. Our 122,000 colleagues serve people in more than 160 countries. JOB DESCRIPTION: Position Overview The Principal Data Engineer (IC) is a senior individual contributor and the accountable technical leader for assigned cross-domain initiatives and enterprise data engineering capabilities. The role owns integrated technical direction and technical outcomes for work spanning multiple data domains, defines and stewards enterprise engineering standards and reference architectures, and drives convergence where duplicated or inconsistent solutions create enterprise cost, risk, or operational burden. The role advises on scope, sequencing, capacity, dependencies, and technical debt, but does not independently commit domain resources or business delivery dates. This position has no people-management responsibility. Enterprise Data operates a domain-aligned model built on Databricks and Unity Catalog. Working with Domain Leaders, Staff Engineers, Platform Engineering, and partner organizations, the role converts ambiguous enterprise needs into executable architecture and carries the most complex or highest-risk work through validation and production. The role remains hands-on through prototyping, reference implementations, critical-path development, design and code review, and production problem solving. This role is based in Madison, WI. Essential Duties Include, but are not limited to, the following: Cross-domain technical leadership and delivery Own the technical outcome of assigned cross-domain initiatives from initial ambigu
Higher-paying openings
Jobs with higher listed pay
Staff Data Scientist, Forecasting
Pinterest · San Francisco, US; Remote, US
From $2M/yr
Data Scientist, Safety
OpenAI · San Francisco, California, United States
$230K – $325K/yr
Data Scientist, Core Experimentation
OpenAI · Seattle, Washington, United States
$293K – $325K/yr
Principal Data Scientist - Safety
Roblox · San Mateo, CA, United States
From $321.2K/yr
Senior / Principal Data Scientist - Discovery
Roblox · San Mateo, CA, United States
From $307.4K/yr
Senior Data Scientist, Consumer Apps
Roblox · San Mateo, CA, United States
From $263.7K/yr
Related career options
Similar roles with stronger pay
Demand 46/100 · 8 jobs
$840K – $840K/yr
Salary →Demand 43/100 · 5 jobs
$840K – $840K/yr
Salary →Demand 43/100 · 6 jobs
$382.5K – $382.5K/yr
Salary →Demand 43/100 · 8 jobs
$300K – $300K/yr
Salary →Demand 42/100 · 7 jobs
$300K – $300K/yr
Salary →Demand 43/100 · 22 jobs
$278.9K – $278.9K/yr
Salary →Other cities to consider
More places hiring for this role
Get new data scientist salary india jobs in United States by email
Daily job updates · Unsubscribe anytime