Green Card Data Scientist Jobs
Data Scientist roles qualify for green card sponsorship under EB-2 for advanced-degree professionals and EB-3 for skilled workers with a bachelor's degree. Employers file a PERM labor certification with DOL before sponsoring your I-140 petition, making this one of the most direct paths to permanent U.S. residency for quantitative professionals.
Find Green Card Data Scientist JobsOverview
Showing 5 of 796+ Data Scientist jobs










See all 796+ Data Scientist Jobs
Sign up for free to unlock all listings, filter by visa type, and get alerts for new Data Scientist roles.
Get Access To All Jobs
INTRODUCTION
At St. Jude Children's Research Hospital, we are committed to accelerating discoveries that improve outcomes for children with catastrophic diseases through innovation, collaboration, and scientific excellence. The Data Scientist, Clinical Machine Learning and Flow Cytometry, will play a critical role in advancing next-generation diagnostic analytics by developing and implementing machine learning solutions for high-dimensional spectral flow cytometry data. Working closely with clinical faculty, laboratory scientists, and multidisciplinary data science teams, this position will help transform measurable residual disease (MRD) detection and clinical flow cytometry interpretation through scalable, reproducible, and clinically validated analytical approaches. The successful candidate will contribute to improving diagnostic accuracy, reducing turnaround times, enhancing laboratory efficiency, and strengthening St. Jude's leadership in precision diagnostics, translational research, and the responsible application of artificial intelligence in healthcare.
ROLE AND RESPONSIBILITIES
The Data Scientist will lead the development, validation, and deployment of machine learning solutions for high-dimensional clinical flow cytometry data. Working in close collaboration with faculty and laboratory leadership in Clinical Immunopathology, the incumbent will design and implement analytical frameworks that support automated identification of rare and clinically relevant cell populations, improve measurable residual disease (MRD) detection, reduce manual interpretation burden, and enhance diagnostic accuracy and reproducibility. The position will support the development of machine learning pipelines for spectral flow cytometry datasets, longitudinal quality monitoring systems, and scalable analytical workflows for clinical laboratory operations.
Job Responsibilities:
- Lead data analysis and deliver high-quality results by formulating advanced and innovative machine learning approaches to address challenging clinical flow cytometry analysis questions.
- Adapt and optimize analytical methodologies to support high-dimensional spectral cytometry and MRD detection initiatives.
- Design, develop, validate, and maintain machine learning pipelines for supervised and unsupervised analysis of flow cytometry data, including clustering, dimensionality reduction, classification, anomaly detection, and predictive modeling.
- Deliver data products, analytical reports, visualizations, and technical documentation that support clinical implementation, regulatory review, and scientific publication.
- Document analytical methods, model performance, validation results, and quality assurance procedures.
- Establish and document protocols, best practices, and reproducible workflows for machine learning applications in clinical flow cytometry and laboratory quality monitoring.
- Develop methods for longitudinal monitoring of assay performance, including statistical process control, drift detection, and quality assessment of instrument, reagent, and workflow variability.
- Recommend opportunities to automate and improve existing analytical workflows and implement enhancements that increase throughput, reproducibility, and operational efficiency.
- Evaluate, benchmark, and test emerging machine learning methods, algorithms, and technologies applicable to biomedical and clinical diagnostics. Develop reusable code, workflows, and software tools that can be leveraged across projects and laboratories.
- Collaborate with pathologists, laboratory scientists, bioinformaticians, statisticians, and data scientists to translate clinical and scientific questions into computational solutions.
- Participate in manuscript preparation, scientific presentations, abstracts, and dissemination of project outcomes to internal and external research communities.
- Lead and participate in interdisciplinary projects involving data science, laboratory medicine, clinical diagnostics, and translational research. Act as project manager when required.
Special Skills, Knowledge and Abilities:
Critical Thinking & Agility:
- Draw insights from large, complex, and heterogeneous datasets.
- Identify root causes of analytical challenges and develop practical solutions.
- Adapt rapidly to evolving technologies, datasets, and clinical requirements.
- Recognize opportunities for innovation and process improvement in diagnostic workflows.
Communication & Influence:
- Communicate complex analytical concepts effectively to scientific, clinical, and operational audiences.
- Collaborate across multidisciplinary teams to achieve project objectives.
- Present findings clearly through reports, publications, and presentations.
- Utilize modern digital collaboration and communication tools effectively.
Results & Execution:
- Maintain focus on project goals amidst competing priorities and evolving requirements.
- Apply analytical rigor to resolve unexpected challenges and optimize outcomes.
- Drive accountability and ownership for delivering impactful results.
- Support implementation of machine learning solutions in operational clinical settings.
Scientific Domain Translation:
- Apply knowledge of hematopathology, immunology, flow cytometry, and clinical laboratory operations to develop meaningful analytical solutions.
- Assess data quality, biological relevance, and model outputs within clinical context.
- Translate scientific and clinical questions into appropriate machine learning frameworks.
- Contribute to publications, presentations, training materials, and educational activities.
Data Science Education & Training:
- Mentor junior analysts, students, and research staff.
- Contribute to educational and professional development activities within the data science community.
- Provide training on machine learning methodologies and analytical best practices.
- Support workshops, seminars, and collaborative learning initiatives.
Machine Learning & Data Science:
- Lead development of predictive and unsupervised learning models using high-dimensional biomedical datasets.
- Design and implement feature engineering, model training, testing, validation, and monitoring frameworks.
- Develop explainable and reproducible machine learning approaches suitable for clinical applications.
- Perform rigorous comparative analyses and benchmarking of analytical methods.
- Develop scalable computational solutions supporting operational and research objectives.
Methodology Development:
- Prototype and optimize analytical workflows using existing and emerging software tools.
- Establish machine learning pipelines for novel data types and clinical applications.
- Evaluate new computational methods and technologies.
- Develop innovative approaches that advance clinical diagnostics and biomedical research.
Data Management & Modeling:
- Design and maintain data structures supporting large-scale flow cytometry datasets.
- Develop data integration, quality control, and data-governance strategies.
- Build relational and non-relational database solutions supporting analytical workflows.
- Implement robust data pipelines and model monitoring systems.
PREFERRED TECHNICAL SKILLS
- Python (required), including pandas, NumPy, SciPy, Scikit-learn, PyTorch, TensorFlow, XGBoost, and LightGBM.
- R and Bioconductor ecosystem.
- Flow cytometry data analysis tools and standards, including FCS file formats, FlowCore, FlowJo integration, Cytobank, Spectre, FlowSOM, UMAP, and t-SNE.
- Statistical modeling, hypothesis testing, longitudinal analysis, and statistical process control.
- Machine learning model development, validation, deployment, and monitoring.
- Data visualization using Plotly, Dash, Streamlit, Shiny, Tableau, or comparable technologies.
- SQL and NoSQL databases.
- Cloud and high-performance computing environments.
- Git-based version control and software development best practices.
- Experience with MLOps, reproducible research workflows, and containerization technologies, including Docker and Singularity.
- Familiarity with healthcare data, laboratory information systems, and clinical validation practices.
MINIMUM REQUIREMENTS
- Bachelor's degree with 10+ years of relevant post-degree work experience in relevant area (e.g., bioinformatics, cheminformatics, statistics/computer science with a background in biological sciences or chemistry) OR Master's degree with 8+ years of relevant experience OR PhD with 5+ years of relevant experience.
- Substantial experience in at least one programming or scripting language and at least one statistical package, with R preferred.
PREFERRED QUALIFICATIONS
- Experience applying machine learning to biomedical, clinical, translational, or laboratory datasets.
- Experience with high-dimensional single-cell or flow cytometry data.
- Experience developing and validating analytical methods in regulated or clinical laboratory environments.
- Demonstrated record of scientific publication and interdisciplinary collaboration.
- Experience translating research algorithms into operational workflows that improve efficiency, quality, or patient care.
COMPENSATION
In recognition of certain U.S. state and municipal pay transparency laws, St. Jude is including a reasonable estimate of the compensation range for this role. This is an estimate offered in good faith and a specific salary offer takes into account factors that are considered in making compensation decisions including but not limited to skill sets, experience and training, licensure and certifications, and other business and organizational needs. It is not typical for an individual to be hired at or near the top of the salary range and compensation decisions are dependent on the facts and circumstances of each case. A reasonable estimate of the current salary range is $125,840 - $238,160 per year for the role of Data Scientist - Clinical Machine Learning & Flow Cytometry. Explore our exceptional benefits! We are committed to a human-centered hiring experience. Technology may support portions of our process, but recruiting decisions involve human review and engagement. Learn more about our approach to AI.
St. Jude is an Equal Opportunity Employer
No Search Firms
St. Jude Children's Research Hospital does not accept unsolicited assistance from search firms for employment opportunities. Please do not call or email. All resumes submitted by search firms to any employee or other representative at St. Jude via email, the internet or in any form and/or method without a valid written search agreement in place and approved by HR will result in no fee being paid in the event the candidate is hired by St. Jude.
See all 796+ Green Card Data Scientist Jobs
Sign up for free to unlock all listings, filter by visa type, and get alerts for new Green Card Data Scientist Jobs.
Get Access To All JobsTips for Finding Green Card Sponsorship as a Data Scientist
Document your degree field precisely
PERM requires your degree to match the role's specialty occupation definition. A degree in statistics, computer science, or applied mathematics strengthens the match. A business or general IT degree may require a credential evaluation to confirm equivalency.
Target employers with PERM filing history
Search OFLC disclosure data for employers who have filed PERM applications under SOC code 15-2051 (Data Scientists). Employers with repeated filings have established internal processes, which shortens your sponsorship timeline significantly.
Search green card sponsoring roles on Migrate Mate
Filter Data Scientist roles by EB-2 or EB-3 sponsorship availability on Migrate Mate. You'll see which employers have active green card pipelines, so you're not starting a sponsorship conversation from scratch during negotiations.
Verify prevailing wage before accepting an offer
Your offered salary must meet the DOL prevailing wage for your job location and experience level. Use the OFLC Wage Search to look up the Level I through Level IV wage for Data Scientists in your target city before the offer stage.
Ask about the EB-2 NIW option if self-petitioning fits
If your research or modeling work has national economic or scientific impact, a National Interest Waiver lets you skip the PERM step entirely and self-petition. USCIS evaluates this under a three-prong test without requiring a specific employer.
Confirm your employer is E-Verify enrolled before PERM starts
DOL requires employers to post a job notice and run recruitment steps before filing PERM. If your employer isn't E-Verify enrolled, some state-level recruitment requirements may complicate the process. Confirm enrollment status during the offer negotiation stage.
Green Card Data Scientist: Frequently Asked Questions
Does a Data Scientist role qualify for EB-2 or EB-3 green card sponsorship?
Data Scientist positions typically qualify for EB-2 if the role requires a master's degree or equivalent, and EB-3 if a bachelor's degree plus experience suffices. Most employers filing PERM for this role use EB-2 because the job normally requires an advanced degree in statistics, computer science, or a quantitative field. Your employer's HR or immigration counsel determines which category fits the specific job description.
How is the green card process different from H-1B sponsorship for Data Scientists?
H-1B visa is a temporary work visa renewed in three-year increments with no path to permanency on its own. The EB-2 and EB-3 green card process, starting with PERM labor certification, leads directly to lawful permanent residency. There's no annual cap lottery at the EB-3 level for most countries, and EB-2 priority dates for countries outside India and China are often current, meaning shorter waits than many candidates expect.
How long does PERM labor certification take for a Data Scientist position?
DOL currently processes most PERM applications within six to eighteen months, depending on whether the application goes through analyst review or audit. Data Scientist roles sometimes draw audits because the job duties overlap with software development roles, which DOL scrutinizes for accurate classification. After PERM certification, your employer files the I-140 petition with USCIS, and then you file for adjustment of status or consular processing depending on your priority date.
What O*NET classification applies to Data Scientist roles, and why does it matter for PERM?
Data Scientists are classified under O*NET SOC code 15-2051. PERM filings must use the correct SOC code to run compliant recruitment and establish the right prevailing wage level. Using a mismatched code, such as filing under a software developer code for a role focused on statistical modeling, can trigger a DOL audit and delay certification by months. Your employer's legal team should verify the SOC match against your actual job duties before filing.
How do I find Data Scientist jobs where the employer will sponsor a green card?
Migrate Mate lets you search Data Scientist roles filtered specifically by EB-2 and EB-3 green card sponsorship availability. This matters because most general job boards don't distinguish between H-1B-only sponsors and employers willing to run a full PERM process. Filtering upfront saves you from roles where sponsorship conversations stall after an offer, which is one of the most common delays for foreign professionals in quantitative fields.