Skip to main content
Posted August 25, 2026

AI Data Scientist

University of Pittsburgh
Pittsburgh, Pennsylvania, United States 15260 Full-Time
Reference: 286699557


AI Data Scientist

SHRS-Office of the Dean - Pennsylvania-Pittsburgh - (26005315)


The University of Pittsburgh School of Health and Rehabilitation Sciences (SHRS) is a nationally renowned leader in the field of health care education, research, and clinical practice preparation. With 14 different disciplines related to health and rehabilitative care, SHRS shapes future generations of health care professionals—therapists, counselors, advocates, scientists, providers, and practitioners—trained to serve the needs of all people regardless of background, levels of health, or mobility. We are built on a legacy of academic excellence and innovation and fueled by passionate educators and researchers, allowing us to meet the health care and rehabilitation needs of today and drive meaningful change in the future. Learn how bold moves SHRS. https://www.shrs.pitt.edu/about/how-bold-moves-shrs

The University of Pittsburgh School of Health and Rehabilitation Sciences (SHRS), in partnership with Pitt’s Department of Biomedical Informatics (DBMI), is seeking a data scientist to oversee a large, complex data set, the Rehabilitation Datamart With Informatics Infrastructure for Research (ReDWINE). ReDWINE draws electronic medical records data from UPMC’s extensive clinical and payor systems and organizes it for analysis by rehabilitation scientists. The data scientist will work under the direction of Dr. Yanshan Wang, Associate Professor, to utilize advanced programming methods for raw data extraction and aggregation, statistical analysis, and document programming to create analytic data files. He/she will be part of a larger team of SHRS research leaders and Pitt informaticists and information technologies specialists who collaborate on ReDWINE’s development and management. The preferred candidate must have a master’s degree, a strong interest in academic research, and experience in managing large data sets, performing statistical analysis, documentation, and reporting.

Experience and knowledge of skillsets weighted more heavily

Experience with database design, database programming skills, SQL skills

Experience designing ETL solutions

Expert in one or more programing languages (Java and Python preferred)

Experience with data model design, data standards, ontology design is a plus

Experience with data warehousing, Hadoop/data lake platform implementation is a plus

Experience with IT platform implementation in a technical and analytical role is a plus

Experience of writing, submitting, and publishing scientific articles


Job Summary

Oversees large, complex data sets and utilizes advanced programming methods and statistical software for raw data extraction and aggregation, statistical analysis, and document programming to create analytic data files. Creates statistical reports; assists with manuscripts, grants, and regulatory submissions. Maintains data infrastructure and tracks data assets. Adheres to all protocols.


Essential Functions

The candidate will design, implement, and maintain the IT infrastructure using enterprise-level Java programming language. Requires experience in Java, SQL, and infrastructure system design.

Ability to track and finish project milestones, research data warehouse development, and report to the stakeholders.

Responsible for leading the clinical data retrieval, data analysis activities will be conducted using customized Python programming language. In addition, the candidate will also summarize the results in scientific articles and submit to academia conferences and journals.

The candidate will communicate with several stakeholders, including rehabilitation researchers to understand the demands, the Pitt internal R3 office to understand the Neptune data warehouse operation, and health informatics faculties for the overall strategy.


Physical Effort

Little physical effort required. Duties are primarily sedentary. May be required to move objects up to 25 pounds occasionally.



Assignment Category: Full-time regular

Job Classification: Staff.Data Scientist III

Job Family: Research

Job Sub Family: Data Science

Campus: Pittsburgh

Minimum Education Level Required: Master's Degree

Minimum Years of Experience Required: 6

Will this position accept substitution in lieu of education or experience: Combination of education and relevant experience will be considered in lieu of education and/ or experience requirement.

Additional details about Required Licensure/Certification: Master's degree, or equivalent experience, in Computer Science, Engineering, Mathematics, or related field. The successful candidate will design, implement, and maintain the IT infrastructure, primarily Java, Python, and/or C/C++. Requires experience in Java, SQL, Python, and the design of database systems.

Work Schedule: Monday - Friday, 8:30 a.m. - 5:00 p.m

Work Arrangement: Hybrid: Combination of On-Campus and Remote work as determined by the department.

Hiring Range: TBD Based Upon Qualifications

Relocation_Offered: No

Visa Sponsorship Provided: No

Background Check: For position finalists, employment with the University will require successful completion of a background check

Child Protection Clearances: Not Applicable

Required Documents: Resume, Cover Letter

Optional Documents: Not Applicable



Equal employment opportunity, including veterans and individuals with disabilities.

PI286699557

Sign up for Job Alerts