UC San Diego Health is seeking a Research Data Analyst 2 to support research-focused sequence analysis and genotyping work across high performance computing and cloud environments. The role builds computational pipelines, performs quality control, and supports reporting and documentation under supervision.
Role Overview
Under general supervision, the Research Data Analyst 2 designs and implements computational workflows for sequencing analysis and genotyping for human and model organism data. Responsibilities include routine quality control, organizing and backing up research data, and developing new pipelines by selecting and testing computational approaches. The position also contributes to statistical and research data reporting tasks requiring careful documentation and clear communication.
Key Responsibilities
- Build computational pipelines for sequence analysis and genotyping for humans and model organisms in HPC and cloud compute environments.
- Run established computational workflows using Bash scripting and bioinformatics tools such as PLINK, GATK, HipSTR, and BEAGLE.
- Perform routine quality control on data inputs and analysis outputs, and assist with data organization and backup.
- Develop new pipelines under supervision, including selecting software and testing different computational approaches.
- Support research data reporting assignments with moderate diversity in scope.
- Use judgment within defined practices and policies to select methods and techniques for solutions.
- Perform other duties as assigned.
Required Qualifications
- Six (6) years of related experience, education, or training, or a Bachelor’s degree in a related area plus two years of related experience/training.
- Working knowledge of research function.
- Working skills in statistical analysis, systems programming, database design, and data security measures.
- Working skills in analysis and consultation.
- Ability to communicate complex information clearly and concisely verbally and in writing.
- Knowledge of bioinformatics software related to sequencing and genotyping, with ability to stay current on developments in the field.
- Proven ability to plan and conduct studies related to genotyping and genotyping quality control, including optimizing parameters in bioinformatics computational pipelines.
- Knowledge of Linux-based workflow development, including Bash scripting.
- Extensive experience developing multistep analysis pipelines/workflows using scripting languages including WDL.
- Demonstrated experience handling big data sets and programming in languages including R and Python.
- Knowledge of Linux, High Performance Computing, and cloud technologies such as Amazon EC2, as well as the All of Us Research workbench.
- Proven ability to manage time effectively and complete assigned project components by deadline.
- Excellent communication skills.
- Working knowledge of sequencing and genotyping-related software including Bcftools, Bedtools, FastQC, Plink, and Vcftools.
- Proven ability to document work and annotate code, including experience with revision control systems such as GIT.
- Experience performing genome-wide analysis of tandem repeat variation in human populations.
Technologies
- Bash scripting
- PLINK
- GATK
- HipSTR
- BEAGLE
- WDL
- R
- Python
- Amazon EC2
- All of Us Research workbench
- Bcftools
- Bedtools
- FastQC
- Vcftools
- GIT
- Linux
- High Performance Computing (HPC)
Location and Employment Details
Location: San Diego, CA (hybrid).
Job type: Contract.
Compensation: USD 36 to 50 per hour.
Special Conditions
Employment is subject to a criminal background check.