top of page

Resume

Work
Experience

April 2024 - current

System and Data Analyst

NorthEast Treatment Centers

  • ​Expertly develop advanced ad-hoc SQL reports and data exports to support agencies with various data analysis tasks, enabling more informed decision-making and improved operational efficiency. Leverage SQL, Credible BI, and Power BI to create customized reports and dashboards that help programs monitor performance, streamline workflows, and optimize service delivery.

  • Leveraged Freshdesk to support agencies with day-to-day system issues. Collaborate with program teams and third-party vendors to troubleshoot and resolve Credible errors, service disruptions, and urgent requests, ensuring smooth operations.

  • Utilized web services and Python to construct a data pipeline, enabling seamless retrieval of agency data from EHR systems to support healthcare analytics

Jun 2022 - Aug 2022

Data Analyst Intern

Spark Therapeutics Inc| Leading gene therapy company

  • ​Visualized the end-to-end data analytics operation process across the Spark clinical trial portfolio by leveraging and advancing competencies in Excel, PowerPoint, SAS, and Spotfire.

  • Brainstormed and executed internship project specifications and design with Sr. Biometrics Data Analyst. Presented internship project to both technical and lay audiences at companywide poster session.

  • Streamlined onboarding for clinical trial dashboard users (physicians, scientists) by creating visual dashboards and presentations, and collaborated with senior analysts and data managers to automate and streamline key trial data processes.

Jan 2022 - April 2022

NLP Data Scientist Intern

Claudius Legal Intelligence  Premier AI platform founded by NSF&Princeton

  • Supported developing of the legal artificial intelligence platform using natural language processing (NLP) models with Python.

  • Automated linguistic feature processing (named entity extraction, POS tagging, tokenization, lemmatization, entity linking) using spaCy on over 8K legal cases, reducing processing time by 30%.

  • Utilized classification ML models to predict sentiments and polarity of tweets corpus referencing Asian hate crimes.

Jun 2021 - Mar 2024

Biostatistics Data Research Scholar

Drexel University College of Medicine

  • ​Spearheading cross-functional collaborations, utilizing SQL for targeted patient data extraction, and integrating Python Pandas to manipulate and analyze the data, enhancing the understanding of opioid overdose patterns.

  • Advancing data-driven insights into influenza trends by integrating structured and unstructured data using PySpark, and pioneering automated, efficient, and precise data processing through an ETL process with stored procedures.

Jun 2021 - Aug 2021

Biostatistics Data Research Scholar
Drexel University College of Medicine 

  • Consulted the Medical School students to provide a reliable and easy-to-use data

  • analytics toolset to enable hypothesis testing. Synthesized and visualized critical experiment results in the PowerBI dashboard for a medical conference.

  • Applied statistical techniques (binomial regression, T-test, variance) in R to perform deep-dive data analysis on large-scale medical datasets. Translated medical questions into statistical requirements and document processes.

May 2016 - May 2017

Co-Founder of Food Delivery App
JinDouYun

  • Co-founded a food delivery app from the ground up, covering the greater Philadelphi Area. Led app software development using Agile DevOps with IT team, built business partnerships, and managed operations serving a 500+ DAU base.

Education

2020 - Now

Drexel University
Master of Science Degree

Major: Business Analysis and Data Science Double Major

Honors: LeBow Alumni Merit Scholarship

GPA: 3.8

2014 - 2018

Temple University 
Bachelor's Degree

Major: Accounting 

Skills
& Expertise

Programming & BI: Python (NumPy, Pandas, SciKit Learn, Matplotlib), R, SQL (MySQL), Excel (Advanced), Tableau

Machine Learning: Linear/Logistic Regression, Classification, Clustering, Deep Learning, PCA, Time Series (ARIMA, HoltWinters, SES), NLP (Sentiment Analytics), Tree-based model (Decision Tree, Random Forest)
Productivity: Agile Software DevOps, Scrum Framework, SDLC, MSFT Office Suite

Big Data: Google Cloud, AWS, Spark, ETL Languages: English, Chinese

bottom of page