Resume
Work
Experience
April 2024 - current
System and Data Analyst
NorthEast Treatment Centers
-
Expertly develop advanced ad-hoc SQL reports and data exports to support agencies with various data analysis tasks, enabling more informed decision-making and improved operational efficiency. Leverage SQL, Credible BI, and Power BI to create customized reports and dashboards that help programs monitor performance, streamline workflows, and optimize service delivery.
-
Leveraged Freshdesk to support agencies with day-to-day system issues. Collaborate with program teams and third-party vendors to troubleshoot and resolve Credible errors, service disruptions, and urgent requests, ensuring smooth operations.
-
Utilized web services and Python to construct a data pipeline, enabling seamless retrieval of agency data from EHR systems to support healthcare analytics
Jun 2022 - Aug 2022
Data Analyst Intern
Spark Therapeutics Inc| Leading gene therapy company
-
Visualized the end-to-end data analytics operation process across the Spark clinical trial portfolio by leveraging and advancing competencies in Excel, PowerPoint, SAS, and Spotfire.
-
Brainstormed and executed internship project specifications and design with Sr. Biometrics Data Analyst. Presented internship project to both technical and lay audiences at companywide poster session.
-
Streamlined onboarding for clinical trial dashboard users (physicians, scientists) by creating visual dashboards and presentations, and collaborated with senior analysts and data managers to automate and streamline key trial data processes.
Jan 2022 - April 2022
NLP Data Scientist Intern
Claudius Legal Intelligence Premier AI platform founded by NSF&Princeton
-
Supported developing of the legal artificial intelligence platform using natural language processing (NLP) models with Python.
-
Automated linguistic feature processing (named entity extraction, POS tagging, tokenization, lemmatization, entity linking) using spaCy on over 8K legal cases, reducing processing time by 30%.
-
Utilized classification ML models to predict sentiments and polarity of tweets corpus referencing Asian hate crimes.
Jun 2021 - Mar 2024
Biostatistics Data Research Scholar
Drexel University College of Medicine
-
Spearheading cross-functional collaborations, utilizing SQL for targeted patient data extraction, and integrating Python Pandas to manipulate and analyze the data, enhancing the understanding of opioid overdose patterns.
-
Advancing data-driven insights into influenza trends by integrating structured and unstructured data using PySpark, and pioneering automated, efficient, and precise data processing through an ETL process with stored procedures.
Jun 2021 - Aug 2021
Biostatistics Data Research Scholar
Drexel University College of Medicine
-
Consulted the Medical School students to provide a reliable and easy-to-use data
-
analytics toolset to enable hypothesis testing. Synthesized and visualized critical experiment results in the PowerBI dashboard for a medical conference.
-
Applied statistical techniques (binomial regression, T-test, variance) in R to perform deep-dive data analysis on large-scale medical datasets. Translated medical questions into statistical requirements and document processes.
May 2016 - May 2017
Co-Founder of Food Delivery App
JinDouYun
-
Co-founded a food delivery app from the ground up, covering the greater Philadelphi Area. Led app software development using Agile DevOps with IT team, built business partnerships, and managed operations serving a 500+ DAU base.
Education
2020 - Now
Drexel University
Master of Science Degree
Major: Business Analysis and Data Science Double Major
Honors: LeBow Alumni Merit Scholarship
GPA: 3.8
2014 - 2018
Temple University
Bachelor's Degree
Major: Accounting
Skills
& Expertise
Programming & BI: Python (NumPy, Pandas, SciKit Learn, Matplotlib), R, SQL (MySQL), Excel (Advanced), Tableau
Machine Learning: Linear/Logistic Regression, Classification, Clustering, Deep Learning, PCA, Time Series (ARIMA, HoltWinters, SES), NLP (Sentiment Analytics), Tree-based model (Decision Tree, Random Forest)
Productivity: Agile Software DevOps, Scrum Framework, SDLC, MSFT Office Suite
Big Data: Google Cloud, AWS, Spark, ETL Languages: English, Chinese