REPOSITORY OVERVIEWLive repository statistics
★ 828Stars
⑂ 196Forks
◯ 0Open issues
◉ 828Watchers
78/100
OPENREPOHUB HEALTH SIGNALHealthy signals
A transparent discovery signal based on current public GitHub metadata.
Recent activity35% weight
100 Community adoption25% weight
58 Maintenance state20% weight
100 License clarity10% weight
0 Project information10% weight
84 This score does not audit code, security, maintainers, documentation quality, or suitability. Verify the repository and its current documentation before adoption.
README preview
Overview
In this repository, you will find the source code to various projects I have been working on or still work-in-progress. The majority of the projects are accompanied by a Medium blog posts at tuannguyen-doan.medium.com. I published almost exclusively on Towards Data Science publication through Medium's Partnership program so please check out these articles as a way to support me and my future projects. Alternatively, you can also find my blog posts at my personal website here.
My interests lie in the intersection of statistical techniques, data visualization and sports (especially football). All the codes are written entirely in Python or R. I don't have a strong preference or attempt to make a concerted effort to code in a specific language/platform. The decision is mostly based on how specific functionalities needed for a project are supported (scraping in Python and data processing with dplyr piping in R).
I. Statistical application:
The statistics of modern football:
A collection of projects that explore the intricate statistical aspect of the Beautiful Game
Statistical theory and its application:
II. External Collaborations:
Published papers:
III. General tutorials with Python and R:
Data visualization:
Machine Learning practicals:
ALGORITHMICALLY RELATEDSimilar Open-Source Projects
Selected from shared topics, language and repository description—not editorial ratings.
This repository contains a collection of Jupyter notebooks, and code snippets that demonstrate various aspects of data science using Python. The notebooks cover a range of topics, including data cleaning and preprocessing, exploratory data analysis, data visualization, statistical analysis, and machine learning.
27/100 healthActive repository
Jupyter NotebookNo license
⑂ 0 forks◯ 0 issuesUpdated Mar 28, 2023
This repository contains the code for my dissertation "How Classification Models Can Improve Thyroid Disease Diagnosis". Scoring a First at University, it shows data cleaning, preprocessing, and visualisation done within Jupyter Notebook.
38/100 healthActive repository
Jupyter NotebookNo license#data-analysis#data-science#data-visualization#datapreprocessing
⑂ 0 forks◯ 0 issuesUpdated Oct 12, 2025