REPOSITORY OVERVIEWLive repository statistics
★ 1Stars
⑂ 0Forks
◯ 0Open issues
◉ 1Watchers
34/100
OPENREPOHUB HEALTH SIGNALLimited signals
A transparent discovery signal based on current public GitHub metadata.
Recent activity35% weight
30 Community adoption25% weight
0 Maintenance state20% weight
100 License clarity10% weight
0 Project information10% weight
35 This score does not audit code, security, maintainers, documentation quality, or suitability. Verify the repository and its current documentation before adoption.
README preview
Efficient-Data-Preprocessing-with-Scikit-Learn-s-Column_Transformer
This repository provides a practical demonstration of how to streamline the data preprocessing workflow in Python using Scikit-Learn's ColumnTransformer. The project contrasts the verbose, step-by-step manual method of preprocessing with the elegant and efficient Column_Transformer approach.
ALGORITHMICALLY RELATEDSimilar Open-Source Projects
Selected from shared topics, language and repository description—not editorial ratings.
This Repository provides a Jupyter Notebook for building a small language model from scratch using 'TinyStories' dataset. Covers data preprocessing, BPE tokenization, binary storage, GPU memory management, and training a Transformer in PyTorch. Generate sample stories to test your model. Ideal for learning NLP and PyTorch.
81/100 healthRecently updatedActive repository
Jupyter NotebookMIT#gpu-computing#llm#small-language-models#tokenization
⑂ 18 forks◯ 0 issuesUpdated 7 days ago
This repository provides Python Jupyter notebook examples to help users work with VSWIR and TIR data from the EMIT and ECOSTRESS missions.
83/100 healthRecently updatedActive repositoryHas homepage
Jupyter NotebookApache-2.0#ecostress#emit#lpdaac
⑂ 105 forks◯ 0 issuesUpdated 5 days ago
Project homepage ↗