in this repo u can look at default template for ds/ml/dl/.. projects or similar
[0] before creating a new project from this template, u need to install the next dependencies
-
$ brew install -U cookiecutter # or $ pip install cookiecutter -
-
mac os
# install $ brew install github/gh/gh # upgrade $ brew update && brew upgrade gh
-
debian/ubuntu linux
- download the
.debfile from the releases page sudo apt install git && sudo dpkg -i gh_*_linux_amd64.debinstall the downloaded file
- download the
-
[1] after go to the directory where u want to create your project and run
cookiecutter gh:vtrokhymenko/dst├── LICENSE <- will be created if u choose
├── README.md <- the main readme
│
├── config <- often it's yaml-files with some parameters
│
├── data
│ ├── external <- data from third party sources
│ ├── interim <- intermediate data that has been transformed
│ ├── processed <- the final, canonical data sets for modeling
│ ├── raw <- the original, immutable data dump
│ └── features <- another
│
├── docs <- a default sphinx project (see sphinx-doc.org for details)
│
├── experiments <- for any experiments
│
├── models <- trained & serialized models, model predictions, or model summaries
│
├── notebooks <- notebooks for research
│ naming convention is a number (for ordering), the creator's initials, and a short `-`
│ delimited description, eg `1.0-jqp-initial-data-exploration`
│
├── references <- data dictionaries, manuals, and all other explanatory materials
│
├── tests <- test for project
│
├── {{ cookiecutter.repo_name }} <- source code
│ ├── __init__.py <- makes src a python module eg propose generate with `mkinit`
│ │
│ ├── data <- scripts to download or generate data
│ │
│ ├── models <- scripts to train models and then use trained models to make predictions
│ │
│ └── visualization <- scripts to create exploratory and results oriented visualizations
│
├── .gitignore <- default for python
│
└── .pre-commit-config.yaml <- custom pcc with `isort`, `pre-commit-hooks`, `flake8`, `black`- cookiecutter-data-science
- cdst by @crplab
- python-package-template by @TezRomacH
- ocean
- kedro
- dvc – open-source version control system for ds projects
- hydra – to configuring complex applications
- dependabot – automated dependency updates
- pre-commit – framework for managing & maintaining multi-language pre-commit hooks
- coding style/review/formatter
- tests
- spellcheckers
@misc{dst,
author = {viktor trokhymenko},
title = {data science template},
year = {2020},
publisher = {github},
howpublished = {\url{https://github.com/vtrokhymenko/dst}}
}
this project is licensed under the terms of the mit license. see the license file for details