An R-package for daily tasks required to handle biological data as well as avoid re-coding of small functions for quick but necessary data management.
-
Updated
Sep 2, 2026 - R
An R-package for daily tasks required to handle biological data as well as avoid re-coding of small functions for quick but necessary data management.
PubMatrixR is an R package that performs systematic literature searches on PubMed and PMC databases using pairwise combinations of search terms. It creates co-occurrence matrices showing the number of publications that mention both terms from two different sets, enabling researchers to explore relationships between genes, diseases, pathways.
Interface for data stream clustering algorithms implemented in the MOA (Massive Online Analysis) framework.
Cet espace est destiné à contenir les ressources du TD d'Outils d'enquête et d'analyses de données avec R.
Homework code for assignment 3 : K-Means Clustering
Desarrollo de algoritmos de Data Mining para encontrar reglas de asociación
This study involves building a logistic regression model to predict flight delays based on a given dataset. The dataset contains flight information such as origin, destination, carrier, departure time, and day of the week. The objective is to evaluate the model's ability to classify flights as "delayed" or "ontime".
Exploration and modeling of white wine preferences using data mining classification to predict excellence based on physicochemical characteristics.
Predicting bank term deposits using classification ML algorithms.
Analyzed the data and found the main causes of employee attrition. Implemented classification models to predict the employees who are most likely to leave. This information can be used to engage them with different strategies and retain them for a longer time.
Final course project under the JHU data science course. This app uses a predictive text model built from the large corpus data. The model was built using the tidyverse package and n – gram function. The app was built using the Shiny package and it allows user to enter string and app will predict the next word.
Data Mining Mid Term Project is a case-study project that focuses on data analysis. With R, we have done our best to present the Data in a more meaningful and impactful way that could improve the financial/ecommerce/market decision making and others. Dataset is scrapped from https://www.flipkart.com/
🧠️🖥️2️⃣️0️⃣️0️⃣️1️⃣️📜️ The document:article subcategory for AI2001.
🧠️🖥️2️⃣️0️⃣️0️⃣️1️⃣️📃️ The document category for AI2001.
in this project, logistic regression, KNN, classification trees, random forests and neural network were used.
5 analytical tasks have been completed using VAT validated gower-PAM clustering, Correspondence Analysis (CA), Asym-Biplot, Multiple Correspondence Analysis (MCA), Chi-Squared test, Regression, and predictive classification models with KNN, SVM, and Random Forest.
To associate your repository with the datamining topic, visit your repo's landing page and select "manage topics."