Stars
Current and Historical Lists of S&P 500 components since 1996
This is part of the datamule project.
HISTDATA - Dataset composed of all FX trading pairs / Crude Oil / Stock Indexes. Simple API to retrieve 1 Minute data (and tick data) Historical FX Prices (up to date).
Awesome-GraphRAG: A curated list of resources (surveys, papers, benchmarks, and opensource projects) on graph-based retrieval-augmented generation.
This repository contains the Hugging Face Agents Course.
Efficient Triton Kernels for LLM Training
Description Describes the IndicNLP corpus and associated datasets
Measure and visualize machine learning model performance without the usual boilerplate.
The Learning Interpretability Tool: Interactively analyze ML models to understand their behavior in an extensible and framework agnostic interface.
This repository contains code and data download scripts for the paper "Using schema.org annotations for training and maintaining product matchers" by Ralph Peeters, Anna Primpeli, Benedikt Wichtlhu…
Bayesian optimization in PyTorch
Repository for all files and code refered to in the seminar titled 'Neural Networks in Action'
⚡ A Fast, Extensible Progress Bar for Python and CLI
🏄 Scalable embedding, reasoning, ranking for images and sentences with CLIP
GNES is Generic Neural Elastic Search, a cloud-native semantic search system based on deep neural network.
NBoost is a scalable, search-api-boosting platform for deploying transformer models to improve the relevance of search results on different platforms (i.e. Elasticsearch)
VIP cheatsheets for Stanford's CS 229 Machine Learning
Your new Mentor for Data Science E-Learning.
Re-implementation of BiDAF(Bidirectional Attention Flow for Machine Comprehension, Minjoon Seo et al., ICLR 2017) on PyTorch.
You were probably looking for our website... this is it. We moved our website here, so you can see the insides of how we work.
Ultimate Tennis Statistics and Tennis Crystal Ball - Tennis Big Data Analysis and Prediction
Alphabetical list of free/public domain datasets with text data for use in Natural Language Processing (NLP)
A collection of all my datasets