-
ServusTeq Software Solutions
- Utrecht, the Netherlands
Stars
Solving the card game 6 nimmt! with reinforcement learning
When AI Fails is a project dedicated to documenting the funny, interesting, and sometimes outright stupid ways in which AI can fail.
Code for Go-Explore: a New Approach for Hard-Exploration Problems
Partially Ai partially human written - uses numpy, matplotlib etc to work with position and velocity vectors
The workflow orchestrator core repository
Terraform templates to deploy open data plaform to the cloud.
A lightweight suite of motion imitation methods for training controllers.
Code for the complete guide to tkinter tutorial
A collection of the code I have written for my YouTube tutorials.
Real-time OLTP system for credit card fraud detection using AWS API Gateway, Kinesis, and RDS PostgreSQL. Features a scalable, serverless pipeline for secure and low-latency transaction processing
Collection of reinforcement learning algorithms
A Beginner's Guide to Variational Inference
henanbjut / RL-Adventure-2
Forked from higgsfield-ai/higgsfieldPyTorch0.4 implementation of: actor critic / proximal policy optimization / acer / ddpg / twin dueling ddpg / soft actor critic / generative adversarial imitation learning / hindsight experience re…
Code for the paper "Exploration by Random Network Distillation"
Long-Term Evolution Project of Reinforcement Learning
RLeXplore provides stable baselines of exploration methods in reinforcement learning, such as intrinsic curiosity module (ICM), random network distillation (RND) and rewarding impact-driven explora…
The state-of-art deep rl algorithms for Montezuma's revenge
Code for "Learning to Reach Goals via Iterated Supervised Learning"
ReKep: Spatio-Temporal Reasoning of Relational Keypoint Constraints for Robotic Manipulation
A training framework for Stable Baselines3 reinforcement learning agents, with hyperparameter optimization and pre-trained agents included.
Streamlining reinforcement learning with RLOps. State-of-the-art RL algorithms and tools, with 10x faster training through evolutionary hyperparameter optimization.
rllab is a framework for developing and evaluating reinforcement learning algorithms, fully compatible with OpenAI Gym.
📌 +2500 elo Machine Learning based chess engine + Game Theory Algorithms , made from basic chess rules + neural networks + evaluation function
Biblioteca para Python con algoritmos de resolución de juegos de mesa (minimax, MCTS, SO-ISMCTS y MO-ISMCTS)
The repository hosting example notebooks for 2024 CZII cryoET machine learning challenge
Fault-tolerant, highly scalable GPU orchestration, and a machine learning framework designed for training models with billions to trillions of parameters