🚀
Senior Research Engineer @ LawZero | Building the next generation of foundation models.
-
LawZero
- Montreal, Quebec, Canada
- https://alexpalms.github.io/
- in/alessandropalmas
- @alexpalms_
- https://github.com/apalmas-saifh
Stars
Toolkit to package and deploy reinforcement learning environments and agents inside reproducible containers
Reinforcement learning framework for decision-level interception prioritization of drone swarms.
Get up and running with Kimi-K2.6, GLM-5.2, MiniMax, DeepSeek, gpt-oss, Qwen, Gemma and other models.
Distributed Reinforcement Learning accelerated by Lightning Fabric
DIAMBRA Arena: a New Reinforcement Learning Platform for Research and Experimentation