-
National Library of Norway AI-Lab
- Oslo
- http://versae.es
- @versae
Stars
Universal Ethernet Direct Connect (UniEDC)
A command-line interface tool for serving LLM using vLLM.
🧠「大模型」2小时完全从0训练64M的小参数LLM!Train a 64M-parameter LLM from scratch in just 2h!
A React component to make correcting automated transcriptions of audio and video easier and faster. By BBC News Labs. - Work in progress
Curated list of datasets and tools for post-training.
Schedule-Free Optimization in PyTorch
Train GEMMA on TPU/GPU! (Codebase for training Gemma-Ko Series)
Official implementation of 'A Large-Scale Exploration of mu-Transfer' (CoRR 2024)
luweigen / whisper_streaming
Forked from ufal/whisper_streamingWhisper realtime streaming for long speech-to-text transcription and translation
Repo for "Monarch Mixer: A Simple Sub-Quadratic GEMM-Based Architecture"
A minimal PyTorch re-implementation of the OpenAI GPT (Generative Pretrained Transformer) training
hamishivi / EasyLM
Forked from young-geng/EasyLMLarge language models (LLMs) made easy, EasyLM is a one stop solution for pre-training, finetuning, evaluating and serving LLMs in JAX/Flax.
Code and documents of LongLoRA and LongAlpaca (ICLR 2024 Oral)
[NeurIPS'23 Oral] Visual Instruction Tuning (LLaVA) built towards GPT-4V level capabilities and beyond.
Extend existing LLMs way beyond the original training length with constant memory usage, without retraining
[ECCV 2024] codes of DiffBIR: Towards Blind Image Restoration with Generative Diffusion Prior
Multipack distributed sampler for fast padding-free training of LLMs
Making large AI models cheaper, faster and more accessible
Positional Skip-wise Training for Efficient Context Window Extension of LLMs to Extremely Length (ICLR 2024)
Make Praat Picture style plots of acoustic data
This repository contains the official implementation of the research paper, "FastViT: A Fast Hybrid Vision Transformer using Structural Reparameterization" ICCV 2023