Skip to content
View NgDMau's full-sized avatar
  • Tokyo, JP

Block or report NgDMau

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse

Starred repositories

Showing results

S-Chain: Structured Visual Chain-of-Thought For Medicine

Python 51 1 Updated Feb 10, 2026

Structured Outputs

Python 15,647 862 Updated Aug 16, 2026

CHiME-9 Task 1 - MCoRec baseline

Python 28 6 Updated Jan 13, 2026

[ICML-2025] LLaMa Adaptor and extensions with non-linear prompt

Python 21 Updated Jul 14, 2025

Canberra Vietnamese-English Corpus (Code-switching)

4 Updated Mar 13, 2021

Fast and Accurate ML in 3 Lines of Code

Python 10,609 1,176 Updated Aug 19, 2026

BABILong is a benchmark for LLM evaluation using the needle-in-a-haystack approach.

Jupyter Notebook 253 27 Updated Jun 1, 2026

Doing simple retrieval from LLM models at various context lengths to measure accuracy

Jupyter Notebook 2,367 247 Updated Jun 8, 2026

Agent Laboratory is an end-to-end autonomous research workflow meant to assist you as the human researcher toward implementing your research ideas

Python 5,800 806 Updated Aug 20, 2025
Python 66 3 Updated Jan 27, 2025

A high-throughput and memory-efficient inference and serving engine for LLMs

Python 89,457 20,925 Updated Aug 19, 2026

Scalable data pre processing and curation toolkit for LLMs

Python 1,723 319 Updated Aug 19, 2026

A curated list of recent and past chart understanding work based on our IEEE TKDE survey paper: From Pixels to Insights: A Survey on Automatic Chart Understanding in the Era of Large Foundation Mod…

242 25 Updated Dec 17, 2025

Code and documentation to train Stanford's Alpaca models, and generate the data.

Python 30,245 3,986 Updated Jul 17, 2024

An open platform for training, serving, and evaluating large language models. Release repo for Vicuna and Chatbot Arena.

Python 39,511 4,783 Updated May 1, 2026

🔥 The open-source no-code platform for web scraping, crawling, search and AI data extraction • Turn websites into structured APIs in minutes 🔥

TypeScript 17,189 1,478 Updated Aug 19, 2026

LLM inference in C/C++

C++ 124,706 21,907 Updated Aug 19, 2026

The AI Scientist: Towards Fully Automated Open-Ended Scientific Discovery 🧑‍🔬

Jupyter Notebook 14,426 2,042 Updated Dec 19, 2025

DANeS is an open-source E-newspaper dataset by collaboration between DATASET JSC (dataset.vn) and AIV Group (aivgroup.vn)

Python 1 Updated Dec 7, 2021

A self-organizing file system with llama 3

TypeScript 5,833 385 Updated Aug 8, 2025

All the resources you need to get to Senior Engineer and beyond

18,053 1,671 Updated Aug 8, 2026

An autoregressive character-level language model for making more things

Python 4,193 1,032 Updated Jun 4, 2024

Minimal Implementation of a D3PM in pytorch

Jupyter Notebook 310 25 Updated Apr 22, 2024

Code for building ConceptNet from raw data.

Roff 2,953 357 Updated Jan 19, 2023

Bud500: A Comprehensive Vietnamese ASR Dataset

71 10 Updated Oct 10, 2025

RWKV (pronounced RwaKuv) is an RNN with great LLM performance, which can also be directly trained like a GPT transformer (parallelizable). We are at RWKV-7 "Goose". So it's combining the best of RN…

Python 14,671 1,019 Updated Aug 14, 2026

LLM and Langchain powered chatbot to handle Google Calendar tasks

Python 181 33 Updated Dec 21, 2023

An API, a query engine, and a database schema for genomic sequences; currently with a focus on SARS-CoV-2

Kotlin 33 6 Updated Aug 19, 2026

Simple and efficient pytorch-native transformer text generation in <1000 LOC of python.

Python 6,245 575 Updated Aug 22, 2025
Next