-
Stanford University
- San Francisco
- https://cs.stanford.edu/~minwoos
Stars
Synthetic Patient Population Simulator
FEMR (Framework for Electronic Medical Records) provides tooling for large-scale, self-supervised learning using electronic health records
Custom eval framework for autoregressive VLMs in MMBU
Schema definitions and Python types for Medical Event Data Standard, a standard for medical event data such as EHR and claims data
The simplest, fastest repository for training/finetuning medium-sized GPTs.
MedEvalKit: A Unified Medical Evaluation Framework
minwoosun / open_clip_bmc
Forked from mlfoundations/open_clipAn open source implementation of CLIP.
🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
This open-source curriculum introduces the fundamentals of Model Context Protocol (MCP) through real-world, cross-language examples in .NET, Java, TypeScript, JavaScript, Rust and Python. Designed …
A Python library for extracting structured information from unstructured text using LLMs with precise source grounding and interactive visualization.
Qwen3-VL is the multimodal large language model series developed by Qwen team, Alibaba Cloud.
Open-source evaluation toolkit of large multi-modality models (LMMs), support 220+ LMMs, 80+ benchmarks
An official implementation of "GOAL⚽: Global-local Object Alignment Learning" (CVPR 2025).
All-in-one training for vision models (YOLO, ViTs, RT-DETR, DINOv3): pretraining, fine-tuning, distillation.
📋 A Python Parser for PubMed Open-Access XML Subset and MEDLINE XML Dataset
Repo for our work "Systematic Evaluation of Large Vision-Language Models for Surgical Artificial Intelligence"
llama3 implementation one matrix multiplication at a time
[CVPR 2025] MicroVQA eval and 🤖RefineBot code for "MicroVQA: A Multimodal Reasoning Benchmark for Microscopy-Based Scientific Research" code for MicroVQA benchmark and RefineBot method
A high-throughput and memory-efficient inference and serving engine for LLMs
[CVPR 2025] BIOMEDICA: An Open Biomedical Image-Caption Archive, Dataset, and Vision-Language Models Derived from Scientific Literature
[CVPR 2025] Custom Open CLIP repo to train biomedical CLIP models
Cookiecutter template for Python packages
The official code to build up dataset PMC-OA
UCE is a zero-shot foundation model for single-cell gene expression data
Confidence intervals for the Cox model test error from cross-validation