Jonathan Lee Martin nybblr
- {}+{}; // => NaN
- https://jonathanleemartin.com
Stars
- All languages
- Bikeshed
- Brightscript
- C
- C#
- C++
- CSS
- Clojure
- CoffeeScript
- Crystal
- Dart
- Elm
- Emacs Lisp
- F#
- Gherkin
- Go
- HTML
- Haskell
- Java
- JavaScript
- Jupyter Notebook
- Kotlin
- Lua
- MDX
- Makefile
- Nix
- Objective-C
- Objective-C++
- PHP
- Perl
- PostScript
- Python
- R
- Ruby
- Rust
- SCSS
- Scala
- Shell
- Svelte
- Swift
- TypeScript
- Vim Script
- Vue
- XSLT
- Zig
Moshi is a speech-text foundation model and full-duplex spoken dialogue framework. It uses Mimi, a state-of-the-art streaming neural audio codec.
Jupyter Notebooks as Markdown Documents, Julia, Python or R scripts
The largest collection of PyTorch image encoders / backbones. Including train, eval, inference, export scripts, and pretrained weights -- ResNet, ResNeXT, EfficientNet, NFNet, Vision Transformer (Vβ¦
The open source AI engineering platform for agents, LLMs, and ML models. MLflow enables teams of all sizes to debug, evaluate, monitor, and optimize production-quality AI applications while controlβ¦
A smaller subset of 10 easily classified classes from Imagenet, and a little more French
PyTorch code and models for VJEPA2 self-supervised learning from video.
Experimental llama.cpp fork for inference research and development
Code Editor for the AI Agents Era - Run an army of Claude Code, Codex, etc. on your machine
An LLM-powered knowledge curation system that researches a topic and generates a full-length report with citations.
An autonomous agent that conducts deep research on any data using any LLM providers
Fully local web research and report writing assistant
A Model Context Protocol (MCP) server that provides web search capabilities through DuckDuckGo, with additional features for content fetching and parsing.
The agent that grows with you
KV cache compression via block-diagonal rotation. Beats TurboQuant: better PPL (6.91 vs 7.07), 28% faster decode, 5.3x faster prefill, 44x fewer params. Drop-in llama.cpp integration.
πͺ¨ why use many token when few token do trick β Claude Code skill that cuts 65% of tokens by talking like caveman
Qwen3-omni is a natively end-to-end, omni-modal LLM developed by the Qwen team at Alibaba Cloud, capable of understanding text, audio, images, and video, as well as generating speech in real time.
This is the Personality Core for GLaDOS, the first steps towards a real-life implementation of the AI from the Portal series by Valve.
Qwen3-TTS is an open-source series of TTS models developed by the Qwen team at Alibaba Cloud, supporting stable, expressive, and streaming speech generation, free-form voice design, and vivid voiceβ¦
Kyutai's Speech-To-Text and Text-To-Speech models based on the Delayed Streams Modeling framework.
TTS model capable of streaming conversational audio in realtime.
Community-built comprehensive 2D content creation appplication for graphic design, digital art, and interactive real-time motion graphics powered by a node-based procedural graphics engine
Reverse engineered Linux driver for the FacetimeHD (Broadcom 1570) PCIe webcam