Highlights
- Pro
Pinned Loading
-
Neural-Edge-Distiller
Neural-Edge-Distiller PublicSynthetic CoT distillation pipeline — transfers structured reasoning from Llama 3.2 8B to 3B via LoRA fine-tuning on Apple Silicon using MLX. Includes data generation, training, benchmarking, and a…
Python
-
RAG-Projects
RAG-Projects PublicA multi-stage RAG (Retrieval-Augmented Generation) system featuring local LLMs (Ollama/Gemma), hybrid search, and semantic re-ranking with Cross-Encoders.
Python
-
beverina-safety-rag
beverina-safety-rag PublicEnd-to-end LLM system for ingredient safety analysis using RAG, FAISS, and QLoRA with local GGUF deployment.
Python 1
-
speculative-llm-engine
speculative-llm-engine PublicLow-latency LLM inference engine using speculative decoding, optimized for multilingual QA workloads.
Python
If the problem persists, check the GitHub status page or contact support.