-
Qwen, Alibaba
- Haidian, Beijing
- naykun.github.io
- @Kun11664638
Stars
Warp is an agentic development environment, born out of the terminal.
⚡️ Free Verified HTTP, SOCKS5, & SOCKS4 Proxy List ⏰ Updated every 30 minutes
Algorithm powering the For You feed on X
Modern CUDA Learn Notes with PyTorch for Beginners, 200+ CUDA Kernels, Tensor Cores, HGEMM, FA-2 MMA.
how to optimize some algorithm in cuda.
A tutorial on RDMA based programming using code examples
rCM & Causal-rCM: Leading and Unified Algorithms/Infrastructures for Bidirectional/Autoregressive Video Diffusion Distillation at Scale
nv-one-logger enables tracking of GPU application progress over time and can help to identify overhead from workload and cluster inefficiencies to provide efficiency metrics.
Qwen-Image is a powerful image generation foundation model capable of complex text rendering and precise image editing.
chat log tool, easily use your own chat data. 聊天记录工具,轻松使用自己的聊天数据
https://wavespeed.ai/ Context parallel attention that accelerates DiT model inference with dynamic caching
Evaluating the faithfulness of long-context language models
Iterable datapipelines for pytorch training.
TexTeller can convert image to latex formulas (image2latex, latex OCR) with higher accuracy and exhibits superior generalization ability, enabling it to cover most usage scenarios.
Expert Kit is an efficient foundation of Expert Parallelism (EP) for MoE model Inference on heterogenous hardware
[NeurIPS 2025] An official implementation of Flow-GRPO: Training Flow Matching Models via Online RL
利用 AI 大模型和自动化工作流,根据主题或关键词一键生成高清短视频。Generate HD short videos from a topic or keyword with an automated AI workflow.
One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks
Puzzles for learning Triton
MoBA: Mixture of Block Attention for Long-Context LLMs
🐳 Efficient Triton implementations for "Native Sparse Attention: Hardware-Aligned and Natively Trainable Sparse Attention"
NVIDIA Cosmos is an open platform of world models, datasets, and tools that enables developers to build Physical AI for robots, autonomous vehicles, smart infrastructure, and more.
[ECCV 2024] FreeInit: Bridging Initialization Gap in Video Diffusion Models
「硬地骇客 - 两个月 $12000 ARR 实践之路」是由 硬地骇客 团队编著,本书是关于 Podwise 产品历程的忠实记录:内容包含 灵感 - 构建 - 发布 - 增长 - 复盘 五个章节。如果你觉得一个人读不够过瘾,欢迎加入「硬地骇客」官方知识星球与专家们一起讨论!Podwise 的故事才刚刚开始,我们也将在星球持续分享我们的认知,成功可能无法复制,但失败一定可以借鉴。现在就点击下方…