Lists (1)
Sort Name ascending (A-Z)
Stars
Orca is the ADE for working with a fleet of parallel agents. Run any coding agent with your own subscription. Available on desktop, mobile and VPS.
AUP Learning Cloud is a customized JupyterHub platform that delivers an intuitive, hands‑on AI learning experience with AMD‑accelerated toolkits.
The Doubleword Inference Stack is the easiest & most performant way to run genAI infrastructure in your private environment.
Robotics Reinforcement learning (RL) Lab on AMD ROCm
An agentic system that auto-optimizes LLM workloads on AMD GPUs.
Turn any codebase, with its docs, SQL schemas, configs, and PDFs, into a queryable knowledge graph. A /graphify skill for Claude Code, Cursor, Codex, and Gemini CLI: local deterministic AST parsing…
MCP server and Claude Code skill for Excalidraw — programmatic canvas toolkit to create, edit, and export diagrams via AI agents with real-time canvas sync.
ExcaliDraw with local support for llms, file managing thanks to ExcaliDash and some custom UI modifications and additions to my own liking including: Localllm support via llamacpp server for mermai…
A comprehensive knowledge management system for Solutions Architects using AI
step by step guides for enterprise AI suite
Claude Code Custom Plugins - Custom plugins for Claude Code CLI
Claude Code skill for generating Excalidraw diagrams
Skill to give Claude Code (and any coding agent) the ability to generate beautiful and practical Excalidraw diagrams.
ArcticInference: vLLM plugin for high-throughput, low-latency inference
This repository provides a set of Ansible playbooks for orchestrating an AMD MI300X GPU in SR-IOV environment using 8 virtual machines, each assigned to a dedicated GPU.
Open-source reporting platform to build and share live dashboards from APIs, SQL and NoSQL databases, with powerful AI assistant, scheduling, and embeddable charts 📈📊
Discovers AMD AINIC's characteristics and labels Kubernetes nodes with the corresponding information
Network Operator simplifies usage of AMD AINICs in a Kubernetes Environment
Instantly calculate the maximum size of quantized language models that can fit in your available RAM, helping you optimize your models for inference.
Open Source Continuous Inference Benchmark Research Platform — Kimi K3 2.8T, MiniMax M3, DeepSeekv4, GLM5 - GB200 NVL72 vs MI355X vs B200 vs GB300 NVL72 & soon™ TPUv6e/v7/Trainium2/3 | 开源持续推理基准研究平台…
Enable true multi gpu capability in Comfy UI using XDiT XFuser and FSDP managed by Ray
Brevitas: neural network quantization in PyTorch