Skip to content
View Xunzhuo's full-sized avatar
🎲
🎲

Block or report Xunzhuo

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Xunzhuo/README.md

Building the Mixture-of-Models for the Next Era of Computing

Pinned Loading

  1. vllm-project/semantic-router vllm-project/semantic-router Public

    A programmable Mixture-of-Models router for heterogeneous LLM inference

    Go 5.9k 975

  2. vllm-project/vllm vllm-project/vllm Public

    A high-throughput and memory-efficient inference and serving engine for LLMs

    Python 92.5k 22.6k

  3. envoyproxy/gateway envoyproxy/gateway Public

    Manages Envoy Proxy as a Standalone or Kubernetes-based Application Gateway

    Go 3k 873

  4. theagentrouter/agent-router theagentrouter/agent-router Public

    Manages Unified Access to Generative AI Services built on Envoy Gateway

    Go 2.1k 384

  5. agentic-in/inferoa agentic-in/inferoa Public

    Inference-native Tokenmaxxing Agent Harness for Loop Engineering

    TypeScript 564 90

  6. agentic-in/elephant-agent agentic-in/elephant-agent Public

    Personal-Model First Self Evolving AI Agent 🐘

    Python 585 65