Skip to content
View zlwang-cs's full-sized avatar
  • University of California, San Diego
  • La Jolla, California

Block or report zlwang-cs

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Content in all repositories owned by your account will be closed.
Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results

Unofficial source-oriented reconstruction and extension of Grok Bot 0.18.0 for macOS

TypeScript 3,514 3,648 Updated Aug 23, 2026

AgentENV (AENV) is a distributed platform for running agent environments at scale.

Rust 3,523 318 Updated Sep 24, 2026

Open-source AI penetration testing tool to find and fix your app’s vulnerabilities.

Python 64,422 7,051 Updated Sep 24, 2026
JavaScript 25 4 Updated Mar 27, 2026

CVE-Factory

C 183 12 Updated Mar 27, 2026

SecCodeBench is a benchmark suite focusing on evaluating the security of code generated by large language models (LLMs).

Python 133 19 Updated Jun 10, 2026

The awesome collection of OpenClaw skills. 5,400+ skills filtered and categorized from the official OpenClaw Skills Registry.🦞

52,761 5,045 Updated Sep 22, 2026

The AI that really does things. Any OS. Any Platform. The lobster way. 🦞

TypeScript 390,352 82,126 Updated Sep 24, 2026

An agent framework for building and evaluating general digital agents.

Python 43 18 Updated Apr 21, 2026

A lightweight, powerful framework for multi-agent workflows

Python 29,669 4,800 Updated Sep 23, 2026

PatchEval: A New Benchmark for Evaluating LLMs on Patching Real-World Vulnerabilities

Python 231 23 Updated Aug 31, 2026

All-in-One Sandbox for AI Agents that combines Browser, Shell, File, MCP and VSCode Server in a single Docker container.

Python 6,000 541 Updated Sep 14, 2026

MCP-Bench: Benchmarking Tool-Using LLM Agents with Complex Real-World Tasks via MCP Servers

Python 509 69 Updated Oct 7, 2025

Official Implementation of "Emergence of Superposition: Unveiling the Training Dynamics of Chain of Continuous Thought"

Python 7 1 Updated Nov 6, 2025

Post-training with Tinker

Python 4,145 542 Updated Sep 23, 2026
Python 182 19 Updated Nov 24, 2025

Official implementation of "Flow Based Policy for Online Reinforcement Learning"

Python 93 9 Updated Oct 29, 2025

SWE-agent takes a GitHub issue and tries to automatically fix it, using your LM of choice. It can also be employed for offensive cybersecurity or competitive coding challenges. [NeurIPS 2024]

Python 20,391 2,233 Updated Sep 21, 2026

[NeurIPS 2025 D&B Spotlight] Scaling Data for SWE-agents

Python 785 132 Updated Sep 21, 2026

Official implementation of paper "Learning to Optimize Multi-objective Alignment Through Dynamic Reward Weighting"

Jupyter Notebook 28 1 Updated Dec 31, 2025

OSS-Fuzz - continuous fuzzing for open source software.

Shell 12,667 2,906 Updated Sep 22, 2026

The open source coding agent.

TypeScript 209,720 27,682 Updated Sep 24, 2026

Sandboxed code execution for AI agents, locally or on the cloud. Massively parallel, easy to extend. Powering SWE-agent and more.

Python 601 121 Updated Sep 21, 2026

SkyRL: A Modular Full-stack RL Library for LLMs

Python 2,345 434 Updated Sep 24, 2026

[ICLR 2026] End-to-End Reinforcement Learning for Multi-Turn Tool-Integrated Reasoning

Python 405 26 Updated Mar 30, 2026
Jupyter Notebook 414 35 Updated Sep 17, 2025

"DeepCode: Open Agentic Coding (Agent Harness & Loop Engineering & Multi-Agent Orchestration)"

Python 16,631 2,163 Updated Sep 22, 2026

The Entropy Mechanism of Reinforcement Learning for Large Language Model Reasoning.

Python 453 15 Updated Jul 11, 2025

Implementation for FP8/INT8 Rollout for RL training without performence drop.

Python 307 23 Updated Nov 7, 2025
Next