Skip to content
View jie311's full-sized avatar

Block or report jie311

Block user

Prevent this user from interacting with your repositories and sending you notifications. Learn more about blocking users.

You must be logged in to block users.

Maximum 250 characters. Please don’t include any personal information such as legal names or email addresses. Markdown is supported. This note will only be visible to you.
Report abuse

Contact GitHub support about this user’s behavior. Learn more about reporting abuse.

Report abuse
Showing results
Python 32 3 Updated Aug 15, 2026

Offical repository for NeurIPS 2025 paper "From Judgment to Interference: Early Stopping LLM Harmful Outputs via Streaming Content Monitoring".

Python 13 1 Updated May 5, 2026
3 Updated Aug 1, 2026

An interface library for RL post training with environments.

Python 2,502 426 Updated Aug 13, 2026

Claude Code for Financial Market

Python 1,653 276 Updated Aug 16, 2026
Python 10 2 Updated Aug 2, 2026
Python 2 1 Updated Apr 9, 2026
7 1 Updated Aug 11, 2026
Python 13 3 Updated May 5, 2026

CLI security scanner built for the agentic era. Detects CI/CD misconfigs, agent permission risks, MCP tool injection, hardcoded secrets, and DMCA-flagged AI dependencies.

JavaScript 823 105 Updated Aug 15, 2026

“ToolHazard: Scalable Red Teaming of LLM Agents via Adversarial Tool-Interactive Environment and Task Synthesis”

Python 6 1 Updated Aug 13, 2026

AI Security Scanner - Test your AI systems for prompt injection and extraction vulnerabilities

TypeScript 723 109 Updated Jul 9, 2026
Python 6 2 Updated Jul 14, 2026

A debugging framework for agentic AI systems: diagnose failures, attribute root causes, recover with evidence, and validate fixes through reruns.

Python 42 5 Updated Aug 9, 2026
Python 1 1 Updated May 5, 2026

SafeLine is a self-hosted WAF(Web Application Firewall) / reverse proxy to protect your web apps from attacks and exploits.

Go 22,369 1,504 Updated Jul 31, 2026

Efficient Multi-Task Safety Classification for Toxicity, Jailbreaks, Hate Speech, and Harmful Content

Python 24 3 Updated May 29, 2026

Echo Agent 是一个可自托管、长期运行、持续学习的 AI Agent,面向个人与团队的私有自动化场景。它可以部署在自有服务器上,统一连接模型、工具、记忆、权限与消息入口。内置四层认知记忆、遗忘曲线与矛盾检测机制,能够在跨会话任务中持续沉淀上下文,并保持长期记忆的质量。针对命令执行、文件操作等高风险行为,它提供基于 LLM 的审批与解释机制,为关键操作建立可审计、可追溯的安全边界。原生…

Python 900 22 Updated Aug 14, 2026

The official implementation of the work "TopicAttack: An Indirect Prompt Injection Attack via Topic Transition"

Python 5 2 Updated Nov 11, 2025

POC管理平台已集成公开POC仓库内容,可自行添加POC到平台实现集中管理,快速搜索,验证

Python 14 3 Updated Aug 6, 2026

Code for paper "ActBench: Self-Evolving Benchmark of Behavioral Safety in Cowork Agents"

Python 4 2 Updated Aug 15, 2026

FinanceHarness: Autonomous Financial Deep Research Framework

Python 84 12 Updated Aug 4, 2026
Python 1 1 Updated Aug 5, 2026
Python 3 1 Updated Jun 10, 2026

The Official Repository for Paper "HarmfulSkillBench: How Do Harmful Skills Weaponize Your Agents?"

Python 16 2 Updated May 2, 2026

Installable AI coding workflow for risk-based routing, scoped sub-agents, and evidence-backed completion.

Shell 26 3 Updated Aug 12, 2026

A lightweight CodeStable distribution

Python 37 2 Updated Aug 1, 2026

Silero VAD: pre-trained enterprise-grade Voice Activity Detector

Python 9,954 825 Updated Jul 16, 2026
Next