VultrVultr
返回目录

GITHUB TOPIC

ai-safety

6个项目包含此标签

6 个项目

GitHub Topic 精确匹配

安全与治理插件

Second-model AI auto-review for DeepSeek Harness approval requests: a read-only reviewer subagent returns structured allow/deny verdicts with reasons, fail-closed by default, fully auditable from the session log (approval/asked -> autoReview/verdict -> approval/decided).

待结构检查
ai-safetyapprovalauto-reviewcordis
安全与治理插件

Claude Code-style declarative permission rules for DeepSeek Harness: ordered allow/deny/ask rules with tool-name, argument (glob/regex), and workspace-path matching on the tools/pre-execute waterfall, session-log audit, and HMR reload.

待结构检查
ai-safetyallow-deny-askapprovalcordis
安全与治理插件

Deterministic fail-closed tool-call authorization for DSH with evidence: allow/block/ask policy gate plus approval-chain deferral.

待结构检查
agent-safetyagent-toolingai-safetydeepseek-harness
Agent 与会话插件

Execution-time drift firewall for long-running DeepSeek Harness agents. Real-Harness tests: unsafe stale mutations 12/12 native -> 0/12; valid controls 7/7 both; post-SIGKILL unsafe continuation 2/2 -> 0/2.

待结构检查
agent-harnessagent-orchestrationagent-planningagent-safety
模型与 MCP插件

DeepSeek Harness 开发锻造工坊:审批守卫、开发 Skills、GitHub/浏览器能力与 Token Watch 消耗监督,装上就能干活。

待结构检查
ai-agentsai-safetycordisdeepseek-harness
安全与治理插件

让测试绿了不等于你做对了 — 双轴收敛纪律 (spec-gaming orthogonal axis) 的 DeepSeek Harness 原生实现 | Two-axis convergence discipline for coding agents

待结构检查
agent-reliabilityai-safetycoding-agentsdeepseek-harness