探索 / Terminal Task Agent
Agent 模板L2高风险

Terminal Task Agent

Terminal Task Agent 是一个Agent 模板,适合代码开发、自动化场景。AgentMaps 当前验证到 L2,风险标记为高风险。

目录
公开基线 · 52

由团队整理的公开基线目录,每条都带验证等级与风险标注。 其中 25 条来自自动化管线

52
推荐分
置信度: 中
65
评测基准分
置信度: 中
需要人工复核L2 · Seed profile normalized into AgentMaps schema.
为什么推荐
推荐分 52,适合 代码开发,配置难度为简单。
最适合
Developers using Codex for coding workflows.
不适合
Production write access without sandboxing or human approval.
验证边界
当前验证到 L2;L3/L4 未覆盖,不能视为完整运行验证。
剩余风险
High-risk permissions require explicit user approval and sandboxing.
下一步
先在 sandbox 复核 token、写权限和网络访问,再进入团队流程。
采用前需要复核
该能力涉及高风险、token、写权限、shell 或 auth-gated 未完全测试。L4 只代表接口解析,不是生产安全批准。

适合做什么

这个条目主要面向 Codex 与 Codex, Claude Code, Self-hosted。适合需要先判断平台适配、配置成本、权限风险和验证证据的用户;正式采用前仍应阅读来源文档并复核所需权限。

使用场景

  • - Review code changes
  • - Inspect repository context
  • - Generate implementation notes
  • - Run repeatable task flows
  • - Coordinate tool calls

最适合

  • - Developers using Codex for coding workflows.
  • - Teams that want visible setup, verification, and risk evidence before adoption.

不适合

  • - Production write access without sandboxing or human approval.
  • - Users who need enterprise SSO controls.

限制

  • - Scenario-level L5-L7 benchmark testing is not part of the current MVP record.
AI-native 运行契约Agentic workflow

采用前先看运行方式

AI 可以辅助筛选、解释和生成试用计划,但不得在未复核前执行写入、shell、外部副作用或生产发布。

可见状态

  • 收集任务上下文与平台约束
  • 预览能力配置、来源和权限边界
  • 在 sandbox 或低权限环境试跑
  • 等待人工复核后再继续

用户控制点

  • 修改任务/平台筛选
  • 加入 Compare
  • 打开来源核验
  • 取消高风险采用
  • 提交 pending/staging 线索

批准门禁

  • 本地目录范围确认
  • 写入或状态修改前批准
  • Shell/命令执行 sandbox 批准

失败恢复

  • 试用失败时回到来源文档、替代方案或 pending/staging 提交线索。

试用验收

  • 能否在 2 分钟内找到匹配任务的候选
  • 能否解释推荐理由和不适用边界
  • 能否识别 token、写权限、shell、网络或本地文件风险
  • 能否区分 L1/L2 证据和完整运行验证
  • 试用失败时是否有回退或人工接管路径

验证证据

验证矩阵展示已通过、部分完成、已跳过和未测试的边界;它不是生产采用批准。

L1 元数据

来源、文档、许可证或包信息是否存在。

已通过
L2 静态审计

静态审计已标记权限和风险边界。

已通过
L3 安装路径

安装路径可检查,但不等于已适配你的环境。

已跳过
L4 接口解析

接口或入口已解析;不等于生产安全批准。

已跳过
  • EN · 自动生成Seed profile normalized into AgentMaps schema.
  • EN · 自动生成Source and documentation fields are present.
  • EN · 自动生成Static risk flags are assigned.
  • EN · 自动生成Install verification pending.
  • EN · 自动生成Interface parsing pending.
  • EN · 自动生成Benchmark score capped at 65 by L2 verification.

评估信任档

启发式估算
需要复核static触发风险 medium

风险发现

  • - EN · 自动生成Shell Access
  • - EN · 自动生成Local File Access

已验证证据

  • - EN · 自动生成Seed profile normalized into AgentMaps schema.
  • - EN · 自动生成Source and documentation fields are present.
  • - EN · 自动生成Static risk flags are assigned.
  • - EN · 自动生成Install verification pending.

评分拆解

安全权限39
执行代理证据68
持续维护78
配置集成81
接口调用质量77
行业标准对齐74
任务适配80
跨平台适配77

得分原因

  • - Clear task fit for the selected scenario.
  • - Static verification evidence is available.
  • - Focused platform fit.

注意事项

  • - Scenario testing is still pending.
  • - High-risk permissions need sandboxing.

替代方案

GitHub MCP Server
MCP Server · 代码开发, 自动化
66
推荐

Connect agents to GitHub repositories, issues, pull requests, and code search.

L4高风险Self-hosted
Playwright MCP
MCP Server · 浏览器, 自动化
84
推荐

Expose browser automation primitives through MCP for web navigation and testing.

L4中风险Self-hosted
Context7 MCP
MCP Server · 代码开发, 研究
87
推荐

Fetch current library documentation and examples directly into coding agents.

L4低风险Cursor