Skip to content
Closed
Show file tree
Hide file tree
Changes from all commits
Commits
Show all changes
70 commits
Select commit Hold shift + click to select a range
a796ec9
fix(goal): immediate goal.updated feedback + mode-aware start button
deepagent-ai Jul 11, 2026
5df891b
feat(v4.0): Event Bus + Router/Scheduler + Multi-Agent Runtime (Waves…
deepagent-ai Jul 11, 2026
63413bd
feat(v4.0): Agent Push policy gate + audit log (Wave 4, §B2/§B4)
deepagent-ai Jul 11, 2026
ee22b9e
feat(v4.0): Observability — trace + metrics (Wave 5, §F)
deepagent-ai Jul 11, 2026
a906fed
feat(v4.0): L/M/N event-wiring policy — panel auto-convene + event vo…
deepagent-ai Jul 11, 2026
6294270
test(v4.0): end-to-end integration + migration/rollback safety (Wave …
deepagent-ai Jul 11, 2026
42483b0
feat(v4.0-beta): event-driven archiver (§L) + Approval Queue (§D2)
deepagent-ai Jul 11, 2026
101f19c
feat(v4.0-beta): advertise V4.0 feature flags via /global/capabilitie…
deepagent-ai Jul 11, 2026
c07672b
feat(v4.0-beta): Oversight HTTP routes — metrics + trace + approval q…
deepagent-ai Jul 11, 2026
b45eb76
feat(v4.0-beta): wire Goal Loop lifecycle → Event Bus + Approval Queu…
deepagent-ai Jul 11, 2026
f502245
feat(v4.0-beta): im_messages V4 columns — event_id + delivery_status …
deepagent-ai Jul 11, 2026
cedd82d
feat(v4.0-beta): §B1 double-write — publish im.message.created on use…
deepagent-ai Jul 11, 2026
ca0489c
feat(v4.0-beta): per-workspace config store (retention/quiet-hours/ra…
deepagent-ai Jul 11, 2026
b297d15
feat(v4.0-beta): §E1 security-gate resolvers + §E3 file-path ACL
deepagent-ai Jul 11, 2026
ce02a64
feat(v4.0-beta): §B3 IM thread / direct message / search / file upload
deepagent-ai Jul 11, 2026
bcf653f
feat(v4.0-beta): §A3 event-retention sweep + §B2/§E4 digest builder
deepagent-ai Jul 11, 2026
a183edf
feat(v4.0-beta): §E2 rate limits + §F1 latency metrics + workspace co…
deepagent-ai Jul 11, 2026
f7946e3
feat(v4.0-beta): §M panel auto-convene consumer + §D autonomy escalation
deepagent-ai Jul 11, 2026
c8d5c38
feat(v4.0-beta): start the V4 event-runtime daemons in production (§A…
deepagent-ai Jul 11, 2026
cf64504
fix(v4.0-beta): make the event-runtime actually functional (review BL…
deepagent-ai Jul 11, 2026
fbdb86c
feat(v4.0-beta): default all V4 flags ON — internal test build activa…
deepagent-ai Jul 11, 2026
346de06
fix(v4.0-beta): §H3 default all V4 flags OFF — production customer-fa…
deepagent-ai Jul 11, 2026
f99f825
fix(v4.0-beta): §E2 make the 1000/min publish rate-limit live + close…
deepagent-ai Jul 11, 2026
c1cf0b3
fix(v4.0-beta): §E1 wire four-layer security gate to fail closed in p…
deepagent-ai Jul 11, 2026
bd7d51f
feat(v4.0-beta): §A1 external webhook ingress — git/ci/pr/monitor pro…
deepagent-ai Jul 11, 2026
79e0022
feat(v4.0-beta): §L publish session.completed so the archiver has a t…
deepagent-ai Jul 11, 2026
0ed487f
feat(v4.0-beta): §A4/§N register production schedules — scheduler goe…
deepagent-ai Jul 11, 2026
7f39de4
feat(v4.0-beta): §L/§M wire dead consumers into prod + fix daemon Ins…
deepagent-ai Jul 12, 2026
a58dfa0
feat(v4.0-beta): §B2/§E3/§E4 wire agent proactive push end-to-end (P2.8)
deepagent-ai Jul 12, 2026
1088e85
feat(v4.0-beta): §C3 real multi-agent isolation — file locks + code-g…
deepagent-ai Jul 12, 2026
746ae82
fix(v4.0-beta): §B3 correctness bugs — FTS injection, thread flag-gat…
deepagent-ai Jul 12, 2026
78c5239
feat(v4.0-beta): §D2 Oversight + §B3 IM V4 frontend — the user-visibl…
deepagent-ai Jul 12, 2026
fbdd873
feat(v4.0-beta): §F2 trace back-half + real artifacts + human_takeove…
deepagent-ai Jul 12, 2026
89489ee
feat(v4.0-beta): §A3/§A4/§C1/§E4 observability + limits + quiet-hours…
deepagent-ai Jul 12, 2026
299459b
fix(v4.0-beta): §E3 resolve real project roots for wrk_ workspaces so…
deepagent-ai Jul 13, 2026
4aef5b8
feat(v4.0-beta): §C1/§A1 built-in agent descriptors — autonomous path…
deepagent-ai Jul 13, 2026
2cb4c37
feat(v4.0-beta): §D2 Rollback oversight surface — the last §D2 surfac…
deepagent-ai Jul 13, 2026
79a7c29
feat(v4.0-beta): §E2 real token budget + §C3.2 worktree isolation + p…
deepagent-ai Jul 13, 2026
1c4b03d
fix(v4.0-beta): §A4 event_dropped distinct + §A3 dlq.alert operationa…
deepagent-ai Jul 13, 2026
b120a0e
feat(v4.1): steering foundation — absorb mid-turn user input at the n…
deepagent-ai Jul 13, 2026
571bae9
feat(v4.1): §goal-steer — a running goal absorbs user guidance betwee…
deepagent-ai Jul 13, 2026
f25400f
feat(v4.1): steering ingress — busy sessions absorb messages as steer…
deepagent-ai Jul 13, 2026
3872717
feat(v4.1): §S2 goal plan hot-edit — user revises a running/paused go…
deepagent-ai Jul 13, 2026
2ae8fdc
test(v4.1): §S3.1 goal plan hot-edit cache red-line regression
deepagent-ai Jul 13, 2026
c512c1a
feat(v4.1): §S3.2 frontend — goal plan hot-edit dialog + busy-steer hint
deepagent-ai Jul 13, 2026
93a692e
fix(v4.1): adversarial-review fixes — 5 confirmed defects across S1/S…
deepagent-ai Jul 13, 2026
37096e5
harden(v4.1): flag-gate goal mutation handlers + audit goal governanc…
deepagent-ai Jul 13, 2026
1783c9d
fix(v4.1): plan-gate deadlock — read-only bash exemption + runtime gr…
deepagent-ai Jul 13, 2026
969b0d2
chore(v4.1): T1.5 delete pure dead exports
deepagent-ai Jul 13, 2026
999d6f4
fix(v4.1): T1.1-T1.3 right-panel dead-ends + i18n cleanup
deepagent-ai Jul 13, 2026
6d5c44c
fix(v4.1): T2.4 goal.completed archive contract — carry sessionID + w…
deepagent-ai Jul 13, 2026
6fc3bcf
fix(v4.1): T2.6 delete dead GoalManager.steerGoal + migrate audit to …
deepagent-ai Jul 13, 2026
4f9ab0b
refactor(v4.1): T2.5 delete dead main-session context curator cluster
deepagent-ai Jul 13, 2026
742f3e9
feat(v4.1): T2.2 expose execution-archive read route + T2.3 delete de…
deepagent-ai Jul 13, 2026
736fe93
feat(v4.1): T2.1 wire expert-panel multi-round debate to the UI
deepagent-ai Jul 13, 2026
909a489
perf(v4.1): T4.1-T4.5 index mtime gate, drop metric, cache guard, run…
deepagent-ai Jul 13, 2026
89fe15e
refactor(v4.1): T3.1-T3.4 right-panel — config table + always-on icon…
deepagent-ai Jul 14, 2026
63f36c2
feat(app): separate server MCP and plugin icons
deepagent-ai Jul 14, 2026
f9bc54f
fix(v4.0-beta): close 10 adversarial-audit findings across core + IM
deepagent-ai Jul 14, 2026
28a4504
fix(v4.0-beta): close 8 seam bugs (port/driver contract violations) p…
deepagent-ai Jul 14, 2026
e5214ca
fix(v4.0-beta): prevent two goal drivers racing one run_context doc
deepagent-ai Jul 14, 2026
1837eb3
chore(v4.0-beta): remove dead public DeepAgentCode API (fork-inherited)
deepagent-ai Jul 14, 2026
430a85e
fix(v4.0-beta): eliminate cross-test state pollution in two process-g…
deepagent-ai Jul 14, 2026
826c1db
test(v4.0-beta): update stale assertions + gateway isolation → core s…
deepagent-ai Jul 14, 2026
e54cc26
test(v4.0-beta): fix environmental deepagent-code test failures → sui…
deepagent-ai Jul 14, 2026
8438504
fix(v4.0-beta): resolve full-suite test failures (1 real bug + stale …
deepagent-ai Jul 14, 2026
98f9268
feat(v4.0): Expert Panel three-state composer control (Off / Single /…
deepagent-ai Jul 14, 2026
3c18ecf
fix(app): apply unified icons across workspace chrome
deepagent-ai Jul 14, 2026
ccffa95
chore(desktop): align feedback entry and app version
deepagent-ai Jul 14, 2026
c773062
docs(v4.0): update EN/ZH README for V4.0 autonomous capabilities
deepagent-ai Jul 14, 2026
File filter

Filter by extension

Filter by extension


Conversations
Failed to load comments.
Loading
Jump to
Jump to file
Failed to load files.
Loading
Diff view
Diff view
23 changes: 19 additions & 4 deletions README.md
Original file line number Diff line number Diff line change
Expand Up @@ -6,7 +6,7 @@
</picture>
</p>

<p align="center"><strong>AI coding agent with persistent memory and control plane</strong></p>
<p align="center"><strong>AI coding agent with persistent memory, autonomous goals, and a control plane</strong></p>

<p align="center">
<a href="README.md">English</a> |
Expand Down Expand Up @@ -39,18 +39,30 @@ The features below start from a real need — something a plain coding agent, op

**What DeepAgent does:** Fork any conversation from a chosen message. The fork opens carrying the parent's memory up to that point, shows a full-width "derived from" marker at the top of its transcript, and nests folder-style under its origin in the session tree (subagents and forks alike, up to three levels deep). Knowledge flows up a scope hierarchy — what one session learns can be promoted to the whole project, and cross-project preferences live at the user-global layer — so switching windows is a clean handoff, not a reset.

### Ask roughly, and let the agent sharpen it
### Pick how much autonomy you want, per task

**The need:** A half-formed prompt gets a half-useful answer, and you don't always know how to phrase what you want.
**The need:** Some asks want a fast answer in your exact words; others want the agent to plan first, or to run on its own until the job is done. One fixed behavior can't serve all three.

**What DeepAgent does:** Two scenario modes on the composer. **Direct** sends your prompt as-is — you own the wording. **Intelligence** refines a rough ask into a sharper prompt, surfaces a draft plan and decision suggestions, and waits for your confirmation before it automates anything. You decide how much the agent shapes the request.
**What DeepAgent does:** Three modes on the composer. **Auto** decides how to plan and act on your request. **Design** explores the problem and proposes a design before building anything. **Loop** turns a request into a supervised goal the agent works toward autonomously — iterating plan → execute → verify until the criteria are met — while you stay in control. You choose the autonomy level per task, not once for the whole tool.

### Go deep on genuinely hard problems

**The need:** Complex work — an architecture decision, a tricky migration, a subtle bug — needs more than a single confident pass. It needs research, a second opinion, and someone actively trying to poke holes.

**What DeepAgent does:** At higher work strengths the primary agent decomposes the task, fans it out to focused subagents that research modules in parallel, synthesizes their findings, and then runs independent reviewers whose job is to *break* the plan rather than agree with it. Fan-out is bounded by a configurable concurrency ceiling, and live subagents surface in a session side panel and inline in the transcript so you can watch and jump into any of them.

### Set a goal and let it run — supervised, not unsupervised

**The need:** Some work is a long haul — a migration, a green-the-suite push, a multi-step feature. You want to hand it off and walk away, but "walk away" can't mean "lose control."

**What DeepAgent does:** Loop mode drives an autonomous goal loop against an objectively decidable finish line (tests pass, no diagnostics, reviewer clean, plan complete). It runs plan → execute → verify → iterate in the background, with hard ceilings on ticks, tokens, wall-clock, and cost so it can never run away. A status bar shows live progress; you can hot-edit the plan mid-run, pause and resume from exactly where it stopped, or take over — a takeover pauses autonomy escalation and hands control back to you. A goal with steps that need a human never reports "done"; it routes to you. Autonomy is a dial you hold, not a switch you flip and hope.

### Get a second opinion that actually argues

**The need:** For a high-stakes decision, one confident answer isn't enough — you want independent experts who see the same evidence, disagree, and defend their positions.

**What DeepAgent does:** Convene an Expert Panel on the current conversation from the composer. Pick single-round for a quick multi-lens review, or multi-round for a real debate: panelists (correctness, security, design, …) each render a verdict, then see each other's *anonymized* opinions and revise across up to three rounds — identity stripped so nobody anchors on "the security expert said." An arbiter synthesizes the surviving verdict. Fan-out and rounds are bounded, and every opinion (including the losers') is archived. The panel convenes on demand, or a running goal loop can convene it at high-risk decision points.

### Chat with your team and your agents in one place

**The need:** Coordinating with teammates and driving agents usually happens in two different tools.
Expand All @@ -69,6 +81,8 @@ Each capability above is served by a control-plane primitive underneath. These a

**Self-learning** — After work lands, the agent proposes candidate knowledge, facts, and methodologies. Promotion is evidence-gated (a test passed, a diagnostic cleared, a validation confirmed) and user-controllable — durable knowledge is carried over deliberately, not silently guessed. Session-stable conclusions consolidate into project memory over time, so the next session starts smarter about *your* codebase.

**Supervised autonomy** — The goal loop, expert panel, and multi-agent fan-out run on an event-driven substrate with the guardrails autonomy needs: hard budget ceilings (ticks/tokens/wall-clock/cost), a stall detector that stops rather than spins, layered permission and safety gates that fail closed, human takeover that pauses escalation, and a full audit trail written back to the document graph. An Agent Dashboard surfaces task success rate, conflicts, and dead-letter events. Autonomy is bounded and observable by construction — never a black box you can't stop.

## Installation

```bash
Expand Down Expand Up @@ -123,6 +137,7 @@ On your next session, when you ask to add rate limiting elsewhere, the agent alr
│ • Domain pack system (composable, auto-activating knowledge)│
│ • Context assembly & admission gates │
│ • Multi-agent orchestration & adversarial review │
│ • Supervised goal loop & expert panel (event-driven) │
│ • Evidence-gated learning & work-strength ladder │
└─────────────────────────────────────────────────────────────┘
Expand Down
23 changes: 19 additions & 4 deletions README.zh.md
Original file line number Diff line number Diff line change
Expand Up @@ -6,7 +6,7 @@
</picture>
</p>

<p align="center"><strong>具备持久记忆与控制平面的 AI 编程智能体</strong></p>
<p align="center"><strong>具备持久记忆、自主目标与控制平面的 AI 编程智能体</strong></p>

<p align="center">
<a href="README.md">English</a> |
Expand Down Expand Up @@ -39,18 +39,30 @@ DeepAgent Code 是一个构建在持久文档记忆之上的 AI 编程智能体

**DeepAgent 的做法:** 从任意一条消息分叉当前对话。分叉出的新对话会继承父对话到该点为止的记忆,在时间线顶部显示一条贯穿窗口的"从对话派生"分割线,并像文件夹一样嵌套挂在来源对话之下(子 agent 与分叉同理,最多三层深)。知识沿作用域层级向上流动——单个会话学到的东西可提升到整个项目,跨项目的偏好则沉淀在用户全局层——所以换窗口是一次干净的交接,而非一次清零重来。

### 想到哪问到哪,让智能体替你打磨
### 自主程度,按任务自己挑

**需求:** 半成品的提问只能换来半有用的回答,而你未必总知道该怎么把想要的东西说清楚
**需求:** 有些提问想要一个快速、原话作答;有些希望智能体先规划;还有些希望它自己跑到把活干完为止。一种固定行为伺候不了这三种

**DeepAgent 的做法:** 输入框上有两种情景模式。**直接**模式原样发送你的提示——措辞由你做主。**智能**模式会把粗糙的想法打磨成更精准的提示,给出草拟的方案和决策建议,并在自动执行任何操作前等你确认。智能体替你塑形到什么程度,由你决定
**DeepAgent 的做法:** 输入框上有三种模式。**自动(Auto)** 由智能体决定如何规划并执行你的请求。**设计(Design)** 先探索问题、给出设计方案,再动手构建。**循环(Loop)** 把请求变成一个受监督的目标,智能体朝它自主推进——不断地"规划 → 执行 → 校验"直到满足完成判据——而掌控权始终在你手里。自主程度按任务来选,而不是为整个工具一次性定死

### 对真正的难题深挖到底

**需求:** 复杂的工作——一个架构决策、一次棘手的迁移、一个隐蔽的 bug——需要的不止一次自信的单程作答。它需要调研、需要第二意见、需要有人主动来挑刺。

**DeepAgent 的做法:** 在更高的工作强度下,主智能体会拆解任务,扇出给专注的子 agent 并行调研各个模块,综合它们的发现,再运行独立的审阅者——审阅者的职责是"击破"方案,而不是附和。扇出受可配置的并发上限约束;运行中的子 agent 会出现在会话侧栏面板和时间线内联卡片里,你可以旁观并随时跳进任意一个。

### 定个目标让它自己跑——是受监督,而非放养

**需求:** 有些工作是场持久战——一次迁移、一轮把测试全刷绿、一个多步骤的特性。你想把它交出去然后走开,但"走开"不能等于"失控"。

**DeepAgent 的做法:** 循环模式驱动一个自主目标回路,朝着一条可客观判定的终点线推进(测试通过、无诊断、审阅无异议、计划完成)。它在后台跑"规划 → 执行 → 校验 → 迭代",并对轮次、令牌、墙钟时间和成本设有硬上限,绝不会失控狂奔。状态条实时显示进度;你可以在运行中热编辑计划、暂停后从中断处精确恢复,或直接接管——接管会暂停自主升级、把控制权交还给你。一个包含需要人工处理步骤的目标绝不会谎报"完成",而是转交给你。自主是你握在手里的旋钮,不是拨一下就听天由命的开关。

### 要一个真会争论的第二意见

**需求:** 面对高风险决策,一个自信的答案不够——你想要一批独立专家,看着同样的证据,各持己见、各自辩护。

**DeepAgent 的做法:** 从输入框就当前对话召集专家团。选单轮做一次快速的多视角评审,或选多轮进行真正的辩论:各位专家(正确性、安全、设计……)先各自给出裁定,然后看到彼此**匿名化**的意见并在最多三轮里修正——身份被抹去,谁都无法因"是安全专家说的"而锚定。一位仲裁者综合出最终存活的裁定。扇出与轮数都有界,每一条意见(包括落败的)都会被归档。专家团可按需召集,运行中的目标回路也能在高风险决策点上召集它。

### 团队与智能体,在同一处协作

**需求:** 和队友协调、驱动智能体,通常发生在两个不同的工具里。
Expand All @@ -69,6 +81,8 @@ DeepAgent Code 是一个构建在持久文档记忆之上的 AI 编程智能体

**自学习** — 工作落地后,智能体会提出候选的知识、事实与方法论。晋升是证据门控的(一项测试通过、一条诊断清零、一次校验确认)且由用户掌控——持久知识是被有意结转的,而非后台悄悄猜出来的。会话中稳定的结论会随时间巩固进项目记忆,于是下一次会话对*你的*代码库上手更聪明。

**受监督的自主** — 目标回路、专家团与多智能体扇出都跑在一套事件驱动的底座上,并带着自主所必需的护栏:硬性预算上限(轮次/令牌/墙钟/成本)、宁停不空转的停滞检测、层层失败即关闭的权限与安全门、暂停升级的人工接管,以及写回文档图的完整审计轨迹。一个 Agent 面板会呈现任务成功率、冲突与死信事件。自主在构造上就是有界且可观测的——绝不是一个你停不下来的黑盒。

## 安装

```bash
Expand Down Expand Up @@ -123,6 +137,7 @@ deepagent-code "为 /api/users 端点添加限流"
│ • 领域包系统(可组合、自动激活的知识) │
│ • 上下文装配与准入门 │
│ • 多智能体编排与对抗式审阅 │
│ • 受监督的目标回路与专家团(事件驱动) │
│ • 证据门控的学习 + 工作强度阶梯 │
└─────────────────────────────────────────────────────────────┘
Expand Down
4 changes: 2 additions & 2 deletions bun.lock

Some generated files are not rendered by default. Learn more about how customized files appear on GitHub.

2 changes: 1 addition & 1 deletion packages/app/package.json
Original file line number Diff line number Diff line change
@@ -1,6 +1,6 @@
{
"name": "@deepagent-code/app",
"version": "1.0.0-beta",
"version": "1.4.0",
"description": "",
"type": "module",
"exports": {
Expand Down
Loading
Loading