Commit graph

135 commits

Author SHA1 Message Date
Kaiyi
df04b8d2fb fix(swarm): resolve reassign orphan row, enrich stall context, decision-aware recovery UI 2026-05-29 23:24:26 +08:00
Kaiyi
53753002b0 feat(tui): surface swarm recovery (retrying/dropped) in the dashboard 2026-05-29 23:03:58 +08:00
Kaiyi
f2cc14889b feat(agent-core): swarm coordinator failure-recovery loop (retry/regenerate/reassign/drop) 2026-05-29 22:48:34 +08:00
Kaiyi
60bc6beeee fix(agent-core): remove NUL byte from swarm stall-hook repeat key
The stall-detection repeat key joined the tool name and canonical args
with a literal NUL (0x00) separator. The control byte caused git to
classify stall-hook.ts as binary, so diffs, blame, and code review on
the file were opaque — which prevented confirming the test history for
this feature. Replace the NUL with a normal space (tool names are
identifiers and never contain spaces, so keys stay collision-free) so
the file is plain UTF-8 text and remains reviewable.

Behavior is unchanged: the key still uniquely combines tool name and
canonical args. Verified by reverting the hook to a no-op stub to show
the three stall-detection test files go red (the discriminating block,
canonical-key, e2e turn-abort, and worker-stall-translation cases all
fail), then restoring the real implementation to confirm they pass —
the failing-first the prior atomic commit never recorded.

Full suite: 5049 passed / 25 skipped; make typecheck clean.
2026-05-29 22:38:30 +08:00
Kaiyi
e88003f856 feat(agent-core): stall-detection hard-stop for swarm workers (repeat-based) 2026-05-29 22:31:18 +08:00
Kaiyi
adb68270f6 feat(tui): show live token counts for running swarm workers 2026-05-29 20:06:11 +08:00
Kaiyi
649596b7ce Merge remote-tracking branch 'origin/main' into kaiyi/karachi 2026-05-29 19:56:58 +08:00
liruifengv
96bbc471c4
feat: add experimental feature-flag system (#205)
Introduce a central, env-driven flag registry in agent-core. Each flag is declared once with an id, full env var name, default, and surface. Within agent-core, flags are consulted through a process-global 'flags' constant that reads live process.env. Resolution precedence: master switch KIMI_CODE_EXPERIMENTAL_FLAG > per-feature KIMI_CODE_EXPERIMENTAL_<NAME> > registry default, with lenient boolean parsing via parseBooleanEnv. FlagId is a literal union derived from the registry for compile-time autocomplete and typo-checking.

SDK boundary: KimiHarness.getExperimentalFlags() returns the resolved values over RPC, and the SDK re-exports only the flag *types* — no runtime value crosses the boundary. The TUI caches that snapshot once at startup and reads it synchronously for command gating.

Gate the plugin system behind the 'plugins' flag, off by default: PluginManager.load() consults flags.enabled('plugins'), so when off no installed plugins are loaded or activated, and the TUI /plugins command is hidden from the palette and resolves as an unknown command.

Tests cover the resolver precedence matrix, registry invariants, the FlagId type guard, the live-env singleton, the plugin-load gate, the getExperimentalFlags RPC, and the TUI command gating.
2026-05-29 19:55:10 +08:00
qer
b9860e9f6e
feat: align datasource plugin with generic workflow (#215) 2026-05-29 19:51:23 +08:00
Kaiyi
c03ba22f05 fix(tui): collapse multi-line swarm task to one line in header and tool description 2026-05-29 19:49:21 +08:00
liruifengv
caaa6d83ee
fix(update): don't report success when native update fails (#214)
* fix(update): don't report success when native update fails

The native auto-updater spawned `bash -c "curl -fsSL … | bash"`. A
pipeline's exit status is that of its last command, so when curl could
not connect (e.g. a dead proxy) it produced no output, the trailing bash
read empty stdin and exited 0, and the whole command looked successful —
printing "Updated … Restart the CLI" while nothing had been installed.

Run the spawned shell with `set -o pipefail` so curl's non-zero status
propagates. installUpdate() then rejects and runUpdatePreflight() warns
and continues on the current version instead of claiming success.

* chore: add changeset for native update fix
2026-05-29 19:45:01 +08:00
Kaiyi
0d11fbc097 fix(tui): match swarm card styling to AgentGroup conventions and fix empty task 2026-05-29 19:42:33 +08:00
_Kerman
2388f20bb3
fix: handle structured context overflow errors (#213) 2026-05-29 19:38:24 +08:00
_Kerman
54590d3d46
fix: back off compaction overflow retries by token budget (#211) 2026-05-29 19:29:00 +08:00
Kaiyi
e873370920 fix(tui): render swarm via the managed tool-call lifecycle to stop duplicate cards 2026-05-29 19:24:46 +08:00
Kaiyi
81749b9c51 fix(tui): count only workers in swarm dashboard, finalize on cancel, clean up on reset 2026-05-29 18:54:54 +08:00
Kaiyi
03e49e5640 feat(tui): render swarm runs as a live dashboard instead of nested tool calls 2026-05-29 18:43:18 +08:00
Kaiyi
7ed20f3749 feat(agent-core): emit structured swarm progress (planned/synthesizing/done) 2026-05-29 18:35:18 +08:00
Kaiyi
3475837c9d feat(tui): add SwarmDashboardComponent 2026-05-29 18:32:35 +08:00
Kaiyi
adc18ad512 feat(tui): add swarm dashboard model and reducer 2026-05-29 18:26:57 +08:00
Kaiyi
8021cecfab fix(agent-core): clarify swarm planner tool guidance, add profileOverride test and changeset 2026-05-29 17:36:18 +08:00
_Kerman
e280f33daf
fix: recover from model token limit errors (#207) 2026-05-29 17:27:59 +08:00
Kai
f3269eacb9
fix(tui): show real terminal status for background agents (#197)
* fix(tui): show real terminal status for background agents

The Agent tool's run_in_background=true call returns a non-error
ToolResult whose body just says "status: running". The transcript
card derived its done/failed badge from that result, so every
terminated background agent — including ones reconcile reclassifies
as lost on resume — kept the green "✓ Completed" label even when
the actual task failed, was killed, or never came back.

Push the real BackgroundTaskInfo.status into the matching Agent
card so the badge reflects what happened. The card's resolver
prefers subagent agentId (live) and falls back to the description
on resume; on resume the apply step also runs after replay
finishes so the agent group can reach the borrowed components.

Also adds an agent-core regression test that pins live, busy,
group, race, and resume scenarios for the bg notification chain.

* fix(tui): also propagate bg agent terminal status to standalone cards

Standalone Agent cards (only one Agent tool call in a step, never
upgraded into an AgentGroupComponent) bypassed the previous
`setBackgroundTaskTerminalStatus` path: the standalone header reads
`getDerivedSubagentPhase`, which still derived `done` from the
non-error spawn-success ToolResult, and the method did not request
a header/content rebuild. Lost/failed/killed bg agents in this
shape still rendered as `✓ Completed`.

Thread the override through `getDerivedSubagentPhase`, populate
`subagentError` with the friendly failure message so both render
paths share one source of truth, and trigger the same header +
content rebuild that `onSubagentFailed` does. Also include the
override in `hasSubagentState` / the subagent-block early-return
so a replayed solo bg agent (no replayed subagent block, no
sub-tool activity) switches to the subagent-aware layout instead
of the generic `Used Agent` rendering.

Adds two standalone-render regression tests so the path no longer
relies on the grouped snapshot to stay correct.

* feat(agent-core): make resume actionable from the lost-task notification

A backgrounded subagent that ends as `lost`/`failed`/`killed` is
already a soft-recoverable thing — `subagentHost.resume` will
reanimate the persisted Agent instance — but the LLM had to dig
through the original spawn-success ToolResult to find the right id
and figure out the recovery shape on its own. The two look-alike
identifiers (the BackgroundManager `task_id` aka `source_id`, and
the `subagentHost` `agent_id`) regularly got confused in practice.

Surface what the model needs at the moment of decision:

  - Add `agent_id` as a top-level `<notification>` attribute for
    agent-* tasks, so the right id is structural, not buried in
    prose. Render path keeps backward-compat by omitting the
    attribute when no agent_id is known (bash tasks, old sessions).
  - On non-success agent terminal states, append a recovery
    paragraph to the body: the precise `Agent(resume=...)` call,
    the disambiguation between `agent_id` and `source_id`, the
    `run_in_background` option, and what state survives the
    restart vs. what may need to be redone.
  - Tighten the spawn-time `resume_hint` with the same
    disambiguation and an explicit pointer at the
    `task.lost`/`task.failed`/`task.killed` recovery trigger.
  - Persist `agent_id` and `subagent_type` in PersistedTask so the
    recovery body still works after a session restart, where
    in-memory `BackgroundTaskInfo.agentId` would otherwise be
    undefined. Optional fields keep the disk schema
    forward/backward compatible — pre-PR records load without
    them and silently fall back to the original short body.

* fix(tui): route bg-agent terminal events by stable agent_id, not description

`tc.subagentAgentId` is left undefined for every backgrounded agent.
`handleSubagentSpawned` early-returns for `runInBackground` before
calling `tc.onSubagentSpawned`, and the wire replay path drops the
`subagent` block entirely (`toolCallFromReplayMessage` returns only
id/name/args). So the `agentId` branch in
`applyBackgroundTaskTerminalStatus` never matched in practice, every
call fell through to the description-based fallback, and the
persisted `agent_id` we added in the previous commit was effectively
dead. That fallback also has a real failure mode: if a foreground
Agent and a backgrounded Agent share the same `args.description`,
the only candidate found is the live (unrelated) card, which gets
incorrectly relabeled as the lost task's terminal state.

Parse `agent_id: agent-N` out of the AgentTool spawn-success
ToolResult body inside `getSubagentAgentId` so the id is always
recoverable, regardless of whether the in-memory subagent metadata
was ever populated. Foreground and backgrounded Agent cards now
carry distinct ids and route correctly.

Also pipe the real `subagent.failed` error through to the parent
card. The background branch of `handleSubagentFailed` previously
only appended the dedicated transcript entry; the parent Agent
card was left with the generic "Background agent failed" written
by the later `background.task.terminated` event. Add an optional
`errorText` to `setBackgroundTaskTerminalStatus` /
`applyBackgroundTaskTerminalStatus` and pass `event.error` through
on the failed branch — the real reason now reaches both the card
and the entry.

* fix(tui): treat agent_id as authoritative when matching bg terminal events

Previously `applyBackgroundTaskTerminalStatus` always tried agent_id
first and then fell back to description match on miss. That fallback
caused two cross-card bugs:

  1. On resume, `applyTerminalBackgroundAgentStatuses` iterates every
     persisted terminal task, including ones whose tool calls fell
     outside the `REPLAY_TURN_LIMIT` window and were never mounted.
     Description fallback could route an old `lost` status onto an
     unrelated recent Agent card sharing the same `args.description`.

  2. During the live spawn → terminate window, the same card briefly
     lives in both `_pendingToolComponents` and `transcriptContainer`.
     A description-only walk visits the same component twice and flags
     itself ambiguous, dropping the otherwise unambiguous update.

When `args.agentId` is provided we now match only by id and skip on
miss. With `getSubagentAgentId` already parsing `agent_id: agent-N`
out of the spawn-success ToolResult, the id path is reliable for
both live and resume even though `tc.subagentAgentId` is never
populated for backgrounded agents. Description fallback is preserved
solely for old pre-PR sessions whose persisted records lack
`agent_id` — same best-effort behavior as before.
2026-05-29 17:26:27 +08:00
_Kerman
07d51e4add
chore(agent-core): move tool services type (#206) 2026-05-29 17:20:57 +08:00
Kaiyi
fc5e4bf787 Merge remote-tracking branch 'origin/main' into kaiyi/karachi 2026-05-29 16:51:11 +08:00
liruifengv
14a0348855
fix(tui): avoid leaking footer when resuming a missing session (#202)
* fix(tui): 避免 resume 不存在的 session 时泄漏 footer

启动期间 footer 在构造时就被挂入渲染树,而 init() 早期的 setAppState
会排出一次渲染,在 await 时真正执行(pi-tui 的 stopped 默认为 false,
未 start 也会 doRender),把 footer 画到终端。resume 不存在的 session
时,init() 随后抛错,这次过早渲染就残留在错误信息上方。

将 footer 改为就绪态 chrome:从 buildLayout 移除挂载,改在 initMainTui
中 init() 成功之后再 mountFooter。致命启动错误会在挂载前抛出,footer
不进渲染树,过早渲染只画空树,错误信息落在干净的行上。

* chore: 精简 changeset 描述

* chore: 精简注释

* test: 适配跨工作目录 resume 校验,补全 workDir mock
2026-05-29 15:38:50 +08:00
Kaiyi
d6a3d91c72 fix(agent-core): enforce swarm worker tool allowlist and propagate abort 2026-05-29 15:33:41 +08:00
Kaiyi
b0b61c27ca feat(tui): add /swarm command that triggers the Swarm tool 2026-05-29 15:22:55 +08:00
qer
3da4daeade
fix(kosong): retry when a response stream is terminated mid-flight (#201)
A mid-stream SSE drop surfaces as a raw undici `TypeError: terminated`, which was classified as a non-retryable generic error and failed the turn on the first attempt. Route raw transport-layer errors through the connection-error heuristic so a dropped stream becomes a retryable APIConnectionError and is retried transparently. User aborts (ESC) are unaffected — the retry loop checks the abort signal before retrying.

Related to #149.
2026-05-29 15:18:55 +08:00
Kaiyi
9c309b19ef feat(agent-core): add Swarm tool wired to SwarmCoordinator with recursion guard 2026-05-29 15:18:44 +08:00
Kaiyi
985fd5c6f6 feat(agent-core): add SwarmCoordinator (plan, parallel workers, synthesize) 2026-05-29 15:14:23 +08:00
_Kerman
5159af341c
fix(agent-core): preserve blocked prompt hook context (#200) 2026-05-29 15:12:36 +08:00
Kaiyi
7591e679f5 feat(agent-core): add swarm types and pure plan-parse/concurrency helpers 2026-05-29 15:12:03 +08:00
_Kerman
8913440541
feat: show warning when resuming across working directories (#118) 2026-05-29 15:11:46 +08:00
liruifengv
588145dc9b
feat(tui): expand and prioritize footer rotating tips (#199) 2026-05-29 15:08:34 +08:00
Kaiyi
0406ad0a9a feat(agent-core): support profileOverride for dynamic-role subagents 2026-05-29 15:07:24 +08:00
_Kerman
3a0e06031a
fix: project persisted context messages (#195) 2026-05-29 14:43:17 +08:00
_Kerman
8c77cfab62
fix(agent-core): handle ripgrep cross-device install (#198) 2026-05-29 14:42:56 +08:00
_Kerman
8de720434f
chore: add docs to pnpm-workspace (#196) 2026-05-29 14:34:05 +08:00
liruifengv
64964a0dda
fix(tui): show plan usage as percent used to match web console (#192)
The `/status` and `/usage` "Plan usage" rows previously rendered the
progress bar from the used ratio while labelling it "X% left", so the
bar direction and the number disagreed.

Display "X% used" instead, aligning the number with the bar, and move
the reset hint to the right without parentheses to mirror the web
console layout.
2026-05-29 13:48:30 +08:00
qer
1873859b0e
refactor(agent-core): slim llm request log line (#190)
Merge turnId/step into a single `turnStep` field ("0.1") and
attempt/maxAttempts into `attempt` ("2/3"), and drop the
messageCount/toolCallCount fields. The per-request `llm request`
line goes from up to 8 fields down to ~3; the `llm config` line
(including thinkingEffort, logged for all providers) is unchanged.
2026-05-29 13:36:34 +08:00
_Kerman
564721fe16
fix: clarify subagent and background task stop messages as user-initiated (#189) 2026-05-29 13:22:34 +08:00
_Kerman
537cf20d18
feat: remove default per-turn step limit of 1000 (#186) 2026-05-29 13:11:45 +08:00
Haozhe
c2bd60fce4
ci(release): reorder Nix installation before setup-node (#188)
Nix's profile prepends its bin directory to PATH, shadowing the Node.js version installed by setup-node. Moving Nix installation before setup-node ensures the Node >=24 PATH entry remains first, fixing `ERR_PNPM_UNSUPPORTED_ENGINE` during release.
2026-05-29 12:51:33 +08:00
Haozhe
092a9a8c8d
test(agent-core): use deterministic jitter id in cron pending-jitter test (#187) 2026-05-29 12:42:58 +08:00
_Kerman
114777e859
refactor(agent-core): split RuntimeConfig into Kaos and ToolServices (#185) 2026-05-29 12:28:47 +08:00
Haozhe
3c18987c0b
ci(release): sync Nix pnpmDeps hash during release PR generation (#184)
- Add version:release script combining changeset version with Nix hash update
- Update release workflow to install Nix and use version:release
- Document updated hash refresh workflow in for-agents/workflows.md
- Include changeset for the fix
2026-05-29 12:25:45 +08:00
qer
68df4b8b84
docs: simplify plugins documentation (#169)
* docs: simplify plugins documentation

* docs: restore plugin caveats lost during simplification

Re-add three behaviors that were dropped from the simplified plugins
docs: the stdio MCP `cwd` must start with `./`, local-path installs run
from the managed copy (so editing the source after install requires a
reinstall), and `/plugins remove` only deletes the install record while
leaving files on disk. Mirror the changes in both en and zh.
2026-05-29 02:21:19 +08:00
liruifengv
681ccc5b85
chore(changelog): sync 0.5.0 and reword 'assistant' to 'agent' (#171) 2026-05-28 22:46:50 +08:00
_Kerman
b5981a523b
feat(agent-core): ModelProvider interface and SingleModelProvider (#167) 2026-05-28 22:27:09 +08:00