loki-mode Agent工作流 是 AI Skill Hub 本期精选AI工具之一。综合评分 8.0 分,整体质量较高。我们强烈推荐将其纳入你的 AI 工具库,帮助提升工作效率。
loki-mode Agent工作流 是一款基于 Python 开发的开源工具,专注于 多智能体、自动化SDLC、工作流编排 等核心功能。作为 GitHub 开源项目,它拥有活跃的社区支持和持续的版本迭代,代码完全透明可审计,支持本地部署以保护数据隐私。无论是个人使用还是集成到企业工作流,都能提供稳定可靠的解决方案。
loki-mode Agent工作流 是一款基于 Python 开发的开源工具,专注于 多智能体、自动化SDLC、工作流编排 等核心功能。作为 GitHub 开源项目,它拥有活跃的社区支持和持续的版本迭代,代码完全透明可审计,支持本地部署以保护数据隐私。无论是个人使用还是集成到企业工作流,都能提供稳定可靠的解决方案。
# 方式一:pip 安装(推荐)
pip install loki-mode
# 方式二:虚拟环境安装(推荐生产环境)
python -m venv .venv
source .venv/bin/activate # Windows: .venv\Scripts\activate
pip install loki-mode
# 方式三:从源码安装(获取最新功能)
git clone https://github.com/asklokesh/loki-mode
cd loki-mode
pip install -e .
# 验证安装
python -c "import loki_mode; print('安装成功')"
# 命令行使用
loki-mode --help
# 基本用法
loki-mode input_file -o output_file
# Python 代码中调用
import loki_mode
# 示例
result = loki_mode.process("input")
print(result)
# loki-mode 配置文件示例(config.yml) app: name: "loki-mode" debug: false log_level: "INFO" # 运行时指定配置文件 loki-mode --config config.yml # 或通过环境变量配置 export LOKI_MODE_API_KEY="your-key" export LOKI_MODE_OUTPUT_DIR="./output"
Most coding agents declare a task done by telling you so in a transcript. The transcript is the agent's own narration; there is nothing to check. Loki Mode takes a different stance: it does not call work done until the work is verified, and every build produces an Evidence Receipt you can re-verify yourself.
The receipt separates two things most tools blur together:
- Facts -- deterministic, non-LLM, and re-derivable by anyone: the git diff (base/head SHAs, file/insertion/deletion counts, a diff_sha256), the test command that ran with its exit code, the build command with its exit code, and each quality-gate verdict. A skeptic can recompute every one of these from the same repo state. - Assessments -- AI judgments such as the review council's verdict. These are labeled explicitly as judgment, not proof, and never make the headline green on their own.
The receipt's headline is computed only from the facts:
- VERIFIED -- tests recorded a real command, ran, and exited 0; the diff is non-empty; nothing was skipped. - VERIFIED WITH GAPS -- some facts checked out, but something was not run or was inconclusive. Every gap is listed by name, so silence never reads as a pass. - NOT VERIFIED -- a test, build, or gate ran and failed (or there was nothing to verify).
This is honesty-of-done, not a claim of perfection. The receipt proves the completion claim is backed by deterministic evidence and is independently re-checkable; it does not claim the generated code is bug-free.
<details> <summary><strong>Verify a receipt yourself -- <code>loki proof</code> commands, tamper/drift checks, proven PRs (advanced)</strong></summary>
_The free, source-available autonomous coding agent by Autonomi. Same Loki CLI, SDK, and MCP for everyone; the commercial editions for teams and enterprises are sold under the Autonomi brand (Autonomi Cloud, Autonomi Enterprise).
Hand it a spec. It does not accept "done" on an empty diff or failing tests.
Website | Documentation | Installation | Changelog
Current release: v9.16.0
</div>
---
bun install -g loki-mode # recommended (npm, Homebrew, Docker below)
<details> <summary>Other install methods</summary>
| Method | Command | Notes |
|---|---|---|
| **Bun (recommended)** | bun install -g loki-mode | Fastest startup for CLI commands. |
| **npm** | npm install -g loki-mode | Works without Bun (bash fallback). Migrate any time with loki self-update --to bun. |
| **Homebrew** | brew tap asklokesh/tap && brew install loki-mode | Auto-installs Bun as a dep. |
| **Docker** | docker pull asklokesh/loki-mode:latest | Bun + Claude CLI pre-installed. See [DOCKER_README.md](DOCKER_README.md). |
Upgrade with loki self-update. Long form: Installation Guide.
</details>
npx loki-mode tour # no install, no API key, no spend, no network
Prints a real Evidence Receipt from a past build, headline and all:
Headline: VERIFIED WITH GAPS
| Fact | Value |
| Files changed | 8 |
| Diff sha256 | c2be6fff3e774c387f276277b25fc424f07b667… |
| Tests | verified (node-test) |
| Build | not_run |
| Security | findings |
| Cost | $10.3218 |
"WITH GAPS" is the point. Build was not run, security has findings, and the receipt says so on its own front page. Recompute the diff hash yourself and check it matches -- you are not asked to trust the agent's self-report.
---
How it works: Drop a spec -- a PRD, GitHub issue, OpenAPI/JSON/YAML, or one-line brief. Loki Mode classifies complexity (run.sh:detect_complexity()), assembles an agent team from 41 specialized agent roles across 8 domains - prompt-defined specifications the orchestrator adopts per phase, with parallel review (blind council) and optional worktree streams on Claude Code, sequential on other providers - and runs autonomous RARV cycles (Reason - Act - Reflect - Verify, seerun.sh:run_autonomous()) with 8 quality gates (seeskills/quality-gates.md). Code is not "done" until it passes automated verification. Output is a Git repo with source, tests, configs, and audit logs.
---
<details> <summary><b>Why verified completion matters</b> -- the failure this exists to fix</summary>
Self-reported completion is the failure users actually hit. A survey of the open issue trackers of seven coding harnesses (OpenHands, Cline, Aider, SWE-agent, Roo-Code, OpenCode, Continue) found the recurring complaint is the agent silently not doing the work -- "always stuck at Preparing write" (opencode#11112, 76 comments), "Continue not making changes to code" (continue#7143), "Agent does not execute functions" (continue#5696). None of those seven publishes a machine-checkable completion artifact.
We measured every named competitor that ships a local CLI -- opencode 1.18.9, aider 0.86.2, codex-cli 0.146.0, Claude Code 2.1.220, cursor-agent -- and none exposes a command that verifies the agent's own output. Rerun it yourself with bash tests/test-competitor-verify-surface.sh.
That is a measurement of the CLI surface, not of whole products: a web UI or an API could expose something --help does not, and Devin and Replit Agent ship no local CLI so they are not covered.
Evaluating this against something else? docs/EVALUATING.md puts a runnable command next to every claim we make, and states plainly what we do not have (no enterprise case studies, no independent benchmark placement, and generation is not air-gapped). It ends with the one question worth asking any agent vendor, including us.
</details>
| Project | Build Time | Complexity |
|---|---|---|
| Landing page with signup form | ~10 min | Simple |
| REST API with JWT auth | ~20 min | Simple |
| Portfolio with animations | ~15 min | Simple |
| SaaS dashboard with analytics | ~25 min | Standard |
| E-commerce store with Stripe | ~45 min | Standard |
| Task manager with kanban board | ~25 min | Standard |
| Chat app with WebSocket | ~30 min | Standard |
| Blog platform with MDX | ~30 min | Standard |
| Microservice architecture | ~2 hours | Complex |
| ML pipeline with monitoring | ~3 hours | Complex |
---
Pass a config file to loki start with --config <path> (aliases: --env-file, --vars), or set LOKI_CONFIG_FILE. The format is detected from the extension or content: .yaml/.yml, .json, or .env (flat LOKI_*=value lines). Values resolve by precedence: a CLI flag beats an ambient env var, which beats the --config file, which beats built-in defaults. Never inline a secret; reference an env var with ${VAR} and the loader expands it at load time (an unset reference is skipped with a warning, and a raw-looking secret literal is flagged). Generate a starter with loki config example.
```bash
dashboard: port: 9000 github: token: ${GITHUB_TOKEN} # expanded from the environment, never stored inline
loki start --config config.yaml ./prd.md
---
<details>
<summary><strong>Configuration env vars (intelligent defaults, opt-out knobs)</strong></summary>
Loki Mode's accuracy and autonomy behaviors are default-on. Each is an opt-out escape hatch, not a setting you have to discover. The most relevant knobs from the v7.41.x accuracy/autonomy hardening:
| Env var | Default | Effect |
|---------|---------|--------|
| `LOKI_REVIEW_INCONCLUSIVE_BLOCK` | `1` | Blocks completion when a code-review round returns zero usable verdicts (an all-empty review proves nothing). Set `0` to record the inconclusive result without blocking. |
| `LOKI_COMPLETION_TEST_CAPTURE` | `1` | Captures fresh test results before the verified-completion evidence gate evaluates. Set `0` to skip the pre-gate capture. |
| `LOKI_AUTO_DOCS` | `true` | Generates the `.loki/docs/` suite before the documentation gate scores it (bounded: once per run when docs are missing, and again only when >10 commits stale). Set `false` to opt out. |
| `LOKI_CAVEMAN` | `1` (on) | Output-token compressor for free-form generation only (never trust-gate subcalls). Set `0` to opt out. |
| `LOKI_CAVEMAN_LEVEL` | inferred | Compression level for the compressor. Auto-inferred per invocation from the run's RARV tier; set explicitly (`lite` / `full` / `ultra`) to override the inference. |
| `LOKI_CONFIDENCE_SPIKE` | `1` (on) | Forces one EXTRA verification pass when the agent's self-reported confidence spikes, instead of trusting the claim. Strictly additive -- it can never skip a gate. Set `0` to opt out; tune with `LOKI_CONFIDENCE_SPIKE_DELTA` (default `40`) and `LOKI_CONFIDENCE_SPIKE_MIN` (default `90`). |
| `LOKI_GOAL_SCORING` | `1` (on) | Flags a goal with no measurable success condition and asks for a threshold, metric, or concrete artifact. Advisory only -- never blocks a build or rewrites the goal. Set `0` to opt out. |
| `LOKI_SMART_RETRY` | `1` (on) | Stops early on a positively-identified permanent failure (bad credentials, unknown model, exhausted quota) rather than burning retries. Unrecognized errors and rate limits still retry as before. Set `0` to retry every failure. |
| `LOKI_SIMPLE` | `0` (off) | EXPERIMENTAL. Strips the coaching half of the system prompt -- the RARV cycle, SDLC phases and memory habits that a frontier model already does natively. Per-iteration state (which gate failed, self-heal output, checklist status) is never touched, because that is information the model cannot derive. Measured at -78% prompt size, ~1562 tokens per iteration, on both the bash and Bun routes. INERT on degraded providers (Codex, Aider): those take an earlier return path whose prompt is already minimal by design, so the flag has nothing to strip there -- a measured zero, not an untested case. Whether it changes build speed or quality is NOT yet measured, so treat it as an experiment, not a tuning knob: run `benchmarks/run-prompt-ablation.sh` on your own workload before adopting it. |
This is a subset. See the [wiki](wiki/Home.md) for the full env-var reference and the RARV-C closure knobs (`LOKI_INJECT_FINDINGS`, `LOKI_OVERRIDE_COUNCIL`, `LOKI_AUTO_LEARNINGS`, `LOKI_HANDOFF_MD`).
</details>
<details>
<summary><strong>BMAD Method Integration</strong></summary>
Loki Mode integrates with the [BMAD Method](https://github.com/bmad-code-org/BMAD-METHOD), a structured AI-driven agile methodology. If your project uses BMAD for requirements elicitation, Loki Mode can consume those artifacts directly:
bash loki start --bmad-project ./my-project
The adapter handles BMAD's frontmatter conventions, FR-format functional requirements, Given/When/Then acceptance criteria, and artifact chain validation. Non-BMAD projects are unaffected -- the integration is opt-in via `--bmad-project`.
See [BMAD Integration Validation](docs/architecture/bmad-integration-validation.md).
</details>
<details>
<summary><strong>Enterprise Features</strong></summary>
Enterprise features are included but require env var activation.
bash export LOKI_ENTERPRISE_AUTH=true # token auth (dashboard/auth.py) export LOKI_OIDC_ISSUER=https://accounts.google.com export LOKI_OIDC_CLIENT_ID=your-client-id # OIDC needs issuer + client id export LOKI_ENTERPRISE_AUDIT=true # force audit logging on export LOKI_TLS_CERT=/path/cert.pem # HTTPS: set BOTH cert and key export LOKI_TLS_KEY=/path/key.pem loki enterprise status ```
Enterprise Architecture | Security | Authentication | Authorization | Metrics | Audit Logging
</details>
<details> <summary><strong>Benchmarks</strong></summary>
Self-reported results from the included test harness. Verification scripts included for reproduction.
| Benchmark | Result | Notes |
|---|---|---|
| HumanEval | 162/164 (98.78%) | Self-reported; harness + results JSON in benchmarks/results/humaneval-loki-results.json. Max 3 retries, RARV self-verification. |
| SWE-bench | Not yet measured | Harness exists and generates patches, but the official SWE-bench evaluator has not been run, so there is no pass-rate to report. Run it yourself: ./benchmarks/run-benchmarks.sh swebench --execute |
See benchmarks/ for methodology.
</details>
<details> <summary><strong>Presentation</strong></summary>

11 slides: Problem, Solution, 41 Agents, RARV Cycle, 8 Quality Gates (HumanEval 98.78%), Multi-Provider, Enterprise Hardening (Live App Preview), Full Lifecycle
</details>
---
export ANTHROPIC_BASE_URL=http://localhost:11434/v1 export LOKI_MODEL_OVERRIDE=<model you have pulled, e.g. the output of ollama list> loki start prd.md
<details> <summary><strong>All commands</strong></summary>
| Command | Description |
|---|---|
loki start [PRD] | Start with optional PRD file (also accepts an issue ref; replaces deprecated loki run). Auto-opens the dashboard in the browser for interactive runs and passes native --effort/--max-budget-usd/--fallback-model for resilience (v7.25.0) |
loki stop | Stop execution |
loki modernize heal <path> | Legacy system healing (archaeology, stabilize, isolate, modernize, validate -- v6.67.0; was: loki heal) |
loki pause / resume | Pause/resume after current session |
loki steer "<note>" | Nudge a running build with a directive (writes .loki/HUMAN_INPUT.md; the loop reads it when LOKI_PROMPT_INJECTION=1) (v8.0.0) |
loki status | Show current status |
loki why | Explain the last outcome; on a stalled run names the real stall reason (proactive stuck-detector + convergence signal) and suggests loki steer (v8.0.0) |
loki cockpit | Live multi-repo status as an inline terminal image (Kitty/iTerm2/WezTerm/Ghostty); text + dashboard fallback elsewhere (v7.126.0) |
loki dashboard | Open web dashboard |
loki preview | Print running app URL and open in browser (Live App Preview, v7.24.0; was: loki open) |
loki web | Launch Purple Lab web UI [DEPRECATED in v7.44.0 -- use loki start which auto-opens the dashboard at http://localhost:57374; for the hosted platform see Autonomi Cloud] |
loki doctor | Check environment and dependencies |
loki plan [PRD] | Pre-execution analysis: complexity, cost, iterations |
loki review [--staged\|--diff] | AI-powered code review with severity filtering |
loki test [--file\|--dir\|--changed] | AI test generation (8 languages, 9 frameworks) |
loki analyze onboard [path] | Project analysis and CLAUDE.md generation (was: loki onboard) |
loki import | Import GitHub issues as tasks |
loki ci | CI/CD quality gate integration |
loki failover | Cross-provider auto-failover management |
loki memory <cmd> | Memory system: index, timeline, search, consolidate |
loki enterprise | Enterprise feature management |
loki version | Show version |
</details>
Run loki --help for all options. Full reference: CLI Reference | Config: config.example.yaml
| Feature | Loki Mode | bolt.new | Replit | Lovable |
|---|---|---|---|---|
| Self-hosted / your keys | Yes | No | No | No |
| Multi-provider failover (5 providers) | Yes | No | No | No |
| 8 quality gates | Yes | No | No | No |
| Blind code review | Yes | No | No | No |
| Enterprise auth (OIDC token + scoped RBAC) | Yes | No | Yes | No |
| Air-gapped deployment | Yes | No | No | No |
| Docker + CI/CD generation | Yes | No | Yes | No |
| Source-available (BUSL-1.1) | Yes | No | No | No |
| Free tier | Source-available | Yes | Yes | Yes |
Among the four tools in this table, Loki Mode is the one that is fully self-hosted, source-available (BUSL-1.1), and includes automated quality verification. Your code, your keys, your infrastructure. We have not surveyed every tool on the market, so read this as a comparison against the named three, not a claim about the whole category.
---
<details> <summary><strong>Provider matrix -- per-provider status, autonomous flags, parallelism, install (includes deprecated Gemini)</strong></summary>
Loki's autonomy and quality loop are the product; the underlying coding CLI is swappable. Loki runs on any of the providers below so you are never locked to one vendor. With LOKI_PROVIDER unset, Loki auto-detects the first installed provider in the order the table lists (claude, cline, codex, aider, opencode); setting it explicitly always wins and is never silently substituted.
| Provider | Status | Autonomous Flag | Parallel Agents | Install |
|---|---|---|---|---|
| **Claude Code** | Active (Tier 1, E2E-verified) | --dangerously-skip-permissions | Yes (10+) | npm i -g @anthropic-ai/claude-code |
| **Cline CLI** | Experimental (Tier 2) | -y | Sequential | npm install -g cline |
| **Codex CLI** | Experimental (Tier 3) | exec --sandbox workspace-write --skip-git-repo-check | Sequential | npm i -g @openai/codex |
| **Aider** | Experimental (Tier 3) | --yes-always | Sequential | pip install aider-chat |
| **opencode** | Experimental | --auto | Sequential | npm install -g opencode-ai |
| **Google Gemini CLI** | REMOVED v7.5.18 | -- | -- | Upstream deprecated; runtime removed. LOKI_PROVIDER=gemini exits with a migration message. |
Status legend: "E2E-verified" means we run real spec-to-code builds on it ourselves. Claude Code is the primary, fully supported provider and the one Loki Mode is built for; it gets full features (subagents, parallelization, MCP, Task tool). "Experimental" means the wiring is in place but we have not produced an end-to-end verified build ourselves; treat as community-tested. Experimental providers run sequentially. Auto-failover switches providers when rate-limited. See Provider Guide.
</details>
---
成熟的多智能体协作框架,完整覆盖SDLC全流程。社区活跃度良好,在自动化开发领域具有实践价值,但需谨慎评估API成本。
该工具使用 NOASSERTION 协议,商用场景请仔细阅读协议条款,必要时咨询法律意见。
AI Skill Hub 为第三方内容聚合平台,本页面信息基于公开数据整理,不对工具功能和质量作任何法律背书。
建议在沙箱或测试环境中充分验证后,再部署至生产环境,并做好必要的安全评估。
📄 NOASSERTION — 请查阅原始协议条款了解具体使用限制。
经综合评估,loki-mode Agent工作流 在AI工具赛道中表现稳健,质量优秀。如果你已有明确的使用需求,可以直接上手体验;如果还在评估阶段,建议对比同类工具后再做决策。
| 原始名称 | loki-mode |
| 原始描述 | 开源AI工作流:Multi-agent autonomous SDLC framework. Spec to deployed app. PRD, GitHub issue, 。⭐929 · Python |
| Topics | 多智能体自动化SDLC工作流编排代码生成CI/CDAI驱动 |
| GitHub | https://github.com/asklokesh/loki-mode |
| License | NOASSERTION |
| 语言 | Python |
收录时间:2026-05-18 · 更新时间:2026-05-19 · License:NOASSERTION · AI Skill Hub 不对第三方内容的准确性作法律背书。