๐ฏ ๆใๅไธช AI ๅฉๆใๅ็บงๆใ7 ไบบ AI ไธไธๅข้ใ
One task โ Multi-role AI collaboration โ One conclusion | V4.5.1 (Approval Gate + Connector Framework + anti-ghost E2E) | V4.5.0 (cross-session continuity + protocol-native skills) | V4.4.0 (5 enhancement modules)
DevSquad ๆฏไธไธชๅค่ง่ฒ AI ไปปๅก็ผๆๅจใๅฝไฝ ๆไบคไธไธชไปปๅกๆถ๏ผๅฎไธๅๆฏๅไธช AI ๅ็ญ๏ผ่ๆฏ่ฎฉ 7 ไธชไธไธ่ง่ฒ๏ผๆถๆๅธใๅฎๅ จไธๅฎถใๆต่ฏๅใๅผๅ่ ็ญ๏ผๅนถ่กๅไฝ๏ผๆๅ็ปๅบ็ป่ฟๅคๆนๅฎกๆ ธ็็ป่ฎบใ
ไผ ็ป AI: ไฝ โโโ ChatGPT โโโ ไธไธชๅ็ญ๏ผๅฏ่ฝไธๅ
จ้ข๏ผ
DevSquad: ไฝ โโโ DevSquad โโโ [ๆถๆๅธ+ๅฎๅ
จ+ๆต่ฏ+ๅผๅ...] โโโ ๅค็ปดๅบฆๅ
ฑ่ฏ็ป่ฎบ
| ็็น | ไผ ็ปๅ AI | DevSquad |
|---|---|---|
| ่ง่งๅไธ | ๅชๆ้็จ่ง่ง | 7 ไธชไธไธ่ง่ฒๅนถ่กๅฎก่ง โ |
| ่ดจ้ไธๅฏๆง | ๅฏ่ฝ้ๆผๅฎๅ จ้ฎ้ข | ๅค็ปดๅบฆไบคๅ้ช่ฏ + ๅ ฑ่ฏๆบๅถ โ |
| ๆ ๅฎก่ฎก่ฟฝ่ธช | ไธ็ฅ้ๅ็ญไพๆฎไปไน | ๅฎๆดๅฎก่ฎก้พ + SHA256 ๅฎๆดๆงๆ ก้ช โ |
| ๅคๆไปปๅกๅดฉๆบ | ้ฟไปปๅกๅฎนๆไธขๅคฑไธไธๆ | Checkpoint ๆญ็น็ปญไผ + ๅทฅไฝๆตๅผๆ โ |
# ๅฎ่ฃ
pip install devsquad
# ่ฟ่ก - ่ฎฉ AI ๅข้ๅธฎไฝ ่ฎพ่ฎก่ฎค่ฏ็ณป็ป
devsquad run "่ฎพ่ฎกไธไธชๅฎๅ
จ็็จๆท่ฎค่ฏ็ณป็ป" --roles architect,security,tester,coder
# ่พๅบ็ปๆๅๆฅๅ๏ผ
# โ
ๆถๆๅธๅปบ่ฎฎ๏ผ้็จ JWT + Refresh Token ๆนๆก...
# โ
ๅฎๅ
จไธๅฎถๅฎกๆฅ๏ผ้้ฒ่ CSRFใXSSใSQL ๆณจๅ
ฅ...
# โ
ๆต่ฏ็ญ็ฅ๏ผๅๅ
ๆต่ฏ่ฆ็็่พพ 90%+...
# โ
ๅผๅๅฎ็ฐ๏ผๆไพๅฎๆดไปฃ็ ๆกๆถ...
# ๐ ๅ
ฑ่ฏ็ป่ฎบ๏ผๆนๆกๅฏ่ก๏ผ้ฃ้ฉๅฏๆง...| ไฝ ็้ๆฑ | ๆจ่ๆนๆก |
|---|---|
| ็ฎๅ้ฎ็ญ๏ผ"Python ๆไนๅ for ๅพช็ฏ๏ผ"๏ผ | ็ดๆฅ็จ ChatGPT/Claude โ |
| ไปฃ็ ็ๆฎตๅฎกๆฅ | DevSquad ๅ่ง่ฒๆจกๅผ โ |
| ๅคๆ็ณป็ป่ฎพ่ฎก๏ผ้่ฆๅค่ง่ง๏ผ | DevSquad ๅค่ง่ฒๅไฝ ๐ฏ |
| ็ไบง็ฏๅข่ชๅจๅๆต็จ | DevSquad + REST API + Dashboard ๐ฏ |
๐ ๆณๆทฑๅ ฅไบ่งฃ๏ผ โ ๅฎๆดๅฟซ้ๅ ฅ้จๆๅ | 187+ ๆจกๅ่ฏฆ็ปๅ่
๐ ็นๅปๅฑๅผ๏ผๅฎๆดๅ่ฝไป็ปไธๆถๆ่ฏฆ่งฃ
DevSquad V4.5.1 (PATCH release, SemVer compliant) introduces 2 new modules and completes 3 ROADMAP items (V451-1, V451-2, V451-7/8/9). All new modules default to safe, backward-compatible behavior โ no API breaking changes. See docs/release_notes/V4.5.1_RELEASE_NOTES.md for full release notes.
- ApprovalGate: User-level approval mechanism for external operations. Fail-closed on callback exceptions. Auto-approve fallback when no callback configured (backward compatible).
- ConnectorFramework: Protocol-based interface for external system integration (GitHub first).
ConnectorProtocol +GitHubConnector(api/cli/simulation modes).simulation=Trueenforced by default in dispatch pipeline. - V451-7 Dashboard browser-level E2E: 11 AppTest cases (Streamlit AppTest replaces Playwright โ avoids heavy browser deps while still being browser-level DOM simulation)
- V451-8 REST API end-to-end user journey E2E: 190 E2E tests covering dispatchโhistoryโrolesโquick dispatchโerror handlingโlifecycleโcross-entry
- V451-9 Connector Framework anti-ghost E2E: 12 E2E tests (AG-1 through AG-8) proving pipeline activation
DevSquad V4.5.0 (merging V4.4.3 + V4.4.4 + V4.5.0 changes) delivers 10 new features for cross-session continuity, protocol-native skill architecture, and action-first reporting. The 7-role AI team orchestrates complex engineering tasks with full audit trails and consensus mechanisms. See docs/VISION.md for the project vision.
- ScratchpadHistoryStore: SQLite-backed cross-session Scratchpad search
- AgentIdentity: Deterministic agent ID for cross-session tracking
- WorkflowTrace: Transparent workflow trace in dispatch reports
- GitContext: Git branch/commit context injection into dispatch
- SkillProvider Protocol: Protocol-native skill architecture (Builtin + MCP providers)
- OutputStyle: Action-first report format (from i-have-adhd insights)
- SessionResume CLI:
devsquad sessions list+dispatch --resume - FileBundler: Deterministic file bundling for review mode (from open-code-review)
- SKILL.md Modular Split: 1216โ282 lines + 3 reference docs (MODULE_REFERENCE / SUB_SKILLS / VERSION_HISTORY)
- VISION Documents: docs/VISION.md + VISION_ORCHESTRATION.md + VISION_AGENT_COLLABORATION.md
- P0-1 RiskRegister: PMP risk management with 7-role weighted assessment (probability ร impact) + 4 response strategies (avoid/transfer/mitigate/accept) +
GateType.RISK_CHECKgate (exposure โฅ 0.36 blocks) - P0-2 ViewpointRegistry: TOGAF architecture viewpoints with 7-role bound formal viewpoints +
is_orthogonal()orthogonality check +check_consistency()conflict detection - P1-1 ErrorBudgetTracker: SRE error budget with SLO 99.9% default +
GateType.ERROR_BUDGETP10 gate (budget exhaustion blocks deployment) +burn_rate()consumption rate - P1-2 GapAnalyzer: TOGAF gap analysis with
analyze(current, target)+prioritize()+generate_roadmap()+suggest_scheduler_decision()driving LoopScheduler - P2-1 DoraMetricsCollector: DORA metrics (Deployment Frequency / Lead Time / Change Failure Rate / MTTR) +
GateType.DORA_CHECKP11 gate (CFR > 15% triggers architecture review) + Elite/High/Medium/Low rating
- Archived orphan i18n docs (docs/i18n/ โ docs/_archive/i18n/)
- Retired CHANGELOG-CN.md (CHANGELOG.md is now SSOT for all languages)
- Consolidated admin credentials to INSTALL.md only (single source of truth)
- Renumbered INSTALL.md methods to continuous 1-7
- Synced version numbers across all external docs (README/SKILL/INSTALL/CLAUDE)
- Multilingual role prompts (EN/CN/JP) for all 7 roles
- Dashboard 6-tab visibility (Overview/Dispatch/Lifecycle/Metrics/Audit/Settings)
- P2 Kanban evaluation (work-in-progress limits + cycle time tracking)
- P3 ITSM evaluation (incident management + change advisory board simulation)
- 13 E2E tests xpass + anti-ghost counters
Every new module includes _call_counter mechanism + E2E anti_ghost test + CI check_module_activation.py verification. Modules must be truly integrated into dispatch pipeline (not just instantiated), with Markdown report sections user-visible. V4.5.1 extends this pattern from V4.4.0 (RiskRegister / ViewpointRegistry / ErrorBudgetTracker / GapAnalyzer / DoraMetricsCollector) to V4.5.1 (ApprovalGate / ConnectorFramework).
- Contract tests: 5.2% (target โฅ5% โ )
- Integration tests: 15.1% (target โฅ15% โ )
- Total tests: 8392+ (CI authoritative)
- E2E coverage: 107 e2e + 1244 integration + 13 V4.4.0 anti-ghost + 12 V4.5.1 anti-ghost
- V4.3.3: P0-P3 enhancement E2E skeletons (xfail TDD for V4.4.0)
- V4.3.2: LLM vs Mock quality gap measurement (calibration gate + thin-slice probe)
- V4.3.0 Phase 3: Quality hardening + user simulation E2E (NPS 9/10)
- V4.3.0 Phase 2: OutputValidator full integration (LLM output safety detection)
- V4.3.0 Phase 1: DependencyHallucinationChecker (anti-slopsquatting supply chain attack)
- V4.3.0 Phase 0: DeploymentComplianceChecker (anti-violation deployment backstop)
- V4.0.0 P1-1 Loop Engineering: Discovery โ Handoff โ Verification โ Persistence โ Scheduling
- V4.0.0 P1-2 UI/UX Patrol: 4-dimension audit + PIL pixel diff visual regression
- V4.0.0 P2-1 Adversarial Verification: red team attack + blue team defense + judge arbitration
- V4.0.0 P2-2 DAG Visualization: Mermaid / JSON / DOT three formats
- V4.0.0 P3-1 Autonomous: plan โ dev โ verify โ fix 4-stage autonomous iteration
- V4.0.0 P3-2 Plugin Hot-Loading: 3 loading paths + path traversal 3-layer protection + reload rollback
8392+ tests passing (CI authoritative).
DevSquad is registered as a TRAE Skill. Simply describe your task in the TRAE IDE chat, and the 7-role team will collaborate automatically. No CLI or API setup needed.
# Interactive setup wizard (1-2 minutes)
python scripts/cli.py init
# Then start collaborating!
devsquad dispatch -t "your task description"# Start MCP server with stdio transport (for IDE integration)
python3 scripts/mcp_server.py
# Or SSE transport (for remote access)
python3 scripts/mcp_server.py --port 8080# Start Streamlit dashboard with authentication
streamlit run scripts/dashboard.py
# Open http://localhost:8501
# Login with default dev credentials (see INSTALL.md "Default credentials" section).
# Change all defaults in production.# Install dependencies
pip install fastapi uvicorn
# Start API server
uvicorn scripts.api_server:app --host 0.0.0.0 --port 8000 --reload
# Access Swagger UI: http://localhost:8000/docs
# Access ReDoc: http://localhost:8000/redocfrom scripts.collaboration.dispatcher import MultiAgentDispatcher
dispatcher = MultiAgentDispatcher()
result = dispatcher.dispatch(
task="Optimize database query performance",
roles=["architect", "security", "tester"],
)
print(result.report)
print(result.consensus)# One-click startup โ 4 phases: env check โ DB init โ frontend build โ service start
./scripts/start.sh
# Launch Streamlit dashboard instead of API server
./scripts/start.sh --dashboard
# Override API port
DEVSQUAD_API_PORT=9000 ./scripts/start.sh
# Show help
./scripts/start.sh --helpstart.sh is the unified entry point introduced in V3.9.2 (P0-2). It validates the environment, initializes the database, builds the frontend, and starts the service in one command. Use requirements.lock alongside it for reproducible builds (pip install -r requirements.lock). V4.1.0 adds Loop Engineering, UI/UX ๅทกๆฃ, Adversarial ้ช่ฏ, DAG ๅฏ่งๅ, Autonomous, and ๆไปถ็ญๅ ่ฝฝ.
| Role | CLI ID | Aliases | Weight | Best For |
|---|---|---|---|---|
| ๐๏ธ Architect | arch |
architect |
1.5 | System design, tech stack, performance/security architecture |
| ๐ Product Manager | pm |
product-manager |
1.2 | Requirements, user stories, acceptance criteria |
| ๐ก๏ธ Security Expert | sec |
security |
1.1 | Threat modeling, vulnerability audit, compliance |
| ๐งช Tester | test |
tester, qa |
1.0 | Test strategy, quality assurance, edge cases |
| ๐ป Coder | coder |
solo-coder, dev |
1.0 | Implementation, code review, performance optimization |
| ๐ง DevOps | infra |
devops |
1.0 | CI/CD, containerization, monitoring, infrastructure |
| ๐จ UI Designer | ui |
ui-designer |
0.9 | UX flow, interaction design, accessibility |
Auto-match: If no roles specified, the dispatcher automatically matches based on task keywords.
DevSquad's 235 modules are organized into 5 capability domains, each solving a specific problem:
่ฎฉ 7 ไธช่ง่ฒ้ซๆๅไฝ็ใๆๆฅไธญๅฟใ
| Module | Purpose | When to Use |
|---|---|---|
| MultiAgentDispatcher | Unified dispatch entry point | All tasks automatically |
| Coordinator | Task decomposition + role assignment | Complex tasks needing breakdown |
| Scratchpad | Shared blackboard for real-time info exchange | Inter-role collaboration |
| ConsensusEngine | Weighted voting + veto + escalation mechanism | Security/architecture disputes |
| BatchScheduler | Parallel/sequential hybrid scheduling | Resource-constrained environments |
Core Workflow:
User Task โ [InputValidator] โ [RoleMatcher] โ [Coordinator Orchestration]
โ [ThreadPoolExecutor Parallel Workers] โ [Scratchpad Real-time Sharing]
โ [ConsensusEngine] โ [ReportFormatter] โ [Structured Report]
้ฒๆญข AI ใๅทๆใๆใๅนป่งใ
| Module | Purpose | When to Use |
|---|---|---|
| InputValidator | Security validation + 40-pattern detection (14 forbidden + 21 prompt injection + 5 suspicious) | Production environments |
| VerificationGate | Mandatory evidence requirements + 7 Red Flags detection | Critical decision scenarios |
| AntiRationalizationEngine | Per-role excuseโrebuttal tables to prevent quality shortcuts | High quality requirements |
| TestQualityGuard | Test quality audit (API validation / anti-pattern detection / dimension coverage) | Pre-release verification |
| PermissionGuard | 4-level safety gate (PLAN/DEFAULT/AUTO/BYPASS) | Security-sensitive tasks |
่ฎฉ็ณป็ปๆดๅฟซใๆด็จณๅฎใๆด็้ฑ
| Module | Purpose | When to Use |
|---|---|---|
| LLMCache | TTL-based LRU cache with disk persistence (60-80% cost reduction) | High-frequency usage |
| LLMRetry | Exponential backoff + circuit breaker + multi-backend fallback | Unstable networks |
| FeedbackControlLoop | Closed-loop feedback control with automatic iteration until quality threshold met | High quality output pursuit |
| ExecutionGuard | Real-time abort guard (timeout/output/keywords) for safe execution | Long-running tasks |
| FallbackBackend | Automatic backend failover with health monitoring | High availability requirements |
็ฅ้็ณป็ปๅจๅไปไนใๅๅพๆไนๆ ท
| Module | Purpose | When to Use |
|---|---|---|
| PerformanceMonitor | P95/P99 response time, CPU/memory tracking, bottleneck detection | Performance tuning |
| UsageTracker | Token/cost usage tracking and reporting | Cost control |
| AuditLogger | SHA256 integrity operation logs with CSV/JSON export (Preview) | Compliance auditing |
| RBAC Engine | 15+ fine-grained permissions, 5 roles (SUPER_ADMIN/ADMIN/OPERATOR/ANALYST/VIEWER) (Preview) | Enterprise access control |
| Multi-Tenancy Manager | 3 isolation levels (strict/moderate/shared), tenant-scoped resources (Preview) | Multi-tenant SaaS |
| Sensitive Data Masker | PII detection and masking (email/phone/ID card/credit card), configurable rules (Preview) | Data compliance |
| HistoryManager | SQLite time-series storage: metrics snapshots, alert history, API logs | Retrospective analysis |
่ๅ ฅไฝ ็็ฐๆๅทฅๅ ท้พ
| Module | Purpose | When to Use |
|---|---|---|
| CLI | Command-line interface with lifecycle commands | Daily developer usage |
| REST API (FastAPI) | 10+ endpoints with OpenAPI/Swagger docs | Microservice integration |
| Dashboard (Streamlit) | Interactive web dashboard with authentication | Operations team visualization |
| MCP Protocol | Integration with TRAE/Claude Code/Cursor | AI Agent ecosystem |
| Docker Support | Multi-stage build for production deployment | Containerized environments |
| GitHub Actions CI | Python 3.10-3.11 matrix testing | CI/CD pipelines |
้ไพตๅ ฅๅผๅ ่ฃ ่ฎพ่ฎก โ ๅฏ้ๅผๅ ณ๏ผ้ถไฟฎๆน็ฐๆๆ ธๅฟ้ป่พ
The 5 cybernetic modules work independently or together without modifying existing core logic:
User Task
โ
[SimilarTaskRecommender] โ Optional: suggest roles from history
โ
[AdaptiveRoleSelector] โ Optional: optimize role selection
โ
[MultiAgentDispatcher]
โ
[FeedbackControlLoop] โ Wrap dispatcher for auto-iteration
โ [each worker step]
[ExecutionGuard] โ Guard each worker execution
โ
[PerformanceFingerprint] โ Record after dispatch completes
- Closed-loop feedback control with automatic iteration until quality threshold met
- Configurable quality gate (
quality_gate) and maximum iterations - Lightweight quality assessment (no LLM calls), supports dry-run mode
- Real-time execution monitoring with 4 abort conditions: timeout, output size, token count, critical keywords
- Lightweight checks (<1ms), zero external dependencies
- Dynamically configurable thresholds
- Unified execution fingerprint recording (fuses 4 data sources)
- Pure Python TF-IDF implementation (no sklearn/numpy), supports English/Chinese mixed content
- JSON persistence to
.devsquad_data/fingerprints/, graceful cold-start degradation
- TF-IDF-based task similarity search with historical success configuration recommendations
- Intelligent role combination recommendation, intent prediction, execution time estimation
- Confidence scoring (high/medium/low), graceful cold-start degradation
- Three-tier selection strategy based on historical success rates
- Configurable minimum success rate and maximum role count
- Supports manual statistics updates and comprehensive role effectiveness reporting
Recommended usage (progressive adoption):
from scripts.collaboration import (
MultiAgentDispatcher, FeedbackControlLoop,
ExecutionGuard, PerformanceFingerprint
)
dispatcher = MultiAgentDispatcher()
guard = ExecutionGuard()
fingerprint = PerformanceFingerprint()
# Option 1: Full cybernetics stack
loop = FeedbackControlLoop(dispatcher, quality_gate=0.7)
result = loop.run("Your task here")
# Option 2: Guard only (minimal adoption)
result = dispatcher.dispatch("Your task")
for w in result.worker_results:
abort, reason = guard.check_abort(w.output, w.duration)
if abort:
print(f"Aborted: {reason}")
# Option 3: Learning only
fingerprint.record_execution("task", result, result.timing, result.matched_roles)
similar = fingerprint.find_similar("new task", top_k=3)All modules are optional switches โ DevSquad works perfectly without them.
โโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโ
โ User Access Layer โ
โ โโโโโโโโโโโโโโโโ โโโโโโโโโโโโโโโโ โโโโโโโโโโโโโโโโ โ
โ โ Streamlit โ โ FastAPI REST โ โ CLI/Notebook โ โ
โ โ Dashboard โ โ API Server โ โ (Existing) โ โ
โ โ (Auth+HTTPS) โ โ (Swagger) โ โ โ โ
โ โโโโโโโโฌโโโโโโโโ โโโโโโโโฌโโโโโโโโ โโโโโโโโโโโโโโโโ โ
โโโโโโโโโโโผโโโโโโโโโโโโโโโโผโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโ
โ โ
โผ โผ
โโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโ
โ Business Logic Layer โ
โ โโโโโโโโโโโโโโโ โโโโโโโโโโโโโโโ โ
โ โAuthManager โ โHistoryMgr โ โ
โ โ(RBAC Auth) โ โ(SQLite TSDB)โ โ
โ โโโโโโโโโโโโโโโ โโโโโโโโโโโโโโโ โ
โ โโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโ โ
โ โ LifecycleProtocol (11-Phase Engine) โ โ
โ โ UnifiedGateEngine + CheckpointManager โ โ
โ โโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโ โ
โโโโโโโโโโโโโโโโโโโโโโโโโโโฌโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโ
โ
โผ
โโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโ
โ Data Persistence Layer โ
โ โโโโโโโโโโโโโโ โโโโโโโโโโโโโโ โโโโโโโโโโโโโโโโโโโโโโโโโโ โ
โ โ SQLite DB โ โ YAML Configโ โ Checkpoint Files โ โ
โ โ (History) โ โ (Deploy) โ โ (Lifecycle State) โ โ
โ โโโโโโโโโโโโโโ โโโโโโโโโโโโโโ โโโโโโโโโโโโโโโโโโโโโโโโโโ โ
โโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโโ
DevSquad provides 8 atomic sub-skills that can be used independently or together. Each sub-skill is a thin wrapper (~50 lines) importing existing core modules โ no duplicated logic.
skills/
โโโ dispatch/ โ DispatchSkill โ MultiAgentDispatcher (7-role orchestration)
โโโ intent/ โ IntentSkill โ IntentWorkflowMapper (6 intents ร 3 languages)
โโโ review/ โ ReviewSkill โ FiveAxisConsensusEngine (5-axis code review)
โโโ security/ โ SecuritySkill โ InputValidator + OperationClassifier + PermissionGuard
โโโ test/ โ TestSkill โ TestQualityGuard + test strategy generation
โโโ retrospective/ โ RetroSkill โ RetrospectiveEngine + pattern extraction
โโโ prototype/ โ PrototypeSkill โ Rapid prototype scaffolding (V4.5.0)
โโโ teach/ โ TeachSkill โ Knowledge transfer & onboarding (V4.5.0)
| Skill | Core Method | Wraps | Mock Mode |
|---|---|---|---|
dispatch |
run(task, roles, mode) |
MultiAgentDispatcher | โ |
intent |
detect(text, lang) |
IntentWorkflowMapper | โ |
review |
review(code) |
FiveAxisConsensusEngine | โ |
security |
scan_input(text) |
InputValidator + OpClassifier | โ |
test |
generate_strategy(module) |
TestQualityGuard | โ |
retrospective |
run_retrospective(results) |
RetrospectiveEngine | โ |
# Direct import (recommended for single skill)
from skills.dispatch.handler import DispatchSkill
result = DispatchSkill().run("Fix login bug", roles=["coder", "tester"])
# Via registry (dynamic discovery)
from skills import get_skill, list_skills
print(list_skills()) # ['dispatch', 'intent', 'review', 'security', 'test', 'retrospective']
skill = get_skill("security")
result = skill.scan_input("DROP TABLE users; --")All sub-skills work without any API key in Mock mode.
Unified Lifecycle Architecture - Resolves CLI 6 commands vs 11-phase lifecycle:
CLI View Layer (6 commands) Core Engine (11 phases)
โโโโโโโโโโโโโโโโโโโโโโโ โโโโโโโโโโโโโโโโโโโโโโโโโโโโ
โ spec โ P1, P2 โโโโView โโโโ P1: Requirements โ
โ plan โ P7 โ Mapping โ P2: Architecture โ
โ build โ P8 โ โ P3: Technical Design โ
โ test โ P9 โ โ ... โ
โ review โ P8,P6 โ โ P10: Deployment โ
โ ship โ P10 โ โ P11: Operations โ
โโโโโโโโโโโโโโโโโโโโโโโ โโโโโโโโโโโโโโโโโโโโโโโโโโโโ
โ โ
UnifiedGateEngine CheckpointManager
(Phase + Worker gates) (Lifecycle state persistence)
Core Components:
- โ LifecycleProtocol - Abstract interface for unified lifecycle management
- โ UnifiedGateEngine - Integrates VerificationGate + Phase transition gates
- โ FullLifecycleAdapter - Complete 11-phase lifecycle with dependency resolution
- โ Enhanced CheckpointManager - Auto save/restore lifecycle state across sessions
- Python 3.10+ (3.10, 3.11 supported, tested in CI)
- pip or pipenv for package management
# Install from PyPI โ zero setup, ready to use
pip install devsquad
# With optional dependencies
pip install "devsquad[api]" # FastAPI + Streamlit dashboard
pip install "devsquad[all]" # All optional featuresgit clone https://github.com/lulin70/DevSquad.git
cd DevSquad
# Install core package (minimal dependencies)
pip install -e .
# Ready to use!
devsquad dispatch -t "Design user authentication system"# Check version
devsquad --version
# Expected: devsquad 4.3.0
# Run tests
pytest tests/ -v --tb=short
# Expected: 7681 passedCreate .devsquad.yaml in your project root:
quality_control:
enabled: true
strict_mode: true
min_quality_score: 85
llm:
backend: auto
base_url: "" # Set via DEVSQUAD_OPENAI_BASE_URL env var
model: "" # Set via DEVSQUAD_OPENAI_MODEL env var
timeout: 120Or use environment variables (higher priority):
# Default: auto tries real backends first, then falls back to mock
export DEVSQUAD_LLM_BACKEND=auto
export DEVSQUAD_OPENAI_BASE_URL=https://api.openai.com/v1
export DEVSQUAD_OPENAI_MODEL=gpt-4
export DEVSQUAD_OPENAI_API_KEY=sk-...Environment Variables Reference:
| Variable | Purpose | Default |
|---|---|---|
DEVSQUAD_LLM_BACKEND |
Default backend type (auto|mock|trae|openai|anthropic|fallback) | auto |
DEVSQUAD_OPENAI_API_KEY |
OpenAI/MOKA AI API key | None |
DEVSQUAD_OPENAI_BASE_URL |
OpenAI-compatible base URL | None |
DEVSQUAD_OPENAI_MODEL |
OpenAI model name | gpt-4 |
DEVSQUAD_ANTHROPIC_API_KEY |
Anthropic API key | None |
DEVSQUAD_ANTHROPIC_BASE_URL |
Anthropic-compatible base URL | None |
DEVSQUAD_ANTHROPIC_MODEL |
Anthropic model name | claude-sonnet-4-20250514 |
DEVSQUAD_LOG_LEVEL |
Logging level | WARNING |
python3 scripts/cli.py --version # Expected: DevSquad 4.1.0
python3 scripts/cli.py status # Expected: System ready
python3 scripts/cli.py roles # Expected: 7 core roles listed# Run all tests (7681 tests passing)
python3 -m pytest tests/ -q --tb=line
# With coverage report
python3 -m pytest tests/ --cov=scripts --cov-report=term-missing| Priority | Scope | Examples | Count |
|---|---|---|---|
| P0 | Quality Framework Core | AntiRationalization, VerificationGate, IntentWorkflowMapper, AuthManager | ~200 |
| P1 | Enhancement Modules | FiveAxisConsensus, OperationClassifier, OutputSlicer | ~150 |
| P1+ | Cybernetics (V3.6.6) | FeedbackControlLoop, ExecutionGuard, PerformanceFingerprint, etc. | 110 |
| P2 | Integration & E2E | Full lifecycle dispatch, cross-module integration | ~200 |
| P3 | Unit per Module | Core dispatcher, RoleMapping, MCEAdapter, LLM backends | ~400+ |
Total: 7681 CI tests / 266 e2e (7681 collected)
Run by priority:
# P0 only (critical path, < 10s)
python3 -m pytest tests/ -k "anti_ratif or verification or intent_workflow or auth" -q
# P0 + P1 (quality + enhancement, < 30s)
python3 -m pytest tests/ -k "anti_ratif or verification or intent or auth or five_axis or operation" -q
# Full suite
python3 -m pytest tests/ -q --tb=line| Document | Description | Language |
|---|---|---|
| QUICKSTART.md | โญ 30 ็งๅฟซ้ๅ ฅ้จๆๅ๏ผๆจ่ๆฐ็จๆท๏ผ | ไธญๆ |
| SKILL.md | ๅฎๆดๆ่ฝๆๅ + 187+ ๆจกๅๅ่ | EN/CN/JP |
| GUIDE.md | ๅฎๅ จ็จๆทๆๅ | ไธญๆ |
| INSTALL.md | ๅฎ่ฃ ๆๅ (Unix + Windows) | EN/CN |
| EXAMPLES.md | ๅฎ้ ไฝฟ็จ็คบไพ | EN |
| CHANGELOG.md | ็ๆฌๅๅฒ่ฎฐๅฝ | EN |
| README-CN.md | ไธญๆ่ฏดๆ | ไธญๆ |
| README-JP.md | ๆฅๆฌ่ช่ชฌๆ | ๆฅๆฌ่ช |
| docs/PRD.md | ไบงๅ้ๆฑๆๆกฃ | ไธญๆ |
| docs/ARCHITECTURE.md | ๆๆฏๆถๆๆๆกฃ | ไธญๆ |
| docs/planning/V43_ROADMAP_PROPOSAL.md | V4.3 ็ปไธๆจ่ฟๆนๆก v1.2๏ผ7-Role ๅ ฑ่ฏ่พพๆ๏ผ | ไธญๆ |
| docs/prd/V4.3.0_PRD.md | V4.3.0 PRD๏ผ้ๆฑ/็จๆทๆ ไบ/้ชๆถๆ ๅ๏ผ | ไธญๆ |
| docs/architecture/V4.3.0_ARCHITECTURE.md | V4.3.0 ๆถๆ่ฎพ่ฎก๏ผๆจกๅ่พน็/ๆฅๅฃๅฅ็บฆ/ไพ่ตๅพ๏ผ | ไธญๆ |
| docs/testing/V4.3.0_TEST_PLAN.md | V4.3.0 ๆต่ฏๆนๆก๏ผๆต่ฏ้ๅญๅก/E2E/็ๅฎ็จๆทๆจกๆ๏ผ | ไธญๆ |
็ๆฌ็ญ็ฅ: V4.3.0 ้ขๅๅธ๏ผๅ จ้จไปฃ็ +ๆๆกฃ+E2E ้ช่ฏ๏ผโ ็จๆท็กฎ่ฎค โ V4.3.0 ๆญฃๅผ็
ๆดๅไธๆน้ข่พๅ ฅ:
- ๆๆฏๅบๆ็ปญๆฒป็๏ผ
todo_drift_monitor+ CI ้ปๅก๏ผ - pickleโJSON ่ฟ็งป๏ผdead code ๅ ้ค + fallback ๅฎๅ จๆถ็ดง + ็งป้ค๏ผ
- ไธๆธธ TraeMultiAgentSkill v2.6-v2.8 ็ฒพ็ปๅๅฏๅ๏ผPonytail ๅๆจกๅผ / LoopKernel ๅ้ / UIUX ๅฎก่ฎก / Dashboard ๅฏ่งๅ๏ผ
V4.3.0 ่ๅด๏ผ9 ้กน๏ผ:
| ID | ๅ็งฐ | ไผๅ ็บง |
|---|---|---|
| P0-1 | pickle dead code ๅ ้ค + fallback ๅฎๅ จๆถ็ดง | P0 |
| P0-2 | todo_drift_monitor.py + CI ้ปๅก + PR template |
P0 |
| P1-1 | Ponytail lite/full ๅๆจกๅผ + DebtCollector + RequirementTracer | P1 |
| P1-4 | LoopKernel RollbackStrategy + ็ฌ็ซ็กฌไธ้ | P1 |
| P1-5 | UIUXAnalyzer ๅญ้กนๅฎก่ฎก + ๆ้่กฅๅ จ | P1 |
| P1-6 | Dashboard ็ถๆๅฏ่งๅ | P1 |
| P2-1 | pickle fallback ็งป้ค | P2 |
| P2-2 | Autonomous SmartConfirmation ๆๆกฃ่กฅๅ จ | P2 |
| P2-4 | V4.3.0 ๅๅธๆๆกฃๅๆญฅ | P2 |
7-Role ๅ ฑ่ฏ: 7/7 APPROVE_WITH_CONCERNS๏ผๆ 10 ้กน่ฐๆดไฟฎ่ฎขๅ่พพๆๅ ฑ่ฏใ่ฏฆ่ง V43_ROADMAP_PROPOSAL.md v1.2ใ
้กน็ฎ็ๅฝๅจๆ: ๆ 11-Phase ๆจกๅๆจ่ฟ๏ผP1 ้ๆฑ โ P2 ๆถๆ โ P3 ๆๆฏ่ฎพ่ฎก โ P7 ๆต่ฏ่ฎกๅ โ P8 ๅฎๆฝ โ P9 ๆต่ฏๆง่ก โ P10 ้จ็ฝฒๅๅธ๏ผ
ๆต่ฏ้ๅญๅกไฟ้: unit โฅ60% / integration 15-25% / e2e โค10% / contract 5-10% / smoke โค5%
- Fork the repository
- Create feature branch (
git checkout -b feature/amazing-feature) - Commit changes (
git commit -m 'Add amazing feature') - Push to branch (
git push origin feature/amazing-feature) - Open a Pull Request
This project is licensed under the MIT License - see the LICENSE file for details.
โญ ๅฆๆ DevSquad ๅฏนไฝ ๆๅธฎๅฉ๏ผ่ฏท็ปไธช Star๏ผโญ
่ฎฉๆดๅคๅผๅ่
ไบซๅๅฐใAI ๅข้ๅไฝใ็ๅ้
๐ Acknowledgments
Inspired by TraeMultiAgentSkill upstream project
Built with โค๏ธ by the DevSquad team
Last updated: 2026-08-05 | Version: V4.5.1 (Approval Gate + Connector Framework + anti-ghost E2E โ 2 new modules, 3 ROADMAP items completed) | V4.5.0 (cross-session continuity + protocol-native skills + action-first reports โ 10 new features) | V4.4.0 (5 enhancement modules: RiskRegister / ViewpointRegistry / ErrorBudgetTracker / GapAnalyzer / DoraMetricsCollector โ see CHANGELOG.md)