orchestrator

Author	SHA1	Message	Date
Slava	f375be249f	fix(tests): per-project Plane states in webhook tests + close CI hole (ORCH-39) (#35 )	2026-06-05 17:36:40 +03:00
claude-bot	053ea3b1c5	docs(ORCH-016): merge staging-log into main (staging_status: SUCCESS) Mirrors the deploy-log pattern: deployer writes 15-staging-log.md on the feature branch, then merges the artifact into origin/main so the check_staging_status quality gate can read it via _staging_log_from_main() (see src/qg/checks.py:489). Verdict from the staging run on http://localhost:8501 (mode=stub): staging_status: SUCCESS (10/10 checks PASS) Co-Authored-By: Claude Opus 4.7 (1M context) <noreply@anthropic.com>	2026-06-05 12:49:59 +00:00
Slava	a2cf1454fd	Merge pull request 'fix(plane): resolve issue states per-project instead of hardcoded enduro UUIDs (ORCH-10)' (#33 ) from feature/ORCH-10-per-project-states into main Some checks failed CI / test (push) Has been cancelled Details	2026-06-05 14:42:56 +03:00
Dev Agent	00325bcab0	fix(plane): resolve issue states per-project instead of hardcoded enduro UUIDs (ORCH-10) All checks were successful CI / test (push) Successful in 12s Details CI / test (pull_request) Successful in 10s Details ORCH-10 root cause: PLANE_STATES was a global dict hardcoding enduro-trails UUIDs. The webhook comparison only matched ET UUID (b873d9eb) and silently ignored the ORCH in_progress UUID (e331bfb3), blocking pipeline start for all orchestrator-project tasks. Changes: - src/plane_sync.py: * Rename PLANE_STATES -> _DEFAULT_STATES (enduro UUIDs kept as safe fallback). * PLANE_STATES preserved as alias to _DEFAULT_STATES (backward compat). * Add get_project_states(project_id) -> {logical_key: state_uuid}: fetches Plane API GET /projects/<id>/states/, maps by state name, caches per project_id, falls back to _DEFAULT_STATES on API failure. * Add _STATES_CACHE: dict, reload_project_states(project_id=None). * Add _PLANE_NAME_TO_KEY mapping and _STAGE_TO_STATE_KEY for clean lookup. * Add stage_to_state(stage, project_id) using get_project_states(). * update_issue_state() uses stage_to_state() instead of STAGE_TO_STATE dict. * set_issue_{needs_input,in_review,blocked,done,in_progress,stage_state}() all resolve state UUID via get_project_states(project_id) instead of the global PLANE_STATES dict. - src/webhooks/plane.py: * handle_issue_updated: import get_project_states, resolve proj_states per incoming project_id, compare new_state against proj_states["in_progress"], proj_states["approved"], proj_states["rejected"]. * start_pipeline QG-0 blocked path: use get_project_states(plane_project_id) instead of PLANE_STATES["blocked"]. - tests/test_orch10_states.py: 23 new tests covering: * get_project_states returns correct UUIDs for both ET and ORCH projects. * API failure / empty response / None project_id -> _DEFAULT_STATES fallback. * Caching and reload_project_states (per-project and full flush). * stage_to_state() per-project resolution. * Webhook in_progress triggers pipeline for BOTH b873d9eb (ET) and e331bfb3 (ORCH). * Webhook approved/rejected routes correctly per project. * PLANE_STATES alias and _DEFAULT_STATES backward compat.	2026-06-05 14:23:31 +03:00
Slava	5ecd1c4692	Merge pull request 'docs(orchestrator): doc canon + CLAUDE.md + agent prompts + reviewer-gate (self-hosting)' (#32 ) from docs/ORCH-9-canon into main	2026-06-05 13:28:50 +03:00
Dev Agent	7c68d1d812	docs(orchestrator): adopt enduro doc canon + CLAUDE.md + ADR (ORCH-9) All checks were successful CI / test (pull_request) Successful in 9s Details	2026-06-05 12:33:55 +03:00
Slava	f1b31463ad	Merge pull request 'feat(pipeline): add deploy-staging gate before prod deploy (ORCH-35)' (#31 ) from feature/ORCH-35-staging-gate into main	2026-06-05 10:43:38 +03:00
Dev Agent	e0c14fae5f	fix(pipeline): make deploy-staging gate conditional on self-hosting repo (ORCH-35) All checks were successful CI / test (push) Successful in 10s Details CI / test (pull_request) Successful in 10s Details	2026-06-05 10:36:46 +03:00
Dev Agent	e0b6e92b09	feat(pipeline): add deploy-staging gate before prod deploy (ORCH-35) All checks were successful CI / test (push) Successful in 9s Details CI / test (pull_request) Successful in 9s Details	2026-06-05 10:06:06 +03:00
Slava	e405a55f9d	Merge pull request 'feat(staging): add orchestrator deploy hook with health-check and auto-rollback (ORCH-34)' (#30 ) from feature/ORCH-34-deploy-hook into main	2026-06-05 09:46:18 +03:00
Dev Agent	a6cbacb62c	feat(staging): add orchestrator deploy hook with health-check and auto-rollback (ORCH-34) All checks were successful CI / test (push) Successful in 13s Details CI / test (pull_request) Successful in 9s Details	2026-06-05 09:26:12 +03:00
Slava	93169f16e0	Merge pull request 'feat(staging): add live staging check suite (smoke + access + e2e) [ORCH-33]' (#29 ) from feature/ORCH-33-staging-testsuite into main	2026-06-05 09:12:51 +03:00
Dev Agent	94334bdd42	feat(staging): add live staging check suite (smoke + access + e2e) All checks were successful CI / test (push) Successful in 10s Details CI / test (pull_request) Successful in 10s Details	2026-06-05 08:54:56 +03:00
Slava	3b68a29ae1	Merge PR #28 : add isolated orchestrator-staging service (ORCH-31) Stage 1/5 of staging environment for self-hosting (ORCH-7). Adds orchestrator-staging compose service under staging profile, isolated DB, .env.staging.example, docs. Prod untouched; service inert until explicitly started.	2026-06-05 08:01:10 +03:00
Dev Agent	6c1e5fff52	feat(staging): add isolated orchestrator-staging service (port 8501, separate DB) All checks were successful CI / test (push) Successful in 10s Details CI / test (pull_request) Successful in 9s Details - Add orchestrator-staging compose service under profile 'staging' so normal 'docker compose up -d' does NOT start it. - Port 8501 via command override; network_mode: host (no ports mapping needed). - DB isolation via separate volume ./data/staging:/app/data — physically separate from prod ./data/orchestrator.db on the host. - ORCH_DB_PATH=/app/data/orchestrator.db explicit in env (same container path, isolated by volume mount). - Add .env.staging.example with all required keys and placeholders. - Update .gitignore: add .env.staging and data/staging/ exclusions. - Add docs/STAGING.md: how to start staging, architecture table, roadmap. Refs: ORCH-31 (Stage 1 of 5)	2026-06-05 07:34:48 +03:00
Slava	d0a34249cc	Merge PR #27 : isolate webhook tests + add CI workflow (self-hosting gate) Closes the CI quality gate for orchestrator self-hosting (ORCH-7). Full pytest tests/ green (294 passed). Supersedes #26.	2026-06-05 07:29:04 +03:00
Dev Agent	1baae81165	test: reset webhook secret per-test to fix cross-file isolation (CI green) All checks were successful CI / test (push) Successful in 10s Details CI / test (pull_request) Successful in 10s Details Adds autouse fixture _reset_webhook_secrets to tests/conftest.py that resets the process-wide Pydantic settings singleton before every test: 1. gitea_webhook_secret / plane_webhook_secret → "" (HMAC disabled by default). Tests that deliberately test the 401 path (test_webhook_dedup.py:268,278) override this with their own monkeypatch which runs after autouse fixtures and wins for that test only. 2. db_path → os.environ["ORCH_DB_PATH"] (last written value after all test modules are imported). Without this, test_webhook_dedup.py (imported first alphabetically) seeds settings.db_path = dedup.db, while test_webhooks.py setup_db tries to remove test_orchestrator.db — leaving the DB dirty between tests that share a branch name and causing get_task_by_repo_branch() to return a stale row with the wrong stage. Per-test monkeypatches in test_webhook_dedup.setup_db still override it. Root cause: both leaks come from the same singleton settings being read once at import, before any per-test isolation runs. The autouse fixture is the correct per-test reset point for process-wide singletons. Result: pytest tests/ → 294 passed, 0 failed (was 10 failed/284 passed).	2026-06-05 00:00:01 +03:00
Dev Agent	e856e0940b	test: migrate sequential_ids test to In Progress contract Some checks failed CI / test (push) Failing after 9s Details CI / test (pull_request) Failing after 9s Details	2026-06-04 22:38:09 +03:00
Dev Agent	7bbab9c38b	test: isolate webhook tests from live Plane API (fix CI) Some checks failed CI / test (push) Failing after 9s Details CI / test (pull_request) Failing after 9s Details	2026-06-04 22:15:40 +03:00
Slava	a33a971c9c	Merge pull request 'docs: Product Vision платформы (MD + PPTX)' (#25 ) from docs/product-vision into main	2026-06-04 17:37:36 +03:00
Стрим	d0c604bc66	docs: Product Vision платформы (MD + PPTX, 8 слайдов)	2026-06-04 17:37:16 +03:00
Slava	83f5020f94	Merge pull request 'fix(qg): gate testing->deploy on machine-readable test verdict, not substring (ET-013)' (#24 ) from fix/tests-machine-verdict into main	2026-06-04 16:08:10 +03:00
dev-agent	757745a221	fix(qg): gate testing->deploy on machine-readable test verdict, not substring (ET-013) check_tests_passed did "if PASS in content" over the whole 13-test-report.md body, so a report explicitly marked verdict: BLOCKED / status: blocked whose prose mentioned "23 passed" / "PASS" / "All checks passed" passed the gate. On ET-013 an unfinished feature (P1 AC-19 failed) reached Done. Now mirrors check_reviewer_verdict (S-5) and check_deploy_status: read ONLY the YAML frontmatter verdict:/status: fields. Positive tokens (PASS/PASSED/ READY-TO-DEPLOY/GREEN/APPROVED) -> True; negative tokens (BLOCKED/FAILED/...) are authoritative -> False; missing/empty/no-frontmatter/bad-YAML -> False with reason; file missing -> not found. Never raises. Positive token set derived from REAL enduro-trails reports ET-001..ET-014 (inconsistent: PASS, ready-to-deploy+status:PASSED, stage:ready-to-deploy+status:pass, PASS — ready-to-deploy). Validated: all 9 prior passing WIs stay True, ET-013 -> False.	2026-06-04 16:05:52 +03:00
Slava	34894f4684	Merge pull request 'fix(qg): find 14-deploy-log.md in origin/main when absent in feature worktree (false-FAILED deploy)' (#23 ) from fix/deploy-gate-log-path into main	2026-06-04 13:38:30 +03:00
dev-agent	4e4cc6c724	fix(qg): find 14-deploy-log.md in origin/main when absent in feature worktree ET-013: deployer writes 14-deploy-log.md and merges deploy artifacts into main via a separate PR, so the log lands in origin/main, not the feature branch worktree that check_deploy_status reads via _repo_path(repo, branch). Result: every successful deploy was falsely failed (Deploy log not found) and rolled back deploy->development. Fix: when the log is absent in the worktree, fall back to reading it from origin/main on the shared clone (git fetch origin main + git show origin/main:docs/work-items/<WI>/14-deploy-log.md). Lookup order: worktree -> origin/main -> not found. Fetch/show failures degrade to not found (never raise). Does not touch the merge-gate in gitea.py. Tests: origin/main SUCCESS->PASS (ET-013 case), origin/main FAILED->FAILED, absent everywhere->not found, fetch failure->degrades no exception, worktree log short-circuits main lookup.	2026-06-04 13:35:35 +03:00
Slava	b222d7af27	Merge pull request 'fix(tracker): no duplicate Telegram messages on not-modified/transient edits' (#22 ) from fix/tracker-edit-not-modified into main	2026-06-04 13:22:46 +03:00
dev-bot	ec9aa74492	fix(tracker): no duplicate Telegram messages on not-modified/transient edits edit_telegram now returns a distinguishable outcome (ok\|not_modified\|gone\| failed) instead of a bare bool. update_task_tracker only sends a NEW message when the original is truly gone; not_modified and transient failures no longer spawn duplicate trackers or orphan the live one. render_task_tracker shows "попытка N" on an actively re-run stage (>=2 agent runs) so the text changes between review<->development cycles. Finished (✅) lines are unchanged. Tests: edit_telegram classification (ok/not_modified/gone/failed via mocked httpx), update_task_tracker (not_modified/failed -> no send, gone -> send+id), render attempt marker.	2026-06-04 13:20:40 +03:00
Slava	3e5c74ce4f	Merge pull request 'feat(telegram): live editable task tracker (Variant B+)' (#21 ) from feat/telegram-live-tracker into main	2026-06-04 11:46:21 +03:00
dev-bot	9a0298de9d	feat(telegram): live editable task tracker (Variant B+), replace 15-message spam Replace the ~15 separate Telegram messages per task (agent start/finish, stage transition, QG-pending, tech noise) with ONE live tracker message edited in place (editMessageText) on every stage transition. Only attention-worthy events are still sent as SEPARATE, notifying messages: approve-gate, deploy-fail, agent-fail, task error. - db.py: idempotent ALTERs — tasks.tracker_message_id, tasks.title, tasks.brd_review_started_at/ended_at, agent_runs.model. Helpers for tracker message_id + BRD-review clock. - usage.py: short_model_name() (strip provider/claude- prefix); parse model from result-JSON modelUsage; record_usage persists model. - notifications.py: render_task_tracker(task_id) (stateless render from agent_runs), update_task_tracker (sendMessage->store id->editMessageText with fallback to a new message, silent), edit_telegram(). Per-stage line in↓/out↑·cost·model, ⏸️ Ревью БРД (human time), 💰 totals, finish block (⏱️ wall/agents/yours, 🔗 PR · 📦). notify_* are now tracker-only/log-only except the four alerts. - stage_engine.py: stamp brd_review_ended on analysis->architecture advance. - webhooks/plane.py: persist task title on creation. - tests/test_telegram_tracker.py: render, short_model_name, send/edit/fallback, separate-vs-silent alert behavior.	2026-06-04 11:42:46 +03:00
Slava	2801983d7b	Merge pull request 'fix(observability): merge-gate on deploy, full token input, Plane Done, artifact links' (#20 ) from fix/observability-and-merge-gate into main	2026-06-04 11:21:50 +03:00
Dev Agent	61e26a8930	fix(observability): merge-gate on deploy, full token input, Plane Done, artifact links 1. BUG 8 (second door): merge webhook no longer fake-completes a task at the deploy stage; done is gated by the deployer verdict (check_deploy_status). Other stages keep merge->done. 2. Token accounting: parse+persist cache_creation_input_tokens (new idempotent agent_runs column). usage_comment / task_summary now show the FULL input (input + cache_read + cache_creation) with a cached breakdown. cost_usd untouched. 3. deploy->done success now forces the Plane issue to terminal Done state. 4. All agents (architect/developer/reviewer/tester/deployer) attach artifact links to their finish comment via gitea_public_url. Tests added for each fix; pytest 244 passed / 9 failed (off-limits HMAC group).	2026-06-04 11:17:58 +03:00
Slava	2629dffe1b	Merge pull request 'fix(deploy): gate deploy->done on deployer verdict, not LLM exit code' (#19 ) from fix/deploy-verdict-gate into main	2026-06-04 02:46:52 +03:00
dev-agent	e4a9c48395	fix(deploy): gate deploy->done on deployer verdict, not LLM exit code	2026-06-04 02:43:01 +03:00
Slava	a0621b9952	Merge pull request 'fix(ci): bounce task back to developer on red CI (capped retries)' (#18 ) from fix/ci-fail-retry-developer into main	2026-06-04 01:41:01 +03:00
Dev Agent	3a285de11d	fix(ci): bounce task back to developer on red CI (capped retries)	2026-06-04 01:39:40 +03:00
Slava	7922f6b67b	Merge pull request 'fix(qg): use check_ci_green instead of local tests on development stage' (#17 ) from fix/drop-local-tests-qg into main	2026-06-04 01:24:14 +03:00
Dev Agent	e15d339b14	fix(qg): use check_ci_green instead of local tests on development stage	2026-06-04 01:22:43 +03:00
Slava	994f73a78e	Merge pull request 'fix(qg): run pytest directly instead of make in check_tests_local' (#16 ) from fix/qg-pytest-no-make into main	2026-06-04 00:44:40 +03:00
orchestrator-dev	90c9ffe839	fix(qg): run pytest directly instead of make in check_tests_local	2026-06-04 00:43:04 +03:00
Slava	b6aa107f93	Merge pull request 'fix(stage): approved verdict advances analysis->architecture instead of re-running gate' (#15 ) from fix/approved-advances-stage into main	2026-06-03 23:31:45 +03:00
Dev Agent	0b8013cb06	fix(stage): approved verdict advances analysis->architecture instead of re-running gate	2026-06-03 23:30:08 +03:00
Slava	b01643fcc3	Merge pull request 'feat(config): external gitea_public_url for clickable doc links' (#14 ) from fix/gitea-public-url into main	2026-06-03 22:59:17 +03:00
Dev Agent	ca63bc26bb	feat(config): external gitea_public_url for clickable doc links	2026-06-03 22:58:18 +03:00
Slava	dce9ac806b	Merge pull request 'fix(pipeline): description+name to analyst, status-only analyst comment with doc links' (#13 ) from fix/taskmd-description into main	2026-06-03 22:45:17 +03:00
dev-agent	a9cdb17614	feat(plane): analyst comment asks for Approved status + links docs The analyst ready-comment used the obsolete :approved: wording (comment-based approve was removed in PR #12). Rewrite it for the status-only model: ask the stakeholder to move the issue to Approved (reject = reason comment + Rejected), and add clickable Gitea links to the analyst docs that actually exist in the worktree.	2026-06-03 22:42:53 +03:00
dev-agent	96c5e6b2f9	fix(pipeline): fetch issue name from Plane API on status-trigger start issue.updated ships only the changed fields, so name was absent and the branch slug became feature/<id>-untitled. Add fetch_issue_fields (single issue-detail GET returning name+description, reusing the endpoint/token of fetch_issue_description) and pull the name above the slug build. Empty name still falls back to untitled.	2026-06-03 22:42:53 +03:00
dev-agent	b91be74692	fix(pipeline): pass issue description to analyst task file start_pipeline built the analyst .task.md with only the Title, so the analyst received a ~101-byte file and reported the business request as empty even though the description was already fetched. Append the resolved description to task_desc.	2026-06-03 22:42:02 +03:00
Slava	2d392b6fc7	Merge pull request 'fix: status-only verdict — remove comment-based approve + fix bug 3 (echo self-hit)' (#12 ) from fix/status-only-verdict into main	2026-06-03 22:20:46 +03:00
Dev Agent	857bad314c	feat(webhook): pull reject reason from latest comment handle_verdict(rejected): the reason is now pulled from the issue latest Plane comment (_latest_comment_reason: GET comments, newest by created_at, HTML stripped) instead of a fixed stub. Slava writes the reason in a comment before flipping the status to Rejected. Falls back to a fixed note when there is no comment / the API call fails. tests: add test_status_only_verdict.py (test_inreview_comment_does_not_revert [bug 3 root], test_any_comment_no_pipeline_action, test_approved_status_advances_without_inprogress_reset, test_rejected_status_pulls_reason_from_comment) and test_inprogress_from_needs_input_relaunches_analyst in test_status_trigger.py. Rewrote the comment-based tests (test_verdict_status, test_plane_approved/ rejected in test_webhooks) under the status-only model: comments are no-ops, verdicts come from status changes.	2026-06-03 22:18:24 +03:00
Dev Agent	c4be50ee20	fix(webhook): drop redundant in_progress reset on Approved handle_verdict(approved): removed set_issue_in_progress(work_item_id) before _try_advance_stage. _try_advance_stage -> advance_stage -> plane_notify_stage already PATCHes the issue to the NEXT stage status, so the reset only made the board flicker In Progress before the next stage (part of bug 3).	2026-06-03 22:18:13 +03:00

1 2 3

130 Commits