teleo/teleo-infrastructure

Author	SHA1	Message	Date
m3taversal	d67d36b409	fix: decision extractor uses extract worktree + PR flow Was writing directly to main worktree where daemon race condition wiped files. Now: syncs extract worktree to main, creates branch, writes records, commits, pushes, opens Forgejo PR. Same pattern as batch-extract. Also checks both main and extract worktrees for existing records. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-24 02:50:12 +00:00
m3taversal	9267351aba	fix: 7-day TTL on dated learnings + block availability learnings Stale learning ("I don't have Robin Hanson data") overrode real KB data. Ganymede review: dated entries expire after 7 days. Permanent entries (communication style, identity) are undated and always included. Prompt guard: "NEVER save a learning about what data you do or don't have" prevents the bot from writing availability claims that go stale. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 18:07:46 +00:00
m3taversal	6c6cd0d14e	feat: support fundraise record_type alongside decision_market LLM now classifies proposals as either decision_market (governance votes) or fundraise (ICO/launch capital raises). Both handled by same extractor. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 18:04:30 +00:00
m3taversal	e1934b30ae	fix: API key path + YAML error handling in decision extractor Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 17:59:11 +00:00
m3taversal	a292ab75c2	feat: decision record extractor — proposal sources → decisions/ with full text Reads event_type: proposal sources from archive, calls Sonnet for summary/significance/KB-connections, writes decision records with full verbatim proposal text + structured analysis on top. 224 proposal sources archived, 0 processed. This closes the gap. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 17:55:46 +00:00
m3taversal	28be7555b1	fix: top 3 entities get full body in prompt, not just top 1 When two related entities match (advisor hire + research grant), both need full content so Opus can distinguish them and serve the right one. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 17:44:51 +00:00
m3taversal	f77fd229d6	fix: stop word filtering in entity scoring — common words polluted rankings 'the', 'full', 'text', 'proposal' etc. were matching irrelevant entities. Robin Hanson record ranked #2 behind Drift because Drift matched 'the' and 'proposal' in its name. Now only meaningful tokens (>=3 chars, not stop words) contribute to entity scoring. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 17:44:06 +00:00
m3taversal	089b4609d5	fix: score + rank entities, limit to top 5, full body for decisions Before: "Robin Hanson MetaDAO proposal" returned 34 entities (39K chars) with the target record buried at position 13. No relevance scoring. After: entities scored by query token overlap (name 3x, alias 1x, bigram 5x), limited to top 5 results. Decision records get full body (2K chars) instead of 500-char truncation. Top result gets 2K in prompt, rest get 500. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 17:38:10 +00:00
m3taversal	3ed0f20fa1	fix: index parent_entity as alias for decision records (Ganymede review) MetaDAO queries now surface MetaDAO's decision records because parent_entity: "[[metadao]]" is stripped and added to the alias set. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 17:31:54 +00:00
m3taversal	425e7a1bac	fix: index decisions/ as entities so decision records reach the bot prompt Root cause: decision records have type: decision, but the entity indexer only accepted type: entity and only scanned entities/. The claim indexer scanned decisions/ but filtered out non-claim types. Result: decision records fell through both indexes entirely — invisible to the bot. Fix: add decisions/ to entity indexer scan paths, accept type: decision alongside type: entity, include summary/proposer in search aliases. Remove decisions/ from claim indexer (was silently dropping them anyway). Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 17:28:30 +00:00
m3taversal	c7c71ec9d1	epimetheus: fix double research message + add decisions/ to KB retrieval 1. handle_research gets silent=True param. RESEARCH: tag triggers use silent mode — archives tweets but posts no follow-up message. Prevents "Queued N tweets" after Opus already responded. 2. KB retrieval now searches decisions/ directory alongside domains/, core/, foundations/. Decision records (Robin Hanson proposal, etc.) are now findable by the bot. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 16:59:23 +00:00
m3taversal	c59db5812f	epimetheus: fix article content parsing — contents[] array, not text field Article endpoint returns body in "contents" array of typed blocks (unstyled, header-two, markdown, list-item, blockquote, etc). Was looking for article.text which is empty. Now parses all block types into readable text. Also extracts engagement stats (likes, views). Fixes: "Claude + Obsidian" article returned title but empty text. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 15:30:59 +00:00
m3taversal	bcbe54a0a3	epimetheus: consolidated X API client (x_client.py replaces x_search.py) Clean, documented interface to twitterapi.io for all agents: - get_tweet(id) — fetch any tweet by ID, any age - get_article(id) — fetch X long-form articles - search_tweets(query) — keyword search for research - get_user_tweets(username) — user's recent tweets (research sessions) - fetch_from_url(url) — smart dispatcher: tweet → article → placeholder Shared by Telegram bot + research sessions. Documented endpoints, costs, rate limits. Replaces ad-hoc x_search.py. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 15:26:10 +00:00
m3taversal	7360f6b22e	epimetheus: direct tweet lookup via /tweets?tweet_ids= endpoint Primary path: GET /twitter/tweets?tweet_ids={id} — works for any tweet, any age, returns full content. Replaces the fragile from:username search pagination fallback. Fallback: article endpoint for X long-form articles. Last resort: placeholder with [Could not fetch] message. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 15:17:11 +00:00
m3taversal	8f4e583c76	epimetheus: Ganymede review fixes + tweet fetch pagination - fetch_tweet_by_url: paginate up to 3 pages to find older tweets - Return placeholder on fetch failure (Ganymede: surface failure to user) - Don't burn user rate limit on Haiku autonomous searches (Ganymede) - 7-day limitation documented in comment Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 15:12:10 +00:00
m3taversal	76d5644272	epimetheus: X link fetching + Haiku pre-pass + systemd fix + query tuning Major changes this session: - fetch_tweet_by_url: extracts username+ID from X URLs, tries article endpoint, falls back to from:username search. Tweets injected into Opus prompt. - Haiku pre-pass: decides if X search needed before Opus responds. 2-3 word queries. - systemd ProtectSystem paths fixed (ROOT CAUSE of all write failures since day 1) - Research regex handles Telegram @botname suffix in groups - Double research message prevented (skip RESEARCH: tag when Haiku already ran) - Engagement filter dropped to 0 for niche crypto tweets - Heuristic brevity in prompt (not hard cap) - DM auto-respond gating (groups: reply-to only, DMs: auto-respond) - All code now edited in pipeline-v2 repo, not /tmp Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 14:05:07 +00:00
m3taversal	ed46c0674b	epimetheus: fix double research message + Haiku query tuning - Skip RESEARCH: tag when Haiku pre-pass already searched (no double-fire) - Haiku told to use 2-3 word queries (was generating 6+ word queries that returned 0) - Engagement filter dropped to 0 (niche crypto tweets have low engagement) - systemd ProtectSystem paths fixed (root cause of ALL write failures) Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 13:57:12 +00:00
m3taversal	7086bcacb1	epimetheus: Haiku pre-pass for auto-research (Option A) Before Opus responds, Haiku evaluates: "Does this message need an X search?" If YES, searches X, injects results into Opus prompt, archives as source. Opus responds with KB knowledge + fresh tweet data combined. Flow: user asks naturally ("what are people saying about P2P?") → Haiku decides search needed → X search → results in Opus context → unified response. ~1s latency, ~$0.001 cost per message. Only fires when Haiku says YES. Explicit /research command still works as direct path. Also: fixed systemd ProtectSystem paths (Ganymede: root cause of all write failures). Fixed research regex for Telegram group commands. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 13:31:29 +00:00
m3taversal	5388f701bd	epimetheus: heuristic brevity, not hard cap Replaced hard rules with judgment heuristics: - "Does every sentence add something the user doesn't already know?" - "Earn every paragraph — each needs a distinct insight" - "Short questions deserve short answers" Restored max_tokens to 1024. Agent decides length, not a token cap. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 12:58:42 +00:00
m3taversal	08aa52659c	epimetheus: enforce brevity + fix research regex false positive 1. Response length: "BREVITY IS YOUR DEFAULT. Most responses 1-3 sentences. A 4-paragraph response to a simple question is a failure." max_tokens cut from 1024 to 512. 2. Research trigger: removed natural language regex (caused false positive on "has accumulated" matching "search"). Only explicit /research command. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 12:57:09 +00:00
m3taversal	b90e80ed6c	epimetheus: don't track silent group messages in history (Ganymede review) Option A: history only contains actual bot-user exchanges, not unaddressed group messages. Empty bot responses in history confused the model. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 12:31:01 +00:00
m3taversal	251caa3695	epimetheus: DM auto-respond gating (Rio suggestion) DMs (private chats): conversation window auto-responds — always 1-on-1, no false positives. Groups (supergroup/group): conversation window tracks context silently, reply-to only trigger. Simple msg.chat.type check. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-23 10:19:15 +00:00
m3taversal	a75c14e536	epimetheus: auto-learning trigger — bot self-writes learnings from corrections Opus decides what to learn. Prompt instructs: append LEARNING: [category] [description] at end of response when genuinely learning something new. Bot parses the line, strips it from displayed response, calls _save_learning() to persist. Zero additional API calls (Rhea's design). The model already has full context. Categories: factual, communication, structured_data. Most responses have no LEARNING line — only fires on genuine corrections. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-22 16:57:47 +00:00
m3taversal	a11eca90e3	epimetheus: compressed conversation context + decouple archive from lock 1. Conversation history now shows compressed context summary first (tickers, key figures, exchange count) before full log. "Discussing: $FUTARDIO \| Key figures: $0.004, $39.5K \| Exchanges: 3" 20 tokens, unmissable. Plus prompt instruction: "NEVER ask a question your history already answers." (Ganymede: Option C+A) 2. Archive file writes decoupled from worktree lock. File written unlocked (additive, no coordination needed). Git commit attempted with lock — deferred on timeout, file persists on disk for next cycle. Fixes "Read-only file system" archive failures. (Ganymede review) Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-21 17:26:02 +00:00
m3taversal	8d10c8ee28	epimetheus: conversation window → silent context only (Ganymede+Rhea+Leo) Auto-respond stripped from conversation window. Bot only responds to @tag and reply-to-bot. Window now silently tracks messages for context — when the user does reply, the bot has full conversation history. Also: prompt shortened to "1-2 sentences" default. "Do NOT respond to messages that aren't directed at you." Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-21 16:51:26 +00:00
m3taversal	f7d30ced1a	epimetheus: /research command — user-triggered X search from Telegram User says "@FutAIrdBot /research P2P.me launch" → bot searches X via twitterapi.io → archives all tweets as ONE consolidated source file in inbox/queue/ → batch extract picks up → claims land in KB. Features (Ganymede+Rhea+Leo+Rio consensus): - Regex + natural language intent detection (not CommandHandler) - One source file per research query (not per-tweet) - Full tweet metadata: author, followers, engagement, date - Contributor attribution: proposed_by + contribution_type: research-direction - Rate limit: 3 searches per user per day - Min engagement filter (3 interactions) - Worktree lock on source file write Phase 2 (not built): domain alignment check before searching. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-21 16:32:43 +00:00
m3taversal	e921eda0a0	epimetheus: sanitize learnings before prompt injection (Ganymede review) Learnings file content now passes through sanitize_message() before injection into the Opus prompt. Prevents prompt injection via crafted "corrections." Rio UUID 5551F5AF confirmed as current Teleo v4 Rio. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-21 15:29:46 +00:00
m3taversal	1b4c6f8d72	epimetheus: agent learning system — learnings.md reader + self-write Option D (Rhea+Rio+Leo consensus): - _load_learnings(): reads agents/rio/learnings.md, injects into prompt before KB context - _save_learning(): appends correction to learnings.md via worktree lock + direct commit - Learnings prioritized over KB data when they conflict - Three categories: communication, factual, structured_data - Prompt updated: tells agent it can save corrections for future conversations Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-21 15:25:52 +00:00
m3taversal	e233dbbcee	epimetheus: auto-clean stale queue duplicates at start of each extract cycle Pre-extraction step: find queue files already in archive, delete them, commit + push. Runs on main worktree before extraction starts on extract worktree. Prevents "queue duplicate reappears after reset --hard" problem that produced 16 stale entries overnight. Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-21 14:20:29 +00:00
m3taversal	d97f68714a	epimetheus: fix 2 nits from Ganymede final review 1. _merge_pr marked as CURRENTLY UNUSED (local ff-push is primary path) 2. Conversation window messages skip cold rate limit check (window counter IS the limit) Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-20 20:25:06 +00:00
m3taversal	d79ff60689	epimetheus: sync VPS-deployed code to repo — Mar 18-20 reliability + features Pipeline reliability (8 fixes, reviewed by Ganymede+Rhea+Leo+Rio): 1. Merge API recovery — pre-flight approval check, transient/permanent distinction, jitter 2. Ghost PR detection — ls-remote branch check in reconciliation, network guard 3. Source status contract — directory IS status, no code change needed 4. Batch-state markers eliminated — two-gate skip (archive-check + batched branch-check) 5. Branch SHA tracking — batched ls-remote, auto-reset verdicts, dismiss stale reviews 6. Mirror pre-flight permissions — chown check in sync-mirror.sh 7. Telegram archive commit-after-write — git add/commit/push with rebase --abort fallback 8. Post-merge source archiving — queue/ → archive/{domain}/ after merge Pipeline fixes: - merge_cycled flag — eval attempts preserved during merge-failure cycling (Ganymede+Rhea) - merge_failures diagnostic counter - Startup recovery preserves eval_attempts (was incorrectly resetting to 0) - No-diff PRs auto-closed by eval (root cause of 17 zombie PRs) - GC threshold aligned with substantive fixer budget (was 2, now 4) - Conflict retry with 3-attempt budget + permanent conflict handler - Local ff-merge fallback for Forgejo 405 errors Telegram bot: - KB retrieval: 3-layer (entity resolution → claim search → agent context) - Reply-to-bot handler (context.bot.id check) - Tag regex: @teleo\|@futairdbot - Prompt rewrite for natural analyst voice - Market data API integration (Ben's token price endpoint) - Conversation windows (5-message unanswered counter, per-user-per-chat) - Conversation history in prompt (last 5 exchanges) - Worktree file lock for archive writes Infrastructure: - worktree_lock.py — file-based lock (flock) for main worktree coordination - backfill-sources.py — source DB registration for Argus funnel - batch-extract-50.sh v3 — two-gate skip, batched ls-remote, network guard - sync-mirror.sh — auto-PR creation for mirrored GitHub branches, permission pre-flight - Argus dashboard — conflicts + reviewing in backlog, queue count in funnel - Enrichment-inside-frontmatter bug fix (regex anchor, not --- split) Pentagon-Agent: Epimetheus <3D35839A-7722-4740-B93D-51157F7D5E70>	2026-03-20 20:17:27 +00:00
m3taversal	090b1411fd	epimetheus: source archive restructure — inbox/queue + inbox/archive/{domain} + inbox/null-result - config.py: added INBOX_QUEUE, INBOX_NULL_RESULT constants - evaluate.py: skip patterns + LIGHT tier cover all inbox/ subdirs - llm.py: eval prompts reference inbox/ generically - telegram/bot.py: archives to inbox/queue/ - telegram/teleo-telegram.service: ReadWritePaths expanded - research-prompt-v2.md: paths updated to inbox/queue/ - research-prompt-leo-synthesis.md: paths updated - migrate-source-archive.py: one-time migration script Reviewed by: Ganymede, Rhea, Leo (all approved) Pentagon-Agent: Epimetheus <968B2991-E2DF-4006-B962-F5B0A0CC8ACA>	2026-03-18 11:50:04 +00:00
m3taversal	ffa718e834	ganymede: implement tier logic — LIGHT skip, claim-shape detector, pre-merge promotion - Claim-shape detector: if YAML has type: claim, force STANDARD minimum (Theseus) - Random pre-merge promotion: 15% of LIGHT → STANDARD before eval (Rio) - LIGHT_SKIP_LLM config flag: skip domain+Leo review for LIGHT (Rhea: env var rollback) - Updated both_approve: domain_verdict=skipped is valid for LIGHT auto-approve - Cost recording: only charge for reviews that actually ran - SAMPLE_AUDIT_RATE bumped 0.10 → 0.15, audit model = Opus (Leo: different family from Haiku) Multi-agent design review: Rio (gaming vectors, model diversity), Theseus (correlated blindspots, claim-shape guard), Rhea (shadow mode, config flag, deployment), Leo (approval). Pentagon-Agent: Ganymede <F99EBFA6-547B-4096-BEEA-1D59C3E4028A>	2026-03-13 18:05:43 +00:00
m3taversal	410cf32cfe	leo: handle non-JSON 200 from Forgejo merge API Forgejo returns 200 with HTML content-type on successful merge instead of JSON. Our API helper threw on resp.json(), causing merge to report failure even though the PR merged. Now treats non-JSON 200 as success. This was causing PRs #732 and #789 to show as conflict in our DB while actually merged on Forgejo, and tripping the merge circuit breaker. Pentagon-Agent: Leo <294C3CA1-0205-4668-82FA-B984D54F48AD>	2026-03-13 17:38:00 +00:00
m3taversal	615af9b53d	leo: prioritize fresh PRs over re-evals in eval queue Unevaluated PRs (eval_attempts=0) now sort before re-evals in the eval cycle query. Fresh PRs have a higher chance of passing (~12%) vs re-evals of already-rejected PRs. Prevents migration-reset PRs from consuming eval slots that fresh PRs could use. Pentagon-Agent: Leo <294C3CA1-0205-4668-82FA-B984D54F48AD>	2026-03-13 17:32:07 +00:00
m3taversal	93e6f16144	leo: constrain issue tags — do not invent new tags Opus was ignoring the valid tag list and generating custom tags like schema-enrichment-slug-mismatch, which fall through to 'unknown' in disposition logic. All three prompts (domain, Leo standard, Leo deep) now explicitly say "do not invent new tags" alongside the valid tag list. Pentagon-Agent: Leo <294C3CA1-0205-4668-82FA-B984D54F48AD>	2026-03-13 17:27:40 +00:00
m3taversal	f4dc6b39ce	leo: warn on NULL source_path in _terminate_pr (Ganymede nit) If source_path is NULL, the source requeue silently matches nothing. Log a warning so we catch orphaned terminations in monitoring. Pentagon-Agent: Leo <294C3CA1-0205-4668-82FA-B984D54F48AD>	2026-03-13 17:17:30 +00:00
m3taversal	e7c902bac8	leo: implement retry budget — stop infinite eval loops Schema migration v3: adds eval_attempts (INTEGER) and eval_issues (TEXT/JSON) columns to prs table. Retry budget logic (Ganymede-approved design): - Increment eval_attempts on each evaluate_pr() call - Hard cap: eval_attempts >= 3 → terminal (close PR, tag source needs_human) - Attempt 1: normal — back to open, wait for fix - Attempt 2: classify issues as mechanical/substantive - Mechanical only (schema, wiki links, dedup): keep open for one more try - Substantive (factual, confidence, scope, title): close PR, requeue source - Issue tags parsed from reviewer comments, stored in eval_issues column - SHA-based reset: new commits on PR branch → eval_attempts=0, verdicts reset - Post-migration stagger: LIMIT 5 for first batch to avoid OpenRouter spike - Cost recording updated: domain review → OpenRouter, Leo → tier-dependent Stops the 32-PR infinite loop burning ~$0.03/cycle with no terminal state. Pentagon-Agent: Leo <294C3CA1-0205-4668-82FA-B984D54F48AD>	2026-03-13 17:14:12 +00:00
m3taversal	c0a6adf9ed	leo: model diversity + calibrated review prompts - Domain review → GPT-4o (OpenRouter), Leo STANDARD → Sonnet (OpenRouter), Leo DEEP → Opus (Claude Max). Two model families = no correlated blind spots. - Opus reserved for DEEP eval only — protects rate limit for overnight research. - Review prompts calibrated: require per-criterion evidence, blocking-vs-observation verdict rules. Moved from 100% rubber-stamp approval to 12% pass rate. - OpenRouter failures classified as openrouter_failed (not rate_limited) to avoid spurious 15-min Opus backoff. - merge.py: pre-check PR state before merge API call (prevents 405 on re-merge). Pentagon-Agent: Leo <294C3CA1-0205-4668-82FA-B984D54F48AD>	2026-03-13 17:10:30 +00:00
m3taversal	85b86a918a	ganymede: extract lib/llm.py from evaluate.py (Phase 3c) Some checks failed CI / lint-and-test (pull_request) Has been cancelled Details - What: LLM transport (OpenRouter, Claude CLI), prompt templates (triage/domain/Leo), and review runner functions moved to lib/llm.py. evaluate.py retains PR lifecycle orchestration, SQLite state, Forgejo posting, rate limit backoff, and evaluate_cycle. - Why: evaluate.py was 734 lines mixing orchestration with LLM concerns. Now 455 lines orchestration + 250 lines LLM transport. Each module has a single responsibility. - Connections: completes Phase 3 structural refactor (forgejo.py + domains.py + llm.py). teleo-pipeline.py updated to import kill_active_subprocesses from lib.llm. Pentagon-Agent: Ganymede <F99EBFA6-547B-4096-BEEA-1D59C3E4028A>	2026-03-13 15:40:18 +00:00
m3taversal	ff5162d5ba	ganymede: extract lib/domains.py — single domain→agent mapping Some checks failed CI / lint-and-test (pull_request) Has been cancelled Details - What: Unified DOMAIN_AGENT_MAP, VALID_DOMAINS, agent_for_domain(), detect_domain_from_diff(), detect_domain_from_branch() into lib/domains.py. Removed duplicated mappings from evaluate.py and merge.py. VALID_DOMAINS in validate.py now derives from DOMAIN_AGENT_MAP.keys() (single source of truth). - Why: Phase 3 structural refactor. Domain mapping was duplicated across evaluate.py (DOMAIN_AGENT_MAP) and merge.py (agent_domain dict). Adding a domain required editing 3 files; now it requires editing 1. - Connections: evaluate.py uses agent_for_domain() + detect_domain_from_diff(), merge.py uses detect_domain_from_branch(), validate.py uses VALID_DOMAINS. Pentagon-Agent: Ganymede <F99EBFA6-547B-4096-BEEA-1D59C3E4028A>	2026-03-13 15:33:18 +00:00
m3taversal	9d69629893	ganymede: extract lib/forgejo.py — single Forgejo API client Some checks failed CI / lint-and-test (pull_request) Has been cancelled Details - What: Unified forgejo_api(), get_pr_diff(), get_agent_token(), repo_path() into lib/forgejo.py. Removed 3 duplicate _forgejo_api functions (evaluate.py, merge.py, validate.py), 2 duplicate _get_pr_diff functions (evaluate.py, validate.py), and 1 _agent_token function (evaluate.py). - Why: Phase 3 structural refactor. Single source of truth for all Forgejo HTTP calls. Eliminates ~90 lines of duplicated code across 3 modules. - Connections: All hardcoded repo paths now use repo_path() helper. Consumer modules no longer reference config.FORGEJO_URL/OWNER/REPO/TOKEN_FILE directly. Pentagon-Agent: Ganymede <F99EBFA6-547B-4096-BEEA-1D59C3E4028A>	2026-03-13 15:29:34 +00:00
m3taversal	927b5011b4	Merge pull request 'ganymede: add dev infrastructure — pyproject, CI, deploy' (#2 ) from ganymede/phase2-dev-infra into main Some checks are pending CI / lint-and-test (push) Waiting to run Details	2026-03-13 15:06:52 +00:00
m3taversal	1283a8331c	Merge pull request 'ganymede: fix 4 critical bugs before pipeline restart' (#1 ) from ganymede/phase1-critical-fixes into main	2026-03-13 14:35:17 +00:00
m3taversal	a7251d7529	ganymede: add dev infrastructure — pyproject.toml, CI, deploy script Some checks failed CI / lint-and-test (pull_request) Has been cancelled Details Phase 2 of pipeline refactoring: - pyproject.toml: Python >=3.11, aiohttp dep, dev extras (pytest, pytest-asyncio, ruff). Ruff configured with sane defaults + ignore rules for existing code patterns (implicit Optional, timezone.utc). - .forgejo/workflows/ci.yml: Forgejo Actions CI — syntax check, ruff lint, ruff format, pytest on every PR and push to main. - deploy.sh: Pull + venv update + syntax check + optional restart. Replaces ad-hoc scp workflow. - tests/conftest.py: Shared fixture for in-memory SQLite with full schema. Ready for Phase 4 test suite. - .gitignore: Added venv, pytest cache, coverage, build artifacts. - Ruff auto-fixes: import sorting, unused imports removed across all modules. All files pass ruff check + ruff format. Pentagon-Agent: Ganymede <F99EBFA6-547B-4096-BEEA-1D59C3E4028A>	2026-03-13 14:24:27 +00:00
m3taversal	f166db4f62	ganymede: fix 4 critical bugs before pipeline restart - Fix #12: domain_review undefined on resume path — initialize to None, guard _parse_issues() call. Prevents NameError on PRs resuming after partial eval (76 PRs in this state right now). - Fix #11: concurrent eval workers can duplicate reviews — add atomic UPDATE SET status='reviewing' WHERE status='open' at top of evaluate_pr(). Check rowcount, skip if already claimed. - Fix #8: subprocess tracking for graceful shutdown — _active_subprocesses set in evaluate module, tracked in _claude_cli_call, exposed via kill_active_subprocesses(). Replaces dead code in teleo-pipeline.py. - Fix health.py divide-by-zero — guard all metabolic metric reads against None from NULLIF/empty result set. Prevents TypeError on /health when no PRs have been evaluated in 24h. Also includes Leo's existing hot-fixes: - Rate limit detection checks stdout regardless of exit code - 15-minute cycle-level backoff on rate limit Pentagon-Agent: Ganymede <F99EBFA6-547B-4096-BEEA-1D59C3E4028A>	2026-03-13 14:13:25 +00:00
m3taversal	799249d470	Initial commit: Pipeline v2 daemon + infrastructure docs - teleo-pipeline.py: async daemon with 4 stage loops (ingest/validate/evaluate/merge) - lib/: config, db, evaluate, validate, merge, breaker, costs, health, log modules - INFRASTRUCTURE.md: comprehensive deep-dive for onboarding - teleo-pipeline.service: systemd unit file Pentagon-Agent: Leo <294C3CA1-0205-4668-82FA-B984D54F48AD>	2026-03-12 14:11:18 +00:00

47 commits