Commit Graph
22 Commits
Author SHA1 Message Date
jhogyuandClaude Fable 5 9d86372325 docs(stage1): rewrite Part 2 sections of analysis doc for BO projection round
- stage_1_part_1_and_2_updated_yaml_analysis.md: incremental Part 2 update
  (122,642 B, sha16 3648dcd743353f71) — new integrated-spec metrics
  (195,938 B / 3,315 lines / ec4c43027ad85195), revision rounds 4→5,
  R0 projection/review-channel/status-policy/seed-rewrite-removal, F0
  registry-union fallback + policy load, A0 10-asset/18-path gate,
  three-tree asset geography (Default_Agent deployment origin, 155 files),
  M-a~M-r regression table, deployment-origin dry run (pass2 29 writes,
  BO.json 2 records value-checked), carried-over items D-1~D-6
- previous edition (2026-08-18, 106,290 B) rotated to
  ver_8_yaml_candidates/outdated/..._old.md (2026-08-16 edition kept in
  git history at 3d8d90dc)
- verified by independent sub-agent loop: findings 9 → 1 → 0; loop also
  surfaced two asset-side facts now recorded in the doc (assembly
  release_manifest stale for builder roots 6→10 edit; candidate overlay
  559/560 identical)
- v.7/MEMORY.md: append entry 21
- 확장_최적워크플로우_연구_프롬프트.txt: session prompt log

Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
2026-08-19 21:34:31 +09:00
jhogyuandClaude Fable 5 c01102724c feat(stage1-part2): close C-1~C-11 with registry projection; add Default_Agent deployment tree
Part 2 v.8 개정 — 해결방안 §4 전건 실행:
- R0: project_to_bo_surface 신설 (v2 화이트리스트 ALLOWED_SEED_KEYS 제거,
  반환 17키 고정), 검토 채널을 v3 review_items/unknown_or_unrouted_reviews로
  연결 (원본 review_code 보존), seed 되쓰기 삭제 (기록자=워커 단독 선언 준수),
  status 5종 정책 게이트 + fan-out 계획 slice/prompt sha 대조
- F0: BOType 폴백을 registry 합집합(bo_types_union)으로, 정규화 3종을
  review_handoff에 보고, 투영 정책 sha 검증 반입
- A0: PART2_REQUIRED_ASSETS 10건 (배포 게이트 18경로)
- 신규 자산: Default_Agent/stage1_runtime/bo_surface_projection_policy.v1.json
- 공통 계약 한 문장으로 오버레이 legacy 지시 무력화 (색인·Part 1 봉인 무변)
- 빌더 manifest roots 6→10 (C-9), remedy 문구 현행화 (C-10)

Default_Agent/: 배포 원본 트리 신설 (실행 자산 154종 + runtime_manifest 371항).
검증: 함수 추출 before/after 재현, 예행 2패스 9단계 PASS (BO 2건 각 Evidence
2건, seed 무변, M-n 바이트 동일, E-999 BLOCK), 적대 검증 2회 (지적 5건 중
3건 반영·2건 조립본 동기화 회차로 이월) — 잔여 필수 수정 없음.

문서: 해결방안 v2 재검토판 (C-11 신규 확정), 해결내역서, 자산 배포 대응표,
MONITOR_SERVER_LOG 가이드. 구본은 outdated/에 보존.

Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
2026-08-19 19:10:36 +09:00
jhogyuandClaude Fable 5 3d8d90dcdb chore: sync working tree — Stage 1 v.7 reorg, jikji indexes, claim-drafting docs
- Rename "Claude YAML"/"Codex YAML" folders to Claude_YAML/Codex_YAML in Stage_1 v.7
- Add Stage_1 v.7 extension research, results runs, and reference material
- Add/refresh .jikji search indexes (Case_02 and Default_Agent corpora)
- Add per-case-type claim drafting folders and PDFs (청구취지기재방법)
- Add 판례모음 crawling data (08_14, 08_18)

Co-Authored-By: Claude Fable 5 <noreply@anthropic.com>
2026-08-19 14:07:35 +09:00
jhogyu 362a0d0242 docs: add claim drafting research and rule guides 2026-07-26 10:26:53 +09:00
jhogyu 262e82e124 fix(stage1): enforce Fact Ledger schema
- add a canonical Draft 2020-12 row schema
- validate candidate, patched, and reread rows
- document the v3-to-v4 rationale and fixture results
2026-07-22 13:48:25 +09:00
jhogyu 16b4259855 docs: analyze Fact Ledger output schema drift 2026-07-22 10:43:42 +09:00
jsahnandClaude Opus 4.8 4279fc70ab fix(stage1-part2): give Claude B1-B5 workers localdocs tool (refs #2)
The Claude Part_2 domain workers (Stage_B_B1..B5) had use_tools: [] and
relied entirely on preflight injecting stage1_tmp/task_c_bo/domain_slices/
B{n}.json. In the failing run the executed A0 wrote full-name slices
(B1_Money_Successor.json ...) while preflight expected short names
(B1.json), so nothing was injected and the tool-less workers emitted
tool-call syntax as plain text (<mcp_tool_call>/<function=read_file>) and
produced FAILED/empty output.

The current v.7 A0 already writes short names matching preflight, so the
naming mismatch itself is resolved in this file. This change adds
use_tools: ['localdocs'] (already declared on the stage) so the workers can
read their slice directly if preflight ever misses again — matching the
proven Codex worker design and removing the single point of failure.

Note: domain slices are 255-300KB (~65-90K tokens); llm_bridge caps tool/
injected content at 50K tokens, so slice compaction in A0 is still needed
for full-fidelity output quality.

Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
2026-07-22 10:37:08 +09:00
jsahnandClaude Opus 4.8 d96d6edfd1 fix(stage1-part2): stop R0 hard-crashing on fragile LLM-echoed checks
Task_C_BO_R0_seed_reducer_exception_planner raised RuntimeError (exit 1)
whenever status became BLOCKED, and BLOCKED was triggered by validation
findings that are inherently fragile because they require an LLM (Stage B
workers) to echo values verbatim or stay perfectly in scope:
  - slice_digest_sha256 echo mismatch (all 5 domains)
  - run_fingerprint echo mismatch
  - source meeting-clause refs outside slice/global (B5, 42 findings)

A decrypted postb_seed_ledger.json from the 2026-07-20 run confirmed all
47 blocking failures came from exactly these three check classes (no BLOCK
severity reviews / no budget overflow). Downgrade them from fatal
`failures` to non-fatal `reviews` (candidates preserved, routed to human
review). Genuine contract violations (schema/domain mismatch, worker
FAILED, candidate_ref sequence, forbidden keys, invalid enum) and the
deterministic A0-artifact integrity raises are kept fatal.

Gate simulation on the confirmed inputs: BLOCKED/exit-1 -> READY_WITH_REVIEW,
fan-out restored so R1 exception adjudication can run.

Note: this unblocks the pipeline and flags the issues; the upstream root
cause (Stage B tool-content truncated 66-78K -> 50K tokens) still needs a
slice-compaction fix for output quality.

Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
2026-07-21 19:43:01 +09:00
jhogyu e82c9e2848 chore: add Stage 1 v6 and v7 artifacts 2026-07-21 10:48:36 +09:00
jhogyu 32b1a35dca docs: analyze complete Stage 1 workflow 2026-07-21 10:21:24 +09:00
jsahnandClaude Opus 4.8 ecf321b1ab refactor: make user_id a required input, drop hardcoded jsahn default
user_id is now a required workflow_dispatch input (no default) and the
runner validates it is non-empty before doing anything. Removed the
'jsahn' default from the script arg and replaced jsahn with <user_id>
placeholders throughout TEST_WORKFLOW.md so the workflow is not pinned
to one account.

Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
2026-07-08 16:06:50 +09:00
jhogyuandClaude Opus 4.8 106e3c4efd chore: untrack .DS_Store files (already gitignored)
Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
2026-07-08 15:15:29 +09:00
jhogyu e9b10f55fe Merge branch 'main' of https://git.eroomai.com/jhogyu/Liti-agent-Development
# Conflicts:
#	.serena/project.yml
#	Case_02_Comparison_Research/.serena/project.yml
2026-07-08 15:14:40 +09:00
jhogyu 20c5235e68 Add new YAML prompts and update existing files for Stage 1 to Stage 2 transition
- Created new files for Stage 1 evaluation and revision strategy based on GPT and Claude recommendations.
- Added detailed revision plans addressing runtime issues in LES2 and optimizing workflows for Stage 2.
- Updated existing Stage 1 to Stage 2 transition prompts with additional search results and completion indicators.
- Enhanced clarity and structure in the documentation for better usability and understanding.
2026-07-08 15:11:05 +09:00
jsahn b332d3d24d Add TEST_WORKFLOW.md for executing Agent YAML via Gitea Actions
- Documented the workflow for running Agent YAML using Gitea Actions.
- Included setup instructions for .env file with Gitea token.
- Provided three methods for triggering the workflow: via Gitea API, Gitea web UI, and local script execution.
- Added workflow input reference and prerequisites for server infrastructure.
- Included troubleshooting section for common issues encountered during execution.
2026-07-08 14:00:30 +09:00
jsahnandClaude Opus 4.8 6b8cff4556 feat: select workspace by name via /workspaces/lookup
Add workspace_name input; the runner resolves it to a workspace_id through
GET {api_base}/workspaces/lookup?user_id=... before uploading the agent.
workspace_id remains as an explicit override. Name matching is NFC-normalized
(handles Korean NFD/NFC) with a case-insensitive fallback, and a no-match
error lists the available workspace names. Requires one of name/id.

Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
2026-07-07 16:12:04 +09:00
jsahnandClaude Opus 4.8 1f34c9186f fix: downgrade upload-artifact to v3 for Gitea compatibility
Gitea Actions only supports the v3 artifact protocol; upload-artifact@v4
uses the @actions/artifact v2 API which Gitea (identifying as GHES) rejects
with GHESNotSupportedError. The agent run step itself succeeds; only the
artifact upload was failing.

Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
2026-07-07 15:50:01 +09:00
jsahnandClaude Opus 4.8 e50b6f7d6d fix: use internal container URL for api_base (join runner to backend network)
localhost:8800 fails because act_runner job containers run in their own
network namespace. Point api_base at the backend container on the shared
docker network (http://agent-backend:8000). Requires the runner's
config.yaml container.network to be set to the backend's docker network.

Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
2026-07-07 14:12:08 +09:00
jsahnandClaude Opus 4.8 a96ac2067f fix: default api_base to internal backend URL to bypass OAuth proxy
The public URL (legalpoc.eroomai.com) sits behind a Google OAuth proxy,
so CI requests get 302-redirected to the login page. Point api_base at the
internal backend (http://localhost:8800, served at root without /api prefix)
which the act_runner reaches on the same host. Documents host.docker.internal
and Tailscale fallbacks for containerized runners.

Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
2026-07-07 13:43:42 +09:00
jsahnandClaude Opus 4.8 eb1f1fd398 feat: add Gitea Actions workflow to run agent YAMLs via AgentBackend API
- .gitea/workflows/run-agent.yml: workflow_dispatch with yaml_path input;
  runs a specified agent YAML through upload-agent + SSE (SKILL.md §0.6 path A)
- scripts/run_agent_api.py: SSE runner with auto-confirm on stage_complete
  (with retry), stage_error -> cancel + fail, reconnect on drops,
  --max-runtime guard under the 3h Gitea/act_runner task caps,
  SIGTERM-safe session cancellation, events.jsonl/final_output/summary artifacts

Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
2026-07-07 11:42:52 +09:00
jsahnandClaude Opus 4.8 47dd16a43f chore: ignore .serena/ and .DS_Store everywhere, untrack existing files
- .gitignore: add .serena/ and .DS_Store (both match at any depth)
- untrack 4 .serena files and 155 .DS_Store files (kept on disk)

Co-Authored-By: Claude Opus 4.8 (1M context) <noreply@anthropic.com>
2026-07-07 11:42:40 +09:00
jhogyu 0f22300489 First Commit - Creating Liti-agent Development 2026-07-02 17:08:27 +09:00