Skip to content

[W1.3] Evaluation (LLM-as-Judge) Realizing-in-CC (Tier A) #310

Description

@julianken

Action ID

W1.3 · Wave 1 · Tier A · Type: satellite_authoring · Estimated: 5h

Scope

Single-file edit; ~750w rendered; worked example PR #290 (2-cycle bot review with R8 second-pass mandatory-find).

Rationale

PR #290 is the canonical LLM-as-judge production case with R8 second-pass; this is the highest-friction Monday test in wave 1 (mint PAT, write 12-rule rubric SKILL.md, gate merge queue) so wave-1 sequencing puts it last after the spine has framed the cost. Forward-link to W2.6 scoper rank-1 article omitted per [draft]-discipline; resolves additively in W2.6.

Files touched

  • /Users/jul/repos/detached-node/src/data/agentic-design-patterns/patterns/evaluation-llm-as-judge.ts

Acceptance criteria

Prerequisites

  • W0.2
  • W1.0

Blocks

  • W2.6
  • W2.9

Rollback plan

Revert single-file edit.

Anti-example evidence (if applicable)

(none)


Tracked in: Epic #302
Source: docs/agentic-bridge/reframe/action-plan-v1.json action W1.3

Metadata

Metadata

Assignees

No one assigned

    Labels

    bot:approvedjulianken-bot has approved this issuestatus:in-progressActively in progresstier:ATier A — full case-studywave:1Wave 1 — foundation + spine

    Projects

    No projects

    Milestone

    No milestone

    Relationships

    None yet

    Development

    No branches or pull requests

    Issue actions