Skip to content

Does the graph lane help on coding sessions, not just LongMemEval? #395

Description

@antejavor

Part of #390

Question

Every measurement so far is on LongMemEval: personal-assistant chat, one user, facts about the user. The context graph's own target is coding-agent sessions. Is there a corpus, even a small hand-labelled one from real Claude Code sessions, on which the typed graph's lanes can be measured? Does the hand vocabulary even fit it?

Output

A corpus proposal with a handful of questions and gold answers, or a finding that the graph lanes can't be judged on coding sessions yet.

No activity

Activity on this issue will appear here.

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions