Skip to content

M1: the catalog model and sync #7

Description

@tamnd

The catalog model and sync. No model calls.

The gate for everything downstream is byte fidelity. A writer that reformats is a writer that produces 548 files of churn on the first run and hides every real change afterwards.

Checklist

  • catalog.py: read and write, segment_id, entry identity on (msgctxt, msgid), header writing
  • Byte-fidelity test: read all 548 files, write them back unchanged, git diff empty
  • memory.py: the store, the four source values, precedence
  • tm rebuild from provenance comments, once M6 writes the comments there is something to rebuild from
  • sync.py: pin to tamnd/python-docs-vi branch 3.15 at da475ff, manifests/upstream.yaml, counts written off the files
  • sync --human: load the 1 435 human msgstr values into the memory as source=human
  • sync diff report: added and orphaned, written to reports/sync.md
  • stale.py: the three staleness causes, upstream text change, glossary bump, prompt change
  • Tests over hand-written fixtures, property test on segment_id stability

Exit

pydocvi sync writes a pin recording 548 files, 87 008 entries and 1 711 382 English words. Round-trip of all 548 files is byte-identical. The memory holds 1 435 human segments. The real numbers go in a comment on this issue.

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    cataloggettext catalogs, entry identity, the translation memorydeterministicNo model calls, reproducible from the corpus alonemilestoneA milestone tracking issue, M0 through M10

    Projects

    No projects

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions