Skip to content

Share validated TGF parsing across training pipelines - #1

Merged
lmlearning merged 1 commit into
mainfrom
fix/shared-tgf-input-validation
Sep 25, 2026
Merged

lmlearning merged 1 commit into
mainfrom
fix/shared-tgf-input-validation

Conversation

@lmlearning

Copy link
Copy Markdown
Owner

Blank lines previously became empty argument identifiers, and the refined parser mishandled tab-separated attacks. Both training entry points now use a dependency-free parser for the unlabelled TGF subset used by the argumentation datasets.

Preserve input order, isolated nodes, self-attacks and the public wrappers' return types. Reject duplicate identifiers, missing/repeated separators, malformed attacks and undeclared endpoints before NetworkX can silently create extra nodes. Add component CI and document the supported subset.

Validation: 10 parser regression tests pass; both training entry points and the shared parser compile. No model training, checkpoint downloads or GPU validation was performed.

@lmlearning
lmlearning merged commit 65a2e8a into main Sep 25, 2026
2 checks passed
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant