Implemented
PR #14 added a fixed OpenShell QLoRA fixture-smoke command, training scripts, model and dataset provenance, and paired base/adapter evaluation wiring. Mock-queue tests cover accepted, partial, failed, and incomplete pairs. The training image was built; the available GGUF is rejected as a training checkpoint.
Remaining work
Run the documented bounded command with a pinned supported Hugging Face checkpoint and verified dataset. Record the produced adapter hash, seed, configuration, model and dataset versions, and paired base/adapter cell outcomes. Confirm that failed or partial training remains ineligible for selection. No live adapter training and paired evaluation has yet been documented.
Acceptance criteria
- A documented command can train one bounded adapter from a verified dataset.
- The produced artifact is evaluated against the base model on the same harness cells.
- Failed or partial training is recorded without making the candidate eligible for selection.
Implemented
PR #14 added a fixed OpenShell QLoRA fixture-smoke command, training scripts, model and dataset provenance, and paired base/adapter evaluation wiring. Mock-queue tests cover accepted, partial, failed, and incomplete pairs. The training image was built; the available GGUF is rejected as a training checkpoint.
Remaining work
Run the documented bounded command with a pinned supported Hugging Face checkpoint and verified dataset. Record the produced adapter hash, seed, configuration, model and dataset versions, and paired base/adapter cell outcomes. Confirm that failed or partial training remains ineligible for selection. No live adapter training and paired evaluation has yet been documented.
Acceptance criteria