Skip to content

RSI: validate live LoRA or QLoRA training and paired evaluation #9

Description

@w4ffl35

Implemented

PR #14 added a fixed OpenShell QLoRA fixture-smoke command, training scripts, model and dataset provenance, and paired base/adapter evaluation wiring. Mock-queue tests cover accepted, partial, failed, and incomplete pairs. The training image was built; the available GGUF is rejected as a training checkpoint.

Remaining work

Run the documented bounded command with a pinned supported Hugging Face checkpoint and verified dataset. Record the produced adapter hash, seed, configuration, model and dataset versions, and paired base/adapter cell outcomes. Confirm that failed or partial training remains ineligible for selection. No live adapter training and paired evaluation has yet been documented.

Acceptance criteria

  • A documented command can train one bounded adapter from a verified dataset.
  • The produced artifact is evaluated against the base model on the same harness cells.
  • Failed or partial training is recorded without making the candidate eligible for selection.

Activity

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Metadata

Metadata

Assignees

No one assigned

    Labels

    enhancementNew feature or request

    Type

    No type

    Projects

    No projects

      Milestone

      No milestone

      Relationships

      None yet

      Development

      No branches or pull requests

      Issue actions