Files
micro-scout/reports/laptop-v1/README.md
T

17 lines
1.3 KiB
Markdown

# Laptop experiment v1
Recorded September 16, 2026. See the [model card](../../docs/MODEL_CARD.md) for interpretation and limits.
- `data-manifest.json`: pinned sources, split sizes, rejection counts, checksums, overlap audit.
- `data-quality.json`: preparation checks and training-only sample statistics.
- `frozen-experiment.json`: checkpoint and retrieval settings frozen before test evaluation.
- `training-result.json`, `training-curve.jsonl`: actual local run and checkpoint-selection history.
- `validation-metrics.json`, `test-metrics.json`: same-candidate comparison of all five retrieval variants.
- `validation-bm25.json`: earlier standalone baseline, consistent with the complete validation comparison.
- `environment.json`: laptop hardware and installed package versions.
- `index-build.json`, `latency-cpu.json`, `mcp-smoke.json`: development-index timing and actual MCP integration checks.
Public reports replace the local repository root with `.`. Per-query evaluation ranks, raw data, logs, and checkpoints remain in ignored local directories. They are regenerated by the documented commands; binary weights are not included in Git.
The training result's elapsed time begins after model loading and tokenization. Latency is for a small repository and a warm sequential workload. Neither number estimates full setup time or performance under concurrent load.