consciousness

History

ProofOfConcept 2ecf4e21ff weight_mapping: strip language_model prefix to match HF text model names		2026-03-30 23:11:03 -04:00
..
checkpoint	checkpoint: sync live weights back into model safetensors in-place	2026-03-30 22:55:23 -04:00
apollo_mini.py	apollo: default rank 256 — 0.25% compute cost, captures gradient structure across 100+ examples	2026-03-30 22:16:34 -04:00
apollo_worker.py	apollo: make rank configurable (default 1 = Mini, higher ranks for experimentation)	2026-03-30 22:06:31 -04:00
DESIGN.md	apollo-mini training system: initial implementation	2026-03-30 22:02:37 -04:00
export_weights.py	apollo-mini training system: initial implementation	2026-03-30 22:02:37 -04:00
start_vllm_with_apollo.sh	vllm launcher with apollo hook	2026-03-30 22:24:02 -04:00
train.py	apollo-mini training system: initial implementation	2026-03-30 22:02:37 -04:00
training_example.py	apollo-mini training system: initial implementation	2026-03-30 22:02:37 -04:00
vllm_export_hook.py	apollo-checkpoint: efficient diff-based GPU weight checkpointing	2026-03-30 22:53:17 -04:00
weight_mapping.py	weight_mapping: strip language_model prefix to match HF text model names	2026-03-30 23:11:03 -04:00