method
active
method:distillation-stageDistillation Stage
Stage 2 of character training: DPO from teacher model to student model to transfer desired behavioral expressions
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- First stage of DiffLogic CA update where each cell gathers information from neighboring cells via logic gate kernels
- Second stage of DiffLogic CA where a DLGN computes each cell's new binary state from perception output and current state
- Stage 3 of character training: SFT on synthetic introspective data generated by post-distillation checkpoint
- The perspective that LLM inference decomposes into distinct computational stages, which the paper extends to looped models
- Prior mechanistic interpretability work reverse-engineering vision models (InceptionV1); the direct predecessor this paper extends to language models
- Change between ordered and disordered macroscopic phases; existence depends on topology
- Formalization of anchoring as posterior allocation over pattern clusters followed by output generation.