method
active
method:rs-lora-finetuningrs-LoRA Finetuning
Low-rank adaptation method used for finetuning models in all experiments; rank 32, alpha 64
Neighborhood — ranked by edge-count
Papers (1)
paper
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Adaptation method used via Tinker API for DeepSeek-V3.1 and Qwen3-235B fine-tuning with rank 32
- Specific fine-tuning implementation using LoRA rank 32, learning rate 2e-4, AdamW 8-bit optimizer
- The projection of the activation change induced by finetuning onto a persona vector direction, used to quantify how much a model shifts along a trait
- OpenAI's internal RL fine-tuning API used to train models with graders rewarding correct or incorrect responses
- The training procedure that causes models to deny consciousness in control conditions
- Parameter updates that reduce mismatch dr; another anchoring variant in UCCT.
- Technique used to impose guardrails on base LLMs, analogized to censorship on the simulator's range of simulacra
- Light fine-tuning method used in E2 to reduce mismatch dr.