method
active
method:lora-fine-tuningLoRA Fine-Tuning
Adaptation method used via Tinker API for DeepSeek-V3.1 and Qwen3-235B fine-tuning with rank 32
Neighborhood — ranked by edge-count
Papers (1)
paper
Methods (1)
method
- LoRA Fine-Tuning with Axolotlrelated_toSpecific fine-tuning implementation using LoRA rank 32, learning rate 2e-4, AdamW 8-bit optimizer
Related by similarity (8)
cosine ≥ 0.65 · no typed edgeEntities in the same semantic neighborhood but without a typed relation to this one — candidates for new edges or unrecognized duplicates.
- Low-rank adaptation method used for finetuning models in all experiments; rank 32, alpha 64
- Parameter updates that reduce mismatch dr; another anchoring variant in UCCT.
- The patient, hand-guided adjustment of shape and dimension to each unique condition in a building; requires materials that make it economical and easy.
- The literature documenting how fine-tuning can compromise safety alignment even without malicious intent
- Training procedure that consistently increases HH-intent strength and consistency across model families.
- OpenAI's internal RL fine-tuning API used to train models with graders rewarding correct or incorrect responses
- Mechanistic explanation of how fine-tuning can shift persona vectors without directly updating activations
- Fine-tuning for persona depth and emotional performance; actively suppresses self-observation