paper
referenced-only
2024
paper:doi-10-48550-arxiv-2405-11143OpenRLHF: An Easy-to-use, Scalable and High-performance RLHF Framework
ByJian Hu·Xibin Wu·Weixun Wang·Songlin Jiang·Dehao Zhang·Yu Cao+4 more
Similar preprints — Semantic Scholar
Cited by (1)
- Open Character Training: Shaping the Persona of AI Assistants through Constitutional AI
Character training—fine-tuning open-weights LLMs to internalize specific personas at a depth that survives adversarial pressure—proves substantially more effective than either system-prompt constraini