parrhesia
Replaces sycophancy with truth-telling in open-weight LLMs via Aristotelian virtue training, a third path beyond RLHF and Constitutional AI. Ships a LoRA adapter, the training method, and a model-agnostic benchmark.
// repository documentation
Was this content helpful?
(0 ratings)
