Popular repositories Loading
-
-
compartmentalized-harm-research
compartmentalized-harm-research PublicCode, protocols, results, and verified artifacts for research on compartmentalized harm in multi-agent systems.
Python
Repositories
Showing 3 of 3 repositories
- compartmentalized-harm-research Public
Code, protocols, results, and verified artifacts for research on compartmentalized harm in multi-agent systems.
- parrhesia Public
Replaces sycophancy with truth-telling in open-weight LLMs via Aristotelian virtue training, a third path beyond RLHF and Constitutional AI. Ships a LoRA adapter, the training method, and a model-agnostic benchmark.
- parrhesiastes Public
People
This organization has no public members. You must be a member to see who’s a part of this organization.
Top languages
Loading…
Most used topics
Loading…