
lineage-eval
Model-agnostic benchmark, generation and LLM-judging harness, and interactive viewer for evaluating selective censorship in language models.
About
Model-agnostic benchmark, generation and LLM-judging harness, and interactive viewer for evaluating selective censorship in language models.
Languages
Contributors1
No features listed.
Comments Theme
Install
pip




