Skip to content
#

deterministic-scoring

Here is 1 public repository matching this topic...

Community-driven behavioral reliability benchmark for LLMs. 231 probes across 19 modules, deterministic scoring, perplexity correlation, layer sensitivity mapping, quant method capture, hardware-stratified community rankings. Every test contributes to the community dataset.

  • Updated Apr 17, 2026
  • Python

Improve this page

Add a description, image, and links to the deterministic-scoring topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the deterministic-scoring topic, visit your repo's landing page and select "manage topics."

Learn more