Skip to content
#

model-auditing

Here are 26 public repositories matching this topic...

A production-grade LLM Evaluation & Benchmarking Framework for systematic model auditing. Features parallel benchmarking, fairness/bias detection, MMLU integration, and a real-time analytics dashboard powered by React and FastAPI.

  • Updated Apr 14, 2026
  • Python

🔬 1- A Human-Centered AI & Data Science hub for rigorous Machine Learning model evaluation, comparison, auditing, and selection, combining performance analysis, validation, hyperparameter optimization, calibration, reproducibility, and responsible real-world applications.

  • Updated Aug 10, 2026
  • Jupyter Notebook

Subgroup-stratified, calibration-aware fairness auditing for ML models: DeLong AUC confidence intervals, per-subgroup calibration error, multiple-comparison-corrected significance, and a novel five-axis cross-platform protocol (CPFE). Grounded in peer-reviewed methods.

  • Updated Aug 1, 2026
  • Python

Improve this page

Add a description, image, and links to the model-auditing topic page so that developers can more easily learn about it.

Curate this topic

Add this topic to your repo

To associate your repository with the model-auditing topic, visit your repo's landing page and select "manage topics."

Learn more