Skip to content
#

model-evaluation

Here are 1,952 public repositories matching this topic...

Open-source, end-to-end platform for evaluating, observing, and improving LLM and AI agent applications. Tracing · Evals · Simulations · Datasets · Gateway · Guardrails. Self-hostable. Apache 2.0.

  • Updated Sep 10, 2026
  • Python

AI-powered NBA game outcome predictor that uses advanced team stats and trend-based features to forecast winners and track model performance

  • Updated Sep 7, 2026
  • Jupyter Notebook
model-serving-minefield

Community registry of LLM serving-path traps that produce confidently wrong measurements: templates, tool parsers, reasoning fields, quant kernel paths, CUDA toolchains, KV allocation, eval harnesses, versioning. Symptom-first, with the check that catches each.

  • Updated Sep 9, 2026
  • Python

Customers in the telecom industry can choose from a variety of service providers and actively switch from one to the next. With the help of ML classification algorithms, we are going to predict the Churn.

  • Updated Dec 29, 2021
  • Jupyter Notebook

An in-depth analysis of audio classification on the RAVDESS dataset. Feature engineering, hyperparameter optimization, model evaluation, and cross-validation with a variety of ML techniques and MLP

  • Updated Nov 5, 2020
  • Jupyter Notebook

Add this topic to your repo

To associate your repository with the model-evaluation topic, visit your repo's landing page and select "manage topics."

Learn more