Skip to content
#

benchmark-dataset

Here are 35 public repositories matching this topic...

AdvSV stands as the first dataset developed specifically for evaluating Speaker Verification (SV) systems against adversarial attacks. It aims to benchmark the robustness of ASV models in the face of such attacks and offers vital resources for researchers to explore the characteristics of adversarial and replay attacks in this domain.

  • Updated Nov 21, 2023
  • HTML
agent-trajectory-sentinel

Real-time detection and repair of LLM agent failures — a one-class behavioural monitor at ~200 µs/step, with 2,823 committed traces.

  • Updated Sep 3, 2026
  • Python

Benchmarking seven forecasting approaches on AgriPriceBD — a novel daily agricultural commodity price dataset of five commodities for Bangladesh. Covers BiLSTM, Transformer, Time2Vec ablation, Prophet and Informer failure analysis, with Diebold-Mariano significance testing.

  • Updated Mar 25, 2026
  • Jupyter Notebook

Add this topic to your repo

To associate your repository with the benchmark-dataset topic, visit your repo's landing page and select "manage topics."

Learn more