[CVPR-2024] The First High Definition (HD) Event based Visual Object Tracking Benchmark Dataset
-
Updated
Mar 25, 2026 - Python
[CVPR-2024] The First High Definition (HD) Event based Visual Object Tracking Benchmark Dataset
[NeurIPS 2024] Official code for HourVideo: 1-Hour Video Language Understanding
The MAMA-MIA Dataset: A Multi-Center Breast Cancer DCE-MRI Public Dataset with Expert Segmentations
[Pattern Recognition 2025] A large-scale benchmark dataset for color-event based visual tracking
This repository contains a gym environment that can be used for developing solvers for robotic 3D bin packing problems.
A High-Quality and Large-Scale Dataset for English-Vietnamese Speech Translation (INTERSPEECH 2022)
[IJCV-2026, arXiv:2408.09764] Event Stream based Human Action Recognition: A High-Definition Benchmark Dataset and Algorithms
Dataset package for facile training and testing of machine learning/AI algorithms that predict drug response in cancer model systems.
AdvSV stands as the first dataset developed specifically for evaluating Speaker Verification (SV) systems against adversarial attacks. It aims to benchmark the robustness of ASV models in the face of such attacks and offers vital resources for researchers to explore the characteristics of adversarial and replay attacks in this domain.
The SWAN-SF dataset is now fully preprocessed, optimized, and ready for binary classification tasks. Our team is excited to release the enhanced version of the SWAN-SF dataset across all five partitions.
[FAccT '25] Characterizing Bias: Benchmarking LLMs in Simplified versus Traditional Chinese
Documentation associated with preparing and formatting datasets LARRY datasets for ML applications with pytorch / pytorch lightning
Real-time detection and repair of LLM agent failures — a one-class behavioural monitor at ~200 µs/step, with 2,823 committed traces.
IAVS: A Multi-Center Dataset and Applicability Evaluation System for Computational Fluid Dynamics-Oriented Intracranial Aneurysm Segmentation (MICCAI 2026)
Code for LEMMA-RCA website
100 long-horizon reverse engineering tasks generated by ProgramSmith.
Reproducible synthetic financial crime data generator for transaction monitoring - customers, accounts, transactions and alerts, ready for analytics and model testing.
A model-agnostic benchmark and agent harness for evaluating AI reasoning capabilities, with a focus on geometry.
Collaborating to improve population dynamics models through benchmark dataset validation
Open topology-grounded benchmark for datacenter RCA, hidden-target localization, and counterfactual remediation validation.
To associate your repository with the benchmark-dataset topic, visit your repo's landing page and select "manage topics."