Benchmark Radar
AI BENCHMARK PROFILE

Prompting-MammAlps

General AIMultimodal PerceptionSearch & RetrievalEPFL CNAI

Prompting-MammAlps is a camera-trap text-to-video retrieval benchmark, evaluating video-language models on fine-grained retrieval of ecological events. It includes a test set of 135 queries and 775 candidate videos, with a proposed method that combines action localization with LLM-based parsing.

Released
2026-07-10
Readiness
Inspectable
Primary field
General AI

Why it matters

Text-to-video retrieval in ecological domains requires spatiotemporal understanding that current VLMs lack. This benchmark provides a standardized evaluation for fine-grained and interpretable retrieval, highlighting the limitations of zero-shot VLMs.

Motivation

Automatically retrieving videos from large camera-trap datasets remains challenging.

Primary resources

Benchmark Radar records only publicly supported details and links back to primary sources for verification.