AIAny
Icon for item

MR-RATE

Paired brain MRI scans and radiology text annotations for multimodal vision–language research. Provides image-level labels and image–text pairs suited for VQA, classification, and image-to-text tasks; CC BY-NC-SA 4.0 and ~10K–100K samples — research/non-commercial use.

Introduction

Clinical MRI datasets with reliable paired text labels are scarce, yet they are essential for training and evaluating multimodal medical models. MR-RATE addresses that gap by providing a curated collection of brain MRI volumes linked to radiology-style annotations and ratings, enabling tasks from visual question answering to image classification and image-to-text generation.

What Sets It Apart
  • Multimodal focus: combines MR image volumes (or slices) with structured text annotations and ratings, so models can learn cross-modal alignment rather than only pixel-level features — useful when evaluating vision–language alignment for clinical findings.
  • Task-ready splits: organized to support VQA, image-to-text, classification, and zero-shot evaluation, which reduces preprocessing overhead for benchmarking clinical foundation models.
  • Practical scale and license: an intermediate-size medical dataset (~10K–100K items) licensed CC BY-NC-SA 4.0, balancing accessibility for academic research with non-commercial constraints.
Who It's For & Trade-offs

Great fit if you are developing or evaluating multimodal medical/vision-language models, probing clinical reasoning in foundation models, or benchmarking VQA and image-to-text approaches on MRI data. Look elsewhere if you need fully de-identified, hospital-grade DICOM metadata for deployment, larger-scale population cohorts, or a permissive commercial license — the CC BY-NC-SA 4.0 terms and dataset scope limit production-use and very large-scale training.

Data & practical notes

The dataset emphasizes clinically relevant MRI content paired with short reports/ratings rather than exhaustive clinical histories. Expect common medical-data caveats: verify provenance, comply with institutional policies for human-data use, and confirm whether provided labels meet your annotation quality requirements before using for model training or evaluation.

Information

Categories

More Items

Hugging Face

Evaluates retrievers and search agents on synthetic multi-hop questions that require assembling a complete set of supporting evidence. Provides English and Russian variants (395 questions each), a fixed dense index embedded with Qwen3-Embedding-8B, and BrowseComp-Plus evaluation integrations.

Hugging Face

Provides re-annotated academic video instruction data for captioning, video QA, and fine-grained motion understanding; rewrites short answers and concise captions into evidence-grounded, instruction-following responses and supplies JSONL annotation files (original videos not included).

Hugging Face

Provides 324 Russian short-answer web-search tasks with gold supporting documents to evaluate fixed-index retrievers and search agents. Tasks span eight topical categories and five retrieval challenge types (multihop, structured evidence, temporal, entity disambiguation, comparative) and use a Qwen3-Embedding-8B index for evaluation.