AIAny
Icon for item

ChartGalaxy

Provides 1.7M+ synthetic and real infographic charts paired with their tabular data for training and evaluating multimodal models on infographic understanding, chart-to-table extraction, chart code generation, and example-based chart synthesis.

Introduction

Most chart datasets focus on plain plots; infographic charts combine visual motifs, icons and rich textual annotations that break assumptions used by many vision–language models. ChartGalaxy addresses this gap by delivering large-scale, design-aware chart data that ties rendered infographic images to the underlying tables and extracted layout templates, so models can learn both visual design cues and the exact data-to-visual mapping.

What Sets It Apart
  • Scale + paired supervision: a multi-million–sample collection of synthetic and real infographic charts where each image is paired with the tabular data that generated it — enabling direct chart↔table supervision for VQA and data extraction tasks, and objective evaluation of code-generation outputs.
  • Design-grounded synthesis: templates, chart types, and style variations are induced from real infographic designs and then used in a human-in-the-loop pipeline to create diverse synthetic charts — so the dataset preserves real-world layout diversity while scaling to millions of examples.
  • Multi-task utility: includes VQA-style QA pairs, layout templates, chart code benchmarks and example-based generation setups, making it useful for fine-tuning LVLMs, benchmarking chart-to-code systems, and example-driven chart synthesis.
Who It's For and Tradeoffs

Great fit if you want to fine-tune or benchmark multimodal models on infographic-style chart understanding, build chart-to-table extractors, or evaluate chart code generation under real design variability. Look elsewhere if you only need plain plotting libraries or small curated scientific plots — ChartGalaxy emphasizes infographic designs (icons, decorative text, complex layouts). Note practical constraints: the Hugging Face card indicates a non-commercial license for the dataset distribution and real-source images may require observing original copyright terms; synthetic portions are intended for broader reuse but verify license details before commercial use.

Information

  • Websitehuggingface.co
  • AuthorsZhen Li, Duan Li, Yukai Guo, Xinyuan Guo, Bowen Li, Lanxi Xiao, Shenyu Qiao, Jiashu Chen, Zijian Wu, Hui Zhang
  • Published date2025/05/07

Categories

More Items

Hugging Face

Provides agentic instruction‑tuning trajectories for software‑engineering tasks, formatted for supervised fine‑tuning and agent training. Contains multi‑file edits, tests, docs and structured agent traces (≈5,115 records, 1.9 GiB). Intended for commercial use; licensed CC‑BY 4.0 with additional permissive licenses.

Hugging Face

Provides raw, unscripted first-person household video footage for training vision and embodied AI models. Released incrementally on Hugging Face in WebDataset shards with metadata parquets under Apache‑2.0; current raw tier contains ~7,834 hours (≈397k videos).

Provides a year-scale multimodal benchmark and evaluation framework for on-device long-term memory in personal assistants, built from real mobile user trajectories. Tests memory construction, retrieval, updating, temporal reasoning, and implicit preference inference, and includes a knowledge-grounded synthesis pipeline to form coherent long-horizon trajectories.