AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Category

Explore by categories

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.
  • All Categories

  • AI Leaderboard

  • AI Agent Tutorials

  • AI Coding Tutorials

  • AI Model

  • AI Agent Papers

  • Chatbot

  • AI Dataset

  • Machine Learning Foundation Books

  • AI Train

  • AI Deploy

  • AI Client

  • Machine Learning Foundation Papers

  • Machine Learning Foundation Tutorials

  • AI Image Demos

  • AI Agent

  • Large Language Model Tutorials

  • Large Language Model Papers

  • Machine Learning Engineering Papers

  • Computer Vision Tutorials

  • Computer Vision Papers

  • Natural Language Processing Papers

  • Reinforcement Learning Papers

  • Speech Technology Papers

  • AI API

  • AI Coding

  • AI Image

  • AI Video

  • MLOps

  • MCP Client

  • MCP Server

  • AI Video Papers

  • AI Audio

  • AI Others

  • AI Infra

  • Embodied AI

Hugging Face
AI Dataset·2026
Icon for item

DECOMEG — Brain Activity During Typing (MEG & EEG)

Jarod Lévy, Mingfang Zhang +5·Basque Center on Cognition, Brain and Language (BCBL), HybridMojo LLC +1

Provides de-identified MEG and EEG recordings of 35 native Spanish speakers typing memorized sentences, with synchronized behavioral logs and standardized event tables. Includes raw .fif and BrainVision files plus MATLAB logs (≈262 GB total); released under CC BY-NC 4.0 for non-commercial research on brain-to-text decoding.

#huggingface#science#ai#nlp#python
Hugging Face
AI Dataset·2026
Icon for item

Complete FABLE.5 Traces 2M

Crownelius

Provides a deduplicated 2.0M-row corpus of FABLE.5 / Mythos agent traces with row-level provenance and session-limit rows removed. Includes canonical Parquet and gzip JSONL exports, SHA256 row hashes, and provenance fields for tracing first-source datasets.

#huggingface#claude-code#vibe-coding#llm#nlp+4
Hugging Face
AI Dataset·2026
Icon for item

AFTER

Julia Belikova, Rauf Parchiev +5

Benchmark for evaluating procedural skill evolution in LLM agents: isolates reusable skill bodies, role-specific work surfaces, and hidden oracle assets to measure whether skill refinements transfer across tasks, roles, and model backbones. Includes 382 workplace tasks, 22 skills, and a controlled evaluation protocol.

#evaluation#agent-skills#huggingface#llm#ai-agent+2
Hugging Face
AI Dataset·2026
Icon for item

GLM-5.2 Agent traces

AletheiaResearch, TeichAI

Provides 319 newline-delimited JSON agent session traces captured from GLM-5.2 using Teich for training agentic models. Preserves reasoning-first assistant fragments, tool-call events, and a dataset-level training-ready tool schema; convertible to OpenAI-style JSONL for SFT/distillation.

#llm#ai-agent#huggingface#json#ai-train+1
Hugging Face
AI Dataset·2026
Icon for item

Nemotron-SFT-SWE-v3.5

NVIDIA

Provides agentic instruction‑tuning trajectories for software‑engineering tasks, formatted for supervised fine‑tuning and agent training. Contains multi‑file edits, tests, docs and structured agent traces (≈5,115 records, 1.9 GiB). Intended for commercial use; licensed CC‑BY 4.0 with additional permissive licenses.

#nvidia#huggingface#swe#software-engineering#coding-agents+6
Hugging Face
AI Dataset·2026
Icon for item

WGO-Bench

Macrodata Labs, InternRobotics +1

Provides a small, manually annotated benchmark for evaluating vision–language models that convert robot and egocentric manipulation videos into timestamped subtask segments and concise action labels. Contains 100 episodes, 743 gold segments, and MP4 bytes embedded per row.

#video#robotics#ai-video#evaluation#huggingface+2
Hugging Face
AI Dataset·2026
Icon for item

Syn4D: A Multiview Synthetic 4D Dataset

Zeren Jiang, Yushi Lan +9·Visual Geometry Group, University of Oxford, Nanyang Technological University +1

Provides multiview synthetic RGB video clips with per-frame depth, instance masks, dense long-range 3D point tracks, camera poses, and SMPL‑X human pose/shape labels for 4D reconstruction, tracking, and geometry-aware novel-view synthesis. Includes ~4.7K clips (1.4M frames) and is licensed for AI training.

#vision#depth#video#multimodal#huggingface+2
Hugging Face
AI Dataset·2026
Icon for item

MatrAIx Persona 1M

MatrAIx2026

Provides 999,847 persona records—599,847 grounded from real sources and 400,000 synthetic—each encoded as 1,290 categorical attributes packed into 645-byte Parquet blobs. Includes a codebook, postings index, and calibration/audit artifacts; decode with pyarrow and persona_codes.schema.json.

#parquet#pandas#polars#huggingface#nlp+2
Hugging Face
AI Dataset·2026
Icon for item

AgentWorldBench

Qwen

Provides 2,170 reference-grounded evaluation samples across seven agent domains (MCP, Search, Terminal, SWE, Android, Web, OS) to score language world models on Format, Factuality, Consistency, Realism and Quality. Includes per-domain JSONL files, judge prompts and an evaluation script for reproducible scoring.

#qwen#evaluation#huggingface#ai-agent#agent-skills+6
Hugging Face
AI Dataset·2026
Icon for item

bigfacing/GOKU-2M

Sen Liang, Cong Wang +9·University of Science and Technology of China, Tencent Hunyuan

Provides ~2 million instruction-aligned video-edit pairs for training and evaluating instruction-based video editing and generation models. Covers multi-task and structural edits (e.g., camera/subject movement), produced via a synthesis pipeline with progressive filtering; licensed CC BY-NC-4.0.

#ai-video#video#huggingface#multimodal#AIGC
Hugging Face
AI Dataset·2026
Icon for item

GOKU-2M

Goku-2M

Multimodal video dataset for text-to-video and video-to-video research: about 2 million short English videos and extracted frames for instruction-based video editing and generation. Hosted on Hugging Face and licensed CC BY‑NC 4.0 (non-commercial).

#huggingface#video#ai-video#multimodal#image+1
Hugging Face
AI Dataset·2026
Icon for item

SVG Generation Benchmark (Static)

Rapidata

Compares 30 frontier LLMs generating static SVG markup from 500 prompts using 1,355,161 human votes across three leaderboards (Preference, Coherence, Alignment); provides raw SVGs, 768×768 rasterized PNGs, and per-comparison human vote records under a CC-BY-4.0 prompt license.

#evaluation#ai-image#image#llm#huggingface+2
  • Previous
  • 1
  • More pages
  • 19
  • 20
  • 21
  • More pages
  • 29
  • Next