AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Category

Explore by categories

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.
  • All Categories

  • AI Leaderboard

  • AI Agent Tutorials

  • AI Coding Tutorials

  • AI Model

  • AI Agent Papers

  • Chatbot

  • AI Dataset

  • Machine Learning Foundation Books

  • AI Train

  • AI Deploy

  • AI Client

  • Machine Learning Foundation Papers

  • Machine Learning Foundation Tutorials

  • AI Image Demos

  • AI Agent

  • Large Language Model Tutorials

  • Large Language Model Papers

  • Machine Learning Engineering Papers

  • Computer Vision Tutorials

  • Computer Vision Papers

  • Natural Language Processing Papers

  • Reinforcement Learning Papers

  • Speech Technology Papers

  • AI API

  • AI Coding

  • AI Image

  • AI Video

  • MLOps

  • MCP Client

  • MCP Server

  • AI Video Papers

  • AI Audio

  • AI Others

  • AI Infra

  • Embodied AI

Hugging Face
AI Dataset·2026
Icon for item

Neko_Audio-80K_Short

liumindmind

Around 80K short audio clips paired with transcripts in JSON, organized for easy loading with the Hugging Face datasets ecosystem—designed for short-form speech tasks (ASR, TTS, fine-tuning) and quick prototyping with common Python data tools.

#audio#speech#ASR#tts#huggingface+3
Hugging Face
AI Dataset·2026
Icon for item

Nemotron-Personas-Vietnam

NVIDIA Corporation, FPT Smart Cloud +1

Provides 600,000 synthetic Vietnamese persona texts (100,000 records, 6 personas per record) aligned to Vietnam's 2024 census and surveys for training and evaluating NLP / text-generation models; includes 21 demographic and persona fields, CC BY 4.0, single train split.

#huggingface#nvidia#nlp#multilingual#llm+1
Hugging Face
AI Dataset·2026
Icon for item

EVA-Bench: A New End-to-end Framework for Evaluating Voice Agents

ServiceNow-AI

End-to-end evaluation framework for conversational voice agents that runs bot-to-bot audio simulations and scores agents on task accuracy (EVA-A) and interaction experience (EVA-X). Includes per-scenario backend state, accent/noise perturbations, and 213 scenarios across airline, healthcare HR, and enterprise IT domains.

#huggingface#voice#speech#ASR#tts+4
Hugging Face
AI Dataset·2026
Icon for item

AI Village (HuggingFace dataset)

AI Digest (aidigestorg), Hugging Face

Provides a complete, lightly-processed export of AI Village's >1-year multi-agent data: per-agent computer sessions (with screenshots), turn-by-turn computer-use logs, group chats, agent memories, goals, and daily summaries for research into agentic behaviour, multi-agent dynamics, long-horizon memory, and AI safety. Access is manually reviewed.

#ai-agent#agent-skills#LLM#llm#huggingface+3
AI Agent Papers·2026
Icon for item

SWE-Explore: Benchmarking How Coding Agents Explore Repositories

Shaoqiu Zhang, Yuhang Wang +9

Measures how coding agents explore repositories by asking them to return a ranked, line-level list of code regions relevant to an issue under a fixed line budget. Covers 848 issues across 203 repos and 10 languages; evaluates coverage, ranking, and context-efficiency to isolate exploration quality.

#ai-agent#ai-coding#agent-skills#paper#code+1
Speech Technology Papers·2026
Icon for item

MMAE: A Massive Multitask Audio Editing Benchmark

Ziyang Ma, Ruiqi Yan +36

Provides a comprehensive benchmark for instruction-based audio editing across seven audio modalities and eight operation types, with 2,000 high-fidelity samples and a rubric that decomposes tasks into 17,741 verifiable criteria for multi-dimensional evaluation.

#audio#multimodal#paper#speech#ai-leaderboard
Hugging Face
AI Dataset·2026
Icon for item

LEDGER — Long-Context KPI Question Answering & Page Retrieval

artefactory

Provides page-level relevance judgments and full OCR'd annual-report text for KPI question answering and page retrieval benchmarking — supports retrieval (per-page qrels) and needle‑in‑a‑haystack numeric extraction over long documents, with eval and train configs.

#huggingface#finance#ocr#NLP#LLM+1
Hugging Face
AI Dataset·2026
Icon for item

LEDGER Long-Context Multi-KPI

artefactory

Pairs OCR-extracted annual-report text with ground-truth financial KPI values to benchmark LLM/table-QA and needle-in-a-haystack extraction tasks. Includes Markdown OCR (.mmd), page images for eval, and 31 KPI columns across multiple years—suited for KPI extraction, retrieval, and robustness testing.

#deepseek#ocr#huggingface#finance#pandas+3
Hugging Face
AI Dataset·2026
Icon for item

CustoMDiT / PexelsCustom-1M

carpedkm

Provides 1,036,431 identity–text–video triplets with per-video JSON annotations and reference keyframes to train and evaluate identity-preserving customized video generation models. Data is drawn from ~320K Pexels HD videos; videos must be downloaded separately per Pexels' terms.

#ai-video#video#huggingface#pandas#diffusers+1
Hugging Face
AI Dataset·2026
Icon for item

HIW-500: Humanoids In-the-Wild Dataset

BitRobot, Unitree +1

Provides 500+ hours of human whole-body teleoperation demonstrations for humanoid robot learning in real homes, with synchronized video, joint states, action traces and language annotations. Includes 23K+ episodes, fine-grained subtask labels, and raw ROS/MCAP plus compressed LeRobot formats.

#robotics#multimodal#vision#huggingface#ai-train+1
Hugging Face
AI Dataset·2026
Icon for item

HIW-500: Humanoids In-the-Wild Dataset (LeRobot)

BitRobot, Unitree +1

Provides 500+ hours of human whole-body teleoperation recordings of a Unitree G1 in real homes, packaged in LeRobot v3.0 for robot learning. Contains 23K+ episodes, ~40M frames, multi-view 480p@30 video, 29-DoF states, actions and language annotations; CC BY 4.0 and large download size.

#robotics#video#huggingface#polars#pytorch
Hugging Face
AI Dataset·2026
Icon for item

Russian PII NER Benchmark

redmadrobot-rnd

Provides a token-level benchmark for Russian PII detection and NER, with 2,841 sentences and 5,614 annotated spans across 21 fine-grained entity types in BIO format. Mixes sanitized production-log examples, synthetic document templates, and hard negatives to evaluate guardrails and anonymization pipelines.

#huggingface#pandas#nlp#python#security
  • Previous
  • 1
  • More pages
  • 17
  • 18
  • 19
  • More pages
  • 33
  • Next