AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Discover the Best AI Resources

Curated essentials, no noise — just what matters

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.

Contents

Hugging Face
AI Dataset·2026
Icon for item

tran-vi-teacher

ngocdang83

Parallel Chinese→Vietnamese dataset of webnovel (xianxia) text provided in JSON for NMT training and teacher-student distillation. In-domain, ~100K–1M examples with CC-BY-4.0 license — useful for fine-tuning or distillation experiments but limited by narrow genre and small download footprint.

#translation#huggingface#nlp#multilingual#pandas+1
GitHub
AI Agent·2026
Icon for item

Defending Code Reference Harness

Anthropic

Runs a multi-stage, Claude-powered pipeline to find, verify, triage, and generate patches for code vulnerabilities, plus interactive skills for threat modeling and customization. Default harness targets C/C++ memory bugs using ASAN inside Docker/gVisor; autonomous runs execute target code and require sandboxing.

#anthropic#claude#claude-code#security#agent-skills+5
Hugging Face
AI Model·2026
Icon for item

google/gemma-4-12B-it

Google DeepMind

Instruction-tuned, unified Gemma 4 12B multimodal model that accepts text, image and audio inputs and generates text outputs locally. Encoder-free design reduces multimodal latency and fits on consumer devices while offering long-context support and native thinking/system-prompt features.

#gemma#google#deepmind#multimodal#transformers+5
Hugging Face
AI Dataset·2026
Icon for item

Qwen3.7 Max Pi Traces

armand0e, TeichAI

Provides raw newline-delimited JSON agent traces where assistant responses were generated by qwen/qwen3.7-max, captured with Teich; includes 47 JSONL files, an embedded tools schema snapshot, and conversion guidance for supervised fine‑tuning and distillation.

#huggingface#llm#ai-agent#agent-skills#ai-train+2
Hugging Face
AI Model·2026
Icon for item

Gemma 4 12B Unified

Google DeepMind

A 12B unified, encoder-free multimodal model that directly ingests text, images and audio and returns text; supports very long contexts (up to 256K tokens), native function-calling/thinking modes, and small-model deployment for local or on-device use.

#gemma#multimodal#transformers#google#deepmind+8
Hugging Face
AI Model·2026
Icon for item

Step 3.7 Flash

stepfun-ai

Processes images and text to produce structured, reasoning-rich text outputs for high-throughput agentic workflows. Sparse MoE design (198B total, ~11B active per token), 256k context window and selectable reasoning levels—optimized for single-pass parsing, verification, and multi-step automation.

#multimodal#llm#transformers#vllm#ai-inference+4
Hugging Face
AI Dataset·2026
Icon for item

Suno AI Music Dataset — Multi-Genre Curated

Kukito·Hugging Face, Suno

A human‑curated corpus of AI‑generated music with MP3s, cover art, exact generation prompts and a 32‑column metadata schema; uses a 70/30 quality vs. mainstream split and a three‑level taxonomy to support fine‑grained audio‑ML, prompt‑fidelity and recommendation research.

#audio#huggingface#AIGC#pandas#polars
Hugging Face
AI Dataset·2026
Icon for item

StreamAudio-2M

zhifeixie

Large streaming-audio dataset for training and evaluating audio-LLMs and audio agents. About 2.28M clips grouped into multi-turn “streams” across six task subsets (ASR, speech translation, audio understanding, voice chat, proactive response, environment-aware); audio shipped as tar shards.

#audio#ASR#translation#speech#voice+2
Hugging Face
AI Audio·2026
Icon for item

MOSS-TTS-v1.5

OpenMOSS-Team

Generates multilingual text-to-speech with zero-shot voice cloning, token-level duration control, and inline pause markers. v1.5 improves multilingual fidelity (with language tags), cloning stability, and long-reference handling—suitable for research and production TTS pipelines.

#speech#audio#voice#multilingual#huggingface+2
Hugging Face
AI Model·2026
Icon for item

Keye-VL-2.0-30B-A3B

Kwai-Keye

Performs hour-scale video understanding and fine-grained temporal localization while exposing agent-style multimodal tool/code/search abilities. Built on a sparse-attention long-context architecture (DSA) and a specialized inference stack—best used in GPU-backed research or production evaluation.

#multimodal#video#deepseek#transformers#huggingface+5
GitHub
AI Agent·2026
Icon for item

ADHD — a skill for coding agents

Udit Akhouri·divergent.sh

Spawns parallel, isolated LLM reasoning frames, then scores, clusters and prunes ideas to avoid premature convergence. Packaged as a reusable Claude/Codex agent skill with CLI and TypeScript APIs for ideation, design decisions and fuzzy debugging.

#agent-skills#coding-agents#claude-code#codex#typescript+6
Hugging Face
AI Model·2026
Icon for item

Mellum2 Thinking

JetBrains

Generates text with explicit chain-of-thought traces for multi-step reasoning and math-heavy tasks, emitting reasoning inside <think>...</think> blocks. Uses a Mixture-of-Experts design and 131k token context for long, verifiable workflows—best when you need inspectable reasoning.

#huggingface#transformers#llm#vllm#foundation-model+1
  • Previous
  • 1
  • More pages
  • 130
  • 131
  • 132
  • More pages
  • 205
  • Next