AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Discover the Best AI Resources

Curated essentials, no noise — just what matters

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.
Hugging Face
AI Dataset·2026
Icon for item

GGT-100K: Generative Ground Truth for Generalizable Real-World Image Restoration

VCLab-PolyU

Provides 100,000 generated low-quality↔high-quality image pairs created with modern multi-frame/multi-modal models to boost generalization of image restoration methods; includes train/test JSONL lists, baseline training code, and pretrained checkpoints under CC BY‑NC‑ND 4.0.

#vision#image#ai-image#huggingface#paper+3
Hugging Face
AI Model·2026
Icon for item

Fara1.5-27B

Microsoft Research AI Frontiers, Microsoft

Automates end-to-end web workflows from browser screenshots by emitting pixel-grounded actions (click, type, scroll, visit, search). Vision-first multimodal agent fine-tuned from Qwen3.5-27B with critical-point safety checks; intended for sandboxed, human-supervised deployments.

#qwen#multimodal#vision#agent-skills#microsoft+6
Hugging Face
AI Model·2026
Icon for item

Miso TTS 8B

MisoLabs

Generates conversational speech and voice continuation from text and optional audio context, outputting Mimi audio codes. Built on a Sesame-style CSM with an 8B Llama-like backbone plus a smaller autoregressive audio decoder. Suited for local TTS inference and voice-cloning workflows.

#pytorch#audio#voice#speech#huggingface+1
Hugging Face
AI Model·2026
Icon for item

NVIDIA Cosmos3-Super-Image2Video

NVIDIA

Generates temporally coherent MP4 videos from a single input image plus text instructions, with configurable resolution, frame count, and optional AAC audio. Optimized for NVIDIA GPU stacks and integrates with vLLM‑Omni and Hugging Face Diffusers for production inference and research workflows.

#nvidia#huggingface#diffusers#ai-video#video+5
Hugging Face
AI Video·2026
Icon for item

LongCat-Video-Avatar-1.5

Meituan LongCat Team

Generates audio-driven avatar videos from text, images, or audio inputs with production-grade stability (accurate lip sync, identity consistency) and an 8-step distillation inference mode for faster serving; suitable for broadcasting, virtual hosts, animation, and multi-person scenarios.

#ai-video#video#audio#transformers#huggingface+4
Hugging Face
AI Model·2026
Icon for item

MiniCPM5-1B

openbmb

A 1.08B-parameter causal LLM engineered for on-device text generation with native long-context (131k tokens) and built-in Think/No-Think modes. It emphasizes tool-calling support, lightweight deployment formats (BF16, GGUF, MLX), and RL+OPD post-training for stronger reasoning and code generation.

#llm#transformers#huggingface#vllm#ollama+3
GitHub
MCP Server·2026
Icon for item

ai-memory

akitaonrails

Provides long-term memory for AI coding agents by compiling sanitized lifecycle observations into a git-versioned Markdown wiki that enables cross-agent handoffs, per-project isolation, and optional vector-backed retrieval.

#mcp#mcp-server#mcp-client#rust#cli+6
Hugging Face
AI Model·2026
Icon for item

Bonsai Image · Ternary 4B (gemlite 2-bit)

Prism ML (prism-ml)

A ternary-weight (~1.58-bit) 4B text-to-image diffusion transformer optimized for NVIDIA GPUs using Gemlite INT2 and HQQ; it reduces the transformer to ~1.21 GB (4.55 GB CUDA payload) and targets 1024×1024 generation with a 4-step FlowMatch-Euler sampler.

#huggingface#ai-image#image#nvidia#ai-inference+3
Hugging Face
AI Dataset·2026
Icon for item

Qwen-Image-Bench

Qwen

Creator-centric benchmark for evaluating text-to-image models with 1,000 bilingual prompts and a 3-level, 56-facet taxonomy. Includes a trained Q-Judger judge model and leaderboard-ready evaluation scripts to surface gaps in real-world fidelity and creative generation.

#huggingface#ai-image#image#vision#multilingual+3
GitHub
AI Client·2026
Icon for item

Kun

KunAgent

Local-first AI agent workspace that unifies coding, writing, design, research and automation under one runtime shared between a desktop GUI and a terminal TUI. Features Agent Graph for long-running, auditable workflows, multi-provider model support, and local-by-default data storage.

#ai-agent#agent-skills#ai-client#ai-tools#terminal+7
GitHub
AI Client·2026
Icon for item

Kimi Code CLI

Moonshot AI

Terminal-native AI coding agent that reads and edits code, runs shell commands, searches files, fetches web pages, and determines next steps from interactive feedback. Delivered as a single-binary TUI with video input, subagents, a plugin marketplace, and IDE (ACP) integration.

#cli#terminal#ai-agent#ai-coding#agent-skills+6
Hugging Face
AI Dataset·2026
Icon for item

WBench

meituan-longcat

Provides a 289-case (1,058-turn) multi-turn benchmark that evaluates interactive video world models across 22 metrics and five dimensions (quality, setting, interaction, consistency, physics). Includes first-/third-person and navigation splits plus a 20-model leaderboard for head-to-head comparisons.

#video#ai-video#vision#physics#huggingface+4
  • Previous
  • 1
  • More pages
  • 129
  • 130
  • 131
  • More pages
  • 205
  • Next

Contents