AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Discover the Best AI Resources

Curated essentials, no noise — just what matters

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.

Contents

Hugging Face
AI Dataset·2026
Icon for item

Prompt Routing Dataset

SupraLabs

Provides labeled prompts with full-reference answers (including chain-of-thought and code blocks) and per-example metadata to train edge routing/orchestrator models that decide whether to handle inputs locally or route them to larger models. Includes complexity scores, coding/math flags, routing justifications, and an automated override rule; suited for fine-tuning small models (50M–1.5B) for edge deployment.

#prompt-engineering#llm#nlp#pandas#polars+4
AI Video Papers·2026
Icon for item

ResearchStudio-Reel: Automate the Last Mile of Research from Paper to Poster, Video, and Blog

Lingao Xiao, Yalun Dai +18

Converts an academic paper into reusable extracted assets and then produces editable poster, synchronized talk video, and bilingual blog via modular generator skills. Key differentiator: a single Paper2Assets extractor shared by three editable generators plus an interactive Paper2Reel viewer that links slides, video, captions and blog while preserving factual consistency and round-tripable PPT/DOCX output.

#agent-skills#claude-code#codex#paper#video+4
Hugging Face
AI Dataset·2026
Icon for item

VideoChat3-Academic2M

MCG-NJU

Provides re-annotated academic video instruction data for captioning, video QA, and fine-grained motion understanding; rewrites short answers and concise captions into evidence-grounded, instruction-following responses and supplies JSONL annotation files (original videos not included).

#video#ai-video#multimodal#vision#huggingface+1
AI Agent Papers·2026
Icon for item

UI-MOPD: Multi-Platform On-Policy Distillation for Continual GUI Agent Learning

Niu Lian, Alan Chen +9

Trains cross-platform GUI agents by combining a Uni-GUI cross-platform dataset with platform-conditioned multi-teacher on-policy distillation, enabling a shared policy to adapt to new platforms while retaining platform-specific behaviors; suitable for research on continual GUI agent learning and cross-platform adaptation.

#RL#multimodal#agent-skills#ai-agent#paper+2
Large Language Model Papers·2026
Icon for item

ResearchStudio-Idea: An Evidence-Grounded Research-Ideation Skill Suite from ML Conference Outcomes

Qihao Zhao, Yangyu Huang +9

Provides a reusable skill suite for evidence-grounded research ideation: Paper-Search for multi-source literature retrieval, Scoop-Check for prior-art collision checking, and IdeaSpark for pattern-guided idea generation, evidence auditing, and idea-card rendering.

#LLM#agent-skills#paper#evaluation#skillkit+2
Reinforcement Learning Papers·2026
Icon for item

Trust Region Policy Distillation

Zhengpeng Xie, Li Lyna Zhang +2

Stabilizes on-policy policy distillation by dynamically constructing a proximal teacher that controls gradient variance. Provides theoretical global convergence and monotonic improvement bounds, and shows improved training stability, sample efficiency, and final performance on mathematical reasoning tasks with zero extra compute overhead.

#RL#paper#algorithms#math
Hugging Face
AI Model·2026
Icon for item

ThinkingCap: Qwen 3.6 27B

Karol Lasocki, Adam Osusky +9·BottleCap AI, Qwen Team

Fine-tuned variant of Qwen3.6-27B that cuts internal reasoning (‘thinking’) token usage by roughly 46% on average while preserving benchmark accuracy and safety behavior. Targets lower latency and inference cost; ships on Hugging Face with GGUF quantizations for local use.

#qwen#transformers#huggingface#llm#multimodal+5
Hugging Face
AI Model·2026
Icon for item

NVIDIA-Nemotron-Labs-3-Puzzle-75B-A9B-NVFP4

NVIDIA, NVIDIA NeMo

Deployment-optimized hybrid MoE LLM (75B total / 9.3B active) produced via Iterative Puzzle compression and Multi-Token Prediction to double server throughput and raise single-GPU concurrency; designed for multilingual reasoning, long-context generation, and high-volume agentic/chat deployments.

#nvidia#huggingface#transformers#pytorch#llm+5
Reinforcement Learning Papers·2026
Icon for item

Weak-to-Strong Generalization via Direct On-Policy Distillation

Shiyuan Feng, Huan-ang Gao +8

Transfers RL-induced policy shifts from a smaller 'weak' teacher to a stronger target by using the teacher's post-/pre-RL log-ratio as a dense implicit reward applied on the student's on-policy states. Enables reuse of RL supervision without running RL rollouts on the target, improving sample/time efficiency.

#RL#LLM#reasoning#qwen#paper+2
AI Video Papers·2026
Icon for item

Light-Omni: Reflex over Reasoning in Agentic Video Understanding with Long-Term Memory

Chang Nie, Jiaju Wei +3

Provides a reflexive agentic framework for long-horizon video understanding that replaces costly iterative reasoning with dual contextual states: a consolidated global multimodal script and parametric latent states for fast retrieval and response, improving speed and memory efficiency.

#video#multimodal#ai-agent#qwen#embeddings+4
Computer Vision Papers·2026
Icon for item

PixWorld: Unifying 3D Scene Generation and Reconstruction in Pixel Space

Sensen Gao, Zhaoqing Wang +4

Trains a single diffusion model that unifies 3D scene reconstruction and generative modeling by operating directly in pixel/rendered-image space. Supervises diffusion on rendered views and adds a geometry-perception loss from a pretrained 3D foundation model, reducing latent information loss and improving 3D fidelity.

#paper#vision#ai-image#image#depth+1
Hugging Face
AI Model·2026
Icon for item

DeepSeek-V4-Flash-GGUF

unsloth·Unsloth (unsloth.ai), DeepSeek-AI

GGUF-format quantized release of DeepSeek‑V4‑Flash for local inference — compatible with llama.cpp and Unsloth runtimes, with guidance for FP4/FP8 mixed precision and Q4/Q8 quantization; tuned for million-token long-context usage.

#deepseek#huggingface#llm#ai-inference#ai-serving+1
  • Previous
  • 1
  • More pages
  • 153
  • 154
  • 155
  • More pages
  • 190
  • Next