AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Discover the Best AI Resources

Curated essentials, no noise — just what matters

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.

Contents

Hugging Face
AI Model·2026
Icon for item

dots3-note Preview

Dots Studio, Xiaohongshu

Multimodal Mixture-of-Experts text-generation model that accepts text, images, video and audio and returns text; preview open-weight release with 280B total params, 16B activated params, up to 512K token context and BF16/FP8 checkpoints under Apache-2.0.

#multimodal#llm#transformers#vllm#safetensors+7
Large Language Model Papers·2026
Icon for item

Motif 3: Technical Report

Junghwan Lim, Joon Son Chung +25

Describes a 314B-parameter decoder-only Mixture-of-Experts language model that activates 13.2B parameters per token for fine-grained sparsity, long-context (up to 256K) and multi-domain capabilities. Emphasizes GDLA architecture, expert balancing, and multi-teacher distillation.

#foundation-model#LLM#long-horizon#distillation#reasoning+4
Hugging Face
AI Model·2026
Icon for item

Muse Glimmer-30B-GGUF

Meta Superintelligence Lab, unsloth

Runs a quantized, locally executable 29.6B multimodal causal language model optimized for agentic workflows. Includes a perception encoder for image+text input, 4-bit quantized weights for 24–32GB devices, a DFlash drafter for speculative decoding, and robust tool-call support.

#transformers#huggingface#multimodal#llm#ai-agent+6
AI Dataset·2026
Icon for item

SWE-Bench ProMax: Benchmarking Agents on Large-Scale Multilingual Code Refactoring

Yuling Shi, Jinghan Xu +13

Provides a curated benchmark of 170 real-world, multilingual code-refactoring instances to evaluate AI coding agents on large-scale, behavior-preserving, cross-file refactors. Each task includes rewritten issue descriptions and manually reviewed test suites to avoid over- and under-constraining evaluations.

#benchmark#benchmarks#ai-coding#coding-agents#multilingual+6
Large Language Model Papers·2026
Icon for item

Macaron-V1: Towards Open Continual Learning with Self-Improvement and Mixture-of-LoRA

Vin Bo, Asher Cai +73·Mind Lab

A research report proposing a continual-learning agent workflow that pairs recursive self-improvement with a Mixture-of-LoRA design: freeze a foundation model, compose specialist LoRA adapters routed per user turn, and support them with long-context RL and post-training infrastructure.

#llm#foundation-model#lora#rl#ai-agent+4
Hugging Face
AI Dataset·2026
Icon for item

MiniMax H3 - 1K

ostris

Provides 1,000 five-second video clips generated by MiniMax H3 for lightweight evaluation of multimodal generation and understanding. Clips are roughly 768p base resolution with diverse aspect ratios and themes, produced with a pruned int8 minimax_h3_fl2va checkpoint at 30 steps.

#video#ai-video#multimodal#huggingface#youtube+2
Hugging Face
AI Video·2026
Icon for item

MiniMax H3 Realism People LoRA

Lovis Odin·fal, MiniMaxAI

A LoRA adapter for MiniMax H3 that improves photorealistic rendering of people—preserving skin texture, coherent micro-expressions, film-style lighting and subtle handheld motion. Trigger word: r34l1sm; intended for text-to-video portrait and close-up shots.

#ai-video#video#multimodal#huggingface#ai-train
Hugging Face
Chatbot·2026
Icon for item

Qwen Sharp Chat Templates

Saga Ishtardottir, froggeric

Provides a drop-in Jinja chat template for Qwen 3.5/3.6/3.8 that reduces reasoning-token waste, enforces a concise terseness system prompt, and preserves in-chat reasoning and tool-call rendering across turns. Terseness is on by default but switchable per request; no model weights are changed.

#qwen#llama.cpp#vllm#huggingface#gguf+5
Hugging Face
AI Model·2026
Icon for item

Ling-3.0-tiny

InclusionAI

Lightweight sparse-MoE LLM (7.9B params, ~1.3B activated per token) designed for hybrid multi-step reasoning and agentic tasks. Uses a KDA–MLA hybrid attention stack and a 128-expert sparse FFN; offered in BF16/FP8/INT4 for local and edge deployment.

#llm#huggingface#vllm#ollama#ai-deploy+6
Hugging Face
AI Dataset·2026
Icon for item

GitSkills

Giuseppe Destefanis, Daniel Graziotin +2·University College London, University of Hohenheim +2

Provides a queryable dataset of 3,797,117 SKILL.md agent-skill files found on public GitHub, deduplicated by content hash and enriched with representative text, front matter, folder composition, repo metadata, and sampled commit history for research.

#agent-skills#github#parquet#sqlite#polars+2
Hugging Face
AI Model·2026
Icon for item

NVIDIA-Nemotron-3.5-Lightning-30B-A3B-BF16

NVIDIA Corporation

A customizable 30B-parameter Mixture-of-Experts LLM (3B active) in BF16 for low-latency, high-throughput agent workflows; supports speculative decoding (MTP/DSpark/DFlash) and up to 1M-token contexts. Released with open weights and recipes under OpenMDW-1.1, intended for post-training, domain adaptation, and research on NVIDIA GPU stacks.

#nvidia#llm#transformers#pytorch#vllm+3
Hugging Face
AI Model·2026
Icon for item

LFM2.5-VL-3B

Liquid AI

Multimodal vision-language model optimized for on-device image+text tasks: image captioning, full-page OCR with layout annotation, grounding/bounding-box prediction, and function calling. Built on the LFM2.5-2.6B backbone with a SigLIP2 NaFlex 400M vision encoder and tuned for low-latency, low-memory edge inference.

#multimodal#vision#ocr#transformers#safetensors+8
  • Previous
  • 1
  • More pages
  • 175
  • 176
  • 177
  • More pages
  • 203
  • Next