AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Discover the Best AI Resources

Curated essentials, no noise — just what matters

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.

Contents

Hugging Face
AI Dataset·2026
Icon for item

ArithMark 3.0

AxiomicLabs

A multiple-choice benchmark for evaluating language-model arithmetic: 1,000 continuation-style elementary word problems (4 choices, balanced labels) organized by topic, grade band, and difficulty. Designed for base-model continuation log-likelihood scoring; released under Apache-2.0.

#evaluation#benchmarks#benchmark#huggingface#nlp+4
Hugging Face
AI Model·2026
Icon for item

Mage-Flow

Zhang Xinjie, Zhang Peng +22·Microsoft

Efficient 4B native-resolution diffusion foundation model for text-to-image generation and instruction-based image editing. Uses a lightweight Mage‑VAE tokenizer and a 4B NR‑MMDiT backbone to produce 512–2048 outputs with low memory and fast inference; ships in Base, RL-aligned and few-step Turbo variants.

#multimodal#ai-image#image#vision#flow-matching+7
Hugging Face
AI Model·2026
Icon for item

unsloth/Laguna-S-2.1-GGUF

unsloth·Unsloth, Poolside

Provides GGUF-format quantized shards of Laguna S 2.1 for local or self-hosted inference—packaged for llama.cpp/llama-server and usable with vLLM/Transformers runtimes; targeted at long-context, agentic coding workloads.

#vllm#transformers#huggingface#llama.cpp#llm+5
Hugging Face
AI Model·2026
Icon for item

fdtn-ai/antares-1b

fdtn-ai

1B-parameter text-generation model tuned for conversational and agentic workflows with a focus on security and vulnerability-detection; suited for low-cost or on-prem/edge deployments and terminal-agent integrations.

#transformers#huggingface#llm#nlp#chatbot+7
Hugging Face
AI Dataset·2026
Icon for item

Asimov Agentic

Pierre Sermanet, Anirudha Majumdar +3·Google, ASIMOV Benchmark

A compact evaluation dataset and harness for testing agentic AI on safety-critical robotics tasks. Includes multimodal episodes in parquet format, task-specific eval scripts (gauge reading, human safety monitoring, VLA estimators), and TFDS/Hugging Face integration for reproducible safety evaluations.

#robotics#evaluation#benchmark#parquet#multimodal+5
Hugging Face
AI Image·2026
Icon for item

Mage-Flow-Edit-Turbo

Xinjie Zhang, Peng Zhang +22·Microsoft

Performs instruction-based image editing from reference images using a 4B native-resolution diffusion transformer; the Turbo variant uses 4-step distillation for interactive latency (≈1.02 s per 1024² edit on A100) while supporting semantic, appearance, structure-aware and restoration edits.

#ai-image#image#multimodal#huggingface#microsoft+5
Hugging Face
AI Model·2026
Icon for item

Nanbeige4.2-3B

Nanbeige, Kanzhun

Compact 3B-scale agentic LLM for multi-step tool use and reasoning, using a Looped Transformer to increase capacity without adding parameters; built for local deployment with configurable "thinking" modes and benchmark gains vs larger open models.

#llm#transformers#huggingface#vllm#llama.cpp+6
Hugging Face
AI Model·2026
Icon for item

Solar Open2 250B — Nota NVFP4

nota-ai, Upstage

4-bit NVFP4 (W4A4) quantized pack of Upstage Solar Open2 250B for vLLM serving on NVIDIA Blackwell GPUs, preserving MoE routing and near-BF16 quality while cutting model size from 500.6 GB to 153.3 GB.

#vllm#llm#huggingface#nvidia#ai-serving+4
Reinforcement Learning Papers·2026
Icon for item

Molt: A Scalable PyTorch-Native Training Framework for Agentic Reinforcement Learning

Jian Hu, Huiying Li +9

A PyTorch-native training framework for agentic reinforcement learning research that keeps researcher-facing code compact and editable. Uses an asynchronous loop to train multimodal and mixture-of-experts policies while never training on tokens the agent didn't generate; matches Megatron-style stacks under a comparable protocol and ships recipes and containers on GitHub.

#pytorch#RL#ai-agent#ai-train#nvidia+4
Hugging Face
AI Model·2026
Icon for item

Solar Open 2 (250B-A15B)

Upstage AI

A 250B-parameter mixture-of-experts LLM that activates 15B parameters per token to lower inference cost for agentic tasks—tool calling, long-context reasoning, and coding. Uses a hybrid softmax+linear-attention stack with 1M-token context and supports English, Korean, and Japanese; requires H200/B200-class GPUs to run efficiently.

#llm#transformers#vllm#multilingual#agent-skills+5
AI Video Papers·2026
Icon for item

Self Gradient Forcing: Native Long Video Extrapolation

Junhao Zhuang, Shiyi Zhang +12

Extrapolates long video sequences from very short contexts by restoring memory-writing supervision in autoregressive video diffusion models using a two-pass Self Gradient Forcing (SGF). SGF records a no-gradient rollout at a sampled denoising exit and then recomputes KV context in a second parallel pass so future losses teach earlier latent writes, enabling minutes-long extrapolation from ~5s windows.

#paper#video#ai-video#vision#diffusers+1
Hugging Face
AI Dataset·2026
Icon for item

Qwen3.8-Max Distillation 50K

r0b0tlab, Alibaba Cloud +1

A curated collection of 49,772 teacher-generated chat traces from qwen3.8-max-preview for supervised fine-tuning and off-policy distillation. Preserves visible chain-of-thought blocks, emphasizes math/code/reasoning mixes, and includes provenance and licensing cautions tied to Alibaba Cloud Model Studio.

#distillation#qwen#reasoning#math#code+6
  • Previous
  • 1
  • More pages
  • 162
  • 163
  • 164
  • More pages
  • 195
  • Next