AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Discover the Best AI Resources

Curated essentials, no noise — just what matters

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.

Contents

Hugging Face
AI Dataset·2026
Icon for item

HelioAI DeepReason 462×105M (Mythos V2 Distill)

HelioAI Labs

Provides 462 unrestricted long-form chain-of-thought reasoning traces distilled from the full Mythos V2 model (≈104.7M characters); intended for long-context evaluation, trace analysis and process-level supervision. License unknown—verify before reuse.

#huggingface#llm#evaluation#security#biology
Hugging Face
AI Model·2026
Icon for item

Cosmos3-Super

NVIDIA

Generates and reasons about multimodal physical-world content—text, images, video, audio, and robot/action trajectories—conditioned on combinations of text, image, video and action inputs. The 64B “Super” variant targets Physical AI use cases and supports vLLM‑Omni, Diffusers, and action prediction.

#nvidia#huggingface#multimodal#robotics#ai-video+5
Large Language Model Papers·2026
Icon for item

Trust Region On-Policy Distillation

Xingrun Xing, Haoqing Wang +3

Proposes TrOPD, a method that restricts token-level on-policy distillation to regions where teacher supervision is reliable to stabilize training under teacher–student distribution mismatch. Adds outlier handling (clipping, masking, forward-KL) and off-policy guidance; shows consistent gains on math reasoning, code generation and general benchmarks.

#LLM#NLP#paper#RL#foundation-model+2
Large Language Model Papers·2026
Icon for item

On the Scaling of PEFT: Towards Million Personal Models of Trillion Parameters

Mind Lab, : +53

Studies small trainable adapters (PEFT) used as persistent personal models on top of large foundation models, analyzing three scaling axes—Scale Up, Scale Down, Scale Out—and introducing MinT, an infrastructure for adapter identity, provenance, evaluation, and serving.

#foundation-model#llm#nlp#paper#ai-train
Hugging Face
AI Video·2026
Icon for item

ByteDance/Bernini-R

ByteDance

Provides the renderer weights and inference code for Bernini’s video renderer, enabling text→video, image→video and video editing inference. Offers a ready diffusers-format bundle or safetensors checkpoints under Apache‑2.0; intended for multi‑GPU/Hopper inference and reproducible research.

#bytedance#huggingface#diffusers#video#ai-video+3
AI Agent Papers·2026
Icon for item

Where Do Deep-Research Agents Go Wrong? Span-Level Error Localization in Agent Trajectories

Jiaming Wang, Ziteng Feng +9

Localizes harmful span-level errors inside long research-agent trajectories to show which trajectory segments make final answers unreliable. Provides a 1,000-instance TELBench of annotated spans and DRIFT, a claim-centric auditing method that improves span-level localization and first-error accuracy by up to 30 percentage points.

#agent-skills#ai-agent#LLM#NLP#paper
Reinforcement Learning Papers·2026
Icon for item

A Local Perturbation Theory for Cross-Domain Interference and Recovery in Multi-Domain RL

Lei Yang, Siyu Ding +1

Analyzes how single-domain RL fine-tuning on LLMs induces cross-domain interference and shows this damage concentrates in a low-dimensional shared conflict subspace; proposes a local perturbation theory and short domain "refresh" procedures that selectively recover earlier domains with minimal collateral loss.

#RL#LLM#paper#NLP#code+1
Hugging Face
AI Dataset·2026
Icon for item

Emerald Bay

Tahoe Therapeutics, tahoebio +1

Provides per-cell transcriptomes and five-day drug-sensitivity readouts for 1.83M single cells across 52 cancer cell lines and 91 drug conditions, with raw counts plus gene, cell-line, drug, and summary metadata for modeling drug response and context-dependent gene function.

#biology#genomics#drug-discovery#chemistry#huggingface+2
Hugging Face
AI Dataset·2026
Icon for item

xlangai/osworld_v2_tasks

xlangai

Provides the gated, official OSWorld 2.0 Python task class files (task_*.py) required to run the benchmark; distributed via a Hugging Face gated dataset to reduce benchmark leakage. Download requires accepting gated access on Hugging Face.

#huggingface#evaluation#agent-skills#ai-agent#json+2
Reinforcement Learning Papers·2026
Icon for item

Harness-1: Reinforcement Learning for Search Agents with State-Externalizing Harnesses

Pengcheng Jiang, Zhiyi Shi +6

A 20B retrieval subagent trained with reinforcement learning inside a stateful search harness that externalizes recoverable search state (candidate pool, curated evidence, verification records). The harness lets the policy focus on semantic search decisions, improving curated recall and transfer robustness.

#RL#ai-agent#agent-skills#vllm#huggingface+1
Computer Vision Papers·2026
Icon for item

Cosmos 3: Omnimodal World Models for Physical AI

Aditi, Niket Agarwal +9

Omnimodal world model that jointly processes and generates text, images, video, audio, and action trajectories for physical AI. Uses a mixture-of-transformers to combine autoregressive reasoning and diffusion-based multimodal generation; released open-source with checkpoints, datasets and benchmarks for robotics and simulation.

#foundation-model#multimodal#video#image#robotics+4
AI Agent Papers·2026
Icon for item

AutoMedBench: Towards Medical AutoResearch with Agentic AI Models

Junqi Liu, Salena Song +13

Workflow-aware benchmark for autonomous medical-AI research that splits agent execution into five stages (Plan, Setup, Validate, Inference, Submit) and evaluates long-horizon runs across segmentation, image enhancement, VQA, report generation, and lesion detection with stage-level scoring.

#vision#multimodal#ai-agent#agent-skills#ai-workflow+2
  • Previous
  • 1
  • More pages
  • 129
  • 130
  • 131
  • More pages
  • 172
  • Next