AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Discover the Best AI Resources

Curated essentials, no noise — just what matters

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.

Contents

Hugging Face
AI Dataset·2026
Icon for item

Q-CARE Benchmark

Jeonghwan Choi, Taewon Yun +4·Korea Advanced Institute of Science and Technology (KAIST), Cluvion

Evaluates retrieval-augmented generation by decomposing user queries into sub-queries and answers into atomic claims, scoring retrieval by query coverage and generation by claim verifiability. Reference-free benchmark with 800 queries, inlined retrieved chunks, and answers from multiple RAG systems; runs locally without API keys.

#RAG#retrieval#benchmark#evaluation#huggingface+5
Hugging Face
AI Dataset·2026
Icon for item

Claude Fable 5 Cursor Traces

TeichAI, Cursor +1

Contains 244 Cursor agent sessions recorded from Claude Fable‑5, formatted for training and research. Sessions include multi-turn assistant/tool interactions and are Teich-compatible; several rows exceed one million characters, so apply explicit tokenization and oversize policies before training.

#cursor#distillation#huggingface#anthropic#claude+3
AI Agent Papers·2026
Icon for item

Agentic Game Development as a Verifiable Trajectory Data Engine for Scaling World Models

Pengfei Zhou, Hexin Wang +6·InfRec, Cardinal AI Lab +3

Proposes treating game development as a recursive data engine and introduces RLHEV (Reinforcement Learning with Human-Engine Verification) to combine dense engine checks (collision, physics, navigability) with human acceptance feedback, producing trajectory data and rewards for post-training world models.

#RL#agent-skills#coding-agents#long-horizon#multimodal+3
Hugging Face
AI Model·2026
Icon for item

unsloth/Qwen3.8-Flash-Next-GGUF

unsloth

GGUF-quantized build of Qwen3.8-Flash-Next for image-text-to-text inference and local deployment. Ships with Unsloth Dynamic 3.0 quantization, thinking-mode controls (preserve_thinking, reasoning_effort), and native long-context support (262k, extensible to 1M with YaRN).

#gguf#qwen#multimodal#vision#llm+7
Speech Technology Papers·2026
Icon for item

VoiceMem: Streaming Dual-Brain Memory for Real-Time Interaction

Zhifei Xie, Jiaqi Lang +8·Nanyang Technological University, National University of Singapore +3

Provides a streaming dual-brain memory for real-time speech agents: an informational left brain for factual retrieval and an affective right brain for persona/emotion, achieving high top-5 accuracy while keeping retrieval latency within VAD budgets (~134 ms).

#voice#speech#audio#ASR#multimodal+4
Hugging Face
AI Dataset·2026
Icon for item

Anthropic Insights Pilot: Partner Cluster Data

Kunal Handa, Miranda Zhang +22·Anthropic, Social and Language Technologies (SALT) Lab, Stanford University +2

Provides aggregated, privacy-preserving cluster outputs from three external research teams' analyses of ~250k Claude/Claude Code conversations; includes per-team CSVs (Stanford, Oxford, METR) for studying human–AI collaboration and model behavior without raw conversations.

#anthropic#claude#LLM#research#pandas+3
AI Agent Papers·2026
Icon for item

JIT-Agent: Scaling Harness Intelligence via Just-in-Time Harness Evolution

Guibin Zhang, Leo Lu +14

Synthesizes, repairs, and self-evolves task-adaptive agent harnesses on demand for off-the-shelf LLM agents, using a trainable harness-intelligence model that distills signals from past configurations. Demonstrates consistent performance gains across benchmarks and model families by producing four-module, composable harnesses.

#ai-agent#agent-skills#llm#deepseek#qwen+5
Hugging Face
AI Model·2026
Icon for item

GLM-5.3-Flash (GGUF)

unsloth, zai-org

Provides a GGUF-quantized build of GLM-5.3-Flash for local text-generation and inference. Key features: 320B total / 18B active parameters, hybrid sparse+linear attention, native multimodal pretraining and Unsloth Dynamic quantization. Best for developers running GGUF local inference workflows.

#gguf#foundation-model#multimodal#llm#transformers+3
Hugging Face
AI Model·2026
Icon for item

Qwen-Drive-1.0-4B

Xin Zhou, Zongchuang Zhao +14

Integrates a pretrained vision–language model with a BEV perception head and a Planning Expert to provide 3D perception, driving VQA and motion planning for autonomous driving while keeping the base VLM architecture unchanged.

#qwen#multimodal#vision#transformers#safetensors+7
AI Agent Papers·2026
Icon for item

What Makes Good Agentic Data? An ACE Lens on Data Generation for LLM Agents

Xingshan Zeng, Zishan Xu +12

Analyzes how to generate useful interaction data for LLM agents and proposes the ACE lens — Accuracy, Complexity, divErsity — while factorizing agentic data as (E, q, τ, v). Surveys verification, difficulty calibration, and coverage strategies and outlines implications for training and benchmarks.

#LLM#agent-skills#evaluation#benchmarks#ai-agent+2
Large Language Model Papers·2026
Icon for item

J-Zero: Unified Challenger--Solver--Judge Co-Evolution from Zero Data

Gyouk Chu, Myeongho Jeon +1·KAIST

A zero-data self-evolution framework that co-trains a Challenger, Solver, and Judge so LLMs can iteratively improve on both verifiable and unverifiable tasks without human labels. Uses role-asymmetry and subtask-amplification preference pairs to train the Judge and sustain improvement.

#paper#LLM#reasoning#evaluation#research+1
Computer Vision Papers·2026
Icon for item

Revisiting Local Context for Long-Horizon Streaming 3D Reconstruction

Jiarong Han, Jincheng Xiong +7·Alibaba Group

Performs causal, bounded‑memory streaming 3D reconstruction by caching KV features from only the preceding 11 frames, predicting a per‑frame point map and adjacent relative pose, and composing these local predictions into a global trajectory; includes a lightweight rotation refiner and composition‑aware loss to limit drift.

#paper#vision#long-horizon#depth#benchmark+3
  • Previous
  • 1
  • More pages
  • 186
  • 187
  • 188
  • More pages
  • 208
  • Next