AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Discover the Best AI Resources

Curated essentials, no noise — just what matters

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.

Contents

AI Video Papers·2026
Icon for item

MemDreamer: Decoupling Perception and Reasoning for Long Video Understanding via Hierarchical Graph Memory and Agentic Retrieval Mechanism

Cong Chen, Guo Gan +8

Decouples perception and reasoning for hours-long videos by streaming inputs into a three-tier Hierarchical Graph Memory and using an agentic Observation–Reason–Action retrieval loop; reduces reasoning context to ~2% of full video while improving benchmark accuracy.

#paper#ai-video#multimodal#GNN#agent-skills+3
Hugging Face
AI Model·2026
Icon for item

unsloth/gemma-4-12B-it-qat-GGUF

unsloth, Google DeepMind

GGUF-format QAT (quantization-aware training) build of Gemma 4 12B that reduces memory needs for local or lightweight inference while preserving near bfloat16 quality. Ready for any-to-any conversational pipelines and ecosystem deployment.

#gemma#huggingface#google#deepmind#transformers+5
Natural Language Processing Papers·2026
Icon for item

Your UnEmbedding Matrix is Secretly a Feature Lens for Text Embeddings

Songhao Wu, Zhongxin Chen +4

Removes the subspace of frequent, uninformative tokens that LLMs inject into text embeddings via the model's unembedding matrix. EmbedFilter is a lightweight linear transform that refines LLM-derived embeddings to improve zero‑shot semantic retrieval, enable dimensionality reduction, and speed up indexing; code on GitHub.

#embeddings#LLM#NLP#paper#github+3
Speech Technology Papers·2026
Icon for item

MMAE: A Massive Multitask Audio Editing Benchmark

Ziyang Ma, Ruiqi Yan +36

Provides a comprehensive benchmark for instruction-based audio editing across seven audio modalities and eight operation types, with 2,000 high-fidelity samples and a rubric that decomposes tasks into 17,741 verifiable criteria for multi-dimensional evaluation.

#audio#multimodal#paper#speech#ai-leaderboard
Computer Vision Papers·2026
Icon for item

AnchorWorld: Embodied Egocentric World Simulation with View-based Evolution Customization

Yu Li, Menghan Xia +9

Simulates egocentric, embodied human–world interactions and enables customizable, self-evolving local scenes by defining anchor views and text-driven evolution. Uses exogenous viewpoints and full-body motion supervision to improve spatial grounding and interaction consistency.

#vision#robotics#paper#multimodal#ai
GitHub
AI Agent·2026
Icon for item

Cloudflare Computer

Cloudflare

Provides a virtual, durable filesystem and pluggable execution runtimes for agents running inside Cloudflare Durable Objects. Offers container, isolate-shell, and isolate-JS backends; preview-stage API with ~10GB Durable Object backing and FUSE-mounted container trade-offs.

#agent-skills#ai-agent#ai-tools#coding-agents#github+4
Large Language Model Papers·2026
Icon for item

On the Geometry of On-Policy Distillation

Zhennan Shen, Yanshu Li +7

Analyzes the parameter-space geometry of on-policy distillation (OPD) for LLM training, showing OPD updates affect fewer weights, avoid principal directions, and rapidly lock into a low-dimensional update subspace. Compares OPD with supervised fine-tuning (SFT) and reinforcement learning with verifiable rewards (RLVR) and studies implications for optimization and objective mixing.

#paper#LLM#RL#NLP#foundation-model+2
Hugging Face
AI Dataset·2026
Icon for item

LEDGER — Long-Context KPI Question Answering & Page Retrieval

artefactory

Provides page-level relevance judgments and full OCR'd annual-report text for KPI question answering and page retrieval benchmarking — supports retrieval (per-page qrels) and needle‑in‑a‑haystack numeric extraction over long documents, with eval and train configs.

#huggingface#finance#ocr#NLP#LLM+1
Hugging Face
AI Dataset·2026
Icon for item

LEDGER Long-Context Multi-KPI

artefactory

Pairs OCR-extracted annual-report text with ground-truth financial KPI values to benchmark LLM/table-QA and needle-in-a-haystack extraction tasks. Includes Markdown OCR (.mmd), page images for eval, and 31 KPI columns across multiple years—suited for KPI extraction, retrieval, and robustness testing.

#deepseek#ocr#huggingface#finance#pandas+3
Hugging Face
AI Model·2026
Icon for item

unsloth/gemma-4-26B-A4B-it-qat-GGUF

unsloth

A GGUF release of Gemma 4 26B A4B (QAT) packaged by Unsloth for local multimodal inference — quantization-aware trained to keep near-bfloat16 quality while significantly lowering memory requirements, compatible with Transformers and Unsloth tooling.

#gemma#huggingface#transformers#llm#vision+3
Hugging Face
AI Model·2026
Icon for item

North Mini Code (CohereLabs/North-Mini-Code-1.0)

CohereLabs

Code-focused sparse Mixture-of-Experts LLM designed for agentic coding and terminal/tool use, offering very long context (256K) and long outputs. Released with open weights under Apache-2.0 and optimized for transformers/vLLM workflows.

#transformers#vllm#opencode#ai-coding#ai-agent+2
Hugging Face
AI Model·2026
Icon for item

huihui-ai/Huihui-gemma-4-12B-it-abliterated

huihui-ai

Experimental, uncensored fine-tune of Google Gemma-4-12B-it that applies an 'abliteration' technique to remove refusal behaviors; intended for research and testing only and carries elevated safety and legal risks.

#gemma#transformers#huggingface#llm#ollama+1
  • Previous
  • 1
  • More pages
  • 133
  • 134
  • 135
  • More pages
  • 173
  • Next