AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Discover the Best AI Resources

Curated essentials, no noise — just what matters

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.

Contents

Hugging Face
AI Model·2026
Icon for item

NuExtract3

numind (NuMind)

Unified 4B vision-language model for document understanding that converts images or text into template-driven structured JSON or clean Markdown. Key features: multimodal inputs (image+text), template-based extraction, reasoning vs non-reasoning modes, and vLLM/OpenAI-compatible deployment for OCR, invoice/forms extraction, and RAG preprocessing.

#transformers#vllm#ocr#multimodal#multilingual+5
GitHub
AI Client·2026
Icon for item

dbx

t8y2, zipg +8·t8y2

Lightweight cross-platform database client that exposes configured database connections to AI coding agents via an MCP server and includes a built-in AI SQL assistant. Ships as a single ~15MB binary, supports 60+ databases, and runs on desktop or Docker.

#mcp#mcp-server#mcp-client#ai-tools#ai-coding+8
GitHub
AI Agent·2026
Icon for item

Google Antigravity SDK

Google LLC, Antigravity Team +1

Lets developers build stateful, tool-enabled Python AI agents that run on Google's Antigravity runtime. Includes built-in tools (file I/O, shell, image generation), a declarative policy/hook system, multimodal input, and MCP integration.

#python#google#gemini#ai-agent#agent-skills+5
Hugging Face
AI Dataset·2026
Icon for item

VisCoR_Contrast (VisCoR-55K Contrastive Pairs)

5551z, Zhiyu Pan +7

Provides ~85K contrastive visual question–answer pairs where each example contains an anchor and a matched counterpart (image, question, answer). Pairs span General, Reasoning, Math, Graph/Chart and OCR categories to help train and evaluate fine‑grained, faithful visual reasoning in VLMs.

#vision#multimodal#image#ocr#huggingface+2
Hugging Face
AI Model·2026
Icon for item

Qwen3.6-27B-Heretic-Uncensored-FINETUNE-NEO-CODE-Di-IMatrix-MAX-GGUF

DavidAU

An uncensored, fine-tuned and GGUF-quantized variant of Qwen3.6-27B tailored for long-context, coding, vision and creative-writing use. Offers multiple NEO-CODE Di-Matrix quants (IQ2/IQ4/Q6/Q8), mmproj vision support and recommended inference settings for local servers.

#huggingface#transformers#llm#multimodal#vision+5
Hugging Face
AI Model·2026
Icon for item

Granite Speech 4.1 2B

IBM Granite Speech Team

Multilingual 2B speech–language model for ASR and bidirectional speech translation (EN, FR, DE, ES, PT, JA), providing punctuation/truecasing, keyword biasing, and a dual-head CTC encoder to boost transcription accuracy.

#ASR#speech#multilingual#transformers#huggingface+3
Hugging Face
AI Dataset·2026
Icon for item

Tachibana4-PREVIEW

sequelbox

Early-preview (≈1.2k rows) dataset of agentic coding prompts and unedited model responses generated by DeepSeek‑V4‑Pro, covering real-world programming tasks across many languages. Intended for research, filtering, and model evaluation rather than production training without review.

#deepseek#ai-coding#agent-skills#huggingface#code+7
Hugging Face
AI Model·2026
Icon for item

Pixal3D

Dong-Yang Li, Wang Zhao +7

Generates high-fidelity 3D assets from a single image by back-projecting pixel-aligned features into 3D, preserving fine geometry and PBR textures; includes inference code and a Hugging Face demo—best suited for single-view object reconstruction.

#huggingface#gitHub#paper#code#ai-image+3
GitHub
AI Agent·2026
Icon for item

deepsec

cramforce, Melkeydev +1·Vercel

Performs agent-driven security scans of codebases using LLM coding agents to find and triage vulnerabilities. Combines fast regex discovery, per-file AI investigation and revalidation, with optional sandboxed parallel execution and Vercel AI Gateway integration for large monorepos.

#security#ai-agent#agent-skills#anthropic#codex+5
Hugging Face
AI Dataset·2026
Icon for item

Thinking with Visual Primitives

NodeLinker, DeepSeek AI

Provides the dataset and accompanying technical report for a DeepSeek project that interleaves spatial markers (points and boxes) into multimodal LLM reasoning. Includes a public subset of data and benchmarks under an MIT license; model weights are not included.

#deepseek#vision#multimodal#foundation-model#llm+3
Hugging Face
AI Model·2026
Icon for item

gemma-4-31B-it-DFlash

z-lab

Draft model for speculative decoding that uses a lightweight block-diffusion drafter to propose multiple tokens in parallel; designed to pair with google/gemma-4-31B-it and accelerate autoregressive text generation (official benchmarks report up to ~5.8× throughput).

#gemma#huggingface#vllm#transformers#llm+3
GitHub
AI Agent·2026
Icon for item

Instatic

CoreBunch

Self-hosted visual CMS that runs as a single Bun server, combining a canvas editor, content engine, media, auth, forms, plugins, and a publisher. Emits semantic HTML and compact CSS and includes a provider-agnostic AI agent that edits pages as real, editable nodes. Best for teams that want full control and simple deployments.

#typescript#plugin#ai-agent#ai-api#postgres+3
  • Previous
  • 1
  • More pages
  • 119
  • 120
  • 121
  • More pages
  • 194
  • Next