AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Discover the Best AI Resources

Curated essentials, no noise — just what matters

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.

Contents

GitHub
AI Agent·2026
Icon for item

Hermes Agent Self-Evolution

Nous Research

Automatically evolves Hermes Agent skills, prompts, tool descriptions and code using DSPy + GEPA — mutating text via API calls, evaluating trace-based failures, and selecting variants that pass tests and human PR review. No GPU training required; runs cost roughly $2–$10 per optimization.

#agent-skills#ai-agent#gitHub#prompt-engineering#mlops+3
GitHub
AI Video·2026
Icon for item

HyperFrames

HeyGen

Author HTML-based video compositions and render deterministic, frame-accurate MP4s with agent-friendly tooling — preview in the browser, drive generation via AI agent skills, and use adapter runtimes (GSAP, Lottie, Three.js).

#video#ai-video#agent-skills#nodejs#javascript+3
GitHub
AI Agent·2026
Icon for item

ARIS (Auto-claude-code-research-in-sleep)

wanshuiyin

Lightweight, Markdown-only skill pack that lets LLM agents autonomously run ML research workflows—literature survey, idea discovery, cross-model review loops, experiment automation and paper writing—designed for Claude Code, Codex CLI, Cursor and local model setups.

#agent-skills#claude-code#codex#mcp#mcp-server+7
Hugging Face
AI Dataset·2026
Icon for item

bones-studio/seed

bones-studio

Provides an annotated multimodal human-motion dataset for language-to-action and robotics research, with BVH and MuJoCo files plus recordings targeted at Unitree-G1 and NVIDIA-SOMA platforms. Covers locomotion, gestures, dance and object interaction with English annotations and 100K–1M samples.

#robotics#video#multimodal#huggingface#ai-video+1
Hugging Face
AI Model·2026
Icon for item

Cosmos3-Nano (NVIDIA)

NVIDIA

Generate text, images, video, audio and action/robot trajectories from combined text, image, video, audio and action inputs. A Mixture-of-Transformers omnimodal foundation model (Cosmos3‑Nano, 16B params) focused on Physical AI (robotics, AV, simulation) and optimized for NVIDIA GPU runtimes.

#nvidia#foundation-model#multimodal#video#robotics+6
Hugging Face
AI Model·2026
Icon for item

google/gemma-4-31B-it

Google DeepMind

Instruction-tuned Gemma 4 31B multimodal model that generates text from text+image inputs with up to 256K context. Dense 31B variant optimized for vision-language understanding, long-context reasoning, and coding; Apache‑2.0 licensed.

#deepmind#google#huggingface#foundation-model#vision+3
Hugging Face
AI Dataset·2026
Icon for item

Roman1111111/claude-opus-4.6-10000x

Roman1111111

JSONL dataset of Claude Opus 4.6 chain-of-thought traces paired with high-difficulty math and logic problems for supervised fine-tuning and distillation; exposes step-by-step reasoning to teach process-oriented problem solving and improve math/logic accuracy in smaller LLMs.

#anthropic#claude#huggingface#LLM#ai-train+2
GitHub
AI Infra·2026
Icon for item

Open Brain

Nate B. Jones, OB1 Team

Provides a single persistent database and open protocol so multiple AI tools share the same memory — built-in vector search, an AI gateway, and capture/skill extensions. Best for teams and power users who want a unified, self-hosted agent memory instead of siloed notes or per-tool caches.

#mcp-server#mcp#embeddings#postgres#ai-agent+3
Hugging Face
AI Dataset·2026
Icon for item

Xperience-10M

Ropedia, Hugging Face

Provides 10 million synchronized egocentric experience episodes with structured 3D/4D multimodal annotations — 2.88B RGB frames, 720M depth frames, 576M pose/mocap frames and ~1PB total. Designed for embodied AI, robotics, and multimodal pretraining; research-only, gated access.

#multimodal#video#robotics#mocap#depth+4
Hugging Face
AI Model·2026
Icon for item

Gemma 4 26B A4B (google/gemma-4-26B-A4B-it)

Google DeepMind

Instruction-tuned Mixture-of-Experts multimodal model that generates text from text+image inputs while activating a 4B subset of parameters for faster inference; supports a 256K context window, multilingual vision-language tasks, and is available under Apache-2.0.

#deepmind#google#huggingface#multimodal#vision+4
GitHub
AI Agent·2026
Icon for item

gstack

Garry Tan

A 23-skill Claude Code toolkit that composes an LLM-driven virtual engineering team (CEO, designer, eng manager, QA, security, release) into slash-command workflows — includes real-browser QA, a persistent GBrain memory, multi-agent integrations, and team auto-update semantics.

#claude-code#anthropic#agent-skills#ai-coding#mcp-server+7
GitHub
AI Agent·2026
Icon for item

pi-autoresearch

davebcn87

Gives the pi terminal AI agent an autonomous experiment loop: propose code changes, run benchmarks, record metrics, auto-commit improvements and revert regressions. Ships a live widget/dashboard, MAD-based confidence scoring, hooks and backpressure checks — made for iterating on speed, bundle size, training loss and build times inside a terminal workflow.

#agent-skills#ai-agent#ai-coding#cli#terminal+3
  • Previous
  • 1
  • More pages
  • 98
  • 99
  • 100
  • More pages
  • 186
  • Next