AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Category

Explore by categories

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.
  • All Categories

  • AI Leaderboard

  • AI Agent Tutorials

  • AI Coding Tutorials

  • AI Model

  • AI Agent Papers

  • Chatbot

  • AI Dataset

  • Machine Learning Foundation Books

  • AI Train

  • AI Deploy

  • AI Client

  • Machine Learning Foundation Papers

  • Machine Learning Foundation Tutorials

  • AI Image Demos

  • AI Agent

  • Large Language Model Tutorials

  • Large Language Model Papers

  • Machine Learning Engineering Papers

  • Computer Vision Tutorials

  • Computer Vision Papers

  • Natural Language Processing Papers

  • Reinforcement Learning Papers

  • Speech Technology Papers

  • AI API

  • AI Coding

  • AI Image

  • AI Video

  • MLOps

  • MCP Client

  • MCP Server

  • AI Video Papers

  • AI Audio

  • AI Others

  • AI Infra

  • Embodied AI

GitHub
MCP Client·2026
Icon for item

MemPalace

MemPalace (GitHub organization)

Stores conversation history verbatim and retrieves it via local semantic search with a structured index (wings/rooms/drawers). Pluggable vector backends and a local-first default mean high recall (benchmarked) without cloud or API keys—useful for agent memory and private RAG.

#mcp#mcp-server#mcp-client#embeddings#cli+3
GitHub
AI Infra·2026
Icon for item

CubeSandbox

Tencent Cloud

Provides hardware-isolated, sub-60ms, ultra-low-overhead sandboxes to run untrusted LLM/agent code. Offers event-level snapshots, kernel-level egress control, credential vaulting, and drop-in E2B SDK compatibility for high-density AI agent deployment.

#rust#ai-agent#ai-deploy#security#mLOps+1
GitHub
AI Agent·2026
Icon for item

Browser Harness

Browser Use

Connects an LLM to a real browser over an editable CDP websocket so the agent can drive clicks, navigation, and generate missing helper code during tasks. The harness self-heals by writing reusable helpers, supports local or cloud browsers, and can optionally record sessions for debugging.

#ai-agent#coding-agents#agent-skills#llm#python+3
GitHub
AI Agent·2026
Icon for item

ADR

Chenning Li, Pan Hu +10·Uber

Monitors and detects risky behavior in enterprise AI agents via high-fidelity telemetry, security benchmarking, and a two-tier detector. Comprises ADR Sensor, ADR-Bench, and ADR Detector; deployed in production at Uber and validated on public benchmarks.

#mcp#mcp-server#ai-agent#agent-skills#benchmark+6
GitHub
AI Infra·2026
Icon for item

FlashKDA

Yutian Chen, Zhiyuan Li +2·Moonshot AI

Provides high-performance CUDA/CUTLASS kernels implementing Kimi Delta Attention (KDA), accelerating KDA prefill on SM90+ (Hopper) GPUs. Integrates as a drop-in backend for flash-linear-attention, supports native variable-length batching, and targets K=V=128; requires CUDA 12.9+/PyTorch 2.4+.

#kimi#cuda#pytorch#ai-inference#benchmarks+2
GitHub
AI Infra·2026
Icon for item

Future AGI

Future AGI

Provides an end-to-end platform to evaluate, observe, protect, and optimize LLM and AI agent deployments. Integrates OpenTelemetry tracing, 50+ evaluation metrics, agent simulations, an OpenAI‑compatible gateway, and guardrails; self‑hostable under Apache 2.0.

#evaluation#LLM#ai-agent#mlops#docker+8
GitHub
AI Infra·2026
Icon for item

AiSOC

beenuar, AiSOC contributors

Ingests and normalizes security telemetry, runs multi-model AI agents to produce replayable investigations and automated triage/response; key features include a step-by-step Investigation Ledger, CI-gated eval harness, and self-hostable deployments.

#mcp-server#mcp#ai-agent#security#llm+9
GitHub
AI Infra·2026
Icon for item

Knowledge Catalog

Google Cloud (Google LLC), GoogleCloudPlatform (GitHub organization)

Provides tools and samples to build context management, enrichment, and retrieval solutions on Google Cloud Knowledge Catalog — an AI-oriented data catalog that builds a dynamic knowledge graph for structured and unstructured data, suitable for RAG and agent workflows.

#google#github#ai#ai-development#RAG+5
GitHub
AI Infra·2026
Icon for item

OpenTelemetry GenAI Semantic Conventions

OpenTelemetry

Defines OpenTelemetry semantic conventions for generative AI telemetry — spans, metrics, and events for GenAI clients, the Model Context Protocol (MCP), and provider-specific integrations. Includes YAML models, human-readable docs, and reference implementations to standardize observability across GenAI deployments.

#mcp#mcp-client#mcp-server#mlops#ai-api+3
GitHub
AI Infra·2026
Icon for item

DwarfStar 4

Salvatore Sanfilippo·antirez (GitHub), DeepSeek

Native local inference engine for DeepSeek V4 Flash (also supports GLM 5.2 and PRO on high‑memory machines). Focused features include model-specific loading, SSD expert streaming, asymmetric routed-expert 2-bit quant support, multi-GPU/tensor/pipeline parallelism, and an OpenAI-compatible server plus a native coding agent.

#deepseek#llama.cpp#metal#cuda#ai-inference+5
GitHub
AI Infra·2026
Icon for item

TokenSpeed

LightSeek Foundation

High-throughput LLM inference engine for agentic workloads, combining a local‑SPMD static compiler for parallelism, a C++ scheduler with a Python execution plane and type‑safe KV‑cache reuse, pluggable high-performance kernels (including an MLA implementation), and a low‑overhead AsyncLLM entrypoint for production GPU inference.

#llm#ai-inference#ai-serving#agent-skills#tensorrt+7
GitHub
AI Infra·2026
Icon for item

Agent Substrate

Runs a Kubernetes-native runtime that multiplexes many stateful agent-like actors onto a small pool of sandboxed worker pods via full-state snapshots and pre-warmed workers, enabling sub-second suspend/resume and 30x+ oversubscription.

#kubernetes#ai-agent#ai-deploy#mLOps#go+5
  • Previous
  • 1
  • 2
  • More pages
  • 17
  • 18
  • 19
  • Next