AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Tag

Explore by tags

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.
  • All

  • 30u30

  • ASR

  • ChatGPT

  • GNN

  • IDE

  • RAG

  • agent-skills

  • ai

  • ai-agent

  • ai-api

  • ai-api-management

  • ai-client

  • ai-coding

  • ai-demos

  • ai-deploy

  • ai-development

  • ai-framework

  • ai-image

  • ai-image-demos

  • ai-inference

  • ai-leaderboard

  • ai-library

  • ai-rank

  • ai-security

  • ai-serving

  • ai-tools

  • ai-train

  • ai-video

  • ai-workflow

  • AIGC

  • algorithms

  • alibaba

  • amazon

  • android

  • anthropic

  • arabic

  • audio

  • aws

  • benchmark

  • benchmarks

  • biology

  • blog

  • book

  • bun

  • bytedance

  • chatbot

  • chatgpt

  • chemistry

  • claude

  • claude-code

  • cli

  • clickhouse

  • code

  • codex

  • coding

  • coding-agents

  • common-crawl

  • copilot

  • course

  • cpu

  • cuda

  • cursor

  • deepmind

  • deepseek

  • depth

  • devops

  • diffusers

  • distillation

  • docker

  • drug-discovery

  • electron

  • embeddings

  • engineering

  • evaluation

  • facebook

  • finance

  • flow-matching

  • foundation

  • foundation-model

  • fp8

  • gcode

  • gcp

  • gemini

  • gemini-cli

  • gemma

  • genomics

  • gguf

  • gitHub

  • github

  • go

  • google

  • gradient-booting

  • grok

  • groq

  • huggingface

  • hy_v4

  • image

  • imatrix

  • ios

  • java

  • javascript

  • json

  • kimi

  • kotlin

  • kubernetes

  • laion

  • llama.cpp

  • LLM

  • llm

  • long-horizon

  • lora

  • mLOps

  • math

  • mcap

  • mcp

  • mcp-client

  • mcp-server

  • meta-ai

  • meta-pytorch

  • metal

  • microsoft

  • mlops

  • mobile

  • mocap

  • moe

  • multilingual

  • multimodal

  • mysql

  • NLP

  • nlp

  • nodejs

  • numpy

  • nvidia

  • ocr

  • ollama

  • openai

  • opencode

  • pandas

  • paper

  • parquet

  • physics

  • pi

  • plugin

  • polars

  • postgres

  • privacy

  • programming

  • prompt-engineering

  • pwa

  • python

  • pytorch

  • qwen

  • react

  • reasoning

  • redis

  • refactoring

  • research

  • retrieval

  • RL

  • rl

  • robotics

  • rust

  • safetensors

  • science

  • security

  • segmentation

  • sft

  • shodan

  • skillkit

  • software-engineering

  • sora

  • speech

  • sqlite

  • ssh

  • stt

  • supabase

  • swe

  • swift

  • tensorrt

  • terminal

  • thinking

  • trae

  • transformers

  • translation

  • tts

  • tutorial

  • typescript

  • unsloth-dynamic

  • vibe-coding

  • video

  • vision

  • vllm

  • voice

  • vue

  • vulkan

  • web-search

  • windsurf

  • xAI

  • xai

  • youtube

GitHub
AI Image·2025
Icon for item

Qwen-Image

QwenLM (Alibaba Group)·Alibaba, Qwen Team

A 20B-parameter MMDiT diffusion model that generates and edits images with accurate embedded text, including dense Chinese and English typography. Handles complex multi-line layouts and identity-preserving edits while keeping text legible.

#foundation-model#ai-image#github#huggingface#pytorch+2
GitHub
AI Image·2025
Icon for item

DINOv3

Meta AI Research (FAIR), Oriane Siméoni +25·Meta AI Research (FAIR)

Self-supervised vision foundation model producing dense, patch-level features that transfer to classification, segmentation, depth, and detection with a frozen backbone. Spans ViT-S (21M) to ViT-7B (6.7B params), plus ConvNeXt and satellite variants.

#vision#pytorch#github#meta-ai#huggingface+3
GitHub
AI Client·2025
Icon for item

CADAM

Adam-CAD, Zach Dive +2

Generate parametric 3D CAD models from natural language and images in the browser, with real-time preview and exports to STL/SCAD. Runs client-side via OpenSCAD WebAssembly, extracts adjustable parameters, and integrates Anthropic for shape generation—suited for rapid prototyping and 3D-printing workflows.

#github#anthropic#ai-image#ai-demos#ai-tools+3
GitHub
AI Video·2025
Icon for item

Project Lyra: Open Generative 3D World Models

NVIDIA Spatial Intelligence Lab (nv-tlabs)

Generates explorable, 3D-consistent virtual worlds from a single image or short video. Includes official implementations of Lyra‑1 (feed‑forward 3D/4D scene generation via video-diffusion self-distillation) and Lyra‑2 (long-horizon, explorable generative 3D worlds). Best for research and creative prototyping; requires substantial GPU compute.

#nvidia#vision#ai-video#ai-image#foundation-model+5
Hugging Face
AI Dataset·2025
Icon for item

malcolmrey's Various AI Model Repository

Malcolm Reynolds

Provides a curated collection of hands-on tutorials, workflows and auxiliary files for training and using generative-model tooling (Stable Diffusion, Flux, WAN). Key items include a WAN 2.1 LoRA training tutorial and an articles collection covering DreamBooth, LoRA, LyCORIS and SDXL.

#lora#diffusers#huggingface#tutorial#ai-workflow+3
GitHub
AI Image·2025
Icon for item

Chandra OCR

Datalab (datalab.to), Vik Paruchuri·Datalab

Converts images and PDFs into structured Markdown, HTML, or JSON while preserving layout, handling tables, math, handwriting, charts, and chemistry diagrams across 90+ languages. Runs locally via HuggingFace or against a vLLM server.

#ocr#vllm#huggingface#pytorch#python+3
GitHub
AI Video·2025
Icon for item

Pixelle-Video

AIDC-AI

Automatically generates complete short-form videos from a single topic: drafts script with an LLM, produces AI images/video, synthesizes multilingual TTS (including voice cloning), adds background music, and composes the final video. Supports local ComfyUI/RunningHub or direct model APIs and customizable templates.

#ai-video#video#ai-image#tts#ai-tools+4
GitHub
AI Image·2025
Icon for item

Next AI Draw.io

DayuanJiang·Independent

Turn plain-English requests into editable draw.io diagrams: the model writes the underlying draw.io XML, which renders live in an embedded canvas. Upload images, PDFs, or text to replicate, refine through chat, and roll back via version history.

#github#ai-tools#ai-image#chatbot#ai-development
GitHub
AI Video·2025
Icon for item

PersonaLive

Zhiyuan Li, Chi-Man Pun +3·University of Macau, Great Bay University +1

Generates real-time, infinite-length portrait video from one reference image on a 12GB GPU. Combines implicit facial signals and 3D keypoints with step-distilled diffusion and autoregressive micro-chunk streaming for low-latency live use.

#ai-video#ai-image#pytorch#AIGC#github+3
GitHub
AI Model·2025
Icon for item

TRELLIS.2 — Native and Compact Structured Latents for 3D Generation

Jianfeng Xiang, Xiaoxue Chen +9·Microsoft

Converts images (and other conditions) into high-fidelity, fully textured 3D assets using a 4B-parameter generative model and a field‑free sparse voxel format (O‑Voxel). Handles arbitrary topology, PBR materials, and near real-time mesh/voxel conversions; requires Linux and an NVIDIA GPU with >=24GB memory.

#microsoft#pytorch#huggingface#nvidia#ai-image+3
GitHub
AI Client·2025
Icon for item

PPT Master — AI generates natively editable PPTX from any document

Hugo He (hugohe3)

Generates natively editable PPTX from PDFs, DOCX, URLs, or Markdown — producing real PowerPoint shapes, text boxes, and charts (not images). An open-source, model-agnostic, local-first pipeline that integrates with multiple AI editors while keeping your data on-device.

#gitHub#python#ai-tools#ai-image#ai-api+1
GitHub
AI Image·2025
Icon for item

Reverse-Engineering SynthID

Alosh Denny

Detects and surgically removes Google's SynthID watermark from images using multi-resolution spectral analysis and a resolution-aware codebook; provides a V3 bypass with high PSNR and strong phase-coherence reduction. Research-focused and intended for analysis/defense, not misrepresentation.

#gemini#ai-image#ai-tools#python#huggingface+4
  • Previous
  • 1
  • More pages
  • 5
  • 6
  • 7
  • More pages
  • 16
  • Next