AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Tag

Explore by tags

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.
  • All

  • 30u30

  • ASR

  • ChatGPT

  • GNN

  • IDE

  • RAG

  • agent-skills

  • ai

  • ai-agent

  • ai-api

  • ai-api-management

  • ai-client

  • ai-coding

  • ai-demos

  • ai-deploy

  • ai-development

  • ai-framework

  • ai-image

  • ai-image-demos

  • ai-inference

  • ai-leaderboard

  • ai-library

  • ai-rank

  • ai-security

  • ai-serving

  • ai-tools

  • ai-train

  • ai-video

  • ai-workflow

  • AIGC

  • algorithms

  • alibaba

  • amazon

  • android

  • anthropic

  • arabic

  • audio

  • aws

  • benchmark

  • benchmarks

  • biology

  • blog

  • book

  • bun

  • bytedance

  • chatbot

  • chatgpt

  • chemistry

  • claude

  • claude-code

  • cli

  • clickhouse

  • code

  • codex

  • coding

  • coding-agents

  • common-crawl

  • copilot

  • course

  • cpu

  • cuda

  • cursor

  • deepmind

  • deepseek

  • depth

  • devops

  • diffusers

  • distillation

  • docker

  • drug-discovery

  • electron

  • embeddings

  • engineering

  • evaluation

  • facebook

  • finance

  • flow-matching

  • foundation

  • foundation-model

  • fp8

  • gcode

  • gcp

  • gemini

  • gemini-cli

  • gemma

  • genomics

  • gguf

  • gitHub

  • github

  • go

  • google

  • gradient-booting

  • grok

  • groq

  • huggingface

  • hy_v4

  • image

  • imatrix

  • ios

  • java

  • javascript

  • json

  • kimi

  • kotlin

  • kubernetes

  • laion

  • llama.cpp

  • LLM

  • llm

  • long-horizon

  • lora

  • mLOps

  • math

  • mcap

  • mcp

  • mcp-client

  • mcp-server

  • meta-ai

  • meta-pytorch

  • metal

  • microsoft

  • mlops

  • mobile

  • mocap

  • moe

  • multilingual

  • multimodal

  • mysql

  • NLP

  • nlp

  • nodejs

  • numpy

  • nvidia

  • ocr

  • ollama

  • openai

  • opencode

  • pandas

  • paper

  • parquet

  • physics

  • pi

  • plugin

  • polars

  • postgres

  • privacy

  • programming

  • prompt-engineering

  • pwa

  • python

  • pytorch

  • qwen

  • react

  • reasoning

  • redis

  • refactoring

  • research

  • retrieval

  • RL

  • rl

  • robotics

  • rust

  • safetensors

  • science

  • security

  • segmentation

  • sft

  • shodan

  • skillkit

  • software-engineering

  • sora

  • speech

  • sqlite

  • ssh

  • stt

  • supabase

  • swe

  • swift

  • tensorrt

  • terminal

  • thinking

  • trae

  • transformers

  • translation

  • tts

  • tutorial

  • typescript

  • unsloth-dynamic

  • vibe-coding

  • video

  • vision

  • vllm

  • voice

  • vue

  • vulkan

  • web-search

  • windsurf

  • xAI

  • xai

  • youtube

GitHub
AI Audio·2025
Icon for item

Supertonic

Supertone Inc.

Delivers multilingual, on-device text-to-speech via ONNX Runtime with prebuilt ONNX assets and cross-platform SDKs (Python, Node, mobile); targets low-latency, privacy-preserving TTS with ready demos and 31-language support in v3.

#audio#speech#multilingual#huggingface#python+8
Hugging Face
AI Dataset·2025
Icon for item

Mental Association Dataset (Rapidata/psychology-association-kiki-bouba-etc)

Rapidata

Collects ~200,000 human responses to 20 visual/semantic association questions (e.g., Bouba–Kiki), with per-response image options and demographic metadata — useful for cross‑cultural perception and evaluation of multimodal systems, but not guaranteed as a rigorously controlled experimental sample.

#huggingface#multimodal#image#pandas#polars+3
GitHub
AI Coding Tutorials·2025
Icon for item

Easy-Vibe

Datawhale (Datawhale China), Sanbu (project lead) +6

A step-by-step, beginner-first programming course that teaches 'vibe coding'—conversational workflows to turn ideas into AI-enabled web and full‑stack prototypes. Features interactive simulated coding, multi-language docs, stage-based projects (from simple demos to SaaS capstones) and advanced agent/Claude Code guidance.

#vibe-coding#course#ai-coding#github#multilingual+3
GitHub
AI Audio·2026
Icon for item

Pocket TTS

Manu Orsini, Simon Rouard +5

Generates low-latency, streaming text-to-speech entirely on CPUs (no GPU or cloud API required), using an ~100M-parameter model with voice cloning and multilingual support. Optimized for low resource use (2 CPU cores, ~200ms to first audio chunk) — suited for local, privacy-sensitive, or embedded TTS.

#pytorch#python#speech#multilingual#cli+4
Hugging Face
AI Dataset·2026
Icon for item

Waxal NLP Datasets

Google Research, Makerere University +6

Provides open ASR and TTS speech data for 24 Sub‑Saharan African languages to train and evaluate speech models. Includes ~1,250 hours of transcribed ASR and ~235 hours of single‑speaker TTS with train/validation/test/unlabeled splits and mixed CC-BY licenses.

#multilingual#audio#speech#ASR#tts+3
Hugging Face
AI Audio·2026
Icon for item

Qwen3-TTS-12Hz-1.7B-CustomVoice

Hangrui Hu, Xinfa Zhu +14·Qwen / QwenLM, Alibaba Group +1

Generates controllable multilingual speech from text with nine predefined timbres and custom-voice control; supports voice design, quick voice cloning and low-latency streaming (first audio packet after a single character), suitable for real-time TTS and voice-design workflows.

#qwen#huggingface#tts#voice#multilingual+4
Hugging Face
AI Dataset·2026
Icon for item

INFINI-NEWS Corpus

Ruggero Marino Lazzaroni, Jana Lasser +1·University of Graz, Hugging Face

Processed, multilingual news corpus of 1.357B articles extracted from Common Crawl CC‑News with per-article WARC provenance. Includes trafilatura-extracted bodies, language labels (GlotLID & CommonLingua), IPTC topic tags, monthly Parquet shards and a companion FM-index for sub-10ms substring queries; bulk text access is gated for academic research.

#huggingface#multilingual#NLP#ai-development
Hugging Face
AI Dataset·2026
Icon for item

Ultra-FineWeb-L3

openbmb

Provides L3 refined synthetic training data by converting high-quality web corpora into Q&A pairs and multi-style rewrites; supplies 400B+ English and 200B+ Chinese tokens for late-stage LLM pretraining and decay-phase training.

#LLM#huggingface#nlp#ai-train#multilingual+1
Hugging Face
AI Dataset·2026
Icon for item

WebWorldData

Qwen

Provides 1.06M web interaction trajectories (state, action, next_state) represented primarily as A11y trees for training browser world models and web agents. Covers diverse real‑web domains, English/Chinese pages, and long contexts (up to 30K tokens); residual PII and dynamic content may limit reproducibility.

#huggingface#ai-agent#agent-skills#llm#multilingual+2
GitHub
AI Video·2026
Icon for item

Seedance 2.0 Skill OS

Iamemily2050

Converts scene intent into production-ready Seedance 2.0 prompts, reference-role mappings, and IP-safe rewrites for multimodal (text/image/audio/video) video generation. Ships as a modular agent-skill OS with multilingual examples, troubleshooting tools, and pro filmmaker handoff artifacts.

#agent-skills#ai-video#video#multimodal#prompt-engineering+5
Hugging Face
AI Model·2026
Icon for item

google/gemma-4-E4B-it

Google DeepMind

An instruction‑tuned Gemma 4 E4B multimodal model on Hugging Face that accepts text, images and audio and generates text; notable for 128K long context support, built-in thinking mode, and an on‑device‑friendly E4B architecture under an Apache‑2.0 license.

#deepmind#google#transformers#huggingface#multimodal+5
GitHub
AI Infra·2026
Icon for item

Pretext

Cheng Lou·Midjourney

Measures multiline text layout and block height without triggering browser reflow: it measures text segments once via Canvas+Intl.Segmenter and caches widths, then computes line breaks with pure arithmetic. Useful for streaming AI text, virtualization, and custom per-line rendering.

#javascript#typescript#nodejs#react#github+3
  • Previous
  • 1
  • 2
  • 3
  • More pages
  • 11
  • 12
  • Next