AIAIAny
  • Search
  • Collection
  • Category
  • Tag
  • Daily AI
AIAIAny

Category

Explore by categories

AIAIAny

Curated AI Resources for Everyone

[email protected]

Powered by airss.app

Product
  • Search
  • Collection
  • Category
  • Tag
Resources
  • Blog
Company
  • Privacy Policy
  • Terms of Service
  • Sitemap
Copyright © 2026 All Rights Reserved.
  • All Categories

  • AI Leaderboard

  • AI Agent Tutorials

  • AI Coding Tutorials

  • AI Model

  • AI Agent Papers

  • Chatbot

  • AI Dataset

  • Machine Learning Foundation Books

  • AI Train

  • AI Deploy

  • AI Client

  • Machine Learning Foundation Papers

  • Machine Learning Foundation Tutorials

  • AI Image Demos

  • AI Agent

  • Large Language Model Tutorials

  • Large Language Model Papers

  • Machine Learning Engineering Papers

  • Computer Vision Tutorials

  • Computer Vision Papers

  • Natural Language Processing Papers

  • Reinforcement Learning Papers

  • Speech Technology Papers

  • AI API

  • AI Coding

  • AI Image

  • AI Video

  • MLOps

  • MCP Client

  • MCP Server

  • AI Video Papers

  • AI Audio

  • AI Others

  • AI Infra

  • Embodied AI

Hugging Face
AI Image·2026
Icon for item

HiDream-O1-Image

HiDream-ai

Generates and edits high-resolution images (up to 2048×2048) from text and reference images, plus subject-driven personalization. Implements a pixel-level unified transformer that encodes raw pixels and text in one token space and includes a reasoning-driven prompt agent for layout and text rendering.

#transformers#multimodal#ai-image#huggingface#gemma+4
Hugging Face
AI Model·2026
Icon for item

Lens: Rethinking Training Efficiency for Foundational Text-to-Image Models

Microsoft

Research-focused text-to-image foundation model that prioritizes training efficiency: a 3.8B-parameter architecture trained on an 800M image-text corpus with mixed-resolution learning, FLUX.2 VAE, RL tuning, and a distilled 4-step Lens-Turbo for fast high-resolution generation.

#microsoft#huggingface#ai-image#image#transformers+3
GitHub
AI Image·2026
Icon for item

无限画布 (infinite-canvas)

Node-based infinite-canvas web workstation for iterative visual creation — integrates image/video generation, reference editing, prompt library, multi-agent assistants, and asset management. Runs in-browser with configurable OpenAI-compatible endpoints; suited for local/personal deployment (AGPL-3.0).

#ai-image#image#ai-agent#mcp-client#typescript+6
Hugging Face
AI Model·2026
Icon for item

Bonsai Image · Ternary 4B (gemlite 2-bit)

Prism ML (prism-ml)

A ternary-weight (~1.58-bit) 4B text-to-image diffusion transformer optimized for NVIDIA GPUs using Gemlite INT2 and HQQ; it reduces the transformer to ~1.21 GB (4.55 GB CUDA payload) and targets 1024×1024 generation with a 4-step FlowMatch-Euler sampler.

#huggingface#ai-image#image#nvidia#ai-inference+3
Hugging Face
AI Image·2026
Icon for item

Krea 2 (Comfy-Org/Krea-2)

Comfy-Org, Krea

Provides ComfyUI-ready repackaged checkpoints of the Krea 2 image model family for local text-to-image workflows. Includes RAW (undistilled base for fine-tuning and LoRA training) and Turbo (8-step distilled checkpoint for fast inference), using a Qwen Image VAE and Qwen3‑VL encoder.

#qwen#diffusers#huggingface#ai-image#image+2
Hugging Face
AI Image·2026
Icon for item

Krea 2 Turbo

Sangwu Lee, Erwann Millon +14·Krea.ai, Inc.

Generates images from natural-language prompts as an 8-step distilled checkpoint of Krea 2, optimized for fast iterative text-to-image workflows with style references and 1K–2K resolution outputs.

#diffusers#ai-image#image#vision#huggingface+5
Hugging Face
AI Image·2026
Icon for item

fal · Krea 2 Style LoRAs

ilkerzgi

Provides 1,503 Krea 2 style LoRAs (original safetensors + ComfyUI builds) trained on fal.ai, each with a short trigger phrase and downloadable weights for quick style transfer or further retraining.

#huggingface#ai-image#ai-tools#diffusers#AIGC+2
Hugging Face
AI Image·2026
Icon for item

Sun Direction LoRA (Flux2Klein 9B)

eric-venti-seeds

Applies or repositions directional sunlight in outdoor images by using a LoRA trained for Flux2Klein 9B to match a reference sun elevation and rotation. Workflow uses an overcast intermediate and a sphere (ball) reference; includes a ComfyUI node and Blender scene for rendering the reference.

#ai-image#image#huggingface#diffusers#ai-demos
Hugging Face
AI Model·2026
Icon for item

Giga-World-1

open-gigaai

Diffusion-based generative model for scene and video synthesis, providing full Diffusers checkpoints and scene LoRA for fast adaptation. Includes Stage‑1 nano (1.3B) and pro (5B) variants and modular transformer/VAE components.

#diffusers#huggingface#lora#pytorch#image+4
Hugging Face
AI Model·2026
Icon for item

Mage-Flow

Zhang Xinjie, Zhang Peng +22·Microsoft

Efficient 4B native-resolution diffusion foundation model for text-to-image generation and instruction-based image editing. Uses a lightweight Mage‑VAE tokenizer and a 4B NR‑MMDiT backbone to produce 512–2048 outputs with low memory and fast inference; ships in Base, RL-aligned and few-step Turbo variants.

#multimodal#ai-image#image#vision#flow-matching+7
Hugging Face
AI Image·2026
Icon for item

Mage-Flow-Edit-Turbo

Xinjie Zhang, Peng Zhang +22·Microsoft

Performs instruction-based image editing from reference images using a 4B native-resolution diffusion transformer; the Turbo variant uses 4-step distillation for interactive latency (≈1.02 s per 1024² edit on A100) while supporting semantic, appearance, structure-aware and restoration edits.

#ai-image#image#multimodal#huggingface#microsoft+5
Hugging Face
AI Image·2026
Icon for item

Kroma v0.1

lodestones

Provides a single-file ComfyUI-compatible LoRA plus full-weight RMSNorm/modulation .diff deltas that reproduce a Krea 2 fine-tune when applied to Krea 2 checkpoints. Rank-256 adapters, ~1.88 GB, MIT-licensed.

#lora#huggingface#diffusers#ai-image#qwen
  • Previous
  • 1
  • 2
  • 3
  • 4
  • 5
  • 6
  • Next