An HDR LoRA fine-tune for Lightricks' LTX-2.3 (22B) that enables image‑conditioned any‑to‑any image-to-video and text-to-video generation. Designed for HDR-aware synthesis workflows; requires the LTX-2.3 base model and a LoRA-capable runtime.
A 284B-parameter Mixture-of-Experts LLM with only 13B activated parameters, designed for 1,000,000-token contexts. Uses hybrid compressed attention and mixed FP4/FP8 precision to reduce long-context KV-cache and per-token FLOPs; aimed at long-document QA, RAG pipelines, and local/high-capacity inference.
Generates conversational and reasoning outputs with support for million‑token contexts; uses a hybrid attention + MoE design to cut long‑context inference FLOPs and KV cache. Suited for long‑document retrieval, coding and complex reasoning; MIT licensed.
A vision-oriented foundation checkpoint for low-latency inference — DeepSeek V4 base in safetensors with FP8 optimizations. Designed for fast image generation and embedding use in inference pipelines; verify license and FP8/runtime compatibility before production use.
Base image-generation foundation model tuned for visual search and prompt-guided synthesis, intended as a compact starting point for local inference or fine-tuning. Emphasizes easy integration into image pipelines and suitability for downstream adaptation.
A 14B dense tri‑mode language model that supports autoregressive, diffusion‑based parallel decoding, and self‑speculation—designed to increase token throughput and acceptance length; best suited for researchers and engineers exploring decode‑efficiency tradeoffs on NVIDIA hardware under the Nemotron Open Model License.
A Qwen-3.6 27B model variant optimized for DFlash (speculative decoding) to reduce generation latency and increase throughput. Focuses on faster inference on serving stacks and is suitable for text-generation endpoints where lower latency and resource efficiency matter.
High-resolution vision transformers pretrained on one billion human images for human-centric tasks such as pose estimation, body-part segmentation, surface-normal and pointmap prediction. Provides multiple backbone sizes and task-specific checkpoints; released under the Sapiens2 license.
Performs task-aware generative video restoration and editing in latent video space — restoration, super-resolution, watermark and subtitle removal — adapting LTX‑2.3 with IC‑Edit/IC‑LoRA adapters to prioritize temporal consistency and occlusion-aware reconstruction.
A lightweight 'drafter' assistant for Gemma 4 31B that generates speculative token drafts to enable up-to-2× decoding speedups while preserving final output quality; compatible with Hugging Face Transformers and any-to-any pipelines.
Generates anime-style images from natural-language prompts with a full fine-tune family built on Z-Image Base — available as Base, 8-step and 4-step distillations, plus AIO and GGUF variants for 8GB/low-VRAM workflows (BF16/FP8 formats).
Open-source Mixture-of-Experts LLM designed for extremely long-context (up to 1M tokens) text generation and agentic workflows; uses a hybrid attention + MTP design to reduce KV-cache footprint while enabling 42B active parameters and FP8 mixed-precision training.