AIAny
Icon for item

Liminal-Dreamcore-1K

Collection of 1,000 AI-generated dreamcore aesthetic images (2K JPEGs, numbered 001–1000) intended for creative prototyping and visual research. Images were produced with GPT Image 2 and released under an MIT license.

Introduction

Why this matters

Dreamcore is a distinct aesthetic niche — liminal, nostalgic, and intentionally uncanny. This dataset provides a compact, consistently generated snapshot of that visual language: 1,000 images produced with the same pipeline (GPT Image 2, 2K resolution, medium quality) and numbered for easy programmatic access. For anyone studying aesthetic style, dataset curation, or creative workflows, a focused synthetic collection like this is useful because it isolates a coherent stylistic space.

What Sets It Apart
  • Consistency by design — each image was produced under a fixed generation pipeline and prompting process, which reduces within-dataset stylistic noise and makes the collection suitable for controlled experiments (e.g., classifier training, style transfer targets).
  • Compact and labeled filenames — 001.jpg through 1000.jpg simplifies indexing, batching, and reproducible sampling for experiments or demos.
  • Permissive reuse — published with an MIT-compatible license on the dataset card, enabling broad reuse and commercial experimentation (still observe caveats below).
Who It's For and Tradeoffs

Great fit if you want a small, coherent corpus to: prototype aesthetic classifiers, fine-tune small generative models for a specific look, build galleries or UI demos, or evaluate how models respond to a narrowly defined visual style. It is less suitable as a general-purpose training corpus: 1,000 images is small for training large models from scratch and the images are synthetic and style-specific, so models trained only on this set will not generalize to diverse, real-world imagery.

Important caveats: the dataset is entirely AI-generated (GPT Image 2) and while the project card states an MIT license, legal and ethical concerns remain — for commercial deployment verify whether any generated images reproduce third-party copyrighted content or sensitive attributes. Also expect generation artifacts and a limited diversity of subjects and scenes inherent to a single aesthetic pipeline.

Where it fits

Think of this collection as a lightweight, plug-and-play asset for aesthetic experimentation rather than a comprehensive vision dataset. Use it for style-focused evaluation, demo content, or as a target domain for transfer learning; avoid using it as the sole training source for models intended to operate on varied real-world images.

Information

Categories

More Items

Hugging Face

Provides 1.3 billion platform-specific video URLs extracted from CommonCrawl along with crawl metadata (no media included), serving as the source corpus for the LAION-BVD multimodal video dataset; distributed on Hugging Face in Parquet format.

Hugging Face

Provides a public test split of multimodal financial GUI interaction examples for evaluating agents that convert instructions and screenshots into grounded UI actions. Includes step-level screenshots, dialogue history, an OpenAI-style computer_use tool schema, and JSON next-action references; training data available on request.

Hugging Face

Contains 40,000 teacher-generated reasoning traces distilled from the Qwen3.8-27B model for supervised fine-tuning and analysis. Covers code, math, science and logic; each example pairs a <think> chain-of-thought with a final response and is distributed in JSONL/Parquet for SFT workflows.