Skip to main content
Browse Models

wavymulder

Analog Diffusion

Released

2022-12-10

Family

Stable Diffusion 1

Type

Fine-Tuned Model

Downloads

External Download

You are about to open a link to an external source. Verify the URL before continuing.

Download

Model Checkpoint

FP16 · analog-diffusion-1.0.safetensors

Model Report

Overview

Analog Diffusion is a text-to-image generative model developed using DreamBooth and trained on a wide array of analog photographs to capture and reproduce the distinctive aesthetic of analog film photography. This model aims to deliver images that reflect the color palette, texture, and subtle imperfections characteristic of traditional film, providing users with an accessible means to simulate analog styles within their synthetic image generation workflows. Detailed documentation and official model releases can be found on the Analog Diffusion Hugging Face page.

Analog Diffusion output collage showing various portraits with an analog film effect

Figure 1. A header collage of eight model outputs, each demonstrating the analog style—featuring both historical and fictional figures, all rendered with distinctive analog film-like softness.

Model Architecture

Analog Diffusion was engineered as a DreamBooth-based model, utilizing Stable Diffusion 1.5 as its foundational architecture. The integration of a Variational Autoencoder (VAE) within this architecture is crucial for encoding images into manageable latent space representations and decoding generated samples back to the image domain. The DreamBooth approach allows Analog Diffusion to specialize in analog visual characteristics by fine-tuning on a dedicated dataset, thereby producing outputs that consistently emulate analog photographic styles.

Training Data and Methodology

The model was trained with a dataset comprised of a diverse range of analog photographic samples. The training methodology centered around DreamBooth fine-tuning, enabling the model to robustly internalize the defining attributes of analog imagery such as grain structures, color shifting, and filmic contrast curves. This targeted approach allows for the simulation of analog effects independent of the subject matter, supporting high versatility in the generated outputs. The creator has shared specifics about the parameters used for example outputs, which include prompt structures and sampler settings.

Key Features and Style Control

The primary feature of Analog Diffusion is its ability to replicate the aesthetic of analog photography on arbitrary text prompts. To invoke this effect, users must include the activation token “analog style” in their prompt. The model provides additional controls for output sharpness and atmospheric haze; by including terms such as “blur” and “haze” in the negative prompt, users can emphasize image clarity, though this may attenuate the analog characteristics. Detailed guidance for prompt engineering is outlined in the model documentation.

Grid of diverse environments showcasing analog style output from Analog Diffusion

Figure 2. Model-generated grid demonstrating the analog style across environments, from natural scenes and interiors to urban and dramatic landscapes. Prompt: 'analog style, snowy house at dusk, Christmas lights; analog style, volcanic eruption night; analog style, cozy attic room', and others.

Output Diversity and Example Results

Analog Diffusion is designed to handle a broad array of subjects, including portraits, characters, animals, and environments. The model’s stylistic treatment is preserved across these varied scenarios, consistently imparting analog hues, contrast profiles, and film-like artifacts. Representative collages shared by the author display outputs such as cinematic portraits, natural vistas, architectural scenes, and animal studies—all unified by the analog signature. Additional uncurated sample batches are available for review through the non-cherrypicked examples archive.

Example collage of character and animal outputs generated by Analog Diffusion

Figure 3. Collage of various subjects generated by Analog Diffusion, including a lion, armored figure, stylized portraits, and an owl, all with strong analog photographic coloration and texture. Prompts included 'analog style, portrait of a lion', 'analog style, person in armor, desert', among others.

Limitations and Considerations

While Analog Diffusion was trained exclusively on analog photographs, it has shown a propensity to generate unintended content in certain prompts. The creator recommends using negative prompting to mitigate the generation of unintended content. Additionally, a trade-off exists between maximizing sharpness (by requesting reduced haze and blur) and the preservation of analog authenticity, as increased clarity may diminish the intended filmic effect. Further technical notes and clarification on usage are detailed in the official documentation and release notes.

Applications

The model is suitable for creative text-to-image tasks where an authentically analog appearance is desired, such as concept art, synthetic photography, and digital moodboarding. The creator provides a user-accessible Gradio-based interface that allows real-time experimentation with prompts and style parameters. For technical integration, the model checkpoint is freely available for direct download, facilitating research and offline inference as required.

External Resources

About Stable Diffusion 1: Stable Diffusion is an open-source text-to-image generative AI model that transforms textual adminDescriptions into corresponding images. Technologically, it employs a latent diffusion model architecture, enhancing computational efficiency by performing diffusion processes in a compressed latent space, which enables high-quality image generation with reduced resource requirements.

More in the Stable Diffusion 1 Family

stabilityai /

Stable Diffusion 1.1

A latent text-to-image diffusion model trained on LAION datasets that generates 512×512 images from natural language prompts using compressed latent space processing.
stabilityai /

Stable Diffusion 1.5

Text-to-image diffusion model trained on LAION dataset subset, generating 512x512 images from natural language prompts using latent space processing.
prompthero /

OpenJourney v4

SD 1.5 fine-tuned on 124k+ additional images generated with Midjourney v4, leading to results that resemble this other closed-source image generation model.
Photographer /

Photon

Photon aims to generate photorealistic and visually appealing images effortlessly.
KandooAI /

Juggernaut

Popular SD 1.5 fine-tune with capability to produce detailed images of a versatile breadth of subjects.
Lykon /

Dreamshaper

SD 1.5 fine-tune with strong art generation ability and a broad generalist capabilities.
SG_161222 /

Realistic Vision

SD 1.5 fine-tune specialized in creating photorealistic portraits of humans.
Meina /

Meina Mix

Model resulting for merging 7 different anime-focused SD 1.5 checkpoints.
epinikion /

epiCRealism

Popular SD 1.5 fine-tune with high competence in translating simple text prompts into realistic images of people.
Lykon /

Absolute Reality

One of the top SD 1.5 variant for generating life-like images of people and objects.
Cyberdelia /

Cyber Realistic

Versatile photorealistic SD 1.5 fine-tune capable of generating a wide range of convincing photographic images.
Merjic /

MajicMIX Realistic

Popular SD 1.5 photorealism fine-tune with training data weighted on people of asian descent.
epinikion /

epiCPhotoGasm

A Stable Diffusion 1.5-based checkpoint model designed for photorealistic image generation with simplified prompting and demographic diversity.
lllyasviel /

ControlNet SD 1.5 Canny

SD 1.5 ControlNet model to replicate the composion of a source image using edge-detection.
lllyasviel /

ControlNet SD 1.5 IP2P

SD 1.5 ControlNet trained with pixel-to-pixel instruction.
lllyasviel /

ControlNet SD 1.5 Depth

SD 1.5 ControlNet model to replicate the depth of a source image.
lllyasviel /

ControlNet SD 1.5 MLSD

SD 1.5 ControlNet model to detect straight-lines, useful for architecture and man-made objects.
lllyasviel /

ControlNet SD 1.5 Normal

SD 1.5 ControlNet model to replicate the depth of a source image, with additional surface details and geometry.
lllyasviel /

ControlNet SD 1.5 Open Pose

SD 1.5 ControlNet model for copying human poses.
lllyasviel /

ControlNet SD 1.5 Scribble

SD 1.5 ControlNet model for converting sketches to images.
lllyasviel /

ControlNet SD 1.5 Segmentation

SD 1.5 ControlNet model for detecting and segmenting distinct parts of images to use in the generation.
lllyasviel /

ControlNet SD 1.5 Soft Edge

SD 1.5 ControlNet model to detect soft-edges, especially useful for recoloring and stylizing.
lllyasviel /

ControlNet SD 1.5 Inpaint

SD 1.5 ControlNet model trained with image inpainting.
lllyasviel /

ControlNet SD 1.5 Line Art

SD 1.5 ControlNet model trained with line art generation.
lllyasviel /

ControlNet SD 1.5 Lineart Anime

SD 1.5 ControlNet model trained with anime line art generation.
lllyasviel /

ControlNet SD 1.5 Shuffle

SD 1.5 ControlNet model trained with image shuffling.
lllyasviel /

ControlNet SD 1.5 Tile

SD 1.5 ControlNet model trained with image tiling.
tencent /

ControlNet 1.5 IP Adapter

SD 1.5 ControlNet model for conditioning on an image prompt.
tencent /

ControlNet 1.5 QR Code

SD 1.5 ControlNet model for generating stylized QR codes.