LIFEHUBBER
Theme

AI Visuals

Find visual AI by what you want to make.

Browse projects for making images and video, building 3D worlds, and handling design or editing work. These are the visual projects from LifeHubber's wider AI Resources list.

Every entry keeps the original source close. LifeHubber is a starting point, not an endorsement or safety check, so review current setup, access, terms, privacy, and what happens to any files you provide before relying on a project.

Search all 38 visual resources

Search by name or task, or filter by source site.

Showing 38 of 38 resources

Advertisements

Advertisements

Choose a visual path

Start with the thing you want to make.

Browse 38 projects here. The full Resource list also covers models, agents, voice, data, and developer tools.

Image and video

Generate and animate visual ideas

Explore projects for creating images, video, motion, avatars, and longer visual sequences.

21

AI Image Prompts Skill

YouMind-OpenLab/ai-image-prompts-skill

GitHub
GitHub stars: 1.1K GitHub forks: 119 Declared license: MIT: MIT Last pushed September 21, 2026: Pushed today
Stats from GitHub

An installable skill for AI assistants that searches a regularly updated library of 10,000+ community image prompts, returns a small set of matching examples with sample images, and can adapt a chosen template to a specific brief.

Cross-model image prompt search, sample images, agent skill

Awesome GPT Image 2

freestylefly/awesome-gpt-image-2

GitHub
GitHub stars: 33.1K GitHub forks: 3.2K Declared license: MIT: MIT Last pushed September 11, 2026: Pushed 10d ago
Stats from GitHub

A prompt-as-code library for GPT Image 2 with a categorized case gallery, reusable visual prompt templates, a live browsing site, and an installable agent skill for turning image ideas into structured prompts.

Structured image prompts, visual templates, agent skill

Boogu Image

boogu-project/Boogu-Image

GitHub
GitHub stars: 987 GitHub forks: 60 Declared license: Apache-2.0: Apache-2.0 Last pushed July 23, 2026: Pushed 2mo ago
Stats from GitHub

A 10B research model family for text-to-image generation and instruction-based image editing, with Base, Turbo, Edit, and Edit-Turbo checkpoints, bilingual Chinese-English text rendering, local inference code, and public demos.

Bilingual image generation and editing

CoMoVi

IGL-HKUST/CoMoVi

GitHub
GitHub stars: 110 GitHub forks: 1 Declared license: Apache-2.0: Apache-2.0 Last pushed April 9, 2026: Pushed 5mo ago
Stats from GitHub

A framework for co-generating 3D human motion and realistic videos, with a focus on motion-conditioned video generation and training workflows.

Human motion and video generation

Dreamverse

hao-ai-lab/FastVideo/apps/dreamverse

GitHub

The FastVideo realtime video generation and editing platform, with backend and web UI setup, local GPU, B200, Docker, Modal, readiness checks, and mock-backend workflows.

Realtime video generation and editing

Fooocus

lllyasviel/Fooocus

GitHub
GitHub stars: 53.1K GitHub forks: 8.6K Declared license: GPL-3.0: GPL-3.0 Last pushed December 1, 2025: Pushed 9mo ago
Stats from GitHub

A local image-generation interface built around prompt-focused SDXL workflows, with Windows downloads, Colab access, inpainting, outpainting, image prompts, and presets.

Image generation UI

Krea 2

krea-ai/krea-2

GitHub
GitHub stars: 852 GitHub forks: 66 Declared license: Apache-2.0: Apache-2.0 Last pushed July 24, 2026: Pushed 1mo ago
Stats from GitHub

A Krea AI open-weight text-to-image model family with an undistilled Raw checkpoint for fine-tuning and LoRA training, an eight-step Turbo checkpoint for inference, official local code, Diffusers, SGLang, ComfyUI, hosted routes, and the Krea 2 Community License.

Image generation and LoRA workflows

Krea 2 Turbo 4-Step LoRA

lvladikov/Krea2-Turbo-Distill-4step-LoRA

Hugging Face
Hugging Face likes: 170 Hugging Face downloads, last 30 days: 31.3K Declared license: other: other Last modified September 15, 2026: Modified 5d ago
Stats from Hugging Face

A community-made distillation LoRA that runs Krea 2 Turbo in four generation steps, with Diffusers weights, a stock ComfyUI workflow, rolling numbered checkpoints, and an independent Hugging Face demo.

Four-step Krea 2 Turbo image generation

Lance

bytedance-research/Lance

Hugging Face
Hugging Face likes: 1.1K Hugging Face downloads, last 30 days: 854 Declared license: Apache-2.0: Apache-2.0 Last modified May 28, 2026: Modified 3mo ago
Stats from Hugging Face

A ByteDance Research unified multimodal model for image and video understanding, generation, and editing, with model files, demos, inference scripts, Gradio setup, benchmark scripts, and a stated 40GB VRAM inference requirement.

Unified image and video model

LongCat-Video-Avatar 1.5

meituan-longcat/LongCat-Video-Avatar-1.5

Hugging Face
Hugging Face likes: 823 Hugging Face downloads, last 30 days: 2K Declared license: MIT: MIT Last modified June 4, 2026: Modified 3mo ago
Stats from Hugging Face

A Meituan LongCat audio-driven avatar video model for single- and multi-person generation, with audio-text-to-video, audio-image-text-to-video, video continuation, model weights, GitHub quickstart, and project-reported evaluation materials.

Avatar video generation

LongLive

NVlabs/LongLive

GitHub
GitHub stars: 2.6K GitHub forks: 255 Declared license: Apache-2.0: Apache-2.0 Last pushed September 7, 2026: Pushed 14d ago
Stats from GitHub

An NVIDIA Labs infrastructure codebase for long video generation, with LongLive 2.0 NVFP4 and FP8 inference paths, parallel training and inference, multi-shot and image-to-video support, async decoding, LongLive-RAG, model links, docs, and configs.

Long video generation infrastructure

LTX-2.5

Lightricks/LTX-2.5

Hugging Face
Hugging Face likes: 4.6K Hugging Face downloads, last 30 days: 1.6M Declared license: other: other Last modified September 1, 2026: Modified 20d ago
Stats from Hugging Face

A Lightricks audio-video model family with gated component weights, synchronized sound, text-to-video, image-to-video, multishot generation, local Python and ComfyUI paths, fine-tuning tools, and a separate hosted API.

Audio-video generation and editing

MiniMax H3

MiniMaxAI/MiniMax-H3

Hugging Face
Hugging Face likes: 5.5K Hugging Face downloads, last 30 days: 4.1M Declared license: other: other Last modified August 13, 2026: Modified 1mo ago
Stats from Hugging Face

A MiniMax audio-video generation system with public H3-Base weights for text, image, video, and audio reference workflows, native stereo output, 4-to-15-second clips, local 768p generation, a hybrid 2K path, and a direct hosted API.

Multimodal audio-video generation

MiniMax H3 Integrations

MiniMax-AI/awesome-minimax-h3-integration

GitHub
GitHub stars: 375 GitHub forks: 28 Last pushed September 17, 2026: Pushed 3d ago
Stats from GitHub

A community-maintained MiniMax H3 integration index that maps checkpoints, hardware and VRAM starting points, runtimes, ComfyUI nodes, prompting tools, acceleration routes, and deployment options.

MiniMax H3 setup and integration map

MiniMax H3 Turbo

lightx2v/Minimax-h3-Turbo

Hugging Face
Hugging Face likes: 974 Hugging Face downloads, last 30 days: 1.6M Declared license: Apache-2.0: Apache-2.0 Last modified September 10, 2026: Modified 11d ago
Stats from Hugging Face

A LightX2V community LoRA project that distills MiniMax H3 into four- and eight-step video-with-audio workflows, with Diffusers and ComfyUI checkpoints, single- and multi-GPU inference code, and a 768p four-step option.

Fewer-step MiniMax H3 generation

PersonaLive

GVCLab/PersonaLive

GitHub
GitHub stars: 3.8K GitHub forks: 544 Declared license: Apache-2.0: Apache-2.0 Last pushed August 28, 2026: Pushed 24d ago
Stats from GitHub

A portrait image-animation framework for live-streaming-style video generation research, with offline and online inference, pretrained weights, a Web UI, and acceleration notes.

Portrait animation, video generation

Qwen-Image-2.1

QwenLM/Qwen-Image-2.1

GitHub
GitHub stars: 836 GitHub forks: 32 Last pushed September 20, 2026: Pushed 1d ago
Stats from GitHub

Qwen's open-weight image model for text-to-image generation, single- and multi-reference editing, marked-region changes, subject extraction, native transparent RGBA output, and 2K workflows.

Image generation, editing, and transparent output

Sana

NVlabs/Sana

GitHub
GitHub stars: 9.1K GitHub forks: 732 Declared license: Apache-2.0: Apache-2.0 Last pushed September 21, 2026: Pushed today
Stats from GitHub

An NVIDIA Labs codebase for efficient high-resolution image and video generation, with Sana, Sana-1.5, Sana-Sprint, Sana-Video, Sana-WM world-model work, training and inference pipelines, model zoo links, and ComfyUI and diffusers paths.

Efficient image and video generation

VDN-H3

OpenVDN/vdn-minimax-h3

GitHub
GitHub stars: 518 GitHub forks: 28 Declared license: Apache-2.0: Apache-2.0 Last pushed September 19, 2026: Pushed 1d ago
Stats from GitHub

A community-made MiniMax H3 derivative that adds a hybrid linear-and-softmax attention branch, eight- and 50-step checkpoints, FP8 single- and multi-GPU inference paths, and the training code behind the conversion.

Hybrid-attention MiniMax H3 acceleration

ViMax

HKUDS/ViMax

GitHub
GitHub stars: 12.4K GitHub forks: 1.9K Declared license: MIT: MIT Last pushed September 20, 2026: Pushed 1d ago
Stats from GitHub

An agentic video-generation framework for turning ideas, scripts, or longer narratives into planned video workflows, with a project-based web interface, interactive agent loop and terminal UI, storyboards, shot planning, previews, render checkpoints, and configurable model providers.

Agentic video generation workflow

Wan-Dancer-14B

Wan-AI/Wan-Dancer-14B

Hugging Face
Hugging Face likes: 186 Hugging Face downloads, last 30 days: 15.8K Declared license: Apache-2.0: Apache-2.0 Last modified July 17, 2026: Modified 2mo ago
Stats from Hugging Face

A music-to-dance video model and framework that uses a reference image, a full music track, and a dance-style prompt, with separate global keyframe planning and local refinement stages for longer outputs.

Music-to-dance video generation

3D and worlds

Build scenes, spaces, and explorable worlds

Find projects for 3D assets, maps, scenes, world generation, and spatial workflows.

8

AniGen

VAST-AI-Research/AniGen

GitHub
GitHub stars: 501 GitHub forks: 42 Last pushed July 15, 2026: Pushed 2mo ago
Stats from GitHub

A framework for generating animatable 3D assets from a single image, with mesh, skeleton, and skinning outputs for downstream animation and simulation workflows.

Animatable 3D asset generation

Arnis

louis-e/arnis

GitHub
GitHub stars: 18K GitHub forks: 1.5K Declared license: Apache-2.0: Apache-2.0 Last pushed September 20, 2026: Pushed 1d ago
Stats from GitHub

Generates real-world locations inside Minecraft with a surprisingly high level of detail.

World generation, mapping

DLP3D.AI

dlp3d-ai/dlp3d.ai

GitHub
GitHub stars: 358 GitHub forks: 38 Declared license: MIT: MIT Last pushed May 25, 2026: Pushed 3mo ago
Stats from GitHub

A real-time framework for voice conversations with customizable 3D characters, combining model providers, speech, memory and reaction logic, facial animation, whole-body motion, and browser playback.

Real-time voice and 3D character framework

Kimodo

nv-tlabs/kimodo

GitHub
GitHub stars: 3.6K GitHub forks: 396 Declared license: Apache-2.0: Apache-2.0 Last pushed July 13, 2026: Pushed 2mo ago
Stats from GitHub

An NVIDIA kinematic motion diffusion model for generating human and humanoid-robot motion from text plus pose, joint, waypoint, and path constraints, with local inference, a timeline demo, motion exports, official docs, and benchmark materials.

Controllable 3D human and humanoid motion generation

LingBot-Map

robbyant/lingbot-map

GitHub
GitHub stars: 17.1K GitHub forks: 1.9K Declared license: Apache-2.0: Apache-2.0 Last pushed September 8, 2026: Pushed 13d ago
Stats from GitHub

A feed-forward 3D foundation model for streaming scene reconstruction, positioned around geometric consistency, long sequences, and efficient real-time inference.

Streaming 3D reconstruction

LingBot-World-V2-1.3B-Causal-Fast

Robbyant/lingbot-world-v2-1.3b-causal-fast

ModelScope

Robbyant's smaller causal-fast checkpoint for action-conditioned interactive video worlds, with a roughly 6.8 GB DiT package, public inference code and paper, differing official GPU setup notes, and declared CC BY-NC-SA 4.0 terms.

Smaller interactive video world model

Lyra

nv-tlabs/lyra

GitHub
GitHub stars: 2.3K GitHub forks: 233 Declared license: Apache-2.0: Apache-2.0 Last pushed July 20, 2026: Pushed 2mo ago
Stats from GitHub

A series of generative 3D world models from NVIDIA, positioned around explorable scenes, 3D consistency, and world-scale generation workflows.

3D world models

TRELLIS.2

microsoft/TRELLIS.2

GitHub
GitHub stars: 11.3K GitHub forks: 1.4K Declared license: MIT: MIT Last pushed July 10, 2026: Pushed 2mo ago
Stats from GitHub

A Microsoft 3D generation model for high-fidelity image-to-3D asset creation, using O-Voxel structured latents, PBR materials, inference code, and training tools.

3D generation, image-to-3D

Design and editing

Shape, edit, and move visual work

Explore tools for design systems, editable interfaces, background removal, and day-to-day visual editing.

9

Archify

tt-a1i/archify

GitHub
GitHub stars: 68.8K GitHub forks: 4.6K Declared license: MIT: MIT Last pushed September 21, 2026: Pushed today
Stats from GitHub

An agent skill and deterministic Node.js renderer for validated architecture, workflow, sequence, data-flow, and lifecycle maps, with typed JSON source, interactive tracing, and portable HTML, image, and motion exports.

Validated technical diagrams for coding agents

FreeCut

walterlow/freecut

GitHub
GitHub stars: 2.2K GitHub forks: 342 Declared license: MIT: MIT Last pushed August 31, 2026: Pushed 20d ago
Stats from GitHub

A browser-based multi-track video editor with local workspace files, on-device transcription and captions, scene analysis, local voice and music tools, effects, subtitles, and in-browser export through modern Chromium APIs.

Local AI video editing, browser media workflow

HTML Anything

nexu-io/html-anything

GitHub
GitHub stars: 8.9K GitHub forks: 866 Declared license: Apache-2.0: Apache-2.0 Last pushed September 15, 2026: Pushed 6d ago
Stats from GitHub

A local agentic HTML editor that uses existing coding-agent CLI sessions to turn Markdown, data, and notes into exportable HTML, PNGs, decks, social cards, data reports, and web prototypes with skill templates and sandboxed preview.

Agentic HTML editor, local agent workflows

Open CoDesign

OpenCoworkAI/open-codesign

GitHub
GitHub stars: 8K GitHub forks: 831 Declared license: MIT: MIT Last pushed September 20, 2026: Pushed today
Stats from GitHub

A local-first desktop design workspace for turning prompts into interactive prototypes, slide decks, marketing pieces, and exportable files through a choice of hosted or local model routes.

Desktop AI workspace for prototypes, decks, and visual assets

Open Design

nexu-io/open-design

GitHub
GitHub stars: 97.4K GitHub forks: 11.3K Declared license: Apache-2.0: Apache-2.0 Last pushed September 21, 2026: Pushed today
Stats from GitHub

A local-first AI design workspace that connects coding-agent CLIs to prototypes, decks, media outputs, design systems, sandboxed previews, and export workflows.

AI design workspace, agent-assisted prototypes

OUI-1

thesysdev/OUI-1

Hugging Face
Hugging Face likes: 127 Hugging Face downloads, last 30 days: 2.5K Declared license: Apache-2.0: Apache-2.0 Last modified September 11, 2026: Modified 10d ago
Stats from Hugging Face

A specialized text-diffusion model for generating OpenUI interface screens from a component library and plain-language brief, with public weights, OpenUI tooling, vLLM and Transformers paths, function calling, and published benchmark code and raw outputs.

OpenUI screen-generation model

Penpot

penpot/penpot

GitHub
GitHub stars: 60.2K GitHub forks: 4.1K Declared license: MPL-2.0: MPL-2.0 Last pushed September 21, 2026: Pushed today
Stats from GitHub

An open-source collaborative design platform with an official MCP server that lets AI agents inspect and modify Penpot files, components, tokens, styles, layouts, and assets through hosted or local connection paths.

AI-connected design systems, design-to-code

UI UX Pro Max

nextlevelbuilder/ui-ux-pro-max-skill

GitHub
GitHub stars: 129.5K GitHub forks: 13.8K Declared license: MIT: MIT Last pushed September 21, 2026: Pushed today
Stats from GitHub

A searchable UI and UX knowledge skill for AI coding assistants, with design-system generation, product-specific style and palette guidance, stack-aware implementation notes, and pre-delivery checks.

AI coding skill, design systems, UI and UX guidance

UX/UI Agent Skills

plugin87/ux-ui-agent-skills

GitHub
GitHub stars: 1.4K GitHub forks: 146 Declared license: MIT: MIT Last pushed September 16, 2026: Pushed 5d ago
Stats from GitHub

A Claude-focused project kit that connects design tokens, component specifications, accessibility guidance, visual-taste references, UX writing, framework adapters, runnable skills, validation gates, rendered critique, and a product starter.

Claude design workflow, tokens, accessibility, critique

Advertisements

Advertisements