LIFEHUBBER
Theme

AI Visuals

Find visual AI by what you want to make.

Browse projects for making images and video, building 3D worlds, and handling design or editing work. These are the visual projects from LifeHubber's wider AI Resources list.

Every entry keeps the original source close. LifeHubber is a starting point, not an endorsement or safety check, so review current setup, access, terms, privacy, and what happens to any files you provide before relying on a project.

Search all 30 visual resources

Search by name or task, or filter by source site.

Showing 30 of 30 resources

Advertisements

Advertisements

Choose a visual path

Start with the thing you want to make.

Browse 30 projects here. The full Resource list also covers models, agents, voice, data, and developer tools.

Image and video

Generate and animate visual ideas

Explore projects for creating images, video, motion, avatars, and longer visual sequences.

18

AI Image Prompts Skill

YouMind-OpenLab/ai-image-prompts-skill

GitHub
GitHub stars: 935 GitHub forks: 106 Declared license: MIT: MIT Last pushed September 1, 2026: Pushed today
Stats from GitHub

An installable skill for AI assistants that searches a regularly updated library of 10,000+ community image prompts, returns a small set of matching examples with sample images, and can adapt a chosen template to a specific brief.

Cross-model image prompt search, sample images, agent skill

Awesome GPT Image 2

freestylefly/awesome-gpt-image-2

GitHub
GitHub stars: 26.8K GitHub forks: 2.6K Declared license: MIT: MIT Last pushed August 30, 2026: Pushed 2d ago
Stats from GitHub

A prompt-as-code library for GPT Image 2 with a categorized case gallery, reusable visual prompt templates, a live browsing site, and an installable agent skill for turning image ideas into structured prompts.

Structured image prompts, visual templates, agent skill

Boogu Image

boogu-project/Boogu-Image

GitHub
GitHub stars: 979 GitHub forks: 58 Declared license: Apache-2.0: Apache-2.0 Last pushed July 23, 2026: Pushed 1mo ago
Stats from GitHub

A 10B research model family for text-to-image generation and instruction-based image editing, with Base, Turbo, Edit, and Edit-Turbo checkpoints, bilingual Chinese-English text rendering, local inference code, and public demos.

Bilingual image generation and editing

CoMoVi

IGL-HKUST/CoMoVi

GitHub
GitHub stars: 110 GitHub forks: 1 Declared license: Apache-2.0: Apache-2.0 Last pushed April 9, 2026: Pushed 4mo ago
Stats from GitHub

A framework for co-generating 3D human motion and realistic videos, with a focus on motion-conditioned video generation and training workflows.

Human motion and video generation

Dreamverse

hao-ai-lab/FastVideo/apps/dreamverse

GitHub

The FastVideo realtime video generation and editing platform, with backend and web UI setup, local GPU, B200, Docker, Modal, readiness checks, and mock-backend workflows.

Realtime video generation and editing

Fooocus

lllyasviel/Fooocus

GitHub
GitHub stars: 52.6K GitHub forks: 8.6K Declared license: GPL-3.0: GPL-3.0 Last pushed December 1, 2025: Pushed 9mo ago
Stats from GitHub

A local image-generation interface built around prompt-focused SDXL workflows, with Windows downloads, Colab access, inpainting, outpainting, image prompts, and presets.

Image generation UI

Krea 2

krea-ai/krea-2

GitHub
GitHub stars: 803 GitHub forks: 63 Declared license: Apache-2.0: Apache-2.0 Last pushed July 24, 2026: Pushed 1mo ago
Stats from GitHub

A Krea AI open-weight text-to-image model family with an undistilled Raw checkpoint for fine-tuning and LoRA training, an eight-step Turbo checkpoint for inference, official local code, Diffusers, SGLang, ComfyUI, hosted routes, and the Krea 2 Community License.

Image generation and LoRA workflows

Krea 2 Turbo 4-Step LoRA

lvladikov/Krea2-Turbo-Distill-4step-LoRA

Hugging Face
Hugging Face likes: 54 Hugging Face downloads, last 30 days: 11.6K Last modified August 27, 2026: Modified 4d ago
Stats from Hugging Face

A community-made distillation LoRA that runs Krea 2 Turbo in four generation steps, with Diffusers weights, a stock ComfyUI workflow, rolling numbered checkpoints, and an independent Hugging Face demo.

Four-step Krea 2 Turbo image generation

Lance

bytedance-research/Lance

Hugging Face
Hugging Face likes: 1.1K Hugging Face downloads, last 30 days: 384 Declared license: Apache-2.0: Apache-2.0 Last modified May 28, 2026: Modified 3mo ago
Stats from Hugging Face

A ByteDance Research unified multimodal model for image and video understanding, generation, and editing, with model files, demos, inference scripts, Gradio setup, benchmark scripts, and a stated 40GB VRAM inference requirement.

Unified image and video model

LongCat-Video-Avatar 1.5

meituan-longcat/LongCat-Video-Avatar-1.5

Hugging Face
Hugging Face likes: 762 Hugging Face downloads, last 30 days: 1.7K Declared license: MIT: MIT Last modified June 4, 2026: Modified 2mo ago
Stats from Hugging Face

A Meituan LongCat audio-driven avatar video model for single- and multi-person generation, with audio-text-to-video, audio-image-text-to-video, video continuation, model weights, GitHub quickstart, and project-reported evaluation materials.

Avatar video generation

LongLive

NVlabs/LongLive

GitHub
GitHub stars: 2.6K GitHub forks: 253 Declared license: Apache-2.0: Apache-2.0 Last pushed August 7, 2026: Pushed 25d ago
Stats from GitHub

An NVIDIA Labs infrastructure codebase for long video generation, with LongLive 2.0 NVFP4 and FP8 inference paths, parallel training and inference, multi-shot and image-to-video support, async decoding, LongLive-RAG, model links, docs, and configs.

Long video generation infrastructure

LTX-2.5

Lightricks/LTX-2.5

Hugging Face
Hugging Face likes: 2.4K Hugging Face downloads, last 30 days: 1.2M Declared license: other: other Last modified September 1, 2026: Modified today
Stats from Hugging Face

A Lightricks audio-video model family with gated component weights, synchronized sound, text-to-video, image-to-video, multishot generation, local Python and ComfyUI paths, fine-tuning tools, and a separate hosted API.

Audio-video generation and editing

MiniMax H3

MiniMaxAI/MiniMax-H3

Hugging Face
Hugging Face likes: 4.7K Hugging Face downloads, last 30 days: 5.4M Declared license: other: other Last modified August 13, 2026: Modified 19d ago
Stats from Hugging Face

A MiniMax audio-video generation system with public H3-Base weights for text, image, video, and audio reference workflows, native stereo output, 4-to-15-second clips, local 768p generation, a hybrid 2K path, and a direct hosted API.

Multimodal audio-video generation

MiniMax H3 Turbo

lightx2v/Minimax-h3-Turbo

Hugging Face
Hugging Face likes: 779 Hugging Face downloads, last 30 days: 977.7K Declared license: Apache-2.0: Apache-2.0 Last modified August 27, 2026: Modified 5d ago
Stats from Hugging Face

A LightX2V community LoRA project that distills MiniMax H3 into four- and eight-step video-with-audio workflows, with Diffusers and ComfyUI checkpoints, single- and multi-GPU inference code, and a 768p four-step option.

Fewer-step MiniMax H3 generation

PersonaLive

GVCLab/PersonaLive

GitHub
GitHub stars: 3.6K GitHub forks: 509 Declared license: Apache-2.0: Apache-2.0 Last pushed August 28, 2026: Pushed 4d ago
Stats from GitHub

A portrait image-animation framework for live-streaming-style video generation research, with offline and online inference, pretrained weights, a Web UI, and acceleration notes.

Portrait animation, video generation

Sana

NVlabs/Sana

GitHub
GitHub stars: 8.9K GitHub forks: 714 Declared license: Apache-2.0: Apache-2.0 Last pushed August 27, 2026: Pushed 5d ago
Stats from GitHub

An NVIDIA Labs codebase for efficient high-resolution image and video generation, with Sana, Sana-1.5, Sana-Sprint, Sana-Video, Sana-WM world-model work, training and inference pipelines, model zoo links, and ComfyUI and diffusers paths.

Efficient image and video generation

ViMax

HKUDS/ViMax

GitHub
GitHub stars: 12.2K GitHub forks: 1.8K Declared license: MIT: MIT Last pushed July 29, 2026: Pushed 1mo ago
Stats from GitHub

An agentic video-generation framework for turning ideas, scripts, or longer narratives into planned video workflows, with a project-based web interface, interactive agent loop and terminal UI, storyboards, shot planning, previews, render checkpoints, and configurable model providers.

Agentic video generation workflow

Wan-Dancer-14B

Wan-AI/Wan-Dancer-14B

Hugging Face
Hugging Face likes: 183 Hugging Face downloads, last 30 days: 6.3K Declared license: Apache-2.0: Apache-2.0 Last modified July 17, 2026: Modified 1mo ago
Stats from Hugging Face

A music-to-dance video model and framework that uses a reference image, a full music track, and a dance-style prompt, with separate global keyframe planning and local refinement stages for longer outputs.

Music-to-dance video generation

3D and worlds

Build scenes, spaces, and explorable worlds

Find projects for 3D assets, maps, scenes, world generation, and spatial workflows.

6

AniGen

VAST-AI-Research/AniGen

GitHub
GitHub stars: 489 GitHub forks: 41 Last pushed July 15, 2026: Pushed 1mo ago
Stats from GitHub

A framework for generating animatable 3D assets from a single image, with mesh, skeleton, and skinning outputs for downstream animation and simulation workflows.

Animatable 3D asset generation

Arnis

louis-e/arnis

GitHub
GitHub stars: 17.7K GitHub forks: 1.5K Declared license: Apache-2.0: Apache-2.0 Last pushed September 1, 2026: Pushed today
Stats from GitHub

Generates real-world locations inside Minecraft with a surprisingly high level of detail.

World generation, mapping

Kimodo

nv-tlabs/kimodo

GitHub
GitHub stars: 3.4K GitHub forks: 377 Declared license: Apache-2.0: Apache-2.0 Last pushed July 13, 2026: Pushed 1mo ago
Stats from GitHub

An NVIDIA kinematic motion diffusion model for generating human and humanoid-robot motion from text plus pose, joint, waypoint, and path constraints, with local inference, a timeline demo, motion exports, official docs, and benchmark materials.

Controllable 3D human and humanoid motion generation

LingBot-Map

robbyant/lingbot-map

GitHub
GitHub stars: 16.8K GitHub forks: 1.9K Declared license: Apache-2.0: Apache-2.0 Last pushed August 31, 2026: Pushed 1d ago
Stats from GitHub

A feed-forward 3D foundation model for streaming scene reconstruction, positioned around geometric consistency, long sequences, and efficient real-time inference.

Streaming 3D reconstruction

Lyra

nv-tlabs/lyra

GitHub
GitHub stars: 2.3K GitHub forks: 234 Declared license: Apache-2.0: Apache-2.0 Last pushed July 20, 2026: Pushed 1mo ago
Stats from GitHub

A series of generative 3D world models from NVIDIA, positioned around explorable scenes, 3D consistency, and world-scale generation workflows.

3D world models

TRELLIS.2

microsoft/TRELLIS.2

GitHub
GitHub stars: 11K GitHub forks: 1.3K Declared license: MIT: MIT Last pushed July 10, 2026: Pushed 1mo ago
Stats from GitHub

A Microsoft 3D generation model for high-fidelity image-to-3D asset creation, using O-Voxel structured latents, PBR materials, inference code, and training tools.

3D generation, image-to-3D

Design and editing

Shape, edit, and move visual work

Explore tools for design systems, editable interfaces, background removal, and day-to-day visual editing.

6

Archify

tt-a1i/archify

GitHub
GitHub stars: 40.7K GitHub forks: 2.6K Declared license: MIT: MIT Last pushed September 1, 2026: Pushed today
Stats from GitHub

An agent skill and deterministic Node.js renderer for validated architecture, workflow, sequence, data-flow, and lifecycle maps, with typed JSON source, interactive tracing, and portable HTML, image, and motion exports.

Validated technical diagrams for coding agents

FreeCut

walterlow/freecut

GitHub
GitHub stars: 2.1K GitHub forks: 332 Declared license: MIT: MIT Last pushed August 31, 2026: Pushed today
Stats from GitHub

A browser-based multi-track video editor with local workspace files, on-device transcription and captions, scene analysis, local voice and music tools, effects, subtitles, and in-browser export through modern Chromium APIs.

Local AI video editing, browser media workflow

HTML Anything

nexu-io/html-anything

GitHub
GitHub stars: 8.6K GitHub forks: 844 Declared license: Apache-2.0: Apache-2.0 Last pushed August 23, 2026: Pushed 9d ago
Stats from GitHub

A local agentic HTML editor that uses existing coding-agent CLI sessions to turn Markdown, data, and notes into exportable HTML, PNGs, decks, social cards, data reports, and web prototypes with skill templates and sandboxed preview.

Agentic HTML editor, local agent workflows

Open Design

nexu-io/open-design

GitHub
GitHub stars: 93.2K GitHub forks: 10.8K Declared license: Apache-2.0: Apache-2.0 Last pushed September 1, 2026: Pushed today
Stats from GitHub

A local-first AI design workspace that connects coding-agent CLIs to prototypes, decks, media outputs, design systems, sandboxed previews, and export workflows.

AI design workspace, agent-assisted prototypes

Penpot

penpot/penpot

GitHub
GitHub stars: 59.4K GitHub forks: 4K Declared license: MPL-2.0: MPL-2.0 Last pushed September 1, 2026: Pushed today
Stats from GitHub

An open-source collaborative design platform with an official MCP server that lets AI agents inspect and modify Penpot files, components, tokens, styles, layouts, and assets through hosted or local connection paths.

AI-connected design systems, design-to-code

UI UX Pro Max

nextlevelbuilder/ui-ux-pro-max-skill

GitHub
GitHub stars: 123.7K GitHub forks: 13.2K Declared license: MIT: MIT Last pushed August 31, 2026: Pushed today
Stats from GitHub

A searchable UI and UX knowledge skill for AI coding assistants, with design-system generation, product-specific style and palette guidance, stack-aware implementation notes, and pre-delivery checks.

AI coding skill, design systems, UI and UX guidance

Advertisements

Advertisements