LIFEHUBBER
Theme

AI Visuals

Find visual AI by what you want to make.

Browse projects for making images and video, building 3D worlds, and handling design or editing work. These are the visual projects from LifeHubber's wider AI Resources list.

Every entry keeps the original source close. LifeHubber is a starting point, not an endorsement or safety check, so review current setup, access, terms, privacy, and what happens to any files you provide before relying on a project.

Search all 22 visual resources

Search by name, task, or source.

Showing 22 of 22 resources

Choose a visual path

Start with the thing you want to make.

Browse 22 projects here. The full Resource list also covers models, agents, voice, data, and developer tools.

Image and video

Generate and animate visual ideas

Explore projects for creating images, video, motion, avatars, and longer visual sequences.

13

Boogu Image

boogu-project/Boogu-Image

GitHub
GitHub stars: 947 GitHub forks: 56 Declared license: Apache-2.0: Apache-2.0 Last pushed July 23, 2026: Pushed 20d ago
Stats from GitHub

A 10B research model family for text-to-image generation and instruction-based image editing, with Base, Turbo, Edit, and Edit-Turbo checkpoints, bilingual Chinese-English text rendering, local inference code, and public demos.

Bilingual image generation and editing

CoMoVi

IGL-HKUST/CoMoVi

GitHub
GitHub stars: 110 GitHub forks: 1 Declared license: Apache-2.0: Apache-2.0 Last pushed April 9, 2026: Pushed 4mo ago
Stats from GitHub

A framework for co-generating 3D human motion and realistic videos, with a focus on motion-conditioned video generation and training workflows.

Human motion and video generation

Dreamverse

hao-ai-lab/FastVideo/apps/dreamverse

GitHub

The FastVideo realtime video generation and editing platform, with backend and web UI setup, local GPU, B200, Docker, Modal, readiness checks, and mock-backend workflows.

Realtime video generation and editing

Fooocus

lllyasviel/Fooocus

GitHub
GitHub stars: 52.2K GitHub forks: 8.5K Declared license: GPL-3.0: GPL-3.0 Last pushed December 1, 2025: Pushed 8mo ago
Stats from GitHub

A local image-generation interface built around prompt-focused SDXL workflows, with Windows downloads, Colab access, inpainting, outpainting, image prompts, and presets.

Image generation UI

Lance

bytedance-research/Lance

Hugging Face
Hugging Face likes: 1.1K Hugging Face downloads, last 30 days: 337 Declared license: Apache-2.0: Apache-2.0 Last modified May 28, 2026: Modified 2mo ago
Stats from Hugging Face

A ByteDance Research unified multimodal model for image and video understanding, generation, and editing, with model files, demos, inference scripts, Gradio setup, benchmark scripts, and a stated 40GB VRAM inference requirement.

Unified image and video model

LongCat-Video-Avatar 1.5

meituan-longcat/LongCat-Video-Avatar-1.5

Hugging Face
Hugging Face likes: 714 Hugging Face downloads, last 30 days: 1.4K Declared license: MIT: MIT Last modified June 4, 2026: Modified 2mo ago
Stats from Hugging Face

A Meituan LongCat audio-driven avatar video model for single- and multi-person generation, with audio-text-to-video, audio-image-text-to-video, video continuation, model weights, GitHub quickstart, and project-reported evaluation materials.

Avatar video generation

LongLive

NVlabs/LongLive

GitHub
GitHub stars: 2.5K GitHub forks: 242 Declared license: Apache-2.0: Apache-2.0 Last pushed August 7, 2026: Pushed 5d ago
Stats from GitHub

An NVIDIA Labs infrastructure codebase for long video generation, with LongLive 2.0 NVFP4 and parallel training/inference support, multi-shot generation, async decoding, model links, docs, configs, and project-reported FPS and VBench results.

Long video generation infrastructure

LTX-2.5

Lightricks/LTX-2.5

Hugging Face
Hugging Face likes: 236 Hugging Face downloads, last 30 days: 39 Declared license: other: other Last modified August 11, 2026: Modified today
Stats from Hugging Face

A Lightricks audio-video model family with gated component weights, synchronized sound, text-to-video, image-to-video, multishot generation, local Python and ComfyUI paths, fine-tuning tools, and a separate hosted API.

Audio-video generation and editing

MiniMax H3

MiniMaxAI/MiniMax-H3

Hugging Face
Hugging Face likes: 3.6K Hugging Face downloads, last 30 days: 59.4K Declared license: other: other Last modified August 11, 2026: Modified 1d ago
Stats from Hugging Face

A MiniMax audio-video generation system with public H3-Base weights for text, image, video, and audio reference workflows, native stereo output, 4-to-15-second clips, local 768p generation, a hybrid 2K path, and a direct hosted API.

Multimodal audio-video generation

PersonaLive

GVCLab/PersonaLive

GitHub
GitHub stars: 3.5K GitHub forks: 496 Declared license: Apache-2.0: Apache-2.0 Last pushed May 15, 2026: Pushed 2mo ago
Stats from GitHub

A portrait image-animation framework for live-streaming-style video generation research, with offline and online inference, pretrained weights, a Web UI, and acceleration notes.

Portrait animation, video generation

Sana

NVlabs/Sana

GitHub
GitHub stars: 8.7K GitHub forks: 699 Declared license: Apache-2.0: Apache-2.0 Last pushed August 11, 2026: Pushed today
Stats from GitHub

An NVIDIA Labs codebase for efficient high-resolution image and video generation, with Sana, Sana-1.5, Sana-Sprint, Sana-Video, training and inference pipelines, model zoo links, ComfyUI and diffusers paths, and newer world-model work.

Efficient image and video generation

ViMax

HKUDS/ViMax

GitHub
GitHub stars: 11.8K GitHub forks: 1.8K Declared license: MIT: MIT Last pushed July 29, 2026: Pushed 14d ago
Stats from GitHub

An agentic video-generation framework for turning ideas, scripts, or longer narratives into planned video workflows, with script generation, storyboards, shot planning, reference selection, consistency checks, and configurable chat, image, and video model providers.

Agentic video generation workflow

Wan-Dancer-14B

Wan-AI/Wan-Dancer-14B

Hugging Face
Hugging Face likes: 183 Hugging Face downloads, last 30 days: 4.6K Declared license: Apache-2.0: Apache-2.0 Last modified July 17, 2026: Modified 26d ago
Stats from Hugging Face

A music-to-dance video model and framework that uses a reference image, a full music track, and a dance-style prompt, with separate global keyframe planning and local refinement stages for longer outputs.

Music-to-dance video generation

3D and worlds

Build scenes, spaces, and explorable worlds

Find projects for 3D assets, maps, scenes, world generation, and spatial workflows.

5

AniGen

VAST-AI-Research/AniGen

GitHub
GitHub stars: 476 GitHub forks: 41 Last pushed July 15, 2026: Pushed 27d ago
Stats from GitHub

A framework for generating animatable 3D assets from a single image, with mesh, skeleton, and skinning outputs for downstream animation and simulation workflows.

Animatable 3D asset generation

Arnis

louis-e/arnis

GitHub
GitHub stars: 17.4K GitHub forks: 1.5K Declared license: Apache-2.0: Apache-2.0 Last pushed August 10, 2026: Pushed 1d ago
Stats from GitHub

Generates real-world locations inside Minecraft with a surprisingly high level of detail.

World generation, mapping

LingBot-Map

robbyant/lingbot-map

GitHub
GitHub stars: 16.4K GitHub forks: 1.8K Declared license: Apache-2.0: Apache-2.0 Last pushed August 8, 2026: Pushed 3d ago
Stats from GitHub

A feed-forward 3D foundation model for streaming scene reconstruction, positioned around geometric consistency, long sequences, and efficient real-time inference.

Streaming 3D reconstruction

Lyra

nv-tlabs/lyra

GitHub
GitHub stars: 2.2K GitHub forks: 227 Declared license: Apache-2.0: Apache-2.0 Last pushed July 20, 2026: Pushed 22d ago
Stats from GitHub

A series of generative 3D world models from NVIDIA, positioned around explorable scenes, 3D consistency, and world-scale generation workflows.

3D world models

TRELLIS.2

microsoft/TRELLIS.2

GitHub
GitHub stars: 10.6K GitHub forks: 1.3K Declared license: MIT: MIT Last pushed July 10, 2026: Pushed 1mo ago
Stats from GitHub

A Microsoft 3D generation model for high-fidelity image-to-3D asset creation, using O-Voxel structured latents, PBR materials, inference code, and training tools.

3D generation, image-to-3D

Design and editing

Shape, edit, and move visual work

Explore tools for design systems, editable interfaces, background removal, and day-to-day visual editing.

4

FreeCut

walterlow/freecut

GitHub
GitHub stars: 2K GitHub forks: 309 Declared license: MIT: MIT Last pushed August 7, 2026: Pushed 4d ago
Stats from GitHub

A browser-based multi-track video editor with local workspace files, on-device transcription and captions, scene analysis, local voice and music tools, effects, subtitles, and in-browser export through modern Chromium APIs.

Local AI video editing, browser media workflow

HTML Anything

nexu-io/html-anything

GitHub
GitHub stars: 8.2K GitHub forks: 805 Declared license: Apache-2.0: Apache-2.0 Last pushed August 11, 2026: Pushed 1d ago
Stats from GitHub

A local agentic HTML editor that uses existing coding-agent CLI sessions to turn Markdown, data, and notes into exportable HTML, PNGs, decks, social cards, data reports, and web prototypes with skill templates and sandboxed preview.

Agentic HTML editor, local agent workflows

Open Design

nexu-io/open-design

GitHub
GitHub stars: 85.2K GitHub forks: 10K Declared license: Apache-2.0: Apache-2.0 Last pushed August 12, 2026: Pushed today
Stats from GitHub

A local-first AI design workspace that connects coding-agent CLIs to prototypes, decks, media outputs, design systems, sandboxed previews, and export workflows.

AI design workspace, agent-assisted prototypes

Penpot

penpot/penpot

GitHub
GitHub stars: 58.4K GitHub forks: 3.9K Declared license: MPL-2.0: MPL-2.0 Last pushed August 11, 2026: Pushed today
Stats from GitHub

An open-source collaborative design platform with an official MCP server that lets AI agents inspect and modify Penpot files, components, tokens, styles, layouts, and assets through hosted or local connection paths.

AI-connected design systems, design-to-code

Also in AI

Follow the next layer.

Keep the thread going with AI Resources for AI projects with original links and practical caveats, AI Guides for decision habits for messy AI choices, AI Pulse for separate public activity signals from tracked AI Resources and AI Ballot.