CoMoVi
IGL-HKUST/CoMoVi
A framework for co-generating 3D human motion and realistic videos, with a focus on motion-conditioned video generation and training workflows.
AI Visuals
Browse projects for making images and video, building 3D worlds, and handling design or editing work. These are the visual projects from LifeHubber's wider AI Resources list.
Every entry keeps the original source close. LifeHubber is a starting point, not an endorsement or safety check, so review current setup, access, terms, privacy, and what happens to any files you provide before relying on a project.
Search all 22 visual resources
Showing 22 of 22 resources
Choose a visual path
Browse 22 projects here. The full Resource list also covers models, agents, voice, data, and developer tools.
Reader signals
Reader marks can point you toward a place to start. Open anything that looks useful for your work and check the source.
IGL-HKUST/CoMoVi
A framework for co-generating 3D human motion and realistic videos, with a focus on motion-conditioned video generation and training workflows.
hao-ai-lab/FastVideo/apps/dreamverse
The FastVideo realtime video generation and editing platform, with backend and web UI setup, local GPU, B200, Docker, Modal, readiness checks, and mock-backend workflows.
nv-tlabs/lyra
A series of generative 3D world models from NVIDIA, positioned around explorable scenes, 3D consistency, and world-scale generation workflows.
penpot/penpot
An open-source collaborative design platform with an official MCP server that lets AI agents inspect and modify Penpot files, components, tokens, styles, layouts, and assets through hosted or local connection paths.
Recently added
Recently added visual projects to open, compare, and check at the source.
Lightricks/LTX-2.5
A Lightricks audio-video model family with gated component weights, synchronized sound, text-to-video, image-to-video, multishot generation, local Python and ComfyUI paths, fine-tuning tools, and a separate hosted API.
boogu-project/Boogu-Image
A 10B research model family for text-to-image generation and instruction-based image editing, with Base, Turbo, Edit, and Edit-Turbo checkpoints, bilingual Chinese-English text rendering, local inference code, and public demos.
MiniMaxAI/MiniMax-H3
A MiniMax audio-video generation system with public H3-Base weights for text, image, video, and audio reference workflows, native stereo output, 4-to-15-second clips, local 768p generation, a hybrid 2K path, and a direct hosted API.
walterlow/freecut
A browser-based multi-track video editor with local workspace files, on-device transcription and captions, scene analysis, local voice and music tools, effects, subtitles, and in-browser export through modern Chromium APIs.
Wan-AI/Wan-Dancer-14B
A music-to-dance video model and framework that uses a reference image, a full music track, and a dance-style prompt, with separate global keyframe planning and local refinement stages for longer outputs.
Image and video
Explore projects for creating images, video, motion, avatars, and longer visual sequences.
boogu-project/Boogu-Image
A 10B research model family for text-to-image generation and instruction-based image editing, with Base, Turbo, Edit, and Edit-Turbo checkpoints, bilingual Chinese-English text rendering, local inference code, and public demos.
IGL-HKUST/CoMoVi
A framework for co-generating 3D human motion and realistic videos, with a focus on motion-conditioned video generation and training workflows.
hao-ai-lab/FastVideo/apps/dreamverse
The FastVideo realtime video generation and editing platform, with backend and web UI setup, local GPU, B200, Docker, Modal, readiness checks, and mock-backend workflows.
lllyasviel/Fooocus
A local image-generation interface built around prompt-focused SDXL workflows, with Windows downloads, Colab access, inpainting, outpainting, image prompts, and presets.
bytedance-research/Lance
A ByteDance Research unified multimodal model for image and video understanding, generation, and editing, with model files, demos, inference scripts, Gradio setup, benchmark scripts, and a stated 40GB VRAM inference requirement.
meituan-longcat/LongCat-Video-Avatar-1.5
A Meituan LongCat audio-driven avatar video model for single- and multi-person generation, with audio-text-to-video, audio-image-text-to-video, video continuation, model weights, GitHub quickstart, and project-reported evaluation materials.
NVlabs/LongLive
An NVIDIA Labs infrastructure codebase for long video generation, with LongLive 2.0 NVFP4 and parallel training/inference support, multi-shot generation, async decoding, model links, docs, configs, and project-reported FPS and VBench results.
Lightricks/LTX-2.5
A Lightricks audio-video model family with gated component weights, synchronized sound, text-to-video, image-to-video, multishot generation, local Python and ComfyUI paths, fine-tuning tools, and a separate hosted API.
MiniMaxAI/MiniMax-H3
A MiniMax audio-video generation system with public H3-Base weights for text, image, video, and audio reference workflows, native stereo output, 4-to-15-second clips, local 768p generation, a hybrid 2K path, and a direct hosted API.
GVCLab/PersonaLive
A portrait image-animation framework for live-streaming-style video generation research, with offline and online inference, pretrained weights, a Web UI, and acceleration notes.
NVlabs/Sana
An NVIDIA Labs codebase for efficient high-resolution image and video generation, with Sana, Sana-1.5, Sana-Sprint, Sana-Video, training and inference pipelines, model zoo links, ComfyUI and diffusers paths, and newer world-model work.
HKUDS/ViMax
An agentic video-generation framework for turning ideas, scripts, or longer narratives into planned video workflows, with script generation, storyboards, shot planning, reference selection, consistency checks, and configurable chat, image, and video model providers.
Wan-AI/Wan-Dancer-14B
A music-to-dance video model and framework that uses a reference image, a full music track, and a dance-style prompt, with separate global keyframe planning and local refinement stages for longer outputs.
3D and worlds
Find projects for 3D assets, maps, scenes, world generation, and spatial workflows.
VAST-AI-Research/AniGen
A framework for generating animatable 3D assets from a single image, with mesh, skeleton, and skinning outputs for downstream animation and simulation workflows.
louis-e/arnis
Generates real-world locations inside Minecraft with a surprisingly high level of detail.
robbyant/lingbot-map
A feed-forward 3D foundation model for streaming scene reconstruction, positioned around geometric consistency, long sequences, and efficient real-time inference.
nv-tlabs/lyra
A series of generative 3D world models from NVIDIA, positioned around explorable scenes, 3D consistency, and world-scale generation workflows.
microsoft/TRELLIS.2
A Microsoft 3D generation model for high-fidelity image-to-3D asset creation, using O-Voxel structured latents, PBR materials, inference code, and training tools.
Design and editing
Explore tools for design systems, editable interfaces, background removal, and day-to-day visual editing.
walterlow/freecut
A browser-based multi-track video editor with local workspace files, on-device transcription and captions, scene analysis, local voice and music tools, effects, subtitles, and in-browser export through modern Chromium APIs.
nexu-io/html-anything
A local agentic HTML editor that uses existing coding-agent CLI sessions to turn Markdown, data, and notes into exportable HTML, PNGs, decks, social cards, data reports, and web prototypes with skill templates and sandboxed preview.
nexu-io/open-design
A local-first AI design workspace that connects coding-agent CLIs to prototypes, decks, media outputs, design systems, sandboxed previews, and export workflows.
penpot/penpot
An open-source collaborative design platform with an official MCP server that lets AI agents inspect and modify Penpot files, components, tokens, styles, layouts, and assets through hosted or local connection paths.
Also in AI
Keep the thread going with AI Resources for AI projects with original links and practical caveats, AI Guides for decision habits for messy AI choices, AI Pulse for separate public activity signals from tracked AI Resources and AI Ballot.