CoMoVi
IGL-HKUST/CoMoVi
A framework for co-generating 3D human motion and realistic videos, with a focus on motion-conditioned video generation and training workflows.
AI Visuals
Browse projects for making images and video, building 3D worlds, and handling design or editing work. These are the visual projects from LifeHubber's wider AI Resources list.
Every entry keeps the original source close. LifeHubber is a starting point, not an endorsement or safety check, so review current setup, access, terms, privacy, and what happens to any files you provide before relying on a project.
Search all 30 visual resources
Showing 30 of 30 resources
Choose a visual path
Browse 30 projects here. The full Resource list also covers models, agents, voice, data, and developer tools.
Reader signals
Reader marks can point you toward a place to start. Open anything that looks useful for your work and check the source.
IGL-HKUST/CoMoVi
A framework for co-generating 3D human motion and realistic videos, with a focus on motion-conditioned video generation and training workflows.
hao-ai-lab/FastVideo/apps/dreamverse
The FastVideo realtime video generation and editing platform, with backend and web UI setup, local GPU, B200, Docker, Modal, readiness checks, and mock-backend workflows.
nv-tlabs/lyra
A series of generative 3D world models from NVIDIA, positioned around explorable scenes, 3D consistency, and world-scale generation workflows.
MiniMaxAI/MiniMax-H3
A MiniMax audio-video generation system with public H3-Base weights for text, image, video, and audio reference workflows, native stereo output, 4-to-15-second clips, local 768p generation, a hybrid 2K path, and a direct hosted API.
penpot/penpot
An open-source collaborative design platform with an official MCP server that lets AI agents inspect and modify Penpot files, components, tokens, styles, layouts, and assets through hosted or local connection paths.
Recently added
Recently added visual projects to open, compare, and check at the source.
tt-a1i/archify
An agent skill and deterministic Node.js renderer for validated architecture, workflow, sequence, data-flow, and lifecycle maps, with typed JSON source, interactive tracing, and portable HTML, image, and motion exports.
YouMind-OpenLab/ai-image-prompts-skill
An installable skill for AI assistants that searches a regularly updated library of 10,000+ community image prompts, returns a small set of matching examples with sample images, and can adapt a chosen template to a specific brief.
lvladikov/Krea2-Turbo-Distill-4step-LoRA
A community-made distillation LoRA that runs Krea 2 Turbo in four generation steps, with Diffusers weights, a stock ComfyUI workflow, rolling numbered checkpoints, and an independent Hugging Face demo.
nv-tlabs/kimodo
An NVIDIA kinematic motion diffusion model for generating human and humanoid-robot motion from text plus pose, joint, waypoint, and path constraints, with local inference, a timeline demo, motion exports, official docs, and benchmark materials.
freestylefly/awesome-gpt-image-2
A prompt-as-code library for GPT Image 2 with a categorized case gallery, reusable visual prompt templates, a live browsing site, and an installable agent skill for turning image ideas into structured prompts.
Image and video
Explore projects for creating images, video, motion, avatars, and longer visual sequences.
YouMind-OpenLab/ai-image-prompts-skill
An installable skill for AI assistants that searches a regularly updated library of 10,000+ community image prompts, returns a small set of matching examples with sample images, and can adapt a chosen template to a specific brief.
freestylefly/awesome-gpt-image-2
A prompt-as-code library for GPT Image 2 with a categorized case gallery, reusable visual prompt templates, a live browsing site, and an installable agent skill for turning image ideas into structured prompts.
boogu-project/Boogu-Image
A 10B research model family for text-to-image generation and instruction-based image editing, with Base, Turbo, Edit, and Edit-Turbo checkpoints, bilingual Chinese-English text rendering, local inference code, and public demos.
IGL-HKUST/CoMoVi
A framework for co-generating 3D human motion and realistic videos, with a focus on motion-conditioned video generation and training workflows.
hao-ai-lab/FastVideo/apps/dreamverse
The FastVideo realtime video generation and editing platform, with backend and web UI setup, local GPU, B200, Docker, Modal, readiness checks, and mock-backend workflows.
lllyasviel/Fooocus
A local image-generation interface built around prompt-focused SDXL workflows, with Windows downloads, Colab access, inpainting, outpainting, image prompts, and presets.
krea-ai/krea-2
A Krea AI open-weight text-to-image model family with an undistilled Raw checkpoint for fine-tuning and LoRA training, an eight-step Turbo checkpoint for inference, official local code, Diffusers, SGLang, ComfyUI, hosted routes, and the Krea 2 Community License.
lvladikov/Krea2-Turbo-Distill-4step-LoRA
A community-made distillation LoRA that runs Krea 2 Turbo in four generation steps, with Diffusers weights, a stock ComfyUI workflow, rolling numbered checkpoints, and an independent Hugging Face demo.
bytedance-research/Lance
A ByteDance Research unified multimodal model for image and video understanding, generation, and editing, with model files, demos, inference scripts, Gradio setup, benchmark scripts, and a stated 40GB VRAM inference requirement.
meituan-longcat/LongCat-Video-Avatar-1.5
A Meituan LongCat audio-driven avatar video model for single- and multi-person generation, with audio-text-to-video, audio-image-text-to-video, video continuation, model weights, GitHub quickstart, and project-reported evaluation materials.
NVlabs/LongLive
An NVIDIA Labs infrastructure codebase for long video generation, with LongLive 2.0 NVFP4 and FP8 inference paths, parallel training and inference, multi-shot and image-to-video support, async decoding, LongLive-RAG, model links, docs, and configs.
Lightricks/LTX-2.5
A Lightricks audio-video model family with gated component weights, synchronized sound, text-to-video, image-to-video, multishot generation, local Python and ComfyUI paths, fine-tuning tools, and a separate hosted API.
MiniMaxAI/MiniMax-H3
A MiniMax audio-video generation system with public H3-Base weights for text, image, video, and audio reference workflows, native stereo output, 4-to-15-second clips, local 768p generation, a hybrid 2K path, and a direct hosted API.
lightx2v/Minimax-h3-Turbo
A LightX2V community LoRA project that distills MiniMax H3 into four- and eight-step video-with-audio workflows, with Diffusers and ComfyUI checkpoints, single- and multi-GPU inference code, and a 768p four-step option.
GVCLab/PersonaLive
A portrait image-animation framework for live-streaming-style video generation research, with offline and online inference, pretrained weights, a Web UI, and acceleration notes.
NVlabs/Sana
An NVIDIA Labs codebase for efficient high-resolution image and video generation, with Sana, Sana-1.5, Sana-Sprint, Sana-Video, Sana-WM world-model work, training and inference pipelines, model zoo links, and ComfyUI and diffusers paths.
HKUDS/ViMax
An agentic video-generation framework for turning ideas, scripts, or longer narratives into planned video workflows, with a project-based web interface, interactive agent loop and terminal UI, storyboards, shot planning, previews, render checkpoints, and configurable model providers.
Wan-AI/Wan-Dancer-14B
A music-to-dance video model and framework that uses a reference image, a full music track, and a dance-style prompt, with separate global keyframe planning and local refinement stages for longer outputs.
3D and worlds
Find projects for 3D assets, maps, scenes, world generation, and spatial workflows.
VAST-AI-Research/AniGen
A framework for generating animatable 3D assets from a single image, with mesh, skeleton, and skinning outputs for downstream animation and simulation workflows.
louis-e/arnis
Generates real-world locations inside Minecraft with a surprisingly high level of detail.
nv-tlabs/kimodo
An NVIDIA kinematic motion diffusion model for generating human and humanoid-robot motion from text plus pose, joint, waypoint, and path constraints, with local inference, a timeline demo, motion exports, official docs, and benchmark materials.
robbyant/lingbot-map
A feed-forward 3D foundation model for streaming scene reconstruction, positioned around geometric consistency, long sequences, and efficient real-time inference.
nv-tlabs/lyra
A series of generative 3D world models from NVIDIA, positioned around explorable scenes, 3D consistency, and world-scale generation workflows.
microsoft/TRELLIS.2
A Microsoft 3D generation model for high-fidelity image-to-3D asset creation, using O-Voxel structured latents, PBR materials, inference code, and training tools.
Design and editing
Explore tools for design systems, editable interfaces, background removal, and day-to-day visual editing.
tt-a1i/archify
An agent skill and deterministic Node.js renderer for validated architecture, workflow, sequence, data-flow, and lifecycle maps, with typed JSON source, interactive tracing, and portable HTML, image, and motion exports.
walterlow/freecut
A browser-based multi-track video editor with local workspace files, on-device transcription and captions, scene analysis, local voice and music tools, effects, subtitles, and in-browser export through modern Chromium APIs.
nexu-io/html-anything
A local agentic HTML editor that uses existing coding-agent CLI sessions to turn Markdown, data, and notes into exportable HTML, PNGs, decks, social cards, data reports, and web prototypes with skill templates and sandboxed preview.
nexu-io/open-design
A local-first AI design workspace that connects coding-agent CLIs to prototypes, decks, media outputs, design systems, sandboxed previews, and export workflows.
penpot/penpot
An open-source collaborative design platform with an official MCP server that lets AI agents inspect and modify Penpot files, components, tokens, styles, layouts, and assets through hosted or local connection paths.
nextlevelbuilder/ui-ux-pro-max-skill
A searchable UI and UX knowledge skill for AI coding assistants, with design-system generation, product-specific style and palette guidance, stack-aware implementation notes, and pre-delivery checks.