1-day current streak·4-day longest streak
-
ComfyUI-fal-API ★ PINNED
Custom nodes for using fal API.
Python ★ 199 2h agoExplain → -
awesome-vlm-architectures ★ PINNED
Famous Vision Language Models and Their Architectures
Markdown ★ 1.3k 6mo agoExplain → -
ComfyUI_VLM_nodes ★ PINNED
Custom ComfyUI nodes for Vision Language Models, Large Language Models, Image to Music, Text to Music, Consistent and Random Creative Prompt Generation
Python ★ 576 6mo agoExplain → -
ComfyUI-Dream-Interpreter ★ PINNED
Dream Interpreter inside ComfyUI
JavaScript ★ 84 6mo agoExplain → -
ComfyUI-Depth-Visualization ★ PINNED
Depth map applied Image viewer inside ComfyUI
JavaScript ★ 68 6mo agoExplain → -
ComfyUI-Texture-Simple ★ PINNED
Visualize your textures inside ComfyUI
JavaScript ★ 57 6mo agoExplain → -
Tile-Upscaler
Image Upscaler with Tile Controlnet Fully Integrated in Huggingface Diffusers
Python ★ 20 6mo agoExplain → -
dualview
Open source side-by-side comparison tool for images, videos, audio, 3D models and documents. GPU-accelerated with 50+ analysis modes, 100+ transitions, and professional export options.
TypeScript ★ 17 6mo agoExplain → -
ComfyUI-Fluxpromptenhancer ⑂
A Prompt Enhancer for flux.1 in ComfyUI
Python ★ 12 6mo agoExplain → -
cuda-course-remotion
No description.
TypeScript ★ 10 4mo agoExplain → -
gpt4o-yellow-tint-cleaner
Cleaning yellow tint that occurs on gpt4o image generation
Python ★ 5 1y agoExplain → -
vision-embeddings
No description.
Python ★ 4 4mo agoExplain → -
MATHESIS
The Mathematical Path to Consciousness — 43 interactive chapters, 100+ exercises, built with Claude 4.6 Opus & Nano Banana Pro
JavaScript ★ 3 5mo agoExplain → -
claude-insights
No description.
Shell ★ 2 6mo agoExplain → -
HunyuanDiT ⑂
Hunyuan-DiT : A Powerful Multi-Resolution Diffusion Transformer with Fine-Grained Chinese Understanding
★ 2 2y agoExplain → -
ControlNeXt ⑂
Controllable video and image Generation, SVD, Animate Anyone, ControlNet, ControlNeXt, LoRA
Python ★ 2 1y agoExplain → -
SimpleSlowVideoUpscaler ⑂
Simple video upscaler extending on the tile upscaler from https://huggingface.co/spaces/gokaygokay/Tile-Upscaler
Python ★ 2 2y agoExplain → -
kimi-micro
No description.
JavaScript ★ 1 10d agoExplain → -
micro-ui
No description.
JavaScript ★ 1 10d agoExplain → -
perigpt
No description.
Python ★ 1 4mo agoExplain → -
autopinn
No description.
Python ★ 1 4mo agoExplain → -
lectures ⑂
Material for cuda-mode lectures
Jupyter Notebook ★ 1 1y agoExplain → -
trellis-wheels
No description.
★ 1 1y agoExplain → -
graph_websearch_agent ⑂
Websearch agent built on the LangGraph framework
★ 1 2y agoExplain → -
ComfyUI-fal-Connector ⑂
No description.
Python ★ 1 2y agoExplain → -
dspy-ollama-colab
dspy with ollama and llamacpp on google colab
Jupyter Notebook ★ 1 2y agoExplain → -
Video-LLaMA ⑂
[EMNLP 2023 Demo] Video-LLaMA: An Instruction-tuned Audio-Visual Language Model for Video Understanding
★ 1 2y agoExplain → -
ai-course-creator-skill
No description.
Python ★ 0 15h agoExplain → -
shader-skills
No description.
Python ★ 0 1mo agoExplain → -
HuggingFace-Space-Healer
No description.
Python ★ 0 2mo agoExplain → -
nanoPINN
No description.
Python ★ 0 4mo agoExplain → -
CogVLM2 ⑂
第二代 CogVLM多模态预训练对话模型
★ 0 2y agoExplain → -
ComfyUI_PuLID_Flux_ll-attn ⑂
No description.
Python ★ 0 1y agoExplain → -
TripoSG ⑂
TripoSG: High-Fidelity 3D Shape Synthesis using Large-Scale Rectified Flow Models
★ 0 1y agoExplain → -
diffusion-self-distillation ⑂
No description.
Python ★ 0 1y agoExplain → -
ComfyUI-PuLID-Flux-Enhanced ⑂
No description.
★ 0 1y agoExplain → -
LatentSync ⑂
Taming Stable Diffusion for Lip Sync!
Python ★ 0 1y agoExplain → -
huggingface.js ⑂
Utilities to use the Hugging Face Hub API
TypeScript ★ 0 1y agoExplain → -
sane-controlnet-aux ⑂
A fork of ComfyUI's ControlNet auxiliary preprocessors for use outside of ComfyUI.
Python ★ 0 1y agoExplain → -
Depth-Anything-V2 ⑂
Depth Anything V2. A More Capable Foundation Model for Monocular Depth Estimation
Python ★ 0 1y agoExplain → -
inference ⑂
A fast, easy-to-use, production-ready inference server for computer vision supporting deployment of many popular model architectures and fine-tuned models.
Python ★ 0 1y agoExplain → -
torchscale ⑂
Foundation Architecture for (M)LLMs
Python ★ 0 1y agoExplain → -
BLIP ⑂
PyTorch code for BLIP: Bootstrapping Language-Image Pre-training for Unified Vision-Language Understanding and Generation
Jupyter Notebook ★ 0 1y agoExplain → -
img2txt-comfyui-nodes ⑂
Implements some of the most popular img2txt models on HF into ComfyUI nodes. Uses questions/conditional-prompts to get descriptions that are suited for being fed back into a txt2img node.
★ 0 2y agoExplain → -
Vitron ⑂
A Unified Pixel-level Vision LLM for Understanding, Generating, Segmenting, Editing
★ 0 2y agoExplain → -
flash-attention-minimal ⑂
Flash Attention in ~100 lines of CUDA (forward pass only)
★ 0 2y agoExplain → -
Reka-Torch ⑂
Implementation of the model: "Reka Core, Flash, and Edge: A Series of Powerful Multimodal Language Models" in PyTorch
★ 0 2y agoExplain → -
siglip ⑂
Projects based on SigLIP (Zhai et. al, 2023) and Hugging Face transformers integration 🤗
★ 0 2y agoExplain → -
awesome ⑂
😎 Awesome lists about all kinds of interesting topics
★ 0 2y agoExplain → -
DeepSeek-VL ⑂
DeepSeek-VL: Towards Real-World Vision-Language Understanding
★ 0 2y agoExplain →
No repos match these filters.