5-day longest streak
-
SenseNova-U1 ★ PINNED ⑂
SenseNova-U series: Native Unified Paradigm with NEO-Unify from the First Principles
★ 2 3mo agoExplain → -
lmms-eval ★ PINNED ⑂
One-for-All Multimodal Evaluation Toolkit Across Text, Image, Video, and Audio Tasks
Python ★ 0 3mo agoExplain → -
lmms-engine ★ PINNED ⑂
A simple, unified multimodal models training engine. Lean, flexible, and built for hacking at scale.
Python ★ 0 4mo agoExplain → -
Awesome-Unified-Multimodal
📖 This is a repository for organizing papers, codes, and other resources related to unified multimodal models.
★ 365 7mo agoExplain → -
Awesome-LVLM-Hallucination
No description.
★ 56 1y agoExplain → -
Purshow_Notes
No description.
TeX ★ 30 4d agoExplain → -
MoE_Notes
No description.
TeX ★ 13 7d agoExplain → -
BDetCLIP
[ICML 2025] Implementation of "Test-Time Multimodal Backdoor Detection by Contrastive Prompting"
Python ★ 8 1y agoExplain → -
Purshow.github.io ⑂
[🔥 Yuwei Niu's Academic Personal Homepage]
HTML ★ 6 3d agoExplain → -
Purshow-0.github.io ⑂
Yuwei Niu's Academic Personal Homepage
★ 1 1y agoExplain → -
Echo-Memory ⑂
Official repo for paper "Echo-Memory: A Controlled Study of Memory in Action World Models"
★ 0 2mo agoExplain → -
Bagel ⑂
Open-source unified multimodal model
★ 0 3mo agoExplain → -
NEO ⑂
NEO Series: Native Vision-Language Models from First Principles
★ 0 4mo agoExplain → -
UniWorld ⑂
UniWorld: High-Resolution Semantic Encoders for Unified Visual Understanding and Generation
★ 0 7mo agoExplain → -
thinking-in-space ⑂
Official repo and evaluation implementation of VSI-Bench
★ 0 1y agoExplain → -
visionbook ⑂
<Foundations of Computer Vision> Book
★ 0 10mo agoExplain → -
BDetCLIP_badCLIP
No description.
Python ★ 0 1y agoExplain →
No repos match these filters.