-
SimpleAR
Pytorch implementation for the paper titled "SimpleAR: Pushing the Frontier of Autoregressive Visual Generation"
Python ★ 431 1y agoExplain → -
M2TR-Multi-modal-Multi-scale-Transformers-for-Deepfake-Detection
No description.
Python ★ 122 4y agoExplain → -
PyDeepFakeDet
PyDeepFakeDet is an integrated and scalable tool for Deepfake detection.
Python ★ 114 3y agoExplain → -
RepWAM
Code for RepWAM: World Action Modeling with Representation Visual-Action Tokenizers
★ 57 1mo agoExplain → -
OmniVid
No description.
Python ★ 57 2y agoExplain → -
STTS
Official PyTorch implementation of the ECCV 2022 paper: Efficient Video Transformers with Spatial-Temporal Token Selection.
Python ★ 52 4y agoExplain → -
ARM
ARM: An AutoRegressive Large Multimodal Model with Discrete Representations
★ 50 1mo agoExplain → -
Objectformer
No description.
Jupyter Notebook ★ 35 2y agoExplain → -
OpenTokenizer
No description.
Python ★ 21 1y agoExplain → -
vllm
No description.
Python ★ 2 1y agoExplain → -
Video-Swin-Transformer ⑂
This is an official implementation for "Video Swin Transformers".
★ 1 4y agoExplain → -
wdrink.github.io
Wang Junke' s personal website.
HTML ★ 0 1mo agoExplain → -
LAVIS ⑂
LAVIS - A One-stop Library for Language-Vision Intelligence
★ 0 3y agoExplain → -
Ask-Anything ⑂
ChatGPT with video understanding! And many more supported LMs such as miniGPT4, StableLM, and MOSS.
★ 0 3y agoExplain → -
audioset-classification ⑂
Train a Deep Learning model to classify audio embeddings on IBM's Deep Learning as a Service (DLaaS) platform - Watson Machine Learning
★ 0 3y agoExplain → -
pytorchvideo ⑂
A deep learning library for video understanding research.
★ 0 4y agoExplain →
No repos match these filters.