AImageLab
Popular repositories Loading
-
dress-code
dress-code PublicDress Code: High-Resolution Multi-Category Virtual Try-On. ECCV 2022
-
meshed-memory-transformer
meshed-memory-transformer Public[CVPR 2020] Meshed-Memory Transformer for Image Captioning
-
multimodal-garment-designer
multimodal-garment-designer Public[ICCV 2023] This is the official repository for the paper "Multimodal Garment Designer: Human-Centric Latent Diffusion Models for Fashion Image Editing"
-
show-control-and-tell
show-control-and-tell Public[CVPR 2019] Show, Control and Tell: A Framework for Generating Controllable and Grounded Captions
-
novelty-detection
novelty-detection PublicLatent space autoregression for novelty detection.
Repositories
- DocAttriBench Public
[BMVC 2026] DocAttriBench: Benchmarking Answer Grounding in Document Visual Question Answering
- CounterVid Public
[EMNLP 2026] Official implementation of CounterVid, a counterfactual video generation framework for mitigating action and temporal hallucinations in video-language models.
- ShieldCLIP Public
Official repository for ShieldCLIP: selective safety alignment for harmful content mitigation in multimodal foundation models. Code, trained models, and dataset access instructions coming soon.
-