FoundationVision/UniRefPublic

NotificationsYou must be signed in to change notification settings
Fork15
Star237

[ICCV2023] Segment Every Reference Object in Spatial and Temporal Spaces

License

MIT license

237 stars 15 forks Branches Tags Activity

Star

Notifications

You must be signed in to change notification settings

Branches Tags

Folders and files

Name		Name	Last commit message	Last commit date
Latest commit History 27 Commits
assets		assets
configs		configs
conversion		conversion
datasets		datasets
demo		demo
detectron2		detectron2
dev		dev
docker		docker
docs		docs
external		external
projects		projects
tests		tests
tools		tools
.gitignore		.gitignore
GETTING_STARTED.md		GETTING_STARTED.md
INSTALL.md		INSTALL.md
LICENSE		LICENSE
MODEL_ZOO.md		MODEL_ZOO.md
README.md		README.md
requirements.txt		requirements.txt
setup.cfg		setup.cfg
setup.py		setup.py

Repository files navigation

UniRef++: Segment Every Reference Object in Spatial and Temporal Spaces

Official implementation ofUniRef++, an extended version of ICCV2023UniRef.

Highlights

UniRef/UniRef++ is a unified model for four object segmentation tasks, namely referring image segmentation (RIS), few-shot segmentation (FSS), referring video object segmentation (RVOS) and video object segmentation (VOS).
At the core of UniRef++ is the UniFusion module for injecting various reference information into network. And we implement it using flash attention with high efficiency.
UniFusion could play as the plug-in component for foundation models likeSAM.

Schedule

Add Training Guide
Add Evaluation Guide
Add Data Preparation
Release Model Checkpoints
Release Code

Results

video_demo.mp4

Referring Image Segmentation

Referring Video Object Segmentation

Video Object Segmentation

Zero-shot Video Segmentation & Few-shot Image Segmentation

Model Zoo

Objects365 Pretraining

Model	Checkpoint
R50	model
Swin-L	model

Imge-joint Training

Model	RefCOCO	FSS-1000	Checkpoint
R50	76.3	85.2	model
Swin-L	79.9	87.7	model

Video-joint Training

The results are reported on the validation set.

Model	RefCOCO	FSS-1000	Ref-Youtube-VOS	Ref-DAVIS17	Youtube-VOS18	DAVIS17	LVOS	Checkpoint
UniRef++-R50	75.6	79.1	61.5	63.5	81.9	81.5	60.1	model
UniRef++-Swin-L	79.1	85.4	66.9	67.2	83.2	83.9	67.2	model

Installation

SeeINSTALL.md

Getting Started

Please seeDATA.md for data preparation.

Please seeEVAL.md for evaluation.

Please seeTRAIN.md for training.

Citation

If you find this project useful in your research, please consider cite:

@article{wu2023uniref++,title={UniRef++: Segment Every Reference Object in Spatial and Temporal Spaces},author={Wu, Jiannan and Jiang, Yi and Yan, Bin and Lu, Huchuan and Yuan, Zehuan and Luo, Ping},journal={arXiv preprint arXiv:2312.15715},year={2023}}

@inproceedings{wu2023uniref,title={Segment Every Reference Object in Spatial and Temporal Spaces},author={Wu, Jiannan and Jiang, Yi and Yan, Bin and Lu, Huchuan and Yuan, Zehuan and Luo, Ping},booktitle={Proceedings of the IEEE/CVF International Conference on Computer Vision},pages={2538--2550},year={2023}}

Acknowledgement

The project is based onUNINEXT codebase. We also refer to the repositoriesDetectron2,Deformable DETR,STCN,SAM. Thanks for their awsome works!

About

[ICCV2023] Segment Every Reference Object in Spatial and Temporal Spaces

Releases

No releases published

Packages

No packages published

Movatterモバイル変換

Navigation Menu

Search code, repositories, users, issues, pull requests...

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

License

Folders and files

Latest commit

History

Repository files navigation

UniRef++: Segment Every Reference Object in Spatial and Temporal Spaces

Highlights

Schedule

Results

Referring Image Segmentation

Referring Video Object Segmentation

Video Object Segmentation

Zero-shot Video Segmentation & Few-shot Image Segmentation

Model Zoo

Objects365 Pretraining

Imge-joint Training

Video-joint Training

Installation

Getting Started

Citation

Acknowledgement

About

Topics

Resources

License

Stars

Watchers

Forks

Releases

Packages

Languages

Movatterモバイル変換

License

FoundationVision/UniRef

Folders and files

Latest commit

History

Repository files navigation

UniRef++: Segment Every Reference Object in Spatial and Temporal Spaces

Highlights

Schedule

Results

Referring Image Segmentation

Referring Video Object Segmentation

Video Object Segmentation

Zero-shot Video Segmentation & Few-shot Image Segmentation

Model Zoo

Objects365 Pretraining

Imge-joint Training

Video-joint Training

Installation

Getting Started

Citation

Acknowledgement

About

Topics

Resources

License

Stars

Watchers

Forks

Releases

Packages0

Languages

Packages