OWL-ViT: Open-World Object Detection with Vision Transformers
OWL-ViT is an open-vocabulary object detector.
Github: https://github.com/google-research/scenic/tree/main/scenic/projects/owl_vit
Paper: https://arxiv.org/abs/2205.06230
Colab: https://colab.research.google.com/github/google-research/scenic/blob/main/scenic/projects/owl_vit/notebooks/OWL_ViT_minimal_example.ipynb
Dataset: https://paperswithcode.com/dataset/objects365
👉 @bigdata_1
Post #677
692

- 👍 1