Skip to main content

Vision tasks

Adapted for this playbook from the 🤗 Transformers documentation by Hugging Face. Official page: https://huggingface.co/docs/transformers/tasks/image_classification. Images © Hugging Face (documentation-images) unless noted. This is not a substitute for the upstream docs — verify against the current version.

Covers Hugging Face pages: tasks/image_classification, semantic_segmentation, object_detection, instance_segmentation, video_classification, zero_shot_object_detection

TaskOfficial guide
Image classificationimage_classification
Semantic segmentationsemantic_segmentation
Object detectionobject_detection
Instance segmentationinstance_segmentation
Video classificationvideo_classification
Zero-shot object detectionzero_shot_object_detection

ViT-related / CLIP imagery — Source: Hugging Face

Source: Hugging Face documentation images.

See flagships ViT and CLIP.

Discussion

Comments​

Share feedback or questions about this page. No account required.

Loading comments…