Image Feature Extraction
Transformers
OpenCLIP
How to use from the
Use from the
Transformers library
# Use a pipeline as a high-level helper
from transformers import pipeline

pipe = pipeline("image-feature-extraction", model="UCSC-VLAA/openvision-vit-huge-patch14-84")
# Load model directly
from transformers import AutoModel
model = AutoModel.from_pretrained("UCSC-VLAA/openvision-vit-huge-patch14-84", device_map="auto")
Quick Links

This repository contains the model weights based on the work described in OpenVision: A Fully-Open, Cost-Effective Family of Advanced Vision Encoders for Multimodal Learning.

Project Page: https://ucsc-vlaa.github.io/OpenVision/

For details on training and usage, please refer to the Github repository.

Downloads last month

-

Downloads are not tracked for this model. How to track
Inference Providers NEW
This model isn't deployed by any Inference Provider. 🙋 Ask for provider support

Collection including UCSC-VLAA/openvision-vit-huge-patch14-84

Paper for UCSC-VLAA/openvision-vit-huge-patch14-84