Download
1 variant available
fp16 SafeTensor
ViT-L-14-GmP-SAE-TE-only.safetensors
Half precision, best balance (pruned) • 308.43 MB
Verified: a year ago
480 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9
(3)
Mar 20, 2025
Other
CLIP-GmP-ViT-L-14 Text - for improved detailed in images containing text.

7.4K0 1 2 3 4 5 6 7 8 9.0 1 2 3 4 5 6 7 8 9K
4760 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9
61.6K0 1 2 3 4 5 6 7 8 90 1 2 3 4 5 6 7 8 9.0 1 2 3 4 5 6 7 8 9K

License:
By expanding the token length, Long-CLIP can process longer text inputs more effectively, capturing more context and details. This is particularly useful for generating images from detailed descriptions, as it allows the model to consider a broader range of information, resulting in higher-quality outputs.
.jpeg)
