Fourier-Qwen2.5-VL-3B-0.67

Official checkpoints for Fourier Compressor: Frequency-Domain Visual Token Compression for Vision-Language Models.

Model Details

Model Base Model Visual Tokens Compression Weights
Fourier-LLaVA-v1.5-7B-256 LLaVA-v1.5-7B 256 55.6% πŸ€— HF
Fourier-LLaVA-v1.5-7B-144 LLaVA-v1.5-7B 144 75.0% πŸ€— HF
Fourier-LLaVA-v1.5-7B-64 LLaVA-v1.5-7B 64 88.9% πŸ€— HF
Fourier-LLaVA-v1.5-7B-36 LLaVA-v1.5-7B 36 93.8% πŸ€— HF
Fourier-LLaVA-v1.5-13B-144 LLaVA-v1.5-13B 144 75.0% πŸ€— HF
Fourier-Qwen2-VL-2B-0.67 Qwen2-VL-2B-Instruct Dynamic 55.6% πŸ€— HF
Fourier-Qwen2.5-VL-3B-0.67 Qwen2.5-VL-3B-Instruct Dynamic 55.6% πŸ€— HF

Links

Downloads last month
15
Safetensors
Model size
4B params
Tensor type
BF16
Β·
Inference Providers NEW
This model isn't deployed by any Inference Provider. πŸ™‹ Ask for provider support

Model tree for whyisverysmart/Fourier-Qwen2.5-VL-3B-0.67

Finetuned
(828)
this model
Quantizations
2 models

Collection including whyisverysmart/Fourier-Qwen2.5-VL-3B-0.67

Paper for whyisverysmart/Fourier-Qwen2.5-VL-3B-0.67