Vision
3 articles
Articles
21 June 2026
MiniCPM-V 4.5: specs, benchmarks, and how to run it on Lyceum
MiniCPM-V 4.5 scores 77.0 on OpenCompass in an efficient 8B package. With its novel 3D-Resampler, it compresses video tokens by 96x, making long-video understanding highly cost-effective.
2 June 2026
Run Vision Language Models on GPU Cloud: VRAM & Setup Guide
Vision language models consume massive VRAM for image tokens. Learn the exact hardware requirements and deployment strategies for production VLMs.
31 May 2026
Multimodal AI Inference on European GPUs: Compliance and Cost Optimization
Running multimodal AI inference at scale exposes the structural flaws of hyperscaler pricing and compliance models. Engineering teams require infrastructure that provides high throughput for complex data types while maintaining strict data residency.