#python#apple_silicon#florence2#idefics#llava#llm#local_ai#mlx#molmo#paligemma#pixtral#vision_framework#vision_language_model#vision_transformer
MLX-VLM lets you run, chat with, and fine-tune Vision Language Models (VLMs) plus audio/video models on your Mac using MLX—install easily with `pip install -U mlx-vlm`. Use CLI for quick text/image/audio generation (e.g., `mlx_vlm.generate --model ... --image photo.jpg`), Gradio UI for chats, Python scripts, or a FastAPI server with OpenAI-compatible endpoints supporting multi-images/videos. Features like TurboQuant cut KV cache memory by 76%, and LoRA/QLoRA fine-tuning works on consumer hardware. You benefit by experimenting with powerful multimodal AI locally—fast, memory-efficient, no cloud costs, perfect for Mac users tweaking models affordably.
https://github.com/Blaizzy/mlx-vlm
⚡️ Advanced Camera Control is now available for #Gen3 Alpha Turbo. Choose both the direction and intensity of how you move through your scenes for even more intention in every shot
VideoGenerator | SUBSCRIBE
⚡️ Camera controls have appeared in Runway #Gen3 alpha!
No official announcement yet, but it looks like they’re rolling it out gradually. 😍
Credits: Pierrick Chevallier
VideoGenerator | SUBSCRIBE
"The Schnitzel Dilemma"
A short film about two colleagues planning their lunch date. Generated with runwayml #Gen3 new Act-One model 😍
VideoGenerator | SUBSCRIBE