MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.
-
Updated
Oct 3, 2026 - Python
MLX-VLM is a package for inference and fine-tuning of Vision Language Models (VLMs) on your Mac using MLX.
Official PyTorch implementation of VisionHOPE: Visual Backbones as Self-Modifying Learning Systems.
Tool for automating common video key-frame extraction, video compression and Image Auto-crop/Image-resize tasks
[CVPR'25 Highlight] The official implementation of "GG-SSMs: Graph-Generating State Space Models"
OpenNCC Kit
Real-time webcam demo with SmolVLM(mlx-community/SmolVLM-Instruct-4bit) and MLX-VLM
Vision framework which brings a more robust, Deep Learning-based approach to some usual OpenCV use cases
Python scripts for use with Turi Create to output a Xcode compatible mlmodel file for use with machine learning object detection with the CoreML or Vision frameworks.
Research the flying-car vision high-tech
On-device Perceive → Reason pipeline for Apple Silicon: Core ML + Vision for perception, a swappable LanguageModel (Apple Foundation Models or Claude) for reasoning. Python conversion/quantization toolkit plus a SwiftUI reference app.
EVE Online 로컬 창을 감시하다 낯선 사람이 들어오면 알리는 macOS 도구 / macOS Local-window alarm for EVE Online
Turn a scrolling chat screen recording into a speaker-separated, chronologically ordered transcript. macOS-native (AVFoundation + Vision), no ffmpeg or dependencies.
Mathematical Foundations of Recursive Cortical Ignition: A Conditional Theory of Sub-Threshold Pattern Coding in Recurrent Visual Cortex
Python wrapper around the Native OCR engines that ship with macOS and Windows
Mac 版微信聊天记录导出工具:读屏 + 自动滚动截图 + 本地 OCR,把当前会话导出成 Word / Markdown,带图片和聊天里的文件。不解密、不读数据库、不联网。
TextLift OCR text extractor
On-device OCR for macOS in 134 lines of Swift over the Vision framework
To associate your repository with the vision-framework topic, visit your repo's landing page and select "manage topics."