BigHugger
sk Skill · junmo-kim

macvis

On-device, zero-token vision for AI agents (mac-local-vision, macOS). Read text from an image / screenshot / PDF (OCR), find the exact click-pixel of a word for E2E/UI assertions, scan QR codes and barcodes, tag/classify an image, group photos by face, flatten a photographed document, extract a document's structured layout, or ask a question about an image. Apple Vision + Foundation Models, fully on-device — no…

installs 8w
0
30-day movement
starts with the next reading
Related entries
1
Connections
0
bashSwift
Host repository
junmo-kim/mac-local-vision
Invocable by
user
Host stars
60
Host language
Swift