Skip to content

Add DeepSight: device-native vision for text-only LLMs - #28

Open
eliaseffects wants to merge 1 commit into
ranpox:mainfrom
eliaseffects:add-deepsight
Open

eliaseffects wants to merge 1 commit into
ranpox:mainfrom
eliaseffects:add-deepsight

Conversation

@eliaseffects

Copy link
Copy Markdown

DeepSight gives text-only LLMs eyes and hands: look/OCR/zoom/locate on images, live screen capture, and desktop automation (click, type, key, scroll, open, focus). Vision is fully on-device: Apple Vision on macOS (zero tokens, zero GPU), PIL + optional Tesseract on Windows. No image data ever leaves the machine. MIT.

Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant