pi-vision-bridge-tomoyo
esm
Vision bridge for text-only models (e.g. DeepSeek V4 Flash): images pasted into the chat and screenshots are described by a cheap vision model (MiniMax M3 via opencode.go) and fed to the main model as text. Registers a describe_image tool for on-demand sc
Version 0.1.0 License MIT
Keywords
pi-packagepi-extensionvisionmultimodalimage-analysisdescribe-imagedeepseekscreenshotocr
INSTALL