dsh-vision-toolkit
★ 648
让纯文本模型更好地做视觉任务:带意图的图片问答、长截图 OCR、UI 还原等。
Vision tasks for text-only models: intent-aware image Q&A, long-screenshot OCR, UI reproduction, grounding, and pixel diff.
第三方插件以你本人的权限运行,精选收录不等于安全审查,安装前看一眼源码。
让纯文本模型更好地做视觉任务:带意图的图片问答、长截图 OCR、UI 还原等。
Vision tasks for text-only models: intent-aware image Q&A, long-screenshot OCR, UI reproduction, grounding, and pixel diff.
第三方插件以你本人的权限运行,精选收录不等于安全审查,安装前看一眼源码。