modlens

技能包模型与接入中文文档MIT
Give a text-only model sight, and just paste the image.

功能特性

  • Evidence, not imagination. Full transcription, reading-order layout regions, entity and relation lists. The model quotes specifics.
  • Install once, use everywhere. Verified on real machines in Claude Code, Codex, Pi, and OpenCode.
  • Open an issue. Bugs, suggestions, confusing errors, unclear docs. Issues are read and shape what gets built next.
  • Fork it. Under MIT your copy is fully yours to modify and publish.

安装命令

dsh plugin add github:liustack/modlens

项目预览

项目简介

The first vision plugin for DeepSeek Harness, and the vision bridge for every text-only coding agent. Paste an image, get structured JSON evidence (OCR, layout, semantics). | 全网第一个 DeepSeek Harness 视觉插件,为 DeepSeek、GLM 等纯文本模型外挂视觉能力,粘贴图片即得结构化 JSON 证据(OCR、版面、语义)。

话题标签

#agent-skills#claude-code#claude-skills#codex#cordis#deepseek#dsh#dsh-plugin#glm#harness#harness-engineering#hermes-agent
← 返回 技能包 列表