npm package discovery and stats viewer.

Discover Tips

  • General search

    [free text search, go nuts!]

  • Package details

    pkg:[package-name]

  • User packages

    @[username]

Sponsor

Optimize Toolset

I’ve always been into building performant and accessible sites, but lately I’ve been taking it extremely seriously. So much so that I’ve been building a tool to help me optimize and monitor the sites that I build to make sure that I’m making an attempt to offer the best experience to those who visit them. If you’re into performant, accessible and SEO friendly sites, you might like it too! You can check it out at Optimize Toolset.

About

Hi, 👋, I’m Ryan Hefner  and I built this site for me, and you! The goal of this site was to provide an easy way for me to check the stats on my npm packages, both for prioritizing issues and updates, and to give me a little kick in the pants to keep up on stuff.

As I was building it, I realized that I was actually using the tool to build the tool, and figured I might as well put this out there and hopefully others will find it to be a fast and useful way to search and browse npm packages as I have.

If you’re interested in other things I’m working on, follow me on Twitter or check out the open source projects I’ve been publishing on GitHub.

I am also working on a Twitter bot for this site to tweet the most popular, newest, random packages from npm. Please follow that account now and it will start sending out packages soon–ish.

Open Software & Tools

This site wouldn’t be possible without the immense generosity and tireless efforts from the people who make contributions to the world and share their work via open source initiatives. Thank you 🙏

© 2026 – Pkg Stats / Ryan Hefner

@showlotus/opencode-image-vision

v1.0.9

Published

OpenCode plugin that adds vision capabilities to text-only LLMs by analyzing pasted images via vision AI providers

Readme

opencode-image-vision

npm version License: MIT Node

OpenCode plugin that gives text-only models (GLM-5, DeepSeek V4, MiniMax, etc.) the ability to understand pasted images. Images are analyzed by a vision model in the background and replaced with text descriptions before the chat model runs — paste → ask → done.

Requires OpenCode 1.14+, Node 18+, and a signed-in vision provider (e.g. glm-4.6v).


Install

Add to ~/.config/opencode/opencode.json (or opencode.jsonc):

{
  "plugin": [
    [
      "@showlotus/opencode-image-vision@latest",
      {
        "model": "zhipuai-coding-plan/glm-4.6v"
      }
    ]
  ]
}

model is required (providerId/modelId). API keys are read from OpenCode's auth.json — no extra key setup.

For local development, use a file:// absolute path (symlinks in plugins/ are skipped):

["file:///Users/YOU/path/to/opencode-image-vision", { "model": "zhipuai-coding-plan/glm-4.6v" }]

Restart OpenCode, paste an image, and ask about it.


Options

| Option | Required | Default | Description | | --------- | -------- | -------- | ----------- | | model | Yes | — | Vision model, e.g. zhipuai-coding-plan/glm-4.6v | | prompt | No | built-in | Analysis prompt | | timeout | No | 120000 | Base timeout for vision API (ms); actual timeout scales with image size up to 300s | | debug | No | false | Log to <tmpdir>/iv-debug.log |

Debug via env: IMAGE_VISION_DEBUG=1 (optional IMAGE_VISION_DEBUG_PATH).

<tmpdir> is the OS temp directory (node:os tmpdir()), which varies by platform:

| OS | Temp dir | Env vars checked by Node (in order) | | ------- | -------------------- | -------------------------------------------------- | | macOS | /var/folders/.../T | $TMPDIR, fallback /tmp | | Linux | /tmp | $TMPDIR, fallback /tmp | | Windows | %LOCALAPPDATA%\Temp| %TEMP%, %TMP%, %LOCALAPPDATA%\Temp, C:\Windows\Temp |

Find the actual path on any system with: node -e "console.log(require('os').tmpdir())".


How it works

The plugin hooks into 4 stages of OpenCode's message lifecycle:

  1. chat.message — Fires when a user sends a message. Detects image parts and sets a flag to trigger tool injection.
  2. chat.params — Injects tool_choice as a fallback hint (actual tool invocation is driven by the transform hook's text instruction).
  3. experimental.chat.messages.transform — Saves each image to a temp file (<tmpdir>/iv-images/<hash>.<ext>) and replaces the image part with a text instruction containing the file path. The model reads the path and calls the tool on its own — no reliance on forced tool injection.
  4. analyze_image tool — Accepts a file_path parameter (SDK schema). Reads the image from disk, runs a child session via the OpenCode SDK against the vision model, and returns the description as tool output. Temp files are preserved for re-analysis in follow-up turns.

The active model automatically skips image processing if it already supports vision. Failed images produce [Analysis failed: reason] and do not block others. Identical images are cached by MD5 hash — cache hits replace the result directly without triggering the tool.


Troubleshooting

  • Plugin not loading — Confirm model is set; check logs for [image-vision] init failed; run opencode auth
  • Timeout on large images — Increase "timeout": 120000 (or higher)
  • No API key — Sign in to the vision provider in OpenCode (opencode auth)
  • Local dev — Use file:///absolute/path, not symlinks; set "debug": true and check <tmpdir>/iv-debug.log

License

MIT — see LICENSE.