242,986 packages matching “vision-language-model”
moondream
v0.2.0 · 10 months ago
Official Node.js client for Moondream, a fast and efficient vision language model.
No known vulnerabilities
@linxin666/dsh-tool-describe-image
v0.3.24 · 14 days ago
Model-facing describe_image tool for the dsh web GUI: gives a text-only model image understanding by asking a vision-language model at an OpenAI-compatible endpoint to describe one image (local path, http(s) URL, or attachment reference). Hot-pluggable —
No known vulnerabilities
dsh-deepseek-vision
v0.1.7 · 1 month ago
Out-of-tree dsh provider plugin: a DeepSeek gateway route that claims image input and transparently describes pasted images through a configured vision-language model (e.g. Qwen-VL) before the text-only DeepSeek wire sees them.
No known vulnerabilities
@phamkhachoabk/dsh-local-vision-describe
v0.1.1 · 17 days ago
Local Vision image descriptions for DeepSeek Harness: a vision-language model on your own machine tells a text-only model what each image shows, beyond the text OCR can read
No known vulnerabilities
@phamkhachoabk/dsh-vision-describe-mlx
v0.1.0 · 18 days ago
Local vision-description provider for DeepSeek Harness: a persistent MLX vision-language model on Apple Silicon, so a text-only model can be told what an image shows
No known vulnerabilities
agent-eyes-mcp
v0.2.0 · 1 month ago
Give text-only LLM agents eyes: an MCP server + CLI that describes images through a vision-language model (VLM) API.
No known vulnerabilities
@neystan/dsh-tool-describe-image
v0.1.16 · 1 month ago
Model-facing describe_image tool for the dsh web GUI: gives a text-only model image understanding by asking a vision-language model at an OpenAI-compatible endpoint to describe one image (local path, http(s) URL, or attachment reference). Hot-pluggable —
No known vulnerabilities
@cohesiumai/modules-vlm
v2.1.2 · 8 months ago
Local Vision-Language Model module for browser-ai - image understanding
No known vulnerabilities
@ai-sdk/provider
v4.0.21 · 4 days ago
No description provided.
No known vulnerabilities
@gestaltrun/dsh-tool-describe-image
v0.3.21-gestaltrun.0 · 22 days ago
Model-facing describe_image tool for the dsh web GUI: gives a text-only model image understanding by asking a vision-language model at an OpenAI-compatible endpoint to describe one image (local path, http(s) URL, or attachment reference). Hot-pluggable —
No known vulnerabilities
@frankledo/receiptnamer
v1.0.0 · 1 month ago
Name scanned receipt PDFs by reading the rendered image with a vision-language model (local Ollama or Anthropic API)
No known vulnerabilities
@mediapipe/tasks-vision
v1.0.1 · 2 months ago
MediaPipe Vision Tasks. See the privacy notice at https://goo.gle/mediapipe-privacy.
No known vulnerabilities
@qvac/vla-ggml
v0.29.1 · 4 days ago
VLA vision-language-action inference addon for QVAC (ggml backend)
No known vulnerabilities
@linxin666/dsh-client-ui-model-capabilities
v0.4.5 · 4 hours ago
Per-model capability declarations (image input and reasoning efforts) for custom pi-ai providers, edited in place on the Models settings cards; writes the official llm-pi-ai settings namespace, no DSH source changes.
No known vulnerabilities
@zeke-02/moondream
v0.2.1 · 8 months ago
Official Node.js client for Moondream, a fast and efficient vision language model.
No known vulnerabilities
@google-cloud/vision
v6.1.1 · 6 days ago
Google Cloud Vision API client for Node.js
No known vulnerabilities
react-native-vision-camera-ocr-plus
v2.0.6 · 1 month ago
React Native Vision Camera plugin for on-device text recognition (OCR) and translation using ML Kit. Maintained fork of react-native-vision-camera-text-recognition
No known vulnerabilities
@sanity/vision
v6.17.0 · 5 days ago
Sanity plugin for running/debugging GROQ-queries against Sanity datasets
No known vulnerabilities
react-native-vision-camera
v5.2.3 · 1 month ago
VisionCamera is the fastest and most powerful Camera for react-native.
No known vulnerabilities
dsh-tool-vision
v0.10.1 · 1 day ago
DeepSeek Harness 外置视觉模型插件:inspect_image 把本地图片或 http(s) 图片 URL 发给任意 OpenAI 兼容端点,视觉模型看图的文字回答直接带回对话;附 Web UI 设置栏。
No known vulnerabilities