glmv-prompt-gen
๐ฏSkillfrom zai-org/glm-skills
Part of the official GLM skills collection for multimodal AI models, this skill generates AI art prompts from visual references for tools like Midjourney, Stable Diffusion, and DALL-E across multiple agent platforms.
Same repository
zai-org/glm-skills(17 items)
Installation
npx vibeindex add zai-org/glm-skills --skill glmv-prompt-gennpx skills add zai-org/glm-skills --skill glmv-prompt-gen~/.claude/skills/glmv-prompt-gen/SKILL.mdSKILL.md
More from this repository10
Official skills for GLM (Zhipu AI) models, providing multimodal capabilities including image/video captioning, document-based writing, object localization, PDF-to-presentation conversion, and resume screening. Compatible with Claude Code, OpenCode, OpenClaw, and other AI coding agents.
A skill from the GLM model family that performs general text extraction from images and PDFs, part of the GLM-OCR suite designed for AI coding agents including Claude Code and OpenCode.
A skill from the GLM model family that leverages GLM-V multimodal capabilities to analyze stock charts, financial documents, and market data for investment research and analysis.
A GLM multimodal skill that converts PDF documents into multi-slide HTML presentations, part of the official GLM skills collection designed for AI coding agents including Claude Code and OpenClaw.
Official skills for the GLM family of models, designed for multiple agent architectures including Claude Code and OpenClaw. Provides multimodal skills for image captioning, document-based writing, object localization, PDF conversion, and resume screening.
An official GLM-OCR skill that extracts mathematical formulas from images and documents into LaTeX format, part of the GLM skills collection designed for agent architectures including Claude Code, OpenCode, and other AI coding agents.
Official skills for the GLM family of models, providing multimodal capabilities like image captioning, object localization, PDF-to-presentation conversion, resume screening, and stock analysis for AI coding agents including Claude Code and OpenCode.
Official skill collection for the GLM family of models, covering multimodal vision (captioning, object grounding, document-to-presentation), OCR (text, formulas, handwriting, tables), and image generation. Compatible with Claude Code, OpenCode, OpenClaw, and other AI coding agents.
An official GLM skill that uses multimodal (GLM-V) capabilities to create frontend visual replicas of existing websites, designed for AI coding agents including Claude Code and OpenCode.
A multimodal captioning skill from GLM Skills that generates captions and descriptions for images, videos, and documents using GLM-V models. Part of a unified skill collection designed for Claude Code, OpenCode, OpenClaw, and other AI coding agents.