LogoAwesome Skills
  • Search
  • Category
  • Tag
  • Blog
LogoAwesome Skills
LogoAwesome Skills

Discover Open-Source Agent Skills for AI Coding Assistants

Product

  • Search
  • Category
  • Tag
  • Blog

Resources

  • Claude Skill Docs
  • Antigravity Skills Docs

Tools

  • Claude Code
  • OpenCode
  • Cursor
  • Codex
  • Antigravity

Company

  • Privacy Policy
  • Terms of Service
  • Sitemap

©2026 Awesome Skills. All rights reserved.

Privacy PolicyTerms
Back to Skills

vision-analysis

Analyze, describe, and extract information from images using the MiniMax vision MCP tool. Use when: user shares an image file path or URL (any message containing .jpg, .jpeg, .png, .gif, .webp, .bmp, or .svg file extension) or uses any of these words/phrases near an image: "analyze", "analyse", "describe", "explain", "understand", "look at", "review", "extract text", "OCR", "what is in", "what's in", "read this image", "see this image", "tell me about", "explain this", "interpret this", in conne

12,732stars1,092forksUpdated 6/22/2026
Developer Tools#image-recognition#ocr#mcp-integration#vision-analysis#minimax#ai-vision

Security Assessment

Low Risk(85/100)

Detected risks:

Secret Exposure([SKILL.md] API_KEY=)
Security Score85/100

About vision-analysis

The vision-analysis skill enables AI agents to analyze, describe, and extract information from images using the MiniMax vision capabilities through the MCP (Model Context Protocol) tool integration. It automatically triggers when users share image files or request image analysis, providing intelligent interpretation of visual content including screenshots, diagrams, charts, UI mockups, and photographs. The skill leverages the MiniMax Token Plan's understand_image tool to deliver comprehensive visual understanding.

This skill supports multiple analysis modes tailored to different use cases: general description for overall image understanding, OCR for text extraction from documents and screenshots, UI review for design critique and feedback on mockups and wireframes, chart data extraction for analyzing graphs and visualizations, and object detection for identifying elements within images. Each mode uses optimized prompting strategies to deliver relevant results based on the specific analysis requirement.

Developers, designers, data analysts, and content creators can use this skill to automate image analysis tasks within their AI-assisted workflows. Common applications include extracting text from screenshots for documentation, reviewing UI designs for accessibility and usability issues, analyzing charts to extract numerical data, identifying objects or activities in photos, and generating detailed descriptions of visual content for reports or accessibility purposes. The skill requires a MiniMax Token Plan subscription and proper MCP configuration to function.

FAQ

What image formats does this skill support?

The skill supports common image formats including JPG, JPEG, PNG, GIF, WebP, BMP, and SVG files. It can analyze images provided as local file paths or URLs.

Do I need a MiniMax subscription to use this skill?

Yes, this skill requires an active MiniMax Token Plan subscription with a valid MINIMAX_API_KEY. The understand_image tool cannot be used with free or other tier API keys.

How do I set up the MiniMax MCP integration?

Configuration depends on your environment. For OpenCode, add the MCP configuration to opencode.json. For Claude Code, use the 'claude mcp add' command. For Cursor, add to MCP settings. After configuration, restart your application and verify with the /mcp command. Detailed setup instructions are available at https://platform.minimaxi.com/docs/token-plan/mcp-guide.

What analysis modes are available and when should I use them?

Five modes are available: 'describe' for general image understanding, 'ocr' for text extraction, 'ui-review' for design critique of mockups and wireframes, 'chart-data' for extracting information from graphs and visualizations, and 'object-detect' for identifying and locating elements within images.

How does the skill automatically trigger?

The skill activates automatically when a message contains an image file extension or when users use trigger phrases like 'analyze', 'describe', 'explain', 'extract text', 'OCR', 'what is in', or 'review' in connection with an image, screenshot, diagram, chart, mockup, or photo.

Install vision-analysis

Download and extract the skill files to your .claude/skills/ directory.

Quick Setup:

  1. Copy the skill folder to .claude/skills/
  2. Claude will automatically detect and use the skill

Repository

minimax-ai/skills

Related Skills

content-pattern-analyzer-sms

513

ai-ui-patterns

249

quick-mockups

7,247

clean-architecture

720