azure-ai-vision
Expert guidance for Azure AI Vision: image analysis, OCR containers, smart-cropping, and video frame processing.
Explore AI agent skills related to 電腦視覺. Browse installable Claude Code and automation skills on Mentalok Skills Hub.
Discover reusable agent skills, browse implementation details, and find the right skill for your workflow.
5 skills found
Expert guidance for Azure AI Vision: image analysis, OCR containers, smart-cropping, and video frame processing.
Search, download, and analyze arXiv academic papers using a hybrid approach of API queries, ar5iv HTML parsing, and browser-based Actionbook automation.
Robot perception system design, configuration, and optimization for cameras, LiDAR, and sensor fusion pipelines. Includes camera calibration, 3D reconstruction, and production deployment best practices.
Find, review, and remove duplicate or near-duplicate images in FiftyOne datasets using computer vision similarity embeddings.
Process and generate multimedia with Google Gemini. Analyze audio, images, videos, and PDFs with high-context windows. Supports transcription, visual QA, OCR, and AI-driven image creation.