glmocr
Extract text from images using GLM-OCR API. Supports images and PDFs with high accuracy OCR, table recognition, formula extraction, and handwriting recognition. Use this skill whenever the user wants …
Browse curated skills with source links, package snapshots, README assets and install signals in one calm, searchable catalog.
Extract text from images using GLM-OCR API. Supports images and PDFs with high accuracy OCR, table recognition, formula extraction, and handwriting recognition. Use this skill whenever the user wants …
Official skill for recognizing and extracting tables from images and PDFs into Markdown format using ZhiPu GLM-OCR API. Supports complex tables, merged cells, and multi-page documents. Use this skill …
Autonomous novel generation system. Creates complete books from premise to final draft with structured workflow including character creation, outlining, chapter planning, drafting, and editing. Use wh…
抖音上升热点选题助手适合内容创作者、运营、电商、营销在用户想知道接下来拍什么、写什么更可能有流量时使用,帮助基于输入材料生成上升热点列表、排名和变化趋势视图、可用于内容规划的选题线索。
Intelligent thesis optimization system for computer science deep learning master's theses. Provides three-dimensional optimization: AI detection reduction, plagiarism reduction, academic polishing. Us…
Full Windows desktop control. Mouse, keyboard, screenshots - interact with any Windows application like a human.
A skill that uses GLM-V native grounding capabilities for coordinate conversion, bounding-box visualization, and more. GLM-V native grounding can locate any target specified by the prompt in an image …
A complete workflow to convert descriptive text into comic pages. Supports two styles: sketch (black and white line drawing) and color painting (color comics). Can read text from direct input or files…
Transform the knowledge system provided by users into a visualized knowledge map, including hierarchical structure diagrams, knowledge point relationship diagrams, and correlation analysis. It is used…
Official skill for recognizing handwritten text from images using ZhiPu GLM-OCR API. Supports various handwriting styles, languages, and mixed handwritten/printed content. Use this skill when the user…
Extract text from images and scanned PDFs using OCR. Supports 100+ languages, table detection, structured output (markdown/JSON), and batch processing.
Universal deep research agent team. 13-agent pipeline for rigorous academic research on any topic. 7 modes: full research, quick brief, paper review, lit-review, fact-check, Socratic guided research d…
Analyze images/videos and generate professional prompts for text-to-image and text-to-video AI tools (Midjourney, Stable Diffusion, DALL-E, Sora, Runway, Kling, Pika). Use when the user wants to gener…
Generate captions (descriptions) for images, videos, and documents using ZhiPu GLM-V multimodal model series. Use this skill whenever the user wants to describe, caption, summarize, or interpret the c…
Official skill for recognizing and extracting mathematical formulas from images and PDFs into LaTeX format using ZhiPu GLM-OCR API. Supports complex equations, inline formulas, and formula blocks. Use…
AI image generation and editing capabilities, based on Nano Banana (Gemini Image), enable text-to-image, image-to-image, and image editing. Suitable for creative design, marketing materials, social me…
小红书短视频运营增长助手适合内容创作者、运营、品牌方、电商在用户提供了小红书笔记链接时使用,帮助基于输入材料生成情绪和舆情视图、用户画像和意图分析、优化与转化建议。
Daily AI and tech news aggregator that fetches summarizes and pushes news from authoritative tech sites. Sources include 机器之心 36氪 TechCrunch The Verge and MIT Technology Review. Use when user asks for…
Video editing tool that requires no ffmpeg installation. All video processing is executed in the cloud - no local ffmpeg installation needed. If both input and output are URLs or Alibaba Cloud OSS, th…
Provide users with a complete solution for intelligent test paper generation based on knowledge points/difficulty, along with answer explanations and performance analysis, covering primary school, mid…
AI Resume Assistant — Polish, customize, export, and score Chinese/English resumes. Supports 40+ check items, a five-dimensional scoring system out of 100 points, and multiple export formats (Word/Mar…
Memory is the cornerstone of intelligent agents. Without it, every interaction starts from zero. This skill covers the architecture of agent memory: short-term (context window), long-term (vector s...
去除文本中的 AI 生成痕迹。适用于编辑或审阅文本,使其听起来更自然、更像人类书写。 基于维基百科的"AI 写作特征"综合指南。检测并修复以下模式:夸大的象征意义、 宣传性语言、以 -ing 结尾的肤浅分析、模糊的归因、破折号过度使用、三段式法则、 AI 词汇、否定式排比、过多的连接性短语。
抖音短视频运营增长助手适合内容创作者、运营、品牌方、电商在用户提供了抖音内容链接时使用,帮助基于输入材料生成情绪和舆情判断、用户画像和意图信号、运营建议和回复建议。