wechat-article-extractor

Extract metadata and content from WeChat Official Account articles. Use when user needs to parse WeChat article URLs (mp.weixin.qq.com), extract article info (title, author, content, publish time, cover image), or convert WeChat articles to structured data. Supports various article types including p

By freestylefly · 3,797 installs

npx skills add freestylefly/wechat-article-extractor-skill --skill wechat-article-extractor

Source repository · Upstream listing

WeChat Article Extractor Extract metadata and content from WeChat Official Account (微信公众号) articles. Capabilities Parse WeChat article URLs ( mp.weixin.qq.com ) Extract article metadata: title, author, description, publish time Extract account info: name, avatar, alias, description Get article content (HTML) Get cover image URL Support multiple article types: post, video, image, voice, text, repost Handle various error cases: deleted content, expired links, access limits Usage Basic Extraction from URL Extraction from HTML Options Response Format Success Response Error Response Error Codes Code Message Description 1000 文章获取失败 General failure 1001 无法获取文章信息 Missing title or publish time 1002 请求失败 HTTP request failed 1003 响应为空 Empty response 1004 访问过于频繁 Rate limited 1005 脚本解析失败 Script parsing error 1006 公众号已迁移 Account migrated 2001 请提供文章内容或链接 Missing input 2002 链接已过期 Link expired 2003 内容涉嫌侵权 Content removed (copyright) 2004 无法获取迁移后的链接 Migration link failed 2005 内容已被发布者删除 Content deleted by author 2006 内容因违规无法查看 Content blocked 2007 内容发送失败 Failed to send 2008 系统出错 System error 2009 不支持的链接 Unsupported URL 2010 内容获取失败 Content fetch failed 2011 涉嫌过度营销 Marketing/spam content 2012 账号已被屏蔽 Account blocked 2013 账号已自主注销 Account deleted 2014 内容被投诉 Content reported 2015 账号处于迁移流程中 Account migrating 2016 冒名侵权 Impersonation Dependencies Required npm packages: cheerio HTML parsing dayjs Date formatting request promise HTTP requests qs Query string parsing lodash.unescape HTML entities Notes Handles various WeChat page structures and anti scraping measures Automatically detects article type from page content Supports extracting from Sogou WeChat search results ( weixin.sogou.com ) Some fields may be null depending on article type and page structure