面向严谨学术论文写作、修订与投稿的 Claude Code skill。跨领域通用,支持按论文设置期刊覆盖规则。Claude Code skill for rigorous academic paper writing, revision, and submission. Field-agnostic with per-paper journal overrides.
仓库/Skill 库
236 个 · 多模态
YuE2 音乐生成,以及面向 ComfyUI 的可编辑钢琴卷帘界面。YuE2 music generation and an editable piano roll for ComfyUI
PCB_lightgraph_portable 是一款点击即可运行的 PCB 图片智能分层与图纸导出的html工具。
OpenRouter 的 MCP server——与 300+ LLM(Claude、Gemini、GPT)对话,分析图像/音频/视频,生成图像/语音/音乐/视频(Veo 3.1、Sora、Seedance、Wan),支持 Claude Desktop、Cursor、Kiro、VS CodeMCP server for OpenRouter — chat with 300+ LLMs (Claude, Gemini, GPT), analyze images / audio / video, generate images / speech / music / video (Veo 3.1, Sora, Seedance, Wan) from Claude Desktop, Cursor, Kiro, VS Code.
受 Premiere 启发、面向 MiniMax H3 Ref2VA 的 PySide6 导演工作室,集成 AI 分镜规划、语义媒体增强、时间线 prompt 调和,以及通过 ComfyUI 实现的镜头感知长视频渲染。Premiere-inspired PySide6 director studio for MiniMax H3 Ref2VA with AI shot planning, semantic media enrichment, timeline prompt reconciliation and shot-aware long-video rendering through ComfyUI.
https://abox.tools/ 的源码——一个面向图像、视频、音频、PDF 与文本的小型 Web 工具集。文件永不离开本机,因为不存在任何能将它们外发的代码路径。Source for https://abox.tools/ — a box of small web tools for images, video, audio, PDFs and text. Your files never leave your machine, because there is no code path that could send them anywhere.
基于 Optical Flow 的 VR 帧生成。Frame generation for VR using Optical Flow
阿杜与 Opus 5.5 精选制作的动画动效模板 Skill。选模板和风格,用 Codex / Claude Code 把口播与素材剪成视频;提供 11,028 个精选 Lottie 动画素材 API。
面向 Linux 的真活体检测人脸解锁:一个 PAM 模块、一个守护进程,以及一个 Omarchy 锁屏指示器。Face unlock for Linux with real liveness detection: a PAM module, a daemon, and an Omarchy lock screen indicator
OneCamp:自托管一体化工作空间(聊天、任务、视频通话、文档、日历与 AI)。OneCamp: self-hosted all-in-one workspace (chat, tasks, video calls, docs, calendar & AI)
开放、结构化的印度一手法律数据:包含最高法院与全部 25 所高等法院的 3250 万判决分块、110 万条立法条款,以及构建数据集的爬虫;采用 CC BY 4.0 许可。Open, structured Indian primary law: 32.5M judgment chunks from the Supreme Court and all 25 High Courts, 1.1M legislation provisions, and the scrapers that build it. CC BY 4.0.
由豆包驱动的语音输入工具,适用于 Linux、Wayland 和 OmarchyDoubao-powered voice input for Linux, Wayland and Omarchy
面向 Polymarket 预测市场的交易机器人——可在终端浏览 CLOB 市场、查看订单簿,并执行 edge detection、流动性提供与跨市场套利策略,支持 paper trading 与风险限额。开源教育工具——不构成投资建议。非官方社区项目,与 Polymarket 无关。Polymarket trading bot for prediction markets — browse CLOB markets, watch the order book in the terminal, run edge detection, liquidity provision and cross-market arbitrage strategies with paper trading and risk limits. Educational open-source toolkit — not financial advice. Unofficial community project, not affiliated with Polymarket.
把 ComfyUI-MiniMaxH3-Easy 与 Goohai-MiniMax-H3_Integration 合并成一个统一的 MiniMax H3 创作工作台(非官方)
高性能视频生成网关与会话协调器。High-performance video generation gateway and session coordinator
交互式标注/分割的文献综述Literature Review for Interactive annotation/segmentation
NVIDIA DLSS 5 神经渲染网络开源复现 OpenDLSS-NR 的 Three.js(TSL / WebGPU)移植版。A Three.js (TSL / WebGPU) port of OpenDLSS-NR, the open-source reimplementation of NVIDIA's DLSS 5 neural rendering network
视觉定位的文献综述Literature review of visual localization.
ComfyUI 中的官方 MiniMax-H3 8 步 PDD Acc LoRA(alibaba-pai):LoRA + 并行解码头库,基于已训练 sigmas 的 euler 调度,8 步生成音视频Official MiniMax-H3 8-step PDD Acc LoRAs (alibaba-pai) in ComfyUI: LoRA + parallel-decoding head bank, euler on trained sigmas, audio+video in 8 steps
MusicGen Pro Trainer 2026:终极 AI 音乐工作室指南MusicGen Pro Trainer 2026: The Ultimate AI Music Studio Guide
高级 WARDOGS 辅助:ESP、自瞄与载具战斗增强 2026Advanced WARDOGS Hacks: ESP, Aimbot & Vehicle Combat Enhancements 2026
MatlowAI 的 MiniMax-H3 ComfyUI 节点:Contact-Sheet diffusion + Motion Lab(针对快速运动的 test-time 去绳状畸变)MatlowAI's MiniMax-H3 ComfyUI nodes: Contact-Sheet diffusion + Motion Lab (test-time de-roping of fast motion)
[ISPRS 2024] 卫星视频单目标跟踪:系统综述与定向目标跟踪基准[ISPRS 2024] Satellite Video Single Object Tracking: A Systematic Review and An Oriented Object Tracking Benchmark
CIT Voice Studio —— 将越南语文本转为语音,完全在本地运行。CIT Voice Stuido - Chuyển văn bản tiếng Việt thành giọng nói, chạy hoàn toàn trên máy bạn.
VisionFoundry:使用合成图像教会 VLM 视觉感知。VisionFoundry: Teaching VLMs Visual Perception with Synthetic Images
受 Discord、Telegram 和 Zoom 启发的开源团队聊天和视频会议工具。OpenSource team chats and video conferences inspired with Discord, Telegram & Zoom
一款把任意图片转换为专业生图提示词的 Chrome / Edge 浏览器插件。支持网页右键识图、本地上传与剪贴板粘贴,生成中文、英文和 JSON 提示词,并提供历史记录与收藏。
面向 AI 辅助研究的科学写作与文档完整性守护——论证经济性、确定性完整性、安全的 Word 编辑。本地 · 确定性 · 零 LLM。Scientific writing & document integrity guard for AI-assisted research — argument economy, deterministic integrity, safe Word editing. Local · Deterministic · Zero LLM.
独立的 Skill 与 Codex 插件,支持 OpenAI 兼容的图像生成、编辑、批处理工作流、QA 以及聚焦画布编辑。Standalone Skill and Codex Plugin for OpenAI-compatible image generation, editing, batch workflows, QA, and focused canvas editing.
证据优先的具身 AI 研究枢纽,覆盖 VLA、世界模型、多模态感知、文献综述、声明审计以及可复用的 Agent skillEvidence-first Embodied AI research hub for VLA, world models, multimodal sensing, literature reviews, claim audits, and reusable agent skills.