π RuView 将现成 WiFi 信号转化为实时空间智能、生命体征监测和存在检测,全程无需任何视频画面。π RuView turns commodity WiFi signals into real-time spatial intelligence, vital sign monitoring, and presence detection — all without a single pixel of video.
仓库/Skill 库
45 个 · 多模态 · 应用 · 快速增长
将任意品牌转化为可滚动 3D 世界落地页的 skillA skill that turn any brand into a scrollable 3D world landing page
使用 Claude Code 或 Codex 生成可投产的 Lottie 动画Generate production-ready Lottie animations with Claude Code or Codex
PaperSpine 是以动机驱动的 Skill,用于研读高质量学术论文、构建论文核心论点,并通过证据感知蓝图、修订矩阵与 LaTeX 安全审计来重写稿件。PaperSpine is a motivation-driven skill for learning from strong academic papers, building a paper’s central argument, and rewriting manuscripts through evidence-aware blueprints, revision matrices, and LaTeX-safe audits.
免费开源项目,将 youtube 视频转化为爆款短视频。集成高光检测、字幕、翻译与配音,一站式内容制作。Free open-source project designed for turning youtube-viedos into viral short videos. Highlight detection, subtitles, translation, voiceover, all in one for your content.
输入主题,输出带解说的解释视频。一款 Claude Code / Codex Skill,可将任意主题转化为黑底动态图形解释视频,配备 TTS 配音、字幕和章节进度条。支持中文或英文;每一帧均通过 Remotion 以代码方式绘制。Topic in, narrated explainer video out. A Claude Code / Codex skill that turns any topic into a black-canvas motion-graphics explainer video with TTS voiceover, subtitles and a chapter progress bar. Chinese or English; every frame drawn in code with Remotion.
Local-first、轻量级、AI-native 的视频工作空间。Local-first, lightweight, AI-native video workspace.
使用 TypeScript 与 Three.js 构建的 Rhine Lab 档案界面。Rhine Lab archive interface built with TypeScript and Three.js
早耳——仅依赖 CPU 的实时多语种语音转文字。支持实时字幕、浏览器仪表盘、说话人标签与翻译,无需 GPU,无需云端。早耳 - Real-time multilingual speech-to-text on CPU only. Live subtitles, browser dashboard, speaker labels, translation. No GPU, no cloud.
本地 AI 歌曲生成器,支持可编辑乐谱 —— 在你的 GPU 上运行 YuE2:带人声的全曲、五线谱、翻唱与精确回放。原生 Windows 应用,无需 Python,安装包自带自动更新。Local AI song generator with an editable score — YuE2 on your GPU: full songs with vocals, sheet music, covers, exact replay. Native Windows app, no Python, installer with auto-update.
代码优先的视频,原生渲染于 WebGPUCode-first video, rendered natively on WebGPU.
基于 DiffusionGemma 构建的开放、Jev 兼容的 System One 决策服务器。Open, Jev-compatible System One decision server on DiffusionGemma
在 Omarchy 上录制会议:将麦克风与电脑音频分别录为两条音轨,在本地机器上转录,包含说话人、章节与播放器。Record meetings on Omarchy: mic and computer audio as two tracks, transcribed on your own machine, with speakers, chapters and a player.
私密、无限制的背景移除,在浏览器中即可运行。Private, unlimited background removal that runs in your browser.
BridgeMind 开源的 AI 视频剪辑桌面应用Open-source AI video clipping desktop app by BridgeMind
个人 Agent 缺失的 GUI。Television 为你和你的 Agent 提供用于创建和处理工件的可视化空间。The missing GUI for personal agents. Television gives you and your agent a visual space for creating and working with artifacts.
面向 Polymarket 预测市场的交易机器人——浏览 CLOB 市场、在终端查看订单簿、运行套利检测、流动性提供与跨市场套利策略,支持模拟交易和风险限额。教育性开源工具包——不构成投资建议。非官方社区项目,与 Polymarket 无关。Polymarket trading bot for prediction markets — browse CLOB markets, watch the order book in the terminal, run edge detection, liquidity provision and cross-market arbitrage strategies with paper trading and risk limits. Educational open-source toolkit — not financial advice. Unofficial community project, not affiliated with Polymarket.
用 AI 可视化每一步——双语可视化讲解、交互式模型与同步旁白。Visualize Every Step with AI — bilingual visual explanations, interactive models and synchronized narration.
使用 Codex + ChatGPT 将参考视频转换为可编辑的动态图形;支持对比、定制并导出 MP4 或 HyperFrames 项目;提供本地 Windows 应用与在线工作台。Turn reference videos into editable motion graphics with Codex + ChatGPT. Compare, customize, and export MP4s or HyperFrames projects. Local Windows app + online studio.
在 Mac 上的虚拟机中运行 Omarchy(OmacVM.app、UTM、VMware Fusion 或 Parallels),获得原生般的体验:Omarchy 栏位于刘海旁,支持触控板手势与类 macOS 滚动,复用 Mac 的 Wi-Fi、音频和键盘。一条命令即可构建和切换功能。Omarchy in a VM on your Mac (OmacVM.app, UTM, VMware Fusion or Parallels), feeling native: Omarchy's bar beside the notch, trackpad gestures, macOS-like scrolling, the Mac's Wi-Fi, audio and keys in Omarchy. One command to build and switch features.
ComfyUI 节点,为 MiniMax H3 而设计,可串联多个视频片段,具备运动上下文、磁盘缓存、动态图像参考、音频参考支持以及最终视频/音频的无缝解码ComfyUI node for MiniMax H3 that chains multiple video clips with motion context, disk caching, dynamic image references, audio reference support, and seamless final video/audio decoding.
面向生成式接口的无障碍 React 加载状态:流式文本、内联活动指示与图像生成。Accessible React loading states for generative interfaces: streamed text, inline activity, and image generation.
语音听写,同时结合你所指屏幕位置的上下文。Voice dictation with the screen context you point at.
用双手操控屏幕——基于摄像头的免穿戴、免手柄手势追踪界面,让你的 AI 直接响应动作Move things on your screen with your bare hands. A webcam-powered, hand-tracked interface for your AI. No headset. No controllers.
DSH 插件:将内核启动页替换为全窗口视频片段,随后渐变过渡到应用。中文说明见 MANUAL.md。DSH plugin: replaces the kernel boot page with a full-window video clip, then dissolves into the app. 中文说明见 MANUAL.md
轻量、快速且准确的土耳其语 TTS。860 万参数,在 Freya-TR-Eval 上 WER 为 0.92%,单 GPU 上首段音频延迟约 4 ms,速度达 1.300× 实时。支持音频流式输出、单 GPU 多调用者批处理,并可在 GPU 或 CPU 上离线运行。Tiny, fast and accurate Turkish text-to-speech. 8.6M parameters, 0.92% WER on Freya-TR-Eval, first audio in ~4 ms and 1,300× real time on one GPU. Streams audio, batches many callers on one GPU, and runs offline on a GPU or CPU.
受 Premiere 启发、面向 MiniMax H3 Ref2VA 的 PySide6 导演工作室,集成 AI 分镜规划、语义媒体增强、时间线 prompt 调和,以及通过 ComfyUI 实现的镜头感知长视频渲染。Premiere-inspired PySide6 director studio for MiniMax H3 Ref2VA with AI shot planning, semantic media enrichment, timeline prompt reconciliation and shot-aware long-video rendering through ComfyUI.
https://abox.tools/ 的源码——一个面向图像、视频、音频、PDF 与文本的小型 Web 工具集。文件永不离开本机,因为不存在任何能将它们外发的代码路径。Source for https://abox.tools/ — a box of small web tools for images, video, audio, PDFs and text. Your files never leave your machine, because there is no code path that could send them anywhere.