围绕 PyTorch 在视觉、文本、强化学习等领域的一组示例。A set of examples around pytorch in Vision, Text, Reinforcement Learning, etc.
仓库/Skill 库
6 个 · 多模态 · 教程
面向 Claude Code 和 Codex 的 AI 视频 skill —— 基于 Remotion 制作电影级产品视频:含 152 张分镜配方卡、209 个动效预览,以及一套开箱即用的模板。AI video skill for Claude Code & Codex — cinematic product videos with Remotion: 152 shot recipe cards, 209 motion previews, a production-ready template
面向编码 Agent 的程序化视频方案——在本地将 HTML 转视频。把 HTML、CSS 与数据渲染为真实 MP4,支持可插拔渲染引擎、21 套模板与 AI 配乐。Apache-2.0,无按次计费。Open Design 团队的官方项目。Programmatic video for coding agents — HTML to video on your laptop. Turn HTML, CSS & data into real MP4s with pluggable render engines, 21 templates, AI soundtrack. Apache-2.0, no per-render fees. An official project by the Open Design team.
本仓库最初是我毕设项目的 PyTorch Lightning 模板,后来演进为包含完整项目实现本身。研究问题是地理定位如何影响遥感视觉-语言模型的训练与推理。注:仓库名称与描述可能变化。This repo originally was a PyTorch Lightning template for my thesis project. It has since evolved to contain the full project implementation itself. The research question is how geolocation influences the training and inference of vision-language models for remote sensing. Note: The repository name and description may change
本地使用 AI 总结视频,从讲座、会议和教程中提取视觉亮点与文本摘要,基于 RTX 40 系列 GPU 的 CUDA 加速Summarize videos locally with AI, extracting visual highlights and text summaries from lectures, meetings, and tutorials using CUDA on RTX 40-series GPUs
基于 GitHub Page 的 2026 年文献综述 Miscanthus 产量制图演示。Demonstration of GitHub Page with Miscanthus Yield mapping for Literature Review 2026