AudioGPT:理解与生成语音、音乐、声音与说话人头像AudioGPT: Understanding and Generating Speech, Music, Sound, and Talking Head
仓库/Skill 库
37 个 · 多模态 · 应用 · 学术写作
AIGCPanel 是一个简单易用的一站式AI数字人系统,支持视频合成、声音合成、声音克隆,简化本地模型管理、一键导入和使用AI模型。
A general-purpose AIGC video engine: script to finished film in one pipeline — dramas, ads, product videos, otome games, and more. | 通用 AIGC 视频引擎 —— 从剧本到成片一条流水线,漫剧、广告、电商、乙游皆可
永久免费开源的 AIGC 课程, 目前已支持Claude Code,Codex,Hermes,OpenClaw,Obsidian,Prompt Engineering, ChatGPT, Midjourney, Runway, Stable Diffusion, AI数字人,AI声音&音乐,开源大模型
📺 基于 Transformer Diffusion 的高分辨率长视频生成端到端解决方案。📺 An End-to-End Solution for High-Resolution and Long Video Generation Based on Transformer Diffusion
一个面向英文学术写作的保守型 AIGC 检测器指导的论文改写 Skill。支持 Turnitin AI、CNKI AIGC、最小化编辑修订、保留学术要素、定性/定量路由,以及逐章降低 AI 写作风险,且不宣称绕过检测器。A conservative AIGC detector-informed thesis rewriting skill for English and Chinese academic writing. Supports Turnitin AI, CNKI AIGC, minimal-edit revision, protected academic elements, qualitative/quantitative routing, and chapter-by-chapter AI-writing risk reduction without detector-bypass claims.
所有人可用的 AI(MoAI)——面向韩语实务领域的 Claude Cowork 与 Claude Code AI harness 与插件市场。覆盖商业计划书、税务、法律、HR、营销、电商、BI、内容等领域,提供 Skill、Agent 与工作流。支持韩语 B2B 场景与办公文档(HWPX/DOCX/XLSX/PPTX/PDF)及 AI 多模态生成(图像/视频/语音)。内置 AI 痕迹审核与韩语 humanize-korean모두의 AI (MoAI) — Claude Cowork & Claude Code 한국 실무 도메인 AI 하네스·플러그인 마켓플레이스. 사업계획서·세무·법률·HR·마케팅·커머스·BI·콘텐츠 도메인 스킬·에이전트·워크플로우. Korean B2B + office docs (HWPX/DOCX/XLSX/PPTX/PDF) + AI media (image/video/voice). AI-slop 검수 + humanize-korean 내장.
[ISPRS 2026] 遥感图像去雾:进展、挑战与前景的系统综述[ISPRS 2026] Remote Sensing Image Dehazing: A Systematic Review of Progress, Challenges, and Prospects
面向严谨学术论文写作、修订与投稿的 Claude Code skill。跨领域通用,支持按论文设置期刊覆盖规则。Claude Code skill for rigorous academic paper writing, revision, and submission. Field-agnostic with per-paper journal overrides.
面向 AI 辅助研究的科学写作与文档完整性守护——论证经济性、确定性完整性、安全的 Word 编辑。本地 · 确定性 · 零 LLM。Scientific writing & document integrity guard for AI-assisted research — argument economy, deterministic integrity, safe Word editing. Local · Deterministic · Zero LLM.
证据优先的具身 AI 研究枢纽,覆盖 VLA、世界模型、多模态感知、文献综述、声明审计以及可复用的 Agent skillEvidence-first Embodied AI research hub for VLA, world models, multimodal sensing, literature reviews, claim audits, and reusable agent skills.
Local-first 学术 PDF 阅读器:离线 CPU LLM 翻译与注释、OCR、ECDICT 词典。Tauri + FastAPI。Local-first academic PDF reader: offline CPU LLM translation & glossing, OCR, ECDICT dictionary. Tauri + FastAPI.
作者感知的科学写作分析:作者空间、目标论文风格、校准的生成证据,以及安全合规的修订Authorship-aware scientific writing analysis: author space, target-paper style, calibrated generation evidence, and integrity-safe revision.
论文《Deep Learning in Dental Image Analysis: A Systematic Review of Datasets, Methodologies, and Emerging Challenges》的官方项目主页。The official project homepage for the paper "Deep Learning in Dental Image Analysis: A Systematic Review of Datasets, Methodologies, and Emerging Challenges"
将任意视频链接转化为博士级研究报告。自动化学术流水线:一键完成下载、转写、检索同行评审文献、核查论据并生成带引用的报告。免费且开源。Turn any video url into a PhD-grade research report. Automated academic pipeline: download, transcribe, search peer-reviewed literature, verify claims, generate cited reports — all with one command. Free & open source.
TTS、语音转换与口语对话 Agent 的活体系统综述Living systematic review of TTS, voice conversion, and spoken conversational agents
面向学术阅读的本地桌面 PDF 阅读器,支持 AI 翻译、OCR、高亮、笔记、翻译历史以及导入/导出。A local desktop PDF reader for academic reading, with AI translation, OCR, highlights, notes, translation history, and import/export support.
《Comparative Analysis of EEG Signal Preprocessing Impact on Generalization of a Transformer-Based EEG Imagined Speech Recognition》研究的可复现研究仓库。Reproducible research repository for 'Comparative Analysis of EEG Signal Preprocessing Impact on Generalization of a Transformer-Based EEG Imagined Speech Recognition' research work
Embodied Daily:每日具身智能 AI 论文推荐(HF Daily Papers + arXiv)Embodied Daily: daily embodied-AI paper recommendations (HF Daily Papers + arXiv)
面向中英双语生物医学研究的术语循证、主张强度与科学写作审计 skill。Evidence-aligned terminology, claim-strength, and scientific-writing audit skill for Chinese and English biomedical research.
本仓库包含论文补充材料,旨在提升系统综述的透明度、可复现性与完整性,涵盖筛选文档、检索策略以及计算机视觉中物理对抗攻击的详细分析。This repository contains the supplementary materials accompanying our paper. The materials are designed to support the transparency, reproducibility, and comprehensiveness of the systematic review, including screening documentation, search strategies, and detailed analysis of physical adversarial attacks in computer vision.
基于 IQ 信号、采用 STFT 频谱图、视觉模型、k-fold 训练与集成推理的多节点 RF 无人机识别可复现研究工具包。A reproducible research toolkit for multi-node RF drone identification from IQ signals using STFT spectrograms, vision models, k-fold training, and ensemble inference.
研究问题与暂定方向Research Question and Provisional Lane
可审计的 Codex skill 与 Python 项目包,用于证据可追溯的研究、系统综述、投稿与修订。Auditable Codex skill and Python project pack for evidence-traceable research, systematic reviews, submission, and revision.
Minh Phan 作品集:机器学习系统、计算机视觉、可复现研究与软件工程Minh Phan portfolio: ML systems, computer vision, reproducible research, and software engineering.
Codex Skill:学术期刊 rebuttal 与修改工作流Codex skill for academic journal rebuttal and revision workflows.
学术 PDF 阅读器:在浏览器中并排展示印刷版页面与重排阅读视图,由同一模型解析并联动。An academic PDF reader: the printed page and a reflowed reading view, side by side, linked by one model parsed in your browser.
Zhixin Li 的学术主页:超声医学、医学影像 AI、多模态学习与可复现研究。Academic homepage of Zhixin Li: ultrasound medicine, medical imaging AI, multimodal learning, and reproducible research.
面向研究生学术英语写作、研究设计、伦理、修改与发表的韩文交互式现场指南Interactive Korean field guide for graduate academic English writing, research design, ethics, revision, and publication
Python 包,作为可复现的科学刺激生成视觉、听觉和视听错觉 —— 17 个已编目错觉导出为 PNG、WAV、GIF、MP4 和 NPZ,附带 SHA-256 清单与确定性随机种子,含 15 个附带源数据边车的发表级图表,测试覆盖率 90% 以上Python package that generates optical, auditory and audio-visual illusions as reproducible scientific stimuli — 17 catalogued illusions exported to PNG, WAV, GIF, MP4 and NPZ with SHA-256 manifests, deterministic seeds, 15 publication figures with source-data sidecars, and 90%+ test coverage.
有证据支撑的学术写作与修订技能:贡献优先的规划、引文支持、手稿审校,以及 LaTeX/Word 交付。Evidence-backed academic writing and revision skill: contribution-first planning, citation support, manuscript audits, and LaTeX/Word delivery.