研究库 开源仓库
Repositories · organized/repo_cards

仓库/Skill 库

236 个 · 多模态

排序 Stars 周增
karim1988781/agri-robotic-imaging-meta-analysis
Jupyter Notebook · 2026-08-21 多模态 数据集 实验 Stars 0 周增 +0

农业机器人成像系统综述与 meta-analysis 的数据与分析代码。Data and analysis code for a systematic review and meta-analysis of agricultural robotic imaging

Jo-YuHuang/medical-image-segmentation-literature-review
未知语言 · 2026-08-25 多模态 收藏榜 研究原型 Stars 0 周增 +0

跨影像模态与临床需求的医学图像分割方法文献综述。Literature review of medical image segmentation methods across imaging modalities and clinical needs

multimodal
JessiePBhalerao/miscanthus-yield-maps-viz-demo
未知语言 · 2026-08-14 多模态 教程 实验 Stars 0 周增 +0

基于 GitHub Page 的 2026 年文献综述 Miscanthus 产量制图演示。Demonstration of GitHub Page with Miscanthus Yield mapping for Literature Review 2026

jakemorgan-research/codex-sci-research-lifecycle
Python · 2026-08-22 多模态 应用 实验 Stars 0 周增 +0

可审计的 Codex skill 与 Python 项目包,用于证据可追溯的研究、系统综述、投稿与修订。Auditable Codex skill and Python project pack for evidence-traceable research, systematic reviews, submission, and revision.

multimodal
Homologic-bid91/voice_clone_lab
Python · 2026-08-11 多模态 应用 实验 Stars 0 周增 +0

使用 Qwen3-TTS 在本地 GPU 上克隆声音并从文本生成语音,提供端到端训练流水线。Clone voices and generate speech from text locally on your GPU using Qwen3-TTS with an end-to-end training pipeline.

engineeringllm-infra
HelgDemidov/refigure
Python · 2026-08-20 多模态 应用 实验 Stars 0 周增 +0

保留图表的转换器 —— DOCX/XLSX 转 Markdown,采用原生 OOXML 图表数据提取(无需光栅化/OCR/VLM),并提供零损耗的复合图表标记。Converters where figures survive — DOCX/XLSX to Markdown with native OOXML chart-data extraction (no rasterize/OCR/VLM) and zero-loss composite-figure markers

ragllm-infra
Harryphan72007/minh-phan-portfolio
TypeScript · 2026-08-26 多模态 应用 实验 Stars 0 周增 +0

Minh Phan 作品集:机器学习系统、计算机视觉、可复现研究与软件工程Minh Phan portfolio: ML systems, computer vision, reproducible research, and software engineering.

multimodal
hanjiarui1025-a11y/rebuttal-revision
未知语言 · 2026-08-16 多模态 应用 实验 Stars 0 周增 +0

Codex Skill:学术期刊 rebuttal 与修改工作流Codex skill for academic journal rebuttal and revision workflows.

multimodal
gong8/paper-reader
HTML · 2026-08-25 多模态 应用 实验 Stars 0 周增 +0

学术 PDF 阅读器:在浏览器中并排展示印刷版页面与重排阅读视图,由同一模型解析并联动。An academic PDF reader: the printed page and a reflowed reading view, side by side, linked by one model parsed in your browser.

galaxy99881/galaxy99881
未知语言 · 2026-08-16 多模态 应用 实验 Stars 0 周增 +0

Zhixin Li 的学术主页:超声医学、医学影像 AI、多模态学习与可复现研究。Academic homepage of Zhixin Li: ultrasound medicine, medical imaging AI, multimodal learning, and reproducible research.

multimodal
fhlyongko/graduate-academic-writing-ebook-kr
TypeScript · 2026-08-09 多模态 应用 实验 Stars 0 周增 +0

面向研究生学术英语写作、研究设计、伦理、修改与发表的韩文交互式现场指南Interactive Korean field guide for graduate academic English writing, research design, ethics, revision, and publication

multimodal
Erikalaylafajri15/MOSS-VL
未知语言 · 2026-09-22 多模态 模型 实验 Stars 0 周增 +0

使用 vision-language model 处理视觉与文本数据,执行多模态推理与图像理解任务。Process visual and textual data with this vision-language model for multimodal reasoning and image understanding tasks.

multimodalllm-infra
docxology/DuckRabbit
Python · 2026-09-27 多模态 应用 实验 Stars 0 周增 +0

Python 包,作为可复现的科学刺激生成视觉、听觉和视听错觉 —— 17 个已编目错觉导出为 PNG、WAV、GIF、MP4 和 NPZ,附带 SHA-256 清单与确定性随机种子,含 15 个附带源数据边车的发表级图表,测试覆盖率 90% 以上Python package that generates optical, auditory and audio-visual illusions as reproducible scientific stimuli — 17 catalogued illusions exported to PNG, WAV, GIF, MP4 and NPZ with SHA-256 manifests, deterministic seeds, 15 publication figures with source-data sidecars, and 90%+ test coverage.

ragmultimodal
Davimoren9040/Youtube-Video-Transcribe-Summarizer-LLM-App
未知语言 · 2026-10-02 多模态 应用 实验 Stars 0 周增 +0

基于 Whisper、Gemini、Streamlit、yt-dlp 与 FFmpeg 构建的 AI 应用,可即时转录并总结任意 YouTube 视频。Transcribe and summarize any YouTube video instantly with AI-powered app using Whisper, Gemini, Streamlit, yt-dlp & FFmpeg.

multimodalllm-infra
BOHARRY/swipeta-research
Python · 2026-08-19 多模态 框架 研究原型 Stars 0 周增 +0

SwipeTA 背后关于图表识别、市场模式与噪声反馈的可复现研究Reproducible research on chart-reading, market patterns, and noisy feedback behind SwipeTA.

benisonodigie-dev/MMLA-systematic-Review
Jupyter Notebook · 2026-08-22 多模态 数据集 实验 Stars 0 周增 +0

MMLA 在发展中经济体与 LMIC 中的综合系统综述。A Comprehensive Systematic review of MMLA in developing economics and LMIC.

asimfish/super_writer
Python · 2026-09-05 多模态 应用 实验 Stars 0 周增 +0

有证据支撑的学术写作与修订技能:贡献优先的规划、引文支持、手稿审校,以及 LaTeX/Word 交付。Evidence-backed academic writing and revision skill: contribution-first planning, citation support, manuscript audits, and LaTeX/Word delivery.

agentmultimodalengineering
Arfayet/Research-Method---Systematic-Review-of-Diffusion-Models-for-MRI-Synthesis-and-Reconstruction
未知语言 · 2026-08-27 多模态 应用 实验 Stars 0 周增 +0
api-evangelist/kibin
未知语言 · 2026-09-23 多模态 应用 实验 Stars 0 周增 +0

Kibin — 由 API Evangelist 出品的公共 API 表面的独立第三方画像。Kibin 是一家运营于 kibin.com 的消费者研究与学术写作公司,自 2011 年起受到学生信赖。其最新产品是一款 AI 驱动的学习伙伴(iOS app),可将讲座、笔记、PDF 与 YouTube 视频转化为个性化学习指南、测验……Kibin — independent third-party profile of a public API surface, by API Evangelist. Kibin is a consumer study and academic-writing company operating at kibin.com, trusted by students since 2011. Its newest product is an AI-powered study partner (iOS app) that turns lectures, notes, PDFs, and YouTube videos into personalized study guides, quizzes, f

multimodal
ajgarciaj/NaViL
未知语言 · 2026-09-19 多模态 应用 实验 Stars 0 周增 +0

🌐 在数据受限条件下重新思考多模态大语言模型的设计与扩展,以 NaViL 通过 Native Training 提升效率与性能🌐 Rethink Multimodal Large Language Models design and scaling under data constraints with NaViL, enhancing efficiency and performance through Native Training.

agentragmultimodalllm-infra