研究库 开源仓库
Repositories · organized/repo_cards

仓库/Skill 库

281 个

排序 Stars 周增
princedasresearch/systematic-review-metaverse-edu--agentic-ai
Python · 2026-09-03 Agent 智能体 应用 实验 Stars 0 周增 +0

生成式、多模态与 Agentic AI 在教育元宇宙中的系统性综述:应用、架构、挑战与未来方向。Systematic review of Generative, Multimodal, and Agentic AI in the Educational Metaverse: applications, architectures, challenges, and future directions.

agentmultimodal
Onlykrizz/prompt-for-code
未知语言 · 2026-10-01 Agent 智能体 教程 实验 Stars 0 周增 +0

🔧 通过灵活的 prompt 模板与任务规范项目简化 AI Agent,实现高效的指令与记忆管理。🔧 Streamline AI agents with this flexible prompt template and task specification project for effective instruction and memory management.

agentmultimodalllm-infra
nick19patrick/qwen36-dual-3090
未知语言 · 2026-09-22 LLM 基础设施 工具 实验 Stars 0 周增 +0

在双 RTX 3090 GPU 上以 262K 上下文长度并发流运行 Qwen3.6-27B,开发工作已迁移至 club-3090。Run Qwen3.6-27B at 262K context with concurrent streams on dual RTX 3090 GPUs. Development moved to club-3090.

multimodalllm-infra
matguo/Phyadv_pipeline
未知语言 · 2026-08-12 多模态 应用 实验 Stars 0 周增 +0

本仓库包含论文补充材料,旨在提升系统综述的透明度、可复现性与完整性,涵盖筛选文档、检索策略以及计算机视觉中物理对抗攻击的详细分析。This repository contains the supplementary materials accompanying our paper. The materials are designed to support the transparency, reproducibility, and comprehensiveness of the systematic review, including screening documentation, search strategies, and detailed analysis of physical adversarial attacks in computer vision.

multimodalrisk
Macr7523/Video-Summarizer
Python · 2026-08-28 多模态 教程 实验 Stars 0 周增 +0

本地使用 AI 总结视频,从讲座、会议和教程中提取视觉亮点与文本摘要,基于 RTX 40 系列 GPU 的 CUDA 加速Summarize videos locally with AI, extracting visual highlights and text summaries from lectures, meetings, and tutorials using CUDA on RTX 40-series GPUs

multimodalllm-infra
lol-dungeonmaster/official-daily-banana
Jupyter Notebook · 2026-09-06 多模态 数据集 实验 Stars 0 周增 +0

每日提示词,激发 nano banana 生成灵感。Daily prompts to inspire nano banana generation.

multimodalllm-infra
littlecookie0722/wairc-2026
Python · 2026-08-23 多模态 应用 实验 Stars 0 周增 +0

基于 IQ 信号、采用 STFT 频谱图、视觉模型、k-fold 训练与集成推理的多节点 RF 无人机识别可复现研究工具包。A reproducible research toolkit for multi-node RF drone identification from IQ signals using STFT spectrograms, vision models, k-fold training, and ensemble inference.

multimodalllm-infra
Lacenedihia/Google-Search-Ranking-Discoverability-Capstone
HTML · 2026-10-05 多模态 应用 实验 Stars 0 周增 +0

研究问题与暂定方向Research Question and Provisional Lane

multimodal
Kvxw1105/kv-agent-vision
Python · 2026-08-11 Agent 智能体 应用 实验 Stars 0 周增 +0

Give pure-text LLM agents eyes: zero-dependency image understanding, OCR & image Q&A via MCP server + standalone Python script. 给纯文本 AI Agent 的图片理解能力(看图/OCR/图片问答)。

agentmultimodalllm-infra
Jo-YuHuang/medical-image-segmentation-literature-review
未知语言 · 2026-08-25 多模态 收藏榜 研究原型 Stars 0 周增 +0

跨影像模态与临床需求的医学图像分割方法文献综述。Literature review of medical image segmentation methods across imaging modalities and clinical needs

multimodal
jakemorgan-research/codex-sci-research-lifecycle
Python · 2026-08-22 多模态 应用 实验 Stars 0 周增 +0

可审计的 Codex skill 与 Python 项目包,用于证据可追溯的研究、系统综述、投稿与修订。Auditable Codex skill and Python project pack for evidence-traceable research, systematic reviews, submission, and revision.

multimodal
iron-XiangruChen/reservoir-image-review-reproducibility
Python · 2026-08-27 评测基准 工具 研究原型 Stars 0 周增 +0

非常规储层图像解释中深度学习系统综述的可复现材料、研究级编码数据、文档与验证脚本Reproducibility materials, study-level coding data, documentation, and verification scripts for a systematic review of deep learning in unconventional reservoir image interpretation.

multimodal
Harryphan72007/minh-phan-portfolio
TypeScript · 2026-08-26 多模态 应用 实验 Stars 0 周增 +0

Minh Phan 作品集:机器学习系统、计算机视觉、可复现研究与软件工程Minh Phan portfolio: ML systems, computer vision, reproducible research, and software engineering.

multimodal
hanjiarui1025-a11y/rebuttal-revision
未知语言 · 2026-08-16 多模态 应用 实验 Stars 0 周增 +0

Codex Skill:学术期刊 rebuttal 与修改工作流Codex skill for academic journal rebuttal and revision workflows.

multimodal
galaxy99881/galaxy99881
未知语言 · 2026-08-16 多模态 应用 实验 Stars 0 周增 +0

Zhixin Li 的学术主页:超声医学、医学影像 AI、多模态学习与可复现研究。Academic homepage of Zhixin Li: ultrasound medicine, medical imaging AI, multimodal learning, and reproducible research.

multimodal
fhlyongko/graduate-academic-writing-ebook-kr
TypeScript · 2026-08-09 多模态 应用 实验 Stars 0 周增 +0

面向研究生学术英语写作、研究设计、伦理、修改与发表的韩文交互式现场指南Interactive Korean field guide for graduate academic English writing, research design, ethics, revision, and publication

multimodal
Erikalaylafajri15/MOSS-VL
未知语言 · 2026-09-22 多模态 模型 实验 Stars 0 周增 +0

使用 vision-language model 处理视觉与文本数据,执行多模态推理与图像理解任务。Process visual and textual data with this vision-language model for multimodal reasoning and image understanding tasks.

multimodalllm-infra
emircbngl/dhm-hybrid
Python · 2026-08-17 Agent 智能体 应用 实验 Stars 0 周增 +0

开源数字全息显微镜(DHM)工作站:离轴全息重建、定量相位成像、自动对焦,以及基于经典 + CNN 混合管线的无参考相位检索。PySide6 GUI,可选 Apple Silicon MLX 加速,并提供面向 AI Agent 的 MCP server。Open-source digital holographic microscopy (DHM) workstation: off-axis hologram reconstruction, quantitative phase imaging, autofocus, and reference-free phase retrieval via a hybrid classical + CNN pipeline. PySide6 GUI, optional Apple Silicon MLX acceleration, and an MCP server for AI agents.

agentragmultimodalengineering
docxology/DuckRabbit
Python · 2026-09-27 多模态 应用 实验 Stars 0 周增 +0

Python 包,作为可复现的科学刺激生成视觉、听觉和视听错觉 —— 17 个已编目错觉导出为 PNG、WAV、GIF、MP4 和 NPZ,附带 SHA-256 清单与确定性随机种子,含 15 个附带源数据边车的发表级图表,测试覆盖率 90% 以上Python package that generates optical, auditory and audio-visual illusions as reproducible scientific stimuli — 17 catalogued illusions exported to PNG, WAV, GIF, MP4 and NPZ with SHA-256 manifests, deterministic seeds, 15 publication figures with source-data sidecars, and 90%+ test coverage.

ragmultimodal
Dedebanded912/course-eligibility-hub
HTML · 2026-10-08 评测基准 应用 实验 Stars 0 周增 +0

即时评估课程资格,输出明确的通过/未通过结果以及定制化的入学测试。Instantly evaluate course eligibility with clear pass/fail results and tailored entry assessments.

multimodaldatabaseriskllm-infra
Davimoren9040/Youtube-Video-Transcribe-Summarizer-LLM-App
未知语言 · 2026-10-02 多模态 应用 实验 Stars 0 周增 +0

基于 Whisper、Gemini、Streamlit、yt-dlp 与 FFmpeg 构建的 AI 应用,可即时转录并总结任意 YouTube 视频。Transcribe and summarize any YouTube video instantly with AI-powered app using Whisper, Gemini, Streamlit, yt-dlp & FFmpeg.

multimodalllm-infra
CSE-Sudo-26/on-care
Dart · 2026-08-20 Agent 智能体 应用 实验 Stars 0 周增 +0

面向 20–30 岁高血压与糖尿病高风险人群的 AI 医疗 — 教练关联的饮食与运动指导平台AI healthcare for 20–30s at risk of hypertension & diabetes — Trainer-Linked Diet & Exercise Coaching Platform

ragmultimodalrisk
ChlorineHi/ST-Path-Survey
未知语言 · 2026-10-08 RAG 检索增强 模型 实验 Stars 0 周增 +0

综述论文 "From Representation Learning to Foundation Models" 的官方仓库。系统性地回顾了空间转录组学与病理学的多模态融合,提出三层分类法(嵌入、模型、知识层级)以及 2018 至 2026 年的演进路线图。Official repository for the survey "From Representation Learning to Foundation Models". A systematic review of multimodal fusion for Spatial Transcriptomics and Pathology, featuring a three-tier taxonomy (Embedding, Model, and Knowledge levels) and an evolutionary roadmap from 2018 to 2026.

multimodal
binglinwang473-hub/original-agent-skills
Python · 2026-08-11 Agent 智能体 应用 实验 Stars 0 周增 +0

binglinwang473-hub 原创 Agent SkillsOriginal Agent Skills by binglinwang473-hub

agentmultimodal
barah123/Claude-Biomedical-Research-Skills
Python · 2026-08-20 Agent 智能体 应用 实验 Stars 0 周增 +0

面向生物医学系统综述的 Claude Code skills 协议:MCP 锚定引用、实证完整性、PRISMA/PROSPERO 系统综述、新颖性核查、Zotero 操作、并行审稿手稿修订,并提供面向健康科学研究策划的生信、临床与科学传播 skill 集合。Claude Code skills protocol for biomedical systematic review: MCP-grounded citations, empirical integrity, PRISMA/PROSPERO systematic reviews, novelty checks, Zotero operations, and parallel-critic manuscript revision — plus a curated set of bioinformatics, clinical, and scientific-communication skills for health-science research.

multimodalengineering
asimfish/super_writer
Python · 2026-09-05 多模态 应用 实验 Stars 0 周增 +0

有证据支撑的学术写作与修订技能:贡献优先的规划、引文支持、手稿审校,以及 LaTeX/Word 交付。Evidence-backed academic writing and revision skill: contribution-first planning, citation support, manuscript audits, and LaTeX/Word delivery.

agentmultimodalengineering
api-evangelist/kibin
未知语言 · 2026-09-23 多模态 应用 实验 Stars 0 周增 +0

Kibin — 由 API Evangelist 出品的公共 API 表面的独立第三方画像。Kibin 是一家运营于 kibin.com 的消费者研究与学术写作公司,自 2011 年起受到学生信赖。其最新产品是一款 AI 驱动的学习伙伴(iOS app),可将讲座、笔记、PDF 与 YouTube 视频转化为个性化学习指南、测验……Kibin — independent third-party profile of a public API surface, by API Evangelist. Kibin is a consumer study and academic-writing company operating at kibin.com, trusted by students since 2011. Its newest product is an AI-powered study partner (iOS app) that turns lectures, notes, PDFs, and YouTube videos into personalized study guides, quizzes, f

multimodal
ajgarciaj/NaViL
未知语言 · 2026-09-19 多模态 应用 实验 Stars 0 周增 +0

🌐 在数据受限条件下重新思考多模态大语言模型的设计与扩展,以 NaViL 通过 Native Training 提升效率与性能🌐 Rethink Multimodal Large Language Models design and scaling under data constraints with NaViL, enhancing efficiency and performance through Native Training.

agentragmultimodalllm-infra
Acadianworkings918/augment-open-patch
TypeScript · 2026-09-19 工程化 工具 实验 Stars 0 周增 +0

自动化监控、打补丁与 Telegram 推送 Augment VSCode 插件更新,附带版本追踪与清理Automate monitoring, patching, and Telegram delivery of Augment VSCode plugin updates with version tracking and cleanup.

ragmultimodalllm-infra