农业机器人成像系统综述与 meta-analysis 的数据与分析代码。Data and analysis code for a systematic review and meta-analysis of agricultural robotic imaging
仓库/Skill 库
236 个 · 多模态
跨影像模态与临床需求的医学图像分割方法文献综述。Literature review of medical image segmentation methods across imaging modalities and clinical needs
基于 GitHub Page 的 2026 年文献综述 Miscanthus 产量制图演示。Demonstration of GitHub Page with Miscanthus Yield mapping for Literature Review 2026
可审计的 Codex skill 与 Python 项目包,用于证据可追溯的研究、系统综述、投稿与修订。Auditable Codex skill and Python project pack for evidence-traceable research, systematic reviews, submission, and revision.
使用 Qwen3-TTS 在本地 GPU 上克隆声音并从文本生成语音,提供端到端训练流水线。Clone voices and generate speech from text locally on your GPU using Qwen3-TTS with an end-to-end training pipeline.
保留图表的转换器 —— DOCX/XLSX 转 Markdown,采用原生 OOXML 图表数据提取(无需光栅化/OCR/VLM),并提供零损耗的复合图表标记。Converters where figures survive — DOCX/XLSX to Markdown with native OOXML chart-data extraction (no rasterize/OCR/VLM) and zero-loss composite-figure markers
Minh Phan 作品集:机器学习系统、计算机视觉、可复现研究与软件工程Minh Phan portfolio: ML systems, computer vision, reproducible research, and software engineering.
Codex Skill:学术期刊 rebuttal 与修改工作流Codex skill for academic journal rebuttal and revision workflows.
学术 PDF 阅读器:在浏览器中并排展示印刷版页面与重排阅读视图,由同一模型解析并联动。An academic PDF reader: the printed page and a reflowed reading view, side by side, linked by one model parsed in your browser.
Zhixin Li 的学术主页:超声医学、医学影像 AI、多模态学习与可复现研究。Academic homepage of Zhixin Li: ultrasound medicine, medical imaging AI, multimodal learning, and reproducible research.
面向研究生学术英语写作、研究设计、伦理、修改与发表的韩文交互式现场指南Interactive Korean field guide for graduate academic English writing, research design, ethics, revision, and publication
使用 vision-language model 处理视觉与文本数据,执行多模态推理与图像理解任务。Process visual and textual data with this vision-language model for multimodal reasoning and image understanding tasks.
Python 包,作为可复现的科学刺激生成视觉、听觉和视听错觉 —— 17 个已编目错觉导出为 PNG、WAV、GIF、MP4 和 NPZ,附带 SHA-256 清单与确定性随机种子,含 15 个附带源数据边车的发表级图表,测试覆盖率 90% 以上Python package that generates optical, auditory and audio-visual illusions as reproducible scientific stimuli — 17 catalogued illusions exported to PNG, WAV, GIF, MP4 and NPZ with SHA-256 manifests, deterministic seeds, 15 publication figures with source-data sidecars, and 90%+ test coverage.
基于 Whisper、Gemini、Streamlit、yt-dlp 与 FFmpeg 构建的 AI 应用,可即时转录并总结任意 YouTube 视频。Transcribe and summarize any YouTube video instantly with AI-powered app using Whisper, Gemini, Streamlit, yt-dlp & FFmpeg.
SwipeTA 背后关于图表识别、市场模式与噪声反馈的可复现研究Reproducible research on chart-reading, market patterns, and noisy feedback behind SwipeTA.
MMLA 在发展中经济体与 LMIC 中的综合系统综述。A Comprehensive Systematic review of MMLA in developing economics and LMIC.
有证据支撑的学术写作与修订技能:贡献优先的规划、引文支持、手稿审校,以及 LaTeX/Word 交付。Evidence-backed academic writing and revision skill: contribution-first planning, citation support, manuscript audits, and LaTeX/Word delivery.
Kibin — 由 API Evangelist 出品的公共 API 表面的独立第三方画像。Kibin 是一家运营于 kibin.com 的消费者研究与学术写作公司,自 2011 年起受到学生信赖。其最新产品是一款 AI 驱动的学习伙伴(iOS app),可将讲座、笔记、PDF 与 YouTube 视频转化为个性化学习指南、测验……Kibin — independent third-party profile of a public API surface, by API Evangelist. Kibin is a consumer study and academic-writing company operating at kibin.com, trusted by students since 2011. Its newest product is an AI-powered study partner (iOS app) that turns lectures, notes, PDFs, and YouTube videos into personalized study guides, quizzes, f
🌐 在数据受限条件下重新思考多模态大语言模型的设计与扩展,以 NaViL 通过 Native Training 提升效率与性能🌐 Rethink Multimodal Large Language Models design and scaling under data constraints with NaViL, enhancing efficiency and performance through Native Training.