用于系统综述与 meta 分析的 R 脚本,可从 DOI 提取全面的元数据。R scripts for extracting comprehensive metadata from DOIs for systematic reviews and meta-analyses
仓库/Skill 库
111 个 · 数据与向量库
GMS 6803 临床研究数据科学课程材料,含幻灯片、代码与实操实验,内容涵盖真实世界健康数据、基于 EHR 的队列构建、SQL、Python、统计分析与临床研究中的机器学习Course materials for GMS 6803 Data Science for Clinical Research, including slides, code, and hands-on labs on real-world health data, EHR-based cohort development, SQL, Python, statistical analysis, and machine learning for clinical research.
关于精神病学障碍中迷幻疗法的活体系统综述,基于 Docker 化 Dash 应用与 PostgreSQL 后端构建。Living systematic review of psychedelic therapy in psychiatric disorders. Built as a dockerized Dash app with PostgreSQL backend.
开源 DBMS 生态的 Reference Quotient 分析与可复现研究流水线。项目列表每月通过 https://github.com/birdflyi/od_label_issue_gen/tree/main/data/database_repo_label_dataframe 中的 csv 文件更新Reference Quotient analysis and reproducible research pipeline for the open-source DBMS ecosystem. The projects list are updated in the csv files in https://github.com/birdflyi/od_label_issue_gen/tree/main/data/database_repo_label_dataframe monthly.
按分支整理的北大课程存档,含精选报告与学术写作作品集。PKU course archives organized by branch, with a curated portfolio of selected reports and academic writing.
自动化 Python 兼容性补丁工具,将 Google Scholar PDF Reader 从 Chromium 移植到基于 Firefox 的浏览器,并使用 IndexedDB 集成自定义离线论文数据库。Automated Python compatibility patcher to port Google Scholar PDF Reader from Chromium to Firefox-based browsers, integrating a custom offline paper database using IndexedDB.
arXiv 论文的端到端推荐引擎:流式接入开放快照,使用量化 ONNX 模型对标题与摘要进行 embedding,构建 FAISS 向量索引,通过 FastAPI 服务提供推荐,并在 Streamlit 面板中交互探索。End-to-end recommendation engine for arXiv research papers: stream an open snapshot, embed titles + abstracts with a quantized ONNX model, build a FAISS vector index, serve recommendations through a FastAPI service, and explore them in a Streamlit dashboard.
SLR Harvester Web 是一个面向 Scopus、PubMed 和 OpenAlex 系统综述的浏览器仪表盘,可持久化记录检索策略、标签与筛选决策,实现完整方法学透明性;注册后可跨设备同步项目,也支持本地文件夹离线使用SLR Harvester Web is a browser-based dashboard for Scopus, PubMed and OpenAlex systematic literature reviews. It persistently logs search strategies, tags, and screening decisions for full methodological transparency. Sign up to sync projects across devices, or work offline with a local folder.
科学写作作品集:摘要、Bulk & scRNA-seq 方法学文档、结果解读与期刊投稿信。Scientific writing portfolio: abstracts, Bulk & scRNA-seq methodology documentation, results interpretation, and journal cover letters
从 Paperimage 实验导出的文献综述论文语料库(含元数据与 PDF)。Literature review paper corpus (metadata + PDFs) exported from Paperimage experiments.
有关基于运动信号的自闭症检测 AI 方法综述的补充数据、检索策略与提取文件。Supplementary data, search strategies, and extraction files for the review on AI methods for autism detection from motion signals.
针对任意 GitHub 仓库的自然语言问答 —— 构建 Neo4j 依赖图与 Qdrant 向量索引,用自然语言即可提问结构性与行为性问题。Natural-language Q&A over any GitHub repo — builds a Neo4j dependency graph + Qdrant vector index so you can ask structural and behavioral questions in plain English
从学术数据库中采集学术元数据,用于组织文献检索与构建研究参考目录。Harvest scholarly metadata from academic databases to organize literature searches and build research bibliographies.
"Domain-Driven Design in Practice: A Large-Scale Empirical Characterisation of the Open-Source Ecosystem" 的复现包——包含数据收集工具、GPT-4o 语义验证流水线,以及针对所有研究问题的 SQL 查询。Replication package for "Domain-Driven Design in Practice: A Large-Scale Empirical Characterisation of the Open-Source Ecosystem" — includes data collection tools, GPT-4o semantic validation pipeline, and SQL queries for all research questions.
开源、本地优先的足球数据、分析与内容工作流。Open-source, local-first football data, analysis, and content workflow.
Mascarene Paleo 是记录马斯卡林群岛古生物与亚化石记录的研发基础设施。该仓库旨在提供关于标本、分类群、地点、出版物、历史资料、数字资源及未解研究问题的透明、可更新、可引用的记录。Mascarene Paleo is a developing research infrastructure for documenting the palaeontological and subfossil record of the Mascarene Islands. This repository is intended to provide a transparent, updateable, and citable record of specimens, taxa, localities, publications, historical sources, digital resources, and unresolved research questions.
GeneTech 14 — 多领域实体集合(6700+ 实体,覆盖 14 个领域)GeneTech 14????? - ?????????? (6700+ entities, 14 domains)
von der Heyde, L., Keusch, F., Buskirk, T. D., & Eck, A. (2026). AI in the Loop?! A Systematic Review of the Use of Large Language Models in Survey and Public Opinion Research. SocArXiv. https://doi.org/10.31235/osf.io/eubj4_v1 的复现材料与文献数据库Replication materials and literature databases for von der Heyde, L., Keusch, F., Buskirk, T. D., & Eck, A. (2026). AI in the Loop?! A Systematic Review of the Use of Large Language Models in Survey and Public Opinion Research. SocArXiv. https://doi.org/10.31235/osf.io/eubj4_v1.
一套可移植、可复现的 PRISMA 2020 系统综述流水线:检索 15 个数据库、执行去重、以完整审计追踪管理人工筛选,并根据决策日志计算 PRISMA 流程计数。A portable, reproducible PRISMA 2020 systematic literature review pipeline. Searches 15 databases, deduplicates, manages human screening with a full audit trail, and computes PRISMA flow counts from the decision log.
引文管理 SaaS:在 1000+ 目录中管理商家信息。Uberall 竞品。Citation Management SaaS: Manage business listings across 1000+ directories. Uberall competitor.
经济研究、计量与统计分析及数据项目作品集。Portfolio of economic research, econometric and statistical analysis, and data projects.
Islam West Africa Collection 的 RDF 词汇表:面向 Omeka S 的 AI 处理溯源与按模型键控的情感标注属性RDF vocabulary for the Islam West Africa Collection: AI processing provenance and model-keyed sentiment annotation properties for Omeka S
将真实世界的研究问题转化为带引用、可关联的 Snowflake 表。Turn real-world research questions into cited, joinable Snowflake tables.
开放研究项目:对乳腺癌、肺癌与黑色素瘤领域近期肿瘤学临床试验进行 meta 分析。Open research project: Meta-analysis of recent oncology clinical trials across breast cancer, lung cancer, and melanoma
纯 Go 实现的分布式列式 SQL 分析引擎——对 S3 上 Parquet 进行向量化执行,无需 JVM、无需 CGo。以 Trino 级别的能力提供相当性能,却只需极小资源占用。原生支持 IPv4/CIDR/MAC 等网络类型,适用于遥测和安全场景。Distributed columnar SQL analytics engine in pure Go — vectorized execution over Parquet on S3, no JVM, no CGo. Trino-class performance on a fraction of the footprint. Network-native IPv4/CIDR/MAC types for telemetry and security workloads.
面向 arXiv 与 Semantic Scholar 的 Scrapy Spider,聚焦非洲 AI 生态 — 导出包含引用次数的丰富 CSV。Spiders Scrapy pour arXiv et Semantic Scholar, centrés sur l'écosystème IA africain — export CSV enrichi du nombre de citations.
将研究压缩包转化为可搜索、可引用的素材。拆分为论文、代码与数据集,索引后即可提出 Deep Research 问题,并引用页码与源行。Self-hosted(Postgres、MinIO、Qdrant)。Turn research zips into searchable, citable materials. Split papers, code, and datasets; index them; ask Deep Research questions that cite pages and source lines. Self-hosted (Postgres, MinIO, Qdrant).
面向基于丙泊酚的 ICU 镇静的计算数字孪生:SQL 数据提取、特征工程、消融实验和论文图表。Computational digital twin for propofol-based ICU sedation: SQL extraction, feature engineering, ablation experiments, and publication figures.
支持神经保留 RARP 术中热规避与能量使用系统综述及 meta 分析的衍生数据与 R 代码。Derived data and R code supporting a systematic review and meta-analysis of thermal avoidance versus energy use during nerve-sparing RARP.
2026 年面向 AI 代码生成的 PostgreSQL 优化终极指南。The Ultimate guide to PostgreSQL Optimization for AI Code Generation in 2026
文章《Approach-avoidance tendencies and physical activity level: A systematic review and meta-analysis》(趋近-回避倾向与体力活动水平:系统综述与 Meta 分析)的代码、数据及相关材料。Codedata, and material for the article: "Approach-avoidance tendencies and physical activity level: A systematic review and meta-analysis"