Repositories · organized/repo_cards

仓库/Skill 库

50 个 · 评测基准 · 应用

排序 Stars 周增
tav0-m/mm-ipsa-research
Python · 2026-08-14 评测基准 应用 实验 Stars 0 周增 +0

围绕矩匹配情景生成与时序投资组合评估的可复现研究。Reproducible research on moment-matching scenario generation and temporal portfolio evaluation

evaluation
synaptiai/flow-harness
TypeScript · 2026-08-20 评测基准 应用 实验 Stars 0 周增 +0

与供应商无关的 coding-agent 框架,具备确定性工作流图、持久化证据与 fail-closed 沙箱执行。Provider-neutral coding-agent harness with deterministic workflow graphs, durable evidence, and fail-closed sandboxed execution

agentllm-infra
Samar-Khalid/Paper-1-NL2SQL-Ambiguity
Python · 2026-08-21 评测基准 应用 实验 Stars 0 周增 +0

论文一:NL2SQL 歧义检测——M1.5 分类法、确定性基线、人工金标评估与可复现研究制品Paper 1: NL2SQL Ambiguity Detection — M1.5 taxonomy, deterministic baseline, human-gold evaluation, reproducible research artifact

evaluationdatabase
ratulhasan001/llm-mental-health-systematic-review-paper
TeX · 2026-08-11 评测基准 应用 研究原型 Stars 0 周增 +0
pot-oie/Face-DeStyle-Pipeline
Python · 2026-08-25 评测基准 应用 实验 Stars 0 周增 +0

面向人脸域去风格化、结构条件化、评估与质量过滤的紧凑型可复现研究流水线。Compact reproducible research pipeline for face-domain destylization, structural conditioning, evaluation, and quality filtering.

evaluationengineering
NIAID-BRC-Codeathons/bvbrc-copilot-workflow-automation
未知语言 · 2026-08-13 评测基准 应用 实验 Stars 0 周增 +0

推进并系统评估 BV-BRC Copilot,从研究问题和已发表方法生成、执行并复现生物信息学工作流。Advancing and systematically evaluating BV-BRC Copilot for generating, executing, and reproducing bioinformatics workflows from research questions and published methods.

ljestaciocerquin/ai-recist-systematic-review
未知语言 · 2026-08-25 评测基准 应用 实验 Stars 0 周增 +0

仓库包含文章与 Excel 文件,收录系统综述《人工智能方法在实体瘤疗效评价标准中的应用》中的 CLAIM 与 FUTURE-AI 评估。Repository containing articles and Excel files with CLAIM and FUTURE-AI assessments for the systematic review Artificial intelligence approaches to the Response Evaluation Criteria in Solid Tumors: a systematic review.

evaluation
hirsaeki/research-loom
Python · 2026-08-25 评测基准 应用 实验 Stars 0 周增 +0

Research Loom 是 profile 驱动的 research harness,用于维持从研究问题与证据到发现、写作与发表的全链路可追溯性Research Loom is a profile-driven research harness for maintaining traceability from research questions and evidence through findings, writing, and publication.

Ginjet/Intelligent-Decision-Driven-Literature-Review-Assistant.
Python · 2026-08-16 评测基准 应用 实验 Stars 0 周增 +0

面向本地学术文档研究:用户对多篇PDF/Markdown论文提问,Agent自主决定检索范围、读取文本/页面/图表证据、判断证据是否充分并生成带引用回答;过程实时可看、结果可离线评测。

agent
fmadore/IWAC-sentiment-analysis
Svelte · 2026-08-12 评测基准 应用 实验 Stars 0 周增 +0

对伊斯兰西非文献集(IWAC)语料库情感分析的交互式可视化,对比 ChatGPT、Gemini 与 Mistral,支持多语言与高级筛选。Interactive visualization of sentiment analysis on the Islam West Africa Collection (IWAC) corpus, comparing ChatGPT, Gemini, and Mistral with multilingual support and advanced filtering.

evaluationllm-infra
Dedebanded912/course-eligibility-hub
HTML · 2026-08-24 评测基准 应用 实验 Stars 0 周增 +0

即时评估课程资格,输出明确的通过/未通过结果以及定制化的入学测试。Instantly evaluate course eligibility with clear pass/fail results and tailored entry assessments.

multimodaldatabaseriskllm-infra
cabbi-bio/miscanthus-yield-maps-review-viz
JavaScript · 2026-08-17 评测基准 应用 实验 Stars 0 周增 +0

GitHub Page 演示:服务于 2026 年文献综述的 Miscanthus 产量地图。Demonstration of GitHub Page with Miscanthus Yield mapping for Literature Review 2026

anas1412/orb-mt5
HTML · 2026-08-23 评测基准 应用 实验 Stars 0 周增 +0

面向 MetaTrader 5 的可配置开盘区间突破 EA,附带可复现的研究 harness;所有参数均为输入项,兼容 Windows 与 Linux/WineConfigurable opening-range breakout EA for MetaTrader 5, with a reproducible research harness. Every parameter is an input; runs on Windows and Linux/Wine.

Abdulrahman-Albeladi/research-toolkit
Python · 2026-08-22 评测基准 应用 实验 Stars 0 周增 +0

可复用的 Python 与 notebook 工作流,用于多语言科研数据准备、统计分析及 AI 文本检测器评估。Reusable Python and notebook workflows for multilingual research-data preparation, statistical analysis, and AI-text detector evaluation.

evaluation