一款隐私优先、可自托管、完全开源的个人知识管理软件,使用 typescript 和 golang 编写。An open-source, privacy-first, self-hosted knowledge workspace where humans and AI agents work together 开源、隐私优先、自托管的知识工作空间,让人与智能体在此协作
仓库/Skill 库
38 个 · 数据与向量库 · 应用
🤖 与你的 SQL 数据库对话 📊。通过 Agentic 检索 🔄 利用 LLM 实现准确的 Text-to-SQL 生成。🤖 Chat with your SQL database 📊. Accurate Text-to-SQL Generation via LLMs using Agentic Retrieval 🔄.
面向 AI agents 的 GenBI(Generative BI),一款开源、可治理的 text-to-SQL 方案,通过开放 context layer 将自然语言问题转化为可信的仪表盘、图表与 SQL,覆盖 20+ 数据源,如 BigQuery、Snowflake、PostgreSQL、ClickHouse、Amazon Redshift、Databricks 等GenBI (Generative BI) for AI agents, an open-source, governed text-to-SQL through an open context layer that turns natural-language questions into trusted dashboards, charts, and SQL across 20+ data sources, such as BigQuery, Snowflake, PostgreSQL, ClickHouse, Amazon Redshift, Databricks and more.
一个轻量、极速的进程内向量数据库。A lightweight, lightning-fast, in-process vector database
开源全球情报平台——实时 OSINT 仪表盘——Palantir 替代方案——按交易量抽取 0.5%,钱包地址 2nZNHm3Lr9umG3DVrzYwHgktwkuKuJRXqqRqs3ewpumpOpen Source Global Intelligence Platform - Real-Time OSINT Dashboard - A Palantir Alternative - We Get 0.5% on Volume Traded 2nZNHm3Lr9umG3DVrzYwHgktwkuKuJRXqqRqs3ewpump
极简、快速的数据库客户端桌面应用。为想要快速查看数据而无需臃肿功能的开发者打造。A minimal, fast, database client desktop application. Built for developers who want to quickly peek at their data without the bloat.
基于 Next.js、LangChain、PostgreSQL + pgvector 构建的 AI 驱动的初创企业加速引擎。可上传、组织并与文档对话。包含预测性缺失文档检测、基于角色的工作流以及页面级洞察提取。AI-powered StartUp Accelerator Engine built with Next.js, LangChain, PostgreSQL + pgvector. Upload, organize, and chat with documents. Includes predictive missing-document detection, role-based workflows, and page-level insight extraction.
为 Claude Code 提供持久化记忆,支持自带数据库Persistent memory for Claude Code. Bring your own database.
面向 agentic AI 与受治理分析的源代码可用的语义与上下文层、规则引擎和语义 sidecar;可将声明式 YAML 模型编译为优化的 SQL、KPI 和语义上下文,覆盖 8 种 SQL 方言Source-available semantic and context layer, rule engine, and semantic sidecar for agentic AI and governed analytics. Compiles declarative YAML models into optimized SQL, KPIs, and semantic context across 8 SQL dialects.
基于 Parquet 构建的嵌入式检索库。快速、高效且可扩展。Embedded retrieval library built on Parquet. Fast, efficient, and scalable.
隐私优先、自托管的加密资产记账应用,零云端数据共享。聚合交易所、钱包及多链历史,计算 FIFO/LIFO/HIFO 成本基础、已实现/未实现盈亏与资本利得,跟踪质押/空投收入,并导出可报税的 CSV/PDF/JSON 报告。加密 SQLite。Python,跨平台。Privacy-first, self-hosted crypto accounting app with zero cloud data sharing. Aggregates exchange, wallet and multi-chain history, calculates FIFO/LIFO/HIFO cost basis, realized/unrealized PnL and capital gains, tracks staking/airdrop income, and exports tax-ready CSV/PDF/JSON reports. Encrypted SQLite. Python, cross-platform.
在 Apple Silicon 上运行 Windows 游戏 —— 免费、开源,并附带开放兼容数据库,告知你哪些游戏真正可用。Run Windows games on Apple Silicon — free, open source, with an open compatibility database that tells you what actually works.
一款面向学术写作的原生 macOS Markdown 编辑器,集成了兼容 Zotero 的文献管理器、表格数据库与任务追踪功能。A native macOS Markdown editor for academic writing, featuring an integrated Zotero-friendly bibliography manager, spreadsheet databases, and task tracking.
一次查询,覆盖全部学术数据库;每次运行归档井然有序。跨 OpenAlex、ADS、arXiv、INSPIRE、Scopus、Semantic Scholar、Crossref 与 WoS 的可复现、合规文献检索One query. Every scholarly database. A tidy archive of every run. Reproducible, ToS-compliant literature search across OpenAlex, ADS, arXiv, INSPIRE, Scopus, Semantic Scholar, Crossref and WoS.
语音 AI API 的开放价格数据库 —— STT、LLM、TTS、S2S、VAD。Open price database for voice AI APIs - STT, LLM, TTS, S2S, VAD
🌟 利用 GPU 加速 Apple Silicon 上的 SQLite 数据库操作,提升分析速度与数据管理效率🌟 Accelerate SQLite database operations on Apple Silicon with GPU power for faster analytics and efficient data management.
开放式数据库工程课程 · 14 个模块共 64 节课 · 210 小时 · 涵盖概念建模、分布式架构、运维及面向 AI 的数据恢复 · 109 项参考文献(带 ISBN、DOI 或标准号),逐课引用并在 CI 中校验 · 5 个零依赖可执行实验 · 256 题自测 📚🗄️ Programa abierto de ingeniería de bases de datos · 64 clases en 14 partes · 210 horas · Del modelado conceptual a la arquitectura distribuida, la operación y la recuperación para IA · 109 fuentes con ISBN, DOI o norma, citadas clase a clase y verificadas en CI · 5 laboratorios ejecutables sin dependencias · Autoevaluación de 256 preguntas 📚
用于系统综述检索的交互式工具:PICO 构建器、数据库语法转换器与邻近算子构建器。Interactive tools for systematic review searching: PICO builder, database syntax converter, and proximity-operator builder.
Chat2DB Community 的 Rust 产品运行时,配套 Java 数据库兼容引擎Rust product runtime with Java database compatibility engine for Chat2DB Community
本仓库包含用于复现论文《Pharmacological treatments for ADHD and comorbid substance use disorder: a systematic review and network meta-analysis》中数据库与分析的 R 代码。This repository holds the R codes for reproducing the databases and analyses in my paper "Pharmacological treatments for ADHD and comorbid substance use disorder: a systematic review and network meta-analysis".
GMS 6803 临床研究数据科学课程材料,含幻灯片、代码与实操实验,内容涵盖真实世界健康数据、基于 EHR 的队列构建、SQL、Python、统计分析与临床研究中的机器学习Course materials for GMS 6803 Data Science for Clinical Research, including slides, code, and hands-on labs on real-world health data, EHR-based cohort development, SQL, Python, statistical analysis, and machine learning for clinical research.
关于精神病学障碍中迷幻疗法的活体系统综述,基于 Docker 化 Dash 应用与 PostgreSQL 后端构建。Living systematic review of psychedelic therapy in psychiatric disorders. Built as a dockerized Dash app with PostgreSQL backend.
开源 DBMS 生态的 Reference Quotient 分析与可复现研究流水线。项目列表每月通过 https://github.com/birdflyi/od_label_issue_gen/tree/main/data/database_repo_label_dataframe 中的 csv 文件更新Reference Quotient analysis and reproducible research pipeline for the open-source DBMS ecosystem. The projects list are updated in the csv files in https://github.com/birdflyi/od_label_issue_gen/tree/main/data/database_repo_label_dataframe monthly.
自动化 Python 兼容性补丁工具,将 Google Scholar PDF Reader 从 Chromium 移植到基于 Firefox 的浏览器,并使用 IndexedDB 集成自定义离线论文数据库。Automated Python compatibility patcher to port Google Scholar PDF Reader from Chromium to Firefox-based browsers, integrating a custom offline paper database using IndexedDB.
针对任意 GitHub 仓库的自然语言问答 —— 构建 Neo4j 依赖图与 Qdrant 向量索引,用自然语言即可提问结构性与行为性问题。Natural-language Q&A over any GitHub repo — builds a Neo4j dependency graph + Qdrant vector index so you can ask structural and behavioral questions in plain English
从学术数据库中采集学术元数据,用于组织文献检索与构建研究参考目录。Harvest scholarly metadata from academic databases to organize literature searches and build research bibliographies.
"Domain-Driven Design in Practice: A Large-Scale Empirical Characterisation of the Open-Source Ecosystem" 的复现包——包含数据收集工具、GPT-4o 语义验证流水线,以及针对所有研究问题的 SQL 查询。Replication package for "Domain-Driven Design in Practice: A Large-Scale Empirical Characterisation of the Open-Source Ecosystem" — includes data collection tools, GPT-4o semantic validation pipeline, and SQL queries for all research questions.
Mascarene Paleo 是记录马斯卡林群岛古生物与亚化石记录的研发基础设施。该仓库旨在提供关于标本、分类群、地点、出版物、历史资料、数字资源及未解研究问题的透明、可更新、可引用的记录。Mascarene Paleo is a developing research infrastructure for documenting the palaeontological and subfossil record of the Mascarene Islands. This repository is intended to provide a transparent, updateable, and citable record of specimens, taxa, localities, publications, historical sources, digital resources, and unresolved research questions.
von der Heyde, L., Keusch, F., Buskirk, T. D., & Eck, A. (2026). AI in the Loop?! A Systematic Review of the Use of Large Language Models in Survey and Public Opinion Research. SocArXiv. https://doi.org/10.31235/osf.io/eubj4_v1 的复现材料与文献数据库Replication materials and literature databases for von der Heyde, L., Keusch, F., Buskirk, T. D., & Eck, A. (2026). AI in the Loop?! A Systematic Review of the Use of Large Language Models in Survey and Public Opinion Research. SocArXiv. https://doi.org/10.31235/osf.io/eubj4_v1.
一套可移植、可复现的 PRISMA 2020 系统综述流水线:检索 15 个数据库、执行去重、以完整审计追踪管理人工筛选,并根据决策日志计算 PRISMA 流程计数。A portable, reproducible PRISMA 2020 systematic literature review pipeline. Searches 15 databases, deduplicates, manages human screening with a full audit trail, and computes PRISMA flow counts from the decision log.
引文管理 SaaS:在 1000+ 目录中管理商家信息。Uberall 竞品。Citation Management SaaS: Manage business listings across 1000+ directories. Uberall competitor.
纯 Go 实现的分布式列式 SQL 分析引擎——对 S3 上 Parquet 进行向量化执行,无需 JVM、无需 CGo。以 Trino 级别的能力提供相当性能,却只需极小资源占用。原生支持 IPv4/CIDR/MAC 等网络类型,适用于遥测和安全场景。Distributed columnar SQL analytics engine in pure Go — vectorized execution over Parquet on S3, no JVM, no CGo. Trino-class performance on a fraction of the footprint. Network-native IPv4/CIDR/MAC types for telemetry and security workloads.
将研究压缩包转化为可搜索、可引用的素材。拆分为论文、代码与数据集,索引后即可提出 Deep Research 问题,并引用页码与源行。Self-hosted(Postgres、MinIO、Qdrant)。Turn research zips into searchable, citable materials. Split papers, code, and datasets; index them; ask Deep Research questions that cite pages and source lines. Self-hosted (Postgres, MinIO, Qdrant).