Transformers:面向文本、视觉、音频及多模态 SOTA 机器学习模型的模型定义框架,同时支持推理与训练。🤗 Transformers: the model-definition framework for state-of-the-art machine learning models in text, vision, audio, and multimodal models, for both inference and training.
仓库/Skill 库
53 个 · 模型 · AI 核心
开源超级 AI 助手与 Agent 编排框架。任务规划、工具与 Skill 调用、基于记忆和知识的自演化。多模型、多通道、轻量、可扩展,一行安装。(原名 chatgpt-on-wechat)Open-source super AI assistant & Agent Harness. Plans tasks, runs tools and skills, self-evolves with memory and knowledge. Multi-model, multi-channel. Lightweight, extensible, one-line install. (formerly chatgpt-on-wechat)
FinGPT:开源金融大语言模型!革命性 🔥 已在 HuggingFace 发布训练好的模型。FinGPT: Open-Source Financial Large Language Models! Revolutionize 🔥 We release the trained model on HuggingFace.
中文LLaMA&Alpaca大语言模型+本地CPU/GPU训练部署 (Chinese LLaMA & Alpaca LLMs)
面向 agents、助手和企业搜索的私有 AI 平台,内置 Agent Builder、深度研究、文档分析、多模型支持,以及为 agents 提供 API 连接能力。Private AI platform for agents, assistants and enterprise search. Built-in Agent Builder, Deep research, Document analysis, Multi-model support, and API connectivity for agents.
💡 面向语义搜索、LLM 编排与语言模型工作流的一体化 AI 框架💡 All-in-one AI framework for semantic search, LLM orchestration and language model workflows
复旦大学开源的工具增强对话语言模型An open-source tool-augmented conversational language model from Fudan University
面向本地部署的高速 LLM 服务High-speed Large Language Model Serving for Local Deployment
提供 AI 应用和模型服务的最简方式 —— 构建模型推理 API、任务队列、LLM 应用、多模型 pipeline 等。The easiest way to serve AI apps and models - Build Model Inference APIs, Job queues, LLM apps, Multi-model pipelines, and more!
LTX-2 音视频生成模型的官方 Python 推理与 LoRA 训练包。Official Python inference and LoRA trainer package for the LTX-2 audio–video generative model.
中文LLaMA-2 & Alpaca-2大模型二期项目 + 64K超长上下文模型 (Chinese LLaMA-2 & Alpaca-2 LLMs with 64K long context models)
面向全模态模型的高效推理框架。A framework for efficient model inference with omni-modality models
用于高性能 AI 模型服务(vLLM、SGLang)和按需 SSH 访问 GPU 实例的 GPU 集群管理器。A GPU cluster manager for high-performance AI model serving (vLLM, SGLang) and on-demand SSH-accessible GPU instances.
高质量且极速的预训练深度学习模型与 demo。Pre-trained Deep Learning models and demos (high quality and extremely fast)
SOTA 低比特 LLM 量化(INT8/FP8/MXFP8/INT4/MXFP4/NVFP4)与稀疏化方案;面向 PyTorch、TensorFlow 与 ONNX Runtime 的领先模型压缩技术SOTA low-bit LLM quantization (INT8/FP8/MXFP8/INT4/MXFP4/NVFP4) & sparsity; leading model compression techniques on PyTorch, TensorFlow, and ONNX Runtime
潜入 Shell 的幽灵。Ante 是一个自包含的 Agent 框架,核心高度优化。体验类似 Claude Code 或 Codex,但无其依赖与模型限制。Ghost in your shell. Ante is a self-contained agent harness with a highly optimized core. It works like Claude Code or Codex, with none of their dependencies or model constraints.
开源 LLM/VLM 负载均衡器与服务平台,用于规模化自托管 LLM(和 VLM)🏓🦙 作为 llm-d、Docker Model Runner 等项目的替代方案,组件更少、部署更简单,基于 ggml 生态构建。支持 CPU 和 GPU。Open-source LLM/VLM load balancer and serving platform for self-hosting LLMs (and VLMs) at scale 🏓🦙 Alternative to projects like llm-d, Docker Model Runner, etc but with less moving parts and simple deployments built around ggml ecosystem. Runs on CPU and GPU.
[GenAI 应用开发框架] 🚀 快速轻松构建 GenAI 应用 💬 在代码中使用结构化数据和链式调用语法与 GenAI Agent 交互 🧩 使用事件驱动的 *TriggerFlow* 管理复杂的 GenAI 业务逻辑 🔀 无需重写代码即可切换任意模型[GenAI Application Development Framework] 🚀 Build GenAI application quick and easy 💬 Easy to interact with GenAI agent in code using structure data and chained-calls syntax 🧩 Use Event-Driven Flow *TriggerFlow* to manage complex GenAI working logic 🔀 Switch to any model without rewrite application code
论文"LAMBDA: A large Model Based Data Agent"的官方仓库。https://www.polyu.edu.hk/ama/cmfai/lambda.htmlThis is the offical repository of paper "LAMBDA: A large Model Based Data Agent". https://www.polyu.edu.hk/ama/cmfai/lambda.html
开源核心的 AI 工作台——覆盖任何模型的 notebooks、Agent、RAG、语音和图像:OpenAI、Anthropic、Google、xAI,或通过 Ollama/vLLM 的本地模型。BSL 1.1,两年后自动转换为 Apache-2.0。当别人的 AI 停止运行时,你的 AI 仍在运行。The open-core AI workbench — notebooks, agents, RAG, voice, and images across any model: OpenAI, Anthropic, Google, xAI, or local via Ollama/vLLM. BSL 1.1, auto-converting to Apache-2.0 on a two-year clock. Your AI keeps running when theirs doesn't.
发布 AI Agent 制品,而非 demo 表演 — 验证器把关、跨模型族运行,自动学习哪个模型最优。Ship AI-agent artifacts, not demo theater — verifier-gated, cross-family runs that learn which model wins.
[CoLM 2026] MoANT 官方代码:面向多任务大语言模型微调的语义感知秩一专家混合模型[CoLM 2026] Official code for MoANT: Mixture-of-Rank-One-Experts with semantic-aware Intuition for Multi-task Large Language Model Finetuning
面向 AI agent 的基于证据的评估——将每条断言与 agent 真实工具输出进行核对(受约束、基于证据的模型判断,而非整体式 LLM 评判的猜测),并附带置信区间。Evidence-grounded evaluation for AI agents — verifies each claim against the agent's real tool outputs (constrained, evidence-grounded model judgment, not holistic LLM-judge guesswork), with confidence intervals.
本地 Windows 桌面监视器,用于查看正在运行的 Claude Code agent —— 实时状态、预估成本、模型、主机以及一键聚焦,按项目分组。只读且完全离线。Local Windows desktop monitor for your running Claude Code agents - live status, estimated cost, model, host and one-click focus, grouped by project. Read-only and fully offline.
SharpAI 是基于 llama.cpp(通过 LlamaSharp)构建的可嵌入 Embedding、补全与模型管理平台,内置 Ollama 兼容的 Web 服务。SharpAI is an embeddable embeddings, completions, and model management platform using llama.cpp via LlamaSharp, with a built-in Ollama-compatible webserver.
运行 Ollama 本地 LLM 服务的 Docker 镜像。默认安全,所有 API 请求需 Bearer token(首次启动时自动生成)。OpenAI 兼容 API。支持首次启动模型预拉取、NVIDIA GPU (CUDA) 加速和持久化模型存储。多架构:amd64、arm64。Docker image to run an Ollama local LLM server. Secure by default, all API requests require a Bearer token (auto-generated on first start). OpenAI-compatible API. Supports first-start model pre-pull, NVIDIA GPU (CUDA) acceleration, and persistent model storage. Multi-arch: amd64, arm64.
Autobots | 由 NVIDIA NIM 驱动的去中心化模型 agent 集群。通过专用 6 文件控制架构编排高精度、端到端可鉴权软件开发流程,具备自动化安全审计、无冗余任务路由与自主状态管理能力。Autobots | A decentralized model agentic swarm powered by NVIDIA NIM. Orchestrating high-precision, end-to-end authenticated software development through a specialized 6-file control architecture. Features automated security auditing, non-redundant task routing, and autonomous state management.
Hearting —— 面向 Claude Code、Codex 与 OpenCode 的 agent 设置。一份可移植契约覆盖三者,支持能力路由、跨运行时密封调度、节点级模型分层与实时 Fleet 视图,前身为 agent_settingHearting — the agent setting for Claude Code, Codex, and OpenCode. One portable contract projected onto all three, with routed capabilities, sealed cross-harness dispatch, per-node model tiers, and a live Fleet view. Formerly agent_setting.
邮件的"大脑":MailFathom 将 IMAP 邮箱转化为自托管、AI 原生服务。邮件同步至自有 PostgreSQL,建立索引以支持搜索与检索,并通过 Model Context Protocol 服务于 AI Agent。当前为只读;后续将支持语义检索、问答与受控写入工具。.NET 10,Apache-2.0。A brain for your mail: MailFathom turns IMAP mailboxes into a self-hosted, AI-native service. Mail synchronizes into your own PostgreSQL, is indexed for search and retrieval, and is served to AI agents over the Model Context Protocol. Read-only today; semantic retrieval, answering, and gated write tools next. .NET 10, Apache-2.0.
你的 agent 能写出变更——却无法告诉你还有哪些内容依赖于它正在修改的代码。本工具从你自己的代码构建产品模型,然后对每个任务打分:变更内容、影响范围、波及的用户旅程与敏感数据——让运算结果直观呈现在屏幕上。本地运行,零依赖,无遥测。包含 MCP。Your agent can write the change — it can't tell you what else depends on the code it's touching. This builds a model of your product from your own code, then scores every task: what it changes, what that reaches, which user journeys and sensitive data are in the blast radius — arithmetic on screen. Local, zero deps, no telemetry. MCP included.
自托管、多用户转录平台:录制或上传音频。支持说话人标记、带时间戳的转录、跨录音识别说话人、摘要、提取行动项,并可使用自有 OpenAI 兼容 LLM 与转录内容对话。你的音频、你的服务器、你的模型。已在笔记本 RTX4070、台式机 RTX3090 与 RTX5090 上测试。Self-hosted, multi-user transcription platform: record or upload audio. Speaker-labeled, timestamped transcripts, Recognize speakers across recordings, Summarize, extract action items and chat over your transcripts with your own OpenAI-compatible LLM. Your Audio, your Server, your Model. Tested on Laptop RTX4070, Desktop RTX3090 and RTX5090
从零开始构建开源预训练 LLM,附带已发布的数据、训练代码、消融实验与结果,以高效提升模型质量。Build open pretraining LLMs from scratch with released data, training code, ablations, and results to improve model quality efficiently
用知识图谱刻画 AI/ML 模型从创建到部署的完整生命周期。Knowledge Graph to capture AI/ML model lifecycle from creation through deployments.
你的 AI Agent 无法掏空的钱包。一个非托管的 Soroban 资金库,限定自主 Agent 的支出上限——由合约而非模型强制执行。The wallet your AI agent can't drain. A non-custodial Soroban treasury that bounds what an autonomous agent may spend - enforced by the contract, not the model.