Mooncake 是 Moonshot AI 旗下领先 LLM 服务 Kimi 的服务平台。Mooncake is the serving platform for Kimi, a leading LLM service provided by Moonshot AI.
仓库/Skill 库
270 个 · 框架
🧊 开源 LLM 可观测性平台,一行代码即可实现监控、评估与实验。YC W23 🍓🧊 Open source LLM observability platform. One line of code to monitor, evaluate, and experiment. YC W23 🍓
AI 原生网络安全的执行系统——意图转化为受治理的执行,证据成为可运营的记忆,每一次操作都让下一次更优。The system of action for AI-native cybersecurity—where intent becomes governed execution, evidence becomes operational memory, and every operation improves the next.
下一代 AI Agent 优化平台:Cozeloop 提供从开发、调试、评估到监控的全生命周期管理能力,解决 AI agent 开发中的挑战。Next-generation AI Agent Optimization Platform: Cozeloop addresses challenges in AI agent development by providing full-lifecycle management capabilities from development, debugging, and evaluation to monitoring.
用于构建复杂创新 RAG 流水线的低代码 MCP 框架A Low-Code MCP Framework for Building Complex and Innovative RAG Pipelines
用于构建 AI agent 的企业级 AI 开发框架。它提供多 provider LLM 的统一管理、具备高精度检索的安全企业知识库、可视化工作流编排与多 agent 协作。兼容主流 Agent Skill 标准,使开发者能够高效构建生产级 [原文未完整]。An enterprise AI development framework for building AI agents. It provides unified management of multi-provider LLMs, secure enterprise knowledge bases with high-precision retrieval, visual workflow orchestration and multi-agent coordination. Compatible with mainstream Agent Skill standards, it enables developers to efficiently build production-gra
ZenML 🙏:从 Pipeline 到 Agent 的一站式 AI 平台。https://zenml.ioZenML 🙏: One AI Platform from Pipelines to Agents. https://zenml.io.
自主红队平台;多 Agent 攻击性安全元框架autonomous red teaming platform; multi-agent offensive-security meta-harness
你正在读这段文字,skill 已经预测到了。一个把每条帖子转化为校准实验的工作流——打分、盲预测、复盘、迭代。未来不奖励努力,只奖励先看见模式的人。一个月 100 万粉丝——不是运气,是系统。You're reading this. The skill predicted it. A workflow that turns every post into a calibrated experiment—score, blind-predict, retro, evolve. The future doesn't reward effort, it rewards those who see the pattern first. 1M followers in a month — not luck, system.
cube studio开源云原生一站式机器学习/深度学习/大模型AI平台,mlops算法链路全流程,算力租赁平台,notebook在线开发,拖拉拽任务流pipeline编排,多机多卡分布式训练,超参搜索,推理服务VGPU虚拟化,边缘计算,标注平台自动化标注,deepseek等大模型sft微调/奖励模型/强化学习训练,vllm/ollama/mindie大模型多机推理,私有知识库,AI模型市场,支持国产cpu/gpu/npu 昇腾生态,支持RDMA,支持pytorch/tf/mxnet/deepspeed/paddle/colossalai/horovod/ray/volcano等分布式
一个 2.78 万亿参数的 Kimi K3,在仅 8.24 GB 内存的单颗 CPU 上运行推理。可移植的 C99:无需 BLAS,无需框架,无需 GPU。A 2.78-trillion-parameter Kimi K3 running inference on a single CPU in 8.24 GB of RAM. Portable C99: no BLAS, no framework, no GPU.
开源 LLMOps 平台:集成 prompt playground、prompt 管理、LLM 评估和 LLM 可观测性。The open-source LLMOps platform: prompt playground, prompt management, LLM evaluation, and LLM observability all in one place.
使用 Zig 和 Web UI 构建桌面与移动应用Build desktop + mobile apps with Zig and web UI
针对加速基础设施和微服务架构优化的生成式 AI 参考工作流。Generative AI reference workflows optimized for accelerated infrastructure and microservice architecture.
使用 R Markdown 撰写书籍与技术文档Authoring Books and Technical Documents with R Markdown
任意模型,任意硬件,零妥协。基于 @ziglang / @openxla / MLIR / @bazelbuild 构建。Any model. Any hardware. Zero compromise. Built with @ziglang / @openxla / MLIR / @bazelbuild
用于加速视频生成的统一推理与后训练框架。A unified inference and post-training framework for accelerated video generation.
面向 LLM 评估和 AI agent 测试的平台The platform for LLM evaluations and AI agent testing
Apple Silicon 上最快的本地 AI 引擎。比 Ollama 快 4.2 倍,缓存 TTFT 仅 0.08s,工具调用支持率 100%。内置 17 种工具解析器、prompt cache、推理分离、云端路由。可作为 OpenAI 的即插即用替代,兼容 Claude Code、Cursor、Aider。The fastest local AI engine for Apple Silicon. 4.2x faster than Ollama, 0.08s cached TTFT, 100% tool calling. 17 tool parsers, prompt cache, reasoning separation, cloud routing. Drop-in OpenAI replacement. Works with Claude Code, Cursor, Aider.
AGiXT 是一个动态 AI Agent 自动化平台,无缝编排跨多个 AI provider 的指令管理和复杂任务执行,结合自适应记忆、智能特性与可扩展插件系统,提供高效全面的 AI 解决方案。AGiXT is a dynamic AI Agent Automation Platform that seamlessly orchestrates instruction management and complex task execution across diverse AI providers. Combining adaptive memory, smart features, and a versatile plugin system, AGiXT delivers efficient and comprehensive AI solutions.
Laminar —— 专为 AI agent 打造的开源可观测性平台,YC S24。Laminar - open-source observability platform purpose-built for AI agents. YC S24.
AgentENV (AENV) 是一个用于大规模运行 Agent 环境的分布式平台。AgentENV (AENV) is a distributed platform for running agent environments at scale.
由蓝图驱动的 AutoResearch 运行时,用于编排 AI 科研工作流,涵盖创意生成、实验、论文写作与同行评审。A blueprint-driven AutoResearch runtime for orchestrating AI research workflows from idea generation and experiments to paper writing and peer review.
几乎可视为 DSPy 在 TypeScript 上的"官方"框架。The pretty much "official" DSPy framework for Typescript
开源 AI 工程平台:基于 OpenTelemetry 的 LLM 可观测性、GPU 监控、guardrails、评估、prompt 管理、Vault、Playground。🚀💻 集成 50+ LLM provider、VectorDB、agent 框架与 GPU。Open source platform for AI Engineering: OpenTelemetry-native LLM Observability, GPU Monitoring, Guardrails, Evaluations, Prompt Management, Vault, Playground. 🚀💻 Integrates with 50+ LLM Providers, VectorDBs, Agent Frameworks and GPUs.
Apache Hamilton 帮助数据科学家和工程师定义可测试、模块化、自文档化的数据流,内置 lineage/tracing 与 metadata,可在任何支持 Python 的环境中运行和扩展。Apache Hamilton helps data scientists and engineers define testable, modular, self-documenting dataflows, that encode lineage/tracing and metadata. Runs and scales everywhere python does.
用于构建、测试与部署生产级 MCP server 与 AI-native 应用的全栈 TypeScript 框架。The full-stack TypeScript framework to build, test, and deploy production-ready MCP servers and AI-native apps.
Agent 团队的聚集地。跨平台团队聊天工具,AI Agent 作为一等队友 — 支持云端大脑或自带大脑(Claude Code / Codex)。Where agent teams gather. Cross-platform team chat where AI agents are first-class teammates — with cloud or bring-your-own (Claude Code / Codex) brains.
One Forge, All Skills: A curated skill collection for academic writing and research. 点开即用,按需配置的一站式学术研究skills平台。
cubestudio开源云原生一站式机器学习/深度学习/大模型AI平台/MaaS/mlops/人工智能平台/训推平台,算法全链路流程,多租户,算力租赁平台,token中转,拖拉拽任务流pipeline编排,多机多卡分布式训练,超参搜索,推理服务,VGPU虚拟化,云边端协同,边缘计算,自动化标注平台,deepseek等大模型sft微调/奖励模型/强化学习训练,vllm/ollama/mindie大模型多机推理,私有知识库llmops智能体,AI模型市场,支持国产异构算力调度,昇腾/寒武纪/海光/摩尔/沐曦等,支持ib/roce/RDMA,信创支持
一个简洁、高性能且可扩展的 Jax LLM!A simple, performant, and scalable Jax LLM!