本文提出一种多任务深度学习的原则性方法,通过考虑各任务的同方差不确定性来加权多个损失函数,从而在分类与回归场景下同时学习具有不同单位或尺度的多种量。A principled approach to multi-task deep learning is proposed which weighs multiple loss functions by considering the homoscedastic uncertainty of each task, allowing us to simultaneously learn various quantities with different units or scales in both classification and regression settings.
新一代 AI 研究与工程知识大盘
从热点资讯进入今天,从深度解读、技术趋势与实操攻略进入可复用知识体系。全天候自主追踪、提炼与评估。
今日热点 AI Hot · 2026-10-09
全部 15 条Tibo 宣布发布 ChatGPT — X:Tibo (@thsottiaux)(3 小时前) Tibo 宣布 ChatGPT 发布,并给出访问地址 https://chatgpt.com/。推文本身仅含发布声明和链接,未提供更多功能或细节。 https://x.com/thsottiaux/status/2108349826727588000
02Claude 九月回顾:Chat 与 Cowork 合一,Claude 5.5 系列模型上线 — Claude:YouTube(RSS)(12 小时前) Claude 发布九月更新回顾,Chat 与 Cowork 合并为一个 Claude,工作流可在云端运行,合上笔记本后继续进行。Claude Docs 支持团队与 Claude 在同一文档中协作编辑,新推出的 Claude 5.5 系列提供 Opus 5.5 处理重任务、Sonnet 5.5 用于快速修改,并可通过 /slides、/docs、/designs 直接生成对应格式。 https://www.youtube.com/shorts/n9WfoNW2XnE
03OpenAI 年化营收被曝接近 500 亿美元,与此前预期差距约 200 亿美元 — IT之家·人工智能(4 小时前) 据金融时报报道,OpenAI 向投资者披露截至 9 月底年化营收逼近 500 亿美元,大幅低于此前外界估算的 700 亿美元,缺口约 200 亿美元。差异源于统计口径不同,Anthropic 计入 AWS 和谷歌云等合作方销售收入,而 OpenAI 剔除该部分。报道发布后美股科技股下跌,纳斯达克 100 指数收跌 1.4%,英伟达跌 2.9%,甲骨文跌 5.5%。 https://www.ithome.com/1/010/751.htm
本周趋势 Top 5
完整趋势榜最新解读 Explainers
全部攻略上新 Guides
全部主题热度 最近 24h
2026-10-09 摘要最新笔记 Notes
全部近期论文卡 Papers
全部Contrastive Captioner (CoCa) 采用极简设计,对图文编码器-解码器基础模型联合使用对比损失与字幕损失进行预训练,从而兼具 CLIP 等对比方法与 SimVLM 等生成方法的能力。Contrastive Captioner (CoCa), a minimalist design to pretrain an image-text encoder-decoder foundation model jointly with contrastive loss and captioning loss, thereby subsuming model capabilities from contrastive approaches like CLIP and generative methods like SimVLM.
本文对度量学习文献进行了系统综述,阐述了每种方法的优缺点,并介绍了近期涌现的一系列强大替代方法,包括非线性度量学习、相似性学习与局部度量学习。A systematic review of the metric learning literature is proposed, highlighting the pros and cons of each approach and presenting a wide range of methods that have recently emerged as powerful alternatives, including nonlinear metric learning, similarity learning and local metric learning.
提出极简的 VLP 模型 Vision-and-Language Transformer (ViLT),其一体化设计将视觉输入处理大幅简化为与文本输入相同的无卷积方式;ViLT 比此前的 VLP 模型快达数十倍,同时下游任务性能具有竞争力甚至更优。A minimal VLP model, Vision-and-Language Transformer (ViLT), monolithic in the sense that the processing of visual inputs is drastically simplified to just the same convolution-free manner that the authors process textual inputs, showing that ViLT is up to tens of times faster than previous VLP models, yet with competitive or better downstream task performance.
在 BIG-bench 上对 OpenAI 的 GPT 模型、Google 内部稠密 Transformer 架构及 Switch 风格稀疏 Transformer 进行评估,模型规模跨越百万至千亿参数,结果显示性能与校准均随规模提升而改善,但绝对水平仍然欠佳。Evaluation of OpenAI's GPT models, Google-internal dense transformer architectures, and Switch-style sparse transformers on BIG-bench, across model sizes spanning millions to hundreds of billions of parameters finds that model performance and calibration both improve with scale, but are poor in absolute terms.
本文提出一种方法,通过多层感知机的一次前馈传播生成独立样本(与近期提出的 GAN 类似),并使用 MMD 学习生成可被解码为样本的 codes。This work forms a method that generates an independent sample via a single feedforward pass through a multilayer perceptron, as in the recently proposed generative adversarial networks, using MMD to learn to generate codes that can then be decoded to produce samples.
高价值仓库 Repositories
全部😎 涵盖各类有趣主题的 Awesome 列表精选😎 Awesome lists about all kinds of interesting topics
一份精选的 Python 框架、库、工具和资源列表An opinionated list of Python frameworks, libraries, tools, and resources
一份可在自有服务器上自托管的自由软件网络服务和 Web 应用清单A list of Free Software network services and web applications which can be hosted on your own servers
一个切实可用的 agentic skills 框架与软件开发方法论An agentic skills framework & software development methodology that works.
Agent harness 性能优化系统。为 Claude Code、Codex、Opencode、Cursor 等提供 Skill、本能、记忆、安全性与研究优先的开发能力。The agent harness performance optimization system. Skills, instincts, memory, security, and research-first development for Claude Code, Codex, Opencode, Cursor and beyond.