arXiv cs.AI发布于 08/26 12:00

A survey detection channel overrides the pixels in an astronomical foundation model, and biases tomographic mean redshifts

(翻译)巡天检测通道覆盖天文学基础模型中的像素,并使层析平均红移产生偏置

arXiv:2608.23626v1 Announce Type: new Abstract: Foundation models for astronomy are trained on survey pixels together with the catalogue products derived from those pixels. Those catalogues are incomplete at a measurable rate, and a model trained on both inherits that incompleteness as a systematic.

查看原文
雷峰网发布于 09/03 19:39

AI 群星闪耀之前,商汤先相信那群少年

2018年,王延森还是一名本科生,做生成对话方向。 那时候CV正热,Transformer浪潮初起。告诉别人自己在做“能聊天的机器人”,对方听完往往礼貌地点点头,表情里写着“挺前沿的,但然后呢?”

查看原文
人人都是产品经理发布于 09/14 16:10

Open AI官网首发提示词教程,教你玩转GPT Image 2.5

GPT Image 2.5 上线,生图能力再突破。本文结合官方指南与实测,拆解手绘角色、带字海报、照片换装、线稿转写实、界面生成五大玩法,手把手教你写提示词、调结果,让设计师和产品经理都能快速上手。

查看原文
arXiv cs.AI发布于 09/25 12:00

DEEPO: Dual-Entropy Enhanced Policy Optimization for Hallucination in MLLMs

(翻译)DEEPO:面向多模态大语言模型幻觉的双熵增强策略优化

arXiv:2609.28570v1 Announce Type: new Abstract: Reinforcement learning (RL) is widely used to sharpen reasoning in multimodal large language models (MLLMs), yet its effect on hallucination is uneven. We trace this to two weak points in the \emph{correction chain} from reward to parameter update. At

查看原文
雷峰网发布于 09/30 11:50

被质疑「容易复制」的视触觉,到底难在哪?

    作者丨向   欣     编辑丨高景辉                                                                        

查看原文
人人都是产品经理发布于 10/10 10:31

一个没卖出去的 AI 产品,我决定把它开源

给一家做装饰的工作室做的 AI 产品,最后没被卖出去,作者决定把它开源。思路只有四步:把过去散落在相册、聊天记录里的作品图回捞汇总,交给视觉模型自动打标签,客户发来新图先拆标签再从案例库里找最相似的几张。识图不难,难的是规范。

查看原文