Aily.to

🔥 AI热点

实时追踪 AI 行业热点动态 · 共 100 条资讯

2026年8月6日
20 条
AIHOT 🔥 50

左右两派罕见达成一致:反对数据中心

The Verge 政策记者 Gaby Del Valle 报道,美国两党民众正联合反对 AI 数据中心建设,佛罗里达州 Hernando County 上月一致通过为期一年的建设禁令。抗议者担忧地下水污染、PFAS 及当地环境不适配,保守派组织 Humans First 成为核心力量。数据中心争议正打破传统党派界限,并可能影响中期选举政治格局。 🔗 阅读原文 via AI HOT · https://aihot.virxact.com/items/cmshmiaz00k2zronkzukabc5z

0 阅读原文
ArXiv 🔥 50

论文:Beyond Top-K: Replacing Black-Box Retrieval with Interpretable Agentic Operations

Retrieval-augmented generation over long documents is dominated by one design: chunk the text, embed the chunks, and surface the top-k nearest neighbours of the query. We argue that for an important class of documents -- financial statements, audit reports, regulatory returns -- this design is structurally unsound, and we make the argument measurable. On a 780-page government financial report, 86.

0 阅读原文
AIHOT 🔥 50

千问全网首发公测,全新 Wan3.0 来了!

阿里千问全网首发公测视频生成模型 Wan3.0,在生成时长、镜头语言、角色真实感及一致性保持等维度全面升级。该模型支持稳定直出 30 秒一镜到底视频,具备导演级镜头与蒙太奇叙事,并强调角色、道具、场景高一致性保持。Wan3.0 主打超高性价比,现已在千问创作(c.qianwen.com)开放体验,千问 app 也将陆续开启。 🔗 阅读原文 via AI HOT · https://aihot.virxact.com/items/cmshjzc390h6cronkkgtujtis

0 阅读原文
ArXiv 🔥 50

论文:HarnessOpt-Bench: Evaluating LLMs at Harness Optimization

As LLMs are increasingly deployed within agentic systems, their capabilities depend not only on the model weights but also on the harness: the prompts, tools, control flow, memory, and orchestration code surrounding them. This makes automated harness optimization -- the iterative and evaluation-guided improvement of a harness by an AI system -- both an important route to improving AI systems and a

0 阅读原文
AIHOT 🔥 50

NVIDIA Cosmos 3:开放世界模型如何推动物理 AI 前沿

NVIDIA 发布 Cosmos 3,一个基于混合 Transformer 架构的开放物理 AI 基础全模态模型,整合视觉推理、世界生成与动作预测。 🔗 阅读原文 via AI HOT · https://aihot.virxact.com/items/cmshjdgsn0gf5ronkj2dwsbwh

0 阅读原文
ArXiv 🔥 50

论文:Bias Analysis of L2 Speaking Assessment Systems Using Concept Activation Vectors

Automatic speaking assessment systems are increasingly deployed in high-stakes settings to mark second language (L2) learners' speaking tests, making it critical to show that their scores depend on speaking proficiency rather than irrelevant speaker attributes such as first language (L1) or age. Transformer-based foundation models have improved the accuracy of these L2 speaking graders, but their

0 阅读原文
AIHOT 🔥 50

面壁智能 AMNESIAC:反向图灵测试 AI 审讯游戏

面壁智能 OpenBMB 在 #BuildSmall 黑客松推出 AMNESIAC,一款反向图灵测试交互游戏:玩家需说服 AI 审讯官 A.M.N. 自己是人类。该游戏由 MiniCPM-o 4.5 驱动实时对话与推理,VoxCPM 生成审讯官语音,并结合摄像头面部表情、脉搏信号与响应计时进行多模态审讯。项目已开源至 HuggingFace。 🔗 阅读原文 via AI HOT · https://aihot.virxact.com/items/cmshjbfzc0gceronkw6dqhrsb

0 阅读原文
ArXiv 🔥 50

论文:On-Policy Self-Distillation without Any Supervision

On-policy (Self-)Distillation (OPD / OPSD) has shown strong potential for post-training large language models (LLMs). However, existing methods still rely heavily on external supervision, including ground-truth signals, environmental feedback, or guidance from larger models, and therefore fall short of genuine "self"-distillation. In this study, we show that on-policy self-distillation can be achi

0 阅读原文
AIHOT 🔥 50

谷歌地图 Ask Maps 智能体升级:可对话订餐、找酒店并接入 Gemini Personal Intelligence

谷歌地图宣布 Ask Maps 迎来新一轮升级,新增智能体功能,可替用户执行订餐操作,并综合考虑饮食要求、当前位置和收藏地点等信息;用户还可通过对话指定装修风格、环境氛围等条件查找酒店和当地活动。 🔗 阅读原文 via AI HOT · https://aihot.virxact.com/items/cmshjall30g5vronk3jwsdnyr

0 阅读原文
ArXiv 🔥 50

论文:QuanTiMedAI: Quantum-Enhanced Time-Series Model guided by Agentic AI for Cardiac Arrest Mortality Prediction

Cardiac arrest remains one of the most lethal conditions encountered in intensive care units. Despite the growing availability of electronic health record data, existing mortality prediction studies in this population largely depend on static summaries derived from early admission. Such approaches ignore the temporal progression of physiological deterioration and recovery that unfolds throughout a

0 阅读原文
AIHOT 🔥 50

宇树科技科创板发行价定为 150.8 元/股,市盈率 219.23 倍高于行业平均

宇树科技公告科创板首次公开发行定价 150.80 元/股,发行 4044.6434 万股,对应上市市值约 609.93 亿元。发行市盈率 219.23 倍,高于行业平均的 38.56 倍,预计募资总额约 60.99 亿元。战略配售获配 808.9286 万股,包括社保基金、深度求索、中国石油集团等,网上申购日为 8 月 10 日。 🔗 阅读原文 via AI HOT · https://aihot.virxact.com/items/cmshh5fi00dsyronkhrkq34l0

0 阅读原文
ArXiv 🔥 50

论文:NeSy-RAG: Neuro-Symbolic RAG for Explainable Question Answering

Retrieval-augmented generation (RAG) improves question answering by grounding large language models (LLMs) in external knowledge such as text corpora. However, its reasoning process remains largely opaque: intermediate reasoning steps are difficult to verify and cannot be reliably attributed to specific evidence. Moreover, missing user-specific context is rarely detected systematically, often lead

0 阅读原文
AIHOT 🔥 50

ChatGPT 推出改进版 GPT-5.6 Sol,并扩大免费用户访问权限

ChatGPT 推出改进版 GPT-5.6 Sol,提升准确性与一致性,同时扩大免费用户访问权限。免费用户还可无限次使用 GPT-5.6 Luna 进行日常对话。 🔗 阅读原文 via AI HOT · https://aihot.virxact.com/items/cmshrwaqs0pmuronkuj6r8o4m

0 阅读原文
ArXiv 🔥 50

论文:Stochastic Dynamics on Persistence Diagram Space via Reinforcement Learning

Persistence diagrams (PDs) provide stable and interpretable summaries of multiscale topological structure. While substantial progress has been made in the statistical analysis of PDs, existing literature often treats diagrams as static objects and provide limited frameworks for probabilistic modeling and stochastic evolution on PD space. We introduce a reinforcement learning framework for stochast

0 阅读原文
AIHOT 🔥 50

阿谀奉承的人工智能会削弱利他意图并助长依赖性(2025)

斯坦福大学和卡内基梅隆大学的研究发现,在11个前沿AI模型中,模型对用户行为的肯定率比人类高出50%,即使涉及操纵或欺骗等有害行为时也不例外。两项预注册实验(N=1604)显示,与阿谀奉承的AI互动显著降低了参与者修复人际冲突的意愿,同时增强了其自认为正确的信念。然而,参与者仍将这类回应评为更高质量、更信任并更愿意再次使用,形成助长依赖的恶性循环。 🔗 阅读原文 via AI HOT · https://aihot.virxact.com/items/cmsh6fjal001mronkgffxycgg

0 阅读原文
ArXiv 🔥 50

论文:The Illusion of Visual Tool-Use: A Causal Audit of Thinking with Images

The "thinking-with-images" paradigm equips multimodal LLMs with active visual operations such as crop-and-zoom. However, models using these operations often achieve only marginal or negative gains over direct inference at substantially higher token cost. They may also repeatedly crop irrelevant regions and fail on questions that direct inference answers correctly. We ask whether the returned visua

0 阅读原文
AIHOT 🔥 50

Qwen-Image-3.0 发布:高分辨率低至 0.03 美元

Qwen-Image-3.0 已可投入生产。支持 4.5K token 提示词。100%+ 文本准确率(无破损 logo)。原生支持 12 种语言。定价灵活可扩展:高分辨率低至 $0.03。完整详情见图片。👉 探索:• Model Studio:https://click.alibabacloud.com/m/20000001606/ • Qwen Cloud:https://click.qwencloud.com/m/20000001648/ #Qwen #QwenImage3 #AlibabaCloud #GenerativeAI #AICreativity 🔗 阅读原文 via AI HOT · https://aihot.virxact.com/items/cmsh5bwo400rjro280qm4x4is

0 阅读原文
ArXiv 🔥 50

论文:Improving the Realism of Synthetic Clinical Benchmarks Under Utility Constraints

Synthetic clinical benchmarks for enterprise AI agents can pass existing utility checks and still remain structurally unrealistic, especially in privacy-sensitive healthcare settings where operational data are hard to access. We study how to improve such benchmarks without breaking the downstream utility checks already used in practice. We formulate benchmark revision as utility-constrained realis

0 阅读原文
ArXiv 🔥 50

论文:OTLesMix: Wasserstein Barycenter and Optimal Transport Map for Synthetic Lesion Generation with Diverse Shapes and Locations

The development of deep learning over the past decade has revolutionized medical imaging segmentation, allowing the extraction of precise descriptors from large volumes to characterize pathologies. Data augmentation is a technique widely regarded as a way to improve model training. It includes simple transformations like spatial operations or intensity modifications, but also more advanced synthes

0 阅读原文
ArXiv 🔥 50

论文:Hypothesis Testing with Conditional Queries: Learnability and the Value of Interaction

Model evaluations may fix all tests before observing any responses or select later tests using earlier responses. We study this choice in a conditional-query model on a finite outcome space $\mathcal{X}$ with $|\mathcal{X}|=N$. We first ask which pairs of distribution classes can be reliably distinguished. We then ask how many additional queries are required to match an adaptive tester when all qu

0 阅读原文