Aily.to

🔥 AI热点

实时追踪 AI 行业热点动态 · 共 100 条资讯

2026年8月6日
12 条
ArXiv 🔥 50

论文:An Optimal Agnostic PAC Algorithm

Let $H\subseteq\{-1,+1\}^X$ be a class of finite VC dimension $d\ge1$. Writing $L$ for the binary risk and $L^*=\min_{h\in H}L(h)$, we construct a learner achieving the statistically optimal risk bound: from an i.i.d.\ sample of size $n$, for every $0<δ\le 1/2$, with probability at least $1-δ$, \[ L(\widehat h) \le L^*+ 7\cdot10^8\left( \sqrt{\frac{L^*(d+\log(1/δ))}{n}} +\frac{d+\log(1/δ)}{n} \rig

0 阅读原文
Product Hunt 🔥 50

Product Hunt:Soloop

Approval-first Agent OS for solo founders Discussion | Link

0 阅读原文
AI Bot 🔥 50

兔展智能推出一站式 AI 设计生产工具 RabbitVis

兔展智能推出图层级一站式AI设计工具 RabbitVis ,基于自研UniWorld-Design视觉大模型,将AI生成与图层级编辑深度融合。产品突破传统AI生图只能生成、难以修改的瓶颈,支持透明素材生成、图像分层拆解与持续编辑,让海报、电商图、杂志内页等设计资产可反复调整与复用。 来源:量子位

0 阅读原文
ArXiv 🔥 50

论文:AV-AIVAT: 74x Cheaper Agent Evaluation with Certified Anytime-Valid Stopping in Imperfect-Information Games

Deciding which of two agents is stronger means playing games until skill outweighs luck, and every game costs money, model inference, or expert time. Since the number of games needed is unknown, fixed-budget evaluations either keep paying after the result is settled or stop before the agents can be told apart, while naive optional stopping with an ordinary confidence interval invalidates the state

0 阅读原文
Product Hunt 🔥 50

Product Hunt:Mem0

Persistent Memory Layer for AI Agents Discussion | Link

0 阅读原文
ArXiv 🔥 50

论文:The Low Frequency Trap: Video Language Models Fail at Simple Event Bookkeeping

Real-world video benchmarks provide broad coverage, but their fixed clips entangle event count, rate, duration, and visual complexity, making failure modes hard to isolate. While existing programmatic benchmarks offer better control, they score only the final answer rather than auditing reported events against executable ground truth. To bridge this gap, we introduce trace-grounded parametric prof

0 阅读原文
Product Hunt 🔥 50

Product Hunt:Shieldstral

Define safety at runtime for text and images Discussion | Link

0 阅读原文
ArXiv 🔥 50

论文:Resourced Authority A Mechanism-Design Model for Participatory Governance of Deployed AI Agents

We give a formal mechanism design model for the continuous participatory governance of a deployed AI agent. The mechanism is built on the principle that governance should control an AI agent through resource allocation so as to make authorization self enforcing via compute budgets. The mechanism seeks to establish the Safe AI paradigm that compute is an effective governance lever. We situate our w

0 阅读原文
GitHub 🔥 50

GitHub 热门:goauthentik/authentik

The authentication glue you need.

0 阅读原文
ArXiv 🔥 50

论文:CalibForge: Adversarial Solver Calibration for Scaling Learnable Terminal Tasks

Training terminal agents requires executable and verifiable tasks that are not merely solvable, but appropriately challenging for learning. Executable validation establishes feasibility, yet does not reveal how a task behaves relative to a given solver setting. In this paper, we present CalibForge, an autonomous terminal-task synthesis system that uses verified solver behavior to revise candidate

0 阅读原文
GitHub 🔥 50

GitHub 热门:google/guava

Google core libraries for Java

0 阅读原文
ArXiv 🔥 50

论文:Challenges in Evaluating Explanation Methods for Static and Evolving Data

This paper addresses the limitations of Explainable Artificial Intelligence (XAI) with respect to insufficient evaluation. They are illustrated through the DetoxAI image recognition system for bias detection and concept unlearning. Then, an example of a human-grounded evaluation of methods for explaining image classification is presented. The paper further explores methods for adapting explanation

0 阅读原文
2026年8月7日
8 条
AIHOT 🔥 50

斯坦福与 Arc Institute 用 AI 设计全新病毒基因组,16 种在实验室成功杀死细菌

斯坦福大学与 Arc Institute 团队用 AI 模型 Evo 从零设计完整病毒基因组,并在实验室构建出 16 种自然界不存在的功能性病毒。Evo 提出 70 万个候选基因组,团队仅筛选最有希望的 285 个序列合成并植入细菌,其中 16 个成功复制并杀死宿主。该研究已通过同行评审发表于《Science》,但 Evo 未接受人类病原体数据训练,且能否推广至其他病毒类群仍是未知数。 🔗 阅读原文 via AI HOT · https://aihot.virxact.com/items/cmsiys4dz1yqironkuitnfi7t

0 阅读原文
GitHub 🔥 50

GitHub 热门:unclebob/swarm-forge

A simple tool for coordinating several AI agents.

0 阅读原文
AIHOT 🔥 50

蚂蚁百灵开源 Ling-3.0-flash:124B 总参数 MoE 模型,支持 API、单机与高性能三种部署

蚂蚁百灵正式开源新一代原生混合推理模型 Ling-3.0-flash,采用 124B 总参数、5.1B 激活参数的 MoE 架构,并提供 FP8、FP4、INT4 等多个版本。 🔗 阅读原文 via AI HOT · https://aihot.virxact.com/items/cmsixbbry1x0qronkhbbjukhr

0 阅读原文
GitHub 🔥 50

GitHub 热门:denoland/celld

self-hosted, distributed Durable Objects

0 阅读原文
AIHOT 🔥 50

Runway 上线 Seedance 2.5,支持 50 个角色参考

Seedance 2.5 现已登陆 Runway。每次生成最多可引用 50 个角色参考,构建充满角色的完整世界;可创作最长 30 秒、带完整音效与对白的片段,再按你的故事需求随意剪辑与延展。 点击下方链接立即开始。 🔗 阅读原文 via AI HOT · https://aihot.virxact.com/items/cmsiujxlq1u0zronkvvklblwk

0 阅读原文
GitHub 🔥 50

GitHub 热门:K2SOsint/Legendary_OSINT

A list of OSINT tools & resources for (fraud-)investigators, CTI-analysts, KYC, AML and more.

0 阅读原文
AIHOT 🔥 50

小红书联合浙大、复旦提出 CULTURE-MT:首个面向社媒翻译的「文化有效性」评测基准,入选 ICML 2026

小红书联合浙江大学、复旦大学提出 CULTURE-MT,这是首个面向中英社媒笔记翻译、兼顾文化符号传递与情感共鸣的评测基准,并首次提出「文化有效性」评估标准与自动评估模型 JUDGER(准确率 86.03%)。 🔗 阅读原文 via AI HOT · https://aihot.virxact.com/items/cmsiry2eo1r4tronkt9qv2twu

0 阅读原文
AIHOT 🔥 50

Krea 推出 Seedance 2.5 视频模型

推出 Seedance 2.5。 30 秒连续视频、完整多镜头序列,以及最多 50 个参考。 立即试用 👇 🔗 阅读原文 via AI HOT · https://aihot.virxact.com/items/cmsimzpmf1ks3ronk4hh1d0z0

0 阅读原文