Ali Ansari, Yasmin Mohammadi, Farnoush Nili, Parsa Esmaeilkhani, Longin Jan Latecki, Eduard Dragut · 2026-07-29 · 9 min AI

ERUnderstand: 评估结构化 ER 图上的视觉语言模型

ERUnderstand: Evaluating Vision-Language Models on Structured ER Diagrams

实体关系图 (ERDs) 是概念数据库设计的核心,,但它们通常仅作为渲染图像而不是主...

Entity-Relationship Diagrams (ERDs) are central to conceptual database design, yet they are typically available only as rendered images rather than ma...

01
Anduel Mehmeti, Gabriella Gigante, Salvatore Venticinque · 2026-07-28 · 3 min AI

用于协助空中交通管制员的可解释强化学习

Explainable Reinforcement Learning for assisting Air Traffic Controllers

为了有效地将人工智能集成到高风险,关键环境中,例如医疗保健,自动驾驶,和航空——并朝着更高的方向前进......

To effectively integrate AI into high-stakes, critical environments such as healthcare, autonomous driving, and aviation--and to advance toward higher...

02
Natan Levy, Harel Berger · 2026-07-27 · 8 min AI

持续保证工业界人工智能代理创建的民主化

Toward Continuous Assurance for the Democratization of AI Agent Creation in Industry

人工智能代理越来越多地由非工程用户通过低代码,、无代码, 和对话式开发环境在组织内部创建......

AI agents are increasingly created inside organizations by non-engineering users through low-code, no-code, and conversational development environment...

03
Gaurav Dadhich · 2026-07-27 · 7 min AI

代理上下文管理: 通过将代理内存和成本视为生命周期和体系结构问题来解决它们

Agentic Context Management: Solving Agent Memory and Cost by Treating Them as Lifecycle and Architecture Problems

生产型 AI 代理的故障较少是由于无法良好推理而导致的,而更常见的是因为他们无法管理推理中的内容......

Production AI agents' failures are less often due to an inability to reason well and more often because they cannot manage what is in their reasoning ...

04
Linjun Li · 2026-07-27 · 9 min AI

相同的危险目标, 相反的建议: 直接暴露与多代理调解

Same Dangerous Objective, Opposite Advice: Direct Exposure versus Multi-Agent Mediation

即使是当前的高能力法学硕士,在直接显示危险目标时也会比其他代理转变并传递其方向时显得更安全......

Even a current high-capability LLM can appear safer when shown a dangerous objective directly than when other agents transform and relay its direction...

05
Fares Fourati, Hinrich Schütze, Eyke Hüllermeier, Iryna Gurevych · 2026-07-27 · 6 min AI

自动化的边界: 人类持续参与理论

The Boundaries of Automation: A Theory of Persistent Human Participation

人工智能的快速进步强化了长期以来对自动化:的追求,尽可能用算法代替人类参与。进出口...

The rapid progress of AI has intensified the long-standing pursuit of automation: replacing human participation with algorithms wherever possible. Imp...

06
Wen Ye, Yuxiao Qu, Aviral Kumar, Xuezhe Ma · 2026-07-27 · 9 min AI

MIRROR: 从其他视角学习多模态推理

MIRROR: Learning from the Other View for Multi-Modal Reasoning

与表现出强大推理能力的大型语言模型 (LLMs) 不同, 视觉语言模型 (VLMs) 即使在...上也很难进行视觉推理,

Unlike large language models (LLMs) that exhibit strong reasoning capabilities, vision-language models (VLMs) struggle with visual reasoning, even on ...

07
Xiao Yu, Baolin Peng, Ruize Xu, Hao Zou, Qianhui Wu, Hao Cheng, Wenlin Yao, Nikhil Singh, Zhou Yu, Jianfeng Gao · 2026-07-26 · 7 min AI

OpenForgeRL: 任何环境下的列车线束原生代理

OpenForgeRL: Train Harness-native Agents in Any Environment

现代 AI 代理依靠复杂的推理工具(例如 Claude Code, Codex, 和 OpenClaw)来驱动多轮推理, 工具使用, 并访问...

Modern AI agents rely on elaborate inference harnesses such as Claude Code, Codex, and OpenClaw to drive multi-turn reasoning, tool use, and access to...

08
Baihui Wang, Bernard Koch · 2026-07-26 · 8 min AI

超越阿谀奉承:法学硕士道德推理中的结构化抵抗和顺从

Beyond Sycophancy: Structured Resistance and Compliance in LLM Moral Reasoning

构建社会校准的大型语言模型,,它可以向他人学习,而不是简单地屈服于他们, 需要的不仅仅是减少阿谀奉承......

Building socially calibrated large language models, which can learn from others without simply yielding to them, requires more than reducing sycophanc...

09
T. Ansah-Narh, Y. Asare Afrane · 2026-07-25 · 8 min AI

加纳时空疟疾发病率的无监督共识异常检测

Unsupervised Consensus-Based Anomaly Detection for Spatiotemporal Malaria Incidence in Ghana

将共识异常检测框架应用于加纳(2014-2023) 的每月疟疾监测数据,以识别非典型传播模式...

A consensus anomaly detection framework was applied to monthly malaria surveillance data from Ghana (2014-2023) to identify atypical transmission patt...

10