Eugene Ng Yi Sheng, Bingquan Shen · 2026-07-13 · 10 min AI

自利代理社会中市场稳定的正式机制: 市场模拟研究

Formal Mechanisms for Market Stability in Self-Interested Agent Societies: A Marketplace Simulation Study

自私的代理人,不受约束,倾向于在重复的社会困境中叛逃,,导致贸易合作收益崩溃。这 ...

Self-interested agents, left unconstrained, tend toward defection in repeated social dilemmas, causing cooperative gains from trade to collapse. This ...

01
Shilin Ou, Yifan Xu, Luyao Zhang · 2026-07-12 · 4 min AI

SolarChain-Eval: 去中心化能源市场中值得信赖的经济主体的物理约束基准

SolarChain-Eval: A Physics-Constrained Benchmark for Trustworthy Economic Agents in Decentralized Energy Markets

随着代理人工智能系统越来越多地应用于网络物理环境,,他们的评估需要评估任务绩效和信任......

As agentic AI systems are increasingly applied to cyber-physical environments, their evaluation requires assessment of both task performance and trust...

02
Yifan Wu, Lizhu Zhang, Yuhang Zhou, Mingyi Wang, Bo Peng, Serena Li, Xiangjun Fan, Zhuokai Zhao · 2026-07-12 · 4 min AI

记住重要时刻: 用于长视野代理的主动内存代理

Remember When It Matters: Proactive Memory Agent for Long-Horizon Agents

在长期任务中,, 决策相关状态通常分散在不断扩展的轨迹, 上,而行动代理必须将其浮现出来并采取行动。作为...

In long-horizon tasks, decision-relevant state is often scattered across an expanding trajectory, while the action agent must surface it and act. As t...

03
Baha Rababah, Shahzeb Qamar, Lorenz Sparrenberg, Rafet Sifa, Murat Kantarcioglu, Cuneyt Gurcan Akcora, Carson K. Leung · 2026-07-12 · 3 min AI

法学硕士中量化效应的等效性:统计特征的错觉

The Illusion of Equivalency: Statistical Characterization of Quantization Effects in LLMs

训练后量化已广泛用于压缩大型语言模型,使其可部署在资源受限的设备上。然而,...

Post-Training Quantization has become widely used to compress large language models to make them deployable on resource-constrained devices. However, ...

04
Emanuele Quinto, Carlo Andrea Rozzi, Francesco Zanitti · 2026-07-11 · 10 min AI

工作流作为知识: 以 LLM 为中介的工作流的语义持久性

Workflow as Knowledge: Semantic Persistence for LLM-Mediated Workflows

大型语言模型 (LLM) 应用程序越来越多地使用显式工作流程来进行工具使用, 检索, 分支, 检查点, 和人工审批。埃西...

Large language model (LLM) applications increasingly use explicit workflows for tool use, retrieval, branching, checkpointing, and human approval. Exi...

05
Siddharth Damodharan, Radhika Gupta, Ali Alshami, Ryan Rabinowitz, Jugal Kalita · 2026-07-11 · 3 min AI

AUTOPILOT VQA: 对以事件为中心的行车记录仪理解的视觉语言模型进行基准测试

AUTOPILOT VQA: Benchmarking Vision-Language Models for Incident-Centric Dashcam Understanding

视觉语言模型,大型语言模型,和多模态大型语言模型的最新进展改善了自动驾驶任务,例如......

Recent advances in Vision-Language Models, Large Language Models, and Multimodal Large Language Models have improved autonomous driving tasks such as ...

06
Kristina Schaaff, Quintus Stierstorfer, Valerie Hekkel · 2026-07-11 · 9 min AI

在高等教育中使用基于人工智能的学习助手: 大规模描述性分析

Using AI-based Learning Assistants in Higher Education: A Large-Scale Descriptive Analysis

在这项研究,中,我们对基于人工智能的学习助手(Syntea)在高等教育中的使用进行了大规模的描述性分析。基于对象...

In this study, we present a large-scale descriptive analysis of the use of an AI-based learning assistant (Syntea) in higher education. Based on objec...

07
Yifan Zhou, Qihao Yang, Yan Li, Donggang Li, Xiru Hu, Hokin Deng, Ziyang Gong, Xuanyi Zhou, Huacan Wang, Xiangchao Yan, Wanghan Xu, Wenlong Zhang, Shaofeng Zhang, Yue Zhou, Yifan Yang, Zhihang Zhong, Xue Yang · 2026-07-11 · 3 min AI

想法有基因组: 基准科学谱系推理和基于谱系的想法生成

Ideas Have Genomes: Benchmarking Scientific Lineage Reasoning and Lineage-Grounded Idea Generation

科学思想很少是从白纸开始的。他们继承了机制, 修复已知的局限性, 并重新组合早期工作, 的各个部分,就像双...

Scientific ideas rarely start from a blank page. They inherit mechanisms, repair known limitations, and recombine pieces of earlier work, much like bi...

08
Yujiao Chen · 2026-07-10 · 10 min AI

机构红队: 部署规则, 不仅仅是模型, 因果塑造多代理人工智能安全

Institutional Red-Teaming: Deployment Rules, Not Just Models, Causally Shape Multi-Agent AI Safety

我们引入机构红队,一种用于测试多智能体AI中的部署规则的评估方法:持有代理,目标,和任务...

We introduce institutional red-teaming, an evaluation methodology for testing deployment rules in multi-agent AI: hold the agents, objectives, and tas...

09
Rachel Gordon | MIT CSAIL · 2026-07-10 · 9 min AI

微型机器人船建造浮动结构

Tiny robot boats build floating structures

大多数人认为海滨是城市的边缘。麻省理工学院的一组研究人员将其视为一个充满活力的, 乐高式建筑工地。他们的新系统...

Most people think of the waterfront as the edge of the city. A team of MIT researchers sees it as a dynamic, Lego-like construction site. Their new sy...

10