Hoang-Loc Cao, Van Pham, Truong Thanh Hung Nguyen, Phuc Truong Loc Nguyen, Phuc Ho, Veronica Whitford, Hung Cao
·
2026-07-19
·
8 min
AI
用于可解释抑郁症症状注释的自我进化的以人为中心的框架
Self-Evolving Human-Centered Framework for Explainable Depression Symptom Annotation
注释质量是为心理健康研究构建可靠且可解释的人工智能 (XAI) 系统的主要瓶颈。在深度...
Annotation quality is a major bottleneck in building reliable and explainable artificial intelligence (XAI) systems for mental health research. In dep...
01
Weimeng Wang, Ziqiang Wang, Zihang Zhan, Chuanpu Fu, Qi Li, Ke Xu
·
2026-07-19
·
3 min
AI
当言语安全但行动致命: 在隐藏状态风险空间中探索超越文本越狱的物理越狱
When Words Are Safe But Actions Kill: Probing Physical Jailbreak Beyond Textual Jailbreak in Hidden-State Risk Space
大型语言模型 (LLMs) 越来越多地充当具体代理的高级规划器,,其中语言上良性的指令可能变得不安全......
Large language models (LLMs) increasingly serve as high-level planners for embodied agents, where linguistically benign instructions can become unsafe...
02
Moein Taherinezhad, Sebastian Maier, Gerardo Vitagliano, Francesco Pierri, Stefan Feuerriegel
·
2026-07-19
·
10 min
AI
AutoSynthesis: 用于自动荟萃分析的代理系统
AutoSynthesis: An agentic system for automated meta-analysis
证据综合对于将初级研究转化为科学, 医学, 教育, 和政策的可靠知识至关重要。然而,定量评估...
Evidence synthesis is crucial for turning primary research into reliable knowledge for science, medicine, education, and policy. Yet, quantitative evi...
03
Qiwei Li, Jorge Ortiz
·
2026-07-19
·
8 min
AI
告诉我为什么 (Ain't 除了堵塞): 对城市驾驶数据的探索性因果分析
teLLMe Why (Ain't Nothing but a Jam): Exploratory Causal Analysis of Urban Driving Data
交通机构现在可以访问大量视频数据来研究安全和拥堵情况。这些数据大部分是观察性的并且...
Traffic agencies now have access to large volumes of video-derived data for studying safety and congestion. Most of these data are observational and c...
04
Yuyao Zhang, Junjie Gao, Zhengxian Wu, Jiaming Fan, Jin Zhang, Shihan Ma, Yao Yao, Weiran Qi, Chuyan Jin, Guiyu Ma, Xingzhong Xu, Kai Yang, Ji-Rong Wen, Zhicheng Dou
·
2026-07-18
·
3 min
AI
SearchOS-V1: 实现强大的开放域信息搜索代理协作
SearchOS-V1: Towards Robust Open-Domain Information-Seeking Agent Collaboration
工具集成大型语言模型的最新进展使网络搜索成为信息查找代理的核心能力。然而,作为交互...
Recent advances in Tool-Integrated Large Language Models have made web search a core capability of information-seeking agents. However, as interaction...
05
Victoria Graf, Hannaneh Hajishirzi, Noah A. Smith, David Kohlbrenner, Kyle Lo
·
2026-07-18
·
9 min
AI
预训练数据可能会通过计算宣传而被毒害
Pretraining Data Can Be Poisoned through Computational Propaganda
破坏预训练数据可能会给 LM 带来难以检测和缓解的有害行为。先前关于中毒预训练数据的工作......
Poisoning pretraining data can introduce harmful behaviors to LMs that are difficult to detect and mitigate. Prior work on poisoning pretraining data ...
06
Michaela Jarvis | MIT Laboratory for Information and Decision Systems
·
2026-07-18
·
10 min
AI
追随问题的引导
Following the questions where they lead
贝利·弗拉尼根(Bailey Flanigan)从小就在威斯康星州,自家’的农田里玩耍,她的选择性,而又广泛的,好奇心引导着她……
Ever since she was a child playing on her family’s farmland in Wisconsin, Bailey Flanigan was guided by her own selective, yet wide-ranging, curiosity...
07
Adam Zewe | MIT News
·
2026-07-17
·
9 min
AI
将 2D 设计转变为 3D 模型以进行快速原型设计的更好方法
A better way to turn 2D designs into 3D models for rapid prototyping
工程师经常使用视觉语言模型来产生新的设计,,例如飞机或汽车部件的设计。为了模拟这些组件将如何...
Engineers often use vision-language models to produce new designs, such as for airplane or automobile components. To simulate how those components wil...
08
Harsha Vardhan Khurdula, Abhinav Kumar Singh, Yoeven D Khemlani, Vineet Agarwal
·
2026-07-16
·
4 min
AI
使用冻结的离散扩散语言模型进行音频本机语音识别
Audio-Native Speech Recognition with a Frozen Discrete-Diffusion Language Model
自动语音识别主要由一次发出一个标记的自回归解码器主导。我们问离散扩散语言模型是否......
Automatic speech recognition is dominated by autoregressive decoders that emit one token at a time. We ask whether a discrete diffusion language model...
09
Junjie Yin, Xinyu Feng
·
2026-07-16
·
4 min
AI
AI 代理是否知道任务何时简单? 进行复杂性感知推理和执行
Do AI Agents Know When a Task Is Simple? Toward Complexity-Aware Reasoning and Execution
大型语言模型 (LLM) 代理越来越多地实现多步骤工程和信息学工作流程的自动化,,但他们很少询问一项任务需要付出多少努力...
Large language model (LLM) agents increasingly automate multi-step engineering and informatics workflows, yet they rarely ask how much effort a task a...
10