论文 精选 65 已解读
Dynamic Master Logic (DML) provides a hierarchical framework for representing system behav…
展开全文 ▼ 领域:cs.AI 作者:Saman Marandi、Yu-Shu Hu、Mohammad Modarres
论文 精选 65 已解读
Agents deployed in enterprise settings must reason across structured APIs and document col…
展开全文 ▼ 领域:cs.AI 作者:Ankita Rajaram Naik、Anupama Murthi、Benjamin Elder
论文 精选 85
Artificial intelligence tools for education and language support are increasingly framed a…
展开全文 ▼ 领域:cs.CL 作者:Avijit Roy、Proma Roy
论文 精选 83
Public procurement involves the allocation of substantial financial resources; therefore, …
展开全文 ▼ 领域:cs.CL 作者:Bryan Torres、Daniel Riofrío、José Vega-Sánchez
论文 精选 82
Multi-agent reinforcement learning for human-AI interaction typically relies on a single l…
展开全文 ▼ 领域:cs.CL 作者:Simon Yu、Nicholas Tomlin、Marwa Abdulhai
论文 精选 65 已解读
Modernizing legacy Fortran is a problem of volume: the transformations are individually ro…
展开全文 ▼ 领域:cs.AI 作者:Yuzhong Shen、Masha Sosonkina、Peng Xu
论文 精选 83
Large language models are increasingly trained and deployed with long contexts that span d…
展开全文 ▼ 领域:cs.CL 作者:Arda Uzunoglu、Benjamin van Durme、Daniel Khashabi
论文 精选 65 已解读
Foundation models for protein structure prediction remain unreliable on certain targets. E…
展开全文 ▼ 领域:cs.AI 作者:Aleksandra Kalisz、Jack Simons、Krisztina Sinkovics
论文 精选 65 已解读
Standard evaluation of large language models assumes stable model rankings across inferenc…
展开全文 ▼ 领域:cs.AI 作者:Rodrigo Guedes de Souza、Alison R. Panisson
论文 精选 83
We present the first systematic study of Massive activations (MAs) in layer-interleaved HL…
展开全文 ▼ 领域:cs.CL 作者:Zunhai Su、Bohan Sun、Xialie Zhuang
论文 精选 65 已解读
Enterprise guideline documents are heterogeneous and multimodal, combining narrative text,…
展开全文 ▼ 领域:cs.AI 作者:Shivali Dalmia、Sumukha Thoppanahalli、Mohammadreza Sediqin
论文 精选 65 已解读
Rubric-based evaluators commonly treat rubrics as prompt context or flat criteria: they sp…
展开全文 ▼ 领域:cs.AI 作者:Xi Chen、Jie Mu、Mo Xuan
论文 精选 65 已解读
AI models have achieved remarkable success across diverse domains, yet the mechanisms unde…
展开全文 ▼ 领域:cs.AI 作者:Mengru Wang、Junfeng Fang、Shuofei Qiao
论文 精选 65 已解读
Agents are increasingly considered for automating network operations and maintenance, wher…
展开全文 ▼ 领域:cs.AI 作者:Xingyu Yan、Tingting Dai、Antonio De Domenico
论文 精选 65 已解读
We propose claim-level falsification as a principle for test-time scaling and instantiate …
展开全文 ▼ 领域:cs.AI 作者:Sen Xu、Wei Wang、Shixi Liu
论文 精选 65 已解读
Tool-using LLM agents are commonly trained and evaluated in environments where tool calls …
展开全文 ▼ 领域:cs.AI 作者:Chaoran Chen、Vy Nguyen、Ziji Zhang
论文 精选 65 已解读
Roles provide an interpretable interface for organizing language-model agents, yet most mu…
展开全文 ▼ 领域:cs.AI 作者:Zhou Liu、Chaoyang Han、Zewei Pan
论文 精选 85
We construct OEIS Open, a benchmark based on 492 open mathematical conjectures from the OE…
展开全文 ▼ 领域:cs.AI 作者:Tom Adamczewski
论文 精选 86
In many practical applications of generative AI systems, from tax rules to airline baggage…
展开全文 ▼ 领域:cs.AI 作者:Rahul Nair、Bastian Lipka、Elizabeth Daly
论文 精选 80
Agent skills are the de facto mechanism for extending LLM agents with reusable guidance. A…
展开全文 ▼ 领域:cs.AI 作者:Gen Dong、Yanjie Gao、Liqun Li
论文 精选 82
Gist-based context compression---summarising older conversation history into compact repre…
展开全文 ▼ 领域:cs.AI 作者:Nicholas E. Kyrkewood
论文 精选 83
The adaptive neuro-fuzzy inference system (ANFIS) is an interpretable reasoning framework …
展开全文 ▼ 领域:cs.AI 作者:Haoran Pei、Zhao Su、Zetao Lin
论文 精选 83
When a coding agent obeys a rule, it may simply have been going to do that anyway. Existin…
展开全文 ▼ 领域:cs.AI 作者:Zining Huang、Haoran Que、Hong Zeng
论文 精选 80
Analogies are quaternary relations of the form "A is to B as C is to D". Among the various…
展开全文 ▼ 领域:cs.AI 作者:Pierre-Alexandre Murena
论文 精选 85
Aligned large language models (LLMs) are expected to exhibit safety behavior based on the …
展开全文 ▼ 领域:cs.AI 作者:Lang Cao