Skip to main content
archive
Search Submit Donate Log in
Press Enter to search · Advanced search

Artificial Intelligence

Authors and titles for September 2026

Total of 6116 entries : 51-150 101-200 201-300 301-400 ... 6101-6116
Showing up to 100 entries per page: fewer | more | all
[51] arXiv:2609.00575 [pdf, html, other]
Title: Residual Sparsification via Output Importance for Compressing Mixture-of-Experts LLMs
Seungwoo Jung, Dohyeok Kwon, Seungmin Cha, Junseok Lee, Yeonho Yoo, Chuck Yoo, Gyeongsik Yang
Comments: Accepted to EMNLP 2026 (Main Conference)
Subjects: Artificial Intelligence (cs.AI)
[52] arXiv:2609.00576 [pdf, html, other]
Title: Consistency Without Alignment: Item-Sensitive Language Models Indistinguishable From Random
Cris Huynh
Comments: 13 pages, 3 figures
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[53] arXiv:2609.00578 [pdf, html, other]
Title: Same Request, Different Boundary: Evaluating Cybersecurity Assistance across Conversational Contexts
Rui Yang, Yang Hong, Yichao Xu, Zhengyu Liu, Ziyang Li, Yinzhi Cao
Comments: 14 pages, 7 figures
Subjects: Artificial Intelligence (cs.AI); Cryptography and Security (cs.CR)
[54] arXiv:2609.00584 [pdf, html, other]
Title: Socrates went Nuclear: Comparing Interaction Strategies for AI systems in a Learning Context using Brain Sensing
Alexandre Clin Deffarges, Nataliya Kosmyna, Pattie Maes
Comments: 11 pages, 2 figures, 3 tables, to appear at The 14th International Conference on Human-Agent Interaction (HAI'26)
Subjects: Artificial Intelligence (cs.AI); Human-Computer Interaction (cs.HC)
[55] arXiv:2609.00621 [pdf, html, other]
Title: Control-Data Flow Separation: Stable Prompt Optimization in Multi-Agent LLMs
Wentao Zhang, Syed Shariyar Murtaza, Junaid Ahmad Bhatti, Utkarsh Soni, Yifan Nie, Eugene Wen, Yuntian Deng
Journal-ref: EMNLP 2026 Findings
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Multiagent Systems (cs.MA)
[56] arXiv:2609.00643 [pdf, html, other]
Title: REVISE: Validity-Guided Recovery for Online Revisions in Agent Workflows
Ruoling Qi, Xuaner Wu, Penghang Liu, Jian Chen, Yirui Liu
Subjects: Artificial Intelligence (cs.AI)
[57] arXiv:2609.00646 [pdf, html, other]
Title: DramaChain Bench: An End-to-End Benchmark for Short-Drama Generation
Haoyuan Shi (1), Mingtao Chen (1), Shuo Jiang (1), Ziyan Chen (1 and 2), Xuyi Sheng (3), Yiming Liu (1), Ying Zhang (1), Miao Wang (1 and 4), Jianxiang Lu (1), Fanyang Lu (1), Songyuanyi Lu (1), Xiele Wu (1), Zhichao Hu (1), Yuhong Liu (1), Richeng Xuan (1) ((1) Hunyuan, Tencent, (2) Beijing Film Academy, (3) Peking University, (4) Shenzhen University)
Comments: 50 pages, 19 figures, 19 tables. Technical report. Haoyuan Shi and Mingtao Chen contributed equally. Project lead: Zhichao Hu. Corresponding author: Richeng Xuan
Subjects: Artificial Intelligence (cs.AI)
[58] arXiv:2609.00652 [pdf, html, other]
Title: Self-Reports Are Not Verification: Environment-Grounded Auditing of LLM Operators in Evolutionary Search
Enrong Pan, Ryan Zhou, Ting Hu
Subjects: Artificial Intelligence (cs.AI); Machine Learning (cs.LG); Neural and Evolutionary Computing (cs.NE)
[59] arXiv:2609.00654 [pdf, html, other]
Title: SciTrue: Reliable Scientific Claim Validation with Frontier and Open Language Models at the NTCIR SciClaimEval Task
Qiming Bao, Neşet Özkan Tan, Siyuan Wang, Mark Gahegan
Comments: To appear in the Proceedings of the 19th NTCIR Conference (NTCIR-19)
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[60] arXiv:2609.00662 [pdf, html, other]
Title: Drift-Aware LLM Routing with Sparse Contexts and Shared Budgets
Cheung Hao Lee, Patrick Wong
Subjects: Artificial Intelligence (cs.AI)
[61] arXiv:2609.00665 [pdf, html, other]
Title: Triple-Bottom-Line Sustainability of Language Models for Edge AI: A Comparison Between SLMs and Quantized LLMs
Jainil Dharmil Shah
Subjects: Artificial Intelligence (cs.AI)
[62] arXiv:2609.00700 [pdf, html, other]
Title: Value Over Language Model: Detecting Original Contribution in Writing
Vibhhu Sharma, Thorsten Joachims, Sarah Dean
Comments: 39 pages
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[63] arXiv:2609.00714 [pdf, html, other]
Title: ChatDev 2.0: A No-Code Multi-Agent Platform for Developing Everything
Yufan Dang, Shu Yao, Bowen Lai, Chenting Xu, Ruijie Shi, Wai-Shing Leung, Huatao Li, Chen Qian, Zhiyuan Liu
Comments: Accepted at EMNLP 2026 Demo Track
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Multiagent Systems (cs.MA)
[64] arXiv:2609.00718 [pdf, html, other]
Title: A Closed-Loop Evaluation of Capability Loss and Recovery in Compressed Driving Policies
Ahmad Alfan Alfian Irfan, Nur Ahmad Khatim, Mansur Arief
Subjects: Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV)
[65] arXiv:2609.00728 [pdf, html, other]
Title: SOVER: Formal Certification of Optimization Reformulations via LLM-Assisted SMT Verification
Swapnil Bhattacharyya, Mayank Baranwal
Comments: Accepted to EMNLP 2026 Findings
Subjects: Artificial Intelligence (cs.AI); Machine Learning (cs.LG); Optimization and Control (math.OC)
[66] arXiv:2609.00731 [pdf, html, other]
Title: Agentic Empirical Asset Pricing: Methodological Foundations
Yingjian Pan, Xiaowei Ding, Kay Giesecke
Comments: 26 pages, 5 figures, 12 tables
Subjects: Artificial Intelligence (cs.AI); Machine Learning (cs.LG); Statistical Finance (q-fin.ST)
[67] arXiv:2609.00738 [pdf, html, other]
Title: Escaping Reasoning Basin Collapse with History-Biased Search
Lu Cheng
Comments: Accepted to NeurIPS'26
Subjects: Artificial Intelligence (cs.AI)
[68] arXiv:2609.00749 [pdf, html, other]
Title: ContextPipe: Database-Inspired Context Assembly for Long-Horizon Agents
Peng Xu, Zuyu Zhang, Yuze Sun, Feng Tian, Long Wang, Chen Zhang
Subjects: Artificial Intelligence (cs.AI); Databases (cs.DB)
[69] arXiv:2609.00755 [pdf, html, other]
Title: S^3martCirc: Self-supervised Smart Circuit Discovery
Wendy Zheng, Yinhan He, Liang Wu, Jundong Li
Subjects: Artificial Intelligence (cs.AI)
[70] arXiv:2609.00763 [pdf, html, other]
Title: Automated Tree Knowledge Graph Construction using Ontology Expansion and Retrieval from Vietnamese History Textbooks
Ket Doan Nguyen, Minh N. H. Nguyen
Subjects: Artificial Intelligence (cs.AI)
[71] arXiv:2609.00768 [pdf, html, other]
Title: Learning What to Practice: Diagnosis-Guided Self-Evolution for Language Models
Xincheng Wei, Yifan Ding, Fucheng Xiong, Yoshua Li, Dongsheng Ma, Rongxiang Weng, Xunliang Cai, Wenjian Ding, Yao Zhang
Subjects: Artificial Intelligence (cs.AI)
[72] arXiv:2609.00782 [pdf, html, other]
Title: When Features Become Instances: Inverted Contrastive Learning for Unsupervised Feature Selection
Utsab Ghosh, Roshni Chakraborty
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[73] arXiv:2609.00787 [pdf, html, other]
Title: StudyBench: Can Self-Evolution Squeeze Textbooks for Olympiad Capability?
Yinghao Chen, Zixi Chen, Bingxiang He, Ziqing Qiao, Huan-ang Gao, Yinuo Xu, Yuxin Zuo, Zeyuan Liu, Yuhao Zhan, Chaojun Xiao
Comments: 9 pages, EMNLP 2026 Findings
Subjects: Artificial Intelligence (cs.AI)
[74] arXiv:2609.00805 [pdf, html, other]
Title: Towards a Reliable and Practical Eval Pipeline
Emma Thuong Nguyen, Abhishek Ghose
Subjects: Artificial Intelligence (cs.AI); Software Engineering (cs.SE)
[75] arXiv:2609.00813 [pdf, html, other]
Title: One Policy, Any Budget: Internalizing Budget-Aware Search via Reinforcement Learning
Xiaowei Sun, Jin Li, Yili Hong, Yikun Fu, Yanghua Xiao
Subjects: Artificial Intelligence (cs.AI)
[76] arXiv:2609.00818 [pdf, html, other]
Title: AnalysisBank: An Expert Analysis Pattern Library for Financial Report Generation
Yajing Yang, Yunshan Ma, Kelvin J.L. Koa, Min-Yen Kan
Comments: Accepted at EMNLP 2026 (main conference, long paper)
Subjects: Artificial Intelligence (cs.AI)
[77] arXiv:2609.00823 [pdf, html, other]
Title: Polished but Unresolved: Identifying Late-Stage Pressure States in Long-Horizon Tool-Use Agents
Haoyang Chen, Yi Liu, Jianzhi Shao, Xiaozhou Xu, Zhe Sun, Wei Hu
Comments: Accepted in the 2026 Conference on Empirical Methods in Natural Language Processing (EMNLP 2026)
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[78] arXiv:2609.00831 [pdf, other]
Title: FLaG: Frequency-Domain Latent-attention Gated Pooling for Token Aggregation
Kewei Li, Rongying Zhang, Xueli Wang, Xiwen Gong, Zhongjian Wang, Qiuchen Zhao, Lan Huang, Ruochi Zhang, Fengfeng Zhou
Subjects: Artificial Intelligence (cs.AI); Biomolecules (q-bio.BM)
[79] arXiv:2609.00845 [pdf, html, other]
Title: Towards Generalizable Visually Grounded Exploration of Household Devices
Linhao Zheng, Zeming Liu, Wangke Chen, Li Zeng, Wanxiang Che, Heyan Huang, Yuhang Guo
Comments: Accepted to Findings of the Association for Computational Linguistics: EMNLP 2026
Subjects: Artificial Intelligence (cs.AI)
[80] arXiv:2609.00858 [pdf, html, other]
Title: Verifiable Disaster Storylines and Causal Knowledge Graphs: A Citation-Grounded Pipeline from Heterogeneous Humanitarian Sources
Ivan Decostanzi, Michele Ronco, Sergio Consoli, Christina Corbane, Lorenzo Bertolini, Indaco Biazzo, Daria Mihaila, Manuel Garcia-Herranz, Felix Schwebel, Yelena Mejova, Kyriaki Kalimeri
Journal-ref: International Conference on Information Technology for Social Good (GoodIT '26), September 02--04, 2026, Pisa, Italy
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[81] arXiv:2609.00859 [pdf, html, other]
Title: Reinforcement Learning Enhanced LLM Agents for Complex Vehicle Routing Problems
Yi Chen, Zikang Yu, Jiahai Wang, Jinbiao Chen, Jianpeng Zhou, Zizhen Zhang
Subjects: Artificial Intelligence (cs.AI)
[82] arXiv:2609.00874 [pdf, html, other]
Title: Beyond the Clock: Measuring the Value of Adaptive Revision
Ayushi Chadha
Comments: 11 pages, 5 figures, Preprint
Subjects: Artificial Intelligence (cs.AI)
[83] arXiv:2609.00875 [pdf, html, other]
Title: FractalNet-Based Heterogeneous Federated Learning for Orbital Edge Intelligence in Satellite Mega-Constellations: A Wildfire Case Study
Sai Puppala, Koushik Sinha
Subjects: Artificial Intelligence (cs.AI); Distributed, Parallel, and Cluster Computing (cs.DC); Emerging Technologies (cs.ET); Machine Learning (cs.LG)
[84] arXiv:2609.00879 [pdf, other]
Title: Towards reliable multimodal disaster severity assessment through preference optimization and explainable vision-language reasoning
Yuanjun Zhang, Fuzel Ahamed Shaik, Suvojit Acharjee, Fahad Khalid, Mourad Oussalah
Comments: Published in Reliability Engineering & System Safety
Journal-ref: Reliability Engineering & System Safety 275 (2026) 112674
Subjects: Artificial Intelligence (cs.AI)
[85] arXiv:2609.00885 [pdf, html, other]
Title: Denoising Diffusion Generative Models Secretly Calculate Attentions
Farzan Haddadi, Leila Monfared, Ebrahim Rezaii, Mohammadreza Malek-Mohammadi, Pejman Zakalvand, Narges Mokhtari
Comments: submitted to IEEE Trans on Pattern Recog. Machine Intellig
Subjects: Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG); Neural and Evolutionary Computing (cs.NE)
[86] arXiv:2609.00891 [pdf, html, other]
Title: CacheBridge: Efficient Cross-Model KV Cache Transfer
Xingyu Qu, Siyuan Lu, Zhiyu Chen, Sheng Wang, Tao Lin
Subjects: Artificial Intelligence (cs.AI)
[87] arXiv:2609.00892 [pdf, html, other]
Title: CARE: Contrastive Anchor-based Rubric Evolution for Large Language Model Post-Training
Siyuan Li, Xinxin Song, Chen Ruinian, Jingjing Fan, Tingxiong Xiao, Yangen Hu, Ke Zeng, Jinli Suo
Comments: EMNLP 2026 MainConference
Subjects: Artificial Intelligence (cs.AI)
[88] arXiv:2609.00904 [pdf, html, other]
Title: In-Context Neurofeedback: Can LLMs Control Their Internal Representations through Privileged Access?
Koshiro Aoki, Ryota Takatsuki, Gouki Minegishi, Yusuke Haruki, Daisuke Kawahara
Subjects: Artificial Intelligence (cs.AI)
[89] arXiv:2609.00918 [pdf, html, other]
Title: RPCBench: A Benchmark for Proactive Premise Critique in LLM-based Recommendation
Zhongru Chen, Yuan Wu, Yi Chang
Comments: 45 pages, 8 figures. Code available at this https URL
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[90] arXiv:2609.00921 [pdf, html, other]
Title: VIBE-Bench: Evaluating Personalized Large Language Models When Profiles Don't Mean Preferences
Yiwen Jiang, Yang Deng, Stephanie Fong, Zimu Wang, Yaling Shen, Wei Feng, Hongxi Yang, Xiangyu Zhao, Zhongxing Xu, Deval Mehta, Xuelian Cheng, Zongyuan Ge
Comments: Accepted at EMNLP 2026 (Findings)
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[91] arXiv:2609.00961 [pdf, html, other]
Title: Few-Shot Out of Domain Intent Detection with Covariance Corrected Mahalanobis Distance
Jayasimha Talur, Oleg Smirnov, Paul Missault
Comments: 1st AAAI Workshop on Uncertainty Reasoning and Quantification in Decision Making
Subjects: Artificial Intelligence (cs.AI)
[92] arXiv:2609.00967 [pdf, html, other]
Title: CoBRA: Learning Tool-Use Boundaries via Counterfactual Margins
Wenhao Zou, Xianglong Liu, Wendong Bi, Hanjie Wang, Simin Zhao, Gong Zhi
Comments: Acceptedy by EMNLP2026
Subjects: Artificial Intelligence (cs.AI)
[93] arXiv:2609.01006 [pdf, html, other]
Title: Figures as Programs: Recursive Generation of Editable Scientific Figures
Yepeng Liu, Dasen Dai, Chengzhi Liu, Yiren Song, Hai Ci, Yu Zhang, Qi Zhang, Mike Zheng Shou, Xin Eric Wang, Yuheng Bu
Subjects: Artificial Intelligence (cs.AI); Graphics (cs.GR)
[94] arXiv:2609.01035 [pdf, html, other]
Title: Spawn Freely, Act Sparingly: Progressive Risk Vesting for Recursive LLM-Agent Trees
Molly Wang (Imperial Business School)
Comments: 8 pages, 2 figures. Theory and synthetic numerical studies; no deployed-agent evaluation
Subjects: Artificial Intelligence (cs.AI); Machine Learning (cs.LG); Probability (math.PR)
[95] arXiv:2609.01038 [pdf, html, other]
Title: Data-Driven Persona-Conditioned Agents for A/B Test Simulation
Ziyad Benomar, Weronika Łajewska, Leonardo Perelli, Saab Mansour
Comments: Work accepted at EMNLP 2026 Industry Track
Subjects: Artificial Intelligence (cs.AI)
[96] arXiv:2609.01045 [pdf, html, other]
Title: AgentFactory: Towards Automated Agentic System Design and Optimization
Enci Zhang, Haofeng Wang, Yuesheng Zhu, Xiaole Cui, Guibo Luo
Subjects: Artificial Intelligence (cs.AI)
[97] arXiv:2609.01049 [pdf, html, other]
Title: QILP-0: Constructing Observational Declarative Twins of Quantum Circuits
Marina de la Cruz Echeandía, César Luis Alonso, Tony Ribeiro, Alfonso Ortega de la Puente
Comments: 39 pages, 7 figures. Submitted to Knowledge-Based Systems
Subjects: Artificial Intelligence (cs.AI); Quantum Physics (quant-ph)
[98] arXiv:2609.01056 [pdf, html, other]
Title: WorldBench: Culturally Grounded Benchmark for Multilingual Agents
Leonardo Ranaldi, Sherrie Shen, Jushi Kai, Alexandra Birch
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[99] arXiv:2609.01057 [pdf, html, other]
Title: User Representation via Cross Multi-source Behavior Pre-training for Mobile Games
Chengqi Yang, Yiran Qiao, Feng Liu, Xingyu Lou, Zijun Zhou, Xiaoyun Mo, Changwang Zhang, Jiayuan Xu, Jun Wang, Xiang Ao
Comments: Accepted by IEEE ICDM 2026, regular paper, 10 pages
Subjects: Artificial Intelligence (cs.AI)
[100] arXiv:2609.01058 [pdf, html, other]
Title: ARISE-RL: Agentic Rubric-Grounded Iterative Self-Evolution with Reinforcement Learning
Fanrui Zhang, Ruixue Ding, Qiang Zhang, Xi Chen, Boli Chen, Shihang Wang, Qiuchen Wang, Hongmin Zhan, Jinxin Bian, Li xingchao, Peijin Zheng, Hao cheng, Pengjun Xie, Kaipeng Zhang, Jiawei Liu, Zheng-Jun Zha
Subjects: Artificial Intelligence (cs.AI)
[101] arXiv:2609.01062 [pdf, html, other]
Title: Space Generative AI with Solar Energy Harvesting
Jierui Zhang, Jianhao Huang, Zhanwei Wang, Kaibin Huang
Comments: 14 pages, 11 figures
Subjects: Artificial Intelligence (cs.AI); Networking and Internet Architecture (cs.NI); Signal Processing (eess.SP)
[102] arXiv:2609.01117 [pdf, html, other]
Title: Latent Recurrent Thoughts: Recurrent Refinement of Proposed Latents for Reasoning with Frozen LLMs
Zhaoliang Chen, Jie Fu
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[103] arXiv:2609.01168 [pdf, html, other]
Title: Jailbreaking Text-to-Image Models Through Cracks: Navigating Heterogeneous Safety Filters via Multi-Agent Debate
Kaiyan Wen, Shijie Zhang, Lu Yu, Guangdong Bai
Comments: 16 pages, 11 figures
Subjects: Artificial Intelligence (cs.AI); Multimedia (cs.MM)
[104] arXiv:2609.01198 [pdf, html, other]
Title: FinLifeBench: Exhaustive Life-Event History and Financial-State Reconstruction from Longitudinal Banking Dialogue
Hangyeul Lee, Juyoung Oh, Jaeyong Ko, Sunmin Kim, Jaeik Park, Hyunkyu Kim, Jungmin Son, Pilsung Kang
Comments: 9 pages, 3 figures, 3 tables
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[105] arXiv:2609.01216 [pdf, html, other]
Title: H2Table: Hierarchical Hypergraph-Enhanced Large Language Models for Complex Table Reasoning
Jia Ling, Yangfan Wang, Chen Tang, Haoming Tan, Yang Yang, Yi Guan, Jingchi Jiang
Subjects: Artificial Intelligence (cs.AI)
[106] arXiv:2609.01217 [pdf, html, other]
Title: Prompt-Robust Language Models: Which Training Strategies Work?
Frederic Sadrieh, Michal Štefánik
Comments: 5 pages, 5 figures, 13 tables. Camera-ready version; Accepted to EMNLP 2026 Findings
Subjects: Artificial Intelligence (cs.AI)
[107] arXiv:2609.01257 [pdf, html, other]
Title: Measuring the Behavioral Fidelity of Long-Horizon Human Activity Simulations
Yi Fei Cheng, Fan Yang, Iremsu Bas, Koichiro Niinuma, Narishige Abe, David Lindlbauer
Subjects: Artificial Intelligence (cs.AI)
[108] arXiv:2609.01260 [pdf, html, other]
Title: Dual Process Motion Planning
Jiayi Yan, Francesco Fabiano, Alessandro Abate
Subjects: Artificial Intelligence (cs.AI); Robotics (cs.RO)
[109] arXiv:2609.01272 [pdf, other]
Title: Making Prospective Memory SLM-Shaped: Typed Intention Stores for Small-Model Agents
Jinqing Zhao, Chengcan Wu
Subjects: Artificial Intelligence (cs.AI)
[110] arXiv:2609.01286 [pdf, html, other]
Title: Analog-DB: An Agent-First Analog Integrated Circuit Database, From Blocks to Systems
Danial Noori Zadeh, Mohamed B. Elamien
Comments: 23 pages, 5 figures, and 13 tables
Subjects: Artificial Intelligence (cs.AI); Hardware Architecture (cs.AR); Signal Processing (eess.SP)
[111] arXiv:2609.01315 [pdf, html, other]
Title: A Composable Evaluation System for Reproducible Omni-Modal Foundation Model Evaluation
Hodong Lee, Sanghee Park, Dohoon Ryu, Jungwhan Kim, Junyeob Kim, Soyoon Kim, Geewook Kim
Comments: 12 pages, 3 figures. Code: this https URL
Subjects: Artificial Intelligence (cs.AI)
[112] arXiv:2609.01320 [pdf, html, other]
Title: Automated Event Log Generation from Unstructured Text Using Finetuned LLMs
Maximilian Seeth, Gabriel Marques Tavares, Daniel Schuster
Subjects: Artificial Intelligence (cs.AI)
[113] arXiv:2609.01337 [pdf, html, other]
Title: LEAP: Likelihood Elicitation and Aggregation for LLM-based Probabilistic Forecasting
Yufei Chen, Yiran Zhao, Xiaogang Xu, Qipeng Xie, Jiafei Wu, Zhe Liu
Comments: Accepted to the 2026 Conference on Empirical Methods in Natural Language Processing (EMNLP 2026). 20 pages, 3 figures, and 15 tables. Code: this https URL
Subjects: Artificial Intelligence (cs.AI)
[114] arXiv:2609.01345 [pdf, html, other]
Title: Cheap Verifiers, Large Blind Spots: Measuring the Reliability Cost of Cost-Saving Cascades
Dushyant Rajput, Nirdesh Chauhan, Siddharth Kosaraju
Subjects: Artificial Intelligence (cs.AI); Cryptography and Security (cs.CR); Machine Learning (cs.LG)
[115] arXiv:2609.01353 [pdf, html, other]
Title: SymFold: Synergizing Evolutionary and Structural Priors for Accurate Protein Inverse Folding
Handong Wang, Jiaxin Qi, Baisheng Lai, Jianqiang Huang
Comments: 13 pages, 5 figures
Subjects: Artificial Intelligence (cs.AI)
[116] arXiv:2609.01360 [pdf, html, other]
Title: EDGE: Error Dependency Graph-Guided Multi-Error Attribution in Multi-Agent LLM Systems
Jun Hou, Priya Pitre, Yi Fang, Xuan Wang
Journal-ref: EMNLP 2026
Subjects: Artificial Intelligence (cs.AI)
[117] arXiv:2609.01408 [pdf, html, other]
Title: Neuro-Symbolic Geometric Abstraction (NeuSOGA): From Observations to Symbolic Mathematical Representations
Qingde Li, Qingqi Hong, Zihan Li, Jie Tian
Comments: 18 pages, 6 figures. Code repository: this https URL
Subjects: Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV); Graphics (cs.GR)
[118] arXiv:2609.01409 [pdf, html, other]
Title: EdiTikZ: Scientific Figure Editing from Revision Trajectories
Christian Greisinger, Zhixue Zhao, Steffen Eger
Comments: 35 pages, 21 figures, and 19 tables. Models and datasets: this https URL . Code: this https URL
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Computer Vision and Pattern Recognition (cs.CV)
[119] arXiv:2609.01466 [pdf, html, other]
Title: Parsing the Stream: A Live Trace Model for Long-Horizon Agents and Their Observers
Egor Pakhomov, Erik Nijkamp
Comments: 22 pages, 4 figures. Code: this https URL ; dataset: this https URL
Subjects: Artificial Intelligence (cs.AI)
[120] arXiv:2609.01481 [pdf, html, other]
Title: Harness-of-Harness: Multi-Day Autonomous Software Development with Continual Improvement
Haoyang Yan, Min-le Su, Hangfan Zhang, Zhanhao Li, Chen Zhang, Shao Zhang, Yang Chen, Lei Bai, Shuyue Hu
Comments: Github: this https URL Project Page: this https URL
Subjects: Artificial Intelligence (cs.AI)
[121] arXiv:2609.01519 [pdf, html, other]
Title: When Guardrails Look Effective: Construct Validity Failures in LLM Agent Commerce Evaluation
Peiying Zhu, Sidi Chang
Comments: Submitted to the NeurIPS 2026 Trust-AI-Eval Workshop. 7 pages, 2 figures, 3 tables. The accompanying artifact is available at this https URL
Subjects: Artificial Intelligence (cs.AI)
[122] arXiv:2609.01526 [pdf, html, other]
Title: EvoSCM: Scientific Belief Revision Through Causal Model Evolution and Experimentation
Qing Zhao, Haowei Li, Weijian Deng, Sibei Yang, Pengxu Wei, Liang Lin
Subjects: Artificial Intelligence (cs.AI)
[123] arXiv:2609.01552 [pdf, html, other]
Title: Can LLMs Discover Scientific Laws in Real and Parallel Worlds?
Yiming Huang, Ziche Liu, Zhuohang Wu, Yiqian Wang, Junxia Cui, Xinkai Zou, Linjun Mao, Nan Huang, Naicheng Yu, Kaijie Zhu, Yue Ma, Kun Zhou, Letian Peng, Jingbo Shang
Comments: 42 pages, 16 figures. Project page: this https URL
Subjects: Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[124] arXiv:2609.01567 [pdf, html, other]
Title: Selective Agent Guidance via Entropy: Learning Autonomous Policies from Imperfect VLM Teachers
Giovanni Bonetta, Matteo Merler, Davide Zago, Rossella Cancelliere, Bernardo Magnini
Comments: 9 pages, 3 figures, 4 tables in the main text, 27 pages, 4 figures, 9 tables including Appendix
Journal-ref: EMNLP 2026 Findings
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL); Machine Learning (cs.LG)
[125] arXiv:2609.01611 [pdf, html, other]
Title: EvalDetectBench: A Benchmark for Measuring Evaluation Awareness in Frontier Language Models
Xinning Li, Kemunto Ochwang'i, Aryasomayajula Ram Bharadwaj, Alexandra Souly, Robert Kirk
Comments: 24 pages, 12 figures, 10 tables. Code: this https URL Data: this https URL
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[126] arXiv:2609.01685 [pdf, other]
Title: Meta-ethics and AI: exploring the novel meta-ethical questions in the era of AI
Shang Lu
Journal-ref: AI and Ethics (2026), 6, Article 281
Subjects: Artificial Intelligence (cs.AI); Computers and Society (cs.CY)
[127] arXiv:2609.01741 [pdf, html, other]
Title: When Can a Machine Trust a Statute? A Survival Certificate for Machine-Extracted Legal Logic
Surya Saka
Comments: 18 pages, 9 figures, 13 tables (6 main text, 7 appendix). Code, data products, and preregistration to be released on GitHub
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[128] arXiv:2609.01814 [pdf, other]
Title: When Does Information Sharing Improve Decentralized Discovery? Aggregation, Independent Rescue, and Equilibrium Selection
Yohei Nakajima
Comments: working paper; deterministic source package; exact bounded evidence
Subjects: Artificial Intelligence (cs.AI); Computer Science and Game Theory (cs.GT)
[129] arXiv:2609.01815 [pdf, html, other]
Title: Induction and Inquiry via Probabilistic Reasoning over Language and Code
Wasu Top Piriyakulkij, Sam Acquaviva, Cassidy Langenfeld, Joshua Tenenbaum, Kevin Ellis
Subjects: Artificial Intelligence (cs.AI)
[130] arXiv:2609.01834 [pdf, html, other]
Title: Architecting Conversational Data Systems for Stateless LLM APIs: The Hydration Proxy Pattern
Joseph Axisa
Comments: 3 pages, 1 table. Presented at the SAO workshop at the 1st ACM Conference on AI and Agentic AI Systems (ACM CAIS 2026)
Journal-ref: SAO Workshop at the 1st ACM Conference on AI and Agentic Systems (ACM CAIS 2026)
Subjects: Artificial Intelligence (cs.AI); Software Engineering (cs.SE)
[131] arXiv:2609.01849 [pdf, html, other]
Title: SSAKG 2.0: An Open-Source Package for Structural Associative Sequence Memory and Context-Based Retrieval
Przemysław Stokłosa, Janusz A. Starzyk, Paweł Raif
Comments: 15 pages, 5 figures
Subjects: Artificial Intelligence (cs.AI)
[132] arXiv:2609.01852 [pdf, html, other]
Title: The Memory Trust Gap: Capability-Dependent Failures in Persistent-Memory Agents
Jundong Hu, Shekar Ramachandran
Comments: Preprint. Under review at a NeurIPS 2026 workshop. 14 pages, 7 figures, 11 tables
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[133] arXiv:2609.01861 [pdf, html, other]
Title: Belief-Calibrated Optimization: An Explicit World Model for Agentic Optimization
Yuhan Chen, Zhihua Tian, Mahavir Dabas, Charith Peris, Rahul Gupta, Ming Jin, Feiyang Kang, Siyuan Zhang, Nan Wang, Ruoxi Jia
Subjects: Artificial Intelligence (cs.AI)
[134] arXiv:2609.01873 [pdf, html, other]
Title: Epistemic Sybil Resistance: Multiplying AI Agents Without Multiplying Evidence
Marc Bara
Comments: 23 pages, 13 figures. Code and data: this https URL
Subjects: Artificial Intelligence (cs.AI); Multiagent Systems (cs.MA)
[135] arXiv:2609.01909 [pdf, html, other]
Title: The Ceiling Is in the Channel: Auditing Learner Gaps and Measurement Frontiers in Clinical Prediction
Sayeed Shafayet Chowdhury, Nusrat Jahan, Snehasis Mukhopadhyay, Shiaofen Fang, Vijay R. Ramakrishnan
Subjects: Artificial Intelligence (cs.AI)
[136] arXiv:2609.01924 [pdf, html, other]
Title: Looped Transformers under the Jacobian Lens: Does the Global Workspace Survive Recurrence?
Wenlong Wang, Fergal Reid
Subjects: Artificial Intelligence (cs.AI)
[137] arXiv:2609.01962 [pdf, html, other]
Title: Post-Training Ternarization of Qwen3-4B Capability, Effective Bit Budget, Storage Compression, and Deployment
Anirudh Malik, M Sparsh Mehra, Poojith Devan
Comments: Weight-only post-training ternarization of a 4B-parameter instruction-tuned language model. Activation quantization and end-to-end generation throughput are outside the scope of the primary evaluation
Subjects: Artificial Intelligence (cs.AI); Machine Learning (cs.LG)
[138] arXiv:2609.01982 [pdf, html, other]
Title: Benchmarking Language Models for Statistical Problem Formulation
Chen Wang, Junzhe Zhao, Xin Cong, Wanlu Deng, Ke Deng
Comments: Accepted for publication at the EMNLP 2026 main conference
Subjects: Artificial Intelligence (cs.AI)
[139] arXiv:2609.01985 [pdf, html, other]
Title: When Agents Implement Systems: A Case Study in Defects, Detection, and Evaluation Rigor
Phanindra Reddy Madduru
Comments: 4 pages
Subjects: Artificial Intelligence (cs.AI); Multiagent Systems (cs.MA)
[140] arXiv:2609.01992 [pdf, html, other]
Title: ClaimReceipt: Verifying Evidence Sufficiency and Coverage in Agent Evaluations
Peiying Zhu, Sidi Chang
Comments: Submitted to Who Verifies the Agents? Toward Reliable Agent Development (NeurIPS 2026 workshop). 8 pages, 1 figure, 7 tables
Subjects: Artificial Intelligence (cs.AI); Cryptography and Security (cs.CR); Multiagent Systems (cs.MA)
[141] arXiv:2609.02029 [pdf, html, other]
Title: HeadWiseKV: Budgeted Per-Head Cache Residency for Hybrid Long-Context Language Models
Renjie Xie, Juncheng Yang, Aoting Hu, Mingxi Zhang, Liyao Wu, Zheheng Hong, Wei Xu
Comments: 14 pages including appendices, 4 figures, 5 tables
Subjects: Artificial Intelligence (cs.AI)
[142] arXiv:2609.02057 [pdf, html, other]
Title: Monitoring Web Agents Without Internal Signals: Observable Trajectories and Key-Step Supervision
Sitong Pan, Yipeng Shen, Yilin Lu, Caiwen Ding, Lu Cheng, Qianwen Wang
Comments: preprint
Subjects: Artificial Intelligence (cs.AI)
[143] arXiv:2609.02059 [pdf, html, other]
Title: DocHop: Benchmarking Out-of-domain Multi-hop Reasoning in Information-Dense Documents
Zhuoran Yu, Le Thien Phuc Nguyen, Jaden Park, Xinyi Gu, Zexue He, Soochahn Lee, Rogerio Feris, Yong Jae Lee
Comments: Accepted by ICML 2026
Subjects: Artificial Intelligence (cs.AI); Computer Vision and Pattern Recognition (cs.CV); Machine Learning (cs.LG)
[144] arXiv:2609.02060 [pdf, html, other]
Title: MineTRACE: An Evidence-Grounded Interactive Reasoning System for Mineral Prospectivity
Yiran Zhang, Jinwen Liu, Daniel Su, Yisu Chen, Qiang Sun, Chris Gonzalez, Eun-Jung Holden, Marco Fiorentini, Wei Liu, Yihao Ding
Comments: EMNLP 2026 Demo
Subjects: Artificial Intelligence (cs.AI)
[145] arXiv:2609.02067 [pdf, html, other]
Title: ToolGate: An Executable Acceptance Pipeline for Tool-Dependent Scientific Benchmark Construction
Ke Zhang, Yankang Liu, Roya Zandi, Maziar Raissi
Comments: 7 pages, 2 figures
Subjects: Artificial Intelligence (cs.AI); Mathematical Software (cs.MS); Software Engineering (cs.SE)
[146] arXiv:2609.02074 [pdf, html, other]
Title: CHIME: Credit-Aware Hierarchical Memory Evolution for Long-Horizon Agentic Planning
Yongshi Ye, Tian Lan, Feihu Jiang, Muyang Ye, Bin Zhu, Qianghuai Jia, Longyue Wang, Zhao Xu, Weihua Luo, Xiaodong Shi
Comments: 7 figures, 7 tables
Subjects: Artificial Intelligence (cs.AI)
[147] arXiv:2609.02092 [pdf, html, other]
Title: Beyond Outcome Gaps: Process-Aware Fairness Diagnosis for LLM-based Multi-Agent Decision Systems
Yiran Zhao, Lu Zhou, Liming Fang, Yufei Chen, Jiafei Wu, Zhe Liu, Xiaogang Xu
Comments: Accepted to EMNLP 2026
Subjects: Artificial Intelligence (cs.AI)
[148] arXiv:2609.02094 [pdf, html, other]
Title: MASkills: Continual Skills Optimization for Multi-Agent LLM Systems
Huaiyuan Yao, Xiaoou Liu, Charles Fleming, Tianlong Chen, Hua Wei
Comments: 14 pages, 4 figures
Journal-ref: EMNLP 2026 Findings
Subjects: Artificial Intelligence (cs.AI); Computation and Language (cs.CL)
[149] arXiv:2609.02095 [pdf, html, other]
Title: READY or Not: Reliable Enterprise Agent Deployment
Veronica Chatrath, Bryan Zhu, Jingxuan Fan, George Pu, Soham Dinesh Tiwari, Soham Dan, Ryan Young, Yuan (Christy)Li, Yuang Yao, Apaar Shanker, Minglai Yang, Daniel Yue Zhang, Yunzhong He, Ying Liu, Chenguang Wang, Zhijun Yin, Yuan Xue
Subjects: Artificial Intelligence (cs.AI)
[150] arXiv:2609.02116 [pdf, html, other]
Title: Semantic Signal-Assisted Inspection and Recovery Allocation in Reverse Logistics
Jiani He, Dingyan Shang, Yihua Xu, Shiqi Huang, Yan Lyu, Jize Li, Shangjing Tang
Comments: Accepted at the IEEE 4th International Conference on Artificial Intelligence, Blockchain, and Internet of Things (AIBThings 2026). 7 pages, 1 figure, 3 tables. Code and benchmark: this https URL
Subjects: Artificial Intelligence (cs.AI)
Total of 6116 entries : 51-150 101-200 201-300 301-400 ... 6101-6116
Showing up to 100 entries per page: fewer | more | all
We gratefully acknowledge support from our major funders, member institutions, , and all contributors.
About · Help · Contact · Subscribe · Copyright · Privacy · Accessibility · Operational Status (opens in new tab)
Major funding support from
Simons Foundation Simons Foundation International Schmidt Sciences