←── back to feed
/topics/arxiv-ai-and-ml-research-papers-september-4

arXiv AI and ML research papers September 4

135 items1 sourcesupdated 15d agotrend 0

On September 4, 2026, arXiv published 20 papers spanning computational linguistics, speech recognition, and legal AI. Key advances include methods for optimizing LLM agent prompts (HARNESSEVO), efficient retrieval-augmented generation (R²Adapter), speculative decoding for faster inference (AdaptiveSpec), and domain-adapted biomedical models (DRET), alongside new benchmarks for misinformation detection (BharatGather), legal issue identification (LexIssue), and document parsing (Jina-OCR-v1).

  • HARNESSEVO decomposes LLM agent harness into four separately evolvable slots: role, task-strategy, tool/format-rules, and reflection/control.
  • R²Adapter routes queries between text and graph RAG strategies to balance reasoning capability and inference latency.
  • AdaptiveSpec enables training-free per-step lossy speculative decoding without fixed token budgets or strict verification rules.
  • DRET injects biomedical domain knowledge into lightweight DistilBERT via priority-based embedding transfer.
  • Jina-OCR-v1 combines 3B mixture-of-experts decoder with FastMTP speculative decoding (K=3 steps) for efficient document parsing on low-budget GPUs.
  • BharatGather dataset targets misinformation detection in Indian public events with socio-cultural context.
[BLG]blog/rss135
Where Does Harness-Optimization Value Live? Localized Gains and the Budget-Splitting Trap in Self-Evolving LLM Agents
arXiv cs.CL · Michael Nguyen, Wei Chen Tan, Nurul Aisyah Hassan, Arvind Raman, Li Hua Lim, Ahmad Faiz Razak · 17d
Bounded Personas Match Retrieval on Classification but Not Regression for a Frozen Agent
arXiv cs.CL · JaeHa Yoon, Minjun Park, Seoyeon Kim, Jiwoo Lee, Hyunwoo Choi, Dohyun Kang · 17d
Counterexamples as Feedback for Agent Self-Correction
arXiv cs.CL · Sidhesh Badrinarayan, Adithya Parthasarathy · 17d
Probe Generalization as Subspace Selection for OOD Deception Detection
arXiv cs.CL · Daniel Yoo, Adrians Skapars · 17d
R$^{2}$Adapter: A Routing and Rewriting Adapter for Efficient Hybrid RAG
arXiv cs.CL · Yucan Guo, Miao Su, Saiping Guan, Long Bai, Zhongni Hou, Zixuan Li, Xiaolong Jin, Jiafeng Guo, Xueqi Cheng · 17d
BharatGather: A Culturally-Informed Benchmark Dataset for Misinformation and Fake News Detection in Indian Public Events
arXiv cs.CL · Parth Bramhecha, Smit Deshmukh, Sairaj Bodhale, Adwait Borate, Raviraj Joshi · 17d
PiPMRE: A Pipeline Based on Language Model for Medical Relation Extraction
arXiv cs.CL · Jiaxin Duan, Fengyu Lu, Junfei Liu · 17d
Margins, Not Windows: Training-Free Per-Step Lossy Speculative Decoding
arXiv cs.CL · Oszk\'ar Urb\'an, Young D. Kwon, Stylianos I. Venieris, Cecilia Mascolo · 17d
Distilled Rapid Embedding Transfer (DRET): Parameter-Efficient Biomedical Domain Adaptation via Priority-Based Embedding Transfer
arXiv cs.CL · Girish Sundaram, Daniel Berleant · 17d
Contamination Inflates Scores but Rarely Reorders Large Language Model Leaderboards
arXiv cs.CL · Xingyao Xiao (Stanford University), Yihong Cheng (City University of Macau) · 17d
Dual-Form ASR: Semantics-Aware Inverse Text Normalization for Chinese Speech Recognition
arXiv cs.CL · Fengrun Zhang, Li Fu, Wangjin Zhou, Lu Fan, Youzheng Wu, Xiaodong He · 17d
RL-ADA: A World-Feedback Framework for Adversarially Robust Enterprise Dialogue Agents
arXiv cs.CL · Ram Narayanan, Harshit Rajgarhia, Abhishek Mukherji · 17d
Listen to the Latents: Self-Correcting Speech Recognition in Large Audio Language Models Through Hidden-State Interactions
arXiv cs.CL · Chan-Jan Hsu, Jaeyeon Kim, Chao-Han Huck Yang, Shinji Watanabe, Hung-yi Lee, Carlos Busso · 17d
Judging LLM-as-a-Judge: Concerning Rubric Artifacts in LLM-based Automated Text Generation Evaluation
arXiv cs.CL · Anshul Bagaria, Sowmya S Sundaram, Gokul S Krishnan, Balaraman Ravindran · 17d
LexIssue: Benchmarking Legal Issue Identification in Chinese Civil Litigation
arXiv cs.CL · Huiyuan Xie, Yuqin Huang, Zhicheng Hao, Yida Cai, Shaochun Wang, Zhenghao Liu, Yuxiao Ye · 17d
Unifying Conformal Language Tasks with In-Context Ensembles
arXiv cs.CL · Xiao Shi Huang, Chen-Yuan Lin, Bruce Kuwahara, Kin Kwan Leung, Jesse C. Cresswell · 17d
SHELF: A Synthetic Harness for Multi-Task Bibliographic Benchmarking
arXiv cs.CL · Michael J. Bommarito II · 17d
Large Language Models in Resolving Contextual Knowledge Conflicts
arXiv cs.CL · Xinye Yang, Zhenyang Liu, Ruisi Li, Yuanyuan Lei · 17d
No country for old linguists: LLM-brain alignment underdetermines neural computation
arXiv cs.CL · Elliot Murphy · 17d
Jina-OCR-v1: Efficient Document Parsing with Speculative Decoding and Dense Verifiable Rewards
arXiv cs.CL · Alejandro Bar\'on Garc\'ia, Feng Wang, Emilia Garcia Casademont, Han Xiao · 17d
MemoryLACE: Memory Lifecycle-Aware Consolidation and Evidence Retrieval
arXiv cs.CL · Meriem Yacoubi, Pia Schmidt, Nenad Petrovic, Ahmed Frikha, Martin Kirchhoff, Alois Knoll · 17d
LLMs Learn Better In-Context from Rules than from Examples
arXiv cs.CL · Xiang Fu, Seungmin Cho, Yukyung Lee, Najoung Kim · 17d
SWIM: Student Writing Simulation via Proficiency-Conditioned Generation
arXiv cs.CL · Heejin Do, Jakub Kontak, Mrinmaya Sachan · 17d
The Analyst in the Prompt: Role, Retrieval, and Memory Biases in LLM Financial Analysis
arXiv cs.CL · Ahmed Asaad, Amr Mohamed, Yang Zhang, Omneya Abdelsalam · 17d
Counterfactual Fairness Audits of Multi-Step Clinical LLM Agents Require a Measured Per-Action Instability Floor
arXiv cs.CL · Rohith Reddy Bellibaltu, Manpreet Singh, Deepak Parashar, Rahul Joshi · 17d
SGD-KV: Summarization Guided KV Cache Compression
arXiv cs.CL · Zeyu Liu, Woomin Song, Xuandi Fu, Sai Muralidhar Jayanthi, Vivek Govindan, Aram Galstyan, Sravan Babu Bodapati, Srikanth Ronanki · 17d
What Else Needs Fixing? Exploring Cost-Effective Test-Time Compute for Revision Propagation in Artifacts Generated Through Conversation
arXiv cs.CL · Daisuke Kikuta · 17d
Contextual Tamil Spelling and Grammar Correction Using Progressively Fine-Tuned Sequence-to-Sequence Transformers
arXiv cs.CL · Karthikeyan A, Jaya Nirmala S, Sangeetha Sivanesan, Indhu R, Pranav Kumar, Bharat Jude Johnson, Vishnu Ram · 17d
PACE: Towards Surfacing Hidden Conflicts in User Requests
arXiv cs.CL · Yoojin Kim, Jihyoung Jang, Hyounghun Kim · 17d
Decoupling Turn-Taking from Semantics: A Decoupled Data Approach for Finite-State-Machine-Based Full-Duplex Dialogue
arXiv cs.CL · Yihang Li, Chenhui Chu · 17d
How Perturbations Propagate: A Multi-Level Analysis of Robustness in Large Language Models
arXiv cs.CL · Dun Li Chan, Emily Liu, Niyathi Allu, Christian Hoang · 17d
Less Is Moral: A CHARMing Framework for Moral Foundations Detection in Endorsement Behaviour
arXiv cs.CL · Huixiang Fu, Marian-Andrei Rizoiu · 17d
FPCO-Dialog: A Multi-Turn False-Premise Benchmark for Correction and Cooperation in Vision-Language Models
arXiv cs.CL · Jiayuan Ma, Yuqi Lu, Weiyang Guo, Chenrui Wang, Junyi Shu, Xuebo Liu, Min Zhang, Jing Li · 17d
Accountable AI with Grounded, Faithful, Consistent, Actionable Rationales: A Case Study in Clinical Trial Matching with VERDICT
arXiv cs.CL · Zikai Zhou, Yufei Jin, Yilin Xu, Yu-Chiang Wang, Chieh-Ju Chao, Monica S. Lam · 17d
FrameBench:A Language Understanding Benchmark Based on Frame Semantics
arXiv cs.CL · Chihiro Yano, Ryohei Sasano · 17d
Chiaroscuro for Emotions: A Contrastive Emotion Benchmark Grounded in Appraisal Theory
arXiv cs.CL · Divyesh Bommana, Mohammad Saim, Tianyu Jiang · 17d
TabScope: Question-Adaptive Scope Selection for Table Question Answering
arXiv cs.CL · Yuxiang Wang, Junhao Gan, Jianzhong Qi · 17d
To What Extent Do Large Language Models Understand Bangla Idioms?
arXiv cs.CL · Mousumi Akter, Md. Faiyaz Abdullah Sayeedi, Nurul Labib Sayeedi, Swakkhar Shatabda · 17d
Lngram v2: Latent N-Gram Memory with Interpretable Discrete Representations
arXiv cs.CL · Yunao Zheng, Bin Wen, Xiaojie Wang · 17d
Random Attention: Rethinking KV Cache Eviction for Efficient Reasoning
arXiv cs.CL · Heng Wang, Jielin Qiu, Wenting Zhao, Cheng Qian, Liangwei Yang, Jiawei Han, Heng Ji, Silvio Savarese, Shelby Heinecke, Huan Wang · 17d
Decoupled Analysis-Judging: An Automated Creativity Evaluator Using LLMs in Complex Multi-step Creativity Tasks
arXiv cs.CL · Xiangyu Wang, Jin Wu, Xiaoyu Li, Chanjin Zheng, Yifeng Zhou · 17d
When Retrieval Helps: Selective Retrieval for Single-Turn Mental-Health QA
arXiv cs.CL · Hyunseo Oh, Chong-Kwon Kim, Yoonhyuk Choi · 17d
When Users Don't Ask: Benchmarking Context-Driven Memory Retrieval in Conversational Agents
arXiv cs.CL · Wen-Yu Chang, Yun-Nung Chen · 17d
Pattern Over-Generalization of Knowledge Graph Embedding
arXiv cs.CL · Junsik Kim, Kangil Kim · 17d
Building and Evaluating Fixed-Voice Thai TTS from Synthetic Speech
arXiv cs.CL · Kunat Pipatanakul, Potsawee Manakul, Warit Sirichotedumrong, Sittipong Sripaisarnmongkol, Pakorn Nathong, Phatrasek Jirabovonvisut · 17d
The Attention Triangle in Audio-Video Models
arXiv cs.AI · Sagi Polaczek, Noa Kraicer, Gal Metzer, Zhuo Ning, Ali Mahdavi-Amiri, Daniel Cohen-Or, Raja Giryes · 17d
KC-Bench: A Dynamic Interactive Benchmark for Evaluating Knowledge Conflicts in LLM Agents
arXiv cs.AI · Yaxing Lyu, Shengjie Zhou, Binbin Toh, Pengyu Zhu, Lijun Li · 17d
A computable representation of the physical laboratory enables verifiable workflows
arXiv cs.AI · Xiaobo Li, Luyao Ge, Xiaohui Li, Lulu Guo, Ming Mao, Jiwang Zheng, Wenting Guan, Xin Yang, Yi Luo, Jun Jiang, Linjiang Chen · 17d
Analysis of Prompt Engineering for Drug Toxicity Prediction
arXiv cs.AI · Mia MacGregor, Aakash Welgamage Don, Mark Bartlett · 17d
Synthetic Semantic Supervision for Contrastive Code Representation Learning in Small Transformers: An Empirical Study
arXiv cs.AI · Kenneth Paulsen, Florian Tambon, Mike Papadakis, Shin Yoo · 17d
Counterfactual Routing Using Integer Programming with Constraint Generation
arXiv cs.AI · Dani\"el Vos, Sterre Lutz · 17d
Artificial Intelligence for Energy Optimization in Data Centers
arXiv cs.AI · Mohammed Basharath Ullah, Summaiya Unnisa Begum, Mohammed Nadeem Ullah · 17d
Proactive Service Agents: A Unified Decision Framework, Methods, and Evaluation
arXiv cs.AI · Yan Tang, Tingyu Cao, Yuanbo Tang, Huaze Tang, Keer Hu · 17d
SimSkill: A Lifelong Learning AI Agent for Autonomous Mastery of Traffic Simulation
arXiv cs.AI · Qi Liu, Qinzheng Wang, Yiming Bie · 17d
Rethinking World Models for Safety-Critical Embodied Systems
arXiv cs.AI · Kailang Ma, Heye Huang, Inhi Kim, Kitae Jang · 17d
DNative-Twin: Decision Graphs and Digital Twins for Reconstructable Agentic Decisions
arXiv cs.AI · Junjie Pang, Zhenzhen Xie, Haoke Han, Ying He, Jing Wang, Gang Liu · 17d
Transfiver: Human-AI Co-Inference through a Shared Editable State
arXiv cs.AI · Minji Park, Seunghyun Yoon, Hyuk Lim · 17d
Govern the Model, Not Only the Data: Storage, Circulation, and Learning in Creative AI
arXiv cs.AI · Phoenix Perry, George Simms, Elizabeth Wilson, Yasmine Boudiaf, Nick Bryan-Kinns, Tega Brain, R. Luke DuBois, Alix Rule, Rachel Meade Smith, Kelani Nichole, Atharva Pravin Pawar, Rebecca Fiebrink · 17d
SVG-Score: Human-Aligned Evaluation of Text-to-SVG Generation
arXiv cs.AI · Marco Cipriano, Leonardo Zini, Alexandra Schild, Valentin Teutschbein, Afsana Mimi, Marcella Cornia, Lorenzo Baraldi, Gerard de Melo · 17d
CauseCollab: Causal Unified and Modality-Agnostic Network for Heterogeneous Collaborative Perception
arXiv cs.AI · Weize Li, Yang Li, Quan Yuan, Xiaoyuan Fu, Guiyang Luo, Jinglin Li · 17d
Semantic Bayesian World Models
arXiv cs.AI · Tommaso Soru · 17d
Adapting to Evolving Requirements: Agentic AI for Retail Supply Chain Operations
arXiv cs.AI · Lei Zheng, Liping Yang, Zihao Li, Guodong Lyu, Chaik Ming Koh, Chung-Piaw Teo · 17d
Bioinfoysis Technical Report
arXiv cs.AI · Qingyang Shao, Xin Zhang, Zhouyang Yuan, Xianying Chen, Yujia Xiang, Zihao Yang, Tong Ye, Yangqi Zhang, Jiakang Xu, Xiaoqing Yan, Xuan Luo, Keyi Li, Enci Fan, Kai Kang, Zhuohan Liu, Xingyu Jin, Chunran Teng, Tao Li, Xinyu Lv, Minghui Wang, Wenfeng Li, Yidan Gao, Siyu Liu, Mingrui Luo, Zhu Liang, Guanren Qiao, Zhiping Xu · 17d
STAIR (STructure Aware Information Retriever): A novel dataset and LLM based retriever for document structure augmentation
arXiv cs.AI · Vineet Kumar, Meghanadh Pulivarthi, vishwajeet kumar, Jaydeep Sen, Riyaz Ahmad Bhat, Sachindra Joshi · 17d
Xiaomi-TabLDM: A Tabular Foundation Model Technical Report
arXiv cs.AI · TabLDM Team, Penghui Wang, Wei Liu, Hong Wang, Chengyue Huang, Yuxi Sun, Zirui Wang, Hongming Huang, Quan Wang, Chunxiao Liu, Erli Meng, Bin Wang · 17d
Inferring Affective Consciousness in an Artificial Agent: A Case Study
arXiv cs.AI · Mark Solms, St John Grimbly, Bruce Bassett, Evert Boonstra, Rowan Hodson, Nicolas Kuske, Kival Mahadew, Benjamin Rosman, Charel van Hoof, Jonathan Shock · 17d
Lose the Order, Keep the Hierarchy: Deordering HTN Plans
arXiv cs.AI · Takudzwa Togarepi, Gaspard Quenard, Damien Pellier, Humbert Fiorino · 17d
Value-Preserving Architectures for Agentic AI Systems
arXiv cs.AI · Alessandro Pesare, Tommaso Dolci, Katja Hose, Emanuel Sallinger · 17d
Speak for Me: Giving LLMs the Situational Awareness to Participate in a Meeting
arXiv cs.AI · Muneeb Khan, Frederic Kirstein, Terry Ruas, Bela Gipp · 17d
Towards Numerical TOHTN Planning with SMT-based HTN-SAT Encoding
arXiv cs.AI · Gaspard Quenard, Takudzwa Togarepi, Damien Pellier, Humbert Fiorino · 17d
More Criticism Does Not Make a Better Review: EquiReview-R
arXiv cs.AI · Zexing Zhang, Jichao Li, Tianyang Lei, Yude Fu, Yang Kewei · 17d
FiMI Banking: A Sovereign Model for Indian Retail Banking
arXiv cs.AI · NPCI AI Research Team, Aman Kumar, Asit Desai, Chandra Bhushan, Harsh Sharma, Harshit Bhushan, Hrithik Kadam, Keyur Doshi, Kolisetty Sai Kapardheeswar, Krishanu Adhikary, Nadeem Shaik, Navya Prakash, Nitin Kukreja, Prashant Devadiga, Shamanth MH, Shantanu Pandey, Suvradip Paul, Yatharth Dedhia · 17d
Interface-Induced Trajectory Censoring
arXiv cs.AI · Wenbo Wang · 17d
Common-Witness Certificates and Sharp Feature Bounds for Counterfactual Image Auditing
arXiv cs.AI · Usef Faghihi, Amir Saki · 17d
Occupancy-based Quantile Risk Control
arXiv stat.ML · Zihao Shi, Huajun Xi, Bingyi Jing, Hongxin Wei · 17d
A Closed-Form Formula for Consistent Lipschitz Regression on Metric Spaces with Sparse Neural Network Realizations
arXiv stat.ML · Ruiyang Hong, Hrad Ghoukasian, Anastasis Kratsios · 17d
ALRA: Adaptive Local Relational Alignment for Logit-Based Pre-training Distillation of Autoregressive Language Models
arXiv stat.ML · Quang Hoang Trung, Quang Huu Hieu, Nguyen Van Hoang Phuc, Vo Nguyen Le Duy · 17d
Towards a Statistical Understanding of Mixture-of-Experts
arXiv stat.ML · Siyuan He, Bokai Yang, Jie Hu, Ziwen Gao, Yuhong Yang · 17d
Towards Scaling Reinforcement Learning to Massive Populations: Learning Mean-Field Representations
arXiv stat.ML · Aditya Makkar, Benjamin Unger, Jeongyeol Kwon, Mathieu Lauri\`ere, Eugene Vinitsky, Yonathan Efroni · 17d
The Geometry of Ignorance: LLMs Know When to Temper Bayesian Priors
arXiv stat.ML · Toni J. B. Liu, Jiajun Bao, Yizhou Liu, Gurbir Arora, Nicolas Boull\'e, Rapha\"el Sarfati, Christopher J. Earls · 17d
Statistical Feature Augmentation for Anomaly Detection in Dynamic Graphs
arXiv stat.ML · Philipp Schlinge, Jean-Luc Schnipper, Martin Atzmueller · 17d
Tail-Likelihood Reinforcement Learning
arXiv stat.ML · Shrinivas Ramasubramanian, Daman Arora, Fahim Tajwar, Guanning Zeng, Qingyang Wu, Zhongzhu Zhou, Chenfeng Xu, Haiwen Feng, Yuda Song, Aarti Singh, Ruslan Salakhutdinov, J. Andrew Bagnell, Jeff Schneider, Andrea Zanette · 17d
TRACE: Spatiotemporal Contact Memory Graph Network Simulator for Granular Dynamics
arXiv stat.ML · Changjian Zhou, Negin Yousefpour, Jie Qi, Junfeng Fang, Guillermo A. Narsilio, Hans Petter Jostad · 17d
No-Regret Bayesian Optimization with Finite-Library Input-Warped Kernels
arXiv stat.ML · Edvin Ketabati Augustinsson, Robert A. Bridges · 17d
Evaluating Graph Neural Networks for Change-Criticality Classification in Maritime Navigation Charts
arXiv stat.ML · Abhishek Potnis, Jacob Arndt · 17d
Causal Foundation Models
arXiv stat.ML · Christopher Stith, Hossein Rahmani, Jesse C. Cresswell · 17d
Discrete Gromov-Wasserstein Duality: Algorithms and Isomorphism Testing
arXiv stat.ML · Gabriel Rioux, Joanna Marks, Riccardo Passeggeri, Ziv Goldfeld · 17d
Spectral characteristics of autoencoder parameters as a vector representation of data
arXiv stat.ML · Maria Nikitina, Anton Bishuk, Oleg Bakhteev · 17d
Correlated initialization of deep residual networks
arXiv stat.ML · Felix Benning, Ivan Nourdin, Giovanni Peccati · 17d
Q-Edge: Symmetry-Reduced Quantum Simulation of Structured Extreme Dependence
arXiv stat.ML · Hongrui Zhang, Paolo Recchia, Ying Chen · 17d
High-Dimensional Learning Dynamics of Attention-Indexed Models
arXiv stat.ML · Yizhou Xu, Margarita Sagitova, Lenka Zdeborov\'a, Florent Krzakala · 17d
Cooperative Multi-Task Semantic Communication for Joint Classification and Regression Tasks
arXiv stat.ML · Ahmad Halimi Razlighi, Mohammad Siddiqur Rahman, Maximilian H. V. Tillmann, Edgar Beck, Armin Dekorsy · 17d
Reliable Selection of Heterogeneous Treatment Effect Estimators
arXiv stat.ML · Jiayi Guo, Zijun Gao · 17d
Active learning for data-driven reduced models of parametric differential systems with Bayesian operator inference
arXiv stat.ML · Shane A. McQuarrie, Mengwu Guo, Anirban Chaudhuri · 17d
Deep networks learn to parse uniform-depth context-free languages from local statistics
arXiv stat.ML · Jack T. Parley, Francesco Cagnetta, Matthieu Wyart · 17d
Semiparametric Inference for Counterfactual Regression under Intervention-Driven Shift
arXiv stat.ML · Kwangho Kim · 17d
Online Learning of Functional Principal Component Analysis for Multidimensional Functional Data
arXiv stat.ML · Muye Nanshan, Nan Zhang, Jiguo Cao · 17d
Data-efficient Kernel Methods for Learning Hamiltonian Systems
arXiv stat.ML · Yasamin Jalalian, Mostafa Samir, Boumediene Hamzi, Peyman Tavallali, Houman Owhadi · 17d
Distributional Treatment Effect Transportability across Heterogeneous Sites
arXiv stat.ML · Borna Bateni, Yubai Yuan, Qi Xu, Annie Qu · 17d
Entropy-Generated Attention Beyond Softmax and Entmax: Kaniadakis and Reciprocal-Symmetric Abe Operators
arXiv stat.ML · Gunn Kim · 17d
Algebraic Invariants of Lightning Self-Attention
arXiv stat.ML · Yulia Alexandr, Hao Duan, Guido Mont\'ufar · 17d
Equation Recast for Canonical Operator Learning Across Parametric PDEs
arXiv cs.LG · Qiyun Cheng, Valentin Duruisseaux, Cesar F. Clauser, Md Hossain Sahadath, Huihua Yang, Shaowu Pan, Nathaniel Ferraro, Anima Anandkumar, Wei Ji, Cristina Rea · 17d
From Euclidean to Graph-Structured Data: A Survey of Collaborative Learning
arXiv cs.LG · R\'emi Bourgerie, \v{S}ar\=unas Girdzijauskas, Viktoria Fodor · 17d
Modern Transformers Are Implicit Hybrids: From Functional Differentiation to Principled Hybrid Architecture Design
arXiv cs.LG · Runlin Shi, Bojian Yin, Guoqi Li · 17d
Mesh-Native Physics-Informed Graph Surrogates for TCAD-in-the-Loop Design Space Exploration
arXiv cs.LG · Leonid Popryho, Ayoub Sadeghi, Inna Partin-Vaisband · 17d
Verify Before You Distill: Prompt-Level Teacher Gating for On-Policy Distillation
arXiv cs.LG · Zhiwei Zhang, Zechen Sun, Fei Zhao, Kang Peng, Bin Liang, Huayu Deng, Yao Hu, Kam-Fai Wong, Mu Chuan · 17d
ObserverBench: Testing Mechanistic Estimates for Intervention and Control
arXiv cs.LG · Vijay Erramilli · 17d
Learnable composition for neural operators
arXiv cs.LG · Zituo Chen, Baiming Zhang, Sili Deng · 17d
LeanStream: A Speculate-and-Refine Streaming Framework for Efficient on-Device LLM Inference
arXiv cs.LG · Renyuan Liu (Richard), Yuyang Leng (Richard), Kaiyan Liu (Richard), Yuzhou Zhong (Richard), Shaohan Hu (Richard), Chun-Fu (Richard), Chen, Peijun Zhao, Heechul Yun, Shuochao Yao · 17d
The Gradient Does Not See Rank: Rank-Indifference in Matrix-CODI on ProsQA
arXiv cs.LG · Samuel Larson (Pebble ML) · 17d
Distilling deep optical flow stereo methods to retrieve dense three-dimensional wind fields
arXiv cs.LG · Thomas J. Vandal, Dong L. Wu, James L. Carr, Derek J. Posselt, Elise Penn, Tristan Ballard, August Posch, Kate Duffy · 17d
Scaling Laws, Tabular Data and Actuarial Ratemaking Models
arXiv cs.LG · Ronald Richman · 17d
Kernel Reboot: Breaking the Boundaries of Neural Tangent Kernels for Neural Fields
arXiv cs.LG · Amir Mallak, Alaa Maalouf, Lior Wolf, Daniela Rus, Dan Rosenbaum · 17d
Routing Is Not Enough: Diagnosing Intra-Adapter Subspace Contention in MoE+LoRA Fine-Tuning
arXiv cs.LG · Mehreen Hossain Chowdhury, Nowshin Mahjabin, Ahmed Shafin Ruhan, Md Azam Hossain, Abu Raihan Mostofa Kamal, Md Tahmid Rahman Laskar · 17d
Frontier LLMs are effective batch optimizers: Assessing reasoning models in continuous and discrete settings
arXiv cs.LG · Frank Hu, Shriram Chennakesavalu, David Graff · 17d
Portable Causal Fairness Across Synthetic Data Generator Families
arXiv cs.LG · Steven Golob, Sikha Pentyala, Martine De Cock · 17d
Language-encoded network topology enables large language models to reason about complex networks
arXiv cs.LG · Ucchwas Talukder Utsha, Sakib Mostafa, James Zou, Md Tauhidul Islam · 17d
The 2026 PNPL Competition: Word Classification and Efficient Cross-Subject Generalisation in LibriBrain100
arXiv cs.LG · Francesco Mantegna, Gereon Elvers, Dulhan Jayalath, Gilad Landau, Tasha Kim, Miran \"Ozdogan, Luisa Kurth, Teyun Kwon, SungJun Cho, Benjamin Ballyk, Alex Fung, Anna Greer, Pratik Somaiya, Christian Herff, Yorguin Mantilla Ramos, Hamza Abdelhedi, Karim Jerbi, Greg Farquhar, Brendan Shillingford, Mark Woolrich, Oiwi Parker Jones · 17d
B2B Customer Conversion Prediction: A Document Representation, Graph Theory, and CatBoost Driven Methodology
arXiv cs.LG · Tianqi Wang, Sheikh Shams Azam, Wan Eih Huang, Anton Wiranata, Christopher G. Brinton, Jan P. Allebach · 17d
FlowBalance: Verifier-Grounded Self-Improvement from On-Policy Reasoning Experience
arXiv cs.LG · Zixun Huang, Kishan Panaganti, Haitao Mi, Leowei Liang · 17d
Selective Hypergraph Refinement for Frozen Graph Clustering
arXiv cs.LG · Zimo Si · 17d
Latent Energy Action Planning with World Models
arXiv cs.LG · Phu Pham, Aniket Bera · 17d
Geometry-Aware Graph Construction via Adaptive Spectral Bandwidth Control
arXiv cs.LG · Ecem Bozkurt, Antonio Ortega · 17d
Risk and Anomaly Identification for Distribution Network Optimal Operation Based on Reinforcement Learning and Uncertainty Quantification
arXiv cs.LG · Ziqi Zhang · 17d
DE-Venus: A Data-Efficient RLVR Framework for Large Language Models
arXiv cs.LG · Shenzhi Yang, Guangcheng Zhu, Kai Tang, Zhengqing Zang, Xing Zheng, Haobo Wang, Yingfan Ma, Bowen Song, Bo Han, Bo An, Lei Feng, Weiqiang Wang, Junbo Zhao, Gang Chen · 17d
A Large Open Multi-Energy Corpus of Soil Compaction Tests, with Machine-Learning Baselines
arXiv cs.LG · Sompote Youwai, Chana Phutthananon, Warat Kongkitkul · 17d
Gradients Know What Outcomes Don't: Unlocking Reinforcement Learning for LLM Reasoning with Gradient-Aligned Rewards
arXiv cs.LG · Leqi Zheng, Jinbo Su, Fang Niu, Chaokun Wang, Weiping Wang, Jiajun Zhang, Shannan Yan, Jie Wu, Zhaolu Kang, Rong Fu, Hang Zhang · 17d
From Zero to Hero: An Open LLM Ecosystem for Armenian
arXiv cs.LG · Erik Arakelyan, Khatun Avetisyan, Meri Davtyan, Heghine Grigoryan, Nane Khachatryan, Hayk Shahsuvaryan, Henrik Sergoyan, Vahan Martirosyan · 17d
Time Without Timesteps: Simulating Coupled Dynamical Systems via Self-Consistency
arXiv cs.LG · Liyu Zerihun, Mark Shinyoung Lee · 17d
SimpleDesign: A Joint Model for Protein Sequence and Structure Codesign
arXiv cs.LG · Jiarui Lu, Yuyang Wang, Yizhe Zhang, Jiatao Gu, Navdeep Jaitly, Joshua M. Susskind, Miguel \'Angel Bautista · 17d
RecurTrace: Adaptive Latent Reasoning with Loop-Time Memory
arXiv cs.LG · Yuxiang Wang, Kunyu Feng, Yingda Shen, Haoning Xu, Junyu Wang, Zhizheng Wu · 17d
TIGPO: Temporal Instance-Graph Policy Optimization for Long-Horizon LLM Agents
arXiv cs.LG · Jinwei Gan · 17d
Inferred Generative-Process Diversity Predicts Correlated Failure Across Language Models
arXiv cs.LG · Ross Tieman, Evan Markou · 17d
TraveL: Transformer-based Multi-view Path Distributional Representation Learning
arXiv cs.LG · Fang He, Tao-yang Fu, Wang-chien Lee · 17d
It's the Problem, Not the Path: Budget and Difficulty Confounds in LLM Reasoning Trajectories
arXiv cs.LG · Yigit Utku Bulut · 17d