Skip to content
TILens What matters today in tech v0.0.5
Theme

Daily edition · AI Research

The daily ledger

TILens turns technical updates into a focused daily brief: official releases, trusted reporting, and practitioner analysis, deduplicated and organized by topic.

20 Aug 2026 edition
AI Research AI

The Diffusion-Attention Connection

arXiv:2604.09560v2 Announce Type: replace Abstract: Softmax attention is the row-normalized operator of a diffusion map: both normalize a learned score into a Markov operator, and differ only in what the score is…

Source: arXiv cs.LG Julio Candanedo
AI Research AI

A Real-Time Tsetlin Machine-based Non-intrusive Load Monitoring System on MCUs

arXiv:2608.18780v1 Announce Type: new Abstract: Non-Intrusive Load Monitoring (NILM) systems estimate individual appliance energy consumption from a single aggregate meter, without requiring separate sensors for each…

Source: arXiv cs.LG Tianhang Tan, Han Wu, Tousif Rahman, Shengyu Duan, Alex Yakovlev, Rishad Shafik
AI Research AI

On the Power of Source Screening for Learning Shared Feature Extractors

arXiv:2602.16125v3 Announce Type: replace Abstract: Learning with shared representation is widely recognized as an effective way to separate commonalities from heterogeneity across various heterogeneous sources. Most…

Source: arXiv cs.LG Leo Muxing Wang, Connor Mclaughlin, Lili Su
AI Research AI

Rethinking Privileged Information in On-Policy Self-Distillation

arXiv:2608.18271v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) trains a student on its own responses using token-level supervision from the same model conditioned on privileged reference information.…

Source: arXiv cs.LG Samyak Shrestha, Alexander Tessier
AI Research AI

GigaBrain-WBC-0.5: A Behavior World Model for Robust Whole-Body Control with Environment Interaction

arXiv:2608.18234v1 Announce Type: cross Abstract: Whole-body motion tracking policies turn a humanoid into a robust control interface: the teleoperator---or an upstream model---only supplies a coarse movement intent,…

Source: arXiv cs.LG Ziyang Cheng, Tianshu Tang, Jinxin Lan, Xinze Chen, Yuhan Gong, Zhichao Liu, Changzhong Wu, Yahao Mao, Zongyan Deng, Mingxuan Ma, Huasen Xi, Yilong Liu, Yutong Wu, Xiaofeng Wang, Yang Wang, Yun Ye, G…
AI Research AI

Approximate Speculative Decoding

arXiv:2608.03447v2 Announce Type: replace Abstract: Speculative decoding accelerates autoregressive generation by verifying a draft block with a target model in parallel. Under standard greedy verification, decoding…

Source: arXiv cs.LG Yuannuo Feng, Zegang Peng, Yuxin Xie, Yubing Ye, Yizhe Chen, Wenshuai Yao, Wenyong Zhou, Wang Kang
AI Research AI

Coordination on a Budget: Federated Active Learning with Few Labels

arXiv:2608.18634v1 Announce Type: new Abstract: Federated Active Learning (FAL) addresses the dual challenges of data privacy and label scarcity, where the absence of a global data view introduces additional hurdles for…

Source: arXiv cs.LG Liam Mohr, Daphna Weinshall
AI Research AI

Denoising-Aware Inversion: Revealing Privacy Risks in Noise-Protected Text Embeddings

arXiv:2608.18610v1 Announce Type: new Abstract: Dense text embeddings are widely used in data mining, retrieval, and downstream machine learning systems due to their compact and semantically rich representations, but…

Source: arXiv cs.LG Yubo Wang, Shujie Cui, James Bailey, Hongzhi Yin, Wenyu Liang, Min Tang, Shiyue Qin, Weiqing Wang
AI Research AI

Sobolev Regularized Score Difference Estimation in Diffusion Models

arXiv:2608.18237v1 Announce Type: cross Abstract: Estimating the difference of two Stein's score functions is a fundamental problem in generative modeling. In particular, score differences arise naturally in transfer…

Source: arXiv cs.LG Chenghan Xie, Jose Blanchet, Renyuan Xu
AI Research AI

A continually expandable foundation model for brain MRI

arXiv:2608.08319v2 Announce Type: replace-cross Abstract: Brain magnetic resonance imaging (MRI) is central to neuroscience and clinical assessment, but models are commonly developed for individual diseases, populations…

Source: arXiv cs.LG Michail Mamalakis, Carmen Jimenez-Mesa, Yonghao Li, Hao Chen, Chao Li, Antonios Mamalakis, John Suckling, Richard Bethlehem, Stephen J. Price, Richard J. Gilbertson, Pietro Lio
AI Research AI

Low-Power, Neuromorphic, Acoustic Anomaly Detection for Persistent Machine Monitoring

arXiv:2608.18341v1 Announce Type: cross Abstract: Persistent acoustic monitoring can detect machine faults without physical contact, but always-on inference is constrained by power, latency, and deployment complexity.…

Source: arXiv cs.LG Steven C. Nesbit (Information Sciences, CAI-3, Los Alamos National Laboratory, Los Alamos, USA), Victor M. Vergara (AeroVironment Inc., Albuquerque, USA), Michael A. Felix (University of New Mexico C…
AI Research AI

Mechanistic Interpretability of Structure-Aware Numerical Reasoning in LLaMA 3.1 8B

arXiv:2608.18419v1 Announce Type: new Abstract: Recent work has shown that large language models (LLMs) exhibit strong numerical sequence modeling capabilities and show promise in time-series prediction. While LLMs…

Source: arXiv cs.LG Rahul Chowdhury, Timothy A Rupprecht, Senhao Cao, Jiahao Liu, Octavia Camps, David Bau, Pu Zhao, Yanzhi Wang
AI Research AI

Training Chemical Plausibility-Aware Large Language Models for Single-Step Retrosynthesis

arXiv:2608.18940v1 Announce Type: new Abstract: Single-step retrosynthesis is a central component of computer-aided synthesis planning, yet its intrinsically one-to-many nature is poorly captured by single-answer…

Source: arXiv cs.LG Bogdan Zagribelnyy, Ivan Ilin, Nikita Bondarev, Maksim Kuznetsov, Mathieu Reymond, Vladimir Aladinskiy, Alex Aliper, Alex Zhavoronkov
AI Research AI

Mechanist: AI as a Scientific Instrument for Discovering the Mechanisms of Intelligence

arXiv:2608.12036v2 Announce Type: replace-cross Abstract: AI models have achieved remarkable success across diverse domains, yet the mechanisms underlying their capabilities and the risks they may pose remain poorly…

Source: arXiv cs.LG Mengru Wang, Junfeng Fang, Shuofei Qiao, Zhenqian Xu, Haoming Xu, Haoxiong Wang, Shumin Deng, Linyi Yang, Zhixiang Cui, Xin Xu, Yunzhi Yao, Buqiang Xu, Fei Shen, Haozhe Luo, Yunxiang Wei, Ningyu Zhan…
AI Research AI

A Unifying Relational Perspective on Expressive Lottery Tickets

arXiv:2608.18819v1 Announce Type: new Abstract: Graph neural networks (GNNs) are widely used, but how parameter sparsity affects the expressivity of relational (RGNNs) and temporal (TGNNs) variants is poorly understood.…

Source: arXiv cs.LG Lorenz Kummer, Samir Moustafa, Anatol Ehrlich, Franka Bause, Marco Nennstiel, Przemys{\l}aw Andrzej Wa{\l}\c{e}ga, Nils Morten Kriege
AI Research AI

Quantum Tensor Network Learning with DMRG

arXiv:2608.18901v1 Announce Type: cross Abstract: Tensor Networks are a relatively new machine learning approach. The architectures proposed initially are inspired by approaches from quantum many-body physics…

Source: arXiv cs.LG Gustav J L J\"ager, Martin B Plenio, Hans-Martin Rieser
AI Research AI

Trust as a Field: A Macroscopic Representation for Vehicular Networks

arXiv:2608.18178v1 Announce Type: cross Abstract: Trust assessment is a fundamental component of cooperative and connected vehicle systems. However, existing approaches operate primarily at the level of individual…

Source: arXiv cs.LG Md Mahmudul Islam, Shaurya Agarwal
AI Research AI

AMPLIFAI: A Multiphase CT Dataset for Benchmarking Clinical Reasoning in LI-RADS Assessment of Liver Lesions

arXiv:2608.14778v2 Announce Type: replace-cross Abstract: Hepatocellular carcinoma (HCC) is the third leading cause of cancer-related mortality worldwide, with early detection improving survival from <20% to >70%. The…

Source: arXiv cs.LG Pranav Kulkarni, Nikhil Shah, Amritansh Suryavanshi, Jana G. Delfino, James Tonascia, Jade Wong-You-Cheong, Barton Lane, Joseph Chirico, Jeffrey D. Hirsch, Ang Li, Heng Huang, Florence X. Doo
AI Research AI

Weak-to-Strong Generalization via Bregman Bias-Variance Decomposition

arXiv:2505.24313v3 Announce Type: replace Abstract: Weak-to-strong generalization (W2SG) is the phenomenon in which a powerful student model, trained on labels produced by a weaker teacher, ultimately outperforms the…

Source: arXiv cs.LG Gengze Xu, Wei Yao, Ziqiao Wang, Yong Liu
AI Research AI

Vector Symbolic Policy Gradient

arXiv:2608.18404v1 Announce Type: new Abstract: We answer this question with Vector-Symbolic Policy Gradient (VSPG), a discrete-action actor that represents each action by a unit-norm hypervector and scores it by…

Source: arXiv cs.LG Ryozo Masukawa, Sanggeon Yun, SungHeon Jeong, Hyunwoo Oh, Raheeb Hassan, Pietro Mercati, Nathaniel D. Bastian, Mahdi Imani, Mohsen Imani
AI Research AI

Reinforced Planning with Latent World Models

arXiv:2608.18669v1 Announce Type: new Abstract: Humans solve complex problems by constructing plans and mentally simulating their outcomes with an internal model of the world. Machine learning has produced world models…

Source: arXiv cs.LG Armin Sommer, Jannik Schilling
AI Research AI

Fair Multi-View Determinantal Coresets via Adaptive NEPv

arXiv:2608.18181v1 Announce Type: cross Abstract: Selecting a small, diverse subset from a large candidate pool often means balancing several incompatible notions of diversity. In trademark curation, for instance, a…

Source: arXiv cs.LG Richard Yi Da Xu
AI Research AI

ChiroEcho: extending automated bat vocalisation classification beyond the learned taxonomy

arXiv:2608.18191v1 Announce Type: new Abstract: Bats are key indicators of ecosystem health and are protected throughout Europe, making reliable population monitoring a conservation priority. Their cryptic nocturnal…

Source: arXiv cs.LG Burooj Ghani, Welmoed Eversteijn, Milan van Hirtum, Juan Sebasti\'an Ca\~nas, Vincent J. Kalkman, Dan Stowell, A. Leonie Baier
AI Research AI

The Road Taken: The Role of Optimizers at the Edge of Stability

arXiv:2608.18415v1 Announce Type: new Abstract: The edge of stability refers to a phenomenon in deep learning with gradient-based optimizers where the Hessian eigenvalues of the loss remain stable above a threshold that…

Source: arXiv cs.LG Jaerin Lee, Kyoung Mu Lee
AI Research AI

WhiteMatter: All-to-All Cross-Layer Connections via KV Mixing

arXiv:2608.18486v1 Announce Type: cross Abstract: In a Transformer, each layer attends to past tokens only through KV produced at its own depth, despite the presence of deeper representations during autoregressive…

Source: arXiv cs.LG Wenbo Zhang, Xiang Ren
AI Research AI

Preference Reasoning under Indeterminacy in Large Language Models

arXiv:2608.18631v1 Announce Type: cross Abstract: As large language models evolve into decision-making agents, the ability to reason over preferences becomes fundamental to alignment, coordination, and collective…

Source: arXiv cs.LG Hadi Hosseini, Samarth Khanna, Xiyuan Wang
AI Research AI

Entropy-Constrained Adaptive Stochastic Quantization

arXiv:2608.18147v1 Announce Type: new Abstract: Adaptive stochastic quantization (ASQ) is a recently introduced quantization approach that optimizes the Mean Squared Error (MSE) for a given input while preserving…

Source: arXiv cs.LG Ran Ben Basat, Yaniv Ben-Itzhak, Michael Mitzenmacher, Shay Vargaftik
AI Research AI

Online Learning for Dynamic Constellation Topologies

arXiv:2603.25954v2 Announce Type: replace Abstract: The use of satellite networks has increased significantly in recent years due to their advantages over purely terrestrial systems, such as higher availability and…

Source: arXiv cs.LG Jo\~ao Norberto, Ricardo Ferreira, Cl\'audia Soares
AI Research AI

Hallucination Detection in Large Language Models Using Diversion Decoding

arXiv:2607.10476v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have emerged as a powerful tool for retrieving knowledge through seamless, human-like interactions. Despite their advanced text…

Source: arXiv cs.LG Basel Abdeen, S M Tahmid Siddiqui, Meah Tahmeed Ahmed, Anoop Singhal, Latifur Khan, Punya Parag Modi, Ehab Al-Shaer
AI Research AI

Performance Drift Detection in Machine Learning as a Service (MLaaS) for IoT Environments

arXiv:2608.18555v1 Announce Type: new Abstract: Machine Learning as a Service (MLaaS) is a powerful cloud paradigm enabling data-driven intelligent applications in Internet of Things (IoT) environments, widely adopted…

Source: arXiv cs.LG Deepak Kanneganti, Sajib Mistry, Sheik Mohammad Mostakim Fattah, Erik Elmroth, Aneesh Krishna, Monowar Bhuyan
AI Research AI

A Configuration-First Framework for Reproducible, Low-Code Machine Learning: a Localization Use Case

arXiv:2510.25692v5 Announce Type: replace-cross Abstract: As machine learning underpins more critical applications, the value of a reported result depends on whether it can be compared and repeated. In practice, this…

Source: arXiv cs.LG Tim Strnad (Jo\v{z}ef Stefan Institute, Slovenia), Bla\v{z} Bertalani\v{c} (Jo\v{z}ef Stefan Institute, Slovenia), Carolina Fortuna (Jo\v{z}ef Stefan Institute, Slovenia)
AI Research AI

LionMuon: Alternating Spectral and Sign Descent for Efficient Training

arXiv:2605.19811v3 Announce Type: replace Abstract: In large-scale optimization, the cheapness and effectiveness of update steps are the most crucial factors for a successful optimizer. Sign-based optimizers like Lion…

Source: arXiv cs.LG Arman Bolatov, Artem Riabinin, Nikita Kornilov, Andrey Veprikov, Samuel Horv\'ath, Martin Tak\'a\v{c}, Aleksandr Beznosikov
AI Research AI

Residual Algebra for Representation-Preserving Learning

arXiv:2608.07349v2 Announce Type: replace Abstract: Learning from heterogeneous representations is often reduced to feature concatenation, erasing which representation produced each error. We propose residual algebra,…

Source: arXiv cs.LG Yao Wu
AI Research AI

Transportable Causal Effect Estimation across Networks under Interference

arXiv:2608.18932v1 Announce Type: new Abstract: Estimating causal effects under network interference typically assumes that the network used for training and the network used for deployment coincide. In practice, an…

Source: arXiv cs.LG Xiaojing Du, Jiuyong Li, Lin Liu, Debo Cheng, Jixue Liu, Thuc Duy Le
AI Research AI

SHANG++: Robust Stochastic Acceleration under Multiplicative Noise

arXiv:2603.09355v2 Announce Type: replace-cross Abstract: Under the multiplicative noise scaling (MNS) condition, original Nesterov acceleration is provably sensitive to noise and may diverge when gradient noise…

Source: arXiv cs.LG Yaxin Yu, Long Chen, Minfu Feng
AI Research AI

Pretraining Reusable Inference Across Views with Synthetic Task Priors

arXiv:2608.19115v1 Announce Type: new Abstract: Modern pretrained encoders make representations from heterogeneous views increasingly reusable, but the procedure that determines view utility and combines evidence is…

Source: arXiv cs.LG Jielong Lu, Zhihao Wu, Jiajun Yu, Zhaoliang Chen, Haishuai Wang
AI Research AI

The Embodiment Gap in Robot Foundation Models

arXiv:2608.18433v1 Announce Type: cross Abstract: Robot foundation models (RFMs), including vision-language-action (VLA) policies, are often discussed through a scaling view: more data, larger models, and broader…

Source: arXiv cs.LG Yukiyasu Domae, Keisuke Shirai, Hanbit Oh, Ryoichi Nakajo, Tomohiro Motoda, Koshi Makihara, Masaki Murooka, Takuma Yagi, Yoshiaki Bando, Ryo Hanai
AI Research AI

Understanding Multilingual Medical ASR Adaptation Through Layer-Wise Analysis

arXiv:2608.18825v1 Announce Type: cross Abstract: Medical automatic speech recognition (MedASR) requires adaptation to specialised terminology, limited annotated clinical data, and multilingual use cases. Although…

Source: arXiv cs.LG Souranil Kahali, Rituparna Bose, Abner Hernandez, Tomas Arias-Vergara, Andreas Maier, Ning Ma, Paula Andrea Perez-Toro
AI Research AI

Physics-Unrolled Neural Operator for Wireless Field Modeling

arXiv:2608.18495v1 Announce Type: new Abstract: Radio maps are essential for wireless decision-making tasks such as access-point placement, coverage planning, and localization, but their fine spatial details are…

Source: arXiv cs.LG Rafid Umayer Murshed, Saif Ur Rahman, Mingyue Tang, Elahe Soltanaghai
AI Research AI

Fast Best-in-Class Regret for Contextual Bandits

arXiv:2510.15483v3 Announce Type: replace-cross Abstract: We study the problem of stochastic contextual bandits in the agnostic setting, where the goal is to compete with the best policy in a given class without…

Source: arXiv cs.LG Samuel Girard, Aurelien Bibaut, Arthur Gretton, Nathan Kallus, Houssam Zenati
AI Research AI

Graphical Design of Interpretable Architectures

arXiv:2608.18936v1 Announce Type: new Abstract: Designing, implementing, and comparing interpretable architectures requires a formal language to represent them. The most common representations fall short in one of two…

Source: arXiv cs.LG Pietro Barbiero
AI Research AI

Geometric Iterative Retrieval for Neural Audio Codec Resynthesis

arXiv:2608.19141v1 Announce Type: cross Abstract: Neural audio codecs based on Residual Vector Quantization (RVQ) have become the dominant discrete representation for token-based general audio generation, yet…

Source: arXiv cs.LG Leo Schmidt-Traub, Fr\'ed\'eric Berdoz, Luca A. Lanzend\"orfer, Roger Wattenhofer
AI Research AI

Conformal Policy Control

arXiv:2603.02196v4 Announce Type: replace-cross Abstract: An agent must try new behaviors to explore and improve. In high-stakes environments, an agent that violates safety constraints may cause harm and must be taken…

Source: arXiv cs.LG Drew Prinster, Clara Fannjiang, Ji Won Park, Kyunghyun Cho, Anqi Liu, Suchi Saria, Samuel Stanton
AI Research AI

Compress and Forget: bitsandbytes Quantization Amplifies Proactive Interference in LLMs

arXiv:2608.18578v1 Announce Type: cross Abstract: Proactive interference (PI) is a documented failure mode in large language models in which retrieval of a repeatedly overwritten value degrades as prior overwrites…

Source: arXiv cs.LG Shayan Shahrabi-Farahani (Shahid Beheshti University, Tehran, Iran), Dara Rahmati (Shahid Beheshti University, Tehran, Iran)
AI Research AI

Redakto - The Incognito Tab for LLMs

arXiv:2608.18260v1 Announce Type: cross Abstract: Large Language Models (LLMs) are being increasingly used in everyday applications. A major challenge in the context of LLMs or Artificial Intelligence (AI) in general is…

Source: arXiv cs.LG Saurav Kumar Saha, Tom R\"ohr, Felix Bie{\ss}mann
AI Research AI

Visual-Prompt Guided Wildlife Instance-Level Recognition

arXiv:2608.18246v1 Announce Type: cross Abstract: Fine-grained wildlife re-identification remains a challenging area in research. Current state-of-the-art approaches apply a detection and re-identification pipeline. We…

Source: arXiv cs.LG Mufhumudzi Muthivhi, Jiahao Huo, Terence van Zyl, Fredrik Gustafsson
AI Research AI

Topology-Aware Differential Privacy in Hierarchical Federated Learning

arXiv:2506.19260v3 Announce Type: replace-cross Abstract: Hierarchical federated learning places regional aggregators between clients and the cloud, so a participant's update is observed only alongside its neighbours'.…

Source: arXiv cs.LG Murtaza Rangwala, Richard O. Sinnott, Rajkumar Buyya
AI Research AI

Lost in Aggregation: How Benchmarks Overlook Irreplaceable Model Strengths

arXiv:2608.18919v1 Announce Type: new Abstract: Tabular machine learning benchmarks typically summarize performance by averaging scores, ranks, or pairwise wins across datasets. Such aggregates are useful for selecting…

Source: arXiv cs.LG Andrej Tschalzev, Stefan L\"udtke, Heiner Stuckenschmidt, Christian Bartelt
AI Research AI

What Makes Software Issue Resolution Tasks Difficult for Agents?

arXiv:2608.18280v1 Announce Type: cross Abstract: Background. Advances in agentic systems are simultaneously, and rapidly, saturating benchmarks. Despite this often discussed phenomena, benchmark scores remain difficult…

Source: arXiv cs.LG Ebtesam Al-Haque, Brittany Johnson
AI Research AI

Off-Manifold Collapse in Guided Protein Language Models

arXiv:2608.18597v1 Announce Type: new Abstract: Protein language models are widely used priors for protein sequence design, and a growing body of work controls them at inference time as an alternative to fine-tuning.…

Source: arXiv cs.LG Shuibai Zhang, Xinchi Liu, Fred Zhangzhi Peng, Zhihan Yang, Shutong Wu, Yingzi Ma, Jiawei Zhang
AI Research AI

Jailbreaking in the Haystack

arXiv:2511.04707v2 Announce Type: replace-cross Abstract: Recent advances in long-context language models (LMs) have enabled million-token inputs, expanding their capabilities across complex tasks like computer-use…

Source: arXiv cs.LG Rishi Rajesh Shah, Chen Henry Wu, Shashwat Saxena, Ziqian Zhong, Alexander Robey, Aditi Raghunathan
AI Research AI

Looped Language Models Improve Compositional Tool Calling

arXiv:2608.18171v1 Announce Type: cross Abstract: Looped language models have shown promising results on reasoning benchmarks, yet their potential for agentic tool use remains largely unexplored. We study this question…

Source: arXiv cs.LG Andrei Cristian Popescu, Haitz S\'aez de Oc\'ariz Borde, Pietro Li\`o
AI Research AI

Optimizing Energy Efficiency and Grid Stability via Public EV Charging Flexibility

arXiv:2608.18126v1 Announce Type: cross Abstract: This study evaluates the potential of electric vehicle (EV) charging flexibility to enhance both energy efficiency and power grid stability. Using real-world data from…

Source: arXiv cs.LG Marek Miltner, Artem Bryksa, Ond\v{r}ej \v{S}togl, Daniel Va\v{s}ata, Magda Friedjungov\'a, Ram Rajagopal, Old\v{r}ich Star\'y
AI Research AI

GraphK: Variable-Size Graph Generation with Efficient Edge Construction

arXiv:2608.18777v1 Announce Type: new Abstract: Graph generation models have advanced significantly with deep learning, yet they remain limited in scalability, flexibility, and ability to model underlying structures. We…

Source: arXiv cs.LG Resul Tugay, Eren Olu\u{g}, Elif Ak, Sule Gunduz Oguducu
AI Research AI

From Inference to Adaptation: A Unified Optimal Transport View of Vision Language Model

arXiv:2608.18339v1 Announce Type: cross Abstract: Vision-language models (VLMs) have demonstrated remarkable zero-shot capabilities yet remain sensitive to real-world distribution shifts during inference. Although…

Source: arXiv cs.LG Qi Yu, Zhichen Zeng, Katherine Tieu, Xiyuan Yang, Ruizhong Qiu, Yuchen Yan, Lihui Liu, Yanjun Zhao, Lingjie Chen, Jingrui He, Hanghang Tong
AI Research AI

Allocating Recurrent Compute in Looped Language Models

arXiv:2608.18230v1 Announce Type: new Abstract: Looped language models improve reasoning and knowledge manipulation by applying shared computation repeatedly. Existing systems usually repeat an entire layer stack,…

Source: arXiv cs.LG Ruhai Lin, Yiyang Guo, Rui-Jie Zhu, Hao Ye, Jason K. Eshraghian
AI Research AI

Tensor Field Models

arXiv:2608.18808v1 Announce Type: new Abstract: This paper introduces Tensor Field Models (TFMs), realization-level Mathematical Structures in which a learned Operator maps a product of admissible component-section…

Source: arXiv cs.LG Alexander Strunk, Roland Assam
AI Research AI

DeGLIF for Label Noise Robust Node Classification using GNNs

arXiv:2506.00244v2 Announce Type: replace Abstract: Noisy labelled datasets are generally inexpensive compared to clean labelled datasets, and the same is true for graph data. In this paper, we propose a denoising…

Source: arXiv cs.LG Pintu Kumar, Nandyala Hemachandra
AI Research AI

SkillNet: Create, Evaluate, and Connect AI Skills

arXiv:2603.04448v2 Announce Type: replace-cross Abstract: Current AI agents can flexibly invoke tools and execute complex tasks, yet their long-term advancement is hindered by the lack of systematic accumulation and…

Source: arXiv cs.LG Yuan Liang, Ruobin Zhong, Haoming Xu, Chen Jiang, Yi Zhong, Runnan Fang, Jia-Chen Gu, Shumin Deng, Yunzhi Yao, Mengru Wang, Shuofei Qiao, Yida Xue, Xin Xu, Tongtong Wu, Kun Wang, Yang Liu, Zhen Bi, J…
AI Research AI

A Cloud-Edge System for Multimodal Clinical Screening in Resource-Constrained Rural Settings

arXiv:2608.12745v2 Announce Type: replace Abstract: Medical AI has demonstrated specialist-level diagnostic accuracy, yet these capabilities remain largely inaccessible in resource-constrained rural settings where…

Source: arXiv cs.LG Hei Ting (Una), Chan, Chenwei Wu, Xueshen Liu, Zesen Zhao, Boyuan Zheng, Luis Filipe Nakayama, Michael G. Morley, Liyue Shen, Jiasi Chen, Z. Morley Mao
AI Research AI

Learning Topological Features of $\widehat Z$-invariants

arXiv:2608.18570v1 Announce Type: cross Abstract: Machine learning and data analysis techniques have recently emerged as powerful tools for identifying patterns and formulating conjectures in mathematical research, most…

Source: arXiv cs.LG Brandon Robinson, Shimal Harichurn, Fabian Ruehle, Sergei Gukov, Rak-Kyeong Seong, Miranda C. N. Cheng
AI Research AI

Progressive Experience Fusion for Multi-Task World Model Control in Endovascular Navigation

arXiv:2608.18647v1 Announce Type: cross Abstract: Autonomous endovascular navigation could support the delivery of mechanical thrombectomy to underserved areas, but controllers must navigate long, multi-stage paths…

Source: arXiv cs.LG Harry Robertshaw, Maxence Boels, Nikola Fischer, Sebastien Ourselin, Christos Bergeles, Alejandro Granados, Thomas C Booth
AI Research AI

A Factor Graph Approach to Scalable Multi-Output Gaussian Process Regression

arXiv:2608.11917v2 Announce Type: replace Abstract: Multi-output Gaussian process regression scales cubically in the number of observations times outputs, and dense kernel-matrix methods need bespoke handling whenever…

Source: arXiv cs.LG Wouter W. L. Nuijten, Esther G. van Pelt, Albert Podusenko, \.Ismail \c{S}en\"oz, Wouter M. Kouw
AI Research AI

Graph-Based Approaches to Learning Epileptogenic Zone Localization Using Stereo-EEG Recordings

arXiv:2608.18887v1 Announce Type: new Abstract: The epileptogenic zone (EZ) is the brain region that generates seizures in an individual, and is the target of epilepsy surgery. Localizing the EZ from stereo-EEG (sEEG)…

Source: arXiv cs.LG Daniel Wendelken (University of Cincinnati, Cincinnati, USA), Brian Ervin (Cincinnati Children's Hospital Medical Center, Cincinnati, USA), Ravindra Arya (Cincinnati Children's Hospital Medical Cente…
AI Research AI

How Quantum Is the Advantage? A Fair, Calibration- and Noise-Aware Benchmark and Attribution Audit of Quantum Machine Learning for Network Intrusion Detection

arXiv:2608.18155v1 Announce Type: cross Abstract: Quantum machine learning (QML) for network intrusion detection (NIDS) is routinely reported to reach near-perfect accuracy, yet the most rigorous studies find that…

Source: arXiv cs.LG Syeda Anshrah Gillani, Mirza Samad Ahmed Baig, Shahid Munir Shah, Asher Ali, Hamzah Siddiqui
AI Research AI

Inference and Uncertainty Quantification for Streaming $r$-PCA

arXiv:2608.18374v1 Announce Type: cross Abstract: We address two open questions in streaming PCA via Oja's algorithm: sharp operator-norm convergence for general rank under sub-Gaussian data, and distributional…

Source: arXiv cs.LG Haoshu Xu, Hongzhe Li
AI Research AI

Learning Canonical Register Automata over Ordered Data Domains

arXiv:2608.18765v1 Announce Type: cross Abstract: Register automata are finite automata equipped with memory that recognize data languages over infinite alphabets. In this work, we investigate active learning algorithms…

Source: arXiv cs.LG Yong Li, Qiyi Tang, Di-De Yen
AI Research AI

AutoOR: Scalably Post-training LLMs to Autoformalize Operations Research Problems

arXiv:2604.16804v3 Announce Type: replace Abstract: Optimization problems are central to decision-making in manufacturing, logistics, scheduling, and other industrial settings. Translating complicated descriptions of…

Source: arXiv cs.LG Sumeet Ramesh Motwani, Chuan Du, Aleksander Petrov, Christopher Davis, Philip Torr, Antonio Papania-Davis, Weishi Yan
AI Research AI

BrainBench: Benchmarking Large Language Models for Comprehensive EEG Understanding

arXiv:2608.04156v2 Announce Type: replace-cross Abstract: Electroencephalography (EEG) analysis extends beyond assigning predefined labels to recordings; it requires workflows connecting natural-language instructions,…

Source: arXiv cs.LG Yangxuan Zhou, Yuning Chen, Chen Wu, Jiquan Wang, Shijian Li, Gang Pan, Sha Zhao
AI Research AI

Accurate Decoding of Natural Sentences from Non-Invasive Brain Recordings

arXiv:2608.18114v1 Announce Type: cross Abstract: Restoring communication for people who have lost the ability to speak or move after a brain injury is a major challenge. While intracranial implants now enable…

Source: arXiv cs.LG Mingfang Zhang, Jarod L\'evy, Cedric Rommel, J\'er\'emy Rapin, Corentin Bel, Julie Bonnaire, Daniel Nieto, Pierre Bourdillon, Svetlana Pinet, St\'ephane d'Ascoli, Thomas Moreau, Jean-R\'emi King
AI Research AI

MITRE-SAGE: A Multi-Agent Cybersecurity Question-Answering Model

arXiv:2608.16921v2 Announce Type: replace-cross Abstract: Effective cybersecurity operations require timely and accurate analysis of large-scale heterogeneous security information; however, analysts increasingly…

Source: arXiv cs.LG Ali Habibzadeh, Farid Feyzi, Reza Ebrahimi Atani
AI Research AI

WorldPack: Dynamic Frame Compression for Long-context Video World Modeling

arXiv:2512.02473v3 Announce Type: replace-cross Abstract: Video world models have attracted significant attention for their ability to produce high-fidelity future visual observations conditioned on past observations…

Source: arXiv cs.LG Yuta Oshima, Yusuke Iwasawa, Masahiro Suzuki, Yutaka Matsuo, Hiroki Furuta
AI Research AI

Think Shallow, Solve Deep: Controlling Recurrent Dynamics for Reliable Test-Time Depth

arXiv:2608.18222v1 Announce Type: new Abstract: Recurrent-depth reasoners aim to solve harder problems by iterating their update longer at test time, but additional iterations can improve, preserve, or degrade an…

Source: arXiv cs.LG Ivan Viakhirev, Kirill Borodin, Amirah Almutairi, Serguei Barannikov, Maxim Abramov, Grach Mkrtchian
AI Research AI

Gated Graph Attention Networks with Learnable Temperature

arXiv:2605.29803v2 Announce Type: replace Abstract: Graph attention networks learn neighbor importance through data-dependent coefficients, but standard layers lack explicit control over unreliable feature dimensions…

Source: arXiv cs.LG Zhongtian Ma, Hao Wu, Yexin Zhang, Qiaosheng Zhang, Zhen Wang
AI Research AI

What is Missing from AI Post-Training AI: An Empirical Analysis

arXiv:2608.19072v1 Announce Type: cross Abstract: Large language model (LLM) agents can now post-train an LLM end-to-end. They can write code, launch training, evaluate checkpoints, and improve downstream performance,…

Source: arXiv cs.LG Joy Jia Yin Lim, Xin Huang, Hao Peng, Yaxi Lu, Xin Cong, Zhong Zhang, Maosong Sun, Yankai Lin
AI Research AI

Model Card for OpenAI Privacy Filter

arXiv:2608.18274v1 Announce Type: cross Abstract: OpenAI Privacy Filter is a compact, bidirectional token-classification model for detecting and redacting personally identifiable information (PII) and secrets in…

Source: arXiv cs.LG Charles de Bourcy, Sahra Ghalebikesabi, Avi Schwarzschild, Alex Gorbachev, Mihai Maruseac, Annie Chu, Vol Kyrylov, Tong Mu, Ally Bennett, Andy Nguyen, Casey Meehan, Jessica Gan Lee, Shane Bauer, Haro…
AI Research AI

Evaluating and Explaining Prompt Sensitivity of LLMs Using Interactions

arXiv:2608.18539v1 Announce Type: new Abstract: The remarkable capabilities of large language models (LLMs) are often undermined by their instability. Even subtle and semantically irrelevant changes in prompts can cause…

Source: arXiv cs.LG Ruiyang Qin, Qingzhuo Wang, Tian Wang, Zhihua Wei, Wen Shen
AI Research AI

Co-RL: Unsupervised Reasoning Emerges from Diverse Cohort in Multi-agent RL

arXiv:2608.17253v2 Announce Type: replace Abstract: Reinforcement learning (RL) has emerged as a powerful approach for improving reasoning in language and vision-language models, yet its strongest successes still depend…

Source: arXiv cs.LG Yunhao Yang, Yuexin Bian, Yunjie Tian, Di Fu, Tianjin Huang, Yuanyuan Shi, Ziang Xiao, Nuno Vasconcelos, Yijiang Li
AI Research AI

Breaking the weakest link to evade vision language models

arXiv:2608.18938v1 Announce Type: cross Abstract: Vision Language Models (VLMs) have recently emerged as a critical component of multimodal AI systems, enabling joint reasoning over visual and textual inputs in…

Source: arXiv cs.LG Ilan Zini, Boussad Addad, Katarzyna Kapusta
AI Research AI

Learning Random Geometric Graphs Drawn in Probabilistic Metric Spaces

arXiv:2608.19082v1 Announce Type: cross Abstract: We present a new data-driven learning of a Random Geometric Graph (RGG) of a multivariate dataset, where the graph is drawn in a probabilistic metric space. This graph…

Source: arXiv cs.LG Dalia Chakrabarty, Kangrui Wang, Chuqiao Zhang, Ye Liu
AI Research AI

Multi-stage neural operator learning with application for convolutions

arXiv:2608.18851v1 Announce Type: new Abstract: Convolution integrals widely exist in applications, and to enable fast and accurate computations, this paper introduces two general multi-stage neural operator learning…

Source: arXiv cs.LG Zhiping Mao, Zhenye Wen, Yong Zhang, Xiaofei Zhao
AI Research AI

Harness Continual Learning: Continual Adaptation Beyond Model Parameters

arXiv:2608.19013v1 Announce Type: new Abstract: Continual learning has largely been model-centric, treating model parameters as the state that changes with sequential experience. Modern agents can also adapt through a…

Source: arXiv cs.LG Borui Kang, Jinrui Gu, Junhan Lv, Wenbin Li, Lei Wang, Yang Gao
AI Research AI

Hierarchical Classification via Cascading Feature Elimination: Application to Human Phenotype Ontology-Aligned Facial Phenotyping (FaceMesh2HPO)

arXiv:2607.05585v2 Announce Type: replace-cross Abstract: FaceMesh2HPO is a framework for classifying facial phenotypic descriptors aligned with the Human Phenotype Ontology (HPO) to support clinical diagnosis. Using…

Source: arXiv cs.LG Fabio Hellmann, Alexander Hustinx, Benjamin D. Solomon, GestaltMatcher Database Consortium, Tzung-Chien Hsieh, Peter Krawitz, Elisabeth Andr\'e
AI Research AI

Converting Expert Deliberation into Financial Signals Through A Context-Aware NLP Pipeline

arXiv:2608.18911v1 Announce Type: new Abstract: We introduce the CDSP (context-conditional deliberation signal pipeline), converting an investment committee's meeting transcripts into structured predictive features.…

Source: arXiv cs.LG Vivek Batra, Kristin Chen, Sanjiv Das, Samuel Judge, Harshad Khadilkar, Sukrit Mittal, Amir Nasrollahzadeh, Daniel Ostrov, Jacob Sisk
AI Research AI

Untrainable elements determine what physical learning remembers

arXiv:2608.00097v2 Announce Type: replace-cross Abstract: Physical learning rules such as equilibrium propagation (EP), coupled learning (CL), and adjoint coupled learning (AL) train resistive networks through local…

Source: arXiv cs.LG Bijaya Dangol
AI Research AI

FlashAttention for Scalable Vector Architectures

arXiv:2608.18656v1 Announce Type: new Abstract: Inference with transformer models on CPUs is increasingly important, especially for Small Language Models (SLMs), where vector architectures are emerging as a promising…

Source: arXiv cs.LG Sonia Rani Gupta, Nikela Papadopoulou, Miquel Peric\`as
AI Research AI

Bernstein-Vazirani Networks: Quantum Machine Learning by Interference

arXiv:2608.19043v1 Announce Type: cross Abstract: We introduce Bernstein-Vazirani Networks (BVNs), a non-variational quantum machine learning framework that leverages quantum interference for supervised learning,…

Source: arXiv cs.LG Natacha Kuete Meli, Tolga Birdal, Prayag Tiwari, Vladislav Golyanik, Michael Moeller
AI Research AI

FiLoRA: Focus-and-Ignore LoRA for Controllable Feature Reliance

arXiv:2602.02060v2 Announce Type: replace Abstract: Multimodal foundation models integrate heterogeneous signals across modalities, yet it remains unclear whether their predictions can be controlled by explicitly…

Source: arXiv cs.LG Hyunsuk Chung, Soyeon Caren Han, Seungyeon Ji, Jinwoo Kim, Eun-Jung Holden, Kyungreem Han
AI Research AI

Assessing Quality of Experience in Natural Language Generation of German Text

arXiv:2608.18888v1 Announce Type: new Abstract: The rapid advancement of Natural Language Generation (NLG) has made the reliable evaluation of generated text increasingly critical, as these systems, such as large…

Source: arXiv cs.CL Dinh Nam Pham, Shushen Manakhimova, Vivien Macketanz, Sebastian M\"oller
AI Research AI

ChainWorld: Composing Long-Horizon Desktop Workloads from Atomic OSWorld Tasks

arXiv:2606.21654v2 Announce Type: replace-cross Abstract: Computer use agents are evaluated almost exclusively on atomic desktop tasks, but realistic desktop work requires sustaining state across multiple objectives. We…

Source: arXiv cs.CL Vincent Siu, Manasi Sharma, Dawn Song, Daniel Yue Zhang, Chenguang Wang, Ying Liu
AI Research AI

AI Can Learn Scientific Taste

arXiv:2603.14473v3 Announce Type: replace Abstract: Scientific discovery depends on expert judgement and foresight, which we call scientific taste: the ability to judge and propose research ideas with the potential for…

Source: arXiv cs.CL Jingqi Tong, Mingzhe Li, Hangcheng Li, Yongzhuo Yang, Yurong Mou, Weijie Ma, Hongji Chen, Xiaoran Liu, Qinyuan Cheng, Ming Zhang, Qiguang Chen, Weifeng Ge, Qipeng Guo, Tianlei Ying, Tianxiang Sun, Yi…
AI Research AI

Cross-Model Memory Transfer via Target-Side Reader Adaptation

arXiv:2608.17050v2 Announce Type: replace Abstract: Methods for improving knowledge use in large language models typically fall into two regimes. Non-parametric retrieval offers flexible access to external knowledge,…

Source: arXiv cs.CL Mingyuan Li, Guangsheng Yu, Xu Wang, Shaoxiong Ji
AI Research AI

Artifact-centered Claim-aware Observability for Autonomous Scientific Agents

arXiv:2608.18312v1 Announce Type: new Abstract: Autonomous scientific agents now increasingly propose ideas, write code, run experiments, analyze results, and even draft papers. Observe and audit those agents are…

Source: arXiv cs.CL Xiangyu Yin, Ming Du, Michael H. Prince, Mathew J. Cherukara
AI Research AI

Metrics That Write Themselves: Evolving an Evaluator from Its Own Blind Spots

arXiv:2608.18744v1 Announce Type: cross Abstract: Agents improve quickly against a reliable automatic metric and stall without one, and the applications that need them most, report generation among them, are the ones…

Source: arXiv cs.CL Xing Zhang, Yanwei Cui, Guanghui Wang, Zhihao Lin, Peiyang He
AI Research AI

Abliteration Mitigation via Refusal Aliases

arXiv:2608.18093v1 Announce Type: new Abstract: Abliteration, the removal of refusal capabilities from large language models by projecting weight matrices orthogonal to an extracted refusal direction, has emerged as a…

Source: arXiv cs.CL Nathan Truong
AI Research AI

Key Coverage Matters: Semi-Structured Extraction of OCR Clinical Reports

arXiv:2605.09440v2 Announce Type: replace Abstract: Clinical reports are often fragmented across healthcare institutions because privacy regulations and data silos limit direct information sharing. When patients seek…

Source: arXiv cs.CL Yu Wang, Yingyun Li, Ying Qin, Haiyang Qian
AI Research AI

OmniAlign: A Unified Multilingual Aligner for Word and Sentence Alignment

arXiv:2608.18474v1 Announce Type: new Abstract: Cross-lingual sequence alignment is fundamental for building and exploiting parallel corpora, spanning mappings from documents and sentences down to words and subwords.…

Source: arXiv cs.CL Mengpeng Yang, Jingxu Yang, Chao Chen, Tian Xia, Yabo Sun, Qiang Liu
AI Research AI

Multimodal Rapport Estimation in Real-World HRI

arXiv:2608.18401v1 Announce Type: cross Abstract: Evaluating interaction quality in real-world HRI is an important challenge. If interaction quality can be estimated reliably, the results can be used to improve dialogue…

Source: arXiv cs.CL Akihiro Sakuramoto, Takato Hayashi, Ryo Miyoshi, Yuki Okafuji, Shogo Okada
AI Research AI

Comment-level Topic Drift Analysis in the Reddit Corpus

arXiv:2608.19133v1 Announce Type: new Abstract: We present a novel application of embedding-based dynamic topic modeling techniques to detect and quantify topic drift at the comment level in a massive corpus. By…

Source: arXiv cs.CL Steven Morse, Daniel Runfola, Trenton W. Ford
AI Research AI

Persona-Guided LLM Agents for Task-Oriented Dialogue

arXiv:2608.18085v1 Announce Type: new Abstract: Prior work has shown that large language models (LLMs) can express diverse personality traits in open-ended text generation. However, it remains unclear whether they can…

Source: arXiv cs.CL Maryam Shoaeinaeini, Brent Harrison, A. B. Siddique
AI Research AI

Backdoor Learning in Language Models and Vision-Language Models

arXiv:2608.18095v1 Announce Type: new Abstract: Recent advances in deep learning have significantly enhanced the capabilities of Natural Language Processing (NLP) and Vision-Language Models (VLMs). However, these…

Source: arXiv cs.CL Weimin Lyu
AI Research AI

From Sequence to Structure: Relational Uncertainty Propagation for LLM Agents

arXiv:2608.16002v2 Announce Type: replace Abstract: Reliable uncertainty quantification (UQ) is essential for deploying large language model (LLM) agents in complex interactive environments. Existing UQ methods largely…

Source: arXiv cs.CL Zhengzhao Ma, Boxi Cao, Yaojie Lu, Hongyu Lin, Xianpei Han, Le Sun
AI Research AI

Self-Improvement of Large Language Models: A Technical Overview and Future Outlook

arXiv:2603.25681v2 Announce Type: replace Abstract: As large language models (LLMs) continue to advance, improving them solely through human supervision is becoming increasingly costly and limited in scalability. As…

Source: arXiv cs.CL Haoyan Yang, Mario Xerri, Solha Park, Huajian Zhang, Yiyang Feng, Sai Akhil Kogilathota, Jiawei Zhou
AI Research AI

ComponentBench: Diagnosing Component-Level Failures in Computer-Use Agents

arXiv:2608.18307v1 Announce Type: cross Abstract: Current evaluation of computer-use agents is split between long-horizon workflow benchmarks and atomic GUI-grounding tests. This leaves an under-instrumented middle…

Source: arXiv cs.CL Tianchen Guan, Xinlei Lin, Royce Cheng-Yue, Xiangjun Wang, Shuyan Zhou
AI Research AI

The Deontic Gap: Large Language Models and the Modal Language of Obligation

arXiv:2608.18144v1 Announce Type: new Abstract: Modal auxiliaries such as must, should, and have to mark necessity and obligation within the contexts of speaker authority and interpersonal stance. We examine whether…

Source: arXiv cs.CL Daniel Hart, Sarah Allred, Joseph Abbas, Morenike Alugo
AI Research AI

When to Call an Apple Red: Humans Follow Introspective Rules, VLMs Don't

arXiv:2604.06422v2 Announce Type: replace Abstract: Understanding when Vision-Language Models (VLMs) will behave unexpectedly, whether models can reliably predict their own behavior, and if models adhere to their…

Source: arXiv cs.CL Jonathan Nemitz, Carsten Eickhoff, Junyi Jessy Li, Kyle Mahowald, Michal Golovanevsky, William Rudman
AI Research AI

KA2L: A Knowledge-Aware Active Learning Framework for LLMs

arXiv:2603.17566v2 Announce Type: replace Abstract: Fine-tuning large language models (LLMs) with high-quality knowledge has been shown to enhance their performance effectively. However, there is a paucity of research…

Source: arXiv cs.CL Haoxuan Yin, Chen Tang, Yangfan Wang, Lian Yan, Jingchi Jiang
AI Research AI

Different Facets of Verbalised Overconfidence: an Interpretability Study

arXiv:2608.18106v1 Announce Type: new Abstract: Large language models tend to overconfidence, giving assertive answers when the evidence suggests hedging or abstention. Using controlled reasoning scenarios that…

Source: arXiv cs.CL Davide Mazzaccara, Leonardo Bertolazzi, Raffaella Bernardi
AI Research AI

GreekBarRetrieval: A Benchmark for Greek Statutory Retrieval

arXiv:2608.18752v1 Announce Type: cross Abstract: Statutory retrieval is necessary for citation-grounded legal question answering, but remains underexplored for Greek. We introduce GreekBarRetrieval, a public retrieval…

Source: arXiv cs.CL Ernest Beta, Odysseas S. Chlapanis, Dimitrios Galanis, Ion Androutsopoulos
AI Research AI

Introducing the Privacy-HSD Trade-off: Hate Speech Detection, but not at the Cost of Privacy

arXiv:2608.19006v1 Announce Type: new Abstract: Hate speech is a real and timely threat that affects a large portion of online users, especially youth and minority groups. While building reliable and robust automatic…

Source: arXiv cs.CL Stephen Meisenbacher, Vlad Garbuz, Chirill Donos, Maxim Dnestreanschii, Gabriel Creanga, Andreea-Elena Bodea, Thomas Lampert, Jana Diesner
AI Research AI

Institutional Newspapers Pipeline: Deriving billions of high quality tokens from historical newspapers

arXiv:2608.18972v1 Announce Type: new Abstract: Historical newspapers are an abundant record of public life, but their dense, irregular and sometimes noisy layouts make computational access to these materials both…

Source: arXiv cs.CL Matteo Cargnelutti, Catherine Brobston, Eben English, Jake Sadow, Kacie Bailey, Greg Leppert, Amanda Watson, Jessica Chapel, Jonathan Zittrain
AI Research AI

SMTrap: Cost-Effective DoS Attacks Against Large Reasoning Models via SMT Conflict Guidance

arXiv:2608.18921v1 Announce Type: new Abstract: Existing LRM-DoS methods rely heavily on model feedback to synthesize attack queries, requiring either repeated queries to the target model or training a dedicated attack…

Source: arXiv cs.CL Jian Yang, Zhenqi Feng, Zhaoyang Yu, Zhaoxin Fan, Kejian Wu, Xiaofeng Wang, Zheng Zhu, Jianjun Huang, Wei You, Bin Liang
AI Research AI

Neurosymbolic Embodied Agents

arXiv:2608.16794v2 Announce Type: replace-cross Abstract: Language and vision-language models generate plausible embodied plans but do not guarantee executability, as their outputs can violate environment dynamics or…

Source: arXiv cs.CL Mohammad Albinhassan, Yuming Feng, Alessandra Russo, Pranava Madhyastha
AI Research AI

Do Large Language Models Hallucinate Electric Fata Morganas?

arXiv:2608.18816v1 Announce Type: new Abstract: AI hallucinations - that is, outputs which are made up, cannot be verified, or contradict the source material - are generally regarded as an engineering flaw to be dealt…

Source: arXiv cs.CL Kristina \v{S}ekrst
AI Research AI

SuTRA : Structurally-Unified Tokenization with Root Awareness

arXiv:2608.18087v1 Announce Type: new Abstract: Existing subword tokenizers optimize statistical compression but ignore morphological structure, particularly the relationship between roots and affixes. This is harmful…

Source: arXiv cs.CL Vaibhav Rathore, Siddhant Gole, Dadhichi Telwadkar, Rooshil Bhatia, Maulik Ruparel, Siddharth Surekha, Neha Bhargava
AI Research AI

Institutional Books - Enriched Text: A customizable multilingual open-source pipeline for denoising, deduplicating, and annotating OCR text at scale

arXiv:2608.19026v1 Announce Type: new Abstract: Released in 2025, Institutional Books: Harvard Library (IB-HL) is a collection of 983,004 volumes (242B o200k_base tokens), originally digitized through Harvard Library's…

Source: arXiv cs.CL David Lowry-Duda, Matteo Cargnelutti, Catherine Brobston, Salwa Ismail, Greg Leppert, Amanda Watson, Jonathan Zittrain
AI Research AI

Corrections of Zipf's and Heaps' Laws Derived from Hapax Rate Models

arXiv:2307.12896v5 Announce Type: replace Abstract: The article introduces corrections to Zipf's and Heaps' laws based on systematic models of the proportion of hapaxes, i.e., words that occur once. The derivation rests…

Source: arXiv cs.CL {\L}ukasz D\k{e}bowski
AI Research AI

SPADE: Self-Play in Adaptive Synthetic Executable Environments

arXiv:2608.19197v1 Announce Type: new Abstract: Continuous self-improvement requires an ever-expanding pool of self-generated, diverse, adaptive goals. For language agents, existing training environment pools…

Source: arXiv cs.CL Bo Liu, Simon Yu, Yiding Jiang, Ao Qu, Andrew Zhao, Zichen Liu, Junsu Kim, Zijian Zhou, Seungone Kim, Tongzheng Ren, Mickel Liu, Hanfei Yu, Zhaorun Chen, Weiyan Shi, Paul Pu Liang, Luke Zettlemoyer,…
AI Research AI

Aslema at NADI 2026: Augmentation through Fewshot for SLU

arXiv:2608.18689v1 Announce Type: new Abstract: We present Aslema, our system for NADI 2026 Shared Task 5, which consists of two subtasks: intent recognition and slot filling. We evaluate four omni LLMs in a zero-shot…

Source: arXiv cs.CL Tajwaar Shafiq, Hunzalah Hassan Bhatti, Shammur Absar Chowdhury, Firoj Alam
AI Research AI

Language Models for Portuguese: A Systematic Mapping Study

arXiv:2608.18138v1 Announce Type: new Abstract: In recent years, the rapid development of language models has transformed the field of Natural Language Processing through a wide range of applications. However, the…

Source: arXiv cs.CL Jhessica Silva, Carlos Caetano, Helena Maia, Breno Bernard Nicolau de Fran\c{c}a, Sandra Avila, Helio Pedrini
AI Research AI

SLAI T-Rex: Full-Parameter Post-training of the DeepSeek-V4 Family on Ascend SuperPOD

arXiv:2607.20145v3 Announce Type: replace Abstract: Full-parameter post-training of trillion-parameter-scale MoE models introduces substantial system-level challenges for large-scale distributed training, including…

Source: arXiv cs.CL Dongfang Li, Xiaodong Luo, Ruoyu Sun, Xuhui Chen, Linyuan Qiu, Jian Meng, Zhengxuan Lu, Yiting Wang, Yucheng Xie, Tao Guo, Tianxiang Fang, Jing Li, Sihang Chen, Shihao Hong, Chang Liu, Weihua Dai, Zi…
AI Research AI

MedUAG: Unified Understanding and Generation for Medical Multimodal Models

arXiv:2608.18937v1 Announce Type: new Abstract: Recent Multimodal Large Language Models (MLLMs) are rapidly evolving into unified understanding and generation (UAG) frameworks. However, extending these unified paradigms…

Source: arXiv cs.CL Zijie Meng, Yuncheng Zhang, Hualiang Wang, Yitian Tang, Xiaotang Gai, Chen Shen, Songtao Jiang, Shaosheng Cao, Jian Wu, Xian Wu, Zuozhu Liu
AI Research AI

Beyond LLM-Based Reasoning: Lightweight GNNs for Agent Failure Attribution

arXiv:2608.18575v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems (MAS) often exhibit complex failure modes, which frequently cause agents to produce incorrect outcomes. This motivates…

Source: arXiv cs.CL Ting-Wei Li, Yuanchen Bei, Xiao Lin, Hanghang Tong
AI Research AI

Self- and Other-Labels Induce Bidirectional Bias in LLM Judges

arXiv:2608.18091v1 Announce Type: new Abstract: As LLM-as-a-judge systems become increasingly widespread, self-preference in LLMs -- the tendency to favor one's own outputs -- raises growing concerns about evaluation…

Source: arXiv cs.CL Songeun Chae, Min Kim, Donghoon Jung, Seojin Choi, Seohyon Jung
AI Research AI

MemFuse: Multi-Source Memory Fusion from Fragmented Observations

arXiv:2608.18704v1 Announce Type: new Abstract: Long-term memory is essential for agents that operate across extended interactions, yet existing memory systems and benchmarks predominantly focus on single-source textual…

Source: arXiv cs.CL Chao Li, Yuanfa Li, Wenhao Wu, Xule Liu, Zhi Wang, Kun Shao
AI Research AI

Safety Alignment Illusion: The Cross-Lingual Safety Gap in LLMs

arXiv:2608.18131v1 Announce Type: cross Abstract: Current safety alignment training for Large Language Models (LLMs) are heavily English-centric. When such safety filters fail for non-English languages, the consequences…

Source: arXiv cs.CL Namya Bhatnagar
AI Research AI

Global Index on Responsible AI 2026 : Conceptual Framework and Methodology

arXiv:2608.18122v1 Announce Type: cross Abstract: This report presents the methodology of the Global Index on Responsible AI (GIRAI), 2nd Edition. This edition refines the 1st Edition by strengthening the distinction…

Source: arXiv cs.AI Fola Adeleke, Rachel Adams, Ayantola Alayande, Daniela Benavente, Ana Florido, Nicol\'as Grossman, Leah Junck
AI Research AI

\textsc{TestifAI}: Tomography-Based Testing for Deep Learning Systems

arXiv:2608.18900v1 Announce Type: new Abstract: As AI systems are increasingly deployed in safety-critical application domains (e.g., autonomous driving), associated risks increase too. Deep learning models underlying…

Source: arXiv cs.AI Arooj Arif, Tobias Hartung, Elena Botoeva, Alexandros Koliousis
AI Research AI

CTIFoundry: An Agent-Native Corpus Scaffold for Cyber Threat Intelligence

arXiv:2608.18613v1 Announce Type: new Abstract: Cyber threat intelligence (CTI) is increasingly consumed not by human analysts but by LLM agents that compose multi-step investigations at query time. The harness side of…

Source: arXiv cs.AI Yutong Cheng, Changze Li, Qian Cui, Wei Ding, Lingzhi Wang, Yan Chen, Peng Gao
AI Research AI

Composed Historical Image Retrieval by Modeling Temporal Representations

arXiv:2608.18694v1 Announce Type: cross Abstract: While time evolves linearly, the geometry of neural embedding spaces is inherently multi-dimensional, often chaotic, and difficult to interpret. In principle, one could…

Source: arXiv cs.AI Adri\`a Molina Rodr\'iguez, Oriol Ramos Terrades, Josep Llad\'os Canet
AI Research AI

Wildfire Suppression: Complexity, Models, and Instances

arXiv:2603.29865v2 Announce Type: replace-cross Abstract: Wildfires cause major losses worldwide, and the frequency of fire-weather conditions is likely to increase in many regions. We study the allocation of…

Source: arXiv cs.AI Gustavo Delazeri, Marcus Ritt
AI Research AI

Learning-State-Aware Dynamic Generative Data Augmentation on Small-Scale Datasets

arXiv:2608.18907v1 Announce Type: cross Abstract: Small-scale image classification is often limited by the scarcity of training data. Generative data augmentation (GDA) based on pretrained generative models has emerged…

Source: arXiv cs.AI Ting Xiang, Chenxi Deng, Jinhui Zhao, Bingting Jiang, Ke Zhang, Changjian Chen, Zhuo Tang
AI Research AI

Detecting Backdoors in Object Detection via Pre-NMS Prediction Distribution Shift

arXiv:2608.19088v1 Announce Type: cross Abstract: Object detection models deployed in safety-critical applications remain vulnerable to backdoor attacks that cause targeted misbehaviors when a hidden trigger is present.…

Source: arXiv cs.AI Longtian Wang, Zhengyu Zhao, Chenhao Lin, Le Yang, Shiwei Wang, Yuhan Zhi, Xiaofei Xie, Chao Shen
AI Research AI

SkillGate: Training In-Policy Skill Selection in Long-Horizon Agents

arXiv:2608.18852v1 Announce Type: new Abstract: Agent frameworks increasingly package procedural knowledge as skills: instruction files an agent reads on demand, while public libraries now hold thousands of them. Which…

Source: arXiv cs.AI Qingyao Li, Wenxiang Jiao, Shuai Shao, Kangning Zhang, Yuan Lu, Yi Guo, Weiwen Liu, Weinan Zhang, Yong Yu
AI Research AI

MBABench: Evaluating LLM Agents on End-to-End Spreadsheet Tasks in Finance

arXiv:2605.22664v5 Announce Type: replace Abstract: LLM agents are increasingly expected to carry out end-to-end workflows, producing complete artifacts from high-level user instructions. To meet enterprise needs,…

Source: arXiv cs.AI Thomson Yen, Julian Poeltl, Harshith Srinivas Gear, Yilin Meng, Joshua Fan, Adam Shen, Yili Liu, Ali Bauyrzhan, Patrick Shea, Siri Du, Haoyang Liu, Daniel Guetta, Hongseok Namkoong
AI Research AI

Eureka: Task-Conditioned Meta-Agent Orchestration for Scientific Discovery

arXiv:2608.19047v1 Announce Type: new Abstract: We present Eureka, a task-conditioned Meta-Agent architecture that compiles long-horizon tasks into dynamic obligation graphs with explicit acceptance semantics. During…

Source: arXiv cs.AI Alizer Wong, Heng Cui, Yi Tan, Xiongchao Zhan, Liang Lin, Yuxiang Guo, Zhaorong Dai, Zixin Zeng, Wenyuan Li
AI Research AI

From Multi-Agent to Single-Agent: When Is Skill Distillation Beneficial?

arXiv:2604.01608v5 Announce Type: replace Abstract: Multi-agent systems (MAS) for structured data-science tasks externalize analytical control through workflows spanning stages, tools, shared state, verification, and…

Source: arXiv cs.AI Binyan Xu, Dong Fang, Haitao Li, Kehuan Zhang
AI Research AI

FM-Bench: A Benchmark for Long-Horizon Management with Competing Agents

arXiv:2608.18423v1 Announce Type: new Abstract: Language model agents now execute bounded tasks reliably. Whether they can sustain effective decision-making over long horizons, where actions have cumulative consequences…

Source: arXiv cs.AI Tianyou Wang, Chongyang Gao, Kezhen Chen, Chen Dong, Yinghao He, Donghan Li, Wangcheng Xu, Hongjiu Zhang, Chi Li
AI Research AI

ContextSniper: AntTrail's Token-Efficient Code Memory for Repository-Level Program Repair

arXiv:2607.01916v5 Announce Type: replace Abstract: Large language model agents can repair real repository issues, but they often spend large context budgets on whole-file reads, broad searches, and long terminal…

Source: arXiv cs.AI Chiwang Luk, Matin Mohammad Najafi, Zhifeng Jia, Wei Yang, Xiuchang Li, Jinwei Zhu, Yang Ren, Lei Chen, Gao Cong
AI Research AI

A Theory of Post-hoc Debate Judgement

arXiv:2608.19002v1 Announce Type: new Abstract: Debates have recently emerged as a useful methodology for agentic AI to improve performance as well as to aid explainability and user engagement. For example,…

Source: arXiv cs.AI Xiang Yin, Adam Dejl, Antonio Rago, Lihu Chen, Francesca Toni
AI Research AI

A Jagged Frontier: Evaluating Robustness of Code Agents to Semantics-Preserving Transformations

arXiv:2608.18389v1 Announce Type: new Abstract: AI code agents are increasingly deployed to resolve real software issues, yet their reliability under superficial code variations remains poorly understood. We evaluate…

Source: arXiv cs.AI Hasan Najib Mahmud (Colorado State University), Shreya Gupta (Microsoft), Isha Chaudhary (University of Illinois Urbana-Champaign), Nathaniel Enis (Colorado State University), Ravi Mangal (Colorado S…
AI Research AI

Interval POMDP Shielding for Imperfect-Perception Agents

arXiv:2604.20728v2 Announce Type: replace Abstract: Autonomous systems that rely on learned perception can make unsafe decisions when sensor readings are misclassified. We study shielding for this setting: given a…

Source: arXiv cs.AI William Scarbro, Ravi Mangal
AI Research AI

DentAgent: Evidence-Centric Multi-Agent Coordination for Multimodal Dental Reasoning

arXiv:2608.18878v1 Announce Type: new Abstract: Oral diseases affect billions of people worldwide, underscoring a pressing need for accurate and reliable dental assessment that integrates heterogeneous evidence from…

Source: arXiv cs.AI Zijie Meng, Xiwei Dai, Yixuan Tang, Jin Hao, Yang Feng, Fudong Zhu, Xiaoqiang Liu, Shaosheng Cao, Zuozhu Liu
AI Research AI

Position: Multi-Agent Systems Should Prioritize Concurrency Control

arXiv:2608.18092v1 Announce Type: new Abstract: LLM-based multi-agent systems (MAS) promise scalable collaboration, yet adding agents often reduces reliability. This position paper argues that many MAS failures are…

Source: arXiv cs.AI Xin Yang, Letian Li, Zimo Ji, Terry Jingchen Zhang, Wenyuan Jiang
AI Research AI

ORBITER: Conflict-Aware Decision-Making for Agentic Last-Mile Delivery

arXiv:2608.18846v1 Announce Type: new Abstract: Last-mile delivery aims to handle dynamically arriving orders with couriers while modeling complex spatial and temporal correlations. Recent learning-based methods model…

Source: arXiv cs.AI Mingzhao Li, Chenxi Liu, Yan Zhao, Hao Miao
AI Research AI

SkillForge: Self-Distilling Agents for Project-Specific Issue Resolution

arXiv:2608.18933v1 Announce Type: cross Abstract: Large language model (LLM) based agents have demonstrated remarkable proficiency in automated software issue resolution, yet they often struggle to resolve issues in a…

Source: arXiv cs.AI Silin Chen, Han Li, Xiaodong Gu, Yuling Shi, Haibing Guan
AI Research AI

A strengthening of the MCFL-ness of $O_2$

arXiv:2608.18813v1 Announce Type: cross Abstract: In the last years, a number of proofs of the fact that $O_2$ is a multiple context-free grammar (MCFG) were given. Such results can be exploited in the fields of both…

Source: arXiv cs.AI Marco B. Caminati
AI Research AI

Accuracy and Robustness of Model Cascades Under Data Perturbations

arXiv:2608.17711v2 Announce Type: replace Abstract: Prediction cascades significantly reduce energy consumption of Artificial Intelligence (AI) models while maintaining high predictive performance. The idea is that easy…

Source: arXiv cs.AI Pallavi Mitra, Jai Kushwaha, Felix Biessmann
AI Research AI

RDFdL: Integrating RDF with Differential Dynamic Logic

arXiv:2608.18165v1 Announce Type: new Abstract: Knowledge graphs modeled in RDF are powerful for describing static knowledge, but they cannot capture or reason about the dynamic behavior of physical systems, e.g.,…

Source: arXiv cs.AI Yuyang Li, Lukas Kubelka, Julia Butte, Tobias K\"afer
AI Research AI

GrabVG: Graph-Attentive Binding for Visual Grounding in UAV Imagery

arXiv:2608.18996v1 Announce Type: cross Abstract: Visual grounding in Unmanned Aerial Vehicle (UAV) imagery aims to localize a target object in complex bird's-eye-view scenes according to a natural language description.…

Source: arXiv cs.AI Chaowei Wang, Yan Di, Jingjun Sun, Baozhe Liu, Jiaxu Tian, Yuheng Li, Guangqian Guo, Shan Gao
AI Research AI

Position: Profiling Game Worlds by Transition Complexity

arXiv:2608.18079v1 Announce Type: new Abstract: Game world modeling (GWM) and reinforcement learning (RL) are often confounded because research papers rarely quantify how difficult the underlying transition prediction…

Source: arXiv cs.AI Lele Cao
AI Research AI

ICICLE: Expanding Retrieval with In-Context Documents

arXiv:2605.26902v3 Announce Type: replace-cross Abstract: Generative retrieval (GR) maps queries directly to document identifiers (docids) using parametric knowledge, However, this design makes corpus expansion costly:…

Source: arXiv cs.AI Yu-Chen Den, Yung-Yu Shih, Zhi Rui Tam, Kuan-Yu Chen, Pu-Jen Cheng, Yun-Nung Chen, Eugene Yang
AI Research AI

Interpretable AI predicts a 2026 summer dry anomaly in central China

arXiv:2608.19163v1 Announce Type: cross Abstract: Seasonal precipitation anomalies are largely regulated by atmospheric circulation, which dynamical models predict with greater reliability than precipitation itself.…

Source: arXiv cs.AI Anran Wang, Wen Shi, Yong Luo, Jianbin Huang, Lijuan Chen, Junhu Zhao, Weixin Jin, Huihui Yuan
AI Research AI

G-ReAct: Graph-Guided Deep Search via Structure-State Co-Evolution

arXiv:2608.01324v3 Announce Type: replace Abstract: Deep search has become a fundamental capability of large language models (LLMs) for solving open-domain complex tasks. However, existing approaches typically rely on…

Source: arXiv cs.AI Shaoxiong Yang, Mengyuan Zhang, Shaojun Lin, Chao Li, Wei Liu, Kun Shao, Jian Luan
AI Research AI

The Lifecycle of LLM-as-a-Judge for Large-Scale Recommendation Explanations

arXiv:2608.18300v1 Announce Type: new Abstract: LLM-as-a-Judge, which leverages a large language model to evaluate natural language generated by another AI application or model, has become a standard, scalable approach…

Source: arXiv cs.AI Emma Yanyang Kong, JJ Tan, Ishan Gupta, Lars Olds, Claire Campbell, David Fagnan, Veli Balin, Rohan Gosain, Louis Garcia, Minsu Jang
AI Research AI

Improving Rural Medication Safety with AI: A Scoping Review

arXiv:2608.18135v1 Announce Type: new Abstract: Introduction: Medication errors (MEs) represent a significant threat to global healthcare systems, contributing to patient harm. Introducing artificial intelligence (AI)…

Source: arXiv cs.AI Jeong-ah Kim, Muhammad Ashad Kabir, Daniel Terry, Maryam Rouhi
AI Research AI

FACET: Preserving Source Intent and Executable State in Terminal Task Synthesis

arXiv:2608.18580v1 Announce Type: new Abstract: Training terminal agents requires scalable executable supervision, yet synthesizing high-quality terminal tasks remains challenging. Each task couples an instruction, an…

Source: arXiv cs.AI Kou Shi, Zun Wang, Qisheng Su, Shiting Huang, Ziao Zhang, Zhen Fang, Qingnan Ren, Jin Liu, Yu Zeng, Yiming Zhao, Lin Chen, Zehui Chen, Feng Zhao
AI Research AI

TrojanGYM: A Detector-in-the-Loop LLM for Adaptive RTL Hardware Trojan Insertion

arXiv:2601.17178v3 Announce Type: replace-cross Abstract: Hardware Trojans (HTs) remain a critical threat because learning-based detectors often overfit to narrow trigger/payload patterns and small, stylized benchmarks.…

Source: arXiv cs.AI Saideep Sreekumar, Zeng Wang, Akashdeep Saha, Weihua Xiao, Minghao Shao, Muhammad Shafique, Ozgur Sinanoglu, Ramesh Karri, Johann Knechtel
AI Research AI

Position: Behavioral Systems Require Behavioral Tests

arXiv:2608.18081v1 Announce Type: new Abstract: Artificial agentic systems increasingly operate as behavioral systems by interacting with dynamic environments, pursuing goals, and adapting over time. Yet, current…

Source: arXiv cs.AI Manuel Cherep, Nikhil Singh, Pattie Maes
AI Research AI

VibeWorlding: Can Multimodal Agents Construct 3D Open Worlds End-to-End?

arXiv:2608.15265v2 Announce Type: replace Abstract: Constructing an interactive 3D open world from a user query is important. However, existing methods are primarily evaluated on idealized, simple queries, making it…

Source: arXiv cs.AI Yansong Ning, Jingwen Ye, Zhongkai Wu, Yang Sun, Yiqin Zhu, Xingyi Li, Weidong Zhang, Hao Liu
AI Research AI

ADEPT: Accelerating Dexterity via Pre-Training and Post-Training using Reinforcement Learning

arXiv:2608.19182v1 Announce Type: cross Abstract: We introduce Accelerating Dexterity via Pre-Training (ADEPT), a large-scale reinforcement learning (RL) framework for learning sim-to-real transferable dexterity across…

Source: arXiv cs.AI Jayjun Lee, Jessica Yin, Asif Rana, Nicholas Blauch, Sam Mady, Mohak Bhardwaj, Nima Fazeli, Nathan Ratliff, Karl Van Wyk, Ankur Handa
AI Research AI

GRIP: Grounded Reasoning via Information-Restricted Premises

arXiv:2608.16776v2 Announce Type: replace Abstract: High-capacity encoders in retrieval-augmented generation (RAG) can let the query dominate the latent state, leaving retrieved evidence functionally irrelevant. We call…

Source: arXiv cs.AI Lirui Teng
AI Research AI

Hybrid ANN-SNN Pipeline with Local Plasticity

arXiv:2606.20151v2 Announce Type: replace-cross Abstract: This work proposes a hybrid ANN-SNN pipeline that effectively leverages the rich embeddings of pretrained artificial neural networks (ANNs) to enable…

Source: arXiv cs.AI Denis Larionov, Khairutin Shtanchaev, Mikhail Kiselev, Mikhail Korovin, Ivan Tugoy
AI Research AI

Teaching agentic AI to learn expert reasoning for rare disease diagnosis

arXiv:2606.16149v4 Announce Type: replace Abstract: Rare disease diagnosis depends on expert reasoning that is scarce and difficult to transfer; off-the-shelf large language models (LLMs) rank the correct disease first…

Source: arXiv cs.AI Minh-Ha Nguyen, Erica Gray, Bryce A. Schuler, Kevin W. Byram, Chih-Ting Yang, Fan Ma, Hua Xu, Wu-Chen Su, Chao Yan, Wei-Qi Wei, Adam Wright, Lisa Bastarache, Josh F. Peterson, Lingyao Li, Siyuan Ma,…
AI Research AI

LEDGER: Claim-to-Evidence Trace Graphs for Auditing LLM Agents

arXiv:2608.18398v1 Announce Type: cross Abstract: Large language model (LLM) agents can now carry out long-horizon technical workflows involving complex tool use, code execution, file edits, and generated artifacts. As…

Source: arXiv cs.AI Daehong Kim, Haichao Miao, Shusen Liu
AI Research AI

The Epistemic Politics of AI Anthropomorphism

arXiv:2608.00961v3 Announce Type: replace-cross Abstract: AI anthropomorphism is typically treated as a problem of user misperception requiring institutional correction. Users who engage in sustained or relational…

Source: arXiv cs.AI Donna M. Bye, Levin Kuhlmann
AI Research AI

Verifiable abstention makes AI leak diagnosis accountable in water distribution networks

arXiv:2608.18836v1 Announce Type: new Abstract: Utilities lose a substantial share of treated water to leakage, yet rarely trust artificial-intelligence localizers to dispatch crews: guessing everywhere cannot justify…

Source: arXiv cs.AI Tianwei Mu, Yue Wang, Mingzhe Yuan, Manhong Huang, Wenhong Wang, Xuerui Yin, Qing Luo, Min Xiao, Hui Yang, Jun Li, Dan Xue
AI Research AI

Syntactic Simplification of OWL Class Expressions

arXiv:2608.18899v1 Announce Type: new Abstract: Class expression learning often produces complex OWL class expressions that are difficult to interpret and reason over. However, by following theoretically grounded…

Source: arXiv cs.AI Alkid Baci, N'Dah Jean Kouagou, Caglar Demir, Axel-Cyrille Ngonga Ngomo
AI Research AI

Planning-aligned Token Compression for Long-Context Autonomous Driving

arXiv:2606.07464v3 Announce Type: replace-cross Abstract: Monolithic vision-action models represent an emerging paradigm in autonomous driving. However, this architecture produces token sequences that quickly exceed…

Source: arXiv cs.AI Zhixuan Liang, Yuxiao Chen, Yurong You, Peter Karkus, Wenhao Ding, Boyi Li, Alexander Popov, Yan Wang, Maximilian Igl, Yiming Li, Danfei Xu, Nikolai Smolyanskiy, Boris Ivanovic, Ping Luo, Marco Pavone
AI Research AI

Counterfactual Contrastive Analysis

arXiv:2608.19032v1 Announce Type: cross Abstract: Visual Counterfactual Explanations (VCEs) aim to explain image classifiers by generating minimally edited and realistic versions of an input image that change the…

Source: arXiv cs.AI Yunlong He, Pietro Gori
AI Research AI

Sleeping Kelly

arXiv:2510.15911v4 Announce Type: replace-cross Abstract: The Sleeping Beauty problem is a problem of imperfect recall that has received considerable attention. One approach to resolving the Sleeping Beauty problem has…

Source: arXiv cs.AI Ben Abramowitz
AI Research AI

DA-WAM: Decision-Aligned Future Latents for Driving World Models

arXiv:2608.19085v1 Announce Type: cross Abstract: Anticipating how scenes evolve under ego actions is fundamental to safe autonomous driving, yet the full potential of world models for decision-making remains…

Source: arXiv cs.AI Ruiguo Zhong, Benshan Ma, Xiaolong Chen, Lang Zhang, Mingyue Feng, Yaonong Wang, Pei Liu, Jun Ma
AI Research AI

Fragility of Value under Imperfect Alignment

arXiv:2607.28881v3 Announce Type: replace Abstract: As more responsibility is placed upon AI systems, it becomes increasingly important to guarantee that these systems are aligned with humanity. A common fear in AI…

Source: arXiv cs.AI Winter Cross
AI Research AI

SeisEvo: Evolution of Seismic Data Reconstruction Algorithms by Agents

arXiv:2608.18272v1 Announce Type: cross Abstract: Classical seismic data reconstruction relies on manually designed structural priors and iterative operators, whose coupled design space is far larger than manual trial…

Source: arXiv cs.AI Yingjie Xu, Siwei Yu, Jianwei Ma
AI Research AI

Finetuning Strategies for Querying Sounds by Vocal Imitation

arXiv:2608.19174v1 Announce Type: cross Abstract: This technical report describes our winning submission to the AES AIMLA 2025 Challenge on querying sound effects by vocal imitation. We investigate two complementary…

Source: arXiv cs.AI Aditya Bhattacharjee, Christos Plachouras, Sungkyun Chang, Emmanouil Benetos
AI Research AI

Large Language Model for Verilog Code Generation: Literature Review and the Road Ahead

arXiv:2512.00020v3 Announce Type: replace-cross Abstract: Code generation has emerged as a critical research area at the intersection of Software Engineering (SE) and Artificial Intelligence (AI), attracting significant…

Source: arXiv cs.AI Guang Yang, Wei Zheng, Xiang Chen, Dong Liang, Peng Hu, Yukui Yang, Shaohang Peng, Zhenghan Li, Jiahui Feng, Xiao Wei, Kexin Sun, Deyuan Ma, Haotian Cheng, Yiheng Shen, Xing Hu, Terry Yue Zhuo, David…
AI Research AI

RULER: Representation-Level Verification of Machine Unlearning

arXiv:2605.27569v3 Announce Type: replace Abstract: Machine unlearning aims to remove the influence of specific training records from a deployed model without retraining from scratch. Current protocols verify this at…

Source: arXiv cs.AI Georgina Cosma, Axel Finke
AI Research AI

One-Stage Object Detectors in Autonomous Driving

arXiv:2608.19014v1 Announce Type: cross Abstract: Autonomous vehicles depend on fast and reliable perception systems to detect surrounding vehicles, pedestrians, cyclists, traffic signs, and other road objects in real…

Source: arXiv cs.AI Jonel Roman, Ryan Sirjue, Peter Nguyen, Daniel Krutky, Juan Jesus, Sudip Dhakal
AI Research AI

How AI Prompts Can Teach Us About the Structure of Human Behavior

arXiv:2608.18265v1 Announce Type: cross Abstract: We introduce a general, easy-to-implement AI-based method for studying the structure and complexity of human behavior. We assign a large language model a ``type vector''…

Source: arXiv cs.AI Matthew O. Jackson, Benjamin S. Manning, Yutong Xie, Walter Yuan, Qiaozhu Mei
AI Research AI

ReasFlow: Assisting Reasoning-Centric Scientific Discovery in Applied Mathematics via a Knowledge-Based Multi-Agent System

arXiv:2607.14178v3 Announce Type: replace Abstract: Recent advances in Large Language Models have fueled autonomous AI agents capable of tackling complex scientific tasks, yet existing automated research systems remain…

Source: arXiv cs.AI Yutong He, Daibo Li, Guohong Li, Jiahe Geng, Zhengyang Huang, Can Ren, Zekun Zhang, Yifan Liu, Shuchen Zhu, Hengrui Zhang, Boao Kong, Ming Sun, Shu Li, Chenyi Li, Jiang Hu, Kun Yuan, Zaiwen Wen, Ping…