TILens Daily Edition 2026-09-01 Filters: topic=ai-research, github=hidden Stats: 1257 articles, 7 sources, 1251 research papers Top topics: AI, AI Research, Python ## News - Everything to know about Fable 5.1, Anthropic's new Claude model Source: the-neuron-ai Topic: AI (+AI Research) URL: https://www.theneuron.ai/news/claude-fable-5-1-cuts-agent-costs-and-safety-friction/ - How to Build a Business Website With AI Source: the-neuron-ai Topic: AI (+AI Research) URL: https://www.theneuron.ai/guides/how-to-build-business-website-with-ai/ - Best AI Video Editing Software for Creators in 2026 Source: the-neuron-ai Topic: AI (+AI Research) URL: https://www.theneuron.ai/newsletter/best-ai-video-editing-software/ - The 8 Best AI App Builders for Non-Coders to Use in 2026 Source: the-neuron-ai Topic: AI (+AI Research) URL: https://www.theneuron.ai/guides/best-ai-app-builders/ - 😺 Runway Solaris treats software like video Source: the-neuron-ai Topic: AI (+AI Research) URL: https://www.theneuron.ai/newsletter/runway-solaris-treats-software-like-video/ - AI Weekly Issue #528: What are companies building with AI? An Applied AI Deep Dive Source: ai-weekly Topic: AI (+AI Research) URL: https://aiweekly.co/issues/applied-ai-deep-dive-what-are-companies-actually-building ## Research - Mapping global methane emissions from space with deep learning Source: google-research-blog Topic: AI Labs (+AI, AI Research) URL: https://research.google/blog/mapping-global-methane-emissions-from-space-with-deep-learning/ - Your JSON Is Valid but Your Data Is Wrong: Five Failure Modes LLM Structured Outputs Won't Catch Source: towards-data-science Topic: AI Research (+AI, Python) URL: https://towardsdatascience.com/your-json-is-valid-but-your-data-is-wrong-five-failure-modes-llm-structured-outputs-wont-catch/ - What We Miss About Missing Values Source: towards-data-science Topic: AI Research (+AI, Python) URL: https://towardsdatascience.com/what-we-miss-about-missing-values/ - Beyond Point Predictions: A Practical Introduction to Bayesian Neural Networks Source: towards-data-science Topic: AI Research (+AI, Python) URL: https://towardsdatascience.com/beyond-point-predictions-a-practical-introduction-to-bayesian-neural-networks/ - 5 AI Skills That Will Keep Data Scientists Relevant in 2027 Source: towards-data-science Topic: AI Research (+AI, Python) URL: https://towardsdatascience.com/5-ai-skills-that-will-keep-data-scientists-relevant-in-2027/ - Estimating Population-Risk Curves Along Nonconvex Gradient Flows from the Training Sample Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30261 - A-MADiff: Attention-Guided Multi-Agent DRL with Diffusion Policies for Memory-Aware Task Orchestration in Mobile AIGC Networks Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29255 - Spectral Analysis for Sparse Matrix Computation: Insights and Potential Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29362 - Multi-Marginal Flow Matching with Adversarially Learnt Interpolants Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2510.01159 - Dec-BFTRL: Squre-Root Regret for Decentralized Online Upper-Linearizable Optimization under Separation Access with Application to Continuous Submodular Maximization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30271 - Error Detection for PET/CT Radiology Reports: Domain-Specific vs Large Language Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30021 - Watch your steps: Dormant Adversarial Behaviors that Activate upon LLM Finetuning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2505.16567 - A Model with No Head and Many Thoughts Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31069 - Quantum-Grassmann-Plucker Token Mixing for Deep Learning-Based Post-Disaster Damage Assessment Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30633 - Language-Informed Flow Matching for Trend-Guided Structure-Based 3D Molecular Generation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31009 - A Lightweight Phenology-Aware YOLOv5 Framework for Tomato Growth Stage Detection in Resource-Constrained Bhutanese Greenhouse Environments Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30088 - Partially Linear Autoencoders for Manifold Learning and Dimensionality Reduction Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29867 - Explainable Machine Learning for Broadband Adoption Disparities: Tract-Level Prediction and SHAP-Based Factor Profiling Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29110 - When Do Larger Batches Help Scale LLM Reinforcement Learning? Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29296 - GRASP: Geometry-aware Residual Alignment for Scalable Pretraining Data Attribution Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.06892 - Trajectory-Initialized Neural Double Q-Routing for Large-Scale Overhead Hoist Transport Systems Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30512 - Multiclass Linear Perceptrons with Multiplicative Margins Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30028 - MoPLEx: Estimating Plackett-Luce Mixture Models for Multi-Objective Alignment Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.25200 - Deep Reinforcement Learning for Dynamic Origin-Destination Matrix Estimation in Microscopic Traffic Simulations Considering Credit Assignment Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2511.06229 - ERR+: Sequential Entropy Resolution for Efficient and Decisive LLM Reasoning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28771 - Unsupervised Latent Space Alignment with Hyperspherical Geodesic Matching Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28840 - Simulation-free and finite-time diffusion model Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.03117 - Amortizing intractable inference in large language models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2310.04363 - Three Steps at a Time: Learning Representations from Action Sequences in Contrastive RL Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30640 - Two Centuries of Sexism in British Parliament: A Computational Analysis of Women's Representation in the Hansard Corpus Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30485 - From Extraction to Governed Memory: Multi-Agent Knowledge Graph Construction with Domain-Expert Review Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28642 - Mitigating Over-Optimization in PRM-Guided Search in Mathematical Reasoning by Optimizing the Guide Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30051 - Fine-Tuning Low-Bit Models with Gradient in Quantized Code Space Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30908 - Do VLMs Share Safety Neurons Across Modalities? Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30750 - LLMODE: Aligning ODEs with LLMs via Gated Token Injection for Irregular Spatio-Temporal Forecasting Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29640 - Automated Testing of LLM-Based Post Hoc Explainers Using Model Checking as an Oracle Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30581 - DASC: Decay-Aware State Compression for Hybrid Linear-Attention Serving Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30386 - Continuity-Free Near-Minimax Leading-Order Regret for CVaR-UCBVI Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28960 - Hybrid Semantic Context-Enhanced Ensemble Learning for Wind Power Ramp-Event Forecasting and Uncertainty-Aware Evaluation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29024 - Drop the Act: Probe-Filtered RL for Faithful Chain-of-Thought Reasoning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.11467 - Bayesian Sparse Low-Rank Adaptation for Large Language Model Uncertainty Estimation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.02182 - Delta-AI: Local objectives for amortized inference in sparse graphical models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2310.02423 - When 3D Gaussian Splatting Recovers Real Surfaces Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30054 - Curvature Cryptanalysis of Smooth Transformer Feed-Forward Networks Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28843 - Adaptive Doubly Robust Off-Policy Evaluation for Ranking Policies under Diverse User Behavior Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29600 - Ceiling-Clipped Acceptance Histograms Indicate Stranded Speed-up in Block-Diffusion Speculative Decoding Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30427 - A Spectral Identifiability Threshold for Dissipative Rate Recovery from Truncated Liouvillian Spectra Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29302 - A Target-Centric Survey of Quantization-Aware Training Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29667 - REALM: Reliable Expertise-Aware Language Model Fine-Tuning from Noisy Annotations Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.17289 - Learning Where Outcomes Change:Credit-Addressable Reasoning for Multimodal Geometry Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30457 - On the Plasticity Collapse in Continual Machine Unlearning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29513 - Designing for the Next Click: Bandits for Real-Time Page Layout Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29850 - TriHead-GAN: A Generative Adversarial Network with Triple-Head Discriminator for Carbon Emission Time Series Generation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.07569 - Brain-Language-Action (BLA) Models: Language-Conditioned EEG for Robotics Control Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28967 - A Zero-shot Generalized Graph Anomaly Detection Framework via Node Reconstruction Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.12673 - Mechanism Shift During Post-training from Autoregressive to Masked Diffusion Language Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.14758 - Normalized Low-Rank Adaptation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31036 - ORDDAR: Observation-Driven Reasoning for Distortion-Resilient Decision, Action, and Cognitive Recovery Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28704 - Deploying DeepSeek 175B Locally on a Single Consumer-Grade RTX 4060 Laptop with 32GB RAM for 200k-Scale Protein-Ligand Virtual Screening Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30877 - Data-to-Energy Stochastic Dynamics Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2509.26364 - Data-Driven Design Optimization of Streaming-Potential-Mediated Electrokinetic Transport of Viscoelastic Fluids in Microchannels Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29939 - Leveraging Turn-taking Dynamics for Intent Recognition in Multi-party Conversations Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28926 - Group Resonance Network: Learnable Prototypes and Multi-Subject Resonance for EEG Emotion Recognition Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.11119 - TSExplorer: An interactive data annotation and exploration tool for time-series data Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30514 - Liquid Gated Attention Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30695 - Rates of Convergence in the Central Limit Theorem for Markov Chains, with an Application to TD Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2401.15719 - How do World Models and Policies Compose in LLM Agents? A Joint Spectral and Behavioral Account Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30067 - VisER: Visual Evidence and Reliance for Object Hallucination Detection in LVLMs Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30480 - Semantics at an Angle: When Cosine Similarity Works Until It Doesn't Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2504.16318 - Learning Generated Controls under Fractured Geometry: Projective Residualization and Variation-Allocation Frontiers Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.14636 - Denoising as Projection: Constrained Optimization with Gradient-Guided Diffusion Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29507 - Trust Under Siege: Label Spoofing Attacks against Machine Learning for Android Malware Detection Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2503.11841 - TSPFN: A Temporal Tabular Foundation Model for Physiological Time Series Classification Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31013 - BAITBENCH: Measuring Agent Reward Hacking with Optional Shortcuts Planted in ML Tasks Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30724 - Knowledge Distillation under Teacher Misspecification: An Order-Parameter Analysis of the Gap between Teacher Mimicry and Task Performance Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29472 - Separable Nonnegative Matrix Factorization Using Powered Ratio-of-Norms Regularization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28799 - Reverse N-Wise Output-Oriented Testing for AI/ML and Quantum Computing Systems Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.14275 - Personalized Group Relative Policy Optimization for Heterogenous Preference Alignment Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.10009 - SingProbe Technical Report Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30703 - Predicting the Unpredictable: LLM-powered Long-term Chaotic Time Series Forecasting under Short-term Observations Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29579 - RL-FAT: Reinforcement Learning for Fair Adversarial Training Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29247 - The Signal in the Noise: An Auditable Reliability Layer for Biomedical Text Classification Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28595 - Subtraction-Based Tumor Segmentation and Lesion-Centered pCR Prediction for the MAMA-MIA Challenge Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29162 - MERIT: Mitigating Exposure Bias in Generative XMC for User-Interest Propensity Modeling Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28931 - Perturbation Sensitivity of Maximum-Likelihood Pairwise Ranking in Computational Decision Systems Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.17805 - Polis: 3D Self-Supervision at City Scale Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29426 - Effective Graph and Rank-based Contextual Embeddings for Textual and Multimedia Data Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29001 - When the Martingale Never Stops Firing: Anytime-Valid Gating on Real Forecast Streams Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30502 - MUSE: A Run-Centric Platform for Multimodal Unified Safety Evaluation of Large Language Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.02482 - Uniform Statistical Convergence of Empirical Sinkhorn Potentials with Exponential and Polynomial Dependence on the Regularization Parameter Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29152 - The Price of Intelligence: A Quality-Adjusted Price Index for AI Services Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29843 - Bergson: An Open Source Library for Data Attribution Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.11660 - Large Discovery Models: Empirically-grounded Model-Based Open-Ended Search Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.15669 - Energy-Based Physics-Informed Form Finding for Clustered Tensegrity Structures Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.12888 - ADMM-Q: An Improved Hessian-based Weight Quantizer for Post-Training Quantization of Large Language Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.11222 - Are Single-Token Sparse Autoencoder Features Causally Necessary? Layer-Depth and SAE-Family Effects Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.20596 - Sharp Restricted Isometry Thresholds for Global Minima of Rank-Restricted Matrix LASSO Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29018 - TPR-Attention for Combinatorial Generalization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30124 - Benchmark Contamination: A Taxonomy Organized by Defeated Mitigation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29463 - One Policy Is Enough: Single-Agent Reinforcement Learning Outperforms Tree Search for Chemistry Tool Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30952 - Learning the Geometry of Admissible Hypotheses through Inductive Bias in Training Distributions Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31028 - Enhancing Bayesian Optimization and Active Learning Through Kernel Diversity Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.24721 - Distributed Semantic Segmentation With Improved Rate-Distortion Trade-Off Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28684 - Segmentation of Bovid Dentition Under Imperfect Annotations: A Comparative Study of Convolutional and Attention Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31052 - Evaluating Tiny Recursive Models Across Training for Code Generation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29376 - VIBE: Video Instruction-aligned Background music gEneration Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30125 - Beyond Churn: Predicting Financial Fragmentation in Retail Banking with Temporal Machine Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30364 - Better World Models Can Lead to Better Post-Training Performance Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2512.03400 - GraM-Diff: A Unified Graph-Mamba Diffusion Framework for EEG-Based Alzheimer's Disease Data Generation and Diagnosis Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29755 - One Capability or Many? Testing the Economic Validity of Frontier AI Evaluation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29420 - Towards Continual Test-Time Adaptation of Vision-Language Models in Open-Vocabulary Semantic Segmentation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29923 - ObjectSplat: Improving Mesh Fidelity and Interactivity for 3D Scenes via Object-Level Mesh Splatting Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30423 - BEACON: Behavioral and Semantic Enrichment of AlphaEarth Embeddings through Tri-Modal Contrastive Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29553 - Constrained Group Relative Policy Optimization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.05863 - Reciprocity Separates Gradient Flow from Rotation in Conservative Physical Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30778 - Uncertainty of Vision Medical Foundation Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30390 - Last Step Matters: Early Uncertainty Cannot Predict Failure in Long-Horizon Agents Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29685 - Which LLM for Which Work? Budgeted Model Allocation under Uncertain Evaluation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29560 - Privacy-Preserving Detection of Rare Disease-Associated Cell Subsets via Secure Multi-Party Computation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.20118 - Sparse Koopman Autoencoders Identify Local Dynamical Regimes in Multibasin Systems Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29057 - Behavioral Latency as Weak Event-Time Supervision for EEG Reaction-Time Decoding Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29428 - Does On-Policy Distillation Really Distill? From Noisy Teacher to Self-Improvement Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31046 - Generalist Graph Anomaly Detection via Prototype-Based Distillation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.26857 - One Adapter, Many Tasks: Task-Conditioned Feature Transformations for Continual Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31096 - Functional Degeneracy in Neural Networks: Measurement and Pruning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30741 - Synthetic Speech, Real Signal: Paralinguistic Preservation and Cross-Lingual Augmentation via Voice Cloning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.22304 - From Location Phrases to Geographic Entities: Task-Adapted Retrieval for People Search Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28965 - It's a matter of timescale: non-linear utility in successor features and multi-objective planning and learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.25723 - Semi-supervised CAPP Transformer Learning via Pseudo-labeling Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.01419 - Learning Simple Test-Time Environments for LLM Web Agents Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29305 - Scaling Automatic Research Agents via World Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.12564 - Revisiting the Provable-Auditable Privacy Gap of DP-SGD Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28934 - T3S: Improving Multi-Task Reinforcement Learning with Task-Specific Feature Selector and Scheduler Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30765 - Frequency-aware forecasting for short-term typhoon gust prediction Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.25604 - Conservative Hybrid Graph Networks for Process Systems with Learned Routing Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28896 - Uncertainty-Driven Replay Memory for Reinforcement Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29860 - Deciding When to Decide: Testing Operational Suboptimality Under Distributional Shift Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29465 - Benchmarking Peptide-Protein Affinity Prediction Across Peptide and Target Shifts Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30175 - Kathleen Remembers: Length-Invariant One-Shot Recall Without Attention Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30376 - The information geometry of product-reference discrete diffusion: Interaction growth complexity and optimal scheduling Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28949 - GRZO: Group-Relative Zeroth-Order Optimization for Large Language Model Fine-Tuning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.02857 - Learning Dynamics of Logits Debiasing for Long-Tailed Semi-Supervised Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30699 - FG$^2$-GDN: Enhancing Long-Context Gated Delta Networks with Doubly Fine-Grained Control Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.19021 - Constant Individual Regret in General Games Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31166 - Test-Time Scaling for Scientific Equation Discovery Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28660 - Explanations, Prompts, and Formalizations: Arguments for New Norms in LLM-Enabled Mathematical Research Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29401 - Hallucination Mitigation for Large Vision-Language Models via Implicit Feature Stabilization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29924 - Securing Time Integrity in Energy IoT Against Clock Drift and Y2K38 Failures Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.23147 - Sensitivity-Constrained Neural Operators for Data-Efficient Forward and Inverse Modeling of Partial Differential Equation Systems Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29888 - CatchBench: When Can an Agent Failure Be Caught? Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.22808 - CrystalGRPO: Target-Aligned and Coverage-Preserving Reinforcement Learning for Flow-Based Crystal Structure Prediction Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.06582 - Beat-Synchronous Tokenization for ECG Transformers Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30367 - Online Estimation of Dynamic Origin-Destination Matrices Using Reinforcement Learning with Link-Flow Propagation Guidance Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30317 - Locked at the Entrance, Open Inside: Where RLVR Narrows the Solution Space Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29188 - Asymmetric On-Policy Distillation: Bridging Exploitation and Imitation at the Token Level Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.06387 - An Efficient Sparse Fine-Tuning with Low Quantization Error via Neural Network Pruning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2502.11439 - Foundation Models Meet Agriculture: Challenges Beyond Pretraining Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30392 - Emergent Misalignment Is Not Magical Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29118 - ScalePRM: Training Process Reward Models by Scaling Verification Compute Without Ground Truth Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2512.03244 - CoMPASS: Collaborative Molecular Property Prediction via Adaptive Small-Large Model Synergy Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30674 - Learning to Evaluate Before Improving: Automatic Rubric Induction for Automatic Research Agents Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31076 - No Equivariant Architecture Covers All Equivariant Attention Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30417 - PathGuide: Dynamic Classifier-Free Guidance via On-Policy Transport Alignment Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29107 - APIFlow-Bench: Measuring Whether Agents Survive Long, Dependent API Workflows Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29128 - HoopMind: A Real-Time Neural Game-Tree System for Opponent-Aware Possession Planning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29563 - Mode Connectivity Beyond Classifiers: Evidence from Generative and Contrastive Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30366 - Beyond Uncertainty: Multi-Solver Disagreement Rewards for Self-Evolving Reasoning Curricula Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30035 - World Model Control by Trajectory Reachability Metrics Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.22164 - When Safety Speaks a Language: A Mechanistic Analysis of Safety-Language Identity Entanglement in LLMs Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29936 - LLM Post-Training as Brownfield Maintenance: An Industrial Perspective on Dataware Engineering Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31102 - ASTRA - Agentic System for Ticket Resolution and Analysis Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28790 - TACS: Trajectory-Aware Candidate Selection for LLM Jailbreak Suffix Optimization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29564 - ToxLens: A Reproducible Graph-Learning Framework for Leakage-Aware, Uncertainty-Calibrated Molecular Toxicity Prediction Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30472 - Forget or Fine-tune? A Comparative Study of Machine Unlearning Strategies for Noisy Label Correction Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30046 - An Open-Source, Event-Driven Pipeline for Cryptocurrency Market Data: Ingestion, Forecasting, and On-Chain Fraud Detection Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29973 - Discrete Compositional Generation via General Soft Operators and Robust Reinforcement Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2506.17007 - Understanding Deep Learning via Notions of Rank Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2408.02111 - Conformalized Large Language Models under Configuration Shift Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.01460 - PQMass: Probabilistic Assessment of the Quality of Generative Models using Probability Mass Estimation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2402.04355 - INTERVenE: Temporal-Abstraction-Interval Based Transformers for Short-Horizon Medical Event Prediction Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29901 - EEGDM: Learning EEG Representation with Latent Diffusion Model Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.20705 - Tensor Methods for Language Models: From Token Representation to Training, Adaptation, Inference, Compression, and Interpretability Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30505 - Tracing Generated Samples to Training-Data Clusters in Flow-Matching Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30081 - SERUM: State Extraction and Refinement for User Modeling Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.29181 - Expected flow networks in stochastic environments and two-player zero-sum games Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2310.02779 - Fine-Grained Multi Image Object Hallucination Benchmark Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30653 - Branch Scaling Manifests as Implicit Architectural Regularization for Improving Generalization in Overparameterized ResNets Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2403.04545 - EDGE: Engine for Deterministic Graph Evaluation through Conversation Simulation from Graph Structured DSL Configuration Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29971 - Beyond Parallel Blindness: Information Floors and Model Gaps in Block Drafting Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.27339 - Structural Hierarchy and Geometry in Molecular Representation Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29886 - A Simple Transformer Pipeline for Full-Key Side-Channel Attacks on Uncropped Datasets Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30105 - Deep graph kernel point processes over networks Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2306.11313 - The Illusion of Replacement: Rethinking Specialized Machine Learning Models in the Foundation Model Era Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28980 - A rigor-matched audit of periodic-step layer skipping for efficient llm inference: conflayers versus swift, with a supplemental analysis of trained routing alternatives Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28846 - Learning Human Health and Diseases from 24-hour Wrist Movement Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29494 - Propensity Straight-Through Gradients for Discrete Stochastic Systems Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.25631 - Mamba-Assisted Non-Markovian Closure for Reduced-Order Modeling Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.05371 - A Unified Framework for Fair and Personalized Decentralized Learning under Communication Constraints Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.26493 - PathBridger: Subgoal Bridges for Offline Goal-Conditioned Reinforcement Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29061 - Safety Screening for Voltage Control in Active Distribution Grids via Distributionally Robust Conformal Screening Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30889 - Generalization as a robust performance property of learning-enabled dynamical systems Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30431 - Uncertainty-Aware Multi-Task Learning for Joint Modulation Recognition and SINR Estimation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28865 - A foundation model with multi-variate parallel attention to generate neuronal activity Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2506.20354 - S3C-LLM: Skill-Code Guided Agentic Language Models for Spectrum-to-Structure Elucidation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30910 - Simulation-Based Evaluation of Energy-Constrained Quantum-Classical Competition Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2308.08025 - Collapsibility of Performance Metrics in Clinical Predictive AI Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30568 - Scalable Clinical Data Infrastructure and Comparative ML Evaluation for Hospitalisation Risk Prediction in Elderly Patients with Multiple Long-Term Conditions using CPRD Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29419 - Controlling Refusal Behavior of LLMs via Stiefel-Constrained Rotation Steering Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30986 - Minerals in the Wild: A Hyperspectral-XRF Dataset for Elemental Composition Estimation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30537 - Autoencoders in Function Space Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2408.01362 - Selective Disclosure of Hidden Directives in Reasoning Models: Behavioral Asymmetry and Steering Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29070 - Towards a Systems Foundation for Agentic Skills: Architecture, Lifecycle, and Security Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29596 - Classification Drives Geographic Bias in Street Scene Segmentation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2412.11061 - ICON Decomposition: Auditing Deep Neural Networks with Multivariate Variance-based Concept-level Explanations Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.26083 - From Uncertainty to Failure Attribution: Self-Diagnosing Models for Failure Attribution under Distribution Shift Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.07953 - MEGA: Message Passing Neural Networks for Multigraphs with EdGe Attributes Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2412.00241 - Error Certificates for KV-Cache Eviction via Randomized Design Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.21475 - MolGA: Molecular Graph Adaptation with Pre-trained 2D Graph Encoder Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2510.07289 - TopGQ: Fast GNN Post-Training Quantization Leveraging Topology Information Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30394 - A Human-in-the-Loop Autonomous Agent for Industry Time Series Forecasting Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30976 - Implementing neural network mixed-effects models in Template Model Builder (TMB) Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31133 - Preference Elicitation for Policy Optimization and Application to Aligning Heart Transplantation with Human Values Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28620 - Temperature-Adaptive Transformed Teacher Matching Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29099 - Unlearning on Spatio-Temporal Graphs through Subgraph Virtual Edge Reconstruction Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29369 - StructSynth: Dependency Graphs as Generation Plans for Low-Data Tabular Synthesis with Language Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.02601 - Learning What Matters: Supervising Global Context Pruning with Causal Evidence Sets Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.21692 - Jigsaw-CRL: Recovering Global Latent Causal Order from Fragmented Multi-Client Interventions Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28991 - Adaptive teachers for amortized samplers Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2410.01432 - Correlation flow governs learning at criticality Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.08350 - When the Strongest Teacher Is Not the Best Teacher: Student-Centric Answer Selection Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.26872 - PRIME: Mitigating Subgroup Optimization Competition in Shared CTR Top Networks with Plug-in Residual Input-Conditioned Mixture of Expert Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30449 - On the Resilience of Text-to-Video Diffusion Models to Hardware Faults Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29598 - Activation Steering Transfer to Agents: One Gain Ratio Does Not Identify Potency and Efficacy Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.09156 - LaRA: Layer-wise Representation Analysis for Detecting Data Contamination in RL Post-Training Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.29888 - Content Exploration Beyond the Feed: Creator Supply and the Shared Corpus Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29430 - Federated Personalization of Early-Exit Networks Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.10015 - SS-ESOAP: Self-Scaled Adaptive Preconditioning for Physics-Informed Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29448 - Driving on Memory Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31029 - Deriving Scaling Laws for OpenEuroLLM Models: Learning Rate, Batch Size and Loss Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28308 - mmIR: Frequency-Space Inverse Rendering for 3D Millimeter-Wave Radar ADC Synthesis Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28913 - A Gold-Standard Study of What Makes a Lightweight Game-Playing Agent Strong Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.06854 - Data Diversity, Not Frequency Invariance: A Controlled and Self-Audited Study of Compression-Robust Deepfake Detection Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28685 - Graph Representational Learning: When Does More Expressivity Hurt Generalization? Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2505.11298 - E-Commerce Bench: Evaluating LLM Agents on Long-Horizon Autonomous Business Operation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30730 - Transformer-Based Flow Shop Scheduling Using MILP-Generated Training Data Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29690 - Adversarial Online Classification with a Preview Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29503 - Context Staircase: Signature-Aligned Dynamics of Token Embeddings under Small Initialization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30315 - Context-Aware Interpretable Representations for Retrieval and Graph Convolutional Network Classification Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29004 - Sparse Competition during Training For the Emergence of Specialized Modules Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30978 - Joint Spatiotemporal Spectral Neural Operators for Learning PDEs on Irregular Domains Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29892 - A Deep Latent Variable Framework for Jointly Modeling Missingness, Measurement Error, and Heterogeneity Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30040 - Reward-Oracle MCTS for Formal Theorem Proving: Sample-Efficient Search and the Need for Kernel-Level Proof Auditing Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28639 - Momentum as Residual-Driven Multiplier Correction for Deep Learning Optimization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.12925 - Using Prosody to Predict Syntactic Structure Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30260 - Approximate Speculative Decoding Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.03447 - Stress-Testing Efficient Responsible-AI Evaluation: When Compute Savings Change Benchmark Conclusions Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31108 - Zero-shot Generalizable Graph Anomaly Detection with Mixture of Riemannian Experts Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.06859 - Reference-Grafting Matches Fine-Tuning at Eliciting Sandbagged Capabilities Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29458 - Where Induction Runs Out: Description-Length Difficulty and the Memorisation Gap in Integer-Sequence Benchmarks Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29411 - On the Instance Hardness as a Decision Criterion in TinyML Systems Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29913 - FAMPWQ: Fisher Information-based Adaptive Mixed Precision Weight Quantization for Effective LLM Inference Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.24945 - PruneShift: A Framework for Evaluating Decision Reliability in Structured Pruning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29765 - APPSolver: Adaptive Patch Partitioning for Point-Wise Ship Flow Prediction on Unstructured Meshes Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29355 - Fully Distributed GNE Algorithms for Multi-Robot Placement without Consensus on Multipliers Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29388 - When Does a Classifier Help an LLM? Classifier-Guided Prompting and Hybrid Classifier-LLM Models for Credit-Default Prediction Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30086 - Mixture-Greedy for Online Generative Model Selection: Is UCB Necessary in Diversity-Aware Multi-Armed Bandits? Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.21716 - Confounding Masquerading as Improvement: A Systematic Evaluation of Offline Reinforcement Learning for Stroke Antithrombotic Treatment in a 129,000-Patient Registry Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30442 - Kascade: A Practical Sparse Attention Method for Long-Context LLM Inference Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2512.16391 - Generalization and Memorization in Rectified Flow Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.13421 - An Identifiability Theory of Masked Prediction: Mode Blindness and Mask Schedules Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.01383 - MOONSHOT : A Framework for Multi-Objective Pruning of Vision and Large Language Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.13287 - Moving the Mean Toward the Known Good, Not Beyond It: What Inference-Time Interventions and Weight Consolidation Buy in Open-Ended Generation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28886 - Validating FKG.in: Soundness Assessment in LLM-Augmented Indian Food Knowledge Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29249 - Target-Aware State-Adaptive $p$-Dirichlet Graph Neural Regression for Non-Invasive Body-Composition Estimation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29496 - Identification of Bivariate Causal Directionality Based on Anticipated Asymmetric Geometries Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.26024 - Task-to-Model Optimization for Enterprise LLM Coding Assistants: A Data-Driven Framework for Cost-Optimal Routing Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.08528 - Quantitative Target Convergence and Uniform-in-Time Propagation of Chaos for Langevin-Regularized SVGD Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28827 - Does Latent Planning Survive Point Clouds? Action-Conditioned JEPA World Models for Geometric Observations Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29434 - $\mathcal{N}_0$-Foundation: Towards the Age of Tactile Intelligence Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29601 - Informative Label Missingness in Multiclass Classification Information Geometry and Excess Risk Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30561 - Cross-lingual Functional Vectors for Emotion Detection in Large Language Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29613 - Flow-JEPA: Flow Matching for Robust Latent Dynamics in JEPA World Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29029 - AdaVLA: Adaptive Step Flow Matching for Training-free Acceleration of Vision-Language-Action Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29208 - The Safety Relay in Roleplay Jailbreaks: A Component-Resolved Causal Analysis of Harm Recognition and Refusal Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30585 - SpecPV: Improving Self-Speculative Decoding for Long-Context Generation via Partial Verification Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2512.02337 - Spatial Entropy based Partitioning for Spatiotemporal Graph Unlearning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29360 - Rotational Equivariance in Machine Learning: A Comprehensive Tutorial Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31045 - Forward-Deployed Full-Stack Engineering for Autonomous Cloud MLOps Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29615 - Efficient GPU Retrieval for Semantic Search Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28968 - Learning Representations through Token Prediction: Geometry, Approximation, and Downstream Guarantees Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30072 - QCell: Recombining and Aligning Cell Queries for Overlapping Instance Segmentation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29253 - PLC-DPO: Posterior Label Correction in Noisy and Ambiguous Preference Optimization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30597 - Strengthening Recursive Constructions for Zero-Error Shannon Capacity Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30273 - Measuring Memory and Generalization as Separable Geometric Channels: The Topo^2 Framework Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30487 - Universal Transformers for Circuit Computations: Perfect Length Generalization in Tiny Transformers Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31067 - A Borel Concept Class of VC Dimension One with a Non-PAC Consistent Learner in ZFC Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30246 - Beyond Token-Level Guidance: Inference-Time Alignment of Specialized LLMs via Cross-Family Representation Steering Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30319 - Towards an Expressivity-Normalized Energy-Demand Comparison of ANNs and SNNs Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29869 - Influence-Directed Distillation: Solving the Diversity Bottleneck in Sampled-Token On-Policy Distillation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29846 - Temporal Analysis of NetFlow Datasets for Network Intrusion Detection Systems Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2503.04404 - Every Layer Counts: An Exponential $L_2$ Depth Hierarchy for ReLU Networks Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.23877 - Machine Learning-Enhanced Tabu Search for Tactical Wireless Network Design Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28627 - DeMMO: Longitudinal and Cross-Disease Modelling of Digital Mobility Outcomes via Multi-Task Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.25073 - Locally-Guided Actor-Critic: Training a Goal-conditioned Actor with a Subgoal-aware Critic Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30406 - Personalized Treatment Outcome Prediction from Scarce Data via Dual-Channel Knowledge Distillation and Adaptive Fusion Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2510.26444 - Fine-Tuning Qwen3-27B for C-to-Rust Code Translation: A Three-Stage Curriculum of Pretraining, Debugging-Aware SFT, and Task-Specific SFT Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.13681 - TopoCompress: Long Context Compression via Graph-Wired Semantic Trajectories Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30811 - Generative Translation Priors: Bayesian Imaging with Cross-Modality Image Translation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28872 - PRACTICE: From Experience to Expertise in Self-Evolving Embodied Agents Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30760 - Seq2Synth: Benchmarking Temporal Fidelity in Synthetic Sequential Tabular Data Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.15606 - Perforated Backpropagation: A Neuroscience Inspired Extension to Artificial Neural Networks Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2501.18018 - Sense Once, Serve Many: Common-Trace Factorized Constrained PPO for Online Sensing-Session Consolidation in Multi-Tenant ISAC Networks Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29256 - Kronecker Factorization Improves Efficiency and Interpretability of Sparse Autoencoders Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2505.22255 - CARE: Context-Aware Ranking Evolution with Executable Scoring Programs for Budgeted Reaction Optimization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.14581 - Least but not Last: Fine-tuning Intermediate Principal Components for Better Performance-Forgetting Trade-Offs Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.03493 - Support Selection Beyond Smooth DAG Exactness: Completion Geometry,Score Margins, and Selective Certificates Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.08103 - TrainSDC: Characterizing and Mitigating Silent Data Corruption in Large Language Model Training Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30769 - MEL: Coordinate-Preserving EEG Tokenization for fMRI Translation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29304 - Audit Me If You Can: Query-Efficient Active Fairness Auditing of Black-Box LLMs Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.03087 - GFlowNets and variational inference Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2210.00580 - When Do Task Vectors Interfere? Mapping the Validity Boundaries of Weight-Space Composition Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.09490 - QueryGraph: Reliable Multi-Tool Query Execution Planning via LLM-Based Graph Generation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.08300 - Adaptive Multi-Branching for Shallow Decision Tree Induction Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29262 - Reading the News: Adapting Large Language Models to Swedish Journalism Through Continued Pre-Training Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30609 - Information-Based Calibration of Uncertainty Quantification in Product-of-Experts Gaussian Process Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29349 - When Design Rules Break: Benchmark Composition Determines Whether Label Informativeness Predicts GNN Aggregator Choice Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.10249 - EvoLen: Evolution-Guided Tokenization for DNA Language Model Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.08698 - Asynchronous Cooperative Online Learning for Multi-Robot Control under Computational Delays Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29562 - Development of an Autonomous AI Coding Agent using Monte Carlo Tree Search (MCTS) and Gemini LLM Frameworks Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29096 - From the Loss Landscape to Diverse Feature Learning in Neural Networks Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28948 - Alert: Learning Trigger Functions for Early Classification of Time Series using Deep-RL Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2502.06584 - Coarse composition suffices: tabular in-context learning for multi-activity antimicrobial peptide profiling Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30337 - A Universal Context-Reuse Layer for Cross-Model KV Sharing Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30963 - Riemannian Optimization for Hadamard Products of Low-Rank Matrices Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.01216 - Beamforming Design Via GNN in mmWave Cell-Free Massive MIMO Using Sub-6 GHz CSI Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30524 - Model Selection and Parameter Estimation of One-Dimensional Gaussian Mixture Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2404.12613 - AutoREC: A reinforcement learning platform for equivalent circuit model generation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.27266 - Universal Redundancies in Time Series Foundation Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.01605 - Signed random Fourier features for fast density estimation with indefinite kernels Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29265 - Singular Curvature in ReLU Training:Differentiation and the Gradient-Flow Limit Need Not Commute Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30960 - Reward-guided Fine-Tuning of One-Step Generative Models via Wasserstein Gradient Flow Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29647 - A Causal Model for Locating and Unlocking Sandbagging in Model Organisms Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29461 - Agnostics: Learning to Code in Any Programming Language via Reinforcement with a Universal Learning Environment Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.04865 - Balancing Privacy, Utility, and Safety in LLM Alignment through Preference Optimization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30141 - On the Complexity of the Compatibility Problem for Succinctly Encoded Conditional Distributions Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31120 - FiLM-GPNet: Geometry-Aware Pseudo-Supervised Phase Restoration with Zero-Shot Generalization for Large Temporal InSAR Stacks Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29384 - Reproducible macroscopic dynamics in a closed-loop human-AI learning system Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30946 - Tracing distinguishability through transformer processing with stochastic LayerNorm Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30720 - Spectral-Embedded Operator Learning for Three-Phase Interfacial Flow: A Ternary Cahn-Hilliard-Navier-Stokes Benchmark Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29069 - Off-Policy Evaluation for Semantic ID Recommenders: Does the Model's Own Code Hierarchy Help? Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28905 - Converse and Collision-Based Achievability for Node Localization with Hybrid Distance-Spectral Graph Positional Encodings Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30152 - Federated Learning for MRI-based BrainAGE: a multicenter study on post-stroke functional outcome prediction Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2506.15626 - ButterMamba: Butterworth-Enhanced Spatial-Temporal Mamba for Efficient Traffic Flow Prediction Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29658 - Tail-Replay: Escaping the Curse of Linear Attention in Prefix Caching for Hybrid LLMs Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30310 - GMTS: Gradient Magnitude-based Token Selection Improves RLVR Training for LLM Reasoning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30632 - Decoupled Structure-Feature Alignment via Alternating Optimization for Graph Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.11577 - A2TTA: Anchored-and-Agile Test-Time Adaptation for Evolving Traffic Sensor Networks Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.25875 - Mol-JEPA: A multimodal Joint Embedding Predictive Architecture for Molecules Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.22642 - Reinforcement Learning on Benign Facts Amplifies Leakage of Memorized Private Data Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.21727 - PAC: Progress-Augmented Advantage Curriculum for Multi-Task Reinforcement Learning of LLMs Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30528 - End-to-End Neural Shrinkage of Indefinite Pairwise Correlation Matrices for Small-Cap-Inclusive Portfolios Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30446 - Compression-Aware Abstention: Teaching LLMs to Refuse When KV-Compression Masks Remove Answer Evidence Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29934 - A Conditional GAN for Tabular Data Generation with Probabilistic Sampling of Latent Subspaces Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.00472 - KV Admission: Learning What to Write for Efficient Long-Context LLM Inference Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2512.17452 - Optimally Selecting Representative Agents from a Metric Space Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29097 - Oculi: A Conversational Agentic Platform for Automated Credit Risk Analysis Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28944 - Sycophantic Agreement Transfers with Neutral Data via Contrastive Preference Optimization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31079 - Continual Test-Time Adaptation via Entropy Sensitivity-Guidance in Strict Online Setting Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29920 - NVE: A Separability and Coverage-Aware Internal Validation Metric for Biclustering Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29045 - SemKV: Semantic Mixed-Precision KV Cache Quantization Guided by the Quality Cliff for Long-Context LLM Inference Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28911 - Inverting Foundation Models of Brain Function with Simulation-Based Inference Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.23865 - Creation begins with understanding: LLMs as strategy designers for privacy-preserving tabular data synthesis Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29674 - PhyloGFN: Phylogenetic inference with generative flow networks Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2310.08774 - Truthful AI Advisors: A Pre-Specified Benchmark for Large Language Model Honesty Under Preference Misalignment Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.01456 - Robust Broad Learning System with Wave Loss for Classification under Data Uncertainty Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29983 - Concepts Whisper: Spectral Anti-Concentration and the Dual Geometry of Transformer Representations Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.01609 - Season-Aware Hybrid Convolutional-Transformer for Antarctic Sea Ice Concentration Forecasting Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30654 - A Cycle-Consistency Constrained Framework for Dynamic Solution Space Reduction in Noninjective Regression Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2507.04659 - CateKV: On Sequential Consistency for Long-Context LLM Inference Acceleration Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30295 - Geometric Attractor Monitoring: A Robust and Frugal Framework for Multi-modal Industrial Robotic Cycles Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30804 - HalluPrism: When Multimodal Uncertainty Should Diagnose, Not Decide Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29193 - Representation Learning with Quantum Signal Processing Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28828 - In-Cell Learning: Language Models That Update Their Own Weights in Sequence Without Changing the File They Ship Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.20873 - LIBERO-Para: A Diagnostic Benchmark and Metrics for Paraphrase Robustness in VLA Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.28301 - Terminal Symmetry as a Carrier of Asymmetric Process Knowledge: Statewise Refinement for Anytime Verified Construction Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.11318 - Grammar of the Wave: Towards Explainable Multivariate Time Series Event Detection via Neuro-Symbolic VLM Agents Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.11479 - TEMPO: Temporally-grounded Multi-task Post-training for Large Audio-Language Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29999 - Equivariant Sheaf Neural Networks: Learning Geometric Transport on Graphs Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28853 - EquiReg: Equivariance Regularized Diffusion for Inverse Problems Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2505.22973 - An Agentic Retrobiosynthesis Framework with Learned Frontier Selection Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30702 - Sharp Approximation Rates for Neural Networks with Affine Latent Parameterizations Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31157 - A Hybrid State-Space Approach for Census-Tract Population Estimation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30094 - RankShift: In-Database Detection and Explanation of Categorical Shifts Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28922 - Graph4BiLO: Graph Neural Network Approximation for Bilevel Mixed-Integer Linear Optimization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30103 - Lies We Can See: Joint Verbal and Non-Verbal Deception by VLM Agents in Embodied Social Interactions Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30428 - Gender, Race, and Intersectional Bias in Resume Screening via Language Model Retrieval Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2407.20371 - You Do Not Fully Utilize Transformer's Representation Capacity Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2502.09245 - SYNAPSE: Neuro-Symbolic Visual Thought-to-Text Decoding via Topological Semantic Denoising Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.27790 - MolLedger: An Additive Graph Neural Network with Chemically Grounded ADME Attributions Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30636 - Structure-Preserving Physics-Informed Neural Network for the Korteweg--de Vries (KdV) Equation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2511.00418 - Conjoint Audio-to-Spikes Encoding and Processing for Efficient Neuromorphic Speech Recognition Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30792 - NeuReasoner: Theory-grounded Mapping of Reasoning Elicitation Boundaries Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.29971 - LoGo: Token-Level Dynamic Local-Global Attention Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29539 - AdaptAV: Continuous Adaption of Vision Models for Autonomous Vehicles Using Cloud-based Oracle Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28673 - State of Health Estimation using Convolutional and Bidirectional LSTM Neural Networks tuned by Bayesian Optimization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30593 - A Unified Perspective on Conformal Prediction and Wasserstein Distributionally Robust Optimization for Uncertainty Quantification Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29789 - Who Should Teach? Confidence-Aware Dual-Teacher Learning for Few-Shot Node Classification on Text-Attributed Graphs Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.22127 - Training-free LLM Verification via Recycling Few-shot Examples Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2506.17251 - Diffusion-Based Refinement for Kilometer-Scale Probabilistic Precipitation Nowcasting Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30205 - BiG-SURE - Bipartite Graph for Semantic Uncertainty and Reliability Estimation of LLMs Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30646 - Q-Strata: Hierarchical Bit Allocation for Mixed-Precision Quantization of Mixture-of-Experts LLMs Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30564 - Hyper-Fold: Exploring the Expressive Limit of Sequence-Geometry Learning for Proteins via Hypergraph Modeling Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29207 - SUB-PLAY: Adversarial Policies against Partially Observed Multi-Agent Reinforcement Learning Systems Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2402.03741 - The Double-Edged Nature of the Rashomon Set for Trustworthy Machine Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2511.21799 - PairAlign: A Framework for Autoregressive Tokenization via Self-Alignment with Applications to Audio Tokenization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.06582 - Structure Aware Neural Architecture Search for Mixture of Experts Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29817 - H-FedSN: Personalized Sparse Networks for Efficient and Accurate Hierarchical Federated Learning for IoT Applications Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2412.06210 - Linguistic Distance Segregates Latent Representations in Automatic Speech Recognition Systems Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30853 - Clustering as Approximation by Constrained Projectors: Theory and Guarantees Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29102 - Locality-Aware Redundancy Pruning for LLM Depth Compression Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.27786 - Machine Learning and ARIMA Model Averaging for Adaptive Public Health Forecasting: Comparative Evaluation and an Ontario COVID-19 Case Study Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.20406 - How Language Models Choose Sides: Internal Representations of Instruction Hierarchy Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28648 - Titans-QFWP: A Regime-Aware Hybrid Quantum Fast Weight Programmer for Portfolio Optimization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29093 - Diffusion-Based Inverse Design of Dielectric Resonator Metasurfaces for Shaping Smart Electromagnetic Environments Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29907 - Benchmarking External Generalization of SPD Matrix Learning for Resting-State fMRI Connectome Prediction Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30418 - Minimax bounds for watermarked and masked recursive discrete distribution estimation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31091 - The PUR-1 Cyber-Physical Digital Twin Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30186 - Spaghetti Architect: A Contamination-Resistant, By-Construction-Labelled, Multi-Language Code Dataset Generator Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.18642 - The Hallucination Signal Is a Mean Shift: Why Simple Probes Suffice Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28930 - Off-Context GRPO: Learning to Reason on Hard Problems using Privileged Information Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.19313 - QuantumBoostNet: Hybrid Classical-Quantum Cardiac View Identification Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.27302 - Evolutionary Soups: Evolving Mixture-of-Experts for Multi-Objective LLM Alignment Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29978 - Learning-Theoretic Foundation for General Coded Computing: The Straggler Setting Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28910 - ARMOR: Manifold-Oriented Training for Adversarially Robust Aerial Object Detection under Data Scarcity Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29510 - Multivariate Scientific Data Compression with Learned Cross-Variable Latent Decorrelation and Autoregressive Entropy Modeling Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30262 - Reinforcement Learning for Symbolic Equation Solving Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30162 - Generation of High-Level Concepts in 3D Scene Graphs via Autoregressive Diffusion Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28733 - Convergence rates for the RMSprop optimizer with full control of the hyperparameters Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30382 - Self-Supervised Pretext Tasks for Infant Cry Analysis: A Controlled Comparison and a Cautionary Result on Donateacry Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30456 - Automated Batch Distillation Process Simulation for a Large Hybrid Dataset for Deep Anomaly Detection Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.09166 - RSPO: Regularized Self-Play Alignment of Large Language Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2503.00030 - Reduced Matrix Multiplication: Input-Adaptive Matrix-Product Reduction for LLM Inference Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.13426 - "Train classical, deploy quantum" requires rethinking generalization Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31117 - MedCache: Efficient and Temporally Valid Memory for Longitudinal Clinical Agents Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29528 - OISD: On-Policy Internal Self-Distillation of Language Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.29089 - FrameScope: Temporal Data Valuation for Stream Active Learning in Autonomous Vehicle Systems Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28672 - PRISM Edit: One Vector for All Temporal Answers Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.11327 - Toward Postural State Classification in Immersive VR with Multimodal Data and Explainability Analysis Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28844 - Emulating the Forced Response of Climate Models with Generative Machine Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.16929 - Supraglacial Lake Fate Is Knowable Long Before the Season Ends Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30113 - AirFM-DDA: Air-Interface Foundation Model in the Delay-Doppler-Angle Domain for AI-Native 6G Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.00020 - Item-Mean Surrogates: Why Richer Persona Data Fail to Improve LLMs as Human Surrogates Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29455 - WiSP: A Working-Set View of Mixture-of-Experts Serving on Extremely Low-Resource Hardware Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.21868 - RSLM: Training-Free Vector Quantization for Approximate Nearest Neighbor Search Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30384 - Strong Drafts Need Compact Memories: Long-Context Speculative Decoding with Compressed KV Cache Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30252 - Efficient Real-Time Adaptation of ROMs for Unsteady Flows Using Data Assimilation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.23188 - Nonparametric Contextual Pricing and Inventory Learning under Censored Demand Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30944 - Learning Materials Properties from Scarce Labels and Unlabeled Crystals Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30682 - Exact Recovery Thresholds for Weighted Data Selection in Vector-Valued Linear Regression Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30254 - V2TATC: A Joint Voice-Trajectory Embedding Framework and Dataset for Air Traffic Controller Situational Awareness Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28981 - Compact and Infinite-Order Error Analysis for Null-Space SVD Estimation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30374 - BCPPO: Bachelier-Inspired Constrained Proximal Policy Optimization for Tail-Risk-Aware Safe Reinforcement Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30283 - CoJEPA: Combining Contrastive Learning and JEPA for Global-Local Music Representations Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30974 - The Intervention Gap in Latent World Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29998 - Large-Scale Bayesian Tensor Reconstruction via Approximate Message Passing Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2505.16305 - OrScale: Orthogonalised Optimization with Layer-Wise Trust-Ratio Scaling Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.07815 - Fairness in multi-class multi-group classification problems via contextial coherent risk measures Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30223 - Unsupervised Multi-Scale Gromov-Wasserstein Hypergraph Alignment Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29635 - Towards Stream Learning on Embedded Systems: Benchmarking the Memory Consumption of Stream Learning Methods Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30923 - Adversarial Calibration Attack on Autonomous Vehicles Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28778 - Kolmogorov--Arnold against bounded translations Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30710 - Learning diverse attacks on large language models for robust red-teaming and safety tuning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2405.18540 - MWIR-4-Plastic: The Identification of Complex End-of-Life Industrial Plastic using Mid-wave Infrared Hyperspectral Imaging and Machine Learning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28874 - SMOTE-VAR: An Uncertainty-Aware Oversampling Method for Predicting Depression Remission in University Students Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30102 - Selection-Aware Stress Testing for Interactive Agents Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30916 - Higher-Dimensional Rotary Position Embedding Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29715 - What It Costs to Compose, Rebuild, and Correct Precomputed Memory Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30647 - Certified Safety Radii in Forecast-Error Space for Wasserstein Distributionally Robust Small Signal Stability-Constrained AC Optimal Power Flow via Lifted Spectrahedral Containment Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30201 - A Few Teacher Steps Go a Long Way: Cost-Efficient On-Policy Data Augmentation for Agent Post-Training Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.04574 - Generative multi-domain transfer learning for fault detection in data-scarce wind turbines Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30323 - Divide-and-Conquer Modeling for the CTF-4-Science Lorenz Benchmark Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.10084 - Learning PDE Time-Stepping with Neural Cellular Automata Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30328 - The Halt Vector: Internalizing a Causal Steering Intervention for Efficient Reasoning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28859 - Wide Learning: Learning to Reach Evidence Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29608 - Uncertainty-Aware End-to-End AI Weather Forecasting: Disentangling Observation and Model Contributions Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30795 - Demystifying Reinforcement Learning Post-Training of Language Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.24949 - Event-triggered Control and Online Learning for Networked Systems under Computational Delays Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29576 - Selection, Representation, and Execution in Sparse Fourier Neural Operators Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30070 - What Emerges and What Breaks in Self-Play Driving Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30819 - TDDM-Melatt: A Decoupled Memory and Diffusion Framework for Generalizable Encrypted Traffic Classification Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30745 - ECA-BLS: An Efficient Complex-Augmented Broad Learning System Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29763 - On the Recoverability of Private Information Unlearning in Large Language Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29943 - Integrating attention into explanation frameworks for language and vision transformers Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.08966 - Uncertainty Makes It Stable: Curiosity-Driven Quantized Mixture-of-Experts Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2511.11743 - Stick to What You Know: A Study of Knowledge-Aligned Supervised Fine-Tuning Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30987 - Event-Driven Language Models with Sparse Neural Activity for Neuromorphic Hardware Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30439 - Motus2: A Self-Evolving General World Model for Dexterous Manipulation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30237 - SHAKE-GNN: Scalable Hierarchical Kirchhoff-Forest Graph Neural Network Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2509.22100 - A Visual Question Answering Model to Automate Nondestructive Evaluation Image Analysis Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29408 - Aligning Multi-Trajectory Supervision with Policy Optimization for VLA Driving Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30122 - Neural ODE enhanced linear mixed effect models for estimating complex association patterns of time-varying covariates with the marker trajectory Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29714 - Partition-Aware Unlearning for Removing Spurious Correlations in Large Vision-Language Models Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29996 - PokaiTrainer: Scaling Belief-State Search to Competitive Pok\'emon VGC Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29197 - AutoScientist-Quant: Self-Evolving Coding Agents for Automatic Research in Quantitative Investment Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28632 - Balance of Benchmarks: Semantic Density Reweighting for Benchmark Multiplicity and Task-Conditioned Evaluation Source: arxiv-cs-lg Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30044 - STAGEET: Stage-wise Typed Edit Tagging for Grammatical Error Correction with Arabic as a Case Study Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28614 - JPO: Juris Policy Optimization for Structured Legal Reasoning in Criminal Judgment Prediction Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29616 - LLM Judges as Raters: A Pre-Registered Audit of Severity, Halo, Reliability, and Version Instability in LLM Essay Scoring on Public Corpora Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29517 - LexRubric: A Rubric-Guided Diagnostic Benchmark for Open-Ended Legal Tasks Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.09389 - When to Adapt: Conditional Memory Adapters for Retention-Preserving Domain Specialization Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29327 - Opinionated, Hesitant and Stressed: Three Studies of How Politicians Speak in Four Slavic Parliaments Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30828 - En-ViMedNER: An English-Vietnamese Parallel Biomedical Corpus with UMLS Semantic Type Annotations Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29890 - Multi-Agent Self-Improving Reinforcement Learning for Video Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28675 - Using Grounded Theory for Agent Behavior Analysis at Scale Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30391 - Multi-Faceted Interactivity Alignment in Full-Duplex Speech Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.11167 - Improving Argument Saliency Coverage in Small LLMs for Long Legal Opinion Summarization via Sequence-Level Distillation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29884 - POLARIS: Guiding Small Models to Write Long Stories Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.04095 - An evidence-guided reinforcement learning method to improve psychiatric reasoning in small language models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.06449 - OmniACBench: A Benchmark for Evaluating Context-Grounded Acoustic Control in Omni-Modal Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.23938 - Surfacing the Unsaid: CUE-Bench for Affective Stance in Chinese Discourse Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.10810 - Do Small Models Use the Law You Give Them? Measuring Context Use on a Bilingual Bangladesh Legal Benchmark Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30327 - MAGneT: Coordinated Multi-Agent Generation of Synthetic Multi-Turn Mental Health Counseling Sessions Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2509.04183 - Asymmetric Within-Document Predictive Learning for Scientific Document Representation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28625 - TCS-BENCH: Benchmarking State-of-the-Art Generative AI Theoretical Computer Science Research Ability Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.09538 - AdaFuse: Adaptive Ensemble Decoding with Test-Time Scaling for LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.06022 - Geometry of Divergence: Tracking Hidden-State Trajectories for Adaptive Multi-Turn Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30650 - When Errors Become Memories: Causal Pathway Tracing in Multi-Turn Memory-Augmented LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30198 - Attention-Discounted Adaptive Sampler for Masked Diffusion Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.10829 - DVBench: Benchmarking MLLMs for Understanding Dynamic Charts and Narratives in Data Videos Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29711 - EmoTrace: An Emotion Trajectory-Centered Framework for Psychological Support Dialogue Generation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.23648 - MultiHashFormer: Hash-based Generative Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.28057 - Social Caption: Evaluating Social Understanding in Multimodal Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.14569 - Do large language models scrutinise what they review? A multimodal audit of scoring calibration, error detection, and author-identity effects Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28626 - Man Made Language Models? Evaluating LLMs' Perpetuation of Masculine Generics Bias Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2502.10577 - Where Does Long-Context Supervision Actually Go? Effective-Context Exposure Balancing Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.10544 - PromptKWS: A Novel Prompt-Guided Open-Vocabulary Keyword Spotting Framework Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28640 - Augmenting Interviewer Judgments of Patient Experience with Automatic Language Analysis Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31007 - Can Large Language Models Identify Meaningful Touchpoints in Conversion Attribution? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28649 - Toward a Cross-Lingual Romanization Ecosystem for Sinitic Languages: A Paired Mandarin-Cantonese Case Study Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29170 - FLAME: A New Dataset on FLemish Accounts of Momentary Experiences Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2504.14707 - Where Do Multilingual Vision-Language Encoders Fail on Low-Resource Languages? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30725 - GUIDE: Generative Unsupervised Chinese Query Correction via Phonetic and Visual Shared-ID Encoding Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.25343 - Graph Evidence Is Not Enough: Diagnosing Native Decoder Use in Graph-Augmented LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30437 - RouteSparse: Input-Conditional Pattern Routing for Budgeted Long-Context Prefilling Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29058 - Beyond NL2Code: A Structured Survey of Multimodal Code Intelligence Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.15932 - Arabic Sentence Segmentation Across Genres and Punctuation Conditions Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.08025 - Evidence-Bounded Mental Health Reasoning from Heterogeneous Speech Protocols Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31014 - Why We Need Speech to Evaluate Speech Translation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.28227 - XQDT: eXplainable and Quantitative Data-Text Alignment Metric with Feedback Signals Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29948 - Mind the Gap: Theory-of-Mind-Grounded Friction for Epistemic Alignment Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30719 - Hi-Q: Hierarchical Evidence-guided Query Refinement for Multi-Hop Question Answering Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30468 - Quantifying and Mitigating Korean Jamo-Level Typographical Vulnerabilities in Large Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30229 - LaMoC: Loss-Aware Modular Compression for LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30226 - COGTRL: Training LLMs for Scientific Discovery Assistance using Cognitive Traces via Reinforcement Learning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30109 - Quantitative Evidence Mining for Plausibility-Aware Biomedical AI Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30393 - Summarization is Not Dead Yet Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.08000 - Learning Personalized Prompts for Healthcare Guidance Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2412.15957 - On the Optimality of Kinship Naming: an Information-theoretic Approach Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2511.19120 - How Human-Like Are Large Language Models? A Register-Aware Linguistic Evaluation Framework Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.23651 - SynthAVE: Scalable Synthetic Labeling for E-Commerce with LLM-Arena Validation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.07469 - MI-Distillation: Selecting from Model-Interpolated Instruct-Reasoning Data Spectrum for Chain-of-Thought Distillation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29623 - Type-Balanced Contextual Learning for Incremental Named Entity Recognition Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31038 - The Language of the Question Selects the Market: Query Language and Exit IP as Separable Factors in Commercial Recommendations from a Generative Search Interface Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30052 - Reinforcing Step-level Reasoning for Effective Self-Correction in LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.11573 - Evaluating LLMs on Conversational Text-to-SQL under Chain Ambiguity and Intent Drift Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29543 - Terminal-Bench-LILT: Multilingual Agentic Coding Benchmark Grounded in Language, Region, and Culture Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28641 - GeoAgent: Evaluating VLM Geolocalization Through Embodied Navigation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29483 - Beyond Factual QA: Mentorship-Oriented Question Answering over Long-Form Multilingual Content Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.17173 - The Emergent Symbolic Structure of Artificial Neural Networks Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29530 - Will the User Ever Know? Covert Indirect Prompt Injection on Tool-Using LLM Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30362 - QUBRIC: Co-Designing Queries and Rubrics for RL Beyond Verifiable Rewards Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.03968 - Vectors from Larger Language Models Predict Human Reading Time and fMRI Data More Poorly when Dimensionality Expansion is Controlled Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2505.12196 - VocalAffectBench: Evaluating Vocal Emotion Recognition in AI Audio Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28932 - A Robust Evaluation of Probe Robustness: Lessons for Reliable OOD Uncertainty Quantification Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.11662 - Graph2Counsel: Clinically Grounded Synthetic Counseling Dialogue Generation from Client Psychological Graphs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.20382 - "Act Like a 5th Grader" is Not Enough: Bounding Knowledge in LLM-Based User Simulators Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30033 - Seeing the Unseen: Visual Similarity for Pixel Language Model Adaptation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30541 - HSRM: Hidden-State Reward Models for Test-Time Verification Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30841 - Labels have Human Values: Value Calibration of Subjective Tasks Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.06631 - Personas Differ from Native-Language Generation: Language Pathways Shape LLM Interpersonal Advice Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30873 - When LLMs Benchmark Themselves: Deconstructing Self-Bias in Automated Evaluation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2509.26600 - From Isolation to Alignment: Unified LoRA for Efficient Multi-Task Learning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.05078 - A.X K2 Technical Report Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30181 - ALTSTEER: Selective Safety Steering for Moving Beyond Hard Refusals to Constructive Alternatives Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30197 - When Does Predictor-Based RL Align with Human Perception? A Study of Subjective Rewards in Codec-Based Speech Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31035 - EvoBrowseComp: Benchmarking Search Agents on Evolving Knowledge Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.13120 - DIASENTINEL: An Auditable Multi-Agent System for Guideline-Grounded Diabetes Risk Screening Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31128 - Beyond Semantic Accuracy: Consequence-Aware Evaluation for Safety-Critical Language Understanding Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.24621 - Low-Resource Preference Adaptation of LLMs via Activation-Based Label Propagation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30902 - Context-Aware Interleaved Batching for WhisperX Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31170 - AlgoWorlds: Benchmarking Tool Use for Global Optimization in Algorithmic Worlds Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29397 - CaRGo-T: Causal Reasoning Graph-of-Thought improves Multimodal Humor Comprehension Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.23172 - Unknown Unknowns: Do Hidden Intentions in LLMs Evade Detection? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.18552 - Aspire: Can Models Self-Evolve from Vague Goals? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31111 - Causal Interventions Reveal Typologically Organized Syntactic Mechanisms in Multilingual Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28924 - Domain Mixture Design via Log-Likelihood Differences for Aligning Language Models with a Target Model Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.16622 - Parametric Multimodal User Memory: Storing What Captions Cannot Carry Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28609 - VibeJam: A User Study Platform for Web Development with Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29889 - Skill or Skip? Learning Selective Skill Invocation in Agentic Tasks via Dual-Granularity Preference Learning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.00510 - Learning from Many Voices: Literary MT Using Multi-Reference Human and Synthetic Data Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2412.18707 - SkillForge: Compositional Skill Synthesis with Verification-in-the-Loop for Generating Formally Verified Dafny Programs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29841 - Co-Evolving Actor-Conditioned Critics for Non-Verifiable Generation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30397 - Political Ideology Shifts in Large Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.16013 - SocialReasonBench: A Video-QA Benchmark for Social Reasoning with Counterfactual Narrative Videos Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30716 - KVoiceBench, KOpenAudioBench, and KMMAU: Agent-Driven Korean Speech Benchmarks for Evaluating SpeechLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.27984 - ACTD: Anchor-Based Cross-Tokenizer Distillation with Residual Regularization Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29662 - CoReflect: A Reflective Co-Evolution Framework for Improving Conversational Evaluation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.12208 - More Capable, Less Faithful: A Multilingual Analysis of Mathematical (Un)Solvability Detection in LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30463 - CLIN: an Objective Framework for Evaluating Creativity in Short Persian Literary Text Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30754 - Beyond Alignment: Value Diversity as a Collective Property in Multicultural Agent Systems Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.05985 - Cloud and On-Premises Deployment of Uzbek Legal RAG via Targeted Retriever Fine-Tuning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29284 - Different Demographic Cues Yield Inconsistent Conclusions About LLM Personalization and Bias Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.18486 - Demand-Side Measurement for Generative Engine Optimization: Constructing and Validating a Million-Persona, Intent-Annotated Buyer Corpus Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30023 - Voices Across Registers: Corpus-Conditioned Vernacular Jailbreaks against Aligned LLMs via Fanfiction Subgenres Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.04483 - One Success Isn't Reliability: Thinkingbox, a Sandbox and Benchmark for Agents in Stateful Business Workflows Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.19741 - Where Identity Lives: Localized, Retain-Free Identity Unlearning in Multimodal Large Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30649 - GPAgentBench-2K: Benchmarking Large Language Model Agents in Complex Clinical Action Space Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30188 - Simulstream: Open-Source Toolkit for Evaluation and Demonstration of Streaming Speech-to-Text Translation Systems Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2512.17648 - Transformer-Encoder Trees for Efficient Multilingual Machine Translation and Speech Translation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2509.17930 - VoiceCodeBench: Evaluating Exact Structured-Token Recovery in Automatic Speech Recognition Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28916 - GRKV: Global Regression for Training-Free KV Cache Compression in Long-Context LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.31105 - How You Ask Shapes What You Get: A Theory-Seeded Measurement of Articulation in Advice-Seeking LLM Conversations Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29591 - ABLE: Representing and Mapping LLMs via Attribution-Based Large-model Embedding Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.07524 - Scoring, Reasoning, and Selecting the Best! Ensembling Large Language Models via a Peer-Review Process Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2512.23213 - Learning Concepts, Not Tokens: Self-Supervised Semantic Alignment for Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.29123 - Detecting Hidden Chain-of-Thought in Large Language Models with Linguistic, Behavioral, and Mechanistic Indicators Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29956 - Semantic Flow Regularization: Teaching LLMs to Generate Diverse Yet Coherent Responses Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.27971 - Reactivating Test-Time Scaling for Plane Geometry Problem Solving Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30156 - HintEval: An Open-Source Python Toolkit for Hint Generation and Hint Evaluation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2502.00857 - When Calibration Rankings Reverse: Accuracy-Controlled Evaluation for Fair Comparison of LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.30814 - DataFoundry: Evolving Data Preparators via Recursive Self-Improvement Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29966 - You Shouldn't Have Asked: A Pragmatics-Inspired Taxonomy for Evaluating LLM Refusals Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30856 - Vocabulary Growth Fundamentals: Bernstein Functions and Hausdorff Sequences Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29449 - GTA-RAG: Graph-Trajectory-Augmented Reinforcement Learning for Multi-Turn Retrieval-Augmented Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.22479 - To Copy or Not to Copy: Copying Is Easier to Induce Than Recall Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.12075 - Test-Time Training with Next-Token Prediction Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.21803 - SinLlama -- A Large Language Model for Sinhala Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.09115 - Configurable Semantic Chunking for Biomedical Information Extraction in Retrieval-Augmented Generation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31139 - Redesigning and Auditing Deep Research Writing for Faithful Reports Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28643 - When History Is Multimodal: Rethinking Context Management for Long-Horizon Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29897 - AIA$^{2}$: Attribute-Agnostic Imbalance Augmentation for Subgroup Robustness Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30297 - Unveiling Language Routing Isolation in Multilingual MoE Models for Interpretable Subnetwork Adaptation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.03592 - Correctness Forensics for Batch Speculative Decoding: Diagnosing the Ragged Tensor Problem Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2510.22876 - From Speech to Subtitles: Evaluating ASR Models in Subtitling Italian Television Programs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2512.19161 - Automated Researchers Can Reliably Mitigate Alignment Failures Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28945 - From Final Artifacts to Trajectories: Retrospective Process Supervision for Evidence-Grounded Long-Form Generation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30461 - From GenAI Virtual Patient Dialogue Logs to Teacher-Interpretable Process Evidence: A Learning Analytics Study in Higher Education Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28619 - LLM Judges Verify Presence, Not Absence: Omission Blindness in AI Clinical Notes and What Recovers It Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31016 - Agents in the Large: Perception-Centered Architecture for Persistent Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30478 - Recognition-Refusal Misalignment in LLMs: Why Models Answer Structurally Unanswerable Questions Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29109 - Annotated Surrogate Retrieval for Polish Statutory Law Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30929 - UReason: Benchmarking Reasoning-to-Generation Alignment in Unified Multimodal Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.08336 - Hidden Threat in Synthetic Data: Covert Targeted Bias Injection through Benign Text Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30619 - Do We Still Need Humans in the Loop? Human vs. LLM Annotation in Active Learning for TikTok Hate Speech Detection Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.13899 - Auditing Chinese Web-scale Corpora via Sampled BPE Token Statistics Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.10678 - Label Semantic Expansion via Label Guided Neural Topic Modeling Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30216 - UTILMEM: Benchmarking Evidence Utilization in Long-Term Conversational Memory Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30508 - Triggering Chain-of-Thought via Latent Feature Interventions in Large Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.08058 - One note in three: a verified census of three deployed AI scribes, and the instrument that counted it Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31017 - Hindsight Memory-PRM: Supervising Memory Management with Auditable Hindsight Credit Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29605 - Are LLMs Ready to Assist Physicians? PhysAssistBench for Interactive Doctor-Patient-EHR Assistance Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.18613 - Beyond Polarization: The Generative Constraint of Chain-of-Thought in Pointwise Reranking Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30398 - Rate-Coding Bundle Memory: A Unified Model of Memory and Control for Symbolic Computation in the Brain Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29189 - AI Can Be Easily Persuaded in Clinical Decision Making Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29453 - Extracting Small Translation Specialists from LLMs by Aggressively Pruning Experts Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.28042 - Generating Clinical Vignettes that Preserve Cognitive Formulations Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29995 - Beyond the Final Layer: Intermediate Representations for Better Multilingual Calibration in Large Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2510.03136 - The Grammar of Transformers: A Systematic Review of Interpretability Research on Syntactic Knowledge in Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.19926 - Detecting and Repairing Hallucinations in Retrieval-Augmented Generation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29307 - CAST: Critique-Aware Supervision for Training Reliable Long-Horizon Tool-Calling Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30147 - MIND: Unified Inquiry and Diagnosis RL with Criteria Grounded Clinical Supports for Psychiatric Consultation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.03677 - QAQ: Bidirectional Semantic Coherence for Selecting High-Quality Synthetic Code Instructions Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.12165 - Distributional Validity and Calibration of a Korean Synthetic Persona Panel for Digital and AI Service Use: A Secondary-Data Validation Against the Korea Media Panel Survey Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28615 - Revisiting Greedy Decoding for Visual Question Answering: A Calibration Perspective Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.23443 - MUDDLE: Measuring Understanding of Documents under Distractor and Length Effects Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29477 - SwarmBench: Can Large Language Models Act as Agent Swarm Orchestrators? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30661 - Look It Up: Analysing Internal Web Search Capabilities of Modern LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2511.18931 - The Depth Flow of Token Representations Is Nonlinear and Does Not Descend Its Own Density Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29706 - PACE-RAG: Patient-Aware Contextual and Evidence-Constrained RAG for Clinical Drug Recommendation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.17356 - CPR for LLMs: Critical-Point Routing against Catastrophic Forgetting in Domain Adaptation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30158 - Faithfulness Is Not Free: Auditing Offline KV-Cache Quantization in Retrieval-Augmented Generation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30996 - Budget-Aware Compression Pipeline for Single-GPU LLM Inference: Methods, Trade-offs, and Coupling Effects Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30076 - WnW: Waxing-and-Waning KV Cache for Long-Form Speech LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.22704 - GraphLit: Learning Text-Enriched Dynamic Character Network Representations for Literary Study Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.28643 - SHADOWBENCH: Toward Reliable Automatic Evaluation of Semantic Alignment in Autoformalization Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29270 - Every Token Leaves a Ripple in the Stream of Thought: Eliciting Model-Internal Token Saliency for Chain-of-Thought Compression Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31066 - Evidence Absence Is Not Evidence Insufficiency: Diagnosing NEI Construction Artifacts in Fact Verification Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.26663 - DIP: Dynamic In-Context Planner For Diffusion Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.03199 - Beyond the Payload: How User Invocation Shapes Coding Agent Vulnerability to Repository Poisoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30686 - GRADE: Probing Knowledge Gaps in LLMs through Gradient Subspace Dynamics Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.02830 - ImageEval 2026: Culturally Grounded Arabic Multimodal Evaluation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30475 - BIRD-History: A Benchmark for History-Driven Text-to-SQL with Fine-Grained Knowledge Annotations Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29345 - Do LLMs Change Their Minds Like Humans? Diagnosing Human--LLM Divergence in Single-Turn Persuasion Judgments Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29803 - Memory-First Fact-Checking: A Knowledge-Graph-Grounded Multi-Agent System for Misinformation Detection Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29617 - Who Flips? Self- and Cross-Model Counterarguments Reveal Answer Instability in LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.16011 - PA3: Policy-Aware Agent Alignment through Chain-of-Thought Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.14602 - How Identity and Opinion Shape Political Sycophancy in LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29198 - Make an Offer They Can't Refuse: Grounding Bayesian Persuasion in Real-World Dialogues without Pre-Commitment Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2510.13387 - SciOrch: Learning to Orchestrate Expert LLMs for Solving Frontier Multimodal Scientific Reasoning Tasks Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.15872 - When Patients Cut In: Extending Clinical Conversational AI Safety to Interruptions Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29241 - Toward Cultural Alignment: Human-Centered Evaluation of Multimodal AI Stories Across Five African Communities Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29209 - GUIDE: Guiding Internal Evidence with Language Instructions Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30712 - Ignorance or Incompetence? Constructing Knowledge-Gated, Verifiable Tasks for LLM Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30322 - Language Proficiency Assessment from Eye Movements in Naturalistic Passage Reading Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30583 - MM-BrowseComp: A Comprehensive Benchmark for Multimodal Browsing Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.13186 - PaperBanana-Interact: Scientific Diagram Refinement with Multi-Turn Human Feedback Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30241 - Thesis Proposal: Toward a Human-Centered and Perspective-Aware Framework for Reproducible ML Evaluation and AI Alignment Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30842 - SemPOI-RL: Aligning LLM Semantic Reasoning for Interpretable Out-of-Town POI Sequential Generation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30399 - Pak3H: Evaluating the Cost of Cultural Mismatch in LLM Alignment with a Human-Contextualized Urdu Benchmark Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30065 - Anchoring Speech with Semantics: A Multimodal Adapter Mechanism for Automatic Speech Recognition in Low-Resource Languages Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29239 - You Know What I Mean: A Benchmark for Agentic Conversational Reference Grounding Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29834 - Beyond Surface Alignment: Grounding the Dynamics of Situational Understanding and Generative Control in LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29610 - StageWell: A Process-Aligned Chinese Corpus for Positive-Psychology Support Dialogue Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29326 - OCR-MetaReasoning Benchmark: Evaluating the Meta-Reasoning Ability of MLLMs in Text-Rich Image Understanding Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30678 - CORE-T: COherent REtrieval of Tables for Text-to-SQL Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.13111 - Towards AI-Assisted Research Writing: Benchmarking LLMs for AI/ML Introduction Generation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.14273 - MA-RAG: Multi-Agent Retrieval-Augmented Generation for Query-Driven Summarization of Longitudinal Parkinson's Disease Assessments Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28624 - PERK: Long-Context Reasoning as Test-Time Learning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2507.06415 - Integrated and Cross-Architecture Interpretation of LLM Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.28006 - When Models Hear What They Expect: Diagnosing Prosodic Heuristics in Multimodal Sarcasm Detection Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30204 - Small Updates, Big Doubts: Does Parameter-Efficient Fine-tuning Enhance Hallucination Detection ? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.11166 - Who Am I? History-Aware Profiles for Student Simulation in Tutoring Dialogues Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.30051 - Writing Style Similarity Reflects Academic Genealogy Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.14843 - When Can We Work in Embedding Space? What Text Embeddings Preserve Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31059 - Entropy-Aware Token Rejection for Improving Speculative Decoding Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2512.23765 - Beyond "To whom it may concern": Tailoring Machine Translation to Audience and Intent Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.03259 - Where Does Robustness Live? Neuron-Guided Adaptation for Retrieval-Augmented Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.02194 - MAPLE: Metadata Conditioned LLM Pretraining for Locale-Aware Question Answering Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.15236 - How Prolific Sellers Self-Present: Dissecting the Communication Patterns of 1.6 Million Reverb Listings Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29952 - Interpretable Predictability-Based AI Text Detection: A Replication Study Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.15034 - REIGN: Refurbished Embeddings with Integrated Guidance Networks for Efficient Context-Length Scaling Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29899 - Emulate or Estimate? The Divergent Strengths of Base and Post-Trained Language Models for Opinion Simulation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.03044 - BLOOM-WILT: Logit Tilting for Behaviour Elicitation in Automated LLM Auditing Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31105 - The Unsampled Truth: Quantifying Prompt Artifacts in LM Psychometrics Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.03357 - Dynamically Allocating Evaluation Effort for Model Ranking Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.03437 - Notes to Self: Can LLMs Benefit from Experiential Abstractions? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.20372 - Rating the Pitch, Not the Product: User Evaluations of LLMs Reflect Expectations More Than Performance Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.05113 - Do Language Models Reason Across Languages? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.06644 - Do MLLMs Really Understand Low-Resource Khmer Documents? A Pilot Study on Khmer Document VQA Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28635 - Revising Context, Shifting Simulated Stance: Auditing LLM-Based Stance Simulation in Online Discussions Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.06443 - SGM: Safety Glasses for Multimodal Large Language Models via Neuron-Level Detoxification Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2512.15052 - Token-Efficient Data Reasoning Agents via Adaptive Structuring of Unstructured Data Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31082 - REER-PT: Reverse-Engineered Reasoning for Perplexity-Guided Pre-training Data Augmentation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30627 - When LLM Meets Tree Search: A Systematic View of Inference as Search in Large Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30395 - Attention Amnesia in Hybrid LLMs: When CoT Fine-Tuning Breaks Long-Range Recall, and How to Fix It Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.11052 - When Less Is More: An Empirical Study of Minimal Responses in Counseling Dialogues and the Behavior of LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.24080 - WebWorld: The Browser as a World Model for Self-Improving Web Code Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30530 - FBHM: Functional Benchmarking and Steering of VLMs for Hateful Meme Detection Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.31349 - IndicQE-APE: A Benchmark for Quality Estimation and Automatic Post-Editing for Indic Languages Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.16344 - A Comprehensive Survey on Linguistic Steganography: Methods, Countermeasures, Evaluation, and Challenges Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29077 - ReTrace: Rejected-Trajectory Conditioning for Speculative Decoding Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29748 - On the Design of Qwen3.8-Next Architecture: Evaluation, Efficiency, and Training Stability Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30320 - MURANO: Design, Run, and Reproduce Mechanistic Interpretability Experiments as Composable Pipelines Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30662 - S3Gym: Can LLMs Turn Self-Testing and Self-Judging into Self-Improvement? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31100 - Beyond Fluency: A Rubric-Based Benchmark for Evaluating Saudi Dialect and Cultural Competence in Large Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29990 - Don't Forget Your Embeddings: Robust Knowledge Erasure via Precise Editing of Embeddings Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.03695 - Beyond Semantic Similarity: Reducing Unnecessary API Calls via Behavior-Aligned Retriever Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.14323 - Beyond Surface Forms: Symbolic Edits as a Test for Logical Reasoning with LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30256 - AtlasNLP: A Country-Aware Atlas of Dataset Representation in NLP Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30107 - Peak-Then-Collapse and the Four Interface Channels of Knowledge-Graph Tool Use Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.26037 - Are Online Skill and Memory Modules Always Worth Their Tokens? A Budget-Constrained Study of Web Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.15017 - Which India Survives Translation? Narrative Homogenisation Across Indian Oral Traditions in LLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.26123 - ReVA: A Region-Aware Visual Assistant for Visually Grounded Question Answering Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28707 - Syntactic Belief Update as the Driver of Garden Path Processing Difficulty Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.27206 - QUACK: Questioning, Understanding, and Auditing Communicated Knowledge in Multimodal Social Deduction Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.27068 - Not All Fallbacks Are Failures: Understanding and Recovering from Fallbacks in Mobile Voice Assistants Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30738 - EvoSkill Injection: Red-Teaming Autonomous Skill Generation and Evolution in Self-Evolving Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30429 - Sequential Trajectories and Simultaneous Blending: Multi-Emotion Modeling for Instruction-Following TTS Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30325 - SemTrace: Source-Grounded Semantic Signatures for Tracing LLM Exposure to Protected Documents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29575 - Heterogeneous Dependency Graph-Guided Attentionfor Patent Representation Learning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.10073 - DuplexGen: Adaptive Synthesis of Human-AI Turn-Taking Dialogues Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.26178 - TRIPPULSE: Multi-Agent Travel Planning with Review-Grounded Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30924 - Tastes without distinction: silicon samples and the synthetic construction of tastes Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.30085 - Can Released LLM Vocabularies Support Token-Level Estimation of Hidden Corpora? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.10690 - When Less is More: Understanding When Token Filtering Helps and Fails in AI-generated Text Detection Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29903 - BioDivergence: A Benchmark and Evaluation Framework for Hidden Contextual Contradictions in Biomedical Abstracts Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.11208 - Relevance as a Vulnerability: How Web Retrieval Degrades Safety Alignment in LLM Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.29224 - Responsible Integration of AI in Cancer Genomics: Barriers, Risks, and Pathways to Trustworthy Clinical Translation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30912 - Token Counts Are Not Model Lineage: A Frozen-Threshold Holdout Study of Black-Box LLM API Fingerprinting Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29930 - Reasoning Beyond Language: A Comprehensive Survey on Latent Chain-of-Thought Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2505.16782 - MemDefrag: Latent Memory Defragmentation for Large Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.05969 - Verification-Aware Training for Speculative Decoding Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30135 - BenGER: Benchmarking LLM Systems on Subsumption-Based Legal Reasoning in German Law Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.28183 - Evaluating the Capabilities of LLMs for Persuasive Dialogue Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29738 - GenRubric: Self-Evolving Rubric Generation for Scalable LLM Evaluation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29856 - HiVe: Beyond Static Prompts for Multitask Learning via Hierarchy-based Vertical Mixture-of-Experts Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29790 - When Chain-of-Thought Fails, the Solution Hides in the Hidden States Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.23351 - Generative vs. Encoder Models for Multilingual NER: A Comprehensive Empirical Study on Naamapadam Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29959 - The First Token Is a Clue: Verbalizing Multi-Token Concepts from the J-lens Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31084 - Small Reward Models via Backward Inference Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.13551 - CARE: Privacy-Compliant Agentic Reasoning with Evidence Discordance Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.01113 - PrivBench: A Holistic and Modular Benchmarking Platform for Evaluating Text-to-Text Privatization Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29624 - Apples to Apples? Towards Comparable Crosslingual Language Model Evaluation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.25089 - LLMs Can't Play Hangman: On the Necessity of a Private Working Memory for Language Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.06973 - Mechanistic Diagnostics of Spatial Lexical Bias in Multimodal Large Language Model Spatial Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.01914 - Lazy Grounding: Attacking Search Agents with Factual Evidence Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30303 - SpanCalib-VLM: Calibrated Hallucination Span Detection in Vision-Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29974 - Conducting Stylistic Analysis of Paintings through an Art-History Agent Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29644 - Sleight of Word Benchmark: Can Language Models Notice If Their Own Output Was Tampered With? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29921 - Skill-Conditioned Gated Self-Distillation for LLM Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.28791 - Privacy-Preserving Generation of Clinical Narratives from Medical Terminologies Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2509.10882 - Manac\'a-1B: An Open, Reproducible Brazilian-Portuguese Language Model and a Tokenizer-Aware, Paired Evaluation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30114 - SafeMath: Inference-time Safety improves Math Accuracy Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.25201 - SynFlow: A Multidimensional Diachronic Semantic Analysis Toolkit Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.19472 - The Hermon Moment: AI Self-Transcendence and Its Human Narration Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30971 - Don't Read Everything: A Curvature-Conditioned Query for Linear Attention Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.01294 - Kinship Data Benchmark for Multi-hop Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.07794 - FocusAgent: Simple Yet Effective Ways of Trimming the Large Context of Web Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2510.03204 - All You Need Is Non-Commutative Words Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29314 - NLP-Driven Knowledge Extraction and Thematic Classification of Translated Ancient Indian Medical Texts Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28608 - Which one is banana man? Evaluating vision-language models in multi-turn pragmatic interpretation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29571 - Modality Fault Lines: Structural Corruptions Reveal Fragile Omni-Modal Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29278 - IndicDetect: Evaluating Cross-Lingual LLM-Generated Text Detection for Hindi, Telugu, and Tamil Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29919 - Cross Lingual Transfer in Tulu Legal Comprehension: Script-Dependent Improvement and RAG-Induced Knowledge Conflict Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28645 - Not All or None: Dynamic Construction of Target-aware Memory Graph for Conversational Stance Detection Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29066 - Vocal Music under Phoneme-Conditional Analysis Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30823 - Proof2Hybrid: Automatic Mathematical Benchmark Synthesis for Proof-Centric Problems Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.02208 - RegDivergence-101: An LLM Benchmark for Cross-Jurisdiction Regulatory Contradiction Detection in Life Sciences Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28607 - Reverse Probing: Supervised Token-level Uncertainty Quantification for Large Language Models in Clinical Text Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.28740 - MedConceal: A Benchmark for Clinical Hidden-Concern Reasoning Under Partial Observability Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.08788 - EmoLASP: Emotion Recognition with Language Models and Answer Set Programming Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29035 - CoVA-SFT: A Large-Scale Dataset for Chain of Visual Abstractions Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28958 - LCoT-GV: Graph Attention Networks for Verifying Long Reasoning Chains in Large Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30679 - Pad\=artha: Ontology-Grounded Fine-Grained NER Benchmark for Classical Sanskrit Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29324 - Looking Again: Measuring Sycophancy in the Reasoning Chains of Multimodal Models Under Pressure Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28623 - Reasoning over Grammar: Can Synthetic Linguistic Reasoning Traces Enhance Low-Resource Machine Translation? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.03782 - Enabling Proactive Spoken Turns via a Generalized Style-Aware Full-Duplex Framework Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28630 - Can MLLMs Critique Like Humans? Evaluating Open-Ended Aesthetic Reasoning in Multimodal Large Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.29689 - MMDS-Bench: Benchmarking Multimodal Large Language Models on Dynamic Stance in Social Media Interactions Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30903 - Investigating Social Bias Changes in Quantized Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.06181 - Unlocking Fine-Grained Translation Quality Estimation in LRMs through Mutually Boosting Implicit and Explicit Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.31378 - SyRuP: Enhancing System-Prompt Following via Reward-Guided Prediction in LLM Decoding Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.23991 - Beyond Consensus: Downward Bias and Role Asymmetry in Multi-Agent LLM Judges for Subjective Evaluation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30373 - WikiSTAR: A System for Shedding Light on the Hidden History of Scientific Wikipedia Articles Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.12441 - Lot Machine: Multimodal Lot Extraction from Auction Catalogs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30510 - PEPPER: Perception-Guided Perturbation for Robust Backdoor Defense in Text-to-Image Diffusion Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2511.16830 - Learning Composable Chains-of-Thought Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2505.22635 - Error-Type-Aware Loss Reweighting for Robust Named Entity Recognition with Noisy LLM Labels Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30827 - Tracing the Latent Threads: A Mechanistic Study of How LLMs Represent and Operationalize Race and Ethnicity Cues Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.12868 - General Phrase Debiaser: Debiasing Masked Language Models at a Multi-Token Level Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2311.13892 - Large Language Models Systematically Favor Popular Options: Evidence and Mitigation Across MCQs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29257 - To Retrieve or To Think? Cross-Boundary Context Evolution for Multi-hop Complex Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.08747 - Can LLMs Take the Pulse of the Economy? A Real-Time Evaluation of LLM Nowcasts on Macroeconomic Indicators Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30110 - Check The Scoreboard: An Analysis of Scoring Schemes on Multiple-Choice Evaluation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29887 - Decomposing Wrong-Consensus Agreement in LLM Self-Consistency Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.18795 - Generative Models Enhanced by Sequence Labelling and Aspect-Code Switching Improve Cross-lingual Aspect-Based Sentiment Analysis Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30425 - OASIS: Optimizing Attacker Sequences for Hard-Label Black-Box Text Attacks Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29568 - Fund2Persona: A Framework for Building and Refining Financial Advisor Personas from Fund Disclosure Data Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.29793 - Switch Attention: Towards Dynamic and Fine-grained Hybrid Transformers Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.26380 - Exploring Autonomous Agentic Data Engineering for Model Specialization Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.30407 - When Similar Means Different: Evaluating LLMs on Arabic--Hebrew Cognates Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.13218 - Decompose, Look, and Reason: Reinforced Latent Reasoning for VLMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.07518 - Calibrating Small Language Models for Claim Check-Worthiness Detection Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30731 - A Hub of Short Rows Inflates Intrinsic Dimension Estimation of Token Embeddings Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29702 - Universal Activation Verbalizer: A Unified Framework for Cross-Model Activation Explanation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.25903 - Learning to Reason and Use Tools through Unsupervised Fine-Tuning in Task-Oriented Dialog Systems Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30426 - EVAR: Evidence-Validated Hypothesis Admission for Budget-Aware Narrative Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29835 - Popular but Wrong: Understanding and Mitigating LLM Overconfidence through Knowledge Popularity Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2505.17537 - HeTGB: A Comprehensive Benchmark for Heterophilic Text-Attributed Graphs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2503.04822 - No Detectable Change in Side-Level WER from Prompt-Level Context: A Preregistered Ablation on a Production Oral-History Corpus Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28875 - Whose Assessment of Distress? Community Perspectives and LLM Alignment on Well-Being Posts Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29446 - LLP: LLM-Based Product Pricing in E-commerce Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2510.09347 - Ontology-Guided Multi-Agent Extraction of Evaluation Objects from Academic Review Texts: Evidence from Chinese Library and Information Science Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29526 - The Fragility of Jailbreak Robustness Across Operational States Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30748 - A^2Agent: Action-Aware Reinforcement Learning for Repository-Level Code Localization Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29831 - Counter with Evidence! A Multi-Agent Memory Efficient Reasoning Framework for Hate Category Informed Counterspeech Generation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.23152 - Arabic Safety Alignment as Selective Refusal: An Empirical Study of SFT, DPO, and Guard Calibration Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29378 - Prefilling-dLLM: Predictive Prefilling for Long-Context Inference in Diffusion Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.10537 - Pearmut: Human Evaluation of Translation Made Trivial Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.02933 - Call Neighbours Yourself: Graph Walks with Destination-Conditioned On-Policy Self-Distillation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29588 - Attribute-Based Activation Steering of LLMs for Group-Specific Explanation Generation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29215 - X-Coder: Advancing Competitive Programming with Synthetic Tasks, Solutions, and Tests Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.06953 - EpiQAL: Benchmarking Large Language Models in Epidemiological Question Answering and Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.03471 - Standardizing Longitudinal Radiology Report Evaluation via Large Language Model Annotation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.16753 - KoALa-Bench: Evaluating Large Audio Language Models on Korean Speech Understanding and Faithfulness Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.19782 - DRIP-R: A Benchmark for Decision-Making and Reasoning Under Real-World Policy Ambiguity in the Retail Domain Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.07699 - PaperGym: Rubric-Centered Evolution for Research-Plan Generation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31119 - Polyglot Teachers: Evaluating Language Models for Multilingual Synthetic Data Generation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.11290 - LITERARYBIGFIVE: Author-Personalized Text Generation in a Unified Interpretable Space Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.23124 - Enhancing Low-Resource Language Reasoning via High-Resource Language Feature Transfer Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30462 - Argument-Aware Semantic Alignment of Normative Texts: A Toulmin-Based Neuro-Symbolic Approach Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29529 - PAUSE: Editable Strategy Artifacts for Long-Form Cultural Story Adaptation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28633 - Intelligent Identification and Repair of Design Defects in BIM via Domain-Specific Large Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28629 - Bye Bye Perspective API: Lessons for Building and Governing Measurement Infrastructure Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.25580 - Dual-Layer Agentic Memory with Fast Write Routing and Slow Consolidation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.22215 - Gurukul AI: An Interactive AI-Driven Educational Platform for Indian Education System Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28611 - Evaluating Multilingual Sentence Embeddings for Translation Error Detection:An English--Greek Contrastive Study Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28776 - SPADE: Self-Play in Adaptive Synthetic Executable Environments Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.19197 - AI Historian: Helping historians organize and verify person-centred temporal clues from dispersed historical narratives Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29133 - R$^2$A: Learning Persona Policies Through Persona Representation Learning and Runtime Alignment Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29798 - EMemBench: Interactive Benchmarking of Episodic Memory for VLM Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.16690 - Latent-Space Intervention for Cross-Lingual Factual Consistency: Consistency Improvements without Accuracy Drops Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28860 - MAIGO: Mitigating Lost-in-Conversation with History-Cleaned On-Policy Self-Distillation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.27186 - Reinforcement Learning Can Amplify Emergent Misalignment from Harmless Rewards Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.31328 - Small Language Models as Judges for Rubric-Based Reinforcement Learning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30005 - Beyond Good Intentions: When Does the Framing of Multilingual and Low-Resource NLP Research Become a Caricature? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30866 - TaxCE : A Framework for Automated Taxonomy Construction and Evaluation at Scale Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30614 - WildSEEK: Evaluating Language Models for Information-Seeking Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30683 - When Hate Meets Facts: LLMs-in-the-Loop for Check-worthiness Detection in Hate Speech Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.25269 - Chain-of-Thought Faithfulness of Reasoning Models Varies with Where and How Preference Cues Are Delivered Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29464 - CogEvol: Towards Efficient and Reliable Learning Environment Generation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30968 - Quantifying Error Tolerance in Synthetic Data: An Atomic-level Operand vs. Operator Perturbation Study Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29144 - SUP-MIMIC: A Multi-Task Clinical Diagnosis Benchmark for Evaluating LLMs' Robustness to Contradictory Evidence Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29582 - Read the Room, Read the Image: Understanding Indirect Speech Acts in Multimodal Visual Contexts Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30270 - Evaluating and Improving LLM Self-Modeling Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30980 - Language-Statistical Analysis of Neural Audio Codec Tokens Across Architectures, Corpora, and Noise Conditions Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31037 - Arkios: An Open Bilingual English-Nepali Language Model Trained From Scratch, with a Devanagari-Aware Tokenizer Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30092 - MemoNoveltyAgent: A Historical Research Memory-Aware Agent Workflow for Paper Novelty Assessment Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.20884 - ManGo: Manga Active Narrative Grounding Optimization Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29865 - Linguistics-Aware Non-Distortionary LLM Watermarking Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.00613 - Turning the Spell Around: Lightweight Alignment Amplification via Rank-One Safety Injection Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.20766 - Extending AI for Research to the Humanities: A Multi-Agent Framework for Evidence-Grounded Scholarship Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.30947 - ContextBias: Controlled Evaluation of Bias Persistence Under Context Shift in Text-to-Image Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29847 - HEAR Who Said What: Unlocking Speaker-Attributed Reasoning via Counterfactual Voice Grounding Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29120 - Randomized YaRN Improves Length Generalization for Long-Context Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.23687 - QQ: A Language Metadata Toolkit for Multilingual NLP Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.00620 - SHERLOC: Structured Diagnostic Localization for Code Repair Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.24820 - Evaluating the Semantic Specificity of Representation Steering in Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29431 - The Differential Reasoning Router: Operationalizing Cost-Aware LLM Annotation in E-commerce Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30224 - A Unifying Perspective on Language Model Representations: From Filler-Role Structure to Mechanistic Interpretability Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29034 - Do Large Language Models Possess a Theory of Mind? A Comparative Evaluation Using the Strange Stories Paradigm Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.18007 - Measuring the Depth of LLM Unlearning via Activation Patching Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.24614 - Anchored Decoding: Provably Reducing Copyright Risk for Any Language Model Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.07120 - Why LLMs Give In: Conversational Factors and Reasoning Behind Medical Sycophancy Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.01017 - What Matters When Building Universal Multilingual Named Entity Recognition Models? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.06347 - Identifying and Mitigating Bottlenecks in Role-Playing Agents: A Systematic Study of Disentangling Character Profile Axes Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.04716 - ArabicDialectSafety: A Dialect-Aware Benchmark for Arabic Content Safety Classification Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.01291 - Improving Information Extraction with Learned Queries Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31058 - Auditing MCQA Benchmarks through Probability Landscapes Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30372 - Physics-R1: An Audited Olympiad Corpus and Released Verifiers for Visual Physics Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.14040 - Stratified Consistency Distillation for Natural Language Formalization Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30258 - Clinically Grounded Privacy Evaluation of Medical LMs Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.09590 - Evaluating and Mitigating Anti-LGBTQ Biases in German and Multilingual Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30884 - Super Library Agent: Joint Generation and Maintenance of Multiple Applications Beyond the Single Codebase Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29310 - Agent Zero Memory: Provenance-Aware Long-Term Memory for LLM Agents Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29606 - GreenBench: Benchmarking Energy Efficiency and Carbon Footprint of Open-Source LLM Inference on Apple Silicon Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28667 - CoCoA: Context-Conditional Cultural Alignment for Large Language Models Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29492 - RENSA: Rich Environment Metadata to Navigate Shared and Distributed Endpoints for Automated Federated SPARQL Query Generation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28963 - ECGQuest: Benchmarking and Fine-Tuning Language Models for Electrocardiography Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30893 - Towards a Joint Khmer Text Recognition and Word Segmentation Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30213 - SIC-Agents: Benchmarking and Building an Adaptive Simulator for Pediatric Serious Illness Communication Training Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29481 - Agentic Chain-of-Thought Steering for Efficient and Controllable LLM Reasoning Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.03965 - A Simple Method to Enhance Pre-trained Language Models with Speech Tokens for Classification Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2512.07571 - ScienceArena: Benchmarking LLMs on Latest Scientific Olympiad Competitions Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30517 - How Order-Sensitive Are LLMs? OrderProbe for Deterministic Structural Reconstruction Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.08626 - Detecting AI Impostors: How Do Middle Schoolers Identify LLM Agents in a Live Collaborative Setting? Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30948 - Detecting and Guiding LLM-Generated Korean Poetry with Interpretable Form-level Features Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28986 - How Mental Health Self-Disclosure Becomes Visible: Evidence from Eight Conditions on Reddit Source: arxiv-cs-cl Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29010 - SPARK: Skeleton-Guided Reasoning Synthesis from Large-Scale Scientific Literature Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30214 - LightNav-0: Eliciting VLM Spatial Intelligence for Generalist Embodied Navigation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30935 - HANIA: Planner-Guided Multimodal Graph Evidence Selection for Grounded Question Answering Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29088 - Scaling Large Reasoning Models beyond Human Supervision: A Path toward Superintelligence Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31075 - Hidden Thoughts Are Not Secret: Reasoning Trace Exposure in LLMs Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.00642 - Diagnose, Then Refine: A Closed-Loop TTS System with AudioLLM-Guided Correction Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28970 - SciAtlas: A Computable Atlas of Science for Knowledge-Grounded AI Research Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.22878 - Capability-Stratified Degradation in Ternary Language Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28809 - Generating Workflow DAGs from Natural Language with Non-Reasoning LLMs Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30250 - Frequency Selective Neural Networks as a Foundation Architecture for Time Series Learning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29012 - Facts Without Rules: Boundary Metadata Collapse in Multi-Agent LLM Handoffs Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29028 - Beyond Dense States: Sparse Transcoders as Causally Testable Operators for LLM Latent Reasoning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.01695 - Agora: Enhancing LLM Agent Reasoning Via Auction-Based Task Allocation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.09600 - Interpreting and Steering for Safe and Correct Code Generation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30025 - How Small Can You Go? LoRA Fine-Tuning 270M-8B Models for Merchant Information Extraction in Financial Transactions Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.08051 - PEAR: Permutation-Equivariant Adaptive Routing Multi-Agent Debate Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.20621 - Summoning the Oracle to Slay It: Mitigating Look-Ahead Bias in Financial Backtesting with Large Language Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.24564 - Evaluating LLM-based AI agents integrated with materials synthesis tools: the case of atomic layer deposition Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29309 - Rethinking the Test-Time Prompt Tuning Objective from the Perspective of Calibration Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30230 - PRISM: Predictive Recomposition via Semantic Latent Decomposition for View-invariant Video Representation Learning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30388 - ReVEL: Multi-Turn Reflective LLM-Guided Heuristic Evolution via Structured Performance Feedback Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.04940 - Perceive to Hypothesize, Verify to Ground: An Agentic Reasoning Framework for Open-World Geo-Localization Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29880 - Selective Forgetting: A Graph-Based Memory Framework for Long-Term LLM Agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28978 - Not Safe for All: Auditing the Dialect Penalty in Text-to-Image Safety Pipelines Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29589 - Co-Annotator: Expert-Distilled ViT and VLM for Visual and Documentation Guidance in Age-Related Macular Degeneration Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30352 - Whole-Slide Image Analysis under Realistic Few-Shot Annotation Protocols Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30420 - SynCrash: A Multi-Stage Pipeline for Zero-Shot Accident Detection and Localization in Traffic Surveillance Video Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29759 - Adopt $\neq$ Adapt: Longitudinal Analyses of LLM Conversations in the Wild Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.29018 - BiasMix-Finance: Post-Generation KYC Guardrails for LLM Portfolio Advice Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28646 - LLM-Based Knowledge Graph Completion Combining Discrete Structural Coding with Similar Entity Information Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30235 - Formal Concept Analysis with Three Types of Negation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29311 - Background-Free Objectness Learning for Class-Agnostic Detection Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29232 - Efficient Geothermal Well-Control Optimization via Diffusion-Surrogate Reinforcement Learning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28791 - Revolutionizing Turn-by-Turn Navigation with Cloud-Edge Deep Learning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29073 - FRAMEWORKERS: A Dynamic Multi-Agent Framework for AI-Generated Video Production Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29814 - From Metaheuristics to Exact Methods: A CP-SAT Approach for Multi-Objective Healthcare Workforce Scheduling Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30419 - CM2: Multimodal Cultural Reasoning via an Integrated Multi-Agent Framework Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30498 - DiagFlowBench: Evaluating How Language Models Handle Off-Procedure Inputs in Grounded Diagnostic Dialogue Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.17904 - Cross-Relational Preference Learning for Better LLM Instruction Following Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29352 - The Policy Deficit in AI x Social-Emotional Learning Research Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29950 - PowerSlider: Exploiting Phase Asymmetry for LLM Serving under Demand Response Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.21719 - RoSe-SLAM: Robust Semantic-Aware Gaussian Splatting SLAM from Dynamic Monocular Videos Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29003 - Unified-MAS: Universally Generating Domain-Specific Nodes for Empowering Automatic Multi-Agent Systems Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.21475 - Evolving Agents in the Dark: Retrospective Harness Optimization via Self-Preference Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.05922 - Hybrid Offline-Online Multi-Agent Decision Transformers for Wireless Resource Management Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28878 - Integrating Triaxial IMU Sensors and Ensemble Learning for Effective Parkinson Disease Severity Classification Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28602 - Predicting Residential Rents in Dakar Using Machine Learning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30865 - Measuring Similarity between Artistic and AI Generated Images using Siamese Neural Networks Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28671 - Computational Depth Measurement in Thermographic Video: Overcoming Spatial Overfitting via Spatio-Temporal Decoupling Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29223 - Learning-Assisted Congestion-Aware Route Scheduling for Semiconductor Fab Material Control Systems Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30520 - TAAL: Mitigating Early Beam Pruning in Generative Recommendation via Temporal Autoregressive Alignment Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29179 - SGE: Semantically-Guided Exploration for Unstructured Environments via Image-Space Waypoint Sampling Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29315 - MR-JEPA: A General Purpose Video Foundation Model for Cardiac MRI Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30975 - ForesightSafety-SAGE:A Fully Automated Scenario Generation and Safety Evaluation Framework for LLM Agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.08531 - FigMirror: Ground It, Code It, Plot It Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28814 - GUI-PRA: Process Reward Agent for GUI Tasks Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2509.23263 - SkillRet: A Large-Scale Benchmark for Skill Retrieval in LLM Agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.05726 - SPRInG: Continual LLM Personalization via Selective Parametric Adaptation and Retrieval-Interpolated Generation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.09974 - Which Rules Matter Now? Policy-Centroid Routing Before an Intelligent System Acts Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30757 - Measure Before You Manage: Evaluating Agent Working Memory in Coding Agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31057 - Stop Before You Fail: Operational Capability Boundaries for Mitigating Unproductive Reasoning in Large Reasoning Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2509.24711 - Agent2UCB: Agentic System for Generative Engine Optimization Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29063 - ScaffoldAgent: Utility-Guided Dynamic Outline Optimization for Open-Ended Deep Research Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.20122 - LOCI: A Locator-Critic with Refinement Loop Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30959 - MusGU+: Toward a Musician-Centered Evaluation Framework and Discovery Tool for Generative Music AI Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30940 - SearchWiki: Learning to Build and Navigate Knowledge Wikis for Active Information Seeking Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29953 - Augmenting Human Performance with an XR Agent Learning from Online Behavior and BCI Evidence Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30369 - SafeAtlas-VL: Beyond Binary Multimodal Safety with Large-Scale Data and Guard Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29098 - Towards Natural Personalization: Evaluating Long-Horizon Preference Following in Personalized User-LLM Interactions Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.04191 - SePArate: Segmenting Patterns from Defects in Wafer Manufacturing Using Weak Supervision Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30410 - DS-Lighting: Making Agent Harnesses Explicit for Data-Science Automation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28590 - IterCAD: An Iterative Multimodal Agent for Visually-Grounded CAD Generation and Editing Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.13368 - AOI-Net: Structural Face AOI-Guided Eye-Gaze Track Representation Learning for Autism Spectrum Disorder Detection Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29289 - VERA: Authority-Preserving Edge Revocation for Federated AI-Agent Workflows Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30091 - Reviving our data foundations is the most disruptive step to data maturity Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29368 - AssetOpsBench: Benchmarking AI Agents for Task Automation in Industrial Asset Operations and Maintenance Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2506.03828 - Review Before Trust: Source-Grounded Integrity Gates for AI-Assisted Personal Health Records Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29965 - RAGDiffusion++: From Macro-Retrieval to Micro-Fidelity Alignment for Garment Generation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29280 - EvoGenUI-Bench: Evaluating LLMs as Multi-Turn Generative UI Assistants Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29387 - DynamicMCPBench: A Trace-Grounded, Effect-Scored Benchmark for LLM Agents over Live MCP Servers Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.20531 - AGRICAM: A Track-Mounted Crop Pollination Monitoring Robot Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29237 - LLMs Interpret, Embeddings Organize, Graphs Emerge: Agent-Driven Compilation of Scientific Knowledge Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29612 - PAGE-RAG: Provenance-Aware Graph Evidence Promotion for Fixed-Budget Multi-hop Retrieval-Augmented Generation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29753 - CrossAudit: A Git-Native, Cross-Vendor Audit Loop for Agentic Science Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28631 - CDEP Agent: Connecting Meteorologically Detected Temporal Compound Events to Real-World Documentary Evidence Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28628 - Beyond Pixels: Exploring DOM Downsampling for LLM-Based Web Agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.04412 - A Multi-Agent Human-LLM Collaborative Framework for Closed-Loop Scientific Literature Summarization Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.01452 - MedAgent-R1: Faithfulness-Aware Reinforcement Learning for Evidence-Grounded Medical Reasoning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30676 - FRAC-MAS: A Safe and Explainable Multi-Agent System for Fracture Diagnosis Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28662 - Science sandboxes measure the scientific capability of AI agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30165 - DiffSAC: Diffusion-guided Sampling for Consensus-based Robust Estimation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30603 - Beyond Helpfulness: A Teaching-over-Solving Diagnostic for Measuring Educational Impact in LLM Tutors Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.16206 - AGM: Achievement-Grounded Memory for Closed-Loop Agents with Frozen VLA Policies Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29537 - Not Your Typical Sycophant: The Elusive Nature of Sycophancy in Large Language Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.15436 - E-SENS: Exclusion-Sensitive Penalization for Negative-Constraint Retrieval Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30130 - MineCEraft: Evaluating Language Models as Construction Engineers in the World of Minecraft Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28884 - ATLAS: Dual-Horizon Diagnostic Evaluation for Industrial Tool-Use Agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30685 - Nested Convex-Body Chasing for Online Optimization with Evolving Feasible Sets Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29074 - HumanStudy-Bench: Towards AI Agent Design for Participant Simulation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.00685 - Generative Retrieval for E-commerce: Jointly Learning Embedding and Codebook with Same Product Cluster Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30606 - Pro-Router: Token-Aware Progressive Model Routing with Adaptive Edge-Cloud Collaboration for Efficient Multimodal LLM Inference Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28726 - AgentRx: Diagnosing AI Agent Failures from Execution Trajectories Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.02475 - CrowdMath: A Dataset of Crowdsourced Mathematical Research Discussions Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.06526 - CAER: Causal Action Effect Reweighting for World Model Training Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30897 - HiRS-Agent: A Hierarchical Multi-Agent System for Reliable Long-Horizon Remote Sensing Task Solving Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30672 - Structured State Reconciliation for Human-AI Task Handover Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28907 - CHASE: How Content Ecosystems Are Reshaped When Ranking Is the Only Target Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30466 - TRACER: Per-Tool Context Retention for LLM Agents via Consequence-Attributed Reinforcement Learning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29363 - Can Agentic Trading Systems Pay for Their Own Intelligence? Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.10286 - RailGen: Improving Railway Intrusion Detection via Agent-Guided Small-Scale Foreign Object Generation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30727 - The Role of Network Topology and Opponent Information in Shaping Cooperation in Multi-Agent Reinforcement Learning Systems Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28977 - Automatic Conversion of NICE Guidelines to an Executable Computational Model Using Large Language Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30022 - Exponential random graph models with soft clique constraints Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30869 - EviAnchor: Mitigating Hallucinations in Large Vision-Language Models via Regional Visual Evidence Compensation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29092 - First-Order Efficiency for Probabilistic Value Estimation via A Statistical Viewpoint Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.02827 - VFR-Audit: Verdict-Level Reliability for Fairness Audits in Hospital Length-of-Stay Prediction Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30846 - Mnemosyne: Agentic Transaction Processing for Validating and Repairing AI-generated Workflows Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.00269 - Training-Free Action Correction for VLA Model Failures via Language Feedback Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29967 - Verification abundance, adjudication scarcity: what happens to mathematical knowledge when proof checking becomes free Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28997 - Autoregressive Mosaics: Probing 2D Spatial Reasoning in Text-Only Language Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30751 - Dynamic Important Example Mining for Reinforcement Finetuning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29252 - Real-Time Video Anomaly Detection Using YOLO Pose Estimation and CLIP-Based Semantic Scoring Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31074 - Plant-Inspired AI: Plants as Inspiration for Novel Problem Formulations, and Two Case Studies Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29356 - Beyond the Answer Key: Robustness Evaluation of Large Language Models for Step-Level Mathematical Verification Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28725 - AutoCRAT: Within-trajectory Joint Control of Stochasticity and Compute for LLM Reasoning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29988 - Beyond Correctness: Validity-Oriented Evaluation of Biomedical LLM Judges Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29127 - WorldCupArena: Fine-Grained Evaluation of Language Models and Deep-Research Agents on Football Forecasting Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.18084 - Cost-efficient Active Learning for Referring Image Segmentation and Grounding Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30621 - When Should Models Change Their Minds? Contextual Belief Management in Large Language Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.30219 - JFTA-Bench: Evaluate LLM's Ability of Tracking and Analyzing Malfunctions Using Fault Trees Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.22978 - Centering before Pruning: Lightweight Geometry Correction for Diversity-Based Visual Token Pruning in LVLMs Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30263 - Messier: A High-Resolution Corpus for Cross-Benchmark Agent Evaluation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.25891 - Reachability-Based Capability Confinement for LLM Agents under Indirect Prompt Injection Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30041 - Multimodal Adaptive Expert Selection with Text Routing and Ordinal Prototype Optimization for Sentiment Analysis Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30726 - Instruction-Tuned Language Models Cannot Sample from Distributions They Can Describe Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.25292 - Agent-Based Model Framework for the North Carolina Modeling Infectious Diseases Program (NC MInD ABM) Overview, Design Concepts, and Details Protocol Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2202.06853 - A Mental Model Based Framework of Trust Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2301.12569 - MultivationBench: A Benchmark for Multimodal Sequential Motivation Reasoning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.26465 - FaVOR: LLM-Based Agentic Framework for Factor Mining via Empirical Validation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30192 - Cross-Regional Grapevine Cold Hardiness Prediction via Learned Multimodal Latent Representations Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31097 - Parallel Time-Band Mixing with Learned Observation-Adding for Robust ASR Front-Ends Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30326 - ImageCAS-X: a dataset and benchmark for coronary artery segmentation and centerline extraction in coronary CT angiography Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30404 - Integrating adaptive human behavior into epidemic models with large language models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29535 - RIFT-Bench: Dynamic Red-teaming For Agentic AI Systems Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.23927 - Can escalation channels redirect reward hacking toward defect disclosure? Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29460 - Spatial Matryoshka Training for Multi-Granularity Visual Document Retrieval Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29951 - Accelerating Unified Multimodal Models with Core-Expansion Routing and Unified Computation Scheduling Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29291 - Scaffolding Foundation Models into Physical-World Agents Pushes the Frontier of Long-Horizon Navigation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30396 - CineForge: Self-Improving Agents for Long-Horizon Video Generation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29621 - Game-Agnostic Value Functions through Automatic JSON Feature Extraction Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30056 - TAMI: Temporally Aligned, Missingness-Aware, and Interpretable Multimodal Fusion for Mental Health Assessment in Older Adults with Mild Cognitive Impairment Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30857 - Defining Operational Conditions for Safety-Critical AI-Based Systems from Data Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.22118 - Perceive Before Reasoning: A Pre-Reasoning Perception Framework for Efficient and Reliable Proactive Mobile Agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.03236 - JudgePanel: A Compact Judge with Panel Deliberation via Adaptive Multi-Reward Reinforcement Learning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29168 - Citation-Closure Retrieval and Per-Rule Attribution for Real-World Regulatory Compliance Question Answering Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.29742 - More Perspectives, Stronger Signals: Multi-Perspective Enhancement and Progressive Fusion for Multimodal Entity Representation Learning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29139 - Extending TotalSegmentator: Predicting Patient and Acquisition Characteristics from CT and MR Images Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29348 - PAVE: Predictive Alignment and Value-Guided Evolution for World-Action Policies Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30378 - PeopleSearchBench: Evaluating AI-Powered People Search Platforms with Criteria-Grounded Verification Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.27476 - GarmentWeaver: Schema-Aware Structured Synthesis for Multimodal Sewing Patterns Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30550 - Wrong Prediction, Right Answer: Recovering Evidence from Collapsed LLM Sequence Scores Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31068 - WebXSkill: Skill Learning for Autonomous Web Agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.13318 - AdaPath: Query-Adaptive Path-Finding via Path-Bank for Multi-Hop Implicit Biomedical KGQA Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30556 - FORESIGHT-9: Prospective and Process-Aware Evaluation of Adaptive Trading Agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29372 - ScenePilot: Grow-and-Repair Policy for Text-Driven 3D Indoor Scene Generation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30307 - CARVE: Verified Expansion for Variable-Length Generation in Diffusion Language Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30922 - MedSegBenchmarker: A Raw-Count-First Framework for Controlled 2D Medical Image Segmentation Benchmarks Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29677 - The reach of a verification tool decides its value: A controlled study of verification surface, artifact quality, and cost in AI coding agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28795 - Dense Clinical Contrasts Enhance Medical Knowledge Updating in Large Language Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30405 - EpaCache: Error-Propagation-Aware Caching for Accelerating Diffusion-Based Visual Generation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29264 - A homotopy-type-theoretic generalization of neurosymbolic inference Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.17851 - A Calibration Audit of Confidence in Feed-Forward 3D Reconstruction Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29705 - CoLa-ICD: A Knowledge-Enhanced Framework for Long-Tail Automated Medical Coding Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30234 - AgenticRag-R1: Agentic Reinforcement Learning with Stack Memory for Multi-Step Reasoning, Retrieval and Memorizing Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29622 - A Generalized Optimization Engine (GOE) for Edge AI Inference Acceleration Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28652 - Disentangling Representation using Attributes-based Gaussian Estimation for Medical Sound Diagnosis Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29026 - Expert-validated STEM QA Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28591 - Reconciling Process Supervision with Outcome-Based Credit in Agentic Policy Optimization Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31077 - Improving Randomized Metric Distortion to 2.3282 Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29308 - Aggregate Disambiguation Systems Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30805 - Free Speech and Artificial Intelligence Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28973 - DART: Draft-Agreement Routing for Training-Free Adaptive Thinking Budgets in Hybrid Reasoning Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.23181 - Imag-Eval: a language-grounded framework for interpretable Text-to-Image instruction following evaluation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29210 - Detect Before You Attribute: Cascade Failure Attribution for Multi-Agent Systems Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29646 - INTRYGUE: Induction-Aware Entropy Gating for Reliable RAG Uncertainty Estimation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.21607 - Leveraging Generative AI to Design Accessible Interactive Visualizations for Undergraduate Mathematics: A Six-Phase Workflow Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28601 - PyKEEN-NSX: A Modular Framework for Static, Dynamic and Schema-Aware Negative Sampling in PyKEEN Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30652 - CDPR: Counterfactual Advantage-based Credit Assignment for Cost-Aware Sequential Medical Diagnosis Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28599 - Predicting Future Organ Dysfunction in ICU Patients Using Temporal Convolutional Networks on MIMIC-IV Data Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29301 - STARLINC: Satellite Trail Artifact Removal using Inter-Frame Correlation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29145 - Agentic AI uncovers conserved cross-tissue protein co-abundance programs inaccessible to single-dataset analysis Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28990 - Learning Action Models with Conditional and Quantified Effects via Uncertainty-Guided Exploration Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30955 - Statutory AI: Aligning Large Language Models With Legal Norms Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28593 - TPvG: A Moral Decision Framework for Large Language Models from One-Shot to Sequential Feedback Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28610 - Benevolent Bias in Multi-Turn Human-Agent Dialogue Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29206 - MMPCBench: Benchmarking Multimodal Large Language Models on Proactive Critique of Flawed Inputs Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29286 - EvoSQL: Memory-Augmented Critic-Generator Co-Evolution for Text-to-SQL Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.20489 - LiteSearch-VL: Small Multimodal Search Agents via Trajectory Distillation and Synthetic Step-DPO Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29357 - Taking the Whys Seriously: Limitations of Counterfactual Explanations in Justification and Recourse Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30956 - Preference Shapes Relevance: Cross-component Hierarchical Semantic Alignment for Personalized Generative Retrieval Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30553 - MobileDreamer: Generative Sketch World Model for GUI Agent Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.04035 - Reasoning and Tool-use Compete in Agentic RL:From Quantifying Interference to Disentangled Tuning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2602.00994 - ActiveAugment: Online Active Learning for Augmentation Selection in Deep Learning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28923 - Understanding Deep Learning via Entropy Space Theory Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29279 - Enhancing SAE-based Steering via Neighbor Integrated Feature Selection Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28806 - Influence Is Not Authority: When Causal Guardrail Signals Make Legitimate Tool Use Look Like an Attack in Tool-Using LLM Agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29942 - CGFM-Nav: Cognitive Graph-Field Memory for Semantic-Guided Lifelong Multimodal Embodied Navigation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29114 - Answer Probing-Guided Search for Diverse Solution Exploration of LLMs Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30345 - When Does Bigger Help? A Controlled Study of LLM Scale for Ontology Learning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31118 - AI Scientist Mission Control (AIMC): Visual Analytics for Human Oversight of Autonomous Scientific Discovery Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28637 - MedTVL: Harnessing Vision and Language for Medical Time Series Classification Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28605 - Lucida: Parse, Generate, and Place for Composable Real-to-Sim Scene Modeling Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30821 - The Race between Agentic AI Capabilities and Data Quality Control in Online Surveys Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28597 - SPADE: A Large Language Model Framework for Soil Moisture Pattern Recognition and Anomaly Detection in Precision Agriculture Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2509.18123 - Training and Agentic Inference Strategies for LLM-based Manim Animation Generation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.18364 - Controllable Memory Usage: Balancing Anchoring and Innovation in Long-Term Human-Agent Interaction Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.05107 - MNIST-PRO: MNIST is Back as a Partially Observable World for AI Agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31022 - Paper Pilot: A Human-in-the-Loop Expert System for Evidence-Traceable Scientific Manuscript Generation in Applied Sciences Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28596 - Defending Wearable VLMs Against Private Attribute Inference Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28691 - Can LLM Agents Discover? Evaluating Creativity on ML Engineering Tasks Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30047 - Handoff Debt: The Rediscovery Cost When Coding Agents Take Over Interrupted Tasks Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.02875 - Zipping the Thought: When and How Compressed Reasoning Data Works in LLM Post-Training Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.28008 - AsyncTool: Evaluating the Asynchronous Function Calling Capability under Multi-Task Scenarios Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.27995 - Explainable Artificial Intelligence (XAI) in Computational Pathology: Definitions, Taxonomy, and Recommendations Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28820 - Where Knowledge and Authority Sit Changes What an Agent Benchmark Can Resolve Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.02975 - Toward Latent Language Model Skills Steering and Optimization: An Empirical Study Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29459 - Reliable Benchmarking of Artifact Detection in Computational Pathology: A Reproducibility and Uncertainty Analysis Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30835 - Retrieval-Augmented LLM Agents: Learning to Learn from Experience Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.18272 - Text-Driven Artistic Staging: Pose, Lighting, and Camera References from Paintings Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28823 - C3-UniMM: Causal Cycle-Consistent Unified Multimodal Modeling via Super Alignment and Shared Decoding Space Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28603 - Safe to Resume? Breaking Execution Continuity of Agent Execution via Rollback Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29381 - Training-Free Hidden-State Refinement for Flow-Matching Image Generators Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29160 - Delegating Before Learning: Where Generative AI Sits in Students' Professional Communication Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28837 - SUN: Persistent Programs For Language-Grounded Control-to-Learning-to-Real Policies Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31167 - Hyper3-CLIP: Hierarchy-Conditioned Hyperbolic Vision-Language Training Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29313 - Ideation Arena: Evaluating LLM Generated Research Ideas with Battle-style Human Expert Assessment Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29696 - DocIntent: Answerability-Guided Agentic Restoration for Real-World Document Visual Question Answering Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29037 - SHAPE of Chain-of-Thought in Math Reasoning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28600 - Self-Correction Can Amplify Hallucinations: Fact-Level Repair with Graph-Based Evidence Routing in Multimodal Generation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.00232 - COMPASS: Grounding Composition-Intent Guidance in Unified Multimodal Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.28696 - ReactBench: A Benchmark for Topological Reasoning in MLLMs on Chemical Reaction Diagrams Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2604.15994 - GuardianAgent: Policy-Conditioned Risk-Adaptive Anonymization with Verified Adversarial Escalation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29251 - RoboPhys-3D: A Comprehensive Embodied World Model Evaluation via 3D Reconstruction Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28718 - PUFFER: Incremental Fuzzy Deduplication for Continuously Evolving Corpora Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28622 - TuringLLM: Efficiently Scaling Foundation Models Toward Physical AI Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30567 - SIR: Self-improving Red-teaming for Compute Use Agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30207 - Learning from What You Retrieve: Online RL Fine-Tuning for Semantic Retrieval Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30753 - A collective capability boundary in frontier large language models on guideline-conformant and case-specific oncology decision-making Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28592 - Masked Distillation: Internalizing the Chain-of-Thought in Language Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2607.22629 - One AI Signal, Many Human Judgments: A Bayesian Cascade Analysis of AI-based Credibility Indicators in Online Information Spread Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30311 - Not the Same Protector: Deployment-Dependent Protective Intervention in LLMs Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29136 - Language-Guided Tuning: Configuration Optimization for Automated ML Research Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2508.15757 - PermitGPT: A Unified Generative-AI Pipeline for Construction Hazard Forecasting, Permit Prediction, and Community Impact Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28728 - Cost-Effective Repository Exploration for Agentic Issue Localization Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29675 - Auditing and Mitigating Privacy Leakage in Cloud-Edge Collaborative Decoding Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29111 - From Analytics to Tumor Boards: An Evidence-Linked Multi-Agent Workflow for Oncology Feature Extraction Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28974 - On the Prospects of Dynamic LLM Conversations in Software Development Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30756 - VeriTrace: Evolving Mental Models for Deep Research Agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.26081 - Towards Cognitive Process-Aware Proactive Writing Support Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30424 - Stride-k Subsampling: Train-Free Audio Token Reduction for Whisper Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30927 - Learning to Follow In-Context Watermark Instructions via Self-Distillation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29030 - PhysWave: Physics-Guided Latent Diffusion Models for Controllable Spatial Audio Generation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29549 - VLA-Trace: Diagnosing Vision-Language-Action Models through Representation and Behavior Tracing Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.30117 - From Question-First to Analyst-First: Domain-Expert Skills and Verified Knowledge Compilation for Proactive Enterprise Analytics Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28594 - Localizing Emergent Failures in Agentic AI: Recovering Minimal Repair Families via Counterfactual Replay Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29228 - Multi-Step Forecasting of Grape Berry Temperature based on LSTM Model with Feed-Forward Attention Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29008 - Self-Specialized Teachers for Domain Post-Training Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28647 - InternReviewer & InternAdvocate: Objective Reward and Evaluation for Agentic Reinforcement Learning in Peer Review and Rebuttal Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28612 - Pretrained, Curriculum-Tuned, and Ensembled: A Tracer-Aware Interactive Segmentation Pipeline for AutoPET V Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30844 - Beyond Ranking Accuracy: Evaluating LLM-Cited Feature Rationales for Next Basket Repurchase Recommendation Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30333 - TaskPress: Query-Agnostic KV Cache Compression via Task-Guided Pruning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.03276 - Agent-as-Peer-Debriefer: A Multi-Agent Framework with Perspective-Based Refinement for Qualitative Analysis Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.24600 - PortBench: A Correlation-Aware, Full-Pipeline Benchmark for LLM-Driven Portfolio Management Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.27887 - Generalizable Multi-Agent Planning from Signal Temporal Logic Specifications via Diffusion Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29490 - RailSyn: Diagnosis-Guided Image Generation for Traceable Data Completion in Railway Foreign Object Detection Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30709 - Discovering Machine Correlates of Consciousness Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28824 - Real-Time Deadlines Reveal Fragile Temporal Adaptation in LLM Strategic Dialogues Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.13206 - Auditing Anonymous AI Models: A Four-Stage Protocol for Black-Box Identity Verification Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31142 - AcrossWAM1.0:A Modular Latent World-Action Stack for Compact Robot Policies Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29937 - OntoAligner-Ensemble: Voting-Based Fusion across Heterogeneous Ontology Alignment Techniques Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.31137 - Useful Memories Become Faulty When Continuously Updated by LLMs Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2605.12978 - Designing an Auditable LLM-Supported Workflow for Qualitative Thematic Analysis Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30543 - SkillZip Pro: Execution-Aware Dynamic Compression of Progressively Loaded Skills for Self-Evolving Agents Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30785 - In LLM Reasoning, there is Irrationality on top of Value Misalignment Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.20624 - Reasoning or Rambling? Exploring the Effect of Thinking on Agent Persuasion Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2509.21054 - Measurement Validity in LLM Cultural Alignment Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29266 - A Large-scale Evaluation of Text-guided Models for Facial Editing Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28802 - RACER: Reinforced Agent Collaboration for Explainable Reasoning on Knowledge Graphs Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29263 - Spec2Twin-Chain: Orchestrating Bi-Level Optimization with LLMs for Blockchain Digital Twin Construction Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30050 - A Composition-Aware Pretraining Framework for Geospatial Foundation Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30817 - SimCRAFT: Distilling Remote Sensing Agents via Synthetic Trajectories and Contextual Retrieval-Augmented Fine-Tuning Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30277 - Applications of Risk Science to AI Fairness Evaluation: Principles, Challenges, and Best Practices Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29478 - Self-Evolving Skills via Surrogate-Guided Solve-and-Reproduce Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28638 - Peer Oversight in Collective Decision Making Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28754 - View-oriented Conversation Compiler for Agent Trace Analysis Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2603.29678 - An Explainable Coherence Score for Detecting Temporal Inconsistencies in Political News Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29175 - DiffPDE: Masked Diffusion Language Models as PDE Solver Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30532 - DERELAB: Probing Defeasible Reasoning and Confirmation Bias in LLMs with a Generative Benchmark Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.30413 - From AGI to ASI Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2606.12683 - Let Prompts Bridge Defense Knowledge: Transferable Graph Purification via Vulnerability-Aware GPL Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29054 - Evaluating the Hidden Costs of Personalization in Large Language Models Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.28833 - AgentLogs: A Dataset for Opening the Black Box of GitHub's Cloud Agent Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2608.29204 - ShardMemo: Scope-Before-Routing for Agentic Memory Retrieval Source: arxiv-cs-ai Topic: AI Research (+AI) URL: https://arxiv.org/abs/2601.21545