Research Paper Cockpit

Daily Digest - 2026-09-21

Papers first seen in this daily snapshot.

Daily Archives

Quick jump into generated daily digests.

Research Workflow

Latest digest: 2026-09-22.

Papers

57 visible entries

arxiv Score 25.8

ArenaFlow: From Trajectory Ranking to Hierarchical Credit Propagation for Open-Ended Agent RL

2026-09-18 · Qiang Zhang, Ruixue Ding, Fanrui Zhang, Xi Chen, Boli Chen, Shihang Wang, Yinfeng Huang, Yi Zheng, Pengjun Xie, Kaipeng Zhang, Jiawei Liu, Zheng-Jun Zha

General AI

Reinforcement learning has substantially improved large language model (LLM) agents in verifiable domains, but remains difficult to apply to open-ended agent tasks, where solutions are diverse and reliable scalar rewards are hard to obtain. Recent pairwise evaluation methods alleviate reward discrimination collapse by …

Review
pending
Role
unreviewed
Read
now
arxiv Score 21.8

AutoViewMem: Self-Configuring Orthogonal Views for Conversational Long-Term Memory

2026-09-18 · Zijie Cao, Xijun Qu, Zhicheng Gu, Xiaoshu Chen, Duanyang Yuan, Yanning Hou, Sihang Zhou, Jianxing Gong, Jian Huang, Yang Mei

Research Track A · General AI

Long-term memory is essential for large language model (LLM) agents to maintain consistency and personalization over extended interactions. Existing memory systems typically rely on fixed granularities or static schemas, but these designs struggle when heterogeneous information, such as preferences, events, constraints…

Review
pending
Role
unreviewed
Read
now
arxiv Score 20.5

EconSkills: Studying Skill Transfer and Retrieval for Web Agents on Live Economic Data

2026-09-17 · Yinzhu Quan, Zefang Liu

Research Track B · General AI

Web agents often revisit the same sites, yet most evaluations discard the procedures learned in earlier successful interactions. We introduce EconSkills, a skill library and evaluation framework that distills verified EconWebArena trajectories into parameterized standard operating procedures for retrieving live economi…

Review
pending
Role
unreviewed
Read
now
arxiv Score 19.8

AVT-Fabric: Active Visuo-Tactile Perception via Adaptive Evidence Selection for Efficient Robotic Fabric Comparison

2026-09-18 · Chang Gao, Zhuo Chen, Suhang Xia, Jihong Zhu, Jiankang Deng, Shan Luo

General AI

Robotic fabric comparison needs to actively combine visual appearance and tactile cues. Here, we present AVT-Fabric, an RGB-first framework that allocates tactile evidence according to the difficulty of each comparison. A dual-scale gate evaluates answer-token confidence and raw logit separation to determine whether an…

Review
pending
Role
unreviewed
Read
now
arxiv Score 18.8

Neuro-Symbolic Agentic AI for Networked Low-Altitude UAVs

2026-09-17 · Yuqi Ping, Tianhao Liang, Nanchi Su, Guangyu Lei, Junwei Wu, Qinyu Zhang, Tingting Zhang

Research Track A · General AI

Networked low-altitude unmanned aerial vehicles (UAVs) need reliable and adaptive decision-making capabilities to operate under uncertain observations, dynamic environments, and intermittent connectivity, while many existing agentic systems remain limited by hallucination risks, data dependence, and weak generalization…

Review
pending
Role
unreviewed
Read
now
arxiv Score 18.8

Think Thrice Before Reranking: Multi-perspective Evidence and Reasoning Integration for Text Reranking

2026-09-17 · Lijun Liu, Zhengzong Chen, Wenyan Li, Yuanyuan Zhao, Fei Huang

General AI

Reasoning-based reranking with Large Language Models (LLMs) has shown promising improvements in text ranking. However, current methods predominantly rely on a single reasoning trajectory, resulting in rankings that are susceptible to reasoning errors and inherently constrained in modeling the multifaceted signals under…

Review
pending
Role
unreviewed
Read
now
arxiv Score 18.8

An Interpretable Memory Decision Controller for LLM Agents Based on Three-Signal Complementarity: Decoupling Confidence and Consistency

2026-09-18 · Yiming Zhang, Jinghong Zhang, Haoran Zhao, Yiren Ma, Chunlei Zhao

General AI

Memory systems for large language models have focused predominantly on efficient retrieval, whereas the decision of whether retrieved memories should be trusted has received comparatively little attention. When the memory store contains conflicting positions, standard retrieval-augmented generation (RAG) blindly inject…

Review
pending
Role
unreviewed
Read
now
arxiv Score 18.8

PRIME: Perception Feedback with Situational Memory Embeddings in VLA Models

2026-09-18 · Erik Deinzer, Naya Baslan, Luca Paparusso, Narunas Vaskevicius, Peter Knott, Luigi Palmieri

General AI

Current Vision-Language-Action (VLA) models for autonomous driving operate primarily through feedforward inference across the perception--reasoning--planning hierarchy. While modern architectures maintain temporal recurrence within the perceptual module, early perception remains blind to downstream reasoning and naviga…

Review
pending
Role
unreviewed
Read
now
arxiv Score 18.8

TrialAtlas: Multi-Agent Research Organization for Clinical Trial Design and Optimization

2026-09-18 · Jiacheng Lin, Zifeng Wang, Zheng Chen, Erick Scott, Ziwei Yang, Fanyang Yu, Sheng Zhong, Jimeng Sun

General AI

Nearly 90% of drugs entering clinical development ultimately fail, despite billions of dollars in investment. Pharmaceutical companies therefore rely on clinical development planning (CDP) and probability of technical and regulatory success assessment to anticipate development risks, yet these decisions remain labor-in…

Review
pending
Role
unreviewed
Read
now
arxiv Score 18.0

An Architecture for Long-Horizon Agents: Levels, Ticks and Cascaded Intelligence

2026-09-17 · Erik Nijkamp, Anurag Koul, Egor Pakhomov, Bo Pang

Research Track A · General AI

Language-model agents are increasingly asked to carry out work spanning days or weeks, such as an operations remediation or a research programme. Such a task outlives any context window, any process and any interval at which a person can attend. In this paper, we argue that a long-horizon agent must run continually wit…

Review
pending
Role
unreviewed
Read
now
arxiv Score 17.0

CGaLore: Curvature-Guided GaLore for Memory-Efficient Continual Adaptation of ASR Foundation Models

2026-09-18 · Steven Vander Eeckt, Hugo Van hamme

Research Track A

Automatic speech recognition models suffer from catastrophic forgetting when adapted to new domains, accents, or downstream tasks. This problem becomes increasingly important with the growing use of speech foundation models, where adaptation should be both memory-efficient and safe, preserving the broad capabilities le…

Review
pending
Role
unreviewed
Read
now
arxiv Score 16.5

Benchmarking World Models for Continual Learning on Compositional Tasks

2026-09-18 · Haoyu Zhou, Joe Watson, Anson Lei, Ingmar Posner

Research Track A · General AI

A desirable property of a world model is the ability to learn continually across tasks, adapting to new environments without forgetting what the agent has already learnt. In particular, the ability to retain and reuse knowledge obtained from prior experiences underpins an agent's ability to efficiently adapt to novel e…

Review
pending
Role
unreviewed
Read
now
arxiv Score 15.8

Designer-RSI: Evolving Procedural Memory from User Traffic for Agentic Graphic Design

2026-09-18 · Hongyang Du, Lan Yan, Christian Flores, Asim Kadav

General AI

Professional graphic design is a long-horizon agentic task in which structured, editable artifacts emerge from many interdependent actions, yet outcomes admit no reliable programmatic oracle. We introduce a continual adaptation framework in which a frozen frontier model operates professional design software through mor…

Review
pending
Role
unreviewed
Read
now
arxiv Score 15.8

MintAct: A Unified Visual Agent for Digital Environments

2026-09-18 · Mingfei Gao, Rui Tian, Haiming Gang, Bohan Zhai, Le Zhang, Yuanzheng Gong, Di Feng, Ege Özsoy, Kaixin Ma, Vishwesh Kirthivasan, Oğuzhan Fatih Kar, Roman Bachmann, Anders Boesen Lindbo Larsen, Afshin Dehghan

Research Track B · General AI

We present MintAct, a family of vision-language models that unifies UI grounding, multi-step navigation across mobile, desktop, and web, and visual tool use, trained at 2B, 4B, and 8B scales. Through careful design of our environments, data, and training recipes, MintAct models match the performance of per-domain speci…

Review
pending
Role
unreviewed
Read
now
arxiv Score 15.8

Supporting Industrial Test-Failure Analysis with LLM-Based Systems: An Experience Report

2026-09-18 · Eric Jansson, Per Strandberg, Thomas Sörensen, Eduard Paul Enoiu, Wasif Afzal

General AI

This study examines tool-augmented Large Language Model (LLM) systems for supporting Root Cause Analysis (RCA) of nightly test failures at Westermo Network Technologies AB. Nightly test executions produce heterogeneous test data and logs that practitioners currently inspect manually across multiple sources. We implemen…

Review
pending
Role
unreviewed
Read
now
huggingface Score 15.0

OmniVChat: Synthesizing, Benchmarking, and Training for Native Audio-Visual Dialogue

2026-09-18 · Haolin He, Yunfei Chu, Qi Chen, Wen Huang, Yuan Feng, Muzhi Zhu, Zheqi Dai, Haoning Xu, Dongchao Yang, Chunyat Wu, Zining Liang, Zhengxi Liu, Xiquan Li, Xie Chen, Xize Cheng, Qize Yang, Jin Xu, Qiuqiang Kong

General AI

We define OmniVChat (Omni Video Chat) as the task of native audio-visual dialogue between a user and an omni model. In OmniVChat, omni models directly and simultaneously receive audio and video from a user and return text. The user's query is embedded in the audio and video, without a separate text question, external c…

Review
pending
Role
unreviewed
Read
now
arxiv Score 14.0

FAN: Foresight Action Normalization for Continual Adaptation of Vision-Language-Action Models

2026-09-18 · Yijun Hong, Jiarun Zhu, Xiaoquan Sun, Le Xu, Qijun He, Xin Jin, Mingqi Yuan, Wenjun Zeng, Jiayu Chen

Research Track A

Vision-Language-Action (VLA) models pre-trained on large-scale, closed datasets have demonstrated remarkable success across diverse robotic manipulation tasks. However, their long-term real-world deployment necessitates continuously acquiring new skills while retaining previously learned capabilities. While pioneering …

Review
pending
Role
unreviewed
Read
now
huggingface Score 14.0

GraphSkillEvo: Evolutionary Optimization of Graph-Structured Agent Skills

2026-09-18 · Rui Sun, Zhi Zheng, Zhenkun Wang, Zhichao Lu

General AI

Skills can improve the performance of Large Language Model (LLM) agents by providing task-specific procedural guidance, while skill optimization further improves their effectiveness through iterative refinement. However, existing skill optimization methods typically represent skills as unstructured natural-language ins…

Review
pending
Role
unreviewed
Read
now
arxiv Score 13.8

SkillIR: Evolving Scene-Aware Skills for Agentic Image Restoration

2026-09-18 · Jie Shao, Shengkai Hu, Xu Zhang, Beihang Song, Yongcheng Jing, Xu Wu, Jun Wan

General AI

This paper studies agentic image restoration, in which multimodal agents coordinate specialized restoration tools to recover images affected by complex degradations. Existing restoration agents often derive complete tool-use plans from the original degraded image or retrieve previously successful trajectories, providin…

Review
pending
Role
unreviewed
Read
now
arxiv Score 12.8

CodeMidas: Scaling Agentic Coding RL Environments from Code Itself

2026-09-18 · Bowen Ye, Lei Li, Shicheng Li, Zihao Yue, Linghao Zhang, Hanglong Lv, Yuanxin Liu, Wenhan Ma, Hao Tian, Rang Li, Jinhao Dong, Yikai Zhao, Xiangwei Deng, Hailin Zhang, Liang Zhao, Qi Liu, Lingpeng Kong, Tong Yang, Fuli Luo

General AI

Training capable coding agents via reinforcement learning (RL) requires diverse tasks with reliable verifiers. Open-source codebases offer a rich source of such tasks, while existing methods typically rely on development artifacts such as issues and commits, limiting the range of tasks that can be extracted. To better …

Review
pending
Role
unreviewed
Read
now
arxiv Score 12.8

NemotronLabs VoiceChat: An Open Full-duplex Speech-to-Speech Model with Tool Calling Capabilities

2026-09-18 · Jagadeesh Balam, Travis Bartley, Edresson Casanova, Sanjay Chauhan, Chen Chen, Zhehuai Chen, Zijia Chen, Francesco Ciannella, Slyne Deng, Mikyas Desta, Harishchandra Dubey, Slim Essid, Nourchene Ferchichi, Boris Ginsburg, Mariana Graterol Fuenmayor, Negar Habibi, Kevin Hu, Anand Joseph, Viraj Karandikar, Myungjong Kim, Viacheslav Klimkov, Seelan Lakshmi Narasimhan, Lily Lee, Jason Li, Eileen Long, Ameya Mahabaleshwarkar, Aditya Malte, Adi Margolin, Sasha Meister, Valentin Mendelev, Oluwatobi Olabiyi, Ankita Pasad, Yifan Peng, Elena Rastorgueva, Jayda Ritchie, Jason Roche, Nikhil Srihari, Yuanhang Su, Yoshi Suhara, Viet Anh Trinh, Jinhan Wang, Piotr Zelasko, Hui Wang, Puhui Meng, Chaosen Zhang, Yunsheng Liu, Shawn Wang, Wenjing Li, Zhonglei He

General AI

We introduce NemotronLabs VoiceChat, an open full-duplex speech-to-speech model with native tool-calling capabilities. NemotronLabs VoiceChat combines a streaming speech encoder and decoder-only language model with parallel specialized output streams for agent text and structured function calls, an auxiliary RNN-T bran…

Review
pending
Role
unreviewed
Read
now
arxiv Score 12.8

PSR: Predictive Sensorimotor Representation Learning for Contact-Rich Manipulation

2026-09-18 · Shengbao Li, Peng Xu, Chao Tang, Hao Wei, Jiaheng Wang, Hong Yin, Jiangtao Chen, Jinxuan Zhu, Zhong Zhou, Mengfan Wang, Tingguang Li

General AI

Contact-rich manipulation requires policies to generate precise actions by reasoning over contact forces, robot configurations, and interaction histories beyond visual observations. Existing methods passively condition on force feedback rather than actively predicting future contact dynamics, limiting their ability to …

Review
pending
Role
unreviewed
Read
now
arxiv Score 11.8

Predictable Failure in Multi-Hop Retrieval: Score-Distributional Confidence Scoring and Abstention

2026-09-18 · Andre Bacellar

General AI

Multi-hop retrieval failures are not uniformly distributed across queries: they cluster in structurally predictable subpopulations. We prove two results formalizing this structure. First (CWAR Reducibility): confident-failure reduction is achievable if and only if retrieval features carry mutual information about succe…

Review
pending
Role
unreviewed
Read
now
huggingface Score 11.4

EvoOntology: A Self-Evolving Ontology Layer for Data Agents

2026-09-14 · Meiduo Chong, Shaolei Zhang, Ju Fan, Xiaoyong Du

General AI

Data agents aim to fulfill natural-language instructions over heterogeneous data, including tables, files, and databases. However, data agents face a challenging agent-data gap: heterogeneous data resides outside the agent, while the agent can access it (e.g., column names and file paths) only through generic tools. Ex…

Review
pending
Role
unreviewed
Read
now
huggingface Score 11.0

Grounded Skill Synthesis from Code at Scale for Agentic Intelligence

2026-09-04 · Yongqi Tong, Pan Wang, Hang Wang, Jianshe Li, Xin Zhang, Jiang-Ming Yang, Wei Wu

General AI

Reusable skills give agents transferable procedural knowledge, making scalable acquisition essential for extending agents beyond prior experience. Existing methods face two limitations: trajectory-based synthesis requires interactions with specific environments, while document-derived skills may lack executable evidenc…

Review
pending
Role
unreviewed
Read
soon
arxiv Score 10.8

APort Vault: Benchmarking AI Agent Payment Authorization with the Open Agent Passport

2026-09-18 · Uchi Uchibeke

General AI

APort Vault is a benchmark for payment authorization in tool-using AI agents. It replays 4,371 attacks written by humans against a live payment agent during a public capture-the-flag event, across 14 models from 8 labs, five policy configurations and two replay tracks, with and without a deterministic pre-action check …

Review
pending
Role
unreviewed
Read
soon
arxiv Score 10.8

Detecting Pretraining Data in Large Language Models from a Free-Energy Perspective

2026-09-18 · Chenye Ke, Zirui Liu, Qi Liu, Yan Zhuang, Jintao Zhang, Zhenya Huang, Shijin Wang

General AI

Detecting pretraining data in large language models is challenging because high likelihood can reflect either training exposure or strong generalization. In the joint space of prediction loss and predictive entropy, a likelihood-only detector uses a horizontal boundary and can mistake predictable non-members for member…

Review
pending
Role
unreviewed
Read
soon
huggingface Score 10.0

MLLMs Hallucinate when Information Distribution Drifts in Synergy Heads

2026-09-05 · Meng'en Qin, Junye Chen, Jucheng Liu, Youlu Xing, Song Wang, Ruize Han

General AI

Multimodal Large Language Models (MLLMs) often struggle with hallucinations, thus hindering their reliable practical applications. Existing attention-based mitigation methods mainly rely on indirect signals (e.g., attention weights) that fail to accurately reflect the actual information shift underlying hallucination g…

Review
pending
Role
unreviewed
Read
soon
huggingface Score 10.0

FRAUDSkill: Structured Frozen-Weight Skill Optimization for Audio Anti-Fraud Detection

2026-09-16 · Chengxian Hu, Zhiming Ma, Mingjun Pan, Yifan Wang, Shun Zhang, Qifan Wang, Zhilei Zhao, Yijin Zhou, Yuxi Zhao, Huiyuan Liu, Peidong Wang, Peng Chen

General AI

Large audio-language models have shown promise for anti-fraud detection by directly processing speech and reasoning over fraud-related evidence. Their deployment, however, requires predictions to follow a predefined label space and a structured decision protocol consisting of service-scenario identification, fraud dete…

Review
pending
Role
unreviewed
Read
soon
huggingface Score 10.0

Geometry of Values: Task Vector Composition for Ethical Preference Alignment in Language Models

2026-09-17 · Utkarsh Agarwal, Monojit Choudhury

General AI

Large Language Models (LLMs) are increasingly deployed in applications that must weigh clashing moral values, yet even strong models exhibit hidden biases and brittle instruction-following across languages. We introduce a 12,000-instance dataset of two-option dilemmas covering pairwise three value conflicts: Honesty vs…

Review
pending
Role
unreviewed
Read
soon
huggingface Score 10.0

Paint-Anything: Unified Any-Color Control for Image Generation and Editing

2026-09-17 · Ji Xie, Dewei Zhou, Xinyu Huang, Zhennan Chen, Xun Wang

General AI

Professional design requires any-color control: the ability to specify an object's target color with any 24-bit hex value for image generation and editing. Prior work has explored color generation, editing, and colorization, but often relies on dedicated color representations or specialized inference procedures. Advanc…

Review
pending
Role
unreviewed
Read
soon
arxiv Score 10.0

Position Paper: Neurotransmitters as a Missing Dimension in Artificial Neural Networks

2026-09-17 · Yupei Li, Manuel Milling, Berrak Sisman, Björn Schuller

Research Track A

Artificial neural networks (ANNs), as core components of modern deep learning (DL) systems, lack the adaptive flexibility and long-term stability exhibited by biological systems. This limitation largely stems from the fact that conventional ANNs rely on uniform, local, and gradient-based parameter updates, while neglec…

Review
pending
Role
unreviewed
Read
soon
huggingface Score 10.0

Calibrating Teacher--Student Discrepancy for On-Policy Distillation

2026-09-18 · Qiangqiang He, Jin Li, MingCai Chen

General AI

On-policy distillation (OPD) improves reasoning models by learning the token-level discrepancy between a stronger teacher and an on-policy student. However, this discrepancy does not purely reflect the capability gap between the teacher and the student: it also contains deviations arising from the teacher itself, which…

Review
pending
Role
unreviewed
Read
soon
arxiv Score 10.0

Learning-Based Augmentation and Adaptation for Grid Sim-to-Real Model Discrepancy

2026-09-18 · Sayak Mukherjee, Kyung-Bin Kwon, Ramij R. Hossain, Marcelo Elizondo

Research Track A

Modern power systems can encounter increased discrepancy between the operators' simulation model and the actual true dynamics of the grid, driven by uncertainties caused by integration of new inverter-based resources (IBRs), large loads, unmodeled dynamics, parameter drifts, etc., to name a few. All of these impact the…

Review
pending
Role
unreviewed
Read
soon
arxiv Score 9.8

MIST: Multimodal Survival Prediction with Genomic-Guided Histology Attention

2026-09-18 · Muhammet Sami Yavuz, Sabri Mustafa Kahya, Richard R. Chen, Jana Lipkova, Benedikt Wiestler

General AI

Multimodal survival models can combine complementary prognostic information from whole-slide images and genomic profiles, but effective fusion remains challenging amid external cohort shift and computational complexity. To address these challenges, we propose MIST, multimodal survival prediction with genomic-guided his…

Review
pending
Role
unreviewed
Read
soon
huggingface Score 9.4

MoME: Mixture-of-Memory Embeddings for Context-Aware Sparse Lookup

2026-09-14 · Muchen Li, Leonid Sigal, Renjie Liao

General AI

Scaling large language models efficiently has motivated sparse capacity mechanisms such as Mixture-of-Experts and, more recently, conditional memory: token-indexed embedding tables that augment the backbone with cheap parametric lookups. Existing memory-embedding methods retrieve via a deterministic function of the sur…

Review
pending
Role
unreviewed
Read
soon
arxiv Score 9.0

A variational model of nonlinear poroelasticity

2026-09-18 · James H. Adler, Xiaozhe Hu, Arkadz Kirshtein

Research Track A

We derive a thermodynamically-consistent model of fluid flow through a poroelastic medium. Starting from elastic and fluid free-energy densities, an energy-dissipation rate, and a kinematic constraint, the force-balance equations are derived using variational principles, with the pressure--density constitutive relation…

Review
pending
Role
unreviewed
Read
soon
arxiv Score 9.0

Gripper-Aware Automatic Dense Packing of Irregular Objects

2026-09-18 · Tianhao Qin, Connor McCann, Berk Calli, Jing Xiao

Research Track A · General AI

Automatic dense packing is widely desired in warehouse operations but remains a fundamental challenge in robotic manipulation. Existing work on irregular-object packing largely targets simulation with idealized contact, treating the object as an isolated rigid body. The gripper often enters as a discrete, post-hoc feas…

Review
pending
Role
unreviewed
Read
soon
arxiv Score 8.8

A Lie Detector Test for Language Models: Reading Knowledge a Model Won't Reveal

2026-09-18 · Hiskias Dingeto

General AI

Large language models can hold knowledge they do not report. A model may sandbag on a capability evaluation, or answer against what it internally knows, and its outputs alone cannot tell whether it is hiding an answer or simply does not have one. We borrow the Concealed Information Test, a forensic method that identifi…

Review
pending
Role
unreviewed
Read
soon
arxiv Score 8.8

A Sim-to-Real Integration Pipeline for Training and Deployment of Chunk-Based VLA Manipulation Policies

2026-09-18 · Mathilde Kappel, Clémence Grislain, Mohamed Chetouani, Olivier Sigaud, Louis Annabi, Faïz Ben Amar, Stéphane Doncieux, Mahdi Khoramshahi

General AI

Vision-Language-Action (VLA) models have become a prominent paradigm for mapping multimodal inputs, including semantic instructions, visual observations of the scene, and proprioceptive observations, to robot actions. Most state-of-the-art models predict actions in the end-effector pose space as sequences of action chu…

Review
pending
Role
unreviewed
Read
soon
arxiv Score 8.8

Beyond Reactive Assistance: PV-Care Using Low-Density EEG and AI to Provide Proactive, Context-Aware Help for MCI

2026-09-18 · Simon L Liu, Manish Kumar Krishne Gowda

General AI

The growing elderly population gives rise to an urgent need for intelligent support systems, particularly for individuals with Mild Cognitive Impairment (MCI). This paper presents PV-Care, a proactive AI-driven assistance scheme that integrates wearable electroencephalogram (EEG) sensing with visual environmental perce…

Review
pending
Role
unreviewed
Read
soon
arxiv Score 8.8

ECG Mirage: Revealing and Mitigating the Underutilisation of ECGs in Vision-Language Models for Clinical Prediction

2026-09-18 · Jinning Liang, Mingcheng Zhu, Tingting Zhu

General AI

Emergency department (ED) decision-making relies on heterogeneous clinical information, including patient history, vital signs, laboratory results, and electrocardiograms (ECGs). Vision--language models (VLMs) can jointly process these modalities, but strong predictive performance does not necessarily imply meaningful …

Review
pending
Role
unreviewed
Read
soon
arxiv Score 8.8

Touvigation: Embodied Adaptive Object Acquisition for Blind and Low-Vision Users in Unfamiliar Indoor Environments

2026-09-18 · George Xi Wang, Xiangyu Li, Shaoyue Wen, Jiaqian Hu, Junan Xie, Yupeng Wang, Ziyue Shi, Qijun Chen, Maaike Bouwmeester, Yuhua Jin, Jing Qian

General AI

Blind and low-vision users often face challenges when locating and physically acquiring objects in unfamiliar indoor environments. Existing vision-language-model-based assistants can provide semantic descriptions but may introduce latency, hallucinations, and guidance that is poorly aligned with embodied action. We pre…

Review
pending
Role
unreviewed
Read
soon
arxiv Score 8.8

Value-Sensitive Delegation in Everyday AI Agent Use: Evidence from OpenClaw

2026-09-18 · Renkai Ma, Ruyuan Wan, Xuan Lu, Fan Yang, Chen Chen, Lingyao Li

General AI

Users increasingly delegate work to autonomous AI agents, yet evaluations typically measure task completion rather than the values users prioritize. Using Value Sensitive Design, we analyzed, with LLM assistance, 73,093 first-person Reddit posts about using OpenClaw, each for its human value, agent aspect, value fulfil…

Review
pending
Role
unreviewed
Read
soon
arxiv Score 8.8

XCalib Depth-Guided Geometric Optimization for Dense Thermal-Visible Video Registration

2026-09-18 · Aurelien Godet, Gabriel Jobert, Mauro Dalla Mura

General AI

Image registration is a vital preprocessing step in multimodal perception tasks, including image fusion, object detection, and semantic segmentation. In Advanced Driver- Assistance Systems (ADAS), spatial misalignment between visible (RGB) and infrared (IR) cameras -caused by non-coincident optical axes and field-of-vi…

Review
pending
Role
unreviewed
Read
soon
huggingface Score 8.0

Refinement Is Inherently Editable: Training-Free Prompt-to-Prompt Image Editing with Generative Refinement Network

2026-09-17 · Yulong Chen, Ziqian Zhang, Haoyu Zhang, Ao He, Senmao Li, Kai Wang

General AI

Text-guided image editing must introduce the requested changes while preserving unrelated source content. Diffusion-based editors rely on spatial controls whose inaccuracies can leave edits incomplete or alter unrelated regions. Causal autoregressive editors face a further constraint: their fixed decoding order limits …

Review
pending
Role
unreviewed
Read
soon
arxiv Score 7.8

BrainWideBench: Benchmarking large-scale pretraining and across-animal transfer in multi-region neural recordings

2026-09-18 · Alexandre Andre, Shivashriganesh P. Mahato, Vinam Arora, Keshav Balaji, Divyansha Lachi, Nanda H. Krishna, Jingyun Xiao, Yizi Zhang, Ximeng Mao, Wenrui Ma, Han Yu, International Brain Laboratory, Daniel Birman, Niccolò Bonacchi, Gaelle A. Chapuis, Joana A. Catarino, Felicia Davatolhagh, Mayo Faulkner, Laura Freitas-Silva, Fei Hu, Julia M. Huntenburg, Anup Khanal, Inês Laranjeira, Petrina Lau, Guido T. Meijer, Nathaniel J. Miska, Jean-Paul Noel, Alejandro Pan-Vazquez, Georg Raiser, Cyrille Rossant, Karolina Z. Socha, Anne E. Urai, Miles J. Wells, Steven J. West, Olivier Winter, Blake Richards, Guillaume Lajoie, Cole Hurwitz, Mehdi Azabou, Matthew R. Whiteway, Liam Paninski, Eva L. Dyer

General AI

Advances in large-scale neural recording have made it possible to collect data across many animals and distributed brain regions, raising the question of whether this scale can be exploited to learn general-purpose neural representations transferable across diverse downstream tasks. Yet, progress toward this goal has b…

Review
pending
Role
unreviewed
Read
soon
arxiv Score 7.8

OmniVBench: A Benchmark and Large-Scale Dataset for Omni Reference-to-Video Generation

2026-09-18 · Wenxue Li, Peiyan Guan, Haoyang Jiang, Junxian Cai, Hualuo Liu, Chunjie Zhang, Chong Guan, Songlian Li, Taiyi Wu, Yongjian Yu, Xiaotong Zhao, Alan Zhao, Eric Liu, Xi Chen, Yu Liu, Lei Zhu

General AI

Reference-to-video (R2V) generation is evolving toward increasingly general and versatile reference control, giving rise to the emerging paradigm of omni R2V generation. However, existing benchmarks fall short of these emerging capabilities: their test cases cover limited reference types and compositions, and their eva…

Review
pending
Role
unreviewed
Read
soon
arxiv Score 7.8

Particle Competition and Cooperation for Robust Graph Convolutional Network Learning Under Label Noise

2026-09-18 · Fabricio Breve

General AI

Graph Convolutional Networks (GCNs) are highly sensitive to label noise, since corrupted supervision can propagate through the graph and degrade learned node representations. This work proposes PCC+GCN, a hybrid framework that uses Particle Competition and Cooperation (PCC) as a graph-based label-refinement stage befor…

Review
pending
Role
unreviewed
Read
soon
huggingface Score 7.0

DeformSmith: Physics Harness-Guided Hierarchical Generation of Deformable Assets for Robot Manipulation

2026-09-17 · Can Li, Jie Gu, Zishun Deng, Jingmin Chen, Lei Sun

General AI

Creating deformable assets for robot manipulation requires jointly specifying their geometry, appearance, and physical properties. This is especially challenging for deformable objects, since text and images provide limited evidence about how they deform and respond to contact, yet these responses directly affect their…

Review
pending
Role
unreviewed
Read
later
arxiv Score 6.8

LLM-Generated Feature Pools for Time Series Anomaly Detection

2026-09-18 · Youssef Attia El Hili, Malik Tiomoko, Corinne Ancourt

General AI

We study how far a simple statistical pipeline can go on univariate time series anomaly detection under a strict selection protocol. The method extracts a small pool of statistics over sliding windows, scores each window with a transductive robust (MAD) model, and selects a feature subset per domain on a held-out tunin…

Review
pending
Role
unreviewed
Read
later
arxiv Score 6.8

Watermarkable Multi-Draft Speculative Sampling via Poisson Processes

2026-09-18 · Yanxiao Liu, Sicheng Wan, Zhan Gao, Deniz Gündüz

General AI

Large language models (LLMs) have achieved state-of-the-art performance across a wide range of tasks, motivating two important aspects of deployment: inference efficiency and output provenance, which can be tackled by speculative sampling and watermarking, respectively. However, recent works have shown that combining t…

Review
pending
Role
unreviewed
Read
later
huggingface Score 6.0

IntBMoE: Integrating Block-Level Conditioning into Expert Composition for Full-Participation Mixture-of-Experts

2026-09-18 · Ran Cheng, Longfei Xu, Zheng Liu, Kaikui Liu, Xiangxiang Chu

General AI

Mixture-of-Experts (MoE) scales capacity, but existing designs cannot set three quantities independently. For a single token, participation is how many experts contribute knowledge to its output, execution is how many are actually computed (compute cost), and materialization is how many expert-sized parameter sets must…

Review
pending
Role
unreviewed
Read
later
arxiv Score 5.8

Auditing bipartite motif interpretations: a worked example with conservation checks and open-path decomposition

2026-09-18 · Tengfei Shao

General AI

Motif profiles of bipartite agent-object networks, such as tourist-site visits and customer-item transactions, are read as evidence about structural roles and about differences between networks, often without asking what the two degree sequences already fix. In a simple bipartite graph the induced k-fan count on one no…

Review
pending
Role
unreviewed
Read
later
arxiv Score 5.8

COMPLEX: A Closed-Form Certified Embedding of Multiparameter Persistence Modules

2026-09-18 · Sushovan Majhi, Atish Mitra, Žiga Virk, Pramita Bagchi

General AI

Every multiparameter persistence vectorization we know of carries a one-sided Lipschitz upper bound and nothing below it: without a lower gauge there is no sense in which the features are faithful, and no per-prediction guarantee can be built on them. This paper supplies the missing side. COMPLEX is a closed-form, trai…

Review
pending
Role
unreviewed
Read
later
arxiv Score 5.8

Can I Trust My Body? A Three-Year Autoethnography of ChatGPT's Place in My Support System for Panic Attacks

2026-09-18 · Dongyijie Primo Pan, Pan Hui, Mirjana Prpa

General AI

People increasingly seek mental health support from large language models, yet little is known about their use across years of recurrent panic. We present a three-year analytic autoethnography of the first author's ChatGPT use while living with panic disorder, drawing on conversations, personal records, and accounts fr…

Review
pending
Role
unreviewed
Read
later
arxiv Score 5.8

Provisional Reachability: Containing Agents by Making Every Crossing Revocable

2026-09-18 · Yoshiaki Takashita

General AI

A companion paper found that what a defender must block over time has units: bits per period [Takashita, 2026a]. This paper sets it. Hold every crossing in escrow for one period, audit each held item independently with probability r, and revoke the window if any audit catches something. An adversary crossing k times, e…

Review
pending
Role
unreviewed
Read
later