WeSearch
Hub / ai-research / arXiv cs.AI
ai-research · source

arXiv cs.AI on WeSearch

Recent ai-research headlines from arXiv cs.AI.

REFORGE: A Method for Benchmarking LLMs' Reverse Engineering Capabilities in Decompiled Binary Function Naming
arXiv cs.AI

REFORGE: A Method for Benchmarking LLMs' Reverse Engineering Capabilities in Decompiled Binary Function Naming

Computer Science > Software Engineering arXiv:2607.07738 (cs) [Submitted on 7 Jul 2026] Title:REFORGE: A Method for…

7/13/2026 · 3 min read · 25 views
A Unified Approach to Interpreting Knowledge Distillation for Large Language Models via Interactions
arXiv cs.AI

A Unified Approach to Interpreting Knowledge Distillation for Large Language Models via Interactions

In this paper, we propose a unified approach to explore the common mechanism of various KD methods using interactions.…

7/13/2026 · 3 min read · 23 views
iLENS: Interpretable LLM-Guided Mixture-of-Experts for Neuroimaging Survival Analysis
arXiv cs.AI

iLENS: Interpretable LLM-Guided Mixture-of-Experts for Neuroimaging Survival Analysis

Predicting AD conversion during the prodromal stage remains critical for disease understanding and patient care. As…

7/13/2026 · 2 min read · 23 views
Signed Symmetric Quantization for Few-Bit Integers
arXiv cs.AI

Signed Symmetric Quantization for Few-Bit Integers

Computer Science > Machine Learning arXiv:2607.08779 (cs) [Submitted on 12 Jun 2026] Title:Signed Symmetric…

7/13/2026 · 3 min read · 21 views
Sticky Routing: Training MoE Models for Memory-Efficient Inference
arXiv cs.AI

Sticky Routing: Training MoE Models for Memory-Efficient Inference

Existing remedies are either system-level (caching heuristics) or post-hoc (router fine-tuning), leaving the root…

7/13/2026 · 2 min read · 24 views
Reward Transport: Property Control in Flow Matching via Noise-Space Alignment
arXiv cs.AI

Reward Transport: Property Control in Flow Matching via Noise-Space Alignment

We show that this coupling can instead serve as an alignment interface: by matching noise and data according to a…

7/13/2026 · 3 min read · 24 views
Director: Accelerating Distributed MoE Serving via Online Proactive Expert Placement
arXiv cs.AI

Director: Accelerating Distributed MoE Serving via Online Proactive Expert Placement

Its efficiency depends on the communication and computation latencies of the GPUs, which are linked to the placement…

7/13/2026 · 3 min read · 23 views
LieBN: Batch Normalization over Lie Groups
arXiv cs.AI

LieBN: Batch Normalization over Lie Groups

Recent advances have extended Deep Neural Networks (DNNs) to operate on manifolds, accompanied by normalization…

7/13/2026 · 3 min read · 22 views
HERO: A Heterogeneity-Aware Benchmark Library for Federated Continual Learning
arXiv cs.AI

HERO: A Heterogeneity-Aware Benchmark Library for Federated Continual Learning

Computer Science > Machine Learning arXiv:2607.08784 (cs) [Submitted on 13 Jun 2026] Title:HERO: A Heterogeneity-Aware…

7/13/2026 · 3 min read · 25 views
DaDaDa: A Dataset for Data Pricing in Data Marketplaces
arXiv cs.AI

DaDaDa: A Dataset for Data Pricing in Data Marketplaces

Recognizing the value of data, data transactions are increasingly common, giving rise to many data marketplaces, e.g.,…

7/13/2026 · 3 min read · 24 views
Accelerating GPU Inference of Large Language Models with Moderately Unstructured Sparse Weight Matrices
arXiv cs.AI

Accelerating GPU Inference of Large Language Models with Moderately Unstructured Sparse Weight Matrices

Pruning techniques that introduce sparsity into weight matrices can accelerate inference. However, maintaining model…

7/13/2026 · 3 min read · 25 views
LLM-Driven Evolutionary Generation of Multi-Objective Bayesian Optimization Algorithms
arXiv cs.AI

LLM-Driven Evolutionary Generation of Multi-Objective Bayesian Optimization Algorithms

We extend the LLaMEA framework to MOBO, using large language models as mutation and crossover operators within…

7/13/2026 · 3 min read · 22 views
EHR-MPC: Inference-Time Control for Sepsis Treatment with Generative Patient Digital Twins
arXiv cs.AI

EHR-MPC: Inference-Time Control for Sepsis Treatment with Generative Patient Digital Twins

Existing reinforcement learning (RL) approaches learn fixed strategies for sepsis treatment, limiting adaptability to…

7/13/2026 · 3 min read · 20 views
Multi-Conditioned Diffusion Synthesis of Sand Boils for Low-Resource Earthen-Levee Inspection
arXiv cs.AI

Multi-Conditioned Diffusion Synthesis of Sand Boils for Low-Resource Earthen-Levee Inspection

We present a diffusion-based synthesis pipeline for low-resource sand-boil imagery. Using Stable Diffusion XL…

7/13/2026 · 3 min read · 23 views
TheBioCollection: Unified Pre-Training Scale LLM Corpus for Biology
arXiv cs.AI

TheBioCollection: Unified Pre-Training Scale LLM Corpus for Biology

However, existing biological resources, such as molecular databases, protein repositories, genomic annotations,…

7/13/2026 · 3 min read · 22 views
Prompt-Driven Exploration
arXiv cs.AI

Prompt-Driven Exploration

Standard methods inject stochasticity in the action space, but such jitter only yields rollouts close to the original.…

7/13/2026 · 3 min read · 54 views
A Novel Parallel QCNN Architecture with Efficient Classical Simulability
arXiv cs.AI

A Novel Parallel QCNN Architecture with Efficient Classical Simulability

Using a novel architecture inspired by previous QCNN and classical convolutional neural network (CNN) implementations,…

7/13/2026 · 3 min read · 20 views
Eluna: An Agentic LLM System for Automating Warehouse Operations with Reasoning and Task Execution
arXiv cs.AI

Eluna: An Agentic LLM System for Automating Warehouse Operations with Reasoning and Task Execution

We present Eluna, a production-deployed agentic system for reliable SOP execution. Eluna is a graph-guided,…

7/13/2026 · 3 min read · 24 views
NL-PAC: Specification Ambiguity and Certified Minimax Risk Floors in LLM-Mediated Supervision
arXiv cs.AI

NL-PAC: Specification Ambiguity and Certified Minimax Risk Floors in LLM-Mediated Supervision

When a specification admits multiple readings but the supervision channel does not reveal which is operative,…

7/13/2026 · 3 min read · 23 views
MultiView-Bench: A Diagnostic Benchmark for World-Centric Multi-View Integration in VLMs
arXiv cs.AI

MultiView-Bench: A Diagnostic Benchmark for World-Centric Multi-View Integration in VLMs

We introduce MultiView-Bench, a diagnostic benchmark expressly designed to evaluate multi-view integration for…

7/13/2026 · 3 min read · 20 views
CLAP: Direct VLM-to-VLA Adaptation via Language-Action Grounding
arXiv cs.AI

CLAP: Direct VLM-to-VLA Adaptation via Language-Action Grounding

Computer Science > Robotics arXiv:2607.08974 (cs) [Submitted on 9 Jul 2026] Title:CLAP: Direct VLM-to-VLA Adaptation…

7/13/2026 · 3 min read · 24 views
The Patchwork Problem in LLM-Generated Code
arXiv cs.AI

The Patchwork Problem in LLM-Generated Code

Computer Science > Software Engineering arXiv:2607.08981 (cs) [Submitted on 9 Jul 2026] Title:The Patchwork Problem in…

7/13/2026 · 3 min read · 23 views
SCATE: Learning to Supervise Coding Agents for Cost-Effective Test Generation
arXiv cs.AI

SCATE: Learning to Supervise Coding Agents for Cost-Effective Test Generation

Currently, mitigating this premature termination requires continuous human-in-the-loop supervision. This heavy…

7/13/2026 · 3 min read · 26 views
AlphaZero in Sparsely Rewarded Games: Limits and Auxiliary Supervision
arXiv cs.AI

AlphaZero in Sparsely Rewarded Games: Limits and Auxiliary Supervision

We study this gap in two oracle-evaluable domains with contrasting structure: Connect Four, a solved partisan game…

7/13/2026 · 3 min read · 27 views
Model Agnostic Graph Prompt Learning for Crystal Property Prediction
arXiv cs.AI

Model Agnostic Graph Prompt Learning for Crystal Property Prediction

These models often encode domain-specific knowledge into their graph encoding modules, which increases their parameter…

7/13/2026 · 3 min read · 27 views
Correlation-Aware Contextual Bandits with Surrogate Rewards for LLM Routing
arXiv cs.AI

Correlation-Aware Contextual Bandits with Surrogate Rewards for LLM Routing

Unlike classical contextual bandits that rely solely on bandit feedback and assume conditional independence across…

7/13/2026 · 3 min read · 27 views
Phone Segmentation and Recognition through Phonological Activation Mapping
arXiv cs.AI

Phone Segmentation and Recognition through Phonological Activation Mapping

Mortensen View a PDF of the paper titled Phone Segmentation and Recognition through Phonological Activation Mapping,…

7/13/2026 · 2 min read · 25 views
Video Generation Models are General-Purpose Vision Learners
arXiv cs.AI

Video Generation Models are General-Purpose Vision Learners

What, then, is the equivalent catalyst needed to achieve a general-purpose model in computer vision? In this paper, we…

7/13/2026 · 3 min read · 25 views
Evolutionary Intelligence for Scientific Discovery: From Evolutionary Computation to Cumulative Discovery Systems
arXiv cs.AI

Evolutionary Intelligence for Scientific Discovery: From Evolutionary Computation to Cumulative Discovery Systems

Evolutionary computation (EC) provides a computational basis for feedback-driven discovery because population-based…

7/13/2026 · 3 min read · 25 views
Quantum Logic as the Logic of Contexts
arXiv cs.AI

Quantum Logic as the Logic of Contexts

We argue for the opposite order of explanation in a finite and fully computable setting. The free orthomodular lattice…

7/13/2026 · 3 min read · 27 views
On Locality and Length Generalization in Visual Reasoning
arXiv cs.AI

On Locality and Length Generalization in Visual Reasoning

This makes human vision distinctly different from most popular computer vision models in use today, which input images…

7/13/2026 · 3 min read · 23 views
Inside the Skill Market: From Software Engineering Activities to Reusable Agent Skills
arXiv cs.AI

Inside the Skill Market: From Software Engineering Activities to Reusable Agent Skills

SE) has continuously evolved through increasingly powerful forms of reuse, from source code and libraries to…

7/13/2026 · 3 min read · 23 views
OmniMapBench: Benchmarking Visual-Centric Reasoning on Diverse Map Documents
arXiv cs.AI

OmniMapBench: Benchmarking Visual-Centric Reasoning on Diverse Map Documents

A critical limitation is identified in many document understanding benchmarks: visual content is often reducible to…

7/13/2026 · 3 min read · 29 views
PRecG: Legal Precedent Retrieval with Graph Neural Networks and Rhetorical Role Segmentation
arXiv cs.AI

PRecG: Legal Precedent Retrieval with Graph Neural Networks and Rhetorical Role Segmentation

Current approaches for automatic precedent retrieval map legal documents to a low-dimensional semantic space and…

7/13/2026 · 3 min read · 30 views
A Coreset Selection Framework with Ensemble Aggregation for Image Classification
arXiv cs.AI

A Coreset Selection Framework with Ensemble Aggregation for Image Classification

Selecting representative training subsets, however, remains challenging: individual sample contributions are unclear,…

7/13/2026 · 3 min read · 31 views
Beyond Metadata: CAPRA for Hidden Subgroup Analysis under Missing Metadata in Medical Imaging
arXiv cs.AI

Beyond Metadata: CAPRA for Hidden Subgroup Analysis under Missing Metadata in Medical Imaging

Once those metadata disappear, clinically critical failure modes can be masked by strong aggregate performance, and…

7/13/2026 · 3 min read · 27 views
Integrating Large Language Models and Graph Convolutional Networks for Semi-Supervised Image Classification
arXiv cs.AI

Integrating Large Language Models and Graph Convolutional Networks for Semi-Supervised Image Classification

Therefore, semi-supervised approaches such as Graph Convolutional Networks (GCNs), which learn from both labeled and…

7/13/2026 · 3 min read · 31 views
Event Stream based Multi-Modal Video Anomaly Detection: A Benchmark Dataset and Algorithms
arXiv cs.AI

Event Stream based Multi-Modal Video Anomaly Detection: A Benchmark Dataset and Algorithms

To address these limitations, we propose EVAD, an event enhanced VAD framework that jointly exploits conventional…

7/13/2026 · 3 min read · 28 views
Augmenting Fundamental Analysis with Large Language Models: A RAG-Based System for Generating Investor Briefs
arXiv cs.AI

Augmenting Fundamental Analysis with Large Language Models: A RAG-Based System for Generating Investor Briefs

Securities and Exchange Commission (SEC) which can be found in EDGAR. We were preprocessing those data and than…

7/13/2026 · 2 min read · 30 views
NormAct: A Benchmark for Hidden Social Norm Compliance in Embodied Planning
arXiv.org

NormAct: A Benchmark for Hidden Social Norm Compliance in Embodied Planning

While explicit goals may render certain actions optimal, implicit social norms often impose hidden constraints.…

6/29/2026 · 3 min read · 45 views
ATOD: Annealed Turn-aware On-policy Distillation for Multi-turn Autonomous Agents
arXiv.org

ATOD: Annealed Turn-aware On-policy Distillation for Multi-turn Autonomous Agents

On-policy distillation (OPD) provides dense teacher guidance and typically improves rapidly in the early stage, but…

6/29/2026 · 3 min read · 37 views
Internalizing the Future: A Unified Agentic Training Paradigm for World Model Planning
arXiv.org

Internalizing the Future: A Unified Agentic Training Paradigm for World Model Planning

The paper introduces a unified training paradigm that equips large language model agents with internal world modeling…

6/29/2026 · 3 min read · 40 views
Towards Reliable and Robust LLM Planning: Symbolic Feedback-Driven Iterative Self-Refinement Framework
arXiv.org

Towards Reliable and Robust LLM Planning: Symbolic Feedback-Driven Iterative Self-Refinement Framework

Planning, a core component of intelligent behavior, remains challenging for LLMs, which often produce infeasible or…

6/29/2026 · 3 min read · 43 views
ToE: A Hierarchical and Explainable Claim Verification Framework with Dynamic Multi-source Evidence Retrieval and Aggregation
arXiv.org

ToE: A Hierarchical and Explainable Claim Verification Framework with Dynamic Multi-source Evidence Retrieval and Aggregation

The paper presents Tree of Evidence (ToE), a hierarchical framework for automated claim verification that builds…

6/29/2026 · 3 min read · 39 views
MER-R1: Multimodal Emotion Reasoning via Slow-Fast Thinking Synergy
arXiv.org

MER-R1: Multimodal Emotion Reasoning via Slow-Fast Thinking Synergy

Specifically, for reasoning-based MLLMs, fast thinking by triggering direct answers often outperforms slow thinking…

6/29/2026 · 3 min read · 44 views
DysLexLens: A Low-Resource LLM Framework for Analysing Dyslexic Learners Insights from Online Forums
arXiv.org

DysLexLens: A Low-Resource LLM Framework for Analysing Dyslexic Learners Insights from Online Forums

However, their lived experiences with these tools remain largely underexamined. This paper proposes DysLexLens, a…

6/29/2026 · 3 min read · 34 views
AI-Model Network: Concept, Current State and Future
arXiv.org

AI-Model Network: Concept, Current State and Future

Computers create the Internet, and the Internet empowers the value of computers. The rapid development of the…

6/29/2026 · 3 min read · 31 views
When Does Personality Composition Matter for Multi-Agent LLM Teams?
arXiv.org

When Does Personality Composition Matter for Multi-Agent LLM Teams?

Computer Science > Artificial Intelligence arXiv:2606.27443 (cs) [Submitted on 25 Jun 2026] Title:When Does…

6/29/2026 · 2 min read · 34 views
Odyssey: Constructing Verifiable Local Truth-Preserving Foundation Models
arXiv.org

Odyssey: Constructing Verifiable Local Truth-Preserving Foundation Models

A foundry is an organized sheaf of knowledge that carries within it an argumentation component. Concrete foundries are…

6/29/2026 · 3 min read · 32 views
Grounded Iterative Language Planning: How Parameterized World Models Reduce Hallucination Propagation in LLM Agents
arXiv.org

Grounded Iterative Language Planning: How Parameterized World Models Reduce Hallucination Propagation in LLM Agents

An agent-based world model calls an LLM API and reasons flexibly in language, but its errors appear as hallucinated…

6/29/2026 · 3 min read · 31 views

How WeSearch handles this source

WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.

Indexing: Allowed Snippet: Allowed AI summary: Limited Retrieval / RAG: Not asserted Model training: Not asserted Commercial reuse: Not permitted

More ai-research sources

Visit arXiv cs.AI directly →