WeSearch
Hub / ai-research / arXiv cs.AI
ai-research · source

arXiv cs.AI on WeSearch

Recent ai-research headlines from arXiv cs.AI.

ADR: An Agentic Detection System for Enterprise Agentic AI Security
arXiv cs.AI

ADR: An Agentic Detection System for Enterprise Agentic AI Security

The ADR system is a new framework designed to enhance the security of AI agents in enterprise settings. It addresses…

5/19/2026 · 3 min read · 24 views
QQJ: Quantifying Qualitative Judgment for Scalable and Human-Aligned Evaluation of Generative AI
arXiv cs.AI

QQJ: Quantifying Qualitative Judgment for Scalable and Human-Aligned Evaluation of Generative AI

The paper introduces Quantifying Qualitative Judgment (QQJ), a new framework for evaluating generative AI. QQJ aims to…

5/19/2026 · 3 min read · 28 views
Heterogeneous Information-Bottleneck Coordination Graphs for Multi-Agent Reinforcement Learning
arXiv cs.AI

Heterogeneous Information-Bottleneck Coordination Graphs for Multi-Agent Reinforcement Learning

The paper introduces Heterogeneous Information-Bottleneck Coordination Graphs (HIBCG) for multi-agent reinforcement…

5/19/2026 · 3 min read · 46 views
Computational Challenges in Token Economics: Bridging Economic Theory and AI System Design
arXiv cs.AI

Computational Challenges in Token Economics: Bridging Economic Theory and AI System Design

The paper discusses the intersection of token economics and AI system design, highlighting the computational…

5/19/2026 · 3 min read · 31 views
Multi-Party Multi-Objective Optimization as Consensus Search: Runtime Analysis of Cross-Party Recombination
arXiv cs.AI

Multi-Party Multi-Objective Optimization as Consensus Search: Runtime Analysis of Cross-Party Recombination

The paper discusses multi-party multi-objective optimization problems (MPMOPs) and their unique requirements for…

5/19/2026 · 3 min read · 28 views
The Capability Paradox: How Smarter Auditors Make Multi-Agent Systems Less Secure
arXiv cs.AI

The Capability Paradox: How Smarter Auditors Make Multi-Agent Systems Less Secure

A recent study explores the security vulnerabilities in multi-agent systems that utilize large language models. The…

5/19/2026 · 3 min read · 27 views
RAG-based EEG-to-Text Translation Using Deep Learning and LLMs
arXiv cs.AI

RAG-based EEG-to-Text Translation Using Deep Learning and LLMs

A new study proposes a retrieval-augmented generation (RAG)-based method for translating EEG signals into text. This…

5/19/2026 · 3 min read · 26 views
Self-supervised Hierarchical Visual Reasoning with World Model
arXiv cs.AI

Self-supervised Hierarchical Visual Reasoning with World Model

The paper introduces ResDreamer, a hierarchical world model designed for self-supervised visual reasoning in 3D…

5/19/2026 · 3 min read · 33 views
Memory-Guided Tree Search with Cross-Branch Knowledge Transfer for LLM Solver Synthesis
arXiv cs.AI

Memory-Guided Tree Search with Cross-Branch Knowledge Transfer for LLM Solver Synthesis

The paper introduces MEMOIR, a memory-guided tree-search framework designed for combinatorial optimization problems.…

5/19/2026 · 3 min read · 25 views
Evaluating Deep Research Agents on Expert Consulting Work: A Benchmark with Verifiers, Rubrics, and Cognitive Traps
arXiv cs.AI

Evaluating Deep Research Agents on Expert Consulting Work: A Benchmark with Verifiers, Rubrics, and Cognitive Traps

A new benchmark has been introduced to evaluate deep research agents (DRAs) on their ability to produce structured…

5/19/2026 · 3 min read · 29 views
Generalization or Memorization? Brittleness Testing for Chess-Trained Language Models
arXiv cs.AI

Generalization or Memorization? Brittleness Testing for Chess-Trained Language Models

The paper discusses the performance of chess-trained language models, particularly focusing on KinGPT, a 25M-parameter…

5/19/2026 · 3 min read · 25 views
ECG-WM: A Physiology-Informed ECG World Model for Clinical Intervention Simulation
arXiv cs.AI

ECG-WM: A Physiology-Informed ECG World Model for Clinical Intervention Simulation

The ECG World Model (ECG-WM) is a new framework designed to enhance the predictive simulation of cardiac…

5/19/2026 · 3 min read · 33 views
NeuSymMS: A Hybrid Neuro-Symbolic Memory System for Persistent, Self-Curating LLM Agents
arXiv cs.AI

NeuSymMS: A Hybrid Neuro-Symbolic Memory System for Persistent, Self-Curating LLM Agents

NeuSymMS is a new hybrid neuro-symbolic memory system designed for large language model agents. It enables these…

5/19/2026 · 2 min read · 25 views
AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment
arXiv cs.AI

AutoRubric-T2I: Robust Rule-Based Reward Model for Text-to-Image Alignment

The paper introduces AutoRubric-T2I, a novel framework for improving text-to-image (T2I) alignment through automatic…

5/19/2026 · 3 min read · 23 views
GraphMind: From Operational Traces to Self-Evolving Workflow Automation
arXiv cs.AI

GraphMind: From Operational Traces to Self-Evolving Workflow Automation

GraphMind is a new system designed to automate complex operational workflows without human input. It constructs,…

5/19/2026 · 3 min read · 25 views
Prediction of Challenging Behaviors Associated with Profound Autism in a Classroom Setting Using Wearable Sensors
arXiv cs.AI

Prediction of Challenging Behaviors Associated with Profound Autism in a Classroom Setting Using Wearable Sensors

A recent study explores the prediction of challenging behaviors in children with profound autism using wearable…

5/19/2026 · 3 min read · 40 views
Episodic-Semantic Memory Architecture for Long-Horizon Scientific Agents
arXiv cs.AI

Episodic-Semantic Memory Architecture for Long-Horizon Scientific Agents

The paper presents a new memory architecture designed for Large Language Models (LLMs) to enhance their performance in…

5/19/2026 · 3 min read · 39 views
WebGameBench: Requirement-to-Application Evaluation for Coding Agents via Browser-Native Games
arXiv cs.AI

WebGameBench: Requirement-to-Application Evaluation for Coding Agents via Browser-Native Games

WebGameBench is a new benchmark designed to evaluate coding agents' ability to create browser-native games from…

5/19/2026 · 3 min read · 34 views
Causal Intervention-Based Memory Selection for Long-Horizon LLM Agents
arXiv cs.AI

Causal Intervention-Based Memory Selection for Long-Horizon LLM Agents

The article discusses a new technique called Causal Memory Intervention (CMI) designed for long-horizon LLM agents.…

5/19/2026 · 3 min read · 29 views
SAPO: Step-Aligned Policy Optimization for Reasoning-Based Generative Recommendation
arXiv cs.AI

SAPO: Step-Aligned Policy Optimization for Reasoning-Based Generative Recommendation

The article discusses a new approach called SAPO, which stands for Step-Aligned Policy Optimization, aimed at…

5/19/2026 · 3 min read · 25 views
Multimodal Cultural Heritage Knowledge Graph Extension with Language and Vision Models
arXiv cs.AI

Multimodal Cultural Heritage Knowledge Graph Extension with Language and Vision Models

The article discusses a new approach to extending cultural heritage knowledge graphs using language and vision models.…

5/19/2026 · 3 min read · 36 views
EGI: A Multimodal Emotional AI Framework for Enhancing Scrum Master Real-time Self-Awareness
arXiv cs.AI

EGI: A Multimodal Emotional AI Framework for Enhancing Scrum Master Real-time Self-Awareness

The paper presents EGI, a multimodal emotional AI framework designed to enhance the real-time self-awareness of Scrum…

5/19/2026 · 2 min read · 40 views
EXG: Self-Evolving Agents with Experience Graphs
arXiv cs.AI

EXG: Self-Evolving Agents with Experience Graphs

The paper introduces EXG, an experience graph framework designed for self-evolving agents. This framework allows…

5/19/2026 · 3 min read · 35 views
Divergence-Suppressing Couplings for Rectified Flow
arXiv cs.AI

Divergence-Suppressing Couplings for Rectified Flow

The paper discusses divergence-suppressing couplings for Rectified Flow, aimed at improving trajectory generation. It…

5/19/2026 · 2 min read · 21 views
Harnessing LLM Agents with Skill Programs
arXiv cs.AI

Harnessing LLM Agents with Skill Programs

The paper introduces HASP, a framework designed to enhance LLM agents by equipping them with executable Program…

5/19/2026 · 3 min read · 28 views
Agents for Experiments, Experiments for Agents: A Design Grammar for AI-Enabled Experimental Science
arXiv cs.AI

Agents for Experiments, Experiments for Agents: A Design Grammar for AI-Enabled Experimental Science

The paper discusses the role of AI systems in experimental science, emphasizing their interaction with humans and…

5/19/2026 · 3 min read · 34 views
Surface-Form Neural Sparse Retrieval: Robust Fuzzy Matching for Industrial Music Search
arXiv cs.AI

Surface-Form Neural Sparse Retrieval: Robust Fuzzy Matching for Industrial Music Search

A new paper presents a robust neural sparse retrieval system aimed at improving music search efficiency. The system…

5/19/2026 · 3 min read · 25 views
Entropy-Gradient Inversion: Moving Toward Internal Mechanism of Large Reasoning Models
arXiv cs.AI

Entropy-Gradient Inversion: Moving Toward Internal Mechanism of Large Reasoning Models

The paper discusses a new approach to understanding the internal mechanisms of Large Reasoning Models (LRMs) through a…

5/19/2026 · 2 min read · 27 views
STRIDE: A Self-Reflective Agent Framework for Reliable Automatic Equation Discovery
arXiv cs.AI

STRIDE: A Self-Reflective Agent Framework for Reliable Automatic Equation Discovery

The article introduces STRIDE, a self-reflective agent framework designed to enhance the reliability of automatic…

5/19/2026 · 2 min read · 29 views
Accelerating AI-Powered Research: The PuppyChatter Framework for Usable and Flexible Tooling
arXiv cs.AI

Accelerating AI-Powered Research: The PuppyChatter Framework for Usable and Flexible Tooling

The PuppyChatter framework aims to simplify the development of AI applications by addressing the complexities…

5/19/2026 · 2 min read · 24 views
Going Headless? On the Boundaries of Vertical AI Firms
arXiv cs.AI

Going Headless? On the Boundaries of Vertical AI Firms

The article discusses the concept of 'going headless' in vertical AI firms, where companies unbundle their services…

5/19/2026 · 3 min read · 27 views
Interactive Evaluation Requires a Design Science
arXiv cs.AI

Interactive Evaluation Requires a Design Science

The paper discusses the evolving landscape of AI evaluation, particularly in the context of interactive benchmarks. It…

5/19/2026 · 3 min read · 28 views
Remembering More, Risking More: Longitudinal Safety Risks in Memory-Equipped LLM Agents
arXiv cs.AI

Remembering More, Risking More: Longitudinal Safety Risks in Memory-Equipped LLM Agents

The paper discusses the safety risks associated with memory-equipped LLM agents over time. It highlights the concept…

5/19/2026 · 3 min read · 25 views
KISS - Knowledge Infrastructure for Scientific Simulation: A Scaffolding for Agentic Earth Science
arXiv cs.AI

KISS - Knowledge Infrastructure for Scientific Simulation: A Scaffolding for Agentic Earth Science

The article discusses the development of the Knowledge Infrastructure for Scientific Simulation (KISS), aimed at…

5/19/2026 · 3 min read · 26 views
PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization
arXiv cs.AI

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization

The article presents a new model called PAIR, which aims to improve multi-turn agent optimization in large language…

5/19/2026 · 3 min read · 17 views
Evaluating Cognitive Age Alignment in Interactive AI Agents
arXiv cs.AI

Evaluating Cognitive Age Alignment in Interactive AI Agents

A new study introduces ChildAgentEval, a benchmark for assessing cognitive age alignment in interactive AI agents.…

5/19/2026 · 2 min read · 32 views
DuIVRS-2: An LLM-based Interactive Voice Response System for Large-scale POI Attribute Acquisition
arXiv cs.AI

DuIVRS-2: An LLM-based Interactive Voice Response System for Large-scale POI Attribute Acquisition

DuIVRS-2 is a new interactive voice response system designed for large-scale Point of Interest (POI) attribute…

5/19/2026 · 3 min read · 26 views
LAST-RAG: Literature-Anchored Stochastic Trajectory Retrieval-Augmented Generation for Knowledge-Conditioned Degradation Model Selection
arXiv cs.AI

LAST-RAG: Literature-Anchored Stochastic Trajectory Retrieval-Augmented Generation for Knowledge-Conditioned Degradation Model Selection

The paper introduces LAST-RAG, a method for selecting degradation models based on observed health indicator…

5/19/2026 · 3 min read · 28 views
Agentic Chunking and Bayesian De-chunking of AI Generated Fuzzy Cognitive Maps: A Model of the Thucydides Trap
arXiv cs.AI

Agentic Chunking and Bayesian De-chunking of AI Generated Fuzzy Cognitive Maps: A Model of the Thucydides Trap

The paper discusses a method for generating feedback causal fuzzy cognitive maps (FCMs) using AI agents. It explores…

5/19/2026 · 3 min read · 36 views
Ethical Hyper-Velocity (EHV): A Provably Deterministic Governance-Aware JIT Compiler Architecture for Agentic Systems
arXiv cs.AI

Ethical Hyper-Velocity (EHV): A Provably Deterministic Governance-Aware JIT Compiler Architecture for Agentic Systems

The article introduces Ethical Hyper-Velocity (EHV), a new framework designed for the formal verification of AI…

5/19/2026 · 3 min read · 29 views
SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain
arXiv cs.AI

SVFSearch: A Multimodal Knowledge-Intensive Benchmark for Short-Video Frame Search in the Gaming Vertical Domain

SVFSearch is a new benchmark designed for short-video frame search specifically in the gaming domain. It includes a…

5/19/2026 · 3 min read · 35 views
Reconciling Contradictory Views on the Effectiveness of SFT in LLMs: An Interaction Perspective
arXiv cs.AI

Reconciling Contradictory Views on the Effectiveness of SFT in LLMs: An Interaction Perspective

The paper investigates the effectiveness of supervised fine-tuning (SFT) in large language models (LLMs) compared to…

5/19/2026 · 2 min read · 19 views
Unleashing LLMs in Bayesian Optimization: Preference-Guided Framework for Scientific Discovery
arXiv cs.AI

Unleashing LLMs in Bayesian Optimization: Preference-Guided Framework for Scientific Discovery

A new framework called LLM-Guided Bayesian Optimization (LGBO) has been proposed to enhance scientific discovery…

5/19/2026 · 3 min read · 32 views
Shared Backbone PPO for Multi-UAV Communication Coverage with Connection Preservation
arXiv cs.AI

Shared Backbone PPO for Multi-UAV Communication Coverage with Connection Preservation

The paper presents a new algorithm called Shared Backbone Proximal Policy Optimization (Shared Backbone PPO) for…

5/19/2026 · 2 min read · 34 views
TeleCom-Bench: How Far Are Large Language Models from Industrial Telecommunication Applications?
arXiv cs.AI

TeleCom-Bench: How Far Are Large Language Models from Industrial Telecommunication Applications?

The paper introduces TeleCom-Bench, a benchmark designed to evaluate the performance of Large Language Models (LLMs)…

5/19/2026 · 3 min read · 33 views
New Insight of Variance reduce in Zero-Order Hard-Thresholding: Mitigating Gradient Error and Expansivity Contradictions
arXiv cs.AI

New Insight of Variance reduce in Zero-Order Hard-Thresholding: Mitigating Gradient Error and Expansivity Contradictions

A new paper presents insights into variance reduction in zero-order hard-thresholding algorithms. The proposed method…

5/19/2026 · 3 min read · 26 views
DocOS: Towards Proactive Document-Guided Actions in GUI Agents
arXiv cs.AI

DocOS: Towards Proactive Document-Guided Actions in GUI Agents

The paper introduces DocOS, a benchmark aimed at enhancing the capabilities of GUI agents through proactive…

5/19/2026 · 3 min read · 22 views
LLM-Guided Communication for Cooperative Multi-Agent Reinforcement Learning
arXiv cs.AI

LLM-Guided Communication for Cooperative Multi-Agent Reinforcement Learning

The paper presents a novel approach to communication in multi-agent reinforcement learning (MARL) called LLM-driven…

5/19/2026 · 2 min read · 41 views
Learning to Solve Compositional Geometry Routing Problems
arXiv cs.AI

Learning to Solve Compositional Geometry Routing Problems

The article discusses a new approach to solving the Compositional Geometry Routing Problem (CGRP), which encompasses…

5/19/2026 · 2 min read · 30 views
Safety Geometry Collapse in Multimodal LLMs and Adaptive Drift Correction
arXiv cs.AI

Safety Geometry Collapse in Multimodal LLMs and Adaptive Drift Correction

The paper discusses the safety challenges faced by multimodal large language models (MLLMs) in transferring safety…

5/19/2026 · 3 min read · 31 views

How WeSearch handles this source

WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.

Indexing: Allowed Snippet: Allowed AI summary: Limited Retrieval / RAG: Not asserted Model training: Not asserted Commercial reuse: Not permitted

More ai-research sources

Visit arXiv cs.AI directly →