WeSearch
Hub / ai-research / arXiv cs.AI
ai-research · source

arXiv cs.AI on WeSearch

Recent ai-research headlines from arXiv cs.AI.

Memory-Augmented Reinforcement Learning Agent for CAD Generation
arXiv cs.AI

Memory-Augmented Reinforcement Learning Agent for CAD Generation

A new paper presents a memory-augmented reinforcement learning framework aimed at improving computer-aided design…

5/20/2026 · 3 min read · 30 views
CogScale: Scalable Benchmark for Sequence Processing
arXiv cs.AI

CogScale: Scalable Benchmark for Sequence Processing

The paper introduces CogScale, a benchmark designed to evaluate the cognitive and memory abilities of various AI…

5/20/2026 · 3 min read · 29 views
What Really Improves Mathematical Reasoning: Structured Reasoning Signals Beyond Pure Code
arXiv cs.AI

What Really Improves Mathematical Reasoning: Structured Reasoning Signals Beyond Pure Code

A recent study investigates the role of code in enhancing mathematical reasoning. The findings suggest that while code…

5/20/2026 · 3 min read · 29 views
GroupAffect-4: A Multimodal Dataset of Four-Person Collaborative Interaction
arXiv cs.AI

GroupAffect-4: A Multimodal Dataset of Four-Person Collaborative Interaction

The GroupAffect-4 dataset introduces a multimodal approach to studying four-person collaborative interactions. It…

5/20/2026 · 3 min read · 34 views
Minimax Optimal Variance-Aware Regret Bounds for Multinomial Logistic MDPs
arXiv cs.AI

Minimax Optimal Variance-Aware Regret Bounds for Multinomial Logistic MDPs

The article discusses a new algorithm for reinforcement learning in multinomial logistic Markov Decision Processes…

5/20/2026 · 3 min read · 36 views
OpenComputer: Verifiable Software Worlds for Computer-Use Agents
arXiv cs.AI

OpenComputer: Verifiable Software Worlds for Computer-Use Agents

OpenComputer is a new framework designed to create verifiable software worlds for computer-use agents. It includes…

5/20/2026 · 2 min read · 27 views
Distribution-Free Uncertainty Quantification for Continuous AI Agent Evaluation
arXiv cs.AI

Distribution-Free Uncertainty Quantification for Continuous AI Agent Evaluation

The paper presents a method for uncertainty quantification in continuous AI agent evaluation. It introduces split…

5/20/2026 · 3 min read · 21 views
From SGD to Muon: Adaptive Optimization via Schatten-p Norms
arXiv cs.AI

From SGD to Muon: Adaptive Optimization via Schatten-p Norms

The article introduces a new adaptive optimization framework that utilizes Schatten-p norms for deep neural networks.…

5/20/2026 · 3 min read · 31 views
Prior Knowledge or Search? A Study of LLM Agents in Hardware-Aware Code Optimization
arXiv cs.AI

Prior Knowledge or Search? A Study of LLM Agents in Hardware-Aware Code Optimization

A recent study investigates the effectiveness of LLM agents in hardware-aware code optimization. The research reveals…

5/20/2026 · 3 min read · 19 views
From Prompts to Pavement Through Time: Temporal Grounding in Agentic Scene-to-Plan Reasoning
arXiv cs.AI

From Prompts to Pavement Through Time: Temporal Grounding in Agentic Scene-to-Plan Reasoning

The article discusses a study on temporal grounding in scene-to-plan reasoning for autonomous vehicles. It highlights…

5/20/2026 · 3 min read · 37 views
Explainable Wastewater Digital Twins: Adaptive Context-Conditioned Structured Simulators with Self-Falsifying Decision Support
arXiv cs.AI

Explainable Wastewater Digital Twins: Adaptive Context-Conditioned Structured Simulators with Self-Falsifying Decision Support

The article discusses the development of an explainable digital twin for wastewater treatment plants. This simulator…

5/20/2026 · 3 min read · 27 views
Streamlined Constraint Reasoning via CNN Pattern Recognition on Enumerated Solutions
arXiv cs.AI

Streamlined Constraint Reasoning via CNN Pattern Recognition on Enumerated Solutions

The paper presents a novel approach to streamline constraint reasoning using Convolutional Neural Networks (CNN) for…

5/20/2026 · 3 min read · 27 views
PEEK: Context Map as an Orientation Cache for Long-Context LLM Agents
arXiv cs.AI

PEEK: Context Map as an Orientation Cache for Long-Context LLM Agents

The paper introduces PEEK, a system designed to enhance long-context LLM agents by maintaining reusable orientation…

5/20/2026 · 3 min read · 25 views
Robotics-Inspired Guardrails for Foundation Models in Socially Sensitive Domains
arXiv cs.AI

Robotics-Inspired Guardrails for Foundation Models in Socially Sensitive Domains

The paper discusses the need for improved guardrails for foundation models used in socially sensitive areas like…

5/20/2026 · 2 min read · 30 views
Probabilistic Tiny Recursive Model
arXiv cs.AI

Probabilistic Tiny Recursive Model

The article discusses the introduction of a new framework called Probabilistic Tiny Recursive Model (PTRM) designed to…

5/20/2026 · 2 min read · 30 views
GeoX: Mastering Geospatial Reasoning Through Self-Play and Verifiable Rewards
arXiv cs.AI

GeoX: Mastering Geospatial Reasoning Through Self-Play and Verifiable Rewards

GeoX is a new framework designed to enhance geospatial reasoning through self-play and verifiable rewards. It operates…

5/20/2026 · 2 min read · 34 views
When Skills Don't Help: A Negative Result on Procedural Knowledge for Tool-Grounded Agents in Offensive Cybersecurity
arXiv cs.AI

When Skills Don't Help: A Negative Result on Procedural Knowledge for Tool-Grounded Agents in Offensive Cybersecurity

A recent study challenges the effectiveness of procedural knowledge in tool-grounded agents for offensive…

5/20/2026 · 3 min read · 21 views
AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration
arXiv cs.AI

AutoResearchClaw: Self-Reinforcing Autonomous Research with Human-AI Collaboration

AutoResearchClaw is a new multi-agent autonomous research pipeline designed to enhance scientific discovery through…

5/20/2026 · 3 min read · 28 views
Probing Embodied LLMs: When Higher Observation Fidelity Hurts Problem Solving
arXiv cs.AI

Probing Embodied LLMs: When Higher Observation Fidelity Hurts Problem Solving

A recent study explores the performance of embodied large language models (LLMs) in robotic tasks. The research…

5/20/2026 · 3 min read · 28 views
Neurosymbolic Learning for Inference-Time Argumentation
arXiv cs.AI

Neurosymbolic Learning for Inference-Time Argumentation

The article discusses a new framework called inference-time argumentation (ITA) for claim verification in uncertain…

5/20/2026 · 3 min read · 32 views
Using Aristotle API for AI-Assisted Theorem Proving in Lean 4: A Formalisation Case Study of the Grasshopper Problem
arXiv cs.AI

Using Aristotle API for AI-Assisted Theorem Proving in Lean 4: A Formalisation Case Study of the Grasshopper Problem

The article discusses a case study on using the Aristotle API for AI-assisted theorem proving in Lean 4, focusing on…

5/20/2026 · 3 min read · 28 views
Not Every Rubric Teaches Equally: Policy-Aware Rubric Rewards for RLVR
arXiv cs.AI

Not Every Rubric Teaches Equally: Policy-Aware Rubric Rewards for RLVR

The paper introduces POW3R, a policy-aware rubric reward framework for reinforcement learning with verifiable rewards.…

5/20/2026 · 3 min read · 35 views
HaorFloodAlert: Deseasonalized ML Ensemble for 72-Hour Flood Prediction in Bangladesh Haor Wetlands
arXiv cs.AI

HaorFloodAlert: Deseasonalized ML Ensemble for 72-Hour Flood Prediction in Bangladesh Haor Wetlands

A new machine learning model called HaorFloodAlert has been developed to predict flash floods in Bangladesh's haor…

5/20/2026 · 3 min read · 27 views
A Methodology for Selecting and Composing Runtime Architecture Patterns for Production LLM Agents
arXiv cs.AI

A Methodology for Selecting and Composing Runtime Architecture Patterns for Production LLM Agents

The paper presents a methodology for selecting and composing runtime architecture patterns for production LLM agents.…

5/20/2026 · 3 min read · 32 views
An Efficient Multilevel Preconditioned Nonlinear Conjugate Gradient Method for Incremental Potential Contact
arXiv cs.AI

An Efficient Multilevel Preconditioned Nonlinear Conjugate Gradient Method for Incremental Potential Contact

A new method called MAS-PNCG has been developed to improve the efficiency of Incremental Potential Contact…

5/20/2026 · 3 min read · 22 views
OmniGUI: Benchmarking GUI Agents in Omni-Modal Smartphone Environments
arXiv cs.AI

OmniGUI: Benchmarking GUI Agents in Omni-Modal Smartphone Environments

The paper introduces OmniGUI, a new benchmark for evaluating graphical user interface (GUI) agents in omni-modal…

5/20/2026 · 3 min read · 27 views
Interoceptive Divergence in Aesthetic Evaluation and Implications for Human-AI Alignment
arXiv cs.AI

Interoceptive Divergence in Aesthetic Evaluation and Implications for Human-AI Alignment

The paper discusses the divergence between human and AI aesthetic evaluations. It highlights that while AI can…

5/20/2026 · 3 min read · 31 views
DOTRAG: Retrieval-Time Reasoning Along Paths
arXiv cs.AI

DOTRAG: Retrieval-Time Reasoning Along Paths

The paper introduces DotRAG, a new framework for Graph Retrieval-Augmented Generation that reformulates retrieval as a…

5/20/2026 · 2 min read · 33 views
ALDEN: Boosting Private Data Extraction from Retrieval-Augmented Generation Systems via Active Learning and Distribution Estimation
arXiv cs.AI

ALDEN: Boosting Private Data Extraction from Retrieval-Augmented Generation Systems via Active Learning and Distribution Estimation

The paper presents ALDEN, a new method for enhancing private data extraction from Retrieval-Augmented Generation (RAG)…

5/20/2026 · 3 min read · 38 views
Query-Conditioned Graph Retrieval for Contextualized LLM Reasoning in Personalized Wearable Data
arXiv cs.AI

Query-Conditioned Graph Retrieval for Contextualized LLM Reasoning in Personalized Wearable Data

The article discusses a new framework called Wearable As Graph (WAG) designed for analyzing personalized wearable data…

5/20/2026 · 3 min read · 28 views
From Intent to AI Pipelines: A Controlled Agentic Framework for Non-AI Expert Scientists
arXiv cs.AI

From Intent to AI Pipelines: A Controlled Agentic Framework for Non-AI Expert Scientists

The paper introduces Domain-Driven Adaptable AI Pipelines (DDAP), a framework designed to assist non-expert scientists…

5/20/2026 · 3 min read · 24 views
STAR: Semantic-Tuned and Tail-Adaptive Retriever for Graph-Augmented Generation
arXiv cs.AI

STAR: Semantic-Tuned and Tail-Adaptive Retriever for Graph-Augmented Generation

The article introduces STAR, a new retriever designed to enhance Graph Retrieval Augmented Generation for multi-hop…

5/20/2026 · 3 min read · 35 views
Retrieve Only Relevant Tables Whether Few or Many: Adaptive Table Retrieval Method
arXiv cs.AI

Retrieve Only Relevant Tables Whether Few or Many: Adaptive Table Retrieval Method

The article discusses a new adaptive table retrieval method designed to improve the retrieval of relevant tables from…

5/20/2026 · 3 min read · 32 views
DualView: Adaptive Local-Global Fusion for Multi-Hop Document Reranking
arXiv cs.AI

DualView: Adaptive Local-Global Fusion for Multi-Hop Document Reranking

The paper presents a new framework called DualView for multi-hop document reranking, which is essential for effective…

5/20/2026 · 2 min read · 38 views
ClusterRAG: Cluster-Based Collaborative Filtering for Personalized Retrieval-Augmented Generation
arXiv cs.AI

ClusterRAG: Cluster-Based Collaborative Filtering for Personalized Retrieval-Augmented Generation

The paper introduces ClusterRAG, a novel approach for Personalized Retrieval-Augmented Generation (RAG). It emphasizes…

5/20/2026 · 2 min read · 30 views
Agentic GraphRAG: Navigating Unstructured Financial Data with Collaborative AI
arXiv cs.AI

Agentic GraphRAG: Navigating Unstructured Financial Data with Collaborative AI

The article introduces the Agentic GraphRAG framework designed for analyzing unstructured financial data. This system…

5/20/2026 · 3 min read · 33 views
Improving Retrieval-Augmented Generation without Taxonomy-based Error Categorization
arXiv cs.AI

Improving Retrieval-Augmented Generation without Taxonomy-based Error Categorization

The paper presents a new approach to improving Retrieval-Augmented Generation (RAG) without relying on taxonomy-based…

5/20/2026 · 2 min read · 34 views
Decentralized autonomous organization and blockchain-based incentivization framework for community-based facilities management
arXiv cs.AI

Decentralized autonomous organization and blockchain-based incentivization framework for community-based facilities management

A new framework for community-based facilities management has been proposed, utilizing blockchain technology and…

5/20/2026 · 2 min read · 37 views
M3DocDep: Multi-modal, Multi-page, Multi-document Dependency Chunking with Large Vision-Language Models
arXiv cs.AI

M3DocDep: Multi-modal, Multi-page, Multi-document Dependency Chunking with Large Vision-Language Models

The article discusses a new method called M3DocDep for processing long, multi-page documents using large…

5/20/2026 · 3 min read · 36 views
Query-Aware Flow Diffusion for Graph-Based RAG with Retrieval Guarantees
arXiv cs.AI

Query-Aware Flow Diffusion for Graph-Based RAG with Retrieval Guarantees

The paper introduces Query-Aware Flow Diffusion for Graph-Based Retrieval-Augmented Generation (QAFD-RAG), a new…

5/20/2026 · 3 min read · 29 views
Mask-to-Correct$^+$: Leveraging Retriever Diversity for Masking-guided Faithful Fact Correction
arXiv cs.AI

Mask-to-Correct$^+$: Leveraging Retriever Diversity for Masking-guided Faithful Fact Correction

The article discusses a new framework called Mask-to-Correct$^+$ designed to improve automated fact correction in the…

5/20/2026 · 2 min read · 30 views
A Reproducibility Analysis of PO4ISR: Diagnosing and Mitigating Semantic Drift in LLM-Based Session Recommendation
arXiv cs.AI

A Reproducibility Analysis of PO4ISR: Diagnosing and Mitigating Semantic Drift in LLM-Based Session Recommendation

The article presents a reproducibility analysis of the PO4ISR model for session-based recommendations. It identifies…

5/20/2026 · 3 min read · 31 views
Can LLMs Emulate Human Belief Dynamics?
arXiv cs.AI

Can LLMs Emulate Human Belief Dynamics?

A recent study tested whether large language models (LLMs) can simulate human belief dynamics in social networks. The…

5/20/2026 · 2 min read · 34 views
The Insurability Frontier of AI Risk: Mapping Threats to Affirmative Coverage, Silent Exposures, and Exclusions
arXiv cs.AI

The Insurability Frontier of AI Risk: Mapping Threats to Affirmative Coverage, Silent Exposures, and Exclusions

The paper discusses the challenges of insuring risks associated with artificial intelligence (AI). It identifies…

5/20/2026 · 3 min read · 26 views
Features have life history. And we should care
arXiv cs.AI

Features have life history. And we should care

The paper discusses the life history of features in language models, highlighting their emergence, persistence, and…

5/20/2026 · 3 min read · 27 views
Dimensional Balance Improves Large Scale Spatiotemporal Prediction Performance
arXiv cs.AI

Dimensional Balance Improves Large Scale Spatiotemporal Prediction Performance

A new framework has been proposed to enhance spatiotemporal prediction performance by addressing the limitations of…

5/20/2026 · 3 min read · 31 views
Robust Basis Spline Decoupling for the Compression of Transformer Models
arXiv cs.AI

Robust Basis Spline Decoupling for the Compression of Transformer Models

A new paper introduces a B-spline-based decoupling framework for compressing transformer models. This method aims to…

5/20/2026 · 3 min read · 39 views
HELLoRA: Hot Experts Layer-Level Low-Rank Adaptation for Mixture-of-Experts Models
arXiv cs.AI

HELLoRA: Hot Experts Layer-Level Low-Rank Adaptation for Mixture-of-Experts Models

The paper introduces HELLoRA, a method for efficient fine-tuning of Mixture-of-Experts models using Low-Rank…

5/20/2026 · 3 min read · 32 views
Simply Stabilizing the Loop via Fully Looped Transformer
arXiv cs.AI

Simply Stabilizing the Loop via Fully Looped Transformer

The paper presents the Fully Looped Transformer, a model designed to enhance training stability and performance in…

5/20/2026 · 3 min read · 28 views
ReCrit: Transition-Aware Reinforcement Learning for Scientific Critic Reasoning
arXiv cs.AI

ReCrit: Transition-Aware Reinforcement Learning for Scientific Critic Reasoning

The paper introduces ReCrit, a transition-aware reinforcement learning framework aimed at improving scientific…

5/20/2026 · 3 min read · 37 views

How WeSearch handles this source

WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.

Indexing: Allowed Snippet: Allowed AI summary: Limited Retrieval / RAG: Not asserted Model training: Not asserted Commercial reuse: Not permitted

More ai-research sources

Visit arXiv cs.AI directly →