WeSearch
Hub / ai-research / arXiv cs.AI
ai-research · source

arXiv cs.AI on WeSearch

Recent ai-research headlines from arXiv cs.AI.

TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning
arXiv cs.AI

TaskGround: Structured Executable Task Inference for Full-Scene Household Reasoning

The paper introduces TaskGround, a framework designed for structured executable task inference in household settings.…

5/19/2026 · 3 min read · 32 views
POST: Prior-Observation Adversarial Learning of Spatio-Temporal Associations for Multivariate Time Series Anomaly Detection
arXiv cs.AI

POST: Prior-Observation Adversarial Learning of Spatio-Temporal Associations for Multivariate Time Series Anomaly Detection

The article discusses a new framework for Multivariate Time Series Anomaly Detection (MTSAD) that integrates…

5/19/2026 · 3 min read · 29 views
Generative AI and the Productivity Divide: Human-AI Complementarities in Education
arXiv cs.AI

Generative AI and the Productivity Divide: Human-AI Complementarities in Education

A recent study explores the impact of Generative AI on productivity in education. The findings indicate that while…

5/19/2026 · 3 min read · 28 views
Evidence-Grounded Frontier Mapping and Agentic Hypothesis Generation in Nanomedicine
arXiv cs.AI

Evidence-Grounded Frontier Mapping and Agentic Hypothesis Generation in Nanomedicine

A new system called pArticleMap has been introduced to enhance research in nanomedicine by mapping literature and…

5/19/2026 · 3 min read · 34 views
Whispers in the Noise: Surrogate-Guided Concept Awakening via a Multi-Agent Framework
arXiv cs.AI

Whispers in the Noise: Surrogate-Guided Concept Awakening via a Multi-Agent Framework

The article discusses a new framework called ConceptAgent designed to address limitations in current concept erasure…

5/19/2026 · 3 min read · 30 views
TRACE: Trajectory Correction from Cross-layer Evidence for Hallucination Reduction
arXiv cs.AI

TRACE: Trajectory Correction from Cross-layer Evidence for Hallucination Reduction

The paper introduces TRACE, a new algorithm designed to reduce hallucinations in language models by utilizing…

5/19/2026 · 3 min read · 31 views
Visualizing the Invisible: Generative Visual Grounding Empowers Universal EEG Understanding in MLLMs
arXiv cs.AI

Visualizing the Invisible: Generative Visual Grounding Empowers Universal EEG Understanding in MLLMs

The paper introduces Generative Visual Grounding (GVG), a framework aimed at enhancing understanding of EEG signals…

5/19/2026 · 3 min read · 35 views
Scalable Environments Drive Generalizable Agents
arXiv cs.AI

Scalable Environments Drive Generalizable Agents

The paper discusses the importance of scalable environments for developing generalizable agents in artificial…

5/19/2026 · 3 min read · 19 views
Pairwise Preference Reward and Group-Based Diversity Enhancement for Superior Open-Ended Generation
arXiv cs.AI

Pairwise Preference Reward and Group-Based Diversity Enhancement for Superior Open-Ended Generation

The paper introduces a new reinforcement learning method called Pairwise Preference Reward and Group-based Diversity…

5/19/2026 · 3 min read · 28 views
Beyond the Cartesian Illusion: Testing Two-Stage Multi-Modal Theory of Mind under Perceptual Bottlenecks
arXiv cs.AI

Beyond the Cartesian Illusion: Testing Two-Stage Multi-Modal Theory of Mind under Perceptual Bottlenecks

The paper explores the limitations of Multi-Modal Large Language Models (MLLMs) in spatial reasoning, particularly…

5/19/2026 · 3 min read · 31 views
DARE-EEG: A Foundation Model for Mining Dual-Aligned Representation of EEG
arXiv cs.AI

DARE-EEG: A Foundation Model for Mining Dual-Aligned Representation of EEG

DARE-EEG is a new foundation model designed for mining dual-aligned representations of EEG data. It addresses the…

5/19/2026 · 3 min read · 26 views
SD-Search: On-Policy Hindsight Self-Distillation for Search-Augmented Reasoning
arXiv cs.AI

SD-Search: On-Policy Hindsight Self-Distillation for Search-Augmented Reasoning

The paper titled 'SD-Search: On-Policy Hindsight Self-Distillation for Search-Augmented Reasoning' introduces a novel…

5/19/2026 · 3 min read · 31 views
Causely: A Causal Intelligence Layer for Enterprise AI A Benchmark Study on SRE and Reliability Workflows
arXiv cs.AI

Causely: A Causal Intelligence Layer for Enterprise AI A Benchmark Study on SRE and Reliability Workflows

The paper introduces Causely, a causal intelligence layer designed to enhance AI agents in Site Reliability…

5/19/2026 · 3 min read · 32 views
QSTRBench: a New Benchmark to Evaluate the Ability of Language Models to Reason with Qualitative Spatial and Temporal Calculi
arXiv cs.AI

QSTRBench: a New Benchmark to Evaluate the Ability of Language Models to Reason with Qualitative Spatial and Temporal Calculi

Researchers have introduced QSTRBench, a new benchmark designed to evaluate the reasoning capabilities of language…

5/19/2026 · 3 min read · 27 views
When Fireflies Cluster; Enhancing Automatic Clustering via Centroid-Guided Firefly Optimization
arXiv cs.AI

When Fireflies Cluster; Enhancing Automatic Clustering via Centroid-Guided Firefly Optimization

A new variant of the Firefly Algorithm has been developed to enhance data clustering capabilities. This algorithm…

5/19/2026 · 2 min read · 32 views
OCCAM: Open-set Causal Concept explAnation and Ontology induction for black-box vision Models
arXiv cs.AI

OCCAM: Open-set Causal Concept explAnation and Ontology induction for black-box vision Models

The paper introduces OCCAM, a framework designed for open-set causal concept explanation and ontology induction in…

5/19/2026 · 2 min read · 28 views
A Practical Noise2Noise Denoising Pipeline for High-Throughput Raman Spectroscopy
arXiv cs.AI

A Practical Noise2Noise Denoising Pipeline for High-Throughput Raman Spectroscopy

A new denoising pipeline for high-throughput Raman spectroscopy has been developed. This method utilizes a…

5/19/2026 · 3 min read · 26 views
AMR-SD: Asymmetric Meta-Reflective Self-Distillation for Token-Level Credit Assignment
arXiv cs.AI

AMR-SD: Asymmetric Meta-Reflective Self-Distillation for Token-Level Credit Assignment

The paper introduces Asymmetric Meta-Reflective Self-Distillation (AMR-SD) as a solution for token-level credit…

5/19/2026 · 3 min read · 33 views
VISAFF: Speaker-Centered Visual Affective Feature Learning for Emotion Recognition in Conversation
arXiv cs.AI

VISAFF: Speaker-Centered Visual Affective Feature Learning for Emotion Recognition in Conversation

The article discusses a new framework called VISAFF for Emotion Recognition in Conversation (ERC). This framework aims…

5/19/2026 · 3 min read · 33 views
Query-Conditioned Knowledge Alignment for Reliable Cross-System Medical Reasoning
arXiv cs.AI

Query-Conditioned Knowledge Alignment for Reliable Cross-System Medical Reasoning

The paper presents a novel approach called Query-Conditioned Entity Alignment (QCEA) for improving cross-system…

5/19/2026 · 3 min read · 20 views
When Outcome Looks Right But Discipline Fails: Trace-Based Evaluation Under Hidden Competitor State
arXiv cs.AI

When Outcome Looks Right But Discipline Fails: Trace-Based Evaluation Under Hidden Competitor State

The paper discusses the limitations of outcome-only evaluations in artificial intelligence, particularly in the…

5/19/2026 · 3 min read · 24 views
DeepSlide: From Artifacts to Presentation Delivery
arXiv cs.AI

DeepSlide: From Artifacts to Presentation Delivery

DeepSlide is a new multi-agent system designed to enhance the presentation delivery process. It focuses on improving…

5/18/2026 · 2 min read · 33 views
SDOF: Taming the Alignment Tax in Multi-Agent Orchestration with State-Constrained Dispatch
arXiv cs.AI

SDOF: Taming the Alignment Tax in Multi-Agent Orchestration with State-Constrained Dispatch

The article discusses the SDOF framework designed to enhance multi-agent orchestration by enforcing state constraints.…

5/18/2026 · 3 min read · 23 views
Does Theory of Mind Improvement Really Benefit Human-AI Interactions? Empirical Findings from Interactive Evaluations
arXiv cs.AI

Does Theory of Mind Improvement Really Benefit Human-AI Interactions? Empirical Findings from Interactive Evaluations

The paper explores the impact of improving Theory of Mind (ToM) capabilities in Large Language Models (LLMs) on…

5/18/2026 · 3 min read · 31 views
SkillSmith: Compiling Agent Skills into Boundary-Guided Runtime Interfaces
arXiv cs.AI

SkillSmith: Compiling Agent Skills into Boundary-Guided Runtime Interfaces

SkillSmith is a new framework designed to optimize the execution of skills in large language model-based agent…

5/18/2026 · 3 min read · 35 views
Fair outputs, Biased Internals: Causal Potency and Asymmetry of Latent Bias in LLMs for High-Stakes Decisions
arXiv cs.AI

Fair outputs, Biased Internals: Causal Potency and Asymmetry of Latent Bias in LLMs for High-Stakes Decisions

A recent study investigates the latent biases in instruction-tuned language models used for high-stakes decisions.…

5/18/2026 · 3 min read · 28 views
CAX-Agent: A Lightweight Agent Harness for Reliable APDL Automation
arXiv cs.AI

CAX-Agent: A Lightweight Agent Harness for Reliable APDL Automation

The paper introduces CAX-Agent, a lightweight agent harness designed for reliable automation in MAPDL finite-element…

5/18/2026 · 3 min read · 23 views
NOVA: Fundamental Limits of Knowledge Discovery Through AI
arXiv cs.AI

NOVA: Fundamental Limits of Knowledge Discovery Through AI

The paper titled 'NOVA: Fundamental Limits of Knowledge Discovery Through AI' explores the capabilities and…

5/18/2026 · 3 min read · 19 views
ICRL: Learning to Internalize Self-Critique with Reinforcement Learning
arXiv cs.AI

ICRL: Learning to Internalize Self-Critique with Reinforcement Learning

The paper presents a novel framework called ICRL, which aims to enhance the self-improvement capabilities of language…

5/18/2026 · 3 min read · 34 views
NIMO Controller: a self-driving laboratory orchestrator based on the Model Context Protocol
arXiv cs.AI

NIMO Controller: a self-driving laboratory orchestrator based on the Model Context Protocol

The NIMO Controller is a proposed self-driving laboratory orchestrator that utilizes the Model Context Protocol to…

5/18/2026 · 3 min read · 28 views
Verifiable Agentic Infrastructure: Proof-Derived Authorization for Sovereign AI Systems
arXiv cs.AI

Verifiable Agentic Infrastructure: Proof-Derived Authorization for Sovereign AI Systems

The paper introduces a Distributed Trust Framework (DTF) designed to enhance authorization for sovereign AI systems.…

5/18/2026 · 3 min read · 32 views
Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution
arXiv cs.AI

Solvita: Enhancing Large Language Models for Competitive Programming via Agentic Evolution

The article introduces Solvita, a new framework designed to enhance large language models for competitive programming.…

5/18/2026 · 3 min read · 38 views
SMCEvolve: Principled Scientific Discovery via Sequential Monte Carlo Evolution
arXiv cs.AI

SMCEvolve: Principled Scientific Discovery via Sequential Monte Carlo Evolution

The paper introduces SMCEvolve, a new framework for automated scientific discovery using Sequential Monte Carlo…

5/18/2026 · 2 min read · 34 views
Context Pruning for Coding Agents via Multi-Rubric Latent Reasoning
arXiv cs.AI

Context Pruning for Coding Agents via Multi-Rubric Latent Reasoning

A new framework called LaMR has been proposed for improving coding agents by enhancing context pruning. This method…

5/18/2026 · 3 min read · 25 views
Zero-Shot Goal Recognition with Large Language Models
arXiv cs.AI

Zero-Shot Goal Recognition with Large Language Models

The paper discusses the evaluation of large language models (LLMs) in the context of zero-shot goal recognition. It…

5/18/2026 · 3 min read · 30 views
Belief Engine: Configurable and Inspectable Stance Dynamics in Multi-Agent LLM Deliberation
arXiv cs.AI

Belief Engine: Configurable and Inspectable Stance Dynamics in Multi-Agent LLM Deliberation

The article introduces the Belief Engine (BE), a new framework designed for multi-agent deliberation using large…

5/18/2026 · 3 min read · 32 views
Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute
arXiv cs.AI

Ensemble Monitoring for AI Control: Diverse Signals Outweigh More Compute

A recent study highlights the importance of ensemble monitoring for AI control, demonstrating that diverse signals are…

5/18/2026 · 3 min read · 29 views
Beyond Partner Diversity: An Influence-Based Team Steering Framework for Zero-Shot Human-Machine Teaming
arXiv cs.AI

Beyond Partner Diversity: An Influence-Based Team Steering Framework for Zero-Shot Human-Machine Teaming

The paper presents a new framework called Influence-Based Team Steering (IBTS) for enhancing human-machine teaming in…

5/18/2026 · 3 min read · 33 views
From LLM-Generated Conjectures to Lean Formalizations: Automated Polynomial Inequality Proving via Sum-of-Squares Certificates
arXiv cs.AI

From LLM-Generated Conjectures to Lean Formalizations: Automated Polynomial Inequality Proving via Sum-of-Squares Certificates

A new framework called NSPI has been proposed for automated polynomial inequality proving. This method combines large…

5/18/2026 · 3 min read · 35 views
X-SYNTH: Beyond Retrieval -- Enterprise Context Synthesis from Observed Human Attention
arXiv cs.AI

X-SYNTH: Beyond Retrieval -- Enterprise Context Synthesis from Observed Human Attention

The paper presents X-SYNTH, a framework for enterprise context synthesis based on observed human attention. It…

5/18/2026 · 3 min read · 33 views
CAPS: Cascaded Adaptive Pairwise Selection for Efficient Parallel Reasoning
arXiv cs.AI

CAPS: Cascaded Adaptive Pairwise Selection for Efficient Parallel Reasoning

The article introduces CAPS, a new framework for efficient parallel reasoning in large language models. CAPS utilizes…

5/18/2026 · 3 min read · 28 views
RTL-BenchMT: Dynamic Maintenance of RTL Generation Benchmark Through Agent-Assisted Analysis and Revision
arXiv cs.AI

RTL-BenchMT: Dynamic Maintenance of RTL Generation Benchmark Through Agent-Assisted Analysis and Revision

The paper introduces RTL-BenchMT, an automated framework for maintaining RTL generation benchmarks. It addresses…

5/18/2026 · 2 min read · 36 views
DRS-GUI: Dynamic Region Search for Training-Free GUI Grounding
arXiv cs.AI

DRS-GUI: Dynamic Region Search for Training-Free GUI Grounding

The paper presents DRS-GUI, a training-free framework for GUI grounding that enhances the performance of Multimodal…

5/18/2026 · 2 min read · 29 views
Position: Artificial Intelligence Needs Meta Intelligence -- the Case for Metacognitive AI
arXiv cs.AI

Position: Artificial Intelligence Needs Meta Intelligence -- the Case for Metacognitive AI

The paper discusses the necessity of metacognition in artificial intelligence design. It proposes that AI systems…

5/18/2026 · 3 min read · 32 views
STAR: A Stage-attributed Triage and Repair framework for RCA Agents in Microservices
arXiv cs.AI

STAR: A Stage-attributed Triage and Repair framework for RCA Agents in Microservices

The article discusses the STAR framework, which aims to enhance the reliability of root cause analysis (RCA) agents in…

5/18/2026 · 3 min read · 36 views
See Before You Code: Learning Visual Priors for Spatially Aware Educational Animation Generation
arXiv cs.AI

See Before You Code: Learning Visual Priors for Spatially Aware Educational Animation Generation

A new framework called OmniManim has been developed to improve the generation of educational animations from code.…

5/18/2026 · 3 min read · 25 views
TopoEvo: A Topology-Aware Self-Evolving Multi-Agent Framework for Root Cause Analysis in Microservices
arXiv cs.AI

TopoEvo: A Topology-Aware Self-Evolving Multi-Agent Framework for Root Cause Analysis in Microservices

The article introduces TopoEvo, a new framework designed for root cause analysis in microservices. This framework…

5/18/2026 · 3 min read · 38 views
ColPackAgent: Agent-Skill-Guided Hard-Particle Monte Carlo Workflows for Colloidal Packing
arXiv cs.AI

ColPackAgent: Agent-Skill-Guided Hard-Particle Monte Carlo Workflows for Colloidal Packing

ColPackAgent is a new framework designed to autonomously run Monte Carlo simulations for colloidal packing. It…

5/18/2026 · 3 min read · 30 views
PRISM: Prompt Reliability via Iterative Simulation and Monitoring for Enterprise Conversational AI
arXiv cs.AI

PRISM: Prompt Reliability via Iterative Simulation and Monitoring for Enterprise Conversational AI

The article introduces PRISM, a framework designed to enhance the reliability of prompts used in enterprise…

5/18/2026 · 3 min read · 25 views
Nudging Beyond the Comfort Zone: Efficient Strategy-Guided Exploration for RLVR
arXiv cs.AI

Nudging Beyond the Comfort Zone: Efficient Strategy-Guided Exploration for RLVR

The article discusses a new framework called NudgeRL for improving exploration in reinforcement learning with…

5/18/2026 · 3 min read · 23 views

How WeSearch handles this source

WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.

Indexing: Allowed Snippet: Allowed AI summary: Limited Retrieval / RAG: Not asserted Model training: Not asserted Commercial reuse: Not permitted

More ai-research sources

Visit arXiv cs.AI directly →