ADR: An Agentic Detection System for Enterprise Agentic AI Security
The ADR system is a new framework designed to enhance the security of AI agents in enterprise settings. It addresses…
Recent ai-research headlines from arXiv cs.AI.

The ADR system is a new framework designed to enhance the security of AI agents in enterprise settings. It addresses…

The paper introduces Quantifying Qualitative Judgment (QQJ), a new framework for evaluating generative AI. QQJ aims to…

The paper introduces Heterogeneous Information-Bottleneck Coordination Graphs (HIBCG) for multi-agent reinforcement…

The paper discusses the intersection of token economics and AI system design, highlighting the computational…

The paper discusses multi-party multi-objective optimization problems (MPMOPs) and their unique requirements for…

A recent study explores the security vulnerabilities in multi-agent systems that utilize large language models. The…

A new study proposes a retrieval-augmented generation (RAG)-based method for translating EEG signals into text. This…

The paper introduces ResDreamer, a hierarchical world model designed for self-supervised visual reasoning in 3D…

The paper introduces MEMOIR, a memory-guided tree-search framework designed for combinatorial optimization problems.…

A new benchmark has been introduced to evaluate deep research agents (DRAs) on their ability to produce structured…

The paper discusses the performance of chess-trained language models, particularly focusing on KinGPT, a 25M-parameter…

The ECG World Model (ECG-WM) is a new framework designed to enhance the predictive simulation of cardiac…

NeuSymMS is a new hybrid neuro-symbolic memory system designed for large language model agents. It enables these…

The paper introduces AutoRubric-T2I, a novel framework for improving text-to-image (T2I) alignment through automatic…

GraphMind is a new system designed to automate complex operational workflows without human input. It constructs,…

A recent study explores the prediction of challenging behaviors in children with profound autism using wearable…

The paper presents a new memory architecture designed for Large Language Models (LLMs) to enhance their performance in…

WebGameBench is a new benchmark designed to evaluate coding agents' ability to create browser-native games from…

The article discusses a new technique called Causal Memory Intervention (CMI) designed for long-horizon LLM agents.…

The article discusses a new approach called SAPO, which stands for Step-Aligned Policy Optimization, aimed at…

The article discusses a new approach to extending cultural heritage knowledge graphs using language and vision models.…

The paper presents EGI, a multimodal emotional AI framework designed to enhance the real-time self-awareness of Scrum…

The paper introduces EXG, an experience graph framework designed for self-evolving agents. This framework allows…

The paper discusses divergence-suppressing couplings for Rectified Flow, aimed at improving trajectory generation. It…

The paper introduces HASP, a framework designed to enhance LLM agents by equipping them with executable Program…

The paper discusses the role of AI systems in experimental science, emphasizing their interaction with humans and…

A new paper presents a robust neural sparse retrieval system aimed at improving music search efficiency. The system…

The paper discusses a new approach to understanding the internal mechanisms of Large Reasoning Models (LRMs) through a…

The article introduces STRIDE, a self-reflective agent framework designed to enhance the reliability of automatic…

The PuppyChatter framework aims to simplify the development of AI applications by addressing the complexities…

The article discusses the concept of 'going headless' in vertical AI firms, where companies unbundle their services…

The paper discusses the evolving landscape of AI evaluation, particularly in the context of interactive benchmarks. It…

The paper discusses the safety risks associated with memory-equipped LLM agents over time. It highlights the concept…

The article discusses the development of the Knowledge Infrastructure for Scientific Simulation (KISS), aimed at…

The article presents a new model called PAIR, which aims to improve multi-turn agent optimization in large language…

A new study introduces ChildAgentEval, a benchmark for assessing cognitive age alignment in interactive AI agents.…

DuIVRS-2 is a new interactive voice response system designed for large-scale Point of Interest (POI) attribute…

The paper introduces LAST-RAG, a method for selecting degradation models based on observed health indicator…

The paper discusses a method for generating feedback causal fuzzy cognitive maps (FCMs) using AI agents. It explores…

The article introduces Ethical Hyper-Velocity (EHV), a new framework designed for the formal verification of AI…

SVFSearch is a new benchmark designed for short-video frame search specifically in the gaming domain. It includes a…

The paper investigates the effectiveness of supervised fine-tuning (SFT) in large language models (LLMs) compared to…

A new framework called LLM-Guided Bayesian Optimization (LGBO) has been proposed to enhance scientific discovery…

The paper presents a new algorithm called Shared Backbone Proximal Policy Optimization (Shared Backbone PPO) for…

The paper introduces TeleCom-Bench, a benchmark designed to evaluate the performance of Large Language Models (LLMs)…

A new paper presents insights into variance reduction in zero-order hard-thresholding algorithms. The proposed method…

The paper introduces DocOS, a benchmark aimed at enhancing the capabilities of GUI agents through proactive…

The paper presents a novel approach to communication in multi-agent reinforcement learning (MARL) called LLM-driven…

The article discusses a new approach to solving the Compositional Geometry Routing Problem (CGRP), which encompasses…

The paper discusses the safety challenges faced by multimodal large language models (MLLMs) in transferring safety…
WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.