Multi-Agent Reinforcement Learning for Safe Autonomous Driving Under Pedestrian Behavioral Uncertainty
A new study explores the use of multi-agent reinforcement learning (MARL) to improve the safety of autonomous vehicles…
Recent ai-research headlines from arXiv cs.AI.

A new study explores the use of multi-agent reinforcement learning (MARL) to improve the safety of autonomous vehicles…

The article introduces FBOS-RL, a new framework for reinforcement learning that enhances training efficiency. It…

The paper titled 'Instance Discrimination for Link Prediction' explores the application of instance discrimination…

The paper discusses a new framework called SELFCI aimed at enhancing contextual integrity in large language models…

The paper titled 'Residual Paving: Diagnosing the Routing Bottleneck in Selective Refusal Editing' presents a new…

A new study presents a pretrained domain-adapted diffusion model for generating heterogeneous PET images from uniform…

Chronicle is a new multimodal foundation model designed for joint understanding of language and time series data. It…

The paper titled 'Catching a Moving Subspace: Low-Rank Bandits Beyond Stationarity' explores the challenges of…

The paper presents Conformal Selective Acting (CSA), a method for risk control in RLVR-trained LLMs. CSA provides a…

A new theoretical model has been proposed to enhance Out-of-Distribution (OOD) generalization in Reinforcement…

The paper introduces M-ORE, a modality-decoupled online recursive editor designed for multimodal large language models…

PolycubeNet introduces a novel dual-latent diffusion model for generating hexahedral meshes from complex CAD…

A new study presents Gated-CNN, a lightweight dual-stream architecture for fall detection using wearable devices. This…

A new framework called Trajectory-Integral Feedback GRPO (TIF-GRPO) has been proposed to enhance the accuracy of…

The paper introduces ClaimDiff-RL, a novel framework for fine-grained caption reinforcement learning that addresses…

The paper introduces Mirage, a framework for auditing visual unlearning in machine learning models. It highlights the…

The paper presents JUDO, a new framework for industrial anomaly detection that integrates domain knowledge into visual…

A new method called Introspective X Training (IXT) has been proposed to enhance the efficiency of large language model…

The article introduces FusionCell, a new dual-modality predictor for standard-cell performance prediction. It…

A new framework called Plug-and-Play Spiking Operators has been proposed to enhance the performance of spiking…

A new paper presents a method for closed-form predictive coding using hierarchical Gaussian filters. This approach…

The paper presents a framework called Quant.npu for efficient mobile NPU inference of large language models through…

The paper titled 'Spectral Unforgetting' addresses the issue of catastrophic forgetting in language models during…

The paper discusses the issue of physical misgeneralization in generative sequence models used for planning motion in…

The paper presents a robust subspace-constrained quadratic model for learning low-dimensional structures from…

The paper presents Co-Fusion4D, a new framework designed to enhance 3D object detection in autonomous driving. It…

The paper introduces Sequential Difference Maximization (SDM), a new gradient-based attack method for evaluating model…

The paper introduces Tiny-Engram, a method for enhancing generative vision models by using trigger-indexed concept…

A recent study explores the advantages of using smaller datasets for training machine learning models. The research…

The paper introduces FullFlow, a method for enhancing text-to-image flow matching models to enable bidirectional…

A new neural network framework has been developed to predict two-particle reduced density matrices (2-RDMs) with a…

The paper introduces the Target-SAT (TSAT) algorithm, which significantly improves the ability to solve random…

The paper presents a method called HF-KCU for causal unlearning in federated learning systems. This method allows for…

The paper discusses full-duplex spoken dialogue models that can listen and speak simultaneously, enhancing interaction…

The article discusses a new approach to knowledge distillation called Consistently Informative Soft-label Temperature…

A new study presents TorchSight, an open-source local system for security document classification. Built around a…

The paper introduces Digit Entropy Loss (DEL) for improving numerical learning in large language models (LLMs). It…

The paper presents a novel training strategy for multimodal semantic segmentation that addresses the issue of missing…

The article introduces SUGAR, a new framework for humanoid loco-manipulation learning that utilizes human videos. This…

The paper explores the conflict between instruction-following and pattern-completion in language models. It examines…

The paper introduces ConceptSeg-R1, a framework for segmenting concepts using meta-reinforcement learning. It…

The paper discusses a novel approach to fMRI data analysis using nonlocal operator learning. It emphasizes the…

The STELLAR model aims to enhance 3D perception for autonomous driving by scaling large models. It incorporates…

The paper discusses the MXFP4 quantization error in reinforcement learning for large language models. It identifies…

The paper discusses the challenge of class imbalance in medical image segmentation, particularly in CT body…

The paper investigates the impact of Chain-of-Thought (CoT) prompting on gender bias in large language models (LLMs).…

Computer Science > Machine Learning arXiv:2605.20440 (cs) [Submitted on 19 May 2026] Title:Group-Algebraic Tensors:…

The paper discusses weight decay regimes in transformers trained on modular arithmetic. It introduces online…

The study explores emotional dynamics in agent-to-agent interactions on the social network Moltbook. It introduces an…

The article presents a comparison of various deep learning architectures for classifying COVID-19 using CT and X-ray…
WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.