WeSearch
Hub / ai-research / arXiv cs.AI
ai-research · source

arXiv cs.AI on WeSearch

Recent ai-research headlines from arXiv cs.AI.

Beyond Control-Flow: Integrating the Resource Perspective into Multi-Collaborative Process Modeling from Text
arXiv cs.AI

Beyond Control-Flow: Integrating the Resource Perspective into Multi-Collaborative Process Modeling from Text

The paper discusses a novel approach to process modeling in Business Process Management (BPM) that integrates resource…

5/26/2026 · 3 min read · 33 views
PALoRA: Projection-Adaptive LoRA for Preserving Reasoning in Large Language Models
arXiv cs.AI

PALoRA: Projection-Adaptive LoRA for Preserving Reasoning in Large Language Models

The paper introduces PALoRA, a framework designed to enhance the integration of new knowledge into Large Language…

5/26/2026 · 3 min read · 35 views
Jailbreak to Protect: Buffering and Reinforcing via Temporary Jailbreaking for Safe Fine-Tuning in Large Language Models
arXiv cs.AI

Jailbreak to Protect: Buffering and Reinforcing via Temporary Jailbreaking for Safe Fine-Tuning in Large Language Models

The paper discusses a new framework for safe fine-tuning of large language models (LLMs) called Buffer-and-Reinforce.…

5/26/2026 · 3 min read · 38 views
Summoning the Oracle to Slay It: Mitigating Look-Ahead Bias in Financial Backtesting with Large Language Models
arXiv cs.AI

Summoning the Oracle to Slay It: Mitigating Look-Ahead Bias in Financial Backtesting with Large Language Models

A new paper proposes a method to mitigate look-ahead bias in financial backtesting using large language models. The…

5/26/2026 · 3 min read · 40 views
Associations between echocardiographic traits and AI-ECG predictions of heart failure
arXiv cs.AI

Associations between echocardiographic traits and AI-ECG predictions of heart failure

A recent study explored the relationship between echocardiographic traits and AI-ECG predictions of heart failure. The…

5/26/2026 · 3 min read · 27 views
HeartBeatAI: An Interpretable and Robust Deep Learning Framework for Multi-Label ECG Arrhythmia Detection
arXiv cs.AI

HeartBeatAI: An Interpretable and Robust Deep Learning Framework for Multi-Label ECG Arrhythmia Detection

HeartBeatAI is a new deep learning framework designed for multi-label ECG arrhythmia detection. It addresses…

5/26/2026 · 2 min read · 28 views
Learning to Reason Efficiently with A* Post-Training
arXiv cs.AI

Learning to Reason Efficiently with A* Post-Training

A recent study explores the use of A* search algorithms to improve reasoning in large language models (LLMs). The…

5/26/2026 · 3 min read · 32 views
Hera: Learning Long-Horizon Coordination for Device-Cloud Collaborative LLM Agents
arXiv cs.AI

Hera: Learning Long-Horizon Coordination for Device-Cloud Collaborative LLM Agents

The article discusses a new approach called Hera for coordinating device-cloud collaborative large language model…

5/26/2026 · 3 min read · 35 views
Agent-as-Peer-Debriefer: A Multi-Agent Framework with Perspective-Based Refinement for Qualitative Analysis
arXiv cs.AI

Agent-as-Peer-Debriefer: A Multi-Agent Framework with Perspective-Based Refinement for Qualitative Analysis

The article presents a new framework called Agent-as-Peer-Debriefer designed to enhance qualitative data analysis…

5/26/2026 · 3 min read · 34 views
Lattice theory and algebraic models for deep convolutional learning based on mathematical morphology
arXiv cs.AI

Lattice theory and algebraic models for deep convolutional learning based on mathematical morphology

A new paper presents an algebraic framework for deep convolutional learning based on lattice theory and mathematical…

5/26/2026 · 3 min read · 33 views
GlobalDentBench: A Multinational Benchmark for Evaluating LLM Clinical Reasoning in Dentistry with Expert Calibration
arXiv cs.AI

GlobalDentBench: A Multinational Benchmark for Evaluating LLM Clinical Reasoning in Dentistry with Expert Calibration

GlobalDentBench is introduced as the first multinational benchmark for evaluating large language models (LLMs) in…

5/26/2026 · 4 min read · 40 views
AVBench: Human-Aligned and Automated Evaluation Benchmark for Audio-Video Generative Models
arXiv cs.AI

AVBench: Human-Aligned and Automated Evaluation Benchmark for Audio-Video Generative Models

AVBench is a newly introduced benchmark aimed at improving the evaluation of audio-video generative models,…

5/26/2026 · 3 min read · 37 views
Beyond Inference-Only Deployment: Comparing Weight-Based Consolidation Against Cascading Compaction
arXiv cs.AI

Beyond Inference-Only Deployment: Comparing Weight-Based Consolidation Against Cascading Compaction

The article discusses a new approach to deploying large language models (LLMs) that goes beyond inference-only…

5/26/2026 · 3 min read · 35 views
Measuring Reasoning Quality in LLMs: A Multi-Dimensional Behavioral Framework
arXiv cs.AI

Measuring Reasoning Quality in LLMs: A Multi-Dimensional Behavioral Framework

A new study proposes a multi-dimensional framework for evaluating reasoning quality in large language models (LLMs).…

5/26/2026 · 3 min read · 29 views
When Mean CE Fails: Median CE Can Better Track Language Model Quality
arXiv cs.AI

When Mean CE Fails: Median CE Can Better Track Language Model Quality

The paper discusses the limitations of mean cross-entropy (CE) as a metric for evaluating language model quality. It…

5/26/2026 · 3 min read · 26 views
Exploration of Perceptual Speech Features for Clinical Decision-Support in Mental Health Care
arXiv cs.AI

Exploration of Perceptual Speech Features for Clinical Decision-Support in Mental Health Care

A new study explores the use of perceptual speech features to support clinical decision-making in mental health care.…

5/26/2026 · 2 min read · 30 views
Emotional intelligence in large language models is fragmented across perception, cognition, and interaction
arXiv cs.AI

Emotional intelligence in large language models is fragmented across perception, cognition, and interaction

A recent study highlights the fragmented nature of emotional intelligence in large language models (LLMs). The…

5/26/2026 · 3 min read · 40 views
MDIA: A Multi-Agent Diagnostic Intelligence Pipeline on HealthBench Professional
arXiv cs.AI

MDIA: A Multi-Agent Diagnostic Intelligence Pipeline on HealthBench Professional

The article discusses the MDIA, a Multi-Agent Diagnostic Intelligence Pipeline designed for clinical reasoning. It…

5/26/2026 · 3 min read · 33 views
Fundamental Limitation in Explaining AI
arXiv cs.AI

Fundamental Limitation in Explaining AI

A recent paper discusses the inherent limitations in explaining AI systems, particularly large-scale models. The…

5/26/2026 · 3 min read · 31 views
Hylos: Operability Contracts for Model-Native Spatial Intelligence
arXiv cs.AI

Hylos: Operability Contracts for Model-Native Spatial Intelligence

The paper titled 'Hylos: Operability Contracts for Model-Native Spatial Intelligence' introduces a new systems…

5/26/2026 · 3 min read · 26 views
Automated Detection and Classification of Delusion-related Content in Naturalistic Audio Diaries Using Multi-Agent Language Models
arXiv cs.AI

Automated Detection and Classification of Delusion-related Content in Naturalistic Audio Diaries Using Multi-Agent Language Models

A new study presents an automated pipeline using multi-agent language models to detect and classify delusion-related…

5/26/2026 · 3 min read · 31 views
Proper Scoring Rules for Agentic Uncertainty Quantification
arXiv cs.AI

Proper Scoring Rules for Agentic Uncertainty Quantification

The paper introduces the Trajectory Proper Score (TPS) for evaluating agentic uncertainty quantification in AI. It…

5/26/2026 · 3 min read · 23 views
Uncertainty Decomposition via Cyclical SG-MCMC and Soft-label Learning for Subjective NLP
arXiv cs.AI

Uncertainty Decomposition via Cyclical SG-MCMC and Soft-label Learning for Subjective NLP

The paper discusses a novel approach to uncertainty decomposition in subjective natural language processing (NLP). It…

5/26/2026 · 2 min read · 34 views
PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback
arXiv cs.AI

PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and Convergent Feedback

The paper titled 'PRIMA: Operational Patterns for Resilient Multi-Agent Research with Verifiable Identity and…

5/26/2026 · 3 min read · 27 views
GRAIL: AI translation for scientists application workflow on satellite data
arXiv cs.AI

GRAIL: AI translation for scientists application workflow on satellite data

The paper introduces GRAIL, an AI translation system designed to assist scientists in converting Python geospatial…

5/26/2026 · 2 min read · 44 views
PANDO: Efficient Multimodal AI Agents via Online Skill Distillation
arXiv cs.AI

PANDO: Efficient Multimodal AI Agents via Online Skill Distillation

The paper introduces PANDO, a framework designed to enhance the efficiency of multimodal AI agents through online…

5/26/2026 · 3 min read · 34 views
CoRe-Code: Collaborative Reinforcement Learning for Code Generation
arXiv cs.AI

CoRe-Code: Collaborative Reinforcement Learning for Code Generation

The paper introduces CoRe-Code, a framework for collaborative reinforcement learning aimed at improving code…

5/26/2026 · 3 min read · 30 views
Agent Manufacturing: Foundation-Model Agents as First-Class Industrial Entities
arXiv cs.AI

Agent Manufacturing: Foundation-Model Agents as First-Class Industrial Entities

The article discusses a new paradigm in manufacturing called Agent Manufacturing, which focuses on the role of…

5/26/2026 · 2 min read · 22 views
Test-Time Deep Thinking to Explore Implicit Rules
arXiv cs.AI

Test-Time Deep Thinking to Explore Implicit Rules

A new framework called Test-Time Exploration (TTExplore) aims to improve the performance of intelligent agents in…

5/26/2026 · 3 min read · 27 views
Geo-Expert: Towards Expert-Level Geological Reasoning via Parameter-Efficient Fine-Tuning
arXiv cs.AI

Geo-Expert: Towards Expert-Level Geological Reasoning via Parameter-Efficient Fine-Tuning

The paper introduces Geo-Expert, a series of parameter-efficient geological language models designed to improve…

5/26/2026 · 2 min read · 30 views
Solving Combinatorial Counting Problems with Weighted First-Order Model Counting
arXiv cs.AI

Solving Combinatorial Counting Problems with Weighted First-Order Model Counting

The paper presents a new approach to solving combinatorial counting problems using a language called Cofola. This…

5/26/2026 · 3 min read · 33 views
Clustering as Reasoning: A $k$-Means Interpretation of Chain-of-Thought Graph Learning
arXiv cs.AI

Clustering as Reasoning: A $k$-Means Interpretation of Chain-of-Thought Graph Learning

The paper presents a novel approach to Chain-of-Thought (CoT) graph learning by interpreting it through the lens of…

5/26/2026 · 3 min read · 34 views
Inverting the Shield: Systematically Generating Safety Tests from Policy Specifications
arXiv cs.AI

Inverting the Shield: Systematically Generating Safety Tests from Policy Specifications

A new framework called POLARIS has been introduced to enhance safety testing for Large Language Models (LLMs). This…

5/26/2026 · 3 min read · 32 views
TaBIIC2: Interactive Building of Ontological Taxonomies using Weighted Self-Organizing Maps
arXiv cs.AI

TaBIIC2: Interactive Building of Ontological Taxonomies using Weighted Self-Organizing Maps

The paper presents TaBIIC2, a tool designed for the interactive construction of ontological taxonomies using weighted…

5/26/2026 · 3 min read · 31 views
ProActor: Timing-Aware Reinforcement Learning for Proactive Task Scheduling Agents
arXiv cs.AI

ProActor: Timing-Aware Reinforcement Learning for Proactive Task Scheduling Agents

The paper introduces ProActor, a framework for proactive task scheduling using timing-aware reinforcement learning. It…

5/26/2026 · 3 min read · 32 views
Noise-Robust Financial Numerical Entity Attribute Tagging
arXiv cs.AI

Noise-Robust Financial Numerical Entity Attribute Tagging

The paper introduces a new method called NORA for improving the understanding of financial numerical entities in…

5/26/2026 · 3 min read · 33 views
Energy Shields for Fairness
arXiv cs.AI

Energy Shields for Fairness

The paper titled 'Energy Shields for Fairness' introduces a novel approach to ensuring runtime fairness in…

5/26/2026 · 3 min read · 34 views
Towards Multi-Turn Dialog Systems for Industrial Asset Operations and Maintenance
arXiv cs.AI

Towards Multi-Turn Dialog Systems for Industrial Asset Operations and Maintenance

The paper presents a multi-turn dialog system tailored for industrial asset operations and maintenance. This system…

5/26/2026 · 2 min read · 34 views
Mitigating Object Hallucinations in Vision-Language Models through Region-Aware Attention Recalibration
arXiv cs.AI

Mitigating Object Hallucinations in Vision-Language Models through Region-Aware Attention Recalibration

A new paper addresses the issue of object hallucination in Large Vision-Language Models (LVLMs). The authors propose a…

5/26/2026 · 3 min read · 32 views
NeurIPS: Neuro-anatomical Inductive Priors for Sphere-based Brain Decoding
arXiv cs.AI

NeurIPS: Neuro-anatomical Inductive Priors for Sphere-based Brain Decoding

The paper presents NeurIPS, a framework designed to enhance surface-based brain decoding by utilizing neuro-anatomical…

5/26/2026 · 3 min read · 37 views
Privacy-Preserving Local Language Models for Longitudinal Data Retrieval in Chronic Dermatologic Disease: Implementation in Pemphigus Patients
arXiv cs.AI

Privacy-Preserving Local Language Models for Longitudinal Data Retrieval in Chronic Dermatologic Disease: Implementation in Pemphigus Patients

A study evaluated the use of a privacy-preserving small language model (SLM) for retrieving clinical data in pemphigus…

5/26/2026 · 3 min read · 36 views
AION: Next-Generation Tasks and Practical Harness for Time Series
arXiv cs.AI

AION: Next-Generation Tasks and Practical Harness for Time Series

The paper titled 'AION: Next-Generation Tasks and Practical Harness for Time Series' presents a new framework for time…

5/26/2026 · 3 min read · 27 views
Evolutionary Enhanced Multi-Agent Reinforcement Learning for Cooperative Air Combat
arXiv cs.AI

Evolutionary Enhanced Multi-Agent Reinforcement Learning for Cooperative Air Combat

The paper presents a new framework for multi-agent reinforcement learning in cooperative air combat scenarios. It…

5/26/2026 · 2 min read · 34 views
RECTOR: Priority-Aware Rule-Based Reranking for Compliance-Aware Autonomous Driving Trajectory Selection
arXiv cs.AI

RECTOR: Priority-Aware Rule-Based Reranking for Compliance-Aware Autonomous Driving Trajectory Selection

The paper presents RECTOR, a rule-based reranking system designed for autonomous driving trajectory selection. It…

5/26/2026 · 3 min read · 35 views
Trust but Verify: Prover-Verifier Deliberation for Selective LLM Prediction
arXiv cs.AI

Trust but Verify: Prover-Verifier Deliberation for Selective LLM Prediction

The paper introduces a new protocol called prover-verifier deliberation (PVD) for improving the reliability of…

5/26/2026 · 3 min read · 32 views
Beyond the Frontier: Stochastic Backtracking for Efficient Test-Time Scaling
arXiv cs.AI

Beyond the Frontier: Stochastic Backtracking for Efficient Test-Time Scaling

The paper introduces a method called stochastic backtracking for improving test-time scaling in language models. This…

5/26/2026 · 3 min read · 33 views
Representation Without Control: Testing the Realization Effect in Language Models
arXiv cs.AI

Representation Without Control: Testing the Realization Effect in Language Models

The paper titled 'Representation Without Control: Testing the Realization Effect in Language Models' explores the…

5/26/2026 · 3 min read · 34 views
SimuWoB: Simulating Real-World Mobile Apps for Fast and Faithful GUI Agent Benchmarking
arXiv cs.AI

SimuWoB: Simulating Real-World Mobile Apps for Fast and Faithful GUI Agent Benchmarking

The paper introduces SimuWoB, a synthetic benchmark designed for evaluating mobile GUI agents. It addresses the…

5/26/2026 · 3 min read · 35 views
SpecAlign: A Semantic Alignment Framework for SystemVerilog Assertion Generation
arXiv cs.AI

SpecAlign: A Semantic Alignment Framework for SystemVerilog Assertion Generation

The paper introduces SpecAlign, a framework designed to enhance the semantic alignment of SystemVerilog Assertions…

5/26/2026 · 2 min read · 29 views
DarkForest: Less Talk, Higher Accuracy for Multi-Agent LLMs
arXiv cs.AI

DarkForest: Less Talk, Higher Accuracy for Multi-Agent LLMs

The paper presents DarkForest, a new framework aimed at improving the accuracy of multi-agent large language models…

5/26/2026 · 3 min read · 26 views

How WeSearch handles this source

WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.

Indexing: Allowed Snippet: Allowed AI summary: Limited Retrieval / RAG: Not asserted Model training: Not asserted Commercial reuse: Not permitted

More ai-research sources

Visit arXiv cs.AI directly →