WeSearch
Hub / ai-research / arXiv cs.AI
ai-research · source

arXiv cs.AI on WeSearch

Recent ai-research headlines from arXiv cs.AI.

CoReVAD: A Contextual Reasoning Framework for Training-Free Video Anomaly Detection
arXiv cs.AI

CoReVAD: A Contextual Reasoning Framework for Training-Free Video Anomaly Detection

The paper presents CoReVAD, a training-free framework for video anomaly detection that leverages a frozen…

5/25/2026 · 3 min read · 28 views
Exploiting Longitudinal Context in Clinician-Verified Interactive Lesion Tracking
arXiv cs.AI

Exploiting Longitudinal Context in Clinician-Verified Interactive Lesion Tracking

A new approach to tracking tumor lesions in CT scans has been proposed, focusing on clinician verification to enhance…

5/25/2026 · 3 min read · 40 views
Defining AI Fatigue in Academic Contexts: Dimensions, Indicators, and a Stage-Based Model Using Grounded Theory
arXiv cs.AI

Defining AI Fatigue in Academic Contexts: Dimensions, Indicators, and a Stage-Based Model Using Grounded Theory

A recent study explores the concept of AI fatigue in academic contexts, highlighting its distinct nature compared to…

5/25/2026 · 3 min read · 34 views
Classical State Preparation for Variational Quantum Algorithms via Reinforcement Learning
arXiv cs.AI

Classical State Preparation for Variational Quantum Algorithms via Reinforcement Learning

The paper presents a new framework called CRiSP, which utilizes reinforcement learning for classical state preparation…

5/25/2026 · 3 min read · 38 views
CALAD: Channel-Aware contrastive Learning for multivariate time series Anomaly Detection
arXiv cs.AI

CALAD: Channel-Aware contrastive Learning for multivariate time series Anomaly Detection

The paper presents CALAD, a novel framework for multivariate time series anomaly detection. It emphasizes…

5/25/2026 · 3 min read · 27 views
Infra-Bayesian Reinforcement Learning Agents Outperform Classical RL For Worst-Case Robustness
arXiv cs.AI

Infra-Bayesian Reinforcement Learning Agents Outperform Classical RL For Worst-Case Robustness

A new study presents infra-Bayesian reinforcement learning agents that outperform traditional methods in scenarios…

5/25/2026 · 3 min read · 35 views
As X, Do Y: How Persona and Task Combine in Instruction-Tuned LLMs
arXiv cs.AI

As X, Do Y: How Persona and Task Combine in Instruction-Tuned LLMs

The paper discusses how persona and task can be combined in instruction-tuned language models. It highlights the…

5/25/2026 · 3 min read · 24 views
Generative AI and the Reorganization of Labor Demand
arXiv cs.AI

Generative AI and the Reorganization of Labor Demand

The paper examines how generative AI is reshaping labor demand across various sectors. It highlights that firms are…

5/25/2026 · 3 min read · 42 views
Autonomous Frontier-Based Exploration with VLM Guidance
arXiv cs.AI

Autonomous Frontier-Based Exploration with VLM Guidance

Aarush Aitha and Avideh Zakhor have proposed a new method for autonomous robotic exploration using Vision-Language…

5/25/2026 · 2 min read · 31 views
PoisonForge: Task-Level Targeted Poisoning Benchmark for Instruction-Tuned LLMs
arXiv cs.AI

PoisonForge: Task-Level Targeted Poisoning Benchmark for Instruction-Tuned LLMs

The article introduces PoisonForge, a benchmark designed to evaluate task-level targeted poisoning in…

5/25/2026 · 3 min read · 33 views
Positional Failures in Long-Context LLMs: A Blind Spot in Reasoning Benchmarks
arXiv cs.AI

Positional Failures in Long-Context LLMs: A Blind Spot in Reasoning Benchmarks

The paper discusses positional failures in long-context large language models (LLMs) and their impact on reasoning…

5/25/2026 · 3 min read · 31 views
Understanding and Improving Noisy Embedding Techniques in Instruction Finetuning
arXiv cs.AI

Understanding and Improving Noisy Embedding Techniques in Instruction Finetuning

The paper discusses advancements in instructional fine-tuning techniques, particularly focusing on noisy embeddings.…

5/25/2026 · 3 min read · 27 views
Scalable Heterogeneous Graph Foundation Models for Data-Driven Optimal Power Flow in Smart Grids
arXiv cs.AI

Scalable Heterogeneous Graph Foundation Models for Data-Driven Optimal Power Flow in Smart Grids

A new paper presents a scalable heterogeneous graph neural network workflow for optimal power flow in smart grids.…

5/25/2026 · 3 min read · 45 views
Adaptive Mass-Segmented KV Compression for Long-Context Reasoning
arXiv cs.AI

Adaptive Mass-Segmented KV Compression for Long-Context Reasoning

The paper presents a new framework called Adaptive Mass-Segmented KV Compression aimed at improving long-context…

5/25/2026 · 3 min read · 32 views
Lipschitz Optimization for Formal Verification of Homographies
arXiv cs.AI

Lipschitz Optimization for Formal Verification of Homographies

The paper presents a formal verification approach for ensuring the robustness of vision neural networks against 3D…

5/25/2026 · 3 min read · 29 views
FastKernels: Benchmarking GPU Kernel Generation in Production
arXiv cs.AI

FastKernels: Benchmarking GPU Kernel Generation in Production

FastKernels introduces a new benchmark for GPU kernel generation that addresses the misalignment between existing…

5/25/2026 · 3 min read · 36 views
PaP-NF: Probabilistic Long-Term Time Series Forecasting via Prefix-as-Prompt Reprogramming and Normalizing Flows
arXiv cs.AI

PaP-NF: Probabilistic Long-Term Time Series Forecasting via Prefix-as-Prompt Reprogramming and Normalizing Flows

The article discusses a new probabilistic forecasting framework called PaP-NF, which is designed for long-term time…

5/25/2026 · 2 min read · 21 views
Are Frontier LLMs Ready for Cybersecurity? Evidence for Vertical Foundation Models from Dual-Mode Vulnerability Benchmarks
arXiv cs.AI

Are Frontier LLMs Ready for Cybersecurity? Evidence for Vertical Foundation Models from Dual-Mode Vulnerability Benchmarks

A recent study evaluates the readiness of frontier large language models (LLMs) for cybersecurity tasks. The findings…

5/25/2026 · 3 min read · 34 views
SimInsert: Seamless Video Object Insertion via Regional Sparse Attention Fusion
arXiv cs.AI

SimInsert: Seamless Video Object Insertion via Regional Sparse Attention Fusion

The paper presents SimInsert, a novel approach to video object insertion that enhances spatio-temporal coherence and…

5/25/2026 · 3 min read · 28 views
Enhancing Deep Neural Network Reliability with Refinement and Calibration
arXiv cs.AI

Enhancing Deep Neural Network Reliability with Refinement and Calibration

A new paper proposes methods to enhance the reliability of deep neural networks (DNNs) by focusing on calibration and…

5/25/2026 · 3 min read · 28 views
Multi-Gate Residuals
arXiv cs.AI

Multi-Gate Residuals

The paper titled 'Multi-Gate Residuals' introduces a new mechanism to address the issue of unbounded activation growth…

5/25/2026 · 2 min read · 23 views
6G Communication Networks Enabling Embodied Agents: Architecture and Prototype
arXiv cs.AI

6G Communication Networks Enabling Embodied Agents: Architecture and Prototype

The article discusses the development of 6G communication networks designed to support embodied agents, which combine…

5/25/2026 · 3 min read · 28 views
Coloring the Noise: Adversarial Sobolev Alignment for Faithful Image Super Resolution
arXiv cs.AI

Coloring the Noise: Adversarial Sobolev Alignment for Faithful Image Super Resolution

The paper presents a new framework called ASASR for image super-resolution that addresses the limitations of existing…

5/25/2026 · 3 min read · 33 views
ChainFlow-VLA: Causal Flow Planning with Vision-Language Models
arXiv cs.AI

ChainFlow-VLA: Causal Flow Planning with Vision-Language Models

The article introduces ChainFlow-VLA, a new approach to causal flow planning that integrates vision-language models.…

5/25/2026 · 3 min read · 21 views
EvalVerse: Pipeline-Aware and Expert-Calibrated Benchmarking for Professional Cinematic Video Generation
arXiv cs.AI

EvalVerse: Pipeline-Aware and Expert-Calibrated Benchmarking for Professional Cinematic Video Generation

EvalVerse introduces a new evaluation framework for cinematic video generation that addresses the limitations of…

5/25/2026 · 3 min read · 32 views
When Good Equations Get Bad Scores: Improving Symbolic Regression Through Better Parameter Optimization
arXiv cs.AI

When Good Equations Get Bad Scores: Improving Symbolic Regression Through Better Parameter Optimization

The paper discusses the challenges in Symbolic Regression (SR) due to the 'Good Structure, Bad Score' phenomenon. It…

5/25/2026 · 3 min read · 26 views
Reinforcement Learning for Microcanonical Graph Ensemble with Assortativity Constraints
arXiv cs.AI

Reinforcement Learning for Microcanonical Graph Ensemble with Assortativity Constraints

A new paper introduces the Deep Microcanonical Graph Generator (DMGG), a reinforcement learning framework for…

5/25/2026 · 3 min read · 32 views
Convergence Without Understanding: When Language Models Agree on Representations but Disagree on Reasoning
arXiv cs.AI

Convergence Without Understanding: When Language Models Agree on Representations but Disagree on Reasoning

The paper explores the convergence of internal representations among large language models while highlighting their…

5/25/2026 · 3 min read · 29 views
Sparse Compositional Flow Matching by geometric assembly from motion primitives
arXiv cs.AI

Sparse Compositional Flow Matching by geometric assembly from motion primitives

The article discusses a new framework for matching embodied trajectories in robotics using motion primitives. This…

5/25/2026 · 3 min read · 28 views
CHASD: Language Increment-Calibrated Contrastive Decoding against Hallucination in LVLMs
arXiv cs.AI

CHASD: Language Increment-Calibrated Contrastive Decoding against Hallucination in LVLMs

The paper presents CHASD, a new framework designed to reduce hallucinations in Large Vision-Language Models (LVLMs).…

5/25/2026 · 3 min read · 30 views
XWind: A Cross-site Router for Large Language Model Inference Serving at Renewable Energy Farms
arXiv cs.AI

XWind: A Cross-site Router for Large Language Model Inference Serving at Renewable Energy Farms

The paper introduces XWind, a cross-site router designed for large language model inference at renewable energy farms.…

5/25/2026 · 3 min read · 31 views
Score-Based One-step MeanFlow Policy Optimization
arXiv cs.AI

Score-Based One-step MeanFlow Policy Optimization

The paper introduces Score-Based One-step MeanFlow Policy Optimization (SOM), an innovative actor-critic algorithm…

5/25/2026 · 2 min read · 39 views
Curriculum reinforcement learning with measurable task representation learning
arXiv cs.AI

Curriculum reinforcement learning with measurable task representation learning

The article discusses a novel approach to curriculum reinforcement learning (CRL) that focuses on measurable task…

5/25/2026 · 3 min read · 34 views
Metacognition as Reward: Reinforcing LLM Reasoning via Knowledge and Regulation Signals
arXiv cs.AI

Metacognition as Reward: Reinforcing LLM Reasoning via Knowledge and Regulation Signals

The paper introduces a new reinforcement learning framework called Metacognition-as-Reward (MaR) aimed at enhancing…

5/25/2026 · 3 min read · 31 views
Every Component is a Lookup: Token Attribution and Composition from a Single Decomposition
arXiv cs.AI

Every Component is a Lookup: Token Attribution and Composition from a Single Decomposition

The paper discusses a method called Unpack for mechanistic interpretability of transformers. It focuses on token…

5/25/2026 · 3 min read · 32 views
Parametric Prior Mapping Framework for Non-stationary Probabilistic Time Series Forecasting
arXiv cs.AI

Parametric Prior Mapping Framework for Non-stationary Probabilistic Time Series Forecasting

The article introduces a new framework called Parametric Prior Mapping (PPM) for forecasting non-stationary…

5/25/2026 · 3 min read · 23 views
Online Hand Gesture Recognition Using 3D Convolutional Neural Networks
arXiv cs.AI

Online Hand Gesture Recognition Using 3D Convolutional Neural Networks

A new system for online hand gesture recognition using 3D convolutional neural networks has been proposed. This system…

5/25/2026 · 2 min read · 33 views
Reflex: Reinforcement Learning with Reflection Symmetry Exploitation in State-Based Continuous Control
arXiv cs.AI

Reflex: Reinforcement Learning with Reflection Symmetry Exploitation in State-Based Continuous Control

The paper introduces Reflex, a new approach to reinforcement learning that utilizes reflection symmetry in state-based…

5/25/2026 · 3 min read · 32 views
Socially fluent AI decouples conversational signals from source identity in online interaction
arXiv cs.AI

Socially fluent AI decouples conversational signals from source identity in online interaction

A recent study explores how socially fluent AI can engage in online conversations, making it difficult for people to…

5/25/2026 · 3 min read · 30 views
SSDAU: Structured Semantic Data Augmentation for Joint Entity and Relation Extraction
arXiv cs.AI

SSDAU: Structured Semantic Data Augmentation for Joint Entity and Relation Extraction

The paper presents a novel method called Structured Semantic Data Augmentation (SSDAU) aimed at improving Joint Entity…

5/25/2026 · 3 min read · 34 views
AI Security Research Should Better Incentivize Defense Research
arXiv cs.AI

AI Security Research Should Better Incentivize Defense Research

A recent paper highlights the imbalance in AI security research, showing a predominance of studies focused on…

5/25/2026 · 2 min read · 28 views
SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation
arXiv cs.AI

SOLAR: A Self-Optimizing Open-Ended Autonomous Agent for Lifelong Learning and Continual Adaptation

The paper presents SOLAR, a self-optimizing autonomous agent designed for lifelong learning and continual adaptation.…

5/22/2026 · 3 min read · 34 views
Tool-Augmented Agent for Closed-loop Optimization,Simulation,and Modeling Orchestration
arXiv cs.AI

Tool-Augmented Agent for Closed-loop Optimization,Simulation,and Modeling Orchestration

The article discusses the introduction of COSMO-Agent, a tool-augmented reinforcement learning framework aimed at…

5/22/2026 · 2 min read · 30 views
OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind
arXiv cs.AI

OSCToM: RL-Guided Adversarial Generation for High-Order Theory of Mind

The paper presents OSCToM, a new approach for modeling nested belief conflicts in language models' Theory of Mind…

5/22/2026 · 3 min read · 30 views
AgentCo-op: Retrieval-Based Synthesis of Interoperable Multi-Agent Workflows
arXiv cs.AI

AgentCo-op: Retrieval-Based Synthesis of Interoperable Multi-Agent Workflows

The article discusses AgentCo-op, a framework designed for synthesizing interoperable multi-agent workflows. It…

5/22/2026 · 3 min read · 32 views
High Quality Embeddings for Horn Logic Reasoning
arXiv cs.AI

High Quality Embeddings for Horn Logic Reasoning

The paper discusses the development of high-quality embeddings for Horn logic reasoning. It introduces various methods…

5/22/2026 · 2 min read · 29 views
$ECUAS_n$: A family of metrics for principled evaluation of uncertainty-augmented systems
arXiv cs.AI

$ECUAS_n$: A family of metrics for principled evaluation of uncertainty-augmented systems

The paper introduces $ECUAS_n$, a new family of metrics designed for evaluating uncertainty-augmented systems in…

5/22/2026 · 3 min read · 26 views
Open-World Evaluations for Measuring Frontier AI Capabilities
arXiv cs.AI

Open-World Evaluations for Measuring Frontier AI Capabilities

The paper discusses the importance of open-world evaluations in measuring AI capabilities. It highlights the…

5/22/2026 · 3 min read · 30 views
AgentAtlas: Beyond Outcome Leaderboards for LLM Agents
arXiv cs.AI

AgentAtlas: Beyond Outcome Leaderboards for LLM Agents

The paper titled 'AgentAtlas: Beyond Outcome Leaderboards for LLM Agents' addresses the limitations of current…

5/22/2026 · 3 min read · 28 views
Personality Engineering with AI Agents: A New Methodology for Negotiation Research
arXiv cs.AI

Personality Engineering with AI Agents: A New Methodology for Negotiation Research

A new methodology called personality engineering utilizes AI agents to enhance negotiation research. This approach…

5/22/2026 · 2 min read · 28 views

How WeSearch handles this source

WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.

Indexing: Allowed Snippet: Allowed AI summary: Limited Retrieval / RAG: Not asserted Model training: Not asserted Commercial reuse: Not permitted

More ai-research sources

Visit arXiv cs.AI directly →