WeSearch
Hub / ai-research / arXiv cs.AI
ai-research · source

arXiv cs.AI on WeSearch

Recent ai-research headlines from arXiv cs.AI.

Theory-optimal Quantization Based on Flatness
arXiv cs.AI

Theory-optimal Quantization Based on Flatness

A new paper introduces a novel quantization framework called Bidirectional Diagonal Quantization (BDQ) aimed at…

5/20/2026 · 3 min read · 24 views
A Nonlinear Complexity Index for Wearable PPG Cardiovascular Stability: Multiscale Validation, Systematic Evaluation Correction, and Bayesian Parameter Optimization
arXiv cs.AI

A Nonlinear Complexity Index for Wearable PPG Cardiovascular Stability: Multiscale Validation, Systematic Evaluation Correction, and Bayesian Parameter Optimization

A new Stability-Constrained Cardiovascular Stability Index (SCSI) has been developed to improve the estimation of…

5/20/2026 · 3 min read · 32 views
PROWL: Prioritized Regret-Driven Optimization for World Model Learning
arXiv cs.AI

PROWL: Prioritized Regret-Driven Optimization for World Model Learning

The paper introduces PROWL, a method for enhancing world model learning through prioritized regret-driven…

5/20/2026 · 3 min read · 28 views
Adaptive Multi-Scale Goodness Aggregation for Forward-Forward Learning
arXiv cs.AI

Adaptive Multi-Scale Goodness Aggregation for Forward-Forward Learning

The paper introduces Adaptive Multi-Scale Goodness Aggregation (AMSGA), an enhancement of the Forward-Forward…

5/20/2026 · 2 min read · 29 views
RecoAtlas: From Semantic Plausibility to Set-Level Utility in LLM Recommendation Agents
arXiv cs.AI

RecoAtlas: From Semantic Plausibility to Set-Level Utility in LLM Recommendation Agents

The paper introduces RecoAtlas, a benchmark and toolkit designed for evaluating LLM recommendation agents. It…

5/20/2026 · 3 min read · 29 views
Block-Based Double Decoders
arXiv cs.AI

Block-Based Double Decoders

The paper introduces a new transformer architecture called block-based double decoders. This model combines the…

5/20/2026 · 2 min read · 36 views
Compositional Literary Primitives in Instruction-Tuned LLMs: Cross-Architectural SAE Features for Self, Style, and Affect
arXiv cs.AI

Compositional Literary Primitives in Instruction-Tuned LLMs: Cross-Architectural SAE Features for Self, Style, and Affect

The paper discusses a compositional architecture of literary primitives in instruction-tuned large language models. It…

5/20/2026 · 3 min read · 24 views
Metric-Gradient Projection for Stable Multi-Agent Policy Learning
arXiv cs.AI

Metric-Gradient Projection for Stable Multi-Agent Policy Learning

The paper introduces a new approach called Hodge-Projected Multi-agent Learning (HPML) aimed at improving stability in…

5/20/2026 · 3 min read · 23 views
D-PACE: Dynamic Position-Aware Cross-Entropy for Parallel Speculative Drafting
arXiv cs.AI

D-PACE: Dynamic Position-Aware Cross-Entropy for Parallel Speculative Drafting

The paper presents D-PACE, a new method for improving speculative decoding in large language models. It introduces a…

5/20/2026 · 3 min read · 30 views
Composition of Memory Experts for Diffusion World Models
arXiv cs.AI

Composition of Memory Experts for Diffusion World Models

The paper discusses a new approach to memory management in diffusion world models. It proposes a framework that…

5/20/2026 · 3 min read · 34 views
Symmetry in the Wild: The Role of Equivariance in Neural Fluid Surrogates
arXiv cs.AI

Symmetry in the Wild: The Role of Equivariance in Neural Fluid Surrogates

The paper discusses the role of equivariance in neural fluid surrogates, which can significantly speed up…

5/20/2026 · 3 min read · 31 views
Emergence of Frontier Superposition: M\"obius attractor and Cascade Supervision
arXiv cs.AI

Emergence of Frontier Superposition: M\"obius attractor and Cascade Supervision

The paper titled 'Emergence of Frontier Superposition: Möbius attractor and Cascade Supervision' presents new findings…

5/20/2026 · 3 min read · 31 views
Hybrid-LoRA: Bridging Full Fine-Tuning and Low-Rank Adaptation for Post-Training
arXiv cs.AI

Hybrid-LoRA: Bridging Full Fine-Tuning and Low-Rank Adaptation for Post-Training

The paper introduces Hybrid-LoRA, a new framework for post-training large language models. This approach combines full…

5/20/2026 · 3 min read · 27 views
Fine-Grained Benchmark Generation for Comprehensive Evaluation of Foundation Models
arXiv cs.AI

Fine-Grained Benchmark Generation for Comprehensive Evaluation of Foundation Models

A new framework for automated benchmark generation has been introduced to improve the evaluation of foundation models.…

5/20/2026 · 3 min read · 26 views
The Routing and Filtering Structure of Attention
arXiv cs.AI

The Routing and Filtering Structure of Attention

The article discusses a new approach to understanding attention mechanisms in machine learning models. It introduces…

5/20/2026 · 3 min read · 27 views
Precision Tracked Transformer via Kalman Filtering, Kriging and Process Noise
arXiv cs.AI

Precision Tracked Transformer via Kalman Filtering, Kriging and Process Noise

The paper introduces the Bayesian Filtering Transformer (BFT), which enhances the traditional Transformer model by…

5/20/2026 · 3 min read · 33 views
Automated Big Data Quality Assessment using Knowledge Graph Embeddings
arXiv cs.AI

Automated Big Data Quality Assessment using Knowledge Graph Embeddings

A new approach to automated data quality assessment using knowledge graph embeddings has been proposed. This method…

5/20/2026 · 3 min read · 32 views
VCR: Learning Valid Contextual Representation for Incomplete Wearable Signals
arXiv cs.AI

VCR: Learning Valid Contextual Representation for Incomplete Wearable Signals

The paper presents VCR, a self-supervised framework designed to handle incomplete wearable signals. It aims to improve…

5/20/2026 · 3 min read · 31 views
Lying Is Just a Phase: The Hidden Alignment Transition in Language Model Scaling
arXiv cs.AI

Lying Is Just a Phase: The Hidden Alignment Transition in Language Model Scaling

The paper discusses the relationship between reasoning and truthfulness in language models as they scale. It…

5/20/2026 · 3 min read · 22 views
An Integrated Forecasting Prototype for Emergency Department Boarding Time to Support Proactive Operational Decision Making
arXiv cs.AI

An Integrated Forecasting Prototype for Emergency Department Boarding Time to Support Proactive Operational Decision Making

A new forecasting prototype has been developed to predict emergency department boarding times, which is crucial for…

5/20/2026 · 3 min read · 30 views
The Growing Pains of Frontier Models: When Leaderboards Stop Separating and What to Measure Next
arXiv cs.AI

The Growing Pains of Frontier Models: When Leaderboards Stop Separating and What to Measure Next

The paper discusses the limitations of current leaderboard systems in evaluating frontier models in machine learning.…

5/20/2026 · 3 min read · 38 views
Graph-Driven Cross-Industry Real-Time Monitoring Framework for Anti-Money Laundering Detection in Converged Mobility-Energy Supply Chain Networks
arXiv cs.AI

Graph-Driven Cross-Industry Real-Time Monitoring Framework for Anti-Money Laundering Detection in Converged Mobility-Energy Supply Chain Networks

A new framework for detecting money laundering in the mobility-energy supply chain has been proposed. This…

5/20/2026 · 3 min read · 33 views
First-Passage Prediction of Grokking Delay: ACalibrated Law under AdamW with Causal Validation
arXiv cs.AI

First-Passage Prediction of Grokking Delay: ACalibrated Law under AdamW with Causal Validation

The article presents a quantitative prediction of grokking delay under the AdamW optimization algorithm. It introduces…

5/20/2026 · 3 min read · 31 views
Lost and Found in Translation: Variational Diagnostics for Neural Codebook Channels
arXiv cs.AI

Lost and Found in Translation: Variational Diagnostics for Neural Codebook Channels

The article discusses a new diagnostic framework for variational autoencoders (VAEs) that addresses the issue of…

5/20/2026 · 3 min read · 35 views
Transformers Linearly Represent Highly Structured World Models
arXiv cs.AI

Transformers Linearly Represent Highly Structured World Models

A recent study investigates how transformers build internal models while solving Sudoku puzzles. The research reveals…

5/20/2026 · 3 min read · 31 views
Exact Linear Attention
arXiv cs.AI

Exact Linear Attention

The paper titled 'Exact Linear Attention' introduces a new mechanism for Transformer attention that achieves linear…

5/20/2026 · 2 min read · 28 views
INSIGHTS: Demonstration-Based Summaries of Time Series Predictors
arXiv cs.AI

INSIGHTS: Demonstration-Based Summaries of Time Series Predictors

The paper introduces INSIGHTS, a model-agnostic approach for providing global explanations of time series models. It…

5/20/2026 · 3 min read · 35 views
KadiAssistant: A conversational AI Agent for information retrieval in Kadi4Mat
arXiv cs.AI

KadiAssistant: A conversational AI Agent for information retrieval in Kadi4Mat

KadiAssistant is a new AI tool designed to enhance information retrieval within the Kadi4Mat research data ecosystem.…

5/20/2026 · 3 min read · 32 views
Robust Checkpoint Selection for Multimodal LLMs via Agentic Evaluation and Stability-Aware Ranking
arXiv cs.AI

Robust Checkpoint Selection for Multimodal LLMs via Agentic Evaluation and Stability-Aware Ranking

The paper discusses the challenges of checkpoint selection for multimodal large language models (MLLMs) due to…

5/20/2026 · 2 min read · 34 views
The 99% Success Paradox: When Near-Perfect Retrieval Equals Random Selection
arXiv cs.AI

The 99% Success Paradox: When Near-Perfect Retrieval Equals Random Selection

The article discusses the limitations of traditional information retrieval systems in the context of large language…

5/20/2026 · 3 min read · 27 views
When Individually Calibrated Models Become Collectively Miscalibrated
arXiv cs.AI

When Individually Calibrated Models Become Collectively Miscalibrated

The paper discusses the phenomenon where individually calibrated models can become collectively miscalibrated in…

5/20/2026 · 3 min read · 24 views
TwinRouterBench: Fast Static and Live Dynamic Evaluation for Realistic Agentic LLM Routing
arXiv cs.AI

TwinRouterBench: Fast Static and Live Dynamic Evaluation for Realistic Agentic LLM Routing

The article introduces TwinRouterBench, a new benchmark for evaluating LLM routing in various applications. It…

5/20/2026 · 3 min read · 29 views
Towards Family-Grouped Hierarchical Federated Learning on Sub-5KB Models: A Feasibility Study of Privacy-Preserving ECG Monitoring for Ultra-Resource-Constrained Wearables
arXiv cs.AI

Towards Family-Grouped Hierarchical Federated Learning on Sub-5KB Models: A Feasibility Study of Privacy-Preserving ECG Monitoring for Ultra-Resource-Constrained Wearables

A new study proposes a Family-Grouped Hierarchical Federated Learning (Family-FL) model for privacy-preserving ECG…

5/20/2026 · 3 min read · 33 views
SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs
arXiv cs.AI

SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs

The paper titled 'SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs' addresses limitations in reinforcement…

5/20/2026 · 3 min read · 32 views
From Sparsity to Simplicity: Enabling Simpler Sequential Replacements via Sparse Attention Distillation
arXiv cs.AI

From Sparsity to Simplicity: Enabling Simpler Sequential Replacements via Sparse Attention Distillation

The paper discusses a method for simplifying the replacement of self-attention mechanisms in transformer models…

5/20/2026 · 3 min read · 32 views
FLUIDSPLAT: Reconstructing Physical Fields from Sparse Sensors via Gaussian Primitives
arXiv cs.AI

FLUIDSPLAT: Reconstructing Physical Fields from Sparse Sensors via Gaussian Primitives

The paper titled FLUIDSPLAT presents a new model for reconstructing physical fields from sparse sensor data. This…

5/20/2026 · 3 min read · 29 views
EVA-0: Test-Time Model Evolution with Only Two Forward Passes per Sample
arXiv cs.AI

EVA-0: Test-Time Model Evolution with Only Two Forward Passes per Sample

The paper presents EVA-0, a novel framework for test-time model evolution that operates with only two forward passes…

5/20/2026 · 3 min read · 17 views
DarkLLM: Learning Language-Driven Adversarial Attacks with Large Language Models
arXiv cs.AI

DarkLLM: Learning Language-Driven Adversarial Attacks with Large Language Models

The paper introduces DarkLLM, a novel framework for generating adversarial attacks using large language models. This…

5/20/2026 · 3 min read · 28 views
MO-CAPO: Multi-Objective Cost-Aware Prompt Optimization
arXiv cs.AI

MO-CAPO: Multi-Objective Cost-Aware Prompt Optimization

The paper introduces MO-CAPO, a novel algorithm for multi-objective prompt optimization in large language models. It…

5/20/2026 · 3 min read · 22 views
Distributional Energy-Based Models for Uncertainty-Aware Structured LLM Reasoning
arXiv cs.AI

Distributional Energy-Based Models for Uncertainty-Aware Structured LLM Reasoning

The paper presents a novel approach to improve the reasoning capabilities of Large Language Models (LLMs) when…

5/20/2026 · 3 min read · 23 views
EUPHORIA: Efficient Universal Planning via Hybrid Optimization for Robust Industrial Robotic Assembly
arXiv cs.AI

EUPHORIA: Efficient Universal Planning via Hybrid Optimization for Robust Industrial Robotic Assembly

EUPHORIA is a new framework designed to enhance robotic assembly in architectural construction. It addresses the…

5/20/2026 · 3 min read · 28 views
AgentWall: A Runtime Safety Layer for Local AI Agents
arXiv cs.AI

AgentWall: A Runtime Safety Layer for Local AI Agents

The paper introduces AgentWall, a runtime safety layer designed for local AI agents. It addresses the critical issue…

5/19/2026 · 3 min read · 33 views
ANNEAL: Adapting LLM Agents via Governed Symbolic Patch Learning
arXiv cs.AI

ANNEAL: Adapting LLM Agents via Governed Symbolic Patch Learning

The paper introduces ANNEAL, a neuro-symbolic agent designed to address recurring failures in LLM-based agents. It…

5/19/2026 · 3 min read · 31 views
From Prompts to Protocols: An AI Agent for Laboratory Automation
arXiv cs.AI

From Prompts to Protocols: An AI Agent for Laboratory Automation

A new AI agent architecture has been developed to automate laboratory protocols, enhancing the efficiency and accuracy…

5/19/2026 · 3 min read · 27 views
Skim: Speculative Execution for Fast and Efficient Web Agents
arXiv cs.AI

Skim: Speculative Execution for Fast and Efficient Web Agents

Skim is a new speculative execution framework designed to enhance the efficiency of web agents. By leveraging…

5/19/2026 · 3 min read · 27 views
Scalable Uncertainty Reasoning in Knowledge Graphs
arXiv cs.AI

Scalable Uncertainty Reasoning in Knowledge Graphs

The paper titled 'Scalable Uncertainty Reasoning in Knowledge Graphs' by Jingcheng Wu addresses the challenges of…

5/19/2026 · 2 min read · 34 views
Counterparty Modeling is Not Strategy: The Limits of LLM Negotiators
arXiv cs.AI

Counterparty Modeling is Not Strategy: The Limits of LLM Negotiators

A recent study examines the capabilities of large language model (LLM) agents in negotiation scenarios. While these…

5/19/2026 · 3 min read · 30 views
PRISMat: Policy-Driven, Permutation-Invariant Autoregressive Material Generation
arXiv cs.AI

PRISMat: Policy-Driven, Permutation-Invariant Autoregressive Material Generation

The paper introduces PRISMat, a new model for material generation that is both cost-effective and…

5/19/2026 · 3 min read · 29 views
TTE-Flash: Accelerating Reasoning-based Multimodal Representations via Think-Then-Embed Tokens
arXiv cs.AI

TTE-Flash: Accelerating Reasoning-based Multimodal Representations via Think-Then-Embed Tokens

The paper titled 'TTE-Flash: Accelerating Reasoning-based Multimodal Representations via Think-Then-Embed Tokens'…

5/19/2026 · 3 min read · 31 views
Sustainable Intelligence for the Wild: Democratizing Ecological Monitoring via Knowledge-Adaptive Edge Expert Agents
arXiv cs.AI

Sustainable Intelligence for the Wild: Democratizing Ecological Monitoring via Knowledge-Adaptive Edge Expert Agents

A new research paper proposes a novel approach to ecological monitoring using knowledge-adaptive edge expert agents.…

5/19/2026 · 3 min read · 26 views

How WeSearch handles this source

WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.

Indexing: Allowed Snippet: Allowed AI summary: Limited Retrieval / RAG: Not asserted Model training: Not asserted Commercial reuse: Not permitted

More ai-research sources

Visit arXiv cs.AI directly →