Theory-optimal Quantization Based on Flatness
A new paper introduces a novel quantization framework called Bidirectional Diagonal Quantization (BDQ) aimed at…
Recent ai-research headlines from arXiv cs.AI.

A new paper introduces a novel quantization framework called Bidirectional Diagonal Quantization (BDQ) aimed at…

A new Stability-Constrained Cardiovascular Stability Index (SCSI) has been developed to improve the estimation of…

The paper introduces PROWL, a method for enhancing world model learning through prioritized regret-driven…

The paper introduces Adaptive Multi-Scale Goodness Aggregation (AMSGA), an enhancement of the Forward-Forward…

The paper introduces RecoAtlas, a benchmark and toolkit designed for evaluating LLM recommendation agents. It…

The paper introduces a new transformer architecture called block-based double decoders. This model combines the…

The paper discusses a compositional architecture of literary primitives in instruction-tuned large language models. It…

The paper introduces a new approach called Hodge-Projected Multi-agent Learning (HPML) aimed at improving stability in…

The paper presents D-PACE, a new method for improving speculative decoding in large language models. It introduces a…

The paper discusses a new approach to memory management in diffusion world models. It proposes a framework that…

The paper discusses the role of equivariance in neural fluid surrogates, which can significantly speed up…

The paper titled 'Emergence of Frontier Superposition: Möbius attractor and Cascade Supervision' presents new findings…

The paper introduces Hybrid-LoRA, a new framework for post-training large language models. This approach combines full…

A new framework for automated benchmark generation has been introduced to improve the evaluation of foundation models.…

The article discusses a new approach to understanding attention mechanisms in machine learning models. It introduces…

The paper introduces the Bayesian Filtering Transformer (BFT), which enhances the traditional Transformer model by…

A new approach to automated data quality assessment using knowledge graph embeddings has been proposed. This method…

The paper presents VCR, a self-supervised framework designed to handle incomplete wearable signals. It aims to improve…

The paper discusses the relationship between reasoning and truthfulness in language models as they scale. It…

A new forecasting prototype has been developed to predict emergency department boarding times, which is crucial for…

The paper discusses the limitations of current leaderboard systems in evaluating frontier models in machine learning.…

A new framework for detecting money laundering in the mobility-energy supply chain has been proposed. This…

The article presents a quantitative prediction of grokking delay under the AdamW optimization algorithm. It introduces…

The article discusses a new diagnostic framework for variational autoencoders (VAEs) that addresses the issue of…

A recent study investigates how transformers build internal models while solving Sudoku puzzles. The research reveals…

The paper titled 'Exact Linear Attention' introduces a new mechanism for Transformer attention that achieves linear…

The paper introduces INSIGHTS, a model-agnostic approach for providing global explanations of time series models. It…

KadiAssistant is a new AI tool designed to enhance information retrieval within the Kadi4Mat research data ecosystem.…

The paper discusses the challenges of checkpoint selection for multimodal large language models (MLLMs) due to…

The article discusses the limitations of traditional information retrieval systems in the context of large language…

The paper discusses the phenomenon where individually calibrated models can become collectively miscalibrated in…

The article introduces TwinRouterBench, a new benchmark for evaluating LLM routing in various applications. It…

A new study proposes a Family-Grouped Hierarchical Federated Learning (Family-FL) model for privacy-preserving ECG…

The paper titled 'SAGE: Shaping Anchors for Guided Exploration in RLVR of LLMs' addresses limitations in reinforcement…

The paper discusses a method for simplifying the replacement of self-attention mechanisms in transformer models…

The paper titled FLUIDSPLAT presents a new model for reconstructing physical fields from sparse sensor data. This…

The paper presents EVA-0, a novel framework for test-time model evolution that operates with only two forward passes…

The paper introduces DarkLLM, a novel framework for generating adversarial attacks using large language models. This…

The paper introduces MO-CAPO, a novel algorithm for multi-objective prompt optimization in large language models. It…

The paper presents a novel approach to improve the reasoning capabilities of Large Language Models (LLMs) when…

EUPHORIA is a new framework designed to enhance robotic assembly in architectural construction. It addresses the…

The paper introduces AgentWall, a runtime safety layer designed for local AI agents. It addresses the critical issue…

The paper introduces ANNEAL, a neuro-symbolic agent designed to address recurring failures in LLM-based agents. It…

A new AI agent architecture has been developed to automate laboratory protocols, enhancing the efficiency and accuracy…

Skim is a new speculative execution framework designed to enhance the efficiency of web agents. By leveraging…

The paper titled 'Scalable Uncertainty Reasoning in Knowledge Graphs' by Jingcheng Wu addresses the challenges of…

A recent study examines the capabilities of large language model (LLM) agents in negotiation scenarios. While these…

The paper introduces PRISMat, a new model for material generation that is both cost-effective and…

The paper titled 'TTE-Flash: Accelerating Reasoning-based Multimodal Representations via Think-Then-Embed Tokens'…

A new research paper proposes a novel approach to ecological monitoring using knowledge-adaptive edge expert agents.…
WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.