LLM Pretraining Shapes a Generalizable Manifold: Insights into Cross-Modal Transfer to Time Series
The paper explores the effectiveness of language-pretrained transformers in forecasting time series data. It…
Recent ai-research headlines from arXiv cs.AI.

The paper explores the effectiveness of language-pretrained transformers in forecasting time series data. It…

The paper introduces Agentic Agile-V, a framework aimed at improving software and hardware development processes…

The research paper presents a comparative analysis of automated image segmentation architectures for predicting…

The article presents EPC-3D-Diff, a new framework for synthesizing CT images from CBCT data. This method addresses…

The article introduces DiffCodeGen, a new method for code generation that enhances efficiency through coverage-guided…

Researchers have introduced a new framework called In-context Training (ICT) for evaluating how language agents can…

The paper presents a novel approach for out-of-distribution (OOD) detection using a multi-encoder fusion of…

ShadeBench is a new benchmark dataset aimed at improving shade simulation in urban environments. It addresses the…

A new study introduces CodecAttack, an advanced method for attacking Audio Large Language Models (Audio LLMs). This…

A recent study evaluates the effectiveness of machine-learning-enhanced non-invasive testing for detecting advanced…

NeuroQA is a newly introduced benchmark aimed at enhancing visual question answering in 3D brain MRI analysis. It…

The article presents a hypothesis called collocational bootstrapping, which explores how statistical signals in…

The paper introduces the Pursuit of Subspaces (PoS) hypothesis, an axiomatic framework for understanding neural…

The paper titled 'Latent Process Generator Matching' introduces a new framework for generative models in machine…

The paper presents advancements in Visual Place Recognition (VPR) through a new method called Weighted Aggregated…

A new method for enhancing reinforcement learning in large language models (LLMs) has been proposed by researchers…

The paper introduces STORM, a new approach for managing state in multi-agent systems to enhance collaboration. STORM…

A recent study challenges the notion that self-training in language models leads to a flattening of language. Instead,…

The paper explores the specialization of experts in Vision Mixture-of-Experts (MoE) models. It highlights the…

A new paper explores lower bounds for advection-diffusion equations using AI-generated proofs. The research presents…

The paper introduces the Autoregressive Video Inverse problem Solver (AVIS), which aims to enhance the efficiency of…

The University of Florida Gators submitted a paper for the AmericasNLP 2026 shared task focusing on cultural image…

The article discusses a new method for improving text-to-image diffusion models in human portrait generation. This…

A new study reveals vulnerabilities in large language models (LLMs) related to optimization techniques. The research…

The paper introduces AVSD, a novel method for self-distillation in language models. AVSD addresses challenges related…

A new framework for pipe routing in aeroengines has been proposed, integrating manufacturability knowledge with…

The paper introduces Predicate Action Skills (PACTS), a new approach to learning complex robot behaviors. PACTS…

The paper introduces AMAR, a lightweight attention-based framework for multi-user activity recognition using Wi-Fi…

The paper introduces Reflector, a framework designed to enhance the safety of Large Language Models (LLMs) against…

A recent study evaluates the effectiveness of AI reviewers in scientific peer review. Conducted by 45 experts, the…

The article presents Dynamic TMoE, a new framework designed for non-stationary time series forecasting. This framework…

The paper titled 'DIVE: Embedding Compression via Self-Limiting Gradient Updates' introduces a new method for…

The paper presents a new method for creating interpretable text representations that are both predictive and…

The paper presents a new cryptographic protocol called Heartbeat-Bound Hierarchical Credentials (HBHC) designed for…

The paper presents Llamas on the Web (LlamaWeb), a WebGPU backend designed for efficient language model inference in…

The paper discusses improvements in Diffusion Transformers (DiTs) through a new method called Diffusion-Adaptive…

The article introduces SCRIBE, a diagnostic framework designed for evaluating automatic speech recognition (ASR) in…

The article discusses a new framework called SAVER designed for multimodal information extraction in social media. It…

The paper presents Adaptive Group Policy Optimization (AGPO), a new method for improving reinforcement learning in…

The paper discusses a new method for designing task vectors in in-context learning (ICL) that aims to improve…

The article discusses the release of TASTE, a dataset designed to evaluate AI-generated graphic design. It includes…

The paper presents a reference monitor designed to prevent data leakage from large language model (LLM) agents. It…

The paper introduces a novel reinforcement learning objective called Distribution-Aware Reward, aimed at improving…

A new paper introduces a method for evaluating reward hacking in autonomous agents. The authors propose embedding…

The paper discusses the concept of verifier strictness in generative verifiers used for step-wise verification. It…

The paper discusses the advantages of Gated Linear Units (GLU) over non-gated structures in machine learning models.…

PACD-Net is a new framework designed to improve glycemic control estimation from self-monitoring of blood glucose…

The paper discusses a new framework for correcting biases in preconditioned language model optimizers. It identifies…

The article introduces ELSA, a new architecture designed for efficient inference in spiking neural networks (SNNs).…

The paper presents Tunable MAGMAX, a model merging framework designed for continual learning (CL) that accommodates…
WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.