REFORGE: A Method for Benchmarking LLMs' Reverse Engineering Capabilities in Decompiled Binary Function Naming
Computer Science > Software Engineering arXiv:2607.07738 (cs) [Submitted on 7 Jul 2026] Title:REFORGE: A Method for…
Recent ai-research headlines from arXiv cs.AI.

Computer Science > Software Engineering arXiv:2607.07738 (cs) [Submitted on 7 Jul 2026] Title:REFORGE: A Method for…

In this paper, we propose a unified approach to explore the common mechanism of various KD methods using interactions.…

Predicting AD conversion during the prodromal stage remains critical for disease understanding and patient care. As…

Computer Science > Machine Learning arXiv:2607.08779 (cs) [Submitted on 12 Jun 2026] Title:Signed Symmetric…

Existing remedies are either system-level (caching heuristics) or post-hoc (router fine-tuning), leaving the root…

We show that this coupling can instead serve as an alignment interface: by matching noise and data according to a…

Its efficiency depends on the communication and computation latencies of the GPUs, which are linked to the placement…

Recent advances have extended Deep Neural Networks (DNNs) to operate on manifolds, accompanied by normalization…

Computer Science > Machine Learning arXiv:2607.08784 (cs) [Submitted on 13 Jun 2026] Title:HERO: A Heterogeneity-Aware…

Recognizing the value of data, data transactions are increasingly common, giving rise to many data marketplaces, e.g.,…

Pruning techniques that introduce sparsity into weight matrices can accelerate inference. However, maintaining model…

We extend the LLaMEA framework to MOBO, using large language models as mutation and crossover operators within…

Existing reinforcement learning (RL) approaches learn fixed strategies for sepsis treatment, limiting adaptability to…

We present a diffusion-based synthesis pipeline for low-resource sand-boil imagery. Using Stable Diffusion XL…

However, existing biological resources, such as molecular databases, protein repositories, genomic annotations,…

Standard methods inject stochasticity in the action space, but such jitter only yields rollouts close to the original.…

Using a novel architecture inspired by previous QCNN and classical convolutional neural network (CNN) implementations,…

We present Eluna, a production-deployed agentic system for reliable SOP execution. Eluna is a graph-guided,…

When a specification admits multiple readings but the supervision channel does not reveal which is operative,…

We introduce MultiView-Bench, a diagnostic benchmark expressly designed to evaluate multi-view integration for…

Computer Science > Robotics arXiv:2607.08974 (cs) [Submitted on 9 Jul 2026] Title:CLAP: Direct VLM-to-VLA Adaptation…

Computer Science > Software Engineering arXiv:2607.08981 (cs) [Submitted on 9 Jul 2026] Title:The Patchwork Problem in…

Currently, mitigating this premature termination requires continuous human-in-the-loop supervision. This heavy…

We study this gap in two oracle-evaluable domains with contrasting structure: Connect Four, a solved partisan game…

These models often encode domain-specific knowledge into their graph encoding modules, which increases their parameter…

Unlike classical contextual bandits that rely solely on bandit feedback and assume conditional independence across…

Mortensen View a PDF of the paper titled Phone Segmentation and Recognition through Phonological Activation Mapping,…

What, then, is the equivalent catalyst needed to achieve a general-purpose model in computer vision? In this paper, we…

Evolutionary computation (EC) provides a computational basis for feedback-driven discovery because population-based…

We argue for the opposite order of explanation in a finite and fully computable setting. The free orthomodular lattice…

This makes human vision distinctly different from most popular computer vision models in use today, which input images…

SE) has continuously evolved through increasingly powerful forms of reuse, from source code and libraries to…

A critical limitation is identified in many document understanding benchmarks: visual content is often reducible to…

Current approaches for automatic precedent retrieval map legal documents to a low-dimensional semantic space and…

Selecting representative training subsets, however, remains challenging: individual sample contributions are unclear,…

Once those metadata disappear, clinically critical failure modes can be masked by strong aggregate performance, and…

Therefore, semi-supervised approaches such as Graph Convolutional Networks (GCNs), which learn from both labeled and…

To address these limitations, we propose EVAD, an event enhanced VAD framework that jointly exploits conventional…

Securities and Exchange Commission (SEC) which can be found in EDGAR. We were preprocessing those data and than…

While explicit goals may render certain actions optimal, implicit social norms often impose hidden constraints.…

On-policy distillation (OPD) provides dense teacher guidance and typically improves rapidly in the early stage, but…

The paper introduces a unified training paradigm that equips large language model agents with internal world modeling…

Planning, a core component of intelligent behavior, remains challenging for LLMs, which often produce infeasible or…

The paper presents Tree of Evidence (ToE), a hierarchical framework for automated claim verification that builds…

Specifically, for reasoning-based MLLMs, fast thinking by triggering direct answers often outperforms slow thinking…

However, their lived experiences with these tools remain largely underexamined. This paper proposes DysLexLens, a…

Computers create the Internet, and the Internet empowers the value of computers. The rapid development of the…

Computer Science > Artificial Intelligence arXiv:2606.27443 (cs) [Submitted on 25 Jun 2026] Title:When Does…

A foundry is an organized sheaf of knowledge that carries within it an argumentation component. Concrete foundries are…

An agent-based world model calls an LLM API and reasons flexibly in language, but its errors appear as hallucinated…
WeSearch's declared handling of arXiv cs.AI's content. Indexing, snippets, summaries, retrieval and training are separate questions — see the rights registry or read this source's machine-readable record.