WeSearch
Hub / Tags / Retraining
TAG · #RETRAINING

Retraining coverage.

Every story in the WeSearch catalog tagged with #retraining, chronological, with view counts. Subscribe to the per-tag RSS feed to follow this topic in your reader of choice.

14 stories tagged with #retraining, in publish-time order across the WeSearch catalog. Tag pages update as new stories ingest.

⌘ RSS feed for this tag →   or   search "Retraining"

RELATED TAGS
#ai2#pretraining2#andrej-karpathy1#anthropic1#openai1#research1#layoffs1#here1#major1#effort1
WASHINGTON EXAMINER

AI layoffs are here. Major retraining effort is needed

America’s economic strength has always come from its ability to transform, innovate, and adapt. From the development of the automobile to the adoption of electricity and even the i…

42 views ·
#layoffs#here#major
CRYPTO BRIEFING

MIT’s MeMo framework boosts LLM performance by 26% without retraining

MIT's MeMo framework trains a compact memory model that boosts LLM performance by up to 26.73% without retraining, with major implications for crypto AI agents.…

28 views ·
#artificial intelligence#machine learning#technology
CRYPTO BRIEFING

MIT’s MeMo boosts LLM performance by 26% without retraining

MIT and NUS researchers developed MeMo, a modular AI framework that boosts LLM performance by 26% without retraining, with major implications for crypto AI.…

24 views ·
#artificial intelligence#machine learning#technology
VENTUREBEAT

MIT's MeMo lets teams swap in a better LLM without retraining — and performance jumps 26%

44 views ·
DEV.TO (TOP)

Karpathy Joined Anthropic to Train Claude Using Claude

Andrej Karpathy joined Anthropic's pretraining team in May 2026. The specific job: use Claude to accelerate the research that makes Claude better.…

30 views ·
#ai#research#pretraining
ARXIV CS.AI

Pretraining Data Exposure in Large Language Models: A Survey of Membership Inference, Data Contamination, and Security Implications

Large Language Models (LLMs) have become the predominant paradigm in NLP, advancing both research and industry. As model sizes and pretraining data grow, concerns about Pretraining…

34 views ·
#artificial intelligence#machine learning#data privacy
ARXIV CS.AI

Memorization Dynamics of Fill-in-the-Middle Pretraining

Fill-in-the-middle (FIM) is a pretraining objective widely used to equip causal language models with infilling ability, yet its effect on verbatim memorization remains underexplore…

26 views ·
#artificial intelligence#machine learning#language models
DEV.TO (TOP)

Stop retraining YOLO: a developer’s guide to zero-shot object detection with generative VLMs

If you have ever maintained a computer vision pipeline in a factory, warehouse, or construction site,...…

35 views ·
#ai#computervision#machinelearning
ARXIV CS.AI

LLM Pretraining Shapes a Generalizable Manifold: Insights into Cross-Modal Transfer to Time Series

Can language-pretrained transformers become effective time-series forecasters, and why? In this paper, we show that cross-modal transfer arises because language pretraining precond…

35 views ·
#machine learning#artificial intelligence#time series
ARXIV CS.AI

Spectral Unforgetting: Post-Hoc Recovery of Damaged Capabilities Without Retraining

Fine-tuning a language model for a target task routinely degrades capabilities the training data never explicitly threatened. We study this phenomenon, known as catastrophic forget…

33 views ·
#machine learning#artificial intelligence#fine-tuning
QUARTZ

OpenAI co-founder Andrej Karpathy joins Anthropic's pretraining team

34 views ·
TECHCRUNCH

OpenAI co-founder Andrej Karpathy joins Anthropic’s pre-training team

Andrej Karpathy has joined Anthropic to work on pre-training. He previously co-founded and worked at OpenAI and led computer vision and AI at Tesla.…

48 views ·
#ai#andrej karpathy#anthropic
ARXIV.ORG

Alignment pretraining: AI discourse creates self-fulfilling (mis)alignment

Pretraining corpora contain extensive discourse about AI systems, yet the causal influence of this discourse on downstream alignment remains poorly understood. If prevailing descri…

24 views ·
#artificial intelligence#machine learning#language models
ARXIV CS.AI

Pretraining Objective Matters in Extreme Low-Data FGVC: A Backbone-Controlled Study

Extreme low-data fine-grained classification is common in expert domains where labeling is expensive, yet practitioners still need principled guidance for selecting pretrained enco…

32 views ·
#computer vision#artificial intelligence#machine learning