https://alignmentpretraining.ai/
Alignment Pretraining: AI Discourse Causes Self-Fulfilling (Mis)alignment
LLMs trained on data about misaligned AIs themselves become less aligned. Luckily, pretraining LLMs with synthetic data about good AIs helps them become more...
alignmentpretrainingdiscoursecausesself
https://proceedings.neurips.cc/paper_files/paper/2022/hash/e9882f7f7c44a10acc01132302bac9d8-Abstract-Conference.html
PyramidCLIP: Hierarchical Feature Alignment for Vision-language Model Pretraining
language modelfeaturealignmentvisionpretraining
https://openreview.net/forum?id=PpSDVE5rAy
TiC-LM: A Multi-Year Benchmark for Continual Pretraining of Language Models | OpenReview
Large language models (LLMs) are trained on data crawled over many years from the web. We investigate how quickly LLMs become outdated over time and how to...
language modelsticlmmultiyear
https://www.osti.gov/pages/biblio/2476269-qarr-fsqa-question-answer-replacement-removal-pretraining-framework-few-shot-question-answering
QARR-FSQA: Question-Answer Replacement and Removal Pretraining Framework for Few-Shot Question...
The U.S. Department of Energy's Office of Scientific and Technical Information
question answerfsqareplacementremovalpretraining
https://github.com/deep-symbolic-mathematics/Multimodal-Math-Pretraining
GitHub - deep-symbolic-mathematics/Multimodal-Math-Pretraining: [ICLR 2024 Spotlight] This is the...
[ICLR 2024 Spotlight] This is the official code for the paper "SNIP: Bridging Mathematical Symbolic and Numeric Realms with Unified Pre-training" -...
symbolic mathematicsgithubdeepmultimodalpretraining
https://www.ijcai.org/proceedings/2024/129
Self-Promoted Clustering-based Contrastive Learning for Brain Networks Pretraining | IJCAI
Electronic proceedings of IJCAI 2024
learning forselfpromotedclusteringbased
https://aclanthology.org/2021.emnlp-main.249/
Frustratingly Simple Pretraining Alternatives to Masked Language Modeling - ACL Anthology
Atsuki Yamaguchi, George Chrysostomou, Katerina Margatina, Nikolaos Aletras. Proceedings of the 2021 Conference on Empirical Methods in Natural Language...
alternatives tolanguage modelingsimplepretrainingmasked
https://www.proceedings.com/079017-3453.html
MATES: Model-Aware Data Selection for Efficient Pretraining with Data Influence Models -...
The world's premier source for conference proceedings, offering Print-on-Demand, DOI, and Content Hosting services.
matesmodelawaredataselection
https://rock-the-prototype.com/en/tag/pretraining-en/
PreTraining Archive - Rock the Prototype - Softwareentwicklung & Prototyping
pretrainingarchiverockprototypesoftwareentwicklung
https://deepai.org/publication/composer-style-classification-of-piano-sheet-music-images-using-language-model-pretraining
Composer Style Classification of Piano Sheet Music Images Using Language Model Pretraining | DeepAI
Jul 29, 2020 - 07/29/20 - This paper studies composer style classification of piano sheet music images. Previous approaches to the composer classification t...
piano sheet musiclanguage modelcomposerstyleclassification
https://www.getorchestra.io/guides/data_and_ai_glossary_pretraining_llms
Pretraining Large Language Models: Key Concepts and Processes | Orchestra
Understand the critical steps in pretraining large language models (LLMs), focusing on data processing, neural network training, and AI-driven analysis for...
large language modelskey conceptspretrainingprocessesorchestra
https://research.google/pubs/analyzing-similarity-metrics-for-data-selection-for-language-model-pretraining/
Analyzing Similarity Metrics for Data Selection for Language Model Pretraining
language modelanalyzingsimilaritymetricsdata
https://www.resumemate.io/jobs/ai-engineer-jobs-in-san-jose/7e7a1e_figureai_4671704006
Helix AI Engineer, Pretraining at Figure | ResumeMate
Apply for Helix AI Engineer, Pretraining at Figure in San Jose, CA. Skills: Python, PyTorch, Machine Learning. Apply now on ResumeMate.
ai engineerhelixpretrainingfigure
https://oecd.ai/en/catalogue/metric-use-cases/mtp-advancing-remote-sensing-foundation-model-via-multi-task-pretraining
MTP: Advancing Remote Sensing Foundation Model via Multi-Task Pretraining - OECD.AI
We propose a novel model-selection method for dynamic real-life networks. Our approach involves training a classifier on a large body of synthetic network...
remote sensingfoundation modelmtpadvancingvia
https://indico.cern.ch/event/1386125/contributions/6139661/
ML4Jets2024 (4-November 8, 2024): Large-Scale Pretraining and Finetuning for Efficient Jet...
The workshop is organised in a hybrid format (see zoom links at the bottom right of this page, visible only to registered participants). We expect speakers to...
large scalenovemberpretrainingfinetuningefficient
https://bytez.com/docs/arxiv/1812.10860/paper
Can You Tell Me How to Get Past Sesame Street? Sentence-Level Pretraining Beyond Language Modeling...
Dec 28, 2018 - Natural language understanding has recently seen a surge of progress with the use of sentence encoders like ELMo (Peters et al., 2018a) and BERT (Devlin et...
tell me howto getsesame streetlanguage modelingpast
https://www.isca-archive.org/interspeech_2023/feng23_interspeech.html
ISCA Archive - Language-Universal Phonetic Representation in Multilingual Speech Pretraining for...
isca archivelanguageuniversalphoneticrepresentation
https://digitalcommons.providence.org/publications/11218/
"Pretraining Patient Foundation Models on Multimodal Patient Journeys" by Daniel P Jeong, Suhana...
By Daniel P Jeong, Suhana Bedi, Cliff Wong, et al., Published on 09/23/25
foundation modelspretrainingpatientmultimodaljourneys
https://huggingface.co/papers/2505.22232
Paper page - Judging Quality Across Languages: A Multilingual Approach to Pretraining Data...
Join the discussion on this paper page
paperjudgingqualityacrosslanguages
https://proceedings.iclr.cc/paper_files/paper/2025/hash/45d74e190008c7bff2845ffc8e3facd3-Abstract-Conference.html
Latent Action Pretraining from Videos
latentactionpretrainingvideos
https://ai.updf.com/paper-detail/moco-cxr-moco-pretraining-improves-representation-and-transferability-of-chest-sowrirajan-yang-5985e77cce2a6a3720682c81dbeaef7b417e0459
MoCo-CXR: MoCo Pretraining Improves Representation and Transferability of Chest X-ray Models
In detecting pleural effusion, it is found that linear models trained on MoCo-CXR-pretrained representations outperform those without Mo co-C XR- pretrained...
x raymocopretrainingimprovesrepresentation
https://publications-cnrc.canada.ca/eng/view/object/?id=ab800d88-7e47-4470-b78d-43446b99a020
Molecular geometry pretraining with SE(3)-invariant denoising distance matching - NRC Publications...
Molecular geometry pretraining with SE(3)-invariant denoising distance matching
moleculargeometrypretrainingseinvariant
https://news.smol.ai/tags/pretraining
Topic: pretraining | AINews
Posts tagged with topic: pretraining
topicpretrainingainews
https://liner.com/review/astt5-structureaware-pretraining-for-code-generation-and-understanding
AST-T5: Structure-Aware Pretraining for Code Generation and Understanding [Quick Review]
Regarding this ICML 2024 paper, this review summarizes AST-T5, a pretraining paradigm leveraging Abstract Syntax Trees for enhanced code generation and und...
code generationquick reviewaststructureaware
https://www.educative.io/courses/generative-ai-essentials/pretraining-paradigms
Understanding Pretraining Paradigms in Foundation AI Models
Learn about key pretraining methods like autoregressive, masked, and contrastive learning shaping foundation models such as GPT and BERT.
ai modelsunderstandingpretrainingparadigmsfoundation
https://lrec.elra.info/lrec2024-main-0031
A CURATEd CATalog: Rethinking the Extraction of Pretraining Corpora for Mid-Resourced Languages -...
curated catalogrethinkingextractionpretrainingcorpora
https://tldr.takara.ai/p/2510.03264
Front-Loading Reasoning: The Synergy between Pretraining and Post-Training Data | Takara TLDR
The prevailing paradigm for enhancing the reasoning abilities of LLMs revolves around post-training on high-quality, reasoning-intensive data. While emerging...
front loadingpost trainingreasoningsynergypretraining
https://arxiv.org/abs/2512.06104v1
[2512.06104v1] ARC-AGI Without Pretraining
Abstract page for arXiv paper 2512.06104v1: ARC-AGI Without Pretraining
arcagiwithoutpretraining
https://www.moaijobs.com/job/helix-ai-engineer-pretraining-figure-1550
Helix AI Engineer, Pretraining
Figure is hiring a Helix AI Engineer, Pretraining.
ai engineerhelixpretraining