← all papers · overview

Towards Tracing Trustworthiness Dynamics: Revisiting Pre-training Period Of Large Language Models

Abstract

Ensuring the trustworthiness of large language models (LLMs) is crucial. Most studies concentrate on fully pre-trained LLMs to better understand and improve LLMs' trustworthiness. In this paper, to reveal the untapped potential of pre-training, we pioneer the exploration of LLMs' trustworthiness during this period, focusing on five key dimensions: reliability, privacy, toxicity, fairness, and robu

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).