← all papers · overview

Jellyfish: A Large Language Model For Data Preprocessing

Abstract

This paper explores the utilization of LLMs for data preprocessing (DP), a crucial step in the data mining pipeline that transforms raw data into a clean format conducive to easy processing. Whereas the use of LLMs has sparked interest in devising universal solutions to DP, recent initiatives in this domain typically rely on GPT APIs, raising inevitable data breach concerns. Unlike these approache

Related papers

Ranked by semantic similarity — how closely each paper's abstract matches this one (100% = near-identical topic).