Point me towards some basic dataset preparation tips for LLM's?

ArtifartX · 1 year ago

Point me towards some basic dataset preparation tips for LLM's?

ArtifartX · 1 year ago

Yea, doing this is part of what spurred the question, because I began to notice some datasets that were very clean and ordered into data pairs, and others that seemed formatted differently, and others still that seemed like they were fed a massive chunk of unstructured text. It made me confused on if there were some sort of standards or not that I was not aware of.