What it is
In this study, we introduce a general method for creating byte-level large language models through byteification that approach the capabilities of subword-based systems. We use a two-stage conversion procedure to retrofit existing subword-based models into byte-level models with minimal extra training.
Every metric behind this entry is listed, with its source, under Sources and data below.
Filed underTopic Modeling, Natural Language Processing Techniques, Software Engineering Research