Entity Embeddings : Perspectives Towards an Omni-Modality Era for Large
Language Models
Eren Unlu; Unver Ciftci
arXiv2023
20
ciftci2023entity
Abstract
Large Language Models (LLMs) are evolving to integrate multiple modalities,
such as text, image, and audio into a unified linguistic space. We envision a
future direction based on this framework where conceptual entities defined in
sequences of text can also be imagined as modalities. Such a formulation has
the potential to overcome the cognitive and computational limitations of
current models. Several illustrative examples of such potential implicit
modalities are given. Along with vast promises of the hypothesized structure,
expected challenges are discussed as well.