NLP Model Overview
Text Classification Models predict the categories a piece of text might belong to.
*all classification variant specifications apply to the NLP model type, with the addition of embeddings
Code Example
TheEmbeddingColumnNames class constructs your embedding objects. You can log them into the platform using a dictionary that maps the embedding feature names to the embedding objects. See our API reference for more details.
- Python Pandas
- Python Single Record
- UI Import JSON Input
- Import for API
Example RowNLP Embedding FeaturesArize supports logging the embedding features associated with the text the model is acting on and the text itself using the See here for more information on embeddings and options for generating them.
Google Colab
EmbeddingColumnNames object.-
The
vector_column_nameshould be the name of the column where the embedding vectors are stored. The embedding vector is the dense vector representation of the unstructured input. ⚠️ Note: embedding features are not sparse vectors. -
The
data_column_nameshould be the name of the column where the raw text associated with the vector is stored. It is the field typically chosen for NLP use cases. The column can contain both strings (full sentences) or a list of strings (token arrays).