Cheap and reliable Node.js hosting starts at $3/month, and $1/month static HTML hosting

Created with love in Canada, visit hostnodejs.com today

Feel like to post an Ad? Learn Details

All Projects → explosion → Spacy Streamlit

explosion / Spacy Streamlit

Licence: mit

👑 spaCy building blocks and visualizers for Streamlit apps

Programming Languages

python

139335 projects - #7 most used programming language

Labels

machine-learning nlp natural-language-processing text-classification named-entity-recognition ner spacy visualizer

Projects that are alternatives of or similar to Spacy Streamlit

Spacy

💫 Industrial-strength Natural Language Processing (NLP) in Python

Stars: ✭ 21,978 (+6005%)

Mutual labels: natural-language-processing, text-classification, named-entity-recognition, spacy

Spacy Lookup

Named Entity Recognition based on dictionaries

Stars: ✭ 212 (-41.11%)

Mutual labels: natural-language-processing, named-entity-recognition, ner, spacy

Bond

BOND: BERT-Assisted Open-Domain Name Entity Recognition with Distant Supervision

Stars: ✭ 96 (-73.33%)

Mutual labels: natural-language-processing, named-entity-recognition, ner

Ncrfpp

NCRF++, a Neural Sequence Labeling Toolkit. Easy use to any sequence labeling tasks (e.g. NER, POS, Segmentation). It includes character LSTM/CNN, word LSTM/CNN and softmax/CRF components.

Stars: ✭ 1,767 (+390.83%)

Mutual labels: natural-language-processing, named-entity-recognition, ner

Bert Sklearn

a sklearn wrapper for Google's BERT model

Stars: ✭ 182 (-49.44%)

Mutual labels: natural-language-processing, named-entity-recognition, ner

Bertweet

BERTweet: A pre-trained language model for English Tweets (EMNLP-2020)

Stars: ✭ 282 (-21.67%)

Mutual labels: text-classification, named-entity-recognition, ner

Text Analytics With Python

Learn how to process, classify, cluster, summarize, understand syntax, semantics and sentiment of text data with the power of Python! This repository contains code and datasets used in my book, "Text Analytics with Python" published by Apress/Springer.

Stars: ✭ 1,132 (+214.44%)

Mutual labels: natural-language-processing, text-classification, spacy

Spark Nlp

State of the Art Natural Language Processing

Stars: ✭ 2,518 (+599.44%)

Mutual labels: natural-language-processing, named-entity-recognition, text-classification

Pytorch Bert Crf Ner

KoBERT와 CRF로 만든 한국어 개체명인식기 (BERT+CRF based Named Entity Recognition model for Korean)

Stars: ✭ 236 (-34.44%)

Mutual labels: natural-language-processing, named-entity-recognition, ner

Snips Nlu

Snips Python library to extract meaning from text

Stars: ✭ 3,583 (+895.28%)

Mutual labels: text-classification, named-entity-recognition, ner

Vncorenlp

A Vietnamese natural language processing toolkit (NAACL 2018)

Stars: ✭ 354 (-1.67%)

Mutual labels: natural-language-processing, named-entity-recognition, ner

Chatbot ner

chatbot_ner: Named Entity Recognition for chatbots.

Stars: ✭ 273 (-24.17%)

Mutual labels: natural-language-processing, named-entity-recognition, ner

Entity Recognition Datasets

A collection of corpora for named entity recognition (NER) and entity recognition tasks. These annotated datasets cover a variety of languages, domains and entity types.

Stars: ✭ 891 (+147.5%)

Mutual labels: natural-language-processing, named-entity-recognition, ner

Turkish Bert Nlp Pipeline

Bert-base NLP pipeline for Turkish, Ner, Sentiment Analysis, Question Answering etc.

Stars: ✭ 85 (-76.39%)

Mutual labels: natural-language-processing, named-entity-recognition, ner

Hanlp

中文分词词性标注命名实体识别依存句法分析成分句法分析语义依存分析语义角色标注指代消解风格转换语义相似度新词发现关键词短语提取自动摘要文本分类聚类拼音简繁转换自然语言处理

Stars: ✭ 24,626 (+6740.56%)

Mutual labels: natural-language-processing, text-classification, named-entity-recognition

Spacy Course

👩‍🏫 Advanced NLP with spaCy: A free online course

Stars: ✭ 1,920 (+433.33%)

Mutual labels: natural-language-processing, named-entity-recognition, spacy

Dan Jurafsky Chris Manning Nlp

My solution to the Natural Language Processing course made by Dan Jurafsky, Chris Manning in Winter 2012.

Stars: ✭ 124 (-65.56%)

Mutual labels: text-classification, named-entity-recognition, ner

Kashgari

Kashgari is a production-level NLP Transfer learning framework built on top of tf.keras for text-labeling and text-classification, includes Word2Vec, BERT, and GPT2 Language Embedding.

Stars: ✭ 2,235 (+520.83%)

Mutual labels: text-classification, named-entity-recognition, ner

anonymization-api

How to build and deploy an anonymization API with FastAPI

Stars: ✭ 51 (-85.83%)

Mutual labels: spacy, named-entity-recognition, ner

NER-and-Linking-of-Ancient-and-Historic-Places

An NER tool for ancient place names based on Pleiades and Spacy.

Stars: ✭ 26 (-92.78%)

Mutual labels: spacy, named-entity-recognition, ner

View All Similar Projects ➔

spacy-streamlit: spaCy building blocks for Streamlit apps

This package contains utilities for visualizing spaCy models and building interactive spaCy-powered apps with Streamlit. It includes various building blocks you can use in your own Streamlit app, like visualizers for syntactic dependencies, named entities, text classification, semantic similarity via word vectors, token attributes, and more.

🚀 Quickstart

You can install spacy-streamlit from pip:

pip install spacy-streamlit --pre

The package includes building blocks that call into Streamlit and set up all the required elements for you. You can either use the individual components directly and combine them with other elements in your app, or call the visualize function to embed the whole visualizer.

Download the English model from spaCy to get started.

python -m spacy download en_core_web_sm

Then put the following example code in a file.

# streamlit_app.py
import spacy_streamlit

models = ["en_core_web_sm", "en_core_web_md"]
default_text = "Sundar Pichai is the CEO of Google."
spacy_streamlit.visualize(models, default_text)

You can then run your app with streamlit run streamlit_app.py. The app should pop up in your web browser. 😀

📦 Example: `01_out-of-the-box.py`

Use the embedded visualizer with custom settings out-of-the-box.

streamlit run https://raw.githubusercontent.com/explosion/spacy-streamlit/master/examples/01_out-of-the-box.py

👑 Example: `02_custom.py`

Use individual components in your existing app.

streamlit run https://raw.githubusercontent.com/explosion/spacy-streamlit/master/examples/02_custom.py

🎛 API

Visualizer components

These functions can be used in your Streamlit app. They call into streamlit under the hood and set up the required elements.

`function` `visualize`

Embed the full visualizer with selected components.

import spacy_streamlit

models = ["en_core_web_sm", "/path/to/model"]
default_text = "Sundar Pichai is the CEO of Google."
visualizers = ["ner", "textcat"]
spacy_streamlit.visualize(models, default_text, visualizers)

Argument	Type	Description
`models`	List[str] / Dict[str, str]	Names of loadable spaCy models (paths or package names). The models become selectable via a dropdown. Can either be a list of names or the names mapped to descriptions to display in the dropdown.
`default_text`	str	Default text to analyze on load. Defaults to `""`.
`default_model`	Optional[str]	Optional name of default model. If not set, the first model in the list of `models` is used.
`visualizers`	List[str]	Names of visualizers to show. Defaults to `["parser", "ner", "textcat", "similarity", "tokens"]`.
`ner_labels`	Optional[List[str]]	NER labels to include. If not set, all labels present in the `"ner"` pipeline component will be used.
`ner_attrs`	List[str]	Span attributes shown in table of named entities. See `visualizer.py` for defaults.
`token_attrs`	List[str]	Token attributes to show in token visualizer. See `visualizer.py` for defaults.
`similarity_texts`	Tuple[str, str]	The default texts to compare in the similarity visualizer. Defaults to `("apple", "orange")`.
`show_json_doc`	bool	Show button to toggle JSON representation of the `Doc`. Defaults to `True`.
`show_meta`	bool	Show button to toggle `meta.json` of the current pipeline. Defaults to `True`.
`show_config`	bool	Show button to toggle `config.cfg` of the current pipeline. Defaults to `True`.
`show_visualizer_select`	bool	Show sidebar dropdown to select visualizers to display (based on enabled visualizers). Defaults to `False`.
`sidebar_title`	Optional[str]	Title shown in the sidebar. Defaults to `None`.
`sidebar_description`	Optional[str]	Description shown in the sidebar. Accepts Markdown-formatted text.
`show_logo`	bool	Show the spaCy logo in the sidebar. Defaults to `True`.
`color`	Optional[str]	Experimental: Primary color to use for some of the main UI elements (`None` to disable hack). Defaults to `"#09A3D5"`.
`get_default_text`	Callable[[Language], str]	Optional callable that takes the currently loaded `nlp` object and returns the default text. Can be used to provide language-specific default texts. If the function returns `None`, the value of `default_text` is used, if available. Defaults to `None`.

`function` `visualize_parser`

Visualize the dependency parse and part-of-speech tags using spaCy's displacy visualizer.

import spacy
from spacy_streamlit import visualize_parser

nlp = spacy.load("en_core_web_sm")
doc = nlp("This is a text")
visualize_parser(doc)

Argument	Type	Description
`doc`	`Doc`	The spaCy `Doc` object to visualize.
keyword-only
`title`	Optional[str]	Title of the visualizer block.
`sidebar_title`	Optional[str]	Title of the config settings in the sidebar.

`function` `visualize_ner`

Visualize the named entities in a Doc using spaCy's displacy visualizer.

import spacy
from spacy_streamlit import visualize_ner

nlp = spacy.load("en_core_web_sm")
doc = nlp("Sundar Pichai is the CEO of Google.")
visualize_ner(doc, labels=nlp.get_pipe("ner").labels)

Argument	Type	Description
`doc`	`Doc`	The spaCy `Doc` object to visualize.
keyword-only
`labels`	Sequence[str]	The labels to show in the labels dropdown.
`attrs`	List[str]	The span attributes to show in entity table.
`show_table`	bool	Whether to show a table of entities and their attributes. Defaults to `True`.
`title`	Optional[str]	Title of the visualizer block.
`sidebar_title`	Optional[str]	Title of the config settings in the sidebar.
`colors`	Dict[str,str]	A dictionary mapping labels to display colors ({"LABEL": "COLOR"})

`function` `visualize_textcat`

Visualize text categories predicted by a trained text classifier.

import spacy
from spacy_streamlit import visualize_textcat

nlp = spacy.load("./my_textcat_model")
doc = nlp("This is a text about a topic")
visualize_textcat(doc)

Argument	Type	Description
`doc`	`Doc`	The spaCy `Doc` object to visualize.
keyword-only
`title`	Optional[str]	Title of the visualizer block.

`visualize_similarity`

Visualize semantic similarity using the model's word vectors. Will show a warning if no vectors are present in the model.

import spacy
from spacy_streamlit import visualize_similarity

nlp = spacy.load("en_core_web_lg")
visualize_similarity(nlp, ("pizza", "fries"))

Argument	Type	Description
`nlp`	`Language`	The loaded `nlp` object with vectors.
`default_texts`	Tuple[str, str]	The default texts to compare on load. Defaults to `("apple", "orange")`.
keyword-only
`threshold`	float	Threshold for what's considered "similar". If the similarity score is greater than the threshold, the result is shown as similar. Defaults to `0.5`.
`title`	Optional[str]	Title of the visualizer block.

`function` `visualize_tokens`

Visualize the tokens in a Doc and their attributes.

import spacy
from spacy_streamlit import visualize_tokens

nlp = spacy.load("en_core_web_sm")
doc = nlp("This is a text")
visualize_tokens(doc, attrs=["text", "pos_", "dep_", "ent_type_"])

Argument	Type	Description
`doc`	`Doc`	The spaCy `Doc` object to visualize.
keyword-only
`attrs`	List[str]	The names of token attributes to use. See `visualizer.py` for defaults.
`title`	Optional[str]	Title of the visualizer block.

Cached helpers

These helpers attempt to cache loaded models and created Doc objects.

`function` `process_text`

Process a text with a model of a given name and create a Doc object. Calls into the load_model helper to load the model.

import streamlit as st
from spacy_streamlit import process_text

spacy_model = st.sidebar.selectbox("Model name", ["en_core_web_sm", "en_core_web_md"])
text = st.text_area("Text to analyze", "This is a text")
doc = process_text(spacy_model, text)

Argument	Type	Description
`model_name`	str	Loadable spaCy model name. Can be path or package name.
`text`	str	The text to process.
RETURNS	`Doc`	The processed document.

`function` `load_model`

Load a spaCy model from a path or installed package and return a loaded nlp object.

import streamlit as st
from spacy_streamlit import load_model

spacy_model = st.sidebar.selectbox("Model name", ["en_core_web_sm", "en_core_web_md"])
nlp = load_model(spacy_model)

Argument	Type	Description
`name`	str	Loadable spaCy model name. Can be path or package name.
RETURNS	`Language`	The loaded `nlp` object.

Note that the project description data, including the texts, logos, images, and/or trademarks, for each open source project belongs to its rightful owner. If you wish to add or remove any projects, please contact us at [email protected].

Stars: ✭ 360

Visit Git Page 🔗Visit User Page 🔗Visit Issues Page (4) 🔗

Cheap and reliable Node.js hosting starts at $3/month, and $1/month static HTML hosting

explosion / Spacy Streamlit

Programming Languages

Labels

Projects that are alternatives of or similar to Spacy Streamlit

spacy-streamlit: spaCy building blocks for Streamlit apps

🚀 Quickstart

📦 Example: 01_out-of-the-box.py

👑 Example: 02_custom.py

🎛 API

Visualizer components

function visualize

function visualize_parser

function visualize_ner

function visualize_textcat

visualize_similarity

function visualize_tokens

Cached helpers

function process_text

function load_model

📦 Example: `01_out-of-the-box.py`

👑 Example: `02_custom.py`

`function` `visualize`

`function` `visualize_parser`

`function` `visualize_ner`

`function` `visualize_textcat`

`visualize_similarity`

`function` `visualize_tokens`

`function` `process_text`

`function` `load_model`