Skip to main content
Home
Contact us
Watch Demo
Contact us
Watch Demo
  • Agentic AI
  • AI Governance
  • Airflow
  • Anaconda
  • Apache Spark
  • Artificial Intelligence
  • Clustering
  • Dask
  • Data Science
Domino's logo

Who is Domino?

Domino Data Lab empowers the largest AI-driven enterprises to build and operate AI at scale. Domino’s Enterprise AI Platform provides an integrated experience encompassing model development, MLOps, collaboration, and governance. With Domino, global enterprises can develop better medicines, grow more productive crops, develop more competitive products, and more. Founded in 2013, Domino is backed by Sequoia Capital, Coatue Management, NVIDIA, Snowflake, and other leading investors.

Watch Demo
  • Platform

      • AI infrastructure
      • Data management
      • AI workbench
      • MLOps
      • AI governance
      • FinOps
      • Pricing
      • Security & compliance
      • What's new
  • Solutions

    • Industries

      • Life sciences
      • Finance
      • Public sector
      • Retail
      • Manufacturing
    • Use Cases

      • Generative AI
      • Cost-effective data science
      • Self-service data science
      • Model risk management
      • Cloud data science
  • Learn

      • Events
      • Blog
      • Podcast
      • Courses and certifications
      • Data Science Dictionary
      • Documentation
      • Support
      • Demo hub
  • Company

      • About
      • Why Domino
      • Careers
      • News and press
      • Partners
      • Customers
      • Contact us

© 2026 Domino Data Lab, Inc. Made in San Francisco.

  • Do not sell my personal information
  • Privacy policy
  • Terms and conditions
  • Security
  • Legal
  • Density-based clustering
  • dplyr
  • Factor analysis
  • Feature
  • Feature Engineering
  • Feature Extraction
  • Feature selection
  • Folium
  • GenomicRanges
  • ggmap
  • ggplot
  • Ground Truth
  • Hash table
  • Hyperparameter Tuning
  • Interpretability
  • Jupyter Notebook
  • Kubernetes
  • LLMOps
  • Machine Learning
  • Machine Learning Algorithms
  • MLOps
  • Model Drift
  • Model Evaluation
  • Model monitoring
  • Model Selection
  • Model Tuning
  • Overfitting
  • Plotly
  • PySpark
  • PyTorch
  • Responsible AI
  • Shiny (in R)
  • sklearn
  • spaCy
  • SR 26-2
  • Statistical Computing Environment (SCE)
  • TensorFlow
  • Underfitting
  • XGBoost
  • spaCy

    What is spaCy?

    spaCy is a free, open-source Python library that provides advanced capabilities to conduct natural language processing (NLP) on large volumes of text at high speed. It helps you build models and production applications that can underpin document analysis, chatbot capabilities, and all other forms of text analysis.

    The two principal authors for spaCy, Matthew Honnibal and Ines Montani, launched the project in 2015. The spaCy framework—along with a growing set of plug-ins and other integrations—provides features for a wide range of natural language tasks. It’s become one of the most widely used natural language libraries in Python for industry use cases, and has quite a large community—and with that, much support for commercialization of research advances as this area continues to evolve rapidly.

    Example of using spacy nlp package
    Example of using spacy nlp package

    Source: spacy.io

    The spaCy Universe offers deep-dives into particular use cases and to see how this field is evolving. Some selections from this “universe” include:

    • Blackstone – parsing unstructured legal texts
    • Kindred – extracting entities from biomedical texts (e.g., Pharma)
    • mordecai – parsing geographic information
    • Prodigy – human-in-the-loop annotation for labeling datasets
    • Rasa NLU – Rasa integration for chat apps
    • spacy-pytorch-transformers to fine-tune (i.e., use transfer learning with) BERT, GPT-2, XLNet, etc.

    SpaCy 3.0

    The latest release, spaCy 3.0, brings many improvements to help build, configure and maintain NLP models, including:

    • Newly trained and retrained transformer-based pipelines that lift accuracy scores significantly
    • Additional configuration capabilities to build your training workflow and tune your training runs
    • A Quickstart Widget to help build your configuration files
    • Easier integration with other tools such as Streamlit, FastAPI, or Ray to build workflows
    • Parallel/Distributed capabilities with Ray for faster training cycles
    • Wrappers that enable you to bring in other frameworks such as PyTorch and TensorFlow

    These features combine to make spaCy better than ever at processing large volumes of text and tuning configurations to match specific use cases in a way that provides better accuracy.

    Summary

    • SpaCy 3.0