Skip to main content

Industry-strength Natural Language Processing extensions for Keras.

Project description

KerasNLP: Modular NLP Workflows for Keras

Python contributions welcome

KerasNLP is a natural language processing library that works natively with TensorFlow, JAX, or PyTorch. Built on multi-backend Keras (Keras 3), these models, layers, metrics, and tokenizers can be trained and serialized in any framework and re-used in another without costly migrations.

KerasNLP supports users through their entire development cycle. Our workflows are built from modular components that have state-of-the-art preset weights when used out-of-the-box and are easily customizable when more control is needed.

This library is an extension of the core Keras API; all high-level modules are Layers or Models that receive that same level of polish as core Keras. If you are familiar with Keras, congratulations! You already understand most of KerasNLP.

See our Getting Started guide to start learning our API. We welcome contributions.

Quick Links

For everyone

For contributors

Installation

To install the latest official release:

pip install keras-nlp --upgrade

To install the latest unreleased changes to the library, we recommend using pip to install directly from the master branch on github:

pip install git+https://github.com/keras-team/keras-nlp.git --upgrade

Quickstart

Fine-tune BERT on a small sentiment analysis task using the keras_nlp.models API:

import os
os.environ["KERAS_BACKEND"] = "jax"  # Or "tensorflow", or "torch".

import keras_nlp
import tensorflow_datasets as tfds

imdb_train, imdb_test = tfds.load(
    "imdb_reviews",
    split=["train", "test"],
    as_supervised=True,
    batch_size=16,
)
# Load a BERT model.
classifier = keras_nlp.models.BertClassifier.from_preset(
    "bert_base_en_uncased", 
    num_classes=2,
    activation="softmax",
)
# Fine-tune on IMDb movie reviews.
classifier.fit(imdb_train, validation_data=imdb_test)
# Predict two new examples.
classifier.predict(["What an amazing movie!", "A total waste of my time."])

For more in depth guides and examples, visit https://keras.io/keras_nlp/.

Configuring your backend

Keras 3 is an upcoming release of the Keras library which supports TensorFlow, Jax or Torch as backends. This is supported today in KerasNLP, but will not be enabled by default until the official release of Keras 3. If you pip install keras-nlp and run a script or notebook without changes, you will be using TensorFlow and Keras 2.

If you would like to enable a preview of the Keras 3 behavior, you can do so by setting the KERAS_BACKEND environment variable. For example:

export KERAS_BACKEND=jax

Or in Colab, with:

import os
os.environ["KERAS_BACKEND"] = "jax"

import keras_nlp

[!IMPORTANT] Make sure to set the KERAS_BACKEND before import any Keras libraries, it will be used to set up Keras when it is first imported.

Until the Keras 3 release, KerasNLP will use a preview of Keras 3 on PyPI named keras-core.

[!IMPORTANT] If you set KERAS_BACKEND variable, you should import keras_core as keras instead of import keras. This is a temporary step until Keras 3 is out!

To restore the default Keras 2 behavior, unset KERAS_BACKEND before importing Keras and KerasNLP.

Compatibility

We follow Semantic Versioning, and plan to provide backwards compatibility guarantees both for code and saved models built with our components. While we continue with pre-release 0.y.z development, we may break compatibility at any time and APIs should not be consider stable.

Disclaimer

KerasNLP provides access to pre-trained models via the keras_nlp.models API. These pre-trained models are provided on an "as is" basis, without warranties or conditions of any kind. The following underlying models are provided by third parties, and subject to separate licenses: BART, DeBERTa, DistilBERT, GPT-2, OPT, RoBERTa, Whisper, and XLM-RoBERTa.

Citing KerasNLP

If KerasNLP helps your research, we appreciate your citations. Here is the BibTeX entry:

@misc{kerasnlp2022,
  title={KerasNLP},
  author={Watson, Matthew, and Qian, Chen, and Bischof, Jonathan and Chollet, 
  Fran\c{c}ois and others},
  year={2022},
  howpublished={\url{https://github.com/keras-team/keras-nlp}},
}

Acknowledgements

Thank you to all of our wonderful contributors!

Project details


Release history Release notifications | RSS feed

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

keras-nlp-nightly-0.7.0.dev2023111002.tar.gz (286.5 kB view details)

Uploaded Source

Built Distribution

File details

Details for the file keras-nlp-nightly-0.7.0.dev2023111002.tar.gz.

File metadata

File hashes

Hashes for keras-nlp-nightly-0.7.0.dev2023111002.tar.gz
Algorithm Hash digest
SHA256 ac83243711f84af9b7f2ef8717118bf04ef3fa12ec0f29cc2725f9c363d19893
MD5 ef0e974f3399b8c5ef9ca4fec95ce7f6
BLAKE2b-256 b9dad9a4f508da117910cd02e02b0edeeba16858add680d20d2c21190ac5c794

See more details on using hashes here.

Provenance

File details

Details for the file keras_nlp_nightly-0.7.0.dev2023111002-py3-none-any.whl.

File metadata

File hashes

Hashes for keras_nlp_nightly-0.7.0.dev2023111002-py3-none-any.whl
Algorithm Hash digest
SHA256 d12c44e80cfe527ce1915b8fd9f127263018361ea3a41c5c34358b2bc55a3d76
MD5 7728b838e2f64038347f47134e6ef83b
BLAKE2b-256 f25ccf3507b7787cf1bfa78745db6a4d029c0e0462cdff4ba744b7a28b6d8f94

See more details on using hashes here.

Provenance

Supported by

AWS AWS Cloud computing and Security Sponsor Datadog Datadog Monitoring Fastly Fastly CDN Google Google Download Analytics Microsoft Microsoft PSF Sponsor Pingdom Pingdom Monitoring Sentry Sentry Error logging StatusPage StatusPage Status page