keras-bert

BERT implemented in Keras

These details have not been verified by PyPI

Project links

Homepage

License
- OSI Approved :: MIT License
Operating System
- OS Independent
Programming Language
- Python :: 2.7
- Python :: 3.6

Project description

Implementation of the paper: BERT: Pre-training of Deep Bidirectional Transformers for Language Understanding

Install

pip install keras-bert

Usage

Train & Use

from keras_bert import get_base_dict, get_model, gen_batch_inputs


# A toy input example
sentence_pairs = [
    [['all', 'work', 'and', 'no', 'play'], ['makes', 'jack', 'a', 'dull', 'boy']],
    [['from', 'the', 'day', 'forth'], ['my', 'arm', 'changed']],
    [['and', 'a', 'voice', 'echoed'], ['power', 'give', 'me', 'more', 'power']],
]


# Build token dictionary
token_dict = get_base_dict()  # A dict that contains some special tokens
for pairs in sentence_pairs:
    for token in pairs[0] + pairs[1]:
        if token not in token_dict:
            token_dict[token] = len(token_dict)
token_list = list(token_dict.keys())  # Used for selecting a random word


# Build & train the model
model = get_model(
    token_num=len(token_dict),
    head_num=5,
    transformer_num=12,
    embed_dim=25,
    feed_forward_dim=100,
    seq_len=20,
    pos_num=20,
    dropout=0.05,
)
model.summary()

def _generator():
    while True:
        yield gen_batch_inputs(
            sentence_pairs,
            token_dict,
            token_list,
            seq_len=20,
            mask_rate=0.3,
            swap_sentence_rate=1.0,
        )

model.fit_generator(
    generator=_generator(),
    steps_per_epoch=1000,
    epochs=100,
    validation_data=_generator(),
    validation_steps=100,
    callbacks=[
        keras.callbacks.EarlyStopping(monitor='val_loss', patience=5)
    ],
)


# Use the trained model
inputs, output_layer = get_model(  # `output_layer` is the last feature extraction layer (the last transformer)
    token_num=len(token_dict),
    head_num=5,
    transformer_num=12,
    embed_dim=25,
    feed_forward_dim=100,
    seq_len=20,
    pos_num=20,
    dropout=0.05,
    training=False,  # The input layers and output layer will be returned if `training` is `False`
)

Custom Feature Extraction

def _custom_layers(inputs):
    return keras.layers.LSTM(units=768, name='LSTM')(inputs)

model = get_model(
    token_num=200,
    embed_dim=768,
    custom_layers=_custom_layers,
)

Project details

These details have not been verified by PyPI

Project links

Homepage

License
- OSI Approved :: MIT License
Operating System
- OS Independent
Programming Language
- Python :: 2.7
- Python :: 3.6

Release history Release notifications | RSS feed

0.89.0

Jan 22, 2022

0.88.0

Jun 19, 2021

0.87.0

Jun 16, 2021

0.86.0

Jul 28, 2020

0.85.0

Jul 8, 2020

0.84.0

Jun 6, 2020

0.83.0

Jun 2, 2020

0.82.0

Jun 2, 2020

0.81.1

Jun 11, 2020

0.81.0

Jan 31, 2020

0.80.0

Oct 12, 2019

0.79.0

Sep 30, 2019

0.78.0

Sep 18, 2019

0.77.0

Sep 3, 2019

0.76.0

Aug 29, 2019

0.75.0

Aug 23, 2019

0.74.0

Aug 21, 2019

0.73.0

Aug 20, 2019

0.72.0

Aug 16, 2019

0.71.0

Jul 31, 2019

0.70.3

Jul 31, 2019

0.70.2

Jul 30, 2019

0.70.1

Jul 23, 2019

0.70.0

Jul 18, 2019

0.69.0

Jul 16, 2019

0.68.2

Jul 14, 2019

0.68.1

Jul 8, 2019

0.68.0

Jul 8, 2019

0.66.0

Jul 5, 2019

0.65.1

Jul 3, 2019

0.65.0

Jun 21, 2019

0.64.0

Jun 20, 2019

0.63.0

Jun 18, 2019

0.62.0

Jun 18, 2019

0.61.0

Jun 13, 2019

0.60.0

Jun 11, 2019

0.59.0

Jun 11, 2019

0.58.0

Jun 10, 2019

0.57.1

Jun 6, 2019

0.57.0

Jun 5, 2019

0.55.2

Jun 4, 2019

0.55.1

May 30, 2019

0.55.0

May 30, 2019

0.54.0

May 30, 2019

0.53.0

May 24, 2019

0.52.0

May 23, 2019

0.51.0

May 22, 2019

0.50.0

May 22, 2019

0.49.0

May 21, 2019

0.47.0

May 21, 2019

0.46.0

May 19, 2019

0.43.0

May 11, 2019

0.42.0

May 11, 2019

0.41.0

Apr 30, 2019

0.40.0

Apr 30, 2019

0.39.0

Apr 16, 2019

0.38.0

Apr 9, 2019

0.37.0

Apr 9, 2019

0.36.1

Mar 27, 2019

0.36.0

Mar 27, 2019

0.35.0

Mar 22, 2019

0.34.1

Mar 21, 2019

0.34.0

Mar 21, 2019

0.33.1

Mar 21, 2019

0.33.0

Mar 19, 2019

0.32.0

Mar 13, 2019

0.31.0

Mar 11, 2019

0.30.0

Feb 28, 2019

0.29.0

Feb 6, 2019

0.28.0

Feb 1, 2019

0.27.0

Jan 31, 2019

0.25.0

Jan 20, 2019

0.24.0

Dec 4, 2018

0.23.0

Nov 29, 2018

0.22.0

Nov 27, 2018

0.21.0

Nov 20, 2018

0.20.0

Nov 14, 2018

0.19.0

Nov 14, 2018

0.18.0

Nov 14, 2018

0.16.0

Nov 14, 2018

0.15.0

Nov 13, 2018

0.14.0

Nov 9, 2018

0.13.0

Nov 7, 2018

0.12.0

Nov 7, 2018

0.11.0

Nov 6, 2018

0.10.0

Nov 1, 2018

0.9.0

Nov 1, 2018

This version

0.8.0

Oct 31, 2018

0.7.0

Oct 29, 2018

0.6.0

Oct 26, 2018

0.5.0

Oct 26, 2018

0.4.0

Oct 26, 2018

0.3.0

Oct 26, 2018

0.2.0

Oct 26, 2018

0.1.0

Oct 26, 2018

Download files

Download the file for your platform. If you're not sure which to choose, learn more about installing packages.

Source Distribution

keras-bert-0.8.0.tar.gz (12.3 kB view details)

Uploaded Oct 31, 2018 Source

File details

Details for the file keras-bert-0.8.0.tar.gz.

File metadata

Download URL: keras-bert-0.8.0.tar.gz
Upload date: Oct 31, 2018
Size: 12.3 kB
Tags: Source
Uploaded using Trusted Publishing? No
Uploaded via: twine/1.11.0 pkginfo/1.4.2 requests/2.18.4 setuptools/28.8.0 requests-toolbelt/0.8.0 tqdm/4.24.0 CPython/3.6.4

File hashes

Hashes for keras-bert-0.8.0.tar.gz
Algorithm	Hash digest
SHA256	`67099c36db8bf7a86c58a58757906ffb717edfcecb8f93fb9f8e4c5e641cc972`
MD5	`4884b1c0e58960a6e1e12e4c316ee1dd`
BLAKE2b-256	`a332750573d4dd6258cd04bc6c1219479ed2662b4d5b4c2f9ffcc48c62231e24`

See more details on using hashes here.

keras-bert 0.8.0

Navigation

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Project description

Install

Usage

Train & Use

Custom Feature Extraction

Project details

Verified details

Maintainers

Unverified details

Project links

Meta

Classifiers

Release history Release notifications | RSS feed

Download files

Source Distribution

File details

File metadata

File hashes