Implementation of segmentation embeddings

Hi,
Thanks for releasing this awesome repo.
In your implementation, I found each input sentence has a separate id in the "bert-type-ids". I wonder if you use these sentences ids to generate the segmentation embeddings?

BERT uses the sum of the token embeddings, the segmentation embeddings, and the position embeddings as input embeddings.  They define there are no more than two input segments and only use 0 and 1 as the segment ids, which means the segment ids greater than 1 have not appeared in the pre-training of BERT. 

Provide feedback

Saved searches

Use saved searches to filter your results more quickly

Implementation of segmentation embeddings #11

Metadata

Assignees

Labels

Type

Projects

Milestone

Relationships

Development

Implementation of segmentation embeddings #11

Description

Metadata

Metadata

Assignees

Labels

Type

Projects

Milestone

Relationships

Development

Issue actions