This model was trained using
SentenceTransformers Cross-Encoder class.
This model was trained on the
Quora Duplicate Questions dataset. The model will predict a score between 0 and 1 how likely the two given questions are duplicates.
Note: The model is not suitable to estimate the similarity of questions, e.g. the two questions "How to learn Java" and "How to learn Python" will result in a rather low score, as these are not duplicates.
1from sentence_transformers import CrossEncoder
2
3model = CrossEncoder('cross-encoder/quora-distilroberta-base')
4scores = model.predict([('Question 1', 'Question 2'), ('Question 3', 'Question 4')])
You can use this model also without sentence_transformers and by just using Transformers AutoModel class