Isn’t this post comparing zero shot Jev to fine tunes of this model for each of the datasets it is tested on? If so seems like fairly impressive results for Jev
Most valuable comment here. You can fine tune BERT to perform really well on classification tasks. What's exciting about Jev is that it generalizes. Laya does not!