2 papers
cs.CL2024
Are ELECTRA's Sentence Embeddings Beyond Repair? The Case of Semantic Textual Similarity
Ivan Rep, David DukiÄ, Jan Å najder
While BERT produces high-quality sentence embeddings, its pre-training computational cost is a significant drawback. In contrast, ELECTRA provides a cost-effective pre-training obj…
cs.CL2024
Looking Right is Sometimes Right: Investigating the Capabilities of Decoder-only LLMs for Sequence Labeling
David DukiÄ, Jan Å najder
Pre-trained language models based on masked language modeling (MLM) excel in natural language understanding (NLU) tasks. While fine-tuned MLM-based encoders consistently outperform…