SciBERT: A Pretrained Language Model for Scientific Text
PulseAugur coverage of SciBERT: A Pretrained Language Model for Scientific Text — every cluster mentioning SciBERT: A Pretrained Language Model for Scientific Text across labs, papers, and developer communities, ranked by signal.
2 day(s) with sentiment data
-
SciBERT model achieves top score in telescope bibliography classification task
Researchers have developed an efficient SciBERT-based method for classifying scientific papers related to telescope bibliographies. Despite facing strict context-length limitations and restricted computational resources…
-
New SNAIL framework automates identification of bioinformatics tools in research papers
Researchers have developed SNAIL, a novel framework for automatically identifying bioinformatics software and database names within scientific literature. This hybrid approach combines lexical pattern recognition with s…
-
New benchmarks and models assess scientific figures using manuscript context
Two new research papers introduce methods for evaluating the quality of scientific figures within the context of their accompanying manuscripts. SciFigAlign and SciFigQual-Bench are proposed benchmarks and associated mo…
-
LLMs evaluated for citation function classification, achieving new SOTA
A new research paper evaluates several large language models (LLMs) for the task of citation function classification, aiming to improve bibliometric analysis. The study achieved new state-of-the-art results on the ACL-A…
-
New research evaluates unsupervised methods for scholarly collaboration recommendations
Researchers have evaluated unsupervised methods for recommending scholarly collaborations based on publication text. The study compared TF-IDF, topic-based models (LDA, BERTopic), and embedding-based retrieval using Sci…
-
New S1-MMAlign dataset boosts AI for scientific figure-text understanding
Researchers have introduced S1-MMAlign, a large-scale dataset designed to improve multimodal understanding in scientific research. The dataset contains over 15.5 million image-text pairs from scientific papers across va…