Open Access. Powered by Scholars. Published by Universities.®

2020

Discipline
Institution
Keyword
Publication
Publication Type
File Type

Articles 91 - 97 of 97

Full-Text Articles in Cataloging and Metadata

Library Experiences Of Technical Services And Special Collections During The Covid-19 Pandemic At Edith Garland Dupré Library, University Of Louisiana At Lafayette, Zachary G. Stein, Sheryl Curry, Janelle Zetty, Michael Mitchell, Scott Jordan Jan 2020

Library Experiences Of Technical Services And Special Collections During The Covid-19 Pandemic At Edith Garland Dupré Library, University Of Louisiana At Lafayette, Zachary G. Stein, Sheryl Curry, Janelle Zetty, Michael Mitchell, Scott Jordan

Faculty Scholarship

The COVID-19 pandemic put a major strain on everyday life. Academic institutions were forced to close, leaving many employees to resort to teleworking. Academic libraries were put in an unusual position, as they strove to provide access to library materials and services, despite limited to no access to the actual sites. This proved to be especially tricky for technical services and special collections departments, which heavily rely on being present at libraries. Therefore, these departments needed to develop solutions that would allow them to continue performing their duties while adhering to stay-at-home orders. With careful preparation and frequent communication, the …


Rose Is A Rose Is A Rose, Janelle Zetty Jan 2020

Rose Is A Rose Is A Rose, Janelle Zetty

Faculty Scholarship

No abstract provided.


Use Of Customized Marc Framework For Recording Books In Koha: With Special References To Public Libraries Of West Bengal, Rounak Biswas Jan 2020

Use Of Customized Marc Framework For Recording Books In Koha: With Special References To Public Libraries Of West Bengal, Rounak Biswas

Library Philosophy and Practice (e-journal)

Library automation means a high degree of mechanization of various routine and repetitive tasks to be performed by library staffs. Presently, public libraries of West Bengal are moving towards automation with koha. Koha, being open source software, allows extensive customization. Most of the libraries are now use their own bibliographic framework for recording different type of documents to save the time of data entry work. This paper highlights the need of customized MARC framework for books in libraries with an instance of use of customized MARC framework in Public libraries of West Bengal.


A Heuristic Baseline Method For Metadata Extraction From Scanned Electronic Theses And Dissertations, Muntabir H. Choudhury, Jian Wu, William A. Ingam, Edward A. Fox Jan 2020

A Heuristic Baseline Method For Metadata Extraction From Scanned Electronic Theses And Dissertations, Muntabir H. Choudhury, Jian Wu, William A. Ingam, Edward A. Fox

Computer Science Faculty Publications

Extracting metadata from scholarly papers is an important text mining problem. Widely used open-source tools such as GROBID are designed for born-digital scholarly papers but often fail for scanned documents, such as Electronic Theses and Dissertations (ETDs). Here we present a preliminary baseline work with a heuristic model to extract metadata from the cover pages of scanned ETDs. The process started with converting scanned pages into images and then text files by applying OCR tools. Then a series of carefully designed regular expressions for each field is applied, capturing patterns for seven metadata fields: titles, authors, years, degrees, academic programs, …


Smartcitecon: Implicit Citation Context Extraction From Academic Literature Using Unsupervised Learning, Chenrui Gao, Haoran Cui, Li Zhang, Jiamin Wang, Wei Lu, Jian Wu Jan 2020

Smartcitecon: Implicit Citation Context Extraction From Academic Literature Using Unsupervised Learning, Chenrui Gao, Haoran Cui, Li Zhang, Jiamin Wang, Wei Lu, Jian Wu

Computer Science Faculty Publications

We introduce SmartCiteCon (SCC), a Java API for extracting both explicit and implicit citation context from academic literature in English. The tool is built on a Support Vector Machine (SVM) model trained on a set of 7,058 manually annotated citation context sentences, curated from 34,000 papers in the ACL Anthology. The model with 19 features achieves F1=85.6%. SCC supports PDF, XML, and JSON files out-of-box, provided that they are conformed to certain schemas. The API supports single document processing and batch processing in parallel. It takes about 12–45 seconds on average depending on the format to process a …


Acknowledgement Entity Recognition In Cord-19 Papers, Jian Wu, Pei Wang, Xin Wei, Sarah Rajtmajer, C. Lee Giles, Christopher Griffin Jan 2020

Acknowledgement Entity Recognition In Cord-19 Papers, Jian Wu, Pei Wang, Xin Wei, Sarah Rajtmajer, C. Lee Giles, Christopher Griffin

Computer Science Faculty Publications

Acknowledgements are ubiquitous in scholarly papers. Existing acknowledgement entity recognition methods assume all named entities are acknowledged. Here, we examine the nuances between acknowledged and named entities by analyzing sentence structure. We develop an acknowledgement extraction system, AckExtract based on open-source text mining software and evaluate our method using manually labeled data. AckExtract uses the PDF of a scholarly paper as input and outputs acknowledgement entities. Results show an overall performance of F1=0.92. We built a supplementary database by linking CORD-19 papers with acknowledgement entities extracted by AckExtract including persons and organizations and find that only up to …


Opening Books And The National Corpus Of Graduate Research, William A. Ingram, Edward A. Fox, Jian Wu Jan 2020

Opening Books And The National Corpus Of Graduate Research, William A. Ingram, Edward A. Fox, Jian Wu

Computer Science Faculty Publications

Virginia Tech University Libraries, in collaboration with Virginia Tech Department of Computer Science and Old Dominion University Department of Computer Science, request $505,214 in grant funding for a 3-year project, the goal of which is to bring computational access to book-length documents, demonstrating that with Electronic Theses and Dissertations (ETDs). The project is motivated by the following library and community needs. (1) Despite huge volumes of book-length documents in digital libraries, there is a lack of models offering effective and efficient computational access to these long documents. (2) Nationwide open access services for ETDs generally function at the metadata level. …