Open Access. Powered by Scholars. Published by Universities.®

Databases and Information Systems

Institution
Keyword
Publication Year
Publication
Publication Type
File Type

Articles 1 - 28 of 28

Full-Text Articles in Cataloging and Metadata

Syntax-Enhanced Boundary-Aware Named Entity Recognition Model, Chuanming Yu, Bin Deng, Zhengang Zhang Jul 2025

Syntax-Enhanced Boundary-Aware Named Entity Recognition Model, Chuanming Yu, Bin Deng, Zhengang Zhang

Journal of Scientific Information Research

[Purpose/significance] This study addresses the issue of inadequate perception of entity boundaries in traditional character-level modeling-based named entity recognition models by integrating syntax information containing entity boundary features into the task using a multi-head graph attention network with dense connections. This integration enhances the effectiveness of named entity recognition.

[Method/process] This study proposes a Syntax-enhanced Boundary-aware Named Entity Recognition Model (SynBNER), which utilizes BERT for text semantic representation and integrates syntax information using a dense-connected graph attention network. This integration incorporates implicit entity boundary information from syntax information into word representations, thereby enhancing the model's entity boundary perception capability.

[Result/conclusion] …


Pulling Up Stakes: Migrating Digital Collections From Contentdm To Digital Commons, Adam C. Northam Sep 2024

Pulling Up Stakes: Migrating Digital Collections From Contentdm To Digital Commons, Adam C. Northam

Velma K. Waters Library Faculty Publications

No abstract provided.


Chatgpt As Metamorphosis Designer For The Future Of Artificial Intelligence (Ai): A Conceptual Investigation, Amarjit Kumar Singh (Library Assistant), Dr. Pankaj Mathur (Deputy Librarian) Mar 2023

Chatgpt As Metamorphosis Designer For The Future Of Artificial Intelligence (Ai): A Conceptual Investigation, Amarjit Kumar Singh (Library Assistant), Dr. Pankaj Mathur (Deputy Librarian)

Library Philosophy and Practice (e-journal)

Abstract

Purpose: The purpose of this research paper is to explore ChatGPT’s potential as an innovative designer tool for the future development of artificial intelligence. Specifically, this conceptual investigation aims to analyze ChatGPT’s capabilities as a tool for designing and developing near about human intelligent systems for futuristic used and developed in the field of Artificial Intelligence (AI). Also with the helps of this paper, researchers are analyzed the strengths and weaknesses of ChatGPT as a tool, and identify possible areas for improvement in its development and implementation. This investigation focused on the various features and functions of ChatGPT that …


Creating Data From Unstructured Text With Context Rule Assisted Machine Learning (Craml), Stephen Meisenbacher, Peter Norlander Dec 2022

Creating Data From Unstructured Text With Context Rule Assisted Machine Learning (Craml), Stephen Meisenbacher, Peter Norlander

School of Business: Faculty Publications and Other Works

Popular approaches to building data from unstructured text come with limitations, such as scalability, interpretability, replicability, and real-world applicability. These can be overcome with Context Rule Assisted Machine Learning (CRAML), a method and no-code suite of software tools that builds structured, labeled datasets which are accurate and reproducible. CRAML enables domain experts to access uncommon constructs within a document corpus in a low-resource, transparent, and flexible manner. CRAML produces document-level datasets for quantitative research and makes qualitative classification schemes scalable over large volumes of text. We demonstrate that the method is useful for bibliographic analysis, transparent analysis of proprietary data, …


Streaminghub: Interactive Stream Analysis Workflows, Yasith Jayawardana, Vikas G. Ashok, Sampath Jayarathna Jan 2022

Streaminghub: Interactive Stream Analysis Workflows, Yasith Jayawardana, Vikas G. Ashok, Sampath Jayarathna

Computer Science Faculty Publications

Reusable data/code and reproducible analyses are foundational to quality research. This aspect, however, is often overlooked when designing interactive stream analysis workflows for time-series data (e.g., eye-tracking data). A mechanism to transmit informative metadata alongside data may allow such workflows to intelligently consume data, propagate metadata to downstream tasks, and thereby auto-generate reusable, reproducible analytic outputs with zero supervision. Moreover, a visual programming interface to design, develop, and execute such workflows may allow rapid prototyping for interdisciplinary research. Capitalizing on these ideas, we propose StreamingHub, a framework to build metadata propagating, interactive stream analysis workflows using visual programming. We conduct …


Mapping Renewal: How An Unexpected Interdisciplinary Collaboration Transformed A Digital Humanities Project, Elise Tanner, Geoffrey Joseph Apr 2021

Mapping Renewal: How An Unexpected Interdisciplinary Collaboration Transformed A Digital Humanities Project, Elise Tanner, Geoffrey Joseph

Digital Initiatives Symposium

Funded by a National Endowment for Humanities (NEH) Humanities Collections and Reference Resources Foundations Grant, the UA Little Rock Center for Arkansas History and Culture’s “Mapping Renewal” pilot project focused on creating access to and providing spatial context to archival materials related to racial segregation and urban renewal in the city of Little Rock, Arkansas, from 1954-1989. An unplanned interdisciplinary collaboration with the UA Little Rock Arkansas Economic Development Institute (AEDI) has proven to be an invaluable partnership. One team member from each department will demonstrate the Mapping Renewal website and discuss how the collaborative process has changed and shaped …


Automatic Metadata Extraction Incorporating Visual Features From Scanned Electronic Theses And Dissertations, Muntabir Hasan Choudhury, Himarsha R. Jayanetti, Jian Wu, William A. Ingram, Edward A. Fox Jan 2021

Automatic Metadata Extraction Incorporating Visual Features From Scanned Electronic Theses And Dissertations, Muntabir Hasan Choudhury, Himarsha R. Jayanetti, Jian Wu, William A. Ingram, Edward A. Fox

Computer Science Faculty Publications

Electronic Theses and Dissertations (ETDs) contain domain knowledge that can be used for many digital library tasks, such as analyzing citation networks and predicting research trends. Automatic metadata extraction is important to build scalable digital library search engines. Most existing methods are designed for born-digital documents, so they often fail to extract metadata from scanned documents such as ETDs. Traditional sequence tagging methods mainly rely on text-based features. In this paper, we propose a conditional random field (CRF) model that combines text-based and visual features. To verify the robustness of our model, we extended an existing corpus and created a …


Opening Books And The National Corpus Of Graduate Research, William A. Ingram, Edward A. Fox, Jian Wu Jan 2020

Opening Books And The National Corpus Of Graduate Research, William A. Ingram, Edward A. Fox, Jian Wu

Computer Science Faculty Publications

Virginia Tech University Libraries, in collaboration with Virginia Tech Department of Computer Science and Old Dominion University Department of Computer Science, request $505,214 in grant funding for a 3-year project, the goal of which is to bring computational access to book-length documents, demonstrating that with Electronic Theses and Dissertations (ETDs). The project is motivated by the following library and community needs. (1) Despite huge volumes of book-length documents in digital libraries, there is a lack of models offering effective and efficient computational access to these long documents. (2) Nationwide open access services for ETDs generally function at the metadata level. …


Data Curation Workshop: Tips And Tools For Today, Matthew M. Benzing Oct 2019

Data Curation Workshop: Tips And Tools For Today, Matthew M. Benzing

Charleston Library Conference

The current state of research data is like a disorganized photo collection: a mix of formats scattered across different media without a lot of authority control. That is changing as the need to make data available to researchers across the world is becoming recognized. Researchers know that their data needs to be maintained and made accessible, but often they do not have the time or the inclination to get involved in all of the details. This provides an excellent opportunity for librarians. Data curation is the process of preparing data to be made available in a repository with the goal …


Astria Ontology: Open, Standards-Based, Data-Aggregated Representation Of Space Objects, Jennie Wolfgang, Kathleen Krysher, Michael Slovenski, Unmil P. Karadkar, Shiva Iyer, Moriba K. Jah Feb 2019

Astria Ontology: Open, Standards-Based, Data-Aggregated Representation Of Space Objects, Jennie Wolfgang, Kathleen Krysher, Michael Slovenski, Unmil P. Karadkar, Shiva Iyer, Moriba K. Jah

Space Traffic Management Conference

The necessity for standards-based ontologies for long-term sustainability of space operations and safety of increasing space flights has been well-established [6, 7]. Current ontologies, such as DARPA’s OrbitOutlook [5], are not publicly available, complicating efforts for their broad adoption. Most sensor data is siloed in proprietary databases [2] and provided only to authorized users, further complicating efforts to create a holistic view of resident space objects (RSOs) in order to enhance space situational awareness (SSA).

The ASTRIA project is developing an open data model with the goal of aggregating data about RSOs, parts, space weather, and governing policies in order …


Creating A Reproducible Metadata Transformation Pipeline Using Technology Best Practices, Cara Key, Mike Waugh Apr 2018

Creating A Reproducible Metadata Transformation Pipeline Using Technology Best Practices, Cara Key, Mike Waugh

Digital Initiatives Symposium

Over the course of two years, a team of librarians and programmers from LSU Libraries migrated the 186 collections of the Louisiana Digital Library from OCLC's CONTENTdm platform over to the open-source Islandora platform.

Early in the process, the team understood the value of creating a reproducible metadata transformation pipeline, because there were so many unknowns at the beginning of the process along with the certainty that mistakes would be made. This presentation will describe how the team used innovative and collaborative tools, such as Trello, Ansible, Vagrant, VirtualBox, git and GitHub to accomplish the task.


A Survey Of Archival Replay Banners, Sawood Alam, Mat Kelly, Michele C. Weigle, Michael L. Nelson Jan 2018

A Survey Of Archival Replay Banners, Sawood Alam, Mat Kelly, Michele C. Weigle, Michael L. Nelson

Computer Science Faculty Publications

We surveyed various archival systems to compare and contrast different techniques used to implement an archival replay banner. We found that inline plain HTML injection is the most common approach, but prone to style conflicts. Iframe-based banners are also very common and while they do not have style conflicts, they suffer from screen real estate wastage and limited design choices. Custom Elements-based banners are promising, but due to being a new web standard, these are not yet widely deployed.


How The University Of California Runs One Repository For Ten Campuses, Katie Fortney Apr 2015

How The University Of California Runs One Repository For Ten Campuses, Katie Fortney

Inaugural CSU IR Conference, 2015

Katie Fortney, JD, MLIS, Copyright Policy & Education Officer, Office of Scholarly Communication, University of California http://osc.universityofcalifornia.edu/


Implementing Metaarchive And Lockss At Digital Commons @Cal Poly, Michele Wyngard Apr 2015

Implementing Metaarchive And Lockss At Digital Commons @Cal Poly, Michele Wyngard

Inaugural CSU IR Conference, 2015

Michele Wyngard, Digital Repository Coordinator, CSU Cal Poly


Using Google Tag Manager And Google Analytics, (Code{4}Lib Journal), Suzanna Conrad Apr 2015

Using Google Tag Manager And Google Analytics, (Code{4}Lib Journal), Suzanna Conrad

Inaugural CSU IR Conference, 2015

Suzanna Conrad, Digital Initiatives Librarian, Cal Poly Pomona


What’S New Since The April 2013 Stim Ir Subcommittee Report To Cold: Hydra, Islandora And Dspace, Aaron Collier, Suzanna Conrad, Carmen Mitchell, Joan Parker, Andrew Weiss, Jeremy C. Shellhase Apr 2015

What’S New Since The April 2013 Stim Ir Subcommittee Report To Cold: Hydra, Islandora And Dspace, Aaron Collier, Suzanna Conrad, Carmen Mitchell, Joan Parker, Andrew Weiss, Jeremy C. Shellhase

Inaugural CSU IR Conference, 2015

Aaron Collier, Digital Repository Services Manager, Chancellor’s Office
Suzanna Conrad, Digital Initiatives Librarian, Cal Poly Pomona
Carmen Mitchell, Institutional Repository Librarian, CSU San Marcos
Joan Parker, Librarian, Moss Landing Marine Laboratories
Andrew Weiss, Digital Services Librarian, CSU Northridge

Jeremy Shellhase, Head of Information Services & Systems Department, Humboldt State University


The State Of Scholarworks, Aaron Collier Apr 2015

The State Of Scholarworks, Aaron Collier

Inaugural CSU IR Conference, 2015

Aaron Collier, Digital Repository Services Manager, Chancellor’s Office


Linked Data Demystified: Practical Efforts To Transform Contentdm Metadata For The Linked Data Cloud, Silvia B. Southwick, Cory K. Lampert Nov 2012

Linked Data Demystified: Practical Efforts To Transform Contentdm Metadata For The Linked Data Cloud, Silvia B. Southwick, Cory K. Lampert

Library Faculty Presentations

The library literature and events like the ALA Annual Conference have been inundated with presentations and articles on linked data. At UNLV Libraries, we understand the importance of linked data in helping to better service our users. We have designed and initiated a pilot project to apply linked data concepts to the practical task of transforming a sample set of our CONTENTdm digital collections data into future-oriented linked data. This presentation will outline rationale for beginning work in linked data and detail the phases we will undertake in the proof of concept project. We hope through this research experiment to …


Evaluating And Implementing Web Scale Discovery Services: Part Two, Jason Vaughan, Tamera Hanken Jul 2011

Evaluating And Implementing Web Scale Discovery Services: Part Two, Jason Vaughan, Tamera Hanken

Library Faculty Presentations

Part Four: Quick Tour of the Current Marketplace:

  • "The Big 5"
  • Similarities and differences

Part Five: It's Not All Sliced Bread:

  • Shortcomings of web scale discovery

Part Six: Implementation (pre launch steps):

  • Selecting and preparing implementation staff
  • Preparing and communicating process/decisions with all staff
  • Working with the vendor (roles, expectations, timeline)
  • Workflow changes and implications (technical services)

Part Seven: Specific implementation tasks, issues, and considerations:

  • Record loading and mapping (catalog content)
  • Harvesting and mapping digital/local content
  • Working with central index data (internal & external content)
  • Web integration and customization
  • Assessment and continuous improvement


Evaluating And Implementing Web Scale Discovery Services: Part One, Jason Vaughan, Tamera Hanken Jul 2011

Evaluating And Implementing Web Scale Discovery Services: Part One, Jason Vaughan, Tamera Hanken

Library Faculty Presentations

Preface: Before Web Scale Discovery

  • A very brief overview

Part 1: What is Web Scale Discovery

  • Content
  • Technology

Part 2: Why is Web Scale Discovery important?

  • What’s the need?
  • How is it different from earlier attempts at broad discovery?

Part 3: A Framework for Evaluating Web Scale Discovery Services

  • What we did at UNLV
  • Other options




Skos And The Semantic Web: Knowledge Organization, Metadata, And Interoperability, Eric A. Robinson Jan 2010

Skos And The Semantic Web: Knowledge Organization, Metadata, And Interoperability, Eric A. Robinson

Other Topics

The Simplified Knowledge Organization System (SKOS) is a Semantic Web framework, based on the Resource Description Framework (RDF) for thesauri, classification schemes and simple ontologies. It allows for machine-actionable description of the structure of these knowledge organization systems (KOS) and provides an excellent tool for addressing interoperability and vocabulary control problems inherent to the rapidly expanding information environment of the Web. This paper discusses the foundations of the SKOS framework and reviews the literature on a variety of SKOS implementations. The limitations of SKOS that have been revealed through its broad application are addressed with brief attention to the proposed …


Reading Over The Shoulder Of The Future At The Library Of Congress, Samuel Gerald Collins Jan 2004

Reading Over The Shoulder Of The Future At The Library Of Congress, Samuel Gerald Collins

Reconstruction: Studies in Contemporary Culture

[First paragraph]

It is 2003 and I am doing some research at the Library of Congress, the de facto national library for the United States and the largest library in the world. Next to me sit some articles I've printed off of online journals on the Defense Advanced Research Project Agency's Total Information Awareness Project, a plan, still in its formative stage, to throw a panopticon net of surveillance across the United States through a combination of language translation technologies, data search and pattern recognition technologies, and advanced collaborative and decision support tools (DARPA). But I am also doing research …


Sistem Arkib Dalam Talian Adt / Ods, Mohd Radzi Nurul Azri Jan 2001

Sistem Arkib Dalam Talian Adt / Ods, Mohd Radzi Nurul Azri

Student Works (2000-2009)

Dokumen ini adalah bertujuan untuk menggariskan pembangunan sistem sebenar di samping menggariskan panduan penggunaan untuk sistem tersebut. Bersama dokumen ini juga telah disertakan panduan pembangunan sistem dan juga senibinanya, Ini penting untuk memastikan sistem yang telah dibangunkan memenuhi kehendak pengguna dan dapat dibangunkan pada masa hadapan. Bagi projek ini sistem yang dibangunkan ialah sistem katalog dokumen bagi bilik dokumen Fakulti Sains Kompuler dan Teknologi Maklumat, Universiti Malaya. Bilik ini menyimpan arkib dokumen yang terdiri daripada buku-buku rujukan manual perisian, manual perkakasan laporan latihan industri dan projek ilmiah tahun akhir. Dokumen ini diarkibkan untuk rujukan masa depan masyarakat fakulti. Skop asas …


Hadiah Sastera Melayu Berasas Web, Shamsuddin Mohamed Zakie Jan 2001

Hadiah Sastera Melayu Berasas Web, Shamsuddin Mohamed Zakie

Student Works (2000-2009)

Projek Hadiah Sastera Melayu berasas web ini berkisar tentang sumber kesusasteraan Melayu. Projek ini amat bermatlamatkan untuk memartabatkan dunia kesusasteraan Melayu supaya dikenali dan memotivasikan para karyawan di Malaysia Supaya terus bergiat dalam bidang ini. Bermotifkan dengan gabungan 2 bidang berbeza iaitu sastera dan teknologi komputer diharap dapat memberi kesan kepada pengguna dan pembangunan sistem agar kedua-dua bidang dapat bergabung untuk kepentingan masyarakat di Malaysia. Teknologi komputer terus berkembang mengikut peredaran zaman. Peranannya yang meluas banyak memberi manfaat kepada berbagai bidang termasuklah bidang kesusasteraan. Tugasan yang dipikul adalah untuk mewujudkan satu sistem pengkalan data didalam persekitaran Internet dan diberi nama …


Sistem Pengurusan Inventori Perisian, Su Lin Tey Jan 2001

Sistem Pengurusan Inventori Perisian, Su Lin Tey

Student Works (2000-2009)

Sistem Pengurusan Inventori Perisian versi 1.0 SPIP v.1.0 telagh dibangunkan atas tujuan mengautomasikan aktiviti-aktiviti pengurusan inventori perisian dan sirkulasi bagi item-item perisian. SPIP adalah satu sistem komputer tersendiri (standalone) yang berupaya menjalankan aktiviti pengkatalogan, peminjaman, dan pemulangan bagi item perisian. Selain daripada itu, ia juga merekodkan maklumat peminjam, menjanakan laoran, dan menjalankan penyelenggaran bagi pengkalan data serta pengguna sistem. SPIP adalah rekabentuk khas kepada pengurusan perisian di Fakulti Sains Komputer dan Teknologi Maklumat, Universiti Malaya. Sistem yang dibangun akan digunakan oleh seorang pengatucara FSKTM, iaitu Puan Azlin. Visual Basic 6.0 digunakan sebagai alat pembangunan perisian yang utama manakala Microsoft Access …


Digital Library Of Thesis: Administration Side, Soo Noan Chew Jan 2001

Digital Library Of Thesis: Administration Side, Soo Noan Chew

Student Works (2000-2009)

Currently, the faculty promoted a plan to rethink the way graduate students present their theses. Traditionally, the theses were produced in paper fonn and storing in library. The access of student work through this method still facing a lot of problems including less timely public access to users, shelf space required for storage, no maintenance, resources are not fully used and cost of preparing theses are high. Thus, Digital Library of Theses is proposed to eliminate all the shortcomings mentioned above. Digital Library of Theses (administration side) is a web-based theses administration system for faculty that can cater about 3 …


Cataloging Expert Systems: Optimism And Frustrated Reality, William Olmstadt Feb 2000

Cataloging Expert Systems: Optimism And Frustrated Reality, William Olmstadt

E-JASL: Electronic Journal of Academic and Special Librarianship (1999-2009, Volumes 1-10)

There is little question that computers have profoundly changed how information professionals work. The process of cataloging and classifying library materials was one of the first activities transformed by information technology. The introduction of the MARC format in the 1960s and the creation of national bibliographic utilities in the 1970s had a lasting impact on cataloging. In the 1980s, the affordability of microcomputers made the computer accessible for cataloging, even to small libraries. This trend toward automating library processes with computers parallels a broader societal interest in the use of computers to organize and store information. Following World War II, …


Sistem Penjana Metadata (Metadata Generator), Zainal Zuraini Jan 2000

Sistem Penjana Metadata (Metadata Generator), Zainal Zuraini

Student Works (2000-2009)

Sistem Penjana Metadata merupakan sistem atas talian yang dinamakan Metadata Generator. Sistem ini dibangunkan bagi keperluan penjanaan metadata dan juga metatag yang mana ia membenarkan metadata tersebut dimasukkan ke dalam laman web. Perkhidmatan yang disediakan oleh sistem ini merupakan salab satu altematif bagi memenuhi keperluan pengorganisasian maklumat yang berkesan. Ini adalah untuk memastikan segala maklumat disusun dengan baik bagi keperluan capaian yang cekap dan lebih efisyen. Sistem yang dibangunkan ini menyediakan dua fungsi utama iaitu menjana metadata dan metatag yang mana metadata yang digunakan adalab mengikut format-format tertentu mengikut pilihan pengguna. Selain itu juga ia menyediakan templat dan juga panduan …