Open Access. Powered by Scholars. Published by Universities.®
Databases and Information Systems Commons™
Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Artificial Intelligence and Robotics (27)
- Library and Information Science (13)
- Social and Behavioral Sciences (13)
- Engineering (8)
- Other Computer Sciences (8)
-
- Software Engineering (8)
- Computer Engineering (7)
- Graphics and Human Computer Interfaces (7)
- Archival Science (6)
- Cataloging and Metadata (6)
- Collection Development and Management (6)
- Data Storage Systems (6)
- Information Security (6)
- Scholarly Communication (6)
- Scholarly Publishing (6)
- Systems Architecture (5)
- Theory and Algorithms (4)
- Arts and Humanities (3)
- Data Science (3)
- Education (3)
- Numerical Analysis and Scientific Computing (3)
- OS and Networks (3)
- Communication (2)
- Communication Technology and New Media (2)
- Computational Engineering (2)
- Digital Humanities (2)
- Engineering Education (2)
- Keyword
-
- Machine Learning (5)
- Natural Language Processing (5)
- Benchmarking (2)
- Blockchain (2)
- Chatbots (2)
-
- Collaborative filtering (2)
- MongoDB (2)
- News Aggregation (2)
- Question Answering System (2)
- Rudy Rucker (2)
- SVM (2)
- Scalability (2)
- Social networks (2)
- Text classification (2)
- Yioop (2)
- "Zip.1" (1)
- 1992 The Happy Muntant Handbook (1)
- Academic Libraries (1)
- Academic Library System (1)
- Apache HBase (1)
- Artificial Intelligence (AI) (1)
- Assembly Language (1)
- Association rule search geometric traversal problem (1)
- AutoComplete (1)
- Automation (1)
- Autosuggest entities solr (1)
- BOING bOING (1)
- Benefits of AI (1)
- Bias in AI Systems (1)
- Big data (1)
- Publication Year
- Publication
- Publication Type
- File Type
Articles 61 - 89 of 89
Full-Text Articles in Databases and Information Systems
Library Writers Reward Project, Saravana Kumar Gajendran
Library Writers Reward Project, Saravana Kumar Gajendran
Master's Projects
Open-source library development exploits the distributed intelligence of participants in Internet communities. Nowadays, contribution to the open-source community is fading [16] (Stackalytics, 2016) as there is not much recognition for library writers. They can start exploring ways to generate revenue as they actively contribute to the open-source community.
This project helps library writers to generate revenue in the form of bitcoins for their contribution. Our solution to generate revenue for library writers is to integrate bitcoin mining with existing JavaScript libraries, such as jQuery. More use of the library leads to more revenue for the library writers. It uses the …
Processing Posting Lists Using Opencl, Radha Kotipalli
Processing Posting Lists Using Opencl, Radha Kotipalli
Master's Projects
One of the main requirements of internet search engines is the ability to retrieve relevant results with faster response times. Yioop is an open source search engine designed and developed in PHP by Dr. Chris Pollett. The goal of this project is to explore the possibilities of enhancing the performance of Yioop by substituting resource-intensive existing PHP functions with C based native PHP extensions and the parallel data processing technology OpenCL. OpenCL leverages the Graphical Processing Unit (GPU) of a computer system for performance improvements.
Some of the critical functions in search engines are resource-intensive in terms of processing power, …
Concept Based Search Engine: Concept Creation, Aishwarya Rastogi
Concept Based Search Engine: Concept Creation, Aishwarya Rastogi
Master's Projects
Data on the internet is increasing exponentially every single second. There are billions and billions of documents on the World Wide Web (The Internet). Each document on the internet contains multiple concepts (an abstract or general idea inferred from specific instances).
In this paper, we show how we created and implemented an algorithm for extracting concepts from a set of documents. These concepts can be used by a search engine for generating search results to cater the needs of the user. The search result will then be more targeted than the usual keyword search.
The main problem was to extract …
Mining Concept In Big Data, Jingjing Yang
Mining Concept In Big Data, Jingjing Yang
Master's Projects
To fruitful using big data, data mining is necessary. There are two well-known methods, one is based on apriori principle, and the other one is based on FP-tree. In this project we explore a new approach that is based on simplicial complex, which is a combinatorial form of polyhedron used in algebraic topology. Our approach, similar to FP-tree, is top down, at the same time, it is based on apriori principle in geometric form, called closed condition in simplicial complex. Our method is almost 300 times faster than FP-growth on a real world database using a SJSU laptop. The database …
Index Strategies For Efficient And Effective Entity Search, Huy T. Vu
Index Strategies For Efficient And Effective Entity Search, Huy T. Vu
Master's Projects
The volume of structured data has rapidly grown in recent years, when data-entity emerged as an abstraction that captures almost every data pieces. As a result, searching for a desired piece of information on the web could be a challenge in term of time and relevancy because the number of matching entities could be very large for a given query. This project concerns with the efficiency and effectiveness of such entity queries. The work contains two major parts: implement inverted indexing strategies so that queries can be searched in minimal time, and rank results based on features that are independent …
An Open Source Advertisement Server, Pushkar Umaranikar
An Open Source Advertisement Server, Pushkar Umaranikar
Master's Projects
This report describes a new online advertisement system and its implementation for the Yioop open source search engine. This system was implemented for my CS298 project. It supports both selling advertisements and displaying them within search results. The selling of advertisement is done using a novel auction system, which we describe in this paper. With this auction system, it is possible to create an advertisement, attach keywords to it, and add it to the advertisement inventory. An advertisement is displayed on a search results page if the search keyword matches the keywords attached to the advertisement. Display of advertisements is …
Context-Based Autosuggest On Graph Data, Hai Nguyen
Context-Based Autosuggest On Graph Data, Hai Nguyen
Master's Projects
Autosuggest is an important feature in any search applications. Currently, most applications only suggest a single term based on how frequent that term appears in the indexed documents or how often it is searched upon. These approaches might not provide the most relevant suggestions because users often enter a series of related query terms to answer a question they have in mind. In this project, we implemented the Smart Solr Suggester plugin using a context-based approach that takes into account the relationships among search keywords. In particular, we used the keywords that the user has chosen so far in the …
A Scalable Search Engine Aggregator, Pooja Mishra
A Scalable Search Engine Aggregator, Pooja Mishra
Master's Projects
The ability to display different media sources in an appropriate way is an integral part of search engines such as Google, Yahoo, and Bing, as well as social networking sites like Facebook, etc. This project explores and implements various media-updating features of the open source search engine Yioop [1]. These include news aggregation, video conversion and email distribution. An older, preexisting news update feature of Yioop was modified and scaled so that it can work on many machines. We redesigned and modified the user interface associated with a distributed news updater feature in Yioop. This project also introduced a video …
How The University Of California Runs One Repository For Ten Campuses, Katie Fortney
How The University Of California Runs One Repository For Ten Campuses, Katie Fortney
Inaugural CSU IR Conference, 2015
Katie Fortney, JD, MLIS, Copyright Policy & Education Officer, Office of Scholarly Communication, University of California http://osc.universityofcalifornia.edu/
Implementing Metaarchive And Lockss At Digital Commons @Cal Poly, Michele Wyngard
Implementing Metaarchive And Lockss At Digital Commons @Cal Poly, Michele Wyngard
Inaugural CSU IR Conference, 2015
Michele Wyngard, Digital Repository Coordinator, CSU Cal Poly
Using Google Tag Manager And Google Analytics, (Code{4}Lib Journal), Suzanna Conrad
Using Google Tag Manager And Google Analytics, (Code{4}Lib Journal), Suzanna Conrad
Inaugural CSU IR Conference, 2015
Suzanna Conrad, Digital Initiatives Librarian, Cal Poly Pomona
What’S New Since The April 2013 Stim Ir Subcommittee Report To Cold: Hydra, Islandora And Dspace, Aaron Collier, Suzanna Conrad, Carmen Mitchell, Joan Parker, Andrew Weiss, Jeremy C. Shellhase
What’S New Since The April 2013 Stim Ir Subcommittee Report To Cold: Hydra, Islandora And Dspace, Aaron Collier, Suzanna Conrad, Carmen Mitchell, Joan Parker, Andrew Weiss, Jeremy C. Shellhase
Inaugural CSU IR Conference, 2015
Aaron Collier, Digital Repository Services Manager, Chancellor’s Office
Suzanna Conrad, Digital Initiatives Librarian, Cal Poly Pomona
Carmen Mitchell, Institutional Repository Librarian, CSU San Marcos
Joan Parker, Librarian, Moss Landing Marine Laboratories
Andrew Weiss, Digital Services Librarian, CSU Northridge
Jeremy Shellhase, Head of Information Services & Systems Department, Humboldt State University
The State Of Scholarworks, Aaron Collier
The State Of Scholarworks, Aaron Collier
Inaugural CSU IR Conference, 2015
Aaron Collier, Digital Repository Services Manager, Chancellor’s Office
A Content-Sensitive Wiki Help System, Eswara Satya Pavan Rajesh Pinapala
A Content-Sensitive Wiki Help System, Eswara Satya Pavan Rajesh Pinapala
Master's Projects
Context-sensitive help is a software application component that enables users to open help pertaining to their state, location, or the action they are performing within the software. Context-sensitive “wiki” help, on the other hand, is help powered by a wiki system with all the features of context-sensitive help. A context-sensitive wiki help system aims to make the context-sensitive help collaborative; in addition to seeking help, users can directly contribute to the help system. I have implemented a context-sensitive wiki help system into Yioop, an open source search engine and software portal created by Dr. Chris Pollett, in order to measure …
A Smart Web Crawler For A Concept Based Semantic Search Engine, Vinay Kancherla
A Smart Web Crawler For A Concept Based Semantic Search Engine, Vinay Kancherla
Master's Projects
The internet is a vast collection of billions of web pages containing terabytes of information arranged in thousands of servers using HTML. The size of this collection itself is a formidable obstacle in retrieving information necessary and relevant. This made search engines an important part of our lives. Search engines strive to retrieve information as relevant as possible to the user. One of the building blocks of search engines is the Web Crawler. A web crawler is a bot that goes around the internet collecting and storing it in a database for further analysis and arrangement of the data.
The …
A Trust-Aware System For Personalized User Recommendations In Social Networks, Magdalini Eirinaki, Malamati Louta, Iraklis Varlamis
A Trust-Aware System For Personalized User Recommendations In Social Networks, Magdalini Eirinaki, Malamati Louta, Iraklis Varlamis
Faculty Publications
Social network analysis has recently gained a lot of interest because of the advent and the increasing popularity of social media, such as blogs, social networking applications, microblogging, or customer review sites. In this environment, trust is becoming an essential quality among user interactions and the recommendation for useful content and trustful users is crucial for all the members of the network. In this paper, we introduce a framework for handling trust in social networks, which is based on a reputation mechanism that captures the implicit and explicit connections between the network members, analyzes the semantics and dynamics of these …
Users Positions In Social Networks, Jasim Qazi
Users Positions In Social Networks, Jasim Qazi
Master's Projects
Social networks are a new phase in human interaction: using
technology to connect people online, the social nehvorks of today have
become a central part of the lives of millions of people. People use social
networks for sharing various infonrmation with their friends and family. This
information can take the forn of text, video, images, sound etc. and it is what
forms the collection of dats in social networks.
As social networks gain popularity and as more and more people start
using social networks, it has become more important now to understand the
inner structures of social networks and understand …
Recipe Suggestion Tool, Sakuntala Padmapriya Gangaraju
Recipe Suggestion Tool, Sakuntala Padmapriya Gangaraju
Master's Projects
ABSTRACT
There is currently a great need for a tool to search cooking recipes based on ingredients. Current search engines do not provide this feature. Most of the recipe search results in current websites are not efficiently clustered based on relevance or categories resulting in a user getting lost in the huge search results presented.
Clustering in information retrieval is used for higher efficiency and better presentation of information to the user. Clustering puts similar documents in the same cluster. If a document is relevant to a query, then the documents in the same cluster are also relevant.
The goal …
Association Rule Mining -- Geometry And Parallel Computing Approach, Dongyi Jia
Association Rule Mining -- Geometry And Parallel Computing Approach, Dongyi Jia
Master's Projects
Mining association rules is a very important aspect in data mining fields. The process to mine association rules not only take much time, but also take huge computing source. How to fast and efficiently find the large itemsets is a crucial point in the association rule algorithms. This paper will focus on two algorithms research and implementation in parallel computing environments. One is Bitmap Combination algorithm, the other is Bitmap FP-Growth algorithm. Compared to Apriori algorithm, both Bitmap Combination and Bitmap FP-Growth algorithms don’t need generate candidate items, avoids costly database scans. Both algorithms need to translate the original database …
Smart Search: A Firefox Add-On To Compute A Web Traffic Ranking, Vijaya Pamidi
Smart Search: A Firefox Add-On To Compute A Web Traffic Ranking, Vijaya Pamidi
Master's Projects
Search engines results are typically ordered according to some notion of importance of a web page as well as relevance of the content of a web page to a query. Web page importance is usually calculated based on some graph theoretic properties of the web. Another common technique to measure page importance is to make use of the traffic that goes to a particular web page as measured by a browser toolbar. Currently, there are some traffic ranking tools available like www.alexa.com, www.ranking.com, www.compete.com that give such analytic as to the number of users who visit a web site. Alexa …
Medical Analysis Question And Answering Application For Internet Enabled Mobile Devices, Loc Nguyen
Medical Analysis Question And Answering Application For Internet Enabled Mobile Devices, Loc Nguyen
Master's Projects
Mobile devices such as smart phones, the iPhone, and the iPad have become more popular in recent years. With access to the Internet through cellular or WIFI networks, these mobile devices can make use of the great source of information available on the Internet. Unlike a desktop or laptop computer, an Internet enabled mobile device is designed to be carried around and available to the owner almost instantly at any moment of the day. Despite having such great advantage and potential, searching for information with a mobile device remains a difficult task. Mobile device users have to juggle between different …
Japanese Kanji Suggestion Tool, Sujata Dongre
Japanese Kanji Suggestion Tool, Sujata Dongre
Master's Projects
Many times, we can see that if we enter a misspelled search term in any of the search engines like Google, it will provide some help with "Did you mean:...." In my project, I am providing some suggestions for the wrong Japanese text entered by a user.
The Japanese language has three types of writing styles - hiragana, katakana and the most difficult, kanji. A single kanji symbol may be used to write one or more different compound words. From the point of view of the reader, kanji symbols are said to have one or more different "readings." Hence, sometimes …
An Online Versions Of Hyperlinked-Induced Topics Search (Hits) Algorithm, Amith Kollam Chandranna
An Online Versions Of Hyperlinked-Induced Topics Search (Hits) Algorithm, Amith Kollam Chandranna
Master's Projects
Generally, search engines perform the ranking of web pages in an offline mode, which is after the web pages have been retrieved and stored in the database. The existing HITS algorithm (Kleinberg, 1996) operates in an offline mode to perform page rank calculation. In this project, we have implemented an online mode of page ranking for this algorithm. This will improve the overall performance of the Search Engine. This report describes the approach used to implement and test this algorithm. Comparison results performed against other existing search engines are also presented in this report. This is helpful in describing the …
Planning Your Way To A More Usable Web Site, Pamela Gore, Sandra Hirsh
Planning Your Way To A More Usable Web Site, Pamela Gore, Sandra Hirsh
Faculty Publications
Planning for long-term periodic usability assessment is therefore as important as adding regularly fresh content and tracking usage. Fortunately, usability assessments need not be time consuming or expensive, unless your site is large and complex and you want to test it thoroughly each time. In a practical sense, usability assessment can reveal problems in the design, navigation, layout, or labeling that prevent users from finding what they need quickly. After analyzing your environment and setting the stage for ongoing usability assessment, it is time to develop the usability assessment plan, which will serve as the blueprint for usability assessment activities …
Entailment Mesh: A Wireless Brainstorming Tool Facilitated By A Server-Side Knowledge Engine, Joel Slayton
Entailment Mesh: A Wireless Brainstorming Tool Facilitated By A Server-Side Knowledge Engine, Joel Slayton
SWITCH
In this article, the author and founder of the CADRE Media Lab at San José State University, Joel Slayton, discusses the art project, "Entailment Mesh," by Gordon Pask and Paul Pangaro. Pask being an expert in the fields of cybernetics and psychology and Pangaro learned computer science and humanities while studying at MIT. The artwork utilizes a framework for a machine conversational and learning system called DoWhatDo developed by the artists at MIT. Slayton argues that the project can be considered both an art piece and a social software that blurs boundaries between structure and process. The concept of the …
The Quest For The Gnarl, Rudy Rucker
The Quest For The Gnarl, Rudy Rucker
SWITCH
The article describes some of the author’s own image-generating computer programs that he describes as “gnarly”. He began writing a simple spirograph program based off simple sine wave function called Spiro. Later transitioned into writing with C and better programs using more nonlinear feedback. Where Spiro is based on a simple sine wave function, Vine uses a nested sine function: the sine of the sine. The need for a more complicated computational approach lead to iteration and parallelism. Julgnarl uses Iteration and Calife uses parallelism. Calife shows one-dimensional cellular automata: spaces in which virtual computers are lined up like beads …
Gnarly Rantings About The Hacker And The Ants, Rudy Rucker
Gnarly Rantings About The Hacker And The Ants, Rudy Rucker
SWITCH
The article is an excerpt from Rucker’s book “The Happy Mutant”. It begins with his reflection of his career with GoMotion. He discusses the relation that he saw between design and cyberspace. Later he discusses his experience with a game a colleague found on the net: a virtual world where player is an ant. He talks about the struggles he goes through in this virtual world because of game difficulty and poor visuals. He ties it all in with how the Silicon Valley works in a similar way, and is filled with hackers and programers all needing each other to …
How I Got Gnarly, Rudy Rucker
How I Got Gnarly, Rudy Rucker
SWITCH
The article describes how Rudy Rucker’s curious interest in celluar automata led to his career in mathematical computer science at San José State University. After conducting interviews on the theory of cellular automata as a freelance writer, he felt compelled to be involved in this great intellectual revolution in computer-aided experimental mathematics. Committed to reinventing himself, Rucker's interactions with mathematicians inspired him to write “Mind Tools”, a book that surveys mathematics from the standpoint that is information. After publishing his book, in 1987, he was eventually offered a position at SJSU in the Mathematics and Computer Science department. With assistance …
The Effect Of Domain Knowledge On Elementary School Children's Search Behavior On An Information Retrieval System: The Science Library Catalog, Sandra Hirsh
Faculty Publications
Few information retrieval systems are designed with children’s special needs and capabilities in mind. We need to learn more about children’s information-seeking behavior in order to provide them with information-based tools which support exploratory learning. This dissertation examines children’s search behavior on a hypertext-based automated library catalog designed for elementary school children. The focus of this research is on the effect of domain knowledge on children’s search performance, search behavior, and learning as they look for science books on this system. Reseaxch has shown that level of domain knowledge in~luences the way people search for information. Data was collected through …