Open Access. Powered by Scholars. Published by Universities.®
Databases and Information Systems Commons™
Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Social and Behavioral Sciences (75)
- Business (68)
- Numerical Analysis and Scientific Computing (42)
- Communication (39)
- Software Engineering (38)
-
- Engineering (30)
- OS and Networks (26)
- Information Security (25)
- Business Administration, Management, and Operations (24)
- Life Sciences (23)
- Medicine and Health Sciences (23)
- Theory and Algorithms (23)
- Artificial Intelligence and Robotics (22)
- Management Sciences and Quantitative Methods (21)
- Bioinformatics (20)
- Communication Technology and New Media (19)
- Graphics and Human Computer Interfaces (19)
- Social Media (19)
- Science and Technology Studies (17)
- Computer Engineering (16)
- Library and Information Science (16)
- Public Affairs, Public Policy and Public Administration (15)
- Health Information Technology (11)
- Technology and Innovation (10)
- Transportation (10)
- E-Commerce (9)
- Education (9)
- Institution
-
- Singapore Management University (201)
- Walden University (47)
- Wright State University (17)
- University of Nebraska at Omaha (11)
- Old Dominion University (8)
-
- Air Force Institute of Technology (4)
- Nova Southeastern University (4)
- San Jose State University (4)
- University of Nebraska - Lincoln (4)
- University of South Florida (4)
- Ateneo de Manila University (3)
- Bryant University (3)
- California Polytechnic State University, San Luis Obispo (3)
- Dakota State University (3)
- Louisiana State University (3)
- New Jersey Institute of Technology (3)
- Purdue University (3)
- Technological University Dublin (3)
- University for Business and Technology in Kosovo (3)
- Brigham Young University (2)
- City University of New York (CUNY) (2)
- Eastern Washington University (2)
- GALILEO, University System of Georgia (2)
- Portland State University (2)
- Southern Methodist University (2)
- The University of Akron (2)
- University of Kentucky (2)
- University of New Mexico (2)
- Virginia Commonwealth University (2)
- Boise State University (1)
- Keyword
-
- Machine learning (9)
- Machine Learning (7)
- Information technology (6)
- Social media (6)
- Artificial intelligence (5)
-
- Cloud computing (5)
- Cybersecurity (5)
- Data mining (5)
- Healthcare (5)
- Leadership (5)
- Online learning (5)
- Personas (5)
- Project Management (5)
- Big data (4)
- Blockchain (4)
- Clustering (4)
- Computer science (4)
- Cryptocurrency (4)
- Deep Learning (4)
- Information Systems (4)
- Information Technology (4)
- Innovation (4)
- Social Media (4)
- Support Vector Machine (4)
- Activity recognition (3)
- Attribute manipulation (3)
- Classification (3)
- Data models (3)
- Data stream (3)
- Database (3)
- Publication
-
- Research Collection School Of Computing and Information Systems (190)
- Walden Dissertations and Doctoral Studies (47)
- Kno.e.sis Publications (17)
- Dissertations and Theses Collection (Open Access) (7)
- Computer Science Faculty Publications (6)
-
- Information Systems and Quantitative Analysis Faculty Publications (5)
- CCAC Theses and Dissertations (4)
- USF Tampa Graduate Theses and Dissertations (4)
- Conference papers (3)
- Department of Information Systems & Computer Science Faculty Publications (3)
- Dissertations (3)
- Faculty Publications (3)
- Masters Theses & Doctoral Dissertations (3)
- The Summer Undergraduate Research Fellowship (SURF) Symposium (3)
- Theses and Dissertations (3)
- UNO Student Research and Creative Activity Fair (3)
- Computer Engineering (2)
- EWU Masters Thesis Collection (2)
- Engineering Management & Systems Engineering Faculty Publications (2)
- International Journal of Business and Technology (2)
- LSU Master's Theses (2)
- MITB Thought Leadership Series (2)
- Marriott Student Review (2)
- Open Educational Resources (2)
- SMU Data Science Review (2)
- Theses/Capstones/Creative Projects (2)
- Williams Honors College, Honors Research Projects (2)
- A with Honors Projects (1)
- Asian Management Insights (1)
- Biology Faculty Publications (1)
- Publication Type
Articles 91 - 120 of 376
Full-Text Articles in Databases and Information Systems
Implicit Linking Of Food Entities In Social Media, Wen Haw Chong, Ee Peng Lim
Implicit Linking Of Food Entities In Social Media, Wen Haw Chong, Ee Peng Lim
Research Collection School Of Computing and Information Systems
Dining is an important part in people’s lives and this explains why food-related microblogs and reviews are popular in social media. Identifying food entities in food-related posts is important to food lover profiling and food (or restaurant) recommendations. In this work, we conduct Implicit Entity Linking (IEL) to link food-related posts to food entities in a knowledge base. In IEL, we link posts even if they do not contain explicit entity mentions. We first show empirically that food venues are entity-focused and associated with a limited number of food entities each. Hence same-venue posts are likely to share common food …
Diversity In Online Advertising: A Case Study Of 69 Brands On Social Media, Jisun An, Ingmar Weber
Diversity In Online Advertising: A Case Study Of 69 Brands On Social Media, Jisun An, Ingmar Weber
Research Collection School Of Computing and Information Systems
Lack of diversity in advertising is a long-standing problem. Despite growing cultural awareness and missed business opportunities, many minorities remain under- or inappropriately represented in advertising. Previous research has studied how people react to culturally embedded ads, but such work focused mostly on print media or television using lab experiments. In this work, we look at diversity in content posted by 69 U.S. brands on two social media platforms, Instagram and Facebook. Using face detection technology, we infer the gender, race, and age of both the faces in the ads and of the users engaging with ads. Using this dataset, …
A Two-Stage Mechanism For Ordinal Peer Assessment, Zhize Li, Le Zhang, Zhixuan Fang, Jian Li
A Two-Stage Mechanism For Ordinal Peer Assessment, Zhize Li, Le Zhang, Zhixuan Fang, Jian Li
Research Collection School Of Computing and Information Systems
Peer assessment is a major method for evaluating the performance of employee, accessing the contributions of individuals within a group, making social decisions and many other scenarios. The idea is to ask the individuals of the same group to assess the performance of the others. Scores or rankings are then determined based on these evaluations. However, peer assessment can be biased and manipulated, especially when there is a conflict of interests. In this paper, we consider the problem of eliciting the underlying ordering (i.e. ground truth) of n strategic agents with respect to their performances, e.g., quality of work, contributions, …
Cognitive Antecedents Of Family Business Bias In Investment Decisions: A Commentary On 'Risky Decisions And The Family Firm Bias: An Experimental Study Based On Prospect Theory, H. Fang, Keng Siau, E. Memili, J. Dou
Cognitive Antecedents Of Family Business Bias In Investment Decisions: A Commentary On 'Risky Decisions And The Family Firm Bias: An Experimental Study Based On Prospect Theory, H. Fang, Keng Siau, E. Memili, J. Dou
Research Collection School Of Computing and Information Systems
Lude and Prügl explored “family business bias,” a cognitive tendency where the family nature of a firm can often reduce investors’ perceived risk in investments. As a result, investors would display lower risk-avoidance in the gain domain and reinforced risk-seeking in the loss domain. We expanded the authors’ work by introducing four cognitive factors (anchoring, representativeness, stereotype heuristic, and information availability) that can explain the underlying mechanisms behind the prevalence of “family business bias” and other cognitive misperceptions surrounding family businesses when it comes to investment decisions.
The Influence Of Conversational Agent Embodiment And Conversational Relevance On Socially Desirable Responding, Ryan M. Schuetzler, Justin Scott Giboney, G. Mark Grimes, Jay F. Nunamaker Jr.
The Influence Of Conversational Agent Embodiment And Conversational Relevance On Socially Desirable Responding, Ryan M. Schuetzler, Justin Scott Giboney, G. Mark Grimes, Jay F. Nunamaker Jr.
Information Systems and Quantitative Analysis Faculty Publications
Conversational agents (CAs) are becoming an increasingly common component in a wide range of information systems. A great deal of research to date has focused on enhancing traits that make CAs more humanlike. However, few studies have examined the influence such traits have on information disclosure. This research builds on self-disclosure, social desirability, and social presence theories to explain how CA anthropomorphism affects disclosure of personally sensitive information. Taken together, these theories suggest that as CAs become more humanlike, the social desirability of user responses will increase. In this study, we use a laboratory experiment to examine the influence of …
A Linked Coptic Dictionary Online, Frank Feder, Maxim Kupreyev, Emma Manning, Caroline T. Schroeder, Amir Zeldes
A Linked Coptic Dictionary Online, Frank Feder, Maxim Kupreyev, Emma Manning, Caroline T. Schroeder, Amir Zeldes
College of the Pacific Faculty Presentations
We describe a new project publishing a freely available online dictionary for Coptic. The dictionary encompasses comprehensive cross-referencing mechanisms, including linking entries to an online scanned edition of Crum’s Coptic Dictionary, internal cross-references and etymological information, translated searchable definitions in English, French and German, and linked corpus data which provides frequencies and corpus look-up for headwords and multiword expressions. Headwords are available for linking in external projects using a REST API. We describe the challenges in encoding our dictionary using TEI XML and implementing linking mechanisms to construct a Web interface querying frequency information, which draw on NLP tools to …
Expected Length Of The Longest Chain In Linear Hashing, Pongthip Srivarangkul, Hemanta K. Maji
Expected Length Of The Longest Chain In Linear Hashing, Pongthip Srivarangkul, Hemanta K. Maji
The Summer Undergraduate Research Fellowship (SURF) Symposium
Hash table with chaining is a data structure that chains objects with identical hash values together with an entry or a memory address. It works by calculating a hash value from an input then placing the input in the hash table entry. When we place two inputs in the same entry, they chain together in a linear linked list. We are interested in the expected length of the longest chain in linear hashing and methods to reduce the length because the worst-case look-up time is directly proportional to it.
The linear hash function used to calculate hash value is defined …
Predict The Failure Of Hydraulic Pumps By Different Machine Learning Algorithms, Yifei Zhou, Monika Ivantysynova, Nathan Keller
Predict The Failure Of Hydraulic Pumps By Different Machine Learning Algorithms, Yifei Zhou, Monika Ivantysynova, Nathan Keller
The Summer Undergraduate Research Fellowship (SURF) Symposium
Pump failure is a general concerned problem in the hydraulic field. Once happening, it will cause a huge property loss and even the life loss. The common methods to prevent the occurrence of pump failure is by preventative maintenance and breakdown maintenance, however, both of them have significant drawbacks. This research focuses on the axial piston pump and provides a new solution by the prognostic of pump failure using the classification of machine learning. Different kinds of sensors (temperature, acceleration and etc.) were installed into a good condition pump and three different kinds of damaged pumps to measure 10 of …
Sort Vs. Hash Join On Knights Landing Architecture, Victor L. Pan, Felix Lin
Sort Vs. Hash Join On Knights Landing Architecture, Victor L. Pan, Felix Lin
The Summer Undergraduate Research Fellowship (SURF) Symposium
With the increasing amount of information stored, there is a need for efficient database algorithms. One of the most important database operations is “join”. This involves combining columns from two tables and grouping common values in the same row in order to minimize redundant data. The two main algorithms used are hash join and sort merge join. Hash join builds a hash table to allow for faster searching. Sort merge join first sorts the two tables to make it more efficient when comparing values. There has been a lot of debate over which approach is superior. At first, hash join …
Zero Textbook Cost Syllabus For Cis 2200h (Introduction To Information Systems And Technologies), Curtis Izen
Zero Textbook Cost Syllabus For Cis 2200h (Introduction To Information Systems And Technologies), Curtis Izen
Open Educational Resources
This course introduces students to information systems in business. Due to the rapid developments in Information Technology (IT) and the dramatic changes brought by these new technologies in the way companies operate, compete and do business, familiarity with information systems has become indispensable for the leaders of today and tomorrow's organizations.
Trajectory-Driven Influential Billboard Placement, Ping Zhang, Zhifeng Bao, Yuchen Li, Guoliang Li, Yipeng Zhang, Zhiyong Peng
Trajectory-Driven Influential Billboard Placement, Ping Zhang, Zhifeng Bao, Yuchen Li, Guoliang Li, Yipeng Zhang, Zhiyong Peng
Research Collection School Of Computing and Information Systems
In this paper we propose and study the problem of trajectory-driven influential billboard placement: given a set of billboards U (each with a location and a cost), a database of trajectories T and a budget L, find a set of billboards within the budget to influence the largest number of trajectories. One core challenge is to identify and reduce the overlap of the influence from different billboards to the same trajectories, while keeping the budget constraint into consideration. We show that this problem is NP-hard and present an enumeration based algorithm with (1−1/e) approximation ratio. However, the enumeration should be …
Transaction Cost Optimization For Online Portfolio Selection, Bin Li, Jialei Wang, Dingjiang Huang, Steven C. H. Hoi
Transaction Cost Optimization For Online Portfolio Selection, Bin Li, Jialei Wang, Dingjiang Huang, Steven C. H. Hoi
Research Collection School Of Computing and Information Systems
To improve existing online portfolio selection strategies in the case of non-zero transaction costs, we propose a novel framework named Transaction Cost Optimization (TCO). The TCO framework incorporates the L1 norm of the difference between two consecutive allocations together with the principles of maximizing expected log return. We further solve the formulation via convex optimization, and obtain two closed-form portfolio update formulas, which follow the same principle as Proportional Portfolio Rebalancing (PPR) in industry. We empirically evaluate the proposed framework using four commonly used data-sets. Although these data-sets do not consider delisted firms and are thus subject to survival bias, …
Customer Segmentation Using Online Platforms: Isolating Behavioral And Demographic Segments For Persona Creation Via Aggregated User Data, Jisun An, Haewoon Kwak, Soon‑Gyo Jung, Joni Salminen, Bernard J. Jansen
Customer Segmentation Using Online Platforms: Isolating Behavioral And Demographic Segments For Persona Creation Via Aggregated User Data, Jisun An, Haewoon Kwak, Soon‑Gyo Jung, Joni Salminen, Bernard J. Jansen
Research Collection School Of Computing and Information Systems
We propose a novel approach for isolating customer segments using online customer data for products that are distributed via online social media platforms. We use non-negative matrix factorization to first identify behavioral customer segments and then to identify demographic customer segments. We employ a methodology for linking the two segments to present integrated and holistic customer segments, also known as personas. Behavioral segments are generated from customer interactions with online content. Demographic segments are generated using the gender, age, and location of these customers. In addition to evaluating our approach, we demonstrate its practicality via a system leveraging these customer …
Proactive And Reactive Resource/Task Allocation For Agent Teams In Uncertain Environments, Pritee Agrawal
Proactive And Reactive Resource/Task Allocation For Agent Teams In Uncertain Environments, Pritee Agrawal
Dissertations and Theses Collection (Open Access)
Synergistic interactions between task/resource allocation and multi-agent coordinated planning/assignment exist in many problem domains such as trans- portation and logistics, disaster rescue, security patrolling, sensor networks, power distribution networks, etc. These domains often feature dynamic environments where allocations of tasks/resources may have complex dependencies and agents may leave the team due to unforeseen conditions (e.g., emergency, accident or violation, damage to agent, reconfiguration of environment).
Secure Enforcement Of Isolation Policy On Multicore Platforms With Virtualization Techniques, Siqi Zhao
Secure Enforcement Of Isolation Policy On Multicore Platforms With Virtualization Techniques, Siqi Zhao
Dissertations and Theses Collection (Open Access)
A number of virtualization based systems have been proposed in the literature as an effective measure against the adversaries with the kernel privilege. However, under a systematic analysis, such systems exhibit vulnerabilities that can still be exploited by such an attacker with the kernel privilege. The fundamental reason is that there is an inherent incompatibility between the tamper-proof requirement and the complete mediation requirement of the reference monitor model. The incompatibility manifests in the virtualization based systems in the form of a discrepancy between the enforcement capability demanded by the high-level policy and the one achievable through the system design …
Unusual Events In Github Repositories, Christoph Treude, Larissa Leite, Maurício Aniche
Unusual Events In Github Repositories, Christoph Treude, Larissa Leite, Maurício Aniche
Research Collection School Of Computing and Information Systems
In large and active software projects, it becomes impractical for a developer to stay aware of all project activity. While it might not be necessary to know about each commit or issue, it is arguably important to know about the ones that are unusual. To investigate this hypothesis, we identified unusual events in 200 GitHub projects using a comprehensive list of ways in which an artifact can be unusual and asked 140 developers responsible for or affected by these events to comment on the usefulness of the corresponding information. Based on 2,096 answers, we identify the subset of unusual events …
Esg And Corporate Financial Performance: Empirical Evidence From China's Listed Power Generation Companies, Changhong Zhao, Yu Guo, Jiahai Yuan, Mengya Wu, Daiyu Li, Yiou Zhou, Jiangang Kang
Esg And Corporate Financial Performance: Empirical Evidence From China's Listed Power Generation Companies, Changhong Zhao, Yu Guo, Jiahai Yuan, Mengya Wu, Daiyu Li, Yiou Zhou, Jiangang Kang
Research Collection School Of Computing and Information Systems
Nowadays, listed companies around the world are shifting from short-term goals of maximizing profits to long-term sustainable environmental, social, and governance (ESG) goals. People have come to realize that ESG has become an important source of the corporate risk and may affect the company's financial performance and profitability. Recent research shows that good ESG performance could improve the financial performance in some countries. Yet, the question of how does ESG affect financial performance has not been thoroughly discussed and studied in China. In this article, we study China's listed power generation groups to explore the relationship between ESG performance and …
Offline Versus Online: A Meaningful Categorization Of Ties For Retweets, Felicia Natali, Feida Zhu
Offline Versus Online: A Meaningful Categorization Of Ties For Retweets, Felicia Natali, Feida Zhu
Research Collection School Of Computing and Information Systems
With the recent proliferation of news being shared through online social networks, it is crucial to determine how news is spread and what drives people to share certain stories. In this paper, we focus on the social networking site Twitter and analyse user’s retweets. We study retweeting patterns between offline and online friends, particularly, how tweet novelty and tweet topic differ between tweets retweeted by offline friends and those retweeted by online friends.
Knowledge As A Bridge: Improving Cross-Domain Answer Selection With External Knowledge, Yang Deng, Ying Shen, Min Yang, Yaliang Li, Nan Du, Wei Fan, Kai Lei
Knowledge As A Bridge: Improving Cross-Domain Answer Selection With External Knowledge, Yang Deng, Ying Shen, Min Yang, Yaliang Li, Nan Du, Wei Fan, Kai Lei
Research Collection School Of Computing and Information Systems
Answer selection is an important but challenging task. Significant progresses have been made in domains where a large amount of labeled training data is available. However, obtaining rich annotated data is a time-consuming and expensive process, creating a substantial barrier for applying answer selection models to a new domain which has limited labeled data. In this paper, we propose Knowledge-aware Attentive Network (KAN), a transfer learning framework for cross-domain answer selection, which uses the knowledge base as a bridge to enable knowledge transfer from the source domain to the target domains. Specifically, we design a knowledge module to integrate the …
Deep Learning For Practical Image Recognition: Case Study On Kaggle Competitions, Xulei Yang, Zeng Zeng, Sin G. Teo, Li Wang, Vijay Chandrasekar, Steven C. H. Hoi
Deep Learning For Practical Image Recognition: Case Study On Kaggle Competitions, Xulei Yang, Zeng Zeng, Sin G. Teo, Li Wang, Vijay Chandrasekar, Steven C. H. Hoi
Research Collection School Of Computing and Information Systems
In past years, deep convolutional neural networks (DCNN) have achieved big successes in image classification and object detection, as demonstrated on ImageNet in academic field. However, There are some unique practical challenges remain for real-world image recognition applications, e.g., small size of the objects, imbalanced data distributions, limited labeled data samples, etc. In this work, we are making efforts to deal with these challenges through a computational framework by incorporating latest developments in deep learning. In terms of two-stage detection scheme, pseudo labeling, data augmentation, cross-validation and ensemble learning, the proposed framework aims to achieve better performances for practical image …
A Characterization Of The Medical-Legal Partnership (Mlp) Of Nebraska Medicine, Jordan Pieper
A Characterization Of The Medical-Legal Partnership (Mlp) Of Nebraska Medicine, Jordan Pieper
Capstone Experience: Master of Public Health
This research study was completed at Legal Aid of Nebraska’s Health, Education, and Law Project through the partnership it has formed working with Nebraska Medicine and Iowa Legal Aid. Traditionally, health and disease have always been viewed exclusively as "healthcare" issues. But with healthcare consistently growing towards holistic approaches to help patients, we now know there are deeper, structural conditions of society that can act as strong driving forces of a person's poor daily living conditions that can negatively impact health. The importance of a Medical-Legal Partnership is that it considers a patient's social determinants of health (SDHs). The goal …
Exact Processing Of Uncertain Top-K Queries In Multi-Criteria Settings, Kyriakos Mouratidis, Bo Tang
Exact Processing Of Uncertain Top-K Queries In Multi-Criteria Settings, Kyriakos Mouratidis, Bo Tang
Research Collection School Of Computing and Information Systems
Traditional rank-aware processing assumes a dataset that contains available options to cover a specific need (e.g., restaurants, hotels, etc) and users who browse that dataset via top-k queries with linear scoring functions, i.e., by ranking the options according to the weighted sum of their attributes, for a set of given weights. In practice, however, user preferences (weights) may only be estimated with bounded accuracy, or may be inherently uncertain due to the inability of a human user to specify exact weight values with absolute accuracy. Motivated by this, we introduce the uncertain top-k query (UTK). Given uncertain preferences, that is, …
Probabilistic Collaborative Representation Learning For Personalized Item Recommendation, Aghiles Salah, Hady W. Lauw
Probabilistic Collaborative Representation Learning For Personalized Item Recommendation, Aghiles Salah, Hady W. Lauw
Research Collection School Of Computing and Information Systems
We present Probabilistic Collaborative Representation Learning (PCRL), a new generative model of user preferences and item contexts. The latter builds on the assumption that relationships among items within contexts (e.g., browsing session, shopping cart, etc.) may underlie various aspects that guide the choices people make. Intuitively, PCRL seeks representations of items reflecting various regularities between them that might be useful at explaining user preferences. Formally, it relies on Bayesian Poisson Factorization to model user-item interactions, and uses a multilayered latent variable architecture to learn representations of items from their contexts. PCRL seamlessly integrates both tasks within a joint framework. However, …
Embedding Wordnet Knowledge For Textual Entailment, Yunshi Lan, Jing Jiang
Embedding Wordnet Knowledge For Textual Entailment, Yunshi Lan, Jing Jiang
Research Collection School Of Computing and Information Systems
In this paper, we study how we can improve a deep learning approach to textual entailment by incorporating lexical entailment relations from WordNet. Our idea is to embed the lexical entailment knowledge contained in WordNet in specially-learned word vectors, which we call “entailment vectors.” We present a standard neural network model and a novel set-theoretic model to learn these entailment vectors from word pairs with known lexical entailment relations derived from WordNet. We further incorporate these entailment vectors into a decomposable attention model for textual entailment and evaluate the model on the SICK and the SNLI dataset. We find that …
Towards An Integrated Framework For Air Quality Monitoring And Exposure Estimation - A Review, Savina Singla, Divya Bansal, Archan Misra, Gaurav Raheja
Towards An Integrated Framework For Air Quality Monitoring And Exposure Estimation - A Review, Savina Singla, Divya Bansal, Archan Misra, Gaurav Raheja
Research Collection School Of Computing and Information Systems
For the health and safety of the public, it is essential to measure spatiotemporal distribution of air pollution in a region and thus monitor air quality in a fine-grain manner. While most of the sensing-based commercial applications available until today have been using fixed environmental sensors, the use of personal devices such as smartphones, smartwatches, and other wearable devices has not been explored in depth. These kinds of devices have an advantage of being with the user continuously, thus providing an ability to generate accurate and well-distributed spatiotemporal air pollution data. In this paper, we review the studies (especially in …
Secure And Efficient Outsourcing Of Large-Scale Overdetermined Systems Of Linear Equations, Shiran Pan, Wen-Tao Zhu, Qiongxiao Wang, Bing Chang
Secure And Efficient Outsourcing Of Large-Scale Overdetermined Systems Of Linear Equations, Shiran Pan, Wen-Tao Zhu, Qiongxiao Wang, Bing Chang
Research Collection School Of Computing and Information Systems
We address overdetermined systems of linear equations, where the number of unknowns is smaller than the number of equations so that only approximate solutions exist instead of exact solutions. Such systems are prevalent in many areas of science and engineering, and finding the optimal solutions is mathematically known as the linear least squares (LLS) problem. Real-world overdetermined systems are often large-scale and computationally expensive to solve. Consequently, we are interested in connecting the LLS problem with cloud computing, where a resource-constrained client outsources the problem to a powerful but untrusted cloud. Among several security considerations is that the input of …
Learning Representations Of Ultrahigh-Dimensional Data For Random Distance-Based Outlier Detection, Guansong Pang, Longbing Cao, Ling Chen, Defu Lian, Huan Liu
Learning Representations Of Ultrahigh-Dimensional Data For Random Distance-Based Outlier Detection, Guansong Pang, Longbing Cao, Ling Chen, Defu Lian, Huan Liu
Research Collection School Of Computing and Information Systems
Learning expressive low-dimensional representations of ultrahigh-dimensional data, e.g., data with thousands/millions of features, has been a major way to enable learning methods to address the curse of dimensionality. However, existing unsupervised representation learning methods mainly focus on preserving the data regularity information and learning the representations independently of subsequent outlier detection methods, which can result in suboptimal and unstable performance of detecting irregularities (i.e., outliers).This paper introduces a ranking model-based framework, called RAMODO, to address this issue. RAMODO unifies representation learning and outlier detection to learn low-dimensional representations that are tailored for a state-of-the-art outlier detection approach - the random …
Neural Collective Entity Linking, Yixin Cao, Lei Hou, Juanzi Li, Zhiyuan Liu
Neural Collective Entity Linking, Yixin Cao, Lei Hou, Juanzi Li, Zhiyuan Liu
Research Collection School Of Computing and Information Systems
Entity Linking aims to link entity mentions in texts to knowledge bases, and neural models have achieved recent success in this task. However, most existing methods rely on local contexts to resolve entities independently, which may usually fail due to the data sparsity of local information. To address this issue, we propose a novel neural model for collective entity linking, named as NCEL. NCEL applies Graph Convolutional Network to integrate both local contextual features and global coherence information for entity linking. To improve the computation efficiency, we approximately perform graph convolution on a subgraph of adjacent entity mentions instead of …
Use Of Artificial Intelligence, Machine Learning, And Autonomous Technologies In The Mining Industry, Z. Hyder, Keng Siau, Fiona Fui-Hoon Nah
Use Of Artificial Intelligence, Machine Learning, And Autonomous Technologies In The Mining Industry, Z. Hyder, Keng Siau, Fiona Fui-Hoon Nah
Research Collection School Of Computing and Information Systems
Mining is an important industrial and economic sector that plays a major role in the economic development of a country and provides many employment opportunities. Implementation of Artificial Intelligence (AI), machine learning, and autonomous technologies in the mining industry started about a decade ago with the first application to autonomous trucks. The autonomous technologies provide many economic benefits to the mining industry through cost reduction, productivity improvement, reduction in exposure of workers to hazardous conditions, continuous production, and improved safety. However, implementation of these technologies has faced economic, financial, technological, workforce, and social challenges. This paper discusses the current status …
Evaluation Criteria For Selecting Nosql Databases In A Single Box Environment, Ryan D. Engle, Brent T. Langhals, Michael R. Grimaila, Douglas D. Hodson
Evaluation Criteria For Selecting Nosql Databases In A Single Box Environment, Ryan D. Engle, Brent T. Langhals, Michael R. Grimaila, Douglas D. Hodson
Faculty Publications
In recent years, NoSQL database systems have become increasingly popular, especially for big data, commercial applications. These systems were designed to overcome the scaling and flexibility limitations plaguing traditional relational database management systems (RDBMSs). Given NoSQL database systems have been typically implemented in large-scale distributed environments serving large numbers of simultaneous users across potentially thousands of geographically separated devices, little consideration has been given to evaluating their value within single-box environments. It is postulated some of the inherent traits of each NoSQL database type may be useful, perhaps even preferable, regardless of scale. Thus, this paper proposes criteria conceived to …