Automatic Emotion Identification From Text,
2015
Wright State University - Main Campus
Automatic Emotion Identification From Text, Wenbo Wang
Kno.e.sis Publications
Emotions are both prevalent in and essential to most aspects of our lives. They in- fluence our decision-making, affect our social relationships and shape our daily behavior. With the rapid growth of emotion-rich textual content, such as microblog posts, blog posts, and forum discussions, there is a growing need to develop algorithms and techniques for identifying people’s emotions expressed in text. It has valuable implications for the studies of suicide prevention, employee productivity, well-being of people, customer relationship management, etc. However, emotion identification is quite challenging partly due to the following reasons: i) It is a multi-class classification problem that …
Era Of Big Data: Danger Of Descrimination,
2015
Sacred Heart University
Era Of Big Data: Danger Of Descrimination, Andra Gumbus, Frances Grodzinsky
WCBT Faculty Publications
We live in a world of data collection where organizations and marketers know our income, our credit rating and history, our love life, race, ethnicity, religion, interests, travel history and plans, hobbies, health concerns, spending habits and millions of other data points about our private lives. This data, mined for our behaviors, habits, likes and dislikes, is referred to as the “creep factor” of big data [1]. It is estimated that data generated worldwide will be 1.3 zettabytes (ZB) by 2016. The rise of computational power plus cheaper and faster devices to capture, collect, store and process data, translates into …
Real-Time Targeted Influence Maximization For Online Advertisements,
2015
Singapore Management University
Real-Time Targeted Influence Maximization For Online Advertisements, Yuchen Li, Dongxiang Zhang, Kian-Lee Tan
Research Collection School Of Computing and Information Systems
Advertising in social network has become a multi-billion dollar industry. A main challenge is to identify key influencers who can effectively contribute to the dissemination of information. Although the influence maximization problem, which finds a seed set of k most influential users based on certain propagation models, has been well studied, it is not target-aware and cannot be directly applied to online advertising. In this paper, we propose a new problem, named Keyword-Based Targeted Influence Maximization (KB-TIM), to find a seed set that maximizes the expected influence over users who are relevant to a given advertisement. To solve the problem, …
Developing Java Programs On Android Mobile Phones Using Speech Recognition,
2015
California State University - San Bernardino
Developing Java Programs On Android Mobile Phones Using Speech Recognition, Santhrushna Gande
Electronic Theses, Projects, and Dissertations
Nowadays Android operating system based mobile phones and tablets are widely used and had millions of users around the world. The popularity of this operating system is due to its multi-tasking, ease of access and diverse device options. “Java Programming Speech Recognition Application” is an Android application used for handicapped individuals who are not able or have difficultation to type on a keyboard. This application allows the user to write a compute program (in Java Language) by dictating the words and without using a keyboard. The user needs to speak out the commands and symbols required for his/her program. The …
Bioinformatics Approaches To Single-Cell Analysis In Developmental Biology,
2015
University of Nebraska-Lincoln
Bioinformatics Approaches To Single-Cell Analysis In Developmental Biology, Dicle Yalcin, Zeynep M. Hakguder, Hasan H. Otu
Department of Electrical and Computer Engineering: Faculty Publications
Individual cells within the same population show various degrees of heterogeneity, which may be better handled with single-cell analysis to address biological and clinical questions. Single-cell analysis is especially important in developmental biology as subtle spatial and temporal differences in cells have significant associations with cell fate decisions during differentiation and with the description of a particular state of a cell exhibiting an aberrant phenotype. Biotechnological advances, especially in the area of microfluidics, have led to a robust, massively parallel and multi-dimensional capturing, sorting, and lysis of single-cells and amplification of related macromolecules, which have enabled the use of imaging …
Cobweb: A Robust Map Update System Using Gps Trajectories,
2015
Singapore Management University
Cobweb: A Robust Map Update System Using Gps Trajectories, Zhangqing Shan, Hao Wu, Weiwei Sun, Baihua Zheng
Research Collection School Of Computing and Information Systems
The accuracy and completeness of a digital map plays a critical role in determining the quality of most location-based services. Unfortunately, road networks change frequently. Consequently, we study the issue of automatic map update in this paper. We propose a system called COBWEB which takes all the unmatched trajectories as input and generates the missing road segments with both the geometry properties and topology features well preserved. We conduct a comprehensive experimental study via real trajectory data generated by roughly 15,000 taxis in Singapore within a 5-month period. Compared with existing work, COBWEB demonstrates a better and more stable performance …
Towards Opinion Summarization From Online Forums,
2015
Singapore Management University
Towards Opinion Summarization From Online Forums, Ding Ying, Jing Jiang
Research Collection School Of Computing and Information Systems
Summarizing opinions expressed in online forums can potentially benefit many people. However, special characteristics of this problem may require changes to standard text summarization techniques. In this work, we present our initial attempt at extractive summarization of opinionated online forum threads. Given the nature of user generated content in online discussion forums, we hypothesize that besides relevance, text quality and subjectivity also play important roles in deciding which sentences are good summary sentences. We therefore construct an annotated corpus to facilitate our study of extractive summarization of online discussion forums. We define a set of features to capture relevance, text …
A Survey On Artificial Intelligence-Based Modeling Techniques For High Speed Milling Processes,
2015
Singapore Management University
A Survey On Artificial Intelligence-Based Modeling Techniques For High Speed Milling Processes, Amin Jahromi Torabi, Meng Joo Er, Xiang Li, Beng Siong Lim, Lianyin Zhai, Richard Jayadi Oentaryo, Gan Oon Peen, Jacek M. Zurada
Research Collection School Of Computing and Information Systems
The process of high speed milling is regarded as one of the most sophisticated and complicated manufacturing operations. In the past four decades, many investigations have been conducted on this process, aiming to better understand its nature and improve the surface quality of the products as well as extending tool life. To achieve these goals, it is necessary to form a general descriptive reference model of the milling process using experimental data, thermomechanical analysis, statistical or artificial intelligence (AI) models. Moreover, increasing demands for more efficient milling processes, qualified surface finishing, and modeling techniques have propelled the development of more …
Latent Factors Meet Homophily In Diffusion Modelling,
2015
Singapore Management University
Latent Factors Meet Homophily In Diffusion Modelling, Duc Minh Luu, Ee-Peng Lim
Research Collection School Of Computing and Information Systems
Diffusion is an important dynamics that helps spreading information within an online social network. While there are already numerous models for single item diffusion, few have studied diffusion of multiple items, especially when items can interact with one another due to their inter-similarity. Moreover, the well-known homophily effect is rarely considered explicitly in the existing diffusion models. This work therefore fills this gap by proposing a novel model called Topic level Interaction Homophily Aware Diffusion (TIHAD) to include both latent factor level interaction among items and homophily factor in diffusion. The model determines item interaction based on latent factors and …
Using Content-Level Structures For Summarizing Microblog Repost Trees,
2015
Singapore Management University
Using Content-Level Structures For Summarizing Microblog Repost Trees, Jing Li, Wei Gao, Zhongyu Wei, Baolin Peng, Kam-Fai Wong
Research Collection School Of Computing and Information Systems
A microblog repost tree provides strong clues on how an event described therein develops. To help social media users capture the main clues of events on microblogging sites, we propose a novel repost tree summarization framework by effectively differentiating two kinds of messages on repost trees called leaders and followers, which are derived from contentlevel structure information, i.e., contents of messages and the reposting relations. To this end, Conditional Random Fields (CRF) model is used to detect leaders across repost tree paths. We then present a variant of random-walk-based summarization model to rank and select salient messages based on the …
Candy Crushing Your Sleep,
2015
Singapore Management University
Candy Crushing Your Sleep, Kasthuri Jeyarajah, Meeralakshi Radhakrishnan, Steven C. H. Hoi, Archan Misra
Research Collection School Of Computing and Information Systems
Growing interest in quantified self has led to the popularity of lifelogging applications. In particular, health and wellness related applications have seen an upsurge with the advent of wearables such as the Fitbit. In this paper, we focus on the quality of sleep that directly impacts the overall wellness of individuals. In particular, in this work, we present a first of its kind study that (1) unobtrusively quantifies the quality of sleep and (2) seeks to identify attributing aspects of our daily lives such as an individual's usage of apps throughout the day and his/her physical environment that may affect …
Mining Revenue-Maximizing Bundling Configuration,
2015
Singapore Management University
Mining Revenue-Maximizing Bundling Configuration, Loc Do, Hady Wirawan Lauw, Ke Wang
Research Collection School Of Computing and Information Systems
With greater prevalence of social media, there is an increasing amount of user-generated data revealing consumer preferences for various products and services. Businesses seek to harness this wealth of data to improve their marketing strategies. Bundling, or selling two or more items for one price is a highly-practiced marketing strategy. In this paper, we address the bundle configuration problem from the data-driven perspective. Given a set of items in a seller’s inventory, we seek to determine which items should belong to which bundle so as to maximize the total revenue, by mining consumer preferences data. We show that this problem …
Maximum Rank Query,
2015
Singapore Management University
Maximum Rank Query, Kyriakos Mouratidis, Jilian Zhang, Hwee Hwa Pang
Research Collection School Of Computing and Information Systems
The top-k query is a common means to shortlist a number of options from a set of alternatives, based on the user's preferences. Typically, these preferences are expressed as a vector of query weights, defined over the options' attributes. The query vector implicitly associates each alternative with a numeric score, and thus imposes a ranking among them. The top-k result includes the k options with the highest scores. In this context, we define the maximum rank query (MaxRank). Given a focal option in a set of alternatives, the MaxRank problem is to compute the highest rank this option may achieve …
Tagcombine: Recommending Tags To Contents In Software Information Sites,
2015
Zhejiang University
Tagcombine: Recommending Tags To Contents In Software Information Sites, Xin Yu Wang, Xin Xia, David Lo
Research Collection School Of Computing and Information Systems
Nowadays, software engineers use a variety of online media to search and become informed of new and interesting technologies, and to learn from and help one another. We refer to these kinds of online media which help software engineers improve their performance in software development, maintenance, and test processes as software information sites. In this paper, we propose TagCombine, an automatic tag recommendation method which analyzes objects in software information sites. TagCombine has three different components: 1) multi-label ranking component which considers tag recommendation as a multi-label learning problem; 2) similarity-based ranking component which recommends tags from similar objects; 3) …
Answering Why-Not Questions On Reverse Top-K Queries,
2015
Zhejiang University
Answering Why-Not Questions On Reverse Top-K Queries, Yunjun Gao, Qing Liu, Gang Chen, Baihua Zheng, Linlin Zhou
Research Collection School Of Computing and Information Systems
Why-not questions, which aim to seek clarifications on the missing tuples for query results, have recently received considerable attention from the database community. In this paper, we systematically explore why-not questions on reverse top-k queries, owing to its importance in multi-criteria decision making. Given an initial reverse top-k query and a missing/why-not weighting vector set Wm that is absent from the query result, why-not questions on reverse top-k queries explain why Wm does not appear in the query result and provide suggestions on how to refine the initial query with minimum penalty to include Wm in the refined query result. …
A Joint Model Of Product Properties, Aspects And Ratings For Online Reviews,
2015
Singapore Management University
A Joint Model Of Product Properties, Aspects And Ratings For Online Reviews, Ding Ying, Jing Jiang
Research Collection School Of Computing and Information Systems
Product review mining is an important task that can benefit both businesses and consumers. Lately a number of models combining collaborative filtering and content analysis to model reviews have been proposed, among which the Hidden Factors as Topics (HFT) model is a notable one. In this work, we propose a new model on top of HFT to separate product properties and aspects. Product properties are intrinsic to certain products (e.g. types of cuisines of restaurants) whereas aspects are dimensions along which products in the same category can be compared (e.g. service quality of restaurants). Our proposed model explicitly separates the …
Did You Expect Your Users To Say This?: Distilling Unexpected Micro-Reviews For Venue Owners,
2015
Singapore Management University
Did You Expect Your Users To Say This?: Distilling Unexpected Micro-Reviews For Venue Owners, Wen-Haw Chong, Bingtian Dai, Ee-Peng Lim
Research Collection School Of Computing and Information Systems
With social media platforms such as Foursquare, users can now generate concise reviews, i.e. micro-reviews, about entities such as venues (or products). From the venue owner's perspective, analysing these micro-reviews will offer interesting insights, useful for event detection and customer relationship management. However not all micro-reviews are equally important, especially since a venue owner should already be familiar with his venue's primary aspects. Instead we envisage that a venue owner will be interested in micro-reviews that are unexpected to him. These can arise in many ways, such as users focusing on easily overlooked aspects (by the venue owner), making comparisons …
Name List Only? Target Entity Disambiguation In Short Texts,
2015
Singapore Management University
Name List Only? Target Entity Disambiguation In Short Texts, Yixin Cao, Juanzi Li, Xiaofei Guo, Shuanhu Bai, Heng Ji, Jie Tang
Research Collection School Of Computing and Information Systems
Target entity disambiguation (TED), the task of identifying target entities of the same domain, has been recognized as a critical step in various important applications. In this paper, we propose a graphbased model called TremenRank to collectively identify target entities in short texts given a name list only. TremenRank propagates trust within the graph, allowing for an arbitrary number of target entities and texts using inverted index technology. Furthermore, we design a multi-layer directed graph to assign different trust levels to short texts for better performance. The experimental results demonstrate that our model outperforms state-of-the-art methods with an average gain …
From Sensors To Sense Making: Leveraging Open-Access Scientific Data To Assess Arctic Maritime Risks,
2015
Dalhousie University
From Sensors To Sense Making: Leveraging Open-Access Scientific Data To Assess Arctic Maritime Risks, Mark A. Stoddard, Melanie Fournier Ph.D, Laurent Etienne Ph.D, Leah Beveridge Ph.D
ShipArc 2015 Conference
No abstract provided.
A System To Support Clerical Review, Correction, And Confirmation Assertions In Entity Identity Information Management,
2015
University of Arkansas Little Rock
A System To Support Clerical Review, Correction, And Confirmation Assertions In Entity Identity Information Management, Cheng Chen
Theses and Dissertations
Clerical review of Entity Resolution(ER) is crucial for maintaining the entity identity integrity of an Entity Identity Information Management (EIIM) system. However, the clerical review process presents several problems. These problems include Entity Identity Structures (EIS) that are difficult to read and interpret, excessive time and effort to review large Identity Knowledgebase (IKB), and the duplication of effort in repeatedly reviewing the same EIS in same EIIM review cycle or across multiple review cycles. Although the original EIIM model envisioned and demonstrated the value of correction assertions, these are applied to correct errors after they have been found. The original …
