Open Access. Powered by Scholars. Published by Universities.®

Computer Sciences Commons™

Open Access. Powered by Scholars. Published by Universities.®

2013

Discipline
Institution
Keyword
Publication
Publication Type
File Type

Articles 901 - 930 of 2092

Full-Text Articles in Computer Sciences

Mining User Relations From Online Discussions Using Sentiment Analysis And Probabilistic Matrix Factorization, Minghui Qiu, Liu Yang, Jing Jiang Jun 2013

Mining User Relations From Online Discussions Using Sentiment Analysis And Probabilistic Matrix Factorization, Minghui Qiu, Liu Yang, Jing Jiang

Research Collection School Of Computing and Information Systems

Advances in sentiment analysis have enabled extraction of user relations implied in online textual exchanges such as forum posts. However, recent studies in this direction only consider direct relation extraction from text. As user interactions can be sparse in online discussions, we propose to apply collaborative filtering through probabilistic matrix factorization to generalize and improve the opinion matrices extracted from forum posts. Experiments with two tasks show that the learned latent factor representation can give good performance on a relation polarity prediction task and improve the performance of a subgroup detection task.


Architectural Control And Value Migration In Layered Ecosystems: The Case Of Open-Source Cloud Management Platforms, Richard Tee, C. Jason Woodard Jun 2013

Architectural Control And Value Migration In Layered Ecosystems: The Case Of Open-Source Cloud Management Platforms, Richard Tee, C. Jason Woodard

Research Collection School Of Computing and Information Systems

Our paper focuses on strategic decision making in layered business ecosystems, highlighting the role of cross-layer interactions in shaping choices about product design and platform governance. Based on evidence from the cloud computing ecosystem, we analyze how concerns about architectural control and expectations regarding future value migration influence the design of product interfaces and the degree of openness to external contributions. We draw on qualitative longitudinal data to trace the development of two open-source platforms for managing cloudbased computing resources. We focus in particular on the emergence of a layered "stack" in which these platforms must compete with both vertically …


Energy-Efficient Collaborative Query Processing Framework For Mobile Sensing Services, Jin Yang, Tianli Mo, Lipyeow Lim, Kai Uwe Sattler, Archan Misra Jun 2013

Energy-Efficient Collaborative Query Processing Framework For Mobile Sensing Services, Jin Yang, Tianli Mo, Lipyeow Lim, Kai Uwe Sattler, Archan Misra

Research Collection School Of Computing and Information Systems

Many emerging context-aware mobile applications involve the execution of continuous queries over sensor data streams generated by a variety of on-board sensors on multiple personal mobile devices (aka smartphones). To reduce the energyoverheads of such large-scale, continuous mobile sensing and query processing, this paper introduces CQP, a collaborative query processing framework that exploits the overlap (in both the sensor sources and the query predicates) across multiple smartphones. The framework automatically identifies the shareable parts of multiple executing queries, and then reduces the overheads of repetitive execution and data transmissions, by having a set of 'leader' mobile nodes execute and disseminate …


When Do Consumers Purchase Online?: Based On Inter-Purchase Time, Youngsoo Kim Jun 2013

When Do Consumers Purchase Online?: Based On Inter-Purchase Time, Youngsoo Kim

Research Collection School Of Computing and Information Systems

This study is motivated by the premise that online consumers can make a purchase at any time of day if they have even a tiny time slot along with Internet access. To identify the increased shopping time flexibility, we first characterize the patterns of online purchase timing in comparison to those in the offline market. The results show (1) the breakdown of purchase timing regularity and (2) the change of weekly spike purchase occurrence. Second, we build online inter-purchase time model and estimate it with the data collected from one of the premier online vendors in Korea. We verify new …


Mitigating Access-Driven Timing Channels In Clouds Using Stopwatch, Peng Li, Debin Gao, Michael K. Reiter Jun 2013

Mitigating Access-Driven Timing Channels In Clouds Using Stopwatch, Peng Li, Debin Gao, Michael K. Reiter

Research Collection School Of Computing and Information Systems

This paper presents StopWatch , a system that defends against timing-based side-channel attacks that arise from coresidency of victims and attackers in infrastructure-as-a-service clouds. StopWatchtriplicates each cloud-resident guest virtual machine (VM) and places replicas so that the three replicas of a guest VM are coresident with nonoverlapping sets of (replicas of) other VMs. StopWatch uses thetiming of I/O events at a VM's replicas collectively to determine the timings observed by each one or by an external observer, so that observable timing behaviors are similarly likely in the absence of any other individual, coresident VM. We detail the design and implementation …


A Latent Variable Model For Viewpoint Discovery From Threaded Forum Posts, Minghui Qiu, Jing Jiang Jun 2013

A Latent Variable Model For Viewpoint Discovery From Threaded Forum Posts, Minghui Qiu, Jing Jiang

Research Collection School Of Computing and Information Systems

No abstract provided.


Mining User Relations From Online Discussions Using Sentiment Analysis And Probabilistic Matrix Factorization, Minghui Qiu, Liu Yang, Jing Jiang Jun 2013

Mining User Relations From Online Discussions Using Sentiment Analysis And Probabilistic Matrix Factorization, Minghui Qiu, Liu Yang, Jing Jiang

Research Collection School Of Computing and Information Systems

No abstract provided.


Approximate Inference In Collective Graphical Models, Daniel Sheldon, Tao Sun, Akshat Kumar, Thomas G. Dietterich Jun 2013

Approximate Inference In Collective Graphical Models, Daniel Sheldon, Tao Sun, Akshat Kumar, Thomas G. Dietterich

Research Collection School Of Computing and Information Systems

We study the problem of approximate inference in collective graphical models (CGMs), which were recently introduced to model the problem of learning and inference with noisy aggregate observations. We first analyze the complexity of inference in CGMs: unlike inference in conventional graphical models, exact inference in CGMs is NP-hard even for tree-structured models. We then develop a tractable convex approximation to the NP-hard MAP inference problem in CGMs, and show how to use MAP inference for approximate marginal inference within the EM framework. We demonstrate empirically that these approximation techniques can reduce the computational cost of inference by two orders …


Enforcing Secure And Privacy-Preserving Information Brokering In Distributed Information Sharing, Fengjun Li, Bo Luo, Peng Liu, Dongwon Lee, Chao-Hsien Chu Jun 2013

Enforcing Secure And Privacy-Preserving Information Brokering In Distributed Information Sharing, Fengjun Li, Bo Luo, Peng Liu, Dongwon Lee, Chao-Hsien Chu

Research Collection School Of Computing and Information Systems

Today’s organizations raise an increasing need for information sharing via on-demand access. Information brokering systems (IBSs) have been proposed to connect large-scale loosely federated data sources via a brokering overlay, in which the brokers make routing decisions to direct client queries to the requested data servers. Many existing IBSs assume that brokers are trusted and thus only adopt server-side access control for data confidentiality. However, privacy of data location and data consumer can still be inferred from metadata (such as query and access control rules) exchanged within the IBS, but little attention has been put on its protection. In this …


Introducing Programmers To Pair Programming: A Controlled Experiment, A. S. M. Sajeev, Subhajit Datta Jun 2013

Introducing Programmers To Pair Programming: A Controlled Experiment, A. S. M. Sajeev, Subhajit Datta

Research Collection School Of Computing and Information Systems

Pair programming is a key characteristic of the Extreme Programming (XP) method. Through a controlled experiment we investigate pair programming behaviour of programmers without prior experience in XP. The factors investigated are: (a) characteristics of pair programming that are less favored (b) perceptions of team effectiveness and how they relate to product quality, and (c) whether it is better to train a pair by giving routine tasks first or by giving complex tasks first. Our results show that: (a) the least liked aspects of pair programming were having to share the screen, keyboard and mouse, and having to switch between …


Improving Internet Security Through Social Information And Social Comparison: A Field Quasi-Experiment, Qian Tang, Leigh L. Linden, John S. Quarterman, Andrew B. Whinston Jun 2013

Improving Internet Security Through Social Information And Social Comparison: A Field Quasi-Experiment, Qian Tang, Leigh L. Linden, John S. Quarterman, Andrew B. Whinston

Research Collection School Of Computing and Information Systems

Cybersecurity is a national priority in this big data era. Because of negative externalities and the resulting lack of economic incentives, companies often underinvest in security controls, despite government and industry recommendations. Although many existing studies on security have explored technical solutions, only a few have looked at the economic motivations. To fill the gap, we propose an approach to increase the incentives of organizations to address security problems. Specifically, we utilize and process existing security vulnerability data, derive explicit security performance information, and disclose the information as feedback to organizations and the public. We regularly release information on the …


Misheard Me Oronyminator: Using Oronyms To Validate The Correctness Of Frequency Dictionaries, Jennifer G. Hughes Jun 2013

Misheard Me Oronyminator: Using Oronyms To Validate The Correctness Of Frequency Dictionaries, Jennifer G. Hughes

Master's Theses

In the field of speech recognition, an algorithm must learn to tell the difference between "a nice rock" and "a gneiss rock". These identical-sounding phrases are called oronyms. Word frequency dictionaries are often used by speech recognition systems to help resolve phonetic sequences with more than one possible orthographic phrase interpretation, by looking up which oronym of the root phonetic sequence contains the most-common words.

Our paper demonstrates a technique used to validate word frequency dictionary values. We chose to use frequency values from the UNISYN dictionary, which tallies each word on a per-occurance basis, using a proprietary text corpus, …


Mantis: A Predictive Driving Directions Recommendation System, Christopher Hoover Jun 2013

Mantis: A Predictive Driving Directions Recommendation System, Christopher Hoover

Master's Theses

This thesis presents Mantis, a system designed to evaluate possible driving routes and recommend the optimal route based on current and predicted travel conditions. The system uses the Bing Maps REST service to obtain a set of routes. Traffic data from the California Department of Transportation’s Performance Measurement System (PeMS) is then used to estimate travel times for these routes. In addition to simple travel time estimation based on instantaneous traffic conditions, Mantis can use historic data to predict traffic speeds at future times. This allows Mantis to more effectively account for regularly repeating traffic patterns such as rush hour, …


An Analysis Of Generational Caching Implemented In A Production Website, Marc E. Zych Jun 2013

An Analysis Of Generational Caching Implemented In A Production Website, Marc E. Zych

Master's Theses

Website scaling has been an issue since the inception of the web. The demand for user generated content and personalized web pages requires the use of a database for a storage engine. Unfortunately, scaling the database to handle large amounts of traffic is still a problem many companies face. One such company is iFixit, a provider of free, publicly-editable, online repair manuals. Like many websites, iFixit uses Memcached to decrease database load and improve response time. However, the caching strategy used is a very ad hoc one and therefore can be greatly improved.

Most research regarding web application caching focuses …


Concept Graphs: Applications To Biomedical Text Categorization And Concept Extraction, Said Bleik May 2013

Concept Graphs: Applications To Biomedical Text Categorization And Concept Extraction, Said Bleik

Dissertations

As science advances, the underlying literature grows rapidly providing valuable knowledge mines for researchers and practitioners. The text content that makes up these knowledge collections is often unstructured and, thus, extracting relevant or novel information could be nontrivial and costly. In addition, human knowledge and expertise are being transformed into structured digital information in the form of vocabulary databases and ontologies. These knowledge bases hold substantial hierarchical and semantic relationships of common domain concepts. Consequently, automating learning tasks could be reinforced with those knowledge bases through constructing human-like representations of knowledge. This allows developing algorithms that simulate the human reasoning …


Novel Color And Local Image Descriptors For Content-Based Image Search, Sugata Banerji May 2013

Novel Color And Local Image Descriptors For Content-Based Image Search, Sugata Banerji

Dissertations

Content-based image classification, search and retrieval is a rapidly-expanding research area. With the advent of inexpensive digital cameras, cheap data storage, fast computing speeds and ever-increasing data transfer rates, millions of images are stored and shared over the Internet every day. This necessitates the development of systems that can classify these images into various categories without human intervention and on being presented a query image, can identify its contents in order to retrieve similar images.

Towards that end, this dissertation focuses on investigating novel image descriptors based on texture, shape, color, and local information for advancing content-based image search. Specifically, …


Introduction (2013), Eric Gossett May 2013

Introduction (2013), Eric Gossett

ACMS Conference Proceedings 2013

Nineteenth Conference of the Association of Christians in the Mathematical Sciences


Paper Abstracts (2013), Association Of Christians In The Mathematical Sciences May 2013

Paper Abstracts (2013), Association Of Christians In The Mathematical Sciences

ACMS Conference Proceedings 2013

Nineteenth Conference of the Association of Christians in the Mathematical Sciences


Genome Wide Search For Pseudo Knotted Non-Coding Rnas, Meghana S. Vasavada May 2013

Genome Wide Search For Pseudo Knotted Non-Coding Rnas, Meghana S. Vasavada

Theses

Non-coding RNAs (ncRNAs) are the functional RNA molecules that are involved in many biological processes including gene regulation, chromosome replication and RNA modification. Searching genomes using computational methods has become an important asset for prediction and annotation of ncRNAs. To annotate an individual genome for a specific family of ncRNAs, a computational tool is interpreted to scan through the genome and align its sequence segments to some structure model for the ncRNA family. With the recent advances in detecting an ncRNA in the genome, heuristic techniques are designed to perform an accurate search and sequence-structure alignment. This study uses a …


A Gpu Program To Compute Snp-Snp Interactions In Genome-Wide Association Studies, Srividya Ramakrishnan May 2013

A Gpu Program To Compute Snp-Snp Interactions In Genome-Wide Association Studies, Srividya Ramakrishnan

Theses

With the recent advances in the next generation sequencing technologies, short read sequences of human genome are made more accessible. Paired end sequencing of short reads is currently the most sensitive method for detecting somatic mutations that arise during tumor development. In this study, a novel approach to optimize the detection of structural variants using a new short read alignment program is presented.

Pairwise interaction effects of the Single Nucleotide Polymorphisms (SNPs) have proven to uncover the underlying complex disease traits. Computing the disease risk based on the interaction effects of SNPs on a case - control study is a …


Rna-Sequence Analysis Of Human Melanoma Cells, Jharna Miya May 2013

Rna-Sequence Analysis Of Human Melanoma Cells, Jharna Miya

Theses

RNA-sequencing refers to the use of high throughput sequencing technologies that are used to sequence cDNA in order to get the complete information of a sample’s RNA content. The objective of this study is to analyze this data in different aspects and to characterize gene expression. Besides this characterization, the data was also used to investigate the effect of sequencing depth on gene expression measurements.

This research focuses on quantitative measurement of expression levels of genes and their transcripts. In this study, complementary DNA fragments of cultured human melanoma cells are sequenced and a total of 139,501,106 million 200-bp reads …


Performance Comparison Of Five Rna-Seq Alignment Tools, Yuanpeng Lu May 2013

Performance Comparison Of Five Rna-Seq Alignment Tools, Yuanpeng Lu

Theses

Aligning millions of short reads to a reference genome is a critical task in high throughput sequencing. In recent years, a large number of mapping algorithms have been developed, all of which have in common that they align a vast number of reads to genomic or transcriptomic sequences. RNA-Seq data is discrete in nature, therefore with reasonable gene models and comparative metrics RNA-Seq data can be simulated to sufficient accuracy to enable meaningful benchmarking of alignment algorithms. To provide guidance in the choice of alignment algorithms, five different alignment tools for RNA-Seq data are evaluated. In order to compare the …


Polyaseeker: A Computational Framework For Identifying Polyadenylation Cleavage Site From Rna-Seq, Xiao Ling May 2013

Polyaseeker: A Computational Framework For Identifying Polyadenylation Cleavage Site From Rna-Seq, Xiao Ling

Theses

Alternative polyadenylation (APA) of mRNA plays a crucial role for post-transcriptional gene regulation. Recently, advances in next generation sequencing technology have made it possible to efficiently characterize the transcriptome and identify the 3’end of polyadenylated RNAs. However, no comprehensive bioi nformatic pipelines have fulfilled this goal. The PolyASeeker, a computational framework for identifying polyadenylation cleavage sites from RNA-Seq data is proposed in this thesis. By using the simulated RNA-seq dataset, a novel method is developed to evaluate the performance of the proposed framework versus the traditional A-stretch approach, and compute accurate Precisions and Recalls that previous estimation could not get. …


Schedule (2013), Association Of Christians In The Mathematical Sciences May 2013

Schedule (2013), Association Of Christians In The Mathematical Sciences

ACMS Conference Proceedings 2013

Nineteenth Conference of the Association of Christians in the Mathematical Sciences


Table Of Contents (2013), Association Of Christians In The Mathematical Sciences May 2013

Table Of Contents (2013), Association Of Christians In The Mathematical Sciences

ACMS Conference Proceedings 2013

Nineteenth Conference of the Association of Christians in the Mathematical Sciences


An Algorithm For Computing Edge Colorings On Regular Bipartite Multigraphs, Andrew S. Hannigan May 2013

An Algorithm For Computing Edge Colorings On Regular Bipartite Multigraphs, Andrew S. Hannigan

Dartmouth College Undergraduate Theses

In this paper, we consider the problem of finding an edge coloring of a d-regular bipartite multigraph with 2n vertices and m=nd edges. The best known deterministic algorithm (by Cole, Ost, and Schirra) takes O(m log d) time to find an edge coloring of such a graph. This bound is achieved by combining an O(m)-time perfect-matching algorithm with the Euler partition algorithm. The O(m) time bound on the Cole, Ost, and Schirra perfect-matching algorithm has been shown to be optimal. In this paper we present an alternative perfect-matching algorithm called QuickMatch. Empirical analysis shows that QuickMatch finds a perfect matching …


19th Conference Of The Associations Of Christians In The Mathematical Sciences, Association Of Christians In The Mathematical Sciences May 2013

19th Conference Of The Associations Of Christians In The Mathematical Sciences, Association Of Christians In The Mathematical Sciences

ACMS Conference Proceedings 2013

Association of Christians in the Mathematical Sciences 19th Biennial Conference Proceedings, May 29 - June 1, 2011, Bethel University.


Iterative Statistical Verification Of Probabilistic Plans, Colin M. Potts May 2013

Iterative Statistical Verification Of Probabilistic Plans, Colin M. Potts

Lawrence University Honors Projects

Artificial intelligence seeks to create intelligent agents. An agent can be anything: an autopilot, a self-driving car, a robot, a person, or even an anti-virus system. While the current state-of-the-art may not achieve intelligence (a rather dubious thing to quantify) it certainly achieves a sense of autonomy. A key aspect of an autonomous system is its ability to maintain and guarantee safety—defined as avoiding some set of undesired outcomes. The piece of software responsible for this is called a planner, which is essentially an automated problem solver. An advantage computer planners have over humans is their ability to consider and …


An Efficient And Probabilistic Secure Bit-Decomposition, Bharath K.K. Samanthula, Hu Chun, Wei Jiang May 2013

An Efficient And Probabilistic Secure Bit-Decomposition, Bharath K.K. Samanthula, Hu Chun, Wei Jiang

Computer Science Faculty Research & Creative Works

Many secure data analysis tasks, such as secure clustering and classification, require efficient mechanisms to convert the intermediate encrypted integers into the corresponding encryptions of bits. The existing bit-decomposition algorithms either do not offer sufficient security or are computationally inefficient. In order to provide better security as well as to improve efficiency, we propose a novel probabilistic-based secure bit-decomposition protocol for values encrypted using public key additive homomorphic encryption schemes. The proposed protocol guarantees security as per the semi-honest security definition of secure multi-party computation (MPC) and is also very efficient compared to the existing method. Our protocol always returns …


Cobweb Theorems With Production Lags And Price Forecasting, Daniel Dufresne, Felisa Vázquez-Abad May 2013

Cobweb Theorems With Production Lags And Price Forecasting, Daniel Dufresne, Felisa Vázquez-Abad

Publications and Research

The classical cobweb theorem is extended to include production lags and price forecasts. Price forecasting based on a longer period has a stabilizing effect on prices. Longer production lags do not necessarily lead to unstable prices; very long lags lead to cycles of constant amplitude. The classical cobweb requires elasticity of demand to be greater than that of supply; this is not necessarily the case in a more general setting. Random shocks are also considered.