A Feasibility Study Of Crowdsourcing And Google Street View To Determine Sidewalk Accessibility,
2012
Singapore Management University
A Feasibility Study Of Crowdsourcing And Google Street View To Determine Sidewalk Accessibility, Kotaro Hara, Victoria Le, Jon Froehlich
Research Collection School Of Computing and Information Systems
We explore the feasibility of using crowd workers from Amazon Mechanical Turk to identify and rank sidewalk accessibility issues from a manually curated database of 100 Google Street View images. We examine the effect of three different interactive labeling interfaces (Point, Rectangle, and Outline) on task accuracy and duration. We close the paper by discussing limitations and opportunities for future work.
Talk Versus Work: Characteristics Of Developer Collaboration On The Jazz Platform,
2012
Singapore Management University
Talk Versus Work: Characteristics Of Developer Collaboration On The Jazz Platform, Subhajit Datta, Renuka Sindhgatta, Bikram Sengupta
Research Collection School Of Computing and Information Systems
IBM's Jazz initiative offers a state-of-the-art collaborative development environment (CDE) facilitating developer interactions around interdependent units of work. In this paper, we analyze development data across two versions of a major IBM product developed on the Jazz platform, covering in total 19 months of development activity, including 17,000+ work items and 61,000+ comments made by more than 190 developers in 35 locations. By examining the relation between developer talk and work, we find evidence that developers maintain a reasonably high level of connectivity with peer developers with whom they share work dependencies, but the span of a developer's communication goes …
Information Retrieval Based Nearest Neighbor Classification For Fine-Grained Bug Severity Prediction,
2012
Singapore Management University
Information Retrieval Based Nearest Neighbor Classification For Fine-Grained Bug Severity Prediction, Yuan Tian, David Lo, Chengnian Sun
Research Collection School Of Computing and Information Systems
Bugs are prevalent in software systems. Some bugs are critical and need to be fixed right away, whereas others are minor and their fixes could be postponed until resources are available. In this work, we propose a new approach leveraging information retrieval, in particular BM25-based document similarity function, to automatically predict the severity of bug reports. Our approach automatically analyzes bug reports reported in the past along with their assigned severity labels, and recommends severity labels to newly reported bug reports. Duplicate bug reports are utilized to determine what bug report features, be it textual, ordinal, or categorical, are important. …
Sensor Openflow: Enabling Software-Defined Wireless Sensor Networks,
2012
Singapore Management University
Sensor Openflow: Enabling Software-Defined Wireless Sensor Networks, Tie Luo, Hwee-Pink Tan, Tony Q. S. Quek
Research Collection School Of Computing and Information Systems
While it has been a belief for over a decade that wireless sensor networks (WSN) are application-specific, we argue that it can lead to resource underutilization and counter-productivity. We also identify two other main problems with WSN: rigidity to policy changes and difficulty to manage. In this paper, we take a radical, yet backward and peer compatible, approach to tackle these problems inherent to WSN. We propose a Software-Defined WSN architecture and address key technical challenges for its core component, Sensor OpenFlow. This work represents the first effort that synergizes software-defined networking and WSN.
Predicting Common Web Application Vulnerabilities From Input Validation And Sanitization Code Patterns,
2012
Singapore Management University
Predicting Common Web Application Vulnerabilities From Input Validation And Sanitization Code Patterns, Lwin Khin Shar, Hee Beng Kuan Tan
Research Collection School Of Computing and Information Systems
Software defect prediction studies have shown that defect predictors built from static code attributes are useful and effective. On the other hand, to mitigate the threats posed by common web application vulnerabilities, many vulnerability detection approaches have been proposed. However, finding alternative solutions to address these risks remains an important research problem. As web applications generally adopt input validation and sanitization routines to prevent web security risks, in this paper, we propose a set of static code attributes that represent the characteristics of these routines for predicting the two most common web application vulnerabilities—SQL injection and cross site scripting. In …
Duplicate Bug Report Detection With A Combination Of Information Retrieval And Topic Modeling,
2012
Iowa State University
Duplicate Bug Report Detection With A Combination Of Information Retrieval And Topic Modeling, Anh Tuan Nguyen, Tung Nguyen, Tien Nguyen, David Lo, Chengnian Sun
Research Collection School Of Computing and Information Systems
Detecting duplicate bug reports helps reduce triaging efforts and save time for developers in fixing the same issues. Among several automated detection approaches, text-based information retrieval (IR) approaches have been shown to outperform others in term of both accuracy and time efficiency. However, those IR-based approaches do not detect well the duplicate reports on the same technical issues written in different descriptive terms. This paper introduces DBTM, a duplicate bug report detection approach that takes advantage of both IR-based features and topic-based features. DBTM models a bug report as a textual document describing certain technical issue(s), and models duplicate bug …
Diversity Maximization Speedup For Fault Localization,
2012
Tsinghua University
Diversity Maximization Speedup For Fault Localization, Liang Gong, David Lo, Lingxiao Jiang, Hongyu Zhang
Research Collection School Of Computing and Information Systems
Fault localization is useful for reducing debugging effort. However, many fault localization techniques require non-trivial number of test cases with oracles, which can determine whether a program behaves correctly for every test input. Test oracle creation is expensive because it can take much manual labeling effort. Given a number of test cases to be executed, it is challenging to minimize the number of test cases requiring manual labeling and in the meantime achieve good fault localization accuracy. To address this challenge, this paper presents a novel test case selection strategy based on Diversity Maximization Speedup (DMS). DMS orders a set …
To What Extent Could We Detect Field Defects? An Empirical Study Of False Negatives In Static Bug Finding Tools,
2012
Singapore Management University
To What Extent Could We Detect Field Defects? An Empirical Study Of False Negatives In Static Bug Finding Tools, Ferdian Thung, Lucia Lucia, David Lo, Lingxiao Jiang, Premkumar Devanbu, Foyzur Rahman
Research Collection School Of Computing and Information Systems
Software defects can cause much loss. Static bug-finding tools are believed to help detect and remove defects. These tools are designed to find programming errors; but, do they in fact help prevent actual defects that occur in the field and reported by users? If these tools had been used, would they have detected these field defects, and generated warnings that would direct programmers to fix them? To answer these questions, we perform an empirical study that investigates the effectiveness of state-of-the-art static bug finding tools on hundreds of reported and fixed defects extracted from three open source programs: Lucene, Rhino, …
Scalable Content Authentication In H.264/Svc Videos Using Perceptual Hashing Based On Dempster-Shafer Theory,
2012
Singapore Management University
Scalable Content Authentication In H.264/Svc Videos Using Perceptual Hashing Based On Dempster-Shafer Theory, Dengpan Ye, Zhuo Wei, Xuhua Ding, Robert H. Deng
Research Collection School Of Computing and Information Systems
The content authenticity of the multimedia delivery is important issue with rapid development and widely used of multimedia technology. Till now many authentication solutions had been proposed, such as cryptology and watermarking based methods. However, in latest heterogeneous network the video stream transmission has b een coded in scalable way such as H.264/SVC, there is still no good authentication solution. In this paper, we firstly summarized related works and p roposed a scalable content authentication scheme using a ratio of different energy (RDE) based perceptual hashing in Q/S dimension, which is used Dempster-Shafer theory and combined with the latest scalable …
The Fat Thumb: Using The Thumb's Contact Size For Single-Handed Mobile Interaction,
2012
Singapore Management University
The Fat Thumb: Using The Thumb's Contact Size For Single-Handed Mobile Interaction, Sebastian Boring, David Ledo, Xiang ‘Anthony’ Chen, Nicolai Marquardt, Anthony Tang, Saul Greenberg
Research Collection School Of Computing and Information Systems
Modern mobile devices allow a rich set of multi-finger interactions that combine modes into a single fluid act, for example, one finger for panning blending into a two-finger pinch gesture for zooming. Such gestures require the use of both hands: one holding the device while the other is interacting. While on the go, however, only one hand may be available to both hold the device and interact with it. This mostly limits interaction to a single-touch (i.e., the thumb), forcing users to switch between input modes explicitly. In this paper, we contribute the Fat Thumb interaction technique, which uses the …
Adaptive In-Network Processing For Bandwidth And Energy Constrained Mission-Oriented Multi-Hop Wireless Networks,
2012
Telcordia Technologies
Adaptive In-Network Processing For Bandwidth And Energy Constrained Mission-Oriented Multi-Hop Wireless Networks, Sharanya Eswaran, James Edwards, Archan Misra, Thomas La Porta
Research Collection School Of Computing and Information Systems
In-network processing, involving operations such as filtering, compression and fusion, is a technique widely used in wireless sensor and ad hoc networks for reducing the communication overhead. In many tactical stream-oriented applications, especially in military scenarios, both link bandwidth and node energy are critically constrained resources. For such applications, in-network processing itself imposes non-negligible computing cost. In this work, we have developed a unified, utility-based closed-loop control framework that permits distributed convergence to both a) the optimal level of compression performed by a forwarding node on streams, and b) the best set of nodes where the operators of the stream …
Observatory Of Trends In Software Related Microblogs,
2012
Singapore Management University
Observatory Of Trends In Software Related Microblogs, Achananuparp Palakorn, Nelman Lubis Ibrahim, Yuan Tian, David Lo, Ee Peng Lim
Research Collection School Of Computing and Information Systems
Microblogging has recently become a popular means to disseminate information among millions of people. Interestingly, software developers also use microblog to communicate with one another. Different from traditional media, microblog users tend to focus on recency and informality of content. Many tweet contents are relatively more personal and Opinionated, compared to that of traditional news report. Thus, by analyzing microblogs, one could get the up-to-date information about what people are interested in or feel toward a particular topic. In this paper, we describe our microblog observatory that aggregates more than 70,000 Twitter feeds, captures software-related tweets, and computes trends from …
Detecting Similar Applications With Collaborative Tagging,
2012
Singapore Management University
Detecting Similar Applications With Collaborative Tagging, Ferdian Thung, David Lo, Lingxiao Jiang
Research Collection School Of Computing and Information Systems
Millions of people, including those in the software engineering communities have turned to microblogging services, such as Twitter, as a means to quickly disseminate information. A number of past studies by Treude et al., Storey, and Yuan et al. have shown that a wealth of interesting information is stored in these microblogs. However, microblogs also contain a large amount of noisy content that are less relevant to software developers in engineering software systems. In this work, we perform a preliminary study to investigate the feasibility of automatic classification of microblogs into two categories: relevant and irrelevant to engineering software systems. …
When Would This Bug Get Reported?,
2012
Singapore Management University
When Would This Bug Get Reported?, Ferdian Thung, David Lo, Lingxiao Jiang, Lucia Lucia, Foyzur Rahman, Premkumar Devanbu
Research Collection School Of Computing and Information Systems
Millions of people, including those in the software engineering communities have turned to microblogging services, such as Twitter, as a means to quickly disseminate information. A number of past studies by Treude et al., Storey, and Yuan et al. have shown that a wealth of interesting information is stored in these microblogs. However, microblogs also contain a large amount of noisy content that are less relevant to software developers in engineering software systems. In this work, we perform a preliminary study to investigate the feasibility of automatic classification of microblogs into two categories: relevant and irrelevant to engineering software systems. …
Kbe-Anonymity: Test Data Anonymization For Evolving Programs,
2012
Singapore Management University
Kbe-Anonymity: Test Data Anonymization For Evolving Programs, Lucia Lucia, David Lo, Lingxiao Jiang, Aditya Budi
Research Collection School Of Computing and Information Systems
High-quality test data that is useful for effective testing is often available on users’ site. However, sharing data owned by users with software vendors may raise privacy concerns. Techniques are needed to enable data sharing among data owners and the vendors without leaking data privacy. Evolving programs bring additional challenges because data may be shared multiple times for every version of a program. When multiple versions of the data are cross-referenced, private information could be inferred. Although there are studies addressing the privacy issue of data sharing for testing and debugging, little work has explicitly addressed the challenges when programs …
Interactive Fault Localization Leveraging Simple User Feedback,
2012
Tsinghua University
Interactive Fault Localization Leveraging Simple User Feedback, Liang Gong, David Lo, Lingxiao Jiang, Hongyu Zhang
Research Collection School Of Computing and Information Systems
Millions of people, including those in the software engineering communities have turned to microblogging services, such as Twitter, as a means to quickly disseminate information. A number of past studies by Treude et al., Storey, and Yuan et al. have shown that a wealth of interesting information is stored in these microblogs. However, microblogs also contain a large amount of noisy content that are less relevant to software developers in engineering software systems. In this work, we perform a preliminary study to investigate the feasibility of automatic classification of microblogs into two categories: relevant and irrelevant to engineering software systems. …
Automatic Classification Of Software Related Microblogs,
2012
Singapore Management University
Automatic Classification Of Software Related Microblogs, Philips Kokoh Prasetyo, David Lo, Achananuparp Palakorn, Yuan Tian, Ee Peng Lim
Research Collection School Of Computing and Information Systems
Millions of people, including those in the software engineering communities have turned to microblogging services, such as Twitter, as a means to quickly disseminate information. A number of past studies by Treude et al., Storey, and Yuan et al. have shown that a wealth of interesting information is stored in these microblogs. However, microblogs also contain a large amount of noisy content that are less relevant to software developers in engineering software systems. In this work, we perform a preliminary study to investigate the feasibility of automatic classification of microblogs into two categories: relevant and irrelevant to engineering software systems. …
A Framework To Annotate The Uncertainty For Geospatial Data,
2012
LSU New Orleans
A Framework To Annotate The Uncertainty For Geospatial Data, Zhao Yang
LSU New Orleans Theses and Dissertations
We have developed a new approach to annotate the uncertainty information of geospatial data. This framework is composed of a geospatial platform and the data with uncertainty. The framework supports geospatial sources such as Geography Markup Language (GML) with uncertainty information. The purpose of this framework is to integrate the uncertainty information of data from the application users and thereby ease the development of processing uncertainty information of geospatial data. Having well organized data and using this framework, the end-users can store the uncertainty information on the current geospatial data structure. For example, a GIS user can share the error …
Measurement-Driven Performance Analysis Of Indoor Femtocellular Networks,
2012
Singapore Management University
Measurement-Driven Performance Analysis Of Indoor Femtocellular Networks, Trung-Tuan Luong, Vigneshwaran Subbaraju, Archan Misra, Srinivasan Seshan
Research Collection School Of Computing and Information Systems
This paper describes initial empirical studies, performed on a 6-node 3G indoor femtocellular testbed, that investigate the impact of pedestrian mobility on network parameters, such as handoff behavior and data throughput. The studies establish that, owing to the small radii of cells, even modest changes in movement speed can have disproportionately large impact on handoff patterns and network throughput. By also revealing a strong temporal dependency effect, the studies motivate the need for algorithms to accurately predict RF signal strength distributions in dynamic indoor environments. We present such an RF prediction algorithm, based on crowd-sourced signal strength readings, and show …
Eliciting A Sensemaking Process From Verbal Protocols Of Reverse Engineers,
2012
711th Human Performance Wing
Eliciting A Sensemaking Process From Verbal Protocols Of Reverse Engineers, Adam R. Bryant, Robert F. Mills, Gilbert L. Peterson, Michael R. Grimaila
Faculty Publications
A process of sensemaking in reverse engineering was elicited from verbal protocols of reverse engineers as they investigated the assembly code of executable programs. Four participants were observed during task performance and verbal protocols were collected and analyzed from two of the participants to determine their problem-solving states and characterize likely transitions between those states. From this analysis, a high-level process of sensemaking is described which represents hypothesis generation and information-seeking behaviors in reverse engineering within a framework of goal-directed planning. Future work in validation and application of the process is discussed.
