Open Access. Powered by Scholars. Published by Universities.®
Graphics and Human Computer Interfaces Commons™
Open Access. Powered by Scholars. Published by Universities.®
- Institution
-
- Singapore Management University (313)
- University of Dayton (48)
- University of Arkansas, Fayetteville (10)
- University of Malaya (8)
- San Jose State University (7)
-
- Technological University Dublin (7)
- City University of New York (CUNY) (6)
- California Polytechnic State University, San Luis Obispo (5)
- Southern Adventist University (5)
- St. Mary's University (5)
- Institute of Business Administration (4)
- University of Nevada, Las Vegas (4)
- California State University, San Bernardino (3)
- Montclair State University (3)
- University of Nebraska at Omaha (3)
- Dakota State University (2)
- Governors State University (2)
- Minnesota State University Moorhead (2)
- Nova Southeastern University (2)
- Old Dominion University (2)
- Rochester Institute of Technology (2)
- The University of Akron (2)
- University of Nebraska - Lincoln (2)
- University of South Carolina (2)
- Arkansas Tech University (1)
- Ateneo de Manila University (1)
- Beirut Arab University (1)
- Bridgewater State University (1)
- Brigham Young University (1)
- Coastal Carolina University (1)
- Keyword
-
- Gamification (8)
- Deep learning (7)
- Visualization (7)
- Machine Learning (6)
- Collaboration (5)
-
- Deep Learning (5)
- Face recognition (5)
- Multimodal (5)
- Usability (5)
- Data visualization (4)
- Database (4)
- Education (4)
- Eye tracking (4)
- Few-shot learning (4)
- Food recognition (4)
- Graph neural networks (4)
- Human-computer interaction (4)
- Knowledge Graph (4)
- Recipe retrieval (4)
- Recommendation (4)
- Trust (4)
- Virtual worlds (4)
- Algorithms (3)
- Avatars (3)
- Click-through data (3)
- Clustering (3)
- Computer science (3)
- Databases (3)
- Design (3)
- E-commerce (3)
- Publication Year
- Publication
-
- Research Collection School Of Computing and Information Systems (310)
- Computer Science Faculty Publications (32)
- MAICS: The Modern Artificial Intelligence and Cognitive Science Conference (12)
- Student Works (2000-2009) (8)
- Graduate Theses and Dissertations (5)
-
- Conference papers (4)
- Articles (3)
- Campus Research Month (3)
- College of Engineering: Graduate Celebration Programs (3)
- Computer Science and Computer Engineering Undergraduate Honors Theses (3)
- Department of Computer Science Faculty Scholarship and Creative Works (3)
- Dissertations and Theses Collection (Open Access) (3)
- International Conference on Information and Communication Technologies (3)
- Master's Projects (3)
- Presentations - 2026 (3)
- All Capstone Projects (2)
- CCAC Theses and Dissertations (2)
- Computer Engineering (2)
- Computer Science Working Papers (2)
- Computer Science and Software Engineering (2)
- Dissertations, Theses, and Capstone Projects (2)
- Electronic Theses, Projects, and Dissertations (2)
- Posters - 2026 (2)
- Publications (2)
- Publications and Research (2)
- Research & Publications (2)
- SWITCH (2)
- Student Academic Conference (2)
- Theses and Dissertations (2)
- Theses/Capstones/Creative Projects (2)
- Publication Type
- File Type
Articles 361 - 390 of 473
Full-Text Articles in Graphics and Human Computer Interfaces
Topic Based Query Suggestions For Video Search, Kong-Wah Wan, Ah-Hwee Tan, Joo-Hwee Lim, Liang-Tien Chia
Topic Based Query Suggestions For Video Search, Kong-Wah Wan, Ah-Hwee Tan, Joo-Hwee Lim, Liang-Tien Chia
Research Collection School Of Computing and Information Systems
Query suggestion is an assistive technology mechanism commonly used in search engines to enable a user to formulate their search queries by predicting or completing the next few query words that the user is likely to type. In most implementations, the suggestions are mined from query log and use some simple measure of query similarity such as query frequency or lexicographical matching. In this paper, we propose an alternative method of presenting query suggestions by their thematic topics. Our method adopts a document-centric approach to mine topics in the corpus, and does not require the availability of a query log. …
Vireo@Trecvid 2011: Instance Search, Semantic Indexing, Multimedia Event Detection And Known-Item Search, Chong-Wah Ngo, Shi-Ai Zhu, Wei Zhang, Chun-Chet Tan, Ting Yao, Lei Pang, Hung-Khoon Tan
Vireo@Trecvid 2011: Instance Search, Semantic Indexing, Multimedia Event Detection And Known-Item Search, Chong-Wah Ngo, Shi-Ai Zhu, Wei Zhang, Chun-Chet Tan, Ting Yao, Lei Pang, Hung-Khoon Tan
Research Collection School Of Computing and Information Systems
The vireo group participated in four tasks: instance search, semantic indexing, multimedia event detection and known-item search. In this paper,we will present our approaches and discuss the evaluation results.
Context-Based Friend Suggestion In Online Photo-Sharing Community, Ting Yao, Chong-Wah Ngo, Tao Mei
Context-Based Friend Suggestion In Online Photo-Sharing Community, Ting Yao, Chong-Wah Ngo, Tao Mei
Research Collection School Of Computing and Information Systems
With the popularity of social media, web users tend to spend more time than before for sharing their experience and interest in online photo-sharing sites. The wide variety of sharing behaviors generate different metadata which pose new opportunities for the discovery of communities. We propose a new approach, named context-based friend suggestion, to leverage the diverse form of contextual cues for more effective friend suggestion in the social media community. Different from existing approaches, we consider both visual and geographical cues, and develop two user-based similarity measurements, i.e., visual similarity and geo similarity for characterizing user relationship. The problem of …
Learning Human Emotion Patterns For Modeling Virtual Humans, Shu Feng, Ah-Hwee Tan
Learning Human Emotion Patterns For Modeling Virtual Humans, Shu Feng, Ah-Hwee Tan
Research Collection School Of Computing and Information Systems
Emotion modeling is a crucial part in modeling virtual humans. Although various emotion models have been proposed, most of them focus on designing specific appraisal rules. As there is no unified framework for emotional appraisal, the appraisal variables have to be defined beforehand and evaluated in a subjective way. In this paper, we propose an emotion model based on machine learning methods by taking the following position: an emotion model should mirror actual human emotion in the real world and connect tightly with human inner states, such as drives, motivations and personalities. Specifically, a self-organizing neural model called Emotional Appraisal …
Automatic Content Generation For Video Self Modeling, Ju Shen, Anusha Raghunathan, Sen-Ching S. Cheung, Ravi R. Patel
Automatic Content Generation For Video Self Modeling, Ju Shen, Anusha Raghunathan, Sen-Ching S. Cheung, Ravi R. Patel
Computer Science Faculty Publications
Video self modeling (VSM) is a behavioral intervention technique in which a learner models a target behavior by watching a video of him or herself. Its effectiveness in rehabilitation and education has been repeatedly demonstrated but technical challenges remain in creating video contents that depict previously unseen behaviors. In this paper, we propose a novel system that re-renders new talking-head sequences suitable to be used for VSM treatment of patients with voice disorder. After the raw footage is captured, a new speech track is either synthesized using text-to-speech or selected based on voice similarity from a database of clean speeches. …
Developing Digital Field Guides For Plants: A Study From The Perspective Of Users, Emily Roseanne Schwarz
Developing Digital Field Guides For Plants: A Study From The Perspective Of Users, Emily Roseanne Schwarz
Master's Theses
A field guide is a tool to identify an object of natural history. Field guides
cover a wide range of topics from plants to fungi, birds to mammals, and shells to minerals. Traditionally, field guides are books, usually small enough to be carried outdoors . They enjoy wide popularity in modern life; almost every American home and library owns at least one field guide, and the same is also true for other areas of the world.
At this time, companies, non-profits, and universities are developing computer
technologies to replace printed field guides for identifying plants. This thesis
examines the state …
Adaptation Of The Nevada Climate Change Data Portal Web Interface To Small-Screen Mobile Devices, Tsvetan Komarov
Adaptation Of The Nevada Climate Change Data Portal Web Interface To Small-Screen Mobile Devices, Tsvetan Komarov
Festival of Communities: UG Symposium (Posters)
Robust and convenient access to the Nevada Climate Change Data Portal is vital for the project’s success, because of the researchers’ need to gather and analyze large volumes of data with minimal effort. However, the current version of the data portal web interface is not optimized for small-screen mobile devices such as mobile phones, PDAs, iPads, NetBooks, and others. The proposed research will address this issue by exploring the current methods for creating a client-aware web interface adaptable to the variety of small-screen devices, designing and implementing the most appropriate solution, and finally, user testing of the implemented solution.
In-Degree Dynamics Of Large-Scale P2p Systems, Zhongmei Yao, Daren B. H. Cline, Dmitri Loguinov
In-Degree Dynamics Of Large-Scale P2p Systems, Zhongmei Yao, Daren B. H. Cline, Dmitri Loguinov
Computer Science Faculty Publications
This paper builds a complete modeling framework for understanding user churn and in-degree dynamics in unstructured P2P systems in which each user can be viewed as a stationary alternating renewal process. While the classical Poisson result on the superposition of n stationary renewal processes for n→∞ requires that each point process become sparser as n increases, it is often difficult to rigorously show this condition in practice. In this paper, we first prove that despite user heterogeneity and non-Poisson arrival dynamics, a superposition of edge-arrival processes to a live user under uniform selection converges to a Poisson process when …
Mobile Visibility Querying For Lbs, James Carswell, Keith Gardiner, Junjun Jin
Mobile Visibility Querying For Lbs, James Carswell, Keith Gardiner, Junjun Jin
Articles
This article describes research carried out in the area of mobile spatial interaction (MSI) and the development of a 3D mobile version of a 2D web-based directional query processor. The TellMe application integrates location (from GPS, GSM, WiFi) and orientation (from magnetometer/accelerometer) sensor technologies into an enhanced spatial query processing module capable of exploiting a mobile device’s position and orientation for querying real-world spatial datasets. This article outlines our technique for combining these technologies and the architecture needed to deploy them on a sensor enabled smartphone (i.e. Nokia Navigator 6210). With all these sensor technologies now available on off-the-shelf devices, …
Understanding Gender Differences In Media Perceptions: A Comparison Of 2d Versus 3d Media, Fiona Fui-Hoon Nah, David Dewester, Brenda Eschenbrenner
Understanding Gender Differences In Media Perceptions: A Comparison Of 2d Versus 3d Media, Fiona Fui-Hoon Nah, David Dewester, Brenda Eschenbrenner
Research Collection School Of Computing and Information Systems
We examine gender differences in 2D versus 3D media perceptions. Using the Hunter-Gatherer Theory of Spatial Gender Differences and Jung’s Theory of Psychological Types, we hypothesize differences in men’s and women’s perceptions of skill, challenge, telepresence, and satisfaction with online experiences in 2D versus 3D media interaction. The findings suggest that even though women perceive lower skill levels and greater challenges in using 2D and 3D media than men, women’s sense of telepresence is higher than men in both 2D and 3D media. Women are also more satisfied with their interaction in 2D and 3D media than men.
Exploiting Intensity Inhomogeneity To Extract Textured Objects From Natural Scenes, Jundi Ding, Jialie Shen, Hwee Hwa Pang, Songcan Chen, Jingyu Yang
Exploiting Intensity Inhomogeneity To Extract Textured Objects From Natural Scenes, Jundi Ding, Jialie Shen, Hwee Hwa Pang, Songcan Chen, Jingyu Yang
Research Collection School Of Computing and Information Systems
Extracting textured objects from natural scenes is a challenging task in computer vision. The main difficulties arise from the intrinsic randomness of natural textures and the high-semblance between the objects and the background. In this paper, we approach the extraction problem with a seeded region-growing framework that purely exploits the statistical properties of intensity inhomogeneity. The pixels in the interior of potential textured regions are first found as texture seeds in an unsupervised manner. The labels of the texture seeds are then propagated through their respective inhomogeneous neighborhoods, to eventually cover the different texture regions in the image. Extensive experiments …
3-D Virtual World Education: An Empirical Comparison With Face-To-Face Classroom, Xiaofeng Chen, Keng Siau, Fiona Fui-Hoon Nah
3-D Virtual World Education: An Empirical Comparison With Face-To-Face Classroom, Xiaofeng Chen, Keng Siau, Fiona Fui-Hoon Nah
Research Collection School Of Computing and Information Systems
3-D virtual worlds are increasing in popularity as a means of pedagogical delivery in higher education. In this research, we assess the relative effectiveness of a 3-D virtual world learning environment, Second Life, and traditional face-to-face learning environment. We also assess the efficacy of instructional strategies in these two learning environments and their effects on interactivity, perceived learning, and satisfaction. Our findings suggest that there is an interaction effect of learning environment and instructional strategy. Pair-wise comparisons indicate that when interactive instructional strategy is used, there is no significant difference for perceived learning and satisfaction between 3-D virtual world and …
Vireo At Trecvid 2010: Semantic Indexing, Known-Item Search, And Content-Based Copy Detection, Chong-Wah Ngo, Shi-Ai Zhu, Hung-Khoon Tan, Wan-Lei Zhao
Vireo At Trecvid 2010: Semantic Indexing, Known-Item Search, And Content-Based Copy Detection, Chong-Wah Ngo, Shi-Ai Zhu, Hung-Khoon Tan, Wan-Lei Zhao
Research Collection School Of Computing and Information Systems
This paper presents our approaches and the comparative analysis of our results for the three TRECVID 2010 tasks that we participated in: semantic indexing, known-item search and content-based copy detection.
Program Transformations For Information Personalization, Saverio Perugini, Naren Ramakrishnan
Program Transformations For Information Personalization, Saverio Perugini, Naren Ramakrishnan
Computer Science Faculty Publications
Personalization constitutes the mechanisms necessary to automatically customize information content, structure, and presentation to the end user to reduce information overload. Unlike traditional approaches to personalization, the central theme of our approach is to model a website as a program and conduct website transformation for personalization by program transformation (e.g., partial evaluation, program slicing). The goal of this paper is study personalization through a program transformation lens and develop a formal model, based on program transformations, for personalized interaction with hierarchical hypermedia. The specific research issues addressed involve identifying and developing program representations and transformations suitable for classes of hierarchical …
Trajectory-Based Visualization Of Web Video Topics, Juan Cao, Chong-Wah Ngo, Yong-Dong Zhang, Dong-Ming Zhang, Liang Ma
Trajectory-Based Visualization Of Web Video Topics, Juan Cao, Chong-Wah Ngo, Yong-Dong Zhang, Dong-Ming Zhang, Liang Ma
Research Collection School Of Computing and Information Systems
While there have been research efforts in organizing largescale web videos into topics, efficient browsing of web video topics remains a challenging problem not yet addressed. The related issues include how to efficiently browse and track the evolution of topics and eventually locate the videos of interest. In this paper, we introduce a novel interface for visualizing video topics as evolution trajectories. The trajectory visualization is capable of highlighting milestone events and depicting the topical hotness over time. The interface also allows multi-level browsing from topics to events and to videos, resulting in search exploration could be more efficiently conducted …
Cast2face: Character Identification In Movie With Actor-Character Correspondence, Mengdi Xu, Xiaotong Yuan, Jialie Shen, Shuicheng Yan
Cast2face: Character Identification In Movie With Actor-Character Correspondence, Mengdi Xu, Xiaotong Yuan, Jialie Shen, Shuicheng Yan
Research Collection School Of Computing and Information Systems
We investigate the problem of automatically identifying characters in a movie with the supervision of actor-character name correspondence provided by the movie cast. Our proposed framework, namely Cast2Face, is featured by: (i) we restrict the names to assign within the set of character names in the cast; (ii) for each character, by using the corresponding actor's name as a key word, we retrieve from Google image search a group of face images to form the gallery set; and (iii) the probe face tracks in the movie are then identified as one of the actors by robust multi-task joint sparse representation …
3dq: Threat Dome Visibility Querying On Mobile Devices, James Carswell, Keith Gardiner, Junjun Yin
3dq: Threat Dome Visibility Querying On Mobile Devices, James Carswell, Keith Gardiner, Junjun Yin
Articles
3DQ (Three Dimensional Query) is our mobile spatial interaction (MSI) prototype for location and orientation aware mobile devices (i.e. today's sensor enabled smartphones). The prototype tailors a military style threat dome query calculation using MSI with hidden query removal functionality for reducing “information overload” on these off-the-shelf devices. The effect gives a more accurate and expected query result for Location-Based Services (LBS) applications by returning information on only those objects visible within a user’s 3D field-of-view. Our standardised XML based request/response design enables any mobile device, regardless of operating system and/or programming language, to access the 3DQ web-service interfaces.
Learning Personal Agents With Adaptive Player Modeling In Virtual Worlds, Yilin Kang, Ah-Hwee Tan
Learning Personal Agents With Adaptive Player Modeling In Virtual Worlds, Yilin Kang, Ah-Hwee Tan
Research Collection School Of Computing and Information Systems
There has been growing interest in creating intelligent agents in virtual worlds that do not follow fixed scripts predefined by the developers, but react accordingly based on actions performed by human players during their interaction. In order to achieve this objective, previous approaches have attempted to model the environment and the user’s context directly. However, a critical component for enabling personalized virtual world experience is missing, namely the capability to adapt over time to the habits and eccentricity of a particular player. To address the above issue, this paper presents a cognitive agent with learning player model capability for personalized …
Automatic Generation Of Semantic Fields For Annotating Web Images, Gang Wang, Tat Seng Chua, Chong-Wah Ngo, Yong Cheng Wang
Automatic Generation Of Semantic Fields For Annotating Web Images, Gang Wang, Tat Seng Chua, Chong-Wah Ngo, Yong Cheng Wang
Research Collection School Of Computing and Information Systems
The overwhelming amounts of multimedia contents have triggered the need for automatically detecting the semantic concepts within the media contents. With the development of photo sharing websites such as Flickr, we are able to obtain millions of images with usersupplied tags. However, user tags tend to be noisy, ambiguous and incomplete. In order to improve the quality of tags to annotate web images, we propose an approach to build Semantic Fields for annotating the web images. The main idea is that the images are more likely to be relevant to a given concept, if several tags to the image belong …
Semantic Context Modeling With Maximal Margin Conditional Random Fields For Automatic Image Annotation, Yu Xiang, Xiangdong Zhou, Zuotao Liu, Tat-Seng Chua, Chong-Wah Ngo
Semantic Context Modeling With Maximal Margin Conditional Random Fields For Automatic Image Annotation, Yu Xiang, Xiangdong Zhou, Zuotao Liu, Tat-Seng Chua, Chong-Wah Ngo
Research Collection School Of Computing and Information Systems
Context modeling for Vision Recognition and Automatic Image Annotation (AIA) has attracted increasing attentions in recent years. For various contextual information and resources, semantic context has been exploited in AIA and brings promising results. However, previous works either casted the problem into structural classification or adopted multi-layer modeling, which suffer from the problems of scalability or model efficiency. In this paper, we propose a novel discriminative Conditional Random Field (CRF) model for semantic context modeling in AIA, which is built over semantic concepts and treats an image as a whole observation without segmentation. Our model captures the interactions between semantic …
Personalization By Website Transformation: Theory And Practice, Saverio Perugini
Personalization By Website Transformation: Theory And Practice, Saverio Perugini
Computer Science Faculty Publications
We present an analysis of a progressive series of out-of-turn transformations on a hierarchical website to personalize a user’s interaction with the site. We formalize the transformation in graph-theoretic terms and describe a toolkit we built that enumerates all of the traversals enabled by every possible complete series of these transformations in any site and computes a variety of metrics while simulating each traversal therein to qualify the relationship between a site’s structure and the cumulative effect of support for the transformation in a site. We employed this toolkit in two websites. The results indicate that the transformation enables users …
Bay Audio Repair Website & Data Management Application, Michael Shelley
Bay Audio Repair Website & Data Management Application, Michael Shelley
Computer Science and Software Engineering
The goal of this senior project was to build a website and software application to receive and manage audio equipment repair requests for a small startup company called Bay Audio Repair (BAR). Furthermore, it allowed me to gain experience in web development and software engineering practices, specifically requirements gathering, design and implementation. The website provides an online interface for BAR’s customers to request repairs and the application allows BAR employees to update the progress of a repair. Several technologies were used in the system’s construction: HTML, XML, PHP, and C#.
Supporting Multiple Paths To Objects In Information Hierarchies: Faceted Classification, Faceted Search, And Symbolic Links, Saverio Perugini
Supporting Multiple Paths To Objects In Information Hierarchies: Faceted Classification, Faceted Search, And Symbolic Links, Saverio Perugini
Computer Science Faculty Publications
We present three fundamental, interrelated approaches to support multiple access paths to each terminal object in information hierarchies: faceted classification, faceted search, and web directories with embedded symbolic links. This survey aims to demonstrate how each approach supports users who seek information from multiple perspectives. We achieve this by exploring each approach, the relationships between these approaches, including tradeoffs, and how they can be used in concert, while focusing on a core set of hypermedia elements common to all. This approach provides a foundation from which to study, understand, and synthesize applications which employ these techniques. This survey does not …
Cbtv: Visualising Case Bases For Similarity Measure Design And Selection, Brian Mac Namee, Sarah Jane Delany
Cbtv: Visualising Case Bases For Similarity Measure Design And Selection, Brian Mac Namee, Sarah Jane Delany
Conference papers
In CBR the design and selection of similarity measures is paramount. Selection can benefit from the use of exploratory visualisation- based techniques in parallel with techniques such as cross-validation ac- curacy comparison. In this paper we present the Case Base Topology Viewer (CBTV) which allows the application of different similarity mea- sures to a case base to be visualised so that system designers can explore the case base and the associated decision boundary space. We show, using a range of datasets and similarity measure types, how the idiosyncrasies of particular similarity measures can be illustrated and compared in CBTV allowing …
Inside The Selection Box: Visualising Active Learning Selection Strategies, Brian Mac Namee, Rong Hu, Sarah Jane Delany
Inside The Selection Box: Visualising Active Learning Selection Strategies, Brian Mac Namee, Rong Hu, Sarah Jane Delany
Conference papers
Visualisations can be used to provide developers with insights into the inner workings of interactive machine learning techniques. In active learning, an inherently interactive machine learning technique, the design of selection strategies is the key research question and this paper demonstrates how spring model based visualisations can be used to provide insight into the precise operation of various selection strategies. Using sample datasets, this paper provides detailed examples of the differences between a range of selection strategies.
Exercise Power Grid Display And Web Interface, Alexander (Alex) Chernetz
Exercise Power Grid Display And Web Interface, Alexander (Alex) Chernetz
Computer Engineering
The 2008-2009 expansion of the Recreation Center at Cal Poly includes three new rooms with cardiovascular fitness equipment. As part of its ongoing commitment to sustainable development, the new machines connect to the main power grid and generate power during a workout. This document explains the process of quantifying and expressing the power generated using two interfaces: an autonomous display designed for a television with a text size and amount of detail adaptable to multiple television sizes and viewing distances, and an interactive, more detailed Web interface accessible with any Java-capable computer system or browser.
Vireo/Dvmm At Trecvid 2009: High-Level Feature Extraction, Automatic Video Search, And Content-Based Copy Detection, Chong-Wah Ngo, Yu-Gang Jiang, Xiao-Yong Wei, Wanlei Zhao, Yang Liu, Jun Wang, Shiai Zhu, Shih-Fu Chang
Vireo/Dvmm At Trecvid 2009: High-Level Feature Extraction, Automatic Video Search, And Content-Based Copy Detection, Chong-Wah Ngo, Yu-Gang Jiang, Xiao-Yong Wei, Wanlei Zhao, Yang Liu, Jun Wang, Shiai Zhu, Shih-Fu Chang
Research Collection School Of Computing and Information Systems
This paper presents overview and comparative analysis of our systems designed for 3 TRECVID 2009 tasks: high-level feature extraction, automatic search, and content-based copy detection.
Distribution-Based Concept Selection For Concept-Based Video Retrieval, Juan Cao, Hongfang Jing, Chong-Wah Ngo, Yongdong Zhang
Distribution-Based Concept Selection For Concept-Based Video Retrieval, Juan Cao, Hongfang Jing, Chong-Wah Ngo, Yongdong Zhang
Research Collection School Of Computing and Information Systems
Query-to-concept mapping plays one of the keys to concept-based video retrieval. Conventional approaches try to find concepts that are likely to co-occur in the relevant shots from the lexical or statistical aspects. However, the high probability of co-occurrence alone cannot ensure its effectiveness to distinguish the relevant shots from the irrelevant ones. In this paper, we propose distribution-based concept selection (DBCS) for query-to-concept mapping by analyzing concept score distributions of within and between relevant and irrelevant sets. In view of the imbalance between relevant and irrelevant examples, two variants of DBCS are proposed respectively by considering the two-sided and onesided …
Scalable Detection Of Partial Near-Duplicate Videos By Visual-Temporal Consistency, Hung-Khoon Tan, Chong-Wah Ngo, Richang Hong, Tat-Seng Chua
Scalable Detection Of Partial Near-Duplicate Videos By Visual-Temporal Consistency, Hung-Khoon Tan, Chong-Wah Ngo, Richang Hong, Tat-Seng Chua
Research Collection School Of Computing and Information Systems
Following the exponential growth of social media, there now exist huge repositories of videos online. Among the huge volumes of videos, there exist large numbers of near-duplicate videos. Most existing techniques either focus on the fast retrieval of full copies or near-duplicates, or consider localization in a heuristic manner. This paper considers the scalable detection and localization of partial near-duplicate videos by jointly considering visual similarity and temporal consistency. Temporal constraints are embedded into a network structure as directed edges. Through the structure, partial alignment is novelly converted into a network flow problem where highly efficient solutions exist. To precisely …
Domain Adaptive Semantic Diffusion For Large Scale Context-Based Video Annotation, Yu-Gang Jiang, Jun Wang, Shih-Fu Chang, Chong-Wah Ngo
Domain Adaptive Semantic Diffusion For Large Scale Context-Based Video Annotation, Yu-Gang Jiang, Jun Wang, Shih-Fu Chang, Chong-Wah Ngo
Research Collection School Of Computing and Information Systems
Learning to cope with domain change has been known as a challenging problem in many real-world applications. This paper proposes a novel and efficient approach, named domain adaptive semantic diffusion (DASD), to exploit semantic context while considering the domain-shift-of-context for large scale video concept annotation. Starting with a large set of concept detectors, the proposed DASD refines the initial annotation results using graph diffusion technique, which preserves the consistency and smoothness of the annotation over a semantic graph. Different from the existing graph learning methods which capture relations among data samples, the semantic graph treats concepts as nodes and the …