Open Access. Powered by Scholars. Published by Universities.®

Engineering

Institution
Keyword
Publication Year
Publication
Publication Type

Articles 211 - 240 of 335

Full-Text Articles in Graphics and Human Computer Interfaces

An Adaptive Educational Game To Help Students Learn How To Solve Systems Of Linear Equations, Wang Zhang Aug 2014

An Adaptive Educational Game To Help Students Learn How To Solve Systems Of Linear Equations, Wang Zhang

Theses and Dissertations

Educational games have been proven to be effective in developing problem solving skills in well-defined domain, such as Math and Physics. In this thesis, an educational game called Matrix was developed to foster problem solving skills in the domain of linear algebra, particularly solving a system of linear equations. Matrix is an adaptive educational game that uses intelligent tutoring modules to guide the student's learning process and provide feedback based on the student's performance. These modules are domain module, student module, pedagogical module and presentation module. The domain module contains all the concepts the student needs to learn and an …


Click-Through-Based Cross-View Learning For Image Search, Yingwei Pan, Ting Yao, Tao Mei, Houqiang Li, Chong-Wah Ngo, Yong Rui Jul 2014

Click-Through-Based Cross-View Learning For Image Search, Yingwei Pan, Ting Yao, Tao Mei, Houqiang Li, Chong-Wah Ngo, Yong Rui

Research Collection School Of Computing and Information Systems

One of the fundamental problems in image search is to rank image documents according to a given textual query. Existing search engines highly depend on surrounding texts for ranking images, or leverage the query-image pairs annotated by human labelers to train a series of ranking functions. However, there are two major limitations: 1) the surrounding texts are often noisy or too few to accurately describe the image content, and 2) the human annotations are resourcefully expensive and thus cannot be scaled up. We demonstrate in this paper that the above two fundamental challenges can be mitigated by jointly exploring the …


Placing Videos On A Semantic Hierarchy For Search Result Navigation, Song Tan, Yu-Gang Jiang, Chong-Wah Ngo Jun 2014

Placing Videos On A Semantic Hierarchy For Search Result Navigation, Song Tan, Yu-Gang Jiang, Chong-Wah Ngo

Research Collection School Of Computing and Information Systems

Organizing video search results in a list view is widely adopted by current commercial search engines, which cannot support efficient browsing for complex search topics that have multiple semantic facets. In this article, we propose to organize video search results in a highly structured way. Specifically, videos are placed on a semantic hierarchy that accurately organizes various facets of a given search topic. To pick the most suitable videos for each node of the hierarchy, we define and utilize three important criteria: relevance, uniqueness, and diversity. Extensive evaluations on a large YouTube video dataset demonstrate the effectiveness of our approach.


Virtual Reality Engine Development, Varun Varahamurthy Jun 2014

Virtual Reality Engine Development, Varun Varahamurthy

Master's Theses

With the advent of modern graphics and computing hardware and cheaper sensor and display technologies, virtual reality is becoming increasingly popular in the fields of gaming, therapy, training and visualization. Earlier attempts at popularizing VR technology were plagued by issues of cost, portability and marketability to the general public. Modern screen technologies make it possible to produce cheap, light head-mounted displays (HMDs) like the Oculus Rift, and modern GPUs make it possible to create and deliver a seamless real-time 3D experience to the user. 3D sensing has found an application in virtual and augmented reality as well, allowing for a …


Coffee: Context Observer For Fast Enthralling Entertainment, Anthony M. Lenz Jun 2014

Coffee: Context Observer For Fast Enthralling Entertainment, Anthony M. Lenz

Master's Theses

Desktops, laptops, smartphones, tablets, and the Kinect, oh my! With so many devices available to the average consumer, the limitations and pitfalls of each interface are becoming more apparent. Swimming in devices, users often have to stop and think about how to interact with each device to accomplish the current tasks at hand. The goal of this thesis is to minimize user cognitive effort in handling multiple devices by creating a context aware hybrid interface. The context aware system will be explored through the hybridization of gesture and touch interfaces using a multi-touch coffee table and the next-generation Microsoft Kinect. …


An Analysis Of Eye Movements With Helmet Mounted Displays, Kalyn A. Tung Mar 2014

An Analysis Of Eye Movements With Helmet Mounted Displays, Kalyn A. Tung

Theses and Dissertations

Helmet or Head-Mounted Displays (HMD) applications have expanded to include a range from advanced military cockpits to consumer glasses. However, users have documented loss of legibility while undergoing vibration. Recent research indicates that undesirable eye movement is related to the vibration frequency a user experiences. In vibrating environments, two competing eye reflexes likely contribute to eye movements. The Vestibulo-ocular Reflex responds to motion sensed in the otoliths while the pursuit reflex is driven by the visual system to maintain the desired image on the fovea. This study attempts to isolate undesirable eye motions that occur while using a HMD by …


Conception, Design And Construction Of A Remote Wifi Vehicle Using Arduino, Clayton Broman Mar 2014

Conception, Design And Construction Of A Remote Wifi Vehicle Using Arduino, Clayton Broman

Physics

The scope of this senior project was to make a wireless vehicle controlled via Internet Protocol. This vehicle operates remotely and without direct line of sight. Commands are sent from a program running on a laptop and transmitted using a wireless router. Visual data is retrieved from a network camera, mounted on the vehicle, in real-time, to see where you are going.


Digital Display With Integrated Computing Circuit, Ronald S. Cok, John W. Harmer, Michael E. Miller Jan 2014

Digital Display With Integrated Computing Circuit, Ronald S. Cok, John W. Harmer, Michael E. Miller

AFIT Patents

A digital display device includes a display substrate; an array of pixels formed on the display substrate; an array of driving circuits located on the display substrate, each driving circuit electrically connected to one or more pixels for controlling a pixel current provided to each pixel; an array of computing circuits located on the display substrate, each computing circuit including circuits for signal or image processing and for communicating with neighboring computing circuits; a plurality of electrical conductors formed on the display substrate and connected to each of the driving circuits and digital computing circuits, wherein each computing circuit is …


Tonescale Compression For Electroluminescent Display, Michael E. Miller, Christopher J. White Nov 2013

Tonescale Compression For Electroluminescent Display, Michael E. Miller, Christopher J. White

AFIT Patents

A method for controlling an electroluminescent display to produce an image for display that has reduced luminance to reduce burn-in on the display while maintaining visible contrast, includes providing the electroluminescent (EL) display having a plurality of EL emitters, the luminance of the light produced by each EL emitter being responsive to a respective drive signal; receiving a respective input image signal for each EL emitter; and transforming the input image signals to a plurality of drive signals that have a reduced peak frame luminance value but maintains contrast in the displayed image to reduce burn-in by adjusting the drive …


Web-Based Visual Analytics For Social Media Data, Jun Xiang Tee, David S. Ebert Oct 2013

Web-Based Visual Analytics For Social Media Data, Jun Xiang Tee, David S. Ebert

The Summer Undergraduate Research Fellowship (SURF) Symposium

Social media data provides valuable information about different events, trends and happenings around the world. Visual data analysis tasks for social media data have large computational and storage space requirements. Due to these restrictions, subdivision of data analysis tools into several layers such as Data, Business Logic or Algorithms, and Presentation Layer is often necessary to make them accessible for variety of clients. On server side, social media data analysis algorithms can be implemented and published in the form of web services. Visual Interface can then be implemented in the form of thin clients that call these web services for …


Image Search By Graph-Based Label Propagation With Image Representation From Dnn, Yingwei Pan, Yao Ting, Kuiyuan Yang, Houqiang Li, Chong-Wah Ngo, Jingdong Wang, Tao Mei Oct 2013

Image Search By Graph-Based Label Propagation With Image Representation From Dnn, Yingwei Pan, Yao Ting, Kuiyuan Yang, Houqiang Li, Chong-Wah Ngo, Jingdong Wang, Tao Mei

Research Collection School Of Computing and Information Systems

Our objective is to estimate the relevance of an image to a query for image search purposes. We address two limitations of the existing image search engines in this paper. First, there is no straightforward way of bridging the gap between semantic textual queries as well as users’ search intents and image visual content. Image search engines therefore primarily rely on static and textual features. Visual features are mainly used to identify potentially useful recurrent patterns or relevant training examples for complementing search by image reranking. Second, image rankers are trained on query-image pairs labeled by human experts, making the …


Error Recovered Hierarchical Classification, Shiai Zhu, Xiao-Yong Wei, Chong-Wah Ngo Oct 2013

Error Recovered Hierarchical Classification, Shiai Zhu, Xiao-Yong Wei, Chong-Wah Ngo

Research Collection School Of Computing and Information Systems

Hierarchical classification (HC) is a popular and efficient way for detecting the semantic concepts from the images. However, the conventional HC, which always selects the branch with the highest classification response to go on, has the risk of propagating serious errors from higher levels of the hierarchy to the lower levels. We argue that the highestresponse-first strategy is too arbitrary, because the candidate nodes are considered individually which ignores the semantic relationship among them. In this paper, we propose a novel method for HC, which is able to utilize the semantic relationship among candidate nodes and their children to recover …


Annotation For Free: Video Tagging By Mining User Search Behavior, Yao Ting, Tao Mei, Chong-Wah Ngo, Shipeng Li Oct 2013

Annotation For Free: Video Tagging By Mining User Search Behavior, Yao Ting, Tao Mei, Chong-Wah Ngo, Shipeng Li

Research Collection School Of Computing and Information Systems

The problem of tagging is mostly considered from the perspectives of machine learning and data-driven philosophy. A fundamental issue that underlies the success of these approaches is the visual similarity, ranging from the nearest neighbor search to manifold learning, to identify similar instances of an example for tag completion. The need to searching for millions of visual examples in high-dimensional feature space, however, makes the task computationally expensive. Moreover, the results can suffer from robustness problem, when the underlying data, such as online videos, are rich of semantics and the similarity is difficult to be learnt from low-level features. This …


Click-Boosting Random Walk For Image Search Reranking, Xiaopeng Yang, Yongdong Zhang, Ting Yao, Zheng-Jun Zha, Chong-Wah Ngo Aug 2013

Click-Boosting Random Walk For Image Search Reranking, Xiaopeng Yang, Yongdong Zhang, Ting Yao, Zheng-Jun Zha, Chong-Wah Ngo

Research Collection School Of Computing and Information Systems

Image reranking is an effective way for improving the retrieval performance of keyword-based image search engines. A fundamental issue underlying the success of existing image reranking approaches is the ability in identifying potentially useful recurrent patterns or relevant training examples from the initial search results. Ideally, these patterns and examples can be leveraged to upgrade the ranks of visually similar images, which are also likely to be relevant. The challenge, nevertheless, originates from the fact that keyword-based queries are used to be ambiguous, resulting in difficulty in predicting the search intention. Mining useful patterns and examples without understanding query is …


Channels: Easy Video Content Consumption, Alexander R. Ledwith Jun 2013

Channels: Easy Video Content Consumption, Alexander R. Ledwith

Computer Engineering

The idea for this project is to take a user’s digital movie and television show library, and organize the individual videos into lists that will continuously and concurrently play like cable TV channels. This means that when a list of channels is switched to, video content will automatically start playing, possibly in the middle, based on a schedule. It should serve the needs of the user by allowing the user to quickly watch any of their existing channels, easily add a new channel based on chosen criteria or manual selection, and easily add new video content. In addition, the product …


Searching Visual Instances With Topology Checking And Context Modeling, Wei Zhang, Chong-Wah Ngo Apr 2013

Searching Visual Instances With Topology Checking And Context Modeling, Wei Zhang, Chong-Wah Ngo

Research Collection School Of Computing and Information Systems

Instance Search (INS) is a realistic problem initiated by TRECVID, which is to retrieve all occurrences of the querying object, location, or person from a large video collection. It is a fundamental problem with many applications, and also a challenging problem different from the traditional concept or near-duplicate (ND) search, since the relevancy is defined at instance level. True responses could exhibit various visual variations, such as being small on the image with different background, or showing a non-homography spatial configuration. Based on the Bag-of-Words model, we propose two techniques tailored for Instance Search. Specifically, we explore the use of …


Circular Reranking For Visual Search, Ting Yao, Chong-Wah Ngo Apr 2013

Circular Reranking For Visual Search, Ting Yao, Chong-Wah Ngo

Research Collection School Of Computing and Information Systems

Search reranking is regarded as a common way to boost retrieval precision. The problem nevertheless is not trivial especially when there are multiple features or modalities to be considered for search, which often happens in image and video retrieval. This paper proposes a new reranking algorithm, named circular reranking, that reinforces the mutual exchange of information across multiple modalities for improving search performance, following the philosophy that strong performing modality could learn from weaker ones, while weak modality does benefit from interacting with stronger ones. Technically, circular reranking conducts multiple runs of random walks through exchanging the ranking scores among …


An Investigation And Analysis Of The Vestibulo-Ocular Reflex In A Vibration Environment, Daniel J. Uribe Mar 2013

An Investigation And Analysis Of The Vestibulo-Ocular Reflex In A Vibration Environment, Daniel J. Uribe

Theses and Dissertations

Forty years of innovations have greatly improved Helmet-Mounted Displays (HMDs) and their integration into military systems. However, a significant issue with HMDs is the effect of vibration and the associated Vestibulo-Ocular Reflex (VOR). When a human’s head is subject to low-frequency vibration, the VOR stabilizes the eye with respect to objects in the external environment. However, this response is inappropriate in HMDs as the display moves with the user’s head and the VOR blurs the image as it is projected on the human retina. Current compensation techniques suggest increasing the size of displayed graphics or text to compensate for the …


Senior Project Report - Doctest, Stephen Weessies Mar 2013

Senior Project Report - Doctest, Stephen Weessies

Computer Engineering

DocTest is a program that, simply put, allows a programmer or user to document STANAG 4586 (a standard for unmanned aerial vehicle interoperability) messages and test the vehicle system at Lockheed Martin [5]. The program is extensible to allow for further development aiding our software team to do what they do best and not get bogged down in tedious but necessary documentation. DocTest is also used to aid in testing, keeping track of the issues and bugs found and creating a document that captures each issue so an issue is not missed or forgotten. This program was made for use …


A Robust Rgb-D Slam System For 3d Environment With Planar Surfaces, Po-Chang Su Jan 2013

A Robust Rgb-D Slam System For 3d Environment With Planar Surfaces, Po-Chang Su

Theses and Dissertations--Electrical and Computer Engineering

Simultaneous localization and mapping is the technique to construct a 3D map of unknown environment. With the increasing popularity of RGB-depth (RGB-D) sensors such as the Microsoft Kinect, there have been much research on capturing and reconstructing 3D environments using a movable RGB-D sensor. The key process behind these kinds of simultaneous location and mapping (SLAM) systems is the iterative closest point or ICP algorithm, which is an iterative algorithm that can estimate the rigid movement of the camera based on the captured 3D point clouds. While ICP is a well-studied algorithm, it is problematic when it is used in …


Electronic Device, Display And Touch-Sensitive User Interface (Dec 2012), Michael E. Miller, Jerald J. Muszak, Michael J. Telek Dec 2012

Electronic Device, Display And Touch-Sensitive User Interface (Dec 2012), Michael E. Miller, Jerald J. Muszak, Michael J. Telek

AFIT Patents

Display devices and methods for operating the same are provided. In one embodiment, the display device has an electronic display having an active area for presenting visual content; a housing holding the electronic display and having an opening allowing a person to view a first portion of the active area; and a bezel about the opening, the bezel covering a second portion of the active area and providing a window through which at least a part of the second portion can be viewed. A sensor system senses when a person is close to touching the bezel or when a person …


Predicting Domain Adaptivity: Redo Or Recycle?, Ting Yao, Chong-Wah Ngo, Shiai Zhu Nov 2012

Predicting Domain Adaptivity: Redo Or Recycle?, Ting Yao, Chong-Wah Ngo, Shiai Zhu

Research Collection School Of Computing and Information Systems

Over the years, the academic researchers have contributed various visual concept classifiers. Nevertheless, given a new dataset, most researchers still prefer to develop large number of classifiers from scratch despite expensive labeling efforts and limited computing resources. A valid question is why not multimedia community “embrace the green” and recycle off-the-shelf classifiers for new dataset. The difficulty originates from the domain gap that there are many different factors that govern the development of a classifier and eventually drive its performance to emphasize certain aspects of dataset. Reapplying a classifier to an unseen dataset may end up GIGO (garbage in, garbage …


Fashionask: Pushing Community Answers To Your Fingertips, Wei Zhang, Lei Pang, Chong-Wah Ngo Nov 2012

Fashionask: Pushing Community Answers To Your Fingertips, Wei Zhang, Lei Pang, Chong-Wah Ngo

Research Collection School Of Computing and Information Systems

We demonstrate a multimedia-based question-answering system, named FashionAsk, by allowing users to ask questions referring to pictures snapped by mobile devices. Specifically, instead of asking verbose questions to depict visual instances, direct pictures are provided as part of questions. To answer these multi-modal questions, FashionAsk performs a large-scale instance search to infer the names of instances, and then matches with similar questions from communitycontributed QA websites as answers. The demonstration is conducted on a million-scale dataset of Web images and QA pairs in the domain of fashion products. Asking a multimedia question through FashionAsk can take as short as five …


Community As A Connector: Associating Faces With Celebrity Names In Web Videos, Zhineng Chen, Chong-Wah Ngo, Juan Cao, Wei Zhang Nov 2012

Community As A Connector: Associating Faces With Celebrity Names In Web Videos, Zhineng Chen, Chong-Wah Ngo, Juan Cao, Wei Zhang

Research Collection School Of Computing and Information Systems

Associating celebrity faces appearing in videos with their names is of increasingly importance with the popularity of both celebrity videos and related queries. However, the problem is not yet seriously studied in Web video domain. This paper proposes a Community connected Celebrity Name-Face Association approach (CCNFA), where the community is regarded as an intermediate connector to facilitate the association. Specifically, with the names and faces extracted from Web videos, C-CNFA decomposes the association task into a three-step framework: community discovering, community matching and celebrity face tagging. To achieve the goal of efficient name-face association under this umbrella, algorithms such as …


Beaglebone Webcam Server, Alexander Corcoran Jun 2012

Beaglebone Webcam Server, Alexander Corcoran

Computer Engineering

The Beaglebone Webcam Server is a Linux based IP webcam, based on an inexpensive ARM development board, which hosts its own web server to display the webcam feed. The server has the ability to either connect to a wired router, or to act as a wireless access point in order for users to connect and control its functions via any Wi-Fi enabled device.


Cogtool-Helper: Leveraging Gui Functional Testing Tools To Generate Predictive Human Performance Models, Amanda Swearngin May 2012

Cogtool-Helper: Leveraging Gui Functional Testing Tools To Generate Predictive Human Performance Models, Amanda Swearngin

School of Computing: Dissertations, Theses, and Student Research

Numerous tools and techniques for human performance modeling have been introduced in the field of human-computer interaction. With such tools comes the ability to model legacy applications. Models can be used to compare design ideas to existing applications, or to evaluate products against those of competitors. One such mod- eling tool, CogTool, allows user interface designers and analysts to mock up design ideas, demonstrate tasks, and obtain human performance predictions for those tasks. This is one step towards a simple and complete analysis process, but it still requires a large amount of manual work. Graphical user interface (GUI) testing tools …


Converting Three-Component To Four-Component Image (2012), Ronald S. Cok, Michael E. Miller May 2012

Converting Three-Component To Four-Component Image (2012), Ronald S. Cok, Michael E. Miller

AFIT Patents

A method of converting a three-or-more-color-component image input signal to an image output signal includes acquiring an input signal having a plurality of pixel signals, each pixel signal having three, or more, color components; determining a residual difference for each color component of each pixel signal; determining a limit value of the residual differences; calculating a common scale factor for each of the color components based upon the limit value; and applying the common scale factor to the image input signal to produce the image output signal.


Extending The Hybridthread Smp Model For Distributed Memory Systems, Eugene Anthony Cartwright Iii May 2012

Extending The Hybridthread Smp Model For Distributed Memory Systems, Eugene Anthony Cartwright Iii

Graduate Theses and Dissertations

Memory Hierarchy is of growing importance in system design today. As Moore's Law allows system designers to include more processors within their designs, data locality becomes a priority. Traditional multiprocessor systems on chip (MPSoC) experience difficulty scaling as the quantity of processors increases. This challenge is common behavior of memory accesses in a shared memory environment and causes a decrease in memory bandwidth as processor numbers increase. In order to provide the necessary levels of scalability, the computer architecture community has sought to decentralize memory accesses by distributing memory throughout the system. Distributed memory offers greater bandwidth due to decoupled …


Electro-Luminescent Display Device, Michael E. Miller, Joseph R. Bietry, Ronald S. Cok Mar 2012

Electro-Luminescent Display Device, Michael E. Miller, Joseph R. Bietry, Ronald S. Cok

AFIT Patents

An electro-luminescent display includes a first array of light-emitting elements. Each of these light-emitting elements has an optical element. A second array of light-emitting elements also includes a second optical element different from the first. One or more row lines are electrically connected to either light-emitting elements in the first array of light-emitting elements or light-emitting elements in the second array of light-emitting elements. One or more column lines provide a data signal to the first and second array of light-emitting elements. A driver circuit delivers common information to the light-emitting elements in both the first and second arrays in …


Stereoscopic Display System With Flexible Rendering Of Disparity Map According To The Stereoscopic Fusing Capability Of The Observer, Elaine W. Jin, Michael E. Miller, Serguei Endrikhovski, Cathleen D. Cerosaletti Jan 2012

Stereoscopic Display System With Flexible Rendering Of Disparity Map According To The Stereoscopic Fusing Capability Of The Observer, Elaine W. Jin, Michael E. Miller, Serguei Endrikhovski, Cathleen D. Cerosaletti

AFIT Patents

A method is provided for customizing scene content, according to a user or a cluster of users, for a given stereoscopic display, including obtaining customization information about the user; obtaining a scene disparity map for a pair of given stereo images and/or a three-dimensional (3D) computer graphic model; and determining an aim disparity range for the user. The method of the present invention also generates a customized disparity map and/or rendering conditions for a three-dimensional (3D) computer graphic model correlating with the user's fusing capability of the given stereoscopic display; and renders or re-renders the stereo images for subsequent display.