Open Access. Powered by Scholars. Published by Universities.®

Computer Sciences Commons

Open Access. Powered by Scholars. Published by Universities.®

Discipline
Institution
Keyword
Publication Year
Publication
Publication Type
File Type

Articles 6391 - 6420 of 63017

Full-Text Articles in Computer Sciences

Direct Range Proofs For Paillier Cryptosystem And Their Applications, Zhikang Xie, Mengling Liu, Haiyang Xue, Man Ho Au, Robert H. Deng, Siu-Ming Yiu Oct 2024

Direct Range Proofs For Paillier Cryptosystem And Their Applications, Zhikang Xie, Mengling Liu, Haiyang Xue, Man Ho Au, Robert H. Deng, Siu-Ming Yiu

Research Collection School Of Computing and Information Systems

The Paillier cryptosystem is renowned for its applications in electronic voting, threshold ECDSA, multi-party computation, and more, largely due to its additive homomorphism. In these applications, range proofs for the Paillier cryptosystem are crucial for maintaining security, because of the mismatch between the message space in the Paillier system and the operation space in application scenarios. In this paper, we present novel range proofs for the Paillier cryptosystem, specifically aimed at optimizing those for both Paillier plaintext and affine operation. We interpret encryptions and affine operations as commitments over integers, as opposed to solely over ZN. Consequently, we propose direct …


A Survey Of Protocol Fuzzing, Xiaohan Zhang, Cen Zhang, Xinghua Li, Zhengjie Du, Bing Mao, Yeting Li, Pan Li Oct 2024

A Survey Of Protocol Fuzzing, Xiaohan Zhang, Cen Zhang, Xinghua Li, Zhengjie Du, Bing Mao, Yeting Li, Pan Li

Research Collection School Of Computing and Information Systems

Communication protocols form the bedrock of our interconnected world, yet vulnerabilities within their implementations pose significant security threats. Recent developments have seen a surge in fuzzing-based research dedicated to uncovering these vulnerabilities within protocol implementations. However, there still lacks a systematic overview of protocol fuzzing for answering the essential questions such as what the unique challenges are, how existing works solve them, and so on. To bridge this gap, we conducted a comprehensive investigation of related works from both academia and industry. Our study includes a detailed summary of the specific challenges in protocol fuzzing and provides a systematic categorization …


Self-Supervised Learning For Time Series Analysis : Taxonomy, Progress, And Prospects, Zhang Kexin, Qingsong Wen, Chaoli Zhang, Rongyao Cai, Ming Jin, Yong Liu, James Y. Zhang, Guansong Pang, Guansong Pang, Pan Shirui Oct 2024

Self-Supervised Learning For Time Series Analysis : Taxonomy, Progress, And Prospects, Zhang Kexin, Qingsong Wen, Chaoli Zhang, Rongyao Cai, Ming Jin, Yong Liu, James Y. Zhang, Guansong Pang, Guansong Pang, Pan Shirui

Research Collection School Of Computing and Information Systems

Self-supervised learning (SSL) has recently achieved impressive performance on various time series tasks. The most prominent advantage of SSL is that it reduces the dependence on labeled data. Based on the pre-training and fine-tuning strategy, even a small amount of labeled data can achieve high performance. Compared with many published self-supervised surveys on computer vision and natural language processing, a comprehensive survey for time series SSL is still missing. To fill this gap, we review current state-of-the-art SSL methods for time series data in this article. To this end, we first comprehensively review existing surveys related to SSL and time …


Demystifying And Extracting Fault-Indicating Information From Logs For Failure Diagnosis, Junjie Huang, Zhihan Jiang, Jinyang Liu, Yintong Huo, Jiazhen Gu, Zhuangbin Chen, Cong Feng, Hui Dong, Zengyin Yang, Michael R. Lyu Oct 2024

Demystifying And Extracting Fault-Indicating Information From Logs For Failure Diagnosis, Junjie Huang, Zhihan Jiang, Jinyang Liu, Yintong Huo, Jiazhen Gu, Zhuangbin Chen, Cong Feng, Hui Dong, Zengyin Yang, Michael R. Lyu

Research Collection School Of Computing and Information Systems

Logs are imperative in the maintenance of online service systems, which often encompass important information for effective failure mitigation. While existing anomaly detection methodologies facilitate the identification of anomalous logs within extensive runtime data, manual investigation of log messages by engineers remains essential to comprehend faults, which is labor-intensive and error-prone. Upon examining the log-based troubleshooting practices at CloudA 1, we find that engineers typically prioritize two categories of log information for diagnosis. These include fault-indicating descriptions, which record abnormal system events, and fault-indicating parameters, which specify the associated entities. Motivated by this finding, we propose an approach to automatically …


Multi-Scale Ai-Assisted Gene Expression Decoding, Fengyao Yan Oct 2024

Multi-Scale Ai-Assisted Gene Expression Decoding, Fengyao Yan

Theses and Dissertations

Genes can be treated as a graph that can be mapped. Tremendous information is coded in genes to ensure a complex functioning organism. Decoding this information is critical to understanding our biology and developing treatments for various diseases including cancer. Deep learning, a new branch of computer science, has gained traction over the past decade. It offers more insight into the data that is processed by the deeplearning models. Our study has shown that deep-learning models can be an effective tool in decoding genetic data such as gene tissue-deconvolution, gene graph mapping and genomic imputation. In tasks such as tissue …


Using Machine Learning And Deep Learning Algorithms To Improve Low Birthweight Prediction, Yang Ren Oct 2024

Using Machine Learning And Deep Learning Algorithms To Improve Low Birthweight Prediction, Yang Ren

Theses and Dissertations

Low birthweight (LBW) is a major public health issue resulting in increased neonatal mortality and long-term health complications. Traditional LBW analysis methods, focusing on incidence rates and risk factors through statistical models, often struggle with complex unseen data, and thus, their effectiveness is limited in early prevention of LBW, requiring more advanced LBW prediction models. Therefore, this dissertation delves into this important research area by proposing and examining novel machine learning (ML) and deep learning (DL) algorithms, aiming to predict LBW more accurately during the early stage of pregnancy. This dissertation consists of three studies, strategically designed to build upon …


Enhancing Microgrid Forecasting Accuracy With Saq-Mtclstm: A Self-Adjusting Quantized Multi-Task Convlstm For Optimized Solar Power And Load Demand Predictions, Ehtisham Lodhi, Nadia Dahmani, Syed Muhammad Salman Bukhari, Sujan Gyawali, Sanjog Thapa, Lin Qiu, Muhammad Hamza Zafar, Naureen Akhtar Oct 2024

Enhancing Microgrid Forecasting Accuracy With Saq-Mtclstm: A Self-Adjusting Quantized Multi-Task Convlstm For Optimized Solar Power And Load Demand Predictions, Ehtisham Lodhi, Nadia Dahmani, Syed Muhammad Salman Bukhari, Sujan Gyawali, Sanjog Thapa, Lin Qiu, Muhammad Hamza Zafar, Naureen Akhtar

All Works

Accurate forecasting of solar power output and load demand is critical for the efficient operation and management of isolated microgrids, where reliability and sustainability are paramount. Traditional methods often struggle with data scarcity, limitations in capturing intricate temporal dynamics, and lack of scalability. This research introduces a novel multi-task learning (MTL) model, the Self-Aware Quantized Multi-Task ConvLSTM (SAQ-MTCLSTM), which addresses these challenges by jointly forecasting solar power and load demand while leveraging shared representations across these interdependent time series. The SAQ-MTCLSTM incorporates a sophisticated architecture that combines convolutional and LSTM layers with self-aware quantization to enhance computational efficiency and model …


Unveiling Similarities In The Code Of Life: A Detailed Exploration Of Dna Sequence Matching Algorithm, Mahmoud Y. Shams, Romany M. Farag, Dalia A. Aldawody, Huda E. Khalid, Ahmed K. Essa, Hazem M. El-Bakry, A. A. Salama Oct 2024

Unveiling Similarities In The Code Of Life: A Detailed Exploration Of Dna Sequence Matching Algorithm, Mahmoud Y. Shams, Romany M. Farag, Dalia A. Aldawody, Huda E. Khalid, Ahmed K. Essa, Hazem M. El-Bakry, A. A. Salama

Neutrosophic Systems with Applications

Identifying similar DNA sequences is crucial in various biological research endeavors. This paper delves into the intricate workings of a specific algorithm designed for this purpose. We provide a systematic explanation, exploring how the algorithm handles user input, reads stored DNA sequences, utilizes the Word2Vec model for vector representation, and calculates sequence similarity using diverse metrics like Cosine Similarity and Neutrosophic Distance. Additionally, the paper explores the incorporation of neutrosophic values to account for uncertainty in the comparisons. Finally, we discuss the extraction of results, including matched sequences, similarity scores, and accuracy measures. This in-depth exploration provides a clear understanding …


An Approach To Multi-Attribute Decision-Making Based On Single-Valued Neutrosophic Hesitant Fuzzy Aczel-Alsina Aggregation Operator, Raiha Imran, Kifayat Ullah, Zeeshan Ali, Maria Akram Oct 2024

An Approach To Multi-Attribute Decision-Making Based On Single-Valued Neutrosophic Hesitant Fuzzy Aczel-Alsina Aggregation Operator, Raiha Imran, Kifayat Ullah, Zeeshan Ali, Maria Akram

Neutrosophic Systems with Applications

A single-valued Neutrosophic hesitant fuzzy set (SVNHFS) is a combination of a single-valued neutrosophic set (SVNS) and hesitant fuzzy set (HFS) that has been developed to address insufficient, unreliable, and vague environments in which each element has several possible options determined by the truthiness, indeterminacy and falsity value. By considering this, in this paper, we have proposed the Aczel-Alsina aggregation operator (AAAO) for SVNHFS, which is more flexible t-norm and t-conorm than the other and due to the flexible nature of parameters to solve Multi-Attribute decision making (MADM) problems. Further, the score function, accuracy function, and certainty function of SVNHFS …


The Promise And Perils Of Smart (City) Bots As Educational Tools, Thomas Menkhoff, Siew Ning Kan, Shaohui Foong Oct 2024

The Promise And Perils Of Smart (City) Bots As Educational Tools, Thomas Menkhoff, Siew Ning Kan, Shaohui Foong

Research Collection Lee Kong Chian School Of Business

Written from an educator’s perspective, the paper aims to examine the increasing use of chatbot technology capable of answering user queries in various fields such as commerce, public service delivery, urban areas, higher education etc. by simulating human conversation through text messages or voice commands. We analyse why smart bots have become popular technologies in the context smart cities; shed light on different types of bots (chatbots, large language models such as ChatGPT, flying bots); and explain how we use bots as teaching tool in an introductory AI for business course. A key argument putting forward is that smart bots …


A Guideline For Open-Source Tools To Make Medical Imaging Data Ready For Artificial Intelligence Applications: A Society Of Imaging Informatics In Medicine (Siim) Survey, Sanaz Vahdati, Bardia Khosravi, Elham Mahmoudi, Kuan Zhang, Pouria Rouzrokh, Shahriar Faghani, Mana Moassefi, Aylin Tahmasebi, Katherine Andriole, Peter Chang, Keyvan Farahani, Mona Flores, Les Folio, Sina Houshmand, Maryellen Giger, Judy Gichoya, Bradley Erickson Oct 2024

A Guideline For Open-Source Tools To Make Medical Imaging Data Ready For Artificial Intelligence Applications: A Society Of Imaging Informatics In Medicine (Siim) Survey, Sanaz Vahdati, Bardia Khosravi, Elham Mahmoudi, Kuan Zhang, Pouria Rouzrokh, Shahriar Faghani, Mana Moassefi, Aylin Tahmasebi, Katherine Andriole, Peter Chang, Keyvan Farahani, Mona Flores, Les Folio, Sina Houshmand, Maryellen Giger, Judy Gichoya, Bradley Erickson

Department of Radiology Faculty Papers

In recent years, the role of Artificial Intelligence (AI) in medical imaging has become increasingly prominent, with the majority of AI applications approved by the FDA being in imaging and radiology in 2023. The surge in AI model development to tackle clinical challenges underscores the necessity for preparing high-quality medical imaging data. Proper data preparation is crucial as it fosters the creation of standardized and reproducible AI models while minimizing biases. Data curation transforms raw data into a valuable, organized, and dependable resource and is a fundamental process to the success of machine learning and analytical projects. Considering the plethora …


Antitrust After The Coming Wave, Daniel A. Crane Oct 2024

Antitrust After The Coming Wave, Daniel A. Crane

Articles

A coming wave of general-purpose technologies, including artificial intelligence ("AI"), robotics, quantum computing, synthetic biology, energy expansion, and nanotechnology, is likely to fundamentally reshape the economy and erode the assumptions on which the antitrust order is predicated. First, AI-driven systems will vastly improve firms' ability to detect (and even program) consumer preferences without the benefit of price signals, which will undermine the traditional information-producing benefit of competitive markets. Similarly, these systems will be able to determine comparative producer efficiency without relying on competitive signals. Second, AI systems will invert the salient characteristics of human managers, whose intentions are opaque but …


An End-To-End Bi-Objective Approach To Deep Graph Partitioning, Pengcheng Wei, Yuan Fang, Zhihao Wen, Zheng Xiao, Binbin Chen Oct 2024

An End-To-End Bi-Objective Approach To Deep Graph Partitioning, Pengcheng Wei, Yuan Fang, Zhihao Wen, Zheng Xiao, Binbin Chen

Research Collection School Of Computing and Information Systems

Graphs are ubiquitous in real-world applications, such as computation graphs and social networks. Partitioning large graphs into smaller, balanced partitions is often essential, with the biobjective graph partitioning problem aiming to minimize both the“cut” across partitions and the imbalance in partition sizes. However, existing heuristic methods face scalability challenges or overlook partition balance, leading to suboptimal results. Recent deep learning approaches, while promising, typically focus only on node-level features and lack a truly end-to-end framework, resulting in limited performance. In this paper, we introduce a novel method based on graph neural networks (GNNs) that leverages multilevel graph features and addresses …


Kpiroot: Efficient Monitoring Metric-Based Root Cause Localization In Large-Scale Cloud Systems, Wenwei Gu, Xinying Sun, Jinyang Liu, Yintong Huo, Zhuangbin Chen, Jianping Zhang, Jiazhen Gu, Yongqiang Yang, Michael R. Lyu Oct 2024

Kpiroot: Efficient Monitoring Metric-Based Root Cause Localization In Large-Scale Cloud Systems, Wenwei Gu, Xinying Sun, Jinyang Liu, Yintong Huo, Zhuangbin Chen, Jianping Zhang, Jiazhen Gu, Yongqiang Yang, Michael R. Lyu

Research Collection School Of Computing and Information Systems

To ensure the reliability of cloud systems, their run-time status reflecting the service quality is periodically monitored with monitoring metrics, i.e., KPIs (key performance indicators). When performance issues happen, root cause localization pinpoints the specific KPIs that are responsible for the degradation of overall service quality, facilitating prompt problem diagnosis and resolution. To this end, existing methods generally locate root-cause KPIs by identifying the KPIs that exhibit a similar anomalous trend to the overall service performance. While straightforward, solely relying on the similarity calculation may be ineffective when dealing with cloud systems with complicated interdependent services. Recent deep learning-based methods …


Transformer-Based Joint Learning Approach For Text Normalization In Vietnamese Automatic Speech Recognition Systems, The Viet Bui, Tho Chi Luong, Oanh Thi Tran Oct 2024

Transformer-Based Joint Learning Approach For Text Normalization In Vietnamese Automatic Speech Recognition Systems, The Viet Bui, Tho Chi Luong, Oanh Thi Tran

Research Collection School Of Computing and Information Systems

In this article, we investigate the task of normalizing transcribed texts in Vietnamese Automatic Speech Recognition (ASR) systems in order to improve user readability and the performance of downstream tasks. This task usually consists of two main sub-tasks: predicting and inserting punctuation (i.e., period, comma); and detecting and standardizing named entities (i.e., numbers, person names) from spoken forms to their appropriate written forms. To achieve these goals, we introduce a complete corpus including of 87,700 sentences and investigate conditional joint learning approaches which globally optimize two sub-tasks simultaneously. The experimental results are quite promising. Overall, the proposed architecture outperformed the …


Data Provenance Via Differential Auditing, Xin Mu, Ming Pang, Feida Zhu Oct 2024

Data Provenance Via Differential Auditing, Xin Mu, Ming Pang, Feida Zhu

Research Collection School Of Computing and Information Systems

With the rising awareness of data assets, data governance, which is to understand where data comes from, how it is collected, and how it is used, has been assuming evergrowing importance. One critical component of data governance gaining increasing attention is auditing machine learning models to determine if specific data has been used for training. Existing auditing techniques, like shadow auditing methods, have shown feasibility under specific conditions such as having access to label information and knowledge of training protocols. However, these conditions are often not met in most real-world applications. In this paper, we introduce a practical framework for …


Hisoma: A Hierarchical Multi-Agent Model Integrating Self-Organizing Neural Networks With Multi-Agent Deep Reinforcement Learning, Minghong Geng, Shubham Pateria, Budhitama Subagdja, Ah-Hwee Tan Oct 2024

Hisoma: A Hierarchical Multi-Agent Model Integrating Self-Organizing Neural Networks With Multi-Agent Deep Reinforcement Learning, Minghong Geng, Shubham Pateria, Budhitama Subagdja, Ah-Hwee Tan

Research Collection School Of Computing and Information Systems

Multi-agent deep reinforcement learning (MADRL) has shown remarkable advancements in the past decade. However, most current MADRL models focus on task-specific short-horizon problems involving a small number of agents, limiting their applicability to long-horizon planning in complex environments. Hierarchical multi-agent models offer a promising solution by organizing agents into different levels, effectively addressing tasks with varying planning horizons. However, these models often face constraints related to the number of agents or levels of hierarchies. This paper introduces HiSOMA, a novel hierarchical multi-agent model designed to handle long-horizon, multi-agent, multi-task decision-making problems. The top-level controller, FALCON, is modeled as a class …


Does Ceo Agreeableness Personality Mitigate Real Earnings Management?, Shan Liu, Xingying Wu, Nan Hu Oct 2024

Does Ceo Agreeableness Personality Mitigate Real Earnings Management?, Shan Liu, Xingying Wu, Nan Hu

Research Collection School Of Computing and Information Systems

Despite efforts to mitigate aggressive financial reporting, earnings management remains challenging to parties interested in inhibiting its dysfunctional effects. Using linguistic algorithms to assess CEO agreeableness personality from their unscripted texts in conference calls, we find that it is a determinant that mitigates a firm's real earnings management. Furthermore, such an effect is more pronounced when firms confront intensive market competition and financial distress and have weaker managerial entrenchment or when CEOs face stronger internal governance. Our findings persist even after we utilize several alternative real earnings management metrics and control other confounding personalities in prior earnings management studies. The …


D2sr: Decentralized Detection, De-Synchronization, And Recovery Of Lidar Interference, Darshana Rathnayake, Hemanth Sabbella, Meera Radhakrishnan, Archan Misra Oct 2024

D2sr: Decentralized Detection, De-Synchronization, And Recovery Of Lidar Interference, Darshana Rathnayake, Hemanth Sabbella, Meera Radhakrishnan, Archan Misra

Research Collection School Of Computing and Information Systems

We address the challenge of multi-LiDAR interference, an issue of growing importance as LiDAR sensors are embedded in a growing set of pervasive devices. We introduce a novel approach named D2SR, enabling decentralized interference detection, mitigation, and recovery without explicit coordination among nearby LiDAR devices. D2SR comprises three stages: (a) Detection, which identifies interfered frames, (b) Mitigation, which performs time-shifting of a LiDAR’s active period to reduce interference, and (c) Recovery, which corrects or reconstructs the depth values in interfered regions of a depth frame. Key contributions include a lightweight interference detection algorithm achieving an F1-score of 92%, a simple …


Motif Graph Neural Network, Xuexin Chen, Ruicui Cai, Yuan Fang, Min Wu, Zijian Li, Zhifeng Hao Oct 2024

Motif Graph Neural Network, Xuexin Chen, Ruicui Cai, Yuan Fang, Min Wu, Zijian Li, Zhifeng Hao

Research Collection School Of Computing and Information Systems

Graphs can model complicated interactions between entities, which naturally emerge in many important applications. These applications can often be cast into standard graph learning tasks, in which a crucial step is to learn low-dimensional graph representations. Graph neural networks (GNNs) are currently the most popular model in graph embedding approaches. However, standard GNNs in the neighborhood aggregation paradigm suffer from limited discriminative power in distinguishing high-order graph structures as opposed to low-order structures. To capture high-order structures, researchers have resorted to motifs and developed motif-based GNNs. However, the existing motif-based GNNs still often suffer from less discriminative power on high-order …


Predicting The Limits: Tailoring Unnoticeable Hand Redirection Offsets In Virtual Reality To Individuals' Perceptual Boundaries, Martin Feick, Kora Persephone Regitz, Lukas Gehrke, André Zenner, Anthony Tang, Tobias Patrick Jungbluth, Maurice Rekrut, Antonio Krüger Oct 2024

Predicting The Limits: Tailoring Unnoticeable Hand Redirection Offsets In Virtual Reality To Individuals' Perceptual Boundaries, Martin Feick, Kora Persephone Regitz, Lukas Gehrke, André Zenner, Anthony Tang, Tobias Patrick Jungbluth, Maurice Rekrut, Antonio Krüger

Research Collection School Of Computing and Information Systems

Many illusion and interaction techniques in Virtual Reality (VR) rely on Hand Redirection (HR), which has proved to be effective as long as the introduced offsets between the position of the real and virtual hand do not noticeably disturb the user experience. Yet calibrating HR offsets is a tedious and time-consuming process involving psychophysical experimentation, and the resulting thresholds are known to be affected by many variables—limiting HR’s practical utility. As a result, there is a clear need for alternative methods that allow tailoring HR to the perceptual boundaries of individual users. We conducted an experiment with 18 participants combining …


Joint Weakly Supervised Image Emotion Analysis Based On Interclass Discrimination And Intraclass Correlation, Xinyue Zhang, Zhaoxia Wang, Guitao Cao, Seng-Beng Ho Oct 2024

Joint Weakly Supervised Image Emotion Analysis Based On Interclass Discrimination And Intraclass Correlation, Xinyue Zhang, Zhaoxia Wang, Guitao Cao, Seng-Beng Ho

Research Collection School Of Computing and Information Systems

Regional information-based image emotion analysis has recently garnered significant attention. However, existing methods often focus on identifying region proposals through layered steps or merely rely on visual saliency. These approaches may lead to an underestimation of emotional categories and a lack of comprehensive interclass discrimination perception and emotional intraclass contextual mining. To address these limitations, we propose a novel approach named InterIntraIEA, which combines interclass discrimination and intraclass correlation joint learning capabilities for image emotion analysis. The proposed method not only employs category-specific dictionary learning for class adaptation, but also models intraclass contextual relationships and perceives correlations at the channel …


Exploring Conversations Between A Practitioner And A Person With Dementia, Kotaro Hara, Rosiana Natalie, Wei Soon Cheong, Jingjing Gu, Qianli Xu Oct 2024

Exploring Conversations Between A Practitioner And A Person With Dementia, Kotaro Hara, Rosiana Natalie, Wei Soon Cheong, Jingjing Gu, Qianli Xu

Research Collection School Of Computing and Information Systems

In social service centers, practitioners engage in conversations with clients with dementia to facilitate their daily activities and provide support when they are distressed. However, the nature of the care demands the practitioner’s active engagement, which becomes difficult to deliver as the number of people who need care expands. Researchers have been investigating the efficacy of developing agents that assume conversational tasks to alleviate this work. To contribute to the future design of agents for caregiving, we collected and analyzed ten conversations between clients with mild dementia and practitioners who provide care. Our analyses of turn-taking dynamics and dialogue acts …


Temporal Relational Graph Convolutional Network Approach To Financial Performance Prediction, Jeyaraman Brindha Priyadarshini, Bing Tian Dai, Yuan Fang Oct 2024

Temporal Relational Graph Convolutional Network Approach To Financial Performance Prediction, Jeyaraman Brindha Priyadarshini, Bing Tian Dai, Yuan Fang

Research Collection School Of Computing and Information Systems

Accurately predicting financial entity performance remains a challenge due to the dynamic nature of financial markets and vast unstructured textual data. Financial knowledge graphs (FKGs) offer a structured representation for tackling this problem by representing complex financial relationships and concepts. However, constructing a comprehensive and accurate financial knowledge graph that captures the temporal dynamics of financial entities is non-trivial. We introduce FintechKG, a comprehensive financial knowledge graph developed through a three-dimensional information extraction process that incorporates commercial entities and temporal dimensions and uses a financial concept taxonomy that ensures financial domain entity and relationship extraction. We propose a temporal and …


Efficient Cascaded Multiscale Adaptive Network For Image Restoration, Yichen Zhou, Pan Zhou, Teck Khim Ng Oct 2024

Efficient Cascaded Multiscale Adaptive Network For Image Restoration, Yichen Zhou, Pan Zhou, Teck Khim Ng

Research Collection School Of Computing and Information Systems

Image restoration, encompassing tasks such as deblurring, denoising, and super-resolution, remains a pivotal area in computer vision. However, efficiently addressing the spatially varying artifacts of various low-quality images with local adaptiveness and handling their degradations at different scales poses significant challenges. To efficiently tackle these issues, we propose the novel Efficient Cascaded Multiscale Adaptive (ECMA) Network. ECMA employs Local Adaptive Module, LAM, which dynamically adjusts convolution kernels across local image regions to efficiently handle varying artifacts. Thus, LAM addresses the local adaptiveness challenge more efficiently than costlier mechanisms like self-attention, due to its less computationally intensive convolutions. To construct a …


Self-Adaptive Fine-Grained Multi-Modal Data Augmentation For Semi-Supervised Muti-Modal Coreference Resolution, Li Zheng, Boyu Chen, Hao Fei, Fei Li, Shengqiong Wu, Lizi Liao, Donghong Ji Oct 2024

Self-Adaptive Fine-Grained Multi-Modal Data Augmentation For Semi-Supervised Muti-Modal Coreference Resolution, Li Zheng, Boyu Chen, Hao Fei, Fei Li, Shengqiong Wu, Lizi Liao, Donghong Ji

Research Collection School Of Computing and Information Systems

Coreference resolution, an essential task in natural language processing, is particularly challenging in multi-modal scenarios where data comes in various forms and modalities. Despite advancements, limitations due to scarce labeled data and underleveraged unlabeled data persist. We address these issues with a self-adaptive fine-grained multi-modal data augmentation framework for semi-supervised MCR, focusing on enriching training data from labeled datasets and tapping into the untapped potential of unlabeled data. Regarding the former issue, we first leverage text coreference resolution datasets and diffusion models,to perform fine-grained text-to-image generation with aligned text entities and image bounding boxes. We then introduce a self-adaptive selection …


Enhancing Recipe Retrieval With Foundation Models: A Data Augmentation Perspective, Fangzhou Song, Bin Zhu, Yanbin Hao, Shuo Wang Oct 2024

Enhancing Recipe Retrieval With Foundation Models: A Data Augmentation Perspective, Fangzhou Song, Bin Zhu, Yanbin Hao, Shuo Wang

Research Collection School Of Computing and Information Systems

Learning recipe and food image representation in common embedding space is non-trivial but crucial for cross-modal recipe retrieval. In this paper, we propose a new perspective for this problem by utilizing foundation models for data augmentation. Leveraging on the remarkable capabilities of foundation models (i.e., Llama2 and SAM), we propose to augment recipe and food image by extracting alignable information related to the counterpart. Specifically, Llama2 is employed to generate a textual description from the recipe, aiming to capture the visual cues of a food image, and SAM is used to produce image segments that correspond to key ingredients in …


Risurconv : Rotation Invariant Surface Attention-Augmented Convolutions For 3d Point Cloud Classification And Segmentation, Zhiyuan Zhang, Licheng Yang, Xiang Zhiyu Oct 2024

Risurconv : Rotation Invariant Surface Attention-Augmented Convolutions For 3d Point Cloud Classification And Segmentation, Zhiyuan Zhang, Licheng Yang, Xiang Zhiyu

Research Collection School Of Computing and Information Systems

Despite the progress on 3D point cloud deep learning, most prior works focus on learning features that are invariant to translation and point permutation, and very limited efforts have been devoted for rotation invariant property. Several recent studies achieve rotation invariance at the cost of lower accuracies. In this work, we close this gap by proposing a novel yet effective rotation invariant architecture for 3D point cloud classification and segmentation. Instead of traditional pointwise operations, we construct local triangle surfaces to capture more detailed surface structure, based on which we can extract highly expressive rotation invariant surface properties which are …


Desk2desk : Optimization-Based Mixed Reality Workspace Integration For Remote Side-By-Side Collaboration, Ludwig Sidenmark, Tianyu Zhang, Leen Al Lababidi, Jiannan Li, Tovi Grossman Oct 2024

Desk2desk : Optimization-Based Mixed Reality Workspace Integration For Remote Side-By-Side Collaboration, Ludwig Sidenmark, Tianyu Zhang, Leen Al Lababidi, Jiannan Li, Tovi Grossman

Research Collection School Of Computing and Information Systems

Mixed Reality enables hybrid workspaces where physical and virtual monitors are adaptively created and moved to suit the current environment and needs. However, in shared settings, individual users’ workspaces are rarely aligned and can vary significantly in the number of monitors, available physical space, and workspace layout, creating inconsistencies between workspaces which may cause confusion and reduce collaboration. We present Desk2Desk, an optimization-based approach for remote collaboration in which the hybrid workspaces of two collaborators are fully integrated to enable immersive side-by-side collaboration. The optimization adjusts each user’s workspace in layout and number of shared monitors and creates a mapping …


Improving Out-Of-Distribution Detection With Disentangled Foreground And Background Features, Choubo Ding, Guansong Pang Oct 2024

Improving Out-Of-Distribution Detection With Disentangled Foreground And Background Features, Choubo Ding, Guansong Pang

Research Collection School Of Computing and Information Systems

Detecting out-of-distribution (OOD) inputs is a principal task for ensuring the safety of deploying deep-neural-network classifiers in open-set scenarios. OOD samples can be drawn from arbitrary distributions and exhibit deviations from in-distribution (ID) data in various dimensions, such as foreground features (e.g., objects in CIFAR100 images vs. those in CIFAR10 images) and background features (e.g., textural images vs. objects in CIFAR10). Existing methods can confound foreground and background features in training, failing to utilize the background features for OOD detection. This paper considers the importance of feature disentanglement in out-of-distribution detection and proposes the simultaneous exploitation of both foreground and …