Defense-To-Attack: Bypassing Weak Defenses Enables Stronger Jailbreaks In Vision-Language Models,
2026
Singapore Management University
Defense-To-Attack: Bypassing Weak Defenses Enables Stronger Jailbreaks In Vision-Language Models, Yunhan Zhao, Xiang Zheng, Yige Li, Xingjun Ma
Research Collection School Of Computing and Information Systems
Despite their superb capabilities, Vision-Language Models (VLMs) have been shown to be vulnerable to jailbreak attacks. While recent jailbreaks have achieved notable progress, their effectiveness and efficiency can still be improved. In this work, we reveal an interesting phenomenon: incorporating weak defense cues into the attack pipeline can significantly enhance both the effectiveness and efficiency of jailbreaks on VLMs. Building on this insight, we propose Defense2Attack, a novel jailbreak method that bypasses the safety guardrails of VLMs by leveraging defensive patterns to guide jailbreak prompt construction. Specifically, Defense2Attack consists of three key components: (1) a visual optimizer that embeds universal …
Closing The Interpretability Gap: Explainable Ml-Based Malware Detection For Defensive Cyberspace Operations,
2026
Washington State University
Closing The Interpretability Gap: Explainable Ml-Based Malware Detection For Defensive Cyberspace Operations, Tashi Stirewalt, Sean Hodgson, Puumaaya Tahiru, Assefaw Gebremedhin
Military Cyber Affairs
This paper presents an end-to-end, explainable malware triage pipeline designed for defense-oriented cyber operations. It combines high-performance static detection methods with analyst-centered interpretability. Utilizing the EMBER 2024 Windows PE subset, we train and evaluate four classifiers and select LightGBM as the production model based on its predictive performance, inference efficiency, and compatibility with exact tree-based attribution. The deployed system consists of four sequential components: PE feature extraction, malware probability scoring, dual explainability (using SHAP and LIME), and large language model (LLM) report generation, all integrated within a Flask web interface. On a temporal test set of 1,080,000 samples, LightGBM achieves …
From Data To Decision-Making: The Role Of Local Digital Twins In Cross-Domain Management Within Municipalities – A Research-In-Progress Study In Veenendaal,
2026
University of Applied Sciences Utrecht
From Data To Decision-Making: The Role Of Local Digital Twins In Cross-Domain Management Within Municipalities – A Research-In-Progress Study In Veenendaal, Diana M.E. Boekman, Koen Smit, Guido Ongena, Rob Peters
Communications of the IIMA
Municipalities are facing increasingly complex, interconnected challenges in areas like housing, climate adaptation, mobility, and social policy. Local Digital Twins (LDTs) are seen as a promising tool to make this complexity more understandable and support decision-making. At the same time, both literature and practice show that few initiatives get past the pilot phase, even though getting through that phase is essential for successful long-term adoption.
This paper presents a research-in-progress study on the development and application of an implementation method for LDT technology within the municipality of Veenendaal, based on human values rather than driven by technological possibilities. Based on …
Neural Symphony Of Flow Experience: Evidence For High-Dimensional Metastable Dynamics,
2026
Singapore Management University
Neural Symphony Of Flow Experience: Evidence For High-Dimensional Metastable Dynamics, Abdelrahman B. M. Eldaly, Kris Zhangguang Kang, Fiona Fui-Hoon Nah, Leanne Lai-Hang Chan, Keng Siau, Xiao Fan Liu, Richard Huskey, Langtao Chen, Tejaswini Yelamanchili, Rene Weber
Research Collection School Of Computing and Information Systems
Flow, an optimal experience characterized by deep immersion and engagement in an activity, has been extensively studied in behavioral research. However, its neural dynamic mechanism remains poorly understood. In a within-subject video gaming experiment, we captured neural activity underlying flow, boredom, and anxiety using a 64-channel electroencephalogram (EEG) system. Compared to boredom and anxiety, flow exhibits the highest global functional connectivity, metastability, and dimensionality of dynamic functional connectivity patterns, suggesting that flow is a highly adaptable process that is supported by high-dimensional neural dynamics. Unlike previous studies that focused on identifying static or localized brain activity, we examine the neural …
Restoring Linguistic Grounding In Vla Models Via Train-Free Attention Recalibration,
2026
Singapore Management University
Restoring Linguistic Grounding In Vla Models Via Train-Free Attention Recalibration, Ninghao Zhang, Bin Zhu, Shijie Zhou, Jingjing Chen
Research Collection School Of Computing and Information Systems
Vision-Language-Action (VLA) models enable robots to perform manipulation tasks directly from natural language instructions and are increasingly viewed as a foundation for generalist robotic policies. However, their reliability under Out-Of-Distribution (OOD) instructions remains underexplored. In this paper, we reveal a critical failure mode in which VLA policies continue executing visually plausible actions even when the language instruction contradicts the scene. We refer to this phenomenon as linguistic blindness, where VLA policies prioritize visual priors over instruction semantics during action generation. To systematically analyze this issue, we introduce ICBench, a diagnostic benchmark constructed from the LIBERO dataset that probes language–action coupling …
Learning 1-Bit Lidar-Based Localization With Auxiliary Objective,
2026
Singapore Management University
Learning 1-Bit Lidar-Based Localization With Auxiliary Objective, Kaijie Yin, Zhiyuan Zhang, Tian Gao, Wentao Zhu, Cheng-Zhong Xu, Hui Kong
Research Collection School Of Computing and Information Systems
6-DoF LiDAR-based localization is a fundamental capability for autonomous systems operating in large-scale outdoor environments. Many deep-learning-based localization methods have achieved promising performance so far. However, as one of the always-on modules competing for limited on-board computational resources, the localization module is expected to consume only a small portion of the overall compute budget. Most existing learning-based methods are still too heavy for this purpose. In contrast, binary neural networks (BNNs) offer an appealing solution, but the 1-bit compression causes severe information loss and performance drop. In this paper, we address this challenge by proposing Binarized LiDAR-based Localization (BiLoc), the …
Generalized Logit Adjustment: Improved Fine-Tuning By Mitigating Label Bias In Zero-Shot Vision Models,
2026
Singapore Management University
Generalized Logit Adjustment: Improved Fine-Tuning By Mitigating Label Bias In Zero-Shot Vision Models, Beier Zhu, Qianru Sun, Xun Yang, Hanwang Zhang
Research Collection School Of Computing and Information Systems
Foundation models like CLIP allow zero-shot transfer on various tasks without additional training data. Yet, the zero-shot performance is less competitive than a fully supervised one. Thus, fine-tuning and ensembling are also commonly adopted to better fit the downstream tasks. However, we argue that such prior work has overlooked the inherent biases in foundation models. Due to the highly imbalanced Web-scale training set, foundation models are inevitably skewed toward frequent semantics, and thus the subsequent fine-tuning or ensembling is still biased. In this study, we systematically examine the biases in foundation models and demonstrate the efficacy of our proposed Generalized …
Ai And The Environment: Solutions For Advancing Technology Safely,
2026
Embry-Riddle Aeronautical University
Ai And The Environment: Solutions For Advancing Technology Safely, Kaitlynn Baker
Discovery Day - Daytona Beach
Since 2022, the world of Artificial Intelligence (AI) has boomed. AI went from a special and rare entity to a commonly used resource available to all through web sites, and phone apps. AI has benefitted everyday activities by making office, class, and personal tasks easier through grammar help, informational citations, and as someone to bounce ideas off of. Additionally, many companies have begun utilizing AI to improve customer service and experience, and train workers more efficiently, therefore, saving thousands of dollars. Despite the benefits humans reap from its use, AI has been harming our environment at growing rates. Data centers …
How Ai Influences The Design Process Of Unmanned Underwater Vehicles’ (Uuvs) 3d Sonar System,
2026
Embry-Riddle Aeronautical University
How Ai Influences The Design Process Of Unmanned Underwater Vehicles’ (Uuvs) 3d Sonar System, Eden Tsouklaris, Abriella Smith, Brianna Broderick, Carissa Aumack, Victoria Cornaro
Discovery Day - Daytona Beach
With the exponential growth of Artificial Intelligence (AI), user interface (UI) designers have explored using AI to shorten design time. This study assessed the effectiveness of UIs designed with AI programs versus manual methods for an Unmanned Underwater Vehicle (UUV) control system. Participants were tasked with designing an interface that would allow submarine operators to monitor and coordinate three UUVs repairing a severed underwater communication cable at a depth of 2,000 meters. The scenario presented several operational challenges (zero visibility, sonar-only perception, data latency, and potential system degradation), requiring participants' designs to maintain spatial awareness and support remote repair tasks. …
Real Bullets, Plastic Guns: Evaluating The Strength Of 3-D Printed Gun Parts,
2026
CUNY John Jay College
Real Bullets, Plastic Guns: Evaluating The Strength Of 3-D Printed Gun Parts, Maria Latenia Mayol
Student Theses
Privately made firearms (PMFs), often referred to as “ghost guns,” are firearms manufactured or assembled by individuals rather than federally licensed manufacturers. Although the terms are frequently used interchangeably, “ghost gun” more specifically describes an unserialized firearm, whereas PMFs include a broader range of firearms produced through nontraditional manufacturing methods. PMFs may be entirely 3-D printed, assembled from partially completed firearm kits, or constructed by integrating additively manufactured components with commercially manufactured firearm parts. The increasing accessibility of additive manufacturing and widespread dissemination of computer-aided design files have raised concerns about concealment, regulation, and forensic evasion, particularly when factory-manufactured components …
Hvi-Cidnet+: Beyond Extreme Darkness For Low-Light Image Enhancement,
2026
Singapore Management University
Hvi-Cidnet+: Beyond Extreme Darkness For Low-Light Image Enhancement, Kangbiao Shi, Xiaowen Ma, Yixu Feng, Tao Hu, Peng Wu, Guansong Pang, Qingsen Yan
Research Collection School Of Computing and Information Systems
Low-Light Image Enhancement (LLIE) aims to recover visually pleasing content and details from degraded low-light images. However, existing RGB-based methods often suffer from color bias and brightness artifacts due to inherent high color sensitivity. Although the HSV color space can decouple brightness and color, it introduces noticeable red and black noise artifacts. To address these challenges, we adopt the Horizontal/Vertical-Intensity (HVI) color space for LLIE, which is defined by the HV color map and learnable intensity. The former enforces small distances for red coordinates to alleviate red noise artifacts, while the latter adaptively compresses low-light regions to suppress black noise …
Task-Aligned Haze Removal With Semantic-Aware Fusion And Contrast Self-Correction,
2026
Singapore Management University
Task-Aligned Haze Removal With Semantic-Aware Fusion And Contrast Self-Correction, Jinbin Wang, Aiping Yang, Guosong Jiang, Wenlong Yu, Dongwei Ren, Qinghua Hu
Research Collection School Of Computing and Information Systems
Adverse haze conditions introduce complex degradations that obscure scene details and distort structural cues critical for object detection, posing persistent challenges for vision‐based sensing systems. Although existing haze removal methods have achieved notable improvements in visual clarity, their optimisation objectives are often misaligned with downstream detection requirements, leading to limited detection performance in real‐world scenarios. To address this issue, this work proposes a task‐aligned weakly supervised haze removal framework, termed Dehaze4Detection, which explicitly aligns low‐level restoration with high‐level detection objectives. The framework incorporates a Semantic‐Aware Multi‐Scale Fusion Module (SMFM) that embeds pixel‐level semantic knowledge into the dehazing process, enabling selective …
Painting A Scene: 3d Painterly Rendering From Curves To Rebelle,
2026
Clemson University
Painting A Scene: 3d Painterly Rendering From Curves To Rebelle, William M. Luttrell
All Theses
Stylized 3D rendering has seen much development and success over the past few years. From Spider-Man: Across the Spider-Verse to The Bad Guys, many studios have developed tools to incorporate stylistic elements from graphic novels, comic books, watercolor paintings, and more into their productions. This stylization process incorporates the pacing, visual style, and themes from the source medium into the animated work, allowing a much greater freedom of expression for artists and directors.
Inspired by these films as well as the needs of the short film Kate Shelley and the Bridge of Darkness currently in production, This paper presents …
"The First Web Novel At 30: The Collection And The Creative Process",
2026
University of Bergen
"The First Web Novel At 30: The Collection And The Creative Process", Robert Arellano, Scott Rettberg
ELO (un)supervised 2026
Summer 2026 marks the 30th anniversary of Sunshine '69, recognized as the first novelistic hypertext fiction published on the web. While the full work remains accessible online—an "(un)supervised" preservation achievement in itself—the archive remains split between boxes and memory. This conversation between the work's creator and a major scholar in electronic literature documents both specific preservation challenges and systemic patterns in what the field chooses to preserve.
Topics include: figuring out web-born composition before established methodologies existed; the three decades of technical decisions that kept a 1996 work alive through format obsolescence and server migrations; and what gets lost …
Research On Image Feature Analysis Of Intangible Cultural Heritage Brocade Integrating Multi-Scale Visual Perception,
2026
1.School of Information Management, Nanjing University, Nanjing 210023
2.Jiangsu Key Laboratory of Data Engineering and Knowledge Service, Nanjing 210023
Research On Image Feature Analysis Of Intangible Cultural Heritage Brocade Integrating Multi-Scale Visual Perception, Ruiyang Yuan, Hao Wang, Shu Zhou, Hui Zhu, Jingwen Qiu
Journal of Scientific Information Research
[Purpose/significance] Addressing the challenges posed by the complex semantic characteristics of intangible cultural heritage brocade imagery, the difficulty in extracting their profound connotations, and the inadequate utilisation of multi-scale features by traditional deep learning models, this paper aims to explore a method for analysing the characteristics of intangible cultural heritage brocade images that integrates multi-scale visual features. [Method/process] This paper constructs a multi-scale feature analysis framework for intangible cultural heritage brocade images (ICH_BC), integrating convolutional neural networks with Transformer architectures. The framework employs ResNet to extract local texture and detail features from brocade images, utilises VIT to capture global structural …
Chi Meta-Project Ecosystem Overview - Spring 2026,
2026
CUNY New York City College of Technology
Chi Meta-Project Ecosystem Overview - Spring 2026, David B. Smith
Publications and Research
This paper offers a high-level account of the Center for Holistic Integration’s (CHI) meta-project ecosystem as visualized in the included system map. CHI provides an organizational structure framed around persistent meta-projects that support and extend individual initiatives across curriculum, scholarly and applied research, infrastructure, artistic production, AI development, cultural inquiry, and external partnerships. Rather than presenting the map as a static inventory of projects, the paper examines how its core domains function as living systems through which knowledge, tools, documentation, participants, and collaborations can accumulate over time. It also considers how CHI-mediated connectivity, institutional integration, and external funding allow the …
Benchmarking Current Progress In 3d Content Generation,
2026
University of Arkansas-Fayetteville
Benchmarking Current Progress In 3d Content Generation, Vuong Ho
Graduate Theses and Dissertations
In recent years, 3D generation has rapidly advanced with the development of powerful generative AI models capable of producing high-quality 3D content from various modalities, including text, images, and multi-view inputs. These advancements have significantly accelerated progress in applications such as gaming, virtual reality, robotics, and digital content creation. Despite this progress, there is still a lack of standardized and fair benchmarking protocols for evaluating 3D generation methods. Existing approaches are often assessed under inconsistent experimental settings, using different datasets, evaluation metrics, and processing pipelines. Such inconsistencies make reliable and objective comparisons difficult, limiting our understanding of the strengths and …
Oscbench: Benchmarking Object State Change In Text-To-Video Generation,
2026
Singapore Management University
Oscbench: Benchmarking Object State Change In Text-To-Video Generation, Xianjing Han, Bin Zhu, Shiqi Hu, Franklin Mingzhe Li, Patrick Carrington, Roger Zimmermann, Jingjing Chen
Research Collection School Of Computing and Information Systems
Text-to-video (T2V) generation models have made rapid progress in producing visually high-quality and temporally coherent videos. However, existing benchmarks primarily focus on perceptual quality, text–video alignment, or physical plausibility, leaving a critical aspect of action understanding largely unexplored: object state change (OSC) explicitly specified in the text prompt. OSC refers to the transformation of an object’s state induced by an action, such as peeling a potato or slicing a lemon. In this paper, we introduce OSCBench, a benchmark specifically designed to assess OSC performance in T2V models. OSCBench is constructed from instructional cooking data and systematically organizes action–object interactions into …
Tranx-Adapter: Bridging Artifacts And Semantics Within Mllms For Robust Ai-Generated Image Detection,
2026
Singapore Management University
Tranx-Adapter: Bridging Artifacts And Semantics Within Mllms For Robust Ai-Generated Image Detection, Wenbin Wang, Yuge Huang, Jianqing Xu, Yue Yu, Jiangtao Yan, Shouhong Ding, Pan Zhou, Yong Luo
Research Collection School Of Computing and Information Systems
Rapid advances in AI-generated image (AIGI) technology enable highly realistic synthesis, threatening public information integrity and security. Recent studies have demonstrated that incorporating texture-level artifact features alongside semantic features into multimodal large language models (MLLMs) can enhance their AIGI detection capability. However, our preliminary analyses reveal that artifact features exhibit high intra-feature similarity, leading to an almost uniform attention map after the softmax operation. This phenomenon causes attention dilution, thereby hindering effective fusion between semantic and artifact features. To overcome this limitation, we propose a lightweight fusion adapter, TranX-Adapter, which integrates a Task-aware Optimal-Transport Fusion that leverages the Jensen-Shannon divergence …
Dual-Diffusional Generative Fashion Recommendation,
2026
Singapore Management University
Dual-Diffusional Generative Fashion Recommendation, Mingzhe Yu, Lei Wu, Qianru Sun, Yunshan Ma
Research Collection School Of Computing and Information Systems
Personalized generative recommender systems have emerged as a promising solution for fashion recommendation. However, existing methods primarily rely on implicit visual embeddings from historical interactions, which often contain preference-irrelevant information and result in insufficient user behavior modeling. Moreover, these models typically generate only item images, providing limited interpretability. To address these limitations, we propose DualFashion, a Dual-Diffusional Generative Fashion Recommendation Architecture that jointly models image and text modalities for personalized and explainable recommendation. DualFashion adopts a dual-diffusion Transformer with image and text branches, where structured attribute-level captions and visual outfit information are jointly used as conditioning signals to model user …
