Activity Transition Graph Generation: How Far Are We?,
2026
Singapore Management University
Activity Transition Graph Generation: How Far Are We?, Jiakun Liu, Peixin Zhang, Han Hu, Yonghui Liu, Wei Minn, Ferdian Thung, Shahar Maoz, Eran Toch, Debin Gao, David Lo
Research Collection School Of Computing and Information Systems
Android applications (i.e., apps) are indispensable nowadays and are getting bigger and bigger with an increasing number offunctionalities. To understand how to access functionalities in an app, prior studies proposed tools to model the transitionsbetween functionalities with the activity transition graph (ATG). ATG is an important data structure and has been used forvarious Android app analyses, including app design, understanding, and testing. However, there is no benchmarking work onATG generation. It is still unclear whether the transitions identified by tools are correct and how many transitions are missed.To fill this gap, we manually identified all transitions in 98 applications to …
Breadquest: Enhancing Roguelike Accessibility Through Procedural Generation And Thematic Design,
2026
California Polytechnic State University, San Luis Obispo
Breadquest: Enhancing Roguelike Accessibility Through Procedural Generation And Thematic Design, Hahns Pena
Computer Science and Software Engineering
BreadQuest is a top-down roguelike dungeon crawler with a whimsical dessert theme that aims to make the genre more accessible while preserving strategic depth and replayability. Players explore procedurally generated dungeons, fight pastry-themed enemies, and collect bakery-inspired items that support a flavor-elemental combat system, with each run offering unique layouts, encounters, and rewards. Built in Unity with a modular, data-driven architecture, the game uses procedural generation techniques like Binary Space Partitioning, Voronoi diagrams, and Perlin noise to create varied and replayable levels. The project emphasizes approachable gameplay, cultural dessert inspiration, and replayability, with success evaluated through playtesting and player feedback.
Ai Interview Helper: A Tool For Assisting Search And Rescue Long-Profile Interviews,
2026
California Polytechnic State University, San Luis Obispo
Ai Interview Helper: A Tool For Assisting Search And Rescue Long-Profile Interviews, Dylan P. Starink
Master's Theses
In Search and Rescue (SAR) operations, time pressure and limited interviewer experience can lead to missed opportunities when interviewing a missing person’s friends and family. This thesis presents a real-time, end-to-end system that provides context-aware follow-up question suggestions as interviews unfold. Leveraging large language models (LLMs) and agentic design patterns, the system is intended to support interviewers by helping them identify relevant follow-up questions and pursue potentially overlooked lines of inquiry.
The system was evaluated through three mock interviews with two SAR interviewer participants across two events. Given the limited sample size, the results provide early insights into the feasibility …
A Pruning-Based Question-Answering For Interactive Video Search: A Simple Baseline,
2026
Singapore Management University
A Pruning-Based Question-Answering For Interactive Video Search: A Simple Baseline, Yu Tong Cheng, Phuong Anh Nguyen, Chong-Wah Ngo
Research Collection School Of Computing and Information Systems
There are various factors affecting the performance of video search. An imprecise query will enlarge search space and reduce the discriminative power of ranking functions. This problem is further exacerbated by the presence of numerous visually or semantically similar videos in large datasets. Consequently, users need to painstakingly browse through many highly similar candidates to locate the search target, leading to increased cognitive load and inefficient searching. Ideally, engaging users through interactive questioning to resolve uncertainties in the search process is an effective strategy for progressively narrowing down the search space. However, despite rapid advances in deep learning, generating informative …
Videocreator: An Agentic System For Multi-Turn Video Production,
2026
Singapore Management University
Videocreator: An Agentic System For Multi-Turn Video Production, Zhengyang Liang, Yan Shu, Cathal Gurrin, Nicu Sebe, Lizi Liao
Research Collection School Of Computing and Information Systems
Recent advances in video generation models enable visually compelling single clips. However, real-world video creation is inherently continuous and iterative: creators refine content over multiple rounds while maintaining narrative, style, and entity consistency. Existing standalone generators are largely stateless and lack memory of previously generated segments, making it difficult to produce a coherent and consistent video project. To address this gap, we present VideoCreator, a unified video agent that integrates generation and understanding with a project-level memory system. VideoCreator leverages understanding capabilities to perform fine-grained analysis of newly produced content and uses persistent memory to retain and reuse prior context …
Sam3-Litetext: An Anatomical Study Of The Sam3 Text Encoder For Efficient Vision-Language Segmentation,
2026
Singapore Management University
Sam3-Litetext: An Anatomical Study Of The Sam3 Text Encoder For Efficient Vision-Language Segmentation, Chengxi Zeng, Yuxuan Jiang, Ge Gao, Shuai Wang, Duolikun Danier, Bin Zhu, Stevan Rudinac, David Bull, Fan Zhang
Research Collection School Of Computing and Information Systems
Vision-language segmentation models such as SAM3 enable flexible, prompt-driven visual grounding, but inherit large, general-purpose text encoders originally designed for open-ended language understanding. In practice, segmentation prompts are short, structured, and semantically constrained, leading to substantial over-provisioning in text encoder capacity and persistent computational and memory overhead. In this paper, we perform a large-scale anatomical analysis of text prompting in vision–language segmentation, covering 404,796 real prompts across multiple benchmarks. Our analysis reveals severe redundancy: most context windows are underutilized, vocabulary usage is highly sparse, and text embeddings lie on a low-dimensional manifold despite high-dimensional representations. Motivated by these findings, we …
Co-Designing With Autistic Livestreamers: Care, Constraints, And Trade-Offs In Livestreaming,
2026
Singapore Management University
Co-Designing With Autistic Livestreamers: Care, Constraints, And Trade-Offs In Livestreaming, Terrance Mok, Anthony Tang, Lora Oehlberg
Research Collection School Of Computing and Information Systems
Autistic livestreamers use platforms like Twitch for social connection, self-expression, and community, but these spaces also impose ongoing social and emotional demands. Prior work has documented these experiences, but less is known about what autistic creators themselves envision for the tools and platforms they use. We address this gap through a Research through Design (RtD) co-design study with three autistic Twitch streamers, using speculative artefacts as discussion prompts to explore how participants reasoned about potential livestreaming technologies. Across three co-design activities, we identify three overarching tensions shaping autistic streaming practice: Expression versus Misinterpretation and Harm; Public Participation versus Control and …
“From Remembering To Shaping”: Narrating Shared Experiences By Co-Designing Cultural Heritage Artifacts In Collaborative Vr,
2026
Singapore Management University
“From Remembering To Shaping”: Narrating Shared Experiences By Co-Designing Cultural Heritage Artifacts In Collaborative Vr, Yushang Yang, Fanxu Meng, Fiona Fui-Hoon Nah, L. C. Ray
Research Collection School Of Computing and Information Systems
The ways people remember and recall places reveal an invisible aspect of cultural heritage (CH), reflecting how individuals and communities relate to these places. Heritage is communal, emerging through collaboratively constructed narratives rather than individual records. To probe how people may share collective memories, we designed an immersive two-person workflow for collaboratively co-designing 3D artifacts and environments in virtual heritage locations, using Generative AI (GenAI) to instantiate these intangible memories. Observations of the co-creation process revealed that participants merged prompts and model placements when negotiating different perspectives. They used spatial operations to compose scenes, and also to express personal and …
Not Too Early, Not All At Once: Design Tensions In Ai-Mediated Self-Disclosure In Online Dating,
2026
Singapore Management University
Not Too Early, Not All At Once: Design Tensions In Ai-Mediated Self-Disclosure In Online Dating, Pei-Hua Tsai, Tianyi Zhang, Emran Bin Elias Poh, Anthony Tang, Yung-Ju Chang
Research Collection School Of Computing and Information Systems
Online dating relies on self-disclosure, yet initial conversations are fragile: users must navigate uncertainty around timing, boundaries, and reciprocity with little shared context. While advances in AI raise the possibility of mediating disclosure, how such support might reshape the experience of early-stage relational disclosure remains underexplored. We conducted 29 semi-structured interviews to examine how daters envision AI-mediated self-disclosure in online dating. Our findings surface recurring design tensions rather than simple opportunities or risks. Participants welcomed guidance that could pace disclosure, support reflection, and reduce social awkwardness, but stressed preserving agency and authorship. They valued interpretive assistance for sense-making of ambiguous …
History To Future: Evolving Agent With Experience And Thought For Zero-Shot Vision-And-Language Navigation,
2026
Singapore Management University
History To Future: Evolving Agent With Experience And Thought For Zero-Shot Vision-And-Language Navigation, Guangzhao Dai, Shuo Wang, Zihan Wang, Guo-Sen Xie, Yang Yang, Jinshan Pan, Qianru Sun, Xiangbo Shu
Research Collection School Of Computing and Information Systems
Vision-and-Language Navigation in Continuous Environment (VLN-CE) requires an agent to follow language instructions to navigate the target destination. With the advancement of large language models (LLMs), recent efforts have explored adapting them for zero-shot VLN-CE, offering a promising solution in addressing the drawbacks of poor generalization in the training-based paradigm. However, existing LLM-based works primarily perform naive reasoning for decision-making and lack feedback, e.g., reviewing historical errors and predicting future potentials. Consequently, it may suffer from continuous failure for those initial error tasks. In this paper, we rethink LLM-based zero-shot VLN-CE and propose a new paradigm, named EvoNav, to improve …
Enhancing Pointing Gestures Of Non-Hmd Users In Asymmetric Collocated Mixed Reality Collaboration,
2026
Singapore Management University
Enhancing Pointing Gestures Of Non-Hmd Users In Asymmetric Collocated Mixed Reality Collaboration, Nam-Dang Vo, Van-Vinh Thai, Anthony Tang, Khanh-Duy Le
Research Collection School Of Computing and Information Systems
A common collocated group setting in mixed-reality (MR) collaboration is a person wearing a MR headset (HMD user) and presenting MR contents to audiences who are not provided with such specialized devices (Non-HMD users). In this setting, while Non-HMD users can view the MR environment shown on a large physical display, it still remains challenging for the HMD user to interpret their pointing gesture when they spatially refer to objects in the MR environment. To address this, we designed and evaluated two pointing techniques—SCREEN and SCREEN+SPACE—that support Non-HMD users in referring to MR content. Screen pointing allows users to refer …
Happycal: Designing Text And Image-Based Supports For Savouring Positive Work Experiences,
2026
Singapore Management University
Happycal: Designing Text And Image-Based Supports For Savouring Positive Work Experiences, Molly Stewart, Minghao Cai, Anthony Tang, Sam Liu, Chris Mosunic, Sowmya Somanath
Research Collection School Of Computing and Information Systems
Savouring positive work experiences can promote positive affect and well-being at work, yet there is limited guidance on how digital applications can support workers to engage in savouring. We developed HappyCal, a work-focused savouring application offering two forms of savouring support: text-based, a common modality in workplace reflection tools, and images, a largely unexplored approach in work-related savouring. We conducted an exploratory qualitative study where participants (N=36) used HappyCal over five days and engaged in savouring through either a text-only modality (n=17) or text input paired with image output (n=19). We found that (1) participants in both groups reported heightened …
Perceptual And Geometric Advances In Crowd Simulation,
2026
New Jersey Institute of Technology
Perceptual And Geometric Advances In Crowd Simulation, Bilas Talukdar
Dissertations
Simulating realistic crowd motion remains a fundamental challenge in computer graphics and multi-agent systems, as it requires modeling both physically plausible interactions and perceptually natural behaviors. Existing crowd simulation methods typically employ simplified geometric abstractions, most commonly circular agent representations, and model navigation using either analytical interaction formulations (e.g., force, velocity, or constraint-based methods) or learned policies derived through reinforcement learning. Despite their effectiveness, these approaches often overlook detailed geometric structure and do not explicitly account for perceptual realism. This dissertation addresses these challenges by improving the realism of virtual crowd simulation through two key advancements: perceptual preference learning and …
Toward Learning-Based Reconstruction And Part Decomposition Of Man-Made 3d Geometry: Neural Implicit Representations And Scalable Supervision,
2026
New Jersey Institute of Technology
Toward Learning-Based Reconstruction And Part Decomposition Of Man-Made 3d Geometry: Neural Implicit Representations And Scalable Supervision, Shen Fan
Dissertations
Digital three-dimensional (3D) models are central to engineering design, analysis, and manufacturing, but learning pipelines for man-made geometry often operate on sampled carriers that do not preserve all of the structure present in exact CAD representations. This dissertation studies learning-based reconstruction and part decomposition for structured man-made 3D geometry, from general object benchmarks to CAD-derived datasets, with a focus on neural implicit representations trained from signed-distance samples, point clouds, and tessellated meshes. The goal is to make these models more accurate, more part-aware, and more consistently supervised.
First, signed distance function (SDF) reconstruction with implicit neural representations is improved through …
Streamlined Biomedical Image Processing Pipelines,
2026
University of Massachusetts Boston
Streamlined Biomedical Image Processing Pipelines, Jiehyun Kim
Graduate Doctoral Dissertations
This dissertation focuses on advancing carotid artery analysis through a series of visualizations and deep learning tools for calcified plaque assessment and related biomedical imaging tasks. Accurate plaque evaluation is essential, but current workflows depend on slow, clinician-dependent manual review. To address these limitations, this work introduces the CACTAS framework, a set of tools and methods that enable fast and reliable plaque segmentation for clinicians.
The first study, the CACTAS-Tool, provides a web-based labeling tool that enables clinicians to label plaque directly in three dimensions through a streamlined one-click interface. This tool significantly reduces the effort required to generate high-quality …
Du Undergraduate Showcase Abstracts: Research, Scholarship, And Creative Works,
2026
University of Denver
Du Undergraduate Showcase Abstracts: Research, Scholarship, And Creative Works, Sophia Wismar, Henry Staats, Allison Metzler, Chloe Puckett, Rachel Levine, Christa Kilpatrick, Scott Wolf, Joe Walsh, Grace Doolittle, John Engebreston, Zoe Lopez, Christopher Aaby, Audrey Duff, Timothy Sisk, Katelyn Lamberton, Angela Narayan, Gilkah Argueta, Habiba Samir, Girena Tesfazghi, Genet Kenore, Sinit Tesfamariam, Effley Brooks, Abi Newell, Megan Doherty, Natalie Baer, Lexi Blood, Talya Riciputi, Jessica Jimenez, Devin Hernandez, Lynn Clark, Taj Kumar, Sunil Kumar, Allen Rutman, Mira Pronobis, Tess Carson, Anna Sher, Frankie Stroud, Tamra Pearson D'Estree, Alyssa Wilson, Emily Melnick, Jenalee Doom, Yihang Gao, Gwendolyn Geiger, Noah Gettle, Scott Nichols, Clare Ayoub, Cara Dienno, Sunny Walker, Zoe Hansen, Maya Wheeler, Addison Rice, Patrick Martin, Sanjana Acharya, Daniel Mcintosh, Amanda Mckellips, Calli Cain, Justin Blake, Peter Sokol-Hessner, Natalie Miller, Max Weisbuch, Sophia Dellota, John Macikas, Charlotte Snow, Mark Siemens, Zoe Lynch, Alex Huffman, Prachi Shah, Jason Roney, Halcyon Levi, Nicole Herzog, Andrea Koly, Daniel Linseman, Annie London, Xi Yang, Avery Zwisler, Jane Smith, Chaz Contag, Michael Kerwin, Lucy Rand, Grace Schroeder, Michelle Rozenman, Nissa Tapper, Guiming Zhang, Mateo Mazariego-Halpern, Keith Meyer, Julie Do, Dakota Park-Ozee, Travis Herink, Kara Neu, Jonathan Plomin, Eve-Odine Duchaufour, Debbie Gale Mitchell, Tennyson Anderson-Stricklin, Lily Treitz, Samantha Rosenberger, Sierra Griffith, Finley Joseph, Daniel Sampson, Emmy Davis, Skyler Kasnoff, Evon Lopez, Vivian Nguyen, Cassy Young, Franklin Sellner, Martin Tobon, Ila Graham, Zach Billings, Holden Hedit, Decatur Boland, Paul Kosempel, Cory Chandler, Jay Mahoney, Sam Dragan, Susan Dagget, Yarrow Ator, Heidi Vuletich, Owen Weber, Andrew Kloeppel, Petersen Gray, Mandi Schaeffer-Fry, Razleen Bassra, Bryanna Rodriguez, Christina Blue, Taubie Sanders, Rachel Epstein, Luke Milburn, Camryn Evans, Ezra Martinez, Mary Westwood, Gabri Notov, Robin Tinghitella, Lilou Cabrol, Eli Barbour, Juliet Mendik, Selma Myers, Zac Wise, Noah Fahlin, Michelle Knowles, Abigail Hopper, Michael Greenberger, Romi Laclair, Sarah Watamura, Sabrina Efroymson, Casey Barker, Sydney Seltzer, Bryn Yehle, Jennifer Hoffman, Sara Garcia, Ryuka Nagamine, Trevor Briggs, Remy Le Boeuf, Elena Krone, Eileen Farrell, Regan O'Rourke, Elena Roel, Greg Mortimer, Ali Ayoub, Stefani Langehennig, Caitlin Turk, Logan Scmid, Stefan Chavez-Norgaard, Karen Kim, Tatiana Peccedi, Courtney Cassidy, John Sebesta, Rhianna Lewis, Janice Bening-Lacek, Vivian Lawless, Mckenna Hanson, Jeffrey Amidon, Riya Joshi, Ram Ambre, Brady Worrell, Perrin Schneider, Ali Azadani, Brooke Agulnek, Lyndsie Salvagio, Elise Siemanowki, Yan Qin, Andre Allen, Melodie Nguyen, Megan Livengood, Abby Reams, Saffron Hartreeve, Bri Wylie, Sarah Brookman, Mariah Loiacono, Green Russo, Abhia Lodhi, Gabrielle Welsh, Nika Spehar, Shahked Levin, Evrim Baykal, Kimberly Chiew, Jocelyn Torres, Kailey Hicks, Mykaela Tanino-Springsteen, Audrey Bellows, Akam Chahal, Madeline Tepper, Shannon Murphy, Alexa Fonseca, Deborah Han, Cassandra Perez, Oluwatoyin Alaba, Julia Roncoroni, Vy Nguyen, Nana Burn, Sarah Sasse, Rubin Tuder, Anthony Gerber, Nancy Lorenzon, Christine Vohwinkel, Camryn Gunter, Tristan Weber, Sam Rommel, Brian Michel, Muskan Fatima, Alannah Oleson, Kira Frey, Edward Garrido, Beckett Morris, Kerstin Haring, Drew Middleton, Abigail Walpert, Liam Dee, Gabby Ishaw, Cole Carnes, Maddie Weiser, Claire Fox, Valeriia Vlasenko, Kateri Mcrae, Riley Smith, Abigail Templin, Kushani Rajapaksha
DU Undergraduate Research Journal Archive
Abstracts from the DU Undergraduate Research Showcase.
Frictional Intelligence,
2026
Rhode Island School of Design
Frictional Intelligence, Posheng Cheng
Masters Theses
This is an experimental interaction design project that challenges anthropomorphism in human-computer interaction. In particular, the recent advancement of artificial intelligence technologies like Large Language Models has taken anthropomorphism to new heights. The conversational chatbot interface of AI prioritizes mimicking an inherently human communication medium to maximize human-likeness. However, anthropomorphism has several downsides. Conversational interfaces obscure the limitations and the tangible cost of the technology. They also imply fictional moral status and human-level cognitive capabilities, which means general public sentiment focuses on the ``overhyped'' excitement and fear rather than on other socio-ethical and capacity questions that are far more urgent …
Conditional Product Sampling For Gaussian Process Implicit Surfaces,
2026
Dartmouth College
Conditional Product Sampling For Gaussian Process Implicit Surfaces, Song Shi
Dartmouth College Master’s Theses
Gaussian Process Implicit Surfaces (GPISes) provide a powerful and unified stochastic geometry representation for rendering surfaces, volumes, and the rich continuum between them. Recent work has shown that GPISes can model a broad space of visual appearances under a unified light transport framework. However, practical rendering with GPISes remains challenging: existing estimators can become inefficient for particular correlation structures, and highly anisotropic or heightfield-like GPISes require specialized treatment to obtain robust variance reduction.
This thesis extends recent work on GPIS rendering by introducing a new next-event estimation (NEE) technique for anisotropic GPISes.We show that standard NEE provides diminishing benefits as …
Approval Motivations In Sharing Humorous Tiktok's,
2026
Sartell High School
Approval Motivations In Sharing Humorous Tiktok's, Mariam Al-Areedy
InnovateHER Meeting 2026
TikTok is a short-form video platform where users create and share content that is often centered around humor, trends, and everyday social experiences. In face-to-face interactions, people typically rely on immediate feedback to navigate conversations, often using approval seeking behaviors to gain positive reactions and rejection-avoidant behaviors to reduce the risk of negative judgement. While these motivations are well-established in in-person settings, less is known about how they function in digital environments like TikTok, where teens privately share humorous content without immediate social cues to guide their interactions. My general hypothesis was that both rejection avoidance and approval-seeking behaviors will …
3d Puzzle Generation Beyond Voxelized Parts,
2026
Yale University
3d Puzzle Generation Beyond Voxelized Parts, Iris Xia
Computer Science Theses
Burr puzzles are interlocking assemblies whose pieces must be inserted and removed through tightly constrained motions. Designing them is difficult because geometric fit, interlocking behavior, and disassembly order are tightly coupled, while existing computational methods remain largely limited to voxelized or template-based constructions.
This work presents a framework for 3D puzzle generation beyond voxelized parts. The method replaces local mobility heuristics with a certified search over geometry edits. Starting from a topological contact specification, it constructs signed distance fields for individual parts, applies complementary local edits, and validates each candidate using exact geometric checks and a kernel disassembly graph. The …
