Open Access. Powered by Scholars. Published by Universities.®

Computer Sciences Commons

Open Access. Powered by Scholars. Published by Universities.®

Databases and Information Systems

Institution
Keyword
Publication Year
Publication
Publication Type
File Type

Articles 91 - 120 of 7250

Full-Text Articles in Computer Sciences

Human Subject Studies For The Alignment Of Llm-As-A-Judge Evaluation Metric For Science News, Gabriel Vega Osborne Mar 2026

Human Subject Studies For The Alignment Of Llm-As-A-Judge Evaluation Metric For Science News, Gabriel Vega Osborne

Knowledge and Creativity Expo

Science news has become an important vehicle to disseminate scientific breakthroughs, discoveries, and technological innovations. With the advancement of large language models and related AI models, it is possible to automatically generate science news from scientific papers, extending the reader population from domain scientists to a broader scope. However, how to evaluate the quality of the generated news warrants research. Traditional token based metrics have been shown to fail to evaluate the semantics and nuances of science news. Inspired by the fact that a major goal of science news is to educate readers with new knowledge, we thus propose knowledge …


Large-Scale File Fragment Classification Via Multi-View Learning, Samuel Hildebrand Mar 2026

Large-Scale File Fragment Classification Via Multi-View Learning, Samuel Hildebrand

LSU Master's Theses

File reassembly is one of the most fundamental tasks in digital forensics, enabling recovery of data from potentially damaged storage media even when file system metadata is unavailable. This thesis reviews more than two decades of work in the realm of file carving, with a particular focus on fragmented file carving, which remains a focus of research, and file fragment classification, a principal component of fragmented file carving. This thesis serves a literature review of both file carving and fragmented file carving, surveys the massive amounts of data needed for the task of fragment classification and the datasets that serve …


Kms-Net: Kolmogorov–Arnold-Based Multi-Scale Attention Network For Cardiac Segmentation, Abid Mehmood, Hassan Ali, David Noule Tolno, Sery Gahouidi Thierry S, Muhammad Saeed, Naeem Ahmed Mar 2026

Kms-Net: Kolmogorov–Arnold-Based Multi-Scale Attention Network For Cardiac Segmentation, Abid Mehmood, Hassan Ali, David Noule Tolno, Sery Gahouidi Thierry S, Muhammad Saeed, Naeem Ahmed

Research & Publications

Accurate segmentation of cardiac structures in 2D echocardiography is essential for diagnosing cardiovascular disease and computing clinical metrics such as chamber volumes and ejection fraction. Conventional U-Net architectures excel at extracting local spatial features but struggle with long-range dependencies inherent in noisy ultrasound images, while pure Transformer-based models capture global context at the expense of fine boundary detail. To address these limitations, we propose KMS-Net, a novel hybrid segmentation architecture that integrates Kolmogorov–Arnold Networks (KANs), a class of learnable, spline-based function approximators that replace fixed activation functions with trainable nonlinear mappings, alongside multi-scale attention mechanisms. Specifically, spline-based KAN layers (grid …


Data Centers In Mountain West Markets, 2026, Cason Noll, Krish Sharma, Maisoon Faris, Olivia K. Cheche, Caitlin J. Saladino, William E. Brown Jr. Mar 2026

Data Centers In Mountain West Markets, 2026, Cason Noll, Krish Sharma, Maisoon Faris, Olivia K. Cheche, Caitlin J. Saladino, William E. Brown Jr.

Transportation & Infrastructure

This fact sheet reports on the distribution and geographic concentration of data centers across the Mountain West states of Arizona, Colorado, Nevada, New Mexico and Utah as of March 6th, 2026. Using data from DataCenterMap, this fact sheet examines the number of data centers in each Mountain West state and further analyzes market-level  distribution, defined as cities within each state where data centers are located. The data are used to compare state totals and to rank Mountain West markets from highest to lowest based on the number of data centers operating in that area.


Navigation Beyond Wayfinding: Robots Collaborating With Visually Impaired Users For Environmental Interactions, Shaojun Cai, Nuwan Janaka, Ashwin Ram, Janidu Shehan, Yingjia Wan, Kotaro Hara, David Hsu Mar 2026

Navigation Beyond Wayfinding: Robots Collaborating With Visually Impaired Users For Environmental Interactions, Shaojun Cai, Nuwan Janaka, Ashwin Ram, Janidu Shehan, Yingjia Wan, Kotaro Hara, David Hsu

Research Collection School Of Computing and Information Systems

Robotic guidance systems have shown promise in supporting blind and visually impaired (BVI) individuals with wayfinding and obstacle avoidance. However, most existing systems assume a clear path and do not support a critical aspect of navigation—environmental interactions that require manipulating objects to enable movement. These interactions are challenging for a human–robot pair because they demand (i) precise localization and manipulation of interaction targets (e.g., pressing elevator buttons) and (ii) dynamic coordination between the user’s and robot’s movements (e.g., pulling out a chair to sit). We present a collaborative human–robot approach that combines our robotic guide dog’s precise sensing and localization …


Addressing Graph Heterogeneity And Heterophily From A Spectral Perspective, Kangkang Lu, Yanhua Yu, Ruopei Guo, Nan Cheng, Zhiyong Huang, Yunshan Ma, Meiyu Liang, Yuling Wang, Xiting Qin, Yimeng Ren, Tat-Seng Chua Mar 2026

Addressing Graph Heterogeneity And Heterophily From A Spectral Perspective, Kangkang Lu, Yanhua Yu, Ruopei Guo, Nan Cheng, Zhiyong Huang, Yunshan Ma, Meiyu Liang, Yuling Wang, Xiting Qin, Yimeng Ren, Tat-Seng Chua

Research Collection School Of Computing and Information Systems

Graph Neural Networks (GNNs) face two key challenges, heterogeneity and heterophily, which often degrade performance. Existing approaches either focus narrowly on specific meta-paths, limiting their expressiveness, or are expressive but cannot effectively leverage higher-order neighbors. In this paper, we propose the Heterogeneous Heterophilic Spectral Graph Neural Network (H2SGNN), which combines local independent filtering to adaptively handle meta-path subgraphs with varying homophily ratios, and global hybrid filtering to capture high-order neighbor interactions with linear computational complexity. On five heterogeneous graph benchmarks—DBLP, ACM, IMDB, AMiner, and Yelp—H2SGNN consistently outperforms strong baselines, for example, achieving +1.0% Macro-F1 and +1.3% Micro-F1 on IMDB. It …


Opencil: Benchmarking Out-Of-Distribution Detection In Class Incremental Learning, Wenjun Miao, Guansong Pang, Trong-Tung Nguyen, Ruohuan Fang, Jin Zheng, Xiao Bai Mar 2026

Opencil: Benchmarking Out-Of-Distribution Detection In Class Incremental Learning, Wenjun Miao, Guansong Pang, Trong-Tung Nguyen, Ruohuan Fang, Jin Zheng, Xiao Bai

Research Collection School Of Computing and Information Systems

Class incremental learning (CIL) aims to learn a model that can not only incrementally accommodate new classes, but also maintain the learned knowledge of old classes. Out-of-distribution (OOD) detection in CIL is to retain this incremental learning ability, while being able to reject unknown samples that are drawn from different distributions of the learned classes. This capability is crucial to the safety of deploying CIL models in open worlds. However, despite remarkable advancements in the respective CIL and OOD detection, there lacks a systematic and large-scale benchmark to assess the capability of advanced CIL models in detecting OOD samples. To …


Stock Market Price Prediction Using Big Data Models Comparison Analysis, Vibhor Pal Mar 2026

Stock Market Price Prediction Using Big Data Models Comparison Analysis, Vibhor Pal

Shelby Hall Graduate Research Forum Posters

The stock market consists of complex financial datasets, and achieving stock price real time prediction needs an efficient big data framework for processing. This paper compares big data distributed data processing frameworks for forecasting stock prices using Graph Neural Networks (GNNs) - Apache Flink and Apache Spark. We analyze 70 publicly traded companies’ monthly data for the last 5 years from Yahoo Finance, ranked by Price-to-Earnings (P/E). In the companies’ datasets, there may be a connection or similarity between companies, and this can lead to similar stocks’ price behavior. These interfirm relationships are maintained by GNNs models, and their output …


Private Set Intersection: A Systematic Review, Yunbo Yang, Defan Zhu, Jianting Ning, Qi Feng, Xiaoguo Li, Yuejia Cheng, Guomin Yang, Kui Ren Mar 2026

Private Set Intersection: A Systematic Review, Yunbo Yang, Defan Zhu, Jianting Ning, Qi Feng, Xiaoguo Li, Yuejia Cheng, Guomin Yang, Kui Ren

Research Collection School Of Computing and Information Systems

Various services, such as search engines, are increasingly deployed in cloud-based and distributed systems. However, data are typically managed by trusted servers, making user privacy and data security critical concerns. Private set intersection (PSI) is a powerful cryptographic primitive that enables multiple parties to compute the intersection of their datasets without revealing private inputs. It has been extensively studied over the past two decades, leading to significant gains in computational and communication efficiency. Yet, in many real-world scenarios, revealing the raw intersection may still leak sensitive information. To address this, numerous PSI variants have been developed to meet different application …


A New Tool For Handling Multiracial And Multi-Identity Data In Health Research, Gabriel J. Merrin Feb 2026

A New Tool For Handling Multiracial And Multi-Identity Data In Health Research, Gabriel J. Merrin

Population Health Research Brief Series

When surveys ask about race or ethnicity, a growing number of Americans select more than one category. The multiracial population now represents over 10% of the U.S. population and is the fastest growing racial group in the country. Yet researchers routinely collapse these individuals into an “other race” category for statistical analysis, rendering specific subgroups invisible. This brief introduces CATAcode, a free software tool that helps researchers systematically explore, document, and prepare check-all-that-apply demographic data for statistical modeling. In a demonstration with over 8,000 high school students, CATAcode revealed 85 distinct racial identity combinations from just eight response options. The …


G-Trac: Graph-Textual Representations Alignment For Cold-Start Recommendations, Li Yang Chang, Yuan Fang, Ming Feng Tsai, Chuan Ju Wang Feb 2026

G-Trac: Graph-Textual Representations Alignment For Cold-Start Recommendations, Li Yang Chang, Yuan Fang, Ming Feng Tsai, Chuan Ju Wang

Research Collection School Of Computing and Information Systems

The cold-start problem remains a significant challenge in recommendation systems, particularly for new users or unseen items with little to no historical data. Existing methods, including graph neural networks, often struggle in such scenarios. Inspired by the success of transformer models in natural language processing, we propose G-TRAC (Graph-Textual Representations Alignment for Cold-start Recommendations), a novel approach that integrates transformer-based textual modeling with graph neural networks. By effectively leveraging both textual and structural information, G-TRAC addresses cold-start challenges more effectively. Extensive experiments demonstrate its ability to enhance recommendation quality and generalize well across diverse scenarios.


Responsible Ai In Teaching And Learning, Asa B. Stone, Mark C. Stone, Derek M. Heeren, Santosh Pitla Feb 2026

Responsible Ai In Teaching And Learning, Asa B. Stone, Mark C. Stone, Derek M. Heeren, Santosh Pitla

PRAIRIE: Pioneering Responsible AI for Research, Innovation, and Education

The rapid adoption of generative artificial intelligence (AI) in higher education presents both transformative opportunities and significant pedagogical risks. While AI tools are becoming embedded in academic and professional environments, their integration into teaching and learning raises critical questions about cognitive engagement, academic integrity, equity, and skill development. This white paper proposes a principled framework for the responsible integration of AI in higher education, grounded in the dual commitment to AI literacy and the cultivation of durable skills.

The framework articulates six core principles: purposefulness; transparency; integrity and attribution; critical AI literacy; equity and access; and privacy and data protection. …


Online Visual Query System For Real-Time Large-Scale Spatio-Temporal Data Explorations With Error Bounds, Xueqi Huang Feb 2026

Online Visual Query System For Real-Time Large-Scale Spatio-Temporal Data Explorations With Error Bounds, Xueqi Huang

Dissertations, Theses, and Capstone Projects

Modern datasets continue to grow in size, dimensionality, and heterogeneity, creating increasing tension between the need for responsive, interactive analysis and the computational cost of accessing, aggregating, and visualizing large volumes of data. Traditional database engines and visualization tools often assume that full data retrieval is feasible or that exact computation is necessary for meaningful insight. In practice, however, analysts frequently benefit from timely, uncertainty-aware approximations than from delayed and exact results. This thesis investigates how data summarization techniques, specifically mergeable sketches can be combined with progressive, out-of-core visualization methods to support interactive exploration of datasets that exceed main memory. …


The Feelit System: Application Content-Aware Perspectives And Challenges On Understanding User Likes In Social Network Posts, Konstantinos Theocharidis, Hady W. Lauw, Panagiotis Karras Feb 2026

The Feelit System: Application Content-Aware Perspectives And Challenges On Understanding User Likes In Social Network Posts, Konstantinos Theocharidis, Hady W. Lauw, Panagiotis Karras

Research Collection School Of Computing and Information Systems

In a series of our prior works, we study influence and subscription maximization problems in social networks that are based on posts having influential content; as content we consider a set of features where each feature corresponds to a specific social network page, whereas influence and subscription relate to gaining the postlike and subscription-to-brand page of targeted users, respectively; subscription is conceptually achieved as repetitive influence on users. So, both influence and subscription depend on content that gains the likes of users; however, to be realistic, modeling and estimating such likes is a complex problem that has not been adequately …


Cellscout: Visual Analytics For Mining Biomarkers In Cell State Discovery, Rui Sheng, Zelin Zang, Jiachen Wang, Yan Luo, Zixin Chen, Yan Zhou, Shaolun Ruan, Huamin Qu Feb 2026

Cellscout: Visual Analytics For Mining Biomarkers In Cell State Discovery, Rui Sheng, Zelin Zang, Jiachen Wang, Yan Luo, Zixin Chen, Yan Zhou, Shaolun Ruan, Huamin Qu

Research Collection School Of Computing and Information Systems

Cell state discovery is crucial for understanding biological systems and enhancing medical outcomes. A key aspect of this process is identifying distinct biomarkers that define specific cell states. However, difficulties arise from the co-discovery process of cell states and biomarkers: biologists often use dimensionality reduction to visualize cells in a two-dimensional space. Then they usually interpret visually clustered cells as distinct states, from which they seek to identify unique biomarkers. However, this assumption is often this assumption often fails to hold due to internal inconsistencies in a cluster, making the process trial-and-error and highly uncertain. Therefore, biologists urgently need effective …


Prompt Tuning Without Labeled Samples For Zero-Shot Node Classification In Text-Attributed Graphs, Sethupathy Parameswaran, Suresh Sundaram, Yuan Fang Feb 2026

Prompt Tuning Without Labeled Samples For Zero-Shot Node Classification In Text-Attributed Graphs, Sethupathy Parameswaran, Suresh Sundaram, Yuan Fang

Research Collection School Of Computing and Information Systems

Node classification is a fundamental problem in information retrieval with many real-world applications, such as community detection in social networks, grouping articles published online and product categorization in e-commerce. Zero-shot node classification in text-attributed graphs (TAGs) presents a significant challenge, particularly due to the absence of labeled data. In this paper, we propose a novel Zero-shot Prompt Tuning (ZPT) framework to address this problem by leveraging a Universal Bimodal Conditional Generator (UBCG). Our approach begins with pre-training a graph-language model to capture both the graph structure and the associated textual descriptions of each node. Following this, a conditional generative model …


Removed: When Taxi Drivers Meet Dynamic Pricing: A Lesson From Singapore's Justgrab Program, Shih-Fen Cheng, Wen-Tai Hsu, Jing Li Feb 2026

Removed: When Taxi Drivers Meet Dynamic Pricing: A Lesson From Singapore's Justgrab Program, Shih-Fen Cheng, Wen-Tai Hsu, Jing Li

Research Collection School Of Economics

This paper studies how dynamic pricing influences taxi drivers’ behaviors using a unique event, the inception of the JustGrab program in Singapore in 2017, which introduces dynamic pricing to some, but not all, taxi drivers. This is the first time in history that traditional taxi drivers have access to dynamic pricing. Using data covering the universe of taxi trips before and after the inception of JustGrab, we find that there is spatial reallocation that directs more taxi drivers to the previously less-served areas, that there is also a temporal reallocation that directs more taxi drivers to rush hours, as well …


Learnable Game-Theoretic Policy Optimization For Data-Centric Self-Explanation Rationalization, Yunxiao Zhao, Zhiqiang Wang, Xingtong Yu, Xiaoli Li, Jiye Liang, Ru Li Feb 2026

Learnable Game-Theoretic Policy Optimization For Data-Centric Self-Explanation Rationalization, Yunxiao Zhao, Zhiqiang Wang, Xingtong Yu, Xiaoli Li, Jiye Liang, Ru Li

Research Collection School Of Computing and Information Systems

Rationalization, a data-centric framework, aims to build self-explanatory models to explain the prediction outcome by generating a subset of human-intelligible pieces of the input data. It involves a cooperative game model where a generator generates the most human-intelligible parts of the input (i.e., rationales), followed by a predictor that makes predictions based on these generated rationales. Conventional rationalization methods typically impose constraints via regularization terms to calibrate or penalize undesired generation. However, these methods are suffering from a problem called mode collapse, in which the predictor produces correct predictions yet the generator consistently outputs rationales with collapsed patterns. Moreover, existing …


Significance, Challenges, And Policy Recommendations For Strengthening Database Development To Support Ai For Science In China, Kaihua Chen, Hongxin Liu, Rui Guo Jan 2026

Significance, Challenges, And Policy Recommendations For Strengthening Database Development To Support Ai For Science In China, Kaihua Chen, Hongxin Liu, Rui Guo

Bulletin of Chinese Academy of Sciences (Chinese Version)

With the rapid emergence of research intelligence driven by big data and artificial intelligence (AI), high-quality, openly shared scientific databases have become a strategic focal point for scientific innovation and enhancing technological competitiveness. Major countries around the world are increasingly recognizing the foundational role of scientific databases in advancing basic research. While continuously strengthening their own scientific data infrastructure through a series of initiatives, they have simultaneously imposed restrictions and suppression on the development of AI technologies in China, including those involving research data. Against this backdrop, building an autonomous and controllable scientific data ecosystem to support research intelligence is …


Data Altruism: Eu Solution And Path Of Localization In China, Teng Wu Jan 2026

Data Altruism: Eu Solution And Path Of Localization In China, Teng Wu

Bulletin of Chinese Academy of Sciences (Chinese Version)

Data altruism transcends the profit-seeking, monopolistic and competitive nature of the market mechanism. It is driven by innovation and encourages data subjects and holders to share information guided by the public interest. As a new type of data application and service model, it aims to achieve a virtuous cycle of the data ecosystem while optimize the utilization of data resources. Tracing back to the source, the theoretical foundation of data altruism from the ethical theory of altruism, and the concept of its budding was supported by the data for good. Analysis shows that the specific scheme of the EU data …


Ransomware As Organization: A Comparative Analysis Of Corporate And Criminal Structures In Conti, George Urling Jan 2026

Ransomware As Organization: A Comparative Analysis Of Corporate And Criminal Structures In Conti, George Urling

Theses, Dissertations and Capstones

Cybercriminal groups continue to pose major threats to global cybersecurity. One of the most common types of cybercriminal groups are, “Ransomware-as-a-Service (RaaS)" groups, who create and sell ransomware. While research is conducted into the development of ransomware, there is limited reporting on the organizational structure and habits of RaaS groups. In 2022, prominent RaaS group Conti had their chat logs leaked, with the logs ranging from 2020 to 2022. This study seeks to provide a deeper understanding of RaaS group structures by utilizing the Conti leaked logs as a case study. The study, entitled “Ransomware as Organization: A Comparative Analysis …


Reinforce Trustworthiness In Multimodal Emotional Support System, Huy M. Le, Dat Tien Nguyen, Ngan T. T. Vo, Tuan D. Q. Nguyen, Nguyen Le Binh, Duy Minh Ho Nguyen, Daniel Sonntag, Lizi Liao, Binh T. Nguyen Jan 2026

Reinforce Trustworthiness In Multimodal Emotional Support System, Huy M. Le, Dat Tien Nguyen, Ngan T. T. Vo, Tuan D. Q. Nguyen, Nguyen Le Binh, Duy Minh Ho Nguyen, Daniel Sonntag, Lizi Liao, Binh T. Nguyen

Research Collection School Of Computing and Information Systems

In today's world, emotional support is increasingly essential, yet it remains challenging for both those seeking help and those offering it. Multimodal approaches to emotional support show great promise by integrating diverse data sources to provide empathetic, contextually relevant responses, fostering more effective interactions. However, current methods have notable limitations, often relying solely on text or converting other data types into text, or providing emotion recognition only, thus overlooking the full potential of multimodal inputs. Moreover, many studies prioritize response generation without accurately identifying critical emotional support elements or ensuring the reliability of outputs. To overcome these issues, we introduce …


Paid Search Marketing Vs. Search Engine Optimization: Analytical Models Of Search Marketing Based On Search Engine Quality, Kai Li, Chunyang Shen, Mei Lin, Zhangxi Lin Jan 2026

Paid Search Marketing Vs. Search Engine Optimization: Analytical Models Of Search Marketing Based On Search Engine Quality, Kai Li, Chunyang Shen, Mei Lin, Zhangxi Lin

Research Collection School Of Computing and Information Systems

As search engines are leading revenue growth in online marketing, search marketing has become a popular area of academic research. Although search engine advertising has interested researchers for decades and much has been learned, one thing that puzzles scholars is why search engine optimization companies are tolerated rather than excluded from the market, even though they capture a significant share of the advertising market. In this paper, we shed light on this phenomenon and establish an analytical model based on organic search quality. Through analysis of the model, we were able to draw several intriguing conclusions. First, there is no …


Design Principles For Customer-Engaging Digital Service Systems: An Action Research Study, Keng Leng Siau, Xiaofeng Chen, Xin Tan Jan 2026

Design Principles For Customer-Engaging Digital Service Systems: An Action Research Study, Keng Leng Siau, Xiaofeng Chen, Xin Tan

Research Collection School Of Computing and Information Systems

Digital services represent a business approach employed by organizations to operate in the digital environment. However, systematic development guidelines for developing quality digital service systems are lacking in the literature. The authors identified four general challenges for developing and implementing customer-engaging digital service systems (CEDSS). By employing the method of canonical action research in a digital service system project, they derived 10 design principles for developing high-quality CEDSS. They empirically evaluated the design principles in the development project and through follow-up focus group sessions. The design principles provide applicable and actionable guidelines for the development of CEDSS.


Tackling The Societal And Regulatory Challenges Of Emerging Technologies: A Case Study Of Deepfake, Jingyao Li Jan 2026

Tackling The Societal And Regulatory Challenges Of Emerging Technologies: A Case Study Of Deepfake, Jingyao Li

2026

Governing emerging technologies such as Artificial Intelligence (AI) poses enduring challenges for policymakers, industries, and societies. Early-stage governance is often hindered by limited understanding of technological implications, rapid innovation cycles, and resistance from powerful industry actors who favor minimal oversight. Yet, timely and effective governance is essential, as new technologies are most malleable in their formative stages. This dissertation examines how emerging technologies can be governed effectively by using deepfakes technology as a focal case. This dissertation comprises three interrelated studies.

The first paper reviews the literature on deepfakes and emerging technology governance, identifying the distinct characteristics of deepfake technology …


Toward Efficient And Scalable Scientific Data Management Through Quality-Oriented Data Compression, Pu Jiao Jan 2026

Toward Efficient And Scalable Scientific Data Management Through Quality-Oriented Data Compression, Pu Jiao

Theses and Dissertations--Computer Science

Scientific simulations and instruments now produce data at rates that overwhelm the storage, memory, and network subsystems of modern high-performance computing (HPC) facilities. Error-bounded lossy compression reduces data movement costs while bounding reconstruction error, yet three barriers limit its adoption in mission-critical workflows: existing compressors cannot guarantee the accuracy of domain-specific quantities of interest (QoIs) derived from compressed data; compression-induced artifacts such as posterization, blocking, and interpolation banding erode user confidence in decompressed fields; and significant compressibility in the quantization index arrays of interpolation-based pipelines remains unexploited. This dissertation addresses all three barriers through four contributions, with the artifact barrier …


Dyno : Dynamic Neurosymbolic Orchestrator For Multi-Agent Systems, Ritvik Garimella, Chathurangi Shyalika, Renjith Prasad, Amit Sheth Jan 2026

Dyno : Dynamic Neurosymbolic Orchestrator For Multi-Agent Systems, Ritvik Garimella, Chathurangi Shyalika, Renjith Prasad, Amit Sheth

Publications

Large Language Model (LLM)-based multi-agent systems (LaMAS) represent an emerging paradigm for tackling complex, multi-step reasoning and decision-making problems. As these systems scale, orchestration, which is the ability to coordinate, manage, and evaluate the interactions among diverse agents, becomes central to their success. While recent orchestrators such as AgentFlow have demonstrated promise in managing communication and task delegation, they remain limited in their ability to understand task semantics, coordinate heterogeneous agent types (e.g., reactive vs. cognitive), and adaptively align outputs with human-defined goals. In this position paper, we introduce the DYNO (Dynamic Neurosymbolic Orchestrator), a system developed as part of …


Leveraging Large Language Models For Career Mobility Analysis: A Study Of Gender, Race, And Job Change Using Us Online Resume Profiles, Palakorn Achananuparp, Ye Xu, Yao Lu, Xavier Jayaraj Siddarth Ashok, Ee-Peng Lim Jan 2026

Leveraging Large Language Models For Career Mobility Analysis: A Study Of Gender, Race, And Job Change Using Us Online Resume Profiles, Palakorn Achananuparp, Ye Xu, Yao Lu, Xavier Jayaraj Siddarth Ashok, Ee-Peng Lim

Research Collection School Of Computing and Information Systems

We present a large-scale analysis of career mobility of college-educated U.S. workers using online resume profiles to investigate how gender, race, and job change options are associated with upward mobility. This study addresses key research questions of how the job changes affect their upward career mobility, and how the outcomes of upward career mobility differ by gender and race. We address data challenges – such as missing demographic attributes, missing wage data, and noisy occupation labels – through various data processing and Artificial Intelligence (AI) methods. In particular, we develop a large language models (LLMs) based occupation classification method known …


Beyond The Lace Index: Benchmarking Machine Learning Architectures And Explaining 30-Day Hospital Readmission Risk With Shap Analysis, Carl E. Hughes Iii Jan 2026

Beyond The Lace Index: Benchmarking Machine Learning Architectures And Explaining 30-Day Hospital Readmission Risk With Shap Analysis, Carl E. Hughes Iii

Williams Honors College, Honors Research Projects

Unplanned 30-day hospital readmission remains a fundamental challenge in US healthcare, associated with increased risk to patient recovery and representing an estimated $52.4 billion in annual expenses (Beauvais et al., 2022). While the rigorously validated LACE index serves as the clinical standard for readmission modeling, its linear structure and four explanatory variables lack the complexity to capture the high-dimensional and interactive nature of patient risk. This study utilizes an admission granularity level cohort of the MIMIC-IV database to develop and compare machine learning architectures against the baseline LACE index. Due to the imbalanced prevalence of readmission, the penalized logistic regression, …


Secure The Database: A Red Team, Blue Team Analysis Of Sql Injection, Andrew N. Miller Jan 2026

Secure The Database: A Red Team, Blue Team Analysis Of Sql Injection, Andrew N. Miller

Williams Honors College, Honors Research Projects

SQL injection (SQLi) attacks are a type of cyberattack that seeks to bypass website logins and gain entry to sensitive information. These pose a significant danger to organizations holding confidential user information. Personally Identifiable Information (PII) like physical addresses, emails, phone numbers, social security numbers are at risk of theft. Login credentials like usernames, passwords, and other sensitive information like financial details and social security numbers are also exposed through SQLi attacks. SQLi attacks harm the confidentiality, integrity, and availability of people’s identity. Additionally, data breaches that reach public battention harm the reputation and trust of organizations. SQLi attacks rank …