Open Access. Powered by Scholars. Published by Universities.®
Programming Languages and Compilers Commons™
Open Access. Powered by Scholars. Published by Universities.®
- Discipline
-
- Software Engineering (44)
- Artificial Intelligence and Robotics (20)
- Databases and Information Systems (18)
- Engineering (16)
- Computer Engineering (15)
-
- Cybersecurity (9)
- Education (9)
- Science and Mathematics Education (9)
- Secondary Education (9)
- Theory and Algorithms (9)
- Numerical Analysis and Scientific Computing (7)
- Data Science (6)
- Information Security (5)
- Other Computer Sciences (5)
- Graphics and Human Computer Interfaces (4)
- Mathematics (4)
- Statistics and Probability (4)
- Digital Communications and Networking (3)
- Electrical and Computer Engineering (3)
- Other Computer Engineering (3)
- Aerospace Engineering (2)
- Applied Mathematics (2)
- Applied Statistics (2)
- Computational Engineering (2)
- Data Storage Systems (2)
- Geography (2)
- Numerical Analysis and Computation (2)
- Institution
-
- Singapore Management University (38)
- City University of New York (CUNY) (14)
- Chapman University (12)
- California State University, San Bernardino (4)
- San Jose State University (3)
-
- California Polytechnic State University, San Luis Obispo (2)
- Georgia Southern University (2)
- Southern Adventist University (2)
- Ursinus College (2)
- Brigham Young University (1)
- Calvin University (1)
- Cleveland State University (1)
- DePaul University (1)
- New Jersey Institute of Technology (1)
- Old Dominion University (1)
- Purdue University (1)
- Southern Methodist University (1)
- Stephen F. Austin State University (1)
- Texas A&M International University (1)
- University of Alabama in Huntsville (1)
- University of Central Florida (1)
- University of Connecticut (1)
- University of Nebraska - Lincoln (1)
- University of North Florida (1)
- University of South Alabama (1)
- Keyword
-
- Software engineering (11)
- Computer science (7)
- Programming languages (7)
- Java (5)
- Introduction to computer programming (4)
-
- Python (4)
- Deep learning (3)
- Object oriented programming (3)
- Sentiment analysis (3)
- Computational thinking (2)
- Computer science education (2)
- Culturally responsive pedagogy (2)
- Culturally sustaining (2)
- Deep Learning (2)
- Elementary (2)
- Equity (2)
- Graph execution (2)
- Imperative programs (2)
- Large Language Models (2)
- Large language models (2)
- Machine Learning (2)
- Node.js (2)
- Performance (2)
- Pull request (2)
- R programming language (2)
- Refactoring (2)
- Reinforcement learning (2)
- Software (2)
- Statistical analysis (2)
- ACL2 (1)
- Publication
-
- Research Collection School Of Computing and Information Systems (35)
- Open Educational Resources (10)
- Journal of Computer Science Integration (9)
- Electronic Theses, Projects, and Dissertations (4)
- Publications and Research (3)
-
- Campus Research Month (2)
- Computational and Data Sciences (PhD) Dissertations (2)
- Dissertations and Theses Collection (Open Access) (2)
- Master's Projects (2)
- Master's Theses (2)
- Mathematics, Computer Science & Statistics Presentations (2)
- College of Computing and Digital Media Dissertations (1)
- College of Graduate Studies: Theses & Dissertations (1)
- Computer Science and Engineering Theses and Dissertations (1)
- Department of Electrical and Computer Engineering Faculty Publications (1)
- Dissertations (1)
- Dissertations, Theses, and Capstone Projects (1)
- Electronic Theses and Dissertations (1)
- Electronic Theses and Dissertations, 2020-2023 (1)
- Engineering Faculty Articles and Research (1)
- Faculty Publications (1)
- Honors College Theses (1)
- Honors Program: Senior Projects (Public) (1)
- Honors Scholar Theses (1)
- Honors Theses (1)
- Library Philosophy and Practice (e-journal) (1)
- Mechanical & Aerospace Engineering Faculty Publications (1)
- Open Educational Resources (OER) (1)
- Research Collection Yong Pung How School Of Law (1)
- Summer Research (1)
- Publication Type
- File Type
Articles 1 - 30 of 95
Full-Text Articles in Programming Languages and Compilers
A Critical Computing Curriculum Design Case: Exploring Tribal Sovereignty For Middle School Students, Kristin A. Searle, Aubrey Rogowski, Colby Tofel-Grehl, Mengying Jiang
A Critical Computing Curriculum Design Case: Exploring Tribal Sovereignty For Middle School Students, Kristin A. Searle, Aubrey Rogowski, Colby Tofel-Grehl, Mengying Jiang
Journal of Computer Science Integration
We report on our efforts to design an integrated computing curriculum for middle school students in Montana that is in line with the Kapor Center’s focus on culturally sustaining-revitalizing pedagogies. Montana provides a unique context for doing this work because a state constitutional mandate requires all K-12 students to learn about tribal histories and cultures through Indian Education For All (IEFA). IEFA centers around seven essential understandings about Indigenous peoples in Montana that are integrated across content areas. In addition, implementation of Montana’s CS standards began in the 2021–2022 school year. In the curricular design, we sought to bring together …
Employing An Abolitionist, Critical Race Pedagogy In Cs: Centering The Voices, Experiences And Technological Innovations Of Black Youth, Tiera Tanksley
Employing An Abolitionist, Critical Race Pedagogy In Cs: Centering The Voices, Experiences And Technological Innovations Of Black Youth, Tiera Tanksley
Journal of Computer Science Integration
This paper proposes a pedagogical extension of culturally responsive praxis called abolitionist, critical race pedagogy in CS. To showcase the power and potentiality of this pedagogy, this paper examines the experiences of 2 cohorts of Black high school students (n = 30) who participated in a critical race technology course that was taught during the dual pandemic of COVID-19 and anti-Black racism. The goal of this summer course was to employ an abolitionist, critical race pedagogy in CS to foster Black students’ ability to critically examine the ubiquity of anti-Black racism within the socio-technical architectures (e.g. code, data, algorithms and …
Culturally Responsive-Sustaining Computational Thinking: Enactment In Elementary Classrooms, Victoria Macann, Aman Yadav
Culturally Responsive-Sustaining Computational Thinking: Enactment In Elementary Classrooms, Victoria Macann, Aman Yadav
Journal of Computer Science Integration
Technology has increasingly permeated many aspects of everyday life and this evolution raises the need for individuals to understand how the digital world works and what opportunities and risks it brings (Nouri, Zhang, Mannila & Norén, 2019). For this to be an experience for everyone, we need to rethink how we integrate computational thinking (CT) and provide teachers with tools to center their students’ identities, experiences, and cultures in the classroom. In this paper, we present two case studies of primary (elementary) teachers from a full primary (student ages 5–13) semi-rural school in the North Island of New Zealand that …
Choosing A Sophisticated, Robust, And Secure Programming Language, J. Simon Richard
Choosing A Sophisticated, Robust, And Secure Programming Language, J. Simon Richard
The Downtown Review: An Interdisciplinary Journal Written and Peer-Reviewed by Mandel Honors College Students at Cleveland State University
This paper explores which programming languages maximize the quality and efficiency of software development projects requiring high levels of sophistication, security, and stability. Of the four languages discussed in this paper—C, C++, Java, and Rust—we conclude that Rust is the best for this application.
Μakka: Mutation Testing For Actor Concurrency In Akka Using Real-World Bugs, Mohsen Moradi Moghadam, Mehdi Bagherzadeh, Raffi Khatchadourian, Hamid Bagheri
Μakka: Mutation Testing For Actor Concurrency In Akka Using Real-World Bugs, Mohsen Moradi Moghadam, Mehdi Bagherzadeh, Raffi Khatchadourian, Hamid Bagheri
Publications and Research
Actor concurrency is becoming increasingly important in the real-world and mission-critical software. This requires these applications to be free from actor bugs, that occur in the real world, and have tests that are effective in finding these bugs. Mutation testing is a well-established technique that transforms an application to induce its likely bugs and evaluate the effectiveness of its tests in finding these bugs. Mutation testing is available for a broad spectrum of applications and their bugs, ranging from web to mobile to machine learning, and is used at scale in companies like Google and Facebook. However, there still is …
Ensuring Non-Repudiation In Long-Distance Constrained Devices, Ethan Blum
Ensuring Non-Repudiation In Long-Distance Constrained Devices, Ethan Blum
Honors Theses
Satellite communication is essential for the exploration and study of space. Satellites allow communications with many devices and systems residing in space and on the surface of celestial bodies from ground stations on Earth. However, with the rise of Ground Station as a Service (GsaaS), the ability to efficiently send action commands to distant satellites must ensure non-repudiation such that an attacker is unable to send malicious commands to distant satellites. Distant satellites are also constrained devices and rely on limited power, meaning security on these devices is minimal. Therefore, this study attempted to propose a novel algorithm to allow …
Llm-Adapters: An Adapter Family For Parameter-Efficient Fine-Tuning Of Large Language Models, Zhiqiang Hu, Lei Wang, Yihuai Lan, Wanyu Xu, Ee-Peng Lim, Lidong Bing, Xing Xu, Soujanya Poria, Roy Ka-Wei Lee
Llm-Adapters: An Adapter Family For Parameter-Efficient Fine-Tuning Of Large Language Models, Zhiqiang Hu, Lei Wang, Yihuai Lan, Wanyu Xu, Ee-Peng Lim, Lidong Bing, Xing Xu, Soujanya Poria, Roy Ka-Wei Lee
Research Collection School Of Computing and Information Systems
The success of large language models (LLMs), like GPT-4 and ChatGPT, has led to the development of numerous cost-effective and accessible alternatives that are created by finetuning open-access LLMs with task-specific data (e.g., ChatDoctor) or instruction data (e.g., Alpaca). Among the various fine-tuning methods, adapter-based parameter-efficient fine-tuning (PEFT) is undoubtedly one of the most attractive topics, as it only requires fine-tuning a few external parameters instead of the entire LLMs while achieving comparable or even better performance. To enable further research on PEFT methods of LLMs, this paper presents LLMAdapters, an easy-to-use framework that integrates various adapters into LLMs and …
Large Language Model Is Not A Good Few-Shot Information Extractor, But A Good Reranker For Hard Samples!, Yubo Ma, Yixin Cao, Yongchin Hong, Aixin Sun
Large Language Model Is Not A Good Few-Shot Information Extractor, But A Good Reranker For Hard Samples!, Yubo Ma, Yixin Cao, Yongchin Hong, Aixin Sun
Research Collection School Of Computing and Information Systems
Large Language Models (LLMs) have made remarkable strides in various tasks. However, whether they are competitive few-shot solvers for information extraction (IE) tasks and surpass fine-tuned small Pre-trained Language Models (SLMs) remains an open problem. This paper aims to provide a thorough answer to this problem, and moreover, to explore an approach towards effective and economical IE systems that combine the strengths of LLMs and SLMs. Through extensive experiments on nine datasets across four IE tasks, we show that LLMs are not effective few-shot information extractors in general, given their unsatisfactory performance in most settings and the high latency and …
Examining The Inter-Consistency Of Large Language Models: An In-Depth Analysis Via Debate, Kai Xiong, Xiao Ding, Yixin Cao, Ting Liu, Bing Qin
Examining The Inter-Consistency Of Large Language Models: An In-Depth Analysis Via Debate, Kai Xiong, Xiao Ding, Yixin Cao, Ting Liu, Bing Qin
Research Collection School Of Computing and Information Systems
Large Language Models (LLMs) have shown impressive capabilities in various applications, but they still face various inconsistency issues. Existing works primarily focus on the inconsistency issues within a single LLM, while we complementarily explore the inter-consistency among multiple LLMs for collaboration. To examine whether LLMs can collaborate effectively to achieve a consensus for a shared goal, we focus on commonsense reasoning, and introduce a formal debate framework (FORD) to conduct a three-stage debate among LLMs with real-world scenarios alignment: fair debate, mismatched debate, and roundtable debate. Through extensive experiments on various datasets, LLMs can effectively collaborate to reach a consensus …
Benchmarking Foundation Models With Language-Model-As-An-Examiner, Yushi Bai, Jiahao Ying, Yixin Cao, Xin Lv, Yuze He, Xiaozhi Wang, Jifan Yu, Kaisheng Zeng, Yijia Xiao, Haozhe Lyu, Jiayin Zhang, Juanzi Li, Lei Hou
Benchmarking Foundation Models With Language-Model-As-An-Examiner, Yushi Bai, Jiahao Ying, Yixin Cao, Xin Lv, Yuze He, Xiaozhi Wang, Jifan Yu, Kaisheng Zeng, Yijia Xiao, Haozhe Lyu, Jiayin Zhang, Juanzi Li, Lei Hou
Research Collection School Of Computing and Information Systems
Numerous benchmarks have been established to assess the performance of foundation models on open-ended question answering, which serves as a comprehensive test of a model’s ability to understand and generate language in a manner similar to humans. Most of these works focus on proposing new datasets, however, we see two main issues within previous benchmarking pipelines, namely testing leakage and evaluation automation. In this paper, we propose a novel benchmarking framework, Language-Model-as-an-Examiner, where the LM serves as a knowledgeable examiner that formulates questions based on its knowledge and evaluates responses in a reference-free manner. Our framework allows for effortless extensibility …
Molca: Molecular Graph-Language Modeling With Cross-Modal Projector And Uni-Modal Adapter, Zhiyuan Liu, Sihang Li, Yanchen Luo, Hao Fei, Yixin Cao, Kenji Kawaguchi, Xiang Wang, Tat-Seng Chua
Molca: Molecular Graph-Language Modeling With Cross-Modal Projector And Uni-Modal Adapter, Zhiyuan Liu, Sihang Li, Yanchen Luo, Hao Fei, Yixin Cao, Kenji Kawaguchi, Xiang Wang, Tat-Seng Chua
Research Collection School Of Computing and Information Systems
Language Models (LMs) have demonstrated impressive molecule understanding ability on various 1D text-related tasks. However, they inherently lack 2D graph perception — a critical ability of human professionals in comprehending molecules’ topological structures. To bridge this gap, we propose MolCA: Molecular Graph-Language Modeling with Cross-Modal Projector and Uni-Modal Adapter. MolCA enables an LM (i.e., Galactica) to understand both text- and graph-based molecular contents via the cross-modal projector. Specifically, the cross-modal projector is implemented as a QFormer to connect a graph encoder’s representation space and an LM’s text space. Further, MolCA employs a uni-modal adapter (i.e., LoRA) for the LM’s efficient …
A Comprehensive Evaluation Of Large Language Models On Legal Judgment Prediction, Ruihao Shui, Yixin Cao, Xiang Wang, Tat-Seng Chua
A Comprehensive Evaluation Of Large Language Models On Legal Judgment Prediction, Ruihao Shui, Yixin Cao, Xiang Wang, Tat-Seng Chua
Research Collection School Of Computing and Information Systems
Large language models (LLMs) have demonstrated great potential for domain-specific applications, such as the law domain. However, recent disputes over GPT-4’s law evaluation raise questions concerning their performance in real-world legal tasks. To systematically investigate their competency in the law, we design practical baseline solutions based on LLMs and test on the task of legal judgment prediction. In our solutions, LLMs can work alone to answer open questions or coordinate with an information retrieval (IR) system to learn from similar cases or solve simplified multi-choice questions. We show that similar cases and multi-choice options, namely label candidates, included in prompts …
Wsdms: Debunk Fake News Via Weakly Supervised Detection Of Misinforming Sentences With Contextualized Social Wisdom, Ruichao Yang, Wei Gao, Jing Ma, Hongzhan Lin, Zhiwei Yang
Wsdms: Debunk Fake News Via Weakly Supervised Detection Of Misinforming Sentences With Contextualized Social Wisdom, Ruichao Yang, Wei Gao, Jing Ma, Hongzhan Lin, Zhiwei Yang
Research Collection School Of Computing and Information Systems
In recent years, we witness the explosion of false and unconfirmed information (i.e., rumors) that went viral on social media and shocked the public. Rumors can trigger versatile, mostly controversial stance expressions among social media users. Rumor verification and stance detection are different yet relevant tasks. Fake news debunking primarily focuses on determining the truthfulness of news articles, which oversimplifies the issue as fake news often combines elements of both truth and falsehood. Thus, it becomes crucial to identify specific instances of misinformation within the articles. In this research, we investigate a novel task in the field of fake news …
Disentangling Transformer Language Models As Superposed Topic Models, Jia Peng Lim, Hady Wirawan Lauw
Disentangling Transformer Language Models As Superposed Topic Models, Jia Peng Lim, Hady Wirawan Lauw
Research Collection School Of Computing and Information Systems
Topic Modelling is an established research area where the quality of a given topic is measured using coherence metrics. Often, we infer topics from Neural Topic Models (NTM) by interpreting their decoder weights, consisting of top-activated words projected from individual neurons. Transformer-based Language Models (TLM) similarly consist of decoder weights. However, due to its hypothesised superposition properties, the final logits originating from the residual path are considered uninterpretable. Therefore, we posit that we can interpret TLM as superposed NTM by proposing a novel weight-based, model-agnostic and corpus-agnostic approach to search and disentangle decoder-only TLM, potentially mapping individual neurons to multiple …
Kape: Knn-Based Performance Testing For Deep Code Search, Yuejun Guo, Qiang Hu, Xiaofei Xie, Cordy Maxime, Mike Papadakis, Yves Le Traon
Kape: Knn-Based Performance Testing For Deep Code Search, Yuejun Guo, Qiang Hu, Xiaofei Xie, Cordy Maxime, Mike Papadakis, Yves Le Traon
Research Collection School Of Computing and Information Systems
Code search is a common yet important activity of software developers. An efficient code search model can largely facilitate the development process and improve the programming quality. Given the superb performance of learning the contextual representations, deep learning models, especially pre-trained language models, have been widely explored for the code search task. However, studies mainly focus on proposing new architectures for ever-better performance on designed test sets but ignore the performance on unseen test data where only natural language queries are available. The same problem in other domains, e.g., CV and NLP, is usually solved by test input selection that …
Attack Prompt Generation For Red Teaming And Defending Large Language Models, Boyi Deng, Wenjie Wang, Fuli Feng, Yang Deng, Qifan Wang, Xiangnan He
Attack Prompt Generation For Red Teaming And Defending Large Language Models, Boyi Deng, Wenjie Wang, Fuli Feng, Yang Deng, Qifan Wang, Xiangnan He
Research Collection School Of Computing and Information Systems
Large language models (LLMs) are susceptible to red teaming attacks, which can induce LLMs to generate harmful content. Previous research constructs attack prompts via manual or automatic methods, which have their own limitations on construction cost and quality. To address these issues, we propose an integrated approach that combines manual and automatic methods to economically generate high-quality attack prompts. Specifically, considering the impressive capabilities of newly emerged LLMs, we propose an attack framework to instruct LLMs to mimic human-generated prompts through in-context learning. Furthermore, we propose a defense framework that fine-tunes victim LLMs through iterative interactions with the attack framework …
Large Language Models As Source Planner For Personalized Knowledge-Grounded Dialogues, Hongru Wang, Minda Hu, Yang Deng, Rui Wang, Fei Mi, Weichao Wang, Yasheng Wang, Wai-Chung Kwan, Irwin King, Kam-Fai Wong
Large Language Models As Source Planner For Personalized Knowledge-Grounded Dialogues, Hongru Wang, Minda Hu, Yang Deng, Rui Wang, Fei Mi, Weichao Wang, Yasheng Wang, Wai-Chung Kwan, Irwin King, Kam-Fai Wong
Research Collection School Of Computing and Information Systems
Open-domain dialogue system usually requires different sources of knowledge to generate more informative and evidential responses. However, existing knowledge-grounded dialogue systems either focus on a single knowledge source or overlook the dependency between multiple sources of knowledge, which may result in generating inconsistent or even paradoxical responses. To incorporate multiple knowledge sources and dependencies between them, we propose SAFARI, a novel framework that leverages the exceptional capabilities of large language models (LLMs) in planning, understanding, and incorporating under both supervised and unsupervised settings. Specifically, SAFARI decouples the knowledge grounding into multiple sources and response generation, which allows easy extension to …
Supporting Software Engineers With Large Language Model-Based Automation, Ting Zhang
Supporting Software Engineers With Large Language Model-Based Automation, Ting Zhang
Dissertations and Theses Collection (Open Access)
In recent years, software engineering (SE) has witnessed significant growth, leading to the creation and sharing of an abundance of software artifacts such as source code, bug reports, and pull requests. Analyzing these artifacts is crucial for comprehending the sentiments of software developers and automating various SE tasks, ultimately leading to more human-centered automated SE and enhancing software development efficiency. However, the diverse and unstructured nature of software text poses a significant challenge to this analysis. In response, researchers have investigated a variety of approaches, including the utilization of natural language processing techniques. The advent of large language models (LLMs), …
Random Variable Spaces: Mathematical Properties And An Extension To Programming Computable Functions, Mohammed Kurd-Misto
Random Variable Spaces: Mathematical Properties And An Extension To Programming Computable Functions, Mohammed Kurd-Misto
Computational and Data Sciences (PhD) Dissertations
This dissertation aims to extend the boundaries of Programming Computable Functions (PCF) by introducing a novel collection of categories referred to as Random Variable Spaces. Originating as a generalization of Quasi-Borel Spaces, Random Variable Spaces are rigorously defined as categories where objects are sets paired with a collection of random variables from an underlying measurable space. These spaces offer a theoretical foundation for extending PCF to natively handle stochastic elements.
The dissertation is structured into seven chapters that provide a multi-disciplinary background, from PCF and Measure Theory to Category Theory with special attention to Monads and the Giry Monad. The …
Hypothyroid Disease Analysis By Using Machine Learning, Sanjana Seelam
Hypothyroid Disease Analysis By Using Machine Learning, Sanjana Seelam
Electronic Theses, Projects, and Dissertations
Thyroid illness frequently manifests as hypothyroidism. It is evident that people with hypothyroidism are primarily female. Because the majority of people are unaware of the illness, it is quickly becoming more serious. It is crucial to catch it early on so that medical professionals can treat it more effectively and prevent it from getting worse. Machine learning illness prediction is a challenging task. Disease prediction is aided greatly by machine learning. Once more, unique feature selection strategies have made the process of disease assumption and prediction easier. To properly monitor and cure this illness, accurate detection is essential. In order …
A Black-Box Attack On Code Models Via Representation Nearest Neighbor Search, Jie Zhang, Wei Ma, Qiang Hu, Shangqing Liu, Xiaofei Xie, Yves Le Traon, Yang Liu
A Black-Box Attack On Code Models Via Representation Nearest Neighbor Search, Jie Zhang, Wei Ma, Qiang Hu, Shangqing Liu, Xiaofei Xie, Yves Le Traon, Yang Liu
Research Collection School Of Computing and Information Systems
Existing methods for generating adversarial code examples face several challenges: limted availability of substitute variables, high verification costs for these substitutes, and the creation of adversarial samples with noticeable perturbations. To address these concerns, our proposed approach, RNNS, uses a search seed based on historical attacks to find potential adversarial substitutes. Rather than directly using the discrete substitutes, they are mapped to a continuous vector space using a pre-trained variable name encoder. Based on the vector representation, RNNS predicts and selects better substitutes for attacks. We evaluated the performance of RNNS across six coding tasks encompassing three programming languages: Java, …
Effective And Efficient Semantic Representations And Their Applications, Chong Cher Chia
Effective And Efficient Semantic Representations And Their Applications, Chong Cher Chia
Dissertations and Theses Collection (Open Access)
The proliferation of affordable and compact digital storage has also led to the creation of enormous databases of information, and much attention has been focused on the problem of processing unorganized and unstructured information into some form from which additional value can be extracted. Contemporary approaches to this problem virtually necessitate the use of complex models running on computational systems due to the sheer volume of information to be processed. While it is possible for the model to be fed the actual data as input, typically a representation of the data is used instead. These representations are therefore of interest, …
Towards Llm-Based Fact Verification On News Claims With A Hierarchical Step-By-Step Prompting Method, Xuan Zhang, Wei Gao
Towards Llm-Based Fact Verification On News Claims With A Hierarchical Step-By-Step Prompting Method, Xuan Zhang, Wei Gao
Research Collection School Of Computing and Information Systems
While large pre-trained language models (LLMs) have shown their impressive capabilities in various NLP tasks, they are still underexplored in the misinformation domain. In this paper, we examine LLMs with in-context learning (ICL) for news claim verification, and find that only with 4-shot demonstration examples, the performance of several prompting methods can be comparable with previous supervised models. To further boost performance, we introduce a Hierarchical Step-by-Step (HiSS) prompting method which directs LLMs to separate a claim into several subclaims and then verify each of them via multiple questionsanswering steps progressively. Experiment results on two public misinformation datasets show that …
Cgt-Gan: Clip-Guided Text Gan For Image Captioning, Jiarui Yu, Haoran Li, Yanbin Hao, Bin Zhu, Tong Xu, Xiangnan He
Cgt-Gan: Clip-Guided Text Gan For Image Captioning, Jiarui Yu, Haoran Li, Yanbin Hao, Bin Zhu, Tong Xu, Xiangnan He
Research Collection School Of Computing and Information Systems
The large-scale visual-language pre-trained model, Contrastive Language-Image Pre-training (CLIP), has significantly improved image captioning for scenarios without human-annotated image-caption pairs. Recent advanced CLIP-based image captioning without human annotations follows a text-only training paradigm, i.e., reconstructing text from shared embedding space. Nevertheless, these approaches are limited by the training/inference gap or huge storage requirements for text embeddings. Given that it is trivial to obtain images in the real world, we propose CLIP-guided text GAN (CgT-GAN), which incorporates images into the training process to enable the model to "see" real visual modality. Particularly, we use adversarial training to teach CgT-GAN to mimic …
Teacher Candidates’ Conceptions And Practices Of Computational Thinking For Equity, Heather F. Clark, Symone A. Gyles, Imelda Nava-Landeros
Teacher Candidates’ Conceptions And Practices Of Computational Thinking For Equity, Heather F. Clark, Symone A. Gyles, Imelda Nava-Landeros
Journal of Computer Science Integration
This study documents novice science and math teachers’ developing pedagogical approaches to integrating computational thinking (CT) and data into their courses to support educational equity and social justice. The 10 novice teacher candidates (TCs) studied were part of an urban teacher residency program that empowered them with an asset-based pedagogy we describe as “CT for Equity.” Drawing on coursework and interviews as data, we asked three questions: What are teachers’ conceptions of CT? What are their CT instructional practices? And how did their students respond to those practices? To explore conceptions of CT, we used Kafai et al.’s (2020) articulation …
Hallucination Detection: Robustly Discerning Reliable Answers In Large Language Models, Yuyuan Chen, Qiang Fu, Yichen Yuan, Zhihao Wen, Ge Fan, Dayiheng Liu, Dongmei Zhang, Zhixu Li, Yanghua Xiao
Hallucination Detection: Robustly Discerning Reliable Answers In Large Language Models, Yuyuan Chen, Qiang Fu, Yichen Yuan, Zhihao Wen, Ge Fan, Dayiheng Liu, Dongmei Zhang, Zhixu Li, Yanghua Xiao
Research Collection School Of Computing and Information Systems
Large language models (LLMs) have gained widespread adoption in various natural language processing tasks, including question answering and dialogue systems. However, a major drawback of LLMs is the issue of hallucination, where they generate unfaithful or inconsistent content that deviates from the input source, leading to severe consequences. In this paper, we propose a robust discriminator named RelD to effectively detect hallucination in LLMs' generated answers. RelD is trained on the constructed RelQA, a bilingual question-answering dialogue dataset along with answers generated by LLMs and a comprehensive set of metrics. Our experimental results demonstrate that the proposed RelD successfully detects …
Using The Typescript Compiler To Fix Erroneous Node.Js Snippets, Brittany Reid, Christoph Treude, Markus Wagner
Using The Typescript Compiler To Fix Erroneous Node.Js Snippets, Brittany Reid, Christoph Treude, Markus Wagner
Research Collection School Of Computing and Information Systems
Most online code snippets do not run. This means that developers looking to reuse code from online sources must manually find and fix errors. We present an approach for automatically evaluating and correcting errors in Node.js code snippets: Node Code Correction (NCC). NCC leverages the ability of the TypeScript compiler to generate errors and inform code corrections through the combination of TypeScript’s builtin codefixes, our own targeted fixes, and deletion of erroneous lines. Compared to existing approaches using linters, our findings suggest that NCC is capable of detecting a larger number of errors per snippet and more error types, and …
Evaluating A Large Language Model’S Ability To Solve Programming Exercises From An Introductory Bioinformatics Course, Stephen R. Piccolo, Paul Denny, Andrew Luxton-Reilly, Samuel H. Payne, Perry G. Ridge
Evaluating A Large Language Model’S Ability To Solve Programming Exercises From An Introductory Bioinformatics Course, Stephen R. Piccolo, Paul Denny, Andrew Luxton-Reilly, Samuel H. Payne, Perry G. Ridge
Faculty Publications
Life scientists frequently write computer code when doing research. Computer programming can aid researchers in performing tasks that are not supported by existing tools. Programming can also help researchers to implement analytical logic in a way that documents their steps and thus enables others to repeat those steps. Many educational resources are available to teach computer programming, but this skill remains challenging for many researchers and students to master. Artificial-intelligence tools like OpenAI’s ChatGPT are able to interpret human-language requests to generate code. Accordingly, we evaluated the extent to which this technology might be used to perform programming tasks described …
Asset-Based Approaches To Multilingual Students’ Computer Science Identity Development, Sharin Rawhiya Jacob, Mark Warschauer
Asset-Based Approaches To Multilingual Students’ Computer Science Identity Development, Sharin Rawhiya Jacob, Mark Warschauer
Journal of Computer Science Integration
While computer science identity development has been examined in several studies, there is much to learn about the development of multilingual students’ computer science (CS) identities. To develop strong CS identities, multilingual students must engage in culturally and linguistically sustaining curriculum, pedagogy, and interaction that draws from their rich and varied resources. This theoretical paper is grounded in a justice-centered, asset-based framework that views the traditions and practices in students’ cultures and communities as strong contributors to knowledge construction in STEM. We draw on multiple studies exploring multilingual student CS identity development to better understand how their personal, familial, community-based, …
Wind River Elementary Computer Science Collaborative: Connecting Computer Science And Indigenous Identities And Knowledges On The Wind River Reservation, Joseph P. Wilson, Kathryn M. Rich, Jared O'Leary, Veronica Miller
Wind River Elementary Computer Science Collaborative: Connecting Computer Science And Indigenous Identities And Knowledges On The Wind River Reservation, Joseph P. Wilson, Kathryn M. Rich, Jared O'Leary, Veronica Miller
Journal of Computer Science Integration
Three Northern Arapaho and Eastern Shoshone–serving districts formed a researcher–practitioner partnership with the Wyoming Department of Education, the American Institutes for Research®, and BootUp Professional Development to advance the computer science (CS) education of their elementary students in ways that strengthen their Indigenous identities and knowledges. In this paper, we share experiences from 2019 to 2022 with our curriculum development, professional development (PD), and classroom implementation. The researcher–practitioner partnership developed student and teacher materials to support elementary CS lessons aligned to Wyoming’s CS standards and “Indian Education for All” social studies standards. Indigenous community members served as experts to codesign …