José Paulo Leal

Cookies Policy

The website need some cookies and similar means to function. If you permit us, we will use those means to collect data on your visits for aggregated statistics to improve our service. Find out More

Institution
Research
Research Domains
Artificial Intelligence

Bioengineering

Communications

Computer Science and Engineering
Photonics

Power and Energy Systems

Robotics

Systems Engineering and Management
RESEARCH CENTERS
Porto, Portugal

+351 222 094 000

info@inesctec.pt
Innovation
Innovation / Tec4

TEC4AGRO-FOOD

TEC4ENERGY

TEC4HEALTH

TEC4INDUSTRY

TEC4SEA

TECPARTNERSHIPS

Available Technologies
Porto, Portugal

+351 222 094 000

info@inesctec.pt
Laboratories
Research Laboratories

iilab
Communication
News

Events

Media

Newsletter
Porto, Portugal

+351 222 094 000

info@inesctec.pt
Work with us
Contacts

Home
People
José Paulo Leal

Read Full presentation

I was born in Portugal in 1964. I graduated in mathematics from the Faculty of Sciences of the University of Porto and earned a Ph.D. in Computer Science from the same institution. My current position is auxiliary professor at the Computer Science department of the Faculty of Sciences of the University of Porto. I am also affiliated with the Center for Research in Advanced Computing Systems (CRACS), an R&D unit of INESCTEC Research Laboratory, where I am an effective member.My main research interests are technology enhanced learning, web adaptability, and semantic web.

Read Full presentation

About

I was born in Portugal in 1964. I graduated in mathematics from the Faculty of Sciences of the University of Porto and earned a Ph.D. in Computer Science from the same institution.
My current position is auxiliary professor at the Computer Science department of the Faculty of Sciences of the University of Porto. I am also affiliated with the Center for Research in Advanced Computing Systems (CRACS), an R&D unit of INESCTEC Research Laboratory, where I am an effective member.
My main research interests are technology enhanced learning, web adaptability, and semantic web.

Interest
Topics

Details

Name
José Paulo Leal
Role
Senior Researcher
Since
01st January 2009

Nationality
Portugal
Centre
Advanced Computing Systems
Contacts
+351220402963
jose.p.leal@inesctec.pt

003

Publications

View all Publications

2025

Incremental Repair Feedback on Automated Assessment of Programming Assignments

Authors
Paiva, JC; Leal, JP; Figueira, A;

Publication
ELECTRONICS

Abstract
Automated assessment tools for programming assignments have become increasingly popular in computing education. These tools offer a cost-effective and highly available way to provide timely and consistent feedback to students. However, when evaluating a logically incorrect source code, there are some reasonable concerns about the formative gap in the feedback generated by such tools compared to that of human teaching assistants. A teaching assistant either pinpoints logical errors, describes how the program fails to perform the proposed task, or suggests possible ways to fix mistakes without revealing the correct code. On the other hand, automated assessment tools typically return a measure of the program's correctness, possibly backed by failing test cases and, only in a few cases, fixes to the program. In this paper, we introduce a tool, AsanasAssist, to generate formative feedback messages to students to repair functionality mistakes in the submitted source code based on the most similar algorithmic strategy solution. These suggestions are delivered with incremental levels of detail according to the student's needs, from identifying the block containing the error to displaying the correct source code. Furthermore, we evaluate how well the automatically generated messages provided by AsanasAssist match those provided by a human teaching assistant. The results demonstrate that the tool achieves feedback comparable to that of a human grader while being able to provide it just in time.

CloseRead Abstract

2025

PAP900: A dataset of semantic relationships between affective words in Portuguese

Authors
dos Santos, AF; Leal, JP; Alves, RA; Jacques, T;

Publication
DATA IN BRIEF

Abstract
The PAP900 dataset centers on the semantic relationship between affective words in Portuguese. It contains 900 word pairs, each annotated by at least 30 human raters for both semantic similarity and semantic relatedness. In addition to the semantic ratings, the dataset includes the word categorization used to build the word pairs and detailed sociodemographic information about annotators, enabling the analysis of the influence of personal factors on the perception of semantic relationships. Furthermore, this article describes in detail the dataset construction process, from word selection to agreement metrics. Data was collected from Portuguese university psychology students, who completed two rounds of questionnaires. In the first round annotators were asked to rate word pairs on either semantic similarity or relatedness. The second round switched the relation type for most annotators, with a small percentage being asked to repeat the same relation. The instructions given emphasized the differences between semantic relatedness and semantic similarity, and provided examples of expected ratings of both. There are few semantic relations datasets in Portuguese, and none focusing on affective words. PAP900 is distributed in distinct formats to be easy to use for both researchers just looking for the final averaged values and for researchers looking to take advantage of the individual ratings, the word categorization and the annotator data. This dataset is a valuable resource for researchers in computational linguistics, natural language processing, psychology, and cognitive science. (c) 2025TheAuthors.

CloseRead Abstract

2024

Comparing Semantic Graph Representations of Source Code: The Case of Automatic Feedback on Programming Assignments

Authors
Paiva, JC; Leal, JP; Figueira, A;

Publication
COMPUTER SCIENCE AND INFORMATION SYSTEMS

Abstract
Static source code analysis techniques are gaining relevance in automated assessment of programming assignments as they can provide less rigorous evaluation and more comprehensive and formative feedback. These techniques focus on source code aspects rather than requiring effective code execution. To this end, syntactic and semantic information encoded in textual data is typically represented internally as graphs, after parsing and other preprocessing stages. Static automated assessment techniques, therefore, draw inferences from intermediate representations to determine the correctness of a solution and derive feedback. Consequently, achieving the most effective semantic graph representation of source code for the specific task is critical, impacting both techniques' accuracy, outcome, and execution time. This paper aims to provide a thorough comparison of the most widespread semantic graph representations for the automated assessment of programming assignments, including usage examples, facets, and costs for each of these representations. A benchmark has been conducted to assess their cost using the Abstract Syntax Tree (AST) as a baseline. The results demonstrate that the Code Property Graph (CPG) is the most feature -rich representation, but also the largest and most space -consuming (about 33% more than AST).

CloseRead Abstract

2024

Clustering source code from automated assessment of programming assignments

Authors
Paiva, JC; Leal, JP; Figueira, A;

Publication
INTERNATIONAL JOURNAL OF DATA SCIENCE AND ANALYTICS

Abstract
Clustering of source code is a technique that can help improve feedback in automated program assessment. Grouping code submissions that contain similar mistakes can, for instance, facilitate the identification of students' difficulties to provide targeted feedback. Moreover, solutions with similar functionality but possibly different coding styles or progress levels can allow personalized feedback to students stuck at some point based on a more developed source code or even detect potential cases of plagiarism. However, existing clustering approaches for source code are mostly inadequate for automated feedback generation or assessment systems in programming education. They either give too much emphasis to syntactical program features, rely on expensive computations over pairs of programs, or require previously collected data. This paper introduces an online approach and implemented tool-AsanasCluster-to cluster source code submissions to programming assignments. The proposed approach relies on program attributes extracted from semantic graph representations of source code, including control and data flow features. The obtained feature vector values are fed into an incremental k-means model. Such a model aims to determine the closest cluster of solutions, as they enter the system, timely, considering clustering is an intermediate step for feedback generation in automated assessment. We have conducted a twofold evaluation of the tool to assess (1) its runtime performance and (2) its precision in separating different algorithmic strategies. To this end, we have applied our clustering approach on a public dataset of real submissions from undergraduate students to programming assignments, measuring the runtimes for the distinct tasks involved: building a model, identifying the closest cluster to a new observation, and recalculating partitions. As for the precision, we partition two groups of programs collected from GitHub. One group contains implementations of two searching algorithms, while the other has implementations of several sorting algorithms. AsanasCluster matches and, in some cases, improves the state-of-the-art clustering tools in terms of runtime performance and precision in identifying different algorithmic strategies. It does so without requiring the execution of the code. Moreover, it is able to start the clustering process from a dataset with only two submissions and continuously partition the observations as they enter the system.

CloseRead Abstract

2024

Authoring Programming Exercises for Automated Assessment Assisted by Generative AI

Authors
Bauer, Y; Leal, JP; Queirós, R;

Publication
5th International Computer Programming Education Conference, ICPEC 2024, June 27-28, 2024, Lisbon, Portugal

Abstract
Generative AI presents both challenges and opportunities for educators. This paper explores its potential for automating the creation of programming exercises designed for automated assessment. Traditionally, creating these exercises is a time-intensive and error-prone task that involves developing exercise statements, solutions, and test cases. This ongoing research analyzes the capabilities of the OpenAI GPT API to automatically create these components. An experiment using the OpenAI GPT API to automatically create 120 programming exercises produced interesting results, such as the difficulties encountered in generating valid JSON formats and creating matching test cases for solution code. Learning from this experiment, an enhanced feature was developed to assist teachers in creating programming exercises and was integrated into Agni, a virtual learning environment (VLE). Despite the challenges in generating entirely correct programming exercises, this approach shows potential for reducing the time required to create exercises, thus significantly aiding teachers. The evaluation of this approach, comparing the efficiency and usefulness of using the OpenAI GPT API or authoring the exercises oneself, is in progress. © Yannik Bauer, José Paulo Leal, and Ricardo Queirós;

CloseRead Abstract

Supervised
thesis

Supervised Thesis

View all Supervised Theses

2023

Reasoning on Semantic Representations of Source Code to Support Programming Education

Author
José Carlos Costa Paiva

Institution
UP-FCUP

2023

Semantic Measures in Large Semantic Graphs

Author
André Fernandes dos Santos

Institution
UP-FCUP

2023

Assessment of simple web applications in a code playground

Author
Luís Miguel Maia da Costa

Institution
UP-FCUP

2023

Narrative extraction from semantic graphs

Author
Daniil Lystopadskyi

Institution
UP-FCUP

2023

Improving Teacher's User Experience in a Virtual Learning Environment

Author
Yannik Bauer

Institution
UP-FCUP

View all Supervised Theses

About

Details

Name

Role

Since

Nationality

Centre

Contacts

FGPE

JuezLTI

FGPEPlus

Incremental Repair Feedback on Automated Assessment of Programming Assignments

PAP900: A dataset of semantic relationships between affective words in Portuguese

Comparing Semantic Graph Representations of Source Code: The Case of Automatic Feedback on Programming Assignments

Clustering source code from automated assessment of programming assignments

Authoring Programming Exercises for Automated Assessment Assisted by Generative AI

Reasoning on Semantic Representations of Source Code to Support Programming Education

Semantic Measures in Large Semantic Graphs

Assessment of simple web applications in a code playground

Narrative extraction from semantic graphs

Improving Teacher's User Experience in a Virtual Learning Environment