EASE 2026
Tue 9 - Fri 12 June 2026 Glasgow, United Kingdom
Wed 10 Jun 2026 11:30 - 11:45 at JMS 745 - Maintenance and Evolution 2 Chair(s): Andrea Capiluppi

AI efficiency has recently taken the spotlight in both academy and industry due to massive model scales, high energy demands, and environmental costs. While reporting Floating Point Operations (FLOPs) is a traditional approach for assessing computational costs, the relationship between FLOPs and execution time is not straightforward, as layers with the same number of FLOPs may not have the same execution time because some operations are more easily parallelized than others. This paper sets out to replicate the original experiments from a study that proposed the α-FLOPs estimation formula to verify whether the results remain applicable on newer, more powerful hardware.

During the replication process, we identify limitations in the replication materials provided by the original study, including a lack of specific dependency details and transparency regarding regression data. Our results validate the thesis that raw FLOPs alone are not an appropriate metric for execution time, as spatial dimensions remain more easily parallelized than kernel dimensions. However, fine-grained measurements reveal that the relationship is much less straightforward than previously shown, with newer hardware exhibiting instabilities and discontinuities in execution time, including jumps and oscillations, that the α-FLOPs formula generally underestimates. Ultimately, this work validates the empirical findings from the original study but shows negative results when applying the α-FLOPs estimation. We also highlight the critical need for complete and accurate replication packages for research on hardware-dependent efficiency assessment and provide a complete replication package for our implementation to facilitate further study.

Wed 10 Jun

Displayed time zone: London change

11:00 - 12:30
Maintenance and Evolution 2Journal First / Research Papers / Industry Papers / Reproducibility and Negative Results at JMS 745
Chair(s): Andrea Capiluppi University of Groningen
11:00
15m
Talk
Toward a Prioritization Approach for Third-Party Software Library Updates
Journal First
Abdalrahman Aburakhia King Fahd University of Petroleum and Minerals, Mohammad Alshayeb King Fahd University of Petroleum & Minerals
11:15
15m
Talk
Many hands make light work: An LLM-based multi-agent system for detecting malicious PyPI packages
Journal First
Muhammad Umar Zeshan University of L’Aquila, Italy, Motunrayo Osatohanmen Ibiyo University of L'Aquila, Claudio Di Sipio University of L'Aquila, Phuong T. Nguyen University of L’Aquila, Davide Di Ruscio University of L'Aquila
11:30
15m
Paper
FLOPs vs Real Work: The Importance of Replication in AI Efficiency Assessment
Reproducibility and Negative Results
Enrique Barba Roque Delft University of Technology, Luís Cruz TU Delft
Pre-print
11:45
15m
Talk
An Empirical Evaluation of Code Smell Detection in Angular Applications
Research Papers
Maykon Nunes Universidade Federal do Ceará, Ivan Machado Federal University of Bahia (UFBA), Carla Ilane Bezerra Federal University of Ceara, Emanuel Coutinho Federal University of Ceará
Pre-print
12:00
10m
Talk
Empirical Evaluation of Lift-and-Shift for Decoupling Drivers in Industrial Legacy Software: Lessons from a CGI Case StudyBest Paper Award - Industry Track
Industry Papers
Bauke van den Berg CGI Netherlands, Andrea Capiluppi University of Groningen
12:10
15m
Talk
Unsafe and Unused? A History of Utility Code in Mature Open Source Projects
Research Papers
Brandon Keller Rochester Institute of Technology, Kaitlin Yandik Rochester Institute of Technology, Angela Ngo Rochester Institute of Technology, Andy Meneely Rochester Institute of Technology
Pre-print