Cross-Domain Evaluation of a Deep Learning-Based Type Inference System

Gruner, Bernd; Sonnekalb, Tim; Heinze, Thomas S.; Brust, Clemens-Alexander

Computer Science > Software Engineering

arXiv:2208.09189v1 (cs)

[Submitted on 19 Aug 2022 (this version), latest version 28 Jul 2023 (v4)]

Title:Cross-Domain Evaluation of a Deep Learning-Based Type Inference System

Authors:Bernd Gruner, Tim Sonnekalb, Thomas S. Heinze, Clemens-Alexander Brust

View PDF

Abstract:Optional type annotations allow for enriching dynamic programming languages with static typing features like better Integrated Development Environment (IDE) support, more precise program analysis, and early detection and prevention of type-related runtime errors. Machine learning-based type inference promises interesting results for automating this task. However, the practical usage of such systems depends on their ability to generalize across different domains, as they are often applied outside their training domain. In this work, we investigate the generalization ability of Type4Py as a representative for state-of-the-art deep learning-based type inference systems, by conducting extensive cross-domain experiments. Thereby, we address the following problems: dataset shifts, out-of-vocabulary words, unknown classes, and rare classes. To perform such experiments, we use the datasets ManyTypes4Py and CrossDomainTypes4Py. The latter we introduce in this paper. Our dataset has over 1,000,000 type annotations and enables cross-domain evaluation of type inference systems in different domains of software projects using data from the two domains web development and scientific calculation. Through our experiments, we detect shifts in the dataset and that it has a long-tailed distribution with many rare and unknown data types which decreases the performance of the deep learning-based type inference system drastically. In this context, we test unsupervised domain adaptation methods and fine-tuning to overcome the issues. Moreover, we investigate the impact of out-of-vocabulary words.

Subjects:	Software Engineering (cs.SE); Machine Learning (cs.LG); Programming Languages (cs.PL)
Cite as:	arXiv:2208.09189 [cs.SE]
	(or arXiv:2208.09189v1 [cs.SE] for this version)
	https://doi.org/10.48550/arXiv.2208.09189

Submission history

From: Bernd Gruner [view email]
[v1] Fri, 19 Aug 2022 07:28:31 UTC (136 KB)
[v2] Wed, 18 Jan 2023 07:03:18 UTC (159 KB)
[v3] Tue, 21 Mar 2023 15:06:37 UTC (138 KB)
[v4] Fri, 28 Jul 2023 08:27:57 UTC (161 KB)

Computer Science > Software Engineering

Title:Cross-Domain Evaluation of a Deep Learning-Based Type Inference System

Submission history

Access Paper:

References & Citations

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators

Computer Science > Software Engineering

Title:Cross-Domain Evaluation of a Deep Learning-Based Type Inference System

Submission history

Access Paper:

References & Citations

BibTeX formatted citation

Bookmark

Bibliographic and Citation Tools

Code, Data and Media Associated with this Article

Demos

Recommenders and Search Tools

arXivLabs: experimental projects with community collaborators