First, Aroma indexes the code corpus as a sparse matrix by parsing each method and creating its parse tree. Then it extracts a set of structural features from the parse tree of each method. Finally, it creates a sparse vector for each method according to its features. The feature vectors for all method bodies become the indexing matrix, which is used for the search retrieval.

To help personalize content, tailor and measure ads and provide a safer experience, we use cookies. By clicking or navigating the site, you agree to allow our collection of information on and off Facebook through cookies. Learn more, including about available controls: Cookie Policy