First IEEE International Workshop on Source Code Analysis and Manipulation Finding Code on the World Wibe Web: A Preliminary Investigation Florence, Italy November 10-December 10 ISBN: 0-7695-1387-5
To find out what kind of design structures programmers really use, we need to examine a wide variety of programs. Unfortunately most program source code is proprietary and is unavailable for analysis. The World Wide Web (Web) potentially can provide a rich source of programs for study. The freely available code on the Web, if in sufficient quality and quantity, can provide a window into software design as it is practiced today. In a preliminary study of source code availability on the Web, we estimate that 4% of URLs contain object-oriented source code, and 9% of URLs contain executable code --- either binary or class files. This represents an enormous resource for program analysis. We can, with some risk of inaccuracy, conservatively project our sampling results to the entire Web. Our estimate is that the Web contains at least 3.4 million files containing either Java, C++, or Perl source code, 20.3 million files containing C source code, and 8.7 million files containing executable code.
Index Terms:
Design, source code analysis, World Wide Web estimation, code on the World Wide Web.
Citation:
James M. Bieman, Vanessa Murdock, "Finding Code on the World Wibe Web: A Preliminary Investigation," scam, pp.0075, First IEEE International Workshop on Source Code Analysis and Manipulation, 2001 Usage of this product signifies your acceptance of the Terms of Use. | |||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||||