Duplications are registered in CPD by storing a Mark. This Mark contains a TokenEntry for where the duplication starts, as well as a line count and the source code that is duplicated. This change adds a beginColumn and endColumn field to each TokenEntry. These are optional fields that can be left empty. Storing these allows us to pinpoint the column position of each token. In addition, an additional TokenEntry is added to the Mark to indicate where the duplication ends. This can then be used to determine the location of the entire duplication; it starts where the first token starts and ends where the last token ends.
PMD
About
PMD is a source code analyzer. It finds common programming flaws like unused variables, empty catch blocks, unnecessary object creation, and so forth. It supports Java, JavaScript, Salesforce.com Apex and Visualforce, Modelica, PLSQL, Apache Velocity, XML, XSL, Scala.
Additionally it includes CPD, the copy-paste-detector. CPD finds duplicated code in C/C++, C#, Dart, Fortran, Go, Groovy, Java, JavaScript, JSP, Kotlin, Lua, Matlab, Modelica, Objective-C, Perl, PHP, PLSQL, Python, Ruby, Salesforce.com Apex, Scala, Swift and Visualforce.
Source and Documentation
Our latest source of PMD can be found on GitHub. Fork us!
The rule designer is developed over at pmd/pmd-designer. Please see its README for developer documentation.
News and Website
More information can be found on our Website and on SourceForge.