Building Patent Document Similarity by Considering the Dependency of Patent Classification Codes
Date Issued
2010
Date
2010
Author(s)
Pu, Kai-Yu
Abstract
A methodology using patent classification codes as feature to calculate the patent document similarity which considered the feature dependency is presented in this study. The patent classification codes which is assigned by patent examiner as selected feature. Because each patent has several patent classification codes means they have feature dependency, and patent classification codes have their class schedule which is a hierarchical schedule. Two dependency factors will be introduced when we calculating the similarities of patent documents: One is the co-occurrence of classification codes which is the patent classification codes pair co-assigned in the same patent called co-occurrence; and another one is the hierarchical relation in patent classification codes schedule where each classification codes belong to is called hierarchical relation. However, some patent classification codes are infrequent so the co-occurrences are not all validity, so this study defines a threshold to filter the invalidity co-occurrences. The validity co-occurrences and hierarchical relations to show the dependencies of patent classification codes which belong to the selected patent data set are visualize by UCINET. Illustration of the systematic methodology is successfully demonstrated by using United State patent classification which is considered the hierarchical relation and co-occurrence as selected feature in some case studies.
Subjects
Patent documents similarity
current USPC
Feature dependency
Hierarchical relation
Co-occurrence
Type
thesis
File(s)![Thumbnail Image]()
Loading...
Name
ntu-99-R97522605-1.pdf
Size
23.53 KB
Format
Adobe PDF
Checksum
(MD5):241d0ec54c73e07c93d97de04b073fe1
