Gateway to Think Tanks
来源类型 | Working Paper |
规范类型 | 报告 |
DOI | 10.3386/w15851 |
来源ID | Working Paper 15851 |
Harmonizing and Combining Large Datasets - An Application to Firm-Level Patent and Accounting Data | |
Grid Thoma; Salvatore Torrisi; Alfonso Gambardella; Dominique Guellec; Bronwyn H. Hall; Dietmar Harhoff | |
发表日期 | 2010-03-25 |
出版年 | 2010 |
语种 | 英语 |
摘要 | This paper discusses methods for the harmonization and combination of large-scale patent and trademark datasets with each other and other sources of data. Dictionary- and rule-based approaches to the consolidation of applicant names in patent data are presented and shown to have both benefits and drawbacks in isolation. We combine the two methods and develop a set of rules and dictionaries to consolidate European, Patent Cooperation Treaty (PCT) and US patent data with firm accounting data. The resulting data encompass about 131,000 patent applicant names from 46 countries, covering 58.8 percent of EPO applications and 50.6 percent of PCT applications by business organizations during the time period from 1979 to 2008. For US data, the resulting dataset includes around 54,000 assignee names and 51.3 percent of US granted patents during approximately the same time period. |
主题 | Econometrics ; Data Collection ; Development and Growth ; Innovation and R& ; D |
URL | https://www.nber.org/papers/w15851 |
来源智库 | National Bureau of Economic Research (United States) |
引用统计 | |
资源类型 | 智库出版物 |
条目标识符 | http://119.78.100.153/handle/2XGU8XDN/573524 |
推荐引用方式 GB/T 7714 | Grid Thoma,Salvatore Torrisi,Alfonso Gambardella,et al. Harmonizing and Combining Large Datasets - An Application to Firm-Level Patent and Accounting Data. 2010. |
条目包含的文件 | ||||||
文件名称/大小 | 资源类型 | 版本类型 | 开放类型 | 使用许可 | ||
w15851.pdf(269KB) | 智库出版物 | 限制开放 | CC BY-NC-SA | 浏览 |
除非特别说明,本系统中所有内容都受版权保护,并保留所有权利。