Lorem Ipsum
Information retrival

Ease of building Patent Search Platform now with Google Patents Public Datasets

Now patent data got cheaper with the launch of new Patents Public Datasets by Google based on the company’s owned enterprise data warehouse BigQuery, which gathers openly available, associated database tables for exact investigation of the worldwide patent framework.

Enterprises often keep up accumulations of private information about patents, for example internal tagging system that compares to particular product offerings, and they need to associate that data with other patent datasets to create reports and examine speculation zones. Now organizations can consolidate their private information with open and paid datasets to ask "what are my active patents and pending patent applications?", "which of my patents in what technological areas are lapsing soon?" or "what are the best organizations that refer to the patents I've labeled with [widget #57]?".

Patent data availability is basis for analyzing new patents, illuminating open approach choices, overseeing corporate interest in protected innovation, and advancing future logical advancement. The developing number of accessible patent information sources implies specialists frequently invest more energy downloading, parsing, stacking, matching up and overseeing nearby databases than leading examination. With these new datasets, specialists and organizations can get to the information they require from different sources in a single place, in this way investing more energy in examination than data preparation.

3
Table IDpatents-public-data:patents.publications
Table Size780 GB
Number of Rows90,740,599
Creation TimeOct 27, 2017, 6:22:47 PM
Last ModifiedOct 27, 2017, 6:22:47 PM
Data LocationUS
LabelsNoneEdit

 

Table Details: publications

Refresh Query Table Copy Table Export Table Delete Table

publication_numberSTRINGNULLABLEPatent publication number (DOCDB compatible), eg: 'US-7650331-B1'
application_numberSTRINGNULLABLEPatent application number (DOCDB compatible), eg: 'US-87124404-A'. This may not always be set.
country_codeSTRINGNULLABLECountry code, eg: 'US', 'EP', etc
kind_codeSTRINGNULLABLEKind code, indicating application, grant, search report, correction, etc. These are different for each country.
application_kindSTRINGNULLABLEHigh-level kind of the application: A=patent; U=utility; P=provision; W= PCT; F=design; T=translation.
application_number_formattedSTRINGNULLABLEApplication number, formatted to the patent office format where possible.
pct_numberSTRINGNULLABLEPCT number for this application if it was part of a PCT filing, eg: 'PCT/EP2008/062623'.
family_idSTRINGNULLABLEFamily ID (simple family). Grouping on family ID will return all publications associated with a simple patent family (all publications share the same priority claims).
title_localizedRECORDREPEATEDThe publication titles in different languages
title_localized.textSTRINGNULLABLELocalized text
title_localized.languageSTRINGNULLABLETwo-letter language code for this text
abstract_localizedRECORDREPEATEDThe publication abstracts in different languages
abstract_localized.textSTRINGNULLABLELocalized text
abstract_localized.languageSTRINGNULLABLETwo-letter language code for this text
claims_localizedRECORDREPEATEDFor US publications only, the claims
claims_localized.textSTRINGNULLABLELocalized text
claims_localized.languageSTRINGNULLABLETwo-letter language code for this text
description_localizedRECORDREPEATEDFor US publications only, the description, limited to the first 9 megabytes
description_localized.textSTRINGNULLABLELocalized text
description_localized.languageSTRINGNULLABLETwo-letter language code for this text
publication_dateINTEGERNULLABLEThe publication date.
filing_dateINTEGERNULLABLEThe filing date.
grant_dateINTEGERNULLABLEThe grant date, or 0 if not granted.
priority_dateINTEGERNULLABLEThe earliest priority date from the priority claims, or the filing date.
priority_claimRECORDREPEATEDThe application numbers of the priority claims of this publication.
priority_claim.publication_numberSTRINGNULLABLESame as [publication_number]
priority_claim.application_numberSTRINGNULLABLESame as [application_number]
priority_claim.npl_textSTRINGNULLABLEFree-text citation (non-patent literature, etc).
priority_claim.typeSTRINGNULLABLEThe type of reference (see parent field for values).
priority_claim.categorySTRINGNULLABLEThe category of reference (see parent field for values).
priority_claim.filing_dateINTEGERNULLABLEThe filing date.
inventorSTRINGREPEATEDThe inventors.
inventor_harmonizedRECORDREPEATEDThe harmonized inventors and their countries.
inventor_harmonized.nameSTRINGNULLABLEName
inventor_harmonized.country_codeSTRINGNULLABLEThe two-letter country code
assigneeSTRINGREPEATEDThe assignees/applicants.
assignee_harmonizedRECORDREPEATEDThe harmonized assignees and their countries.
assignee_harmonized.nameSTRINGNULLABLEName
assignee_harmonized.country_codeSTRINGNULLABLEThe two-letter country code
examinerRECORDREPEATEDThe examiner of this publication and their countries.
examiner.nameSTRINGNULLABLEName
examiner.departmentSTRINGNULLABLEThe examiner's department
examiner.levelSTRINGNULLABLEThe examiner's level
uspcRECORDREPEATEDThe US Patent Classification (USPC) codes.
uspc.codeSTRINGNULLABLEClassification code
uspc.inventiveBOOLEANNULLABLEIs this classification inventive/main?
uspc.firstBOOLEANNULLABLEIs this classification the first/primary?
uspc.treeSTRINGREPEATEDThe full classification tree from the root to this code
ipcRECORDREPEATEDThe International Patent Classification (IPC) codes.
ipc.codeSTRINGNULLABLEClassification code
ipc.inventiveBOOLEANNULLABLEIs this classification inventive/main?
ipc.firstBOOLEANNULLABLEIs this classification the first/primary?
ipc.treeSTRINGREPEATEDThe full classification tree from the root to this code
cpcRECORDREPEATEDThe Cooperative Patent Classification (CPC) codes.
cpc.codeSTRINGNULLABLEClassification code
cpc.inventiveBOOLEANNULLABLEIs this classification inventive/main?
cpc.firstBOOLEANNULLABLEIs this classification the first/primary?
cpc.treeSTRINGREPEATEDThe full classification tree from the root to this code
fiRECORDREPEATEDThe FI classification codes.
fi.codeSTRINGNULLABLEClassification code
fi.inventiveBOOLEANNULLABLEIs this classification inventive/main?
fi.firstBOOLEANNULLABLEIs this classification the first/primary?
fi.treeSTRINGREPEATEDThe full classification tree from the root to this code
ftermRECORDREPEATEDThe F-term classification codes.
fterm.codeSTRINGNULLABLEClassification code
fterm.inventiveBOOLEANNULLABLEIs this classification inventive/main?
fterm.firstBOOLEANNULLABLEIs this classification the first/primary?
fterm.treeSTRINGREPEATEDThe full classification tree from the root to this code
citationRECORDREPEATEDThe citations of this publication. Category is one of {CH2 = Chapter 2; SUP = Supplementary search report ; ISR = International search report ; SEA = Search report; APP = Applicant; EXA = Examiner; OPP = Opposition; 115 = article 115; PRS = Pre-grant pre-search; APL = Appealed; FOP = Filed opposition}, Type is one of {A = technological background; D = document cited in application; E = earlier patent document; 1 = document cited for other reasons; O = Non-written disclosure; P = Intermediate document; T = theory or principle; X = relevant if taken alone; Y = relevant if combined with other documents}
citation.publication_numberSTRINGNULLABLESame as [publication_number]
citation.application_numberSTRINGNULLABLESame as [application_number]
citation.npl_textSTRINGNULLABLEFree-text citation (non-patent literature, etc).
citation.typeSTRINGNULLABLEThe type of reference (see parent field for values).
citation.categorySTRINGNULLABLEThe category of reference (see parent field for values).
citation.filing_dateINTEGERNULLABLEThe filing date.
entity_statusSTRINGNULLABLEThe USPTO entity status (large, small).
art_unitSTRINGNULLABLEThe USPTO art unit performing the examination (2159, etc).

These datasets incorporates Google Patents Public Data table containing worldwide bibliographic information on more than 90 million patent publications from 17 countries and US full text, provided by IFI CLAIMS Patent Services. Along with this Google is also providing a Google Patents Research Data table containing English machine translations for all titles and abstracts from Google Translate, similarity vectors, extracted top terms, and more. Common research datasets from patents, chemistry, and litigation have also been uploaded. Users can get to data gathered by different analysts and patent information suppliers in a similar database, and blend them with private information to create reports or research queries with the full opportunity of SQL, without setting up their very own database.

1

Commercial Data providers are also making their patent data available for purchase in BigQuery, starting with IFI CLAIMS Patent Data Enrichments including legal status information and standardized assignee names. Accessing these datasets through BigQuery gives users an up-to-date database managed by data providers, so users get the flexibility of a database without the engineering cost of maintaining one. Getting to these datasets through BigQuery surrenders clients a to-date database oversaw by information suppliers, so clients get the adaptability of a database without the designing expense of looking after one.

2

Several third party tools such as Tableau and Looker that can access BigQuery can also be employed which provide much easier interface for accessing database than SQL. For corporate having classified data that cannot leave their network, some of these tools can be used to fetch from the BigQuery and process that in conjunction to sensitive data.

BigQuery for Data Providers

For data providers, BigQuery is an extraordinary approach to pitch information in a right away helpful configuration to clients. The commonplace choices for information dissemination are either in bulk format through CSV/XML downloads, or through a web interface, yet both have drawbacks. Bulk format permit adaptability to the detriment of the client programming and keeping up their own databases, while web interfaces are anything but difficult to get to, however can't undoubtedly be reached out with new paid or private wellsprings of information, and have a settled arrangement of conceivable approaches to question and show the information. Presently clients can get a similar adaptability of a database with the simple access of a web interface to associate private information and show it in dashboards and other visualization tools.

Source : https://cloud.google.com/blog/big-data/2017/10/google-patents-public-datasets-connecting-public-paid-and-private-patent-data