Movatterモバイル変換

Data Commons

From Wikipedia, the free encyclopedia

Knowledge repository integrating open datasets

Data Commons

Results for a query in Data Commons
Founder	Ramanathan V. Guha
Key people	Prem Ramaswami (Head of Data Commons)
Parent	Google
URL	datacommons.org
Launched	May 2018; 7 years ago (2018-05)

Data Commons is an open-source platform^[1] created byGoogle^[2] that provides anopen knowledge graph, combining economic, scientific and other public datasets into a unified view.^[3]Ramanathan V. Guha, a creator of web standards includingRDF,^[4]RSS, andSchema.org,^[5] founded the project,^[6] which is now led by Prem Ramaswami.^[7]

The Data Commons website was launched in May 2018 with an initial dataset consisting offact-checking data published inSchema.org "ClaimReview" format by several fact checkers from theInternational Fact-Checking Network.^[8]^[9] Google has worked with partners such as theUnited Nations (UN) to populate the repository,^[2] which also includes data from theUnited States Census, theWorld Bank, theUS Bureau of Labor Statistics,^[10]Wikipedia, theNational Oceanic and Atmospheric Administration and theFederal Bureau of Investigation.^[11]

The service expanded during 2019 to include anRDF-style knowledge graph populated from a number of largely statistical open datasets. The service was announced to a wider audience in 2019.^[12] In 2020 the service improved its coverage of non-US datasets, while also increasing its coverage ofbioinformatics andcoronavirus.^[13] In 2023, the service relaunched with a natural-language front end powered by alarge language model.^[2] It also launched as the back end to the UN data portal withSustainable Development Goals data.^[14]

Features

[edit]

Data Commons places more emphasis on statistical data than is common forlinked data andknowledge graph initiatives. It includes geographical, demographic, weather and real estate data alongside other categories,^[3] describing states, Congressional districts, and cities in the United States as well as biological specimens, power plants, and elements of thehuman genome via theEncyclopedia of DNA Elements (ENCODE) project.^[11] It represents data assemantic triples each of which can have its own provenance.^[3] It centers on the entity-oriented integration of statistical observations from a variety of public datasets. Although it supports a subset of the W3CSPARQL query language,^[15] itsAPIs^[16] also include tools — such as aPandas dataframe interface — oriented towards data science, statistics and data visualization.

Data Commons is integrative, meaning that it does not provide a hosting platform for different datasets, but rather attempts to consolidate much of the information provided by the datasets into a single data graph.

Technology

[edit]

Data Commons is built on agraph data-model. The graph can be accessed through a browser interface and several APIs,^[3]^[11] and is expanded through loading data (typically CSV andMCF-based templates).^[17] The graph can be accessed by natural language queries inGoogle Search.^[18] The data vocabulary used to define the datacommons.org graph is based uponSchema.org.^[3] In particular the Schema.org terms StatisticalPopulation^[19] and Observation^[20] were proposed to Schema.org to support datacommons-like use cases.^[21]

Software from the project is available onGitHub underApache 2 license.^[22]

References

[edit]

^"Custom Data Commons".Docs - Data Commons. Retrieved16 July 2024.
^^a ^b ^c"Data Commons is using AI to make the world's public data more accessible and helpful".Google. 13 September 2023. Retrieved16 July 2024.
^^a ^b ^c ^d ^eFensel, Dieter; Şimşek, Umutcan; Angele, Kevin; Huaman, Elwin; Kärle, Elias; Panasiuk, Oleksandra; Toma, Ioan; Umbrich, Jürgen; Wahler, Alexander (2020),"Introduction: What Is a Knowledge Graph?",Knowledge Graphs, Cham: Springer International Publishing, pp. 1–10,doi:10.1007/978-3-030-37439-6_1,ISBN 978-3-030-37438-9,S2CID 213620389, retrieved2020-10-16
^Guns, Raf (2013). "Tracing the origins of the semantic web".Journal of the American Society for Information Science and Technology.64 (10):2173–2181.doi:10.1002/asi.22907.hdl:10067/1111170151162165141.
^Funke, Daniel (7 December 2017)."This website helps you find related fact checks - and it was built by a 17-year-old".Poynter. Retrieved16 July 2024.
^Guha, Ramanathan V. (15 October 2020)."Data Commons, now accessible on Google Search".docs.datacommons.org. Retrieved2020-10-16.
^O'Donnell, James (12 September 2024)."Google's new tool lets large language models fact-check their responses".MIT Technology Review. Retrieved17 September 2024.
^"Fact Checks".datacommons.org. 29 March 2019. Retrieved14 October 2020.
^Jiang, Shan; Baumgartner, Simon; Ittycheriah, Abe; Yu, Cong (2020-04-20)."Factoring Fact-Checks: Structured Information Extraction from Fact-Checking Articles".Proceedings of the Web Conference 2020. WWW '20. Taipei Taiwan: ACM. pp. 1592–1603.doi:10.1145/3366423.3380231.ISBN 978-1-4503-7023-3.S2CID 215882520.
^Raghavan, Prabhakar (2020-10-15)."How AI is powering a more helpful Google".Google. Retrieved2020-10-16.
^^a ^b ^cSheth, Amit; Padhee, Swati; Gyrard, Amelie; Sheth, Amit (2019-07-01). "Knowledge Graphs and Knowledge Networks: The Story in Brief".IEEE Internet Computing.23 (4):67–75.arXiv:2003.03623.Bibcode:2019IIC....23d..67S.doi:10.1109/MIC.2019.2928449.ISSN 1089-7801.S2CID 204820800.
^Luong, Daphne; Chou, Charina (5 March 2019)."Doing our part to share open data responsibly".The Keyword. Retrieved14 October 2020.
^Ramasubramanian, Sowmya (21 September 2020)."Google's open source data to study impact of COVID-19".The Hindu. Retrieved14 October 2020.
^Manyika, James (19 September 2023)."Using data and AI to track progress toward the UN Global Goals".Google. Retrieved22 July 2024.
^"Query the Data Commons Knowledge Graph using SPARQL".datacommons.org. Retrieved14 October 2020.
^"Overview".datacommons.org. Retrieved14 October 2020.
^"Contributing to Data Commons – Adding datasets".datacommons.org. Data Commons. Archived fromthe original on 2020-09-19. Retrieved2020-10-14.
^Guha, Ramanathan V. (15 October 2020)."Data Commons, now accessible on Google Search".docs.datacommons.org. Retrieved2020-10-16.
^"StatisticalPopulation type at Schema.org".schema.org. Retrieved14 October 2020.
^"Observation type at Schema.org".schema.org. Retrieved14 October 2020.
^"Proposal for representing Aggregate Statistical Data".GitHub – Schema.org repository. 25 June 2019. Retrieved14 October 2020.
^"datacommons.org GitHub".GitHub.

External links

[edit]

Google

a subsidiary ofAlphabet

Company

Divisions

Subsidiaries

Active

Defunct

Programs

Events

Infrastructure

People

Current	Krishna Bharat Vint Cerf Jeff Dean John Doerr Sanjay Ghemawat Al Gore John L. Hennessy Urs Hölzle Salar Kamangar Ray Kurzweil Ann Mather Alan Mulally Rick Osterloh Sundar Pichai (CEO) Ruth Porat (CFO) Rajen Sheth Hal Varian Neal Mohan
Former	Andy Bechtolsheim Sergey Brin (co-founder) David Cheriton Matt Cutts David Drummond Alan Eustace Timnit Gebru Omid Kordestani Paul Otellini Larry Page (co-founder) Patrick Pichette Eric Schmidt Ram Shriram Amit Singhal Shirley M. Tilghman Rachel Whetstone Susan Wojcicki

Criticism

General	Censorship DeGoogle FairSearch "Google's Ideological Echo Chamber" No Tech for Apartheid Privacy concerns Street View YouTube Trade unions Alphabet Workers Union YouTube copyright issues
Incidents	Backdoor advertisement controversy Blocking of YouTube videos in Germany Data breach Elsagate Fantastic Adventures scandal Kohistan video case Reactions toInnocence of Muslims San Francisco tech bus protests Services outages Slovenian government incident Walkouts YouTube headquarters shooting

Other

Development

Software

A–C	Accelerated Linear Algebra AMP Actions on Google ALTS American Fuzzy Lop Android Cloud to Device Messaging Android Debug Bridge Android NDK Android Runtime Android SDK Android Studio Angular AngularJS Apache Beam APIs App Engine App Inventor App Maker App Runtime for Chrome AppJet Apps Script AppSheet ARCore Base Bazel BeyondCorp Bigtable BigQuery Bionic Blockly Borg Caja Cameyo Chart API Charts Chrome Frame Chromium Blink Closure Tools Cloud Connect Cloud Dataflow Cloud Datastore Cloud Messaging Cloud Shell Cloud Storage Code Search Compute Engine Cpplint
D–N	Dalvik Data Protocol Dialogflow Exposure Notification Fast Pair Fastboot Federated Learning of Cohorts File System Firebase Firebase Studio Firebase Cloud Messaging FlatBuffers Flutter Freebase Gadgets Ganeti Gears Gerrit GLOP gRPC Gson Guava Guetzli Guice gVisor GYP JAX Jetpack Compose Keyhole Markup Language Kubernetes Kythe LevelDB Lighthouse Looker Studio lmctfy MapReduce Mashup Editor Matter Mobile Services Namebench Native Client Neatx Neural Machine Translation Nomulus
O–Z	Open Location Code OpenRefine OpenSocial Optimize OR-Tools Pack PageSpeed Piper Plugin for Eclipse Polymer Programmable Search Engine Project Shield Public DNS reCAPTCHA RenderScript SafetyNet SageTV Schema.org Search Console Shell Sitemaps Skia Graphics Engine Spanner Sputnik Stackdriver Swiffy Tango TensorFlow Tesseract Test Translator Toolkit Urchin UTM parameters V8 VirusTotal VisBug Wave Federation Protocol Weave Web Accelerator Web Designer Web Server Web Toolkit Webdriver Torso WebRTC

Operating systems

Machine learning models

Neural networks

Computer programs

Formats and codecs

Programming languages

Search algorithms

Domain names

Typefaces

Software

A	Aardvark Account Dashboard Takeout Ad Manager AdMob Ads AdSense Affiliate Network Alerts Allo Analytics Antigravity Android Auto Android Beam Answers Apture Arts & Culture Assistant Attribution Authenticator
B	BebaPay BeatThatQuote.com Beam Blog Search Blogger Body Bookmarks Books Ngram Viewer Browser Sync Building Maker Bump BumpTop Buzz
C	Calendar Cast Catalogs Chat Checkout Chrome Chrome Apps Chrome Experiments Chrome Remote Desktop Chrome Web Store Classroom Cloud Print Cloud Search Contacts Contributor Crowdsource Currents (social app) Currents (news app)
D	Data Commons Dataset Search Desktop Dictionary Dinosaur Game Directory Docs Docs Editors Domains Drawings Drive Duo
E	Earth Etherpad Expeditions Express
F	Family Link Fast Flip FeedBurner fflick Fi Wireless Finance Files Find Hub Fit Flights Flu Trends Fonts Forms Friend Connect Fusion Tables
G	Gboard Gemini Nano Banana Gesture Search Gizmo5 Google+ Gmail Goggles GOOG-411 Grasshopper Groups
H	Hangouts Helpouts Home
I	iGoogle Images Image Labeler Image Swirl Inbox by Gmail Input Tools Japanese Input Pinyin Insights for Search
J	Jaiku Jamboard
K	Kaggle Keep Knol
L	Labs Latitude Lens Like.com Live Transcribe Lively
M	Map Maker Maps Maps Navigation Marketing Platform Meet Messages Moderator My Tracks
N	Nearby Share News News & Weather News Archive Notebook NotebookLM Now
O	Offers One One Pass Opinion Rewards Orkut Oyster
P	Panoramio PaperofRecord.com Patents Page Creator Pay (mobile app) Pay (payment method) Pay Send People Cards Person Finder Personalized Search Photomath Photos Picasa Picasa Web Albums Picnik Pixel Camera Play Play Books Play Games Play Music Play Newsstand Play Pass Play Services Podcasts Poly Postini PostRank Primer Public Alerts Public Data Explorer
Q	Question Hub Quick, Draw! Quick Search Box Quick Share Quickoffice
R	Read Along Reader Reply
S	Safe Browsing SageTV Santa Tracker Schemer Scholar Search AI Overviews Knowledge Graph SafeSearch Searchwiki Sheets Shoploop Shopping Sidewiki Sites Slides Snapseed Socratic Softcard Songza Sound Amplifier Spaces Sparrow (chatbot) Sparrow (email client) Speech Recognition & Synthesis Squared Stadia Station Store Street View Surveys Sync
T	Tables Talk TalkBack Tasks Tenor Tez Tilt Brush Toolbar Toontastic 3D Translate Travel Trendalyzer Trends TV
U	URL Shortener
V	Video Vids Voice Voice Access Voice Search
W	Wallet Wave Waze WDYL Web Light Where Is My Train Widevine Wiz Word Lens Workspace Workspace Marketplace
Y	YouTube YouTube Kids YouTube Music YouTube Premium YouTube Shorts YouTube Studio YouTube TV YouTube VR

Hardware

Pixel

Smartphones	Pixel (2016) Pixel 2 (2017) Pixel 3 (2018) Pixel 3a (2019) Pixel 4 (2019) Pixel 4a (2020) Pixel 5 (2020) Pixel 5a (2021) Pixel 6 (2021) Pixel 6a (2022) Pixel 7 (2022) Pixel 7a (2023) Pixel Fold (2023) Pixel 8 (2023) Pixel 8a (2024) Pixel 9 (2024) Pixel 9 Pro Fold (2024) Pixel 9a (2025) Pixel 10 (2025) Pixel 10 Pro Fold (2025)
Smartwatches	Pixel Watch (2022) Pixel Watch 2 (2023) Pixel Watch 3 (2024) Pixel Watch 4 (2025)
Tablets	Pixel C (2015) Pixel Slate (2018) Pixel Tablet (2023)
Laptops	Chromebook Pixel (2013–2015) Pixelbook (2017) Pixelbook Go (2019)
Other	Pixel Buds (2017–present)

Nexus

Smartphones	Nexus One (2010) Nexus S (2010) Galaxy Nexus (2011) Nexus 4 (2012) Nexus 5 (2013) Nexus 6 (2014) Nexus 5X (2015) Nexus 6P (2015)
Tablets	Nexus 7 (2012) Nexus 10 (2012) Nexus 7 (2013) Nexus 9 (2014)
Other	Nexus Q (2012) Nexus Player (2014)

Other

v t e Litigation
Advertising	Feldman v. Google, Inc. (2007) Rescuecom Corp. v. Google Inc. (2009) Goddard v. Google, Inc. (2009) Rosetta Stone Ltd. v. Google, Inc. (2012) Google, Inc. v. American Blind & Wallpaper Factory, Inc. (2017) Jedi Blue
Antitrust	European Union (2010–present) United States v. Adobe Systems, Inc., Apple Inc., Google Inc., Intel Corporation, Intuit, Inc., and Pixar (2011) Umar Javeed, Sukarma Thapar, Aaqib Javeed vs. Google LLC and Ors. (2019) United States v. Google LLC (2020) United States v. Google LLC (2023)
Intellectual property	Perfect 10, Inc. v. Amazon.com, Inc. (2007) Viacom International, Inc. v. YouTube, Inc. (2010) Lenz v. Universal Music Corp.(2015) Authors Guild, Inc. v. Google, Inc. (2015) Field v. Google, Inc. (2016) Google LLC v. Oracle America, Inc. (2021) Smartphone patent wars
Privacy	Rocky Mountain Bank v. Google, Inc. (2009) Hibnick v. Google, Inc. (2010) United States v. Google Inc. (2012) Judgement of the German Federal Court of Justice on Google's autocomplete function (2013) Joffe v. Google, Inc. (2013) Mosley v SARL Google (2013) Google Spain v AEPD and Mario Costeja González (2014) Frank v. Gaos (2019)
Other	Garcia v. Google, Inc. (2015) Google LLC v Defteros (2020) Epic Games v. Google (2021) Gonzalez v. Google LLC (2022)

Concepts

Products

Android	Booting process Custom distributions Features Recovery mode Software development
Street View coverage	Africa Antarctica Asia Israel Europe North America Canada United States Oceania South America Argentina Chile Colombia
YouTube	Copyright strike Education Features Moderation Most-disliked videos Most-liked videos Most-subscribed channels Most-viewed channels Most-viewed videos Arabic music videos Chinese music videos French music videos Indian videos Pakistani videos Official channel Social impact YouTube Premium original programming
Other	Gmail interface Maps pin Most downloaded Google Play applications Stadia games

Documentaries

Books

Popular culture

Google Feud
Google Me (film)
"Google Me" (Kim Zolciak song)
"Google Me" (Teyana Taylor song)
Is Google Making Us Stupid?
Proceratium google
Matt Nathanson: Live at Google
The Billion Dollar Code
The Internship
Where on Google Earth is Carmen Sandiego?

Other

Italics denotediscontinued products.

Retrieved from "https://en.wikipedia.org/w/index.php?title=Data_Commons&oldid=1322884219"

Categories:

Hidden categories:

[8]ページ先頭