LOD2 – Creating Knowledge out of Interlinked Data

Creating Knowledge out of Interlinked Data

LOD2 is a large-scale integrating project co-funded by the European Commission within the FP7 Information and Communication Technologies Work Programme (Grant Agreement No. 257943). Commencing in September 2010, this 4-year project comprises leading Linked Open Data technology researchers, companies, and service providers – 15 partners from across 11 European countries, plus one associated partner from Korea – and is coordinated by the AKSW research group at the University of Leipzig.

Over the past years the semantic web activity has gained momentum with the widespread publishing of structured data as RDF. The Linked Data paradigm has evolved from a practical research idea into a very promising candidate for addressing one of the biggest challenges in the area of intelligent information management: the exploitation of the Web as a platform for data and information integration in addition to document search. With partners among those who initiated and strongly supported the Linked Open Data initiative, LOD2 tackles these challenges by developing:

  1. enterprise-ready tools and methodologies for exposing and managing very large amounts of structured information on the Data Web,
  2. a testbed and bootstrap network of high-quality multi-domain, multi-lingual ontologies from sources such as Wikipedia and OpenStreetMap,
  3. algorithms based on machine learning for automatically interlinking and fusing data from the Web,
  4. standards and methods for reliably tracking provenance, ensuring privacy and data security as well as for assessing the quality of information,
  5. adaptive tools for searching, browsing, and authoring of Linked Data.

The resulting tools, methods and data sets are integrated and syndicated with large-scale, existing applications, with the benefits showcased in the three application scenarios of media and publishing, corporate data intranets and eGovernment.

Project at a glance

  • Funding: FP7 ICT Work Programme, Grant Agreement No. 257943
  • Duration: September 2010 – 2014 (4 years)
  • Consortium: 15 partners in 11 European countries, one associated partner in Korea
  • Coordination: AKSW research group, Universität Leipzig
  • Scenarios: media & publishing, corporate data intranets, eGovernment

Project structure · Work packages · Deliverables · Consortium

Results and technology

The project produced an integrated LOD2 technology stack – a set of Linked Data tools that can be used together to publish, interlink, and consume structured data on the Web. The stack bundles established components such as OntoWiki, Silk, DL-Learner, LIMES, Sparallax and OpenLink Virtuoso, distributed as Debian packages and virtual machines.

Technology stack overview · Results · Publications

Latest from the LOD2 blog

First release of the LOD2 Stack – the initial public release of the integrated stack, with installation instructions for the Debian packages and the virtual machine images.

PUBLINK Linked Data Starter Service call 2013 – the consultancy service that helped organisations take their first steps towards publishing Linked Open Data.

Big Data RDF store benchmarking experiences – results from benchmarking the triple stores used in the LOD2 knowledge store cluster.

All blog posts · LOD2 webinar series

Systems Notes

Occasional notes on the machinery underneath – how processors, memory and the conventions around them actually behave.

  • The Call Stack – what the hardware does on a call, why the stack grows down, and why measuring its worst case on a device is harder than it sounds.
  • The Program Counter – the one register a processor cannot execute without, and what interrupts and resets really do to it.
  • Read-Only Memory – mask ROM to flash, why flash cannot set a single bit on its own, and why the ROM in a microcontroller is almost always writable.

February 2021 archive · tagged memory, EEPROM, stack, call stack, control stack, run-time stack

Participate

The PUBLINK open call programme offered funded consultancy to public bodies, SMEs and research groups wanting to publish Linked Open Data. Results of the funded calls are documented in the 2010 results and the 2011 winners.

Contact the consortium · Imprint


Testimonials

Dr. Mateja Verlič (Zemanta d.o.o., R&D)

We read a lot about people doing important things, however this time we – LOD2 partners – are the ones changing the history of Web by revolutionizing information use and reuse, contributing to semantic data standards and pushing the limits of almost every WWW-related aspect (sc… read

Dr. Jens Lehmann (Universität Leipzig, Research Group Leader)

The size and number of Semantic Web knowledge bases published as Linked Open Data has been growing tremendously over the past years. The LOD2 project will be a key factor for sustaining this momentum. More importantly, the quality of knowledge bases and the scalability of methods accessing… read

Jindřich Mynarz (University of Economics, Czech Republic)

Apart from being legally open, linked open data is an open technology. Due to its non-exclusive, non-proprietary, well-formalized, and standards-based nature, linked open data supports a wide spectrum of uses. It does not exclude any application from using it and thus it is open to be mixe… read

Dr. Mladen Stanojevic (Institute Mihajlo Pupin, Serbia)

The richness of any semantic model and the usability of represented data is dependent on the links between them. LOD2 is an important step in this direction that will enable an efficient exploration and processing of vast quantity of open data on the Web.

read

Mun Yong Yi (Korea Advanced Institute of Science and Technology (KAIST))

Data becomes more meaningful and powerful as they are linked and integrated. As I learn more about what LOD can do, I am convinced that it will not only change the future of the Internet but also the quality of human life. I am glad that I am participating in this project, which will defin… read

Orri Erling (OpenLink Software, Virtuoso Program Manager)

Value from information increasingly depends on integration. RDF is a great model for this. LOD2 will make RDF a cost competitive alternative in the database space, without compromising ad hoc flexibility and expressive power. read

Vojtech Svatek (University of Economics, Czech Republic)

As researcher in ontological engineering, I am excited to see LOD2 provide tools that make the generation of semantically structured data easy and thus widespread. I believe that ontological research is deemed to build upon the world views already expressed by means of simpl… read

Andreas Blumauer (Semantic Web Company, CEO)

15 years ago we all were excited when we published HTML for the first time and it didn't take a long time until all of us were "on the internet". Now we are starting to publish data on the web. Based on semantic web technologies professional data management will be possible in distributed … read

Kingsley Idehen (OpenLink Software, CEO)

Three years ago, OpenLink Software enthusiastically contributed the prowess of Virtuoso to the grassroots effort that lead to DBpedia and the Linked Open Data cloud that coalesced around it. Today, we are both honored and enthusiastic about Virtuoso's critical infrastructure role in this n… read

Bastiaan Deblieck (TenForce, Partner and Business Development Manager)

An internet of data opens up tremendous opportunities for our corporate and government customers. We intend to be on the forefront of this evolution. read

Christian Dirschl (Wolters Kluwer, Content Architect)

Linked (Open) Data will change the existing publishing paradigms! Creating high quality content for professional usage will remain an important factor in future publishing, but additional access points and new usage environments will equally define its success. read

Dr. Giovanni Tummarello (National University of Ireland, Galway, Research Unit Leader)

Semantic Markups on the Web could drive information reuse to enable scenarios and applications which we can now only dream of. The idea is extraordinarely compelling, but we know now it won't simply realize itself. The LOD2 project is now a great opportunity for inspired and coordinated re… read

Gregory Grefenstette (Exalead, Chief Science Officer)

Enterpise search is all about providing correct, complete, and appropriate information to the employee and decision maker. LOD2 promises to not only allow internal company information to be linked up to the growing amount of Open Data on the web, but to also provide the mechanisms for val… read

Hugh Williams (OpenLink Software)

It is exciting to see the LOD2 project finally kick off in its quest to take the Linked Open Data cloud to the next level of scalability, performance and integration for the exploitation of the Web as a viable platform for enterprise level data and information integration. read

Martin Kaltenböck (Semantic Web Company, CFO)

Linked (Open) Data technologies offer a new way of data integration for the enterprise! Smooth interoperability between internal data sets can reduce costs as well as the enrichment of these data sets by external data can support new market intelligence paradigms for a better decision making. read

Dr. Peter Boncz (Centrum Wiskunde & Informatica)

The publishing of ever more datasets by e.g. governments adds value for many key applications including business intelligence, which will drive the Linked Open Data (LOD) paradigm going forward. In the LOD2 project, CWI is working to increase the scalability and performance of querying int… read

Dr. Sören Auer (Universität Leipzig, LOD2 coordinator)

The Linked Data paradigm is a simple and efficient way for integration of heterogeneous information on the Web. Ultimately, we will, for example, be able to search for a new appartment and a close-by available spot in child care in one go. read

Tassilo Pellegrini (Semantic Web Company, R&D)

Semantic interoperability changes the technological and economic nature of metadata opening up exciting opportunities for value creation in various comercial and non-commercial areas. Linked Data is the blueprint for this new ecosystem and it will change the way we think about and use the web today. read

Wouter Dewanckel (TenForce, WP Project Leader)

It is an honor to take part in this challenging integration project to create solutions that can generate business value out of emerging technologies. read

News

IEEE publication of ?Virtuoso, a Hybrid RDBMS/Graph Column Store?

Apr 23, 2012 4:55:17 PM

My article, Virtuoso, a Hybrid RDBMS/Graph Column Store (PDF), can be found in Volume 35, Number 1, March 2012 (PDF) of the Bulletin of the IEEE Computer Society Technical Committee on Data Engineering (also known as the IEEE Data Engineering Bulletin). Abstract: We discuss applying column store techniques to both graph (RDF) and relational data for mixed workloads ranging from lookup to analytics in the context of the OpenLink Virtuoso DBMS. In so doing, we need to obtain the excellent m ...

News from the CKAN team 19 April 2012

Apr 19, 2012 4:36:28 PM

Here’s an update on what the CKAN team have been doing lately. It’s been a while since the last one and we’ve been busy, so there’s plenty to report. Features Adrià has implemented a great map view in Recline, CKAN’s built-in data viewer. If some structured data resource contains latitude and longitude information, this will enable it to be viewed on a map from within CKAN. A sneak preview is here (select the ‘map’ view). Ross has done some work on a ‘Related Stuff’ extensi ...

ICDE 2012 (post 6 of 6) - Science Data Panel

Apr 17, 2012 9:36:28 PM

Michael Stonebraker chaired a panel on the future of science data at ICDE 2012 last week. Other participants were Jeremy Kepner from MIT Lincoln Labs, Anastasia Ailamaki from EPFL, and Alex Szalay from Johns Hopkins University. This is the thrust of what was said, noted from memory. My comments follow after the synopsis. Jeremy Kepner: When Java was new we saw it as the coming thing and figured that in HPC we should find space for this. When MapReduce and Hadoop came along, we saw this as a se ...