]> <p>OpenLink Software is a United Kingdom SME founded in 1992. It operates a business development and sales subsidiary in the United States. Most product development takes place within the EU, across the UK, Netherlands and Bulgaria. The company is a leading provider of high-performance, scalable, and secure technology covering: data access drivers, data virtualization middleware; native database management (RDBMS or RDF based Graph Store), and enterprise collaboration. Respective product portfolio offerings include: OpenLink High-Performance Drivers for ODBC, JDBC, ADO.NET, OLE-DB, and XMLA; Virtuoso Universal Server; and OpenLink Data Spaces for socially enhanced personal and/or enterprise collaboration.</p> <p>OpenLink has extensive experience in scalable RDF triple (and quad) stores as a result of extending its native Virtuoso database/SQL engine to incorporate SPARQL query language support. Virtuoso also includes powerful Linked Data Deployment and RDF to RDBMS transformation functionality. It offers management and creation of physical and virtual triples in conjunctions with the ability to declaratively produce RDF views of SQL Data (SQL to RDF mapping). OpenLink possess pioneering experience at both the applications and data management levels within the Semantic Web technology realm.</p> <p>OpenLink is a W3C member, an active participant in the W3C Semantic Web Education and Outreach (SWEO) Interest Group, a key member of the Linking Open Data project, and timeless supporter of the Open Data Movement. OpenLink is a DBpedia project co-creator and has hosted live instances of the DBpedia database since project inception.</p> <p>As a consortium member, OpenLink will leverage its experience with local and distributed query processing, SQL and SPARQL query optimization, RDF store and heterogeneous data integration technologies en route to developing an integrated database backbone for the LOD2 project.</p> OpenLink contributes in particular to developing the scalable LOD2 knowledge store (WP2); to track and inference about data provenance and reliability; to support personalized views on knowledge and spatial data; alerts on data; to contribute to standardization activities regarding the integration of semantic and spatial technologies. 5 WP1 – Requirements, Design and LOD2 Stack Prototype Objectives of this work package are: (1) to develop use case specifications and to collect user requirements by consulting the communities of practice relevant for the LOD2 use cases and additional prospective application scenarios, (2) to identify technical constraints as well as standards, (3) to produce the architecture and the LOD2 Stack design and (4) to produce an early prototype of the LOD2 Stack in the first year. 1 WP6 – Interfaces, Component Integration & LOD2 Stack This work package will continue the prototyping activity under WP1/Task 1.4 by fully integrating the individual components developed in WP2-5 into a ready-to-use LOD2 Stack and associated APIs. The primary goal of the LOD2 Stack integration is to enable communities of practice to rapidly create domain specific Linked Data applications. Consequently, the LOD2 Stack will support the whole life cycle of Linked Data from creation over enrichment, interlinking, fusing to maintenance. The stack will be very versatile, for all functionality we will define clear interfaces, which enable the plugging in of alternative third-party implementations. We will also provide a stack configurator, which enables potential user to create their own personalized version of the LOD2 Stack, which contains only those functions relevant for their usage scenario. 6 WP2 – Storing and Querying Very Large Knowledge bases This work package implements the LOD2 knowledge store component needed for managing the Web of Linked Data as a vast database. The starting point is OpenLink Virtuoso and MonetDB on the database side and Sindice on the information retrieval side. The present data volumes handled are around 10 billion RDF triples and the target is over the 1 trillion triples. The approach is scale-out out (physical) complemented with significant scalability improvements in the RDF engine (logical): Physically, when data grows, servers can be added and data redistributed without interruption of service, ibid for server failure. Logically, there is no point answering questions nobody is asking. Therefore the base data is kept as RDF with text indexing and search ranking. Additional inference results or indices for caching joins are made as a by-product of querying; further exploitation of such (partially) materialized inferences through the graph at run-time is to exploit structural correlations in the graphs. 2 WP3 – Knowledge Base Creation, Enrichment and Repair WP3 contains tasks focused on the transformation of legacy data to RDF and Linked Data and furthermore on the improvement of existing or extracted data especially with respect to schema enrichment and ontology repair. It is complementary to WP4, which is concerned with interlinking several knowledge bases and providing unified views of them. Tasks concerning the triplification of data will be grounded on existing techniques and know-how of the consortium and will be refined during the lifetime of this project and integrated into the LOD2 Stack. Legacy data triplification represents the entry point for legacy systems to participate in the LOD cloud. The members of the Consortium are leading in the development of transformational tools such as Virtuoso Sponger, RDF Views, D2R server, Triplify, and the DBpedia framework, which have received high acceptance in the Linked Data community. 3 WP4 – Reuse, Interlinking and Knowledge Fusion While WP3 is concerned with making legacy data available via URLs – a prerequisite – and enrichment of knowledge bases, this WP addresses automatic and semi-automatic link creation with minimal human interaction, evolvement of knowledge bases under the aspect of linkage and schema mapping combined with Data Fusion. 4 WP8 – Use Case 2: LOD2 for Enterprise Data Web This use case will be driven by Exalead, one of the leading enterprise search providers worldwide. We will deploy the LOD2 platform in a real corporate environment with high semantic information integration requirements and needs. Based on the authoring tools and semantic GUIs developed for the LOD2 Stack, we will implement a set of procedures for data input that integrate semantic features and annotations early in the ingestion process of data. We will also suggest a set of best practices for storing, managing, accessing and exchanging data. Moreover, a quality framework will be defined to measure the benefit of using LOD2 components in the corporate. This framework will consist of a set of quality measures to define the gain in precision, recall of searching and browsing the corporate data. Other measurements will be investigated like the impact of the LOD2 platform in the activity of the corporation (income, costs, efficiency, etc). 8 WP10 – Training, Dissemination, Community Building, Fertilization The general aim of this work package is to establish a worldwide focal point for academic and industry parties interested in contributing to or taking advantage of the novel Linked Data methodologies and components, which will emerge in the project. 10 Realizing the vision of LOD2 together with the ones of the three use cases will have significant socio-economic impact. Standardization of such an architecture and exploitation of knowledge and technical results (and related IPR) is covered in this work package. 11 The project management will entail strategic, project-wide as well as day-to-day central management and coordination activities. The several different management boards which will be established in the consortium will be responsible for decisions and activities of different scope and level according to their function. 12 <p>DBpedia is a community effort to extract structured information from Wikipedia and to make this information available on the Web. It currently already contains a tremendous amount of valuable knowledge extracted from Wikipedia. The DBpedia knowledge base will be used for evaluation LOD2&rsquo;s interlinking, fusing, aggregation and visualization components. The DBpedia multi-domain ontology will be used as background-knowledge for the LOD2 applications (WP7, WP8 and WP9), and as an alignment and annotation ontology for LOD in general.</p> DBpedia is a community effort to extract structured information from Wikipedia and to make this information available on the Web. It currently already contains a tremendous amount of valuable knowledge extracted from Wikipedia. The DBpedia knowledge base will be used for evaluation LOD2’s interlinking, fusing, aggregation and visualization components. The DBpedia multi-domain ontology will be used as background-knowledge for the LOD2 applications (WP7, WP8 and WP9), and as an alignment and annotation ontology for LOD in general. <p>Virtuoso is an innovative industry standards compliant platform for native data, information, and knowledge management. It implements and supports a broad spectrum of query languages, data access interfaces, protocols, and data representation formats that includes: SQL, SPARQL, ODBC, JDBC, HTTP, WebDAV, XML, RDF, RDFa, and many more.</p> <p>In addition to its core data management capabilities, Virtuoso also delivers sophisticated Data Virtualization functionality enabling the construction of federated views -- which may or may not be materialized -- over heterogeneous RDBMS, Web Services, and Hypermedia data sources.</p> <p>The open-source edition of Virtuoso, which includes a scalable high-performance RDF Quad Store, will be the basis for the LOD2 Stack's knowledge store.</p> Virtuoso is a knowledge store and virtualization platform that transparently integrates Data, Services, and Business Processes across the enterprise. Its product architecture enables it to deliver traditionally distinct server functionality within a single system offering along the following lines: Data Management & Integration (SQL, XML and EII), Application Integration (Web Services & SOA), Process Management & Integration (BPEL), Distributed Collaborative Applications. The open-source data integration server and the highly efficient and scalable RDF triple store implementation in Virtuoso will be the basis for the knowledge store component in the LOD2 Stack. 2011-08-31 2013-08-31 D2.1.1 – Initial LOD cloud hosted on the LOD2 Knowledge Store Cluster 2010-11-30 D2.1.3 – Intermediate hosted LOD cloud (50B) and knowledge store evaluation 2011-11-30 D2.1.3 - D2.1.5 are performance and deployment targets. The performance and deployment targets are validated with real and synthetic workloads as described in Task 2.1. The baseline is the present LOD cache of about 8 billion triples, expanded with other data as it becomes available. D2.1.4 – Intermediate hosted LOD cloud (250B) and knowledge store evaluation 2012-11-30 D2.1.3 - D2.1.5 are performance and deployment targets. The performance and deployment targets are validated with real and synthetic workloads as described in Task 2.1. The baseline is the present LOD cache of about 8 billion triples, expanded with other data as it becomes available. D2.1.5 – Final hosted LOD cloud (1T) and knowledge store evaluation 2013-11-30 This deliverable is an update of the copy of the Linked Open Data cloud (including all major data sets) hosted on the LOD2 Knowledge Store Cluster as initially deployed in <a href="../../Deliverable/D2.1.1">D2.1.1</>. The deliverable will include benchmark results for major RDF benchmarks, which demonstrate the scalability improvements of the LOD2 Knowledge Store component throughout the project. For this particular deliverable we envision the scalability to reach 1 trillion triples. 2011-08-31 As servers are added or removed, data is migrated to keep a given level of redundancy and load balance. An initial implementation of reusable intermediate join and owl:sameAs inferences is also included in this deliverable. D2.6 – LOD2 knowledge store release with integrated bulk processing features 2012-08-31 This release will add map-reduce style function-shipping-to-cluster-nodes functionality to the knowledge store. D2.8 – LOD2 knowledge store release with basic geographical indexing 2013-08-31 This release of the LOD2 knowledge store will contain functionality that integrates spatial criteria into search and demonstrates improved scale of spatial queries from adaptive caching. <p>The Web-based Systems Group at Freie Universität Berlin explores technical and economic questions concerning the development of global, decentralized information environments. Its current focus lies on the publication and interlinking of structured data on the Web using Semantic Web technologies.</p> <p>The group has initialized several widely used open source software projects including D2RQ, D2R Server, RAP – RDF API for PHP, and NG4G – Named Graphs API for Jena. The group contributes to several open data publishing efforts including DBpedia and the W3C Linking-Open-Data project which aims at interlinking large numbers of data sources on the Web. The Web-based Systems Group is active within the World Wide Web Consortium where it has contributed to the SPARQL recommendation and participates in the Semantic Web Education and Outreach activity. The group maintains strong research ties with the Massachusetts Institute of Technology, Hewlett-Packard Labs, and the Open Archives Initiative.</p> Freie University Berlin will bring in expertise, tools and outreach capabilities to LOD2: (1) FUB has developed and maintains the Silk – Link Discovery Framework and Link Quality Assurance Workbench, which will be significantly extended and integrated into the LOD2 Stack. (2) FUB has developed D2R Server, the most widely used tool for publishing relational databases as Linked Data on the Web. D2R Server will be used for the domain complementation task in WP3 and will be included together with Pubby and Silk into the LOD2 Stack (WP6). (3) Within WP10 Training, Dissemination, community building, FUB will use its existing community building (initiator of W3C LOD) and outreach capabilities (Linked Data on the Web (LDOW) workshop series, Semantic Web Challenge competition series) to maximize the impact of LOD2. 4 <p>Semantic Web Company (SWC), based in Vienna, is an SME, which offers consulting services in the fields of semantic web technologies, knowledge management systems and social software since 2005. Besides being one of the first European suppliers of consulting services in semantic web technologies and related topics, SWC is the founder of the I-SEMANTICS conference series, one of Europe’s largest industry outreach events in the field of semantic systems and knowledge management, which will take place for the 6th time in 2010. The mission of SWC is to narrow the gap between industry and academia R&D in semantic systems and to accelerate the time to market process of semantic web-technologies. By providing practice-oriented knowledge transfer SWC help companies and public organizations to integrate and apply semantic technologies for various purposes such as data integration, resource management and knowledge management. SWC communicates benefits and risks of semantic applications in corporate environments to developers and decision makers and serves organizations from the life sciences, insurance, energy and the media. SWC has built a special media focus over the last years specializing in business development aspects through semantic (web) technologies. SWC acts as a hub of a wide network of industry partners (software industry, end-user organizations), other Networks (Know Center Graz, Salzburg Research, Austrian Computer Society, W3C, PWM - Platform for Knowledge Management) and academic and research partners throughout Europe.</p> SWC adopt the tool Poolparty, a self-developed modeling tool for corporate thesauri, as a component for the LOD2 stack. SWC will provide its expertise in technology assessment and business development when it comes to evaluate the economic rationale, organisational effects and the commercial potential of semantic web technologies. SWC will investigate governance and regulatory issues such as competition, IPR issues and privacy thereby contributing to the use cases in WP 7, 8 & 9. Our expertise in communicating the principles, technologies and especially business needs for semantic web applications on the one hand and the „translation“ of practical problems into technological concepts on the other hand will help to disseminate the findings beyond the scope of the use cases. 6