The NGI Search 2nd Open Call Results have been Announced!

ngi-search 17 de octubre de 2023
The NGI Search 2nd Open Call Results have been Announced!

The NGI Search 2nd Open Call Results have been Announced! πŸ₯³ πŸŽ‰

The NGI Search project, co-funded by the European Union's Horizon 2020 research and innovation program, has unveiled the outcomes of its second open call.

This initiative aims to assist a diverse range of beneficiaries, including researchers, developers, internet activists, hackers, SMEs, and startups, in adopting and developing solutions for data and resource discovery on the Internet.

Eighty-eight proposals were submitted for this call, resulting in the selection of eleven proposals that have sought financial support. These beneficiaries will receive technical and business support, from Open Source Licensing to Market Readiness.

πŸ”For more details, please click here.

image
Meet the selected projects.πŸ‘‡πŸ‘‡

βœ…Verif.ai

Verif.ai is an AI system designed to verify and document the accuracy of facts in AI-generated texts. It enables users to find verified and trustworthy answers to their questions quickly and to reduce the spread of misinformation and false health information on the web. A system capable of providing verifiable and referenced answers can facilitate and accelerate progress in many areas of medicine and life sciences, such as discovering new biological targets, assessing the veracity of hypotheses generated from real data, summarising regulatory documents or improving the work of product and solution managers, thereby increasing user productivity.

The model and code will be released under an MIT open-source licence, allowing commercial and non-commercial use. The model will be published on HuggingFace, while the code will be on GitHub.

βœ…Science-Checker. Reloaded.

The project addresses the problem of spreading fake news in the scientific literature to promote transparency and build trust in online information.

The group's mission within the NGI Search framework is to combine their current automatic fact-checking web application, Science Checker, with new pipelines using Natural Language to extract and select answers from several scientific publications.

The result will be a novel architecture that explores scientific literature through natural language queries, without personal information collection and counting on links to verified documents. In addition, the users can inspect the eventual debate around their questions.

βœ…ADITIV - Asset Discovery In The Internet of Value

Blockchain is increasingly gaining importance as a siloed technology for implementing on-chain dApps and as an infrastructure element in the global internet-based computing landscape. In particular, uniquely identifying assets allows more straightforward implementation of traceability and auditability in any tokenisation process within the Internet of Value, thus increasing trust, privacy, and transparency regarding products and processes involving physical or digital assets. This grant proposal focuses on developing an open "search engine indexing mechanism" that leverages digital identifiers and asset metadata whose states are bilaterally synchronised between on-chain and off-chain systems.

ADITIV will provide an open-source "search engine indexing mechanism" that leverages digital identifiers and asset metadata whose states are bilaterally synchronised between on-chain and off-chain systems.

βœ…MindBugs Discovery

MindBugs Discovery is an AI-powered knowledge graph that visualises connections in the world of disinformation. It started during the AI4Media programme and the IPI hackathon and was showcased at the IPI World Congress, with excellent feedback from the journalistic community. It offers an innovative solution for navigating and understanding the complex disinformation landscape. The tool reveals the narrative, topic, and entities targeted by inputting a statement. Interactive charts display origins, evolution, locations targeted, and temporal trends.

The project will be released as an open-source initiative. This means that the AI algorithms, knowledge graph, and dataset will be freely available for others to use and enhance.

βœ…Debunker-Assistant

Debunker-Assistant (D-A) is a citizen-based AI tool which supports the analysis and detection of online misinformation. The tool takes as an input the link of a news article and returns its misinformation profile based on NLP and NA features, part of which are designed by involving non-expert citizens. D-A works with Italian and English news, and it is not bound to a specific topic: it is designed as a general-purpose tool that can extract relevant features for assessing the quality of information.

Inside GitHub, the organisation will collect among its repos:

πŸ“Œ The NLP (Natural Language Processing) and NA (Network Analysis) application code
πŸ“Œ The code of the interface library that will be created
πŸ“Œ The definitions of usable API
πŸ“Œ The site that presents the project and collects the community's contributions that will be born around it.

βœ…ALLMA: an AI-based LLM for Android

This project will Compare and evaluate open-source LLMs against criteria for:

πŸ‘‰ Privacy protection of users
πŸ‘‰ Usability/quality
πŸ‘‰ Collaborative nature: open-source, open data, open standards and community involvement

It will also integrate the best open-source LLM from (1) under the working name 'ALLMA' into /e/OS. /e/OS is a fully open-source, 'privacy-by-design' fork of the Android Open Source Project that combines mainstream usability with privacy (absolutely zero data collection on users) and open-source. The focus will be the same for ALLMA: making it useful for mainstream users while championing its open-source and privacy-safe nature.

All the work with /e/OS is published open-source at e Β· GitLab under GPL v3 for new projects and the same open-source license as the origin project for forked projects.

βœ…SWH Scanner

SWH Scanner is a free software edited by the Software Heritage Foundation (Software Heritage is responsible for INRIA, the French National Institute for Research in Computer Science and Control). It is currently a multi-platform CLI that scans a source code project to discover files and directories in the Software Heritage archive and output a summary of the results to different formats.

Shortly, SWH Scanner should become a solution to identify open source content provenance and potential security vulnerabilities, licensing issues, or outdated components in the software being developed or used.

The project is open source, like the Software Heritage archive platform. The SWH-Scanner project is released under GNU GENERAL PUBLIC LICENSE (GPL).

βœ…TEThYS

TEThYS is made of two parts: a pipeline for ingesting massive data corpora, built upon state-of-the-art technologies (including large language models), and extracting from them highly relevant topics, clustered along orthogonal dimensions; and an interactive dashboard, active after data preparation, supporting topic visualisation as word clouds and their exploration through user-friendly interaction.

The TEThYS concept is fully demonstrated in the CorToViz prototype that explores the CORD-19 dataset collected during the pandemic, focused on COVID-19 and the SARS-CoV-2 virus. An unbounded number of interesting domains could be analysed using the TEThYS approach, including climate change and controversial debates on social media.

NGI SEARCH Support will allow the development of the testbed into a solid architecture (reaching a mature TRL 5) and an understanding of the crucial aspects of approaching TRL 6 being ready for addressing the market.

The code of the prototypal architecture of CorToViz is open source on GitHub, under license BSD 3-clause, which permits distribution, changes, and commercial/private use. TEThYS will be implemented and published on a similar repository, and all NGI Search quality criteria requirements will be followed.

βœ…WAISE - Wiki Artificial Intelligence Search Engine

WAISE aims to create an open-source AI Search Server which adds an efficient, conversational layer on top of XWiki or any other CMS by running LLMs on consumer-grade or Cloud GPU hardware using the OpenAI API and the LocalAI framework. The project will, in particular, perform the following tasks:

πŸ“Œ Design and implement an API for computing, indexing and querying vector embeddings, which considers content access rights and responds to users in natural language.
πŸ“Œ Create a qualitative and quantitative benchmark of open-source LLM models on question-answering capabilities, content summarisation and content generation.
πŸ“Œ Create a versatile and easy-to-use web UI for conducting advanced conversations with the WAISE server, which will index the content of the XWiki server or any site.
πŸ“Œ Integrate the search and question-answering capabilities into Element Chat (Matrix protocol) by building an AI Search on a Matrix Chat Bot, which communicates with the WAISE server.

✨All the XWiki source code is available under the LGPLv2 open-source license.✨

βœ…The World Literature KG

World Literature Knowledge Graph (WL-KG) is a knowledge base aimed at exploring and mitigating the underrepresentation of non-Western writers.

The resource relies on an interoperable semantic model to compare the degree of inclusivity of different platforms (Eg, Open Library, Goodreads). WL-KG is accessible through a graph-based visualisation platform designed to encourage serendipic exploration. The existing resource will be improved with a set of procedures to automatically gather knowledge from new sources, which will be tested on 3 new platforms. The project already has a Github Folder where the ontologies have been released. Additionally, the first version of the Knowledge Graph is publicly available through SPARQL Endpoint. Code and an updated version of the Knowledge Graph will be released in this folder.

βœ…Chat-EUR-Lex

This proposal aims to revolutionise the accessibility of the EUR-Lex normative database, a vital source of EU legislation, employing cutting-edge AI techniques, including Chat-Based Large Language Models (Chat LLMs) and Retrieval Augmented Generation (RAG). The objective is to create an AI-powered interface capable of understanding complex legal texts, providing simplified explanations, and conducting interactive, context-specific discussions. The ability to deliver understandable and accurate legal insights in real time will significantly reduce the barrier to understanding EU law for citizens and businesses alike. This application of AI contributes to the EU's vision of promoting digital transformation, transparency, and inclusiveness, thus fostering a well-informed and participatory European community.

All the code produced will be published in the linked GitHub repository. The GitHub repository will also contain automated scripts and instructions to test, maintain, add features and deploy the application.

✨Ready to be the next NGI Search trailblazer?✨

NGI Search is gearing up for more Open Calls, even though the most recent one, the third edition, closed on October 2nd. But fear not;there are two more opportunities on the horizon.

The NGI Search project is an endeavour for all those who can convincingly demonstrate their commitment to transforming how we utilise, experience, explore, and uncover data and resources on the Internet and the web.

Remember to become part of the NGI Search Community and keep you updated by the project's Twitter (X) and LinkedIn for the latest news.