vera.ai: VERification Assisted by Artificial Intelligence

  • Thessaloniki, Greece
  • September 2022
  • Information Technologies Institute of Centre for Research and Technology-Hellas (CERTH-ITI)
Online disinformation and fake media content have emerged as a serious threat to democracy, economy and society. vera.ai seeks to build professional trustworthy AI solutions against advanced disinformation techniques, co-created with and for media professionals & researchers and to also set the foundation for future research in the area of AI against disinformation. Key novel characteristics of the AI models will be fairness, transparency (incl. explainability), robustness against concept drifts, continuous adaptation to disinformation evolution through a fact-checker-in-the-loop approach, and ability to handle multimodal and multilingual content.
  • Open source
  • 60/40/0 (% m/f/d)
  • Public

Project stage (in ):

Research/planning

Implemented by:

Inhouse

Industrial Sectors:

Information and communication, Professional, scientific and technical activities

Usage of AI:

Natural Language Processing, Computer Vision, Speech and Sound Recognition, Data Management and Analysis, Information Retrieval, Generative Models

Generation of AI:

Traditional Machine Learning (Linear Regression, CART, SVM, etc.), Deep Learning (CNN, Transformers, etc.)

Model training:

Supervised Learning, Semi-supervised Learning, Unsupervised Learning, Reinforcement Learning, Transfer Learning

Contact

Responsible Person

Dr. Symeon Papadopoulos

Motivation and values

What in your view defines the public interest and how does your project meet this purpose?

Recent major events and social trends, such as elections across the globe, the COVID pandemic, and a global rise in populism, along with impressive advances in the field of Artificial Intelligence (AI), such as the automated creation of highly convincing imagery and text (also known as deepfakes), have made clear that the challenge of disinformation can have grave impact on our societies. Media and news professionals, who are still seen as the gatekeepers of accurate and high-quality information, are in dire need of sophisticated solutions to make their ever increasing (both in volume and complexity) verification work more efficient and manageable. To this end, vera.ai seeks to build trustworthy AI solutions against advanced disinformation techniques, co-created with and for media professionals and set the foundation for future research in the area of AI against disinformation. Key novel characteristics of the AI models will be fairness, transparency (including explainability), robustness to new data, and continuous adaptation to new disinformation techniques.

How did the idea of your project come about?

Internal in our organization

What is the goal of your project in relation to the public interest?

Deliver multilingual and multimodal trustworthy AI methods for content analysis, enhancement, and evidence retrieval to assist in disinformation detection and content verification. Deliver multimodal trustworthy AI tools for the detection of deepfake, synthetic media and manipulated content, including AI generated and manipulated content, namely images, videos, audio and text. Enable the discovery, tracking, and impact measurement of disinformation narratives and campaigns across social platforms, modalities, and languages, through integrated AI and network science methods. Provide an intelligent verification and debunk authoring assistant, based on chatbot NLP technology, to explain the AI outputs and guide its professional users through the verification process. Pursue a fact-checker-in-the-loop approach to seamlessly gather new actionable feedback as a side effect of verification workflows. Ensure adoption and sustainability of the new AI tools in real-world applications through integration in leading verification tools (Truly Media and the InVID-WeVerify verification plugin) and AI platforms (AI4EU, ELG) with established large stakeholder communities.

Did you follow one or more guidelines for ethical AI, and if yes, which one?

European Union - Ethics Guidelines for Trustworthy AI

Guiding Values

What are your top 5 guiding-values for the project? Top 1

Verify multimodal content

What measures do you use to implement this value? Top 1

Bias-free and explainable AI methods for verification of images, videos, and audio content; explainable multilingual credibility assessment and retrieval of relevant evidence from authoritative sources; multimodal deepfake and manipulation analysis; cross-modal detection of decontextualised content. To ensure the AI models continuously adapt to new kinds of disinformation, vera.ai will also implement a factchecker-in-the-loop approach for seamless gathering of new, human-tagged training data.

What are your top 5 guiding-values for the project? Top 2

Analyse disinformation agents

What measures do you use to implement this value? Top 2

A network-science approach and tools for analysing communities and spreaders using graph neural networks; deep learning methods for extracting source credibility signals following the W3C guidelines; visualisations of spreader networks and accounts through Truly Media and the verification plugin.

What are your top 5 guiding-values for the project? Top 3

Uncover disinformation campaigns

What measures do you use to implement this value? Top 3

Deep learning methods for near-duplicate detection of images and videos and clustering of topically and semantically similar messages; knowledge graph models and spatio-temporal analysis of disinformation campaigns; network science methods for discovery of coordinated sharing behaviour.

What are your top 5 guiding-values for the project? Top 4

Assess and reduce impact

What measures do you use to implement this value? Top 4

Development of a first-of-a-kind AI-based debunk writing assistant; research audits of the effectiveness of social platforms and their algorithms on reducing the impact and spread of mis- and disinformation; research into the effectiveness of debunking and fact-checking.

What are your top 5 guiding-values for the project? Top 5

Assist professionals

What measures do you use to implement this value? Top 5

Development and integration of novel AI-based verification and debunk authoring assistant that incorporates a chatbot advising professionals on how best to use the new AI tools and interpret their outputs and offers assistance with debunk authoring through automated evidence gathering and snippet authoring; also encouraging users to flag mistakes in the AI outputs, which will then be used for retraining.

Design & Safeguards

Which stakeholders were involved in the process of development and implementation of the project?

Developer, External domain expert, Academic researcher, Industry partners

How did you engage relevant stakeholders?

Workshops, Survey, Interviews, Co-creation process

Did you apply specific methods of participatory design, and if yes, which ones?

Yes following a co-creation methodology.

Have the project results been validated by third parties?

No

Have the design and the results of the project been made transparent to the public?

Yes

In what way have the design and the results of the project been made transparent to the public?

By academic researchers

Involving the people who will be affected by the project is a necessary part of the project design

5

What technical and organizational safeguards have you implemented to protect personal data and mitigate possible harms? Choose all that apply

Data minimization (incl. not gathering personal data), Authentication and access management, User control (consent, update, retract)

Do you take measures in regards to ecological sustainability of your project? Please elaborate.

vera.ai acknowledges that the benefits of fighting disinformation using AI should not come at disproportionately large environmental costs as a result of massive AI model training and deployment. To this end, care will be taken to select good trade-offs between accuracy and model size (which translates into higher energy consumption), employing more compact architectures whenever possible.

Does your project rely partly or fully on the use of open data and/or publish results in an open data set? Please elaborate.

Project partners may re-use existing open datasets available online in compliance with their license and applying GDPR provisions. vera.ai datasets will be made available as soon as they are collated and cleared through ethics and data management.