59 items found

Organisations: SoBigData Catalogue Tags: Web data

Filter Results
    • PDF
      The resource: 'Misinformation Detection ...' is not accessible as guest user. You must login to access it!
  • Method

    Quantum Distance-Based Classifier

    The Quantum Distance-Based Classifier is a technique inspired by the classical k-Nearest Neighbors that leverages quantum properties to perform prediction.
  • Application

    The news tells us about peace - dashboard

    Previous research demonstrates that the official Global Peace Index (GPI) can be captured at a higher frequency through GDELT, a digital news database. We have created a...
    • PDF
      The resource: 'DEAP-FAKED: Knowledge ...' is not accessible as guest user. You must login to access it!
  • Experiment

    Physical Activity Levels and Perceived Changes in the Context of Intra-EEA Mi...

    As mobility within the European Economic Area (EEA) is on the rise, it is important to understand migrants’ health-related behaviors (such as physical activity [PA]) within...
  • Experiment

    Using Computer Vision Techniques to Study Images from the Web [Video Tutorial]

    The video tutorial discusses using computer vision techniques to study online images, including case studies, research methods, challenges faced, and lessons learned.
    • The resource: ' Using Computer Vision ...' is not accessible as guest user. You must login to access it!
  • Experiment

    Analysing Meme Collections with the Computer Vision Network Approach [Video T...

    The video tutorial presents techniques for analysing meme collections with the computer vision network approach, including seven steps for network building and interpretation,...
    • The resource: 'Analysing Meme Collections ...' is not accessible as guest user. You must login to access it!
  • Experiment

    Making meme collections [Video Tutorial]

    The video tutorial discusses making meme collections, including the meme as a technical collection of objects, automated visual analysis, and meme collection distinctiveness.
    • The resource: 'Making meme collections ...' is not accessible as guest user. You must login to access it!
  • Dataset

    Multi-Task Faces (MTF) dataset

    The Multi-Task Faces (MTF) dataset consists of cropped human faces for classification tasks or other research purposes. Each image in the dataset is labelled according to four...
    • ZIP
      The resource: 'MTF_dataset_20230701' is not accessible as guest user. You must login to access it!
  • Access required...

    ×

    Method

    Private Boilernet

    Deploys an artificial neural network to remove the boilerplate from HTML files. Annotates the text content in the file or extracts the text from the HTML file.
  • Dataset

    CoPhIR

    The CoPhIR (Content-based Photo Image Retrieval) Test-Collection has been developed to make significant tests on the scalability of the SAPIR project infrastructure (SAPIR:...
    • The resource: 'cophir.isti.cnr.it' is not accessible as guest user. You must login to access it!
  • Dataset

    The Italian Music Dataset

    The dataset is built by exploiting the Spotify and SoundCloud APIs. It is composed of over 14,500 different songs of both famous and less famous Italian musicians. Each song...
    • JSON
      The resource: 'Dataset' is not accessible as guest user. You must login to access it!
  • Dataset

    GERDAQ Dataset

    This is a benchmark dataset of annotated search-engine queries. Mentions of entities in search-engine queries are tagged with the entity they refer to. Wikipedia is used as...
    • XML
      The resource: 'GERDAQ dataset' is not accessible as guest user. You must login to access it!
  • Method

    ArchiveSpark

    ArchiveSpark is an Apache Spark framework for easy data access, processing, extraction as well as derivation for Web archives and archival collections. It has a simple and...
    • The resource: 'ArchiveSpark on GitHub' is not accessible as guest user. You must login to access it!
  • Dataset

    German Academic Web

    The dataset contains regular crawls of the websites for German academic institutions.
  • Dataset

    MSN Search query log

    The data consists of an MSN Search query log excerpt with 15 million queries, from US users, sampled over one month of activity. Data attributes made available per query: 1)...
  • Dataset

    Product Reviews for Ordinal Quantification

    This data set comprises a labeled training set, validation samples, and testing samples for ordinal quantification. It appears in our research paper "Ordinal Quantification...
    • The resource: 'Zenodo link' is not accessible as guest user. You must login to access it!
  • Dataset

    Wikipedia Word Embeddings

    Embeddings were created through applying word2vec skipgram to a corpus of wikipedia non-stub articles from a December 2015 English dump with the following parameters: -cbow 0...
    • The resource: 'Embeddings' is not accessible as guest user. You must login to access it!
  • Method

    The Propagation of Misinformation in Social Media

    There is growing awareness about how social media circulate extreme viewpoints and turn up the temperature of public debate. Posts that exhibit agitation garner...
    • The resource: 'The Propagation of ...' is not accessible as guest user. You must login to access it!