Sentence embedding

In natural language processing, a sentence embedding refers to a numeric representation of a sentence in the form of a vector of real numbers which encodes meaningful semantic information. State of the art embeddings are based on the learned hidden layer representation of dedicated sentence transformer models. BERT pioneered an approach involving the use of a dedicated [CLS] token preprended to the beginning of each sentence inputted into the model; the final hidden state vector of this token encodes information about the sentence and can be fine-tuned for use in sentence classification tasks. In practice however, BERT's sentence embedding with the [CLS] token achieves poor performance, often worse than simply averaging non-contextual word embeddings. SBERT later achieved superior sentence embedding performance by fine tuning BERT's [CLS] token embeddings through the usage of a siamese neural network architecture on the SNLI dataset. Other approaches are loosely based on the idea of distributional semantics applied to sentences. Skip-Thought trains an encoder-decoder structure for the task of neighboring sentences predictions. Though this has been shown to achieve worse performance than approaches such as InferSent or SBERT. An alternative direction is to aggregate word embeddings, such as those returned by Word2vec, into sentence embeddings. The most straightforward approach is to simply compute the average of word vectors, known as continuous bag-of-words (CBOW). However, more elaborate solutions based on word vector quantization have also been proposed. One such approach is the vector of locally aggregated word embeddings (VLAWE), which demonstrated performance improvements in downstream text classification tasks. In recent years, sentence embedding has seen a growing level of interest due to its applications in natural language queryable knowledge bases through the usage of vector indexing for semantic search. LangChain for instance utilizes sentence transformers for purposes of indexing documents.

Examining European Press Coverage of the Covid-19 No-Vax Movement: An NLP Framework

Daniel Gatica-Perez

This paper examines how the European press dealt with the no-vax reactions against the Covid-19 vaccine and the dis- and misinformation associated with this movement. Using a curated dataset of 1786 articles from 19 European newspapers on the anti-vaccine ...

ASSOC COMPUTING MACHINERY2023

Prompt–RSVQA: Prompting visual context to a language model for Remote Sensing Visual Question Answering

Devis Tuia, Christel Marie Tartini-Chappuis, Sylvain Lobry, Valérie Zermatten

Remote sensing visual question answering (RQA) was recently proposed with the aim of interfacing natural language and vision to ease the access of information contained in Earth Observation data for a wide audience, which is granted by simple questions in ...

2022

Examining European Press Coverage of the Covid-19 No-Vax Movement: An NLP Framework

Daniel Gatica-Perez

ASSOC COMPUTING MACHINERY2023

Prompt–RSVQA: Prompting visual context to a language model for Remote Sensing Visual Question Answering

Devis Tuia, Christel Marie Tartini-Chappuis, Sylvain Lobry, Valérie Zermatten

2022

Examining European Press Coverage of the Covid-19 No-Vax Movement: An NLP Framework

Interpretable Representation Learning and Evaluation for Abstractive Summarization

Prompt–RSVQA: Prompting visual context to a language model for Remote Sensing Visual Question Answering

Graph Chatbot

Examining European Press Coverage of the Covid-19 No-Vax Movement: An NLP Framework

Interpretable Representation Learning and Evaluation for Abstractive Summarization

Prompt–RSVQA: Prompting visual context to a language model for Remote Sensing Visual Question Answering