Abstract:Representation learning has been proven to play an important role in the unprecedented success of machine learning models in numerous tasks, such as machine translation, face recognition and recommendation. The majority of existing representation learning approaches often require a large number of consistent and noise-free labels. However, due to various reasons such as budget constraints and privacy concerns, labels are very limited in many real-world scenarios. Directly applying standard representation learning approaches on small labeled data sets will easily run into over-fitting problems and lead to sub-optimal solutions. Even worse, in some domains such as education, the limited labels are usually annotated by multiple workers with diverse expertise, which yields noises and inconsistency in such crowdsourcing settings. In this paper, we propose a novel framework which aims to learn effective representations from limited data with crowdsourced labels. Specifically, we design a grouping based deep neural network to learn embeddings from a limited number of training samples and present a Bayesian confidence estimator to capture the inconsistency among crowdsourced labels. Furthermore, to expedite the training process, we develop a hard example selection procedure to adaptively pick up training examples that are misclassified by the model. Extensive experiments conducted on three real-world data sets demonstrate the superiority of our framework on learning representations from limited data with crowdsourced labels, comparing with various state-of-the-art baselines. In addition, we provide a comprehensive analysis on each of the main components of our proposed framework and also introduce the promising results it achieved in our real production to fully understand the proposed framework. To encourage reproducible results, we make our code available online at https://github.com/tal-ai/RECLE .

Legislator Representation Learning with Social Context and Expert Knowledge

PAR: Political Actor Representation Learning with Social Context and Expert Knowledge

Align Voting Behavior with Public Statements for Legislator Representation Learning.

Joint Representation Learning of Legislator and Legislation for Roll Call Prediction

Political Actor Agent: Simulating Legislative System for Roll Call Votes Prediction with Large Language Models

L(u)PIN: LLM-based Political Ideology Nowcasting

UPPAM: A Unified Pre-training Architecture for Political Actor Modeling based on Language

Understanding Political Polarization via Jointly Modeling Users, Connections and Multimodal Contents on Heterogeneous Graphs

Hybrid Representation Learning via Epistemic Graph

Representation Learning in Heterogeneous Professional Social Networks with Ambiguous Social Connections

Unifying Local and Global Knowledge: Empowering Large Language Models As Political Experts with Knowledge Graphs

Representative Social Choice: From Learning Theory to AI Alignment

Simulating The U.S. Senate: An LLM-Driven Agent Approach to Modeling Legislative Behavior and Bipartisanship

Modeling Ideological Salience and Framing in Polarized Online Groups with Graph Neural Networks and Structured Sparsity

Analysis of multiview legislative networks with structured matrix factorization: Does Twitter influence translate to the real world?

How Predictable is Your State? Leveraging Lexical and Contextual Information for Predicting Legislative Floor Action at the State Level

Representation Learning from Limited Educational Data with Crowdsourced Labels

XAI in Computational Linguistics: Understanding Political Leanings in the Slovenian Parliament

A Two-Step Method for Classifying Political Partisanship Using Deep Learning Models

Relational Learning Analysis of Social Politics using Knowledge Graph Embedding

Using Graph-Aware Reinforcement Learning to Identify Winning Strategies in Diplomacy Games (Student Abstract)