Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
119 tokens/sec
GPT-4o
56 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
6 tokens/sec
GPT-4.1 Pro
47 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Self-supervised Hypergraphs for Learning Multiple World Interpretations (2308.07615v2)

Published 15 Aug 2023 in cs.CV

Abstract: We present a method for learning multiple scene representations given a small labeled set, by exploiting the relationships between such representations in the form of a multi-task hypergraph. We also show how we can use the hypergraph to improve a powerful pretrained VisTransformer model without any additional labeled data. In our hypergraph, each node is an interpretation layer (e.g., depth or segmentation) of the scene. Within each hyperedge, one or several input nodes predict the layer at the output node. Thus, each node could be an input node in some hyperedges and an output node in others. In this way, multiple paths can reach the same node, to form ensembles from which we obtain robust pseudolabels, which allow self-supervised learning in the hypergraph. We test different ensemble models and different types of hyperedges and show superior performance to other multi-task graph models in the field. We also introduce Dronescapes, a large video dataset captured with UAVs in different complex real-world scenes, with multiple representations, suitable for multi-task learning.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (8)
  1. Alina Marcu (11 papers)
  2. Mihai Pirvu (3 papers)
  3. Dragos Costea (9 papers)
  4. Emanuela Haller (9 papers)
  5. Emil Slusanschi (4 papers)
  6. Ahmed Nabil Belbachir (8 papers)
  7. Rahul Sukthankar (39 papers)
  8. Marius Leordeanu (47 papers)
Citations (4)

Summary

We haven't generated a summary for this paper yet.