Papers
Topics
Authors
Recent
Gemini 2.5 Flash
Gemini 2.5 Flash
80 tokens/sec
GPT-4o
59 tokens/sec
Gemini 2.5 Pro Pro
43 tokens/sec
o3 Pro
7 tokens/sec
GPT-4.1 Pro
50 tokens/sec
DeepSeek R1 via Azure Pro
28 tokens/sec
2000 character limit reached

Pose-dIVE: Pose-Diversified Augmentation with Diffusion Model for Person Re-Identification (2406.16042v2)

Published 23 Jun 2024 in cs.CV

Abstract: Person re-identification (Re-ID) often faces challenges due to variations in human poses and camera viewpoints, which significantly affect the appearance of individuals across images. Existing datasets frequently lack diversity and scalability in these aspects, hindering the generalization of Re-ID models to new camera systems. We propose Pose-dIVE, a novel data augmentation approach that incorporates sparse and underrepresented human pose and camera viewpoint examples into the training data, addressing the limited diversity in the original training data distribution. Our objective is to augment the training dataset to enable existing Re-ID models to learn features unbiased by human pose and camera viewpoint variations. To achieve this, we leverage the knowledge of pre-trained large-scale diffusion models. By conditioning the diffusion model on both the human pose and camera viewpoint concurrently through the SMPL model, we generate training data with diverse human poses and camera viewpoints. Experimental results demonstrate the effectiveness of our method in addressing human pose bias and enhancing the generalizability of Re-ID models compared to other data augmentation-based Re-ID approaches.

User Edit Pencil Streamline Icon: https://streamlinehq.com
Authors (11)
  1. Inès Hyeonsu Kim (4 papers)
  2. Joungbin Lee (4 papers)
  3. Soowon Son (3 papers)
  4. Woojeong Jin (17 papers)
  5. Kyusun Cho (5 papers)
  6. Junyoung Seo (14 papers)
  7. Min-Seop Kwak (7 papers)
  8. Seokju Cho (19 papers)
  9. Jeongyeol Baek (3 papers)
  10. Byeongwon Lee (4 papers)
  11. Seungryong Kim (103 papers)