Switch language한국어
Back to the list

LaCoVL-FER: Landmark-Guided Contrastive Learning Network with Vision-Language Enhancement for Facial Expression Recognition

TL;DR AI

Key summary

2 min read
  1. Researchers introduced LaCoVL-FER, a facial expression recognition model that combines facial landmarks with CLIP-based vision-language features.

  2. The method uses landmark-guided adaptive encoding and vision-language enhancement to fuse geometric and semantic cues.

  3. On RAF-DB, FERPlus, and AffectNet, the model delivered state-of-the-art performance.

  4. The approach is designed to improve robustness under pose changes, occlusion, and lighting variation.

Read the original