Switch language한국어
Back to the list

Geometry Conflict: Explaining and Controlling Forgetting in LLM Continual Post-Training

TL;DR AI

Key summary

2 min read
  1. Researchers propose a geometry-based view of catastrophic forgetting in continually post-trained LLMs, arguing that update conflicts are relative to the model’s current state.

  2. They validate the idea on Qwen3 models and introduce GCWM, which uses geometry conflict to gate Wasserstein merge corrections.

  3. The method improves performance in continual learning settings and helps preserve prior capabilities without storing replay data.

Read the original