1. a model that works for one person/task may not work for another;
2. there are many models (DeepSeek, Qwen, GPT, Claude, Gemini, etc.) that are released every 6 months or so;
3. it takes time to use, test, and evaluate the suitability of a new model and not everyone has an automated evaluation process for their use cases.
Thus, if you find a model that works for you then you are not going to spend more time evaluating a model that may not work, or may only do so when time permits.