At the time the whole thing was led by Yann LeCun who seemed to spend more time arguing with people on Twitter than figuring out new techniques to make Llama the best. Meanwhile Deepseek was figuring out large scale RL on kneecapped hardware like H800s and how to scale architectures an order of magnitude bigger with MoE.
Meanwhile Chinese labs were forced to innovate with more efficient models.
He was not in charge of the Meta LLM stuff, from anything i've read over the past few years.
Which brings me to my other point. If you want to compete with serious labs (OpenAI, Anthropic, Deepseek, Alibaba) you need to be smart and focused, not messing around.