OpenAI, Google, Anthropic admit they can’t scale up their chatbots any further

MajorHavoc@programming.dev · edit-2 30 days ago

OpenAI, Google, Anthropic admit they can’t scale up their chatbots any further

31337@sh.itjust.works · 27 days ago

Hmm. I just assumed 14B was distilled from 72B, because that’s what I thought llama was doing, and that would just make sense. On further research it’s not clear if llama did the traditional teacher method or just trained the smaller models on synthetic data generated from a large model. I suppose training smaller models on a larger amount of data generated by larger models is similar though. It does seem like Qwen was also trained on synthetic data, because it sometimes thinks it’s Claude, lol.

Thanks for the tip on Medius. Just tried it out, and it does seem better than Qwen 14B.

OpenAI, Google, Anthropic admit they can’t scale up their chatbots any further

OpenAI, Google, Anthropic admit they can’t scale up their chatbots any further

OpenAI, Google, Anthropic admit they can’t scale up their chatbots any further – Pivot to AI