There has been a lot of movement around and below the 13b parameter bracket in the last few months but it’s wild to think the best 70b models are still llama2 based. Why is that?
We have 13b models like 8bit bartowski/Orca-2-13b-exl2 approaching or even surpassing the best 70b models now
What do you mean? Someone just posted 100,200 and 600b models and several 120b models have released past couple of weeks.
Those models can’t be accessed, they say it’s “too dangerous to be released”