What open-source LLMs are you using in 2024?

@Blaed · 1 year ago

What open-source LLMs are you using in 2024?

@xodoh74984 · 1 year ago

This one is only 7B parameters, but it punches far above its weight for such a little model:
https://huggingface.co/berkeley-nest/Starling-LM-7B-alpha

My personal setup is capable of running larger models, but for everyday use like summarization and brainstorming, I find myself coming back to Starling the most. Since it’s so small, it runs inference blazing fast on my hardware. I don’t rely on it for writing code. Deepseek-Coder-33B is my pick for that.

Others have said Starling’s overall performance rivals LLaMA 70B. YMMV.

@Blaed · 1 year ago

What sort of tokens per second are you seeing with your hardware? Mind sharing some notes on what you’re running there? Super curious!