Hacker Newsnew | past | comments | ask | show | jobs | submitlogin

I have been using open source LLMs locally for the past 5 months like Qwen, Llama, DeepSeek and others, and I have also noticed that current local models feel significantly dumber than closed source commercial models like ChatGPT, Claude, and Gemini. One main reason I think is the amount, variety, and quality of original authentic data on which they are being trained on, and also the training method plays a significant role in the performance difference between local open source models and closed source commercial models.


They are also an order of magnitude difference in number of parameters and that matters a lot.




Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: