Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
|
lovelydata's comments
login
lovelydata
10 months ago
|
parent
|
context
[–]
| on:
Ask HN: Who uses open LLMs and coding assistants l...
llama.cpp + Qwen3-4B running on older PC with AMD Radeon GPU (Vulcan). Users connect via web UI. Usually around 30 tokens/sec. Usable.
NicoJuicy
10 months ago
|
parent
[–]
What do they use it for? It's a very small model
embedding-shape
10 months ago
|
root
|
parent
[–]
Autocomplete words, I'd wager, as yeah, super tiny model that can barely output coherent output in many cases.
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: