Hacker News
new
top
best
ask
show
job
Show HN: mlxsh A lightweight CLI to serve multiple local LLMs on Apple
(
github.com
)
2 points
by
ansnadeem
3 hours ago
1 comment
nittanymount
an hour ago
machine memory is the bottleneck/limit for running local models, how is the memory handled? did not find notes on that on the repo readme page. could you explain, thanks.