Hacker News
new
top
best
ask
show
job
The Next 3x in Inference Won't Come from Faster Kernels
(
twitter.com
)
2 points
by
floathub
5 hours ago
1 comment
floathub
5 hours ago
One unique idea for solving the chip shortage: pack more models into each existing GPU by having them share the resources. :-)