Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
c0rruptbytes
17 days ago
|
parent
|
context
|
favorite
| on:
Show HN: Running 104GB Qwen3.8-Flash-Next on 48GB ...
so many inference project, omlx already supports all of this and has a 1000 people trying to optimize it constantly
carloslfu
17 days ago
|
next
[–]
Both projects are different in scope. Think of slotstream as optimizing for memory and for this specific model for now, my intention is not to build an inference engine the same as oMLX
carloslfu
17 days ago
|
prev
[–]
Interesting! I'll check it out
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search: