Hacker News
new
|
past
|
comments
|
ask
|
show
|
jobs
|
submit
login
esquire_900
4 months ago
|
parent
|
context
|
favorite
| on:
Mamba-3
This is sort of what their first sentence states? Except your line implies that they are fast in training and inference, they imply they are focusing on inference and are dropping training speed for it.
It's a nice opening as it is imo
cubefox
4 months ago
[–]
They don't say anything about dropping training speed.
estearum
4 months ago
|
parent
[–]
> a departure from Mamba-2, which optimized for training speed.
?
cubefox
4 months ago
|
root
|
parent
[–]
Yes? Mamba-2 optimized for training speed compared to Mamba-1. Mamba-3 adds optimization for inference. These are pretty much version numbers.
Consider applying for YC's Fall 2026 batch!
Applications
are open till July 27.
Guidelines
|
FAQ
|
Lists
|
API
|
Security
|
Legal
|
Apply to YC
|
Contact
Search:
It's a nice opening as it is imo