There's some speculation that Jev is an open weight model with novel post-training (RLCD). So, if these folks have competitive accuracy with just the base model, it may raise some questions about the necessity of Jev's architecture. You generally don't want to find yourself competing only on price.
> it may raise some questions about the necessity of Jev's architecture
When I hear “architecture” I am thinking number of parameters and latency.
When I hear “accuracy” I think training recipe, data, and (later) number of parameters.
So when you say that Jev’s architecture may not be necessary, the evidence I expect to see is comparable quality at comparable latency. Not equal quality at 2x latency and 4x the cost.
m4y0u · · focus · HN ↗
kylecazar · · focus · HN ↗
Fyi, I haven't tested this yet.
janalsncm · · focus · HN ↗
When I hear “architecture” I am thinking number of parameters and latency.
When I hear “accuracy” I think training recipe, data, and (later) number of parameters.
So when you say that Jev’s architecture may not be necessary, the evidence I expect to see is comparable quality at comparable latency. Not equal quality at 2x latency and 4x the cost.
kylecazar · · focus · HN ↗
Hence my note about the risk of cost being the only moat.