dluan an hour ago
jmspring 11 minutes ago
kamranjon 2 hours ago
Anyhow, this kinda reminds me of that quote about architecture: "We replaced our monolith with micro services so that every outage could be more like a murder mystery."
slashdave 36 minutes ago
meander_water 39 minutes ago
tj800x 2 hours ago
hmokiguess an hour ago
https://www.ycombinator.com/companies?query=tracerml
I don't see it?
2 hours ago
Comment deletedAlifatisk 2 hours ago
jmaw 2 hours ago
I have often wondered how tools like GHCP choose the best model for the job when set to "auto".
indiantinker 33 minutes ago
yonatan8070 2 hours ago
As I understand it, in an MoE model, you essentially have hundreds of smaller sub-models ("experts") that are good at different tasks, and for every generated token, a single "master" model chooses which ones are most relevant to participate, and you only activate them.
janalsncm an hour ago
So “1/3 the cost” really depends.
maxdo an hour ago
ninjahawk1 an hour ago
I go to the website…and it’s a sign up. I expected a repo. Otherwise how do I use it? As a SaaS? Yeah right.
Oh well I guess at least the benchmarks are good…I find the benchmarks and many are either not present or are not what the title claims.
My main question is how this has so many updoots from HN, probably the passerby not looking closer for sure.
I mean no offense and I really do wish you best on this, but it seems like what we used to call back in the day, vaporware.
jacobgold an hour ago
But we get ~$2500/mo worth of Fable credits for $200/mo on Anthropic pan? I'm still confused why people (who don't have to use API billing) are chasing open weight models based on cost.
bbstats an hour ago
bnjemian 2 hours ago
While an LLM isn’t what you’d traditionally consider a weak learner, the theorems on learning systems clearly point to them being so in this context. The feigned surprise at combining them to yield better results seems disingenuous.
Even so, the work to predict which models are best suited for which task, how to delegate, and how to combine their outputs is interesting, especially if you’re placing a cost minimization objective on it. That said, this isn’t too far off from what many AI labs are already doing.
wizche 2 hours ago
fneddy 2 hours ago
codekansas 2 hours ago
jambalaya8 2 hours ago
j45 2 hours ago
ototot 2 hours ago
theneocorner 2 hours ago
Comment deleted