undefined | Better HN

0 pointsJumpCrisscross10d ago0 comments

I guess so. 4.8 + 4.8 > Fable 5 is interesting, though not particularly game changing. (The others all fuse frontier models. Which is an argument for using those frontier models more. Not less.)

0 comments

1 comments · 1 top-level

pants210d ago

Yeah, all that's really saying is a weaker model with a better harness can beat a stronger model with a worse harness, specifically on the DRACO benchmark

This isn't really a surprising result. Needs more evidence to make a broader claim.

j / k navigate · click thread line to collapse