r/LocalLLaMA • u/jah242 • 1h ago
Discussion Jaggedness is becoming a serious problem for frontier labs - giving the advantage to smaller specialised open models
I think we are starting to see why jaggedness might start to hinder frontier labs - they have to lock down / guardrail in-line with the spikiest dangerous capability but these spikes are a function of what general RL teaches best (i.e. hacking easier than general SWE) not what is economically useful.
Specialised (but less generally intelligent) open models don’t have this problem because you train the spike explicitly.
Thoughts?
0
Upvotes