Also if anyone ever says the works "maximizes my comprehension/sec" in real life I will absolutely punch them in the face for no reason other than the betterment of the universe.
this auto-routing vs Hierarchical Reasoning Model? Frankly, I am not expert and not sure if it's more relevant to deployed model efficiency per unit of intelligence AND/OR a faster RL w/ algo-Abstraction embedded. (the obvious con is that it's "just" sudoku and mazes, but that would be a bit US/Mag7 biased/vested considering how quickly it learns vs SOTA).
Beautifully written and thoughtful.
Also if anyone ever says the works "maximizes my comprehension/sec" in real life I will absolutely punch them in the face for no reason other than the betterment of the universe.
this auto-routing vs Hierarchical Reasoning Model? Frankly, I am not expert and not sure if it's more relevant to deployed model efficiency per unit of intelligence AND/OR a faster RL w/ algo-Abstraction embedded. (the obvious con is that it's "just" sudoku and mazes, but that would be a bit US/Mag7 biased/vested considering how quickly it learns vs SOTA).
https://www.sapient.inc/blog/5