A curated reading list of research in Adaptive Computation, Inference-Time Computation & Mixture of Experts (MoE).
Exploration into the proposed "Self Reasoning Tokens" by Felipe Bonetto
Yet another random morning idea to be quickly tried and architecture shared if it works; to allow the transformer to pause for any amount of time on any token