Cool idea. There are a lot of interesting harness designs that are starting to emerge around tool calling and code execution. RLM is one of them.
But so is this Speculative Programmatic Tool Calling approach (from the same author of RLM).
A general class of technique for speculating on tool calls during code generation in a harness and queuing them early to overlap with token generation + REPL execution time.
Post #4432
621