Inputs
When the cache does not fit, the model recomputes the prefix instead of evicting the other branch’s slot, so an oversized setting degrades speed rather than failing the run.
Outputs
This documentation was AI-generated. If you find any errors or have suggestions for improvement, please feel free to contribute! Edit on GitHub
Source fingerprint (SHA-256):
0c10cdb465d1ee4063273ffbb4913def3830f7e329694cd0a2e292d6f3c37ae4