Forecast how long an LLM reply will run, before Enter
Token Forecaster puts a range on an LLM reply before you press Enter: the usual length, and a worst case that held for 90.6% of 4,146 unseen calls. While the reply streams, it shows whether it is running long. It only watches and never changes the request. Learns from your local history. Open source, MIT.
Token Forecaster launched on Product Hunt on 2026-09-25 and has 65 upvotes and 2 comments so far.
This tool is open source.