Model should be able to understand where logical sentence ends, to stop buffering, and optionally rewrite some of the test that has already been output.
Model should be able to understand where logical sentence ends, to stop buffering, and optionally rewrite some of the test that has already been output.
IIRC that is exactly how Dragon Naturally speaking did it decades ago.