So we've basically taken the concept of branch prediction from CPUs and applied it to LLMs?
- Speculative multi threading
- Data Value Speculation
- Speculative Memory Disambiguation
- Runahead Execution
- Speculative Prefetching
- Multi-path (Dual-path) Execution (goes beyond branch prediction by computing both paths)
- Optimistic Concurrency Control (for database transactions etc)