LLMs' fanciness and increasing intelligence don't change the reality—ultimately, they're just API calls. As such, they suffer from the same problems as API calls, such as throttling, timeouts, and exponential delays. 🔌 ⚡️
➡️ But you can use clever ways to make LLM calls faster.