We’re launching Lyria 3.5 in Google Flow Music, with advances across musicality, lyrics, vocals, and creative control
deepmind.google·6d ago

TL;DR
Large language models (LLMs) have faced a computational bottleneck that limits their efficiency and energy use. The startup Subquadratic claims to have developed a new method that reduces the number of computations required by transformers, making LLMs faster and cheaper.
✦ Why It Matters
Engineers can explore Subquadratic's method to enhance LLM efficiency and reduce operational costs.
Key Takeaways
How It Works
Subquadratic's breakthrough involves optimizing the computation process in transformer models, which are foundational to LLMs. By reducing the number of calculations needed to generate responses, the model can operate more efficiently, leading to faster outputs and lower operational costs.
Related