Tag
#inference speed
3 stories taggedinference speed.

Frontier Labs
Liquid AI's New Draft Models Make Its LFMs Up to 3.2x Faster, No Quality Trade-off
A technique called speculative decoding lets a small helper model do the heavy lifting so the main model just checks the work. The result: dramatically faster output on everything from a data-centre GPU to a MacBook.
4 min read

Explained
Apple's AI Shortcut: How a 'Draft and Check' Trick Makes Reasoning Models Twice as Fast
Apple ML Research has built a smarter way to speed up AI thinking, one that checks meaning instead of counting exact words. It could cut the cost of running powerful AI in half.
3 min read

Science & Space
Apple Researchers Found a Smarter Way to Make AI Video Generation Much Faster
A new technique from Apple ML Research cuts the time it takes AI to turn text prompts into video, by teaching the model to skip calculations it doesn't actually need.
3 min read