The biggest change: we integrated model layer streaming across all local inference pipelines, cutting peak VRAM usage enough to run on 16 GB VRAM machines. This has been one of the most requested changes since launch, and it’s live now.
What else is in 1.0.3:
The VRAM reduction is the one we’re most excited about. The higher VRAM requirement locked out a lot of capable desktop hardware. If your GPU kept you on the sideline, try it now and let us know how it works for you on GitHub.
Already using Desktop? The update downloads automatically.
New here? Download
submitted by /u/ltx_model
[link] [comments]
Directly added to your Samsung Wallet account, it’s yet another cash-back credit card, this time…
A new modeling study finds that weak AI regulation may be worse than no regulation…
Large language models (LLMs), the computational algorithms underpinning ChatGPT, Gemini and other artificial intelligence (AI)-powered…
Our remote team clocked serious hours walking, working, and sometimes jogging to find the best…
Another powerful new artificial intelligence model from China took the U.S. tech industry by surprise…
In this article, you will learn what prompt injection and tool misuse are in the…