The biggest change: we integrated model layer streaming across all local inference pipelines, cutting peak VRAM usage enough to run on 16 GB VRAM machines. This has been one of the most requested changes since launch, and it’s live now.
What else is in 1.0.3:
The VRAM reduction is the one we’re most excited about. The higher VRAM requirement locked out a lot of capable desktop hardware. If your GPU kept you on the sideline, try it now and let us know how it works for you on GitHub.
Already using Desktop? The update downloads automatically.
New here? Download
submitted by /u/ltx_model
[link] [comments]
Model Context Protocol (MCP) servers allow foundation models to access external data and tools, supporting…
It’s one of the best times of the year to buy a mattress, and our…
Physicists have demonstrated a new way to entangle distant quantum bits without the constant measurements…
Cats, dogs, wolves and other agile four-legged animals can easily jump and squeeze through tight…
For reference here's the post about REFMOD by it's creator (u/LuisaPinguinnn). I basically used an…
After Donald Trump’s executive order demanding the name change, Google is the first major online…