Deepseek V4 then V4.1 Flash have been my workhorse models for months, both locally & on the APIs (but recently, GLM 5.3f, Mimo 2.6f, & Qwen 3.8f Next have been great alternatives for local hosting)

1302 views Pinned
Nick Antonaccio
Nick AntonaccioAdmin
Aug 30, 2026 at 15:38 (edited, 1 revision)
#21

I'm starting to use GLM 5.3 more lately - it may end up replacing Deepseek V4 Flash for serious local work, although it does require a cluster of at least 2 DGX Sparks.

See https://aibynick.com/thread/72

Nick Antonaccio
Nick AntonaccioAdmin
Sep 04, 2026 at 11:02
#22

This is a great article about optimizing Deepseek V4 Flash performance on 2 clustered DGX Spark machines:

https://www.xda-developers.com/running-284-billion-parameter-model-two-machines-matches-cloud/

Please login to post a reply.

© 2026 AI By Nick.