跳到正文
r/LocalLLaMA· /u/-dysangel-·· 8 小时前AI 评分22

用 GLM 5.3 Flash 在双 DGX Sparks 上全本地跑出跑酷小游戏

Fully local little parkour sim

AI 导读

用户用 GLM 5.3 Flash 在 2 台 DGX Sparks 上全本地搭建了一个跑酷小游戏,vLLM 采用 TP2 配置,prefill 约 1500t/s、decode 约 40t/s @ 100k。以 Claude Code 作为脚手架,上下文 260k,作者称该模型编码能力介于 GLM 5.1 与 5.3 之间,视觉与 3D 理解不错,交互速度稳定。

正文
Fully local little parkour sim

I vibed this up this weekend, fully local, with GLM 5.3 Flash running on 2x DGX Sparks.

vllm TP2 recipe: https://github.com/tonyd2wild/GLM-5.3-Flash-NVFP4-DFlash2-2x-DGX-Spark

Prefill: ~1500t/s
Decode: ~40t/s @ 100k

Using Claude Code as the scaffold with 260k context size.

I'm really impressed with this model. Feels somewhere between GLM 5.1 and 5.3 in terms of coding depending on the task. Good vision and 3D understanding. Solid interactive speeds. I feel like I've finally reached a "good enough" setup at home, and looking forward to things only getting better from here.

submitted by /u/-dysangel-
[link] [留言]

来源:r/LocalLLaMA · reddit.com