NVIDIA AI 指出可在 DGX Spark 上运行 DeepSeek V4 Flash 模型,达到 1000 tok/s 预填充和 59 tok/s 多智能体服务,一条命令安装 · AIWatch