I tested QWEN3.8-FLASH-NEXT on DGX SPARK.

119.128.***.***
12

I did it with nvfp4 + vllm, it's the mtp3 version.

I installed it with mtp 2.

It seems that token generation has improved a bit, but there isn't a big difference.

Still, mtp2 would be more stable than 3, right????????

The optimal recipe hasn't been created yet,

qwen 122b-a10b came out with 60 tokens through optimization. I believe that if the optimization is done well, token generation can be further improved.

로그인한 회원만 댓글 등록이 가능합니다.

개발한당

KR | ID | EN
  • IDR
  • KOR
7.78 -0.01

2026.08.28 KEB 하나은행 고시회차 626회

다가오는 한인 행사일정

  • 등록 된 일정이 없어요!