---
format: "aidr-story-markdown/v1"
id: "2655af8eb0bc114b5bc4b3a5da09da4d508b4569e0bf29a7c1484a8026cfc7b9"
canonical_url: "https://aidr.today/2655af8e?lang=vi"
title: "DFlash 2 đạt tốc độ 70 Tok/s trên MacBook Pro M5 Max, nhanh hơn gấp 4,6 lần"
lang: "vi"
requested_lang: "vi"
available_langs: ["en","vi"]
translation_fallback: null
fallback_fields: []
published_at: "2026-08-18T23:50:02.000Z"
category: "Infra"
topics: ["qwen","inference","apple","hardware"]
source_urls: ["https://huggingnews.com/ai/dflash-2-hits-70-toks-on-m5-max-macbook-pro-in-46x-speed-jump-01fefd72"]
summary: "Inco AI và Z Lab vừa ra mắt engine suy luận mới giúp tăng tốc chạy các LLM trên phần cứng của Apple. Với DFlash 2, mô hình Qwen3.8-27B có thể đạt được..."
---

# DFlash 2 đạt tốc độ 70 Tok/s trên MacBook Pro M5 Max, nhanh hơn gấp 4,6 lần

> [Open the canonical story](<https://aidr.today/2655af8e?lang=vi>)

**Published:** 2026-08-18T23:50:02.000Z
**Category:** Infra
**Topics:** qwen, inference, apple, hardware

## Summary

Inco AI và Z Lab vừa ra mắt engine suy luận mới giúp tăng tốc chạy các LLM trên phần cứng của Apple\. Với DFlash 2, mô hình Qwen3\.8\-27B có thể đạt được\.\.\.

## Sources

- [Story source](<https://huggingnews.com/ai/dflash-2-hits-70-toks-on-m5-max-macbook-pro-in-46x-speed-jump-01fefd72>)

