Thứ Tư, 19 thg 8, 2026

Inco AI vừa trình làng DFlash 2, giải pháp mới giúp đẩy nhanh tốc độ inference AI ngay trên thiết bị cục bộ lên tới 4,6 lần so với thế hệ trước.

Đăng nhập để góp ý, chỉnh sửa
Nguồn chính
SOURCE@lliebenwein23:16 18 thg 8one of many innovations powering the inference stack at @inco_ai
SUPPORT@xianbao_qian11:35 19 thg 8DFlash 2 is available on @vllm_project
SUPPORT@nielsrogge10:53 19 thg 8I've made the write-up available on Papers with Code for anyone to learn more
SOURCE@zhijianliu_22:08 18 thg 8seeded at Z Lab and upgraded at Inco AI
SUPPORT@jun_song23:00 18 thg 8faster than Fable or Sol
SOURCE@elliotarledge8:40 19 thg 8Up to 4.6× the speed of autoregressive decoding, with the same output.
SOURCEhuggingnews