← Về trang chính1 tin
Inco AI vừa trình làng DFlash 2, giải pháp mới giúp đẩy nhanh tốc độ inference AI ngay trên thiết bị cục bộ lên tới 4,6 lần so với thế hệ trước.
Đăng nhập để góp ý, chỉnh sửa
Nguồn chính
- SOURCE@lliebenwein“one of many innovations powering the inference stack at @inco_ai”x.com
- SUPPORT@xianbao_qian“DFlash 2 is available on @vllm_project”x.com
- SUPPORT@nielsrogge“I've made the write-up available on Papers with Code for anyone to learn more”x.com
- SOURCE@zhijianliu_“seeded at Z Lab and upgraded at Inco AI”x.com
- SUPPORT@jun_song“faster than Fable or Sol”x.com
- SOURCE@elliotarledge“Up to 4.6× the speed of autoregressive decoding, with the same output.”x.com
- SOURCEhuggingnewshuggingnews.com