Mô hình Mercury 2.5 mới từ Inception có giá 0,04 USD cho mỗi triệu input token và 0,15 USD cho mỗi triệu output token khi phát hành rộng rãi. Mô hình hiện đã có mặt trên OpenRouter...

Đăng nhập để góp ý, chỉnh sửa

Nguồn chính

  1. SOURCE@ollama“DeepSeek-V4.1-Flash is now being rolled out on Ollama's cloud, starting with Max and Team accounts.”x.com
  2. SUPPORT@askalphaxiv“Together, these reduce global KV cache to 890 bytes/token, ~4x smaller than V4-Flash, and persistent KV by ~8x, while supporting 1M token contexts with much stronger overall performance.”x.com
  3. SUPPORT@teortaxestex“DeepSeek V4.1 Flash takes first place on AutomationBench-AA with 69%, equal to GPT-6 Astra (69%) and slightly above Grok 4.6 (67%).”x.com
  4. SUPPORT@ekzhang1“DeepSeek V4.1 has mathematically different outputs depending on whether you prefix cache hit”x.com
  5. SUPPORT@ollama“DeepSeek-V4.1-Flash is now rolling out for Pro plan subscribers.”x.com
  6. SOURCE@askvenice“fully private.”x.com
  7. SOURCE@justinsuntron“Supporting a 1M context window and up to 384K output”x.com
  8. SUPPORT@therundownai“It beats GPT-5.6 Sol and Claude Opus 5 on several agentic, coding, and cyber benchmarks.”x.com
Bản Markdown