DeepSeek v4.1 Flash(twitter.com)
720 points by Liwink 9 hours ago | 378 comments
tl;dr: DeepSeek has announced DeepSeek-V4.1-Flash, the smallest model in a new architecture family featuring native visual understanding. The company claims improvements in capability, inference speed, and throughput, with the architecture designed to scale to larger models.
HN Discussion:
  • DeepSeek's transparency and technical detail is refreshing compared to safety-focused competitors
  • Admiration for DeepSeek's fearless innovation and clever research ideas at frontier scale
  • Model size has grown too large to reasonably run locally, undermining the 'flash' label
  • Pricing has increased and may not be justified by performance versus competitors
  • Practical strengths noted: strong performance, fewer refusals, and useful as a backup model