HOOSHWARE
Live
LIVE
0
LIVEConnecting to Hooshware intelligence
00:00:00
AI Heat
LOW
Monitoring0 sources

Hooshware

ESTABLISHING NEURAL LINK...
Back to News Radar
DeepSeek AI Released DeepSeek-V4.1-Flash with 1M Context, FP4 KV Cache, and Cross-Layer Attention Reuse
Model Release1w ago

DeepSeek AI Released DeepSeek-V4.1-Flash with 1M Context, FP4 KV Cache, and Cross-Layer Attention Reuse

DeepSeek AI has released DeepSeek-V4.1-Flash, a multimodal Mixture-of-Experts model featuring a 1M-token context window, 552B backbone parameters, 196B Engram…

#DeepSeek#LLM#Mixture-of-Experts#Model Release#AI Agents#AI#Release#Released

Related ecosystem entities

Observed across 2 sources

1 editorial report · 1 verified social mention. The most authoritative report leads while later evidence completes the story.

MarkTechPostPrimary source

DeepSeek AI Released DeepSeek-V4.1-Flash with 1M Context, FP4 KV Cache, and Cross-Layer Attention Reuse

Open report

Social corroboration

telegram1w ago

🐋 DeepSeek just made its AI architecture much cheaper DeepSeek has introduced V4.1-Flash, a new multimodal model designed to deliver more intelligence while using dramatically less compute and memory. The model has 552B parameters, but only 8B are active for input and 16B for output thanks to a new Causal Encoder–Decoder architecture. DeepSeek says its combination of new pre-training and large-scale RL can outperform even its flagship V4-Pro on several benchmarks. The bigger breakthrough may be efficiency. V4.1-Flash uses just 1/4 of the HBM and 1/8 of the SSD storage needed for its previous-generation KV cache. That matters a lot for AI agents, where cached context can become a major part of inference costs. DeepSeek is also cutting prices, with off-peak rates set at 50% of peak pricing. Source. @aipost 🏴

Open mention
BTC
SYNCING
ETH
SYNCING
NVDA
SYNCING
MSFT
SYNCING