NDTVCenter·
IPO-Bound DeepSeek Launches V4.1-Flash As Smallest Model In New AI Architecture
The company claimed that it had greatly reduced inference costs for AI agents by reducing this memory footprint to 890 bytes per token from 3,514 bytes in the prior Flash version.
Read at NDTV →




