AI情报Sep 18, 2026AI情报
文章
Cactus Compute's Needle 3 claims 8-29MB models can match DeepSeek V4 Flash on downstream tasks after one epoch of fine-tuning.
The company says intelligence laddering yields 2-20 layer subnetworks, suitable for tiny devices.
AI 生成:摘要由 AI 根据所链接的来源撰写。 我们如何使用 AI
AI 生成来源: Cactus Compute
01
来源简报
Cactus Compute's Needle 3 claims 8-29MB models can match DeepSeek V4 Flash on downstream tasks after one epoch of fine-tuning. The company says intelligence laddering yields 2-20 layer subnetworks, suitable for tiny devices.