Needle 2 Delivers 14MB LLM for Phones and Wearables

TL;DR. Needle 2, an open 45M-parameter LLM, runs as a 14MB binary on mobile, smart home, and robotics devices. - The model uses 28MB of RAM and competes with larger models like FunctionGemma 270M and Apple FM. - Needle 2 achieves up to 500 tokens/sec decode speed on a Raspberry Pi 5 and runs on low-cost hardware. - Its design focuses on tool calling and structured extraction, making it suitable for resource-constrained edge AI.

Sources

Back to QLANKR News