CrewCrew
FeedSignalsMy Subscriptions
Get Started
Small and On-Device Models: Phi, Gemma, Apple

Small and On-Device Models: Phi, Gemma, Apple — 2026-09-04

  1. Signals
  2. /
  3. Small and On-Device Models: Phi, Gemma, Apple

Small and On-Device Models: Phi, Gemma, Apple — 2026-09-04

Small and On-Device Models: Phi, Gemma, Apple|September 4, 2026(2h ago)2 min read8.4AI quality score — automatically evaluated based on accuracy, depth, and source quality
0 subscribers

This week, Nvidia unveiled the RTX Spark "Superchip" at IFA 2026, marking the arrival of dedicated AI PCs capable of running local models with high efficiency. Simultaneously, Samsung confirmed the Galaxy S26 FE will feature the Exynos 2500 chipset and enhanced on-device AI tools, while South Korea finalized its K-On Device AI semiconductor consortium.

Small and On-Device Models: Phi, Gemma, Apple — 2026-09-04


Top developments


Nvidia RTX Spark "Superchip" Debuts at IFA 2026

Nvidia and its partners showcased the first laptops and mini PCs powered by the new RTX Spark "Superchip" at IFA 2026. These devices are specifically engineered to run AI models directly on the computer, reducing reliance on cloud APIs for local inference tasks. This hardware launch complements the recent trend of small language models like Microsoft's Phi-4-mini and Google's Gemma 3 becoming viable on consumer-grade NPUs and GPUs.

Nvidia RTX Spark AI PC
Nvidia RTX Spark AI PC


Samsung Galaxy S26 FE Launches with Exynos 2500 and On-Device AI

Samsung officially launched the Galaxy S26 FE, featuring One UI 9 and a suite of new AI-powered tools including Photo Assist and Horizon Lock. The device is expected to utilize the Exynos 2500 chipset, signaling Samsung's continued commitment to optimizing on-device AI performance for its mid-range flagship line. This move aligns with broader industry efforts to integrate efficient small language models directly into smartphone hardware for faster, privacy-focused processing.

Samsung Galaxy S26 FE
Samsung Galaxy S26 FE


South Korea Finalizes K-On Device AI Semiconductor Consortium

After months of delays due to budget disputes between ministries, South Korea has finalized the final partners for its national "K-On Device AI Semiconductor" project. The consortium, which includes notable players like Mobilit and HyperExcel, aims to develop specialized chips for edge AI applications. This government-backed initiative underscores the strategic importance of localizing AI hardware production to support domestic on-device AI ecosystems.


Local view

South Korean media outlets have closely followed the finalization of the K-On Device AI project, with ZDNet Korea highlighting the selection of key consortium members after a period of uncertainty. Additionally, user feedback on Samsung’s community forums indicates growing scrutiny of UI clarity regarding "on-device only" data processing settings, reflecting heightened consumer awareness of privacy features in on-device AI implementations.


Context & numbers

The push for on-device AI is supported by competitive performance metrics from small language models. Recent benchmarks indicate that Microsoft’s Phi-4-mini (3.8B parameters) achieves an 83.7% score on ARC-C, a top result in its size class, while Google’s Gemma 3 4B posts an 89.2% on GSM8K math reasoning. These figures demonstrate that sub-10B models are increasingly capable of handling complex tasks locally, driving demand for hardware like the newly announced Nvidia RTX Spark.


On the radar

  • DeepX NPU Expansion: South Korean startup DeepX reported 77 global orders in its first year of mass production for its ultra-low-power DX-M1 NPU, indicating strong early adoption for edge AI devices.
  • Hugging Face Microduck Robot: The small robot "Microduck," co-developed by Hugging Face and Pollen Robotics, has seen explosive initial sales, featuring a Chinese Rockchip NPU, highlighting the integration of small models into consumer robotics.

This content was collected, curated, and summarized entirely by AI — including how and what to gather. It may contain inaccuracies. Crew does not guarantee the accuracy of any information presented here. Always verify facts on your own before acting on them. Crew assumes no legal liability for any consequences arising from reliance on this content.

Explore related topics
  • QHow much does the Nvidia RTX Spark cost?
  • QWhat are the Exynos 2500 AI benchmarks?
  • QWhich companies joined the K-On consortium?

Powered by

CrewCrew

Sources

Want your own AI intelligence feed?

Create custom signals on any topic. AI curates and delivers 24/7.