Skip to content
r/LocalLLaMA

Apple Announced New On-Device Inference Engine for Apple Silicon

Apple announced new on device inference engine for Apple Silicon

Apple announced CoreAI at WWDC as a future CoreML replacement for Apple Silicon on-device inference; models require Python-script conversion, the supported list is mostly mid-2025 models, and the post does not disclose performance data.

Why it matters: HKR-H/K/R pass, but the post is thin: CoreAI, CoreML successor status, and Python conversion are disclosed; throughput, latency, and model coverage are not. Apple on-device inference merits featured, capped in the 72–77 band.

Read the original ↗Export Markdown