Hugging Face Blog
Releasing Swift Transformers: Run On-Device LLMs in Apple Devices
Read the original on Hugging Face Blog →The Flow has not summarised this story yet — read it at Hugging Face Blog.
The Flow has not summarised this story yet — read it at Hugging Face Blog.
arXiv:2607. 00501v1 Announce Type: cross Abstract: We present BaseRT, a native Metal inference runtime for large language models (LLMs) on Apple Silicon, and report the highest inference throughput on this hardware to date.
A Round Up And Comparison of 10 Open-Weight LLM Releases in Spring 2026