Swift-1.5-Qwen3.8-27b-oQ8e-mtp achieves 34.8 tokens per second on Apple M5 Max
A Reddit user shared benchmark results for the Swift-1.5-Qwen3.8-27b-oQ8e-mtp model on an Apple M5 Max device on September 29, 2026. The model reached 34.8 tokens per second in a coding scenario, generating a playable game in a single HTML file without edits. The benchmark tool and parameters are detailed for reproducibility, though no official confirmation or context from Qwen or Apple is…
Key points
- Swift-1.5-Qwen3.8-27b-oQ8e-mtp hit **34.8 tokens per second** on Apple M5 Max in a coding task
- Median speed across five runs was **34.5 tokens per second**, with **KV cache** settings affecting results
The story so far
2 episodes →- Swift-1.5-Qwen3.8-27b-oQ8e-mtp achieves 34.8 tokens per second on Apple M5 Maxthis story
Swift-1.5-Qwen3.8-27b-oQ8e-mtp on Apple M5 Max — 34.8 tok/s — llm-bench.io
llm-bench.io · 29 September 2026Loading the full article…
This text was published by llm-bench.io. It is reproduced here with attribution so you can read it in full; the rights remain with the publisher. Read it at the source ↗
Coverage and discussion
1source- Reddit discussionreddit.com
The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us. Published by Martin K., who runs Digest AI and handles corrections.
More in Research
All →- Anthropic’s Frontier Red Team finds GLM-5.3 and Claude Mythos Preview can hijack code · 1 src
- Zhipu automates infrastructure with GLM-5.3’s outer RSI loop in under two weeks · 3 src
- Ohio State and Google launch AI research hub with DeepMind models and Gemini access · 1 src
- General Intuition raises $220M at $6.2B valuation for AI agents trained on gameplay footage · 1 src
- Anthropic’s AI finds a CRISPR-like gene-editing system in phages · 1 src
Comments
via GitHub Discussions