DFlash Team Advances into Mac; 27B Model Achieves 144 Tokens per Second
2026-09-19 16:25:47
According to CoinMeta, the DFlash team announced their entry into the Mac platform, where the 27B model achieved a peak performance of 144 token /s on M5 Max MacBook Pro. The team's open-source local inference engine, Splash, has been integrated into LM Studio version 0.4.25 and supports Qwen3.8-27B and Qwen3.6-35B-a3b models. By processing multiple token in parallel, the overall performance has been improved by reducing the computation required for individual generation. During horizontal testing with 48GB of M5 Pro, the single-threaded short context speed reached 74 token /s, and with four-way concurrency, the total throughput reached 170 token /s, which is 3.9 times that of the second-place competitor. Splash is open-source and not tied to any particular LM Studio; it requires a M3 or newer Mac, macOS with a version above 26.4, and at least 36GB of unified memory.
Source:Internet
This content is for market information only and does not constitute investment advice.
Follow WalletJYS official accounts to stay updated

Hot Articles
Refresh

Legit Bitcoin Trading Apps 2026: Top 4 Safe & Regulated Picks
09-18 18:52

Zcash Jumps 23% After Fed Hike, Beats Bitcoin: How Far Can It Go?
09-17 18:03

Which Crypto Wallet Is Best? 2026 Ranking & Review
09-16 18:34

What Is CASHCAT Coin? Why It Leads On-Chain Despite 26% Weekly Drop
09-15 16:04

Top 10 Crypto Trading Apps Sept 2026: Ranking & Review
09-14 17:24



