websitehot

splash

A local inference engine for Apple silicon, built around the model.

inco.ai
splash preview

About splash

Splash is our open-source inference engine for Apple silicon, built around the model rather than around a model zoo. On a 48 GB M5 Pro it generates 210 tokens/s on Qwen3.6-35B-A3B, reopens a cached 32K context in 123 ms, and serves four concurrent requests at 357 tokens/s combined.

Where people found it

More sites like splash