Supported
Executed by AI Stays Local on our own hardware and accepted. Evidence is published on the benchmarks page.
Models
A model is supported when AI Stays Local has executed it and accepted the result. An adapter existing in an upstream engine is not support; it is an adapter. Every entry states total and active parameters separately, because they are different numbers.
Executed by AI Stays Local on our own hardware and accepted. Evidence is published on the benchmarks page.
Being brought up and measured now. It is not a capability of the product, and it may not survive validation.
Intended for a future release. Nothing has been run yet.
Demonstrated by an upstream project, not reproduced by AI Stays Local. Evidence that the physics works, not evidence that we can do it.
LimitationsA small model. Strong on everyday drafting, summarising and questions about your own documents; it will lose to a large hosted model on hard reasoning.
LimitationsUnder validation on GPU. It is not operational, we do not claim it, and no third-party benchmark for it is published here as ours.
LimitationsUpstream research evidence that expert streaming scales. It needs a great deal of local storage, reported generation is far slower than interactive cloud chat, and it is not a capability of this product.
The upstream entry is included because it shows the approach is physically possible, not because it is something this product does. It has not been reproduced by AI Stays Local on any hardware.