18 models · checked 20 September 2026
Quantised decision model builds
Conversions of an existing model into a format you can run locally. Same weights, different packaging.
rlcd-modernbert-151m
heman10x
The smallest decision model we track, at 151M parameters on a GLiClass/ModernBERT encoder. If it holds accuracy, it changes what this class of model costs to run.
151M · knowledgator/gliclass-modern-base-v2.0 · apache-2.0
Community cardlaya-onnx
Mattepiu
A ONNX build of laya, for running typed decisions on runtimes without PyTorch instead of a hosted API.
convaiinnovations/laya · apache-2.0
Community carddecider-2b-GGUF
cosetoenor
A GGUF build of decider-2b, for running typed decisions on llama.cpp instead of a hosted API.
Mapika/decider-2b · apache-2.0
Community carddecider-2b-vision-GGUF
mradermacher
A GGUF build of decider-2b-vision, for running typed decisions on llama.cpp instead of a hosted API.
Mapika/decider-2b-vision · apache-2.0
Community cardlaya-multilingual-coreml
aac6fef
A Core ML build of laya-multilingual, for running typed decisions on on-device Apple runtimes instead of a hosted API.
convaiinnovations/laya-multilingual · apache-2.0
Community cardlaya-GGUF
mys
A GGUF build of laya, for running typed decisions on llama.cpp instead of a hosted API.
convaiinnovations/laya · apache-2.0
Community cardlaya-multilingual-GGUF
mys
A GGUF build of laya-multilingual, for running typed decisions on llama.cpp instead of a hosted API.
convaiinnovations/laya-multilingual · apache-2.0
Community cardlaya-onnx-fp16
sevenreasons
A ONNX build of laya, for running typed decisions on runtimes without PyTorch instead of a hosted API.
convaiinnovations/laya · apache-2.0
Community carddecider-0.8b-GGUF
mradermacher
A GGUF build of decider-0.8b, for running typed decisions on llama.cpp instead of a hosted API.
Mapika/decider-0.8b · apache-2.0
Community cardgemma-e2b-rlcd
larkooo
A MLX build of gemma-4-E2B-it, for running typed decisions on Apple Silicon instead of a hosted API.
5.1B · google/gemma-4-E2B-it · apache-2.0
Community cardlaya-coreml
aac6fef
A Core ML build of laya, for running typed decisions on on-device Apple runtimes instead of a hosted API.
convaiinnovations/laya · apache-2.0
Community cardlaya-typed-decisions-GGUF
mys
A GGUF build of laya-typed-decisions, for running typed decisions on llama.cpp instead of a hosted API.
convaiinnovations/laya-typed-decisions · apache-2.0
Community cardMobiMind-Decider-7B-GGUF
mradermacher
A GGUF build of MobiMind-Decider-7B, for running typed decisions on llama.cpp instead of a hosted API.
IPADS-SAI/MobiMind-Decider-7B · apache-2.0
Community cardlaya-multilingual-coreml-ane
aac6fef
A Core ML build of laya-multilingual, for running typed decisions on on-device Apple runtimes instead of a hosted API.
convaiinnovations/laya-multilingual · apache-2.0
Community cardlaya-multilingual-coreml-ane-w8
aac6fef
A Core ML build of laya-multilingual, for running typed decisions on on-device Apple runtimes instead of a hosted API.
convaiinnovations/laya-multilingual · apache-2.0
Community cardlaya-typed-decisions-coreml
aac6fef
A Core ML build of laya-typed-decisions, for running typed decisions on on-device Apple runtimes instead of a hosted API.
convaiinnovations/laya-typed-decisions · apache-2.0
Community cardlaya-multilingual-coreml-snake
aac6fef
A Core ML build of laya-multilingual, for running typed decisions on on-device Apple runtimes instead of a hosted API.
convaiinnovations/laya-multilingual · apache-2.0
Community carddecider-35b-a3b-nvfp4
Mapika
A typed-decision model from Mapika, 18B parameters.
18B · Mapika/decider-35b-a3b · apache-2.0
Find out what your data actually supports
We turn your historical decisions into a frozen test set, then tell you which model, which threshold, and how much of it you can safely automate.