⬢github Python · 6.4K ★ +8 since we first saw it · pushed 4 d ago · Apache-2.0
mizorewww/laya-mlx
Native MLX runtime for Laya typed decision models — 7–14 ms short decisions on M3 Max. No text generation, PyTorch, or cloud API.
Laya-MLX is a native Apple Silicon (MLX) runtime for Laya typed decision models: instead of generating text, it returns structured answers (e.g. picking a department from a choice list) for short English questions. It runs fully locally with no PyTorch, Transformers, or cloud API, achieving ~13 ms median latency (7.4 ms multilingual) on an M3 Max, with a pip-installable Python API and a Snake-game demo.
Why now: It's trending among newly starred repos, likely due to the low-latency local typed-decision demos (Snake at 75 moves/s) and benchmark claims versus text-generation approaches.
Who it is for: Developers on Apple Silicon Macs who need fast, local, structured classification/routing decisions without LLM text generation.
apple-silicondecision-modelinferencelayalocal-aimachine-learningmlxmodernbertsystem-onetyped-decisions
Stars over our 24 snapshots: 6.4K to 6.4K, since 4 h ago.
Where people talked about it
- ⬢github new repos, most starred 5 min ago
API: https://socialmediatrends-api.osmike.com/v1/repos/mizorewww/laya-mlx