Laya, a 322M-parameter decision engine, gains 26,639 GitHub stars in nine days
Laya, an open-source decision engine released on September 18, 2026, has rapidly gained traction with 26,639 GitHub stars and 2,326 forks within nine days. Developed by NandhaKishorM and licensed under Apache-2.0, the tool offers a non-autoregressive alternative to using large language models for simple classification tasks. Instead of generating text tokens, Laya performs a single forward pass…
Key points
- Laya gained 26,639 GitHub stars and 2,326 forks in nine days since its September 18, 2026 release.
- The 322M-parameter multilingual model answers typed questions in one forward pass with zero output tokens.
- Docs warn checkpoints are over-confident; users must fit temperatures on held-out data to trust probabilities.
The project ships three checkpoints on Hugging Face, including a 421M-parameter English model and a 322M-parameter multilingual model supporting over 100 languages. The repository claims a latency of 33ms per question on a T4 GPU, dropping to 7.2ms when batched. A hands-on test on a CPU-only VM confirmed that the system runs locally without API keys or GPU requirements, though CPU inference is slower at roughly 0.7 seconds per question. The tool includes a built-in abstention mechanism that flags low-confidence answers, allowing developers to route uncertain cases to human review or larger models.
While the speed and cost efficiency are significant, the documentation warns that the shipped checkpoints are over-confident and require users to fit temperature parameters on their own data before relying on the probability scores. Laya is positioned for high-volume, bounded decisions like ticket triage and moderation, where deterministic, fast responses are preferred over the nuanced reasoning of frontier models.
Model page: Laya →
Laya: replace LLM-as-a-judge with a 322M-parameter decision engine (26,639 stars in 9 days, hands-on test)
aifrontierpost.com · 27 September 2026
Loading the full article…
This text was published by aifrontierpost.com and written by Marcus Doyle. It is reproduced here with attribution so you can read it in full; the rights remain with the publisher. Read it at the source ↗
Coverage and discussion
1source- Reddit discussionreddit.com
The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us. Published by Martin K., who runs Digest AI and handles corrections.
More in Research
All →- Simon Willison reviews 2026 LLM trends including Claude Opus 4.5, GPT-5.1, and coding agents · 1 src
- Anthropic opens biology lab to test AI-driven scientific discovery · 2 src
- OpenAI data shows AI automates execution but not strategic decisions · 3 src
- GraphRAG with TypeSafe Jev: A System One Approach to Scalable Knowledge Graphs · 1 src
- Synthetic data defined, uses, risks, and best practices · 1 src
Comments
via GitHub Discussions