DigestAI news desk

Cut through the AI noise.

Research10 min read

Laya, a 322M-parameter decision engine, gains 26,639 GitHub stars in nine days

Laya, an open-source decision engine released on September 18, 2026, has rapidly gained traction with 26,639 GitHub stars and 2,326 forks within nine days. Developed by NandhaKishorM and licensed under Apache-2.0, the tool offers a non-autoregressive alternative to using large language models for simple classification tasks. Instead of generating text tokens, Laya performs a single forward pass…

1 source

Key points

  • Laya gained 26,639 GitHub stars and 2,326 forks in nine days since its September 18, 2026 release.
  • The 322M-parameter multilingual model answers typed questions in one forward pass with zero output tokens.
  • Docs warn checkpoints are over-confident; users must fit temperatures on held-out data to trust probabilities.

The project ships three checkpoints on Hugging Face, including a 421M-parameter English model and a 322M-parameter multilingual model supporting over 100 languages. The repository claims a latency of 33ms per question on a T4 GPU, dropping to 7.2ms when batched. A hands-on test on a CPU-only VM confirmed that the system runs locally without API keys or GPU requirements, though CPU inference is slower at roughly 0.7 seconds per question. The tool includes a built-in abstention mechanism that flags low-confidence answers, allowing developers to route uncertain cases to human review or larger models.

While the speed and cost efficiency are significant, the documentation warns that the shipped checkpoints are over-confident and require users to fit temperature parameters on their own data before relying on the probability scores. Laya is positioned for high-volume, bounded decisions like ticket triage and moderation, where deterministic, fast responses are preferred over the nuanced reasoning of frontier models.

Model page: Laya →

Full story from aifrontierpost.com · by Marcus Doyle · via Reddit AI communitiesOpen source ↗

Laya: replace LLM-as-a-judge with a 322M-parameter decision engine (26,639 stars in 9 days, hands-on test)

aifrontierpost.com · 27 September 2026

Loading the full article…

This text was published by aifrontierpost.com and written by Marcus Doyle. It is reproduced here with attribution so you can read it in full; the rights remain with the publisher. Read it at the source ↗

Coverage and discussion

1source
Topics · follow one to build your own front page
Hugging FaceLayaModernBERT-largeNandhaKishorM

The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us. Published by Martin K., who runs Digest AI and handles corrections.

Comments

via GitHub Discussions

More in Research

All →

Related stories