SmolLM 360M
2 stories mentioning SmolLM 360M, newest first, each with its sources and discussion. Follow to see new ones on your front page.
-
Finetuned 1.5B Qwen to generate bash commands at gpt-4o level using 400k synthetic examples + Fully opensource finetune dataset
Finetuned 1.5B Qwen to generate bash commands at gpt-4o level using 400k synthetic examples + Fully opensource finetune dataset How it started Despite using LLMs for most of the coding, there was always one thing I…
1 sourcedirac.run -
Study finds token-level entropy blind in small language models, semantic entropy helps
A new arXiv paper investigates whether entropy‑based confidence signals can improve the accuracy of small language models (SLMs) with fewer than 3 billion parameters that run on consumer hardware. The authors evaluate…
1 source primary sourcearXiv cs.CL
Questions about SmolLM 360M
What is the latest news about SmolLM 360M?
Finetuned 1.5B Qwen to generate bash commands at gpt-4o level using 400k synthetic examples + Fully opensource finetune dataset (5 October 2026).