Finetuned 1.5B Qwen to generate bash commands at gpt-4o level using 400k synthetic examples + Fully opensource finetune dataset
Finetuned 1.5B Qwen to generate bash commands at gpt-4o level using 400k synthetic examples + Fully opensource finetune dataset How it started
Key points
- 401,975 synthetic request/command pairs used for fine‑tuning
- Model scored 63.7 % on ALFA‑updated 300‑task benchmark
Despite using LLMs for most of the coding, there was always one thing I kept Googling: Bash commands. It's quite flow-breaking to pause work, open Google, type the full query, go to Stack Overflow or similar, and look up the syntax I wanted. Intuitively, this always felt like something a small model would perform well at because you have a finite set of commands, well-defined syntax, and easy-to-generate training data. So one day I decided to actually find out. Over the course of the experiment, I tried six different models: SmolLM 135M, SmolLM 360M, Qwen3-0.6B, Qwen2.5-Coder-1.5B, Qwen3.5-0.8B, and Qwen3.5-2B.
Finetuned 1.5B Qwen to generate bash commands at gpt-4o level using 400k synthetic examples + Fully opensource finetune dataset
dirac.run · 5 October 2026Loading the full article…
This text was published by dirac.run and written by Max Trivedi. It is reproduced here with attribution so you can read it in full; the rights remain with the publisher. Read it at the source ↗
Coverage and discussion
1source- Reddit discussionreddit.com
The headline, key points and digest above were generated by Digest AI's editorial model from the linked sources. Automated summaries can contain errors: the sources are the record. Spotted a mistake? Tell us. Published by Martin K., who runs Digest AI and handles corrections.
More in Generative AI & Models
All →- Aleph Alpha releases Kolibri-1: 78B MoE model with 3.46B active parameters for German and English · 5 src
- TypeSafe AI's Jev decision model matches Claude Sonnet 5 accuracy at lower cost and latency · 4 src
- Yandex introduces Sona, a generative recommender that replaces recommendation cascade · 1 src
- Anthropic releases Opus 5.5 and Sonnet 5.5; OpenAI launches GPT-6 Sol and Luna tiers · 6 src
- Anthropic cuts AI costs with faster, cheaper Claude Sonnet 5.5 · 31 src
Comments
via GitHub Discussions