{"version":1,"type":"story","url":"https://digestai.news/story/tuning-qwen-3-8-27b-and-omp-as-a-coding-agent-on-2-3090s","json":"https://digestai.news/story/tuning-qwen-3-8-27b-and-omp-as-a-coding-agent-on-2-3090s.json","markdown":"https://digestai.news/story/tuning-qwen-3-8-27b-and-omp-as-a-coding-agent-on-2-3090s.md","slug":"tuning-qwen-3-8-27b-and-omp-as-a-coding-agent-on-2-3090s","headline":"Tuning Qwen 3.8 27B and OMP as a coding agent on 2× 3090s","summary":"The hardware setup consists of two used RTX 3090 GPUs running vLLM. By default, omp is tuned for hosted cloud models, which can cause local setups to experience severe latency or truncated file writes.\n\nTo optimize performance, the developer adjusted five key settings in omp. These changes include setting explicit effort levels for all roles, establishing a thinking budget of 6,000 to 7,500 tokens, raising the maximum token limit to 32,768, saving tool outputs over 10 KB to files, and limiting parallel subagents to four.","keyPoints":["Setting a thinking budget of 6,000-7,500 tokens prevented the model from wasting time on excessive reasoning.","Raising maxTokens to 32,768 stopped the agent from saving half-written files due to truncated outputs."],"whyItMatters":"This guide shows how developers can run capable, private coding agents locally on consumer hardware, avoiding recurring cloud API fees and data privacy concerns.","category":{"slug":"marketing","name":"Marketing & Small Business","url":"https://digestai.news/category/marketing"},"entities":{"companies":[],"models":["Qwen3.8-27B","Qwen3.8 Flash Next"],"people":[]},"firstPublishedAt":"2026-09-13T00:00:00Z","updatedAt":"2026-09-13T00:00:00Z","sourceCount":1,"hasPrimarySource":false,"sources":[{"outlet":"doug.sh","title":"Tuning Qwen 3.8 27B and OMP as a coding agent on 2× 3090s","url":"https://doug.sh/posts/tuning-a-local-coding-agent-oh-my-pi","publishedAt":"2026-09-13T00:00:00Z","type":"press","primary":false,"lead":true}],"sourceNotes":null,"discussions":[{"site":"Reddit","url":"https://www.reddit.com/r/LocalLLaMA/comments/1wk8jef/tuning_qwen_38_27b_and_omp_as_a_coding_agent_on_2/","points":null}],"thread":null,"cite":{"text":"Digest AI, \"Tuning Qwen 3.8 27B and OMP as a coding agent on 2× 3090s\", 13 September 2026, https://digestai.news/story/tuning-qwen-3-8-27b-and-omp-as-a-coding-agent-on-2-3090s","publisher":"Digest AI","title":"Tuning Qwen 3.8 27B and OMP as a coding agent on 2× 3090s","datePublished":"2026-09-13T00:00:00Z","url":"https://digestai.news/story/tuning-qwen-3-8-27b-and-omp-as-a-coding-agent-on-2-3090s"},"generatedBy":"Written by Digest AI's editorial model from the linked sources; the sources are the record.","license":"Headlines, digests and key points are written by Digest AI and may be quoted with a link to the story page. Linked articles belong to their publishers. Terms: https://digestai.news/terms#reuse"}