Skip to content

Train lm_head LoRA in the gpt-oss preset to match Tinker SDK defaults - #6

Draft
kevintli wants to merge 1 commit into
devin/1790748364-miles-cli-overridesfrom
devin/1790748858-gpt-oss-lm-head
Draft

kevintli wants to merge 1 commit into
devin/1790748364-miles-cli-overridesfrom
devin/1790748858-gpt-oss-lm-head

Conversation

@kevintli

@kevintli kevintli commented Sep 30, 2026 •

Copy link
Copy Markdown

Summary

Adds lm_head as a target module for the gpt-oss-20b-lora-64k preset. This makes us consistent with the Tinker SDK defaults and the training recipe in https://github.com/jasper-lu/sec-search-rl.

Note: it seems the core issue here is that Miles doesn't support a multi-LoRA setup where different LoRA clients have different settings on which adapters they want to train. This forces us to define presets on the Spindle server side that exactly match what users choose for their training run on the client side (so you can't turn lm_head on/off on demand for example).

Filed a follow-up issue for this here with a proposal from Devin, open to further discussion: #12

@devin-ai-integration

Copy link
Copy Markdown

I'll fix CI failures and address comments from users with write access that start with 'Devin'.

  • Disable automatic comment, CI, and merge conflict monitoring

Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
Sign up for free to join this conversation on GitHub. Already have an account? Sign in to comment

Labels

None yet

Projects

None yet

Development

Successfully merging this pull request may close these issues.

1 participant