Based on my experience the limitation can be followed too literally (literal 5 limit) and/or applied on lists it shouldn't be applied on (lists that influence further results and shouldn't be really limited, grouped, etc at that point as they might be relevant later on).
In multiple occasions it didn't do what IMO the original intention was. It doesn't just split a list - that is bigger than 5 items - into "do now" vs "later," or "must" vs "nice to have.". It just literally "capped the list at 5 items". It should be called more like "split up", "create sections", "group", etc.
Also on top of this not every question has 5 (or less) valid answer, and there are lists that must have all the actually relevant items, even if it means having more than 5. I totally agree that the endless lists some newer models produce are bad, but intentionally dropping very relevant findings aren't really better either - and with the current wording it did happen.
For my environment I completely removed that section, and it still brings most of the benefit - with better accuracy - so that is an option as well, but there might be less-nuclear options than completely removing that section:
- the wording should be tweaked to drop anything barely resembling "limiting", and probably explicitly scoped to influence only the final output
- the item count limit should be made dynamic to be adjusted by the LLM based on the context/task itself - or at least increasesed to 10 so cuts less by accident
My point is that - if I get it right - this skill should be about influencing the output the user gets, not the internal thought process. Sure burning tokens on something that won't be shown/used later on is a waste, but I doubt we can influence the internal workings of a thought process without affecting result/finding validity/completeness, so instructions that might be interpreted as such should be made less ambigous and more focused about the scope.
I wanted to talk about this before I do any PR - or you might take it yourself.
|
### 9. Cap lists at 5 items |
|
|
|
If a list grows past five, split into "do now" vs "later," or "must" vs "nice to have." Five items ranked beats ten unranked. |
|
|
Based on my experience the limitation can be followed too literally (literal 5 limit) and/or applied on lists it shouldn't be applied on (lists that influence further results and shouldn't be really limited, grouped, etc at that point as they might be relevant later on).
In multiple occasions it didn't do what IMO the original intention was. It doesn't just split a list - that is bigger than 5 items - into "do now" vs "later," or "must" vs "nice to have.". It just literally "capped the list at 5 items". It should be called more like "split up", "create sections", "group", etc.
Also on top of this not every question has 5 (or less) valid answer, and there are lists that must have all the actually relevant items, even if it means having more than 5. I totally agree that the endless lists some newer models produce are bad, but intentionally dropping very relevant findings aren't really better either - and with the current wording it did happen.
For my environment I completely removed that section, and it still brings most of the benefit - with better accuracy - so that is an option as well, but there might be less-nuclear options than completely removing that section:
My point is that - if I get it right - this skill should be about influencing the output the user gets, not the internal thought process. Sure burning tokens on something that won't be shown/used later on is a waste, but I doubt we can influence the internal workings of a thought process without affecting result/finding validity/completeness, so instructions that might be interpreted as such should be made less ambigous and more focused about the scope.
I wanted to talk about this before I do any PR - or you might take it yourself.
i-have-adhd/skills/i-have-adhd/SKILL.md
Lines 105 to 108 in d05af1e