I’ve been doing the same thing, giving the same tasks to Qwen 3.8 27B and Opus, and the main difference is that Qwen does not consider edge cases which Opus catches. It’s good at the happy path, but even when hinting that there are uncovered edge cases and gotchas it’s oblivious to it. So I feel like I need a bigger model to do planning/review.
I noticed Claude code does something similar when "Exploring" the code base, it spawns a subagent using Heroku to get the summary and file locations, which is both faster and it doesn't pollute the context as much.
I wonder how it compares to an vector indexing approach.
10 years worth of Claude Max today. Also - Anthropic recently removed a model I relied on and isn't giving it back. As a non-US citizen, I would rather pay in advance but be sure, I will keep having access to inference on my own terms.
I noticed Fable was quite a bit terser, and I think it's due to changes in the system prompt [0]. They're literally saying "just give me the TLDR" and "give brief updates". You can tweak a lot of that with an AGENTS.md.
Thunderbird does not seem to have auto-categorization from what I see, just filtering. Neither does KMail. Unless you’re referring to some addons? For apple mail you have to add it on each client. And a lot of comments are about how to disable it because it categorizes wrong.
Reddit “best” sorting is pretty much like instagram and TikTok now, have to make sure it on hot/top, otherwise it’ll show you “related” things from subreddits you never subscribed to.
reply