LLM stuff
The scary part here is just how long it can remain coherent as its contexts fills up. Even 3.6 would get confused quite quickly and need to be compacted after a few hours. This thing managed to ingest files from like three codebases and still have space for generating a weird setup for testing a react-native component.
LLM stuff
I think part of this is that I've gotten better at working within it's limitations and strengths as I've been updating my "agent" to be more effective.
LLM stuff
I haven't had much of a need for skills so far because the models are pretty capable even with minimal prompting and I found that skills often fill up the context more than is strictly necessary. I was thinking for ergonomics I'd have users define skills in their "documents" folder to make it more intuitive for graphical interface users.