• CombatWombat@feddit.online
    link
    fedilink
    English
    arrow-up
    1
    ·
    13 hours ago

    Okay, so the author suggests when you run out of tokens that you should learn what other people do

    what should happen to the roles that currently by design have little to do when they see at screen the feared 5h:100% 7d:100%? Are they allowed to pick up a book and study something? Go and learn what they colleagues do?

    Ane he thinks this is a good use of your time because it’s unrealistic to expect you to learn what you do?

    Let’s be very pragmatic: when you have a handful of (sub)agents running in parallel, especially with models that don’t even show their reasoning beyond short occasional summaries, you’d spend hours just to figure out the simplest or most approachable of these workstreams. You try to understand something, do a bit of manual work, and hope you don’t break the internal consistency the agent was following.

  • nymnympseudonym@piefed.social
    link
    fedilink
    English
    arrow-up
    2
    ·
    16 hours ago

    10x token usage reduction is often achievable.

    Not doing stupid shit like asking an LLM to scan a million line file when you could grep, breaking skills files into progressively-discoverable chunks, running normal algorithms where a full LLM isn’t necessary.

    And then using an open source model to really reduce the cost

      • nymnympseudonym@piefed.social
        link
        fedilink
        English
        arrow-up
        1
        ·
        15 hours ago

        Depends on how things play out. A combination of prices, regulations, and public sentiment leads a lot of petro-companies to offer energy-saving advice (even if it hardly offsets the rest of the negatives of the business, they do wind up doing this)