• Jrockwar@feddit.uk
    link
    fedilink
    English
    arrow-up
    6
    ·
    2 months ago

    But this is only because of execs’ stupidity.

    For a simple task that AI can actually do, say a boring text processing task, the API calls don’t just cost less than my salary, they typically cost less than keeping the monitor on during the time it would take me to do it.

    However when companies are stupid and decide to do things such as “tokenmaxxing” or leaderboards for who can waste more AI compute, you end up with things like calling multi-trillion models to do number calculations, or passing a 300k token context into every turn of the chat and giving it your entire codebase to change two lines of code.

    This one I blame squarely on stupid CEOs and execs. The smaller version of Gemma 4 is low-powered enough that can run on phones, and can output usable results for many use cases.