• Pommes_für_dein_Balg ( mech@feddit.org ) 
        link
        fedilink
        arrow-up
        18
        ·
        edit-2
        8 months ago

        And everyone promotes them for tasks they aren’t experts in.
        Managers think they could replace devs, but never a manager.
        Devs think they could replace management but never a senior developer.
        Storyboard drawers think they can write screenplays. Screenplay writers think they can draw storyboards. Etc.
        As an expert, you know how shit AI is in your own field, but surely those other jobs are simple enough to be replaced.

          • mrgoosmoos ( mrgoosmoos@lemmy.ca ) 
            link
            fedilink
            English
            arrow-up
            3
            ·
            8 months ago

            some management, sure. just like some of your coworkers could probably be replaced by AI. but not the competent ones, and not the essential ones.

            and personally, I’d still rather work with an incompetent person who can improve than four incompetent chatbots

            although I’d rather work with no incompetence at all

            • Rivalarrival ( Rivalarrival@lemmy.today ) 
              link
              fedilink
              English
              arrow-up
              4
              ·
              8 months ago

              The principal task of a competent manager is, primarily, intervening between incompetent upper managers and actual workers. Replacing the incompetent manager removes the need for the competent one.

      • This.

        They’re incredibly useful, but you have to treat their output as disposable and untrustworthy. They’re reinforcement trained to generate a solution, regardless of if it’s right, because it’s impossible to AI evaluate that these solutions are correct at scale.

        If you’re writing some core code: you can use an agent to review it, refactor parts, stump the original version, infill methods, and to run your test/benchmark scripts.

        but you still have to manage it, edit it, make sure it’s not recreating the same code in 6 existing modules, generating faked tests, etc.


        As an example this week on my side project I had Claude Opus write some benchmarks. Total throwaway code.

        It actually took my input files, generated a static binary payload from it using numpy, and loaded that into my app’s memory (on its own that’s really cool), then it ran my one function and declared the whole system 100x faster than comparable libraries that parse the original data. Not a fair test at all, nor was it a useful test.

        You cannot trust this software.

        You’ll see these games metrics, gamed tests, duplicate parallel implementations, etc.