Brand MU Day
    • Categories
    • Recent
    • Tags
    • Popular
    • Users
    • Groups
    • Register
    • Login

    AI Megathread

    Scheduled Pinned Locked Moved No Escape from Reality
    400 Posts 52 Posters 124.0k Views
    Loading More Posts
    • Oldest to Newest
    • Newest to Oldest
    • Most Votes
    Reply
    • Reply as topic
    Log in to reply
    This topic has been deleted. Only users with topic management privileges can see it.
    • PavelP
      Pavel @dvoraen
      last edited by

      @dvoraen And somehow someone still needs to know COBOL.

      He/Him. Opinions and views are solely my own unless specifically stated otherwise.
      BE AN ADULT

      FaradayF 1 Reply Last reply Reply Quote 0
      • FaradayF
        Faraday @Pavel
        last edited by

        @Pavel Which GenAI will almost certainly never be able to because there isn’t enough COBOL stuff out there for it to steal for training data.

        PavelP 1 Reply Last reply Reply Quote 1
        • PavelP
          Pavel @Faraday
          last edited by

          @Faraday All COBOL knowledge is held exclusively by two men, both called Steve. They’re not allowed to travel at the same time, to avoid the risk of all worldly knowledge of COBOL being lost in the same incident.

          He/Him. Opinions and views are solely my own unless specifically stated otherwise.
          BE AN ADULT

          1 Reply Last reply Reply Quote 1
          • R
            Rathenhope @Hobbie
            last edited by Rathenhope

            @Hobbie As another person who works in fintech, I have been incredibly lucky on this score - I’m the most senior tech other than the CTO, and both of us are incredibly sceptical of LLMs, so while there has been the occasional push to Do More With AI we’ve been able to stand firm and not bring it in to general use.

            And we both got vindicated this quarter when two potential clients (the biggest we’d have) both went “we consider any use of AI to be high risk and we don’t want client information anywhere near it” and we went “excellent that’s our philosophy too.”

            That said I’ve been banned from talking about LLMs on our all-hands calls as it invariably turns into a 20 minute rant about why they are bad for our purposes and the calls are only mean to be 30 minutes long.

            HobbieH 1 Reply Last reply Reply Quote 8
            • FaradayF
              Faraday
              last edited by

              This English professor asserting that em-dashes are the biggest tell-tale sign of her students using AI is what I’m talking about when I complain about people picking on the dashes.

              1 Reply Last reply Reply Quote 1
              • hellfrogH
                hellfrog
                last edited by

                bring back downvotes

                fr fr
                (she/her)

                1 Reply Last reply Reply Quote 3
                • HobbieH
                  Hobbie @Rathenhope
                  last edited by

                  @Rathenhope said in AI Megathread:

                  two potential clients (the biggest we’d have) both went “we consider any use of AI to be high risk and we don’t want client information anywhere near it”

                  Just casually send me the names of these clients so I can get our CEO to try and sell to them, I really want to see that metaphorical ice bucket tipped all over him when he hears that.

                  1 Reply Last reply Reply Quote 2
                  • FaradayF
                    Faraday
                    last edited by

                    FYI, AI detectors are still kinda crappy.

                    Can AI Detectors Be Trusted? The Authors Guild Put Five of Them to the Test

                    “Our test confirms that while a couple of AI detection tools accurately identify human-authored text, some commonly used consumer-facing AI detection tools are wildly inaccurate, which presents a major risk for authors. Moreover, these tools change constantly—updated models, shifting benchmarks, evolving AI outputs—and their accuracy at any given moment cannot be assumed.”

                    Pangram Flagged My Own Writing as AI

                    Regarding false positive benchmarks: “that benchmark tests pure human text and pure AI text under controlled conditions. It doesn’t describe real-world use cases”

                    TezT 1 Reply Last reply Reply Quote 0
                    • TezT
                      Tez Administrators @Faraday
                      last edited by Tez

                      @Faraday

                      I don’t agree with your conclusions. From the same article:

                      Can AI detectors be trusted?

                      The results varied widely, and in some cases, dramatically.
                      Pangram and Originality.ai were the most reliable performers. Pangram returned 0 percent across all ten articles. Originality.ai returned 0 percent on eight of ten, with 1 percent on the remaining two. Both tools correctly identified every piece as human written.
                      Grammarly performed nearly as well, returning 0 percent on eight articles and flagging two at 7 percent and 9 percent respectively—low enough that neither would likely trigger concern in practice.

                      3 out of 5 did okay? They are tools to use alongside human reasoning. But demonstrably they aren’t garbage.

                      she/they

                      FaradayF 1 Reply Last reply Reply Quote 0
                      • FaradayF
                        Faraday @Tez
                        last edited by

                        @Tez said in AI Megathread:

                        3 out of 5 did okay? They are tools to use alongside human reasoning. But demonstrably they aren’t garbage.

                        One of the tools you’re saying did “okay” in the first article is the same one the second article raises concerns about. Do they sometimes work? Sure. Are they reliable across a wide variety of use cases? No.

                        TezT 1 Reply Last reply Reply Quote 0
                        • TezT
                          Tez Administrators @Faraday
                          last edited by

                          @Faraday IDK, I’d just continue to call them imperfect tools. That just puts it at 1 false flag in 11 samples instead of it’s 0 false flags out of 10 samples. Still not garbage.

                          I think there are real issues in these tools, but I think it is a mistake to write it all off as garbage. Use it as a tool alongside human judgment. Try to use the best tool you can.

                          she/they

                          1 Reply Last reply Reply Quote 0
                          • First post
                            Last post