Stack Overflow seeks rebrand as traffic continues to plummet – which is bad news for developers

mesa@piefed.social · 7 months ago

Stack Overflow seeks rebrand as traffic continues to plummet – which is bad news for developers

anotherandrew@mbin.mixdown.ca · 7 months ago

Grab a copy of the stackoverflow database and use it locally, or train your own local LLM on the datastore.

And if you can, donate to the Internet Archive – those people do really important work in today’s age of killing off old information and constant enshittification.

madame_gaymes@programming.dev · 7 months ago

Came here to say something similar about a local archive.

You can also use the app Kiwix to make it a little easier to download/search (and grab several other doc archives like Python PEP and Wikipedia)

anotherandrew@mbin.mixdown.ca · 7 months ago

Completely forgot about kiwix; I have that on my ipad and laptop, along with Dash which is like a modern day HELPPC.COM if anyone remembers that thing…

madame_gaymes@programming.dev · 7 months ago

I didn’t know about Dash, but it sounds pretty great. Appears to be Mac only, though, and requires a subscription for the latest version.

Also found someone that appears to have converted HelpPC to HTML. Can’t speak to the legitimacy of it, though.

https://www.stanislavs.org/helppc/

Xanthobilly@lemmy.world · 7 months ago

I used Kiwix to grab a copy of Wikipedia.

ABetterTomorrow@lemm.ee · 7 months ago

Damnnnnn, truth bombs!

melroy@kbin.melroy.org · 7 months ago

Bad news. Since AI can only answer what it knows. If you have a question that is legit but not yet part of stackoverflow, you get a bad AI response.

In that case you can ask it on the stackoverflow website. But due to the fact that everybody now only rely on AI stackoverflow is dead. Well there you go, you just killed the source of truth.

cecilkorik@lemmy.ca · 7 months ago

Which is eventually going to cause AI model collapse, since AI no longer has any source of truth to train on. This is such an interesting technology being used in such a stupid and irresponsible way.

melroy@kbin.melroy.org · 7 months ago

Exactly my point. So what you see now is Ai is generating Ai content used for training. Also known as synthetic data… I know right?

anotherandrew@mbin.mixdown.ca · 7 months ago

I don’t know if it’s just my age/experience or some kind of innate “horse sense” But I tend to do alright with detecting shit responses, whether they be human trolls or an LLM that is lying through its virtual teeth. I don’t see that as bad news, I see it as understanding the limitations of the system. Perhaps with a reasonable prompt an LLM can be more honest about when it’s hallucinating?

mbtrhcs@feddit.org · 7 months ago

I don’t know if it’s just my age/experience or some kind of innate “horse sense” But I tend to do alright with detecting shit responses, whether they be human trolls or an LLM that is lying through its virtual teeth

I’m not sure how you would do that if you are asking about something you don’t have expertise in yet, as it takes the exact same authoritative tone no matter whether the information is real.

Perhaps with a reasonable prompt an LLM can be more honest about when it’s hallucinating?

So far, research suggests this is not possible (unsurprisingly, given the nature of LLMs). Introspective outputs, such as certainty or justifications for decisions, do not map closely to the LLM’s actual internal state.

anotherandrew@mbin.mixdown.ca · 7 months ago

I’m not sure how you would do that if you are asking about something you don’t have expertise in yet, as it takes the exact same authoritative tone no matter whether the information is real.

I agree – That’s why I’m chalking it up to some kind of healthy sense of skepticism when it comes to trusting authoritative-sounding answers by themselves. e.g. “ok that sounds plausible, let’s see if we can find supporting information on this answer elsewhere or, maybe ask the same question a different way to see if the new answer(s) seem to line up.”

So far, research suggests this is not possible (unsurprisingly, given the nature of LLMs). Introspective outputs, such as certainty or justifications for decisions, do not map closely to the LLM’s actual internal state.

Interesting – I still see them largely as black boxes so reading about how people smarter than me describe the processes is fascinating.

mbtrhcs@feddit.org · 7 months ago

let’s see if we can find supporting information on this answer elsewhere or, maybe ask the same question a different way to see if the new answer(s) seem to line up

Yeah, that’s probably the best way to go about it, but still requires some foundational knowledge on your part. For example, in a recent study I worked on we found that programming students struggle hard when the LLM output is wrong and they don’t know enough to understand why. They then tend to trust the LLM anyways and end up prompting variations of the same thing over and over again to no avail. Other studies similarly found that while good students can work faster with AI, many others are actually worse off due to being misled.

I still see them largely as black boxes

The crazy part is that they are, even for the researchers that came up with them. Sure we can understand how the data flows from input to output, but realistically not a single person in the world could look at all of the weights in an LLM and tell you what it has learned. Basically everything we know about their capabilities on tasks is based on just trying it out and seeing how well it works. Hell, even “prompt engineers” are making a lot of their decisions based on vibes only.

Stack Overflow seeks rebrand as traffic continues to plummet – which is bad news for developers

Stack Overflow seeks rebrand as traffic continues to plummet – which is bad news for developers

Stack Overflow seeks rebrand as traffic continues to plummet – which is bad news for developers • DEVCLASS