I'm tired of LLM bullshitting. So I fixed it.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

I'm tired of LLM bullshitting. So I fixed it.

floquant@lemmy.dbzer0.com · 5 months ago

Holy shit I’m glad to be on the autistic side of the internet.

Thank you for proving that fucking JSON text files are all you need and not “just a couple billion more parameters bro”

Awesome work, all the kudos.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

FrankLaskey@lemmy.ml · 5 months ago

This is very cool. Will dig into it a bit more later but do you have any data on how much it reduces hallucinations or mistakes? I’m sure that’s not easy to come by but figured I would ask. And would this prevent you from still using the built-in web search in OWUI to augment the context if desired?

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

7toed@midwest.social · 5 months ago

abilterated one

Please elaborate, that alone piqued my curiosity. Pardon me if I couldve searched

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

7toed@midwest.social · 5 months ago

Thank you again for your explainations. After being washed up with everything AI, I’m genuinely excited to set this up. I know what I’m doing today! I will surely be back

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

BaroqueInMind@piefed.social · 5 months ago

I have no remarks, just really amused with your writing in your repo.

Going to build a Docker and self host this shit you made and enjoy your hard work.

Thank you for this!

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

Diurnambule@jlai.lu · 5 months ago

Same sentiment. Tonight it run on my systems XD.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

PolarKraken@lemmy.dbzer0.com · 5 months ago

This sounds really interesting, I’m looking forward to reading the comments here in detail and looking at the project, might even end up incorporating it into my own!

I’m working on something that addresses the same problem in a different way, the problem of constraining or delineating the specifically non-deterministic behavior one wants to involve in a complex workflow. Your approach is interesting and has a lot of conceptual overlap with mine, regarding things like strictly defining compliance criteria and rejecting noncompliant outputs, and chaining discrete steps into a packaged kind of “super step” that integrates non-deterministic substeps into a somewhat more deterministic output, etc.

How involved was it to build it to comply with the OpenAI API format? I haven’t looked into that myself but may.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

PolarKraken@lemmy.dbzer0.com · 5 months ago

The very hardest part of designing software, and especially designing abstractions that aim to streamline use of other tools, is deciding exactly where you draw the line(s) between intended flexibility (user should be able and find it easy to do what they want), and opinionated “do it my way here, and I’ll constrain options for doing otherwise”.

You have very clear and thoughtful lines drawn here, about where the flexibility starts and ends, and where the opinionated “this is the point of the package/approach, so do it this way” parts are, too.

Sincerely that’s a big compliment and something I see as a strong signal about your software design instincts. Well done! (I haven’t played with it yet, to be clear, lol)

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

recklessengagement@lemmy.world · 5 months ago

I strongly feel that the best way to improve the useability of LLMs is through better human-written tooling/software. Unfortunately most of the people promoting LLMs are tools themselves and all their software is vibe-coded.

Thank you for this. I will test it on my local install this weekend.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

WolfLink@sh.itjust.works · 5 months ago

I’m probably going to give this a try, but I think you should make it clearer for those who aren’t going to dig through the code that it’s still LLMs all the way down and can still have issues - it’s just there are LLMs double-checking other LLMs work to try to find those issues. There are still no guarantees since it’s still all LLMs.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

skisnow@lemmy.ca · 5 months ago

I haven’t tried this tool specifically, but I do on occasion ask both Gemini and ChatGPT’s search-connected models to cite sources when claiming stuff and it doesn’t seem to even slightly stop them bullshitting and claiming a source says something that it doesn’t.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

skisnow@lemmy.ca · 5 months ago

How does having a key solve anything? Its not that the source doesn’t exist, it’s that the source says something different to the LLM’s interpretation of it.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

skisnow@lemmy.ca · 5 months ago

The hash proves which bytes the answer was grounded in, should I ever want to check it. If the model misreads or misinterprets, you can point to the source and say “the mistake is here, not in my memory of what the source was.”.

Eh. This reads very much like your headline is massively over-promising clickbait. If your fix for an LLM bullshitting is that you have to check all its sources then you haven’t fixed LLM bullshitting

If it does that more than twice, straight in the bin. I have zero chill any more.

That’s… not how any of this works…

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

rollin@piefed.social · 5 months ago

At first blush, this looks great to me. Are there limitations with what models it will work with? In particular, can you use this on a lightweight model that will run in 16 Gb RAM to prevent it hallucinating? I’ve experimented a little with running ollama as an NPC AI for Skyrim - I’d love to be able to ask random passers-by if they know where the nearest blacksmith is for instance. It was just far too unreliable, and worse it was always confidently unreliable.

This sounds like it could really help these kinds of uses. Sadly I’m away from home for a while so I don’t know when I’ll get a chance to get back on my home rig.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

rollin@piefed.social · 5 months ago

I never knew LLMs can run on such low-spec machines now! That’s amazing. You said elsewhere you’re using Qwen3-4B (abliterated), and I found a page saying that there are Qwen3 models that will run on “Virtually any modern PC or Mac; integrated graphics are sufficient. Mobile phones”

Is there still a big advantage to using Nvidia GPUs? Is your card Nvidia?

My home machine that I’ve installed ollama on (and which I can’t access in the immediate future) has an AMD card, but I’m now toying with putting it on my laptop, which is very midrange and has Intel Arc graphics (which performs a whole lot better than I was expecting in games)

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

Angel Mountain@feddit.nl · 5 months ago

Super interesting build

And if programming doesn’t pan out please start writing for a magazine, love your style (or was this written with your AI?)

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

Karkitoo@lemmy.ml · edit-2 5 months ago

meat popsicle

( ͡° ͜ʖ ͡°)

Anyway, the other person is right. Your writing style is great !

I successfully read your whole post and even the README. Probably the random outbursts grabbed my attention back to te text.

Anyway version 2, this Is a very cool idea ! I cannot wait to either :

incorporate it to my workflows
let it sit in a tab to never be touched ever again
tgeoryceaft, do tests and request features so much as to burnout

Last but not least, thank you for not using github as your primary repo

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

Karkitoo@lemmy.ml · 5 months ago

Don’t spam my Github inbox plz

I can spam your codeberg’s then ? :)

About the random outburst: caused by TOO MUCH FUCKING CHATGPT WASTING HOURS OF MY FUCKING LIFE, LEADING ME DOWN BLIND ALLEYWAYS, YOU FUCKING PIEC… …sorry, sorry…

Understandable, have a great day.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

CIA_chatbot@lemmy.world · 5 months ago

Doing gods work

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

CIA_chatbot@lemmy.world · 5 months ago

Friendship drive activated.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

null@piefed.nullspace.lol · 5 months ago

This is awesome. Definitely gonna dig into this later.

pineapple@lemmy.ml · 5 months ago

This is amazing! I will either abandon all my other commitments and install this tomorrow or I will maybe hopefully get it done in the next 5 years.

Likely accurate jokes aside this will be a perfect match with my obsidian volt as well as researching things much more quickly.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

Domi@lemmy.secnd.me · 5 months ago

I have a Strix Halo machine with 128GB VRAM so I’m definitely going to give this a try with gpt-oss-120b this weekend.

recklessengagement@lemmy.world · 5 months ago

Strix halo gang. Out of curiosity, what OS are you using?

Domi@lemmy.secnd.me · 5 months ago

Fedora 43 with the Rawhide kernel.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

Domi@lemmy.secnd.me · 5 months ago

Of course, self hosted with llama-swap and llama.cpp. :)

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

Domi@lemmy.secnd.me · 5 months ago

gpt-oss is pretty much unusable without custom system prompt.

Sycophancy turned to 11, bullet points everywhere and you get a summary for the summary of the summary.

sp3ctr4l@lemmy.dbzer0.com · 5 months ago

This seems astonishingly more useful than the current paradigm, this is genuinely incredible!

I mean, fellow Autist here, so I guess I am also… biased towards… facts…

But anyway, … I am currently uh, running on Bazzite.

I have been using Alpaca so far, and have been successfully running Qwen3 8B through it… your system would address a lot of problems I have had to figurr out my own workarounds for.

I am guessing this is not available as a flatpak, lol.

I would feel terrible to ask you to do anything more after all of this work, but if anyone does actually set up a podman installable container for this that actually properly grabs all required dependencies, please let me know!

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

sp3ctr4l@lemmy.dbzer0.com · 5 months ago

Oh I entirely believe you.

Hell hath no wrath like an annoyed high functioning autist.

I’ve … had my own 6 month black out periods where I came up with something extremely comprehensive and ‘neat’ before.

Seriously, bootstrapping all this is incredibly impressive.

I would… hope that you can find collaborators, to keep this thing alive in the event you get into a car accident (metaphorical or literal), or, you know, are completely burnt out after this.

… but yeah, it is… yet another immensely ironic aspect of being autistic that we’ve been treated and maligned as robots our whole lives, and then when the normies think they’ve actually built the AI from sci fi, no, turns out its basically extremely talented at making up bullshit and fudging the details and being a hypocrite, which… appalls the normies when they have to look into a hyperpowered mirror of themselves.

And then, of course, to actually fix this, its some random autist no one has ever heard of (apologies if you are famous and i am unaware of this), who is putting in an enormous of effort, that… most likely, will not be widely recognized.

… fucking normies man.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

Fmstrat@lemmy.world · 5 months ago

No promises, but if I end up running this it will be by putting it in a container. If I do, then I’ll put a PR on Codeberg with a Docker Compose file (compatible with Podman on Bazzite).

@SuspciousCarrot78@lemmy.world

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

itkovian@lemmy.world · 5 months ago

Based AF. Can anyone more knowledgeable explain how it works? I am not able to understand.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

itkovian@lemmy.world · 5 months ago

As I understand it, it corrects the output of LLMs. If so, how does it actually work?

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

itkovian@lemmy.world · 5 months ago

That is much clearer. Thank you for making this. It actually makes LLMs useful with much lesser downsides.

SuspciousCarrot78@lemmy.world · edit-2 2 months ago

[deleted by user]

itkovian@lemmy.world · 5 months ago

Will do.

I'm tired of LLM bullshitting. So I fixed it.

I'm tired of LLM bullshitting. So I fixed it.

llama-conductor