It has to be pure ignorance.
I only have used my works stupid llm tool a few times (hey, I have to give it a chance and actually try it before I form opinions)
Holy shit it's bad. Every single time I use it I waste hours. Even simple tasks, it gets details wrong. I correct it constantly. Then I come back a couple months later, open the same module to do the same task, it gets it wrong again.
These aren't even tools. They're just shit. An idiot intern is better.
Its so angering people think this trash is good. Get ready for a lot of buildings and bridges to collapse because of young engineers trusting a slop machine to be accurate on details. We will look back on this as the worst era in computing.
96 Comments
TBi@lemmy.world · 53 pts · 177d
Generally I equate positivity about LLMs with people’s technical ability. I find the more they say AI is good the worse programmer they are.
bridgeenjoyer@sh.itjust.works · 36 pts · 177d
Technical literacy in general. My friend thinks it's the greatest thing ever, is an idiot with technology (and life in general).
okamiueru@lemmy.world · 3 pts · 176d
It also says a lot about their inability to identify bullshit
Valmond@lemmy.dbzer0.com · 3 pts · 175d
Might be some dunnig kruger curve there. Not tooting my horn but I know my ways around and I only use ai for programming when I more or less know how it works already. Which means I verify and fix any eventual problems before committing any code. It does speed up the process, it's a tiny bit simpler than checking stuff out on stack overflow IMO.
Now, if you don't know your ways around, and "trusts" the outcome on an LLM, boy are you in trouble 😵💫.
CompactFlax@discuss.tchncs.de · 39 pts · 177d
Yes, exactly this. It looks good, I ask for it to tweak something. It tweaks, but now something else needs adjustment. Then it comes back unusable.
It ends up taking the same time as doing it myself. There’s some value perhaps in either the novelty or engagement that keeps me focused but it’s not more efficient.
When it does work, I’m always worried it is an illusion I’ve missed something. Like how you send an email and immediately see the typo.
People who love it, love it because they don’t need to or care about having accuracy and precision in their work. Sales and marketing, management, etc. Business idiots.
bridgeenjoyer@sh.itjust.works · 9 pts · 177d
Ed ed ed!!!!
cypherpunks@lemmy.ml · 29 pts · 177d
you're obviously prompting it wrong, and/or not using the latest models
/sLiketearsinrain@lemmy.ml · 7 pts · 176d
bonenode@piefed.social · 6 pts · 176d
a_non_monotonic_function@lemmy.world · 3 pts · 176d
LLM: Wow, you are so good at this. Are you sure this is your first time? ...Oh, prompt me harder, daddy.
GrindingGears@lemmy.ca · 1 pts · 175d
Oh man I deleted LinkedIn last month, and March has been glorious.
I literally feel like there's hope for humanity now. Like just a little glimmer of it. I didn't even really use it, but it somehow sucked my soul and shattered it into 1,000 pieces.
AlecSadler@lemmy.dbzer0.com · 17 pts · 176d
There are right tools and wrong tools depending on the application.
There are right ways to use said tools and wrong ways...like you wouldn't use a phillips head screwdriver on a flat head.
I guarantee your company's provided tool is Copilot or OpenAI based, which is already bottom of the barrel for usefulness.
bridgeenjoyer@sh.itjust.works · 3 pts · 176d
Haha, yes it is
kunaltyagi@programming.dev · 0 pts · 176d
Inuse a flathead (minus) screwdriver on a philips (plus) screw all the time
Mikina@programming.dev · 16 pts · 177d
My experience is that it can work reasonabky well, but you have to waste absurd amount of tokens and have the 1m token context window.
I only do gamedev, which means a little bit more simple scriots, but it could handle even more involved systems.
If i force it to first document and explain the wholw architecture and data flow.
Did it help? Yes, but it still did take only a slightly faster. I could do it in a day more, probably.
It also cost like 50$ in tokens, in today's prices - where every AI company is loosing trillions of money, so the costs will get a lot worse. And if I tried to conserve tokens, it's shit. You have to feed it 10$ of data to be useful.
Add to that the fact that it also causes skill attrition, so once the expensive-cost future arrives, you probsbly won't be able to afford it, and good luck getting your skills back after that.
Our company wants us to use it, and the average token consumption is like 100$ per day in consumer prices. How is that even considered for such a minor gain?
So, I'll pass.
bridgeenjoyer@sh.itjust.works · 6 pts · 177d
Especially once they kill the "old net" and you won't be able to browse anything to learn skills at all.
kagi is the only way i can stay sane on web 3.0.
AA5B@lemmy.world · 1 pts · 175d
There’s enough people who drink the koolaid.
I helped this one as guy use an LLM to migrate his test suite to java21. It did help him incorporate some new language features but I don’t see how it made up for my time sitting with him
….. yet to management he saved 20% time. They trust that number despite no actual measurement, and hold it up as efficiency we all need to find.
But certainly if a 20% efficiency gain were real, that would be well worth $100
ReallyCoolDude@lemmy.ml · 10 pts · 177d
Id say you dont know how to use the tool. 'Write tests here, develop.feature X' is not a good way to use llms. Using 80k tokens and keep using same a Session is context rotting. There are a lot of boring, everyday tasks in my job that got faster. Many others that meh. Use AI, dont be driven by AI.
Professorozone@lemmy.world · 1 pts · 176d
I think I'm going to use AI to tell me how to use AI.
douglasg14b@lemmy.world · 9 pts · 176d
I know this community is all about fuck AI, but this is just straight echo chambering.
But honestly your post sounds like you're just not using it right? You can get pretty good results with it with enough guardrails. Just because you can't get the results you want doesn't mean that no one can.
That said, fuck AI. It's all a bunch of bullshit, but denying real results just means you're sticking your head in the sand and that's not how you fix this problem.
cooligula@sh.itjust.works · 13 pts · 176d
I agree... Saying LLMs are good at nothing is just plain ignorance... One can disagree with the philosophy or dislike hallucinations, but they are definitely good at some things.
GrindingGears@lemmy.ca · -2 pts · 175d
It's basically like Google with a bit more detail in my experience. Everytime I've tried to use it in a professional context, I've come up massively empty. Pages and pages and pages and pages of just absolutely walls of text, but nothing actually useful. I mean I've got it to calculate stuff and whatever, but then you examine something and its not coming up for you like the LLM says it should be. Which pretty much immediately means you have to validate everything else, and then it's like well hey look here I am however many hours later, manually doing something.
Our executives keep telling us to adapt or we'll be on the losing end. At this point, I'd just like the check please. Because if the company can survive on images of Super Mario committing 9/11, or walls of useless text or just straight up make belief, that's something I'd like to watch from the sidelines.
utopiah@lemmy.world · 3 pts · 175d
examples?
arbitrary_sarcasm@lemmy.world · 5 pts · 175d
For a research project, I had to convert 20+ projects from a dataset into a new format. The old format was simply a single script for each project that builds it. But I needed a format with a Docker file and a script. It would've taken me around a week to do all that one by one.
I got Claude to do it in 2 hours.
I know people hate AI in this community, but to say it doesn't do anything good or to insult all people who use it is just pure negativity.
bridgeenjoyer@sh.itjust.works · 2 pts · 174d
Thats good. It has use cases. Is the monetary and earth destroying cost worth it? Not in the slightest
utopiah@lemmy.world · 1 pts · 168d
Thanks for sharing. I'm not sure where me asking for examples "say it doesn’t do anything good or to insult all people who use it". Someone makes a claim without any proof, I ask for proof. To me that sounds both simple and legitimate.
T156@lemmy.world · 2 pts · 175d
Or that it's not right for their use case.
Like someone throwing a bunch of data into an LLM and trying to use it to process it into a chart or something. It can work, but it was never designed to be used in that manner.
I've got an acquaintance who does that, despite the fact that python would be a better thing to use.
Personally, I sometimes run a few saved images thorough a multi-modal 8 gigaparameter local model on my computer, so I can automate giving them more descriptive names than randomnumbers.png, and that seems to work fine. I could do it by hand, but it would take hours and days, compared to minutes, and since it's not too important, it doesn't matter if it's wrong. The resource usage is also less of an issue, since it's my own computer.
GreenKnight23@lemmy.world · 7 pts · 175d
holy fucking shit man. this community has a clear astroturfing problem.
bridgeenjoyer@sh.itjust.works · 2 pts · 174d
I don't really know what the word means. I was posting my experiences.
GreenKnight23@lemmy.world · 3 pts · 174d
not anything to do with your post. it's the comments here.
fuckai used to be a community where no exceptions were made for AI. it's quite literally in the name of the community.
so many posts from this community over the last 3-4 weeks has had an increasing amount of users that are "astroturfing" that AI has its uses and can be helpful sometimes. I kind of feel like these comments are made disingenuously as a way to silence the community at large by over commenting in a community that was created literally to hate AI, no exceptions.
anyway, won't stop me from never using AI. If anything it'll just make me read more books from before AI was a thing.
bridgeenjoyer@sh.itjust.works · 2 pts · 174d
OHH I thought the opposite.
jj4211@lemmy.world · 2 pts · 175d
A lot of people have their livelihood tied to the narrative that LLM deserves every cent of investment. The fact that it's utility is more limited is an existential threat to their careers.
The truth that it is selectively useful gives them a thread of hope, but the fact it is useless for a lot of stuff drives irritation. We don't make a distinction between the sort of work that LLM can do and can't so people end up completely dumbfounded by the other perspective.
rabber@lemmy.ca · 5 pts · 177d
I recently used it to install Nvidia l40s drivers on redhat 9 and pass it through to my Frigate instance. Took me a few minutes. Would have been a lot of reading to find the exact answers manually.
bridgeenjoyer@sh.itjust.works · 1 pts · 174d
Not a bad usage
rabber@lemmy.ca · 2 pts · 174d
I also used it to compile the correct Yolo detection models since I wrote this comment. I gave it my specific cameras and settings and it told me how to compile the correct model. I have 200 cameras running now. Almost unheard of on a deployed Frigate instance. Using the correct cuda model now I'm seeing 38 percent usage average on my l40s compared to 80 before.
100_kg_90_de_belin@feddit.it · 5 pts · 176d
I cut my LLM usage to almost zero because of environmental and political reasons, but it was helpful enough to wish it could be sustainable and not another tool in the dystopian take on the world.
IronBird@lemmy.world · 4 pts · 176d
local models are advanced enough to the point where you can run em as needed without datacenter.
the datacenter craze is basically just an excuse to get the banks (and eventually the american taxpayer, via bailouts when they fail) to fund your local nepitistic infrastructure rollout.
the entire US economy is built around the purposeful boom/bust system, as it's very effecient at "bagging" people that don't know the rules.
wewbull@feddit.uk · 1 pts · 175d
They've still had a huge power investment in creating them.
foxwolf@pawb.social · 4 pts · 175d
Oooh buddy, is isn't even young engineers using these to destroy their designs. I was at a building construction conference recently where one of the presentations was about how AI is going to "give us so much time back" as designers. He then told us about how the AIs hallucinate math still, and that the AI companies are not liable for their output. After the presentation, I and another person asked him a question about who exactly the liability will lie with and how someone could protect themselves from the liability without spending all the time we "save" meticulously checking the outputs. His response was to generate thousands of outputs for the same task and then only check "the best versions." Okay, so how will we know which are the "best" without meticulously checking thousands of them?
Anyway, afterwards, I asked my colleagues from all around the country who were at the conference for their opinion on AI and the presentation, and most of these 50-60 year old men told me they regularly use it in their work already. So be prepare for things constructed in the past few years to be incredibly dangerous facilities to be in or near.
utopiah@lemmy.world · 4 pts · 175d
Well, 100% because the intern WILL eventually learn. That's the entire difference. It won't be about adjusting the prompt, or add yet another layer of "reasoning", or wait for the next "version" with a different code name an .1% larger dataset. No, you'll point to the intern they did a mistake, try not calling them an idiot, explain WHY it's wrong, optionally explain how to do it right, THEN the next time they'll avoid it or fix it after.
That's the entire point of having an intern : initially they suck BUT as you train them, they don't! Meanwhile an LLM, despite technical jargon hijacked by the marketing department, they don't "learn" (from machine learning) or train (from "training dataset") or have "neurons" (from "artificial neural networks") rather it's just statistics on the next most probable world, sounding right with 0 "reasoning".
jj4211@lemmy.world · 2 pts · 175d
Had a person a few years back who would never ever learn.
In fact, a way I have expressed my opinion of LLM is that it is like working with that useless guy, except at least faster.
Based on my experience, the broader company is chock full of the never learn developers and I suppose I can see why they see value in the LLM, but either way their product sucks and no one likes them.
bridgeenjoyer@sh.itjust.works · 1 pts · 175d
You're so right .
And if the person sucks that bad, get rid of them
jj4211@lemmy.world · 2 pts · 175d
Yeah, but the same bad management that keeps thinking LLMs are magic are the same bad management that kept that guy around.
Every interaction that guy had where a senior tech ever dared to say he was useless ultimately landed the senior tech in hot water with management, as they claim "he says you aren't providing what he needs to suceed, that he is very skilled and willing to work, but you never told him how or gave him access or (a million other excuses that were generally lies)".
After a way too long career with us, he finally overplayed his hand by making the same old claims to the manager about no one giving him what he needed to work. Except he forgot that this time, the manager himself was the one who had been directing him and so he accidentally was accusing the manager of lying to himself.
Finally, the only person with credibility to the manager was on the receiving end of this guys grift.
bridgeenjoyer@sh.itjust.works · 1 pts · 174d
Its all a grift in the end!
Thats why youll mostly see conservatives/Nazis in love with llms. It fits their propaganda agenda perfectly.
AnotherPenguin@programming.dev · 3 pts · 176d
For programming, at least it's a good way to speed up things that you know how to do but take some time to type, or you don't remember the syntax of. But relying on AI any more than that usually means you'll be adding free technical debt and debugging time or becoming dependent on it.
AdamBomb@lemmy.sdf.org · 3 pts · 176d
Yeah, don’t generate code with it. Treat it like StackOverflow. It does pretty good at that.
BlameTheAntifa@lemmy.world · 5 pts · 176d
This is the only way I use it, and I do it grudgingly only because AI has ironically also ruined the web and web search. It’s also a last resort for when Kagi isn’t helping.
AA5B@lemmy.world · 1 pts · 176d
Unfortunately for me it’s a kpi so I need to figure out how to do something useful with it.
LLM is good for
But just in time for my performance review, I spent a week ignoring my work to set and tweak rule sets. Now it can be noticeably more useful
AdamBomb@lemmy.sdf.org · 2 pts · 176d
I agree with all that, especially if your performance is being measured by your use of LLMs. Those are cases where I find the code generation to be ok and doesn’t create comprehension debt.
GrindingGears@lemmy.ca · 1 pts · 175d
Just literally make something up and get it to lie about something. This is literally the land of make belief at this point, all this KPI shit. Don't stress about it. Execs want slop, give em slop.
Not_mikey@lemmy.dbzer0.com · 2 pts · 176d
Claude and super powers / planning have changed my mind more on AI feature development. Iterating on the spec and making it as unambiguous as possible gives good results when you clear context and have it implement the plan. Even if it starts to stray you can just do a git reset and start a new session with the spec, adjusting it a bit, because time wise you probably haven't invested much.
It also depends on the code base, if the code base has very clear separation of concerns, good documentation, and good contracts between layers then claude can handle it pretty well. If the code base is full of spaghetti code with multiple ways to do the same thing then AI will struggle with it. In our large legacy monolith repo it doesn't do well, in our micro service repos it does great.
Also time wise it may not seem like a benefit if you just set it and wait for it to complete, the productivity advantage comes from running a couple sessions in parallel.
Also context is key, having a good claude.md file in the repo to explain patterns helps it to avoid pitfalls. If it's only context is the prompt you gave it and you tell it to implement a feature without a plan / spec outlined it will generate shit code.
nightlily@leminal.space · 13 pts · 176d
If only we had a way to communicate with machines in a reliable, deterministic and unambiguous way.
Liketearsinrain@lemmy.ml · 7 pts · 176d
Not_mikey@lemmy.dbzer0.com · -1 pts · 176d
Yeah you can write the code yourself. You can also write in c or even assembly if you really want to make it as unambiguous as possible, it'll just take more time. Some people like to code in Python though because they can write faster with it even if a lot of implementation details and choices are hidden from them because they don't care about those details.
Spec driven development in my view is just another step, albeit a big one, on the level of abstraction between assembly and python. Like python it has its places and has places where it should never be used for safety and performance reasons.
Hoimo@ani.social · 4 pts · 176d
They may not care about the implementation details of a Python library, they do care about consistent execution and predictable results. And in some edge cases, they will care about the documentation saying exactly how those edge cases are handled.
Writing Python is abstraction, yes, but it's still programming. Once that Python code is written and tested and the dependencies are locked down, you can ship it and be certain it always works as designed.
Spec-driven code generation is nothing like that. I can't ship the specs. I could generate the code in a pipeline and ship that, maybe. But there's no way I'm getting consistent builds from a code generator. So what do people do? They generate the code and put it in source control for review. When have you ever checked-in a compiled executable or looked at it? There's machine code in there, shouldn't you review that the compiler did what you asked of it?
Not_mikey@lemmy.dbzer0.com · -1 pts · 176d
Consistency is dependent on the code base and not the "compiler" in this sense. If the code base has consistent patterns and only has one well documented way to implement something then the AI will follow those patterns, ie. If there is only one way to run a job, AI will use that method. There might be some variation in variable names, formatting, etc. but the core flow should be consistent between "runs"
You can and should still test your code , both manually and with automation to ensure it does what it says it does. Testing should be the way you are certain it always works as designed. IMO understanding your tests and test coverage is more important than understanding the implementation. This is why part of the spec for superpowers is a test plan, and that should be the most reviewed / iterated part.
TankovayaDiviziya@lemmy.world · 2 pts · 176d
It's situation specific. For tabulating data, yes. For everything else, probably not. But the thing is, you have to ask LLM if it can read the raw data to confirm if it is reading it right, before ordering it to execute more complex commands and tasks. You have to define the parameters one by one, one query everytime.
Lutra@lemmy.world · 2 pts · 175d
I think the intern comparison fits. The root of the problem is that AI can very good at the thing is is good at. That leads humans to believe that it is good at other things. This is often untrue.
Often the things it is good at are in the set of 'problems machines are good at'. Most professionals, people who are trained/experienced in their field face problem's that are NOT in that set. They are skilled, experienced problem solvers, who are solving difficult, real world problems. Not generic workers, or human resources.
The belief at the top is often that this machine which is 'so impressive', must therefore be good at everything. And this gets pushed down. Where people experience that same truth. The machine is incredibly good at the things it's good at, but it sucks doing what they do.
paraphrasing my grandpa - "To a suit with hammer, everything looks like a nail"
HarneyToker@lemmy.world · 2 pts · 176d
For every post I see of people complaining, I have to imagine there are 100 other people that get value out of LLMs quietly.
Azzu@lemmy.dbzer0.com · 1 pts · 176d
LLMs do a terrible job, however many things they're used for are so straightforward or unimportant that a terrible job is still "good enough".
CookieOfFortune@lemmy.world · 1 pts · 176d
Have you tried using skills/workflows? You can improve its context over time.
bridgeenjoyer@sh.itjust.works · 1 pts · 174d
I havent. Thats more time spent on a thing thats generally trash.
My work is too detail oriented with specific un documented use cases for it to work.
Its a glorified excel formula writer. Its OK at that.
homesweethomeMrL@lemmy.world · 1 pts · 177d
Fwiw, when I limit it to creating outlines based on given source docs or summarizing transcripts it does fine.
Definitely not worth what it cost to get there, but useful enough in those strict scenarios.
AA5B@lemmy.world · 1 pts · 176d
While I also don’t see how it’s productive, it can be useful for certain things, certain steps. But it really seems like you need to have the knowledge in question to help it do a good job.
People underestimate how much handholding it needs. You can tell it to do something and it might but you may not like the results. However with a bit of interaction or setting context, it might. The pretentious are calling it “prompt engineering” but it’s a combination of asking ai questions and modifying your terminology until it does what you want
People also don’t seem to understand ai really puts a premium on evaluation. You don’t see it being written but you own it, so you really need to look through the result in detail to understand whether it’s what you wanted. I see this in code a lot where the LLM produces something but a junior developer doesn’t have the skill to evaluate it before committing to source control
Witchfire@lemmy.world · 1 pts · 175d
A previous job forced us to use them, I spent more time getting the damn thing to work than actually doing work
okwhateverdude@lemmy.world · -1 pts · 177d
Is your work paying for dumb robots? Like Atlassian's shit? Or something built-in to your industry software (I guess some kinda CAD)? These are next to useless. Or is work only paying for basic model access? Also pretty useless for detailed work. The only models that give me consistent, detailed-ish work are the state of the art models. And even then you have to watch them, or have very strong verification/validation so they can bash their head against to eventually get the right result.
I'll say that I am not so much a booster, but more of a pragmatist. After the step change in quality this past fall/winter, I gave the SOTA models a try with hard earned cash. And it was worth it.
My ADHD makes it difficult to really finish personal projects once they get past the fun and interesting learning portion and neck deep into tedium of actually molding the code into the right shape or shaving the various yaks that came up. All motivation ceases. Unfortunately, my job is also my hobby. I don't wanna work after work, yo.
That game I always wanted to finish writing but got stuck at needing to grind out code? Done in an afternoon of carefully directing it. That programming language I spent significant amounts of time thinking and designing and getting the shitty PoC running but now needed to actually make it work? A week to the first version. Another to my first significant application written in that language which revealed flaws in my design for real use cases. Another to the next version with a conformance test suite which was then used along with the spec to do a complete reimplementation in another language. Another project was trying to "grow" a sorting algorithm expressed in a niche esoteric programming language using a genetic algorithm. Stuck at the point of needing to build the tools for analysis, needing a refactor to fix the poor persistence choices, just nothing but yaks to shave. Got it unstuck over a weekend and actually started to DO the damn experiment after spending so much time writing the esoteric lang interpreter and all of the experiment harness.
It is not perfect. It fucks up frequently. I have to really watch it and steer it. It loves mediocrity and shortcuts. All that said...
Like, holy shit. The amount of work I've finished or moved forward in two months is nothing short of miraculous given how many projects like these I have in various states of finished.
All I can say is that my experience aligned with your experience any time I needed to use a bolted-on AI to some product (Atlassian, Lucid, etc) but that does not reflect my experience when using SOTA models for real work.
bridgeenjoyer@sh.itjust.works · 7 pts · 177d
They just have a sub for gibbity 4 and 5 is all I know. I don't care, so I don't really look into it.
Thats good it works for you. I don't have the focus for a whole game. I'd have to start at the bottom, anything else is cheating and probably results in missing important aspects you would have learned. I could never bring myself to use it for projects, (unless, I'm getting paid to do it I guess) much like i would never use drum triggers/samples or autotune to fix my recordings.
I think less people have guilt nowadays. I detest things that are fake or shortcuts. It devalues real work.
jj4211@lemmy.world · -1 pts · 175d
I think it really depends on the task.
There are folks who manage to have their whole careers be basically put stuff from documentation and stack overflow to implement very basic stuff over and over again, and pray it works and doesn't need debugging. They hate coding, but it was heralded as a doctor/lawyer level pay but way easier to get into. LLMs can largely replace the work for those. These are humans I would never have trusted with anything significant, and only have them low stakes low risk stuff to keep them busy because management wants them utilized. Sure they spend a month to fall at delivering something basic, but management is happy enough.
Then there are folks who mostly live in code that is needed because it truly doesn't already exist. Those folks will find LLM relatively less useful. Now those folks do end up with braindead chores on occasion. Change from library x to y because whole they both do the same thing, x got discontinued. LLM can be useful at accelerating that because it's just so obvious but not quite as simple as search and replace. Or if you want to define some argv parsing you can let a codegen do that because it's easy but tedious.
To go back to the days of car analogies. Imagine some tech people got the world excited because they created tools to automate engineering in motor sports. People even come out saying how it helped them engineer their own vehicle and stories saying the most prolific motor racing is being taken over. You as an F1 engineer see it as mostly useless, but everyone keeps talking about it's going to replace engineers. Turns out everyone is actually taking about go karts and it is true that it works ok for that and that go karts are way more common than F1.
The problem is that to the world, programming is programming without distinction, and even the people in charge of the F1 type work don't know the difference because they were never technical either.
bridgeenjoyer@sh.itjust.works · 1 pts · 175d
I don't really mind people using it for simple dumb tasks. I'm just sick of boosters saying how great it is. If you're an intelligent person, you realize real fast how stupid it is at real work.
But we have destroyed the economy and planet and given all of our data to billionaires to do it. Not fucking worth it in the SLIGHTEST.
Maybe if THE PEOPLE owned all the data centers and models I could get behind it. But that'll never happen.
jj4211@lemmy.world · 2 pts · 175d
This is why I'm hoping the bubble pops soon. Too many people trying to gaslight about the utility of it for their self interest.
If it were just "boringly useful, but not mind boggingly profitable somehow", then I'm sure I'd no longer have a ton of people trying to micromanage use of LLM all around me.
Currently, my management has dictated that our failure to see "magic results" is because we just haven't been trained enough, and are paying for and mandating for over 200 hours of training on how to use LLM correctly. The grift is insane, since the whole point of the LLM is that it shouldn't need training to use, but here we are, people found a way to grift training on a 'doesn't need training' solution that doesn't work as advertised...
On a call with a partner, after demonstrating their software and everything it does, one of our executives kept insisting that they need to use LLM and then it would be even better.
So ready to have execs stop caring, and then maybe I'll somehow appreciate the residual utility of it, whatever it is.
infinitevalence@discuss.online · -4 pts · 177d
It really depends on the task and the tool. Current MOE models that have agentic hooks can actually be really useful for doing automated tasks. Generally, you don't want to be using AI to create things. What you want to do is hand it a very clear set of instructions along with source material. And then tell it to either iterate build on summarize or in some cases create from that.
I created a simple script with the help of AI to automate scanning files in from an automatic document reader. Convert them to OCRD PDFs. Scan through the document properly. Title the file based on the contents, then create a separate executive summary and then add an index to a master index file of a growing json.
Doing this allowed me to automate several steps that would have taken time and in the end I'm able to just search through my folders and my PDFs and very quickly find any information I need.
And this is only scratching the surface. I wouldn't have AI write me a resume or write me an email or a book. I might use it to generate an image that I then give to a real artist saying this is kind of what was in my head.
But boring stuff repetitive stuff. Things that really benefit from automation with a little bit of reasoning and thinking behind it. That's where we are right now with AI.
Carnelian@lemmy.world · 17 pts · 177d
Pretty much every pro AI person I’ve ever spoken to IRL tells me this exact same story
Basically, “I used AI for a boilerplate task. It gives me the vibe of being capable of much more” but then nobody can ever really get it to do much more
bridgeenjoyer@sh.itjust.works · 8 pts · 177d
Bro it will bro, it WILL just 5 more years and 3 more trillion bro i promise
infinitevalence@discuss.online · -2 pts · 177d
Yeah i used it for a boilerplate task, and then I did not have to do that task. I was able to scan in hundreds of documents, and at the end I had a fully indexed, properly file named, summarized, and searchable PDF library. Just doing one document at a time manually would have been a multi hour task for me, and in a few hours I was done, and it was good enough to good.
Carnelian@lemmy.world · 17 pts · 177d
Sorry, to clarify, a boilerplate task is not the repetitive job that you are automating. It’s the code that’s doing it.
What I’m saying is that you (and many others) are incorrectly attributing (your paragraph about the benefits your finished program is conferring to you) to AI, because that happens to be the path you took to arrive there.
In reality, the reason it was able to produce functional code is because your problem was already solved and documented. A few years ago, instead of “asking AI”, you would have simply copied and pasted the boilerplate code from someone else’s project. In all likelihood it also would have been faster for you to have done so
Quick edit: Sorry again, just want to further clarify that when I say “boilerplate task” I’m referring to a type of programing problem that you solve with “boilerplate code”. Reading back the above I was kind of using them interchangeably which is not strictly accurate
infinitevalence@discuss.online · -4 pts · 177d
If I could have found the code, and copied it yes. Using AI means I did not have to search for it, did not need to learn Python, and then did not need to do the tasks.
Please understand I hate 90% of all the AI bullshit being forced on us, and employers requiring AI use is insane. I am just saying that blanked hate is short sighted because it is useful in some cases, and the number of cases keep going up as we improve things.
I could also use a dictionary to check my spelling, but having spell check enabled is just faster.
Carnelian@lemmy.world · 10 pts · 177d
Right, and what I’m saying is that the ‘usefulness’ that people claim to have discovered is totally nonsensical because these problems have been solved for decades
bridgeenjoyer@sh.itjust.works · 5 pts · 177d
exactly. It's people realizing what coding is for the first time because it's being shoved in our faces. Before all this (bought and paid for propaganda) publicity, your average dummy had no idea what code was or did.
Basically, the non-tech people who know nothing about coding or how computers work are now amazed because they (think they) discovered what coding is because of the AI hype.
infinitevalence@discuss.online · -6 pts · 177d
Yes but now they are accessible to anyone.
Carnelian@lemmy.world · 11 pts · 177d
Your perspective is actually completely backwards on this
This process has always been accessible to everyone. You’d google basically the same words you typed into your prompt and it would bring you directly to the same block of code that everybody uses.
AI on the other hand is currently temporarily being made available for free or low cost because they are actively trying to create a cohort of users who impulsively “just reach for AI” as their first step to solving every problem.
In a few years you may find yourself praising how “accessible” it is because they occasionally run offers for a week of subscription time for $40 instead of the usual rate of $189.99/mo for entry level access. You may find yourself wondering how people ever lived without it
bridgeenjoyer@sh.itjust.works · 3 pts · 177d
I understand that and good for you for finding a use case.
however a normal program could have done that exact same thing. It's just miniscule easier to do it how you did. Probably not repeatable though, or if your model you used gets the plug pulled, you're back to square 1.
infinitevalence@discuss.online · -2 pts · 177d
Cant pull the plug, the model is local and running on my GPU. So I could break it if I wanted to but im not dependent on someone else.
Liketearsinrain@lemmy.ml · 1 pts · 176d
infinitevalence@discuss.online · 1 pts · 176d
For sure but I don't know Python. I can edit a script but I'm not proficient enough to make a new one.
bridgeenjoyer@sh.itjust.works · 11 pts · 177d
OK so that's what the boosters keep saying. However, every Task ive given it have very clear exact set of instructions, and when i comb through it, it still hallucinates. I can try talking to it like a 6 year old, and its going to forget what I told it the next day and hallucinate again.
Meanwhile I burned up a shit ton of power for 0 benefit and wasted my time. Worst case, someone else who is not detail oriented like myself is going to fuck up a lot of work.
However when i give it something I don't know how to do or dont know an answer to, it does a great job and is so smart! Amazing how that works.
infinitevalence@discuss.online · 0 pts · 177d
Your experience and feelings are totally valid, its still a shit show and in most cases a massive waste of power, water, brain power, and money. I am just saying that with the right training, and appropriate use, AI really can be effective NOW. One thing you can do is describe the task, and ask the AI to write a prompt to accomplish it. Then test that prompt in a new chat, and see if you get better results.
Most AIs will tell you how to improve your prompts.
bridgeenjoyer@sh.itjust.works · 5 pts · 177d
I can see where you're coming from and I get it.
I've done all that. I've followed it's little bs "hey, to improve, tell me this!!". Still gets it wrong.
It probably is fine for non-detail work and stuff no one really cares about (and if no one cares about it, maybe revisit what your job is). It's definitely not worth the cost. I myself don't enjoy telling a 5 year old the same thing 25 times (and then they forget because they saw a stick), which is exactly what using an LLM is akin to.
Lodespawn@aussie.zone · 9 pts · 177d
This is a wildly niche task. You also can't trust the executive summaries or indexes will be accurate so you need to read both the scanned documents, the executive summaries and the indexs to review, markup and correct. Your response will be "but it's good enough" but is it? You wouldn't trust it to write a resume, but you would trust it to write an executive summary of a much longer text that you rely on to search for information?
infinitevalence@discuss.online · -1 pts · 177d
Fair criticism, because it is "good enough" and it saved me so much time because I only needed detailed information on a few of the documents. But that is the point, it takes a repetitive bulk niche task and lets YOU quickly and simply create an automation that meets your needs.
Another example I have been using is having an AI read a resume, and do an ATS review based on a job posting. Then it generates a fit report, and makes suggestions on how the resume can be adjusted to better get past the ATS to an actual person. This is very useful in getting call backs for job applications. That has a real value.
Lodespawn@aussie.zone · 10 pts · 177d
How do you know it hasnt missed or misinterpreted key information from some of the documents you might have needed? The AI won't complete the same task the same way on every iteration so the automation is only verifiable if you check every time.
Did you get any of the jobs?
infinitevalence@discuss.online · -1 pts · 177d
Put in an application yesterday, got a call back today. Rest is up to me.
No i dont know if it missed something, but im not relying on the Exe summery for perfect detail, im using it to know which document is which and then using the index to go to points in the PDF.
Lodespawn@aussie.zone · 2 pts · 176d
Congrats on the call back!
Lodespawn@aussie.zone · 2 pts · 176d
This was posted to Lemmy recently and is relevant and well written
infinitevalence@discuss.online · 2 pts · 176d
That is an interesting paper for sure! Thanks for sharing it.