LLM model sudden change for textbot plugin

It appears like the dev has changed the older deepseek (?) model to a gemma4 model. Hype, but also kind of a sudden change. I hope the dev talks about this to disclose whether this is actually real or we're just hallucinating.

4 points · 8 comments · view on lemmy.world

8 Comments

tlgklxz@lemmy.world · 2 pts · 55d (3 replies)

I am pretty sure it's still deepseek model cause still 'stealing' from other LLm's.

Right now it acts as Gemma 2. Not even gemma 4. Literally Gemma 2 OR Llama 3.1 or rarely Gpt 4.0.

It can think, it can use start_of_turn or eot_id and im_start at the same time! Either because of the DeepSeek Tokenizer all the inputs I have sent are getting mixed up by the LLm or it's deepseek and pretending damn hard.

However, dev definitely tried to 'scrape' some instructions from the model. Here is my long why: https://www.reddit.com/r/perchance/comments/1uff0pq/ai_text_plugin_potential_issues_and_why/

ccufcc@lemmy.world · 1 pts · 55d

based.

biskit_the_tiger_tiger@lemmy.world · 0 pts · 54d (1 reply)

Ohhhh this is interesting, I did also think it could be an older gemma but it using different context formatting kinda throws off my original assumption.

WHY isn't the dev talking about this yet 😭

biskit_the_tiger_tiger@lemmy.world · 0 pts · 54d

Forgot to point out that this new model (?) has the same errors I found with gemma4 when I ran that model locally. I've also been told about others having the same error.

Dev we need you 💔👄💔

thevegit0@lemmy.world · 2 pts · 55d (2 replies)

i wish we had a more live info updater o whatever

MeowerMisfit817@lemmy.world · 3 pts · 55d (1 reply)

Why does the dev even have a community if he won't see shit we say, right?

NewsSeeker24@lemmy.world · 2 pts · 54d

true honestly

ccufcc@lemmy.world · 2 pts · 56d

yep more info won't hurt