It appears like the dev has changed the older deepseek (?) model to a gemma4 model. Hype, but also kind of a sudden change. I hope the dev talks about this to disclose whether this is actually real or we're just hallucinating.
It appears like the dev has changed the older deepseek (?) model to a gemma4 model. Hype, but also kind of a sudden change. I hope the dev talks about this to disclose whether this is actually real or we're just hallucinating.
8 Comments
tlgklxz@lemmy.world · 2 pts · 55d
I am pretty sure it's still deepseek model cause still 'stealing' from other LLm's.
Right now it acts as Gemma 2. Not even gemma 4. Literally Gemma 2 OR Llama 3.1 or rarely Gpt 4.0.
It can think, it can use start_of_turn or eot_id and im_start at the same time! Either because of the DeepSeek Tokenizer all the inputs I have sent are getting mixed up by the LLm or it's deepseek and pretending damn hard.
However, dev definitely tried to 'scrape' some instructions from the model. Here is my long why: https://www.reddit.com/r/perchance/comments/1uff0pq/ai_text_plugin_potential_issues_and_why/
ccufcc@lemmy.world · 1 pts · 55d
based.
biskit_the_tiger_tiger@lemmy.world · 0 pts · 54d
Ohhhh this is interesting, I did also think it could be an older gemma but it using different context formatting kinda throws off my original assumption.
WHY isn't the dev talking about this yet ðŸ˜
biskit_the_tiger_tiger@lemmy.world · 0 pts · 54d
Forgot to point out that this new model (?) has the same errors I found with gemma4 when I ran that model locally. I've also been told about others having the same error.
Dev we need you 💔👄💔
thevegit0@lemmy.world · 2 pts · 55d
i wish we had a more live info updater o whatever
MeowerMisfit817@lemmy.world · 3 pts · 55d
Why does the dev even have a community if he won't see shit we say, right?
NewsSeeker24@lemmy.world · 2 pts · 54d
true honestly
ccufcc@lemmy.world · 2 pts · 56d
yep more info won't hurt