Elektrine lite

← Feed

@tay@tech.lgbt

Post #2094895

2026-05-06 12:58 UTC

@dalias @lazza @Vivaldi well, i think the reason it's in the browser itself is because a) these files are, as mentioned, massive, so you don't want to have each site store their own, and b) i don't know if the WebGPU APIs are there yet for doing LLM inference at comparable speed i'm not opposed to the APIs in principle - LLM technology is simply not going away, and there are actually decent use cases for them, and I oppose the current status quo of just shipping it all to OpenAI or Anthropic's cloud server My biggest concern is that no two LLM models will ever behave in the same way as each other, so sites & users that expect Google's Gemini model, wouldn't have the same experience as if say Safari had this with one of their on device models. And maybe by some pure miracle we could convince all the implementations to standardise on one model (not happening) - you can't ever update that model as newer ones are developed without breaking those expectations (also why the extension model wouldn't really work)

Replies (2)

  • @dalias@hachyderm.io 2026-05-06 13:01

    @tay @lazza @Vivaldi Fuck off slop apologist. Yes it is going away. We're making it go away.

    Open ##2094896

  • @tael@yiff.life 2026-05-06 13:19

    @tay I'm not interested in using it and therefore it has no place on my system. I don't care if websites might want to use it in the future. I didn't consent to that and disagree that it wouldn't be better for each website to store their own; there are, after all, many more users than there are websites.

    Open ##2094909