Top
Best
New

Posted by logicallee 3 hours ago

MicroLLM Lab – Try 7 tiny LLM's in the browser(stateofutopia.com)
90 points | 36 comments
demibabs 1 hour ago|
Cool project, but I'd really suggest looking at the UI.

The text is too small and it's way too dense with information in general. Considering how simple this product is to use, it's kinda crazy that I have to scroll through over a page length of (mostly useless, AI-generated) information before getting to the actual interface.

Also what is going on with the footer (why does it link back to the site itself, why is it telling me to "serve over HTTP").

logicallee 48 minutes ago||
Thank you for the feedback, I'll think about how to incorporate the changes you've suggested.
phist_mcgee 1 hour ago||
This is the future of software, sloppy ui.
mgaunard 9 minutes ago||
> how would you compare your capabilities to that of claude fable 5.1 by anthropic

> Comparing your capabilities to that of claude fable 5.1 would be very similar. Both are stories about a clown who is a cunning and manipulative animal. They both have problems and challenges, and they face challenges to overcome them. However, they have different goals and approaches. The clown

I think that tells you all you need to know.

tolugenius 2 hours ago||
I did the default arithmetic with PetitGPT research-v1

>What is 2+2?

Answer

> To find 2 + 2, we need to add 2 to both sides of the equation.

> 2 + 2 = 4

> So, 2 + 2 = 4 + 2.

Brilliant

dotancohen 52 minutes ago||
That's not incorrect.

LLMs produce semantically correct sentences, not factually correct statements. Have we forgotten this so soon?

anyfoo 45 minutes ago||
It is, in every sense, incorrect. Which statement in this short snippet is "semantically correct"? (Better LLMs get this right, of course.)
tecleandor 36 minutes ago|||
GPT-2 124M is terrible...

  > what is 2+2?
Answer:

  > 3+3? 4+4? 5+6? 7+8?
  > Reply ~18000 0 ~10 min 2 By : 1-1: I'm a beginner. 3x2 is my best option, but if you're not sure about the other options then just go for it and try again
logicallee 4 minutes ago|||
GPT-2 is an interesting one because it is a February 2019 model. (You can see some information about it below the card if you click on the card.)

That was 2-3 years before the big "ChatGPT moment" (the highly coherent ChatGPT research preview was released in November 2022, I think it was ChatGPT 3.5). Back in 2019 the models really were not producing very coherent output. Now you can see it for yourself right in your browser :) Everything has come a really long way since then!

krackers 33 minutes ago|||
It's not instruct tuned looks like? It's closer to a base model rather than a chatbot.
logicallee 3 minutes ago||
That one is a 2019 model :) Years before the ChatGPT public preview.
logicallee 42 minutes ago||
I got the correct output for PetitGPT research-v1: https://ibb.co/0pP9DS2T
micw 2 minutes ago||
It's LLMs, not LLM's ;-)
kenzic 1 hour ago||
Really cool project. Giving web apps direct access to on-device models is something I’m excited about, and it’s cool to see the different approaches.

I’ve been working on a related proposal called the Web Models API, which explores a browser standard for an API that runs open-weight models on-device. Would love your thoughts: https://www.webmodels.dev

logicallee 1 hour ago|
I read your proposal, I think it's great! Where will the navigator get the model if the user agrees to download it? For this demonstration I just serve the models on my own server, but for larger models it may be an issue as they may not have direct download links even if they are open weights.
kenzic 1 hour ago||
Great question. Right now there isn’t a definitive answer, but it’s something that needs to be worked out. There would likely be a registry. The question is how to keep model IDs consistent: does each browser manage its own registry, or is there one shared across browsers?
logicallee 28 minutes ago||
Since you're asking for some kinds of permissions anyway, you could ask if the user is willing to also seed the model, p2p. (However, seeding files is not as popular as it used to be, many residential Internet connections don't have good upload.) If you have the capacity for it, your site webmodels.dev could act as a tracker and initial seed for any models. Then it could be the one central registry. It might get to be too much for you though, a lot of the open weights models are huge.
kenzic 25 minutes ago||
Interesting idea. I hadn't considered that. Thanks
willaaam 1 hour ago||
I like it as I'm vibecoding an (airgappable) browser AI workspace myself, but in terms of putting the models to use, just exposing the chat interface feels a bit limiting to me.

My take on this concept: https://github.com/willaaam/gemma-4-E2B-webgpu-vision

bhouston 2 hours ago||
I built something like this just last month, using a few of the same models, but I used ThreeJS's Three-Shading-Language abstraction to do it: https://three-llm.ben3d.ca/?model=qwen3.5-0.8b
vs4vijay 1 hour ago||
I have been experimenting with something similar here - https://sonistellar.com/lab/
jellyfiz 2 hours ago||
Similarly, if someone wants to hack some LLMs in their browser and burn some cycles, please feel free to try to break them here:

https://ai-attacks.neal.codes/

inventor7777 1 hour ago|
PetitGPT told me that

> "2+2 is 2."

Otherwise, a very neat demo. As others have said, the UI is VERY confusing, way too much stuff going on.

dvh 1 hour ago|
It is. For very small values of 2.
More comments...