Seems legit

The Picard Maneuver@piefed.world · 15 days ago

Seems legit

DarkCloud@lemmy.world · 15 days ago

You can get offline versions of LLMs.

criss_cross@lemmy.world · 15 days ago

And gpt-oss is an offline version of chatgpt

linkinkampf19 🖤🩶🤍💜🇺🇦@lemmy.world · 15 days ago

First thing that came to mind: GPT4All

Ghostalmedia@lemmy.world · 15 days ago

I mean, most people have a local LLM in their pocket right now.

sp3ctr4l@lemmy.dbzer0.com · 14 days ago

Unless I am missing something:

Most people do not have a local LLM in their pocket right now.

Most people have a client app that talks to a remote LLM, which ‘lives’ in an ecologically and economically dubious mega-datacenter, in their pocket right now.

GamingChairModel@lemmy.world · 14 days ago

Plenty of the AI functions on phones are on-device. I know the iPhone is capable of several text-based processing (summarizing, translating) offline, and they have an API for third party developers to use on-device models. And the Pixels have Gemini Nano on-device for certain offline functions.

tetris11@feddit.uk · 14 days ago

My phone does speech-to-text flawlessly offline, it’s a crazy useful little LLM tool

sp3ctr4l@lemmy.dbzer0.com · 14 days ago

Oh!

Well, I didn’t know that.

I’m too poor to be able to afford such fancy phones.

Ghostalmedia@lemmy.world · 14 days ago

Gemini nano, Apple Intelligence On-device, etc.

khepri@lemmy.world · edit-2 15 days ago

Could you crunch an LLM into 700Mb that was still functional? Cause this looks like a fun thing to actually do as a joke.

Edit, I bet I could get https://huggingface.co/distilbert/distilgpt2 to run off a CD. How many tps am I gonna get guys 🤣

yellow [she/her]@lemmy.blahaj.zone · 15 days ago

Qwen3-0.6B is about 400 MB at Q4 and is surprisingly coherent for what it is.

khepri@lemmy.world · 15 days ago

That’s so crazy that an LLM capable of doing anything at all can be that small! That’s leaves room for like an entire .avi episode of family guy at dvd resolution on there, which is the natural choice for the remaining space of course

tetris11@feddit.uk · 14 days ago

a 4k episode of family guy using H265 (HEVC) and assuming not too many cutaway gags could produce a file about 240MB. You could probably fit a 480i episode of south park in the remaining 60MB

khepri@lemmy.world · 15 days ago

Wow, just popped it onto my very slow desktop and this little model rips haha. I really think tiny LLMs with a good LoRA on top are going to be a huge deal going forward

lime!@feddit.nu · edit-2 15 days ago

there’s also tinyllama, which is somewhere around 600MB. it’s hilariously inept. it’s like someone jpeg-compressed a robot.

also you’re only gonna load off of that cd once so it’ll perform fine.

tomiant@piefed.social · 15 days ago

FCKGW-RHQQ2-YXRKT-8TG6W-2B7Q8

Eager Eagle@lemmy.world · edit-2 15 days ago

make sure to disconnect the internet first

Ghostalmedia@lemmy.world · 15 days ago

CrAcKeD

NullPointerException@lemmy.ca · 15 days ago

That’s just Dr Sbaitso.

Björn@swg-empire.de · 15 days ago

It’s just audio of French farting cats.

Lemmyoutofhere@lemmy.ca · 15 days ago

Le pfffft.

Akasazh@feddit.nl · 14 days ago

My bet was on porn.

Or a copy of an old Encarta cd-rom

SSUPII@sopuli.xyz · edit-2 15 days ago

If we assume a CD, you can probably fit a 256M parameters model in it. But it will LOAD.

MacN'Cheezus@lemmy.today · 15 days ago

DVDs exist. They can fit approx. 7B params, enough to be somewhat productive.

MidsizedSedan@lemmy.world · 15 days ago

Isn’t it possible to download all of wikipedia, and it being surprisenly a small file size? Can it fit on a CD?

SSUPII@sopuli.xyz · 15 days ago

No

(English) 24,05GB without media. Adding media adds 428,36TB.

Axolotl@feddit.it · 15 days ago

No, you really can’t; It’s like 43 gb the text only version

puppycat (she/her)@lemmy.blahaj.zone · 15 days ago

yes you really can; it’s like 20-25 gb depending on how recent of a copy you have. I’ve been seeding wikipedia for almost a year and it barely takes any space on my computer

AmbiguousProps@lemmy.today · 15 days ago

It could fit on a BDXL disc.

masterspace@lemmy.ca · 15 days ago

You can fit text-only wikipedia on a normal Blu Ray as it’s only about 24GB. You can also easily fit Llama 3.1 or any of the other open, offline capable ai models as they’re only about 4GB.

Uriel238 [all pronouns]@lemmy.blahaj.zone · 15 days ago

Offline LLMs exist but tend to have a few terabytes of base data just to get started (e.g. before LORAs)

nomorebillboards@lemmy.world · 15 days ago

I thought it was more like 10-20GB to start out with a usable (but somewhat stupid) model.

Are you confusing the size of the dataset with the size of the model?

SubArcticTundra@lemmy.ml · 15 days ago

Does anyone know of any OSS LLMs that can search the web the way ChatGPT can?

yellow [she/her]@lemmy.blahaj.zone · edit-2 15 days ago

It’s not the LLM that does the web searching, but the software stack around it. On its own, an LLM is just a text completer. What you’d need a frontend like OpenWebUI or Perplexica that would ask the LLM for, say five internet search queries that could return useful information for the prompt, throw those queries into SearxNG, and then pipe the results into the LLM’s context for it to be used.

As for the models themselves, any decently-sized one that was released fairly recently would work. If you’re looking specifically for open-source rather than open-weight models (meaning that the training data and methodologies were also released rather than just the model weights), GPT-OSS 20B/120B and the OLMo models are recent standouts there. If not, the Qwen3 series are pretty good. (There are other good models out there, this is just what I remember off the top of my head.)

SubArcticTundra@lemmy.ml · 15 days ago

Thank you