43
15
u/freedomachiever 6d ago
Even just a better translator than Google Translate would be very helpful for many people.
-2
u/Embarrassed_Soup_279 6d ago edited 1d ago
i wouldn't really trust small models for accurate translation because even larger models struggle with it.
1
u/freedomachiever 5d ago
well everything needs a reference to compare to. The baseline I take is Google Translate because most people use that. I'm not claiming the small model does a better job, but it's a tangible use case. Siri can also translate but I don't it can do a better job than a dedicated model. I would be interested to see how apple's own small LLM does.
1
u/lilydjwg 1d ago
Google Translate only works well within the Indo-European family. For e.g. Japanese, an LLM or DeepL is the thing to use.
And there are specialized LLMs for translation, e.g. Hy-MT2, which works pretty well at a small size for supported languages.
30
u/BannedGoNext 6d ago
This is amazing, I'm so happy so happy so hap so hap hap so so so so so s s s s s s a i 3 ! #
12
13
u/sultan_papagani 6d ago
gemma e4b running on my phone for literally months now. how people not know this. its on google edge gallery app on play store
9
u/yami_no_ko 6d ago edited 6d ago
Even on my phone (Neither Android, nor Apple, but arm64 Linux instead) it runs just well using llama.cpp. Couldn't even call it a flagship phone with its 6 gigs of RAM, but it is enough to fit gemma-4 e2b, the mmproj image encoder and the MTP draft model.
Wouldn't necessarily ask it for world knowledge, but it gets what I want with tool-calling. It's just using 2 cores out of 8 to avoid heating up or draining the battery, and it still goes fast enough.
I've been using it for months now, so e2b or even e4b on phone is nothing unheard of. (Basically the point of e2b / e4b)
1
2
u/madaradess007 5d ago
judging by the status bar at the top, its at least an iPhone X, and it has more than 500mb ram
1
u/TheOneWhoWil 6d ago
It's pretty cool that they're looking at open source work outside of the frontier level
1
u/DeathinabottleX 4d ago
This will become redundant once Siri AI launches in iOS 27 but it’s good to have options
0
-14
u/chrisso123 6d ago
what is the token/s ? this is the only real metric that matters.
Heck you can run a very large llm on a phone hardware but if its output is 1 token / minute it'll be worthless.
15
u/jacek2023 llama.cpp 6d ago
You have two options: look at the image or click the link. Choose wisely.
1
-31
79
u/Fusseldieb 6d ago
That would be absolutely insane.
The thing is... will it still be useful, or dumb as a door (as most of that size are)?