Technology

76248 readers

4718 users here now

This is a most excellent place for technology news and articles.

Our Rules

Follow the lemmy.world rules.
Only tech related news or articles.
Be excellent to each other!
Mod approved content bots can post up to 10 articles per day.
Threads asking for personal tech support may be deleted.
Politics threads may be removed.
No memes allowed as posts, OK to post as comments.
Only approved bots from the list below, this includes using AI responses and summaries. To ask if your bot can be added please contact a mod.
Check for duplicates before posting, duplicates may be removed
Accounts 7 days and younger will have their posts automatically removed.

Approved Bots

founded 2 years ago

MODERATORS

L3s@lemmy.world

enu@lemmy.world

technopagan@lemmy.world

L4s@lemmy.world

L3s@hackingne.ws

L4s@hackingne.ws

Hackers can read private AI-assistant chats even though they’re encrypted (arstechnica.com)

submitted 2 years ago by floofloof@lemmy.ca to c/technology@lemmy.world

5 comments fedilink hide all child comments

you are viewing a single comment's thread
view the rest of the comments

[–] kevincox@lemmy.ml 38 points 2 years ago (1 children)

This is pretty clever. As I understand it.

Because LLMs are slow most of them stream the response to the user.
The response is streamed as text, but generated in tokens.
This means that each "chunk" leaks the length of the text corresponding to the token.
You can then use heuristics to guess the text of the response based on the token lengths.

This is a good reminder any time you are sending content in small chunks over an encrypted channel, many encrypted channels don't provide protection against size leaks by default.

It seems there are a few easy solutions to this:

Send the token IDs (as fixed-size integers) over the network rather than the text.
Pad the text representations of the tokens to a fixed length.
Batch the tokens more (and maybe add padding) to produce bigger chunks and obscure individual token size.

These still all leak the approximate length of the response, but that is probably acceptable.

[–] PlexSheep@feddit.de 7 points 2 years ago (1 children)

That actually is really really interesting. Thanks for giving the tldr. Do token lengths vary that much?

[–] kevincox@lemmy.ml 6 points 2 years ago

Absolutely. They are sort of a compression scheme so the tokens contain different numbers of characters based on how frequent that string is. So common words like "the" will typically be one token, or maybe even common phrases like "I am". On the other hand rare punctuation such as "~" may be its own token. There will also be tokens for many common prefixes and suffixes such as "non" and "n't". The tokens of each model are different but they definitely vary in length.