jcoglan on Nostr: somehow I only recently learned that LLMs need the whole conversation fed back to ...
somehow I only recently learned that LLMs need the whole conversation fed back to them on each prompt so their i/o cost scales as O(n^2), something that would be considered completely unacceptable in almost any other production network-accessible software
Published at
2026-07-23 08:35:04 UTCEvent JSON
{
"id": "397d9f46f8e075fcd3afd8715ce9b0f738862aee22326688b05d25ea78cdb0f1",
"pubkey": "5b3ebe2b1d86dcf255dc3720a2cf89fc42c3fd192bc5133fd56793577675db2a",
"created_at": 1784795704,
"kind": 1,
"tags": [
[
"proxy",
"https://mastodon.social/@jcoglan/116968371274567873",
"web"
],
[
"proxy",
"https://mastodon.social/users/jcoglan/statuses/116968371274567873",
"activitypub"
],
[
"L",
"pink.momostr"
],
[
"l",
"pink.momostr.activitypub:https://mastodon.social/users/jcoglan/statuses/116968371274567873",
"pink.momostr"
],
[
"-"
]
],
"content": "somehow I only recently learned that LLMs need the whole conversation fed back to them on each prompt so their i/o cost scales as O(n^2), something that would be considered completely unacceptable in almost any other production network-accessible software",
"sig": "5f000ae22fb8d3f5d42489a3cbbad42ed7fba5ed9a765f71a10421e901212c05b24e73687c67d7afcdb6f2785ba778a6defd8fb90531ef26076ed3238806c2ca"
}