• lad@programming.dev
      link
      fedilink
      English
      arrow-up
      1
      ·
      edit-2
      1 分钟前

      That’s nice, albeit I want to point out for anyone wondering that this is only conjectured and not guaranteed:

      One of the properties that π is conjectured to have is that it is normal, which is to say that its digits are all distributed evenly, with the implication that it is a disjunctive sequence, meaning that all possible finite sequences of digits will be present somewhere in it.

      There is no guarantee for any specific sequence to appear in π, but for short chunks chances are better (it’s not really a probability, but it’s simpler to say and I can’t explain in details anyway). That’s because (from wiki):

      It is widely believed that the (computable) numbers √2, π, and e are normal, but a proof remains elusive.

  • Shanmugha@lemmy.world
    link
    fedilink
    arrow-up
    9
    ·
    2 小时前

    Here is another one, no compression:

    • read image
    • send a prompt “regenerate image, here is pixel-by-pixel description”
    • enjoy the result

    (sarcasm)

  • TrickDacy@lemmy.world
    link
    fedilink
    arrow-up
    15
    ·
    4 小时前

    a typical jpeg of 20 Mb

    Uh what. Literally should be the highest fucking possible quality from a $10K camera if it’s that big. My raw images aren’t even that big usually.

    • bstix@feddit.dk
      link
      fedilink
      arrow-up
      3
      ·
      33 分钟前

      Raw images is where you go wrong. You need at least 10 mb of metadata tags to achieve professional levels of file sizes. How can you even look at a picture without having a full description of all your childhood memories that lead you to take this beautiful picture of yesterdays mac’'n’cheese dinner. This is why we need more data centers.

    • PieMePlenty@lemmy.world
      link
      fedilink
      arrow-up
      9
      ·
      edit-2
      2 小时前

      Its not typical, but you can get 20 Mb+ jpegs out of an entry level 18MP dlsr. Especially if there’s lots of color and at like 5500x3300 resolutions and created with 100% quality preset.
      I checked my immich and I have some (and larger), but yeah, not exactly typical.

  • chunes@lemmy.world
    link
    fedilink
    arrow-up
    1
    ·
    2 小时前

    There is a future where AI can simply look up the DNA for any individual you specify and it can perfectly reconstruct what they look like, at any age.

  • OldGrayDog@fedinsfw.app
    link
    fedilink
    English
    arrow-up
    40
    ·
    8 小时前

    I’ve read that the Trump administration is hiring him to archive all of the Epstein files using that format.

  • MalReynolds@slrpnk.net
    link
    fedilink
    English
    arrow-up
    231
    ·
    10 小时前

    Trying to find it funny, but in 2026 it’s way too close to the bone and just makes me sad.

    • jafra@slrpnk.net
      link
      fedilink
      arrow-up
      16
      ·
      edit-2
      10 小时前

      Yeah. I first thought it’s about jpeg, then i read ai and i wasnt sure if its worrying stupidity reporting or bad satire. Edit: i thought '92 i mean

  • ShellMonkey@piefed.socdojo.com
    link
    fedilink
    English
    arrow-up
    31
    ·
    8 小时前

    On a similar note, I saw a story a bit back of someone saving input tokens by feeding the bot an image of a wall of text rather than the text itself and having it read the image via OCR.

    Satire and reality are too hard too distinguish these days.

  • DaddleDew@lemmy.world
    link
    fedilink
    arrow-up
    129
    ·
    edit-2
    10 小时前

    Amateur. I can compress entire seasons of a TV series to a few bytes. All I have to do is type its title in Netflix and then BAM, gigabytes of video come out.

  • AItoothbrush@lemmy.zip
    link
    fedilink
    English
    arrow-up
    15
    ·
    8 小时前

    The fuuuucking annoying part is that these weight trained models would be perfect for translation models, compression, etc. An llm is already kind of a really efficient lossy compressor but you could actually make it lossless and an actual compressor if used properly. But instead people are literally telling llms to translate instead of training models that are for translating. The technology isnt the problem itself, its the industry and capitalism.

    • ChaoticNeutralCzech@feddit.org
      link
      fedilink
      English
      arrow-up
      2
      ·
      edit-2
      3 小时前

      There are not many places where 95% compression of UTF-8 plain text could outweigh needing a 30GB model in memory and a significant fraction of current LLM inference cost to decompress it − and good luck convincing librarians to adopt it.