Join the discussion

Write your take first — we'll ask for email only when you're ready to publish.

  • Hacker News
  • The single best animated data visualization to demonstrate the stochastic nature of LLM models: an animation of the probability of solving a long multiplication problem over several runs. [0]

    [0] https://adamsohn.com/reasoning-grid/#walk-the-surface

  • Chanterelles and false chanterelles are especially tricky. They fruit at the same time in the same areas. I have to do a 2nd pass after I have washed and dried my haul.

    https://www.mushroom-appreciation.com/wp-content/uploads/202...

  • Is it possible to reliably (>99%) identify mushrooms from a single image, with any method?

    All image datasets have an intrinsic minimum error that comes from the limit of information in the data, regardless of what model or method you use to analyze it.

    You might not be able to tell apart two similar-looking mushroom species just by looking.

  • There are old mushroom hunters, and there are bold mushroom hunters.

    There are no old bold mushroom hunters.

  • I've used Gemini quiet a bit for foraging. It's definitely a "trust, but verify" situation. It's really good at at least getting you in the right lane to verify, rather than having to comb hundreds of pictures.

    So even if it's wrong, a short verification usually reveals why it was wrong. I haven't had a situation yet where it was catastrophically wrong (i.e. the mushroom it called out looks nothing like the mushroom imaged). It also will generally warn of lookalikes.

    I also use it a lot for general plant ID, and it's really impressive there too, with much lower stakes (I'm not eating those).

  • Ha! I just had a very interesting conversation with ChatGPT about a mushroom that I found in my front lawn. I am 99% sure it was an edible oyster mushroom which included analysis the of the size, growing substrate (dead ash tree root which turns out to be a useful identifying characteristic -- I learned a lot about mycology in a short time). Plus I don't use fertilizers, herbicide, pesticides, fungicides (obviously) on my lawn so I think it would have been quite tasty. Did I eat it? Hell, no. But then I got to thinking, would I trust a human to correctly identify it? I am actually not sure I would either. Penalty is just too great for misidentification LLM or not.
  • Trusting an LLM with a decision that can potentially be fatal is a good Darwin test I suppose.

    Sadly I don’t think the general populace understands that LLMs are unreliable. (And even people in tech can vastly overestimate the capabilities.. at least judging by the insane spending on them).

Explore Birbla archives