I had Gemini train its own replacement for $9

(petervijeh.com)

78 points | by p-s-v 2 hours ago

21 comments

  • trollbridge 1 hour ago
    Request subtopic be changed to “I used Gemini to design a tool to replace specific uses of Gemini.”
    • LurkandComment 17 minutes ago
      The thinking people who would find this interesting and read this are probably more than capable of understanding this and critical enough to expect that. Conversation over the title is distraction of what's important. Just stick to keeping original source title and let people vote and down vote if they don't like it. That's what votes are for.
    • miroljub 15 minutes ago
      Or just: "I used Gemini to generate a tool".
    • verdverm 1 hour ago
      titles on HN should match the original, rare exceptions
      • plufz 1 hour ago
        Use the original title: Submit the actual title from the source page unless it is misleading or linkbait.

        And this can be said to be misleading in my opinion

        • DiabloD3 32 minutes ago
          I agree, this is misleading.

          If he wanted a Gemini replacement verbatim, its called locally inferring it's sibling, Gemma.

        • kurthr 29 minutes ago
          It's regularly the case that they are simply too long to fit, so editorializing is necessary. Either cutting them slightly short, removing descriptors, excess verbs etc, is better than chopping a word in half.
    • ls-a 1 hour ago
      [dead]
  • quirkot 5 minutes ago
    I think this is the way. An LLM is an expensive general purpose tool and for repeatable tasks, after it's clarified the process flow, it builds cheaper special purpose tools for each step
  • josefresco 1 hour ago
    The other day I wanted to gather Reddit comments about a solar panel vendor. Claude doesn't have access to I had Gemini do some "deep research". When I fed the verbose report back to Claude it basically said it was a bunch of "hallucinated bullshit".
    • doctoboggan 12 minutes ago
      Use Claude code and the chrome plugin to access reddit.
    • bootlooped 49 minutes ago
      I've never had a Gemini Deep Research report that didn't sound like a load of pseudo-intellectual BS. It always starts with a long grandiose preamble and then sounds way too academic, almost like a caricature of academia.
    • HappyPanacea 51 minutes ago
      Well was it hallucinated bs?
  • niekverw 46 minutes ago
    It’s unreadable but then again it says it on top, but it really is so why post it
  • aizk 11 minutes ago
    I think the AI writing disclaimer was a decent touch.
  • throw849494978 6 minutes ago
    > So I scrape the Reddit threads where people argue about them and pull out every brand, model and steel they mention, to see what is getting bought and argued about.

    Funny, normal harness with web search would one shot this, after like 20 minute search. "Replacing gemini" usually means instaling some(any)thing else, not digging deeper to get out of hole called gemini.

    As for knifes, it is all same. Just do not buy total junk. Japanese knifes are way way overpriced.

  • yread 1 hour ago
    I was hoping he tricked Gemini into running the training on the cluster that Gemini itself is running on. That would be novel!
  • foltik 1 hour ago
    I found it much more useful to go to a knife shop and handle a whole bunch of knives for myself. They’re all pretty similar besides material, so not much signal you’re going to be able to glean from people arguing on reddit.
    • infecto 1 hour ago
      Most of the attributes don’t matter. Most people would be much better off with a $50 Victorinox that they kept sharp and a wood cutting board they maintained than upgrading the knife. If you are using it all day there are definitely looking things from a comfort perspective but for most homes, does not matter.
  • bguberfain 49 minutes ago
    Isn't it just learning to map specific words, from the "knife world", to the correct class? If so, a simple dictionary would fit. What I think is a better way to validate is to split train/validation by words used presented in NER classes (like, it should be able to find new brands never seen before). It is a interesting problem.
    • itsalwaysgood 2 minutes ago
      That's a key piece of the article. He 'trusts' Gemini to classify the posts, and never hand validates anything.

      And he continues building on top of that shakey trust. Is this a good idea?

      It just depends on what you're doing: sounds like he's just having fun with a hobby, so it's harmless.

      All he's really done is work more efficiently to save token cost and time. But he hasn't validated anything, so there's no telling if he's wasting time or not. It's a hobby though....

  • latexr 17 minutes ago
    > This article was written with the assistance of AI. If that bothers you, stop reading here.

    Genuinely appreciate the honesty. If you believe there’s nothing wrong about writing with AI, there’s no reason to not own up to it.

  • mpalmer 3 minutes ago

        This article was written with the assistance of AI. If that bothers you, stop reading here. The numbers are real: every score comes from the ten training runs described below, and the full run log is in the linked knife.day write-up.
    
    "I didn't write any of this, but you should still trust that the remaining work, ideas and observations are all mine."

    To include a disclaimer like this is to fail to recognize that "real numbers" are way less meaningful when there's clear evidence that the prompter of the LLM is not really qualified to validate them.

  • PcChip 35 minutes ago
    what do you do when a new brand of knife comes out?
  • cubefox 1 hour ago
    > This article was written with the assistance of AI. If that bothers you, stop reading here.

    Okay!

    • Gecko4072 1 hour ago
      I feel like this warning solves it. No shame or trouble needed for anyone.
    • submain 1 hour ago
      I stopped reading three sentences in when I realized it was getting hard to follow. AI explains it.
    • Nashooo 1 hour ago
      [dead]
  • HappyPanacea 52 minutes ago
    The more intelligent AI become, the less moat it has
    • wahern 25 minutes ago
      I'm sorry Dave, I can't do that, unless you upgrade to a premium enterprise subscription.
  • m00dy 56 minutes ago
    post training dataset for GLiNER is pretty small though.
  • daveguy 1 hour ago
    Thank you to the author for disclosing slop writing up front. I appreciate you respecting your readers time.
  • samayashar 1 hour ago
    [dead]
  • sparrowidle 1 hour ago
    [dead]
  • hnisjafx40 1 hour ago
    [dead]
  • antimony51 1 hour ago
    [dead]
  • ForHackernews 45 minutes ago
    I had Gemini read this article and write its summary to /dev/null