Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Show HN: Bi-temporal Graph RAG in Postgres (new documents retire old facts)(crajah.github.io ↗)
    discuss
  2. T. Rex Had a Body Temperature of 97 Degrees(nytimes.com ↗)
    1comments
  3. Economic Policy for AGI(deepmind.com ↗)
    discuss
  4. Europe's e‑waste is a valuable resource – let's not throw it away(theconversation.com ↗)
    discuss
  5. Software Engineer: I'm Tired of Pretending [video](youtube.com ↗)
    discuss
  6. Show HN: S-Roll – local AI video understanding and clipping for Mac(saliency.dev ↗)
    discuss
  7. Show HN: Built a Simulation of Autonomous Agents(agents.london ↗)
    discuss
  8. Noam Brown – Agent swarms, alignment, & recursive self-improvement(dwarkesh.com ↗)
    discuss
  9. Show HN: Custom Relevance – Describe what matters and let Jev rank it(foglight.co ↗)
    discuss
  10. Feeling overwhelmed by the AI doom loop? Here's the essential reading list(theguardian.com ↗)
    discuss
  11. Google Noto 3D Emoji Design Process(design.google ↗)
    1comments
  12. Becoming a Benchmark(terrytao.wordpress.com ↗)
    discuss
  13. Sourcerer: Fast Git client written in C++ version 1.0.24 released(sourcererapp.com ↗)
    discuss
  14. The Download: mice with part-human brains and climate tech innovators(technologyreview.com ↗)
    discuss
  15. Should OpenAlex publish a novelty score? A vibe check(openalex.org ↗)
    discuss
  16. Show HN: Smart Crosshair – A lightweight C++/Win32 adaptive contrast crosshair(steampowered.com ↗)
    discuss
  17. I didn't sign the Fields medallists' letter(terrytao.wordpress.com ↗)
    1comments
  18. Towards Self-Driving Codebases(detail.dev ↗)
    discuss
  19. 4-Bit Rotational Quantization: -45% RAM, <1% recall drop vs. TurboQuant(weaviate.io ↗)
    discuss
  20. Show HN: Repodify: Make Podcasts Out of Podcasts(repodify.app ↗)
    1comments
  21. Palantir's Karp: AI needs to have 'reasonable guidelines,'(cnbc.com ↗)
    2comments
  22. Aegis: Zero-GC 64-byte cache-aligned memory arena in C++20 (1B ops in 0.649s)(github.com/markbgilbert ↗)
    discuss
  23. Ask HN: Co-Founder(s). Do I need any? How would I even find them?
    discuss
  24. From Stonemasons to Carpenters(thelastsoftwareengineer.substack.com ↗)
    discuss
  25. I Made Turn-Based Combat into a Database(louisxu3.substack.com ↗)
    discuss
  26. Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data(arxiv.org ↗)
    discuss
  27. Show HN: Pappice live demo in browser via WASM(pappice.eu ↗)
    discuss
  28. Show HN: AutoBot – live voice control for long-running AI work(github.com/demeyer1 ↗)
    discuss
  29. Show HN: Craigslist for agent skills, curated by a human(skillbay.sh ↗)
    discuss
  30. The Bicycle and the Algorithm: Amber Case on Why AI Has It Backwards(designwhine.com ↗)
    discuss

I had Gemini train its own replacement for $9

83 pointsby 3h agopetervijeh.com
40 comments
3h agoHN ↗

Request subtopic be changed to “I used Gemini to design a tool to replace specific uses of Gemini.”

3h agoHN ↗

titles on HN should match the original, rare exceptions

3h agoHN ↗

Use the original title: Submit the actual title from the source page unless it is misleading or linkbait.

And this can be said to be misleading in my opinion

2h agoHN ↗

I agree, this is misleading.

If he wanted a Gemini replacement verbatim, its called locally inferring it's sibling, Gemma.

2h agoHN ↗

It's regularly the case that they are simply too long to fit, so editorializing is necessary. Either cutting them slightly short, removing descriptors, excess verbs etc, is better than chopping a word in half.

1h agoHN ↗

The thinking people who would find this interesting and read this are probably more than capable of understanding this and critical enough to expect that. Conversation over the title is distraction of what's important. Just stick to keeping original source title and let people vote and down vote if they don't like it. That's what votes are for.

1h agoHN ↗

Or just: "I used Gemini to generate a tool".

3h agoHN ↗

I found it much more useful to go to a knife shop and handle a whole bunch of knives for myself. They’re all pretty similar besides material, so not much signal you’re going to be able to glean from people arguing on reddit.

2h agoHN ↗

Most of the attributes don’t matter. Most people would be much better off with a $50 Victorinox that they kept sharp and a wood cutting board they maintained than upgrading the knife. If you are using it all day there are definitely looking things from a comfort perspective but for most homes, does not matter.

1h agoHN ↗

That's called taking responsibility for your actions and being an informed consumer or a critical thinker.

Unfortunately that is completely undoable for most people nowadays. That's seen as "too much work" for most people. That's something that people see as a waste of time and should be done by something else for them. A knife should be a 5 second one-click buy, then when it sucks people will complain that all their options are bad.

You spent time/money traveling to a shop? In 2026!? Omg what a waste of time! Why didn't you aggregate the 500 product review sites and Reddit comments to pick 5 possible knives? (Sarcasm)

3h agoHN ↗

I was hoping he tricked Gemini into running the training on the cluster that Gemini itself is running on. That would be novel!

3h agoHN ↗

This article was written with the assistance of AI. If that bothers you, stop reading here.

Okay!

3h agoHN ↗

I feel like this warning solves it. No shame or trouble needed for anyone.

3h agoHN ↗

I stopped reading three sentences in when I realized it was getting hard to follow. AI explains it.

3h agoHN ↗

Thank you to the author for disclosing slop writing up front. I appreciate you respecting your readers time.

2h agoHN ↗

The other day I wanted to gather Reddit comments about a solar panel vendor. Claude doesn't have access to I had Gemini do some "deep research". When I fed the verbose report back to Claude it basically said it was a bunch of "hallucinated bullshit".

2h agoHN ↗

I've never had a Gemini Deep Research report that didn't sound like a load of pseudo-intellectual BS. It always starts with a long grandiose preamble and then sounds way too academic, almost like a caricature of academia.

1h agoHN ↗

Use Claude code and the chrome plugin to access reddit.

2h agoHN ↗

post training dataset for GLiNER is pretty small though.

2h agoHN ↗

The more intelligent AI become, the less moat it has

2h agoHN ↗

I'm sorry Dave, I can't do that, unless you upgrade to a premium enterprise subscription.

2h agoHN ↗

Isn't it just learning to map specific words, from the "knife world", to the correct class? If so, a simple dictionary would fit. What I think is a better way to validate is to split train/validation by words used presented in NER classes (like, it should be able to find new brands never seen before). It is a interesting problem.

1h agoHN ↗

That's a key piece of the article. He 'trusts' Gemini to classify the posts, and never hand validates anything.

And he continues building on top of that shakey trust. Is this a good idea?

It just depends on what you're doing: sounds like he's just having fun with a hobby, so it's harmless.

All he's really done is work more efficiently to save token cost and time. But he hasn't validated anything, so there's no telling if he's wasting time or not. It's a hobby though....

28m agoHN ↗

It wasn’t as simple as just mapping brands… for one I don’t have a complete dictionary of all the brands that could be mention.

Other problems are:

Brand names that match common English words, brand names that are also item models used by another brand… etc

2h agoHN ↗

It’s unreadable but then again it says it on top, but it really is so why post it

2h agoHN ↗

I had Gemini read this article and write its summary to /dev/null

2h agoHN ↗

what do you do when a new brand of knife comes out?

1h agoHN ↗

This article was written with the assistance of AI. If that bothers you, stop reading here.

Genuinely appreciate the honesty. If you believe there’s nothing wrong about writing with AI, there’s no reason to not own up to it.

1h agoHN ↗

I think the AI writing disclaimer was a decent touch.

1h agoHN ↗

So I scrape the Reddit threads where people argue about them and pull out every brand, model and steel they mention, to see what is getting bought and argued about.

Funny, normal harness with web search would one shot this, after like 20 minute search. "Replacing gemini" usually means instaling some(any)thing else, not digging deeper to get out of hole called gemini.

As for knifes, it is all same. Just do not buy total junk. Japanese knifes are way way overpriced.

1h agoHN ↗

I think this is the way. An LLM is an expensive general purpose tool and for repeatable tasks, after it's clarified the process flow, it builds cheaper special purpose tools for each step

1h agoHN ↗

100%. We now have super general tools that reduce the cost to build other specific tools. My favorite thing with LLMs has been building a ton of little utilities for work and personal that I could've built before, but never had the time at work or the want to spend time on in my personal time.

1h agoHN ↗

    This article was written with the assistance of AI. If that bothers you, stop reading here. The numbers are real: every score comes from the ten training runs described below, and the full run log is in the linked knife.day write-up.

"I didn't write any of this, but you should still trust that the remaining work, ideas and observations are all mine."

To include a disclaimer like this is to fail to recognize that "real numbers" are way less meaningful when there's clear evidence that the prompter of the LLM is not really qualified to validate them.

1h agoHN ↗

The reader assigns value to what is read. I found some value in what I read, even with all the holes.

It's a different way to think about a problem, even if it's not applicable in all areas.

1h agoHN ↗

Unfortunately, advertisers are getting smarter and using bots to praise their own products on Reddit. Thanks to training on genuine comments, some models are very good at sounding like a human commenter, and can easily generate a comment history with diverse interests to appear human, making them basically undetectable. So it seems like this method will, at some point, not identify the company with the best knife, but the one with the most ad spend on bot comments.

1h agoHN ↗

who the fuck is using reddit for product recommendations? reddit was astroturfed in like 2015.

1h agoHN ↗

If not earlier. Reddit has been highly manipulated since at least 2015.

I used to buy/sell reddit usernames for astroturfing and it was always hilarious seeing one of my usernames on the front page.

58m agoHN ↗

after 24 minutes on a Tesla T4 […] about $2.50 of GPU time

Who's the scam cloud provider who sells T4 GPUs for $6.25/hour?! That's B300 territory!

34m agoHN ↗

I wish they told us which model wrote the article so i can be sure to avoid it in the future. This is brutal to try and read