Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. One engineer shipped 2k PRs a month to production. Verification is the key(thenewstack.io ↗)
    discuss
  2. A Hurricane of Pseudoscience(fasterik.substack.com ↗)
    discuss
  3. Ilha Da Queimada Grande(wikipedia.org ↗)
    discuss
  4. Google will no longer satisfy Zero Data Retention through OpenRouter routing(reddit.com ↗)
    discuss
  5. Zenodo performance: an update on the current situation(zenodo.org ↗)
    discuss
  6. Argentic – Stateless L402 proxy gate for machine data monetization(github.com/amarongrantham-dev ↗)
    discuss
  7. Apple Explains Why the iPhone 18 Pro Has a Variable Aperture(petapixel.com ↗)
    discuss
  8. Key symbols we lost to time, pt. 2: The Mac side(aresluna.org ↗)
    discuss
  9. Computers turn 1's and 0's into Text and Images [video](youtube.com ↗)
    discuss
  10. Pakistan 2076 Security, State and the Future(hassanali110.substack.com ↗)
    discuss
  11. Paris probes smart glasses use for suspected sexual harassment(siliconrepublic.com ↗)
    discuss
  12. Open WebUI: Self-Hosted AI Platform(openwebui.com ↗)
    discuss
  13. The Effect of CRTs on Pixel Art(datagubbe.se ↗)
    discuss
  14. Polyscope – Agentic coding environment for Laravel(getpolyscope.com ↗)
    discuss
  15. Secure VMs for Kubernetes: Hardening Kata Containers(srcreigh.ca ↗)
    discuss
  16. Mac Mini Alternatives for Local LLMs: M6, M5 and Strix Halo(terminalbytes.com ↗)
    1comments
  17. Enthusiast replaces ripped-off CPU pin in substrate – chip boots and hits 33% OC(tomshardware.com ↗)
    discuss
  18. Ask HN: How do non-resident founders meet others in SF?
    discuss
  19. China Stockpiled Oil, and Now It Could Dominate the Energy Landscape(nytimes.com ↗)
    1comments
  20. Multi-agent workflows to reproduce error logs and open PRs(nishantjani.com ↗)
    discuss
  21. The innovators under 35 shaping climate tech(technologyreview.com ↗)
    discuss
  22. Show HN: Reader – I made a Mac workspace for books, browser tabs, notes, and AI(github.com/marvy101 ↗)
    discuss
  23. Does AI Assistance Enhance or Erode Expertise? Evidence from 3-Month Experiment(nber.org ↗)
    discuss
  24. Jev example use cases from community(jevable.com ↗)
    discuss
  25. Checkout offered by banks and credit unions. Alternative to OneLink(paze.com ↗)
    discuss
  26. Montessori gave us great start. Still switched to public school for first grade(businessinsider.com ↗)
    discuss
  27. The paradox at the heart of AI and science – Terence Tao [video](youtube.com ↗)
    discuss
  28. Lawsuit: Illegal Agreement of Anthropic, OpenAI, SpaceXAI, Google on AI Slowdown(independent.co.uk ↗)
    discuss
  29. Unblock AI Shield – Non-invasive perimeter hardening with zero-disk RAM retentio(unblock-shield.vercel.app ↗)
    discuss
  30. My company let candidates use AI tools in interviews. Here's what I saw(reddit.com ↗)
    discuss

Tin: full-text search for Postgres

94 pointsby 3h agoplanetscale.com
45 comments
2h agoHN ↗

Becoming more the norm for them, Neki is the same.

Immediately rules out ever using them (though I don't currently have any problems that would benefit from that level of scale currently, have in the past though).

Postgres's license allows this but for me (personally) it leaves a bad taste.

Also it's not really "full-text search for Postgres" it's "full-text search for our hosted version of Postgres" so the title is a little misleading.

1h agoHN ↗

Why would you need super fast search for local testing?

1h agoHN ↗

It's Postgres search any coding agent can switch you to something else in 5 minutes

1h agoHN ↗

Why don’t you like Postgres’ license? It’s as permissive as a license gets.

1h agoHN ↗

Building non-open extensions on top of it, it's not the license I don't like, the bad taste is that they use something open extend it and keep part of it closed.

The license allows it but on the flip side it's vendor lock-in predicated on using something open as the base.

Fully proprietary no issue with that, full open, no issue with that, building proprietary on top of open is where the bad taste comes in.

For completeness, it's not them specifically either, the other cloud companies do similar things and I suspect in part the reason they don't open these extensions up is because the others will but then they are doing the same thing themselves.

1h agoHN ↗

When one develops using open-source software they have an obligation to follow the licenses. They also have a moral obligation to be respectful of the work upon which they’re building. And they have a social obligation to help improve that software where they can.

Those that develop on top of open-source have no obligation to give you their work for free.

You’d be surprised as to the amount of open-source contributions TIN drove towards Postgres, LLVM, and pgrx. And you’d be speechless at the amount of upstream work across all sorts of open-source PlanetScale does. Postgres 18.6, for example, is better for you today, in part, because of TIN. You’re welcome.

56m agoHN ↗

You’re welcome.

I didn't say thank you and don't presume I would, Planetscale acting in their own self interest by improving postgres upstream isn't deserving of thanks, any more than Intel upstreaming a bunch of Linux kernel work is or myriad other examples.

Corporations acting in their self interest isn't worth giving thanks for, neither is the work of the people paid to do work on their behalf.

I don't expect my employer to thank me, I expect them to pay me, I don't expect users of software I was paid to write to thank me because I did it for the money not out of altruism towards those users.

46m agoHN ↗

It’s clear that we’re on opposite ends of open-source ideology.

Good luck out there! 2026 is wild times!!

41m agoHN ↗

You as well, and yes it's very "may you live in interesting times" at the moment everywhere.

46m agoHN ↗

Everything else aside

Those that develop on top of open-source have no obligation to give you their work for free.

This is a pretty narrow view of open source.

Just as a simple example, GPLv3 and AGPLv3 are both considered Open Source and, depending on how you hold them, may obligate releasing work to customers essentially for free.

34m agoHN ↗

I thought I was clear:

When one develops using open-source software they have an obligation to follow the licenses.

Obviously what follows from one of those licenses is what you say.

I’m glad we agree!

1h agoHN ↗

The problem is that they don’t support bare metal. I’d love to use PlanetScale in our bare metal servers.

1h agoHN ↗

we support bare metal inside AWS, GCP, and very soon Azure

2h agoHN ↗

I struggle with FTS inside SQL (SQLite and MSSQL). There is often a fairly significant impedance mismatch between the relational concerns and how the documents need to be stored.

I've always preferred to use SQL as the system of record and then build/maintain an external Lucene index. Do we think these integral FTS capabilities are at the point where a hybrid architecture doesn't make sense anymore? How much customization exists in this provider?

1h agoHN ↗

I’m one of TIN’s developers and if you google my username you’ll see I’ve been in this space for a long time.

The answer to your first question is simply: yes

As far as your second question, what customization do you need that you believe TIN or PlanetScale doesn’t provide? These are things we can do, with alacrity.

1h agoHN ↗

in my experience it’s pretty common to find big inverted indexes for text directly in the database - not necessarily large docs but certainly free text records in volume. using bm25 and unicode’s breakiterator is a very good way to build it. like putting lucene in the database basically - makes a lot of sense when the database is already large. places that bend over backwards to move search out of the db are usually trying to avoid having a very large db (and often end up with one anyway, getting the worst of both worlds)

1h agoHN ↗

They also end up with all the infrastructure and processes necessary to keep the external search system in sync, resync/reindex, pkey shipping back to their source of truth in queries, application-side joins and enrichment between both sources. It’s brutal.

Having everything in one place eliminates entire classes of development and especially operational problems.

1h agoHN ↗

I think what we're seeing with every database company providing new full-text search capabilities is an example of AI coding productivity showing up in the real world.

It started with paradeDB and pg_search https://www.paradedb.com/blog/introducing-search

Timescale has pg_textsearch https://github.com/timescale/pg_textsearch

Neon and Databricks have Lakebase Search https://docs.databricks.com/aws/en/oltp/projects/lakebase-se...

Now PlanetScale.

AFAIK all of these are implementations of the BM25 algorithm. You can just tell an agent to read about BM25 and implement it in your system of choice. Cool to see. Seems like there's still a lot of juice to be squeezed out of how it's architected and integrated into each system, but you can't help but wonder if this will lead to aggressive commodification

1h agoHN ↗

ParadeDB's implementation builds on the Tantivy crate, which predates AI coding.

1h agoHN ↗

There is a lot of truth to this, but it's also very much down to domain experts being able to do this to move faster.

Planetscale (assuming they used a agentic development practice) will have pulled this off, to the level of performance that they have, because they have a team of very highly experienced Postgres developers. Their knowlage of Postgres internals will have given them the insights needed to steer the models to a plan that used the architecture as described in the post. That's not something a model can do on its own*

World experts + LLMs = moving mountains.

(* we're obviously seeing something a little different from inside the research teams in the labs. They are showing that the models, when you burn the level of tokens only they can, are able to do novel things from the models own insights.)

53m agoHN ↗

Seems like a lot of this knowledge was encoded into the blog post. I wonder if given this post and access to a planet scale instance to compare with, how close an agentic agent could get.

47m agoHN ↗

I think we can frame it as LLMs materializing existing potential. It seems like there needs to be an underlying potential to tap into, without which, the results could be slop.

1h agoHN ↗

What’s funny is I worked with a company with planet in the name Who could really use a full text search that was great in the Postgres

1h agoHN ↗

Interestingly enough SQLites FTS supports Lucene queries out of the box with great performance characteristics. IIRC only writes become pretty slow after a while. I’ve always wondered what exactly would prevent PostgreSQL from strapping that implementation into its own database. My experience with ts_query hasn’t been particularly rosy. It can be better than LIKE but only marginally so and at the cost of insane index sizes… If this extension becomes open source and we can test it out in the real world I’m sure there’s a sweet spot

1h agoHN ↗

SQLite FTS relies on shadow B-trees under single-writer locks. Postgres index access methods must map postings directly to physical ctid tuples, surviving MVCC visibility checks and heap tuple churn.

1h agoHN ↗

What are the advantages of Tin over using ts_vector with gin and gist indexes?

24m agoHN ↗

From the benchmarks deep in the document, TIN is much faster than built in text search (tested against GIN, which is itself much faster than GiST for text search.)

58m agoHN ↗

It’s just me or there are others who keep seeing these updates and think mongodb had all of this years ago?

Seriously so happy to be running our production stack on mongo.

52m agoHN ↗

Is this an ad? Postgres has had search for more than a decade.

26m agoHN ↗

Postgres has had search for more than a decade.

More than two decades (it moved to core from contrib in version 8.3 in 2008, but it was available in contrib since 7.4 in 2003.)

53m agoHN ↗

I want to try Planetscale... but we're addicted to (and totally dependent on) Neon's branching model. They really got us hooked on that!

53m agoHN ↗

I suppose it all depends on the scale of your project, but I've had pretty good luck using both MySQL's and SQLite's FTS capabilities. Surprised to hear that Open Source champion Postgres didn't have up-to-snuff FTS up to now...?

30m agoHN ↗

It has FTS built in and has had it for a VERY long time.

52m agoHN ↗

Please read the Postgres manual. It has incredible built-in search capability.

34m agoHN ↗

I think it’s fantastic.

You want to use this thing instead?

31m agoHN ↗

If you read deep into this, they claim much better performance than the built in search; they also imply that the built-in search is missing features they provide but don’t make clear which ones (I think it is just support in the same index for queries covering other conditions on other columns, because every other feature they claim seems to line up with the built in search features, which have been around for about 20 years.)

8m agoHN ↗

is this 21st century embrace, extend, extinguish

I assume here you're talking about Amazon's modus operandi?

29m agoHN ↗

Postgres does have pg_fts (tsvector/tsquery/tsrank) which is a quite sophisticated full text search package integrated with functional indexing and query optimization. Why would I use something vibecoded that isn't part of core Postgres instead?