Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Owed a billion dollars in Nvidia stock (colo.to)
    135comments
  2. Thinking Fast and Slow in AI: The Role of Metacognition (arxiv.org)
    1comments
  3. Nissan's third generation e-POWER powertrain (nissan-global.com)
    16comments
  4. Ember-1 (fireworks.ai)
    196comments
  5. When did Google get so weird? (sancho.bearblog.dev)
    514comments
  6. Alan Kay's answer to “Did the ENIAC have a BIOS”? (quora.com)
    31comments
  7. There is more to code review than (automatable) detection (adaptivecapacitylabs.com)
    49comments
  8. The state of SIMD in Rust in 2026 (shnatsel.github.io)
    23comments
  9. Musk, the Movie (bleeckerstreetmedia.com)
    31comments
  10. Malleable software: Restoring user agency in a world of locked-down apps (2025) (inkandswitch.com)
    3comments
  11. Show HN: Lofi Cities – Pixel-art city nights with browser-generated lofi (loficities.com)
    87comments
  12. As A.I. makes law firms more efficient, clients ask: 'Where's my discount?' (nytimes.com)
    56comments
  13. Don't couple your Go code to GitHub (iain.rocks)
    85comments
  14. Lunar Terminator Paradox (secretsauce.net)
    38comments
  15. Microsoft drops Copilot+ branding from its new laptops (tomshardware.com)
    8comments
  16. Guitar amp and effects pedal built on the Waveshare ESP32-S3-Touch-AMOLED-2.06 (github.com/dashersw)
    11comments
  17. Self-Hosting on the Dark Web (alvarezrosa.com)
    33comments
  18. What I did at Recurse Center (thill.me)
    25comments
  19. Imp is a full port of DSPy to the BEAM (github.com/deepfates)
    6comments
  20. TabPFN and TabICL vs. tuned XGBoost: the model that doesn't train won 14/14 (efraingaray.com)
    3comments
  21. In an $80 motel room, a discovery to shed light on the origins of life (nytimes.com)
    83comments
  22. Deterministic Concurrency [video] (youtube.com)
    —discuss
  23. Replacing the old battery on rechargeable bike lights (jvns.ca)
    81comments
  24. Previously unheard recordings of John Coltrane, captured by Frank Tiberi (jazzwise.com)
    28comments
  25. Oral history of John Chowning, inventor of FM synthesis [video] (youtube.com)
    13comments
  26. Writing Efficient C++ Code (2013) (asawicki.info)
    85comments
  27. A New Experiment Meta-Strategy (chillphysicsenjoyer.substack.com)
    —discuss
  28. Research finds 485 chemicals in US pesticide products linked to breast cancer (theguardian.com)
    21comments
  29. The Cartesian Hand: In-Hand Manipulation with All-Linear Fingers (generalroboticslab.com)
    11comments
  30. Fakecloud: Local AWS cloud emulator for integration tests (fakecloud.dev)
    63comments

TabPFN and TabICL vs. tuned XGBoost: the model that doesn't train won 14/14

5 pointsby 2h agoefraingaray.com
3 comments
28m agoHN ↗

Tabular foundation models are one of those things where when you first look into it, it does not make sense as to why they would work so well but it does.

In drug property prediction domain, tabular foundation models coupled with another foundation model for molecules are pretty close to being the state of art.

Btw the article makes heavy use of AI or is written in that way A lot of unnecessary dramatic flair that gets very tiring

24m agoHN ↗

True.

Why "Every number is a measured one" and not "every parameter is measured?"

Stopped reading at that point.

LLMs are like autotune. Imagine the Rolling Stones auto-tuned.

19m agoHN ↗

  XGBoost’s search optimized accuracy, and afterwards I also compare by area under the curve. Which means the “fourteen of fourteen on AUC” is against a boosting model that was not tuned for that metric. Tuning it for AUC would probably improve it there; I did not measure that.

So, not a fair test?

I also did not get why Xgboost had to count its training time for the inference. You only train once. I guess in some scenario, where someone says, "I need the best model now, you have five minutes on this singular dataset", but I have never been in that situation.

I would feel better if scripts were released, because I am fairly dubious. I take it as a given that a tabular model has been pre-trained on all of the public benchmark datasets, but that is what it is.

The slop was so meandering, I am not sure what is truth or not.