Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Owed a billion dollars in Nvidia stock (colo.to)
    155comments
  2. Thinking Fast and Slow in AI: The Role of Metacognition (arxiv.org)
    2comments
  3. Nissan's third generation e-POWER powertrain (nissan-global.com)
    37comments
  4. Ember-1 (fireworks.ai)
    199comments
  5. When did Google get so weird? (sancho.bearblog.dev)
    537comments
  6. Alan Kay's answer to “Did the ENIAC have a BIOS”? (quora.com)
    33comments
  7. There is more to code review than (automatable) detection (adaptivecapacitylabs.com)
    54comments
  8. Deterministic Concurrency [video] (youtube.com)
    —discuss
  9. The state of SIMD in Rust in 2026 (shnatsel.github.io)
    23comments
  10. Don't couple your Go code to GitHub (iain.rocks)
    93comments
  11. Malleable software: Restoring user agency in a world of locked-down apps (2025) (inkandswitch.com)
    6comments
  12. Self-Hosting on the Dark Web (alvarezrosa.com)
    35comments
  13. Microsoft drops Copilot+ branding from its new laptops (tomshardware.com)
    15comments
  14. Show HN: Lofi Cities – Pixel-art city nights with browser-generated lofi (loficities.com)
    88comments
  15. Guitar amp and effects pedal built on the Waveshare ESP32-S3-Touch-AMOLED-2.06 (github.com/dashersw)
    16comments
  16. What I did at Recurse Center (thill.me)
    25comments
  17. Was Silent Reading Unusual During Augustine's Time? (historyofinformation.com)
    1comments
  18. Reading’s Bayeux Tapestry (diamondgeezer.blogspot.com)
    8comments
  19. Imp is a full port of DSPy to the BEAM (github.com/deepfates)
    6comments
  20. TabPFN and TabICL vs. tuned XGBoost: the model that doesn't train won 14/14 (efraingaray.com)
    3comments
  21. In an $80 motel room, a discovery to shed light on the origins of life (nytimes.com)
    84comments
  22. Replacing the old battery on rechargeable bike lights (jvns.ca)
    83comments
  23. Oral history of John Chowning, inventor of FM synthesis [video] (youtube.com)
    15comments
  24. Previously unheard recordings of John Coltrane, captured by Frank Tiberi (jazzwise.com)
    28comments
  25. Lunar Terminator Paradox (secretsauce.net)
    41comments
  26. Goodbye to the Hard Parts That Never Mattered (jakegoldsborough.com)
    6comments
  27. Writing Efficient C++ Code (2013) (asawicki.info)
    88comments
  28. Research finds 485 chemicals in US pesticide products linked to breast cancer (theguardian.com)
    22comments
  29. The Cartesian Hand: In-Hand Manipulation with All-Linear Fingers (generalroboticslab.com)
    11comments
  30. A New Experiment Meta-Strategy (chillphysicsenjoyer.substack.com)
    —discuss

TabPFN and TabICL vs. tuned XGBoost: the model that doesn't train won 14/14

8 pointsby 2h agoefraingaray.com
3 comments
1h agoHN ↗

Tabular foundation models are one of those things where when you first look into it, it does not make sense as to why they would work so well but it does.

In drug property prediction domain, tabular foundation models coupled with another foundation model for molecules are pretty close to being the state of art.

Btw the article makes heavy use of AI or is written in that way A lot of unnecessary dramatic flair that gets very tiring

1h agoHN ↗

True.

Why "Every number is a measured one" and not "every parameter is measured?"

Stopped reading at that point.

LLMs are like autotune. Imagine the Rolling Stones auto-tuned.

1h agoHN ↗

  XGBoost’s search optimized accuracy, and afterwards I also compare by area under the curve. Which means the “fourteen of fourteen on AUC” is against a boosting model that was not tuned for that metric. Tuning it for AUC would probably improve it there; I did not measure that.

So, not a fair test?

I also did not get why Xgboost had to count its training time for the inference. You only train once. I guess in some scenario, where someone says, "I need the best model now, you have five minutes on this singular dataset", but I have never been in that situation.

I would feel better if scripts were released, because I am fairly dubious. I take it as a given that a tabular model has been pre-trained on all of the public benchmark datasets, but that is what it is.

The slop was so meandering, I am not sure what is truth or not.