Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Fujitsu launches made-in-Japan next-generation CPU FUJITSU-MONAKA(global.fujitsu ↗)
    107comments
  2. Rate limits on GitLab.com are changing(about.gitlab.com ↗)
    37comments
  3. Artificial intelligence now beats some of the best human forecasters(economist.com ↗)
    22comments
  4. CrowdSec Source Code Leak(crowdsec.net ↗)
    1comments
  5. LLM Classification Is Feature Engineering(minimallysufficient.com ↗)
    8comments
  6. Whoisinspace.com/(whoisinspace.com ↗)
    5comments
  7. Launch HN: Skillsync (YC W26) – AI chat sessions made portable across agents(skillsync.com ↗)
    discuss
  8. One Year of Sponsored Servo Development(servo.org ↗)
    123comments
  9. Vinix – A modern operating system written in V(vinix-os.org ↗)
    7comments
  10. hister(github.com/asciimoo ↗)
    discuss
  11. Show HN: Die With Me – Claude and Codex rate limits as AIM away messages(diewithme.co ↗)
    2comments
  12. CCC invites all model citizens to 40C3(ccc.de ↗)
    64comments
  13. Nvidia announces native GPU programming in Rust(nvidia.com ↗)
    352comments
  14. Show HN: Share your AI Setup, Learn from others(mysetup.ai ↗)
    30comments
  15. My temporary PHP fix from 2014 has nearly 20M installs. Today I'm deprecating it(jakeasmith.com ↗)
    70comments
  16. The Relation Between Mathematics and Physics by Paul Dirac (1939)(cam.ac.uk ↗)
    36comments
  17. Keys Not Included: recovering the signing keys for US driver's license barcodes(ryan.science ↗)
    125comments
  18. Mastering Layout Engines in Graphviz: Dot vs. Neato vs. Twopi vs. Circo(visual-paradigm.com ↗)
    1comments
  19. GLM Built Its Own Inference Infrastructure(z.ai ↗)
    201comments
  20. Ask HN: How to recover Google auth after phone stolen?
    11comments
  21. Show HN: I built a new version of my fun spatial 3D online meeting app(flat.social ↗)
    48comments
  22. Grand MS-DOS Gaming General MIDI Showdown(johnnovak.net ↗)
    discuss
  23. Xiaomi Mimo 2.6 live post-training dashboard(xiaomi.com ↗)
    147comments
  24. Better Vector Search for Long Documents: Chunking Inside Manticore Search(manticoresearch.com ↗)
    10comments
  25. Lucasart's Afterlife(togameforlife.wordpress.com ↗)
    40comments
  26. Online Z3 Guide(microsoft.github.io ↗)
    14comments
  27. Cloudflare/Security-Audit-Skill(github.com/cloudflare ↗)
    33comments
  28. Show HN: An e-ink frame that hears birds and draws them as 1800s illustrations(github.com/arnegiacomo ↗)
    250comments
  29. Comparison of Malloc() Algorithms(egbert.net ↗)
    32comments
  30. Why Germany Is Building an Ark for U.S. Climate Data(yale.edu ↗)
    1comments

LLM Classification Is Feature Engineering

27 pointsby 1h agominimallysufficient.com
7 comments
46m agoHN ↗

I don't understand the point they are trying to make.

It's very often (always?) the case that something general also solves particular problems.

    A sorting algorithm is an implementation of min()

    A parser also is a syntax checker.

    A route planner is a reachability checker. 

    A computer algebra system is a basic arithmetic calculator.

    A general constraint solver is a Soduku hint maker.

It's true that LLM output can be used as an input to another classifier, this is also true of any classifier. The improvement on top of the straight LLM classification is relatively small, and I would argue that working on the prompt or just including in the prompt for the LLM what features might be useful to consider would likely work even better.

Fundamentally I read this article as: We want to build a simpler, dumbed down clone of Mathematica, so we cobbled together the following pieces... We also needed a way to do arithmetic, so we also include a copy of Mathematica to do basic arithmetic.

43m agoHN ↗

I took the point as: don't make the LLM the classifier. Use it to turn messy input into useful features, then let a normal model make the actual decision. That gives you thresholds/calibration you can inspect.

What I'm not sure about is how stable those features are when you switch the underlying LLM or model version.

29m agoHN ↗

Why does he first ask to label "ironic" or "not", and then answer the feature questions? Wouldn't it be better to reverse the order?

28m agoHN ↗

Why not use a text embedder for the unstructured data and concatenate with the structured data?

For example you could freeze most of the layers of the embedder but let the final ones learn. Then you wouldn’t need to do either feature or prompt engineering?

21m agoHN ↗

Calibration / Threshold Control

The amount of thinking is relatively calibrated. Ask an obvious classification, you get an instant answer. Ask a tricky one, much more thinking.

20m agoHN ↗

I really wish people would define terms when using math. What is y? What is LLM(x)? Presumably it evaluates to some real number so that it can be fed to the logistic sigmoid function. If it is the logistic function, then why does beta going to infinity matter? It seems to just collapse the output of the sigmoid function to 1 and make the value of LLM(x) meaningless instead of their claim that it recovers the LLM classifier. What is the function I()?

Maybe these are well understood terms in some field? Maybe I'm just lost?