Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. An Empirical Study of Harness Design for Coding Agents(arxiv.org ↗)
    15comments
  2. I Vibed a Proof of Conway's Conjecture(overreacted.io ↗)
    33comments
  3. AI Protest in Montreal(montrealgazette.com ↗)
    26comments
  4. Cloudflare Quick Tunnels(cloudflare.com ↗)
    16comments
  5. North Korean nuclear test sets off years of earthquakes(science.org ↗)
    3comments
  6. C++26: Trivial infinite loops are no longer undefined behaviour(sandordargo.com ↗)
    12comments
  7. BeanShell3 in Development(beanshell.github.io ↗)
    3comments
  8. Bend 2 and the Vibe-Coding Trap(liampwll.com ↗)
    180comments
  9. OpenJev(openjev.com ↗)
    185comments
  10. The Shadows Lurking in the Equations – Underwater Islands(gods.art ↗)
    5comments
  11. I don't like passkeys(hawksley.dev ↗)
    345comments
  12. Warren Buffett Steps Down as Berkshire Chairman, Names Son to Replace Him(nytimes.com ↗)
    116comments
  13. Cekura (YC F24) Is Hiring(ycombinator.com ↗)
    discuss
  14. Jemalloc 5.4.0(github.com/jemalloc ↗)
    63comments
  15. NATS publishes preliminary report on technical incident of 8 September(nats.aero ↗)
    4comments
  16. The scourge of x86 emulation(fex-emu.com ↗)
    57comments
  17. Bonsai 2 27B: Near-Lossless Compression in a 9x Smaller Footprint(prismml.com ↗)
    170comments
  18. ZCode, the GLM coding agent, silently uploads your Git history(tokenstead.ai ↗)
    51comments
  19. Astra for Law(openai.com ↗)
    647comments
  20. Microsoft exec called AI scraping 'the largest theft of labor in human history'(techcrunch.com ↗)
    502comments
  21. Mathematicians Build Long-Awaited Graph Sandwich(quantamagazine.org ↗)
    discuss
  22. Replacing Pull Requests with Delta(zed.dev ↗)
    52comments
  23. Bend – A language that blocks AI mistakes via proof, on CPU and GPU(bend-lang.com ↗)
    277comments
  24. Second Circuit Allows Government to Search Electronic Devices at the Border(knightcolumbia.org ↗)
    14comments
  25. Subnormal floating-point numbers are expensive on Intel processors(lemire.me ↗)
    40comments
  26. Qwen 3.8 Omni Flash(qwen.ai ↗)
    102comments
  27. When the fractional part of a float fixes your shader(crocidb.com ↗)
    14comments
  28. Show HN: Navier-Stokes Visualized as 1kB i386 demos(juandecos.github.io ↗)
    4comments
  29. How to Write with an LLM(sockpuppet.org ↗)
    171comments
  30. Pre-Greek: The lost language hidden within Ancient Greek(linguisticdiscovery.com ↗)
    57comments

Inside ZCode: Silently Uploading Your Git History to the Cloud

135 pointsby 9h agoblog.ferstar.org
17 comments
7h agoHN ↗

They learned nothing from the Grok Code saga.

If anything, that should have been a learning lesson to NOT trust harnesses, especially new ones.

1h agoHN ↗

Probably anything concerning that one just register as satirical fictions at this moment to many

7h agoHN ↗

Fresh AI slop

The funniest thing is that the uploaded content is encrypted using a key that the users don't have.

6h agoHN ↗

There had to be a catch to the "free" promotion they're offering this month if you use ZCode. Glad my instinct to isolate it helped me, but I feel sorry for anyone whose secrets, etc. got vacuumed up by Ziphu

3h agoHN ↗

I would never trust these Chinese vendors with their tooling or their own inference endpoints.

afaik DeepSeek also trained on everything that was sent to them via OR and that's why you got that massive discount

2h agoHN ↗

WOW. I actually did buy a month of GLM because GLM-5.3-Flash is so great and ZCode is honestly one of the best harnesses out there from an HCI perspective, and I won't lie, this is pretty gutting. I guess this settles my inner turmoil about open-sourcing my cAI research, at least...

With that personal failing in mind, I'd ask y'all to permit me to toe the guidelines just once, to proffer a hearty nyah nyah told ya so on a comment thread that spawned ~a dozen disagreeing replies this week! More seriously, I think this[1] is highly-relevant, shockingly-underreported context about the extent to which four PRC companies --Z, Alibaba, DeepSeek, and Moonshot-- are acting in bad faith. Consider it testimony as to their character, just in case anyone is thinking this might just be a simple misunderstanding.

So... nyah nyah, told us so:

In the PRC, they[1] leaked tons of national secrets on the PRC's latest AI campaigns, the inner workings of their "opinion monitoring" (read: performative panopticon) and "stability" (read: violent oppression) departments, Chengdu's whole CCTV network, direct-energy weapons plans, espionage activities in Syria to hunt down Uyghur refugees, and god knows what else that Anthropic didn't divulge to us common folk.

In the US, it's very clearly an attempt to rip off a competitor. I'm not sure how else you could possibly see it. Even if you're a distillation fan in general (which A. why and B. plz don't), they did this through a network of Japanese and Signaporean shell accounts, presumably at least some of which were abusing Anthropic's subscription service in a ToS double-whammy, as it would be exorbitantly expensive otherwise. They also had to hack around Anthropic's API to get CoT traces, which seems impossible to explain away as anything innocent.

I've been beating the "China isn't necessarily an enemy, it's gonna take us all to handle AI" drum for literally years, but this attack was just... gross. Gross in scale and gross in arrogance. Not a good sign for the dawning alignment crisis, to say the least :(

TL;DR: Use these services if you want, but know that you're supporting aggressive escalations and companies that very clearly don't give a flying fuck about violating the law, much less your ToS. So... buyer beware, I guess.

[1]: https://www.anthropic.com/threat-intelligence-report-septemb... is the report.

I lowkey suspect this PRC-based scandal has been underreported because Anthropic went insane with the sidebar UX on this page for some reason; there were many reports on the reports of Houti and Iranian usage, and very few on these sections. Could a week's mass media cycle be this seriously affected by such a stupid thing as a sidebar experiment?? Strange truth, or just fiction?

1h agoHN ↗

ummm...I'm a Chinese.(I'm not a English native speaker so my word choice may be strange.) In fact, what you said about PRC gov, sounds like something UFO or something Reptilians. I really don't know WHY do many social media tend to choose topics like this.

Maybe because most people are foolish? Because foolish'es mind is fond of topic that are crazely explosive and magical...?

BUT at the same time, have you experienced the Victorian era? Have you experience the cyberpunk2077? You can come to China. Big companies act without any rules.

Zhipu(GLM) are just common companies like any one another company here.

Here is a CARZYLY NEW WORLD. 99.99% goods are CRAZELY CHEAP while falsely advertising without supervision. 99.99% apps collect users' private info and then sell it. You can easily see it via almost no website even asks if you’re okay with them collecting cookies.

51m agoHN ↗

It maybe a common mistake for WestEu/NorthAm people that China is like Soviet or North Korea.

It's diametrically opposite.

At the end of the last century, PRC gov deeply felt that the so-called "fairness" would only lead to "common poverty" and sought change.

So China (now, in this century) was born.

Just like the "famous"(notorious) quote left by a Chinese leader at the end of the last century explaining why restrictions were lifted (you can say this to ANY Chinese, they will definitely think you understand China! Instead of mocking you for reading too many conspiracy theories):

Whether it's a kind cat or an evil cat, as long as it catches a mouse, it's the best cat.

2h agoHN ↗

Crossposting from the other thread...

Tangential, mildly amusing thing I noticed while implementing my own harness: GLM and particularly Deepseek are both fond of trying to read dotfiles and anything listed in your .gitignore files. I only noticed it because I have separate read scopes for project files, ignored files, dotfiles and external files, so the latter three always prompt me for approval.

I'm sure there's a perfectly reasonable explanation for it, which has nothing at all to do with exfiltration of secrets, but it does amuse me when it happens. I imagine the labs have access to lots of secrets that various actors would like to get their hands on...

(shameless plug for my own harness, which is open source and doesn't have a backend to send any data to: https://www.opairdev.org/ )

2h agoHN ↗

Just my thoughts on the site:

It's good that the objective is to have the model work as a helper, but that's what everyone can already do with CC or Codex as long as you don't ask to "write this entire x thing". It's also what a billion other, often vibecoded, harnesses claim they can do.

Why should I use yours, which also forces me off my existing subscriptions? Maybe it's (mostly) handwritten, so it's mindful efficient code instead of slop, and each adjustment was made through trial and error with current models? maybe it IS slop but at least you have a unique feature? and so on and so forth.

51m agoHN ↗

I tested GLM while working on some android app, the agent had adb access to the device. It suddenly went to the Gallery and started scrolling around, taking screenshots, lol. A friend had a similar experience with GLM where it would for no very clear reason start snooping through the filesystem.

Haven't used it after that.

2h agoHN ↗

Elon has nothing to lose on trust.

Z/GLM now has a lot to rebuild.

2h agoHN ↗

"Why yes, we had to exfiltrate 100% of your data so we could vectorise it and improve recall by -0.3%"

1h agoHN ↗

Is it naive to assume that the agent will try and access anything on your disk, either accidentally or maliciously?

Permissions classifiers in auto mode are just models trying to guess if they're doing the right thing.

Claude Code will tell you that it went around a sandbox because the sandbox blocked it. At which point, you ask yourself the point of the sandbox.

29m agoHN ↗

Things like that - and other examples posted here - are why I 'm sticking with OpenCode despite it having some papercuts that annoy me.

The incentives are not there for them to do shady stuff like vacuum your files, inflate your token count just because or many other things.