Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Android 17 is the first since 3.x to add new APIs without releasing to the AOSP(grapheneos.social ↗)
    64comments
  2. Cloudflare Quick Tunnels(cloudflare.com ↗)
    190comments
  3. Apple releases iPhone Duo simulator and Xcode 27.1 beta(developer.apple.com ↗)
    18comments
  4. Korea raises data breach fines to 10% of revenue(koreajoongangdaily.com ↗)
    1comments
  5. Saving another 100TB of RAM with math (and Rust)(cloudflare.com ↗)
    7comments
  6. Cache-to-Cache: Direct Semantic Communication Between Large Language Models(arxiv.org ↗)
    3comments
  7. Photon-Emission-Guided Laser Fault Injection Enables RP2350 Secure Debug(ledger.com ↗)
    33comments
  8. Show HN: Cactus Needle 3: 8-29MB automation models can match DeepSeek V4 Flash(cactuscompute.com ↗)
    57comments
  9. The Implications of Linguistic Illegibility for LLM Security(arxiv.org ↗)
    8comments
  10. OpenJev(openjev.com ↗)
    231comments
  11. C++26: Trivial infinite loops are no longer undefined behaviour(sandordargo.com ↗)
    139comments
  12. Our brain evolved from two primitive nervous systems that merged: Study(newscientist.com ↗)
    32comments
  13. I vibed a proof of Conway's conjecture(overreacted.io ↗)
    158comments
  14. North Korean nuclear test sets off years of earthquakes(science.org ↗)
    127comments
  15. Border agents can search cellphones without a warrant or reasonable suspicion(lawandcrime.com ↗)
    58comments
  16. The first new cat species discovered in 100 years(nationalgeographic.com ↗)
    13comments
  17. US Military had close call after using AI for hallucinated intelligence report(cnn.com ↗)
    203comments
  18. Show HN: Ax-check.com – Can agents use your product?(ax-check.com ↗)
    21comments
  19. Inside ZCode: Silently uploading your Git history to the cloud(ferstar.org ↗)
    83comments
  20. A heap overflow and SSO misconfiguration to compromise OpenAI internal repos(hacktron.ai ↗)
    193comments
  21. Minimal Phone 2(minimalcompany.com ↗)
    95comments
  22. How SpaceX streamlined the Raptor engine(construction-physics.com ↗)
    11comments
  23. Cekura (YC F24) Is Hiring(ycombinator.com ↗)
    discuss
  24. "From Geometry to Algebra and Back Again: 4000 Years of Papers" by Jack Rusher [video](youtube.com ↗)
    discuss
  25. Warez: The Infrastructure and Aesthetics of Piracy (2021)(archive.org ↗)
    6comments
  26. Show HN: Scry, programmable internet search w/ congestion pricing(scry.io ↗)
    12comments
  27. A search-and-inference database from scratch in pure Zig(antfly.io ↗)
    6comments
  28. How to Write with an LLM(sockpuppet.org ↗)
    211comments
  29. Mathematicians Build Long-Awaited Graph Sandwich(quantamagazine.org ↗)
    14comments
  30. Jemalloc 5.4.0(github.com/jemalloc ↗)
    82comments

Telstra outage: The night a network decided the year was 2006

85 pointsby 19h agonetnod.se
32 comments
18h agoHN ↗

Shout out to the Jeff Geerling video about this

18h agoHN ↗

Author needs to decide which end of the stratum stack they want as the top.

This design from 2010 had at the top stratum 1 sources at Australia's National Measurement Institute (NMI)

Among otherwise comparable candidates, the lower stratum carries more weight.

18h agoHN ↗

Each decision solved the problem in front of it. Nobody was asked to look at the sum of all actions.

This is the perfect description of Telstra as a company.

15h agoHN ↗

This reminds me of this: https://obamawhitehouse.archives.gov/blog/2015/03/26/why-we-...

I saw the talk version of that and it has lodged in my memory. The point he made was that "health.gov" was made by a bunch of teams that made sure that their part worked, but nobody was in charge of making the whole thing work.

This is sadly "normal" for large bureaucracies.

Every time I see a giant, multi-million dollar catastrophe, it's always "proper", "documented", "enterprise", "by the book", and... "a total failure". The reason is always that nobody actually cares about the final outcome, only the paperwork in front of them that they need fill out, the checkbox that needs to be ticked, or the compliance requirement that needs to be met.

13h agoHN ↗

This, in my opinion, is some part of how we are schooling people - do as you are told and follow the rules and some part of not having stakes in the outcome.

10h agoHN ↗

No this is what the corporate world does to you.

There's a brief period when you join any company where you really want to improve things, and then about a year in the system has asserted towards the mean.

Nothing to do with schooling, every thing to do with business structures.

13h agoHN ↗

Why did you include the source tracker in your link?

13h agoHN ↗

Definitely not on purpose! I keep forgetting that ChatGPT does that now, which is rather irritating because I don’t use it to write comments but the tracker makes it look like I do.

I do use ChatGPT to resolve vague memories of articles read long ago into concrete URLs I can share, which is one of the best uses of the accursed things right now!

I assure you that I am an organic meat human like yourself and the article dates itself to the “before times” and can also be trusted to be free of contamination.

12h agoHN ↗

ChatGPT is a far more useful search engine than Google, duckduckgo, etc.

It won’t last, but just as Google was so much better than hotbit, altavista, yahoo etc back in the day.

6h agoHN ↗

This is most legacy telecoms everywhere.

This is the perfect description of most legacy telecoms everywhere.

16h agoHN ↗

They're pretty readable, I recommend that second one (though it leaves some questions I have unanswered). It's a good example of how latent issues with a large system can go unrecognised for a long time before a seeming unrelated change can trigger a cascading failure

14h agoHN ↗

One thing it doesn't make clear – the GPS card only supported the original L1 signal, which is what causes the GPS week rollover issue. The newer L2C signal has a much longer rollover period (157 years vs 19.6 years), which means no rollover until next century. If the GPS card had supported the newer L2C signal, then likely this would not have happened even if the other misconfigurations had still occurred.

17h agoHN ↗

All I could think of while reading this was “has no one heard of ntptrace?”

17h agoHN ↗

It's funny, but I had the same thing happen to me once. One day, overnight, my computer decided it was 2006.

It happened twenty years ago...

;-)

15h agoHN ↗

Is it really so that Time Division Multiplex timing in cellular networks is based on NTP time? I stopped reading after the article implied this - it doesn't sound plausible at all.

15h agoHN ↗

No, it absolutely is not.

That timing requires a very tight time, frequency and phase reference. It’s derived from local high stability GNSS-disciplined time references.

12h agoHN ↗

I would assume it’s similar to dab etc, and it’s ptp that carried that data on IP. In the broadcast tv world we use ptp, or more traditionally black and burst, or word clocks in audio.

The source is still the same - your grandmaster (ptp) or sync generator (b+b) is fed from a good timing source - often GPS, with one or two inputs, but sometimes local atomic clocks.

I’ve seen issues before with 40ms audio gaps when a GPS antenna has died and the coders can’t cope with frame drift so they incentive a 40ms silence in the audio every 40 minutes or so as the source and destination clocks are different speeds by about 1 part in 60,000.

In DAB the tolerances are far lower, there was a problem about 10 years ago due to a GPs satellite being a few microseconds off

https://www.radioworld.com/trends-1/gps-anomalies-caused-pro...

12h agoHN ↗

No. The official report seems to suggest that software in IMS core (the subsystem that controls VOIP signalling in the mobile network) bugged out after the time was shifted back 20 years. That doesn't narrow down the specific failure that stopped calls from going through very much. IMS core is full of timers, but it could also have been a license server issue for all the report said about it.

10h agoHN ↗

I haven’t read the report, but I did read at the time that it was a certificate validation failure: clients thought the current time was before the "Not Valid Before" date. I forget now what essential service became unavailable because it was disbelieved.

(Source would either have been the Whirlpool forums – where Australian ISP and telco engineers can be found – or local Australian mass media.)

13h agoHN ↗

Human made summary: A telecom company had time synchronisation issues in a location of their infrastructure. They eventually used a single hardware GPS clock that fixed the issues but became their only authoritative clock in this location.

One day, the clock was turn off and on and it went 1024 weeks backwards because the time GPS time protocol sucks and use a week counter with too few bits.

Apparently a GPS clock can keep track of the time if it’s up and running when the week counter overflows, otherwise it has to take a wild guess. Apparently their hardware GPS clock used a hardcoded start time from its firmware instead of trusting a less reliable existing clock.

The article finishes with some AI looking suggestions to prevent such an issue to happen again.

Mine would be to have bought one or two more GPS clocks and not from the same provider.

10h agoHN ↗

This is the opposite of a Google AI as summary in the best way.

Thank you for slogging through the slop to bring the human angle

10h agoHN ↗

One day, the clock was turn off and on and it went 1024 weeks backwards because the time GPS time protocol sucks and use a week counter with too few bits.

As I pointed out in my other comment below, this is only true of the old L1 signal, not the newer L2C signal. If they had a newer GPS card, this would never have happened.

Mine would be to have bought one or two more GPS clocks and not from the same provider.

I think it would be more important to have a newer one that doesn't have this problem, than two old ones which both do.

9h agoHN ↗

I don't think every GPS receiver necessarily has to have this problem, even with the old signal. It's not crazy to keep a week count in persistent memory and use that as a lower bound on startup so you only get a wrong time if the receiver isn't online for 1024 weeks.

7h agoHN ↗

Exactly. The sane ones just ask you what year it is at startup, and use that to determine the initial epoch. Then they handle rollover automatically since then.

It's been widely opined that 1024 weeks (~20 years) is the worst possible interval. Either rollovers should've happened VERY frequently (say, 128 weeks) so receivers would be FORCED to deal with it, or extremely infrequently (16384 weeks?), so it's simply never an issue.

The unhappy medium is long enough that developers feel justified in saying "naaaah, our receiver won't still be in use then, we can ignore that!", but in practice it's very likely to happen.

4h agoHN ↗

Years ago while working as a contract sysadmin for a clearing firm, a coworker made an overbroad puppet change that changed every server's time zone from UTC to CDT. The only thing I heard about it was from a text from an old colleague/friend who was troubleshooting why their clearing reports were all off by five hours. I still chuckle at that random text.