Hacker News

Top stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Astra for Law(openai.com ↗)
    206comments
  2. Bend – A language that blocks AI mistakes via proof, on CPU and GPU(bend-lang.com ↗)
    100comments
  3. Bonsai 2 27B: Near-Lossless Compression in a 9x Smaller Footprint(prismml.com ↗)
    25comments
  4. Hister: A private search engine for the pages you visit and the files you keep(github.com/asciimoo ↗)
    120comments
  5. Wax motor(wikipedia.org ↗)
    37comments
  6. Fujitsu launches made-in-Japan next-generation CPU FUJITSU-MONAKA(global.fujitsu ↗)
    174comments
  7. Sex, AI, and the Apocalypse(iankduncan.com ↗)
    22comments
  8. Flet 1.0 – Build cross-platform apps in Python(flet.dev ↗)
    3comments
  9. CrowdSec Source Code Leak(crowdsec.net ↗)
    34comments
  10. How to Write with an LLM(sockpuppet.org ↗)
    4comments
  11. Rate limits on GitLab.com are changing(about.gitlab.com ↗)
    103comments
  12. Infinite-Parameter LLMs: Generating and Adapting Weights from Live Data(arxiv.org ↗)
    26comments
  13. Everybody's Lost Their Minds(netmeister.org ↗)
    169comments
  14. Diplodocus, Long Thought Exclusively American, Turns Up in Spain(sci.news ↗)
    3comments
  15. The American Religion of Self-Storage Facilities(newyorker.com ↗)
    292comments
  16. How Uber Protects Against Retry Storms(uber.com ↗)
    6comments
  17. How GLM built its own inference infrastructure(z.ai ↗)
    254comments
  18. Zettascale (YC S24) Is Hiring ASIC/FPGA Engineers to Build Chips for ASI(zscc.ai ↗)
    discuss
  19. TSMC revealing details about next gen A14 node(mapyourshow.com ↗)
    27comments
  20. Why I didn’t sign the Fields medallists’ letter(gowers.wordpress.com ↗)
    240comments
  21. Show HN: Snapdrop: Instantly share files between devices. No setup, no signup(snapdrop.me ↗)
    6comments
  22. How do we prevent mathemathics from devolving into the Medieval Era of secrecy?(mathoverflow.net ↗)
    29comments
  23. Running Ubuntu on the Lenovo IdeaPad Duet(vhaudiquet.fr ↗)
    18comments
  24. One year of sponsored Servo development(servo.org ↗)
    136comments
  25. Canto: A speech model built for the real world(wisprflow.ai ↗)
    11comments
  26. André Weil and the Hodge Conjecture(jiahao116.github.io ↗)
    8comments
  27. CCC invites all model citizens to 40C3(ccc.de ↗)
    175comments
  28. Launch HN: Skillsync (YC W26) – AI chat sessions made portable across agents
    43comments
  29. Towards Self-Driving Codebases(detail.dev ↗)
    74comments
  30. Show HN: Share your AI Setup, Learn from others(mysetup.ai ↗)
    82comments

Canto: A speech model built for the real world

27 pointsby 4h agowisprflow.ai
11 comments
2h agoHN ↗

They need to show some examples. You beat all the top models on your private data set? Show at least a couple examples - audio and transcripts - from examples that your model got right and others got wrong.

1h agoHN ↗

after reading the article I still have no idea how their thing performs, or if I should care how it performs since a majority of voice benchmarks still don't map to human evals

2h agoHN ↗

Congrats Wisprflow, love to see this. Especially learning the users nuances and corrections, i think that may be novel in the dictation app space. Things like Handy have the ability to specify common typos but some method of automatic learning is new.

This is a pretty hot space right now. For me, it's all incredible for two things:

- I get my unabridged thoughts down on the page substantially quicker and cleaner using dictation. I believe dictation is the perfect first draft tool, and an amazing way to interact with AI too as you can dump tonnes of personal opinion in to every prompt without the overhead of keyboard interface.

- Accessibility! I know a couple of folks whos ability to use a computer has been tremendously elevated by the recent improvements in dictation. Due to mobility issues, they feel largely locked out of interacting online and tools like Handy and Wisprflow have been a game changer for them.

Very excited about all this, keep it coming!

24m agoHN ↗

Is Handy still the best option for local?

Wispr lagged my PC to hell; plus, the cost. The quality was excellent, though.

5m agoHN ↗

There are lots of good option (and they all use the same set of models, so really take your pick), but Handy works very well.

2h agoHN ↗

Wonder if they are aware of Cantonese, the language that is often abbreviated as Canto.

1h agoHN ↗

Probably this is after the italian word for singing

1h agoHN ↗

Or Spanish, specifically it's the first person conjugation, meaning "I sing" or "I'm singing"

1h agoHN ↗

Is it actually better than Microsoft's MAI-Transcribe-2? That generally seems like the best model right now and it's not included in their benchmarks.

I switched from Superwhisper->WisprFlow->Spokenly->Fieldwork and found WisprFlow the least accurate of the 4.