Hacker News

New stories

Live mirror
30 storiesupdated just nowView source snapshot
  1. Microdnf: Minimal Python-Free Dnf(github.com/rpm-software-management)
    discuss
  2. Getting the most out of Opus 5.5 in Claude and Claude Code(claude.dev)
    discuss
  3. Show HN: DynamicNotch – An interactive, customizable notch utility for macOS(github.com/hitjack007)
    discuss
  4. Diesel prices could crush Republicans in the heartland(natesilver.net)
    discuss
  5. Meta's Muse Drags Down Stocks That Depend on 'Consumer Inertia'(bloomberg.com)
    discuss
  6. Ask HN: When is fine-tuning a small LLM worth it?
    discuss
  7. Are we going to use the same Desktop UX forever? [video](youtube.com)
    1comments
  8. Better prompt caching for GPT‑6(openai.com)
    discuss
  9. S&P Global Enters Agreement to Acquire OpenZeppelin(openzeppelin.com)
    discuss
  10. Z.ai says sorry for slurping up your code, open sources ZCode(theregister.com)
    discuss
  11. Bootstrapping Frontier AI Governance by Mutualizing Risk(lawfaremedia.org)
    discuss
  12. 'Same sense of urgency': UN chief compares AI risk to nuclear weapons(smh.com.au)
    1comments
  13. Agent's Memory Needs a Retention Policy(memanto.ai)
    discuss
  14. Hayek's Federalism and the Making of European Integration [pdf](cosmosandtaxis.org)
    1comments
  15. Are Meta's smart glasses training AI for robots?(proton.me)
    discuss
  16. Show HN: Shrewd – what I learned distilling LLM labels into local classifiers(github.com/sshah03)
    discuss
  17. Show HN: Riftri – copy-on-write Git worktrees for running agents in parallel(twitter.com/kinfisht)
    discuss
  18. The JavaScript Midlife Crisis(maroun-baydoun.com)
    2comments
  19. Ask HN: Ideas for getting more GitHub stars?
    1comments
  20. Open source tool for building hierarchical agent loops(github.com/plasma-ai)
    discuss
  21. How the Barcode Was Invented(worldofsystems.co)
    1comments
  22. I built a tool that reverse engineers competitors SEO(rankshawk.com)
    1comments
  23. Egress Testing – Server that records inbound TCP and UDP packets(portleak.link)
    1comments
  24. A16Z is challenging Silicon Valley's love for drop-outs by launching a school(techcrunch.com)
    2comments
  25. Reading specs beats reading code(technative.eu)
    1comments
  26. Native apps written in TypeScript and CSS(github.com/geastack)
    1comments
  27. Agent manager now with mouse support(agent-manager.dev)
    2comments
  28. Obscura: The first VPN that can't log your activity(obscura.com)
    14comments
  29. AI Leaders Are Standing on a Liability Landmine(wsj.com)
    2comments
  30. WebMCP Integration at Stripe(twitter.com/stevekaliski)
    discuss

Unreal Agent

40 pointsby 2h agounreallabs.ai
27 comments
1h agoHN ↗

I don't really understand how async tool calls translate to token savings

It says that it removes tokens wasted while a model is waiting on synchronous tool calls. What tokens exactly is Pi using when waiting?

59m agoHN ↗

I think a running agent periodically checks if a process it started has finished.

1h agoHN ↗

I think this is what I’ve been waiting for! Live interactions with models are fundamentally asynchronous! This will be great for interactivity.

What a good time for harness design. Just while OpenCode is growing up a bit and focusing on their harness. I’m delighted to see people focusing on good general solid harness design principles.

Someone make an OpenCode API (the best general agent end UI API I know) compatible server for it!

Sucks that I haven’t upgraded my client to V2 yet and I don’t really want to build on an outdated API ...

Edit: It would be cool to get streaming responses :) But I understand (and actually applaud) that the authors seem to have been very focused on the core mechanics.

Also note that this works with the responses API!

1h agoHN ↗

The headline graph is kind of bizarre.

For some reason they're comparing their harness running on Astra xhigh to Codex with Astra max?

---

Also worth noting that OpenAI just added support for async tool calling to their harness, which isn't 1:1 with this approach, but is slowly ramping up in being able to provide something similar.

A big part of why Codex uses so many tokens is that it basically hot loops on polling tasks it starts for... absolutely no good reason: https://www.reddit.com/r/codex/comments/1wdlp7q/weve_discove...

I fixed it on my fork of Codex too, also back in Jan/Feb – I keep this patch rebased, for anyone who wants it: https://github.com/tekacs/codex/commit/9ffcf8db9078eae43d411...

It results in token savings similar in scale to those displayed here by Unreal.

---

My harness has used a slightly fancier version of the approach that Unreal is using since ~Feb, and... it definitely works excellently, but it's also assuredly smoother with Astra and other recent models that are more aware of async tool calling.

9m agoHN ↗

To reduce the codex polling I slapped this into my ~/.codex/config.toml:

  [features.multi_agent_v2]
  enabled = true
  wait_agent_enabled = true
  min_wait_timeout_ms = 10000 # 10s
  default_wait_timeout_ms = 300000 # 5m
  max_wait_timeout_ms = 3600000 # 1h

Seems to do the job and reduce usage; I just ran Astra for ~5 hours (using a goal) and it used the last 30% of my usage. And now they released GPT-6 Sol and Luna (which is basically 5.6 Sol and Luna, but a bit better and also 50% cheaper) ;_;

58m agoHN ↗

Honestly the fact that we're still modeling agent harnesses like chat and strapping them into shell sessions instead of building async actor systems with sandboxed OS functionality access actors is bananas to me. So much easy wins can be had just by building on right abstractions... Hopefully will get enough time to play with this idea soon on my own.

10m agoHN ↗

Do you have any examples of this approach?

57m agoHN ↗

Sounds like a trademark issue when Epic ships a wildly popular Unreal Engine

55m agoHN ↗

But I don't think you can trademark a generic word like Unreal...

40m agoHN ↗

"But I don't think you can trademark a generic word like Unreal..."

Sure you can, trademarks are contextual. If they were a landscaping business it wouldn't matter. But within the same industry absolutely.

36m agoHN ↗

They've trademarked it for quite a few closely-related domains: https://tsdr.uspto.gov/#caseNumber=87709072&caseSearchType=U...

Computer software, namely, game engine software for video game development and operation; Computer software, namely, software development tools for the creation of computer-generated imagery and graphics for the production of video games; Computer software, namely, software development tools for the creation of computer-generated imagery and graphics for the production of content for virtual worlds and 3D platforms; Computer software, namely, software development tools for the creation of computer-generated imagery and graphics for the production of motion pictures, television shows, videos, 3D animations, 3D simulations, 3D visualizations, virtual reality motion pictures, virtual reality television shows; Computer software, namely, software development tools for the creation of computer-generated imagery and graphics for the production of virtual reality video games; Virtual reality game software; Virtual reality software for creating multimedia content; Augmented reality game software; Augmented reality software for use in mobile devices for integrating electronic data with real world environments for the purposes of entertainment

10m agoHN ↗

Those do all seem directly tied to what Unreal Engine does i.e. graphics

6m agoHN ↗

Unreal is trademarked. This is not in question. And personally I 100% thought this was an agent specifically for dealing with the Unreal Engine, and was interpreting all of that information in that context. Very weird name for a company/product.

46m agoHN ↗

Mildly disappointed that this has nothing to do with Unreal Engine, the popular game engine.

41m agoHN ↗

Yeah, I was hoping it would be something to make it easier for agents to interface with Unreal Engine games.

38m agoHN ↗

I was imagining a talking head avatar rendered in Unreal Engine, with lip sync and facial expressions driven by a multimodal LLM that produces the speech.

45m agoHN ↗

omp.sh does this way better by just allowing structural toolcall execution in eval with python/js.

16m agoHN ↗

I think the biggest selling point of a codex sub vs a claude sub is that you can use codex subs in any harness you want.

18m agoHN ↗

This feels like it's begging for a lawsuit from Epic.

17m agoHN ↗

I think this space is very untapped. Models are interesting, but I am absolutely obsessed with some things I've been researching/working on for the past few years:

Fractal tool discovery: tool taxonomy where an agent can "drill deeper" to find what specific tool it's looking for. Helps if/when polluting context with a zillion (mostly unnecessary) tools.

Leveraging splay trees: this is my favorite data structure and I think relatively unused in the context of agents/harnesses. A lot of times, recently-used workflows/tool-chains will be used again, so having those at the top of the search hierarchy is an awesome optimization.

Virtual containerized notebooks: models working in sandboxed (WASI) Python notebooks is incredible. Even local models (if given enough time) will usually converge on a good solution. Being able to mount tools/resources/fs is again, imo quite untapped. Some problems here are running native things (thing numpy/pandas) in containers is a nightmare (or impossible).

Anyway, happy to see other folks seriously doing stuff in this space. If anyone wants to collaborate on anything don't hesitate to reach out :) I'm also actively looking for a job or some contract gigs.

Fun times ahead.

12m agoHN ↗

This is not an Unreal Engine agents? Very disappointed