- 77comments
- 11comments
- 107comments
- 5comments
- 2comments
- 8comments
- 48comments
- 10comments
- 270comments
- 29comments
- 28comments
- 66comments
- —discuss
- 12comments
- 413comments
- 43comments
- 239comments
- 883comments
- 295comments
- 12comments
- 60comments
- 104comments
- 68comments
- 59comments
- 70comments
- 11comments
- 102comments
- 64comments
- 49comments
- 23comments
From https://github.com/akahkhanna/groundtruth :
Which other agent IDEs does or could this work with?
Is this also where to attach provenance metadata and sign the agent trace, and debug?
Re: MCPSnoop, proxies like Aegis and LiteLLM, and awesome-auditable-ai: https://news.ycombinator.com/item?id=48777144#48779413
There are web standards for signing metadata with interoperable schema as linked data:
W3C JSON-LD/YAML-LD + W3C PROV + W3C DID + W3C VC Verifiable Claims
"Follow up to verify that the work was actually satisfactorily completed"
Are there other sound management practices that aren't yet effectively implemented in current gen agents?
I've had some success with using a more expensive model to plan a decent work breakdown structure in a document and then a low-cost model to implement the plan and keep following-up, but wonder about subsequent need for code review and code quality.
Can a lower-cost model verify completion? Iff tests and test coverage and e2e tests?
The same oracle / model routing and partitioning problem