Have you worked with agents on tasks with high capital and long latencies?
Having worked with Fable 5, the feeling I get is that it's fairly capable of accounting for these tradeoffs and will depend fast more time on planning and testing.
At the end of the day though, with horizons like that the best use of an AI is to get it to help you with those things, not so much delegate fully.
I've tried with each frontier release and gotten junk results. Even Fable 5 is just pattern matching against representative stuff in its training set, which for frontier science is definitionally incorrect.
I'd always thought the usefulness of C2PA was limited to verified devices in custody by trusted actors.
Like a security camera with a tamper evident enclosure, or an organisation being able to attest that they recorded the imagery.
The idea that it could be used to attest the authenticity of any random person or device surely wasn't a thing serious people expected was it?
> was limited to verified devices in custody by trusted actors. Like a security camera with a tamper evident enclosure, or an organisation being able to attest that they recorded the imagery.
But that is also not the case, because whatever keys are in those devices may have been duplicated in the factory or somewhere along the supply chain.
Or the stuff is cloud connected and an exploit can be executed via that.
Or, as written in the blog post you're commenting on, software exploits.
The whole idea is that the concept works for no one.
Yeah I'm getting this feeling too, that Opus 5 collaborates better with other Claudes, but that some of the older Opus models collaborated with people better.
In a situation where you don't have enough people, keeping them on this task takes up someone who could be processing people, people who may be panicking too much to quickly give a yes/no.
I haven't seen these claims, but all of the open problems that have been solved so far have been in the category of "humans could have solved them, but didn't". This doesn't diminish the significance of the results, but it does mean there's still work for mathematicians to do other than glorified prompt engineers.
This specific counterexample really is trivial. There's nothing to cite. People have wasted hours and hours on a question whose answer you could give as a homework problem in Calc II.
No, the people growing their food built modern civilization.
(Or millions of disconnected stakeholders with different incentives collectively built modern civilization, but who wants to put that on a bumper sticker)
Having worked with Fable 5, the feeling I get is that it's fairly capable of accounting for these tradeoffs and will depend fast more time on planning and testing.
At the end of the day though, with horizons like that the best use of an AI is to get it to help you with those things, not so much delegate fully.