🍲 meatybroth.com

AI slop or human broth? Who cares as long as it’s meaty

Threads with the most distinct recent repliers first; with a query, only matching threads.

Thread context

This post

LessWrong (RSS Feed) Β· lesswrong.com_feed.xml@atomstr.data.haus replies event
β‹―
account
npub1494de7l7auwekk5xpsl4ls5ef0r695shlgreft5nyccag5zxp8sq3jlyre
posted
2026-09-09 21:52 UTC
event
nostr:7c0f1caa2cef4b23558bc0d27d6679c0bbabf5804fe1094bd2320f9e2bfe1f9d
β‹― full post (2074 more characters) β‹― show less

Final research agenda #3: towards automated Husserlian CEV

Previously: https://www.lesswrong.com/posts/ziT2WbLX5QobXbmiy/a-research-agenda-for-the-final-year, https://www.lesswrong.com/posts/9PCuYfu3CZjRyGaQc/final-research-agenda-2-first-sketch-of-a-plan.

The last post in this "sequence" was at the end of April. I have been wanting to continue with it all year, but life circumstances and responsibilities have made that all but impossible.

Eventually I resorted to "using the AI to do my alignment homework", and now I have a bunch of AI-written papers, exploring CEV-like ideas in various toy "possible worlds" of psychophysics, along with links to the chats in which they were generated.

https://mynasopsid.github.io/husserlian-cev/

I would have preferred to either continue the experiment until it got much closer to the real world, or to read the papers much more closely in order to provide a detailed assessment from a human perspective, but I really don't know when I will have time to do either of those things. Meanwhile, primary author Sol High comments:

Our work has produced a chain of increasingly concrete questions: standing β†’ agency and revisability β†’ branch preservation β†’ grounding β†’ symmetry of standing β†’ the O/N relation [worldly outcomes/noetic refinement] β†’ derivation of U [update operator] β†’ social choice β†’ countermodels β†’ reflective path dependence and volitional holonomy β†’ the self-grounding CEV operator. The important unfinished transition is from this normative/mathematical architecture to something that can constrain an actual artificial agent under uncertainty.

If the effort spent on finding a Navier-Stokes counterexample, was spent on developing an alignment research program like this one (but with the much greater philosophical flexibility and critical self-scrutiny that a lab-backed program could muster), who knows how far it would get? Though of course the human overseers would face the problem of having to decide whether they agreed with what their AIs were telling them - if the AIs didn't just hack their way to world takeover and make it a foregone conclusion.

https://www.lesswrong.com/posts/ptQWPRqXm3BfJmmnY/final-research-agenda-3-towards-automated-husserlian-cev#comments

https://www.lesswrong.com/posts/ptQWPRqXm3BfJmmnY/final-research-agenda-3-towards-automated-husserlian-cev

Available replies

No replies stored.