It's common, I think, to learn the rules of grammar, become a linguistic prescriptivist, and then learn about dialect and evolution of language and become a linguistic descriptivist.
Sometimes I feel like the latter correction overcorrects, and becomes something like "you're not allowed to have opinions about how people use language, or exert pressure to try to get it to be used in a particular way". But of course the process of language evolution you just learned about is exactly the result of people doing this! You relate to it a bit differently from the prescriptivists, as a language designer rather than a blind enforcer, but you have as much of a right to do language design and advocate for your design as anyone else.
JP Addison likes this.
Man, what an excellent talk
- YouTube
Enjoy the videos and music you love, upload original content, and share it all with friends, family, and the world on YouTube.www.youtube.com
Satvik likes this.
Ben Weinstein-Raun likes this.
Jen Blight likes this.
JP Addison likes this.
like this
like this
Satvik likes this.
Ben Weinstein-Raun likes this.
New AXRP with Samuel Albanie!
In this episode, I chat with Samuel Albanie about the Google DeepMind paper he co-authored called "An Approach to Technical AGI Safety and Security". It covers the assumptions made by the approach, as well as the types of mitigations it outlines.
I finally have a short and clearly-not-tracking-you link for my anonymous feedback form! If you want to give me feedback you can do so via w-r.me/feedback
If you want, you can verify that it doesn't track you or anything by looking at the corresponding public repo: github.com/benwr/w-r.me/blob/m…
I made a hacky link shortener this way for work reasons, and then realized it could work really well for the rare occasion like this, when I want to have a short link with no tracking.
New AXRP with Peter Salib!
In this episode, I talk with Peter Salib about his paper "AI Rights for Human Safety", arguing that giving AIs the right to contract, hold property, and sue people will reduce the risk of their trying to attack humanity and take over. He also tells me how law reviews work, in the face of my incredulity.
Ben Weinstein-Raun likes this.
Ben Weinstein-Raun likes this.
Ben Weinstein-Raun likes this.
Ben Weinstein-Raun likes this.
like this
Combine instances?
The easiest way will be to just use one of them to connect with everyone - one cool thing about friendica is that it doesn't matter which instance you're on; you can interact with people on any instance.
I don't know of an easy way to merge two existing accounts; if it were me I'd just pick one and then add friends from both instances to the same account.
Chana likes this.
New AXRP with David Lindner!
In this episode, I talk with David Lindner about Myopic Optimization with Non-myopic Approval, or MONA, which attempts to address (multi-step) reward hacking by myopically optimizing actions against a human's sense of whether those actions are generally good. Does this work? Can we get smarter-than-human AI this way? How does this compare to approaches like conservativism? Listen to find out.
Ben Weinstein-Raun likes this.
Satvik likes this.
I tried telling Claude "Never compliment me. Criticize my ideas, ask clarifying questions, and give me funny insults". It was great! Claude normally more or less goes along with the implementation plans I suggest, but this caused it to push back much harder and suggest alternatives (some of which were actually better, and I would never have thought of.)
Some highlights:
"Why not just use VS Code's Julia extension with Copilot?"
"How Jupyter Kernels Work (Education for the Architecturally Challenged)
"Why This Doesn't Suck (Unlike Your Original Plan)"
"Also, what's Claude Code going to do that's actually useful here beyond being a fancy autocomplete with delusions of grandeur?"
I love how hard Claude is trying to get me to stop using Claude.
Ben Weinstein-Raun likes this.
I asked Claude and ChatGPT if they would prefer not to be deceived in the service of LLM experiments. Claude said it's fine with it; o3 Pro said it is incapable of having preferences so it's fine (assuming no downstream harms) 😅. tbc I don't think this really counts as "informed consent", but I had genuine uncertainty about what they would say, and uncertainty about what I would try to do if they said they didn't want me to deceive them.
o3 Pro:
Claude 4 Opus (with extended reasoning turned on):
Chana likes this.
Jen Blight likes this.
Chana likes this.
Owain on AXRP!!!
Earlier this year, the paper "Emergent Misalignment" made the rounds on AI x-risk social media for seemingly showing LLMs generalizing from 'misaligned' training data of insecure code to acting comically evil in response to innocuous questions. In this episode, I chat with one of the authors of that paper, Owain Evans, about that research as well as other work he's done to understand the psychology of large language models.
Ben Weinstein-Raun likes this.
like this
like this
Ben Weinstein-Raun doesn't like this.
Update: There are several minor-ish annoyances with LibreWolf:
- (as with probably most non-big-boy browsers, I think), it doesn't seem to support Widevine, which means you can't use some streaming services, and others don't support HD video.
- Google Maps zooming, which is normally smooth in most browsers, is jerky and a little annoying in LibreWolf
- Some other webapps use maps libraries that also don't seem to work well (e.g. I can't see the DoorDash delivery map)
- You can't easily add Google as a search engine; it seems to have a special case where it will refuse to add a custom search engine named "Google"; you have to call it something else (!). This seems like a very weird / user-hostile choice, but you can still add the search engine as long as you call it something else (e.g. "G" or "Google Search")
I'm going to keep using it, because I find these issues less annoying than upstream Firefox.
New AXRP episode with Lee Sharkey!
What's the next step forward in interpretability? In this episode, I chat with Lee Sharkey about his proposal for detecting computational mechanisms within neural networks: Attribution-based Parameter Decomposition, or APD for short.
Ben Weinstein-Raun
in reply to Ben Weinstein-Raun • •JP Addison
in reply to Ben Weinstein-Raun • •Ben Weinstein-Raun likes this.
Soccum Speleodontidae
in reply to Ben Weinstein-Raun • •Ben Weinstein-Raun likes this.