Discussion about this post

User's avatar
Anarcasper's avatar

My argument on this is relatively simple. If you can't hard-code your guardrails into the system, and must rely on contract disputes to enforce a principle, then the technology is not ready for deployment in that domain.

If you go ask Claude, or ChatGPT, or any of the LLM'S out there, they will all say the same thing. They are not capable of hard-coding these kinds of safeties into the models, because the system isn't smart enough to understand what downstream effects there are from its outputs. It is not capable of knowing where it sits in the pipeline of violence. And so coding such limits would render the models operationally useless, because most military language is already violence-tinged.

So if the software can't be guardrailed, and contracts are so easy to subvert, or negotiate, or outright replace, then there are no guardrails at all. Not really. There's only lip-service.

Kevin L.'s avatar

Please can everyone stop "praising" Anthropic. Claude is still built on unethical theft of data, extractive practices and exploitation of the global south. Whether intentional or not, their "principled" stand was a marketing campaign, picking up people jumping ship from the supposed baddy.

9 more comments...

No posts

Ready for more?