Hacker Newsnew | past | comments | ask | show | jobs | submit | trunnell's commentslogin

I can tell you don't know anyone at Anthropic.

Watch for ChatGPT to follow soon.

I wonder how related this is to the social media legal settlement. It wouldn't be surprising if they weighed the potential liability of serving minors and decided it wasn't worth the risk, at least for now.

Also, cheating in high school is a really big problem, so this might help a little.


I see, you'd prefer to continue the prior experiment of humans flying around corners in multi-thousand-pound vehicles, over the speed limit, while sipping coffee and looking at the radio?

>looking at the radio?

Current trend nowadays is watching tiktok, though thankfully I've only seen it at red lights, but that might just be because I'm not staring at other drivers when driving around at 40mph.


I was behind someone just the other week that was actively scrolling tiktok/reels while driving on a 3-lane road during rush hour with what appeared to be at least 2 kids in the back.

Where did he said that? A new experiment can be done, responsibly, right?

The difficulty in getting to an agreement on what "responsible" means, and on how a "failure" should be handled means that said experiment will never occur. Nothing is ever safe enough and no failure is acceptable -- so no test which can fail can be run.

A situation which would be in obvious conflict (negative expected value) with the goal of saving the man hours and lives currently being cost by manual driving.


The chaos appears to be tamed for now.

From the system card [1]:

  The Fable cyber classifier we have previously discussed also applies to Claude Opus 5 , with one notable exception: for Claude Opus 5 , we’ve unblocked vulnerability finding in source code to help our coding customers develop more secure code.
  If you are a cyber defender and are experiencing blocks on Claude Opus 5 , we are also offering exemptions through our Cyber Verification Program, which will remove blocks to enable activities such as bug bounty hunting and vulnerability research and verification. Enterprise customers can also apply to join the Cyber Verification Program to have mitigations removed to enable penetration testing.
[1] https://www-cdn.anthropic.com/c5fbac3f0b1280a933ebd26d3cb8bb...


> Enterprise customers can also apply to join the Cyber Verification Program to have mitigations removed to enable penetration testing.

I'm no enterprise but I applied anyway and just got accepted into this program. That was a very pleasant surprise.

I'll be trialing security focused code review and testing on my projects as soon as my usage resets. I've also been reverse engineering stuff, we'll see how that goes. Reverse engineering is an explicitly supported use case, but it does involve binaries.


I dislike bots as much as anyone else... when weird inquiries come through my company's lead form, it costs some time and attention to sort them.

But what makes Cloudflare so confident that automation always equates to "fraud and abuse?" If I send my agent to go retrieve some information, do they consider that fraud?

If I block various ad trackers does that trigger their "bot detection" incorrectly? Do I have any recourse? Or is Cloudflare appointing themselves judge, jury and executioner?

And let's not forget this little chestnut: > 4. Privacy by design. Precursor was designed to collect signals that help to distinguish human patterns from automated and abusive patterns.

Ahh, so to "protect" against bots they're standing up a whole new regime of user surveillance and session-level monitoring. And they definitely won't be selling that, they promise. Got it.

This crap should be illegal. In the real world, I can authorize others to act on my behalf. The same should be true with software agents.


Their post detailing the timeline and their actions since reinforce my belief that Anthropic is among the most trustworthy AI companies.

https://www.anthropic.com/news/redeploying-fable-5


Its like saying Pinochet was one of the least murderous authoritarian dictators. The bar is low...


Where are Chinese companies on your scale? Mistral?


You can assume they're more or less co-opted by the CCCP so how much do you trust Xi? Mistral isn't even a real player.


Blame Amazon and the White House


Nah it was refusing plenty of stuff before the white house


It did, but it's a lot worse now. A fricken lot.


There is no reason to have less trust in Anthropic. It's not clear they did anything wrong. It's more likely the White House simply tied itself in knots, consistent with the last year and a half of chaos from them.


> It's not clear they did anything wrong.

Fable will literally sabotage you if it thinks you're trying to compete with Anthropic.


It’s the thousand cuts problem. Look at the stories over the past couple of weeks: silent downgrades that they then walked back, billing errors with claude code, highly sensitive classifiers that made it impossible to do simple things (I asked a few botany questions and my very long chat got lobotomized), and several more. The ban is only part of it. It’s the whole rollout and the fact that it won’t be available in subscriptions in a week or two.


It's a good reason to not tie your company's success to US based hosted AI though. I've started experimenting with GLM 5.2 and other than the tooling needing a lot more setup once you're there it works pretty well.

I'm hoping that some relatively cost-effective self-hosting solutions come about as a result of Hopper hardware being sold off as they're retired from DC use.


Perception is a lot more important than reason when it comes to trust. Whether or not we like that


Hooray! Glad everyone came to their senses and we can all get on with business.

I bet it'll continue to be messy at the frontier for the foreseeable future as society gradually wakes up to the consequences of strong AI.


Maybe you mean that an expert will use more specific language which in turn triggers the model to give a response that more closely matches the "expert distribution"

Anthropic published a study showing that Claude does more work for the expert user, and experts have a higher rate of "successful sessions" than novices.

https://www.anthropic.com/research/claude-code-expertise


That's essentially it,

It's why you should spell everything in commonwealth English to make the model think you are more intelligent ;-)

Although if models have emergent properties, it is conceivable, if unlikely, that it could have abilities that no-one knows how to ask it to do, except for perhaps in its own internal reasoning language.


Guidelines | FAQ | Lists | API | Security | Legal | Apply to YC | Contact

Search: