• mobyduck648@lemmy.world
    link
    fedilink
    English
    arrow-up
    3
    ·
    7 hours ago

    It depends, if Anthropic keep nerfing Claude’s capabilities they’re going to lose to their competitors. I’ve been doing defensive cybersecurity work for the last couple of weeks and Claude, a tool which should be ideal for automating the tedious parts of this, has been a complete pain in the arse.

    Claude: I’ve found a security bug in your codebase that nobody knows about!

    Me: Oh shit that’s not good, can I see it?

    Claude: <blocked by safety classifier>

    Repeat ad nauseam.

    The prissy little bastard’s recalcitrance made me to do a lot of stuff by hand anyway, but instead of being bored I’m now actively irritated by the bot telling me what I can and can’t do. Obviously can’t use them at work for regulatory reasons, but when I was doing similar stuff at home to secure a side project which has to live on the public internet GLM and DeepSeek have no such qualms and honestly I’m convinced the whole ‘DeepSeek is the budget option’ thing is mostly behavioural economics; the cheapness to me feels like efficiency gains more than nerfing the model.