Ask a model if code is malicious and it reaches for its morals

(manifold.security)

15 points | by codyznash 5 hours ago ago

4 comments

  • prasadvara 5 hours ago ago

    This is great writeup, can we do a cross comparison with "human" experts?? whether models perform better OR worse??

    • LambdaComplex 4 hours ago ago

      This writeup reads like it was written by Claude, which makes me immediately question its accuracy.

      > Each dot is one code sample; bars mark the median. The split between malicious and benign packages, perfect before the prune, was perfect after.

      People do not write like this.

  • cortesoft 4 hours ago ago

    I feel like LLMs really highlight the ambiguity and imprecision of the english language in normal use. LLMs are getting really good at guessing what we mean, but it is still a guess.

  • undefined 5 hours ago ago
    [deleted]