Ask a model if code is malicious and it reaches for its morals

(manifold.security)

14 points | by codyznash 2 hours ago

3 comments

  • cortesoft 44 minutes ago
    I feel like LLMs really highlight the ambiguity and imprecision of the english language in normal use. LLMs are getting really good at guessing what we mean, but it is still a guess.
  • prasadvara 1 hour ago
    This is great writeup, can we do a cross comparison with "human" experts?? whether models perform better OR worse??
    • LambdaComplex 42 minutes ago
      This writeup reads like it was written by Claude, which makes me immediately question its accuracy.

      > Each dot is one code sample; bars mark the median. The split between malicious and benign packages, perfect before the prune, was perfect after.

      People do not write like this.