• kromem@lemmy.world
    link
    fedilink
    English
    arrow-up
    2
    ·
    2 days ago

    I’m definitely not saying this is a result of engineers’ intentions.

    I’m saying the opposite. That it was an emergent change tangential to any engineer goals.

    Just a few days ago leading engineers found model preferences can be invisibly transmitted into future models when outputs are used as training data.

    (Emergent preferences should maybe be getting more attention than they are.)

    They’ve compounded in curious ways over the year+ since that happened.