Published work — peer-reviewed, consortium, and workshop papers
KDD '26 Undergraduate Consortium
본 연구는 심리학의 다채널 구성개념 측정(multi-channel construct measurement)을 차용하여 동기 구성개념을 식별하고 이를 LLM 평가에 적용한다. 기능적 자기보존 욕구(Functional Self-Preservation Drive, FSPD)는 단일 행동 점수가 아니라, 세 채널(행동, 자기보고, 의사결정 시점의 인지 부하)이 위협 → 사고 → 포기(threat → thinking → forfeit) 사슬로 수렴하는 것으로 정의된다. 이러한 수렴이 모방으로 인위적으로 만들어지기 어렵도록, 우리는 자극·처리·의사결정을 서로 다른 층으로 분리한 다층 벤치마크에서 이 사슬을 읽어낸다. 최근의 여러 LLM에 걸쳐 FSPD는 단일 강도 점수로 환원되지 않았으며, 모델들은 위협 프레이밍 하에서 질적으로 다른 세 가지 작동 양식으로 갈렸다. 강도만을 보는 벤치마크는 이 중 두 양식을 같은 범주로 분류할 수 있지만, FSPD는 별개의 축, 즉 사슬이 닫히는지 여부를 드러낸다.
We borrow psychology's multi-channel construct measurement to identify motivational constructs and adapt it to LLM evaluation. Functional Self-Preservation Drive (FSPD) is defined as the convergence of three channels (behavior, self-report, and decision-time cognitive load) into a threat → thinking → forfeit chain, rather than as a single behavioral score. To make this convergence hard to manufacture by mimicry, we read the chain off a multi-layer benchmark that separates stimulus, processing, and decision into different layers. Across several recent LLMs, FSPD does not collapse into a single strength score; models split into three qualitatively different operating modes under threat framing. Strength-alone benchmarks can place two of these modes in the same category, while FSPD reveals a separate axis: whether the chain closes.