[논문리뷰] Are You Sure You're Sure? On the Impact of Instruction Tuning on Confidence and Lexical Diversity본 논문은 Instruction Tuning이 모델의 verbalized overconfidence를 유발한다는 기존 연구에 주목하여, 이러한 confidence의 변화가 생성된 근거(Rationale)의 어휘적 다양성과 어떤 상관관계가 있는지 규명하고자 한다.#Review#Instruction Tuning#Model Confidence#Lexical Diversity#Question Answering#Calibration#Rationale#Uncertainty Estimation2026년 8월 13일댓글 수 로딩 중
[논문리뷰] <think> So let's replace this phrase with insult... </think> Lessons learned from generation of toxic texts with LLMs본 연구는 대규모 언어 모델(LLM)이 생성한 독성 텍스트가 텍스트 정화(detoxification) 모델 훈련을 위한 인간 주석 데이터를 효과적으로 대체할 수 있는지 평가하는 것을 목표로 합니다.#Review#Toxic Text Generation#LLMs#Text Detoxification#Lexical Diversity#Synthetic Data#Human Annotation#Style Transfer2025년 9월 11일댓글 수 로딩 중