์ด์ง๋ฅ์ผ๋ก ํฅํ๋ ๊ธธ, GPT-4 ํ๋ จ์ํค๋ CriticGPT
OpenAI๋ GPT-4์ ์ค๋ฅ๋ฅผ ์ฐพ๊ธฐ ์ํด CriticGPT๋ผ๋ ๋ชจ๋ธ ๊ฐ๋ฐ
CriticGPT๋ GPT-4 ๋ฒ ์ด์ค๋ก ChatGPT์ ์ฝ๋๋ฅผ ๋นํํด์ฃผ๋ GPT
CriticGPT๋ GPT-4์ ์๋ต์ ๊ฒํ ํ๊ณ ์ค๋ฅ๋ฅผ ์ง์ ํ๋ ์ญํ ์ ํจ
์ฐ๊ตฌ ๊ฒฐ๊ณผ, CriticGPT๊ฐ ์ฌ๋๊ณผ ํ๋ ฅํ์ ๋ ์ค๋ฅ๋ฅผ ๋ ์ ํํ๊ฒ ์ฐพ์๋
CriticGPT๋ ์๋ต์ ๋ ผ๋ฆฌ์ ์ผ๊ด์ฑ์ ๊ฒํ ํ๊ณ , ๋ชจํธํ ํํ์ ์ง์ ํ๋ฉฐ, ์ ๋ณด์ ์ ํ์ฑ์ ํ๊ฐํจ
ํ์ง๋ง ๊ธด ํ ์คํธ๋ ๋ณต์กํ ์ค๋ฅ๋ฅผ ์ฒ๋ฆฌํ๋ ๋ฐ๋ ํ๊ณ๊ฐ ์์
OpenAi์ธก์ ์์ผ๋ก ์ด ๋ชจ๋ธ์ ๊ฐ์ ํด ๋๊ฐ ๊ณํ์ด๋ผ ๋ฐํ
์ต๊ทผ Frontier AI ์ฐ๊ตฌ์๋ค์์ ์ ์ผ ๋ง์ด ๋ ผ์๋๋ ์ฃผ์ ์ค ํ๋๋ ๊ฒฐ๊ตญ ์ด๋ ์๊ฐ ๊ฐ์ข ๋ณ๋ชฉ์ ๋ซ๊ณ AGI โ ASI๋ก ๊ฐ๊ธฐ ์ํด AI๋ก AI๋ฅผ ๊ฐ๋ฅด์น๋ ๋ฃจํ๋ฅผ ์ฐ๊ฒฐํด์ผ ํ๋ค๋ ์ฃผ์
์คํAI์ ๋ณธ๋ฌธ์๋ ๋์์๋ฏ ์ด๋ ์์ผ๋ก ์ด์ง๋ฅ์ ์๋์ '์ธ๊ฐ๋ณด๋ค ๋ ๋์ ์ง๋ฅ์ ์ด๋ป๊ฒ ๊ด๋ฆฌ๊ฐ๋ ํ ๊ฒ์ธ๊ฐ'์ ๋ํ ์ฃผ์ ์ ๋ํ ์ฐ๊ตฌ
*This is a step towards being able to evaluate outputs from advanced AI systems that can be difficult for people to rate without better tools.*
์ง๊ธ๊น์ง ๊ด๋ฒ์ํ๊ฒ ํ์ ์ ๊ฐ๋ฅ์ผ ํ๋ RLHF๋ ํ๊ณ์ ๋ถ๋ชํ ๊ฒ์ด๊ธฐ ๋๋ฌธ CriticGPT๋ ๊ฒฐ๊ณผ์ ์ผ๋ก ์ฌ๋๋ง ํ๋ ๊ฒ๋ณด๋ค ๋ ๋์ ๊ฒฐ๊ณผ๋ฅผ ๋๊ณ ์คํAI๋ ํฅํ ์ด๋ฅผ ํ์ฅํ์ฌ ์ค์ ๋ก ๋ชจ๋ธ์ ์ ์ฉํ๊ฒ ๋ค๋ ๋ง์ ๋จ๊น
*In order to align AI systems that are increasingly complex, weโll need better tools. In our research on CriticGPT, we found that applying RLHF to GPT-4 has promise to help humans produce better RLHF data for GPT-4. We are planning to scale this work further and put it into practice.*
https://openai.com/index/finding-gpt4s-mistakes-with-gpt-4/
OpenAI๋ GPT-4์ ์ค๋ฅ๋ฅผ ์ฐพ๊ธฐ ์ํด CriticGPT๋ผ๋ ๋ชจ๋ธ ๊ฐ๋ฐ
CriticGPT๋ GPT-4 ๋ฒ ์ด์ค๋ก ChatGPT์ ์ฝ๋๋ฅผ ๋นํํด์ฃผ๋ GPT
CriticGPT๋ GPT-4์ ์๋ต์ ๊ฒํ ํ๊ณ ์ค๋ฅ๋ฅผ ์ง์ ํ๋ ์ญํ ์ ํจ
์ฐ๊ตฌ ๊ฒฐ๊ณผ, CriticGPT๊ฐ ์ฌ๋๊ณผ ํ๋ ฅํ์ ๋ ์ค๋ฅ๋ฅผ ๋ ์ ํํ๊ฒ ์ฐพ์๋
CriticGPT๋ ์๋ต์ ๋ ผ๋ฆฌ์ ์ผ๊ด์ฑ์ ๊ฒํ ํ๊ณ , ๋ชจํธํ ํํ์ ์ง์ ํ๋ฉฐ, ์ ๋ณด์ ์ ํ์ฑ์ ํ๊ฐํจ
ํ์ง๋ง ๊ธด ํ ์คํธ๋ ๋ณต์กํ ์ค๋ฅ๋ฅผ ์ฒ๋ฆฌํ๋ ๋ฐ๋ ํ๊ณ๊ฐ ์์
OpenAi์ธก์ ์์ผ๋ก ์ด ๋ชจ๋ธ์ ๊ฐ์ ํด ๋๊ฐ ๊ณํ์ด๋ผ ๋ฐํ
์ต๊ทผ Frontier AI ์ฐ๊ตฌ์๋ค์์ ์ ์ผ ๋ง์ด ๋ ผ์๋๋ ์ฃผ์ ์ค ํ๋๋ ๊ฒฐ๊ตญ ์ด๋ ์๊ฐ ๊ฐ์ข ๋ณ๋ชฉ์ ๋ซ๊ณ AGI โ ASI๋ก ๊ฐ๊ธฐ ์ํด AI๋ก AI๋ฅผ ๊ฐ๋ฅด์น๋ ๋ฃจํ๋ฅผ ์ฐ๊ฒฐํด์ผ ํ๋ค๋ ์ฃผ์
์คํAI์ ๋ณธ๋ฌธ์๋ ๋์์๋ฏ ์ด๋ ์์ผ๋ก ์ด์ง๋ฅ์ ์๋์ '์ธ๊ฐ๋ณด๋ค ๋ ๋์ ์ง๋ฅ์ ์ด๋ป๊ฒ ๊ด๋ฆฌ๊ฐ๋ ํ ๊ฒ์ธ๊ฐ'์ ๋ํ ์ฃผ์ ์ ๋ํ ์ฐ๊ตฌ
*This is a step towards being able to evaluate outputs from advanced AI systems that can be difficult for people to rate without better tools.*
์ง๊ธ๊น์ง ๊ด๋ฒ์ํ๊ฒ ํ์ ์ ๊ฐ๋ฅ์ผ ํ๋ RLHF๋ ํ๊ณ์ ๋ถ๋ชํ ๊ฒ์ด๊ธฐ ๋๋ฌธ CriticGPT๋ ๊ฒฐ๊ณผ์ ์ผ๋ก ์ฌ๋๋ง ํ๋ ๊ฒ๋ณด๋ค ๋ ๋์ ๊ฒฐ๊ณผ๋ฅผ ๋๊ณ ์คํAI๋ ํฅํ ์ด๋ฅผ ํ์ฅํ์ฌ ์ค์ ๋ก ๋ชจ๋ธ์ ์ ์ฉํ๊ฒ ๋ค๋ ๋ง์ ๋จ๊น
*In order to align AI systems that are increasingly complex, weโll need better tools. In our research on CriticGPT, we found that applying RLHF to GPT-4 has promise to help humans produce better RLHF data for GPT-4. We are planning to scale this work further and put it into practice.*
https://openai.com/index/finding-gpt4s-mistakes-with-gpt-4/
๋ ผ๋ฌธ์ถ - dc App
์์ผ๋ก ํ๊ฒ ๋ค๋ใ ๊ฑฐ๊ตฌ๋.. ๋ค์ gpt๋ก ๋์ด๊ฐ์ผ ํ๋ ๊ณ ๋ฏผํ๋๋ ์ผ๋จ์ ํด๋ก๋๋ค! - dc App
๊ทผ๋ฐ ์ ํ์ฑ์ ํ๊ฐํ๋ AI์ ์ ํ์ฑ์ ๋๊ฐ ํ๊ฐํ๋๊ฑฐ์ ์ AI๊ฐ '์ด๊ฑด ์ฌ๋ฐ๋ฅธ ์ ๋ณด๋ค'
๋ผ๊ณ ํ๋ฉด ๊ทธ๊ฒ ์ฐ์ธ์ ์งญ์ธ์ง ์ด์ผ์์ด ๊ฑ ๋ฏฟ๋๊ฑฐ์?
์ผ๋จ์ ์ ์ฑํ๊ฐ ํด์ผ ํ ๋ฏ? ์๋๋ฉด, ์ ๋ต์ด ๋ผ๋ฒจ ๋ ์ฝ๋ ๋ฌธ์ ๋ค์ ์ ๋ ฅ ์ํค๊ณ ๊ฐ์ ์ํจ ๊ฑธ ํ๊ฐ์งํ๋ก ํํํ ์๋ ์์๊ฑฐ๊ณ . ์์ ๋ง ์ ๋๋ ๊ฑด ์๋๊ฒ, openAI๊ฐ GPT ๋ง๋ค ๋ RLHF ์ผ๋ฏ, ์คํธ๋กํฝ๋ ์๊ธฐ๋ค ์ฑ๋ด ๊ฐ์ ํ ๋ ๋ค๋ฅธ ์ฑ๋ด์ ํตํด์ ๊ฒํ ํ๋ค๊ณ ํ์๊ฑฐ์.
๋ธ๋ก์ฒด์ธ ์จ์ผ์ง - dc App
์ธ๋์ธ 5์ฒ๋ช ์ผ๋ก ์ ์๊ฒ์ฌํจ
๊ฒฐ๊ณผ๋ฅผ ๋ณด๊ณ ํ๋จํด์ผ์ง. ์ฌ๋์ ์ ํ์ฑ์ ๋๊ฐ ํ๊ฐํจ? ์ธํฐ๋ท ์ ๋ณด ์ ๋ถ ์ถ์ฒ๋ ๊ทผ๊ฑฐ ์ฐพ์๋ณด๋ฉด์ ํ๊ฐํจ?
์๋ค ๋ ผ๋ฌธ์ ๊ณต๊ฐํ๊ธด ํ๋๊ตฌ๋
ai๋ก ai๋ฅผ ๊ฐ๋ฅด์น๋ ์ง๋ฅํญ๋ฐ์ ์ํด์ ํ์ํ ์ค๋น์ธ๊ฐ
์ ์ด์ ์ง๋ํ์ต ์์ด์ง๋