Diego Almeida, a former OpenAI researcher who helped invent the reinforcement learning techniques behind today’s most popular chatbots, recently made a surprising observation: despite their massive success, language models have failed to deliver true software automation. Chatbots are excellent at talking to humans, but computer programs speak an entirely different language. When developers try to…
ChatGPT නිර්මාණය කිරීමට සහ වර්තමාන AI යුගයේ ගාමක බලවේගය වන RLHF (Reinforcement Learning from Human Feedback) තාක්ෂණය ගොඩනැගීමට දායක වූ ප්රධාන පර්යේෂකයෙකු වන ඩියාගෝ අල්මේදා (Diego Almeida) මෑතකදී අපූරු ප්රකාශයක් කළේය. ඔහු පැවසුවේ ChatGPT වැනි භාෂා ආකෘති (LLMs) මිනිසුන් සමඟ කතාබස් කිරීමට ඉතා දක්ෂ වුවද, සැබෑ ‘ස්වයංක්රීයකරණයක්’ (automation) ඇති කිරීමට ඒවා අසමත් වී ඇති බවයි.…