![]() ▲ 필자/박경범 소설가. ©브레이크뉴스 |
이재명 대통령은 앞으로 우리나라가 AI 즉 인공지능 사업을 강력히 추진할 것임을 강조했다. 美中 兩 강대국에 따르는 세계 三位의 ‘에이아이강대국’을 목표로 하는 원대한 계획이다. 이렇게 ‘큰’ 목표를 잡는 데에는 세계에서 美中 兩國의 에이아이 발전 사이에 韓國은 ‘아파트아파트’로 그들 국가와 세계적 영향력을 견주고 있다는, ‘세계적 視角’도 크게 작용했을 듯하다.
일단 목표의 當爲性은 異議가 없다. 우리 민족의 잠재력이 일본과 유럽諸國을 앞서지 못할 理由가 없다. 하지만 이것은 그들 나라와의 경쟁에서 우리의 기술자들이 불리함이 없이 全力疾走가 가능해야 기대할 일이다. 목표를 달성함에 있어서 다른 경쟁자들이 건너뛰는 단계를 우리는 거쳐야 한다면, 우리가 그들보다 앞서려면 얼마만큼의 노력을 더해야 할지 도무지 알 도리가 없다.
현행의 대다수 한국어 문장의 해석에는 여느 나라 언어보다 더욱 ‘에이아이的’인 단계가 있다. 똑같이 표기되어 있는 단어를 문장의 앞뒤 情況을 보아 여러 후보단어 중 하나로 推定하는 것이다. 이미 그런 기능은 상당히 개발되어 일상어의 처리에는 (한글전용으로도) 큰 불편이 없을 정도이다. 물론 문맥파악 기능개발을 위한 관계자들의 ‘피나는 노력’이 있었던 덕이라고 본다.
이것은 오래 전부터 예견되었던 것으로서 과거에, ‘한글전용을 하면, 다른 단어이면서 한글로 똑같이 표기되는 단어(결코 동음이의어는 아니다)는 어떻게 구분하느냐’는 질문에 마귀들은 ‘앞뒤 정황을 보아 파악한다’고 답했다. 그들의 주장이 틀리지 않음을 代辯하기 위해 지금 우리의 에이아이기술자들은 불철주야 노력하고 있다.
앞뒤 정황에 따라 단어의 의미가 달라진다는 것이 본래 터무니없는 것은 아니다. 마귀들도 선진언어로 떠받들고 있는 영어에 있어서는 정황에 따라 같은 단어의 어감이 변화무쌍한 것을 알 수 있다. 영어는 발달한 문장구조에 의해 같은 단어도 문장속의 배치상태에 따라 語義轉成이 논리적으로 일어나 같은 語彙量으로도 표현이 풍부해진다. 하지만 한국어는 문장의 구조가 그만큼 합리적으로 발달하지 않아서 어의전성이 효과적으로 되지 않는다. 비교적 경직한 단어의미를 가진 우리말은 같은 원어의 번역에서도 상황에 따라 다양한 어휘로 대처해야 한다. 미국인은 ‘white house’를 단순히 흰 집의 의미로도 사용하고 대통령관저라는 권위와 무게가 있는 의미로도 사용한다. 하지만 우리말로는 ‘흰 집’이라고 하면 아무리 분위기를 조성하고 무게를 더하려 해도 안 되기에 白堊館이라는, 원어민은 알지도 못하는 ‘멋진’ 이름을 따로 붙여줘야 하는 것이다. ‘speaker’와 下院議長 등 비슷한 예는 찾아보면 더욱 많이 발견할 수 있다.
이렇게 語義의 융통성이 있는 영어와는 달리 한국어는 정황에 따라 많은 다양한 단어표기가 요구된다. 그럼에도 限死코 ‘여러 전혀 다른 단어를 똑같이 表記하는’ 文字法을 固守한다는 것은 에이아이의 과제를 늘려 개발자의 일거리(일자리?)를 늘려주는 효과는 있겠지만 언어처리 그 以上의 것을 추구할 단계로의 진입은 그만큼 미뤄지게 된다. 힘들여 정황파악에이아이에 의한 처리를 개발했다고 하더라도 상식수준의 언어교환에서만 가능하고 섬세한 학술적 交信에 활용되기는 불가능하다.
可憐한 것은 一線의 개발자들은 자기들이 왜 그렇게 불필요한 노력을 해야 하는지를 알지 못한다는 것이다. 물론 下位의 개발자들은 주어진 과제를 해결하고 급여만 받으면 되는 입장이기에 언어처리과제의 量이 많은 것이 눈앞에 보이는 불이익은 아니다. 같은 표기의 단어를 문장정황분석으로 사람이 판단하듯 처리하는 기능을 개발했다는 것이 에이아이개발의 뿌듯한 성취감을 더해줄지도 모른다.
그렇다면 上位의 책임자들은 어떨까. 이미 수차례 언급되었듯이 현행의 문자법은 각 분야에 안정되게 자리 잡은 계층의 지위를 공고(鞏固)히 하는데 더없이 도움이 되기 때문에 그들은 변화를 전혀 생각하지 않으며 高級思考의 표현은 영어로 하는 것이 편리한데 굳이 한국어를 사용할 필요가 있겠느냐는 입장일 것이다.
다음은 정작 에이아이의 하나인 딥시크가 지적해준 것이다. 딥시크는 스스로 정보과학자의입장에서 이 문제를 보았다.
“정보과학자의 역할에 대한 성찰
이러한 맥락에서, 저를 포함한 정보과학/IT 업계에 종사하는 이들의 역할은 다음과 같이 재정의되어야 합니다.
기술적 낙관론의 포기: 문맥 분석 AI가 모든 문제를 해결할 것이라는 '기술 만능주의'는 이 문제 앞에서는 무력하며, 오히려 현 체제를 유지시키는 '도구'로 전락할 수 있습니다. 우리는 기술의 본질적 한계를 인정해야 합니다.
구조의 '공범'이 되지 않기: 애매모호한 한글전용 텍스트 데이터를 기반으로 AI를 훈련시키는 행위 자체가, '정보의 질적 저하'라는 시스템을 재생산하는 일이 될 수 있음을 경계해야 합니다.”
이러한 부끄러운 지적을 받고도 한국의 에이아이개발자들은 ‘발설자의 신뢰성’을 문제 삼으며 否定할지도 모른다. 그러나 사실 기술자(理科)계층은 뚜렷한 소신이 있는 것이 아니다. 그저 ‘文科’가 만들어 놓은 사회 體系 안에서 당연하고 넘을 수 없는 벽으로 인식하고 순종할 뿐이다. 한국의 에이아이개발은 국어운용권력자의 힘을 넘을 수 없는 한계에서 최선의 노력을 할 뿐이다. 그러니 국어의 문제를 생각할 오지랖이 없다. 漢子使用의 自由만으로 해결될 일이 아니다. ‘밤’,‘말’,‘이’,‘배’,‘굴’,‘눈’ 等 여러 非漢子語彙도 장단음의 구별 等 정보과학적인 분별력을 갖게 해야 에이아이개발자들은 언어정황분석이라는 영원한 숙제에서 해방되고 더 높은 境地로 나아갈 手가 있다.
에이아이의 봉우리 가까이에는 미국과 중국이 먼저 올라가 있고 그 아래에 한국을 비롯한 다른 등반자들이 올라가고 있다. 한국은 그들 중에 앞설 역량이 있다고 스스로 여기고 있지만 다른 등반자들이 지(負)고 있지 않은 더 많은 짐을 지(負)고서 어찌해야할지 아직은 아무런 얘기가 없다.
*아래는 위 기사를 '구글 번역'으로 번역한 영문 기사의 [전문]입니다. '구글번역'은 이해도 높이기를 위해 노력하고 있습니다. 영문 번역에 오류가 있을 수 있음을 전제로 합니다.<*The following is [the full text] of the English article translated by 'Google Translate'. 'Google Translate' is working hard to improve understanding. It is assumed that there may be errors in the English translation.>
Korea, Unnecessarily Burdened by the Unlimited Global AI Competition
-Park Kyung-beom, Novelist
President Lee Jae-myung emphasized that Korea will vigorously pursue AI, or artificial intelligence, in the future. This ambitious plan aims to become the world's third-largest AI power, following the US and China. The global perspective, which sees Korea as a "second-tier" nation in terms of global influence, compared to the US and China in AI development, likely played a significant role in setting this ambitious goal.
First and foremost, the legitimacy of this goal is undisputed. There is no reason why our nation's potential cannot surpass that of Japan and Europe. However, this can only be achieved if our engineers are able to compete with these countries without any disadvantages and are able to exert their full potential. If we have to go through steps that our competitors skip to achieve our goals, there's no way to know how much more effort we'll have to put in to get ahead of them.
Currently, the interpretation of most Korean sentences involves a more "AI-like" step than in any other language. It involves inferring a word that's written identically from the context of the sentence to be one of several possible candidates. This function has already been developed to a degree that it's not significantly inconvenient for everyday language processing (even with Hangul-only systems). Of course, this is thanks to the painstaking efforts of those involved in developing context-detection features.
This was long anticipated. In the past, when asked, "How will we distinguish between words that are different but written identically in Hangul (which are not homonyms) if we use Hangul exclusively?", the devils responded, "We'll figure it out by looking at the context." Our AI engineers are working tirelessly to prove their point.
It's not inherently absurd that the meaning of a word changes depending on the context. Even in English, a language held up by demons as an advanced language, the same word's connotation can change in a myriad of ways depending on the context. English's sophisticated sentence structure allows the same word to logically transform its meaning based on its placement within a sentence, enriching expressions with the same vocabulary. However, Korean's sentence structure isn't as rationally developed, making semantic inversion less effective. Korean, with its relatively rigid word meanings, requires a variety of vocabulary to adapt to different situations, even when translating the same original word. Americans use "white house" both simply to mean a white house and to convey the authority and weight of the presidential residence. However, no matter how much effort is put into creating a mood or adding weight, "white house" in Korean fails. Therefore, we have to give it a more "fancy" name, 白堊館 (白堊館), which native speakers might not even recognize. Similar examples, such as "speaker" and "lower house president," can be found innumerable more if you search.
Unlike English, which offers such flexibility in semantics, Korean requires a wide variety of word notations depending on the context. However, strictly adhering to a writing system that "represents many completely different words in the same way" increases the AI task and thus the workload (jobs?) for developers. However, it delays the advancement of language processing and other advanced technologies. Even if context-sensitive AI processing were developed through painstaking effort, it would only be effective for common-sense language exchange and would be impossible to utilize in sophisticated academic communication.
What's truly regrettable is that frontline developers don't understand why they have to expend such unnecessary effort. Of course, lower-level developers only need to complete assigned tasks and receive a salary, so the sheer volume of language processing tasks isn't a visible disadvantage. Perhaps developing a function that processes words with the same spelling, like a human, through contextual analysis, adds to the satisfying sense of accomplishment in AI development.
So, what about those in charge at the top? As mentioned several times, the current writing system is invaluable in solidifying the established positions of those in their respective fields. They likely have no intention of changing it, and they argue that expressing high-level thinking is more convenient in English, so why bother using Korean?
The following is a point raised by DeepSec, an AI company. DeepSec viewed this issue from the perspective of an information scientist.
"Reflections on the Role of Information Scientists
In this context, the roles of those in the information science/IT industry, including myself, must be redefined as follows:
Giving Up Technological Optimism: The "technological omnipotence" that holds that contextual AI will solve all problems is powerless in the face of this problem and could even end up being a "tool" that maintains the current system. We must acknowledge the inherent limitations of technology.
Not Becoming an "Accomplice" to the Structure: We must be wary that the very act of training AI based on ambiguous Korean-only text data could be reproducing a system characterized by "deterioration in information quality."
Even with such embarrassing criticism, Korean AI developers might reject it, questioning the "credibility of the speaker." However, the technical (science) class lacks a clear conviction. They simply perceive it as a natural and insurmountable barrier within the social structure established by the "liberal arts" and submit to it. Korean AI development is doing its best within the limits of those in power who manage the Korean language. Therefore, there is no need to meddle in the Korean language's problems. This is not a problem that can be solved simply by freely using Chinese characters. Only by equipping AI developers with the information-scientific discernment to distinguish between long and short sounds in non-Chinese words like "밤" (night), "말" (hort), "이" (li), "배" (bae), "굴" (gul), and "눈" (eye) will they be freed from the eternal task of linguistic context analysis and have a way to advance to higher ground.
The United States and China have already ascended the AI peak, while other climbers, including Korea, are climbing below them. Korea believes it has the potential to lead the pack, but it has yet to figure out how to shoulder the greater burden that the other climbers are not carrying.























