언어-유희 편집부 · 2026-08-11 · EN / KR
Combining Google Gemini Context Caching: AI Pre-Learns In-Game Text to Extract Glossaries, Powering End-to-End Localization & LQA Workbench 'Lexicontexta' Beta Launch
The Perfect Fusion of Lexicon and Context — Overcoming Conventional MT Limitations with a 2-Pass Automated Glossary Pipeline, PM-Dedicated Source QA & Sampling Inspections, and Regression Translation for Temperature Control
Google Gemini 컨텍스트 캐싱 결합… 작업 시작 전 게임 내 텍스트를 미리 학습한 AI가 용어집 생성, 번역부터 LQA 검수까지 모두 다룰 수 있는 AI 현지화 워크벤치 ‘Lexicontexta’ 베타 오픈
용어집(Lexicon)과 문맥(Context)의 완성형 결합 — 자동 용어 추출 2-Pass 파이프라인, PM 전용 사전 QA 및 표본 검사, Temperature 창의성 제어와 회귀 번역으로 기존 범용 AI 번역의 한계 극복
The Perfect Fusion of Lexicon and Context > Overcoming Conventional MT Limitations with a 2-Pass Automated Glossary Pipeline, PM-Dedicated Source QA & Sampling Inspections, and Regression Translation for Temperature Control
While advances in generative AI have significantly boosted machine translation fluency, practitioners translating game scripts, academic papers, and technical documents still face deep challenges. Conventional AI translators and general-purpose LLMs struggle to maintain global context continuously, leading to inconsistent proper nouns, fluctuating character tones across dialogue lines, and frequent corruption of structural tags or equations.
To solve these critical industry pain points, Lexicontexta, an AI translation & LQA (Language Quality Assurance) workbench tailored for game and technical document localization, has officially launched its public Beta service.
From Full Text Scanning to Refinement: 'Intelligent Automated Glossary Pipeline'
The core feature built with utmost dedication by the Lexicontexta development team is the Automated Glossary Extraction Engine. Moving beyond conventional translators that merely count raw word frequencies, Lexicontexta incorporates an advanced 3-step hybrid pipeline:
- Precision Candidate Scan: The engine scans the entire document to extract initial candidates for proper nouns and domain-specific vocabulary.
- AI Contextual Screening: Powered by Google Gemini, the model filters candidates based on background lore and narrative context, retaining only truly meaningful terms.
- Pre-translation Directive Confirmation: Users can review, refine, and set target-language term translations prior to initiating full project translation.
Notably, generated glossaries can be extracted and exported standalone (as XLSX files) without running full document translation, making it an independent tool for building Localization Kits (Loc Kits).
'Source QA' & 'Smart LQA' for Localization PMs and Agencies: Quality Verification Without Target Language Fluency
Lexicontexta supports precision QA modes that dramatically reduce workloads for Project Managers (PMs) and localization agencies.
- Source QA: Before assigning files to translators, the AI detects and cleans up typos, broken tags, mismatched placeholders, and formatting errors directly within the source text.
- Smart LQA & Sampling Inspection: Even if PMs lack expertise in specific target languages, the AI automatically categorizes and grades mistranslations, glossary violations, tag corruptions, and typos by severity. PMs can generate structured feedback reports for translators instantly. Furthermore, for large projects where full review is impractical, the Sampling Inspection feature allows PMs to quickly evaluate the quality of deliverables from new linguists or vendor agencies prior to full acceptance.
Defending Against Hallucinations with Temperature Control and 'Regression Translation'
Increasing a model's creativity parameter (Temperature) produces more fluid, literary phrasing, but can occasionally trigger hallucinations or tag loss. Lexicontexta incorporates a Regression Translation loop to mitigate high-temperature side effects.
By pinpointing only rows flagged by Smart LQA, it executes surgical re-translation, rectifying errors without altering valid, unaffected segments.
Additionally, Lexicontexta features a zero-cost Local Translation Memory (TM) for exact matches alongside native integration with Google Gemini's Context Caching, reducing AI data ingestion costs by up to 90%.
Enterprise IP Security & Flexible Prepaid Credit Model
Lexicontexta preserves complex document structures including {0} placeholders, LaTeX equations, footnotes, and hyperlinks across game scripts (XLSX/CSV) and technical/academic documents (DOCX/MD). Source files are parsed locally within the browser without being stored on remote servers, and enterprise-grade Google Cloud API pipelines guarantee that transmitted data is never used for AI model training.
Instead of rigid monthly subscriptions, Lexicontexta employs a flexible Prepaid Credit system aligned with project-based budgets. Users can pair various Gemini models (Flash Lite, Flash, Pro) with custom Temperature and Thinking levels.
"Lexicontexta is not just an API wrapper outputting raw translations; it is a comprehensive workbench that grants PMs and translators full control—from source cleanup and automated glossary extraction to tone alignment, translation, and LQA," said a Lexicontexta spokesperson. "Through this Beta launch, we aim to deliver a leap in productivity for localization teams fatigued by manual review."
For more information and Beta access, visit the official website at https://lexicontexta.com.
용어집(Lexicon)과 문맥(Context)의 완성형 결합 > 자동 용어 추출 2-Pass 파이프라인, PM 전용 사전 QA 및 표본 검사, Temperature 창의성 제어와 회귀 번역으로 기존 범용 AI 번역의 한계 극복
생성형 AI의 발전으로 기계번역의 문장력은 향상되었지만, 정작 게임 스크립트나 학술·기술 문서를 번역하는 현업 실무자들의 고민은 깊어지고 있다. 기존 AI 번역기나 범용 LLM은 전체 맥락을 지속적으로 유지하지 못해 고유명사가 통일되지 않거나, 캐릭터의 어투가 문장마다 흔들리고, 각종 구조 태그나 수식이 파괴되는 문제가 빈번하게 발생하기 때문이다.
이러한 현장의 페인 포인트(Pain Point)를 해결하기 위해, 게임 및 전문 문서 현지화에 특화된 AI 번역 & LQA(Language Quality Assurance) 워크벤치 ‘Lexicontexta(렉시콘텍스타)’가 베타(Beta) 서비스를 공식 시작했다.
텍스트 전수 분석부터 정제까지… ‘지능형 자동 용어집(Glossary) 파이프라인’
Lexicontexta 개발진이 가장 심혈을 기울여 구축한 대표 기능은 ‘자동 용어집 추출 엔진’이다. 기존 번역기들이 단순 단어 빈도수만 보여주던 방식에서 벗어나, 고도화된 3단계 하이브리드 파이프라인을 탑재했다.
- 텍스트 후보군 정밀 스캔: 업로드된 문서 전체의 고유명사와 핵심 어휘 후보군을 엔진이 1차 추출한다.
- AI 검토 및 문맥 심사: Gemini 모델이 배경지식과 맥락을 바탕으로 실제 용어집에 포함할 단어만 유의미하게 선별한다.
- 번역 방침 사전 확정: 각 용어를 어떤 타깃 언어로 정제하여 번역할지 유저가 사전에 다듬고 프로젝트를 시작할 수 있다.
특히, 생성된 용어집은 전체 번역 작업을 진행하지 않더라도 단독 extraction & export(XLSX 다운로드)가 가능해, 기존 프로젝트의 로컬라이제이션 키트(Loc Kit) 구축용 독립 도구로도 유용하게 활용할 수 있다.
현지화 PM 및 에이전시를 위한 ‘Source QA’ & ‘Smart LQA’… 언어 몰라도 검수·피드백 가능
Lexicontexta는 프로젝트 매니저(PM)와 에이전시 담당자의 업무 부담을 비약적으로 줄여주는 정밀 검수 모드를 지원한다.
- 원문 QA(Source QA): 번역가에게 파일을 넘기기 전, 소스 텍스트에 포함된 오탈자, 태그 파괴, 짝이 맞지 않는 플레이스홀더, 수식 표기 오류 등을 AI가 미리 탐지하여 정돈한다.
- Smart LQA & 표본검사: PM이 타깃 언어(Target Language)에 대한 전문 지식이 없더라도 AI가 오역, 용어집 위반, 태그 오류, 단순 오탈자를 카테고리/심각도별로 자동 분류해 준다. 번역가에게 정교한 오답 노트와 피드백을 즉시 전달할 수 있는 것은 물론, 전체 전수 검사가 부담스러울 경우 ‘표본 검사(Sampling Inspection)’ 기능을 통해 신규 번역가나 외부 에이전시 작업물의 퀄리티를 사전에 빠르게 평가할 수 있다.
Temperature 창의성 제어와 ‘회귀 번역(Regression)’으로 할루시네이션 완벽 방어
번역 스타일의 유연함을 극대화하기 위해 모델의 창의성 수치(Temperature)를 높일 경우, 문장은 부드러워지지만 간혹 환각(Hallucination)이나 태그 유실이 발생할 수 있다. Lexicontexta는 이러한 고온도(High Temperature) 번역의 부작용을 완벽히 통제하는 '회귀 번역(Regression)' 루프를 탑재했다.
Smart LQA를 통해 결함이 탐지된 행만 핀포인트로 집어내 정밀 재번역을 실행함으로써, 멀쩡한 전체 셀을 건드리지 않고 문제가 발생한 부분만 정확하게 보정한다.
더불어, 번역 시작 전 원문과 완전 동일한 문장을 비용 제로(0 크레딧)로 자동 처리해 주는 '로컬 Translation Memory(TM)'와 구글 Gemini의 컨텍스트 캐싱(Context Caching)을 네이티브로 연동하여, 전체 번역시 AI에 데이터를 주입하는 데 들어가는 비용을 기존 대비 최대 90%까지 절감했다.
기업 자산 보안 유지 & 합리적인 선불 크레딧 방식
게임 스크립트(XLSX/CSV)뿐만 아니라 논문, 매뉴얼, 기술 스펙 문서(DOCX/MD) 작업 시 {0} 플레이스홀더, LaTeX 수식, 각주, 하이퍼링크 등 복잡한 구조를 완벽 보존한다. 원본 파일은 서버에 저장되지 않고 브라우저 로컬 환경에서 파싱되며, Enterprise급 Google API 아키텍처를 적용해 전송 데이터의 AI 모델 학습을 엄격히 차단했다.
Lexicontexta는 매달 정기 결제되는 구독제가 아닌, 프로젝트 예산에 맞춰 충전해 쓰는 ‘선불 크레딧(Prepaid Credit)’ 정책을 도입했다. 사용자는 Gemini Flash Lite, Flash, Pro 등 목적에 맞는 AI 모델과 Temperature, Thinking 레벨을 조합해 자유롭게 작업할 수 있다.
Lexicontexta 개발팀 관계자는 “Lexicontexta는 단순히 번역 결과를 뱉어내는 툴이 아니라, 현지화 PM과 번역가가 작업 전 원문 교정부터 자동 용어 추출, 톤 정립, 번역, 그리고 퀄리티 검수(LQA)까지 완벽하게 제어할 수 있는 종합 워크벤치”라며, “이번 Beta 오픈을 통해 수동 검증에 지쳤던 현 현지화 팀들에게 획기적인 생산성 향상을 제공할 것”이라고 밝혔다.
Lexicontexta 서비스에 대한 자세한 정보와 Beta 참여는 공식 웹사이트(https://lexicontexta.com)에서 확인할 수 있다.