← News

언어-유희 편집부 · 2026-08-11 · EN / KR

Combining Google Gemini Context Caching: AI Pre-Learns In-Game Text to Extract Glossaries, Powering End-to-End Localization & LQA Workbench 'Lexicontexta' Beta Launch

The Perfect Fusion of Lexicon and Context — Overcoming Conventional MT Limitations with a 2-Pass Automated Glossary Pipeline, PM-Dedicated Source QA & Sampling Inspections, and Regression Translation for Temperature Control

The Perfect Fusion of Lexicon and Context > Overcoming Conventional MT Limitations with a 2-Pass Automated Glossary Pipeline, PM-Dedicated Source QA & Sampling Inspections, and Regression Translation for Temperature Control

While advances in generative AI have significantly boosted machine translation fluency, practitioners translating game scripts, academic papers, and technical documents still face deep challenges. Conventional AI translators and general-purpose LLMs struggle to maintain global context continuously, leading to inconsistent proper nouns, fluctuating character tones across dialogue lines, and frequent corruption of structural tags or equations.

To solve these critical industry pain points, Lexicontexta, an AI translation & LQA (Language Quality Assurance) workbench tailored for game and technical document localization, has officially launched its public Beta service.

From Full Text Scanning to Refinement: 'Intelligent Automated Glossary Pipeline'

The core feature built with utmost dedication by the Lexicontexta development team is the Automated Glossary Extraction Engine. Moving beyond conventional translators that merely count raw word frequencies, Lexicontexta incorporates an advanced 3-step hybrid pipeline:

  1. Precision Candidate Scan: The engine scans the entire document to extract initial candidates for proper nouns and domain-specific vocabulary.
  2. AI Contextual Screening: Powered by Google Gemini, the model filters candidates based on background lore and narrative context, retaining only truly meaningful terms.
  3. Pre-translation Directive Confirmation: Users can review, refine, and set target-language term translations prior to initiating full project translation.

Notably, generated glossaries can be extracted and exported standalone (as XLSX files) without running full document translation, making it an independent tool for building Localization Kits (Loc Kits).

'Source QA' & 'Smart LQA' for Localization PMs and Agencies: Quality Verification Without Target Language Fluency

Lexicontexta supports precision QA modes that dramatically reduce workloads for Project Managers (PMs) and localization agencies.

  • Source QA: Before assigning files to translators, the AI detects and cleans up typos, broken tags, mismatched placeholders, and formatting errors directly within the source text.
  • Smart LQA & Sampling Inspection: Even if PMs lack expertise in specific target languages, the AI automatically categorizes and grades mistranslations, glossary violations, tag corruptions, and typos by severity. PMs can generate structured feedback reports for translators instantly. Furthermore, for large projects where full review is impractical, the Sampling Inspection feature allows PMs to quickly evaluate the quality of deliverables from new linguists or vendor agencies prior to full acceptance.

Defending Against Hallucinations with Temperature Control and 'Regression Translation'

Increasing a model's creativity parameter (Temperature) produces more fluid, literary phrasing, but can occasionally trigger hallucinations or tag loss. Lexicontexta incorporates a Regression Translation loop to mitigate high-temperature side effects.

By pinpointing only rows flagged by Smart LQA, it executes surgical re-translation, rectifying errors without altering valid, unaffected segments.

Additionally, Lexicontexta features a zero-cost Local Translation Memory (TM) for exact matches alongside native integration with Google Gemini's Context Caching, reducing AI data ingestion costs by up to 90%.

Enterprise IP Security & Flexible Prepaid Credit Model

Lexicontexta preserves complex document structures including {0} placeholders, LaTeX equations, footnotes, and hyperlinks across game scripts (XLSX/CSV) and technical/academic documents (DOCX/MD). Source files are parsed locally within the browser without being stored on remote servers, and enterprise-grade Google Cloud API pipelines guarantee that transmitted data is never used for AI model training.

Instead of rigid monthly subscriptions, Lexicontexta employs a flexible Prepaid Credit system aligned with project-based budgets. Users can pair various Gemini models (Flash Lite, Flash, Pro) with custom Temperature and Thinking levels.

"Lexicontexta is not just an API wrapper outputting raw translations; it is a comprehensive workbench that grants PMs and translators full control—from source cleanup and automated glossary extraction to tone alignment, translation, and LQA," said a Lexicontexta spokesperson. "Through this Beta launch, we aim to deliver a leap in productivity for localization teams fatigued by manual review."

For more information and Beta access, visit the official website at https://lexicontexta.com.