Enterprise Ready · 100% Private

Private Machine Translation for High-Stakes Domains

Private engines trained on your domain. Tags and terminology stay intact — even when the source dataset is small.

Translation Drop Zone
ISO 27001 Aligned
GDPR Sub-Processor
Zero Tag Loss

How it works

01 / 05

We clean your data

Hyper-Tuned Precision

BLEU Score Comparison

Real client results on held-out domain data. Select your domain to see how LexaForte performs against the strongest available baseline.

DeepL Pro

Commercial baseline

32.0 BLEU

Google AutoML

Custom trained baseline

56.5 BLEU
Best in Domain

LexaForte

Hybrid NMT + LLM · EN → RU

73.5 BLEU

Real client result on held-out Chemistry & Pharma data. Commercial-grade quality from small, specialized datasets.

Scores measured with SacreBLEU on held-out domain data. LexaForte engines are fully private, tag-safe, and deployable on your infrastructure or ours.

Why Hybrid Wins

NMT vs LLM vs LexaForte Hybrid

Same technical segment. Three engines. See exactly what each one does to your content.

Source

Measure <chem_tag>H2SO4</chem_tag> purity at 20°C (target: 0,5 M).

Standard NMT

Fast · Literal · Fragile

Messen Sie H2SO4 Reinheit bei 20°C (Ziel: 0,5 M).

  • ✗ Tags completely dropped
  • ✗ Literal verb, no domain term
  • ✗ Wrong decimal & unit format

Pure LLM

Fluent · Unstable · Expensive

Bestimmen Sie die Reinheit von H2SO4 bei 20 Grad Celsius (Zielwert: 0.5 Mol).

  • ✗ Tags removed / rewritten
  • ~ Style improved, subscripts lost
  • ✗ Unit altered, formula risk

LexaForte Hybrid

Precise · Locked · Controlled

Bestimmung der <chem_tag>H2SO4</chem_tag>-Reinheit bei 20 °C (Sollwert: 0,5 mol/l).

  • ✓ Tags fully preserved
  • ✓ Formula & subscripts locked
  • ✓ Corporate term + SI unit applied

NMT alone drops structure. LLM alone invents. Hybrid keeps the speed of NMT and the control of a locked terminology layer.

Security & Topology

Isolated Dual-Server Network

The processing engine is physically isolated from the public internet. Public requests never hit model memory directly.

Server 01 · Gateway Helsinki · CCX23

API gatekeeper & portals

Public face of the stack: marketing site, client portals, Nginx TLS termination, and reverse proxy to the engine. No models live here.

  • Hosts Drop Zone, feedback, and api.lexaforte.com
  • Forwards only proxied API traffic — keys are enforced on the engine
  • No glossaries, no training data, no CTranslate2 binaries on this box
Public endpoint Live
IP firewalled
Server 02 · Engine Falkenstein · AX41

Isolated CTranslate2 core

Dedicated host for float16 engines, client shields, LRU cache, and optional LLM boost. Port 8000 is locked to the gateway.

  • IP lockdown. Accepts traffic only from Server 01
  • No public ingress. Not reachable from the open web
  • Dedicated option. Private GPU instance available if you want the LLM hop on your hardware
Inference network Isolated host

Data Architecture

LexaForte TMX Purifier

Proprietary cleaning engine covering ~20 scripts — Latin, Cyrillic, CJK, Ethiopic and more. Noise is removed before a single training step runs.

Layer 01 2M segs / 5s

Vectorized Fast-Pass

Rust / C++ multi-threaded pass that vaporizes control characters, PUA Unicode, broken text, duplicate segments and OCR artefacts.

~2,000,000 rows per 5 seconds
Layer 02 Multi-core

Multiprocess Heuristics

Source and target checked together across CPU cores: number consistency, script-leak rates, bracket symmetry, and tag balance.

HTML / XML tags and placeholders normalized
Layer 03 Zero write

Synthetic Dry-Run

Purification is simulated with nothing written to disk. Every kept or killed pair is logged with a reason, then locked into a client JSON rule set.

Company-specific JSON policy for the real run

Run Layer 1 on your file

Upload a TMX or TSV for an instant diagnostic audit.

Drop TMX / TSV / parallel file here

or

Core Advantages

Built for High-Stakes Content

Every layer of LexaForte was engineered to eliminate the chronic failure modes of commercial engines.

Zero Tag Loss

HTML, placeholders, chemical subscripts and proprietary tokens are natively protected. No post-processing tax.

Dual-Tier Terminology

Static lock-in during training + dynamic runtime hot-fixes. Product names stay correct forever.

Isolated EU Engine

NMT runs on a firewalled production host. Training sets are deleted after fit. Shared-price LLM boost uses OpenRouter; a dedicated GPU instance is available when you want the model on your path.

Dirty-Data Mastery

Industrial purification engine that repairs rather than discards. Top metrics from small or messy memories.

Living Model

2 free hot-fixes every month. Feedback → permanent improvement. The engine gets better the longer you use it.

Private or Shared

Deploy on our infrastructure or your own. Lightweight ~150 MB engines. Full control when you need it.

Living Engine

The Model Gets Better Every Month

Feedback is applied within 24–48 hours and locked into the engine, so the same class of error does not return.

Source · Software / UI domain

Click <ui_btn>Run Diagnostic</ui_btn> to calibrate pressure sensor.

1

Initial Model

Base engine

~35%

Time-to-Edit

Issues present

Klicken Sie Run Diagnostic um Drucksensor zu kalibrieren.

  • • UI tag destroyed
  • • Button untranslated
  • • Awkward phrasing
2

+ Hot-fix 1

Tag shielding

~12%

Time-to-Edit

Major errors fixed

Klicken Sie auf <ui_btn>Diagnose ausführen</ui_btn>, um den Drucksensor zu kalibrieren.

  • • UI tags restored & locked
  • • Button string translated
  • • Style still imperfect
3

+ Hot-fix 2

Corporate style

<3%

Time-to-Edit

Near-final quality

<ui_btn>Diagnose starten</ui_btn> wählen, um den Drucksensor zu kalibrieren.

  • • Corporate verb “wählen”
  • • Button label refined
  • • Fully aligned

Ongoing

Locked in runtime

~0%

Recurring effort

Permanent

UI tags, button glossary and corporate phrasing stay locked in the model. Matching UI strings need virtually no post-editing.

Financial Impact

Post-Editing Savings Calculator

Modelled residual effort after 2–3 feedback cycles: ~12%. Adjust the numbers to match your reality.

Your Numbers

2,000,000
200k 15M
$0.10
$0.04 $0.25
38%
15% 60%

LexaForte residual effort after 2–3 iterations is modelled at ~12% (conservative). Character-to-word ratio: 6 : 1.

Estimated Monthly Savings

Words per month

333,333

Hours saved

Money saved

Annual impact (×12)

Based on ~12% residual PE effort after 2–3 feedback cycles. Real results are frequently even lower.

Transparent Pricing

Simple, Usage-Based

Pay only for characters you translate. No seats, no platform fees, no surprises.

Direct API

For teams connecting natively

$50 / 1M characters
  • Pay-as-you-go
  • 2 free hot-fixes / month
  • 1 full retrain every 6 months
  • No monthly minimum
Most common

Turnkey / CAT

Via partner platforms or connectors

$50 / 1M characters
  • Same rate
  • $200 / month minimum commitment
  • Covers first 4M characters
  • Full maintenance included

Onboarding Roadmap

From Zero to Production

01

Prototype

Days 1–3

Bespoke engine for your domain & language pair. Data sterilization + baseline training.

02

Optimization

Days 3–7

2–3 targeted performance loops using your feedback. Terminology and style locked in.

03

Approval

Quality sign-off

Final check against your internal benchmarks. Commercial terms aligned.

04

Go-Live

Production

API keys, endpoint routing, CAT connectors. Continuous improvement begins.

Zero risk. If the engine still misses your quality bar after the improvement loops, you pay nothing.

Risk-Free Evaluation

Launch Your Risk-FreeEvaluation Phase

Bespoke prototype engine delivered within 1–3 business days. Send your sample files. Receive production-ready translations + full telemetry.

Translation Drop Zone