The Aligner pairs your original documents with their existing translations, sentence by sentence, and gives you a single file of aligned pairs. Use it to turn translations you already have (from before wxrks, or from another tool) into a Translation Memory (TM), a bilingual spreadsheet, or a parallel corpus (a collection of texts paired with their translations, used for example to train or evaluate AI models).
Who is this for? Account Admins and Project Managers. The Aligner appears only for these two roles. Clients and Vendors don't see it in the menu.
What you'll need:
The original documents (the source) and their translations (the targets).
Files in one of these formats:
.txt,.md,.json,.html,.htm,.xml,.docx.
Key concepts
Term | What it means |
Source files | Your original documents, all in one language. |
Target batch | A group of translated files in one target language. Add one batch per language: one for French, one for Portuguese, and so on. |
Aligned unit | One source sentence paired with its translation. It's the equivalent of one segment in a TM. |
Confidence | A score from 0 to 1 that the AI model gives each aligned unit, showing how sure it is that the two sentences are translations of each other. |
Confidence Threshold | The minimum confidence a unit needs to be kept in the output. Units below it are discarded. |
Step 1: Open The Aligner
In the sidebar, click Tools, then The Aligner.
Step 2: Add your source files
In the Source Files panel, click Add Files and select your original documents. You can select several at once.
Check the Source Language field. It starts as Auto-detect. After the upload, wxrks detects the language of your files and selects it, for example English. If the detected language is wrong, pick the right one from the list.
ℹ️ Note: Only the accepted file types listed in the panel can be added. If you select other files, The Aligner rejects them with Not accepted: {file names} and adds nothing. Text files (.txt, .md, .json, .html, .htm, .xml) must be saved with UTF-8 encoding.
Step 3: Add a target batch for each language
In the Target Batches panel, Batch 1 is already there. Click Add Files on that batch and select the translations for one language.
Check the batch's Target Language field. wxrks detects the batch language after the upload, just like for the source. Pick a different language from the list if needed.
Have translations in another language? Click Add Batch and repeat. Click Show or Hide to expand or collapse a batch. The trash icon removes a batch, but the last remaining batch can't be removed.
💡 Tip: Detection selects the base language, such as Portuguese, and the output file is tagged with that code (pt). If the TM you'll import into uses a regional variant, select it instead, for example Portuguese (Brazil), so the file is tagged pt-BR. Do the same for the source language.
How files are paired
Within each batch, The Aligner first pairs source and target files by file name. It ignores upper and lower case, the file extension, and a language code at the end of the name. For example, all of these pair with each other:
Source file | Target file in the Portuguese batch |
|
|
|
|
Files left without a name match are then combined, on each side, and aligned as one block of text.
💡 Tip: Give each translation the same name as its original, plus a language code. Name-matched pairs align document to document, which gives more accurate results than combined, unmatched text.
⚠️ Warning: If, after name matching, only one side of a batch still has files left over (for example, an extra target file with no source to pair with), that content is skipped. A warning in the results tells you how many documents were skipped.
Step 4 (optional): Adjust the settings
Click Settings in the top-right corner of the page.
Setting | Options | Default |
Segmentation Mode | Language-specific or default: splits text into sentences with rules made for each language, falling back to the general rules when a language has none. | Language-specific or default |
Confidence Threshold | Any value from 0 to 1. | 0.80 |
Click Save. These settings, and the output format, are saved in your browser. They're still there the next time you open The Aligner on the same browser.
ℹ️ Note: The Aligner uses its own built-in sentence rules. Custom Segmentation Rules configured in your account don't apply here.
Choosing a Confidence Threshold
The threshold is a trade-off: raise it to keep only pairs the AI model is sure about, lower it to keep more pairs and review them yourself.
Example: a run produces 120 aligned units. 95 score 0.80 or higher, 18 score between 0.50 and 0.79, and 7 score 0.20.
At 0.80, the file contains 95 units and 25 are discarded. Use this when the output goes straight into a TM.
At 0.50, the file contains 113 units. Use this when you'll review the pairs in a spreadsheet before importing them.
⚠️ Warning: Avoid setting the threshold to 0.20 or lower. When the AI model can't align part of a document, The Aligner falls back to pairing sentences in order (first with first, second with second) and gives those pairs a confidence of 0.20. These pairs are often wrong. At the default 0.80 they're discarded. At 0.20 or lower they end up in your output, and wrong pairs in a TM become wrong matches in future projects.
Step 5: Run the alignment
In the Output panel, choose a Format:
TMX (Translation Memory eXchange): the standard TM file format. Choose it to import the result into a TM.
XLSX: a spreadsheet with a
sourceTextcolumn and, for each target language, a column for the translation and one for its confidence. Choose it to review the pairs.JSON: the same data in a structured format, for developers or scripts.
Click Run Alignment. The button changes to Running....
Follow the Run Log at the bottom of the page. It shows each stage (validate, extract, detect-language, match-files, align, export), with the newest entry at the top.
To cancel, click Stop. The run ends and the page shows Alignment was stopped.
Step 6: Review and download the result
When the run finishes, the Output panel shows Alignment complete: {N} units ({format}) and a summary:
Confidence threshold: the threshold this run used.
Included and Discarded: how many units passed the threshold and how many were dropped. {N} in the heading is the total of both, so the downloaded file contains only the Included units.
Unmatched source segments: source sentences the AI model couldn't pair with any translation. They're left out of the output.
Source Language and Batch languages: the languages used for each side.
Warnings (only when there are any): anything that needs your attention, such as skipped content or a language with no specific sentence rules.
ℹ️ Note: Units are counted per language. Ten source sentences aligned into Portuguese and French give Alignment complete: 20 units (TMX).
Click Download to save the file. It's named after the run's format, for example alignment.tmx. All batches go into one file: each source sentence appears once, with its translation in every target language you aligned.
ℹ️ Note: The downloaded file always uses the format of the run, even if you change Format afterwards. To get a different format, choose it and click Run Alignment again.
⚠️ Warning: Download the file before you leave or reload the page. The result isn't saved to your account, and once you leave the page the Download button is gone.
Step 7: Import the result into a Translation Memory
Open the TM you want to fill and click Import.
Under Languages, choose which languages from the file to import:
ALL (default): imports every language in the file. Target languages the TM doesn't have yet are added to the TM.
SELECTED: imports only the Source Language and Target Language shown. They start as the TM's own languages. Other languages in the file are skipped.
Click Choose File, select the
.tmxfile you downloaded, and click Upload.
Example: you aligned 10 English sentences into Portuguese and French, and you import the file into an English (United States) → Portuguese (Brazil) TM.
Languages option | What the TM ends up with |
ALL | 30 segments in EN-US → FR-FR, PT-BR. French is now one of the TM's target languages. |
SELECTED (English → Portuguese) | 20 segments in EN-US → PT-BR. The French translations are skipped. |
⚠️ Warning: If the TM should keep only its own language pair, choose SELECTED before you upload. With ALL, every extra language in the file becomes part of the TM.
For more on importing, see Import in All about Translation Memories.
Troubleshooting
Message | What it means and what to do |
Add at least one source file. | No source files were uploaded. Add them in Source Files (Step 2). |
Batch {N} has no files. | A batch is empty. Add files to it, or remove it with the trash icon. |
Not accepted: {file names}. Accepted file types: … | You selected a file type The Aligner can't read. Convert the file to one of the accepted types, such as |
Could not read {file name} as text. Save it with UTF-8 encoding and upload it again. | The text file uses another encoding. Open it in a text editor, save it as UTF-8, and add it again. |
{N} source segments had no match in the target files and were left out of the output. (warning) | Part of the source has no translation in the target files, for example a paragraph that was never translated. Check that the target files are complete. |
Skipped unmatched-only content in Batch … (warning) | Files in this batch had nothing left on the other side to pair with. Check the file names (see How files are paired in Step 3), or add the missing source or target file. |
No SRX file found for … language …; using fallback segmentation. (warning) | wxrks has no language-specific sentence rules for this language, so it used general rules. The run still completes. Review the output if the language splits sentences unusually. |
Alignment complete: 0 units | Nothing could be aligned. Check that the files contain text and are translations of each other, and read the Warnings for skipped content. |
Included is much lower than Discarded | Many pairs scored below the threshold. Check that the files really are translations of each other, then try a lower Confidence Threshold (see Step 4), staying above 0.20. |
The Aligner isn't in the Tools menu | Your role isn't Account Admin or Project Manager. Ask an Account Admin to check your role. |






