|
1. Rememo Me? (Premium Content)
|
For the last few months I have been keeping tabs on MemoQ's latest
releases, which have come out with refreshing regularity. Every so often there
was enough interesting material to write about, but I always held out. Until
now.
This last week MemoQ 3.5 was released, and here
are the new features according to Kilgray's (the Hungarian company behind MemoQ):
·
Longest substring concordance
·
Wildcard concordance and wildcard in
terms
·
STAR Transit filter
·
Bi-directional language enhancements
·
Horizontal edit view
·
XML preview feature
·
PowerPoint 2007 filter
·
Drastic server speed improvements
I downloaded the new version and specifically looked at the first three
features, which I found truly ground-breaking.
Let's start with the first feature first, the oddly named "longest
substring concordance" (or, just as odd: LSC). This is an attempt to automate
concordance searches according to user-defined parameters (under Tools> Options>
Subsegment leverage you can set the minimum number of concordance hits
and/or the minimum number in words/letters/characters) and without interrupting
the workflow. In the Translation results pane there are now not only
translation memory matches, terminology matches, and assembled rows (available
since version 3.2) but also ominous-looking matches of subsets of the string
that needs to be translated with empty targets. These are the LSC matches.
Clicking on them will produce the Concordance dialog, which will list all
(or a predefined number of) the appearances of that particular substring in the
translation memory within the context of the complete strings in the TM.
Confused?
Imagine you have this (real-life) sentence:
The starter motor rotates
the engine during the start sequence, by driving it through the reduction gear
unit assembly.
MemoQ might find that
"the reduction gear unit" is a worthwhile subsegment and will display this in the Translation
results pane with an empty target. Double-clicking on it would bring up the
Concordance dialog with these options:
The starter motor rotates the engine during the start
sequence, by driving it through the reduction gear unit assembly.
A coupling assembly mechanically connects the main
output drive shaft of the reduction gear unit to the driven unit.
Individual accessory drive pads for the main lube oil
pump are also incorporated on the reduction gear unit.
with their respective translations. If you wanted to use any of those,
you would just need to highlight the part of the translation you want to use
and select Insert selected.
I really like this feature because it offers a new granularity to TM
materials without being obtrusive (the program does not slow you down by
displaying an additional dialog like a comparable feature in Trados, plus
you are free to use the displayed option or not), it enables easy paste access
to the desired translation, and MemoQ's superior search-and-lookup speed
allows it to operate without even seeming to slow down the search process too
much.
Speaking of the Concordance dialog, that same dialog is also used
for the enhanced concordance features. (A "concordance search" is the
process of manually highlighting one term or phrase in the source segment,
pressing a shortcut key -- in the case of MemoQ, it's Ctrl+K -- and accessing all occurrences
of that in the TM.) What makes MemoQ's concordance feature attractive is
the possible automatic addition of a wildcard character. In the Concordance
dialog you can select the option Add wildcard to selected text, and for
the current and next search(es) MemoQ automatically adds an asterisk (*)
to each of the terms. This will make MemoQ look for 0 or more additional
characters to the term in question, something that is particularly helpful for
languages with heavy flexion. (Plus, I really like this because I always forget
where the asterisk is on the German keyboard and I get tired of switching back
to the English keyboard to enter it manually.)
The last feature that really caught my interest was the Transit
compatibility feature. Now, Transit is a great program, but it's really
different, and many folks just don't want to spend the time to learn to use the
free Satellite edition (even if they might miss out on something). For
those, this feature will be very useful. It allows you to import a PXF file --
this is a Transit-specific package file with the translation files, reference
material (i.e., TM), and glossary data -- extract the translation
files as well as the TM content, translate it, and then send it back to your
client as a TXF file -- Transit's return package format.
In general it works very well. There are a few glitches -- a couple of
strings (out of a few thousand in a test run) were unduly protected as tags,
and I had to switch my import mode to get all my data, but it's impressive that
the MemoQ developers were able to use Transit logic to
"harvest" reference material that is readily available in the TM you
selected when you started translating. What this feature is not able to do is
automatically import the TermStar glossaries. This is somewhat unfortunate
because Transit projects tend to be very terminology-heavy because of
its excellent terminology tool.
Here are some other caveats with the new version: Word 2007 is
still not supported -- though PowerPoint and Excel 2007 are --
and TBX, the termbase exchange standard, is also not supported yet.
This morning I had a chance to talk with the owner of a translation
agency who has been using the server edition of MemoQ for a while, just
to get an idea of what his take on performance and user acceptance was. He was
very positive overall. Compared to other server-based products (Idiom WorldServer,
Logoport, and Across) he reported equal or better server response
times. He also liked the option for translators to check out resources to work
offline or the online document storage. This last feature allows translators
not only to share translation memories and terminology databases, but also the
actual documents, which -- optionally -- can be server-based as well. Thus,
multiple users can have access to large documents that can be translated and
edited at the same time for faster turnaround. His assessment of that process
was kind of interesting: He ran into problems with translators not working
successively from top to bottom through documents, creating havoc for the poor
editors; however, that seems to be less a technical limitation than an
organizational one.
As far as user acceptance, he acknowledged that not all his translators
were super-eager to adopt a new tool at the drop of a hat, but it helped that
they did not have to pay for the program. With MemoQ's mobile licensing
concept, he is able to assign temporary full licenses to his users. Interestingly
-- and he was not the first to mention it -- Déjà Vu users in particular
are struggling with the idea of using a different tool.
Speaking of licensing, I am interested to see how MemoQ's
licensing scheme will be adopted once it's time for it. Last fall Kilgray adopted
a new system in which all upgrades are free for a year after purchase, no
matter how major or minor they might be; after that year, a 20% annual fee is
applicable for further upgrades. This is certainly not an uncommon practice in
the software industry in general, but as far as I know it is unusual for the
individual user sector in our industry. We'll see what happens come fall of
2009.
|
2. Passolo's New Look
| |
At the beginning of this week, SDL released a new Passolo
website in SDL look (I'm very sorry to see the old Passolo website go away!)
and, more importantly, a new version of Passolo. Most of
you know that Passolo is a localization tool, i.e., a software
translation tool, and the strongest contender to Alchemy Catalyst, whose
new version I wrote about in a recent newsletter. (Of course, you can see both Passolo
and Catalyst going "head to head" in the new addition to TranslatorsTraining -- see
above.)
The other day I talked to Florian Sachse, Passolo's
head developer, about the new features in the new version, and there are a few
that I thought were really nice.
First of all, there is a new look and feel to Passolo.
Not that that's really important, but it's not insignificant. Passolo
always tended to look a tad bit geeky and that has changed now. The panes have
a fresher look and are all freely dockable, which means that you can put them
anywhere on your screen(s) you like. It has a bit of the feel of the new Trados
Studio -- with which it will be completely compatible once that is
officially released. There are also some things that make it a little less
right-click-intensive: for instance, you can now just drag and drop files into Passolo
to start the processing.
And that's it with the new features!
Just kidding.
There are some more heavyweight features as well.
If you have followed the reviews of TEnTs and
localization tools in the last few newsletters, you will have realized that the
overarching theme for tools this year is subsegment search. The new and
upcoming version of Trados will do it, Transit does it, MemoQ
does it, Catalyst does it, and the interesting thing is that they all do
it differently. Well, Passolo did not want to be left out. It also does
it, it also does it differently and it also does it well.
The new subsegment search is made possible through a
major new indexing feature called QuickIndex. If you have worked with Passolo
before you will have noticed major slowdowns when it came to searches through
very large Passolo glossaries. With QuickIndex the searches are
very fast, almost instantaneous, in fact -- even with double-digit MBs worth of
glossaries. But not only does it give you back matches for the complete string
as it did before, it now also looks at subsegments on the source level for
matches in the glossaries. If it finds them, it marks them similarly to the way
that MultiTerm matches used to be (and are still) marked in Passolo
and Trados. This is especially helpful, of course, when it comes to
software glossaries, where it is the rule rather than the exception that short
commands, menu names, and other controls are used in longer phrases over and
over again.
Another new feature is the "Translation
History" feature. This is essentially a project-internal database that
stores all modifications to any string over any length of time. These
modifications could be linguistic but they can also be functional, such as
sizing coordinates. It is possible to view the history of any string on the fly
and with a single mouse-click revert to any item in the history. Also recorded is
information like date and user name. Since at some point the history could get
too long and memory-intensive, it is possible to delete it and start from
scratch -- for instance, after the localization of a finalized version and before
the start of the next version with a blank history. Nice.
Another translation feature that I liked was the
"Translation Helpers" (accessible in the Options dialog). This
allows you to set different sources for different processes. For instance, this
means that you could say that you would like to use a particular Trados
TM for pretranslation, another TM for fuzzy matches, a particular glossary for concordance
searches, and a MultiTerm termbase for terminology. Or you can have one
or several sources for all.
And, ta-da, Passolo now has assignable keyboard
shortcuts. No need to say why this is helpful and why I like it.
And for the software developers: There is now support of
.NET 3.5 and WPF (Silverlight) files.
Here are a few things I was disappointed about: A
number of TEnTs now offer the translation of Java .properties files with
an automatic tagging of HTML code. Passolo still does not do the
automatic tagging of HTML code in .properties files -- and I don't quite
understand why. I also wish that the subsegmenting feature described above had
been carried just a little bit further with the ability to automatically
assemble the translation of the detected subsegments -- but I can imagine that
this will be implemented in a later version.
|
| ADVERTISEMENT |
Confused about translation
technology?
TranslatorsTraining provides clarity. Save time
and money by checking out our wide selection of comparative CAT tool video
tutorials.
P.S. Sign up now and receive a free one-year
subscription to the Premium Edition of the Tool Kit
newsletter. |
3.
Excel, Once and for All
| |
This is the last time that I will say anything about Excel!
Shortly after I sent out the Erratum Edition concerning the Excel trick,
French colleague and Tool Kit reader Gilbert Liotard sent this message:
A word of caution for
people using multiple instances of Excel and performing copying/pasting
operations: there is a 256-character limit from one cell to the other one between
sheets running in separate instances of Excel (this limitation does not
exist between sheets within the same instance of the application). I would
clearly let everyone know of this limitation before they wonder why they lost
some data in the shuffle.
First I tried to just ignore his email. No way, I thought. Then I
quickly Googled this issue but didn't come up with anything. So I just tried to
forget it. That worked for about a day. So I tried it out and -- ouch --
discovered that Gilbert was right, at least for Excel 2003 and below.
So, if you have Excel files open in more than one session and you try to
copy complete cells containing more than 256 characters from one Excel
file to the other, they will be truncated (many of us painfully remember that this
was a limitation per se with Excel 95). Now, if you instead copy the actual
content of individual cells in Excel's Formula Bar or from within
a cell, you won't have problems pasting that into another Excel file,
but that's not how we usually do it, I guess.
So, if you have used that little trick that I mentioned two weeks ago
and you need to copy and paste between Excel files, open each of the
files with the File> New command from within Excel. That way
they will run in the same instance and you won't have those problems.
And of course, if you have Excel 2007, you won't have those
issues anyway, but then you also did not need the silly trick in the first
place.
Oh, and I asked Gilbert for some kind of
"official" example -- he did not have any, "just painful
experiences." Now we know that his pain was for a good cause!
|
4.
Three Cool Things from this Past Week
| |
1. It's probably not appropriate
to find this interactive NY Times mood barometer "cool,"
but it's remarkably insightful and a telling sign of our troubled economic
times.
2. Pipl is a very powerful search engine to
quickly find people who, like me, don't like "social networks." It
uses the very ominously named "deep web," but I don't care what it is
-- it sure offers results. Some people might think it's scary how much
information is out there about you in such a readily available fashion. I choose
to find it cool
3. And here is a cool geeky piece of information
that I liked: For those who still like the command prompt (essentially the DOS
window), there was (and still is) a little utility for Windows XP that
allows you to right-click on any Windows Explorer folder and select Open
Command Window Here so that you don't have to navigate there by typing your
heart out. Now, in Windows Vista, you don't need that little friend.
Instead, you can simply start a Command Prompt window with a desired
folder preselected by holding down the Shift
key, right-click the desired folder, and select Start Command Window Here.
|
| ADVERTISEMENT |
Change the way you translate with MemoQ 3.5
Kilgray Translation Technologies is proud to announce the release of MemoQ 3.5. This upgrade includes automatic subsegment suggestion, full STAR Transit-compatibility, enhanced support for Arabic & Hebrew, a horizontal translation editor and many more features. To learn more, visit www.kilgray.com.
|
| 5. Developers Are not
Fuzzy, Part I | |
This was the plan: For weeks now I have been planning to write
about a Russian developer-written TEnT that has some . . . let's say . . . intricacies that no other tool
has. But first the website was down for a while, then there were other problems
with the download, and now that I finally have my hands on the tool there is still
one bug preventing me from running it properly. I've sent out a call for help,
but since this newsletter is late as it is, let's talk about the tool next
time.
As it happens, though, there
is another tool that I've just come across (thanks to Piotr Graff), so we can
have a quick look at that:
Another tool written by
a developer is the Resource Translation Toolkit, developed
by no one less than Mike Funduc, whom many of you know as the brains behind the
well-liked search utility Search & Replace (which, by
the way, is completely Unicode-enabled now, including the ability to search in
Asian languages!).
For his own software localization efforts he took a quick glance at the
existing tools landscape and decided to try to develop a tool on his own. The
tool that he developed is interesting in that it is different from anything I
have seen so far in the translation memory arena. It shows how a developer
rather than a translator might approach translation.
As of now it translates only two software resource file types -- PO and
RC files -- and uses TMX or PO/POT files as translation memories. But instead
of working on the translatable file, the translator actually keeps on translating
the translation memory, to which new strings can be added by simply importing the
RC or PO file to it. In this process all non-language-related information,
including hotkey characters (letters that are preceded by & and result in
an underline in the finalized software, which in turn shows the Alt+[letter]
combination that will serve as a hotkey), are stripped so the translator does
not have to worry about it.
Once all the necessary strings in the translation memory are translated,
they are applied to the translatable file and a new, translated version of the RC
or PO file is automatically translated.
I know this sounds really complicated, but it
really isn't once you see it work. It's actually rather elegant. The drawback of this tool -- aside from the few
supported file formats -- is that there is no fuzzy matching. In addition, the concept
of stripping and automatically adding back the hotkey ampersands only works
when the same letter occurs in the translated string as well; otherwise it is
simply dropped. So there needs to be a bit of QA once the translation is done.
I could imagine that this problem will be fixed quickly, and I think then it
will be an interesting tool for companies (like Mike Funduc's) with a limited
number of software products to translate and volunteer translators without
access to full-fledged TEnTs or localization tools. And, according to the
master himself, there most likely will also be more software file formats (such
as Java .properties and .resx files) supported in future versions.
|
The Last Word on the Tool Kit
|
|
If you would like to promote this newsletter by placing a link on your website, I will in turn mention your website in a future edition of the Tool Kit. Just paste the code you find here into the HTML code of your webpage, and the little icon that is displayed on that page with a link to my website will be displayed. © 2009 International Writers' Group | |