ToolkitSmall

A computer newsletter for translation professionals

Issue 9-4-137
(the one hundred thirty-seventh edition)
Contents
1. Rememo Me? (Premium Content)
2. Passolo's New Look
3. Excel, Once and for All
4. Three Cool Things from this Past Week
5. Developers Are not Fuzzy, Part I
The Last Word on the Tool Kit
Hooray!
We've been talking about it for a while, and now we have finally released it: TranslatorsTraining now offers not only videos of more than a dozen TEnTs (translation environment tools), allowing you to learn about each and compare them on a one-to-one basis, but now also the same for Passolo and Catalyst, the two major localization tools. As for the TEnTs, we again prepared files, wrote a script, and asked the makers of Passolo and Catalyst to prepare detailed videos for us. We then edited and narrated these (this time it's the slightly robotized but still nightingale-like voice of my wife). If you ever wanted to know exactly what localization tools actually do, or why there have to be localization tools in addition to TEnTs, or if you ever wanted to compare the two big guys in that branch of our industry on an even playing field, here is your chance. And the price for a one-year subscription is still ridiculously cheap: 34.99 Euro.
1. Rememo Me? (Premium Content)

For the last few months I have been keeping tabs on MemoQ's latest releases, which have come out with refreshing regularity. Every so often there was enough interesting material to write about, but I always held out. Until now.

This last week MemoQ 3.5 was released, and here are the new features according to Kilgray's (the Hungarian company behind MemoQ):

·         Longest substring concordance

·         Wildcard concordance and wildcard in terms

·         STAR Transit filter

·         Bi-directional language enhancements

·         Horizontal edit view

·         XML preview feature

·         PowerPoint 2007 filter

·         Drastic server speed improvements

I downloaded the new version and specifically looked at the first three features, which I found truly ground-breaking.

Let's start with the first feature first, the oddly named "longest substring concordance" (or, just as odd: LSC). This is an attempt to automate concordance searches according to user-defined parameters (under Tools> Options> Subsegment leverage you can set the minimum number of concordance hits and/or the minimum number in words/letters/characters) and without interrupting the workflow. In the Translation results pane there are now not only translation memory matches, terminology matches, and assembled rows (available since version 3.2) but also ominous-looking matches of subsets of the string that needs to be translated with empty targets. These are the LSC matches. Clicking on them will produce the Concordance dialog, which will list all (or a predefined number of) the appearances of that particular substring in the translation memory within the context of the complete strings in the TM.

Confused?

Imagine you have this (real-life) sentence:

The starter motor rotates the engine during the start sequence, by driving it through the reduction gear unit assembly.

MemoQ might find that "the reduction gear unit" is a worthwhile subsegment and will display this in the Translation results pane with an empty target. Double-clicking on it would bring up the Concordance dialog with these options:

The starter motor rotates the engine during the start sequence, by driving it through the reduction gear unit assembly.

A coupling assembly mechanically connects the main output drive shaft of the reduction gear unit to the driven unit.

Individual accessory drive pads for the main lube oil pump are also incorporated on the reduction gear unit.

with their respective translations. If you wanted to use any of those, you would just need to highlight the part of the translation you want to use and select Insert selected.

I really like this feature because it offers a new granularity to TM materials without being obtrusive (the program does not slow you down by displaying an additional dialog like a comparable feature in Trados, plus you are free to use the displayed option or not), it enables easy paste access to the desired translation, and MemoQ's superior search-and-lookup speed allows it to operate without even seeming to slow down the search process too much.

Speaking of the Concordance dialog, that same dialog is also used for the enhanced concordance features. (A "concordance search" is the process of manually highlighting one term or phrase in the source segment, pressing a shortcut key -- in the case of MemoQ, it's Ctrl+K -- and accessing all occurrences of that in the TM.) What makes MemoQ's concordance feature attractive is the possible automatic addition of a wildcard character. In the Concordance dialog you can select the option Add wildcard to selected text, and for the current and next search(es) MemoQ automatically adds an asterisk (*) to each of the terms. This will make MemoQ look for 0 or more additional characters to the term in question, something that is particularly helpful for languages with heavy flexion. (Plus, I really like this because I always forget where the asterisk is on the German keyboard and I get tired of switching back to the English keyboard to enter it manually.)

The last feature that really caught my interest was the Transit compatibility feature. Now, Transit is a great program, but it's really different, and many folks just don't want to spend the time to learn to use the free Satellite edition (even if they might miss out on something). For those, this feature will be very useful. It allows you to import a PXF file -- this is a Transit-specific package file with the translation files, reference material (i.e., TM), and glossary data -- extract the translation files as well as the TM content, translate it, and then send it back to your client as a TXF file -- Transit's return package format.

In general it works very well. There are a few glitches -- a couple of strings (out of a few thousand in a test run) were unduly protected as tags, and I had to switch my import mode to get all my data, but it's impressive that the MemoQ developers were able to use Transit logic to "harvest" reference material that is readily available in the TM you selected when you started translating. What this feature is not able to do is automatically import the TermStar glossaries. This is somewhat unfortunate because Transit projects tend to be very terminology-heavy because of its excellent terminology tool.

Here are some other caveats with the new version: Word 2007 is still not supported -- though PowerPoint and Excel 2007 are -- and TBX, the termbase exchange standard, is also not supported yet.

This morning I had a chance to talk with the owner of a translation agency who has been using the server edition of MemoQ for a while, just to get an idea of what his take on performance and user acceptance was. He was very positive overall. Compared to other server-based products (Idiom WorldServer, Logoport, and Across) he reported equal or better server response times. He also liked the option for translators to check out resources to work offline or the online document storage. This last feature allows translators not only to share translation memories and terminology databases, but also the actual documents, which -- optionally -- can be server-based as well. Thus, multiple users can have access to large documents that can be translated and edited at the same time for faster turnaround. His assessment of that process was kind of interesting: He ran into problems with translators not working successively from top to bottom through documents, creating havoc for the poor editors; however, that seems to be less a technical limitation than an organizational one.

As far as user acceptance, he acknowledged that not all his translators were super-eager to adopt a new tool at the drop of a hat, but it helped that they did not have to pay for the program. With MemoQ's mobile licensing concept, he is able to assign temporary full licenses to his users. Interestingly -- and he was not the first to mention it -- Déjà Vu users in particular are struggling with the idea of using a different tool.

Speaking of licensing, I am interested to see how MemoQ's licensing scheme will be adopted once it's time for it. Last fall Kilgray adopted a new system in which all upgrades are free for a year after purchase, no matter how major or minor they might be; after that year, a 20% annual fee is applicable for further upgrades. This is certainly not an uncommon practice in the software industry in general, but as far as I know it is unusual for the individual user sector in our industry. We'll see what happens come fall of 2009.

2. Passolo's New Look

At the beginning of this week, SDL released a new Passolo website in SDL look (I'm very sorry to see the old Passolo website go away!) and, more importantly, a new version of Passolo. Most of you know that Passolo is a localization tool, i.e., a software translation tool, and the strongest contender to Alchemy Catalyst, whose new version I wrote about in a recent newsletter. (Of course, you can see both Passolo and Catalyst going "head to head" in the new addition to TranslatorsTraining -- see above.)

The other day I talked to Florian Sachse, Passolo's head developer, about the new features in the new version, and there are a few that I thought were really nice.

First of all, there is a new look and feel to Passolo. Not that that's really important, but it's not insignificant. Passolo always tended to look a tad bit geeky and that has changed now. The panes have a fresher look and are all freely dockable, which means that you can put them anywhere on your screen(s) you like. It has a bit of the feel of the new Trados Studio -- with which it will be completely compatible once that is officially released. There are also some things that make it a little less right-click-intensive: for instance, you can now just drag and drop files into Passolo to start the processing.

And that's it with the new features!

Just kidding.

There are some more heavyweight features as well.

If you have followed the reviews of TEnTs and localization tools in the last few newsletters, you will have realized that the overarching theme for tools this year is subsegment search. The new and upcoming version of Trados will do it, Transit does it, MemoQ does it, Catalyst does it, and the interesting thing is that they all do it differently. Well, Passolo did not want to be left out. It also does it, it also does it differently and it also does it well.

The new subsegment search is made possible through a major new indexing feature called QuickIndex. If you have worked with Passolo before you will have noticed major slowdowns when it came to searches through very large Passolo glossaries. With QuickIndex the searches are very fast, almost instantaneous, in fact -- even with double-digit MBs worth of glossaries. But not only does it give you back matches for the complete string as it did before, it now also looks at subsegments on the source level for matches in the glossaries. If it finds them, it marks them similarly to the way that MultiTerm matches used to be (and are still) marked in Passolo and Trados. This is especially helpful, of course, when it comes to software glossaries, where it is the rule rather than the exception that short commands, menu names, and other controls are used in longer phrases over and over again.

Another new feature is the "Translation History" feature. This is essentially a project-internal database that stores all modifications to any string over any length of time. These modifications could be linguistic but they can also be functional, such as sizing coordinates. It is possible to view the history of any string on the fly and with a single mouse-click revert to any item in the history. Also recorded is information like date and user name. Since at some point the history could get too long and memory-intensive, it is possible to delete it and start from scratch -- for instance, after the localization of a finalized version and before the start of the next version with a blank history. Nice.

Another translation feature that I liked was the "Translation Helpers" (accessible in the Options dialog). This allows you to set different sources for different processes. For instance, this means that you could say that you would like to use a particular Trados TM for pretranslation, another TM for fuzzy matches, a particular glossary for concordance searches, and a MultiTerm termbase for terminology. Or you can have one or several sources for all.

And, ta-da, Passolo now has assignable keyboard shortcuts. No need to say why this is helpful and why I like it.

And for the software developers: There is now support of .NET 3.5 and WPF (Silverlight) files.

Here are a few things I was disappointed about: A number of TEnTs now offer the translation of Java .properties files with an automatic tagging of HTML code. Passolo still does not do the automatic tagging of HTML code in .properties files -- and I don't quite understand why. I also wish that the subsegmenting feature described above had been carried just a little bit further with the ability to automatically assemble the translation of the detected subsegments -- but I can imagine that this will be implemented in a later version.

ADVERTISEMENT

 Confused about translation technology?

TranslatorsTraining provides clarity. Save time and money by checking out our wide selection of comparative CAT tool video tutorials.

P.S. Sign up now and receive a free one-year subscription to the Premium Edition of the Tool Kit newsletter.

3. Excel, Once and for All

This is the last time that I will say anything about Excel! Shortly after I sent out the Erratum Edition concerning the Excel trick, French colleague and Tool Kit reader Gilbert Liotard sent this message:

A word of caution for people using multiple instances of Excel and performing copying/pasting operations: there is a 256-character limit from one cell to the other one between sheets running in separate instances of Excel (this limitation does not exist between sheets within the same instance of the application). I would clearly let everyone know of this limitation before they wonder why they lost some data in the shuffle.

First I tried to just ignore his email. No way, I thought. Then I quickly Googled this issue but didn't come up with anything. So I just tried to forget it. That worked for about a day. So I tried it out and -- ouch -- discovered that Gilbert was right, at least for Excel 2003 and below. So, if you have Excel files open in more than one session and you try to copy complete cells containing more than 256 characters from one Excel file to the other, they will be truncated (many of us painfully remember that this was a limitation per se with Excel 95). Now, if you instead copy the actual content of individual cells in Excel's Formula Bar or from within a cell, you won't have problems pasting that into another Excel file, but that's not how we usually do it, I guess.

So, if you have used that little trick that I mentioned two weeks ago and you need to copy and paste between Excel files, open each of the files with the File> New command from within Excel. That way they will run in the same instance and you won't have those problems.

And of course, if you have Excel 2007, you won't have those issues anyway, but then you also did not need the silly trick in the first place.

Oh, and I asked Gilbert for some kind of "official" example -- he did not have any, "just painful experiences." Now we know that his pain was for a good cause!
4. Three Cool Things from this Past Week

1.  It's probably not appropriate to find this interactive NY Times mood barometer "cool," but it's remarkably insightful and a telling sign of our troubled economic times.

2.  Pipl is a very powerful search engine to quickly find people who, like me, don't like "social networks." It uses the very ominously named "deep web," but I don't care what it is -- it sure offers results. Some people might think it's scary how much information is out there about you in such a readily available fashion. I choose to find it cool

3. And here is a cool geeky piece of information that I liked: For those who still like the command prompt (essentially the DOS window), there was (and still is) a little utility for Windows XP that allows you to right-click on any Windows Explorer folder and select Open Command Window Here so that you don't have to navigate there by typing your heart out. Now, in Windows Vista, you don't need that little friend. Instead, you can simply start a Command Prompt window with a desired folder preselected by holding down the Shift key, right-click the desired folder, and select Start Command Window Here.
ADVERTISEMENT
Change the way you translate with MemoQ 3.5
 
Kilgray Translation Technologies is proud to announce the release of MemoQ 3.5. This upgrade includes automatic subsegment suggestion, full STAR Transit-compatibility, enhanced support for Arabic & Hebrew, a horizontal translation editor and many more features. To learn more, visit www.kilgray.com.
5. Developers Are not Fuzzy, Part I

This was the plan: For weeks now I have been planning to write about a Russian developer-written TEnT that has some . . .  let's say . . . intricacies that no other tool has. But first the website was down for a while, then there were other problems with the download, and now that I finally have my hands on the tool there is still one bug preventing me from running it properly. I've sent out a call for help, but since this newsletter is late as it is, let's talk about the tool next time.

As it happens, though, there is another tool that I've just come across (thanks to Piotr Graff), so we can have a quick look at that:

Another tool written by a developer is the Resource Translation Toolkit, developed by no one less than Mike Funduc, whom many of you know as the brains behind the well-liked search utility Search & Replace (which, by the way, is completely Unicode-enabled now, including the ability to search in Asian languages!).

For his own software localization efforts he took a quick glance at the existing tools landscape and decided to try to develop a tool on his own. The tool that he developed is interesting in that it is different from anything I have seen so far in the translation memory arena. It shows how a developer rather than a translator might approach translation.

As of now it translates only two software resource file types -- PO and RC files -- and uses TMX or PO/POT files as translation memories. But instead of working on the translatable file, the translator actually keeps on translating the translation memory, to which new strings can be added by simply importing the RC or PO file to it. In this process all non-language-related information, including hotkey characters (letters that are preceded by & and result in an underline in the finalized software, which in turn shows the Alt+[letter] combination that will serve as a hotkey), are stripped so the translator does not have to worry about it.

Once all the necessary strings in the translation memory are translated, they are applied to the translatable file and a new, translated version of the RC or PO file is automatically translated.

I know this sounds really complicated, but it really isn't once you see it work. It's actually rather elegant. The drawback of this tool -- aside from the few supported file formats -- is that there is no fuzzy matching. In addition, the concept of stripping and automatically adding back the hotkey ampersands only works when the same letter occurs in the translated string as well; otherwise it is simply dropped. So there needs to be a bit of QA once the translation is done. I could imagine that this problem will be fixed quickly, and I think then it will be an interesting tool for companies (like Mike Funduc's) with a limited number of software products to translate and volunteer translators without access to full-fledged TEnTs or localization tools. And, according to the master himself, there most likely will also be more software file formats (such as Java .properties and .resx files) supported in future versions.
The Last Word on the Tool Kit

If you would like to promote this newsletter by placing a link on your website, I will in turn mention your website in a future edition of the Tool Kit. Just paste the code you find here into the HTML code of your webpage, and the little icon that is displayed on that page with a link to my website will be displayed.

© 2009 International Writers' Group