Jskad

Author	SHA1	Message	Date
dchandler	689c1910aa	To deal with java.swing.text.rtf bugs regarding hexadecimal escape sequences, I've created RTFFixerInputStream. It turns illegal hexadecimal escapes into Unicode escapes.	2003-06-29 02:30:08 +00:00
dchandler	0b849aed97	Fixed comments w.r.t. javadoc warnings.	2003-06-29 02:22:20 +00:00
dchandler	4e279defb4	Fixed a couple of array bounds checks. Added support for two more oddballs. Deprecated the oddball lookup method because it drops up to 30 glyphs in TibetanMachine. The correct solution is to transform the RTF before Java's busted RTF readers ever see it. \'97 becomes \u151, e.g.	2003-06-28 16:33:58 +00:00
dchandler	2a359c45ef	Bad conversions were not leaving the unconvertable characters at the beginning of the document as they should and as they are documented to. They now do, and they bracket the bad characters with the TM or TMW for U+0F3C on the left and the TM or TMW for U+0F3D on the right. Some cleanup.	2003-06-28 16:20:19 +00:00
dchandler	c39d8d6326	My earlier code cleanup introduced this bug; TMW->TM conversion was busted.	2003-06-26 22:48:51 +00:00
dchandler	25510542b2	Now with a nicer error message in one case.	2003-06-26 22:48:05 +00:00
dchandler	c34259b105	Code cleanup.	2003-06-25 01:04:24 +00:00
dchandler	9e6c3009ac	Added an About button. Code cleanup. Changed the Cancel button to the Close button.	2003-06-25 00:49:11 +00:00
dchandler	569fba6467	Made the comments in the my_thdl_preferences.txt file use standard line separators.	2003-06-25 00:03:46 +00:00
dchandler	0f3c4174b6	Made the comments in the my_thdl_preferences.txt file more useful.	2003-06-24 23:48:00 +00:00
dchandler	c67ddb2d6c	Use Ximalaya, not Arial Unicode MS, by default.	2003-06-24 12:51:32 +00:00
dchandler	33beb7b782	Bye bye debugging output.	2003-06-24 12:23:37 +00:00
dchandler	f547734043	Added Than's converter GUI code; adapted it to work with Jskad's converters. TMW->Unicode now uses Ximalaya by default.	2003-06-24 03:02:29 +00:00
dchandler	19d7cabfe6	Forget the final=faster myth.	2003-06-24 03:01:13 +00:00
dchandler	917864574c	Fixed a logic bug in mapTMWtoTM and mapTMtoTMW. You can now specify which Unicode font to use via 'java -Dthdl.tmw.to.unicode.font=Ximalaya ...'.	2003-06-23 01:58:11 +00:00
dchandler	b6d8fd89f9	When errors in (all but TMW->Wylie and Wylie->TMW) conversion occur, the troublesome glyphs are now put at the beginning of the document AFTER AN ACHEN. This makes a glyph like \tmw7095 visible atop the achen. Major fix to the handling of paragraphs in conversion; we were (for whatever reason) dropping paragraphs before.	2003-06-23 01:24:02 +00:00
dchandler	0f77b32add	It's not in use, but there's code to make Jskad a TM carrier just as it is a TMW carrier.	2003-06-22 22:11:40 +00:00
dchandler	1f4343bed0	TMW->TM, TM->TMW, and TMW->Unicode conversions are all (at least 2) orders of magnitude faster.	2003-06-22 22:10:58 +00:00
dchandler	afe73c2228	The pseudo-file '-', referring to standard input, is now accepted as a command-line argument.	2003-06-22 21:05:16 +00:00
dchandler	900f7492b0	'ant clean check' was failing because I hadn't updated the --find-some-non-tmw and --find-all-non-tmw baselines. Code cleanup.	2003-06-22 16:11:58 +00:00
dchandler	66287f3cc9	Small TMW->Wylie performance improvements. TMW->Wylie is much faster than TMW->Unicode etc.; this is because many fewer replacements are made (i.e., more text is replaced each time a replacement is performed). I must find a way to still preserve formatting but do many fewer replacements in TMW->{Unicode,TM} and TM->TMW.	2003-06-22 04:32:59 +00:00
dchandler	6540b260bd	Fixes a (small, I think) TMW->Unicode performance glitch. I was inserting 5 characters at a time and then skipping ahead just one position. I don't think this affected correctness. I believe there's still a terrible (exponential?) slowdown as the input file gets bigger, however. Perhaps not -- but we run through the first 1000 TMW glyphs in 6 seconds, the 20th thousand takes at least 60 seconds. Is TMW->Wylie faster than TMW->Unicode? If so, why? Thought: don't use a DuffPane within TibetanConverter -- it can only add overhead, right? My hprof profile said that the conversion was taking just a couple of percent of the work; the rest was going to display-related stuff that you should only see if you were displaying the document. I'm not!	2003-06-22 04:08:33 +00:00
dchandler	dfe64a1927	Added --find-some-non-tm and --find-all-non-tm modes to the converter to help ensure worry-free TM->TMW conversions.	2003-06-22 00:14:18 +00:00
dchandler	80101666c7	Included a fix from WylieWord's tibwn.ini. Removed some needless trailing tildes.	2003-06-21 02:35:21 +00:00
dchandler	9a41f512d9	It used to be the case that you could select 'Close', and then when asked "do you want to save?" you could press yes and then press cancel and Jskad would still exit. That's no longer the case. Added File->Exit to Jskad.	2003-06-21 02:07:51 +00:00
dchandler	45b87b0fb4	In Jskad, you can now clear the preferences and return to default values.	2003-06-21 01:26:17 +00:00
eg3p	fbb6245fdb	Added cut() and copy() methods to override JTextPane's methods of same name.	2003-06-20 15:27:20 +00:00
dchandler	149eecc4fb	Added a thdl-nightly-build target for nightly builds on iris.lib.virginia.edu. This will need to be edited to use the real path on iris.	2003-06-19 01:58:59 +00:00
dchandler	1ae7948fed	Nightly builds are now done using pserver CVS updates. The 'dc-nightly-build' target no longer does a cvs update, because that's not always going to work.	2003-06-19 01:11:33 +00:00
dchandler	5067683121	Edward corrected me; he had intended to have M map to 7.91, not 7.90.	2003-06-17 01:46:19 +00:00
dchandler	6712b47e13	Added an option to control the Unicode font for TMW->Unicode conversions.	2003-06-15 20:28:56 +00:00
dchandler	ced830a7d3	Renamed TMW_RTF_TO_THDL_WYLIE TibetanConverter.	2003-06-15 19:19:23 +00:00
dchandler	34a7b5da9b	This converter now performs TMW->Unicode conversions.	2003-06-15 18:38:42 +00:00
dchandler	da70434e52	Jskad now allows for TMW->Unicode conversion.	2003-06-15 16:27:36 +00:00
dchandler	af5b95b08d	A TMW->Unicode table is here. Note these issues, however: Is the EWTS '_' to be represented as U+0020, or is it a wider space? Does TMW9.42, Dza, map to U+0F5F,U+0F39? Does TMW6.60, r+y, map to U+0F62,U+0FBB or to U+0F6A,U+0FBB? (Likewise with r+w, TMW6.61, TMW6.62, etc.) Is U+0F7E a bindu? What Unicode does TMW7.96 map to, for example? What does TMW7.91 map to? Should TMW8.97 and TMW8.98 map to swastiskas elsewhere in Unicode? If so, which codepoints? Likewise with TMW9.60, a Chinese character. Does TMW7.68 map to U+0F39? Does TMW7.74, the ITHI secret sign, have a Unicode mapping? f68,fa0,f80,f72 comes close, but fa0 would be too large, wouldn't it? What Unicode does TMW9.61 map to? Is it for sequences like f40,f7c,f60,f72? Or is it for f60,f72,f7c?	2003-06-15 03:25:45 +00:00
dchandler	b387c512e9	Fixed two bugs.	2003-06-15 03:08:57 +00:00
dchandler	189fef9aec	Made Jskad smart enough to handle a few more EWTS characters; some it can only convert to Wylie, others are live key sequences. This will make converting the shechen documents go more smoothly.	2003-06-09 13:35:43 +00:00
dchandler	09a55110b7	Handles more TibetanMachine oddballs.	2003-06-09 02:01:13 +00:00
dchandler	b9219640e5	Handles more TibetanMachine oddballs.	2003-06-09 01:53:01 +00:00
dchandler	e97e1c8464	Handles more TibetanMachine oddballs.	2003-06-09 01:20:32 +00:00
dchandler	651a599188	Fixed usage info.	2003-06-08 23:23:12 +00:00
dchandler	70b31558fa	Tried to fix a crashing bug that happened when you converted TM->TMW and then tried to convert that TMW to Wylie. I swear it's Java's problem (see the ugly stack trace in the code and decide for yourself), and I tried replacing rather than inserting-and-then-removing, but it didn't work. I've left these things as options.	2003-06-08 23:12:52 +00:00
dchandler	212414edef	TMW_RTF_TO_THDL_WYLIE now converts TM->TMW.	2003-06-08 22:43:27 +00:00
dchandler	32831b698f	If bad (oddball) TM glyphs appear, then converting to TMW causes, by default, all oddballs to appear once in the resulting document. This'll help me find the correct glyphs for the oddballs, and it'll prevent the average user from converting a document with oddballs.	2003-06-08 22:37:38 +00:00
dchandler	d45f5ab8c8	Improved performance (I suppose).	2003-06-03 23:49:34 +00:00
dchandler	7d768c9e06	Fixed a crashing bug that happened upon converting wylie to tibetan.	2003-06-03 23:45:15 +00:00
dchandler	0f724989b5	The Wylie 'M' used to map to TMW7.91, when it should map to TMW7.90. I've fixed that. I've also added a couple of Unicode mappings to give a flavor for how multi-codepoint mappings will be represented. TM->TMW conversion takes about 1 second per thousand glyphs on my PIII-550.	2003-06-01 23:05:32 +00:00
dchandler	54ca37c824	The Wylie 'M' used to map to TMW7.91, when it should map to TMW7.90. I've fixed that. I've also added a couple of Unicode mappings to give a flavor for how multi-codepoint mappings will be represented.	2003-06-01 19:14:08 +00:00
dchandler	e2caf99085	Some code cleanup. tibwn.ini must now have, in the Unicode column, either nothing, or 0FXX(,0FXX)*. E.g., 0F04,0F05 is valid. Debugging code ensures this is the case.	2003-06-01 18:09:49 +00:00
dchandler	1f6bb07d53	Fixes bogus Unicode mappings mentioned in http://sourceforge.net/tracker/index.php?func=detail&aid=746871&group_id=61934&atid=502515.	2003-06-01 04:02:04 +00:00

1 2 3 4 5 ...

328 commits