summaryrefslogtreecommitdiff
path: root/src/backend/utils/mb
Commit message (Collapse)AuthorAgeFilesLines
* pgindent run on all C files. Java run to follow. initdb/regressionBruce Momjian2001-10-256-305/+561
| | | | tests pass.
* Ok, here is the modified encoding table (column1 is the standard name,Tatsuo Ishii2001-10-162-30/+33
| | | | | | | | | | | | | | | | | | | | | | | 2 is our "official" name, and 3 is alias). If there's no objection, I will change them. ASCII SQL_ASCII UTF-8 UNICODE UTF_8 MULE-INTERNAL MULE_INTERNAL ISO-8859-1 LATIN1 ISO_8859_1 ISO-8859-2 LATIN2 ISO_8859_2 ISO-8859-3 LATIN3 ISO_8859_3 ISO-8859-4 LATIN4 ISO_8859_4 ISO-8859-5 ISO_8859_5 ISO-8859-6 ISO_8859_6 ISO-8859-7 ISO_8859_7 ISO-8859-8 ISO_8859_8 ISO-8859-9 LATIN5 ISO_8859_9 ISO-8859-10 LATIN6 ISO_8859_10 ISO-8859-13 LATIN7 ISO_8859_13 ISO-8859-14 LATIN8 ISO_8859_14 ISO-8859-15 LATIN9 ISO_8859_15 ISO-8859-16 LATIN10 ISO_8859_16
* Add UTF-8 char >= 0x10000 checkTatsuo Ishii2001-10-151-2/+10
|
* Add a new function "pg_client_encoding" which returns the current clientTatsuo Ishii2001-10-121-1/+8
| | | | | | side encoding name. This is necessary for client API's such as JDBC to perform correct encoding conversions. See my email "[HACKERS] pg_client_encoding" 10 Sep 2001.
* Add support for ISO-8859-6 to 16Tatsuo Ishii2001-10-1122-21/+2284
|
* Fix bug in mic2ascii(). It does not handle correctly if none ASCIITatsuo Ishii2001-09-251-2/+2
| | | | chars are in the input.
* Add pg_database_encoding_max_length() function.Tatsuo Ishii2001-09-231-1/+11
|
* Remove test driversTatsuo Ishii2001-09-226-405/+3
| | | | Also fix comment in conv.c.
* Fix type_maximum_size() to give the right answer in MULTIBYTE cases.Tom Lane2001-09-212-20/+37
| | | | Avoid use of prototype-less function pointers in MB code.
* Remove old file.Peter Eisentraut2001-09-191-130/+0
|
* Implement following item in TODO:Tatsuo Ishii2001-09-112-55/+120
| | | | * Reject character sequences those are not valid in their charset
* Backout Karel's patchTatsuo Ishii2001-09-091-22/+6
|
* > > A simple and robus solution is in the begin of mbutils.c set defaultBruce Momjian2001-09-081-7/+11
| | | | | | | | | | > > ClientEncoding to SQL_ASCII (like default DatabaseEncoding). Bruce, can > > you change it? It's one line change. Again thanks. Forget it! A default client encoding must be set by actual database encoding... Please apply the small attached patch that solve it better. Karel Zak
* Remove file, per Karel.Bruce Momjian2001-09-071-130/+0
|
* Remove variable length macros used in debugging, per Karel.Bruce Momjian2001-09-071-28/+10
|
* Remove unused files for Karel's patch.Bruce Momjian2001-09-074-519/+0
|
* Remove common.c, removed in Karal's patch.Bruce Momjian2001-09-071-99/+0
|
* Add missing files.Tatsuo Ishii2001-09-078-0/+1117
|
* Commit Karel's patch.Tatsuo Ishii2001-09-066-208/+220
| | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | | ------------------------------------------------------------------- Subject: Re: [PATCHES] encoding names From: Karel Zak <zakkr@zf.jcu.cz> To: Peter Eisentraut <peter_e@gmx.net> Cc: pgsql-patches <pgsql-patches@postgresql.org> Date: Fri, 31 Aug 2001 17:24:38 +0200 On Thu, Aug 30, 2001 at 01:30:40AM +0200, Peter Eisentraut wrote: > > - convert encoding 'name' to 'id' > > I thought we decided not to add functions returning "new" names until we > know exactly what the new names should be, and pending schema Ok, the patch not to add functions. > better > > ...(): encoding name too long Fixed. I found new bug in command/variable.c in parse_client_encoding(), nobody probably never see this error: if (pg_set_client_encoding(encoding)) { elog(ERROR, "Conversion between %s and %s is not supported", value, GetDatabaseEncodingName()); } because pg_set_client_encoding() returns -1 for error and 0 as true. It's fixed too. IMHO it can be apply. Karel PS: * following files are renamed: src/utils/mb/Unicode/KOI8_to_utf8.map --> src/utils/mb/Unicode/koi8r_to_utf8.map src/utils/mb/Unicode/WIN_to_utf8.map --> src/utils/mb/Unicode/win1251_to_utf8.map src/utils/mb/Unicode/utf8_to_KOI8.map --> src/utils/mb/Unicode/utf8_to_koi8r.map src/utils/mb/Unicode/utf8_to_WIN.map --> src/utils/mb/Unicode/utf8_to_win1251.map * new file: src/utils/mb/encname.c * removed file: src/utils/mb/common.c -- Karel Zak <zakkr@zf.jcu.cz> http://home.zf.jcu.cz/~zakkr/ C, PostgreSQL, PHP, WWW, http://docs.linux.cz, http://mape.jcu.cz
* Add conver/convert2 functions. They are similar to the SQL99's convert.Tatsuo Ishii2001-08-151-75/+193
|
* TODO item:Tatsuo Ishii2001-07-151-5/+27
| | | | * Make n of CHAR(n)/VARCHAR(n) the number of letters, not bytes
* Fix a message error in utf_to_localTatsuo Ishii2001-05-281-2/+2
|
* BTW it does not add encodign it just patches existing one (KOI8) toBruce Momjian2001-05-032-16/+16
| | | | | | | support two - KOI8-R and KOI8-U (latter is superset of the former if not to take to the account pseudographics) Andy Rysin
* Add missing Unicode support for Cyrillic encodings.Tatsuo Ishii2001-04-299-7/+972
| | | | Patches contributed by Victor Wagner.
* Add a crash gurard to pg_encoding_mblen in case of an invalid encodingTatsuo Ishii2001-04-191-2/+2
| | | | given.
* Correction for mathematical properties in Unicode converison maps.Tatsuo Ishii2001-04-1618-34/+350
| | | | Patches contributed by Eiji Tokuya (e-tokuya@sankyo-unyu.co.jp)
* getdatabaseencoding() and PG_encoding_to_char() were being sloppy aboutTom Lane2001-04-162-4/+7
| | | | | converting char* strings to type 'name'. Imagine my surprise when 7.1 release coredumped upon start when compiled --enable-multibyte ...
* Fix unportable assumptions about alignment of local char[n] variables.Tom Lane2001-03-251-8/+5
|
* pgindent run. Make it all clean.Bruce Momjian2001-03-225-184/+202
|
* Modify wchar conversion routines to not fetch the next byte past the endTom Lane2001-03-082-35/+33
| | | | | | | | | | | | | of a counted input string. Marinos Yannikos' recent crash report turns out to be due to applying pg_ascii2wchar_with_len to a TEXT object that is smack up against the end of memory. This is the second just-barely- reproducible bug report I have seen that traces to some bit of code fetching one more byte than it is allowed to. Let's be more careful out there, boys and girls. While at it, I changed the code to not risk a similar crash when there is a truncated multibyte character at the end of an input string. The output in this case might not be the most reasonable output possible; if anyone wants to improve it further, step right up...
* Enhanced UTF-8/SJIS mapping generator, contributed byTatsuo Ishii2001-02-231-25/+38
| | | | Eiji Tokuya" <e-tokuya@Mail.Sankyo-Unyu.co.jp>
* Unicode <-> SJIS new mapping tables (based on CP932.TXT) contributed byTatsuo Ishii2001-02-152-9/+1309
| | | | Eiji Tokuya" <e-tokuya@Mail.Sankyo-Unyu.co.jp>
* Move pg_encoding_mblen() from common.c to wchar.c.Tatsuo Ishii2001-02-112-9/+11
|
* conv.c did not compile anymore. Fix wrong header file inclusion.Tatsuo Ishii2001-02-111-4/+2
|
* Restructure the key include files per recent pghackers discussion: thereTom Lane2001-02-1011-20/+31
| | | | | | | | | | | are now separate files "postgres.h" and "postgres_fe.h", which are meant to be the primary include files for backend .c files and frontend .c files respectively. By default, only include files meant for frontend use are installed into the installation include directory. There is a new make target 'make install-all-headers' that adds the whole content of the src/include tree to the installed fileset, for use by people who want to develop server-side code without keeping the complete source tree on hand. Cleaned up a whole lot of crufty and inconsistent header inclusions.
* Fix a bug in conversion from big5 to EUC_TW (CNS 11643-1992 Plane 3)Tatsuo Ishii2000-12-091-2/+2
| | | | Thanks Chih-Chang Hsieh <cch@cc.kmu.edu.tw> for finding the bug.
* Make all commands that link a program look likePeter Eisentraut2000-11-301-4/+4
| | | | | | | $(CC) $(CFLAGS) $(LDFLAGS) <object files> <extra-libraries> $(LIBS) -o $@ This form seemed to be the most portable, readable, and logical, but in any case it's better than having a dozen different ones in the tree.
* Unicode conversion fix suggested by Jan Varga...Tatsuo Ishii2000-11-269-11/+565
| | | | | | | | | | | | | | | | | | | | -------------------------------------------------- Subject: Bug in unicode conversion ... From: Jan Varga <varga@utcru.sk> To: t-ishii@sra.co.jp Date: Sat, 18 Nov 2000 17:41:20 +0100 (CET) Hi, I tried this new feature in PostgreSQL. I found one bug. Script UCS_to_8859.pl skips input lines which 1. code <0x80 or 2. ucs <0x100 I think second one is not good idea because some codes in ISO8859-2 have ucs <0x100 (e.g. 0xE9 - 0x00E9) --------------------------------------------------
* Fix bugs in EUC_TW support. This fix includes patches contributedTatsuo Ishii2000-11-171-4/+11
| | | | | by Chih-Chang Hsi. See "A Patch for MIC to EUC_TW code converting in mb support" posting in pgsql-patches list dated 09 Nov 2000.
* Extend CREATE DATABASE to allow selection of a template database to beTom Lane2000-11-141-15/+2
| | | | | | | | | | cloned, rather than always cloning template1. Modify initdb to generate two identical databases rather than one, template0 and template1. Connections to template0 are disallowed, so that it will always remain in its virgin as-initdb'd state. pg_dumpall now dumps databases with restore commands that say CREATE DATABASE foo WITH TEMPLATE = template0. This allows proper behavior when there is user-added data in template1. initdb forced!
* Add support for code conversion between Unicode and other encodings.Tatsuo Ishii2000-10-3038-26759/+141337
| | | | | | Supported encodings are: EUC_JP, EUC_CN, EUC_KR, EUC_TW, Shift JIS, Big5, ISO8859-[1-5]. TODO: testings! and documentations...
* Remove gcc-only macro definitionTatsuo Ishii2000-10-271-1/+13
|
* Support SET/SHOW/RESET client_encoding and server_encoding even whenTom Lane2000-10-253-88/+4
| | | | | | MULTIBYTE support is not compiled (you just can't set them to anything but SQL_ASCII). This should reduce interoperability problems between MB-enabled clients and non-MB-enabled servers.
* Add support for VPATH builds, that is, building somewhere else than in thePeter Eisentraut2000-10-201-37/+11
| | | | | | | | | source directory. This involves mostly makefiles using $(srcdir) when they might have used ".". (Regression tests don't work with this, yet.) Sort out usage of CPPFLAGS, CFLAGS (and CXXFLAGS). Add "override" keyword in most places, to preserve necessary flags even when the user overrode the flags.
* Support for conversion between UNICODE and other encodingsTatsuo Ishii2000-10-1214-438/+27865
| | | | | currently ISO8859-[1-5] and EUC_JP are supported. support for other encodings will be coming soon.
* Fix relative path references so that make knowns which dependencies referPeter Eisentraut2000-08-311-4/+4
| | | | | to one another. Sort out builddir vs srcdir variable namings. Remove some now obsoleted make variables.
* Change pg_mblen and pg_encoding_mblen return types from voidTatsuo Ishii2000-08-274-89/+64
| | | | to int so that they return the number of whcars.
* First phase of memory management rewrite (see backend/utils/mmgr/READMETom Lane2000-06-281-13/+16
| | | | | | | | | | | | | for details). It doesn't really do that much yet, since there are no short-term memory contexts in the executor, but the infrastructure is in place and long-term contexts are handled reasonably. A few long- standing bugs have been fixed, such as 'VACUUM; anything' in a single query string crashing. Also, out-of-memory is now considered a recoverable ERROR, not FATAL. Eliminate a large amount of crufty, now-dead code in and around memory management. Fix problem with holding off SIGTRAP, SIGSEGV, etc in postmaster and backend startup.
* Another batch of fmgr updates. I think I have gotten all old-styleTom Lane2000-06-132-5/+27
| | | | | functions that take pass-by-value datatypes. Should be ready for port testing ...
* Generated header files parse.h and fmgroids.h are now copied intoTom Lane2000-05-291-3/+1
| | | | | the src/include tree, so that -I backend is no longer necessary anywhere. Also, clean up some bit rot in contrib tree.