Parent Directory
|
Revision Log
|
Patch
| revision 1.9 by wakaba, Tue Oct 14 04:32:49 2008 UTC | revision 1.44 by wakaba, Sat Sep 5 11:31:58 2009 UTC | |
|---|---|---|
| # | Line 1 | Line 1 |
| 1 | 2009-09-05 Wakaba <[email protected]> | |
| 2 | ||
| 3 | * Tokenizer.pm.src: Changed to keep non-normal character | |
| 4 | references as is (HTML5 revision 3374). | |
| 5 | ||
| 6 | 2009-09-05 Wakaba <[email protected]> | |
| 7 | ||
| 8 | * Tokenizer.pm.src: Discard unclosed tags (HTML5 revision 2990). | |
| 9 | ||
| 10 | 2009-09-05 Wakaba <[email protected]> | |
| 11 | ||
| 12 | * Tokenizer.pm.src (_get_next_token): Implemented the "comment end | |
| 13 | space state" (HTML5 revision 3195). | |
| 14 | ||
| 15 | 2009-09-05 Wakaba <[email protected]> | |
| 16 | ||
| 17 | * Tokenizer.pm.src (_get_next_token): Implemented the "comment end | |
| 18 | bang state" (HTML5 revision 3191). | |
| 19 | ||
| 20 | 2009-08-16 Wakaba <[email protected]> | |
| 21 | ||
| 22 | * Tokenizer.pm.src: Any "<" character in attribute names become | |
| 23 | parse error (HTML5 revision 3354). | |
| 24 | ||
| 25 | 2009-08-16 Wakaba <[email protected]> | |
| 26 | ||
| 27 | * Tokenizer.pm.src: Lowercase-fold doctype names (HTML5 revision | |
| 28 | 2501, cf. HTML5 revision 3571). | |
| 29 | ||
| 30 | 2009-07-05 Wakaba <[email protected]> | |
| 31 | ||
| 32 | * Tokenizer.pm.src: Reduced the number of parse errors on broken | |
| 33 | DOCTYPE (HTML5 revision 3121). | |
| 34 | ||
| 35 | 2009-07-03 Wakaba <[email protected]> | |
| 36 | ||
| 37 | * Tokenizer.pm.src: Reduced a parse error (HTML5 revision 3194). | |
| 38 | ||
| 39 | 2009-07-03 Wakaba <[email protected]> | |
| 40 | ||
| 41 | * Tokenizer.pm.src: "<" in unquoted attribute values is now | |
| 42 | treated as parse error (HTML5 revision 3206). | |
| 43 | ||
| 44 | 2008-11-07 Wakaba <[email protected]> | |
| 45 | ||
| 46 | * Dumper.pm (dumptree): Support for namespace abbreviation for | |
| 47 | SWML namespaces. | |
| 48 | ||
| 49 | 2008-10-19 Wakaba <[email protected]> | |
| 50 | ||
| 51 | * Tokenizer.pm.src: Normalize white space characters in attribute | |
| 52 | value literals in XML documents. Don't apply character reference | |
| 53 | mapping table for non-NULL non-surrogate code points. | |
| 54 | ||
| 55 | 2008-10-19 Wakaba <[email protected]> | |
| 56 | ||
| 57 | * Tokenizer.pm.src: Set the "stop_processing" flag true when a | |
| 58 | parameter entity occurs in a standalone="no" document. | |
| 59 | ||
| 60 | 2008-10-19 Wakaba <[email protected]> | |
| 61 | ||
| 62 | * Tokenizer.pm.src: Column number counting fixed. | |
| 63 | ||
| 64 | 2008-10-19 Wakaba <[email protected]> | |
| 65 | ||
| 66 | * Tokenizer.pm.src: Raise a parse error for '&' that does not | |
| 67 | introduce a reference in XML. Support for non-ASCII entity | |
| 68 | reference names. | |
| 69 | ||
| 70 | 2008-10-19 Wakaba <[email protected]> | |
| 71 | ||
| 72 | * Tokenizer.pm.src: Make uppercase "&#X" in XML a parse error. | |
| 73 | Remove the limitation of entity name length. Enable replacement | |
| 74 | of text-only general entities. Raise a parse error for an | |
| 75 | unparsed entity reference. Raise a parse error for a general | |
| 76 | entity reference to an undefined entity. | |
| 77 | ||
| 78 | 2008-10-19 Wakaba <[email protected]> | |
| 79 | ||
| 80 | * Tokenizer.pm.src: Support for <!ELEMENT>. | |
| 81 | (AFTER_NOTATION_NAME_STATE): Renamed as |AFTER_MD_DEF_STATE| (i.e. | |
| 82 | after markup declaration definition state). | |
| 83 | ||
| 84 | 2008-10-19 Wakaba <[email protected]> | |
| 85 | ||
| 86 | * Tokenizer.pm.src: Support for EntityValue. | |
| 87 | ||
| 88 | 2008-10-19 Wakaba <[email protected]> | |
| 89 | ||
| 90 | * Dumper.pm: Dump text content of Entity nodes. | |
| 91 | ||
| 92 | * Tokenizer.pm.src: Support for <!ENTITY ... NDATA>. | |
| 93 | ||
| 94 | 2008-10-19 Wakaba <[email protected]> | |
| 95 | ||
| 96 | * Tokenizer.pm.src (_get_next_token): Make keywords 'ENTITY', | |
| 97 | 'ELEMENT', 'ATTLIST', and 'NOTATION' ASCII case-insensitive. | |
| 98 | ||
| 99 | 2008-10-18 Wakaba <[email protected]> | |
| 100 | ||
| 101 | * Tokenizer.pm.src: Modifies PUBLIC/SYSTEM identifier tokenizer | |
| 102 | states such that <!ENTITY> and <!NOTATION> can be tokenized by | |
| 103 | those states as well. | |
| 104 | (BOGUS_MD_STATE): A new state; used for bogus markup declarations, | |
| 105 | in favor of BOGUS_COMMENT_S |