--- test/html-webhacc/error-description-source.xml 2008/07/27 10:33:45 1.26 +++ test/html-webhacc/error-description-source.xml 2008/08/15 14:11:13 1.32 @@ -22,8 +22,9 @@

HTML5 Character Encoding Errors

- - Character encoding $0 + + Character encoding {text} is not allowed for HTML document.

The character encoding used for the document is not allowed @@ -31,9 +32,9 @@ - - Character encoding $0 + + Character encoding {text} should not be used for HTML document.

The character encoding used for the document is not recommended @@ -42,19 +43,20 @@ - - Use of UTF-8 is encouraged. + + Use of UTF-8 is encouraged (this document + is encoded in {text}).

Use of UTF-8 as the character encoding of the document is encouraged, though the use of another character encoding is still conforming.

- + Conformance for character encoding requirements - cannot be checked. + cannot be checked, since the input is not a byte stream.

The conformance checker cannot detect whether the input document met the requirements on character encoding, since the document @@ -64,8 +66,8 @@ - + There is no character encoding declaration. @@ -85,11 +87,11 @@ - + No character encoding metadata is found in lower‐level protocol nor is there BOM, while - character encoding $0 + character encoding {text} is not a superset of ASCII.

The document is not labeled with character encoding name @@ -115,11 +117,50 @@ - + + Character encoding of this document is sniffed + as {text} (Sniffed because no explicit specification + for the character encoding of this document is found in the transfer + procotol headers). + + + + Character encoding of this document is defaulted + to {text} because no explicit specification + for the character encoding of this document is found in the transfer + procotol headers. + + + + Since no decoder for the document character + encoding is found, decoder for the character encoding + {text} is used. Checking results might be + wrong. + + + + Conformance error checking for the character + encoding {text} is not supported. + + + + Sniffed character encoding + {text} is same as the character encoding specified + in the character encoding declaration. This is not an + error. + + + While parsing the document as - $0, a character encoding declaration specifying - character encoding as $1 is found. The document + {text}, a character encoding declaration specifying + a different character encoding is found. The document is reparsed.

While parsing a document in a character encoding, @@ -145,6 +186,19 @@ + + + The NULL character + is not allowed. + + + + Code point {text} is + not allowed. + +

@@ -179,8 +233,24 @@ + + Attribute name cannot contain characters + ", ', and =. + + + + Attribute value must be quoted by " + or ' if it contains a ", ', or + = character. + + + class="tokenize-error" + modules="HTML::Parser"> The & character must be escaped as &amp;. @@ -217,7 +287,8 @@ + class="tokenize-error" + modules="HTML::Parser"> A </ string is not followed by a tag name. @@ -240,7 +311,8 @@ + class="tokenize-error" + modules="HTML::Parser"> A < character is not followed by tag name or by a ! character. @@ -256,7 +328,8 @@ + class="tokenize-error" + modules="HTML::Parser"> The decimal representation of the code position of a character must be specified after &#. @@ -289,7 +362,8 @@ + class="tokenize-error" + modules="HTML::Parser"> The hexadecimal representation of the code position of a character must be specified after &#x. @@ -311,7 +385,8 @@ + class="tokenize-error" + modules="HTML::Parser"> String <! is not followed by --. @@ -345,7 +420,8 @@ + class="tokenize-error" + modules="HTML::Parser"> String </ is not followed by tag name. @@ -366,8 +442,24 @@ + + Character reference to + {text} is not allowed. + + + + Character reference to + U+000D (CARRIAGE RETURN) + is not allowed. + + + class="tokenize-error" + modules="HTML::Parser"> There is a -- sequence in a comment. @@ -384,9 +476,10 @@ + class="tokenize-error" + modules="HTML::Parser"> There are two attributes with name - $0. + {text}.

There are more than one attributes with the same name in a tag. The document is non-conforming.

@@ -396,8 +489,36 @@
+ + Empty start tag (<>) is not + allowed. + + + + Empty end tag (</>) is not + allowed. + + + + End tag cannot have attributes. + + + + Character reference to + {text} is not allowed. + + + class="tokenize-error" + modules="HTML::Parser"> Polytheistic slash (/>) cannot be used for this element. @@ -443,11 +564,55 @@ + + After the string <!DOCTYPE , the + document type name must be specified. + + + + After the keyword PUBLIC, no + oublic identifier is specified. + + + + Character reference must be closed by a + ; character. + + + + After the string <!DOCTYPE, there + must be at least a white space character before the document type + name. + + + + Attributes must be separeted by at least a + white space character. + + + + After the keyword SYSTEM, no + system identifier is specified. + + class="tokenize-error" + modules="HTML::Parser"> Processing instruction - (<?...>) cannot be used. + (<?...>) is not allowed in HTML + document.

Processing instructions (<?...?>), including XML declaration (<?xml ...?>) @@ -495,15 +660,135 @@ + + There is a bogus string after the document type + name. + + + + There is a bogus string after the keyword + PUBLIC. + + + + There is a bogus string after the public + identifier. + + + + There is a bogus string after the keyword + SYSTEM. + + + + There is a bogus string after the system + identifier. + + + + Attribute value is not closed by a quotation + mark. + + + + Comment is not closed by a string + -->. + + + + The DOCTYPE is not closed by a + > character. + + + + The public identifier literal is not closed by a + quotation mark. + + + + The system identifier literal is not closed by a + quotation mark. + + + + Tag is not closed by a > + character. + +

HTML5 Parse Errors in Tree Construction Stage

+ + Start tag <{text}> is + not allowed after the body is closed. + + + + End tag </{text}> is + not allowed after the body is closed. + + + + Non‐white‐space characters are not allowed + after the body is closed. + + + + Start tag <{text}> is + not allowed after the frameset is closed. + + + + End tag </{text}> is + not allowed after the frameset is closed. + + + + Non‐white‐space characters are not allowed + after the frame is closed. + + - The $0 element cannot be - inserted between head and body elements. + The {text} element cannot be + inserted between head and body + elements.

A start tag appears after the head element is closed but before the body element is opened. @@ -511,22 +796,37 @@ - - A DOCTYPE appears after any - element or data character has been seen. - -

A DOCTYPE appears after any element or data character - has been seen. The document is non-conforming.

- -

The DOCTYPE must be placed before any - tag, reference, or data character. Only white space characters - and comments can be inserted before the DOCTYPE.

-
+ + Start tag <{text}> is + not allowed after the html is closed. + + + + End tag </{text}> is + not allowed after the html is closed. + + + + Non‐white‐space characters are not allowed + after the html is closed. + + + + The image element is + obsolete. + class="parse-error" + modules="HTML::Parser"> Anchor cannot be nested.

HTML a elements cannot be nested. @@ -538,8 +838,9 @@ - Tag <$0> + class="parse-error" + modules="HTML::Parser"> + Start tag <{text}> is not allowed in the body element.

The start or end tag of an element, which @@ -549,8 +850,58 @@ + + Some element is not closed before the end of + file. + + + + The button element cannot be + nested. + + + + Element is not closed before the end of + file. + + + + Start tag <form> is + not allowed in a form element. + + + + Start tag <{text}> is + not allowed in a framset element. + + + + End tag </{text}> is + not allowed in a frameset element. + + + + Non‐white‐space characters are not allowed + in a frameset element. + + + class="parse-error" + modules="HTML::Parser"> Start tag <head> is not allowed in the head element. @@ -563,9 +914,85 @@ + + A DOCTYPE appears after any + element or data character has been seen. + + + +

A DOCTYPE appears after any element or data character + has been seen. The document is non-conforming.

+ +

The DOCTYPE must be placed before any + tag, reference, or data character. Only white space characters + and comments can be inserted before the DOCTYPE.

+
+
+ + + The nobr element cannot be + nested. + + + + The {text} element is not + allowed in a noscript element in the + head element. + + + + An end tag </{text}> + appers before the noscript element is closed. + + + + A noscript element is not closed + before the end of file. + + + + Non‐white‐space characters are not allowed + in a noscript element in the head + element. + + + + Element is not closed before the end of + file. + + + + Start tag <{text}> + is not allowed in a select element. + + + + End tag </{text}> + is not allowed in a select element. + + - Tag <$0> + class="parse-error" + modules="HTML::Parser"> + Start tag <{text}> is not allowed in a table element.

The start or end tag of an element, which @@ -581,13 +1008,21 @@ - - Data character is not allowed in - table. + + End tag </{text}> + is not allowed in a