genshi/mirror: markup/core.py annotate

annotate markup/core.py @ 112:5f9af749341c trunk

Docstring typo fix.

author	cmlenz
date	Mon, 31 Jul 2006 22:08:32 +0000
parents	2368c3becc52
children	d10fbba1d5e0

rev	line source
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	1 # -- coding: utf-8 --
5479aae32f5a Initial import. cmlenz parents: diff changeset	2 #
66 59eb24184e9c Switch copyright to Edgewall and URLs to markup.edgewall.org. cmlenz parents: 54 diff changeset	3 # Copyright (C) 2006 Edgewall Software
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	4 # All rights reserved.
5479aae32f5a Initial import. cmlenz parents: diff changeset	5 #
5479aae32f5a Initial import. cmlenz parents: diff changeset	6 # This software is licensed as described in the file COPYING, which
5479aae32f5a Initial import. cmlenz parents: diff changeset	7 # you should have received as part of this distribution. The terms
66 59eb24184e9c Switch copyright to Edgewall and URLs to markup.edgewall.org. cmlenz parents: 54 diff changeset	8 # are also available at http://markup.edgewall.org/wiki/License.
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	9 #
5479aae32f5a Initial import. cmlenz parents: diff changeset	10 # This software consists of voluntary contributions made by many
5479aae32f5a Initial import. cmlenz parents: diff changeset	11 # individuals. For the exact contribution history, see the revision
66 59eb24184e9c Switch copyright to Edgewall and URLs to markup.edgewall.org. cmlenz parents: 54 diff changeset	12 # history and logs, available at http://markup.edgewall.org/log/.
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	13
5479aae32f5a Initial import. cmlenz parents: diff changeset	14 """Core classes for markup processing."""
5479aae32f5a Initial import. cmlenz parents: diff changeset	15
5479aae32f5a Initial import. cmlenz parents: diff changeset	16 import htmlentitydefs
5479aae32f5a Initial import. cmlenz parents: diff changeset	17 import re
5479aae32f5a Initial import. cmlenz parents: diff changeset	18 from StringIO import StringIO
5479aae32f5a Initial import. cmlenz parents: diff changeset	19
5479aae32f5a Initial import. cmlenz parents: diff changeset	20 __all__ = ['Stream', 'Markup', 'escape', 'unescape', 'Namespace', 'QName']
5479aae32f5a Initial import. cmlenz parents: diff changeset	21
5479aae32f5a Initial import. cmlenz parents: diff changeset	22
17 74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	23 class StreamEventKind(str):
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	24 """A kind of event on an XML stream."""
5479aae32f5a Initial import. cmlenz parents: diff changeset	25
5479aae32f5a Initial import. cmlenz parents: diff changeset	26
5479aae32f5a Initial import. cmlenz parents: diff changeset	27 class Stream(object):
5479aae32f5a Initial import. cmlenz parents: diff changeset	28 """Represents a stream of markup events.
5479aae32f5a Initial import. cmlenz parents: diff changeset	29
5479aae32f5a Initial import. cmlenz parents: diff changeset	30 This class is basically an iterator over the events.
5479aae32f5a Initial import. cmlenz parents: diff changeset	31
5479aae32f5a Initial import. cmlenz parents: diff changeset	32 Also provided are ways to serialize the stream to text. The `serialize()`
5479aae32f5a Initial import. cmlenz parents: diff changeset	33 method will return an iterator over generated strings, while `render()`
5479aae32f5a Initial import. cmlenz parents: diff changeset	34 returns the complete generated text at once. Both accept various parameters
5479aae32f5a Initial import. cmlenz parents: diff changeset	35 that impact the way the stream is serialized.
5479aae32f5a Initial import. cmlenz parents: diff changeset	36
5479aae32f5a Initial import. cmlenz parents: diff changeset	37 Stream events are tuples of the form:
5479aae32f5a Initial import. cmlenz parents: diff changeset	38
5479aae32f5a Initial import. cmlenz parents: diff changeset	39 (kind, data, position)
5479aae32f5a Initial import. cmlenz parents: diff changeset	40
5479aae32f5a Initial import. cmlenz parents: diff changeset	41 where `kind` is the event kind (such as `START`, `END`, `TEXT`, etc), `data`
5479aae32f5a Initial import. cmlenz parents: diff changeset	42 depends on the kind of event, and `position` is a `(line, offset)` tuple
5479aae32f5a Initial import. cmlenz parents: diff changeset	43 that contains the location of the original element or text in the input.
5479aae32f5a Initial import. cmlenz parents: diff changeset	44 """
5479aae32f5a Initial import. cmlenz parents: diff changeset	45 __slots__ = ['events']
5479aae32f5a Initial import. cmlenz parents: diff changeset	46
17 74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	47 START = StreamEventKind('START') # a start tag
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	48 END = StreamEventKind('END') # an end tag
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	49 TEXT = StreamEventKind('TEXT') # literal text
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	50 PROLOG = StreamEventKind('PROLOG') # XML prolog
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	51 DOCTYPE = StreamEventKind('DOCTYPE') # doctype declaration
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	52 START_NS = StreamEventKind('START-NS') # start namespace mapping
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	53 END_NS = StreamEventKind('END-NS') # end namespace mapping
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	54 PI = StreamEventKind('PI') # processing instruction
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	55 COMMENT = StreamEventKind('COMMENT') # comment
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	56
5479aae32f5a Initial import. cmlenz parents: diff changeset	57 def __init__(self, events):
5479aae32f5a Initial import. cmlenz parents: diff changeset	58 """Initialize the stream with a sequence of markup events.
5479aae32f5a Initial import. cmlenz parents: diff changeset	59
27 b4f78c05e5c9 * Fix the boilerplate in the Python source files. cmlenz parents: 18 diff changeset	60 @param events: a sequence or iterable providing the events
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	61 """
5479aae32f5a Initial import. cmlenz parents: diff changeset	62 self.events = events
5479aae32f5a Initial import. cmlenz parents: diff changeset	63
5479aae32f5a Initial import. cmlenz parents: diff changeset	64 def __iter__(self):
5479aae32f5a Initial import. cmlenz parents: diff changeset	65 return iter(self.events)
5479aae32f5a Initial import. cmlenz parents: diff changeset	66
17 74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	67 def render(self, method='xml', encoding='utf-8', filters=None, **kwargs):
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	68 """Return a string representation of the stream.
5479aae32f5a Initial import. cmlenz parents: diff changeset	69
5479aae32f5a Initial import. cmlenz parents: diff changeset	70 @param method: determines how the stream is serialized; can be either
96 fa08aef181a2 Add an XHTML serialization method. Now really need to get rid of some code duplication in the `markup.output` module. cmlenz parents: 91 diff changeset	71 "xml", "xhtml", or "html", or a custom `Serializer`
fa08aef181a2 Add an XHTML serialization method. Now really need to get rid of some code duplication in the `markup.output` module. cmlenz parents: 91 diff changeset	72 subclass
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	73 @param encoding: how the output string should be encoded; if set to
5479aae32f5a Initial import. cmlenz parents: diff changeset	74 `None`, this method returns a `unicode` object
5479aae32f5a Initial import. cmlenz parents: diff changeset	75
5479aae32f5a Initial import. cmlenz parents: diff changeset	76 Any additional keyword arguments are passed to the serializer, and thus
5479aae32f5a Initial import. cmlenz parents: diff changeset	77 depend on the `method` parameter value.
5479aae32f5a Initial import. cmlenz parents: diff changeset	78 """
17 74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	79 generator = self.serialize(method=method, filters=filters, **kwargs)
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	80 output = u''.join(list(generator))
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	81 if encoding is not None:
9 5dc4bfe67c20 Actually use the specified encoding in `Stream.render()`. cmlenz parents: 8 diff changeset	82 return output.encode(encoding)
8 3710e3d0d4a2 `Stream.render()` was masking `TypeError`s (fix based on suggestion by Matt Good). cmlenz parents: 6 diff changeset	83 return output
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	84
5479aae32f5a Initial import. cmlenz parents: diff changeset	85 def select(self, path):
5479aae32f5a Initial import. cmlenz parents: diff changeset	86 """Return a new stream that contains the events matching the given
5479aae32f5a Initial import. cmlenz parents: diff changeset	87 XPath expression.
5479aae32f5a Initial import. cmlenz parents: diff changeset	88
5479aae32f5a Initial import. cmlenz parents: diff changeset	89 @param path: a string containing the XPath expression
5479aae32f5a Initial import. cmlenz parents: diff changeset	90 """
5479aae32f5a Initial import. cmlenz parents: diff changeset	91 from markup.path import Path
17 74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	92 return Path(path).select(self)
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	93
17 74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	94 def serialize(self, method='xml', filters=None, **kwargs):
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	95 """Generate strings corresponding to a specific serialization of the
5479aae32f5a Initial import. cmlenz parents: diff changeset	96 stream.
5479aae32f5a Initial import. cmlenz parents: diff changeset	97
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	98 Unlike the `render()` method, this method is a generator that returns
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	99 the serialized output incrementally, as opposed to returning a single
5479aae32f5a Initial import. cmlenz parents: diff changeset	100 string.
5479aae32f5a Initial import. cmlenz parents: diff changeset	101
5479aae32f5a Initial import. cmlenz parents: diff changeset	102 @param method: determines how the stream is serialized; can be either
96 fa08aef181a2 Add an XHTML serialization method. Now really need to get rid of some code duplication in the `markup.output` module. cmlenz parents: 91 diff changeset	103 "xml", "xhtml", or "html", or a custom `Serializer`
fa08aef181a2 Add an XHTML serialization method. Now really need to get rid of some code duplication in the `markup.output` module. cmlenz parents: 91 diff changeset	104 subclass
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	105 @param filters: list of filters to apply to the stream before
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	106 serialization. The default is to apply whitespace
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	107 reduction using `markup.filters.WhitespaceFilter`.
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	108 """
17 74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	109 from markup.filters import WhitespaceFilter
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	110 from markup import output
5479aae32f5a Initial import. cmlenz parents: diff changeset	111 cls = method
5479aae32f5a Initial import. cmlenz parents: diff changeset	112 if isinstance(method, basestring):
96 fa08aef181a2 Add an XHTML serialization method. Now really need to get rid of some code duplication in the `markup.output` module. cmlenz parents: 91 diff changeset	113 cls = {'xml': output.XMLSerializer,
fa08aef181a2 Add an XHTML serialization method. Now really need to get rid of some code duplication in the `markup.output` module. cmlenz parents: 91 diff changeset	114 'xhtml': output.XHTMLSerializer,
fa08aef181a2 Add an XHTML serialization method. Now really need to get rid of some code duplication in the `markup.output` module. cmlenz parents: 91 diff changeset	115 'html': output.HTMLSerializer}[method]
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	116 else:
27 b4f78c05e5c9 * Fix the boilerplate in the Python source files. cmlenz parents: 18 diff changeset	117 assert issubclass(cls, output.Serializer)
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	118 serializer = cls(**kwargs)
17 74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	119
111 2368c3becc52 Some fixes and more unit tests for the XPath engine. cmlenz parents: 100 diff changeset	120 stream = _ensure(self)
17 74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	121 if filters is None:
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	122 filters = [WhitespaceFilter()]
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	123 for filter_ in filters:
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	124 stream = filter_(iter(stream))
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	125
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	126 return serializer.serialize(stream)
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	127
5479aae32f5a Initial import. cmlenz parents: diff changeset	128 def __str__(self):
5479aae32f5a Initial import. cmlenz parents: diff changeset	129 return self.render()
5479aae32f5a Initial import. cmlenz parents: diff changeset	130
5479aae32f5a Initial import. cmlenz parents: diff changeset	131 def __unicode__(self):
5479aae32f5a Initial import. cmlenz parents: diff changeset	132 return self.render(encoding=None)
5479aae32f5a Initial import. cmlenz parents: diff changeset	133
5479aae32f5a Initial import. cmlenz parents: diff changeset	134
69 c40a5dcd2b55 A couple of minor performance improvements. cmlenz parents: 66 diff changeset	135 START = Stream.START
c40a5dcd2b55 A couple of minor performance improvements. cmlenz parents: 66 diff changeset	136 END = Stream.END
c40a5dcd2b55 A couple of minor performance improvements. cmlenz parents: 66 diff changeset	137 TEXT = Stream.TEXT
c40a5dcd2b55 A couple of minor performance improvements. cmlenz parents: 66 diff changeset	138 PROLOG = Stream.PROLOG
c40a5dcd2b55 A couple of minor performance improvements. cmlenz parents: 66 diff changeset	139 DOCTYPE = Stream.DOCTYPE
c40a5dcd2b55 A couple of minor performance improvements. cmlenz parents: 66 diff changeset	140 START_NS = Stream.START_NS
c40a5dcd2b55 A couple of minor performance improvements. cmlenz parents: 66 diff changeset	141 END_NS = Stream.END_NS
c40a5dcd2b55 A couple of minor performance improvements. cmlenz parents: 66 diff changeset	142 PI = Stream.PI
c40a5dcd2b55 A couple of minor performance improvements. cmlenz parents: 66 diff changeset	143 COMMENT = Stream.COMMENT
c40a5dcd2b55 A couple of minor performance improvements. cmlenz parents: 66 diff changeset	144
111 2368c3becc52 Some fixes and more unit tests for the XPath engine. cmlenz parents: 100 diff changeset	145 def _ensure(stream):
2368c3becc52 Some fixes and more unit tests for the XPath engine. cmlenz parents: 100 diff changeset	146 """Ensure that every item on the stream is actually a markup event."""
2368c3becc52 Some fixes and more unit tests for the XPath engine. cmlenz parents: 100 diff changeset	147 for event in stream:
2368c3becc52 Some fixes and more unit tests for the XPath engine. cmlenz parents: 100 diff changeset	148 try:
2368c3becc52 Some fixes and more unit tests for the XPath engine. cmlenz parents: 100 diff changeset	149 kind, data, pos = event
2368c3becc52 Some fixes and more unit tests for the XPath engine. cmlenz parents: 100 diff changeset	150 except ValueError:
2368c3becc52 Some fixes and more unit tests for the XPath engine. cmlenz parents: 100 diff changeset	151 kind, data, pos = event.totuple()
2368c3becc52 Some fixes and more unit tests for the XPath engine. cmlenz parents: 100 diff changeset	152 yield kind, data, pos
2368c3becc52 Some fixes and more unit tests for the XPath engine. cmlenz parents: 100 diff changeset	153
69 c40a5dcd2b55 A couple of minor performance improvements. cmlenz parents: 66 diff changeset	154
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	155 class Attributes(list):
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	156 """Sequence type that stores the attributes of an element.
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	157
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	158 The order of the attributes is preserved, while accessing and manipulating
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	159 attributes by name is also supported.
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	160
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	161 >>> attrs = Attributes([('href', '#'), ('title', 'Foo')])
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	162 >>> attrs
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	163 [(u'href', '#'), (u'title', 'Foo')]
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	164
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	165 >>> 'href' in attrs
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	166 True
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	167 >>> 'tabindex' in attrs
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	168 False
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	169
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	170 >>> attrs.get(u'title')
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	171 'Foo'
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	172 >>> attrs.set(u'title', 'Bar')
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	173 >>> attrs
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	174 [(u'href', '#'), (u'title', 'Bar')]
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	175 >>> attrs.remove(u'title')
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	176 >>> attrs
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	177 [(u'href', '#')]
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	178
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	179 New attributes added using the `set()` method are appended to the end of
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	180 the list:
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	181
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	182 >>> attrs.set(u'accesskey', 'k')
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	183 >>> attrs
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	184 [(u'href', '#'), (u'accesskey', 'k')]
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	185 """
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	186 __slots__ = []
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	187
5479aae32f5a Initial import. cmlenz parents: diff changeset	188 def __init__(self, attrib=None):
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	189 """Create the `Attributes` instance.
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	190
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	191 If the `attrib` parameter is provided, it is expected to be a sequence
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	192 of `(name, value)` tuples.
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	193 """
27 b4f78c05e5c9 * Fix the boilerplate in the Python source files. cmlenz parents: 18 diff changeset	194 if attrib is None:
b4f78c05e5c9 * Fix the boilerplate in the Python source files. cmlenz parents: 18 diff changeset	195 attrib = []
b4f78c05e5c9 * Fix the boilerplate in the Python source files. cmlenz parents: 18 diff changeset	196 list.__init__(self, [(QName(name), value) for name, value in attrib])
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	197
5479aae32f5a Initial import. cmlenz parents: diff changeset	198 def __contains__(self, name):
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	199 """Return whether the list includes an attribute with the specified
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	200 name.
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	201 """
27 b4f78c05e5c9 * Fix the boilerplate in the Python source files. cmlenz parents: 18 diff changeset	202 return name in [attr for attr, _ in self]
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	203
5479aae32f5a Initial import. cmlenz parents: diff changeset	204 def get(self, name, default=None):
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	205 """Return the value of the attribute with the specified name, or the
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	206 value of the `default` parameter if no such attribute is found.
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	207 """
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	208 for attr, value in self:
5479aae32f5a Initial import. cmlenz parents: diff changeset	209 if attr == name:
5479aae32f5a Initial import. cmlenz parents: diff changeset	210 return value
5479aae32f5a Initial import. cmlenz parents: diff changeset	211 return default
5479aae32f5a Initial import. cmlenz parents: diff changeset	212
5 dbb08edbc615 Improved `py:attrs` directive so that it removes existing attributes if they evaluate to `None` (AFAICT matching Kid behavior). cmlenz parents: 1 diff changeset	213 def remove(self, name):
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	214 """Removes the attribute with the specified name.
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	215
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	216 If no such attribute is found, this method does nothing.
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	217 """
5 dbb08edbc615 Improved `py:attrs` directive so that it removes existing attributes if they evaluate to `None` (AFAICT matching Kid behavior). cmlenz parents: 1 diff changeset	218 for idx, (attr, _) in enumerate(self):
dbb08edbc615 Improved `py:attrs` directive so that it removes existing attributes if they evaluate to `None` (AFAICT matching Kid behavior). cmlenz parents: 1 diff changeset	219 if attr == name:
dbb08edbc615 Improved `py:attrs` directive so that it removes existing attributes if they evaluate to `None` (AFAICT matching Kid behavior). cmlenz parents: 1 diff changeset	220 del self[idx]
dbb08edbc615 Improved `py:attrs` directive so that it removes existing attributes if they evaluate to `None` (AFAICT matching Kid behavior). cmlenz parents: 1 diff changeset	221 break
dbb08edbc615 Improved `py:attrs` directive so that it removes existing attributes if they evaluate to `None` (AFAICT matching Kid behavior). cmlenz parents: 1 diff changeset	222
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	223 def set(self, name, value):
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	224 """Sets the specified attribute to the given value.
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	225
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	226 If an attribute with the specified name is already in the list, the
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	227 value of the existing entry is updated. Otherwise, a new attribute is
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	228 appended to the end of the list.
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	229 """
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	230 for idx, (attr, _) in enumerate(self):
5479aae32f5a Initial import. cmlenz parents: diff changeset	231 if attr == name:
5479aae32f5a Initial import. cmlenz parents: diff changeset	232 self[idx] = (attr, value)
5479aae32f5a Initial import. cmlenz parents: diff changeset	233 break
5479aae32f5a Initial import. cmlenz parents: diff changeset	234 else:
5479aae32f5a Initial import. cmlenz parents: diff changeset	235 self.append((QName(name), value))
5479aae32f5a Initial import. cmlenz parents: diff changeset	236
77 f5ec6d4a61e4 * Simplify implementation of the individual XPath tests (use closures instead of callable classes) cmlenz parents: 73 diff changeset	237 def totuple(self):
f5ec6d4a61e4 * Simplify implementation of the individual XPath tests (use closures instead of callable classes) cmlenz parents: 73 diff changeset	238 return TEXT, u''.join([x[1] for x in self]), (None, -1, -1)
f5ec6d4a61e4 * Simplify implementation of the individual XPath tests (use closures instead of callable classes) cmlenz parents: 73 diff changeset	239
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	240
5479aae32f5a Initial import. cmlenz parents: diff changeset	241 class Markup(unicode):
5479aae32f5a Initial import. cmlenz parents: diff changeset	242 """Marks a string as being safe for inclusion in HTML/XML output without
5479aae32f5a Initial import. cmlenz parents: diff changeset	243 needing to be escaped.
5479aae32f5a Initial import. cmlenz parents: diff changeset	244 """
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	245 __slots__ = []
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	246
27 b4f78c05e5c9 * Fix the boilerplate in the Python source files. cmlenz parents: 18 diff changeset	247 def __new__(cls, text='', *args):
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	248 if args:
5479aae32f5a Initial import. cmlenz parents: diff changeset	249 text %= tuple([escape(arg) for arg in args])
27 b4f78c05e5c9 * Fix the boilerplate in the Python source files. cmlenz parents: 18 diff changeset	250 return unicode.__new__(cls, text)
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	251
5479aae32f5a Initial import. cmlenz parents: diff changeset	252 def __add__(self, other):
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	253 return Markup(unicode(self) + escape(other))
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	254
5479aae32f5a Initial import. cmlenz parents: diff changeset	255 def __mod__(self, args):
5479aae32f5a Initial import. cmlenz parents: diff changeset	256 if not isinstance(args, (list, tuple)):
5479aae32f5a Initial import. cmlenz parents: diff changeset	257 args = [args]
5479aae32f5a Initial import. cmlenz parents: diff changeset	258 return Markup(unicode.__mod__(self,
5479aae32f5a Initial import. cmlenz parents: diff changeset	259 tuple([escape(arg) for arg in args])))
5479aae32f5a Initial import. cmlenz parents: diff changeset	260
5479aae32f5a Initial import. cmlenz parents: diff changeset	261 def __mul__(self, num):
5479aae32f5a Initial import. cmlenz parents: diff changeset	262 return Markup(unicode(self) * num)
5479aae32f5a Initial import. cmlenz parents: diff changeset	263
17 74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	264 def __repr__(self):
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	265 return '<%s "%s">' % (self.__class__.__name__, self)
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	266
54 1f3cd91325d9 Fix a number of escaping problems: cmlenz parents: 34 diff changeset	267 def join(self, seq, escape_quotes=True):
1f3cd91325d9 Fix a number of escaping problems: cmlenz parents: 34 diff changeset	268 return Markup(unicode(self).join([escape(item, quotes=escape_quotes)
34 3421dd98f015 quotes should not be escaped inside text nodes mgood parents: 27 diff changeset	269 for item in seq]))
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	270
5479aae32f5a Initial import. cmlenz parents: diff changeset	271 def stripentities(self, keepxmlentities=False):
5479aae32f5a Initial import. cmlenz parents: diff changeset	272 """Return a copy of the text with any character or numeric entities
5479aae32f5a Initial import. cmlenz parents: diff changeset	273 replaced by the equivalent UTF-8 characters.
5479aae32f5a Initial import. cmlenz parents: diff changeset	274
5479aae32f5a Initial import. cmlenz parents: diff changeset	275 If the `keepxmlentities` parameter is provided and evaluates to `True`,
17 74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	276 the core XML entities (&, ', >, < and ") are not
74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	277 stripped.
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	278 """
5479aae32f5a Initial import. cmlenz parents: diff changeset	279 def _replace_entity(match):
5479aae32f5a Initial import. cmlenz parents: diff changeset	280 if match.group(1): # numeric entity
5479aae32f5a Initial import. cmlenz parents: diff changeset	281 ref = match.group(1)
5479aae32f5a Initial import. cmlenz parents: diff changeset	282 if ref.startswith('x'):
5479aae32f5a Initial import. cmlenz parents: diff changeset	283 ref = int(ref[1:], 16)
5479aae32f5a Initial import. cmlenz parents: diff changeset	284 else:
5479aae32f5a Initial import. cmlenz parents: diff changeset	285 ref = int(ref, 10)
5479aae32f5a Initial import. cmlenz parents: diff changeset	286 return unichr(ref)
5479aae32f5a Initial import. cmlenz parents: diff changeset	287 else: # character entity
5479aae32f5a Initial import. cmlenz parents: diff changeset	288 ref = match.group(2)
27 b4f78c05e5c9 * Fix the boilerplate in the Python source files. cmlenz parents: 18 diff changeset	289 if keepxmlentities and ref in ('amp', 'apos', 'gt', 'lt',
b4f78c05e5c9 * Fix the boilerplate in the Python source files. cmlenz parents: 18 diff changeset	290 'quot'):
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	291 return '&%s;' % ref
5479aae32f5a Initial import. cmlenz parents: diff changeset	292 try:
5479aae32f5a Initial import. cmlenz parents: diff changeset	293 codepoint = htmlentitydefs.name2codepoint[ref]
5479aae32f5a Initial import. cmlenz parents: diff changeset	294 return unichr(codepoint)
5479aae32f5a Initial import. cmlenz parents: diff changeset	295 except KeyError:
5479aae32f5a Initial import. cmlenz parents: diff changeset	296 if keepxmlentities:
5479aae32f5a Initial import. cmlenz parents: diff changeset	297 return '&%s;' % ref
5479aae32f5a Initial import. cmlenz parents: diff changeset	298 else:
5479aae32f5a Initial import. cmlenz parents: diff changeset	299 return ref
5479aae32f5a Initial import. cmlenz parents: diff changeset	300 return Markup(re.sub(r'&(?:#((?:\d+)\|(?:[xX][0-9a-fA-F]+));?\|(\w+);)',
5479aae32f5a Initial import. cmlenz parents: diff changeset	301 _replace_entity, self))
5479aae32f5a Initial import. cmlenz parents: diff changeset	302
5479aae32f5a Initial import. cmlenz parents: diff changeset	303 def striptags(self):
5479aae32f5a Initial import. cmlenz parents: diff changeset	304 """Return a copy of the text with all XML/HTML tags removed."""
5479aae32f5a Initial import. cmlenz parents: diff changeset	305 return Markup(re.sub(r'<[^>]*?>', '', self))
5479aae32f5a Initial import. cmlenz parents: diff changeset	306
5479aae32f5a Initial import. cmlenz parents: diff changeset	307 def escape(cls, text, quotes=True):
5479aae32f5a Initial import. cmlenz parents: diff changeset	308 """Create a Markup instance from a string and escape special characters
5479aae32f5a Initial import. cmlenz parents: diff changeset	309 it may contain (<, >, & and \").
5479aae32f5a Initial import. cmlenz parents: diff changeset	310
5479aae32f5a Initial import. cmlenz parents: diff changeset	311 If the `quotes` parameter is set to `False`, the \" character is left
5479aae32f5a Initial import. cmlenz parents: diff changeset	312 as is. Escaping quotes is generally only required for strings that are
5479aae32f5a Initial import. cmlenz parents: diff changeset	313 to be used in attribute values.
5479aae32f5a Initial import. cmlenz parents: diff changeset	314 """
73 1da51d718391 Some more performance tweaks. cmlenz parents: 69 diff changeset	315 if not text:
1da51d718391 Some more performance tweaks. cmlenz parents: 69 diff changeset	316 return cls()
1da51d718391 Some more performance tweaks. cmlenz parents: 69 diff changeset	317 if type(text) is cls:
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	318 return text
73 1da51d718391 Some more performance tweaks. cmlenz parents: 69 diff changeset	319 text = unicode(text).replace('&', '&') \
1da51d718391 Some more performance tweaks. cmlenz parents: 69 diff changeset	320 .replace('<', '<') \
1da51d718391 Some more performance tweaks. cmlenz parents: 69 diff changeset	321 .replace('>', '>')
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	322 if quotes:
5479aae32f5a Initial import. cmlenz parents: diff changeset	323 text = text.replace('"', '"')
5479aae32f5a Initial import. cmlenz parents: diff changeset	324 return cls(text)
5479aae32f5a Initial import. cmlenz parents: diff changeset	325 escape = classmethod(escape)
5479aae32f5a Initial import. cmlenz parents: diff changeset	326
5479aae32f5a Initial import. cmlenz parents: diff changeset	327 def unescape(self):
5479aae32f5a Initial import. cmlenz parents: diff changeset	328 """Reverse-escapes &, <, > and \" and returns a `unicode` object."""
5479aae32f5a Initial import. cmlenz parents: diff changeset	329 if not self:
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	330 return u''
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	331 return unicode(self).replace('"', '"') \
5479aae32f5a Initial import. cmlenz parents: diff changeset	332 .replace('>', '>') \
5479aae32f5a Initial import. cmlenz parents: diff changeset	333 .replace('<', '<') \
5479aae32f5a Initial import. cmlenz parents: diff changeset	334 .replace('&', '&')
5479aae32f5a Initial import. cmlenz parents: diff changeset	335
5479aae32f5a Initial import. cmlenz parents: diff changeset	336 def plaintext(self, keeplinebreaks=True):
6 71e8e645fe81 Simplified implementation of `py:content` directive. cmlenz parents: 5 diff changeset	337 """Returns the text as a `unicode` string with all entities and tags
71e8e645fe81 Simplified implementation of `py:content` directive. cmlenz parents: 5 diff changeset	338 removed.
71e8e645fe81 Simplified implementation of `py:content` directive. cmlenz parents: 5 diff changeset	339 """
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	340 text = unicode(self.striptags().stripentities())
5479aae32f5a Initial import. cmlenz parents: diff changeset	341 if not keeplinebreaks:
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	342 text = text.replace(u'\n', u' ')
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	343 return text
5479aae32f5a Initial import. cmlenz parents: diff changeset	344
5479aae32f5a Initial import. cmlenz parents: diff changeset	345 def sanitize(self):
5479aae32f5a Initial import. cmlenz parents: diff changeset	346 from markup.filters import HTMLSanitizer
5479aae32f5a Initial import. cmlenz parents: diff changeset	347 from markup.input import HTMLParser
17 74cc70129d04 Refactoring to address #6: all match templates are now processed by a single filter, which means that match templates added by included templates are properly applied. A side effect of this refactoring is that `Context` objects may not be reused across multiple template processing runs. cmlenz parents: 10 diff changeset	348 text = StringIO(self.stripentities(keepxmlentities=True))
91 a71a58df6bf5 Some subtle fixes to generation and sanitization. cmlenz parents: 77 diff changeset	349 return Markup(Stream(HTMLSanitizer()(HTMLParser(text))))
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	350
5479aae32f5a Initial import. cmlenz parents: diff changeset	351
5479aae32f5a Initial import. cmlenz parents: diff changeset	352 escape = Markup.escape
5479aae32f5a Initial import. cmlenz parents: diff changeset	353
5479aae32f5a Initial import. cmlenz parents: diff changeset	354 def unescape(text):
5479aae32f5a Initial import. cmlenz parents: diff changeset	355 """Reverse-escapes &, <, > and \" and returns a `unicode` object."""
5479aae32f5a Initial import. cmlenz parents: diff changeset	356 if not isinstance(text, Markup):
5479aae32f5a Initial import. cmlenz parents: diff changeset	357 return text
5479aae32f5a Initial import. cmlenz parents: diff changeset	358 return text.unescape()
5479aae32f5a Initial import. cmlenz parents: diff changeset	359
5479aae32f5a Initial import. cmlenz parents: diff changeset	360
5479aae32f5a Initial import. cmlenz parents: diff changeset	361 class Namespace(object):
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	362 """Utility class creating and testing elements with a namespace.
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	363
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	364 Internally, namespace URIs are encoded in the `QName` of any element or
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	365 attribute, the namespace URI being enclosed in curly braces. This class
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	366 helps create and test these strings.
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	367
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	368 A `Namespace` object is instantiated with the namespace URI.
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	369
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	370 >>> html = Namespace('http://www.w3.org/1999/xhtml')
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	371 >>> html
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	372 <Namespace "http://www.w3.org/1999/xhtml">
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	373 >>> html.uri
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	374 u'http://www.w3.org/1999/xhtml'
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	375
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	376 The `Namespace` object can than be used to generate `QName` objects with
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	377 that namespace:
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	378
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	379 >>> html.body
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	380 u'{http://www.w3.org/1999/xhtml}body'
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	381 >>> html.body.localname
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	382 u'body'
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	383 >>> html.body.namespace
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	384 u'http://www.w3.org/1999/xhtml'
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	385
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	386 The same works using item access notation, which is useful for element or
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	387 attribute names that are not valid Python identifiers:
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	388
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	389 >>> html['body']
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	390 u'{http://www.w3.org/1999/xhtml}body'
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	391
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	392 A `Namespace` object can also be used to test whether a specific `QName`
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	393 belongs to that namespace using the `in` operator:
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	394
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	395 >>> qname = html.body
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	396 >>> qname in html
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	397 True
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	398 >>> qname in Namespace('http://www.w3.org/2002/06/xhtml2')
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	399 False
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	400 """
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	401 def __init__(self, uri):
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	402 self.uri = unicode(uri)
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	403
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	404 def __contains__(self, qname):
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	405 return qname.namespace == self.uri
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	406
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	407 def __eq__(self, other):
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	408 if isinstance(other, Namespace):
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	409 return self.uri == other.uri
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	410 return self.uri == other
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	411
5479aae32f5a Initial import. cmlenz parents: diff changeset	412 def __getitem__(self, name):
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	413 return QName(self.uri + u'}' + name)
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	414 __getattr__ = __getitem__
5479aae32f5a Initial import. cmlenz parents: diff changeset	415
5479aae32f5a Initial import. cmlenz parents: diff changeset	416 def __repr__(self):
5479aae32f5a Initial import. cmlenz parents: diff changeset	417 return '<Namespace "%s">' % self.uri
5479aae32f5a Initial import. cmlenz parents: diff changeset	418
5479aae32f5a Initial import. cmlenz parents: diff changeset	419 def __str__(self):
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	420 return self.uri.encode('utf-8')
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	421
5479aae32f5a Initial import. cmlenz parents: diff changeset	422 def __unicode__(self):
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	423 return self.uri
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	424
5479aae32f5a Initial import. cmlenz parents: diff changeset	425
5479aae32f5a Initial import. cmlenz parents: diff changeset	426 class QName(unicode):
5479aae32f5a Initial import. cmlenz parents: diff changeset	427 """A qualified element or attribute name.
5479aae32f5a Initial import. cmlenz parents: diff changeset	428
5479aae32f5a Initial import. cmlenz parents: diff changeset	429 The unicode value of instances of this class contains the qualified name of
5479aae32f5a Initial import. cmlenz parents: diff changeset	430 the element or attribute, in the form `{namespace}localname`. The namespace
5479aae32f5a Initial import. cmlenz parents: diff changeset	431 URI can be obtained through the additional `namespace` attribute, while the
5479aae32f5a Initial import. cmlenz parents: diff changeset	432 local name can be accessed through the `localname` attribute.
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	433
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	434 >>> qname = QName('foo')
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	435 >>> qname
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	436 u'foo'
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	437 >>> qname.localname
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	438 u'foo'
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	439 >>> qname.namespace
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	440
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	441 >>> qname = QName('http://www.w3.org/1999/xhtml}body')
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	442 >>> qname
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	443 u'{http://www.w3.org/1999/xhtml}body'
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	444 >>> qname.localname
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	445 u'body'
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	446 >>> qname.namespace
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	447 u'http://www.w3.org/1999/xhtml'
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	448 """
5479aae32f5a Initial import. cmlenz parents: diff changeset	449 __slots__ = ['namespace', 'localname']
5479aae32f5a Initial import. cmlenz parents: diff changeset	450
5479aae32f5a Initial import. cmlenz parents: diff changeset	451 def __new__(cls, qname):
100 a519f581a1b1 Ported [111] to trunk. cmlenz parents: 96 diff changeset	452 if type(qname) is cls:
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	453 return qname
5479aae32f5a Initial import. cmlenz parents: diff changeset	454
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	455 parts = qname.split(u'}', 1)
100 a519f581a1b1 Ported [111] to trunk. cmlenz parents: 96 diff changeset	456 if len(parts) > 1:
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	457 self = unicode.__new__(cls, u'{' + qname)
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	458 self.namespace = unicode(parts[0])
5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	459 self.localname = unicode(parts[1])
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	460 else:
5479aae32f5a Initial import. cmlenz parents: diff changeset	461 self = unicode.__new__(cls, qname)
5479aae32f5a Initial import. cmlenz parents: diff changeset	462 self.namespace = None
18 5420cfe42d36 Actually make use of the `markup.core.Namespace` class, and add a couple of doctests. cmlenz parents: 17 diff changeset	463 self.localname = unicode(qname)
1 5479aae32f5a Initial import. cmlenz parents: diff changeset	464 return self

Mercurial > genshi > mirror

annotate markup/core.py @ 112:5f9af749341c trunk