wikimedia/mediawiki-extensions-VisualEditor - fanwikis.org Git Server

wikimedia/mediawiki-extensions-VisualEditor

mirror of https://gerrit.wikimedia.org/r/mediawiki/extensions/VisualEditor synced 2024-12-02 01:46:47 +00:00

Author	SHA1	Message	Date
Catrope	04516bb02e	Whitespace preservation was broken after the first run The first run of getDomFromData() would preserve whitespace just fine, but it blanked out the .veInternal.whitespace[1] element in certain cases, contaminating the linear model and making the whitespace data inconsistent. Subsequent runs of getDomFromData() would then refuse to serialize that whitespace because the information about it was inconsistent. In getDomFromData(), we sometimes unset .veInternal.whitespace[1] (i.e. set it to undefined) to prevent double processing. Because we're potentially going to modify .veInternal, don't assign it by reference, but copy the object. Added tests asserting that the linear model is unchanged after calling getDomFromData(), because that function should never modify linear model data. This test failed in 4 cases (all whitespace-related) before I added the copyObject() call. Bug: 43543 Change-Id: Ic4c93510518163894201a693ab50331413715967	2013-04-17 11:28:05 +00:00
Ed Sanders	6dacc54954	Hybridise MWTemplateNode * Create MWTemplateBlockNode & MWTemplateInlineNode (in ce and dm) * Move Alien's 'isInline' code to ve.dm.Node.static.isHybridInline * Move definition of ce.AlienBlock/InlineNode inside ce.Aline.js file to match dm.AlienBlock/InlineNode and MWTemplate * Duplicate AlienBlock/Inline styles for templates * Create test case for inline templates * Count test cases in ve.dm.Converter.test.js automatically Change-Id: Id9bc7f049ea974dd5e7f8b7a66080939e0948bbd	2013-04-14 02:34:18 +00:00
Catrope	54a232a92b	Allow nodes to handle their own children For data->DOM, this is easy: .toDataElements() can optionally return an array instead of an object, and that will be treated as the data to insert. If this happens, the converter won't descend. The node handler can recursively invoke the converter if it needs to (although I suspect the current implementation is broken when converting block content in an inline context). For DOM->data, this is a bit more complex. The node sets .static.handlesOwnChildren = true; , which triggers the converter to pass a data slice rather than a single data element, and not to descend. The node handler can invoke the converter to recursively convert DOM subtrees to data. ve.dm.Converter (data->DOM): * Renamed createDataElement() to createDataElements() ** .toDataElement() may return element or array, handle this * Renamed childDataElement to childDataElements, is now an array * Actually alienate if .toDataElement() returns null ** Shockingly, this claimed to be supported before but wasn't * Rather than pushing to data, concat to it ** Add closing if needed * Don't descend if .toDataElement() returned an array of length >1, or if the node has .handlesOwnChildren = true ve.dm.Converter (DOM->data): * Split getDomSubtreeFromData() and getDomFromData() * When converting a node that handles its own children, pass in a data slice and skip over that data Change-Id: I196cb4c0895cbf0b428a189adb61b56565573ab3	2013-04-11 22:41:18 +00:00
Catrope	1b5a376c28	Allow hybrids across Model subclasses A node could already implement a toDataElements() function that returns a data element of another node type, but it couldn't return an annotation or a meta item. This is fixed now, and any dm.Model subclass can now morph into any other dm.Model subclass. I didn't originally plan to do this today at all, but doing this now makes my upcoming converter changes easier. Surprise feature! Change-Id: Ief6ac302094df084221a5a97c32a522b929c2960	2013-04-11 11:12:44 -07:00
Catrope	76b080dce1	Pass the converter object to the node handler in toDataElement() This will allow node handlers to recursively invoke getDataFromDomRecursion() Change-Id: I12cd4b31614a549bfbe8fbdc7d0607ece32aa98a	2013-04-11 11:12:44 -07:00
Catrope	daaf255f13	Make getDataFromDomRecursion() use a context stack to pass context info This will allow toDataElement() functions to just call this function with a DOM element, rather than having to have all the recursion context data to pass in. Also expose this information using getters. Change-Id: I89574c42385267e08704f018c0892d63014376a6	2013-04-11 11:12:39 -07:00
Ed Sanders	62c06d0253	Create GeneratedContentNode which can store rendered HTML in IV store AlienNode is now a subclass of GCNode, but doesn't use the IV store yet. Bug: 46571 Change-Id: If0717afdf557a2aa681d1bae3a6e98299631091a	2013-04-10 19:34:19 +01:00
Catrope	27875c8220	Reduce code duplication for annotation rendering ve.dm.Converter and ve.ce.ContentBranchNode were duplicating a fair bit of logic for annotation rendering. Moved the annotation opening and closing logic into ve.dm.Converter.openAndCloseAnnotations, and implemented both annotation rendering code paths in terms of that function with callbacks for caller-specific behavior. Change-Id: I7cba7d2fda7002287b07949a1b8120ba80bfe854	2013-04-09 23:38:03 +00:00
Catrope	0b55bb8cdc	Move common Node/Annotation/MetaItem code into ve.dm.Model ve.dm.Model is now the common base class for these three. ve.dm.Node inherited from ve.Node before, so it now uses it as a mixin instead. This required changing ve.Node's usage of ve.EventEmitter from inhertiance to a mixin as well, because inherited methods apparently don't get mixed in correctly. * Change annotation terminology from linmodAnnotation to element for consistency with Node, MetaItem and Model * Reimplement getClonedElement() in Node for .internal treatment Change-Id: Ifd3922af23557c0b0f8984d36b31c8a1e2ec497e	2013-04-09 12:05:05 -07:00
Catrope	2eb0d2a6b2	Great Annotation Refactor of 2013 This changes the annotation API to be the same as the node API, sans a few boolean flags that don't apply. The APIs were different, but there was really no good reason why, so this makes things simpler for API users. It also means we'll be able to factor a bunch of things out because they're now duplicated between nodes, meta items and annotations. Linear model annotations are now objects with 'type' and 'attributes' properties (rather than 'name' and 'data'), for consistency with elements. They now also contain html/0/* attributes for HTML attribute preservation, which obsoletes the htmlTagName and htmlAttributes properties. dm.Annotation subclasses take a reference to such an object and implement conversion using .static.toDataElement and .static.toDomElements just like nodes do. The custom .getHash() functions are no longer necessary because of the way HTML attribute preservation was reimplemented. CE rendering has been moved out of dm.Annotation (it never made sense to have CE rendering functions in DM classes, this was bothering me) and into separate ce.Annotation subclasses. These are very similar to CE nodes in that they have a this.$ generated based on something in the DM; the main difference is that nodes listen to events and update themselves, whereas annotations are static and are simply destroyed and rebuilt when they change. This change also adds whitelisted HTML attribute rendering for annotations, as well as class="ve-ce-FooAnnotation" attributes. Now that annotation classes produce real DOM nodes rather than weird objects describing HTML tags, we can't generate HTML as a string in ce.ContentBranchNode anymore. getRenderedContents() has been rewritten to be much more similar to the way the converter renders annotations; in fact, significant parts of it were copied from the converter, so that should be factored out in the future. This change actually fixes an annotation rendering discrepancy between ce.ContentBranchNode and dm.Converter; see the diff of ve.ce.ContentBranchNode.test.js. ve.ce.MWEntityNode.js: * Remove stray property ve.dm.MWExternalLinkAnnotation.js: * Store 'rel' attribute ve.dm.TextStyleAnnotation.js: * Put all the conversion logic in the abstract base class ve.dm.Converter.js: * Also feed annotations through getDomElementsFromDataElement() and createDataElement() ve.dm.Node.js: * Fix undocumented property ve.ce.ContentBranchNode.test.js: * Add descriptive messages for each test case * Compare DOM trees, not HTML strings * Compare without all the class="ve-ce-WhateverAnnotation" clutter ve.ui.LinkInspector.js: * Replace direct .getHash() calls (evil!) with ve.getHash() Bug: 46464 Bug: 44808 Change-Id: I31991488579b8cce6d98ed8b29b486ba5ec38cdc	2013-04-08 18:10:16 -07:00
Ed Sanders	fdf30b1ac8	Store data in LinearData class with an index-value store for objects Created an IndexValueStore class which can store any object and return an integer index to its hash map. Linear data is now stored in ve.dm.LinearData instances. Two subclasses for element and meta data contain methods specific to those data types (ElementLinearData and MetaLinearData). The static methods in ve.dm.Document that inspected data at a given offset are now instance methods of ve.dm.ElementLinearData. AnnotationSets (which are no longer OrderedHashSets) have been moved to /dm and also have to be instantiated with a pointer the store. Bug: 46320 Change-Id: I249a5d48726093d1cb3e36351893f4bff85f52e2	2013-03-30 10:06:34 +00:00
Catrope	a835c03bc1	Change MetaNodes to MetaItems Rather than meta-things being special kinds of nodes, they are now a separate class of things (MetaItems) along with Nodes and Annotations. * Created a generic ve.dm.MetaItem that meta items inherit from. There will be actual instances of this class as well in the upcoming meta group code. * Renamed MetaNode to AlienMetaItem, MWMetaNode to MWMetaItem, 'metaBlock'/'metaInline' to 'alienMeta' * Created a MetaItemFactory, handle meta items in the ModelRegistry * Kill ve.dm.Node.static.isMeta, now obsolete ve.dm.Converter: * Pass in the MetaItemFactory * Look up data element types in the ModelRegistry rather than the NodeFactory, because they can be either nodes or meta items * Document createDataElement() and make explicit that modelClass can be either a node or a meta item * Handle meta items in getDataFromDom() * In getDomFromData(), check the MetaItemFactory as well as the NodeFactory Change-Id: I893709c6f3aa00f85c1b905b70f9f4e597bdeada	2013-03-14 23:35:50 -07:00
Ed Sanders	9a7b8aacf8	Only unwrap { generated: wrapper } based on context. Wrapper paragraphs should only be unwrapped if they are the first element in their parent - or if there is a block level element separating them from the previous unwrapped paragraph. Empty paragraphs should only be unwrapped if they are empty and the last element in their parent. Also in this commit is a simple test for IndentationAction.decrease(). Bug: 45590 Change-Id: I1f47d12db6d57d984fd4607f667a3b62c53f3dd6	2013-03-13 00:42:16 +00:00
Catrope	1463d03f44	Change one last .storeHTMLAttributes to .storeHtmlAttributes Change-Id: I161b2e8bf22c3784ca660ab0961a528b23601022	2013-02-22 16:13:47 -08:00
Catrope	04c72f6871	Add MWMetaNode to clean up <meta>/<link> hack in the converter Change-Id: I4c69bff4981eef78415b43d31c3fd2ee271450ef	2013-02-22 15:21:40 -08:00
Catrope	2e36f1542b	(bug 45062) Implement the new node API in the converter This changes the node API to work with multiple elements, so we can support about groups. Instead of passing in and returning single DOM elements, we use arrays of DOM elements. ve.dm.Converter: * Pass modelRegistry into the constructor * Remove onNodeRegister handler and its data * Remove getDataElementFromDomElement() and getDataAnnotationFromDomElement(). Most logic moved into getDataFromDom(), some into createDataElement() * Remove createAlien(), replaced with createDataElement( ve.dm.AlienNode, ... ) * Replace doAboutGrouping() (which wrapped about groups) with getAboutGroup() (which returns an array of the nodes in the group) * Put in a hack so <meta>/<link> elements with an mw: property aren't alienated * Remove about group wrapping behavior in favor of just outputting multiple nodes ve.dm.AlienNode.js: * For multi-element aliens, only choose inline if all elements are inline, not just the first one ve.dm.example.js: * Add html/0 stuff for meta nodes * Fix test case to reflect new alien behavior Change-Id: I40dcc27430f778bc00a44b91b7d824bfb2718be6	2013-02-22 15:21:40 -08:00
Catrope	5e16141750	Change context.wrapping to context.inWrapper It's what the docs say it should be, so let's call it that :) Change-Id: Ia10861ce77872243beb3d3a1886877824103f6c9	2013-02-22 15:21:40 -08:00
Catrope	bbe3783d58	Add .static.storeHtmlAttributes Defaults to true, but set to false, so we don't do redundant work for aliens. Change-Id: If35db3a67afd78124b4b2b46bb78ad60cbac46f5	2013-02-22 15:21:31 -08:00
James D. Forrester	82114467f1	Bump copyright notice year range to -2013 over -2012 199 files touched. Whee! Change-Id: Id82ce4a32f833406db4a1cc585674f2bdb39ba0d	2013-02-19 15:37:34 -08:00
Catrope	3035c311ec	Make the converter work with full HTML documents rather than fragments The Parsoid output will also be expected to be a full HTML document. For backwards compatibility, we allow for the Parsoid output to be a document fragment as well. We don't send a full document back yet, also for b/c -- we'll change this later once Parsoid has been updated in production. ve.dm.Converter.js: * Make getDataFromDom() accept a document rather than a node Split off the recursion (which does use nodes) into its own function For now we just convert the <body>. In the future, we'll want to do things with the <head> as well * Pass the document around so we can use it when creating elements * Make getDomFromData() return a document rather than a <div> ve.init.mw.Target.js: * Store a document (this.doc) rather than a DOM node (this.dom) * Pass around documents rather than DOM nodes * Detect whether the Parsoid output is an HTML document or a fragment using a hacky regex * When submitting to Parsoid, submit the innerHTML of the <body> ve.init.mw.ViewPageTarget.js: * s/dom/doc/ * Store body.innerHTML in this.originalHtml ve.Surface.js: * s/dom/doc/ demos/ve/index.php: * Don't wrap HTML in <div> * Pass HTML document rather than DOM node to ve.Surface ve.dm.Converter.test.js: * Construct a document from the test HTML, rather than a <div> ve.dm.example.js: * Wrap the HTML in the converter test cases in <body> tags to prevent misinterpretation (HTML fragments starting with comments, <meta>, <link> and whitespace are problematic) Change-Id: I82fdad0a099febc5e658486cbf8becfcdbc85a2d	2013-02-19 10:38:39 -08:00
Catrope	591f2e7396	Change the HTML attribute prefix from html/ to html/0/ This means that <p data-foo="bar"> will now be converted to a paragraph with attributes {"html/0/data-foo":"bar"} rather than {"html/foo":"bar"} This paves the way for multi-element node (about group) handling in the node API: nodes representing multiple DOM elements will have html/i/attr to represent an attribute of the i'th DOM element. Change-Id: Iea52bdccd721942ca708c8f9f47e934524809845	2013-02-06 12:00:43 -08:00
Catrope	57ad316988	Fix bug where inline nodes didn't trigger wrapping When encountering an inline node (i.e. content node that's not a text node) within a branch node that's not a content branch node, the converter should start a wrapper. But it doesn't do this, it only opens wrappers for text nodes and annotations. Fixed this in the converter, added a test for it, and fixed an existing test that asserted the broken behavior. Change-Id: I6e143e21e68b68f0d85b8772e24a2d3a5d465410	2013-02-01 16:06:17 -08:00
Catrope	a97f777685	Introduce context object in getDataFromDom() * Introduce context object as specified for ve.dm.Node.static.toDataElement() * Remove wrapping variable in favor of context.wrapping * Remove wrappingIsOurs in favor of context.canCloseWrapper * Introduce originallyExpectingContent and use it to repopulate context.expectingContent after closing a wrapper * Replace most uses of branchHasContent with context.expectingContent ** Except for two cases where we need originallyExpectingContent These changes fix a case where a metaBlock was generated in an inline position. Updated the tests to reflect this. Change-Id: I6baf6053f8a3a0b7d91487f812b9235a7b2b3db1	2013-01-31 15:00:00 -08:00
Catrope	1927330352	Use AnnotationSet rather than array in getDataFromDom() Change-Id: Ie5f1f94bd62739ccf74c027dd0ba418fe2d91fa6	2013-01-31 15:00:00 -08:00
Catrope	99df776543	Allow matchTagNames = null in ve.dm.Converter This won't usefully register the node with the converter right now, but we need to allow this because the ModelFactory tests will need to have stub nodes with tag-only matches. Change-Id: I023cc8ff647363ab55c73dff39b17ca47e9e6681	2013-01-22 17:45:43 -08:00
Catrope	aa372b6c16	Actually use this.nodeFactory and this.annotationFactory in ve.dm.Converter Change-Id: I138a437d2e64577ad905ff70ecedf1eb7e6c8360	2013-01-22 15:55:11 -08:00
Catrope	de6193734d	Add annotation-like static properties to nodes Add static properties for matching, data<->DOM conversion, and name. Use matchTagNames, toDataElement and toDOMElement. name isn't used yet. Change-Id: I5e7df3303bbd65e6968e931b568c23d76003a9a4	2013-01-18 14:51:40 -08:00
Trevor Parscal	8d33a3de0d	Major Documentation Cleanup * Made method descriptions imperative: "Do this" rather than "Does this" * Changed use of "this object" to "the object" in method documentation * Added missing documentation * Fixed incorrect documentation * Fixed incorrect debug method names (as in those VeDmClassName tags we add to functions so they make sense when dumped into in the console) * Normalized use of package names throughout * Normalized class descriptions * Removed incorrect @abstract tags * Added missing @method tags * Lots of other minor cleanup Change-Id: I4ea66a2dd107613e2ea3a5f56ff54d675d72957e	2013-01-16 15:37:59 -08:00
jenkins-bot	d349f18d4b	Merge "(bug 43056) Inline tags like <span> are block-alienated sometimes"	2013-01-08 20:41:28 +00:00
Timo Tijhof	b11bbed7a6	JSDuck: Generated code documentation! See CODING.md for how to run it. Mistakes fixed: * Warning: Unknown type function -> Function * Warning: Unknown type DOMElement -> HTMLElement * Warning: Unknown type DOM Node -> HTMLElement * Warning: Unknown type Integer -> Mixed * Warning: Unknown type Command -> ve.Command * Warning: Unknown type any -> number * Warning: Unknown type ve.Transaction -> ve.dm.Transaction * Warning: Unknown type ve.dm.AnnotationSet -> ve.AnnotationSet * Warning: Unknown type false -> boolean * Warning: Unknown type ve.dm.AlienNode ve.dm doesn't have a generic AlienNode like ve.ce -> Unknown type ve.dm.AlienInlineNode\|ve.dm.AlienBlockNode * Warning: Unknown type ve.ve.Surface -> ve.ce.Surface * ve.example.lookupNode: -> Last @param should be @return * ve.dm.Transaction.prototype.pushReplace: -> @param {Array] should be @param {Array} * Warning: ve.BranchNode.js:27: {@link ve.Node#hasChildren} links to non-existing member -> (removed) * Warning: ve.LeafNode.js:21: {@link ve.Node#hasChildren} links to non-existing member -> (removed) Differences fixed: * Variadic arguments are like @param {Type...} [name] instead of @param {Type} [name...] * Convert all file headers from /** to /! because JSDuck tries to parse all /* blocks and fails to parse with all sorts of errors for "Global property", "Unnamed property", and "Duplicate property". Find: \/\\([^@]+)(@copyright) Replace: /!$1$2 Indented blocks are considered code examples. A few methods had documentation with numbered lists that were indented, which have now been updated to not be intended. * The free-form text descriptions are parsed with Markdown, which requires lists to be separated from paragraphs by an empty line. And we should use `backticks` instead of {braces} for inline code in text paragraphs. * Doc blocks for classes and their constructor have to be in the correct order (@constructor, @param, @return must be before @class, @abstract, @extends etc.) * `@extends Class` must not have Class {wrapped} * @throws must start with a {Type} * @example means something else. It is used for an inline demo iframe, not code block. For that simply indent with spaces. * @member means something else. Non-function properties are marked with @property, not @member. * To create a link to a class or member, in most cases the name is enough to create a link. E.g. Foo, Foo.bar, Foo.bar#quux, where a hash stands for "instance member", so Foo.bar#quux, links to Foo.bar.prototype.quux (the is not supported, as "prototype" is considered an implementation detail, it only indexes class name and method name). If the magic linker doesn't work for some case, the verbose syntax is {@link #target label}. * @property can't have sub-properties (nested @param and @return values are supported, only @static @property can't be nested). We only have one case of this, which can be worked around by moving those in a new virtual class. The code is unaltered (only moved down so that it isn't with the scope of the main @class block). ve.dm.TransactionProcessor.processors. New: * @mixins: Classes mixed into the current class. * @event: Events that can be emitted by a class. These are also inherited by subclasses. (+ @param, @return and @preventable). So ve.Node#event-attach is inherited to ve.dm.BreakNode, just like @method is. * @singleton: Plain objects such as ve, ve.dm, ve.ce were missing documentation causing a tree error. Documented those as a JSDuck singleton, which they but just weren't documented yet. NB: Members of @singleton don't need @static (if present, triggers a compiler warning). * @chainable: Shorthand for "@return this". We were using "@return {classname}" which is ambiguous (returns the same instance or another instance?), @chainable is specifically for "@return this". Creates proper labels in the generated HTML pages. Removed: * @mixin: (not to be confused with @mixins). Not supported by JSDuck. Every class is standalone anyway. Where needed marked them @class + @abstract instead. Change-Id: I6a7c9e8ee8f995731bc205d666167874eb2ebe23	2013-01-05 01:16:32 +01:00
Catrope	7bcf35e0e8	(bug 43056) Inline tags like <span> are block-alienated sometimes This happens when the <span> is the start of unwrapped content. The converter logic to look at the tag name in wrapping mode doesn't kick in because we're not yet in wrapping mode at that point. The core issue was that previously, we relied on the document structure/state to choose between alienBlock and alienInline, and only used the tag name where the document structure was ambiguous (wrapping). Changed this to be the other way around: we now rely primarily on the tag name, and if that doesn't match what we expect based on the document structure, we work around that if possible. Specifically: * inline tag in our wrapper --> inline alien * block tag in our wrapper --> close wrapper, block alien * inline tag in wrapper that's not ours --> inline alien * block tag in wrapper that's not ours --> inline alien * inline tag in structural location --> open wrapper, inline alien * block tag in structural location --> block alien * inline tag in content location --> inline alien * block tag in content location --> inline alien only in the fourth and the last case do we need to use the "wrong" alien type to preserve document validity, and it will always be inline where block was expected, which should reduce UI issues. The condensed version of the above, which is used in the code, is: * If in a non-wrapper content location, use inline * If in a wrapper that's not ours, use inline * Otherwise, decide based on tag name * Open or close wrapper if needed ve.dm.Converter: * Replace isInline logic in createAlien() with the above * Factor out code to start wrapping (was duplicated) into startWrapping() * Call startWrapping() if createAlien() returns an alienInline and we're in a structural location Tests: * Add test cases with aliens at the start and end of unwrapped content ** The first one failed prior to these changes and now passes, the second one was already passing * Fix about group test case, was exhibiting the bug that this commit fixes Change-Id: I657aa0ff5bc2b57cd48ef8a99c8ca930936c03b8	2012-12-22 12:27:11 +01:00
Timo Tijhof	4fa57b469a	Phase out $.toJSON, use JSON.stringify directly. Although $.toJSON optimises heavily for modern browsers (it becomes a direct reference to JSON.stringify), we still load the extra plugin. JSON is specified as part of ECMAScript 5, but most browsers supported this one before they supported the rest of ES5. http://caniuse.com/#search=JSON Cut off for native JSON is IE7, Firefox 3.0 (3.6 supports it) and Safari 3. Not any of our concern as VE will most likely never support those (certainly not at this point in time, and less likely as time goes on). Change-Id: I4e8f26ac94763fa38d29e41264de0247f53a21e5	2012-12-13 01:33:46 +01:00
Catrope	045b597253	Fix the "list of US Presidents" bug I noticed this bug on [[List of Presidents of the United States]]. When there's HTML that looks like "<td>Foo\n<meta/></td>", the converter will collect the newline in wrappedWhitespace, then attempt to splice it out and store it in internal data. But instead, it ends up splicing out the /metaBlock element, which causes strange unbalanced input, which causes an empty table in the node tree. Change-Id: I79ed2fa9a834cc42759c7d21250d8842f563d38f	2012-12-11 11:23:31 -08:00
Catrope	085a6f0985	(bug 42487) Don't crash the converter for "<span>\n<p>Foo</p></span>" The converter was misbehaving when handling <p>s inside <span>s. This can't be expressed in the linmod, but it would try to anyway. <span><p> would result in too many paragraph closing elements, leading to an exception in ve.dm.Document complaining about unbalanced input. <span>\n<p> would result in an exception in the converter itself while trying to perform whitespace preservation on the newline. This change makes the converter detect these scenarios and alienate the offending node. So <span><p>Foo</p></span> converts to a wrapper paragraph containing an alienInline whose HTML is "<p>Foo</p>" and which is annotated with a TextStyleSpanAnnotation. ve.dm.Converter.getDomFromData(): * Change the criteria for alienBlock vs alienInline Only infer from the node type if we're in wrapping mode AND we're at the same level where the wrapping started (wrappingIsOurs). If the latter isn't the case, we can't split the wrapper in the block case because we're at the wrong level. Use alienInline not only if the branch is a content branch, but also if there are active annotations. This catches e.g. <li><b><p> (and generally <span><p> on the top level). * Before converting a child element, check that the child isn't "bad". Bad children are non-content children in content branches, and non-content children encountered within a wrapper that we can't split. Only good children are converted, and bad children are alienated (cue Santa/Sinterklaas jokes). * Add childIsContent and rename branchIsContent to branchHasContent Change-Id: If420ae80ab0777424a9a5517335ef9d0170e87ae	2012-12-05 17:20:07 -08:00
Catrope	e95cc34978	(bug 42469) Leading newlines in <pre>s get eaten HTML DOM has annoying behavior for <pre>s where .innerHTML eats the first newline in a <pre>. Work around this by explicitly adding a newline in the data->DOM converter if the <pre> already contained a newline. There is a separate bug in Parsoid that causes the newline to be lost anyway, filed as bug 42666 Change-Id: Ia26cd4a4c61afbe439b0562deb7f24ee8b8147d7	2012-12-03 17:14:33 -08:00
Catrope	e123a39b4e	Handle annotated inline nodes in the converter Was broken both on the way in and on the way out. * Move alien restoration (data->DOM) out of the main getDomFromData() function and into getDomElementFromDataElement(). This means the comment about District 9 is gone (sniff), but moving this here ensures all code paths hit it (previously, it was assumed annotated nodes could never be aliens). * In the DOM->data converter, add annotation application to getDataElementFromDomElement() (for content nodes) and createAlien() (for aliens). Previously, these nodes would not get annotations. ** ve.AnnotationSet doesn't have a constructor that takes an array, we should fix that. Change-Id: I65f8e9a322111ca3af275bf9997b0b1e7ee93769	2012-11-27 14:41:40 -08:00
Inez Korczyński	a9082e6dde	Only apply HTML attributes to DOM nodes that are "safe" * Added whitelist argument to setDomAttributes which allows filtering of attributes being set * Added prefix argument to ve.dm.Node.getAttributes to allow extracting a subset of attributes by name prefix * Added a whitelist to ve.ce.Node which was extracted from MediaWiki's Sanitizer class * Replaced attribute copying code with a call to setDomAttributes using the whitelist argument, passing in attributes from a call to ve.dm.Node.getAttributes using the prefix argument Also… * Removed comment in constructor of ve.ce.Node, documentation for properties is usually in the getters/setters, and already was in this case * Renamed ve.setDOMAttributes to ve.setDomAttributes * Renamed ve.getDOMAttributes to ve.getDomAttributes * Renamed ve.getDOMText to ve.getDomText * Renamed ve.getDOMHash to ve.getDomHash * Updated all callers of renamed methods Change-Id: Id556172d5d18ea431044b9d402400e1f0e67a293	2012-11-27 14:34:29 -08:00
Trevor Parscal	b6139ba65e	Merge "(bug 42124) Store comments in the meta-linmod"	2012-11-21 22:12:41 +00:00
Catrope	bf7b243627	(bug 42121) Change markers lost for first paragraph on new page When editing a new page, or loading an empty page into the editor, the converter generates a paragraph so the document isn't completely empty. This paragraph is then unwrapped on the way out, potentially destroying change markers and generally producing strange HTML output. Mark this paragraph with generated=empty rather than generated=wrapper, and only unwrap it on the way out if it's still empty. This means we cleanly round-trip empty documents (and empty list items and the like), but if the user enters text, we create a paragraph like we're supposed to. Change-Id: Id0241221a67b769445676b833b5741320d99ea5f	2012-11-21 13:54:52 -08:00
Catrope	662880605c	(bug 42119) Handle alienation in wrapping mode properly When alienating in wrapping mode, we need to look at the type of tag to decide whether to create a wrapped alienInline, or to interrupt the paragraph for an alienBlock. This was being done just fine for the general alienation case (unrecognized tag), but not for the special cases (mw:unrecognized, about groups). * Centralize the logic for ending a wrapper in stopWrapping() * Move the wrapping-contingent block/inline detection logic into createAlien() * Simplify the terrible if statement to decide whether a future decision requires us to stop wrapping. Instead, detect the cases in each code path separately and call stopWrapping() as appropriate * Add tests Change-Id: I4054584ae05e7d5daa71edead3e6a6588cf5d3bb	2012-11-21 13:42:13 -08:00
Catrope	1234a702c9	(bug 42218) Add MWEntityNode <span typeof="mw:Entity"> tags are now correctly represented in the model, and rendered in CE. There are still issues with cursor movement etc. in CE. Because the prioritization mechanism for annotations vs nodes is broken in the current "node API", I had to hack two special cases for mw:Entity into the converter. I also had to change the converter to ignore the children of inline nodes (this was a legitimate bug, but had never come up before). Change-Id: Ib9f70437c58b4ca06aa09f7272bf51d9c41b18f2	2012-11-20 16:19:55 -08:00
Catrope	3a047e0208	(bug 42124) Store comments in the meta-linmod * Make converter generate meta nodes with 'style': 'comment' * Handle style==='comment' in MetaBlockNode toDOM converter * Add some comments to the meta test case ** Update other tests accordingly * Change getDomElementSummary() to actually assert presence of comment nodes (specifically, all non-text child nodes) Change-Id: Ieef9418f4c47df3541477d9420aa2ab8df6e3df1	2012-11-19 20:01:09 -08:00
Trevor Parscal	2c8411eb62	(bug 41947) Propagate change markers when unwrapping generated nodes Editing the text of a list item results in a change marker on the paragraph within that list item. However, that paragraph usually isn't present in the HTML, so the converter unwraps it when converting back to HTML, and the change markers are lost. Instead, transfer the change markers to the <li>. Change-Id: Id675075d19c08d69bc8e990174841dc393b749fc	2012-11-16 15:39:35 -08:00
Catrope	3acc6cb8f4	Disable change marking by default It's causing problems with Parsoid in production Change-Id: Id47493baafe1ec7f7c0e2bbdb2ea60a82913dfaf	2012-11-14 11:58:32 -08:00
Catrope	d4ea93b872	Add basic support for about groups About groups are HTML structures like the following: <div about="#mwt1">....</div> <span about="#mwt1">...</span> <div about="#mwt1">...</div> When about groups are alienated, they are now merged into one alien node, rather than producing a separate alien node for each sibling. This is very basic about group handling, because it only works for groups of directly adjacent siblings (text nodes are permitted in between, but nothing else) assumes all about groups are aliens (which is currently true). * Before processing an element in the DOM->data converter, perform about grouping on its children. This temporarily wraps about groups in <div data-ve-aboutgroup="value of about attribute"> * Extended createAlien() to handle single nodes as well as wrappers holding multiple nodes. * In the data->DOM converter, temporarily wrap multi-node aliens in <div data-ve-multi-child-alien-wrapper="true"> . This makes the rest of the algorithm easier. Change-Id: I2df5f62bc222b570fc11a89fe43d353f8363ead8	2012-11-07 18:13:50 -08:00
Catrope	1f01100eb9	Flag pre nodes as having significant whitespace This causes the converter not to strip inner whitespace in them, and causes CE to suppress the whitespace mangling logic that is normally applied (↵ for newlines, ➞ for tabs, alternating  s for spaces). Change-Id: I738a750c91a4ca4836c485e282865bb7525bf30a	2012-11-07 12:10:58 -08:00
Catrope	04a999f991	Add change marking for Parsoid's benefit * Add map of change markers per offset to Transaction * Map is populated by TransactionProcessor * Markers are reversed on rollback * Removals aren't marked, Parsoid can detect these using DSR discontinuities Change-Id: I2290886ab411c6ad6162044ed85c091313613e51	2012-11-06 10:11:11 -08:00
Catrope	84e598953a	Wrap inline elements properly The HTML "1<br/>2" was being converted to a linmod that looked like "<p>1</p><br></br><p>2</p>". This commit fixes the wrapping logic such that the result is "<p>1<br></br>2</p>" instead. In general, inline nodes (content nodes) should not interrupt the wrapping, but block nodes should. This creates a problem for alien nodes: normally, we determine whether an alien node is a block alien or an inline alien based on context, but if we're in wrapping mode we're unsure of the context. We can't tell the difference between "1<tt>Foo</tt>2" (should be wrapped as one, because tt is inline) and "1<figure></figure>2" (1 and 2 should be wrapped separately, because figure is block) using context alone, so in these cases (and ONLY in these cases) we look up whether the HTML tag in question is an inline tag or a block tag and use that to decide. Change-Id: I75e7f3da387dd401d9b93e09a21751951eccbb83	2012-10-17 13:50:29 -07:00
Catrope	735ee449e3	New annotation API: ve.dm.Converter integration The annotation-related code in the converter is greatly simplified because the API itself takes care of almost everything already. Change-Id: Ib48f52bad6b650a05dc4e7ef82db4158c19b3cf5	2012-10-12 15:07:28 -07:00
Trevor Parscal	daa7e76807	Merge "Add a node type for meta nodes"	2012-09-18 18:15:46 +00:00

1 2