wikimedia/mediawiki-extensions-DiscussionTools

mirror of https://gerrit.wikimedia.org/r/mediawiki/extensions/DiscussionTools synced 2024-12-18 19:12:16 +00:00

Author	SHA1	Message	Date
Bartosz Dziewoński	c7723baf72	CommentParser: Replace uses of Title with TitleValue Another small step towards removing the reliance on global state. Change-Id: Ifb4a5bcbef6606d02f1c7aa7385d72822cb0bad0	2022-03-18 18:24:34 +00:00
Bartosz Dziewoński	b68832ace0	Fix parsing of non-English titles in tests We were calling Title::newFromText() before setupEnv(), which meant that the title for each test case was parsed using the default rules for English, rather than the rules for the specified wiki. This only makes a practical difference for tests with self-links. Changed the only such test to demonstrate the fix. Change-Id: I45561f1c9f0d149e2b743f0000b742bf6fc014af	2022-03-18 18:24:07 +00:00
Bartosz Dziewoński	77614a2d02	tests: Fix root node / container handling Since times immemorial, and for reasons lost to history, our test code was adding an extra <div> wrapper before parsing the HTML used for tests. This wasn't a problem, until now, because I want to add some tests for T303396 that need to check that the real wrappers present in some test cases are handled correctly. Changes to test cases mostly remove a leading "0/" from serialized ranges, corresponding to removing the extra wrapper. Change-Id: Ia50e3590538c8cd274b02d2a937ba1a3fbb4ac89	2022-03-10 18:43:58 +01:00
Bartosz Dziewoński	4c29304484	CommentParser: Avoid using a dynamic undeclared property Change-Id: Iefa8dea83bc0d31b9c6b3509189eeaa652dd9ea0	2022-03-08 23:30:11 +00:00
Bartosz Dziewoński	063174e71c	Use `instanceof` for checking for text/element nodes in PHP It is friendlier for static analysis tools like Phan, which can't infer anything from the `->nodeType === …` checks, and we were already using it in most places. Fix newly revealed Phan failures (and one unneeded suppression). Change-Id: Id789f05e16a210f7ba22ca7514587c392fac0741	2022-03-08 23:28:39 +00:00
jenkins-bot	3c91a800ed	Merge "Improve detecting already signed comments"	2022-03-02 14:14:13 +00:00
Ed Sanders	dc8b4e8d4f	Highlight all comments since the oldest in a thread bundle For topic subscriptions, further restrict this to comments in the same thread. Bug: T302014 Change-Id: Ifba218871122901031a891034e709b886fc406da	2022-02-28 21:58:10 +00:00
Bartosz Dziewoński	0ecc8a4c05	Improve detecting already signed comments Previously, we required a signature at the end of the comment. This was a pretty rough heuristic that did not correctly handle many comments that we would consider entirely properly signed in CommentParser (e.g. comments wrapped in formatting like <small>…</small>, comments with a post-scriptum or in parentheses, or comments generated by various templates). Now we process the user input using the same code that adds reply links, and only add a signature when we detect that there really isn't a signature (including template-generated), or if the signature is in the wrong place and would result in the reply link showing up in the wrong place as well (not at the end of the comment). Bug: T278442 Bug: T268558 Bug: T278355 Bug: T291421 Bug: T282983 Change-Id: I46b6110af328ebdf93b7dfc2bd941e04391a1599	2022-02-21 21:21:26 +00:00
Bartosz Dziewoński	85165543f4	CommentParser: Inject a forgotten service Also sort alphabetically. Change-Id: I9e77c4aa1fba930f382e3c4f17ac0504c2f06668	2022-02-21 20:15:54 +01:00
Bartosz Dziewoński	8e44b43df0	Split off ThreadItemSet from CommentParser Goal: ----- Finishing the work from Iadb7757debe000025e52770ca51ebcf24ca8ee66 by changing CommentParser::parse() to return a data object, instead of the whole parser. Changes: -------- ThreadItemSet.php: ThreadItemSet.js: * New data class to access the results of parsing a discussion. Most methods and properties are moved from CommentParser with no changes. CommentParser.php: Parser.js: * parse() returns a new ThreadItemSet. * Remove methods moved to ThreadItemSet. * Placeholder headings are generated slightly differently, as we process things in a different order. * Grouping threads and computing IDs/names is no longer lazy. We always needed IDs/names anyway. * computeId() explicitly uses a ThreadItemSet to check the existing IDs when de-duplicating. controller.js: * Move the code for turning some nodes annotated by CommentFormatter into a ThreadItemSet (previously a Parser) from controller#init to ThreadItemSet.static.newFromAnnotatedNodes, and rewrite it to handle assigning parents/replies and recalculating legacy IDs more nicely. * mw.dt.pageThreads is now a ThreadItemSet. Change-Id: I49bfe019aa460651447fd383f73eafa9d7180a92	2022-02-21 16:22:32 +00:00
Bartosz Dziewoński	4613ae78e7	Change CommentParser into a service Goal: ----- To have a method like CommentParser::parse(), which just takes a node to parse and a title and returns plain data, so that we don't need to keep track of the config to construct a CommentParser object (the required config like content language is provided by services) and we don't need to keep that object around after parsing. Changes: -------- CommentParser.php: * …is now a service. Constructor only takes services as arguments. The node and title are passed to a new parse() method. * parse() should return plain data, but I split this part to a separate patch for ease of review: I49bfe019aa460651447fd383f73eafa9d7180a92. * CommentParser still cheats and accesses global state in a few places, e.g. calling Title::makeTitleSafe or CommentUtils::getTitleFromUrl, so we can't turn its tests into true unit tests. This work is left for future commits. LanguageData.php: * …is now a service, instead of a static class. Parser.js: * …is not a real service, but it's changed to behave in a similar way. Constructor takes only the required config as argument, and node and title are instead passed to a new parse() method. CommentParserTest.php: parser.test.js: * Can be simplified, now that we don't need a useless node and title to test internal methods that don't use them. testUtils.js: * Can be simplified, now that we don't need to override internal ResourceLoader stuff just to change the parser config. Change-Id: Iadb7757debe000025e52770ca51ebcf24ca8ee66	2022-02-19 19:51:57 +01:00
Bartosz Dziewoński	99b5de8038	Split Data class into ResourceLoaderData and LanguageData The Data class contained utilities for two unrelated purposes. Split each half to a separate class. Notably, this improves the signature of the getLocalData() function. Change-Id: Icde615fb9d483fee1f352c34909b37f8ffde8081	2022-02-19 19:37:34 +01:00
Bartosz Dziewoński	ae9f26a9e5	Various code quality tweaks (suggested by PhpStorm) composer.json: * Document required PHP extensions Parser.js: * Remove incorrect param documentation * Fix some typos in comments (missing parentheses) CommentParser.php: * Fix some typos in comments (missing parentheses) ImmutableRange.php: * Remove unused property * Add a `throw` to indicate that code path is unreachable SubscribedNewCommentPresentationModel.php: * Add missing `return false` CommentParserTest.php: * Remove unnecessary pass-by-reference CommentModifierTest.php: * Remove unused variable CommentParserTest.php: * Don't construct Element objects directly. PHP's DOMElement allows it, but Parsoid/Dodo's doesn't, and we use the latter for static analysis. This generates all kinds of confusing warnings. Change-Id: Ia9598ebea0e99830dd485296e94a9d96acc4b258	2022-02-19 19:36:52 +01:00
Bartosz Dziewoński	165ca9b847	Improve CommentModifier::addReply() API for re-use and testing Goal: To be able to re-use or test the transformations we previously performed in addWikitextReply() / addHtmlReply(), without requiring a Comment object or adding the result as a reply. Change-Id: I040c4be9b6b9bddba661f30fd0566f8850673074	2022-02-03 21:12:48 +00:00
Bartosz Dziewoński	b7cbd714ca	Add tests for bullet indentation Bug: T259864 Change-Id: If38016564b67ee7217fe7328b40973aa244ff467	2022-01-14 00:27:04 +00:00
jenkins-bot	7f329ca9a2	Merge "Enable wikis to customize the syntax used for replies"	2022-01-12 21:32:49 +00:00
Ed Sanders	34011b7a07	Parser: Pass in title of page being parsed Will be used to parse selflinks in the future. Change-Id: I2bc29d1c5c69cb6309f582f162f9af7d96ce8913	2022-01-12 21:17:59 +00:00
Ed Sanders	1fed7115f4	Tests: Add original titles to test cases These are not used for anything yet, but soon the parser will want to know the title of the page it is parsing. Change-Id: I02fa5d63fae78f3e92032d93bc27ac5c744faecb	2022-01-12 22:16:03 +01:00
Bartosz Dziewoński	7b1053300a	Enable wikis to customize the syntax used for replies The following values for configuration variables are supported: $wgDiscussionToolsReplyIndentation = 'invisible'; (default) $wgDiscussionToolsReplyIndentation = 'bullet'; Bug: T259864 Change-Id: Icefad79630adc6ed35687498614e6a03ede1451b	2022-01-12 20:54:04 +00:00
jenkins-bot	cd8f426ad4	Merge "Add missing typehints"	2021-12-02 21:13:49 +00:00
Bartosz Dziewoński	f68f91e883	Set $wgUsePigLatinVariant = false while running tests Data used for the tests assumes there are no variants for English, and some tests fail when there are. Correct behavior with language variants is tested using other languages. Change-Id: I348a0ba0389c2a18644ce5e05c7f37d8f26a8c55	2021-12-01 23:25:30 +01:00
Ed Sanders	8e4f08182e	Add missing typehints Change-Id: Ia25c5bea1834a3fdd26f32a9d5ed097789329824	2021-12-01 14:57:09 +00:00
Bartosz Dziewoński	0d57aa9762	Automatic topic subscriptions (on any edit) Bug: T284836 Change-Id: Ia42ad087218fd91a0cdd1664157d1049738e3c01	2021-11-15 22:45:42 +01:00
Ed Sanders	0fba9b0048	Suppress events from comments that are more than 10 minutes old Bug: T290803 Change-Id: Ic0e23f439eef8a1b785f408d4557bec0abe9104b	2021-11-09 16:37:46 +00:00
Ed Sanders	a86d308d66	CommentItem.php: Store timestamp object instead of string We do something similar in CommentItem.js with a moment object. The object can be converted to a string when required. Change-Id: Id7221e9201db0d89c3b771574634c878c9515ca0	2021-11-09 16:37:45 +00:00
Alexander Vorwerk	0935bb1271	MediaWikiTestCase -> MediaWikiIntegrationTestCase MediaWikiTestCase has been renamed to MediaWikiIntegrationTestCase in 1.34. Bug: T293043 Change-Id: I485c5c5f0376ab60cdec49e934c6e7eea8c9feb5	2021-10-12 00:40:27 +02:00
Bartosz Dziewoński	c1f4668806	Change CommentParser and ImmutableRange to use offsets in codepoints instead of bytes The PHP DOM extension measures lengths and offsets in Unicode codepoints. Our PHP code used UTF-8 bytes, causing some offsets to be slightly off. Now it mostly uses Unicode codepoints as well (we're forced to use bytes in a few places, because preg_match returns offsets in bytes). In practice, this had no visible effect to the user. It caused the markers `<span data-mw-comment-end="..."></span>` to be placed at the end of their container instead of the correct position when the timestamp contained multibyte characters (e.g. "ź" in Polish); but the correct position is usually at the end of the container anyway. In the test cases, the only difference is placing these markers before a trailing line break inside `<p>...</p>` tags rather than before it. The patch also accidentally fixes another bug, where element nodes with no children (mostly <img>) were incorrectly excluded when calling cloneContents(), because they were treated as if they were text nodes. Change-Id: Iccdccf1078598f4b62cab96225e9c85a4c0e93ee	2021-09-27 19:04:16 +00:00
Bartosz Dziewoński	a6a547f2b2	Add some tests covering ThreadItem::getHTML() and related methods * ThreadItem::getText * CommentItem::getBodyText (used when generating notifications) * ThreadItem::getHTML (may soon be used in API) * CommentItem::getBodyHTML (may soon be used in API) * ImmutableRange::cloneContents (the common implementation for all of the above) The outputs are only lightly reviewed. This is mostly meant to document the current behavior rather than the expected behavior, to avoid making unintentional changes while refactoring. Change-Id: I14471ee4969aa3d0b5577d9de2a6d4462fab4d09	2021-08-24 07:54:09 +02:00
Bartosz Dziewoński	ad04b24ffd	Create a hidden revision tag for talk page comments Bug: T262107 Depends-On: I21159d03eebaf46ad94f4273ba698a59b8019185 Change-Id: Iceddfaf6a4bcc5e8b5c85c8cd5638bf14aa7db03	2021-08-16 15:42:51 +00:00
Bartosz Dziewoński	47510a22f3	EventDispatcher: Fix ignoring level 3+ headings The code (prior to `d25825a754`) assumed that level 3+ headings would always follow a level 2 heading or the placeholder heading, but we don't generate a placeholder heading if there are no comments in section zero. Add more tests to confirm that comments under level 3+ headings (that are not sub-headings of level 2), and level 1 headings, are ignored when generating notifications, and do not mess with normal headings. Bug: T288775 Change-Id: Ic57b56752a4797cb01234f66e0ed7b849752bd70	2021-08-16 15:42:06 +00:00
Bartosz Dziewoński	b46893eb7d	Remove pointless uses of preserveWhiteSpace property This DOMDocument property has no effect, because we do not use DOMDocument methods for parsing HTML, but rather DOMUtils::parseHTML() provided by Parsoid. Change-Id: I1d9e73e53f2d44f41cf9dcda4f06ac8647671096	2021-08-09 23:45:48 +02:00
jenkins-bot	10c23d0eb1	Merge "Deal with document body consistently"	2021-08-06 03:08:28 +00:00
Bartosz Dziewoński	8de8d80cde	Deal with document body consistently Use `DOMCompat::getBody( ... )` as a nicer getter than `->getElementsByTagName( 'body' )->item( 0 )`. Remove overly defensive checks and redundant annotations on its return value. Since we're dealing with HTML documents throughout, the document body is guaranteed to exist. We previously needed some of them to convince Phan when it thought the body may be null, but this seems to no longer be needed. Change-Id: If7aee7b6adbfa78269c7ba28b26a6eaa21fe935b	2021-08-03 15:12:55 +02:00
Bartosz Dziewoński	80704b6e80	Test cases for interactions with events generated by base Echo Adding test cases in a separate commit to make it easier to review how the test results change after I98fbca8e. * For mentions, the 'mentioned-users' extra parameter is copied to our event (which is then used to avoid duplicate notifications). * For user talk page edit, nothing special happens right now (we use the target page title to avoid duplicate notifications, but this is not apparent from the test case, since page titles are not present). Bug: T281590 Bug: T253082 Change-Id: I153e7735f63f1e2643ed881281d807313cd699c3	2021-08-01 12:27:33 +02:00
Bartosz Dziewoński	78cb03c471	Test cases for comments posted in close succession Adding test cases in a separate commit to make it easier to review how the test results change. As expected, in every case, no notifications are generated right now. Bug: T285528 Change-Id: I25308754112c521d2db8c54ef0c82373456d9e31	2021-08-01 12:27:33 +02:00
C. Scott Ananian	25272e7a4a	Don't refer directly to PHP `dom` extension classes; avoid nonstandard behavior These changes ensure that DiscussionTools is independent of DOM library choice, and will not break if/when Parsoid switches to an alternate (more standards-compliant) DOM library. We run `phan` against the Dodo standards-compliant DOM library, so this ends up flagging uses of non-standard PHP extensions to the DOM. These will be suppressed for now with a "Nonstandard DOM" comment that can be grepped for, since they will eventually will need to be rewritten or worked around. Most frequent issues: * Node::nodeValue and Node::textContent and Element::getAttribute() can return null in a spec-compliant implementation. Add `?? ''` to make spec-compliant results consistent w/ what PHP returns. * DOMXPath doesn't accept anything except DOMDocument. These uses should be replaced with DOMCompat::querySelectorAll() or similar (which end up using DOMXPath under the covers for DOMDocument any way, but are implemented more efficiently in a spec-compliant implementation). * A couple of times we have code like: `while ($node->firstChild!==null) { $node = $node->firstChild; }` and phan's analysis isn't strong enough to determine that $node is still non-null after the while. This same issue should appear with DOMDocument but phan doesn't complain for some reason. One apparently legit issue: * Node::insertBefore() is once called in a funny way which leans on the fact that the second option is optional in PHP. This seems to be a workaround for an ancient PHP bug, and can probably be safely removed. Bug: T287611 Bug: T217867 Change-Id: I3c4f41c3819770f85d68157c9f690d650b7266a3	2021-07-30 18:15:40 -04:00
C. Scott Ananian	5203d30ea6	Use DOMCompat::newDocument() to create a new Document For compatibility with Parsoid's document abstraction (Parsoid may switch to an alternate DOM library in the future), don't explicitly create a new document object using `new DOMDocument`; instead use the Parsoid wrapper `DOMCompat::newDocument()`. This ensures that the Document object created will be compatible with Parsoid. There are a number of other subtle dependencies on the PHP `dom` extension in DiscussionTools, like explicit `instanceof` tests; those will be tweaked in a follow-up patch (I3c4f41c3819770f85d68157c9f690d650b7266a3) since they do not affect correctness so long as Parsoid is aliasing Document to a subclass of the built-in DOMDocument. Similarly, the Phan warnings we suppress do not cause runtime errors (because of the fixes included in c5265341afd9efde6b54ba56dc009aab88eff83c) but phan will be happier once the follow-up patch lands and aligns all the DOM types. Bug: T287611 Depends-On: If0671255779571a91d3472a9d90d0f2d69dd1f7d Change-Id: Ib98bd5b76de7a0d32a29840d1ce04379c72ef486	2021-07-30 18:15:11 -04:00
Bartosz Dziewoński	d0e4aeaecb	Fix notifications when new comment is under subheading The user interface only allows you to subscribe to level 2 headings. But we would generate events for whatever heading was the closest, If it was e.g. level 3, no one would receive that notification. Now we generate events for the closest level 2 heading, or we don't generate the event at all if there isn't one (if the only headings are of level 3 and below, or level 1, or if the comment is added before the first heading on the page). Bug: T286736 Change-Id: Iae99853070e353ab81c9cc29ef1d53c877adfc66	2021-07-24 05:28:10 +02:00
Bartosz Dziewoński	801b57b0f4	Add PHPUnit integration tests for EventDispatcher Bug: T286608 Change-Id: I711483be80d455f4439e96d37844ee4552619a92	2021-07-24 05:28:04 +02:00
libraryupgrader	b0884b177c	build: Updating dependencies composer: * mediawiki/mediawiki-codesniffer: 36.0.0 → 37.0.0 npm: * postcss: 7.0.35 → 7.0.36 * https://npmjs.com/advisories/1693 (CVE-2021-23368) * glob-parent: 5.1.1 → 5.1.2 * https://npmjs.com/advisories/1751 (CVE-2020-28469) * trim-newlines: 3.0.0 → 3.0.1 * https://npmjs.com/advisories/1753 (CVE-2021-33623) Change-Id: I7a71e23da561599da417db3b3077b78d91173bbc	2021-07-22 16:29:04 +00:00
Bartosz Dziewoński	9c8d709b8a	Use placeholder localisation messages in CommentFormatter tests Otherwise they will fail whenever translations are updated (and they are failing right now). Change-Id: I849c57b86d36fb6c7739cc31a74df741e08462f4	2021-06-02 21:46:36 +02:00
libraryupgrader	12fb65b9f1	build: Updating composer dependencies * mediawiki/mediawiki-codesniffer: 35.0.0 → 36.0.0 * php-parallel-lint/php-parallel-lint: 1.2.0 → 1.3.0 Change-Id: I5c152292e83e7f3441e2c08b7d0ad23ac90f194b	2021-05-05 11:14:52 +00:00
Bartosz Dziewoński	475aa80057	Fetch user's topic subscriptions on the page in a single query Previously, we have made a query per each topic on the page. Bug: T281000 Change-Id: I1029e62a65fc191ca37e1178ea7ffc55afafa1b9	2021-04-28 21:54:26 +00:00
Ed Sanders	722a4e5198	Avoid splitting ParserCache on user language Bug: T280295 Change-Id: I87eab83803d24c11db4d723377bf7b40390b2e70	2021-04-21 11:57:30 +00:00
Bartosz Dziewoński	5103e651be	Add tests for CommentFormatter::postprocessTopicSubscription Change-Id: Ief9648b8805fadcc170c54b627eb669cc8b907b6	2021-04-21 11:57:25 +00:00
Bartosz Dziewoński	4bbfe6cb5d	Rename CommentFormatter::addReplyLinks Bug: T280351 Change-Id: I0d7627d63407e11cca6091f78e4d440eec6efa91	2021-04-21 11:24:03 +00:00
Bartosz Dziewoński	42ce942c86	Introduce comment "names" to identify comments across revisions/pages The existing comment IDs can't be used to find the same comment on a different revision or page (when it's transcluded), because they depend on the comment's parent and its position on the page. Comment names depend only on the author and timestamp. The trade-off is that they can't distinguish comments posted within the same minute, or in the same edit, so we will still need the IDs sometimes. Prefer using comment names when replying, if they're not ambiguous. This fixes T273413 and T275821. Heading names depend on the author and timestamp of the oldest comment. This way we don't have to detect changes to the heading text, but we can't distinguish headings without any comments. Bug: T274685 Bug: T273413 Bug: T275821 Change-Id: Id85c50ba38d1e532cec106708c077b908a3fcd49	2021-03-23 16:08:42 +00:00
Bartosz Dziewoński	a103abb8ae	Ignore warnings about legacy IDs in tests Change-Id: I3c74b4e65aac9b84494917547cce7eb6a75995b4	2021-03-18 20:42:03 +01:00
Bartosz Dziewoński	44f2209abf	Trim signatures when added in an empty existing node, too Add unit tests for appendSignature(). Bug: T276612 Change-Id: Ic44c52f4d54492e092f9396c626380e2637b6f0f	2021-03-08 23:38:46 +00:00
Bartosz Dziewoński	5a07139249	CommentFormatterTest: Avoid re-serializing the HTML The code we're testing already produces a string of serialized HTML, no need to parse and re-serialize it. Also, we recently learned that the precise format matters here (T274709), and now this test actually covers the fix for that bug. Follow-up to `5b26e9664b`. As a downside, this test might now spuriously fail if the format of the output of Parsoid's XMLSerializer changes. Hopefully that won't happen too often. Change-Id: I69b514f545e47dcb437fb39a83edb8e2f19ed99b	2021-03-01 21:30:28 +01:00

1 2 3

112 commits