From mboxrd@z Thu Jan 1 00:00:00 1970 Path: news.gmane.io!.POSTED.blaine.gmane.org!not-for-mail From: Stefan Kangas Newsgroups: gmane.emacs.bugs Subject: bug#48211: 28.0.50; eww strips whitespace between elements Date: Mon, 3 May 2021 19:51:06 -0500 Message-ID: References: <87y2cvl6eg.fsf@tcd.ie> Mime-Version: 1.0 Content-Type: text/plain; charset="UTF-8" Content-Transfer-Encoding: quoted-printable Injection-Info: ciao.gmane.io; posting-host="blaine.gmane.org:116.202.254.214"; logging-data="6625"; mail-complaints-to="usenet@ciao.gmane.io" Cc: Lars Ingebrigtsen , 48211@debbugs.gnu.org To: "Basil L. Contovounesios" Original-X-From: bug-gnu-emacs-bounces+geb-bug-gnu-emacs=m.gmane-mx.org@gnu.org Tue May 04 02:52:12 2021 Return-path: Envelope-to: geb-bug-gnu-emacs@m.gmane-mx.org Original-Received: from lists.gnu.org ([209.51.188.17]) by ciao.gmane.io with esmtps (TLS1.2:ECDHE_RSA_AES_256_GCM_SHA384:256) (Exim 4.92) (envelope-from ) id 1ldjIK-0001aQ-B6 for geb-bug-gnu-emacs@m.gmane-mx.org; Tue, 04 May 2021 02:52:12 +0200 Original-Received: from localhost ([::1]:33144 helo=lists1p.gnu.org) by lists.gnu.org with esmtp (Exim 4.90_1) (envelope-from ) id 1ldjIJ-00059q-Dm for geb-bug-gnu-emacs@m.gmane-mx.org; Mon, 03 May 2021 20:52:11 -0400 Original-Received: from eggs.gnu.org ([2001:470:142:3::10]:54968) by lists.gnu.org with esmtps (TLS1.2:ECDHE_RSA_AES_256_GCM_SHA384:256) (Exim 4.90_1) (envelope-from ) id 1ldjIA-00057C-IX for bug-gnu-emacs@gnu.org; Mon, 03 May 2021 20:52:02 -0400 Original-Received: from debbugs.gnu.org ([209.51.188.43]:38164) by eggs.gnu.org with esmtps (TLS1.2:ECDHE_RSA_AES_128_GCM_SHA256:128) (Exim 4.90_1) (envelope-from ) id 1ldjIA-00023K-9Y for bug-gnu-emacs@gnu.org; Mon, 03 May 2021 20:52:02 -0400 Original-Received: from Debian-debbugs by debbugs.gnu.org with local (Exim 4.84_2) (envelope-from ) id 1ldjIA-00016e-8D for bug-gnu-emacs@gnu.org; Mon, 03 May 2021 20:52:02 -0400 X-Loop: help-debbugs@gnu.org Resent-From: Stefan Kangas Original-Sender: "Debbugs-submit" Resent-CC: bug-gnu-emacs@gnu.org Resent-Date: Tue, 04 May 2021 00:52:02 +0000 Resent-Message-ID: Resent-Sender: help-debbugs@gnu.org X-GNU-PR-Message: followup 48211 X-GNU-PR-Package: emacs Original-Received: via spool by 48211-submit@debbugs.gnu.org id=B48211.16200894754222 (code B ref 48211); Tue, 04 May 2021 00:52:02 +0000 Original-Received: (at 48211) by debbugs.gnu.org; 4 May 2021 00:51:15 +0000 Original-Received: from localhost ([127.0.0.1]:49703 helo=debbugs.gnu.org) by debbugs.gnu.org with esmtp (Exim 4.84_2) (envelope-from ) id 1ldjHP-000162-CY for submit@debbugs.gnu.org; Mon, 03 May 2021 20:51:15 -0400 Original-Received: from mail-pl1-f178.google.com ([209.85.214.178]:46702) by debbugs.gnu.org with esmtp (Exim 4.84_2) (envelope-from ) id 1ldjHN-00015u-74 for 48211@debbugs.gnu.org; Mon, 03 May 2021 20:51:13 -0400 Original-Received: by mail-pl1-f178.google.com with SMTP id s20so3857412plr.13 for <48211@debbugs.gnu.org>; Mon, 03 May 2021 17:51:13 -0700 (PDT) X-Google-DKIM-Signature: v=1; a=rsa-sha256; c=relaxed/relaxed; d=1e100.net; s=20161025; h=x-gm-message-state:from:in-reply-to:references:mime-version:date :message-id:subject:to:cc:content-transfer-encoding; bh=yF4ftS2VNtPYLPylv7b6y7fLPuHnJAlC8xbcdjCog3s=; b=Rkkyxa9yK8+gLcZYJgY0wvdgxfaijw3WzAaNtCF/vvkBoYzD9lm3AoX/28TSxTa7zd oobTdfF4FkfFQ/OWuYdXMBj8DEjZl8JlVAjzGzYDWa3k4DVoJR0Rn/Lnk5I0E4K7BA0n 87jW1dOca8p5OPhpUPzeuHjGMMkPus6NbNuXZY5SKz4wNi0cQW3FXV8dBvGRgjMZJWm9 0EHhgQ8Zgh+bXWPDR4/tPNYFM047zTCs94CS5U2xUt3uWaeMO5XXg4vNtRW0S5MzX5B1 98ePo27PjWBBK2N8n7NkotyKBELpx12ewJN1FNOKSBnuf5pVQGomKNlfsBpn+GnxKRhn E9uA== X-Gm-Message-State: AOAM530d355bOTeuhyZHKjHhUWLGdbTi65SayvTJrezMsUy3+bdc8HVd MI2xTPKygTP1zMttrckhksZiyNUPz8xjVAP5u1E= X-Google-Smtp-Source: ABdhPJymChpB0l3tQs1BlJltnQzemjIujJe1lXEwXb11RugSWraMT39rVQ+kwzjJYHfvoi3YKGq5JmCLU3U5wtPvGeE= X-Received: by 2002:a17:902:24d:b029:ee:df5b:286 with SMTP id 71-20020a170902024db02900eedf5b0286mr5324720plc.39.1620089467404; Mon, 03 May 2021 17:51:07 -0700 (PDT) Original-Received: from 753933720722 named unknown by gmailapi.google.com with HTTPREST; Mon, 3 May 2021 19:51:06 -0500 In-Reply-To: X-BeenThere: debbugs-submit@debbugs.gnu.org X-Mailman-Version: 2.1.18 Precedence: list X-BeenThere: bug-gnu-emacs@gnu.org List-Id: "Bug reports for GNU Emacs, the Swiss army knife of text editors" List-Unsubscribe: , List-Archive: List-Post: List-Help: List-Subscribe: , Errors-To: bug-gnu-emacs-bounces+geb-bug-gnu-emacs=m.gmane-mx.org@gnu.org Original-Sender: "bug-gnu-emacs" Xref: news.gmane.io gmane.emacs.bugs:205569 Archived-At: Stefan Kangas writes: > FWIW, the below diff works around this bug for me. > > diff --git a/lisp/net/shr.el b/lisp/net/shr.el > index cbdeb65ba8..3eb3a5bc49 100644 > --- a/lisp/net/shr.el > +++ b/lisp/net/shr.el > @@ -1485,6 +1485,12 @@ shr-tag-tt > ;; The `tt' tag is deprecated in favor of `code'. > (shr-tag-code dom)) > > +(defun shr-tag-mark (dom) > + (shr-generic dom) > + ;; Hack to work around bug in libxml2 (Bug#48211): > + ;; https://gitlab.gnome.org/GNOME/libxml2/-/issues/247 > + (insert " ")) > + > (defun shr-tag-ins (cont) > (let* ((start (point)) > (color "green") Well, I should moderate that statement. It doesn't exactly fix the bug as I'm now getting this instead: 1. f. Unidad ling=C3=BC=C3=ADstica , dotada generalmente de significado= , que se separa de las dem=C3=A1s mediante pausas potenciales en la pronunciaci=C3=B3n y blancos en la escritura . 2. f. Representaci=C3=B3n gr=C3=A1fica de la palabra hablada . 3. f. Facultad de hablar . IOW, whitespace is added even if the following character is punctuation...