webread not yielding actual website

Jakob Sievers

13 Jul. 2022

0 Antworten

Aktualisiert 14 Jul. 2022

9 Ansichten (30 Tage)

Melden Sie sich an, um diese Frage zu beantworten.

Follow Question

Melden Sie sich an, um diese Frage zu beantworten.

Follow Question

Ältere Kommentare anzeigen

0 Stimmen

Hi there. I'm trying to learn how to extract information from websites. As an example, i'm trying to extract text from Facebook posts but webread gives me something which appears to be quite different from what I'm actually seeing on the website. I'm a complete Noob at this particular type of task and so I was hoping I could get some pointers concerning how to get the text, as I see it, rather than some obscured version. Thanks in advance!

3 Kommentare
1 älteren Kommentar anzeigen 1 älteren Kommentar ausblenden

DGM am 14 Jul. 2022

Bearbeitet: DGM am 14 Jul. 2022

Considering the source, I'm going to guess it's dynamic content.

https://www.mathworks.com/matlabcentral/answers/1750720-webread-not-returning-full-html-contents

Without knowing what page and what content exactly is being targeted, it's hard to be sure.

Jakob Sievers am 14 Jul. 2022

@DGM: reading through the references in the thread you're referring to, I think it may be the exact problem that the stuff you're seeing on sites like Facebook is created not by basic HTML but by tons of scripts and such, which webread then is not able to extract.

Is there no way to dig deeper than webread, using matlab? I'd really like to stay on the Matlab platform, which I'm most familiar with, before considering other alternatives

Melden Sie sich an, um zu kommentieren.

Melden Sie sich an, um diese Frage zu beantworten.

Follow Question