regexp for html tags with Matlab

Question

I'm looking for a way to use regexp in order to remove all html tags from a string.
So if I have <HTML><b><FONT color="red" size="3">Hello</FONT></b></HTML> I would like to get the hello from it.

I know it will probably look like nested tags, but it's not really, because all I want to do here is to remove anything between two <>.

I'm using Matlab for doing so, but the regexp is the exact same, so feel free to contribute any help.
Thank you.

ilalex · Accepted Answer

My solution is:

>> str='<HTML><b><FONT color="red" size="3">Hello</FONT></b></HTML>';
>> regexprep(str, '<.*?>','')

ans =

Hello

stema · Answer

To match such a tag

<[^>]*>

See online here at Rubular

regexp for html tags with Matlab

Tags:

regex

parsing

tags

matlab

shahar_m

2 Answers

ilalex

stema

Recent Activity

Donate For Us

regexp for html tags with Matlab

Tags:

regex

parsing

tags

matlab

shahar_m

2 Answers

ilalex

stema

Related questions

Recent Activity

Donate For Us