c# - Regex to extract attribute value

Question

Ask a Question

Welcome To Ask or Share your Answers For Others

c# - Regex to extract attribute value

asked Oct 24, 2021 in Technique[技术] by 深蓝 (71.8m points)

What would be a quick way to extract the value of the title attributes for an HTML table:

...
<li><a href="/wiki/Proclo" title="Proclo">Proclo</a></li>
<li><a href="/wiki/Proclus" title="Proclus">Proclus</a></li>
<li><a href="/wiki/Ptolemy" title="Ptolemy">Ptolemy</a></li>
<li><a href="/wiki/Pythagoras" title="Pythagoras">Pythagoras</a></li></ul><h3>S</h3>
...

so it would return Proclo, Proclus, Ptolemy, Pythagoras,.... in strings for each line. I'm reading the file using a StreamReader. I'm using C#.

Thank you.

See Question&Answers more detail:os

与恶龙缠斗过久,自身亦成为恶龙；凝视深渊过久,深渊将回以凝视…

265 views

1 Answer

深蓝 · Answer 1 · 2021-10-23T19:10:22+0000

This C# regex will find all title values:

(?<=title=")[^"]*

The C# code is like this:

Regex regex = new Regex(@"(?<=title="")[^""]*");
Match match = regex.Match(input);
string title = match.Value;

The regex uses positive lookbehind to find the position where the title value starts. It then matches everything up to the ending double quote.

Categories

c# - Regex to extract attribute value

Please log in or register to add a comment.

Please log in or register to answer this question.

1 Answer

Please log in or register to add a comment.

Just Browsing Browsing

Most popular tags