awk - Print whole lines, when find duplicate

Question

Ask a Question

Welcome To Ask or Share your Answers For Others

awk - Print whole lines, when find duplicate

asked Jan 31, 2022 in Technique[技术] by 深蓝 (71.8m points)

This is fragment of my input:

DGD3 SOL10
DGD53 SOL15
DGD100 SOL15
DGD92 SOL20
DGD41 SOL22
DGD62 SOL35
DGD13 SOL40
DGD13 SOL40

My expected output

DGD53 SOL15
DGD100 SOL15
DGD13 SOL40
DGD13 SOL40

In my data I have sometimes SOL duplicates (not more than two repetitions not for example three times some SOL in a file but only duplicates). SOL is in my second column ($2). So I need a program which print whole line (DGD and SOL) when I find duplicate SOL ($2). Could you help me?

See Question&Answers more detail:os

与恶龙缠斗过久,自身亦成为恶龙；凝视深渊过久,深渊将回以凝视…

197 views

1 Answer

深蓝 · Answer 1 · 2022-01-31T07:21:34+0000

Adding one more way in awkish style, where to get all value count in first read of Input_file and print all values as per their count in 2nd read. Fair warning this may not be fast as other 2 solutions but should be simple from understanding purposes.

awk '
FNR==NR{
  count[$2]++
  next
}
(count[$2]>1)
' Input_file  Input_file

Categories

awk - Print whole lines, when find duplicate

Please log in or register to add a comment.

Please log in or register to answer this question.

1 Answer

Please log in or register to add a comment.

Just Browsing Browsing

Most popular tags