
Nuestra entrada se parece a
2012-04-17 [GBPGBP]
2012-04-13 [GBP GBP]
2012-04-13 [GBP]
2012-04-11 [GBPGBP]
2012-04-11 [GBP GBP]
2012-04-10 [GBPGBP]
2012-04-06 [GBP GBP GBP]
2012-04-17 [GBPGBP]
2012-04-13 [GBP CDN]
2012-04-13 [GBP]
2012-04-11 [GBPCDN]
2012-04-11 [GBP DL DL]
2012-04-10 [PSGBP]
2012-04-06 [PS PS]
Y nos gustaría obtener resultados como
2012-04-17 [GBP]
2012-04-13 [GBP]
2012-04-13 [GBP]
2012-04-11 [GBP]
2012-04-11 [GBP]
2012-04-10 [GBP]
2012-04-06 [GBP]
2012-04-17 [GBP]
2012-04-13 [GBP CDN]
2012-04-13 [GBP]
2012-04-11 [GBPCDN]
2012-04-11 [GBP DL]
2012-04-10 [PSGBP]
2012-04-06 [PS]
Básicamente elimine cualquier cadena repetida entre corchetes. ¿Alguna sugerencia?
Respuesta1
sed -e ': a' -e 's/\(\[[^][]*\)\([A-Z][A-Z][A-Z]*\)\([^][]*\)\2/\1\2\3/' -e 't a'
: a
establece una etiqueta al comienzo del guión.s/\(wibble\)\(foo\)\(bar\)\2/\1\2\3/
reemplaza wibblefoobarfoo por wibblefoobar.[A-Z][A-Z][A-Z]*
coincide con dos o más letrast a
vuelve a la etiquetaa
si el comando anteriors
realizó un reemplazo.