행의 다음 부분을 3열 파일의 현재 행에 병합합니다.

Question 1

wordtype게시한 예제 입력에 표시된 것처럼 입력이 및 필드에 대해 정렬된다고 가정합니다 .

$ cat tst.awk
BEGIN { FS=" @@@ "; ORS="" }
{ curr = $1 FS $2 }
curr != prev {
    printf "%s%s", ORS, $0
    prev = curr
    ORS = RS
    next
}
{ printf " ;;; %s", $NF }
END { print "" }

$ awk -f tst.awk file
word0 @@@ type2 @@@ sentence0
word1 @@@ type1 @@@ sentence1 ;;; sentence2 ;;; sentence3
word1 @@@ type2 @@@ sentence4
word2 @@@ type1 @@@ sentence5

위의 코드는 awk를 사용하는 모든 UNIX 시스템의 모든 쉘에서 작동하고 한 번에 한 줄만 메모리에 저장하며 입력과 동일한 순서로 출력을 생성합니다.

Answer

wordtype게시한 예제 입력에 표시된 것처럼 입력이 및 필드에 대해 정렬된다고 가정합니다 .

$ cat tst.awk
BEGIN { FS=" @@@ "; ORS="" }
{ curr = $1 FS $2 }
curr != prev {
    printf "%s%s", ORS, $0
    prev = curr
    ORS = RS
    next
}
{ printf " ;;; %s", $NF }
END { print "" }

$ awk -f tst.awk file
word0 @@@ type2 @@@ sentence0
word1 @@@ type1 @@@ sentence1 ;;; sentence2 ;;; sentence3
word1 @@@ type2 @@@ sentence4
word2 @@@ type1 @@@ sentence5

위의 코드는 awk를 사용하는 모든 UNIX 시스템의 모든 쉘에서 작동하고 한 번에 한 줄만 메모리에 저장하며 입력과 동일한 순서로 출력을 생성합니다.

Question 2

이것은 awk의 방법입니다.

$ awk -F'@@@' '{ $1 in a ? a[$1][$2]=a[$1][$2]" ;;; "$3 : a[$1][$2]=$3}END{for(word in a){for (type in a[word]){print word,FS,type,FS,a[word][type]} }}' file 
word0  @@@  type2  @@@  sentence0
word1  @@@  type1  @@@  sentence1 ;;;  sentence2 ;;;  sentence3
word1  @@@  type2  @@@  ;;;  sentence4
word2  @@@  type1  @@@  sentence5

또는 더 명확하게 말하면 다음과 같습니다.

awk -F'@@@' '{ 
                if($1 in a){ 
                    a[$1][$2]=a[$1][$2]" ;;; "$3
                }
                else{
                    a[$1][$2]=$3
                }
             }
             END{
                 for(word in a){
                     for (type in a[word]){
                         print word,FS,type,FS,a[word][type]
                     }
                 }
             }' file

이를 위해서는 awkLinux 시스템의 기본 구현인 GNU awk()와 같은 다차원 배열을 이해하는 구현이 필요합니다 gawk.awk

Answer

이것은 awk의 방법입니다.

$ awk -F'@@@' '{ $1 in a ? a[$1][$2]=a[$1][$2]" ;;; "$3 : a[$1][$2]=$3}END{for(word in a){for (type in a[word]){print word,FS,type,FS,a[word][type]} }}' file 
word0  @@@  type2  @@@  sentence0
word1  @@@  type1  @@@  sentence1 ;;;  sentence2 ;;;  sentence3
word1  @@@  type2  @@@  ;;;  sentence4
word2  @@@  type1  @@@  sentence5

또는 더 명확하게 말하면 다음과 같습니다.

awk -F'@@@' '{ 
                if($1 in a){ 
                    a[$1][$2]=a[$1][$2]" ;;; "$3
                }
                else{
                    a[$1][$2]=$3
                }
             }
             END{
                 for(word in a){
                     for (type in a[word]){
                         print word,FS,type,FS,a[word][type]
                     }
                 }
             }' file

이를 위해서는 awkLinux 시스템의 기본 구현인 GNU awk()와 같은 다차원 배열을 이해하는 구현이 필요합니다 gawk.awk

행의 다음 부분을 3열 파일의 현재 행에 병합합니다.

답변1

답변2

관련 정보