Professional Documents
Culture Documents
Bf6db30ed01 - Amr Spec
Bf6db30ed01 - Amr Spec
Bf6db30ed01 - Amr Spec
0 (2000-01)
Technical Specification
Universal Mobile Telecommunications System (UMTS); Mandatory Speech Codec speech processing functions; AMR Speech Codec Frame Structure (3G TS 26.101 version 3.0.0 Release 1999)
Reference
DTS/TSGS-0426101U
Keywords
UMTS
Office address
650 Route des Lucioles - Sophia Antipolis Valbonne - FRANCE Tel.: +33 4 92 94 42 00 Fax: +33 4 93 65 47 16
Siret N 348 623 562 00017 - NAF 742 C Association but non lucratif enregistre la Sous-Prfecture de Grasse (06) N 7803/88
Internet
secretariat@etsi.fr
Individual copies of this ETSI deliverable can be downloaded from
http://www.etsi.org
If you find errors in the present document, send your comment to: editor@etsi.fr
Important notice This ETSI deliverable may be made available in more than one electronic version or in print. In any case of existing or perceived difference in contents between such versions, the reference version is the Portable Document Format (PDF). In case of dispute, the reference shall be the printing on ETSI printers of the PDF version kept on a specific network drive within ETSI Secretariat.
Copyright Notification No part may be reproduced except as authorized by written permission. The copyright and the foregoing restriction extend to reproduction in all media.
European Telecommunications Standards Institute 2000. All rights reserved.
ETSI
Foreword
This Technical Specification (TS) has been produced by the ETSI 3rd Generation Partnership Project (3GPP). The present document may refer to technical specifications or reports using their 3GPP identities or GSM identities. These should be interpreted as being references to the corresponding ETSI deliverables. The mapping of document identities is as follows: For 3GPP documents: 3G TS | TR nn.nnn "<title>" (with or without the prefix 3G) is equivalent to ETSI TS | TR 1nn nnn "[Digital cellular telecommunications system (Phase 2+) (GSM);] Universal Mobile Telecommunications System; <title> For GSM document identities of type "GSM xx.yy", e.g. GSM 01.04, the corresponding ETSI document identity may be found in the Cross Reference List on www.etsi.org/key
ETSI
Contents
Intellectual Property Rights ............................................................................................................................... 4 Foreword ............................................................................................................................................................ 4 1 2 3
3.1 3.2
4
4.1 4.1.1 4.1.2 4.1.3 4.1.4 4.2 4.2.1 4.2.2 4.2.3 4.3
Annex A: Annex B:
AMR Interface Format 2 (with octet alignment) Tables for AMR Core Frame bit ordering
11 15
History.............................................................................................................................................................. 18
ETSI 3GPP
ETSI 3GPP
Scope
The present document describes a generic frame format for the Adaptive Multi-Rate (AMR) speech codec. This format shall be used as a common reference point when interfacing speech frames between different elements of the 3G system and between different systems. Appropriate mappings to and from this generic frame format will be used within and between each system element. Annex A describes a second frame format which shall be used when octet alignment of AMR frames is required.
Normative references
This TS incorporates by dated and undated reference, provisions from other publications. These normative references are cited at the appropriate places in the text and the publications are listed hereafter. For dated references, subsequent amendments to or revisions of any of these publications apply to this TS only when incorporated in it by amendment or revision. For undated references, the latest edition of the publication referred to applies. [1] [2] [3] TS 26.090: AMR Speech Codec; Speech Transcoding Functions". TS 26.093: AMR Speech Codec; Source Controlled Rate Operation". TS 26.092: AMR Speech Codec; Comfort Noise Aspects".
3
3.1
AMR mode:
AMR codec mode: Same as AMR mode RX_TYPE: TX_TYPE: Classification of the received frame as defined in [2]. Classification of the transmitted frame as defined in [2].
3.2
CRC FQI LSB MSB RX SCR SID TX
Abbreviations
Cyclic Redundancy Check Frame Quality Indicator Least Significant Bit Most Significant Bit Receive Source Controlled Rate operation Silence Descriptor Transmit
For the purposes of the present document, the following abbreviations apply:
ETSI 3GPP
This section describes the generic frame format for both the speech and comfort noise frames of the AMR speech codec. This format is referred to as AMR Interface Format 1 (AMR IF1). Annex A describes AMR Interface Format 2 (AMR IF2). Each AMR codec mode follows the generic frame structure depicted in Figure 1 below. The frame is divided into three parts: AMR Header, AMR Auxiliary Information, and AMR Core Frame. The AMR Header part includes the Frame Type and the Frame Quality Indicator fields. The AMR auxiliary information part includes the Mode Indication, Mode Request, and Codec CRC fields. The AMR Core Frame part consists of the speech parameter bits or, in case of a comfort noise frame, the comfort noise parameter bits. In case of a comfort noise frame, the comfort noise parameters replace Class A bits of AMR Core Frame while Class B and C bits are omitted.
Frame Type (4 bits) Frame Quality Indicator (1 bit) Mode Indication (3 bits) Mode Request (3 bits) Codec CRC (8 bits) Class A bits Class B bits Class C bits
AMR Header AMR Auxiliary Information (for Tandem Free Operation, Mode Adaptation, and Error Detection) AMR Core Frame (speech or comfort noise data)
4.1
4.1.1
Table 1a below defines the 4-bit Frame Type field. Frame Type can indicate the use of one of the eight AMR codec modes, one of four different comfort noise frames, or an empty frame. In addition, three Frame Type Indices are reserved for future use. The same table is reused for the Mode Indication and Mode Request fields which are 3-bit fields each and use only the Frame Type Indices 07 to specify one of the eight AMR codec modes.
ETSI 3GPP
Frame content (AMR mode, comfort noise, or other) 4.75 kbit/s 5.15 kbit/s 5.90 kbit/s 6.70 kbit/s (PDC-EFR) 7.40 kbit/s (IS-641) 7.95 kbit/s 10.2 kbit/s 12.2 kbit/s (GSM EFR) AMR Comfort Noise Frame GSM-EFR Comfort Noise Frame IS-641 Comfort Noise frame PDC-EFR Comfort Noise Frame For future use No transmission/No reception
Table 1a. Interpretation of Frame Type Index used for the Frame Type, Mode Indication, and Mode Request fields of AMR Header
4.1.2
The content of the Frame Quality Indicator field is defined in Table 1b. The field length is one bit which indicates whether the data in the frame contains errors. Frame Quality Indicator (FQI) 0 1 Quality of data Corrupted frame (bits may be used to assist error concealment) Good frame
4.1.3
Table 1c shows how the AMR Header data (FQI and Frame Type) maps to the TX_TYPE and RX_TYPE frames defined in [2].
ETSI 3GPP
Comment
The specific Frame Type Index depends on the bit-rate being used. The specific Frame Type Index depends on the bit-rate being used. The corrupted data may be used to assist error concealment. SID_FIRST and SID_UPDATE are differentiated using Class A bits.
1 0 1 0 1
8 8 8 9-11 9-11 15
Typically a non-transmitted frame or an erased or stolen frame with no data usable to assist error concealment.
Table 1c. Mapping of Frame Quality Indicator and Frame Type to TX_TYPE and RX_TYPE [2], respectively.
4.1.4
Codec CRC
AMR codec frames with Frame Type 0..11 are associated with an 8-bit CRC for error-detection purposes. The Codec CRC field of AMR Auxiliary Information in Figure 1 contains the value of this CRC. These eight parity bits are generated by the cyclic generator polynomial G(x)=D8 + D6 + D5 + D4 + 1 which is computed over all Class A bits of AMR Core Frame. Class A bits for Frame Types 0..7 are defined in Section 4.2.2 (for speech bits) and for Frame Types 8..11 in Section 4.2.3 (for comfort noise bits). When Frame Type Index of Table 1a is 15 the CRC field is not included in the AMR frame.
4.2
This section contains the description of AMR Core Frame of Figure 1. The descriptions for AMR Core Frame with speech bits and with comfort noise bit are given separately.
4.2.1
This section describes how AMR Core Frame carries the coded speech data. The bits produced by the speech encoder are denoted as {s(1),s(2),...,s(K)}, where K refers to the number of bits produced by the speech encoder as shown in Table 2 below. The notation s(i) follows that of [1]. The speech encoder output bits are ordered according to their subjective importance. This bit ordering can be utilized for error protection purposes when the speech data is, for example, carried over a radio interface. Tables B.1 to B.8 in Annex B define the AMR IF1 bit ordering for all the eight AMR codec modes. In these tables the speech bits are numbered in the order they are produced by the corresponding speech encoder as described in the relevant tables of TS 26.090 [1]. The reordered bits are denoted below, in the order of decreasing importance, as {d(0),d(1),...,d(K-1)}. The ordering algorithm is described in pseudo code as for j = 0 to K-1 d(j) := s(tablem(j)+1);
ETSI 3GPP
where tablem(j) refers to the relevant table in Annex B depending on the AMR mode m=0..7. The Annex B tables should be read line by line from left to right. The first element of the table has the index 0.
4.2.2
The reordered bits are further divided into three indicative classes according to their subjective importance. This class division is only informative and provides supporting information for mapping this generic format into specifics formats. The three different importance classes can then be subject to different error protection in the network The importance classes are Class A, Class B, and Class C. Class A contains the bits most sensitive to errors and any error in these bits typically results in a corrupted speech frame which should not be decoded without applying appropriate error concealment. This class is protected by the CRC in AMR Auxiliary Information. Classes B and C contain bits where increasing error rates gradually reduce the speech quality, but decoding of an erroneous speech frame is usually possible without annoying artifacts. Class B bits are more sensitive to errors than Class C bits. The importance ordering applies also within the three different classes and there are no significant step-wise changes in subjective importance between neighboring bits at the class borders. The number of speech bits in each class (Class A, Class B, and Class C) for each AMR mode is shown in Table 2 below. The classification in Table 2 and the importance ordering d(j), together, are sufficient to assign all speech bits to their correct classes. For example, when the AMR codec mode is 4.75, then the Class A bits are d(1)..d(42), Class B bits are d(43)..d(95), and there are no Class C bits. Frame Type Index 0 1 2 3 4 5 6 7 AMR codec mode 4.75 5.15 5.90 6.70 7.40 7.95 10.2 12.2 Total number of bits 95 103 118 134 148 159 204 244 Class A Class B Class C
42 49 55 58 61 75 65 81
53 54 63 76 87 84 99 103
0 0 0 0 0 0 40 60
Table 2. Number of bits in Classes A, B, and C for each AMR codec mode
4.2.3
The AMR Core Frame content for the additional frame types with Type Indices 8-15 in Table 1a are described in this section. These mainly consist of the frames related to Source Controlled Rate Operation specified in [2]. The data content (comfort boise bits) of the additional frame types is carried in AMR Core Frame. The comfort noise bits are all mapped to Class A of AMR Core Frame and Classes B and C are not used. This is a notation convention only and the class division has no meaning for comfort noise bits. The number of bits in each class (Class A, Class B, and Class C) for the AMR comfort noise bits (Frame Type Index 8) is shown in Table 3. The contents of SID_UPDATE and SID_FIRST are divided into three parts (SID Type Indicator, Mode Indication, and Comfort Noise Parameters) as defined in [2].
ETSI 3GPP
10
FQI
Class A
Class B
Class C
8 8 8
1 1 0
39 39 39
1 (= 1) 1 (= 0) 1
0 0 0
0 0 0
Table 3. Bit classification for Frame Type 8 (AMR Comfort Noise Frame) The number of bits in each class (Class A, Class B, and Class C) for the comfort noise bits of Frame Types 9-11 is shown in Table 4. TABLE FOR FURTHER STUDY
4.3
The compound AMR frame is formed as a concatenation of AMR Header, AMR Auxiliary Information, and AMR Core Frame, in this order. The first bit of the AMR frame is the first bit of the Frame Type field. The last bit of the AMR frame is the last bit of AMR Core Frame which is the last bit of speech bits or the last bit of comfort noise bits. Table 5 below summarizes all possible AMR frame format combinations in terms of number of bits in each field. Frame Type Index 0 1 2 3 4 5 6 7 8 9 10 11 12 13 14 15 Frame Type 4 4 4 4 4 4 4 4 4 4 4 4
Not used Not used Not used
Mode Indication 3 3 3 3 3 3 3 3 3 3 3 3
Mode Request 3 3 3 3 3 3 3 3 3 3 3 3
Codec CRC 8 8 8 8 8 8 8 8 8 8 8 8
Class A
Class B
Class C
Total
42 49 55 58 61 75 65 81 39 43 38 37
53 54 63 76 87 84 99 103 0 0 0 0
0 0 0 0 0 0 40 60 0 0 0 0
Table 5. Number of bits for different fields in different AMR frame compositions
ETSI 3GPP
11
Frame Type (4 bits) Class A bits Class B bits Class C bits AMR Core Frame (speech or comfort noise data)
Bit Stuffing
Figure A.1. Frame structure for AMR IF2 Table A.1a shows an example how the AMR 6.7 kbit/s mode is mapped into AMR IF2. The four LSBs of the first octet b0 consist of the Frame Type Index 3 for the AMR 6.7 kbit/s mode (see Table 1a in AMR IF1 specification). This data field is followed by the 134 AMR Core Frame speech bits (d(0)d(133)) which consist of 58 Class A bits and 76 Class B bits as described in Table 2 for AMR IF1. This results in a total of 138 bits and 6 bits are needed for Bit Stuffing to arrive to the closest multiple of 8 which is 144 bits. Octet b0 b1 b2 b17 MSB d(3) d(11) UB d(2) d(10) UB d(1) d(9) UB Octet structure d(0) d(8) UB 0 d(7) UB 0 d(6) UB 1 d(5) d(133) LSB 1 d(4) d(12) d(132)
Table A.1a: Example mapping of the AMR speech coding mode 6.7kbit/s into AMR IF2. The bits used for Bit Stuffing are denoted as UB (for "unused bit"). Table A.1b shows the composition of AMR IF2 frames for all Frame Types in terms of how many bits are used for each field of Figure A.1. Tables A.2 to A.5 specify how the AMR Core Frame comfort noise bits of Frame Types 8-11 are mapped to AMR IF2. Table A.6 specifies the mapping for an empty frame ("no transmission").
Frame Type
Number of bits
Number of Bits in
Bit Stuffing
Number of
ETSI 3GPP
12
Index 0 1 2 3 4 5 6 7 8 9 10 11 12-14 15
in Frame Type 4 4 4 4 4 4 4 4 4 4 4 4 4
AMR Core Frame 95 103 118 134 148 159 204 244 39 43 38 37 0 5 5 6 6 0 5 0 0 5 1 6 7 4
octets (N) 13 14 16 18 19 21 26 31 6 6 6 6 1
Table A.1b: Mapping of the AMR speech coding modes defined in TS 26.090 to mode index bits in transmitted octets.
Transmitted Octets 1
MSB Index of 1st LSF subvector s4 Index of 2nd LSF subvector s12
Mapping of bits
LSB
3 index of 2nd LSF subvector s20 4 index of 3rd LSF subvector s28 SID type bit t1 6 Stuffing bits UB UB UB UB UB smi(2) smi(1) smi(0) s27 s26 s25 s24 s21 index of 3rd LSF subvector s31 s30 s29 Speech_Mode_Indication s23 s22 S19 s18 s17 s16 s15 s14 s13
Table A.2: Mapping of comfort noise bits from TS 26.092 to octets for Frame Type 8 (Bits from s1 to s35 refer to TS 26.092). Definitions of additional descriptor bits needed for the silence descriptor in the table are as follows: SID-type (t1) is {0=SID_FIRST, 1=SID_UPDATE }, Speech Mode Indication (smi(0)- smi(2)) is the AMR codec mode according to the first eight entries in Table 1a.
Transmitted Octets
MSB
Mapping of bits
LSB
ETSI 3GPP
13
1 s4 2 s12 3 Index of 3rd LSF submatrix s20 4 index of 4th LSF submatrix s28 5 s36 Stuffing bits UB s27 s26 index of 5th LSF submatrix s35 s34 s25 s19 s18 s17 s16 sign of 3rd LSF submatrix s24 Index of 2nd LSF submatrix s15 s14 s13 Index of 2nd LSF submatrix s11 s10 s9 s8 index of 1st LSF subMatrix s7 s6 s5 Index of 1st LSF subMatrix S3 s2 s1 mi(3) Mode_Index mi(2) mi(1) mi(0)
s32
index of 4th LSF submatrix s31 s30 s29 index of 5th LSF submatrix s87 s38 s37
s91
s90
s89
s88
Table A.3: Mapping of comfort noise bits from GSM 06.60 (parameters also described in GSM 06.62) to octets for Frame Type 9 (Bits from s1 to s91 refer to GSM 06.60)
Transmitted Octets 1
MSB
Mapping of bits
LSB
Index of 1st LSF subvector cn3 2 Index of 2nd LSF subvector cn11 3 Index of 3rd LSF subvector cn19 cn18 Random Excitation Gain cn27 cn26 Index of 1st RESC parameter cn35 6 Stuffing bits UB UB UB UB UB cn34 cn17 cn16 Cn10 cn9 cn8 cn7 cn2 cn1 cn0 mi(3)
cn25
cn24
cn21
cn20
cn33
cn32
cn31
cn30
UB
Table A.4: Mapping of comfort noise bits from IS641-A to octets for Frame Type 10 (Bits from cn0 to cn37 refer to IS-641-A). Transmitted Octets MSB Mapping of bits LSB
ETSI 3GPP
14
3 index of 2nd LSF subvector s20 4 index of 3rd LSF subvector s28 5 SID type t1 6 Stuffing bits UB UB UB UB UB UB UB SID type t2 s35 s34 frame energy s33 s32 s31 s30 s27 s26 s25 s24 s23 s22 s21 Index of 3rd LSF subvector s29 s19 s18 s17 s16 s15 s14 s13
Table A.5: Mapping of comfort noise bits from ARIB xx to octets for Frame Type 11 (Bits from s1 to s35 refer to ARIB xx). Definition of additional descriptor bits needed for the table is as follows: SID-type is {0=POST0, 1=POST1(SID_UPDATE), 2=PRE, 3=POST1_BAD }, where LSB of SID_type is t1 and MSB of SID-type is t2.
Transmitted Octets 1
MSB
Frame content
LSB
Table A.6: Mapping of the "no transmission" frame (Frame Type 15) to octets.
ETSI 3GPP
15
Table B.1: Ordering of the speech encoder bits for the 4.75 kbit/s mode: table1(j)
7 13 27 83 28 48 30 57 55 90 53
6 12 46 82 47 67 49 77 74 31 72
5 11 65 81 66 17 68 35 93 50 91
4 10 84 102 85 20 86 54 32 69
3 9 45 101 18 22 19 73 51 88
2 8 44 100 41 40 16 92 33 37
1 23 43 42 60 59 87 76 52 56
0 24 64 61 79 78 39 96 70 75
15 25 63 80 98 97 38 95 71 94
14 26 62 99 29 21 58 36 89 34
Table B.2: Ordering of the speech encoder bits for the 5.15 kbit/s mode: table2(j)
1 9 27 96 53 47 23 64 100 57 101 59
4 11 73 117 99 68 22 65 40 82 42 84
5 12 26 31 32 93 18 44 61 103 63 105
3 14 72 77 78 114 17 90 86 35 88 37
6 10 30 52 33 46 20 25 107 56 109 58
7 16 76 98 79 67 24 45 39 81 41 83
13 74 97 70 69 113 43 91 85 34 87
Table B.3: Ordering of the speech encoder bits for the 5.9 kbit/s mode: table3(j)
ETSI 3GPP
16
4 15 26 85 76 87 17 101 128 96 35 98 40 90
5 14 30 111 130 23 24 54 74 39 89 43 94
6 10 84 48 52 53 25 79 49 64 114 68 119
13 28 16 73 77 78 50 108 45 93 34 97 37
Table B.4: Ordering of the speech encoder bits for the 6.7 kbit/s mode: table4(j)
Table B.5: Ordering of the speech encoder bits for the 7.4 kbit/s mode: table5(j)
Table B.6: Ordering of the speech encoder bits for the 7.95 kbit/s mode: table6(j)
ETSI 3GPP
17
6 13 30 161 158 163 18 79 166 202 83 103 131 143 183 39 59 196 134 106 46
3 10 116 68 201 109 24 126 71 76 84 127 147 181 185 50 43 178 144 99
2 9 117 69 32 155 25 125 113 165 94 128 148 180 193 40 60 187 145 62
1 8 118 108 33 198 37 124 110 81 101 138 142 182 175 52 44 188 105 47
0 26 119 111 121 19 36 123 114 82 102 137 150 172 192 42 54 151 90 57
16 27 120 112 122 23 35 169 159 92 96 139 132 184 176 41 194 136 100 64
15 28 72 154 74 21 34 168 156 91 104 129 149 174 186 51 179 146 107 45
Table B.7: Ordering of the speech encoder bits for the 10.2 kbit/s mode: table7(j)
0 10 19 141 146 200 136 241 198 153 55 162 36 60 108 119 163 211 222 77 126 134 180 229 237
1 11 20 39 44 48 189 91 29 203 105 212 37 66 107 118 169 210 221 82 125 133 185 228 236
2 12 21 142 147 98 239 194 30 89 158 63 54 65 106 157 168 209 73 81 124 176 184 227 96
3 13 22 40 45 151 87 92 31 139 208 113 53 64 112 156 167 215 72 80 129 175 183 232 199
4 14 24 143 148 201 137 195 32 192 90 166 52 70 111 155 173 214 71 85 128 174 188 231
5 23 25 41 46 49 190 93 33 242 140 216 58 69 110 161 172 213 76 84 127 179 187 230
6 15 26 144 149 99 240 196 34 51 193 67 57 68 116 160 171 219 75 83 132 178 186 235
7 16 27 42 47 152 88 94 35 101 243 117 56 104 115 159 207 218 74 123 131 177 226 234
8 17 28 145 97 202 138 197 50 154 59 170 62 103 114 165 206 217 79 122 130 182 225 233
9 18 38 43 150 86 191 95 100 204 109 220 61 102 120 164 205 223 78 121 135 181 224 238
Table B.8: Ordering of the speech encoder bits for the 12.2 kbit/s mode: table8(j)
ETSI 3GPP
18
History
Document history
V3.0.0 January 2000 Publication
ETSI