Audio (audio) compression comparison [mp3, wma, ogg, atrac] Part 2


Free Download Mp4Gain
picture

Audio (audio) compression comparison [mp3, wma, ogg, atrac] Part 2

AUDIO COMPRESSION

[Sound source used and points of interest]
・ 1kHz sine wave
Check for noise or correction. Investigate if abnormal sounds are mixed by emphasis or noise different from the originally generated range.

Audio Compression

· White noise
Check the frequency characteristics. Use sounds that are emitted at the same level for all sounds from 0 to 20 kHz and see if they are reproduced correctly.

·music
Use real music and investigate the differences with the original.

[Bitrate Settings]
Fixed bit rate: 96kHz, 128kHz, 256kHz, 320kHz Variable bit rate: 96-160kHz, 192-320kHz.
However, depending on the software, 320kHz cannot be set fixed and 350 can be set, or the upper and lower limit bits cannot be specified in the variable, and the sound quality standard can be specified in several steps ( medium sound quality, high sound quality). quality).be. Also, there are some that are configured with average bitrate instead of variable bitrate, so understand that it’s not a completely fair comparison.

[Software used [encoder]]
・ MP3 system
Afternoon Koda Ver.3.11a [gogo.exe ver.3.11]
Lame Ivy Frontend Encoder Ver.2.91 [Lame.exe Ver.3.93]
B’s GOLD Ver.7.12 [Unknown]
RipAudiCO Ver.3.70 [leme_enc.dll Ver.3.93]

・OGG system
oggdropXPd Ver.1.7.11 [Unknown]
B’s GOLD Ver.7.12 [Unknown]

・ WMA
B’s GOLD Ver.7.12 [Unknown]
(For WMA, I tried 3 types of software in my environment, but the result was exactly the same (maybe the encoder itself uses the same thing?) And it corresponds by software Since the bitrate range was narrow, only used a type).

・ ATRAC
nothing special. For ATRAC, we recorded analog from a CD player to an MD deck, optically connected an MD deck to a PC, and measured what was captured by WAV.

· To measure
Wave Space Ver.1.31

【others】
Although it is different from the main theme, I converted it to WAV for the visual measurement of each standard (because WaveSpace only supports wav), but the position where the sound of the WAVized data ends and the total playback. We discovered that there was a difference in time. , so we also investigated it.

3.Hardware 3.
Originally, the equipment used should be described in detail here, but the hardware environment is different for each individual, and this survey is only a guide in the first place, and it will be different if other people do the same. is a possibility of results, I will omit the detailed description of the hardware. (The thing is that I don’t have enough equipment to publish)

【Results of the test】
See the following for a summary of the results of each survey.
・White noise measurement result
・1kHz sine wave measurement result
・Music measurement results
・Simple file size and comment list

[Discussion]
ah There seems to be no big difference in file size (between the same bitrate)
stomach. Sound quality appears to be MP3 < WMA < OGG at low bit rates
(MP3: 128 = WMA: 96 < OGG: 96).
Hare. There is little difference at high bit rates
(there is a slight difference in the treble range, but it seems you won’t notice the difference unless you’re in a very good environment).
Worker. The difference in the encoder software was more than I expected
(especially in MP3)

“My conclusion”
[Less than 128]
If you’re worried about popularization (compatibility), [WMA] is good, and if you basically use it alone, [OGG] is good.
(I am worried about the amount of noise or the correction, but I sacrificed a bit on the sound quality anyway, so I chose the one that covers up to the high range as it is. Also, due to the relationship between ① and ②, mp3 is another with the same sound quality.The file size will be larger than

[With 256 and more]
The variable bit rate (192-320) of [Afternoon Koda] is good.
The fixed 320 is good for sound quality, but there is little difference between the fixed 256 and variable high-quality sound, and it seems that you can barely understand it even if you listen to it. If the sound quality is about the same, the smaller the file size, the better.

[Other impressions]
About OGG
I had high expectations for OGG, but I was concerned about the measurement result at 1 kHz, whether it was noise or correction. However, I find the relationship between sound quality (wide playback band) at low bit rates and file size to be excellent. At high bit rates the sound quality and file size are about the same as MP3s so I think MP3s are advantageous considering the penetration rate but I think they are doing pretty well considering the fact that they have just been developed. expected in the future


Free Download Mp4Gain
picture


Mp4Gain Main Window
picture


Mp4Gain Features
picture


Free Download Mp4Gain
picture

Audio (audio) compression comparison [mp3, wma, ogg, atrac]

Audio (audio) compression comparison [mp3, wma, ogg, atrac]

Compressed Audio

MP3-typed audio, etc., for storing music that was recorded on cassette tapes, music borrowed from CD rental shops or purchased music CDs, or for easy enjoyment with a portable player or car.

compressed audio

More and more people are recording with compression technology. However, there are many standards such as WMA recommended by Microsoft as well as MP3 when it comes to audio compression. Also, since the sound quality and compression rate of each standard change depending on the bit rate setting and the like, there is a wide variety of compression methods depending on the combination of the standard and the setting.

So, I wanted to check what the sound quality and file size would be when recording with which standard and with which settings, and select the standard that suits my purpose, so I took this survey. However, due to the investigation of the ideas of fans, the software and equipment used were covered by those that are freely obtainable in hand or on the net, so the result may be different from the original performance. , but it is only one. Take it as an example.
Since this test focuses on sound quality, it does not test at a low bit rate, which deteriorates sound quality.

Finally, in conducting this survey, I referenced many documents on the Internet. We would like to express our gratitude to each person (individual/corporation) for facilitating us to review materials that have been researched and created with considerable effort from their respective points of view. The sites I mainly referred to will be featured at the bottom of this page, so I recommend that those who are viewing this also take a look.

[Survey outline]
1. 1. Destination standards
As mentioned above, there are many audio compression standards, but here we have limited them to MP3, WMA, OGG, and ATRAC. The standards and reasons for the survey are shown below.

・MP3 ( Moving Picture Experts Group 1 Audio Layer – 3 )
I chose it because it is probably the best known and most popular standard and there are many compatible players for the same reason.

・WMA ( Windows Media Audio ) _ _
It is widely known alongside MP3. Recently, it has become compatible with car audio and DVD players. Also, according to a theory, the same bitrate is rumored to have higher sound quality and compression than MP3, so I chose it.

・OGG (Ogg Vorbis)
It may not be familiar to you yet, but although MP3 requires a license, the number of compatible players is gradually increasing due to the fact that it is unlicensed but offers high sound quality and high compression. Since it is (apparently) high-performance and license-free, it is easy to develop encoders and playback software, so we chose it with the expectation that it will spread in the future.

・ATRAC ( Advanced TR Transform Acoustic Coding ) _ _
This name may not be familiar to you, but you can understand the standard adopted by MD. Many people think that MD has the same high sound quality as CD, and since it is widely used as a storage medium for music, it was used as a reference for comparison.

・ Reason for not targeting other standards
There are many compression standards in addition to the above, but there are few compatible software and players, and considering the interaction with others (although I cannot say publicly), I judged that the comparison with the three types above is adequate. In addition, there is a standard called OpenMG (ATRAC3) recommended by SONY, etc., and there is no need to adopt other than SONY in mobile players, etc., but there are still few (limited) supported players, and recording is done. except for VAIO users, since it is difficult to do so, it was excluded from the target.

2. 2. Survey method
The three types of sounds selected for the survey were converted to various bit rates of each standard, visually compared to the original sounds, and listened to and evaluated. Also, I heard rumors that although the standard is the same, there are differences depending on the conversion software, so I used various types of software (encoder). the detail is just below.

Data lost due to compression is irreversible

Data lost due to compression is irreversible

Audio Compression

In this series, we will focus on the basic knowledge about “sound” that is necessary for video production, and we will make it easy to understand by omitting small and difficult things as much as possible, such as a little general knowledge and sound, including music. . I look forward to delivering it, so I look forward to working with you!

Audio Compression

Now, let’s talk about the first memorable event under the name [Digital Audio Basics]. There are several types of digital audio. Among them, I have summarized the main ones.

[Format types and functions]
◉ Uncompressed format: linear PCM (WAV, BWF, AIFF)
→ The most basic format for digital audio. BWF is a commercial WAV that can contain metadata.

◉ Lossy compression format: P3, AAC (MP4), MQA, etc.
→ Format used mainly for general purposes. In many cases, the information in the uncompressed data is shrunk and compressed. The data capacity is reduced, but the sound quality also deteriorates accordingly. MQA is a new format that is irreversible in terms of data, but reversible in terms of sound quality.

◉ Lossless compression format: FLAC, ALAC, etc.
→ Format mainly used for high-quality listening. It has the reversibility of being able to reproduce exactly the same sound quality as before compression, but the data capacity is not that small.

◉ Others: DSD (DSF, DSDIFF, etc.)
→ It is also called 1-bit audio, but since the concept is fundamentally different from multi-bit audio like linear PCM, it can be compared to “24bit” WAV, etc. in the same line I have not. Currently, it is one of the highest quality formats, but it has the weakness of not being editable.

How is it? I think there are several things, from the familiar ones to the ones you see for the first time, but among them, the one that is most suitable for today’s video production is “Linear PCM”! The reason is as follows.

1. Since it is an uncompressed format, it has excellent sound quality.

2. You can edit like cut and paste.

3. The digital voice tracker is the most popular Ma ‘around the world because the bet, any device, can be managed by software.

Since MP3 and AAC (MP4) are compressed formats, there is a considerable loss in sound quality. Depending on the compression ratio, it may not be obvious at first glance, but it is not suitable as processing-based material such as video production and music production. FLAC and ALAC are lossless compression formats that do not deteriorate sound quality, but do not significantly reduce capacity, and there is no software that can be edited natively (without conversion to other formats), so it is still unsuitable for the production. . DSD was adopted from SACD which appeared in 1999, and is said to be the most analog digital audio today, and it has a smooth texture that is different from linear PCM in terms of sound quality. This format has finally attracted attention in recent years, but due to its mechanism, it has the weakness that it cannot be edited as is, so on the production site, mainly one-shot music recording (recording without editing) and mixing (long-playing recording without editing) and mixing (often used as a master recorder when combining multiple sounds into one stereo or surround sound (also called track down). “Almost Ichi 択 linear PCM” video production, I think I could understand that you can refer to. Of course, if the compressed format does not make you uncomfortable, you can use it, but consider it as an emergency. If you still want quality, you must use linear PCM. The data lost by compression is irreversible. The file that will be the master of the work must be of the highest possible quality. By the way, whether you use WAV or AIFF, the sound quality is almost the same. However, co Considering compatibility, even Mac users can be relieved to use WAV for data transfer.

“16 bit / 44.1 kHz” is
the lowest line of CD quality

Now let’s dive a little deeper into linear PCM. There are “number of quantization bits” (bit depth) and “sample rate” (sample rate) that represent linear PCM specifications. Have you ever seen the notation “16 bit / 44.1 kHz”? This means that the original (analog) audio is sampled (digitized) 44,100 times per second at the 16-bit volume stage (2 raised to 16 = 65,536)! Still, I think it’s “what is this?”, So I tried to sum up the points by comparing it to the video!

What is the best way to use compressed sound sources like MP3, AAC and WMA correctly?

What is the best way to use compressed sound sources like MP3, AAC and WMA correctly?

Audio Compression

When listening to music on a smartphone or iPod, what you seem to know but not understand is digitally compressed sound sources like MP3, AAC, and WMA. Let’s think again about “in what format” and “how much bit rate” is good.

audio compression mp3 acc wma

◆ World standard MP3, Apple standard AAC, Windows standard WMA
You all know that there are various formats of “digital sound sources”.

The best known is the WAV format, which is also used for CDs. Since it is an uncompressed format, there is no deterioration in sound quality and it is very versatile, but the capacity is not small, just over 50MB in 5 minutes.

Therefore, when used with a portable music player such as a smartphone, iPod, or Walkman, it is common to convert (= encode) from WAV to compressed sound sources such as MP3, AAC (M4A / M4P), and WMA.

By the way, compressed sound sources are used from the beginning for download distribution like iTunes. AAC for iTunes, MP3 for Amazon, and WMA for major national distribution sites are mainstream.

・ MP3 …… The oldest compression format established in 1995. There are many supported products, and it is the de facto standard that can be used in any case. “MP4” is a video standard, so don’t get it confused.

・ AAC (M4A / M4P) …… A standard established after MP3, which is a standard format for Apple products such as iPod and iPhone. M4P is a file protected by copyright. AAC is also used for audio on digital terrestrial broadcasts and digital BS on television.

・ WMA …… A format advocated by Microsoft. It has a strong affinity for Windows and many products are also used in voice recorders.

Sample rate and bit rate Part 2

Sample rate and bit rate Part 2

Sample Rate  Bit Rate

Listen and compare

sample rate and bit rate

Why don’t you really ask? In my memory, when I checked it in the past, I remember that it was difficult to distinguish it from the original sound (PCM) at 128 kbps of AAC under the conditions in the table above. I think this varies from person to person, and although I am involved with the audio and sound, I am aware that my ears are not a big problem, so even at a slightly higher rate, it is the same as the sound. original. I’m sure there are people who can tell the difference. At the low 32 kbps, you can clearly see the difference in sound quality. In terms of music, you can understand the metallic sound of the drum hi-hat.
Personally, I think that 44.1 Hz 16-bit (stereo) music CDs can be saved even at 128 kbps (1/10 compression or less) without losing sound quality. About 128 kbps is enough for my ears for both MP3 and AAC.

The bit rate is the compression rate
What happens if you set the encoding bit rate to 256 kbps for 16 kHz audio (monaural with 16 quantization bits)? .. .. Since the compression rate is 100%, it will be the same as the original sound. The sound quality should be the same as the original sound, but it may cause strange behavior depending on the encoders that are available for free (a configuration error may occur).

Sampling frequency Number of quantization bits Number of channels Original sound bit rate (PCM) Remarks
32 kHz 16 1 512 kbps Super Wide Band
24 kHz 16 1 384 kbps
16 kHz 16 1 256 kbps Broadband
8 kHz 16 1 128 kbps Narrowband
Regarding lossy compression of AAC and MP3, I think it is the result of research on how to encode at a low rate, so I personally think that setting a bitrate of 50% or more is not good. Lossless is recommended for compression ratios around 50% (lossless compression, MPEG-4 ALS, etc.). If you only think about saving, even if you compress it as is in PCM, it seems like it’s about half for audio with quiet sections. For lossy compression AAC, MP3, etc., if sound quality is important, about 15-20%, and if high compression is important, about 10% is sufficient sound quality.
Also, for audio purposes less than 10% and 5% is fine, but for audio it is recommended to lower the sample rate rather than suppress the bit rate to 48 kHz or 44.1 kHz (8 kHz or 16 kHz).

Stereo M / S (middle side)
The left and right signals are sum / difference signals. When encoding the sum signal (L + R) and the difference signal (LR) of both channels, the code is used when the correlation between channels is high, such as in stereo. The conversion efficiency is improved. For example, you can improve the coding efficiency of musical voices (L / R in phase, same amplitude).

Intensity stereo
When listening to high frequencies, the bit rate is reduced by combining the high frequency information (quantization coefficient) into one using the property that it is more susceptible to loudness than the L / R time difference.

In the end
Although bit rate may seem like a measure of sound quality, the digital audio field does not specify an encoded bit rate that exceeds the original sound bit rate. In short, I think it is important to use the proper bitrate for each encoder (encoder).

Sample rate and bit rate

Sample rate and bit rate

Sample Rates and Bit Depth

The compression ratio of audio encoding is determined by the bit rate at the time of encoding.

Sample Rate and Bit Depth

Last time I mainly wrote about the original sound bit rate (PCM), but this time I would like to write about the bit rate and compression rate of the encoding.

Specifically, setting a lower bitrate will increase the compression ratio and reduce the size of the file when it is saved. As I wrote last time, the bit rate of the sound source (PCM) before compression is as follows.

PCM bit rate = sample rate (Hz) x number of quantization bits x number of channels
For example, a music CD has the following 44.1 kHz stereo bit rate.

Music CD bit rate: 44100Hz x 16bit x 2ch (stereo) = 1411.2kbps
If it is encoded with MP3, AAC, etc., for example 256 kbps, the compression rate (assuming the original sound is 100%) is approximately 18% and the file size is 1/5 or less.

Encode Music CDs at 256 kbps: 256 kbps / 1,411.2 kbps = approximately 18%
If it’s 4 minutes of music, the file size is as follows.

Original sound: 1,411.2 kbps x 240 seconds = approximately 40.4 MB
Encode at 256 kbps: 256 kbps x 240 seconds = approximately 7.3 MB (+ header)
If a song is about 4 minutes long, 16 songs can be saved on CD650MB as original sound, but if it is encoded at 256 kbps as MP3 or AAC, 89 songs can be recorded.

Original sound: CD650MB / 40.4MB = about 16 songs
256 kbps encoded: CD650MB / 7.3MB = approximately 89 songs
If you check the web, you can compare the sound quality due to the difference in the bit rate. I think all the conditions are the same except the bit rate, but first of all there is a difference in the sound quality depending on the sample rate of the original sound source (PCM) and the number of quantization bits (the bit rate of the original sound changes). At the time of analog to digital conversion (ADC), the sound quality is determined by the conditions. No matter how high the bit rate is encoded for a sound source in poor condition, the sound quality is still poor. Even with the same bit rate, the compression rate changes depending on the number of channels (stereo or monaural). Therefore, strictly speaking, the evaluation of the sound quality cannot be judged only by the difference in the bit rate.
For example, when 48 kHz and 44.1 kHz 16-bit PCM is encoded at 32 kbps to 320 kbps, the compression ratio is as follows.

16-bit PCM compression ratio (when original sound is 100%)
Encoded bit rate 48 kHz stereo (1,536 kbps) 48 kHz monaural (768 kbps) 44.1 kHz stereo (1,411.2 kbps) 44.1 kHz monaural (705.6 kbps)
320 kbps 320/1536 = about 21% About 42% 320 / 1,411.2 = about 23% About 45%
256 kbps 256/1536 = about 17% About 33% 256 / 1,411.2 = about 18% About 36%
192 kbps 192/1536 = about 13% About 25% 192 / 1,411.2 = about 14% About 27%
160 kbps 160/1536 = about 10% About 21% 160 / 1,411.2 = about 11% About 23%
128 kbps 128/1536 = about 8% About 17% 128 / 1,411.2 = about 9% About 18%
64 kbps 64/1536 = about 4% About 8% 64 / 1,411.2 = about 5% About 9%
32 kbps 32/1536 = about 2% About 4% 32 / 1,411.2 = about 2% About 5%
Comparison with the original sound
It’s a bit of a twisted idea, but for example, which one is closer to the original sound, stereo or monaural in the above conditions?
Considering the compression ratio, it is the latter. Of course, stereo is superior to monaural in terms of expression, like expressing the depth of sound, so it makes sense to compare this and evaluate the sound quality, but in encoding, compression is done efficiently using stereo. Since there are algorithms (Stereo M / S and Stereo Intensity), the quality is not half that of monaural and the stereo is compressed efficiently.

What is the best way to use compressed sound sources like MP3, AAC and WMA correctly?

What is the best way to use compressed sound sources like MP3, AAC and WMA correctly?

Audio Compression

When listening to music on a smartphone or iPod, what you seem to know but not understand is digitally compressed sound sources like MP3, AAC, and WMA. Let’s think again about “in what format” and “how much bit rate” is good.

You all know that there are various formats of “digital sound sources”.

The best known is the WAV format, which is also used for CDs. Since it is an uncompressed format, there is no deterioration in sound quality and it is very versatile, but the capacity is not small, just over 50MB in 5 minutes.

Therefore, when used with a portable music player such as a smartphone, iPod, or Walkman, it is common to convert (= encode) from WAV to compressed sound sources such as MP3, AAC (M4A / M4P), and WMA.

By the way, compressed sound sources are used from the beginning for download distribution like iTunes. AAC for iTunes, MP3 for Amazon, and WMA for major national distribution sites are mainstream.

・ MP3 …… The oldest compression format established in 1995. There are many supported products, and it is the de facto standard that can be used in any case. “MP4” is a video standard, so don’t get it confused.

・ AAC (M4A / M4P) …… A standard established after MP3, which is a standard format for Apple products such as iPod and iPhone. M4P is a file protected by copyright. AAC is also used for audio on digital terrestrial broadcasts and digital BS on television.

・ WMA …… A format advocated by Microsoft. It has a strong affinity for Windows and many products are also used in voice recorders.

Based on these characteristics, let’s consider the compression format depending on the device used.

Methods of compression and compression of audio signals Part 3

Methods of compression and compression of audio signals Part 3

Audio Compression

The most popular compression format today is MP3.

The MP3 (MPEG Layer 3) format was developed, after several intermediate formats, by the Fraunhofer Institute in Germany. Actually, the .MP3 format relies on fooling the human ear. After some research, it turned out that human hearing tends to adapt to the appearance of new sounds, which is expressed in an increase in the hearing threshold. Therefore, some sounds are capable of masking (that is, making them subjectively inaudible) others. So in this format, some of the sounds that, according to the corresponding theory, are made inaudible, are simply removed from the general sound. The resulting “semi-finished product” is then encoded using the Hoffman method. Be sure to note that in the MP3 format, programs that compress the sound of the original are not standardized, that is, each competent programmer can implement their own compression scheme. And only the decoders obey the standards, which leads to the fact that the quality of MP3 playback does not always depend on the player that plays this file. Due to the different abilities and predilections of implementers of various encoders, some of them are better at handling symphonic music, some at rock and metal, some at rap and rave, etc.

JointStereo, which is one of the features of MP3, means that instead of encoding stereo as two independent channels, it encodes the call. center channel and the difference from the original stereo channels. Many stereo channel audio components are the same, and encoding them on the common channel allows you to free up additional bandwidth for more detailed encoding of the difference, leading to improved quality.

Be sure to mention the variable bit rate or VBR. This means that the encoder changes the compression ratio on the fly, depending on the nature of the sound. This approach results in a reduction in the final file size or, if quality requirements increase, the same file size produces better sound.

MP3 Pro – Introduced in 2001, the MP3 Pro codec was developed by Coding Technologies in association with Thomson Multimedia. It is MP3 based and as a result it turned out to be fully MP3 backward compatible and only partially forward compatible. It uses SBR (Spectral Band Replication) technology, so the codec provides good quality at low bit rates. However, the encoding quality at medium to high bit rates is inferior to almost all other codecs. As a result, MP3 Pro is used more for streaming on the Internet and demonstrating snippets of new musical compositions.

The MPEG-4 audio standard does not require a single or small set of highly efficient compression schemes, but rather a complex set to perform a wide range of operations, from low-quality speech coding to high-quality music and audio synthesis.

The MPEG-4 family of audio coding algorithms ranges from low quality voice (up to 2 kbps) to high quality audio (64 kbps per channel and higher).

RAW – Yes, it is not just the image format in which some digital cameras write photographs. In fact, RAW is the so-called. “Pure digitization”, which does not contain a title and contains only a sequence of samples of a sound wave. Typically, the scan is stored in 16-bit format.

Shorten is one of the first lossless codecs to appear. For a long time the project “slept sweetly.” However, in 2007, it began to develop again.

TTA (True Audio) – Finally about the most interesting. TTA is being developed by a team of our compatriots. And, I must say, the result of their work is impressive. All in order.

The codec is still quite young, but despite this it contains all the necessary features. We won’t list them again, we’ll just note that the format only lacks support for streaming audio over the network.

The format is open, as well as the source codes of the encoder program. There are compiled versions for Mac and Linux. There should be no compatibility issues during playback either, because there are already plugins for all popular players, as well as DirectShow filters for Windows Media Player. There is a plugin for Adobe Audition, which is important for musicians. For the past 4 years, hardware support has even appeared on players!

WAV – This is the primary audio format for many, many digital audio playback systems and is used as a standard audio file format on personal computers.

Compression and compression methods for audio signals Part 2

Compression and compression methods for audio signals Part 2

audio compression

FLAC is a member of the Xiph.Org codec family. By the way, it also includes the well-known ogg vorbis, one of the best lossy music compression algorithms. As a container for audio data, of course, OGG (files with the extension .ogg) and another open source container – Matroska (files with the extension .mka) are used.

It should be noted right away that both the FLAC format and algorithm are fully open. They are not patented, so they can be used completely free of charge in any program. This is the reason for the wide support for FLAC in players – any serious gamer has a plugin for FLAC. In addition, there are hardware mp3 players that support the FLAC codec.

The FLAC encoder is compiled for most platforms in use, so there should be no compatibility issues on alternative Windows operating systems.

FLAC supports tags in its own “FlacTags” format. There is the ability to encode multi-channel audio, a great advantage over Monkey’s Audio. The format supports any sample rate in the range of 1 Hz (!) To 65,535 Hz. Audio bit depth from 4 (!) To 32 bits.

FLAC is believed to be the most efficient use of system resources when decoding (playing) audio compared to other lossless codecs. Unfortunately, this is achieved at the expense of a significant increase in encoding (compression) time.

The FLAC website is regularly updated and new versions of the codec are released. Overall, FLAC is without a doubt the leader in terms of development activity. This may make it the main format in the future. Well, let’s see …

FLAC is the best option for storing high quality music.

MIDI (Musical Instrument Digital Interface) is a standard for hardware and software that allows you to play (and record) music by executing / recording special commands, as well as the format of the files that contain those commands. The playback device or program is called a MIDI synthesizer (sequencer) and is actually an automatic musical instrument.

Unlike other formats, it does not store the digitized sound, but sets of commands (played notes, links to played instruments, variable sound parameter values) that can be played differently depending on the playback device. The convenience of the MIDI format as a data representation format enables devices that produce automatic arrangements according to given chords, as well as 3D sound visualization applications. Additionally, these files tend to be orders of magnitude smaller than digitized audio of comparable quality.

Monkey’s Audio is a popular lossless digital audio encoding format. Distributed for free along with open source and a suite of encoding and playback software, as well as plugins for popular players. Monkey’s audio files use the following extensions: .ape to store audio and .apl to store metadata. Despite being open source, Monkey’s Audio is not free, as its license imposes significant restrictions on its use.

Audio files compressed with the Monkey audio codec have the extension ‘APE’; As you can see, the monkeys are present not only in the logo or the name (from English monkey: monkey, primate).

The average bit rate in an audio file is 600 to 700 kbps; compare with 128 kbps in MP3. Average compression is 40-50%, depending on the genre of music: if classical or jazz pieces are compressed in the best way, then compositions in the style of trash-metal or something similar “electronic noise” will show the worst result. . For codecs with acceptable quality loss, compression is approximately 80%.

There are four levels of compression. Maximum compression may seem like the only correct solution, although the compression time is quite long. However, you must also take into account the resource consumption of the system that plays the file; for the most compressed file, it is relatively high.

The .APE format provides tag support for searching for songs in your music collection. Another advantage is the verification of the integrity of the file during decoding. Recovery of original compressed .APE wav files is supported.

Monkey’s Audio has a graphical interface for Windows, in other words, a convenient window program to manage the encoding process. The rest of the codecs require the use of the command line or third-party interfaces.

Compression and compression methods of audio signals

Compression and compression methods of audio signals (types, differences, use)

Audio Compression

Basics of the analog-to-digital conversion principle, sound conversion and compression method, existing sound storage formats. Programs to convert and process sound and audio files. Application of these programs in linguistic research.

Bit rate is the amount of information per unit of time. In general, the bit rate is the number of bits that we spend encoding a sound with a duration of 1 second.

Analog-to-digital converter (ADC): A device that converts an input analog signal into a binary code (digital signal). The reverse conversion is done using a DAC (digital-to-analog converter, DAC). Typically, an ADC is an electronic device that converts voltage into a binary digital code. However, some non-electronic devices with digital output must also be classified as ADCs, such as some types of angle-to-code converters. The simplest one-bit binary ADC is a comparator.

The circuit to convert an audio signal from analog to digital:

Sampling is the transformation of continuous images and sound into a set of discrete values ​​in the form of codes.

Quantization is the process of aligning a set of musical notes to a grid.

Compression (compression) of audio data is a process of lowering the bit rate by reducing the statistical and psychoacoustic redundancy of a digital audio signal.

The underlying idea behind all lossy audio compression techniques is to neglect the subtle details of the original sound that are beyond the reach of the human ear.

Codec (CoDec) is an abbreviation for compressor and decompressor. Basically, a codec is a collection of files, drivers, and libraries required to package a video or audio file into a compressed format and play the compressed file.

Formats:

AAC (Advanced Audio Coding) is an audio file format with less quality loss when encoding than MP3 of the same size. The format also allows you to compress without losing the quality of the source (ALAC AAC profile).

AAC (Advanced Audio Coding) was originally created as a successor to MP3 with improved encoding quality. The AAC format, officially known as ISO / IEC 13818-7, was released in 1997 as the new seventh part of the MPEG-2 family. There is also the AAC format known as MPEG-4

Apple AIFF: This file type is standard for Apple Macintosh systems and sound processing systems built on top of it. Apple AIFF stands for Audio Interchange File Format, an audio interchange file format, it is somewhat similar to WAV. Its peculiarity is that it allows you to place additional information along with the sound wave, in particular WaveTable samples (examples of the instrument sound together with synthesizer parameters), which improves the quality of the final result. Although today Apple computers are capable of playing files of almost any format, including MP3.

FLAC (Free Lossless Audio Codec) is a popular free codec for audio compression. Unlike lossy Ogg Vorbis, MP3 and AAC codecs, it does not remove any information from the audio stream and is suitable for both daily listening and archiving of audio collection. Today, the FLAC format is compatible with many audio applications.