|
|
|
题名
|
作者
|
年代
|
出处
|
被引量
|
| 1 | Platform for dynamic virtual auditory environment real-time rendering system显示文摘This paper reports the recent works and progress on a PC and C++ language-based virtual auditory environment(VAE) system platform.By tracing the temporary location and orientation of listener's head and dynamically simulating the acoustic propagation from sound source to two ears,the system is capable of recreating free-field virtual sources at various directions and distances as well as auditory perception in reflective environment via headphone presentation.Schemes for improving VAE performance,including PCA-based(principal components analysis) near-field virtual source synthesis,simulating six degrees of freedom of head movement,are proposed.Especially,the PCA-based scheme greatly reduces the computational cost of multiple virtual sources synthesis.Test demonstrates that the system exhibits improved performances as compared with some existing systems.It is able to simultaneously render up to 280 virtual sources using conventional scheme,and 4500 virtual sources using the PCA-based scheme.A set of psychoacoustic experiments also validate the performance of the system,and at the same time,provide some preliminary results on the research of binaural hearing.The functions of the VAE system is being extended and the system serves as a flexible and powerful platform for future binaural hearing researches and virtual reality applications. | ZHANG ChengYun XIE BoSun | 2013 | Chinese Science Bulletin2013,58,3: | 10 |
| 2 | Head-related transfer function database and its analyses显示文摘Based on the measurements from 52 Chinese subjects (26 males and 26 females), a high-spatial-resolution head-related transfer function (HRTF) database with corre- sponding anthropometric parameters is established. By using the database, cues relating to sound source localization, including interaural time difference (ITD), interaural level difference (ILD), and spectral features introduced by pinna, are analyzed. Moreover, the statistical relationship between ITD and anthropometric parameters is estimated. It is proved that the mean values of maximum ITD for male and female are significantly different, so are those for Chinese and western sub- jects. The difference in ITD is due to the difference in individual anthropometric parameters. It is further proved that the spectral features introduced by pinna strongly depend on individual; while at high frequencies (f≥ 5.5 kHz), HRTFs are left-right asymmetric. This work is instructive and helpful for the research on bin- aural hearing and applications on virtual auditory in future. | XIE BoSun ZHONG XiaoLi RAO Dan LIANG ZhiQiang | 2007 | Science China(Physics,Mechanics & Astronomy)2007,50,3: | 10 |
| 3 | Analyse and sound image localization experiment study on multi-channel planar surround sound system显示文摘In this paper the method of approximate expansion is used to analyse a perfect planar surround sound system, resulting in an order of new and upgrade systems. First reproductinn signals of the perfect system and the characteristics of different orders systems are analysed. The independent transmission signals and decoding (reproduction) equation of the systexns are given. The compatibility among different orders systems and the problem of simplifying output channels are discussed. The problem of signal picking up, recording,transmitting and the possibility of putting the systems into practical use are studied. A sound hoage localization experiment for the systems is carried out in order to study haage localization in relaion to the numbers of transmission signals and output channels. The experimental result is consistemt with the theoretical result. This work lay down a base for practical use. | XIE Bosun and XIE Xingfu (Applied Physics Dept. South china Universityof Technology,Guangzhou 510641) | 1996 | Chinese Journal of Acoustics1996,15,1: | 5 |
| 4 | The cross-correlation of signals and spatial impression in surround sound reproduction显示文摘恰好,在喂信号和由左被创造的听觉的空间印象(ASI ) 的跨关联的系数之间的关系离开了在 5.1 隧道包围并且恰好包围扬声器包围健全系统被 psychoacoustic 实验调查。结果为由左右或左右的前面复制显示出那包围扬声器对,听觉的来源宽度(ASW ) 能被控制喂信号到某程度的 crosscorrelation 系数拓宽。在 ASW 和跨关联的系数之间的量的关系频率依赖者。为由一双侧面的扬声器复制,然而, ASW 不能被控制喂信号的跨关联的系数改变。为由前面复制并且仅仅用中央频率同时并且为粉红色的噪音和八音度噪音包围扬声器对 1kHz,听众包封(LEV ) 的一种强壮的感觉能被控制适当地喂信号的 crosscorrelation 系数获得。为有在 2 kHz 和 4 kHz 的中央频率的八音度乐队噪音,然而, LEV 不能被控制喂信号的 crosscorrelation 系数获得。在之间没有唯一的关系的进一步理论的计算和大小表演内部听觉跨关联(IACC ) 并且在 5.1 条隧道的 ASW 包围健全复制,它可能由于 IACC 的算法计算。进一步试验性的确认被需要为评估 ASI 调查 IACC 的适用性。现在的结果将对有用实际包围记录的健全编程和评估。 | SHI Bei XIE Bosun | 2010 | Chinese Journal of Acoustics2010,29,3: | 4 |
| 5 | Analysis of the spatial discrimination threshold of head-related transfer function magnitude显示文摘A binaural-loudness-model-based method for evaluating the spatial discrimination threshold of magnitudes of head-related transfer function(HRTF) is proposed.As the input of the binaural loudness model,the HRTF magnitude variations caused by spatial position variations were firstly calculated from a high-resolution HRTF dataset.Then,three perceptualrelevant parameters,namely interaural loudness level difference,binaural loudness level spectra,and total binaural loudness level,were derived from the binaural loudness model.Finally,the spatial discrimination thresholds of HRTF magnitude were evaluated according to just-noticedifference of the above-mentioned perceptual-relevant parameters.A series of psychoacoustic experiments was also conducted to obtain the spatial discrimination threshold of HRTF magnitudes.Results indicate that the threshold derived from the proposed binaural-loudness-modelbased method is consistent with that obtained from the traditional psychoacoustic experiment,validating the effectiveness of the proposed method. | LIU Yu XIE Bosun YU Guangzheng RUI Yuanqing | 2016 | Chinese Journal of Acoustics2016,35,1: | 2 |
| 6 | Spatial symmetry of head-related transfer function显示文摘为分析头相关的转移功能(HRTF ) 的空间对称的方法被建议。HRTF 的对称上的解剖结构的影响用在 KEMAR 人体模型和人的题目上测量的 HRTF 被调查。结果为 KEMAR 人体模型显示出那, pinnae 破坏在 5 ~ 6 kHz 上面的 HRTF 的前面背对称,当因为耳朵的地点,为人的题目频率把 kHz 归结为 2.5 时。而且在低、中部的频率, HRTF 是近似左右的对称。当当频率增加,好解剖 leftright 差别引起的不对称现象出现时。开始的频率和在 HRTF 的左右的不对称现象的程度依靠个人。分析表明当前的有两耳的模型在是有效的 HRTF 和频率范围的空间对称的特征。 | ZHONG Xiaoli XIE Bosun | 2007 | Chinese Journal of Acoustics2007,26,1: | 2 |
| 7 | Maximal azimuthal resolution needed in measurements of head-related transfer functions 显示文摘 | ZHONG Xiaoli XIE Bosun | 2009 | Journal of the Acoustical Society of America2009,125,4: | 1 |
| 8 | Spatial interpolation of HRTFs and signal mixing for multichannel surround sound显示文摘从空间采样的点, HRTF (头相关的转移功能) 和混合为的信号的空间插值多信道(包围) 声音被分析。首先,他们是算术地相等的,这被证明。为 HRTF 插值的不同方法等价于为多信道的声音混合方法的不同信号。然后,为为立体声的声音多信道的声音和正弦的法律混合的信号的更严格的推导被给。它被指出那试着重建由邻近的线性插值的侧面的 HRTF 是错误的。并且为精确健全图象本地化, HRTF 的邻近的线性插值的常规方程被修订。最后,一些方法在 HRTF 的分析使用了,这也被指出,多信道的声音能互相被用于参考。 | XIE Bosun | 2006 | Chinese Journal of Acoustics2006,25,4: | 1 |
| 9 | Head-related transfer function database and its analyses显示文摘 | XIE Bosun ZHONG Xiaoli RAO Dan | 2006 | Science in China Series G :Physics Mechanics & Astronomy2006,36,5: | 1 |
| 10 | On the low frequency characteristics of head-related transfer function显示文摘在低频率改正测量头相关的转移功能(HRTF ) 的一个方法被建议。由在低频率从球形的头模型分析 HRTF,在 400 Hz 的频率下面, HRTF 的大小是将近不变的,阶段为远、近的地两个都是频率的线性功能,这被证明。因此,如果在 400 Hz 上面的 HRTF 被实验精确地测量,它能由理论模型在低频率改正 HRTF。计算和主观实验的结果证明建议方法的可行性。 | XIE Bosun | 2009 | Chinese Journal of Acoustics2009,28,2: | 1 |
| 11 | Analysis with binaural auditory model and experiment on the timbre of Ambisonics recording and reproduction显示文摘A scheme for analyzing the timbre in spatial sound with binaural auditory model is proposed and the Ambisonics is taken as an example for analysis.Ambisonics is a spatial sound system based on physical sound field reconstruction.The errors and timbre colorations in the final reconstructed sound field depend on the spatial aliasing errors on both the recording and reproducing stages of Ambisonics.The binaural loudness level spectra in Ambisonics reconstruction is calculated by using Moore's revised loudness model and then compared with the result of real sound source,so as to evaluate the timbre coloration in Ambisonics quantitatively.The results indicate that,in the case of ideal independent signals,the high-frequency limit and radius of region without perceived timbre coloration increase with the order of Ambisonics.On the other hand,in the case of recording by microphone array,once the high-frequency limit of microphone array exceeds that of sound field reconstruction,array recording influences little on the binaural loudness level spectra and thus timbre in final reconstruction up to the highfrequency limit of reproduction.Based on the binaural auditory model analysis,a scheme for optimizing design of Ambisonics recording and reproduction is also suggested.The subjective experiment yields consistent results with those of binaural model,thus verifies the effectiveness of the model analysis. | LIU Yang XIE Bosun | 2015 | Chinese Journal of Acoustics2015,34,4: | 1 |
| 12 | 6.1 channel general planar surround sound system显示文摘A new 6.1 channel surround sound system and its two signal mixing methods are proposed. Theoretical and experimental results show that the system is able to recreate 360° sound image in horizontal plane. Especiallys compared with current 5.1 channel system, lateral and rear image of the new system is improved obviously. Therefore it is suitable to be used as a general surround sound system. It is also proved that, the new system is fully compatible with 5.1 channel system, and current methods are available to record 6.1 channel signals. | XIE Bosun (Applied Phgsics Dept., South China University of Technology Guangzhou 510641) | 2001 | Chinese Journal of Acoustics2001,20,2: | 1 |
| 13 | A head-related transfer function model for fast synthesizing multiple virtual sound sources显示文摘A head-related transfer function(HRTF) model for fast and real-time synthesizing multiple virtual sound sources is proposed.A head-related impulse response(HRIR,time-domain version of HRTF) is first decomposed by a two-level wavelet packet and then represented by a model composed of subband filters and reconstruction filters.The coefficients of the subband filters are the zero interpolation of the wavelet coefficients of the HRIR.The coefficients of the reconstruction filters can be calculated from the wavelet function.The model is simplified by applying a threshold method to reduce the wavelet coefficients.The calculated results indicate that for a model with 30 wavelet coefficients,the error of reconstructed HRIR is about 1%.And the result of a psychoacoustic test shows that a model with 35 wavelet coefficients is perceptually indistinguishable from the original HRIR.When multiple virtual sound sources are synthesized simultaneously,the computational cost of the proposed model is much less than the traditional HRTF filters. | LIANG Zhiqiang XIE Bosun | 2013 | Chinese Journal of Acoustics2013,32,2: | 1 |
| 14 | Virtual reproduction of surround sound in frontal space using four loudspeakers显示文摘By considering the contribution of dynamic cue to auditory vertical localization,a method for virtual reproduction of surround sound in frontal space using four actual loudspeakers is proposed.The four actual loudspeakers are arranged in the left-front and right-front directions in the horizontal plane,as well as left-front-up,right-front-up directions in a higher elevation plane,respectively.Transaural signal processing is used to convert the multichannel sound signals to the signals for the four actual loudspeakers.Virtual reproduction of 9.1 channel sound is taken as an example.The analysis of binaural pressures and corresponding localization cues using head-related transfer functions indicates that the current method is able to recreate correct interaural time difference and its dynamic variation with head turning,and thus able to create appropriate binaural cue for lateral localization and dynamic cue for vertical localization.Results of psychoacoustic experiment indicate that the current method is able to recreate stable horizontal and vertical virtual source within the frontal-hemispherical directions.Therefore,combined with transaural processing,the four loudspeakers arrangement is enough to reproduce the vertical localization information in frontal space and realize the down-mixing and simplification of multichannel spatial surround sound. | XIE Bosun LIU Lulu ZHANG Chengyun | 2021 | Chinese Journal of Acoustics2021,40,2: | 1 |
| 15 | Design and validation on a multiple sound source fast-measurement system of near-field head-related transfer functions显示文摘 | YU Guangzheng LIU Yu XIE Bosun | 2018 | Chinese Journal of Acoustics2018,37,2: | 1 |
| 16 | Spatial Sound——History,Principle,Progress and Challenge显示文摘The aim of spatial sound or spatial audio is to reproduce the spatial information of sound,so as to recreate the desired spatial auditory events or perceptions.Recently,spatial sound becomes a hot topic in the fields of acoustics,signal processing,and communication.A series of spatial sound techniques have been developed and applied to a wide area of scientific research,engineering,and amusement.The history,principle,progress and applications of spatial sound technique are comprehensively reviewed in this article.Especially,spatial sound techniques based on different principles are united within the framework of spatial sampling and reconstruction theorem of sound field.The challenges and prospects of spatial sound are also addressed. | XIE Bosun | 2020 | Chinese Journal of Electronics2020,29,3: | 0 |
| 17 | Interchannel phase difference and stereo sound image localization显示文摘By considering higher order approximation to the interaural phase difference, a more general localization equation for stereo sound image with interchannel phase difference is derived. At very low frequency or low interchannel phase difference, the equation can be simplified to Makita theory. In general, image position is obviously affected by frequency.It is shown that image position varying with freqllency is the main reason for image width broadening in stereo reproduction with interchannel phase difference. And an extra interaural sound level difference caused by interchannel phase difference is the main reason for image naturalness degrading. In practice, it is necessary to reduce the interchannel phase difference,at least, to less than 60°. | XIE Bosun(Applied Physics Dept., South China University of Technology Guangzhou .510641) | 1998 | Chinese Journal of Acoustics1998,17,1: | 0 |
| 18 | An individualized interaural time difference model based on spherical harmonic function expansion显示文摘A spatially continuous and individualized interaural time difference(ITD) model is proposed based on spherical harmonic expansion,and the spatial sampling theorem of ITD is derived.Analyses on ITDs from 52 human subjects demonstrate that the proposed ITD model, which is represented as a weighted sum of spherical harmonics up to a degree of 6,is adequate for describing the spatial characteristics of ITD including left-right symmetry and front-back asymmetry.The spatially continuous ITD for a specific subject can be accurately reconstructed from 39 measured ITDs at different spatial directions,or approximately evaluated from four head- and pinna-related anthropometric parameters.The proposed ITD model is superior to the existing ITD models in aspects of model structure and accuracy,and is helpful to simplification of ITD measurement as well as anthropometry-based customization of ITD. | ZHONG Xiaoli XIE Bosun | 2013 | Chinese Journal of Acoustics2013,32,3: | 0 |
| 19 | Influence of the Number of Loudspeakers on the Timbre in Horizontal and Mixed-Order Ambisoncis Reproduction显示文摘Ambisonics is a series of flexible spatial sound reproduction systems based on spatial harmonics decomposition of sound field.Traditional horizontal and spatial Ambisonics reconstruct horizontal and spatial sound field with certain order of spatial harmonics,respectively.Both the Shannon-Nyquist spatial sampling frequency limit for accurately reconstructing sound field and the complexity of system increase with the increasing order of Ambisonics.Based on the fact that the horizontal localization resolution of human hearing is higher than vertical resolution,mixed-order Ambisonics(MOA)reconstructs horizontal sound field with higher order spatial harmonics,while reconstructs vertical sound field with lower order spatial harmonics,and thereby reaches a compromise between the perceptual performance and the complexity of system.For a given order horizontal Ambisoncis or MOA reproduction,the number of horizontal loudspeakers is flexible,providing that it exceeds some low limit.By using Moore’s revised loudness model,the present work analyzes the influence of the number of horizontal loudspeakers on timbre both in horizontal Ambisonics and MOA reproduction.The binaural loudness level spectra(BLLS)of Ambisoncis reproduction are calculated and then compared with those of target sound field.The results indicate that below the Shannon-Nyquist limit of spatial sampling,increasing the number of horizontal loudspeakers influence little on BLLS then timbre.Above the limit,however,the BLLS for Ambisoncis reproduction deviate from those of target sound field.The extent of deviation depends on both the direction of target sound field and the number of loudspeakers.Increasing the number of horizontal loudspeakers may increase the change of BLLS then timbre in some cases,but reduce the change in some other cases.For MOA,the influence of the number of horizontal loudspeakers on BLLS and timbre reduces when virtual source departs from horizontal plane to the high or low elevation.The subjective evaluation experiment also validates the analysis. | Haiming Mai Bosun Xie Jianliang Jiang Dan Rao Yang Liu | 2019 | Sound & Vibration2019,53,3: | 0 |
| 20 | Multiple Scattering Between Adjacent Sound Sources in Head-Related Transfer Function Measurement System显示文摘To accelerate head-related transfer functions(HRTFs)measurement,two or more independent sound sources are usually employed in the measurement system.However,the multiple scattering between adjacent sound sources may influence the accuracy of measurement.On the other hand,the directivity of sound source could induce measurement error.Therefore,a model consisting of two spherical sound sources with approximate omni-directivity and a rigid-spherical head is proposed to evaluate the errors in HRTF measurement caused by multiple scattering between sources.An example of analysis using multipole re-expansion indicates that the error of ipsilateral HRTFs are within the bound of±1.0 dB below a frequency of 20 kHz,provided that the sound source radius does not exceed 0.025 m,the source distance relative to head center is not less than 0.5 m,and the angular interval between two adjacent sources is not less than 20 degrees.Similar conclusions under different conditions can also be analyzed and discussed by using this calculation method.Furthermore,the results are verified by measurements of HRTFs for a rigid sphere and a KEMAR artificial head. | Guangzheng Yu Yu Liu Bosun Xie Huali Zhou | 2019 | Sound & Vibration2019,53,4: | 0 |