Learning High-Frequency Functions Made Easy with Sinusoidal Positional Encoding

التفاصيل البيبلوغرافية
العنوان: Learning High-Frequency Functions Made Easy with Sinusoidal Positional Encoding
المؤلفون: Sun, Chuanhao, Yuan, Zhihang, Xu, Kai, Mai, Luo, N, Siddharth, Chen, Shuo, Marina, Mahesh K.
سنة النشر: 2024
المجموعة: Computer Science
مصطلحات موضوعية: Computer Science - Machine Learning
الوصف: Fourier features based positional encoding (PE) is commonly used in machine learning tasks that involve learning high-frequency features from low-dimensional inputs, such as 3D view synthesis and time series regression with neural tangent kernels. Despite their effectiveness, existing PEs require manual, empirical adjustment of crucial hyperparameters, specifically the Fourier features, tailored to each unique task. Further, PEs face challenges in efficiently learning high-frequency functions, particularly in tasks with limited data. In this paper, we introduce sinusoidal PE (SPE), designed to efficiently learn adaptive frequency features closely aligned with the true underlying function. Our experiments demonstrate that SPE, without hyperparameter tuning, consistently achieves enhanced fidelity and faster training across various tasks, including 3D view synthesis, Text-to-Speech generation, and 1D regression. SPE is implemented as a direct replacement for existing PEs. Its plug-and-play nature lets numerous tasks easily adopt and benefit from SPE.
Comment: 16 pages, Conference, Accepted by ICML 2024
نوع الوثيقة: Working Paper
URL الوصول: http://arxiv.org/abs/2407.09370
رقم الأكسشن: edsarx.2407.09370
قاعدة البيانات: arXiv